REVIEW 2 cited by
Towards Comprehensive Vietnamese Retrieval-Augmented Generation and Large Language Models
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
This paper presents our contributions towards advancing the state of Vietnamese language understanding and generation through the development and dissemination of open datasets and pre-trained models for Vietnamese Retrieval-Augmented Generation (RAG) and Large Language Models (LLMs).
Forward citations
Cited by 2 Pith papers
-
VN-MTEB: Vietnamese Massive Text Embedding Benchmark
VN-MTEB is a new 41-dataset Vietnamese benchmark for text embeddings, built by machine-translating MTEB datasets with embedding-based and LLM-based quality filters.
-
Optimizing Legal Document Retrieval in Vietnamese with Semi-Hard Negative Mining
A lightweight Bi-Encoder plus Cross-Encoder pipeline with random top-candidate negative sampling achieves 79.1% MRR@10 on Vietnamese legal retrieval.
Discussion (0). Continue with ORCID to comment.