Query Expension for Better Query Embedding using LLMs
-
Updated
Feb 18, 2025 - Python
Query Expension for Better Query Embedding using LLMs
Code and models for the paper "Questions Are All You Need to Train a Dense Passage Retriever (TACL 2023)"
SPRINT Toolkit helps you evaluate diverse neural sparse models easily using a single click on any IR dataset.
Evaluation of BEIR Datasets using ColBERT retrieval model
LoRA fine-tuning of bi-encoder retrievers with hard negatives and cross-encoder distillation, evaluated on NFCorpus.
A genral RAG Search chatbot, with SoTA RAG techniques such as HyDE, Hybrid retrieval with BM25 + RRF and Cross encoder reranking. Evaluated on the BEIR scifact dataset and compared all the different pipelines i tried along the way
Retrieval benchmark that runs in CI without an API key: 4 strategies over BEIR corpora, scored with IR metrics, bootstrap intervals and paired significance tests
Research-grade hybrid retrieval API — BM25 + FAISS + CDF calibration + entropy-weighted fusion + cross-encoder reranking. Benchmarked on BEIR SciFact with bootstrap significance tests.
SERA-VQ: Discrete codes for extreme embedding compression — outperforms PCA+int8 at low memory budgets on BEIR/SciFact
端到端 Hybrid RAG 系统 (BEIR/SciFact):BM25 + Dense + Weighted Fusion + RAGAS 评估 + LLM-as-judge 稳定性研究
Dense retrieval + cross-encoder reranking pipeline benchmarked on BEIR datasets (NDCG, Recall@K, MRR)
A RAG system that replaces standard BM25/FAISS retrieval with a fully learned neural retrieval stack - including a fine-tuned bi-encoder, a cross-encoder reranker, ColBERT-style late interaction scoring, and a locally hosted LLM generator. Built entirely with free and open-source tools.
Is a ternary (2/9)-sigma zero band worth its keep for compressed retrieval? 9 preregistered experiments, honest negatives included.
Cross-Family LLM-Judge Agreement for Institutional RAG: 5 families, 9 judges. Validated on TREC RAG 2024 (kappa=0.4941) + BEIR scifact.
CPU-only retrieval benchmark comparing BM25, dense embeddings, RRF hybrid, and cross-encoder reranking on a video-game corpus -- with blind graded relevance judgments and a quality-vs-cost Pareto analysis.
Official implementation of Density2R, an efficient zero-shot document re-ranking method for RAG, information retrieval, and LLM-based reranking.
Scripts to convert the LegalBench-RAG dataset into the standard IR format
Corpus-level evaluation of diversity-aware retrieval for RAG on BEIR HotpotQA using BM25, DPR, Contriever, ColBERTv2, MMR, clustering, and DPP.
RAG pipeline with a reproducible eval harness: hybrid retrieval, reranking, hallucination metrics, and CI regression gates on retrieval quality
Comparative study of transformer and non-transformer encoder architectures for dense retrieval — mapping the latency × accuracy Pareto frontier on BEIR. Solo research, PES University CSE 2026.
Add a description, image, and links to the beir topic page so that developers can more easily learn about it.
To associate your repository with the beir topic, visit your repo's landing page and select "manage topics."