Benchmark comparing Naive RAG, Light RAG, and HippoRAG2-Hier on Japanese river & sediment-control technical standards. 6 conditions (3 RAG × 2 LLMs), AI-as-Judge scoring.
bm25 civil-engineering faiss japanese-nlp rag lightgbm-regressor colbert llm retrieval-augmented-generation ollama hipporag lambda-rank ai-as-judge hipporag2 rag-calibration contextual-late-interaction
-
Updated
Jun 29, 2026 - Python