RAG Evaluation

RAG: Retrieval-Augmented Generation

Latest papers 224

All topics
CardsList
  1. Evaluating Chunking Strategies for Retrieval-Augmented Generation on Academic Texts

    Jul 2, 2026Valentin J. J. Kreileder, Johannes Reisinger, Andreas FischerDocument ChunkingRetrieval-Augmented Generation

  2. Dual-Confidence Contrastive Decoding for Retrieval-Augmented Generation

    Jul 1, 2026Raymond Li, Md Tawkat Islam Khondaker, Amirhossein Abaskohi +3Retrieval-Augmented GenerationKnowledge Conflicts in Language Models

  3. 5ting at SemEval-2026 Task 8: Strong End-to-End Multi-Turn RAG via LLM-Based Reranking and Faithfulness Control

    Jun 27, 2026Thien-Qua-T-Nguyen, Chi Hoang, Nguyen Tran +3Retrieval-Augmented GenerationMulti-Turn Dialogue Evaluation

  4. MKG-RAG-Bench: Benchmarking Retrieval in Multimodal Knowledge Graph-Augmented Generation

    Jun 24, 2026Xiaochen Wang, Bao Hoang, Han Liu +2Multimodal RAGRetrieval-Augmented Generation

  5. ProvenAI: Provenance-Native Traces of Evidence in Generated Answers

    Jun 24, 2026Mohammad Faizan, Dalal AlharthiMulti-Hop QAData Provenance

  6. Is GraphRAG Needed? From Basic RAG to Graph-/Agentic Solutions with Context Optimization

    Jun 24, 2026Long Chen, Ryan Razkenari, Yuxuan Zhou +5Retrieval-Augmented GenerationGraph Retrieval-Augmented Generation

  7. To Isolate or to Score? Model-Adaptive Assessment for Cost-Efficient Multi-Agent RAG

    Jun 23, 2026Jungseob Lee, Chanjun Park, Heuiseok LimRetrieval-Augmented GenerationAdaptive Model Routing

  8. When Confidence Takes the Wrong Path: Diagnosing Retrieval-State Lock-In in RAG

    Jun 22, 2026Sahib JulkaRetrieval-Augmented GenerationRAG Evaluation

  9. Fixed RAG Compression Collapses Measured Reader Scaling

    Jun 20, 2026Sugam Panthi, Rabab AbdelfattahRetrieval-Augmented GenerationLLM Evaluation

  10. FactoryLLM: A Safe and Open-Source AI Playground for Evaluating LLMs in Smart Factories

    Jun 12, 2026Yash Pulse, Yong-Bin Kang, Abhik Banerjee +2ManufacturingLLM Evaluation

  11. Carolina Guide: A Multi-Agent RAG System with Institutional Guardrails for Academic Policy Assistance

    Jun 11, 2026Ben Torsion, Jun ZhouMulti-Agent LLM SystemsLLM Guardrails

  12. Rethinking RAG in Long Videos: What to Retrieve and How to Use It?

    Jun 11, 2026Yuho Lee, Jisu Shin, Nicole Hee-Yeon Kim +5Multimodal RAGEgocentric Video Understanding

  13. SafeLLM: Extraction as a Hallucination-Resistant Alternative to Rewriting in Safety-Critical Settings

    Jun 11, 2026Julia Ive, Felix Jozsa, Evridiki Georgaki +6Retrieval-Augmented GenerationLLM Grounding

  14. How Fine-Grained Should a RAG Benchmark Be? A Hierarchical Framework for Synthetic Question Generation

    Jun 11, 2026Chase M. Fensore, Kaustubh Dhole, Jason Fan +2Benchmark DesignSynthetic Data Generation

  15. Findings of the MAGMaR 2026 Shared Task

    Jun 10, 2026Alexander Martin, Dengjia Zhang, Joel Brogan +7Multimodal RAGMultimodal Grounding

  16. When More Documents Hurt RAG: Mitigating Vector Search Dilution with Domain-Scoped, Model-Agnostic Retrieval

    Jun 9, 2026Nabaraj Subedi, Ahmed Abdelaty, Shivanand Venkanna SheshappanavarRetrieval-Augmented GenerationAgentic RAG

  17. Energy-Efficient On-Device RAG on a Mobile NPU: System Design and Benchmark on Snapdragon X Elite

    Jun 9, 2026Zhiyuan Cheng, Longying LaiRetrieval-Augmented GenerationAI Accelerator Inference

  18. AbstRAG: Learning to Abstract for Retrieval Problems

    Jun 8, 2026Lei Xu, Xin Quan, Daniel Pedronette +1Retrieval-Augmented GenerationAdaptive Retrieval