ChatQA: Surpassing GPT-4 on Conversational QA and RAG
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Zihan, Ping, Wei, Roy, Rajarshi, Xu, Peng, Lee, Chankyu, Shoeybi, Mohammad, Catanzaro, Bryan |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ChatQA 2: Bridging the Gap to Proprietary LLMs in Long Context and RAG Capabilities
by: Xu, Peng, et al.
Published: (2024)
by: Xu, Peng, et al.
Published: (2024)
NV-Embed: Improved Techniques for Training LLMs as Generalist Embedding Models
by: Lee, Chankyu, et al.
Published: (2024)
by: Lee, Chankyu, et al.
Published: (2024)
MM-Embed: Universal Multimodal Retrieval with Multimodal LLMs
by: Lin, Sheng-Chieh, et al.
Published: (2024)
by: Lin, Sheng-Chieh, et al.
Published: (2024)
RankRAG: Unifying Context Ranking with Retrieval-Augmented Generation in LLMs
by: Yu, Yue, et al.
Published: (2024)
by: Yu, Yue, et al.
Published: (2024)
Evidence Contextualization and Counterfactual Attribution for Conversational QA over Heterogeneous Data with RAG Systems
by: Roy, Rishiraj Saha, et al.
Published: (2024)
by: Roy, Rishiraj Saha, et al.
Published: (2024)
RAGONITE: Iterative Retrieval on Induced Databases and Verbalized RDF for Conversational QA over KGs with RAG
by: Roy, Rishiraj Saha, et al.
Published: (2024)
by: Roy, Rishiraj Saha, et al.
Published: (2024)
EnronQA: Towards Personalized RAG over Private Documents
by: Ryan, Michael J., et al.
Published: (2025)
by: Ryan, Michael J., et al.
Published: (2025)
Comprehensive Comparison of RAG Methods Across Multi-Domain Conversational QA
by: Alushi, Klejda, et al.
Published: (2026)
by: Alushi, Klejda, et al.
Published: (2026)
InstructRetro: Instruction Tuning post Retrieval-Augmented Pretraining
by: Wang, Boxin, et al.
Published: (2023)
by: Wang, Boxin, et al.
Published: (2023)
Retrieval meets Long Context Large Language Models
by: Xu, Peng, et al.
Published: (2023)
by: Xu, Peng, et al.
Published: (2023)
ExpertGenQA: Open-ended QA generation in Specialized Domains
by: Shahgir, Haz Sameen, et al.
Published: (2025)
by: Shahgir, Haz Sameen, et al.
Published: (2025)
TA-Mem: Tool-Augmented Autonomous Memory Retrieval for LLM in Long-Term Conversational QA
by: Yuan, Mengwei, et al.
Published: (2026)
by: Yuan, Mengwei, et al.
Published: (2026)
Cross-Language Approach for Quranic QA
by: Oshallah, Islam, et al.
Published: (2025)
by: Oshallah, Islam, et al.
Published: (2025)
Wrong Answers Can Also Be Useful: PlausibleQA -- A Large-Scale QA Dataset with Answer Plausibility Scores
by: Mozafari, Jamshid, et al.
Published: (2025)
by: Mozafari, Jamshid, et al.
Published: (2025)
Beyond Retrieval: Ensembling Cross-Encoders and GPT Rerankers with LLMs for Biomedical QA
by: Verma, Shashank, et al.
Published: (2025)
by: Verma, Shashank, et al.
Published: (2025)
Neural at ArchEHR-QA 2026: One Method Fits All: Unified Prompt Optimization for Clinical QA over EHRs
by: Majeedi, Abrar, et al.
Published: (2026)
by: Majeedi, Abrar, et al.
Published: (2026)
Don't Retrieve, Navigate: Distilling Enterprise Knowledge into Navigable Agent Skills for QA and RAG
by: Sun, Yiqun, et al.
Published: (2026)
by: Sun, Yiqun, et al.
Published: (2026)
CRAFT: Training-Free Cascaded Retrieval for Tabular QA
by: Singh, Adarsh, et al.
Published: (2025)
by: Singh, Adarsh, et al.
Published: (2025)
TComQA: Extracting Temporal Commonsense from Text
by: Nair, Lekshmi R, et al.
Published: (2025)
by: Nair, Lekshmi R, et al.
Published: (2025)
Enhancing Retrieval in QA Systems with Derived Feature Association
by: Shah, Keyush, et al.
Published: (2024)
by: Shah, Keyush, et al.
Published: (2024)
ConvKGYarn: Spinning Configurable and Scalable Conversational Knowledge Graph QA datasets with Large Language Models
by: Pradeep, Ronak, et al.
Published: (2024)
by: Pradeep, Ronak, et al.
Published: (2024)
PluriHopRAG: Exhaustive, Recall-Sensitive QA Through Corpus-Specific Document Structure Learning
by: Sveistrys, Mykolas, et al.
Published: (2025)
by: Sveistrys, Mykolas, et al.
Published: (2025)
ArabicaQA: A Comprehensive Dataset for Arabic Question Answering
by: Abdallah, Abdelrahman, et al.
Published: (2024)
by: Abdallah, Abdelrahman, et al.
Published: (2024)
Unlocking Electronic Health Records: A Hybrid Graph RAG Approach to Safe Clinical AI for Patient QA
by: Thio, Samuel, et al.
Published: (2025)
by: Thio, Samuel, et al.
Published: (2025)
Few-Shot Prompting for Extractive Quranic QA with Instruction-Tuned LLMs
by: Basem, Mohamed, et al.
Published: (2025)
by: Basem, Mohamed, et al.
Published: (2025)
ASTRA-QA: A Benchmark for Abstract Question Answering over Documents
by: Wang, Shu, et al.
Published: (2026)
by: Wang, Shu, et al.
Published: (2026)
Few-Shot Multilingual Open-Domain QA from 5 Examples
by: Jiang, Fan, et al.
Published: (2025)
by: Jiang, Fan, et al.
Published: (2025)
Knowledge in Triples for LLMs: Enhancing Table QA Accuracy with Semantic Extraction
by: Sholehrasa, Hossein, et al.
Published: (2024)
by: Sholehrasa, Hossein, et al.
Published: (2024)
Data-Driven Function Calling Improvements in Large Language Model for Online Financial QA
by: Tang, Xing, et al.
Published: (2026)
by: Tang, Xing, et al.
Published: (2026)
Evaluation of retrieval-based QA on QUEST-LOFT
by: Scales, Nathan, et al.
Published: (2025)
by: Scales, Nathan, et al.
Published: (2025)
DragonVerseQA: Open-Domain Long-Form Context-Aware Question-Answering
by: Lahiri, Aritra Kumar, et al.
Published: (2024)
by: Lahiri, Aritra Kumar, et al.
Published: (2024)
Two-Stage Quranic QA via Ensemble Retrieval and Instruction-Tuned Answer Extraction
by: Basem, Mohamed, et al.
Published: (2025)
by: Basem, Mohamed, et al.
Published: (2025)
ChatR1: Reinforcement Learning for Conversational Reasoning and Retrieval Augmented Question Answering
by: Lupart, Simon, et al.
Published: (2025)
by: Lupart, Simon, et al.
Published: (2025)
Optimized Quran Passage Retrieval Using an Expanded QA Dataset and Fine-Tuned Language Models
by: Basem, Mohamed, et al.
Published: (2024)
by: Basem, Mohamed, et al.
Published: (2024)
BioPulse-QA: A Dynamic Biomedical Question-Answering Benchmark for Evaluating Factuality, Robustness, and Bias in Large Language Models
by: Bhattarai, Kriti, et al.
Published: (2026)
by: Bhattarai, Kriti, et al.
Published: (2026)
RecGPT: Generative Personalized Prompts for Sequential Recommendation via ChatGPT Training Paradigm
by: Zhang, Yabin, et al.
Published: (2024)
by: Zhang, Yabin, et al.
Published: (2024)
AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy
by: Liu, Zihan, et al.
Published: (2025)
by: Liu, Zihan, et al.
Published: (2025)
The Reasoning Bottleneck in Graph-RAG: Structured Prompting and Context Compression for Multi-Hop QA
by: Zarrinkia, Yasaman, et al.
Published: (2026)
by: Zarrinkia, Yasaman, et al.
Published: (2026)
AceReason-Nemotron: Advancing Math and Code Reasoning through Reinforcement Learning
by: Chen, Yang, et al.
Published: (2025)
by: Chen, Yang, et al.
Published: (2025)
AMAQA: A Metadata-based QA Dataset for RAG Systems
by: Bruni, Davide, et al.
Published: (2025)
by: Bruni, Davide, et al.
Published: (2025)
Similar Items
-
ChatQA 2: Bridging the Gap to Proprietary LLMs in Long Context and RAG Capabilities
by: Xu, Peng, et al.
Published: (2024) -
NV-Embed: Improved Techniques for Training LLMs as Generalist Embedding Models
by: Lee, Chankyu, et al.
Published: (2024) -
MM-Embed: Universal Multimodal Retrieval with Multimodal LLMs
by: Lin, Sheng-Chieh, et al.
Published: (2024) -
RankRAG: Unifying Context Ranking with Retrieval-Augmented Generation in LLMs
by: Yu, Yue, et al.
Published: (2024) -
Evidence Contextualization and Counterfactual Attribution for Conversational QA over Heterogeneous Data with RAG Systems
by: Roy, Rishiraj Saha, et al.
Published: (2024)