CXMArena: Unified Dataset to benchmark performance in realistic CXM Scenarios
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Garg, Raghav, Sharma, Kapil, Gupta, Karan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
AfroXLMR-Comet: Multilingual Knowledge Distillation with Attention Matching for Low-Resource languages
von: Raju, Joshua Sakthivel, et al.
Veröffentlicht: (2025)
von: Raju, Joshua Sakthivel, et al.
Veröffentlicht: (2025)
Information Extraction in Low-Resource Scenarios: Survey and Perspective
von: Deng, Shumin, et al.
Veröffentlicht: (2022)
von: Deng, Shumin, et al.
Veröffentlicht: (2022)
Rethinking Legal Judgement Prediction in a Realistic Scenario in the Era of Large Language Models
von: Nigam, Shubham Kumar, et al.
Veröffentlicht: (2024)
von: Nigam, Shubham Kumar, et al.
Veröffentlicht: (2024)
Introducing Super RAGs in Mistral 8x7B-v1
von: Thakur, Ayush, et al.
Veröffentlicht: (2024)
von: Thakur, Ayush, et al.
Veröffentlicht: (2024)
Explainable Lane Change Prediction for Near-Crash Scenarios Using Knowledge Graph Embeddings and Retrieval Augmented Generation
von: Manzour, M., et al.
Veröffentlicht: (2025)
von: Manzour, M., et al.
Veröffentlicht: (2025)
DENSE: Longitudinal Progress Note Generation with Temporal Modeling of Heterogeneous Clinical Notes Across Hospital Visits
von: Keerthana, Garapati, et al.
Veröffentlicht: (2025)
von: Keerthana, Garapati, et al.
Veröffentlicht: (2025)
OmniRetrieval: Unified Retrieval across Heterogeneous Knowledge Sources
von: Baek, Jinheon, et al.
Veröffentlicht: (2026)
von: Baek, Jinheon, et al.
Veröffentlicht: (2026)
Towards A Unified View of Answer Calibration for Multi-Step Reasoning
von: Deng, Shumin, et al.
Veröffentlicht: (2023)
von: Deng, Shumin, et al.
Veröffentlicht: (2023)
RankRAG: Unifying Context Ranking with Retrieval-Augmented Generation in LLMs
von: Yu, Yue, et al.
Veröffentlicht: (2024)
von: Yu, Yue, et al.
Veröffentlicht: (2024)
BookWorm: A Dataset for Character Description and Analysis
von: Papoudakis, Argyrios, et al.
Veröffentlicht: (2024)
von: Papoudakis, Argyrios, et al.
Veröffentlicht: (2024)
RAMQA: A Unified Framework for Retrieval-Augmented Multi-Modal Question Answering
von: Bai, Yang, et al.
Veröffentlicht: (2025)
von: Bai, Yang, et al.
Veröffentlicht: (2025)
REGEN: A Dataset and Benchmarks with Natural Language Critiques and Narratives
von: Su, Kun, et al.
Veröffentlicht: (2025)
von: Su, Kun, et al.
Veröffentlicht: (2025)
UniGLM: Training One Unified Language Model for Text-Attributed Graph Embedding
von: Fang, Yi, et al.
Veröffentlicht: (2024)
von: Fang, Yi, et al.
Veröffentlicht: (2024)
InstructIE: A Bilingual Instruction-based Information Extraction Dataset
von: Gui, Honghao, et al.
Veröffentlicht: (2023)
von: Gui, Honghao, et al.
Veröffentlicht: (2023)
LML-DAP: Language Model Learning a Dataset for Data-Augmented Prediction
von: Vadlapati, Praneeth
Veröffentlicht: (2024)
von: Vadlapati, Praneeth
Veröffentlicht: (2024)
HealthBranches: Synthesizing Clinically-Grounded Question Answering Datasets via Decision Pathways
von: Cosentino, Cristian, et al.
Veröffentlicht: (2025)
von: Cosentino, Cristian, et al.
Veröffentlicht: (2025)
Succeeding at Scale: Automated Dataset Construction and Query-Side Adaptation for Multi-Tenant Search
von: Jain, Prateek, et al.
Veröffentlicht: (2026)
von: Jain, Prateek, et al.
Veröffentlicht: (2026)
Multi-EuP: The Multilingual European Parliament Dataset for Analysis of Bias in Information Retrieval
von: Yang, Jinrui, et al.
Veröffentlicht: (2023)
von: Yang, Jinrui, et al.
Veröffentlicht: (2023)
Scaling Test-Time Inference with Policy-Optimized, Dynamic Retrieval-Augmented Generation via KV Caching and Decoding
von: Srinivas, Sakhinana Sagar, et al.
Veröffentlicht: (2025)
von: Srinivas, Sakhinana Sagar, et al.
Veröffentlicht: (2025)
OAEI-LLM-T: A TBox Benchmark Dataset for Understanding Large Language Model Hallucinations in Ontology Matching
von: Qiang, Zhangcheng, et al.
Veröffentlicht: (2025)
von: Qiang, Zhangcheng, et al.
Veröffentlicht: (2025)
Cluster-based Adaptive Retrieval: Dynamic Context Selection for RAG Applications
von: Xu, Yifan, et al.
Veröffentlicht: (2025)
von: Xu, Yifan, et al.
Veröffentlicht: (2025)
IndicMedDialog: A Parallel Multi-Turn Medical Dialogue Dataset for Accessible Healthcare in Indic Languages
von: Nigam, Shubham Kumar, et al.
Veröffentlicht: (2026)
von: Nigam, Shubham Kumar, et al.
Veröffentlicht: (2026)
CKnowEdit: A New Chinese Knowledge Editing Dataset for Linguistics, Facts, and Logic Error Correction in LLMs
von: Fang, Jizhan, et al.
Veröffentlicht: (2024)
von: Fang, Jizhan, et al.
Veröffentlicht: (2024)
Human-Inspired Memory Architecture for LLM Agents
von: Kerestecioglu, Doga, et al.
Veröffentlicht: (2026)
von: Kerestecioglu, Doga, et al.
Veröffentlicht: (2026)
Towards Robust Evaluation: A Comprehensive Taxonomy of Datasets and Metrics for Open Domain Question Answering in the Era of Large Language Models
von: Srivastava, Akchay, et al.
Veröffentlicht: (2024)
von: Srivastava, Akchay, et al.
Veröffentlicht: (2024)
NyayaAnumana & INLegalLlama: The Largest Indian Legal Judgment Prediction Dataset and Specialized Language Model for Enhanced Decision Analysis
von: Nigam, Shubham Kumar, et al.
Veröffentlicht: (2024)
von: Nigam, Shubham Kumar, et al.
Veröffentlicht: (2024)
LegalSeg: Unlocking the Structure of Indian Legal Judgments Through Rhetorical Role Classification
von: Nigam, Shubham Kumar, et al.
Veröffentlicht: (2025)
von: Nigam, Shubham Kumar, et al.
Veröffentlicht: (2025)
Semantic In-Domain Product Identification for Search Queries
von: Sharma, Sanat, et al.
Veröffentlicht: (2024)
von: Sharma, Sanat, et al.
Veröffentlicht: (2024)
Unified Hallucination Detection for Multimodal Large Language Models
von: Chen, Xiang, et al.
Veröffentlicht: (2024)
von: Chen, Xiang, et al.
Veröffentlicht: (2024)
DuoLens: A Framework for Robust Detection of Machine-Generated Multilingual Text and Code
von: Agrawal, Shriyansh, et al.
Veröffentlicht: (2025)
von: Agrawal, Shriyansh, et al.
Veröffentlicht: (2025)
Retrieval Augmented Generation for Domain-specific Question Answering
von: Sharma, Sanat, et al.
Veröffentlicht: (2024)
von: Sharma, Sanat, et al.
Veröffentlicht: (2024)
OneGen: Efficient One-Pass Unified Generation and Retrieval for LLMs
von: Zhang, Jintian, et al.
Veröffentlicht: (2024)
von: Zhang, Jintian, et al.
Veröffentlicht: (2024)
Personalized Product Search Ranking: A Multi-Task Learning Approach with Tabular and Non-Tabular Data
von: Morishetti, Lalitesh, et al.
Veröffentlicht: (2025)
von: Morishetti, Lalitesh, et al.
Veröffentlicht: (2025)
RUST-BENCH: Benchmarking LLM Reasoning on Unstructured Text within Structured Tables
von: Abhyankar, Nikhil, et al.
Veröffentlicht: (2025)
von: Abhyankar, Nikhil, et al.
Veröffentlicht: (2025)
TableGuard -- Securing Structured & Unstructured Data
von: Sharma, Anantha, et al.
Veröffentlicht: (2024)
von: Sharma, Anantha, et al.
Veröffentlicht: (2024)
Retrieval-Augmented Generation as Noisy In-Context Learning: A Unified Theory and Risk Bounds
von: Guo, Yang, et al.
Veröffentlicht: (2025)
von: Guo, Yang, et al.
Veröffentlicht: (2025)
StealthRank: LLM Ranking Manipulation via Stealthy Prompt Optimization
von: Tang, Yiming, et al.
Veröffentlicht: (2025)
von: Tang, Yiming, et al.
Veröffentlicht: (2025)
Collab-RAG: Boosting Retrieval-Augmented Generation for Complex Question Answering via White-Box and Black-Box LLM Collaboration
von: Xu, Ran, et al.
Veröffentlicht: (2025)
von: Xu, Ran, et al.
Veröffentlicht: (2025)
Small Encoders Can Rival Large Decoders in Detecting Groundedness
von: Abbes, Istabrak, et al.
Veröffentlicht: (2025)
von: Abbes, Istabrak, et al.
Veröffentlicht: (2025)
LORE: A Large Generative Model for Search Relevance
von: Lu, Chenji, et al.
Veröffentlicht: (2025)
von: Lu, Chenji, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
AfroXLMR-Comet: Multilingual Knowledge Distillation with Attention Matching for Low-Resource languages
von: Raju, Joshua Sakthivel, et al.
Veröffentlicht: (2025) -
Information Extraction in Low-Resource Scenarios: Survey and Perspective
von: Deng, Shumin, et al.
Veröffentlicht: (2022) -
Rethinking Legal Judgement Prediction in a Realistic Scenario in the Era of Large Language Models
von: Nigam, Shubham Kumar, et al.
Veröffentlicht: (2024) -
Introducing Super RAGs in Mistral 8x7B-v1
von: Thakur, Ayush, et al.
Veröffentlicht: (2024) -
Explainable Lane Change Prediction for Near-Crash Scenarios Using Knowledge Graph Embeddings and Retrieval Augmented Generation
von: Manzour, M., et al.
Veröffentlicht: (2025)