RTI-Bench: A Structured Dataset for Indian Right-to-Information Decision Analysis
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Bose, Joy |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
IMLJD: A Computational Dataset for Indian Matrimonial Litigation Analysis
von: Bose, Joy
Veröffentlicht: (2026)
von: Bose, Joy
Veröffentlicht: (2026)
NewsLens: A Multi-Agent Framework for Adversarial News Bias Navigation
von: Bose, Joy
Veröffentlicht: (2026)
von: Bose, Joy
Veröffentlicht: (2026)
Falkor-IRAC: Graph-Constrained Generation for Verified Legal Reasoning in Indian Judicial AI
von: Bose, Joy
Veröffentlicht: (2026)
von: Bose, Joy
Veröffentlicht: (2026)
Right for Right Reasons: Large Language Models for Verifiable Commonsense Knowledge Graph Question Answering
von: Toroghi, Armin, et al.
Veröffentlicht: (2024)
von: Toroghi, Armin, et al.
Veröffentlicht: (2024)
DimStance: Multilingual Datasets for Dimensional Stance Analysis
von: Becker, Jonas, et al.
Veröffentlicht: (2026)
von: Becker, Jonas, et al.
Veröffentlicht: (2026)
LCFO: Long Context and Long Form Output Dataset and Benchmarking
von: Costa-jussà, Marta R., et al.
Veröffentlicht: (2024)
von: Costa-jussà, Marta R., et al.
Veröffentlicht: (2024)
CRISP: Persistent Concept Unlearning via Sparse Autoencoders
von: Ashuach, Tomer, et al.
Veröffentlicht: (2025)
von: Ashuach, Tomer, et al.
Veröffentlicht: (2025)
SeLeRoSa: Sentence-Level Romanian Satire Detection Dataset
von: Smădu, Răzvan-Alexandru, et al.
Veröffentlicht: (2025)
von: Smădu, Răzvan-Alexandru, et al.
Veröffentlicht: (2025)
DimABSA: Building Multilingual and Multidomain Datasets for Dimensional Aspect-Based Sentiment Analysis
von: Lee, Lung-Hao, et al.
Veröffentlicht: (2026)
von: Lee, Lung-Hao, et al.
Veröffentlicht: (2026)
ManagerBench: Evaluating the Safety-Pragmatism Trade-off in Autonomous LLMs
von: Simhi, Adi, et al.
Veröffentlicht: (2025)
von: Simhi, Adi, et al.
Veröffentlicht: (2025)
EduGuardBench: A Holistic Benchmark for Evaluating the Pedagogical Fidelity and Adversarial Safety of LLMs as Simulated Teachers
von: Jiang, Yilin, et al.
Veröffentlicht: (2025)
von: Jiang, Yilin, et al.
Veröffentlicht: (2025)
EVM-QuestBench: An Execution-Grounded Benchmark for Natural-Language Transaction Code Generation
von: Yang, Pei, et al.
Veröffentlicht: (2026)
von: Yang, Pei, et al.
Veröffentlicht: (2026)
IWLV-Ramayana: A Sarga-Aligned Parallel Corpus of Valmiki's Ramayana Across Indian Languages
von: VP, Sumesh
Veröffentlicht: (2026)
von: VP, Sumesh
Veröffentlicht: (2026)
OPOR-Bench: Evaluating Large Language Models on Online Public Opinion Report Generation
von: Yu, Jinzheng, et al.
Veröffentlicht: (2025)
von: Yu, Jinzheng, et al.
Veröffentlicht: (2025)
MAWARITH: A Dataset and Benchmark for Legal Inheritance Reasoning with LLMs
von: Bouchekif, Abdessalam, et al.
Veröffentlicht: (2026)
von: Bouchekif, Abdessalam, et al.
Veröffentlicht: (2026)
A Dataset for Metaphor Detection in Early Medieval Hebrew Poetry
von: Toker, Michael, et al.
Veröffentlicht: (2024)
von: Toker, Michael, et al.
Veröffentlicht: (2024)
PILA: A Historical-Linguistic Dataset of Proto-Italic and Latin
von: Bothwell, Stephen, et al.
Veröffentlicht: (2024)
von: Bothwell, Stephen, et al.
Veröffentlicht: (2024)
ML-Promise: A Multilingual Dataset for Corporate Promise Verification
von: Seki, Yohei, et al.
Veröffentlicht: (2024)
von: Seki, Yohei, et al.
Veröffentlicht: (2024)
EMO-KNOW: A Large Scale Dataset on Emotion and Emotion-cause
von: Nguyen, Mia Huong, et al.
Veröffentlicht: (2024)
von: Nguyen, Mia Huong, et al.
Veröffentlicht: (2024)
Locations of Characters in Narratives: Andersen and Persuasion Datasets
von: Ozyurt, Batuhan, et al.
Veröffentlicht: (2025)
von: Ozyurt, Batuhan, et al.
Veröffentlicht: (2025)
Y-NQ: English-Yorùbá Evaluation dataset for Open-Book Reading Comprehension and Text Generation
von: Costa-jussà, Marta R., et al.
Veröffentlicht: (2024)
von: Costa-jussà, Marta R., et al.
Veröffentlicht: (2024)
Tracking Semantic Change in Slovene: A Novel Dataset and Optimal Transport-Based Distance
von: Pranjić, Marko, et al.
Veröffentlicht: (2024)
von: Pranjić, Marko, et al.
Veröffentlicht: (2024)
I run as fast as a rabbit, can you? A Multilingual Simile Dialogue Dataset
von: Ma, Longxuan, et al.
Veröffentlicht: (2023)
von: Ma, Longxuan, et al.
Veröffentlicht: (2023)
Text-Based Approaches to Item Difficulty Modeling in Large-Scale Assessments: A Systematic Review
von: Peters, Sydney, et al.
Veröffentlicht: (2025)
von: Peters, Sydney, et al.
Veröffentlicht: (2025)
PaperAudit-Bench: Benchmarking Error Detection in Research Papers for Critical Automated Peer Review
von: Tu, Songjun, et al.
Veröffentlicht: (2026)
von: Tu, Songjun, et al.
Veröffentlicht: (2026)
Scaling Synthetic Logical Reasoning Datasets with Context-Sensitive Declarative Grammars
von: Sileo, Damien
Veröffentlicht: (2024)
von: Sileo, Damien
Veröffentlicht: (2024)
BlasBench: An Open Benchmark for Irish Speech Recognition
von: Raj, Jyoutir, et al.
Veröffentlicht: (2026)
von: Raj, Jyoutir, et al.
Veröffentlicht: (2026)
The GDN-CC Dataset: Automatic Corpus Clarification for AI-enhanced Democratic Citizen Consultations
von: Lequeu, Pierre-Antoine, et al.
Veröffentlicht: (2026)
von: Lequeu, Pierre-Antoine, et al.
Veröffentlicht: (2026)
2M-BELEBELE: Highly Multilingual Speech and American Sign Language Comprehension Dataset
von: Costa-jussà, Marta R., et al.
Veröffentlicht: (2024)
von: Costa-jussà, Marta R., et al.
Veröffentlicht: (2024)
RomanLens: The Role Of Latent Romanization In Multilinguality In LLMs
von: Saji, Alan, et al.
Veröffentlicht: (2025)
von: Saji, Alan, et al.
Veröffentlicht: (2025)
XferBench: a Data-Driven Benchmark for Emergent Language
von: Boldt, Brendon, et al.
Veröffentlicht: (2024)
von: Boldt, Brendon, et al.
Veröffentlicht: (2024)
LegalBench-BR: A Benchmark for Evaluating Large Language Models on Brazilian Legal Decision Classification
von: Neto, Pedro Barbosa de Carvalho
Veröffentlicht: (2026)
von: Neto, Pedro Barbosa de Carvalho
Veröffentlicht: (2026)
UA-Legal-Bench: A Benchmark for Evaluating Large Language Models on Ukrainian Legal Reasoning
von: Ovcharov, Volodymyr
Veröffentlicht: (2026)
von: Ovcharov, Volodymyr
Veröffentlicht: (2026)
Learning to Generate Structured Output with Schema Reinforcement Learning
von: Lu, Yaxi, et al.
Veröffentlicht: (2025)
von: Lu, Yaxi, et al.
Veröffentlicht: (2025)
DecisionBench: A Benchmark for Emergent Delegation in Long-Horizon Agentic Workflows
von: Gao, Yuxuan, et al.
Veröffentlicht: (2026)
von: Gao, Yuxuan, et al.
Veröffentlicht: (2026)
Schema as Parameterized Tools for Universal Information Extraction
von: Liang, Sheng, et al.
Veröffentlicht: (2025)
von: Liang, Sheng, et al.
Veröffentlicht: (2025)
Tokenization and Morphology in Multilingual Language Models: A Comparative Analysis of mT5 and ByT5
von: Dang, Thao Anh, et al.
Veröffentlicht: (2024)
von: Dang, Thao Anh, et al.
Veröffentlicht: (2024)
CR-LT-KGQA: A Knowledge Graph Question Answering Dataset Requiring Commonsense Reasoning and Long-Tail Knowledge
von: Guo, Willis, et al.
Veröffentlicht: (2024)
von: Guo, Willis, et al.
Veröffentlicht: (2024)
Extracting Structured Insights from Financial News: An Augmented LLM Driven Approach
von: Dolphin, Rian, et al.
Veröffentlicht: (2024)
von: Dolphin, Rian, et al.
Veröffentlicht: (2024)
VotIE: Information Extraction from Meeting Minutes
von: Evans, José Pedro, et al.
Veröffentlicht: (2026)
von: Evans, José Pedro, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
IMLJD: A Computational Dataset for Indian Matrimonial Litigation Analysis
von: Bose, Joy
Veröffentlicht: (2026) -
NewsLens: A Multi-Agent Framework for Adversarial News Bias Navigation
von: Bose, Joy
Veröffentlicht: (2026) -
Falkor-IRAC: Graph-Constrained Generation for Verified Legal Reasoning in Indian Judicial AI
von: Bose, Joy
Veröffentlicht: (2026) -
Right for Right Reasons: Large Language Models for Verifiable Commonsense Knowledge Graph Question Answering
von: Toroghi, Armin, et al.
Veröffentlicht: (2024) -
DimStance: Multilingual Datasets for Dimensional Stance Analysis
von: Becker, Jonas, et al.
Veröffentlicht: (2026)