BHRAM-IL: A Benchmark for Hallucination Recognition and Assessment in Multiple Indian Languages
Fuente:
arXiv
Salvato in:
| Autori principali: | Terdalkar, Hrishikesh, Bhojani, Kirtan, Dongare, Aryan, Behera, Omm Aditya |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Gyan: An Explainable Neuro-Symbolic Language Model
di: Srinivasan, Venkat, et al.
Pubblicazione: (2026)
di: Srinivasan, Venkat, et al.
Pubblicazione: (2026)
Performance of Large Language Models in Supporting Medical Diagnosis and Treatment
di: Sousa, Diogo, et al.
Pubblicazione: (2025)
di: Sousa, Diogo, et al.
Pubblicazione: (2025)
The Battle of LLMs: A Comparative Study in Conversational QA Tasks
di: Rangapur, Aryan, et al.
Pubblicazione: (2024)
di: Rangapur, Aryan, et al.
Pubblicazione: (2024)
HATL: Hierarchical Adaptive-Transfer Learning Framework for Sign Language Machine Translation
di: Shahin, Nada, et al.
Pubblicazione: (2026)
di: Shahin, Nada, et al.
Pubblicazione: (2026)
Graph Repairs with Large Language Models: An Empirical Study
di: Terdalkar, Hrishikesh, et al.
Pubblicazione: (2025)
di: Terdalkar, Hrishikesh, et al.
Pubblicazione: (2025)
DeepContext: Stateful Real-Time Detection of Multi-Turn Adversarial Intent Drift in LLMs
di: Albrethsen, Justin, et al.
Pubblicazione: (2026)
di: Albrethsen, Justin, et al.
Pubblicazione: (2026)
Persistent Identity in AI Agents: A Multi-Anchor Architecture for Resilient Memory and Continuity
di: Menon, Prahlad G.
Pubblicazione: (2026)
di: Menon, Prahlad G.
Pubblicazione: (2026)
Internal APIs Are All You Need: Shadow APIs, Shared Discovery, and the Case Against Browser-First Agent Architectures
di: Tham, Lewis, et al.
Pubblicazione: (2026)
di: Tham, Lewis, et al.
Pubblicazione: (2026)
How Good is ChatGPT in Giving Adaptive Guidance Using Knowledge Graphs in E-Learning Environments?
di: Ocheja, Patrick, et al.
Pubblicazione: (2024)
di: Ocheja, Patrick, et al.
Pubblicazione: (2024)
Real-Time Black-Box Optimization for Dynamic Discrete Environments Using Embedded Ising Machines
di: Kashimata, Tomoya, et al.
Pubblicazione: (2025)
di: Kashimata, Tomoya, et al.
Pubblicazione: (2025)
iTrash: Incentivized Token Rewards for Automated Sorting and Handling
di: Ortega, Pablo, et al.
Pubblicazione: (2025)
di: Ortega, Pablo, et al.
Pubblicazione: (2025)
ResBench: Benchmarking LLM-Generated FPGA Designs with Resource Awareness
di: Guo, Ce, et al.
Pubblicazione: (2025)
di: Guo, Ce, et al.
Pubblicazione: (2025)
UA-Legal-Bench: A Benchmark for Evaluating Large Language Models on Ukrainian Legal Reasoning
di: Ovcharov, Volodymyr
Pubblicazione: (2026)
di: Ovcharov, Volodymyr
Pubblicazione: (2026)
Rethinking Spatio-Temporal Anomaly Detection: A Vision for Causality-Driven Cybersecurity
di: Malarkkan, Arun Vignesh, et al.
Pubblicazione: (2025)
di: Malarkkan, Arun Vignesh, et al.
Pubblicazione: (2025)
Inside the Scaffold: A Source-Code Taxonomy of Coding Agent Architectures
di: Rombaut, Benjamin
Pubblicazione: (2026)
di: Rombaut, Benjamin
Pubblicazione: (2026)
The Stage Comes to You: A Real-Time Tele-Immersive System with 3D Point Clouds and Vibrotactile Feedback
di: Matsumoto, Takahiro, et al.
Pubblicazione: (2025)
di: Matsumoto, Takahiro, et al.
Pubblicazione: (2025)
Robustness of Large Language Models to Perturbations in Text
di: Singh, Ayush, et al.
Pubblicazione: (2024)
di: Singh, Ayush, et al.
Pubblicazione: (2024)
NeuroChat: A Neuroadaptive AI Chatbot for Customizing Learning Experiences
di: Baradari, Dünya, et al.
Pubblicazione: (2025)
di: Baradari, Dünya, et al.
Pubblicazione: (2025)
Text-Based Approaches to Item Difficulty Modeling in Large-Scale Assessments: A Systematic Review
di: Peters, Sydney, et al.
Pubblicazione: (2025)
di: Peters, Sydney, et al.
Pubblicazione: (2025)
Evaluation of QCNN-LSTM for Disability Forecasting in Multiple Sclerosis Using Sequential Multisequence MRI
di: Mayfield, John D., et al.
Pubblicazione: (2024)
di: Mayfield, John D., et al.
Pubblicazione: (2024)
Machine Translation Hallucination Detection for Low and High Resource Languages using Large Language Models
di: Benkirane, Kenza, et al.
Pubblicazione: (2024)
di: Benkirane, Kenza, et al.
Pubblicazione: (2024)
A Case Study of Cross-Lingual Zero-Shot Generalization for Classical Languages in LLMs
di: Akavarapu, V. S. D. S. Mahesh, et al.
Pubblicazione: (2025)
di: Akavarapu, V. S. D. S. Mahesh, et al.
Pubblicazione: (2025)
YETI (YET to Intervene) Proactive Interventions by Multimodal AI Agents in Augmented Reality Tasks
di: Bandyopadhyay, Saptarashmi, et al.
Pubblicazione: (2025)
di: Bandyopadhyay, Saptarashmi, et al.
Pubblicazione: (2025)
Red Teaming for Large Language Models At Scale: Tackling Hallucinations on Mathematics Tasks
di: Buszydlik, Aleksander, et al.
Pubblicazione: (2023)
di: Buszydlik, Aleksander, et al.
Pubblicazione: (2023)
From Noise to Insights: Enhancing Supply Chain Decision Support through AI-Based Survey Integrity Analytics
di: Mani, Bhubalan
Pubblicazione: (2026)
di: Mani, Bhubalan
Pubblicazione: (2026)
Adaptive Activation Cancellation for Hallucination Mitigation in Large Language Models
di: Yocam, Eric, et al.
Pubblicazione: (2026)
di: Yocam, Eric, et al.
Pubblicazione: (2026)
Automating Feedback Analysis in Surgical Training: Detection, Categorization, and Assessment
di: Nasriddinov, Firdavs, et al.
Pubblicazione: (2024)
di: Nasriddinov, Firdavs, et al.
Pubblicazione: (2024)
Fine-Tuned Large Language Models for Logical Translation: Reducing Hallucinations with Lang2Logic
di: Pan, Muyu, et al.
Pubblicazione: (2025)
di: Pan, Muyu, et al.
Pubblicazione: (2025)
Predictive Simultaneous Interpretation: Harnessing Large Language Models for Democratizing Real-Time Multilingual Communication
di: Iida, Kurando, et al.
Pubblicazione: (2024)
di: Iida, Kurando, et al.
Pubblicazione: (2024)
High-Throughput Phenotyping of Clinical Text Using Large Language Models
di: Hier, Daniel B., et al.
Pubblicazione: (2024)
di: Hier, Daniel B., et al.
Pubblicazione: (2024)
LightNobel: Improving Sequence Length Limitation in Protein Structure Prediction Model via Adaptive Activation Quantization
di: Han, Seunghee, et al.
Pubblicazione: (2025)
di: Han, Seunghee, et al.
Pubblicazione: (2025)
Low-Resource Court Judgment Summarization for Common Law Systems
di: Liu, Shuaiqi, et al.
Pubblicazione: (2024)
di: Liu, Shuaiqi, et al.
Pubblicazione: (2024)
Council Mode: A Heterogeneous Multi-Agent Consensus Framework for Reducing LLM Hallucination and Bias
di: Wu, Shuai, et al.
Pubblicazione: (2026)
di: Wu, Shuai, et al.
Pubblicazione: (2026)
Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models
di: Gokdemir, Ozan, et al.
Pubblicazione: (2025)
di: Gokdemir, Ozan, et al.
Pubblicazione: (2025)
MedHal: An Evaluation Dataset for Medical Hallucination Detection
di: Mehenni, Gaya, et al.
Pubblicazione: (2025)
di: Mehenni, Gaya, et al.
Pubblicazione: (2025)
Listen to the Layers: Mitigating Hallucinations with Inter-Layer Disagreement
di: Subbalakshmi, Koduvayur, et al.
Pubblicazione: (2026)
di: Subbalakshmi, Koduvayur, et al.
Pubblicazione: (2026)
Beyond Hallucinations: A Composite Score for Measuring Reliability in Open-Source Large Language Models
di: Salla, Rohit Kumar, et al.
Pubblicazione: (2025)
di: Salla, Rohit Kumar, et al.
Pubblicazione: (2025)
Hallucination or Creativity: How to Evaluate AI-Generated Scientific Stories?
di: Argese, Alex, et al.
Pubblicazione: (2026)
di: Argese, Alex, et al.
Pubblicazione: (2026)
PolyGLU: State-Conditional Activation Routing in Transformer Feed-Forward Networks
di: Medeiros, Daniel Nobrega
Pubblicazione: (2026)
di: Medeiros, Daniel Nobrega
Pubblicazione: (2026)
EQ-Bench: An Emotional Intelligence Benchmark for Large Language Models
di: Paech, Samuel J.
Pubblicazione: (2023)
di: Paech, Samuel J.
Pubblicazione: (2023)
Documenti analoghi
-
Gyan: An Explainable Neuro-Symbolic Language Model
di: Srinivasan, Venkat, et al.
Pubblicazione: (2026) -
Performance of Large Language Models in Supporting Medical Diagnosis and Treatment
di: Sousa, Diogo, et al.
Pubblicazione: (2025) -
The Battle of LLMs: A Comparative Study in Conversational QA Tasks
di: Rangapur, Aryan, et al.
Pubblicazione: (2024) -
HATL: Hierarchical Adaptive-Transfer Learning Framework for Sign Language Machine Translation
di: Shahin, Nada, et al.
Pubblicazione: (2026) -
Graph Repairs with Large Language Models: An Empirical Study
di: Terdalkar, Hrishikesh, et al.
Pubblicazione: (2025)