HiFACTMix: A Code-Mixed Benchmark and Graph-Aware Model for EvidenceBased Political Claim Verification in Hinglish
Fuente:
arXiv
Guardado en:
| Autores principales: | Thakur, Rakesh, Sharma, Sneha, Chopra, Gauri |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Code-Mix Sentiment Analysis on Hinglish Tweets
por: Garg, Aashi, et al.
Publicado: (2026)
por: Garg, Aashi, et al.
Publicado: (2026)
TrueGradeAI: Retrieval-Augmented and Bias-Resistant AI for Transparent and Explainable Digital Assessments
por: Thakur, Rakesh, et al.
Publicado: (2025)
por: Thakur, Rakesh, et al.
Publicado: (2025)
Leveraging Weakly Annotated Data for Hate Speech Detection in Code-Mixed Hinglish: A Feasibility-Driven Transfer Learning Approach with Large Language Models
por: Yadav, Sargam, et al.
Publicado: (2024)
por: Yadav, Sargam, et al.
Publicado: (2024)
A Claim Decomposition Benchmark for Long-form Answer Verification
por: Zhang, Zhihao, et al.
Publicado: (2024)
por: Zhang, Zhihao, et al.
Publicado: (2024)
The Alignment Bottleneck in Decomposition-Based Claim Verification
por: Akhter, Mahmud Elahi, et al.
Publicado: (2026)
por: Akhter, Mahmud Elahi, et al.
Publicado: (2026)
MArgE: Meshing Argumentative Evidence from Multiple Large Language Models for Justifiable Claim Verification
por: Ng, Ming Pok, et al.
Publicado: (2025)
por: Ng, Ming Pok, et al.
Publicado: (2025)
Quantum-RAG and PunGPT2: Advancing Low-Resource Language Generation and Retrieval for the Punjabi Language
por: Singh, Jaskaranjeet, et al.
Publicado: (2025)
por: Singh, Jaskaranjeet, et al.
Publicado: (2025)
Optimizing Decomposition for Optimal Claim Verification
por: Lu, Yining, et al.
Publicado: (2025)
por: Lu, Yining, et al.
Publicado: (2025)
Claim Verification in the Age of Large Language Models: A Survey
por: Dmonte, Alphaeus, et al.
Publicado: (2024)
por: Dmonte, Alphaeus, et al.
Publicado: (2024)
Robust Claim Verification Through Fact Detection
por: Jafari, Nazanin, et al.
Publicado: (2024)
por: Jafari, Nazanin, et al.
Publicado: (2024)
CAPTURE: Context-Aware Prompt Injection Testing and Robustness Enhancement
por: Kholkar, Gauri, et al.
Publicado: (2025)
por: Kholkar, Gauri, et al.
Publicado: (2025)
ThinknCheck: Grounded Claim Verification with Compact, Reasoning-Driven, and Interpretable Models
por: Rao, Delip, et al.
Publicado: (2026)
por: Rao, Delip, et al.
Publicado: (2026)
Combating Biomedical Misinformation through Multi-modal Claim Detection and Evidence-based Verification
por: Barone, Mariano, et al.
Publicado: (2025)
por: Barone, Mariano, et al.
Publicado: (2025)
Evergreen: Efficient Claim Verification for Semantic Aggregates
por: Lee, Alexander W., et al.
Publicado: (2026)
por: Lee, Alexander W., et al.
Publicado: (2026)
Verify-in-the-Graph: Entity Disambiguation Enhancement for Complex Claim Verification with Interactive Graph Representation
por: Pham, Hoang, et al.
Publicado: (2025)
por: Pham, Hoang, et al.
Publicado: (2025)
ClaimPKG: Enhancing Claim Verification via Pseudo-Subgraph Generation with Lightweight Specialized LLM
por: Pham, Hoang, et al.
Publicado: (2025)
por: Pham, Hoang, et al.
Publicado: (2025)
Can AI Validate Science? Benchmarking LLMs for Accurate Scientific Claim $\rightarrow$ Evidence Reasoning
por: Javaji, Shashidhar Reddy, et al.
Publicado: (2025)
por: Javaji, Shashidhar Reddy, et al.
Publicado: (2025)
Distill and Align Decomposition for Enhanced Claim Verification
por: Magomere, Jabez, et al.
Publicado: (2026)
por: Magomere, Jabez, et al.
Publicado: (2026)
When Verification Fails: How Compositionally Infeasible Claims Escape Rejection
por: Liu, Muxin, et al.
Publicado: (2026)
por: Liu, Muxin, et al.
Publicado: (2026)
Step-by-Step Fact Verification System for Medical Claims with Explainable Reasoning
por: Vladika, Juraj, et al.
Publicado: (2025)
por: Vladika, Juraj, et al.
Publicado: (2025)
HealthFC: Verifying Health Claims with Evidence-Based Medical Fact-Checking
por: Vladika, Juraj, et al.
Publicado: (2023)
por: Vladika, Juraj, et al.
Publicado: (2023)
Argumentative Large Language Models for Explainable and Contestable Claim Verification
por: Freedman, Gabriel, et al.
Publicado: (2024)
por: Freedman, Gabriel, et al.
Publicado: (2024)
BiDeV: Bilateral Defusing Verification for Complex Claim Fact-Checking
por: Liu, Yuxuan, et al.
Publicado: (2025)
por: Liu, Yuxuan, et al.
Publicado: (2025)
Foundation Models to Unlock Real-World Evidence from Nationwide Medical Claims
por: Ma, Fan, et al.
Publicado: (2026)
por: Ma, Fan, et al.
Publicado: (2026)
DynaCode: A Dynamic Complexity-Aware Code Benchmark for Evaluating Large Language Models in Code Generation
por: Hu, Wenhao, et al.
Publicado: (2025)
por: Hu, Wenhao, et al.
Publicado: (2025)
Socio-Culturally Aware Evaluation Framework for LLM-Based Content Moderation
por: Kumar, Shanu, et al.
Publicado: (2024)
por: Kumar, Shanu, et al.
Publicado: (2024)
ClaimBrush: A Novel Framework for Automated Patent Claim Refinement Based on Large Language Models
por: Kawano, Seiya, et al.
Publicado: (2024)
por: Kawano, Seiya, et al.
Publicado: (2024)
Fine-grained Claim-level RAG Benchmark for Law
por: Das, Souvick, et al.
Publicado: (2026)
por: Das, Souvick, et al.
Publicado: (2026)
MKE-Coder: Multi-Axial Knowledge with Evidence Verification in ICD Coding for Chinese EMRs
por: You, Xinxin, et al.
Publicado: (2025)
por: You, Xinxin, et al.
Publicado: (2025)
HiBench: Benchmarking LLMs Capability on Hierarchical Structure Reasoning
por: Jiang, Zhuohang, et al.
Publicado: (2025)
por: Jiang, Zhuohang, et al.
Publicado: (2025)
Comparing Knowledge Sources for Open-Domain Scientific Claim Verification
por: Vladika, Juraj, et al.
Publicado: (2024)
por: Vladika, Juraj, et al.
Publicado: (2024)
Entropic Claim Resolution: Uncertainty-Driven Evidence Selection for RAG
por: Di Gioia, Davide
Publicado: (2026)
por: Di Gioia, Davide
Publicado: (2026)
Dafny as Verification-Aware Intermediate Language for Code Generation
por: Li, Yue Chen, et al.
Publicado: (2025)
por: Li, Yue Chen, et al.
Publicado: (2025)
MAPLE: Micro Analysis of Pairwise Language Evolution for Few-Shot Claim Verification
por: Zeng, Xia, et al.
Publicado: (2024)
por: Zeng, Xia, et al.
Publicado: (2024)
PoliticsBench: Benchmarking Political Values in Large Language Models with Multi-Turn Roleplay
por: Khetan, Rohan, et al.
Publicado: (2026)
por: Khetan, Rohan, et al.
Publicado: (2026)
Internal Planning in Language Models: Characterizing Horizon and Branch Awareness
por: Ustaomeroglu, Muhammed, et al.
Publicado: (2025)
por: Ustaomeroglu, Muhammed, et al.
Publicado: (2025)
Policy-as-Prompt: Turning AI Governance Rules into Guardrails for AI Agents
por: Kholkar, Gauri, et al.
Publicado: (2025)
por: Kholkar, Gauri, et al.
Publicado: (2025)
CRScore: Grounding Automated Evaluation of Code Review Comments in Code Claims and Smells
por: Naik, Atharva, et al.
Publicado: (2024)
por: Naik, Atharva, et al.
Publicado: (2024)
ClaimIQ at CheckThat! 2025: Comparing Prompted and Fine-Tuned Language Models for Verifying Numerical Claims
por: Anik, Anirban Saha, et al.
Publicado: (2025)
por: Anik, Anirban Saha, et al.
Publicado: (2025)
Courtroom-Style Multi-Agent Debate with Progressive RAG and Role-Switching for Controversial Claim Verification
por: Chowdhury, Masnun Nuha, et al.
Publicado: (2026)
por: Chowdhury, Masnun Nuha, et al.
Publicado: (2026)
Ejemplares similares
-
Code-Mix Sentiment Analysis on Hinglish Tweets
por: Garg, Aashi, et al.
Publicado: (2026) -
TrueGradeAI: Retrieval-Augmented and Bias-Resistant AI for Transparent and Explainable Digital Assessments
por: Thakur, Rakesh, et al.
Publicado: (2025) -
Leveraging Weakly Annotated Data for Hate Speech Detection in Code-Mixed Hinglish: A Feasibility-Driven Transfer Learning Approach with Large Language Models
por: Yadav, Sargam, et al.
Publicado: (2024) -
A Claim Decomposition Benchmark for Long-form Answer Verification
por: Zhang, Zhihao, et al.
Publicado: (2024) -
The Alignment Bottleneck in Decomposition-Based Claim Verification
por: Akhter, Mahmud Elahi, et al.
Publicado: (2026)