Truth Knows No Language: Evaluating Truthfulness Beyond English
Fuente:
arXiv
Salvato in:
| Autori principali: | Figueras, Blanca Calvo, Sagarzazu, Eneko, Etxaniz, Julen, Barnes, Jeremy, Gamallo, Pablo, de-Dios-Flores, Iria, Agerri, Rodrigo |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Critical Questions Generation: Motivation and Challenges
di: Figueras, Blanca Calvo, et al.
Pubblicazione: (2024)
di: Figueras, Blanca Calvo, et al.
Pubblicazione: (2024)
Benchmarking Critical Questions Generation: A Challenging Reasoning Task for Large Language Models
di: Figueras, Banca Calvo, et al.
Pubblicazione: (2025)
di: Figueras, Banca Calvo, et al.
Pubblicazione: (2025)
Lost in Variation? Evaluating NLI Performance in Basque and Spanish Geographical Variants
di: Bengoetxea, Jaione, et al.
Pubblicazione: (2025)
di: Bengoetxea, Jaione, et al.
Pubblicazione: (2025)
Political Leaning Inference through Plurinational Scenarios
di: de Landa, Joseba Fernandez, et al.
Pubblicazione: (2024)
di: de Landa, Joseba Fernandez, et al.
Pubblicazione: (2024)
Physical Commonsense Reasoning for Lower-Resourced Languages and Dialects: a Study on Basque
di: Bengoetxea, Jaione, et al.
Pubblicazione: (2026)
di: Bengoetxea, Jaione, et al.
Pubblicazione: (2026)
Unveiling the Truth and Facilitating Change: Towards Agent-based Large-scale Social Movement Simulation
di: Mou, Xinyi, et al.
Pubblicazione: (2024)
di: Mou, Xinyi, et al.
Pubblicazione: (2024)
Open Generative Large Language Models for Galician
di: Gamallo, Pablo, et al.
Pubblicazione: (2024)
di: Gamallo, Pablo, et al.
Pubblicazione: (2024)
The Consensus Trap: Dissecting Subjectivity and the "Ground Truth" Illusion in Data Annotation
di: Munir, Sheza, et al.
Pubblicazione: (2026)
di: Munir, Sheza, et al.
Pubblicazione: (2026)
XNLIeu: a dataset for cross-lingual NLI in Basque
di: Heredia, Maite, et al.
Pubblicazione: (2024)
di: Heredia, Maite, et al.
Pubblicazione: (2024)
BertaQA: How Much Do Language Models Know About Local Culture?
di: Etxaniz, Julen, et al.
Pubblicazione: (2024)
di: Etxaniz, Julen, et al.
Pubblicazione: (2024)
Scaling Truth: The Confidence Paradox in AI Fact-Checking
di: Qazi, Ihsan A., et al.
Pubblicazione: (2025)
di: Qazi, Ihsan A., et al.
Pubblicazione: (2025)
Multimodal Large Language Models for Low-Resource Languages: A Case Study for Basque
di: Arana, Lukas, et al.
Pubblicazione: (2025)
di: Arana, Lukas, et al.
Pubblicazione: (2025)
A Catalog of Basque Dialectal Resources: Online Collections and Standard-to-Dialectal Adaptations
di: Bengoetxea, Jaione, et al.
Pubblicazione: (2026)
di: Bengoetxea, Jaione, et al.
Pubblicazione: (2026)
HiTZ at VarDial 2025 NorSID: Overcoming Data Scarcity with Language Transfer and Automatic Data Annotation
di: Bengoetxea, Jaione, et al.
Pubblicazione: (2024)
di: Bengoetxea, Jaione, et al.
Pubblicazione: (2024)
Latxa: An Open Language Model and Evaluation Suite for Basque
di: Etxaniz, Julen, et al.
Pubblicazione: (2024)
di: Etxaniz, Julen, et al.
Pubblicazione: (2024)
Truth Sleuth and Trend Bender: AI Agents to fact-check YouTube videos and influence opinions
di: Logé, Cécile, et al.
Pubblicazione: (2025)
di: Logé, Cécile, et al.
Pubblicazione: (2025)
LLM-Generated Fake News Induces Truth Decay in News Ecosystem: A Case Study on Neural News Recommendation
di: Hu, Beizhe, et al.
Pubblicazione: (2025)
di: Hu, Beizhe, et al.
Pubblicazione: (2025)
Evaluating Shortest Edit Script Methods for Contextual Lemmatization
di: Toporkov, Olia, et al.
Pubblicazione: (2024)
di: Toporkov, Olia, et al.
Pubblicazione: (2024)
TruthEval: A Dataset to Evaluate LLM Truthfulness and Reliability
di: Khatun, Aisha, et al.
Pubblicazione: (2024)
di: Khatun, Aisha, et al.
Pubblicazione: (2024)
Beyond English: Unveiling Multilingual Bias in LLM Copyright Compliance
di: Chen, Yupeng, et al.
Pubblicazione: (2025)
di: Chen, Yupeng, et al.
Pubblicazione: (2025)
Beyond Single Ground Truth: Reference Monism as Epistemic Injustice in ASR Evaluation
di: Choi, Anna Seo Gyeong, et al.
Pubblicazione: (2026)
di: Choi, Anna Seo Gyeong, et al.
Pubblicazione: (2026)
KatotohananQA: Evaluating Truthfulness of Large Language Models in Filipino
di: Nery, Lorenzo Alfred, et al.
Pubblicazione: (2025)
di: Nery, Lorenzo Alfred, et al.
Pubblicazione: (2025)
Truth Neurons
di: Li, Haohang, et al.
Pubblicazione: (2025)
di: Li, Haohang, et al.
Pubblicazione: (2025)
Beyond Agreement: Rethinking Ground Truth in Educational AI Annotation
di: Thomas, Danielle R., et al.
Pubblicazione: (2025)
di: Thomas, Danielle R., et al.
Pubblicazione: (2025)
The Grammar of Transformers: A Systematic Review of Interpretability Research on Syntactic Knowledge in Language Models
di: Graichen, Nora, et al.
Pubblicazione: (2026)
di: Graichen, Nora, et al.
Pubblicazione: (2026)
Evaluating Large Language Models for Detecting Antisemitism
di: Patel, Jay, et al.
Pubblicazione: (2025)
di: Patel, Jay, et al.
Pubblicazione: (2025)
Meaning Beyond Truth Conditions: Evaluating Discourse Level Understanding via Anaphora Accessibility
di: Zhu, Xiaomeng, et al.
Pubblicazione: (2025)
di: Zhu, Xiaomeng, et al.
Pubblicazione: (2025)
To Build Our Future, We Must Know Our Past: Contextualizing Paradigm Shifts in Natural Language Processing
di: Gururaja, Sireesh, et al.
Pubblicazione: (2023)
di: Gururaja, Sireesh, et al.
Pubblicazione: (2023)
Do Llamas Work in English? On the Latent Language of Multilingual Transformers
di: Wendler, Chris, et al.
Pubblicazione: (2024)
di: Wendler, Chris, et al.
Pubblicazione: (2024)
Preserving Historical Truth: Detecting Historical Revisionism in Large Language Models
di: Ortu, Francesco, et al.
Pubblicazione: (2026)
di: Ortu, Francesco, et al.
Pubblicazione: (2026)
Emergence of Linear Truth Encodings in Language Models
di: Ravfogel, Shauli, et al.
Pubblicazione: (2025)
di: Ravfogel, Shauli, et al.
Pubblicazione: (2025)
Credibility Governance: A Social Mechanism for Collective Self-Correction under Weak Truth Signals
di: He, Wanying, et al.
Pubblicazione: (2026)
di: He, Wanying, et al.
Pubblicazione: (2026)
TruthStance: An Annotated Dataset of Conversations on Truth Social
di: Ameen, Fathima, et al.
Pubblicazione: (2026)
di: Ameen, Fathima, et al.
Pubblicazione: (2026)
Do LLMs Really Know What They Don't Know? Internal States Mainly Reflect Knowledge Recall Rather Than Truthfulness
di: Cheang, Chi Seng, et al.
Pubblicazione: (2025)
di: Cheang, Chi Seng, et al.
Pubblicazione: (2025)
TruthTorchLM: A Comprehensive Library for Predicting Truthfulness in LLM Outputs
di: Yaldiz, Duygu Nur, et al.
Pubblicazione: (2025)
di: Yaldiz, Duygu Nur, et al.
Pubblicazione: (2025)
A LLM-Based Ranking Method for the Evaluation of Automatic Counter-Narrative Generation
di: Zubiaga, Irune, et al.
Pubblicazione: (2024)
di: Zubiaga, Irune, et al.
Pubblicazione: (2024)
A Platform for Generating Educational Activities to Teach English as a Second Language
di: Rosá, Aiala, et al.
Pubblicazione: (2025)
di: Rosá, Aiala, et al.
Pubblicazione: (2025)
Cross-lingual Argument Mining in the Medical Domain
di: Yeginbergen, Anar, et al.
Pubblicazione: (2023)
di: Yeginbergen, Anar, et al.
Pubblicazione: (2023)
Lemma Dilemma: On Lemma Generation Without Domain- or Language-Specific Training Data
di: Toporkov, Olia, et al.
Pubblicazione: (2025)
di: Toporkov, Olia, et al.
Pubblicazione: (2025)
MedExpQA: Multilingual Benchmarking of Large Language Models for Medical Question Answering
di: Alonso, Iñigo, et al.
Pubblicazione: (2024)
di: Alonso, Iñigo, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Critical Questions Generation: Motivation and Challenges
di: Figueras, Blanca Calvo, et al.
Pubblicazione: (2024) -
Benchmarking Critical Questions Generation: A Challenging Reasoning Task for Large Language Models
di: Figueras, Banca Calvo, et al.
Pubblicazione: (2025) -
Lost in Variation? Evaluating NLI Performance in Basque and Spanish Geographical Variants
di: Bengoetxea, Jaione, et al.
Pubblicazione: (2025) -
Political Leaning Inference through Plurinational Scenarios
di: de Landa, Joseba Fernandez, et al.
Pubblicazione: (2024) -
Physical Commonsense Reasoning for Lower-Resourced Languages and Dialects: a Study on Basque
di: Bengoetxea, Jaione, et al.
Pubblicazione: (2026)