FFE-Hallu:Hallucinations in Fixed Figurative Expressions:Benchmark of Idioms and Proverbs in the Persian Language
Fuente:
arXiv
Salvato in:
| Autori principali: | Hosseini, Faezeh, Yousefzadeh, Mohammadali, Yaghoobzadeh, Yadollah |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Large Language Models for Persian $ \leftrightarrow $ English Idiom Translation
di: Rezaeimanesh, Sara, et al.
Pubblicazione: (2024)
di: Rezaeimanesh, Sara, et al.
Pubblicazione: (2024)
PerHalluEval: Persian Hallucination Evaluation Benchmark for Large Language Models
di: Hosseini, Mohammad, et al.
Pubblicazione: (2025)
di: Hosseini, Mohammad, et al.
Pubblicazione: (2025)
GhazalBench: Usage-Grounded Evaluation of LLMs on Persian Ghazals
di: Kalhor, Ghazal, et al.
Pubblicazione: (2026)
di: Kalhor, Ghazal, et al.
Pubblicazione: (2026)
Comparative Study of Multilingual Idioms and Similes in Large Language Models
di: Khoshtab, Paria, et al.
Pubblicazione: (2024)
di: Khoshtab, Paria, et al.
Pubblicazione: (2024)
Evaluating the Creativity of LLMs in Persian Literary Text Generation
di: Tourajmehr, Armin, et al.
Pubblicazione: (2025)
di: Tourajmehr, Armin, et al.
Pubblicazione: (2025)
Counterfactuals As a Means for Evaluating Faithfulness of Attribution Methods in Autoregressive Language Models
di: Kamahi, Sepehr, et al.
Pubblicazione: (2024)
di: Kamahi, Sepehr, et al.
Pubblicazione: (2024)
Extending LLMs to New Languages: A Case Study of Llama and Persian Adaptation
di: Sani, Samin Mahdizadeh, et al.
Pubblicazione: (2024)
di: Sani, Samin Mahdizadeh, et al.
Pubblicazione: (2024)
Layer-wise Positional Bias in Short-Context Language Modeling
di: Rahimi, Maryam, et al.
Pubblicazione: (2026)
di: Rahimi, Maryam, et al.
Pubblicazione: (2026)
Explanations of Large Language Models Explain Language Representations in the Brain
di: Rahimi, Maryam, et al.
Pubblicazione: (2025)
di: Rahimi, Maryam, et al.
Pubblicazione: (2025)
HalluCana: Fixing LLM Hallucination with A Canary Lookahead
di: Li, Tianyi, et al.
Pubblicazione: (2024)
di: Li, Tianyi, et al.
Pubblicazione: (2024)
HalluLens: LLM Hallucination Benchmark
di: Bang, Yejin, et al.
Pubblicazione: (2025)
di: Bang, Yejin, et al.
Pubblicazione: (2025)
HalluScore: Large Language Model Hallucination Question Answering Benchmark
di: Alansari, Aisha, et al.
Pubblicazione: (2026)
di: Alansari, Aisha, et al.
Pubblicazione: (2026)
SOI Matters: Analyzing Multi-Setting Training Dynamics in Pretrained Language Models via Subsets of Interest
di: Vassef, Shayan, et al.
Pubblicazione: (2025)
di: Vassef, Shayan, et al.
Pubblicazione: (2025)
PerCul: A Story-Driven Cultural Evaluation of LLMs in Persian
di: Monazzah, Erfan Moosavi, et al.
Pubblicazione: (2025)
di: Monazzah, Erfan Moosavi, et al.
Pubblicazione: (2025)
uTeBC-NLP at SemEval-2024 Task 9: Can LLMs be Lateral Thinkers?
di: Sadeghi, Pouya, et al.
Pubblicazione: (2024)
di: Sadeghi, Pouya, et al.
Pubblicazione: (2024)
MasalBench: A Benchmark for Contextual and Cross-Cultural Understanding of Persian Proverbs in LLMs
di: Kalhor, Ghazal, et al.
Pubblicazione: (2026)
di: Kalhor, Ghazal, et al.
Pubblicazione: (2026)
HalluEntity: Benchmarking and Understanding Entity-Level Hallucination Detection
di: Yeh, Min-Hsuan, et al.
Pubblicazione: (2025)
di: Yeh, Min-Hsuan, et al.
Pubblicazione: (2025)
Benchmarking Large Language Models for Persian: A Preliminary Study Focusing on ChatGPT
di: Abaskohi, Amirhossein, et al.
Pubblicazione: (2024)
di: Abaskohi, Amirhossein, et al.
Pubblicazione: (2024)
GenKnowSub: Improving Modularity and Reusability of LLMs through General Knowledge Subtraction
di: Bagherifard, Mohammadtaha, et al.
Pubblicazione: (2025)
di: Bagherifard, Mohammadtaha, et al.
Pubblicazione: (2025)
AgentHallu: Benchmarking Automated Hallucination Attribution of LLM-based Agents
di: Liu, Xuannan, et al.
Pubblicazione: (2026)
di: Liu, Xuannan, et al.
Pubblicazione: (2026)
HalluHard: A Hard Multi-Turn Hallucination Benchmark
di: Fan, Dongyang, et al.
Pubblicazione: (2026)
di: Fan, Dongyang, et al.
Pubblicazione: (2026)
NLP Datasets for Idiom and Figurative Language Tasks
di: Matheny, Blake, et al.
Pubblicazione: (2025)
di: Matheny, Blake, et al.
Pubblicazione: (2025)
HalluDetect: Detecting, Mitigating, and Benchmarking Hallucinations in Conversational Systems in the Legal Domain
di: Anaokar, Spandan, et al.
Pubblicazione: (2025)
di: Anaokar, Spandan, et al.
Pubblicazione: (2025)
Proverbs Run in Pairs: Evaluating Proverb Translation Capability of Large Language Model
di: Wang, Minghan, et al.
Pubblicazione: (2025)
di: Wang, Minghan, et al.
Pubblicazione: (2025)
MedHallu: A Comprehensive Benchmark for Detecting Medical Hallucinations in Large Language Models
di: Pandit, Shrey, et al.
Pubblicazione: (2025)
di: Pandit, Shrey, et al.
Pubblicazione: (2025)
HalluDial: A Large-Scale Benchmark for Automatic Dialogue-Level Hallucination Evaluation
di: Luo, Wen, et al.
Pubblicazione: (2024)
di: Luo, Wen, et al.
Pubblicazione: (2024)
HalluVerse25: Fine-grained Multilingual Benchmark Dataset for LLM Hallucinations
di: Abdaljalil, Samir, et al.
Pubblicazione: (2025)
di: Abdaljalil, Samir, et al.
Pubblicazione: (2025)
HalluZig: Hallucination Detection using Zigzag Persistence
di: Samaga, Shreyas N., et al.
Pubblicazione: (2026)
di: Samaga, Shreyas N., et al.
Pubblicazione: (2026)
A Dual-Axis Taxonomy of Knowledge Editing for LLMs: From Mechanisms to Functions
di: Salehoof, Amir Mohammad, et al.
Pubblicazione: (2025)
di: Salehoof, Amir Mohammad, et al.
Pubblicazione: (2025)
Synthia: Scalable Grounded Persona Generation from Social Media Data
di: Rahimzadeh, Vahid, et al.
Pubblicazione: (2025)
di: Rahimzadeh, Vahid, et al.
Pubblicazione: (2025)
HalluWorld: A Controlled Benchmark for Hallucination via Reference World Models
di: Liu, Emmy, et al.
Pubblicazione: (2026)
di: Liu, Emmy, et al.
Pubblicazione: (2026)
HalluCounter: Reference-free LLM Hallucination Detection in the Wild!
di: Urlana, Ashok, et al.
Pubblicazione: (2025)
di: Urlana, Ashok, et al.
Pubblicazione: (2025)
HalluClean: A Unified Framework to Combat Hallucinations in LLMs
di: Zhao, Yaxin, et al.
Pubblicazione: (2025)
di: Zhao, Yaxin, et al.
Pubblicazione: (2025)
FinReflectKG -- HalluBench: GraphRAG Hallucination Benchmark for Financial Question Answering Systems
di: Kumar, Mahesh, et al.
Pubblicazione: (2026)
di: Kumar, Mahesh, et al.
Pubblicazione: (2026)
HalluMix: A Task-Agnostic, Multi-Domain Benchmark for Real-World Hallucination Detection
di: Emery, Deanna, et al.
Pubblicazione: (2025)
di: Emery, Deanna, et al.
Pubblicazione: (2025)
PolyFrame at MWE-2026 AdMIRe 2: When Words Are Not Enough: Multimodal Idiom Disambiguation
di: Hosseini-Kivanani, Nina
Pubblicazione: (2026)
di: Hosseini-Kivanani, Nina
Pubblicazione: (2026)
HalluSAE: Detecting Hallucinations in Large Language Models via Sparse Auto-Encoders
di: Chen, Boshui, et al.
Pubblicazione: (2026)
di: Chen, Boshui, et al.
Pubblicazione: (2026)
Jawaher: A Multidialectal Dataset of Arabic Proverbs for LLM Benchmarking
di: Magdy, Samar M., et al.
Pubblicazione: (2025)
di: Magdy, Samar M., et al.
Pubblicazione: (2025)
HalluScan: A Systematic Benchmark for Detecting and Mitigating Hallucinations in Instruction-Following LLMs
di: Cherif, Ahmed
Pubblicazione: (2026)
di: Cherif, Ahmed
Pubblicazione: (2026)
HalluShift++: Bridging Language and Vision through Internal Representation Shifts for Hierarchical Hallucinations in MLLMs
di: Nath, Sujoy, et al.
Pubblicazione: (2025)
di: Nath, Sujoy, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Large Language Models for Persian $ \leftrightarrow $ English Idiom Translation
di: Rezaeimanesh, Sara, et al.
Pubblicazione: (2024) -
PerHalluEval: Persian Hallucination Evaluation Benchmark for Large Language Models
di: Hosseini, Mohammad, et al.
Pubblicazione: (2025) -
GhazalBench: Usage-Grounded Evaluation of LLMs on Persian Ghazals
di: Kalhor, Ghazal, et al.
Pubblicazione: (2026) -
Comparative Study of Multilingual Idioms and Similes in Large Language Models
di: Khoshtab, Paria, et al.
Pubblicazione: (2024) -
Evaluating the Creativity of LLMs in Persian Literary Text Generation
di: Tourajmehr, Armin, et al.
Pubblicazione: (2025)