Not All Needles Are Found: How Fact Distribution and Don't Make It Up Prompts Shape Literal Extraction, Logical Inference, and Hallucination Risks in Long-Context LLMs
Fuente:
arXiv
Salvato in:
| Autori principali: | Ebrahimzadeh, Amirali, Salili, Seyyed M. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Don't Use LLMs to Make Relevance Judgments
di: Soboroff, Ian
Pubblicazione: (2024)
di: Soboroff, Ian
Pubblicazione: (2024)
Don't Let It Hallucinate: Premise Verification via Retrieval-Augmented Logical Reasoning
di: Qin, Yuehan, et al.
Pubblicazione: (2025)
di: Qin, Yuehan, et al.
Pubblicazione: (2025)
Lost in Literalism: How Supervised Training Shapes Translationese in LLMs
di: Li, Yafu, et al.
Pubblicazione: (2025)
di: Li, Yafu, et al.
Pubblicazione: (2025)
Don't Fight Hallucinations, Use Them: Estimating Image Realism using NLI over Atomic Facts
di: Rykov, Elisei, et al.
Pubblicazione: (2025)
di: Rykov, Elisei, et al.
Pubblicazione: (2025)
Attention Is All You Need But You Don't Need All Of It For Inference of Large Language Models
di: Tyukin, Georgy, et al.
Pubblicazione: (2024)
di: Tyukin, Georgy, et al.
Pubblicazione: (2024)
Bayesian Mixture-of-Experts: Towards Making LLMs Know What They Don't Know
di: Li, Albus Yizhuo
Pubblicazione: (2025)
di: Li, Albus Yizhuo
Pubblicazione: (2025)
Don't Retrieve, Generate: Prompting LLMs for Synthetic Training Data in Dense Retrieval
di: Sinha, Aarush
Pubblicazione: (2025)
di: Sinha, Aarush
Pubblicazione: (2025)
Don't Break the Cache: An Evaluation of Prompt Caching for Long-Horizon Agentic Tasks
di: Lumer, Elias, et al.
Pubblicazione: (2026)
di: Lumer, Elias, et al.
Pubblicazione: (2026)
All the Same The Words Don't Go Away
di: Emerson, Caryl
Pubblicazione: (2019)
di: Emerson, Caryl
Pubblicazione: (2019)
Don’t Test Twice, It’s All Right
di: Zoe Raglow, et al.
Pubblicazione: (2025)
di: Zoe Raglow, et al.
Pubblicazione: (2025)
Essential Tremor Therapies Don't Make the GRADE
di: Ludy C. Shih, et al.
Pubblicazione: (2026)
di: Ludy C. Shih, et al.
Pubblicazione: (2026)
Insights into LLM Long-Context Failures: When Transformers Know but Don't Tell
di: Lu, Taiming, et al.
Pubblicazione: (2024)
di: Lu, Taiming, et al.
Pubblicazione: (2024)
Don't Make the LLM Read the Graph: Make the Graph Think
di: Sun, Yuqi, et al.
Pubblicazione: (2026)
di: Sun, Yuqi, et al.
Pubblicazione: (2026)
You Don't Need Prompt Engineering Anymore: The Prompting Inversion
di: Khan, Imran
Pubblicazione: (2025)
di: Khan, Imran
Pubblicazione: (2025)
Don't Judge a Book by its Cover: Testing LLMs' Robustness Under Logical Obfuscation
di: Borah, Abhilekh, et al.
Pubblicazione: (2026)
di: Borah, Abhilekh, et al.
Pubblicazione: (2026)
Context Shapes LLMs Retrieval-Augmented Fact-Checking Effectiveness
di: Bernardelle, Pietro, et al.
Pubblicazione: (2026)
di: Bernardelle, Pietro, et al.
Pubblicazione: (2026)
You Don't Need All Attentions: Distributed Dynamic Fine-Tuning for Foundation Models
di: Ding, Shiwei, et al.
Pubblicazione: (2025)
di: Ding, Shiwei, et al.
Pubblicazione: (2025)
Don't Persist All : Efficient Persistent Data Structures
di: Mahapatra, Pratyush, et al.
Pubblicazione: (2019)
di: Mahapatra, Pratyush, et al.
Pubblicazione: (2019)
Why Don't Prompt-Based Fairness Metrics Correlate?
di: Zayed, Abdelrahman, et al.
Pubblicazione: (2024)
di: Zayed, Abdelrahman, et al.
Pubblicazione: (2024)
What Prompts Don't Say: Understanding and Managing Underspecification in LLM Prompts
di: Yang, Chenyang, et al.
Pubblicazione: (2025)
di: Yang, Chenyang, et al.
Pubblicazione: (2025)
Minimal Model Reasoning in Description Logics: Don't Try This at Home!
di: Di Stefano, Federica, et al.
Pubblicazione: (2025)
di: Di Stefano, Federica, et al.
Pubblicazione: (2025)
Don't Start Over: A Cost-Effective Framework for Migrating Personalized Prompts Between LLMs
di: Zhao, Ziyi, et al.
Pubblicazione: (2026)
di: Zhao, Ziyi, et al.
Pubblicazione: (2026)
Unequal Political Representation of Core and Periphery: How Regions That Don’t Matter Vote for Parties That Don’t Bother
di: Evert Meijers, et al.
Pubblicazione: (2025)
di: Evert Meijers, et al.
Pubblicazione: (2025)
Don't Teach Students How to Organize Knowledge
di: Rebecca Weaver
Pubblicazione: (2025)
di: Rebecca Weaver
Pubblicazione: (2025)
Transformers Don't In-Context Learn Least Squares Regression
di: Hill, Joshua, et al.
Pubblicazione: (2025)
di: Hill, Joshua, et al.
Pubblicazione: (2025)
100-LongBench: Are de facto Long-Context Benchmarks Literally Evaluating Long-Context Ability?
di: Yang, Wang, et al.
Pubblicazione: (2025)
di: Yang, Wang, et al.
Pubblicazione: (2025)
sDPO: Don't Use Your Data All at Once
di: Kim, Dahyun, et al.
Pubblicazione: (2024)
di: Kim, Dahyun, et al.
Pubblicazione: (2024)
Long-horizon Embodied Planning with Implicit Logical Inference and Hallucination Mitigation
di: Liu, Siyuan, et al.
Pubblicazione: (2024)
di: Liu, Siyuan, et al.
Pubblicazione: (2024)
Out-of-Context Abduction: LLMs Make Inferences About Procedural Data Leveraging Declarative Facts in Earlier Training Data
di: Imran, Sohaib, et al.
Pubblicazione: (2025)
di: Imran, Sohaib, et al.
Pubblicazione: (2025)
Don't Get Too Excited -- Eliciting Emotions in LLMs
di: Fazzi, Gino Franco, et al.
Pubblicazione: (2025)
di: Fazzi, Gino Franco, et al.
Pubblicazione: (2025)
Don't be salesmen
Pubblicazione: (1997)
Pubblicazione: (1997)
Don't Ignore Dual Logic Ability of LLMs while Privatizing: A Data-Intensive Analysis in Medical Domain
di: Du, Yanrui, et al.
Pubblicazione: (2023)
di: Du, Yanrui, et al.
Pubblicazione: (2023)
FactSelfCheck: Fact-Level Black-Box Hallucination Detection for LLMs
di: Sawczyn, Albert, et al.
Pubblicazione: (2025)
di: Sawczyn, Albert, et al.
Pubblicazione: (2025)
Don't Hallucinate, Abstain: Identifying LLM Knowledge Gaps via Multi-LLM Collaboration
di: Feng, Shangbin, et al.
Pubblicazione: (2024)
di: Feng, Shangbin, et al.
Pubblicazione: (2024)
NoLiMa: Long-Context Evaluation Beyond Literal Matching
di: Modarressi, Ali, et al.
Pubblicazione: (2025)
di: Modarressi, Ali, et al.
Pubblicazione: (2025)
Technology: "Don't Make Me Think"--A Plea for Simplicity and Transparency
di: Bell, Colleen
Pubblicazione: (2006)
di: Bell, Colleen
Pubblicazione: (2006)
When Corrections Don't Stick: How Attitudes and Knowledge Shape the Persistence of Misinformation About Genetically Modified Food
di: Sebastian Scholz, et al.
Pubblicazione: (2026)
di: Sebastian Scholz, et al.
Pubblicazione: (2026)
Don't Stop Me Now: Embedding Based Scheduling for LLMs
di: Shahout, Rana, et al.
Pubblicazione: (2024)
di: Shahout, Rana, et al.
Pubblicazione: (2024)
Don’t Give It Away: The Hidden Risks of Generative AI for Scientists
di: Brandão, Anarosa
Pubblicazione: (2026)
di: Brandão, Anarosa
Pubblicazione: (2026)
LLM Cyber Evaluations Don't Capture Real-World Risk
di: Lukošiūtė, Kamilė, et al.
Pubblicazione: (2025)
di: Lukošiūtė, Kamilė, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Don't Use LLMs to Make Relevance Judgments
di: Soboroff, Ian
Pubblicazione: (2024) -
Don't Let It Hallucinate: Premise Verification via Retrieval-Augmented Logical Reasoning
di: Qin, Yuehan, et al.
Pubblicazione: (2025) -
Lost in Literalism: How Supervised Training Shapes Translationese in LLMs
di: Li, Yafu, et al.
Pubblicazione: (2025) -
Don't Fight Hallucinations, Use Them: Estimating Image Realism using NLI over Atomic Facts
di: Rykov, Elisei, et al.
Pubblicazione: (2025) -
Attention Is All You Need But You Don't Need All Of It For Inference of Large Language Models
di: Tyukin, Georgy, et al.
Pubblicazione: (2024)