Can LLMs Take Retrieved Information with a Grain of Salt?
Fuente:
arXiv
Salvato in:
| Autori principali: | Shayegh, Behzad, Ahmed, Mohamed Osama, Tung, Fred, Feng, Leo |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Tree-Averaging Algorithms for Ensemble-Based Unsupervised Discontinuous Constituency Parsing
di: Shayegh, Behzad, et al.
Pubblicazione: (2024)
di: Shayegh, Behzad, et al.
Pubblicazione: (2024)
Ensemble Distillation for Unsupervised Constituency Parsing
di: Shayegh, Behzad, et al.
Pubblicazione: (2023)
di: Shayegh, Behzad, et al.
Pubblicazione: (2023)
Error Diversity Matters: An Error-Resistant Ensemble Method for Unsupervised Dependency Parsing
di: Shayegh, Behzad, et al.
Pubblicazione: (2024)
di: Shayegh, Behzad, et al.
Pubblicazione: (2024)
Feeding Two Birds or Favoring One? Adequacy-Fluency Tradeoffs in Evaluation and Meta-Evaluation of Machine Translation
di: Shayegh, Behzad, et al.
Pubblicazione: (2025)
di: Shayegh, Behzad, et al.
Pubblicazione: (2025)
EBBS: An Ensemble with Bi-Level Beam Search for Zero-Shot Machine Translation
di: Wen, Yuqiao, et al.
Pubblicazione: (2024)
di: Wen, Yuqiao, et al.
Pubblicazione: (2024)
GPT-HateCheck: Can LLMs Write Better Functional Tests for Hate Speech Detection?
di: Jin, Yiping, et al.
Pubblicazione: (2024)
di: Jin, Yiping, et al.
Pubblicazione: (2024)
Memory Efficient Neural Processes via Constant Memory Attention Block
di: Feng, Leo, et al.
Pubblicazione: (2023)
di: Feng, Leo, et al.
Pubblicazione: (2023)
A Dual-Path Architecture for Scaling Compute and Capacity in LLMs
di: Frey, Markus, et al.
Pubblicazione: (2026)
di: Frey, Markus, et al.
Pubblicazione: (2026)
SLIM-LLMs: Modeling of Style-Sensory Language RelationshipsThrough Low-Dimensional Representations
di: Khalid, Osama, et al.
Pubblicazione: (2025)
di: Khalid, Osama, et al.
Pubblicazione: (2025)
When to Retrieve: Teaching LLMs to Utilize Information Retrieval Effectively
di: Labruna, Tiziano, et al.
Pubblicazione: (2024)
di: Labruna, Tiziano, et al.
Pubblicazione: (2024)
Fingerprinting LLMs via Prompt Injection
di: Hu, Yuepeng, et al.
Pubblicazione: (2025)
di: Hu, Yuepeng, et al.
Pubblicazione: (2025)
Signal in the Noise: Decoding the Reality of Airline Service Quality with Large Language Models
di: Dawoud, Ahmed, et al.
Pubblicazione: (2026)
di: Dawoud, Ahmed, et al.
Pubblicazione: (2026)
How Well Can LLMs Echo Us? Evaluating AI Chatbots' Role-Play Ability with ECHO
di: Ng, Man Tik, et al.
Pubblicazione: (2024)
di: Ng, Man Tik, et al.
Pubblicazione: (2024)
From Reviews to Requirements: Can LLMs Generate Human-Like User Stories?
di: Sakib, Shadman, et al.
Pubblicazione: (2026)
di: Sakib, Shadman, et al.
Pubblicazione: (2026)
Can Multimodal LLMs do Visual Temporal Understanding and Reasoning? The answer is No!
di: Imam, Mohamed Fazli, et al.
Pubblicazione: (2025)
di: Imam, Mohamed Fazli, et al.
Pubblicazione: (2025)
Generative Dense Retrieval: Memory Can Be a Burden
di: Yuan, Peiwen, et al.
Pubblicazione: (2024)
di: Yuan, Peiwen, et al.
Pubblicazione: (2024)
LLMs Can Compensate for Deficiencies in Visual Representations
di: Takishita, Sho, et al.
Pubblicazione: (2025)
di: Takishita, Sho, et al.
Pubblicazione: (2025)
Lexicon-Enriched Graph Modeling for Arabic Document Readability Prediction
di: Elchafei, Passant, et al.
Pubblicazione: (2025)
di: Elchafei, Passant, et al.
Pubblicazione: (2025)
LLMs Can Generate a Better Answer by Aggregating Their Own Responses
di: Li, Zichong, et al.
Pubblicazione: (2025)
di: Li, Zichong, et al.
Pubblicazione: (2025)
Can We Further Elicit Reasoning in LLMs? Critic-Guided Planning with Retrieval-Augmentation for Solving Challenging Tasks
di: Li, Xingxuan, et al.
Pubblicazione: (2024)
di: Li, Xingxuan, et al.
Pubblicazione: (2024)
Knowledge Graph Analysis of Legal Understanding and Violations in LLMs
di: Jha, Abha, et al.
Pubblicazione: (2025)
di: Jha, Abha, et al.
Pubblicazione: (2025)
Data Checklist: On Unit-Testing Datasets with Usable Information
di: Zhang, Heidi C., et al.
Pubblicazione: (2024)
di: Zhang, Heidi C., et al.
Pubblicazione: (2024)
Let LLMs Take on the Latest Challenges! A Chinese Dynamic Question Answering Benchmark
di: Xu, Zhikun, et al.
Pubblicazione: (2024)
di: Xu, Zhikun, et al.
Pubblicazione: (2024)
Automating Legal Interpretation with LLMs: Retrieval, Generation, and Evaluation
di: Luo, Kangcheng, et al.
Pubblicazione: (2025)
di: Luo, Kangcheng, et al.
Pubblicazione: (2025)
When Do LLMs Need Retrieval Augmentation? Mitigating LLMs' Overconfidence Helps Retrieval Augmentation
di: Ni, Shiyu, et al.
Pubblicazione: (2024)
di: Ni, Shiyu, et al.
Pubblicazione: (2024)
PhageBench: Can LLMs Understand Raw Bacteriophage Genomes?
di: Hou, Yusen, et al.
Pubblicazione: (2026)
di: Hou, Yusen, et al.
Pubblicazione: (2026)
Can LLMs Rank the Harmfulness of Smaller LLMs? We are Not There Yet
di: Atil, Berk, et al.
Pubblicazione: (2025)
di: Atil, Berk, et al.
Pubblicazione: (2025)
The AI Co-Ethnographer: How Far Can Automation Take Qualitative Research?
di: Retkowski, Fabian, et al.
Pubblicazione: (2025)
di: Retkowski, Fabian, et al.
Pubblicazione: (2025)
Do Methods to Jailbreak and Defend LLMs Generalize Across Languages?
di: Atil, Berk, et al.
Pubblicazione: (2025)
di: Atil, Berk, et al.
Pubblicazione: (2025)
Can LLMs Detect Their Own Hallucinations?
di: Kadotani, Sora, et al.
Pubblicazione: (2025)
di: Kadotani, Sora, et al.
Pubblicazione: (2025)
Can Editing LLMs Inject Harm?
di: Chen, Canyu, et al.
Pubblicazione: (2024)
di: Chen, Canyu, et al.
Pubblicazione: (2024)
Can LLMs Reason in the Wild with Programs?
di: Yang, Yuan, et al.
Pubblicazione: (2024)
di: Yang, Yuan, et al.
Pubblicazione: (2024)
Funny or Persuasive, but Not Both: Evaluating Fine-Grained Multi-Concept Control in LLMs
di: Labroo, Arya, et al.
Pubblicazione: (2026)
di: Labroo, Arya, et al.
Pubblicazione: (2026)
UniToMBench: Integrating Perspective-Taking to Improve Theory of Mind in LLMs
di: Thiyagarajan, Prameshwar, et al.
Pubblicazione: (2025)
di: Thiyagarajan, Prameshwar, et al.
Pubblicazione: (2025)
It Helps to Take a Second Opinion: Teaching Smaller LLMs to Deliberate Mutually via Selective Rationale Optimisation
di: Patnaik, Sohan, et al.
Pubblicazione: (2025)
di: Patnaik, Sohan, et al.
Pubblicazione: (2025)
Can LLMs Express Their Uncertainty? An Empirical Evaluation of Confidence Elicitation in LLMs
di: Xiong, Miao, et al.
Pubblicazione: (2023)
di: Xiong, Miao, et al.
Pubblicazione: (2023)
Extracting Unlearned Information from LLMs with Activation Steering
di: Seyitoğlu, Atakan, et al.
Pubblicazione: (2024)
di: Seyitoğlu, Atakan, et al.
Pubblicazione: (2024)
LLMs Know What They Need: Leveraging a Missing Information Guided Framework to Empower Retrieval-Augmented Generation
di: Wang, Keheng, et al.
Pubblicazione: (2024)
di: Wang, Keheng, et al.
Pubblicazione: (2024)
Can Language Models Take A Hint? Prompting for Controllable Contextualized Commonsense Inference
di: Colon-Hernandez, Pedro, et al.
Pubblicazione: (2024)
di: Colon-Hernandez, Pedro, et al.
Pubblicazione: (2024)
With a Grain of SALT: Are LLMs Fair Across Social Dimensions?
di: Arif, Samee, et al.
Pubblicazione: (2024)
di: Arif, Samee, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Tree-Averaging Algorithms for Ensemble-Based Unsupervised Discontinuous Constituency Parsing
di: Shayegh, Behzad, et al.
Pubblicazione: (2024) -
Ensemble Distillation for Unsupervised Constituency Parsing
di: Shayegh, Behzad, et al.
Pubblicazione: (2023) -
Error Diversity Matters: An Error-Resistant Ensemble Method for Unsupervised Dependency Parsing
di: Shayegh, Behzad, et al.
Pubblicazione: (2024) -
Feeding Two Birds or Favoring One? Adequacy-Fluency Tradeoffs in Evaluation and Meta-Evaluation of Machine Translation
di: Shayegh, Behzad, et al.
Pubblicazione: (2025) -
EBBS: An Ensemble with Bi-Level Beam Search for Zero-Shot Machine Translation
di: Wen, Yuqiao, et al.
Pubblicazione: (2024)