How Grounded is Wikipedia? A Study on Structured Evidential Support and Retrieval
Fuente:
arXiv
Salvato in:
| Autori principali: | Walden, William, Ricci, Kathryn, Wanner, Miriam, Jiang, Zhengping, May, Chandler, Zhou, Rongkun, Van Durme, Benjamin |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
CLAIMCHECK: How Grounded are LLM Critiques of Scientific Papers?
di: Ou, Jiefu, et al.
Pubblicazione: (2025)
di: Ou, Jiefu, et al.
Pubblicazione: (2025)
A Closer Look at Claim Decomposition
di: Wanner, Miriam, et al.
Pubblicazione: (2024)
di: Wanner, Miriam, et al.
Pubblicazione: (2024)
Weird Generalization is Weirdly Brittle
di: Wanner, Miriam, et al.
Pubblicazione: (2026)
di: Wanner, Miriam, et al.
Pubblicazione: (2026)
Reasoning Models Will Sometimes Lie About Their Reasoning
di: Walden, William, et al.
Pubblicazione: (2026)
di: Walden, William, et al.
Pubblicazione: (2026)
DnDScore: Decontextualization and Decomposition for Factuality Verification in Long-Form Text Generation
di: Wanner, Miriam, et al.
Pubblicazione: (2024)
di: Wanner, Miriam, et al.
Pubblicazione: (2024)
MegaWika 2: A More Comprehensive Multilingual Collection of Articles and their Sources
di: Barham, Samuel, et al.
Pubblicazione: (2025)
di: Barham, Samuel, et al.
Pubblicazione: (2025)
Conformal Linguistic Calibration: Trading-off between Factuality and Specificity
di: Jiang, Zhengping, et al.
Pubblicazione: (2025)
di: Jiang, Zhengping, et al.
Pubblicazione: (2025)
Core: Robust Factual Precision with Informative Sub-Claim Identification
di: Jiang, Zhengping, et al.
Pubblicazione: (2024)
di: Jiang, Zhengping, et al.
Pubblicazione: (2024)
All Claims Are Equal, but Some Claims Are More Equal Than Others: Importance-Sensitive Factuality Evaluation of LLM Generations
di: Wanner, Miriam, et al.
Pubblicazione: (2025)
di: Wanner, Miriam, et al.
Pubblicazione: (2025)
Rank1: Test-Time Compute for Reranking in Information Retrieval
di: Weller, Orion, et al.
Pubblicazione: (2025)
di: Weller, Orion, et al.
Pubblicazione: (2025)
Always Tell Me The Odds: Fine-grained Conditional Probability Estimation
di: Wang, Liaoyaqi, et al.
Pubblicazione: (2025)
di: Wang, Liaoyaqi, et al.
Pubblicazione: (2025)
RE-AdaptIR: Improving Information Retrieval through Reverse Engineered Adaptation
di: Fleshman, William, et al.
Pubblicazione: (2024)
di: Fleshman, William, et al.
Pubblicazione: (2024)
Seeing Through the MiRAGE: Evaluating Multimodal Retrieval Augmented Generation
di: Martin, Alexander, et al.
Pubblicazione: (2025)
di: Martin, Alexander, et al.
Pubblicazione: (2025)
NELLIE: A Neuro-Symbolic Inference Engine for Grounded, Compositional, and Explainable Reasoning
di: Weir, Nathaniel, et al.
Pubblicazione: (2022)
di: Weir, Nathaniel, et al.
Pubblicazione: (2022)
RORA: Robust Free-Text Rationale Evaluation
di: Jiang, Zhengping, et al.
Pubblicazione: (2024)
di: Jiang, Zhengping, et al.
Pubblicazione: (2024)
NevIR: Negation in Neural Information Retrieval
di: Weller, Orion, et al.
Pubblicazione: (2023)
di: Weller, Orion, et al.
Pubblicazione: (2023)
Bonsai: Interpretable Tree-Adaptive Grounded Reasoning
di: Sanders, Kate, et al.
Pubblicazione: (2025)
di: Sanders, Kate, et al.
Pubblicazione: (2025)
SEQR: Secure and Efficient QR-based LoRA Routing
di: Fleshman, William, et al.
Pubblicazione: (2025)
di: Fleshman, William, et al.
Pubblicazione: (2025)
LoRA-Augmented Generation (LAG) for Knowledge-Intensive Language Tasks
di: Fleshman, William, et al.
Pubblicazione: (2025)
di: Fleshman, William, et al.
Pubblicazione: (2025)
RE-Adapt: Reverse Engineered Adaptation of Large Language Models
di: Fleshman, William, et al.
Pubblicazione: (2024)
di: Fleshman, William, et al.
Pubblicazione: (2024)
SpectR: Dynamically Composing LM Experts with Spectral Routing
di: Fleshman, William, et al.
Pubblicazione: (2025)
di: Fleshman, William, et al.
Pubblicazione: (2025)
Is That Your Final Answer? Test-Time Scaling Improves Selective Question Answering
di: Jurayj, William, et al.
Pubblicazione: (2025)
di: Jurayj, William, et al.
Pubblicazione: (2025)
Seq vs Seq: An Open Suite of Paired Encoders and Decoders
di: Weller, Orion, et al.
Pubblicazione: (2025)
di: Weller, Orion, et al.
Pubblicazione: (2025)
Multi-Field Adaptive Retrieval
di: Li, Millicent, et al.
Pubblicazione: (2024)
di: Li, Millicent, et al.
Pubblicazione: (2024)
Compressed Chain of Thought: Efficient Reasoning Through Dense Representations
di: Cheng, Jeffrey, et al.
Pubblicazione: (2024)
di: Cheng, Jeffrey, et al.
Pubblicazione: (2024)
Rank-K: Test-Time Reasoning for Listwise Reranking
di: Yang, Eugene, et al.
Pubblicazione: (2025)
di: Yang, Eugene, et al.
Pubblicazione: (2025)
WorldAPIs: The World Is Worth How Many APIs? A Thought Experiment
di: Ou, Jiefu, et al.
Pubblicazione: (2024)
di: Ou, Jiefu, et al.
Pubblicazione: (2024)
Language Models and Logic Programs for Trustworthy Tax Reasoning
di: Jurayj, William, et al.
Pubblicazione: (2025)
di: Jurayj, William, et al.
Pubblicazione: (2025)
Compactor: Calibrated Query-Agnostic KV Cache Compression with Approximate Leverage Scores
di: Chari, Vivek, et al.
Pubblicazione: (2025)
di: Chari, Vivek, et al.
Pubblicazione: (2025)
LLMs Provide Unstable Answers to Legal Questions
di: Blair-Stanek, Andrew, et al.
Pubblicazione: (2025)
di: Blair-Stanek, Andrew, et al.
Pubblicazione: (2025)
WikiVideo: Article Generation from Multiple Videos
di: Martin, Alexander, et al.
Pubblicazione: (2025)
di: Martin, Alexander, et al.
Pubblicazione: (2025)
Investigating Retrieval-Augmented Generation Systems on Unanswerable, Uncheatable, Realistic, Multi-hop Queries
di: Liu, Gabrielle Kaili-May, et al.
Pubblicazione: (2025)
di: Liu, Gabrielle Kaili-May, et al.
Pubblicazione: (2025)
SocialNLI: A Dialogue-Centric Social Inference Dataset
di: Deo, Akhil, et al.
Pubblicazione: (2025)
di: Deo, Akhil, et al.
Pubblicazione: (2025)
Does Local News Stay Local?: Online Content Shifts in Sinclair-Acquired Stations
di: Wanner, Miriam, et al.
Pubblicazione: (2025)
di: Wanner, Miriam, et al.
Pubblicazione: (2025)
Learning to Retrieve Iteratively for In-Context Learning
di: Chen, Yunmo, et al.
Pubblicazione: (2024)
di: Chen, Yunmo, et al.
Pubblicazione: (2024)
Zero and Few-shot Semantic Parsing with Ambiguous Inputs
di: Stengel-Eskin, Elias, et al.
Pubblicazione: (2023)
di: Stengel-Eskin, Elias, et al.
Pubblicazione: (2023)
LM Agents for Coordinating Multi-User Information Gathering
di: Jhamtani, Harsh, et al.
Pubblicazione: (2025)
di: Jhamtani, Harsh, et al.
Pubblicazione: (2025)
The NLP Task Effectiveness of Long-Range Transformers
di: Qin, Guanghui, et al.
Pubblicazione: (2022)
di: Qin, Guanghui, et al.
Pubblicazione: (2022)
KV-Distill: Nearly Lossless Learnable Context Compression for LLMs
di: Chari, Vivek, et al.
Pubblicazione: (2025)
di: Chari, Vivek, et al.
Pubblicazione: (2025)
CLERC: A Dataset for Legal Case Retrieval and Retrieval-Augmented Analysis Generation
di: Hou, Abe Bohan, et al.
Pubblicazione: (2024)
di: Hou, Abe Bohan, et al.
Pubblicazione: (2024)
Documenti analoghi
-
CLAIMCHECK: How Grounded are LLM Critiques of Scientific Papers?
di: Ou, Jiefu, et al.
Pubblicazione: (2025) -
A Closer Look at Claim Decomposition
di: Wanner, Miriam, et al.
Pubblicazione: (2024) -
Weird Generalization is Weirdly Brittle
di: Wanner, Miriam, et al.
Pubblicazione: (2026) -
Reasoning Models Will Sometimes Lie About Their Reasoning
di: Walden, William, et al.
Pubblicazione: (2026) -
DnDScore: Decontextualization and Decomposition for Factuality Verification in Long-Form Text Generation
di: Wanner, Miriam, et al.
Pubblicazione: (2024)