Salvato in:
| Autori principali: | Zhang, Zhengxuan, Liang, Zhuowen, Wu, Yin, Lin, Teng, Luo, Yuyu, Tang, Nan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2504.10036 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Long-Document QA with Chain-of-Structured-Thought and Fine-Tuned SLMs
di: Liang, Zhuowen, et al.
Pubblicazione: (2026)
di: Liang, Zhuowen, et al.
Pubblicazione: (2026)
EXCLAIM: An Explainable Cross-Modal Agentic System for Misinformation Detection with Hierarchical Retrieval
di: Wu, Yin, et al.
Pubblicazione: (2025)
di: Wu, Yin, et al.
Pubblicazione: (2025)
Fine-Grained Knowledge Structuring and Retrieval for Visual Question Answering
di: Zhang, Zhengxuan, et al.
Pubblicazione: (2025)
di: Zhang, Zhengxuan, et al.
Pubblicazione: (2025)
SRAG: Structured Retrieval-Augmented Generation for Multi-Entity Question Answering over Wikipedia Graph
di: Lin, Teng, et al.
Pubblicazione: (2025)
di: Lin, Teng, et al.
Pubblicazione: (2025)
MEBench: Benchmarking Large Language Models for Cross-Document Multi-Entity Question Answering
di: Lin, Teng, et al.
Pubblicazione: (2025)
di: Lin, Teng, et al.
Pubblicazione: (2025)
DocSage: An Information Structuring Agent for Multi-Doc Multi-Entity Question Answering
di: Lin, Teng, et al.
Pubblicazione: (2026)
di: Lin, Teng, et al.
Pubblicazione: (2026)
In-Context Sharpness as Alerts: An Inner Representation Perspective for Hallucination Mitigation
di: Chen, Shiqi, et al.
Pubblicazione: (2024)
di: Chen, Shiqi, et al.
Pubblicazione: (2024)
SketchFill: Sketch-Guided Code Generation for Imputing Derived Missing Values
di: Zhang, Yunfan, et al.
Pubblicazione: (2024)
di: Zhang, Yunfan, et al.
Pubblicazione: (2024)
TransXSSM: A Hybrid Transformer State Space Model with Unified Rotary Position Embedding
di: Wu, Bingheng, et al.
Pubblicazione: (2025)
di: Wu, Bingheng, et al.
Pubblicazione: (2025)
LLMs Encode Harmfulness and Refusal Separately
di: Zhao, Jiachen, et al.
Pubblicazione: (2025)
di: Zhao, Jiachen, et al.
Pubblicazione: (2025)
TableTale: Reviving the Narrative Interplay Between Data Tables and Text in Scientific Papers
di: Wang, Liangwei, et al.
Pubblicazione: (2026)
di: Wang, Liangwei, et al.
Pubblicazione: (2026)
Are Large Language Models Good Statisticians?
di: Zhu, Yizhang, et al.
Pubblicazione: (2024)
di: Zhu, Yizhang, et al.
Pubblicazione: (2024)
ChartInsights: Evaluating Multimodal Large Language Models for Low-Level Chart Question Answering
di: Wu, Yifan, et al.
Pubblicazione: (2024)
di: Wu, Yifan, et al.
Pubblicazione: (2024)
EllieSQL: Cost-Efficient Text-to-SQL with Complexity-Aware Routing
di: Zhu, Yizhang, et al.
Pubblicazione: (2025)
di: Zhu, Yizhang, et al.
Pubblicazione: (2025)
LightKGG: Simple and Efficient Knowledge Graph Generation from Textual Data
di: Lin, Teng
Pubblicazione: (2025)
di: Lin, Teng
Pubblicazione: (2025)
AnnoRetrieve: Efficient Structured Retrieval for Unstructured Document Analysis
di: Lin, Teng, et al.
Pubblicazione: (2026)
di: Lin, Teng, et al.
Pubblicazione: (2026)
Oolong: Investigating What Makes Transfer Learning Hard with Controlled Studies
di: Wu, Zhengxuan, et al.
Pubblicazione: (2022)
di: Wu, Zhengxuan, et al.
Pubblicazione: (2022)
RoleBreak: Character Hallucination as a Jailbreak Attack in Role-Playing Systems
di: Tang, Yihong, et al.
Pubblicazione: (2024)
di: Tang, Yihong, et al.
Pubblicazione: (2024)
D$^2$HScore: Reasoning-Aware Hallucination Detection via Semantic Breadth and Depth Analysis in LLMs
di: Ding, Yue, et al.
Pubblicazione: (2025)
di: Ding, Yue, et al.
Pubblicazione: (2025)
Can LLMs Clean Up Your Mess? A Survey of Application-Ready Data Preparation with LLMs
di: Zhou, Wei, et al.
Pubblicazione: (2026)
di: Zhou, Wei, et al.
Pubblicazione: (2026)
Challenging Multilingual LLMs: A New Taxonomy and Benchmark for Unraveling Hallucination in Translation
di: Wu, Xinwei, et al.
Pubblicazione: (2025)
di: Wu, Xinwei, et al.
Pubblicazione: (2025)
ReCOGS: How Incidental Details of a Logical Form Overshadow an Evaluation of Semantic Interpretation
di: Wu, Zhengxuan, et al.
Pubblicazione: (2023)
di: Wu, Zhengxuan, et al.
Pubblicazione: (2023)
Shared Imagination: LLMs Hallucinate Alike
di: Zhou, Yilun, et al.
Pubblicazione: (2024)
di: Zhou, Yilun, et al.
Pubblicazione: (2024)
Wisdom is Knowing What not to Say: Hallucination-Free LLMs Unlearning via Attention Shifting
di: Tan, Chenchen, et al.
Pubblicazione: (2025)
di: Tan, Chenchen, et al.
Pubblicazione: (2025)
Weak-eval-Strong: Evaluating and Eliciting Lateral Thinking of LLMs with Situation Puzzles
di: Chen, Qi, et al.
Pubblicazione: (2024)
di: Chen, Qi, et al.
Pubblicazione: (2024)
HART: Data-Driven Hallucination Attribution and Evidence-Based Tracing for Large Language Models
di: Liang, Shize, et al.
Pubblicazione: (2026)
di: Liang, Shize, et al.
Pubblicazione: (2026)
SHARP: Unlocking Interactive Hallucination via Stance Transfer in Role-Playing LLMs
di: Kong, Chuyi, et al.
Pubblicazione: (2024)
di: Kong, Chuyi, et al.
Pubblicazione: (2024)
Can LLMs Generate and Solve Linguistic Olympiad Puzzles?
di: Majmudar, Neh, et al.
Pubblicazione: (2025)
di: Majmudar, Neh, et al.
Pubblicazione: (2025)
Learning to Trust Your Feelings: Leveraging Self-awareness in LLMs for Hallucination Mitigation
di: Liang, Yuxin, et al.
Pubblicazione: (2024)
di: Liang, Yuxin, et al.
Pubblicazione: (2024)
Understanding New-Knowledge-Induced Factual Hallucinations in LLMs: Analysis and Interpretation
di: Dang, Renfei, et al.
Pubblicazione: (2025)
di: Dang, Renfei, et al.
Pubblicazione: (2025)
Where Fake Citations Are Made: Tracing Field-Level Hallucination to Specific Neurons in LLMs
di: Chen, Yuefei, et al.
Pubblicazione: (2026)
di: Chen, Yuefei, et al.
Pubblicazione: (2026)
nvBench 2.0: Resolving Ambiguity in Text-to-Visualization through Stepwise Reasoning
di: Luo, Tianqi, et al.
Pubblicazione: (2025)
di: Luo, Tianqi, et al.
Pubblicazione: (2025)
How Chinese are Chinese Language Models? The Puzzling Lack of Language Policy in China's LLMs
di: Wen-Yi, Andrea W, et al.
Pubblicazione: (2024)
di: Wen-Yi, Andrea W, et al.
Pubblicazione: (2024)
Atom of Thoughts for Markov LLM Test-Time Scaling
di: Teng, Fengwei, et al.
Pubblicazione: (2025)
di: Teng, Fengwei, et al.
Pubblicazione: (2025)
A Strategic Coordination Framework of Small LLMs Matches Large LLMs in Data Synthesis
di: Gao, Xin, et al.
Pubblicazione: (2025)
di: Gao, Xin, et al.
Pubblicazione: (2025)
Geometric Uncertainty for Detecting and Correcting Hallucinations in LLMs
di: Phillips, Edward, et al.
Pubblicazione: (2025)
di: Phillips, Edward, et al.
Pubblicazione: (2025)
NoisyAG-News: A Benchmark for Addressing Instance-Dependent Noise in Text Classification
di: Huang, Hongfei, et al.
Pubblicazione: (2024)
di: Huang, Hongfei, et al.
Pubblicazione: (2024)
PlotCraft: Pushing the Limits of LLMs for Complex and Interactive Data Visualization
di: Zhang, Jiajun, et al.
Pubblicazione: (2025)
di: Zhang, Jiajun, et al.
Pubblicazione: (2025)
Counterfactual Debating with Preset Stances for Hallucination Elimination of LLMs
di: Fang, Yi, et al.
Pubblicazione: (2024)
di: Fang, Yi, et al.
Pubblicazione: (2024)
DCR: Quantifying Data Contamination in LLMs Evaluation
di: Xu, Cheng, et al.
Pubblicazione: (2025)
di: Xu, Cheng, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Long-Document QA with Chain-of-Structured-Thought and Fine-Tuned SLMs
di: Liang, Zhuowen, et al.
Pubblicazione: (2026) -
EXCLAIM: An Explainable Cross-Modal Agentic System for Misinformation Detection with Hierarchical Retrieval
di: Wu, Yin, et al.
Pubblicazione: (2025) -
Fine-Grained Knowledge Structuring and Retrieval for Visual Question Answering
di: Zhang, Zhengxuan, et al.
Pubblicazione: (2025) -
SRAG: Structured Retrieval-Augmented Generation for Multi-Entity Question Answering over Wikipedia Graph
di: Lin, Teng, et al.
Pubblicazione: (2025) -
MEBench: Benchmarking Large Language Models for Cross-Document Multi-Entity Question Answering
di: Lin, Teng, et al.
Pubblicazione: (2025)