Towards A Unified View of Answer Calibration for Multi-Step Reasoning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Deng, Shumin, Zhang, Ningyu, Oo, Nay, Hooi, Bryan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Automating Steering for Safe Multimodal Large Language Models
von: Wu, Lyucheng, et al.
Veröffentlicht: (2025)
von: Wu, Lyucheng, et al.
Veröffentlicht: (2025)
Information Extraction in Low-Resource Scenarios: Survey and Perspective
von: Deng, Shumin, et al.
Veröffentlicht: (2022)
von: Deng, Shumin, et al.
Veröffentlicht: (2022)
ChineseHarm-Bench: A Chinese Harmful Content Detection Benchmark
von: Liu, Kangwei, et al.
Veröffentlicht: (2025)
von: Liu, Kangwei, et al.
Veröffentlicht: (2025)
Why Steering Works: Toward a Unified View of Language Model Parameter Dynamics
von: Xu, Ziwen, et al.
Veröffentlicht: (2026)
von: Xu, Ziwen, et al.
Veröffentlicht: (2026)
LLMs for Knowledge Graph Construction and Reasoning: Recent Capabilities and Future Opportunities
von: Zhu, Yuqi, et al.
Veröffentlicht: (2023)
von: Zhu, Yuqi, et al.
Veröffentlicht: (2023)
From Data to Behavior: Predicting Unintended Model Behaviors Before Training
von: Wang, Mengru, et al.
Veröffentlicht: (2026)
von: Wang, Mengru, et al.
Veröffentlicht: (2026)
LightThinker: Thinking Step-by-Step Compression
von: Zhang, Jintian, et al.
Veröffentlicht: (2025)
von: Zhang, Jintian, et al.
Veröffentlicht: (2025)
Editing Conceptual Knowledge for Large Language Models
von: Wang, Xiaohan, et al.
Veröffentlicht: (2024)
von: Wang, Xiaohan, et al.
Veröffentlicht: (2024)
InnoEval: On Research Idea Evaluation as a Knowledge-Grounded, Multi-Perspective Reasoning Problem
von: Qiao, Shuofei, et al.
Veröffentlicht: (2026)
von: Qiao, Shuofei, et al.
Veröffentlicht: (2026)
Beyond Prompt Engineering: Robust Behavior Control in LLMs via Steering Target Atoms
von: Wang, Mengru, et al.
Veröffentlicht: (2025)
von: Wang, Mengru, et al.
Veröffentlicht: (2025)
Knowledge Circuits in Pretrained Transformers
von: Yao, Yunzhi, et al.
Veröffentlicht: (2024)
von: Yao, Yunzhi, et al.
Veröffentlicht: (2024)
QAGCN: Answering Multi-Relation Questions via Single-Step Implicit Reasoning over Knowledge Graphs
von: Wang, Ruijie, et al.
Veröffentlicht: (2022)
von: Wang, Ruijie, et al.
Veröffentlicht: (2022)
CaKE: Circuit-aware Editing Enables Generalizable Knowledge Learners
von: Yao, Yunzhi, et al.
Veröffentlicht: (2025)
von: Yao, Yunzhi, et al.
Veröffentlicht: (2025)
OneGen: Efficient One-Pass Unified Generation and Retrieval for LLMs
von: Zhang, Jintian, et al.
Veröffentlicht: (2024)
von: Zhang, Jintian, et al.
Veröffentlicht: (2024)
Verif.ai: Towards an Open-Source Scientific Generative Question-Answering System with Referenced and Verifiable Answers
von: Košprdić, Miloš, et al.
Veröffentlicht: (2024)
von: Košprdić, Miloš, et al.
Veröffentlicht: (2024)
OneEdit: A Neural-Symbolic Collaboratively Knowledge Editing System
von: Zhang, Ningyu, et al.
Veröffentlicht: (2024)
von: Zhang, Ningyu, et al.
Veröffentlicht: (2024)
Unified Hallucination Detection for Multimodal Large Language Models
von: Chen, Xiang, et al.
Veröffentlicht: (2024)
von: Chen, Xiang, et al.
Veröffentlicht: (2024)
LightThinker++: From Reasoning Compression to Memory Management
von: Zhu, Yuqi, et al.
Veröffentlicht: (2026)
von: Zhu, Yuqi, et al.
Veröffentlicht: (2026)
RankRAG: Unifying Context Ranking with Retrieval-Augmented Generation in LLMs
von: Yu, Yue, et al.
Veröffentlicht: (2024)
von: Yu, Yue, et al.
Veröffentlicht: (2024)
Exploring Collaboration Mechanisms for LLM Agents: A Social Psychology View
von: Zhang, Jintian, et al.
Veröffentlicht: (2023)
von: Zhang, Jintian, et al.
Veröffentlicht: (2023)
CKnowEdit: A New Chinese Knowledge Editing Dataset for Linguistics, Facts, and Logic Error Correction in LLMs
von: Fang, Jizhan, et al.
Veröffentlicht: (2024)
von: Fang, Jizhan, et al.
Veröffentlicht: (2024)
Multilingual Non-Factoid Question Answering with Answer Paragraph Selection
von: Mishra, Ritwik, et al.
Veröffentlicht: (2024)
von: Mishra, Ritwik, et al.
Veröffentlicht: (2024)
ReasonIR: Training Retrievers for Reasoning Tasks
von: Shao, Rulin, et al.
Veröffentlicht: (2025)
von: Shao, Rulin, et al.
Veröffentlicht: (2025)
InstructIE: A Bilingual Instruction-based Information Extraction Dataset
von: Gui, Honghao, et al.
Veröffentlicht: (2023)
von: Gui, Honghao, et al.
Veröffentlicht: (2023)
RAMQA: A Unified Framework for Retrieval-Augmented Multi-Modal Question Answering
von: Bai, Yang, et al.
Veröffentlicht: (2025)
von: Bai, Yang, et al.
Veröffentlicht: (2025)
SciAtlas: A Large-Scale Knowledge Graph for Automated Scientific Research
von: Qiao, Shuofei, et al.
Veröffentlicht: (2026)
von: Qiao, Shuofei, et al.
Veröffentlicht: (2026)
ReCode: Updating Code API Knowledge with Reinforcement Learning
von: Wu, Haoze, et al.
Veröffentlicht: (2025)
von: Wu, Haoze, et al.
Veröffentlicht: (2025)
Scaling Generalist Data-Analytic Agents
von: Qiao, Shuofei, et al.
Veröffentlicht: (2025)
von: Qiao, Shuofei, et al.
Veröffentlicht: (2025)
Retrieval-augmented Prompt Learning for Pre-trained Foundation Models
von: Chen, Xiang, et al.
Veröffentlicht: (2025)
von: Chen, Xiang, et al.
Veröffentlicht: (2025)
ReLearn: Unlearning via Learning for Large Language Models
von: Xu, Haoming, et al.
Veröffentlicht: (2025)
von: Xu, Haoming, et al.
Veröffentlicht: (2025)
SciNets: Graph-Constrained Multi-Hop Reasoning for Scientific Literature Synthesis
von: Dubey, Sauhard
Veröffentlicht: (2025)
von: Dubey, Sauhard
Veröffentlicht: (2025)
Graph-based Confidence Calibration for Large Language Models
von: Li, Yukun, et al.
Veröffentlicht: (2024)
von: Li, Yukun, et al.
Veröffentlicht: (2024)
IEPile: Unearthing Large-Scale Schema-Based Information Extraction Corpus
von: Gui, Honghao, et al.
Veröffentlicht: (2024)
von: Gui, Honghao, et al.
Veröffentlicht: (2024)
CuriousLLM: Elevating Multi-Document Question Answering with LLM-Enhanced Knowledge Graph Reasoning
von: Yang, Zukang, et al.
Veröffentlicht: (2024)
von: Yang, Zukang, et al.
Veröffentlicht: (2024)
KnowPhish: Large Language Models Meet Multimodal Knowledge Graphs for Enhancing Reference-Based Phishing Detection
von: Li, Yuexin, et al.
Veröffentlicht: (2024)
von: Li, Yuexin, et al.
Veröffentlicht: (2024)
ReasonRank: Empowering Passage Ranking with Strong Reasoning Ability
von: Liu, Wenhan, et al.
Veröffentlicht: (2025)
von: Liu, Wenhan, et al.
Veröffentlicht: (2025)
DR-Venus: Towards Frontier Edge-Scale Deep Research Agents with Only 10K Open Data
von: Venus Team, et al.
Veröffentlicht: (2026)
von: Venus Team, et al.
Veröffentlicht: (2026)
Leveraging LLM Reasoning Enhances Personalized Recommender Systems
von: Tsai, Alicia Y., et al.
Veröffentlicht: (2024)
von: Tsai, Alicia Y., et al.
Veröffentlicht: (2024)
Eliciting In-context Retrieval and Reasoning for Long-context Large Language Models
von: Qiu, Yifu, et al.
Veröffentlicht: (2025)
von: Qiu, Yifu, et al.
Veröffentlicht: (2025)
CXMArena: Unified Dataset to benchmark performance in realistic CXM Scenarios
von: Garg, Raghav, et al.
Veröffentlicht: (2025)
von: Garg, Raghav, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Automating Steering for Safe Multimodal Large Language Models
von: Wu, Lyucheng, et al.
Veröffentlicht: (2025) -
Information Extraction in Low-Resource Scenarios: Survey and Perspective
von: Deng, Shumin, et al.
Veröffentlicht: (2022) -
ChineseHarm-Bench: A Chinese Harmful Content Detection Benchmark
von: Liu, Kangwei, et al.
Veröffentlicht: (2025) -
Why Steering Works: Toward a Unified View of Language Model Parameter Dynamics
von: Xu, Ziwen, et al.
Veröffentlicht: (2026) -
LLMs for Knowledge Graph Construction and Reasoning: Recent Capabilities and Future Opportunities
von: Zhu, Yuqi, et al.
Veröffentlicht: (2023)