KnowCoder-A1: Incentivizing Agentic Reasoning Capability with Outcome Supervision for KBQA
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Zhuo, Wang, Fei, Li, Zixuan, Zhang, Zhao, Ding, Weiwei, Yang, Chuanguang, Xu, Yongjun, Jin, Xiaolong, Guo, Jiafeng |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
KnowCoder-X: Boosting Multilingual Information Extraction via Code
by: Zuo, Yuxin, et al.
Published: (2024)
by: Zuo, Yuxin, et al.
Published: (2024)
KnowCoder-V2: Deep Knowledge Analysis
by: Li, Zixuan, et al.
Published: (2025)
by: Li, Zixuan, et al.
Published: (2025)
KnowCoder: Coding Structured Knowledge into LLMs for Universal Information Extraction
by: Li, Zixuan, et al.
Published: (2024)
by: Li, Zixuan, et al.
Published: (2024)
Self-Improvement Programming for Temporal Knowledge Graph Question Answering
by: Chen, Zhuo, et al.
Published: (2024)
by: Chen, Zhuo, et al.
Published: (2024)
Temporal Knowledge Graph Question Answering: A Survey
by: Su, Miao, et al.
Published: (2024)
by: Su, Miao, et al.
Published: (2024)
Incentivizing Strong Reasoning from Weak Supervision
by: Yuan, Yige, et al.
Published: (2025)
by: Yuan, Yige, et al.
Published: (2025)
StruProKGR: A Structural and Probabilistic Framework for Sparse Knowledge Graph Reasoning
by: Guo, Yucan, et al.
Published: (2025)
by: Guo, Yucan, et al.
Published: (2025)
KBQA-o1: Agentic Knowledge Base Question Answering with Monte Carlo Tree Search
by: Luo, Haoran, et al.
Published: (2025)
by: Luo, Haoran, et al.
Published: (2025)
LogicReward: Incentivizing LLM Reasoning via Step-Wise Logical Supervision
by: Xu, Jundong, et al.
Published: (2025)
by: Xu, Jundong, et al.
Published: (2025)
G2S: A General-to-Specific Learning Framework for Temporal Knowledge Graph Forecasting with Large Language Models
by: Bai, Long, et al.
Published: (2025)
by: Bai, Long, et al.
Published: (2025)
Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models
by: Huang, Wenxuan, et al.
Published: (2025)
by: Huang, Wenxuan, et al.
Published: (2025)
Class-Incremental Few-Shot Event Detection
by: Zhao, Kailin, et al.
Published: (2024)
by: Zhao, Kailin, et al.
Published: (2024)
Internalizing Outcome Supervision into Process Supervision: A New Paradigm for Reinforcement Learning for Reasoning
by: Ding, Fei, et al.
Published: (2026)
by: Ding, Fei, et al.
Published: (2026)
An In-Context Schema Understanding Method for Knowledge Base Question Answering
by: Liu, Yantao, et al.
Published: (2023)
by: Liu, Yantao, et al.
Published: (2023)
Towards Robust Universal Information Extraction: Benchmark, Evaluation, and Solution
by: Zhu, Jizhao, et al.
Published: (2025)
by: Zhu, Jizhao, et al.
Published: (2025)
Towards Event Extraction with Massive Types: LLM-based Collaborative Annotation and Partitioning Extraction
by: Liu, Wenxuan, et al.
Published: (2025)
by: Liu, Wenxuan, et al.
Published: (2025)
Context Reasoner: Incentivizing Reasoning Capability for Contextualized Privacy and Safety Compliance via Reinforcement Learning
by: Hu, Wenbin, et al.
Published: (2025)
by: Hu, Wenbin, et al.
Published: (2025)
ZeroSearch: Incentivize the Search Capability of LLMs without Searching
by: Sun, Hao, et al.
Published: (2025)
by: Sun, Hao, et al.
Published: (2025)
Incentivizing Agentic Reasoning in LLM Judges via Tool-Integrated Reinforcement Learning
by: Xu, Ran, et al.
Published: (2025)
by: Xu, Ran, et al.
Published: (2025)
AlchemistCoder: Harmonizing and Eliciting Code Capability by Hindsight Tuning on Multi-source Data
by: Song, Zifan, et al.
Published: (2024)
by: Song, Zifan, et al.
Published: (2024)
MCTS-KBQA: Monte Carlo Tree Search for Knowledge Base Question Answering
by: Xiong, Guanming, et al.
Published: (2025)
by: Xiong, Guanming, et al.
Published: (2025)
Right Is Not Enough: The Pitfalls of Outcome Supervision in Training LLMs for Math Reasoning
by: Guo, Jiaxing, et al.
Published: (2025)
by: Guo, Jiaxing, et al.
Published: (2025)
R1-T1: Fully Incentivizing Translation Capability in LLMs via Reasoning Learning
by: He, Minggui, et al.
Published: (2025)
by: He, Minggui, et al.
Published: (2025)
DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
by: DeepSeek-AI, et al.
Published: (2025)
by: DeepSeek-AI, et al.
Published: (2025)
SemCoder: Training Code Language Models with Comprehensive Semantics Reasoning
by: Ding, Yangruibo, et al.
Published: (2024)
by: Ding, Yangruibo, et al.
Published: (2024)
Nested Event Extraction upon Pivot Element Recogniton
by: Ren, Weicheng, et al.
Published: (2023)
by: Ren, Weicheng, et al.
Published: (2023)
From Web Search towards Agentic Deep Research: Incentivizing Search with Reasoning Agents
by: Zhang, Weizhi, et al.
Published: (2025)
by: Zhang, Weizhi, et al.
Published: (2025)
LB-KBQA: Large-language-model and BERT based Knowledge-Based Question and Answering System
by: Zhao, Yan, et al.
Published: (2024)
by: Zhao, Yan, et al.
Published: (2024)
Iterative Repair with Weak Verifiers for Few-shot Transfer in KBQA with Unanswerability
by: Sawhney, Riya, et al.
Published: (2024)
by: Sawhney, Riya, et al.
Published: (2024)
RadFabric: Agentic AI System with Reasoning Capability for Radiology
by: Chen, Wenting, et al.
Published: (2025)
by: Chen, Wenting, et al.
Published: (2025)
Agentic Reasoning: A Streamlined Framework for Enhancing LLM Reasoning with Agentic Tools
by: Wu, Junde, et al.
Published: (2025)
by: Wu, Junde, et al.
Published: (2025)
WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning
by: Zhuang, Yuchen, et al.
Published: (2025)
by: Zhuang, Yuchen, et al.
Published: (2025)
R1-ShareVL: Incentivizing Reasoning Capability of Multimodal Large Language Models via Share-GRPO
by: Yao, Huanjin, et al.
Published: (2025)
by: Yao, Huanjin, et al.
Published: (2025)
KBQA-R1: Reinforcing Large Language Models for Knowledge Base Question Answering
by: Sun, Xin, et al.
Published: (2025)
by: Sun, Xin, et al.
Published: (2025)
Interactive-KBQA: Multi-Turn Interactions for Knowledge Base Question Answering with Large Language Models
by: Xiong, Guanming, et al.
Published: (2024)
by: Xiong, Guanming, et al.
Published: (2024)
CLAUSE: Agentic Neuro-Symbolic Knowledge Graph Reasoning via Dynamic Learnable Context Engineering
by: Zhao, Yang, et al.
Published: (2025)
by: Zhao, Yang, et al.
Published: (2025)
MIND: From Passive Mimicry to Active Reasoning through Capability-Aware Multi-Perspective CoT Distillation
by: Cui, Jin, et al.
Published: (2026)
by: Cui, Jin, et al.
Published: (2026)
Selective Temporal Knowledge Graph Reasoning
by: Hou, Zhongni, et al.
Published: (2024)
by: Hou, Zhongni, et al.
Published: (2024)
MORE: Multi-mOdal REtrieval Augmented Generative Commonsense Reasoning
by: Cui, Wanqing, et al.
Published: (2024)
by: Cui, Wanqing, et al.
Published: (2024)
Qwen2.5-Coder Technical Report
by: Hui, Binyuan, et al.
Published: (2024)
by: Hui, Binyuan, et al.
Published: (2024)
Similar Items
-
KnowCoder-X: Boosting Multilingual Information Extraction via Code
by: Zuo, Yuxin, et al.
Published: (2024) -
KnowCoder-V2: Deep Knowledge Analysis
by: Li, Zixuan, et al.
Published: (2025) -
KnowCoder: Coding Structured Knowledge into LLMs for Universal Information Extraction
by: Li, Zixuan, et al.
Published: (2024) -
Self-Improvement Programming for Temporal Knowledge Graph Question Answering
by: Chen, Zhuo, et al.
Published: (2024) -
Temporal Knowledge Graph Question Answering: A Survey
by: Su, Miao, et al.
Published: (2024)