TaTToo: Tool-Grounded Thinking PRM for Test-Time Scaling in Tabular Reasoning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zou, Jiaru, Roy, Soumya, Verma, Vinay Kumar, Wang, Ziyi, Wipf, David, Lu, Pan, Negi, Sumit, Zou, James, He, Jingrui |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Rethinking Test Time Scaling for Flow-Matching Generative Models
von: Yu, Qingtao, et al.
Veröffentlicht: (2025)
von: Yu, Qingtao, et al.
Veröffentlicht: (2025)
ReasonFlux-PRM: Trajectory-Aware PRMs for Long Chain-of-Thought Reasoning in LLMs
von: Zou, Jiaru, et al.
Veröffentlicht: (2025)
von: Zou, Jiaru, et al.
Veröffentlicht: (2025)
AutoTool: Dynamic Tool Selection and Integration for Agentic Reasoning
von: Zou, Jiaru, et al.
Veröffentlicht: (2025)
von: Zou, Jiaru, et al.
Veröffentlicht: (2025)
ADROIT: A Self-Supervised Framework for Learning Robust Representations for Active Learning
von: Banerjee, Soumya, et al.
Veröffentlicht: (2025)
von: Banerjee, Soumya, et al.
Veröffentlicht: (2025)
DIVERSED: Relaxed Speculative Decoding via Dynamic Ensemble Verification
von: Wang, Ziyi, et al.
Veröffentlicht: (2026)
von: Wang, Ziyi, et al.
Veröffentlicht: (2026)
GenPRM: Scaling Test-Time Compute of Process Reward Models via Generative Reasoning
von: Zhao, Jian, et al.
Veröffentlicht: (2025)
von: Zhao, Jian, et al.
Veröffentlicht: (2025)
Think Right, Not More: Test-Time Scaling for Numerical Claim Verification
von: Chungkham, Primakov, et al.
Veröffentlicht: (2025)
von: Chungkham, Primakov, et al.
Veröffentlicht: (2025)
Thinking While Listening: Simple Test Time Scaling For Audio Classification
von: Verma, Prateek, et al.
Veröffentlicht: (2025)
von: Verma, Prateek, et al.
Veröffentlicht: (2025)
Does Thinking More always Help? Mirage of Test-Time Scaling in Reasoning Models
von: Ghosal, Soumya Suvra, et al.
Veröffentlicht: (2025)
von: Ghosal, Soumya Suvra, et al.
Veröffentlicht: (2025)
Optimal Aggregation of LLM and PRM Signals for Efficient Test-Time Scaling
von: Kuang, Peng, et al.
Veröffentlicht: (2025)
von: Kuang, Peng, et al.
Veröffentlicht: (2025)
TIM-PRM: Verifying multimodal reasoning with Tool-Integrated PRM
von: Kuang, Peng, et al.
Veröffentlicht: (2025)
von: Kuang, Peng, et al.
Veröffentlicht: (2025)
NOVO: Unlearning-Compliant Vision Transformers
von: Roy, Soumya, et al.
Veröffentlicht: (2025)
von: Roy, Soumya, et al.
Veröffentlicht: (2025)
ContextPRM: Leveraging Contextual Coherence for multi-domain Test-Time Scaling
von: Zhang, Haotian, et al.
Veröffentlicht: (2025)
von: Zhang, Haotian, et al.
Veröffentlicht: (2025)
Thinking vs. Doing: Agents that Reason by Scaling Test-Time Interaction
von: Shen, Junhong, et al.
Veröffentlicht: (2025)
von: Shen, Junhong, et al.
Veröffentlicht: (2025)
GM-PRM: A Generative Multimodal Process Reward Model for Multimodal Mathematical Reasoning
von: Zhang, Jianghangfan, et al.
Veröffentlicht: (2025)
von: Zhang, Jianghangfan, et al.
Veröffentlicht: (2025)
STEM-POM: Evaluating Language Models Math-Symbol Reasoning in Document Parsing
von: Zou, Jiaru, et al.
Veröffentlicht: (2024)
von: Zou, Jiaru, et al.
Veröffentlicht: (2024)
LLM-Forest: Ensemble Learning of LLMs with Graph-Augmented Prompts for Data Imputation
von: He, Xinrui, et al.
Veröffentlicht: (2024)
von: He, Xinrui, et al.
Veröffentlicht: (2024)
VisRef: Visual Refocusing while Thinking Improves Test-Time Scaling in Multi-Modal Large Reasoning Models
von: Ghosal, Soumya Suvra, et al.
Veröffentlicht: (2026)
von: Ghosal, Soumya Suvra, et al.
Veröffentlicht: (2026)
Good Learners Think Their Thinking: Generative PRM Makes Large Reasoning Model More Efficient Math Learner
von: He, Tao, et al.
Veröffentlicht: (2025)
von: He, Tao, et al.
Veröffentlicht: (2025)
Demystifying Reinforcement Learning in Agentic Reasoning
von: Yu, Zhaochen, et al.
Veröffentlicht: (2025)
von: Yu, Zhaochen, et al.
Veröffentlicht: (2025)
Orion-MSP: Multi-Scale Sparse Attention for Tabular In-Context Learning
von: Bouadi, Mohamed, et al.
Veröffentlicht: (2025)
von: Bouadi, Mohamed, et al.
Veröffentlicht: (2025)
ToolPRM: Fine-Grained Inference Scaling of Structured Outputs for Function Calling
von: Lin, Jianghao, et al.
Veröffentlicht: (2025)
von: Lin, Jianghao, et al.
Veröffentlicht: (2025)
Transformer Copilot: Learning from The Mistake Log in LLM Fine-tuning
von: Zou, Jiaru, et al.
Veröffentlicht: (2025)
von: Zou, Jiaru, et al.
Veröffentlicht: (2025)
Influence-Preserving Proxies for Gradient-Based Data Selection in LLM Fine-tuning
von: Chen, Sirui, et al.
Veröffentlicht: (2026)
von: Chen, Sirui, et al.
Veröffentlicht: (2026)
MatryoshkaThinking: Recursive Test-Time Scaling Enables Efficient Reasoning
von: Chen, Hongwei, et al.
Veröffentlicht: (2025)
von: Chen, Hongwei, et al.
Veröffentlicht: (2025)
Towards Thinking-Optimal Scaling of Test-Time Compute for LLM Reasoning
von: Yang, Wenkai, et al.
Veröffentlicht: (2025)
von: Yang, Wenkai, et al.
Veröffentlicht: (2025)
AdaFuse: Adaptive Ensemble Decoding with Test-Time Scaling for LLMs
von: Cui, Chengming, et al.
Veröffentlicht: (2026)
von: Cui, Chengming, et al.
Veröffentlicht: (2026)
PRM-BAS: Enhancing Multimodal Reasoning through PRM-guided Beam Annealing Search
von: Hu, Pengfei, et al.
Veröffentlicht: (2025)
von: Hu, Pengfei, et al.
Veröffentlicht: (2025)
AvaTaR: Optimizing LLM Agents for Tool Usage via Contrastive Reasoning
von: Wu, Shirley, et al.
Veröffentlicht: (2024)
von: Wu, Shirley, et al.
Veröffentlicht: (2024)
OctoTools: An Agentic Framework with Extensible Tools for Complex Reasoning
von: Lu, Pan, et al.
Veröffentlicht: (2025)
von: Lu, Pan, et al.
Veröffentlicht: (2025)
EvoSelect: Data-Efficient LLM Evolution for Targeted Task Adaptation
von: Li, Ting-Wei, et al.
Veröffentlicht: (2026)
von: Li, Ting-Wei, et al.
Veröffentlicht: (2026)
PaperMind: Benchmarking Agentic Reasoning and Critique over Scientific Papers in Multimodal LLMs
von: Zhao, Yanjun, et al.
Veröffentlicht: (2026)
von: Zhao, Yanjun, et al.
Veröffentlicht: (2026)
MC-Search: Evaluating and Enhancing Multimodal Agentic Search with Structured Long Reasoning Chains
von: Ning, Xuying, et al.
Veröffentlicht: (2026)
von: Ning, Xuying, et al.
Veröffentlicht: (2026)
GroundedPRM: Tree-Guided and Fidelity-Aware Process Reward Modeling for Step-Level Reasoning
von: Zhang, Yao, et al.
Veröffentlicht: (2025)
von: Zhang, Yao, et al.
Veröffentlicht: (2025)
Recursive Multi-Agent Systems
von: Yang, Xiyuan, et al.
Veröffentlicht: (2026)
von: Yang, Xiyuan, et al.
Veröffentlicht: (2026)
MIRAGE: Scaling Test-Time Inference with Parallel Graph-Retrieval-Augmented Reasoning Chains
von: Wei, Kaiwen, et al.
Veröffentlicht: (2025)
von: Wei, Kaiwen, et al.
Veröffentlicht: (2025)
Language in the Flow of Time: Time-Series-Paired Texts Weaved into a Unified Temporal Narrative
von: Li, Zihao, et al.
Veröffentlicht: (2025)
von: Li, Zihao, et al.
Veröffentlicht: (2025)
Thinking Long, but Short: Stable Sequential Test-Time Scaling for Large Reasoning Models
von: Metel, Michael R., et al.
Veröffentlicht: (2026)
von: Metel, Michael R., et al.
Veröffentlicht: (2026)
Thinking in Frames: How Visual Context and Test-Time Scaling Empower Video Reasoning
von: Li, Chengzu, et al.
Veröffentlicht: (2026)
von: Li, Chengzu, et al.
Veröffentlicht: (2026)
Start Small, Think Big: Curriculum-based Relative Policy Optimization for Visual Grounding
von: Yan, Qingyang, et al.
Veröffentlicht: (2025)
von: Yan, Qingyang, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Rethinking Test Time Scaling for Flow-Matching Generative Models
von: Yu, Qingtao, et al.
Veröffentlicht: (2025) -
ReasonFlux-PRM: Trajectory-Aware PRMs for Long Chain-of-Thought Reasoning in LLMs
von: Zou, Jiaru, et al.
Veröffentlicht: (2025) -
AutoTool: Dynamic Tool Selection and Integration for Agentic Reasoning
von: Zou, Jiaru, et al.
Veröffentlicht: (2025) -
ADROIT: A Self-Supervised Framework for Learning Robust Representations for Active Learning
von: Banerjee, Soumya, et al.
Veröffentlicht: (2025) -
DIVERSED: Relaxed Speculative Decoding via Dynamic Ensemble Verification
von: Wang, Ziyi, et al.
Veröffentlicht: (2026)