AlphaApollo: A System for Deep Agentic Reasoning
Fuente:
arXiv
Salvato in:
| Autori principali: | Zhou, Zhanke, Cao, Chentao, Feng, Xiao, Li, Xuan, Li, Zongze, Lu, Xiangyu, Yao, Jiangchao, Huang, Weikai, Cheng, Tian, Zhang, Jianghangfan, Jiang, Tangyu, Xu, Linrui, Zheng, Yiming, Miranda, Brando, Liu, Tongliang, Koyejo, Sanmi, Sugiyama, Masashi, Han, Bo |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
From Passive to Active Reasoning: Can Large Language Models Ask the Right Questions under Incomplete Information?
di: Zhou, Zhanke, et al.
Pubblicazione: (2025)
di: Zhou, Zhanke, et al.
Pubblicazione: (2025)
Reference-guided Policy Optimization for Molecular Optimization via LLM Reasoning
di: Li, Xuan, et al.
Pubblicazione: (2026)
di: Li, Xuan, et al.
Pubblicazione: (2026)
DeepInception: Hypnotize Large Language Model to Be Jailbreaker
di: Li, Xuan, et al.
Pubblicazione: (2023)
di: Li, Xuan, et al.
Pubblicazione: (2023)
Pretraining Scaling Laws for Generative Evaluations of Language Models
di: Schaeffer, Rylan, et al.
Pubblicazione: (2025)
di: Schaeffer, Rylan, et al.
Pubblicazione: (2025)
Landscape of Thoughts: Visualizing the Reasoning Process of Large Language Models
di: Zhou, Zhanke, et al.
Pubblicazione: (2025)
di: Zhou, Zhanke, et al.
Pubblicazione: (2025)
Pantograph: A Machine-to-Machine Interaction Interface for Advanced Theorem Proving, High Level Reasoning, and Data Extraction in Lean 4
di: Aniva, Leni, et al.
Pubblicazione: (2024)
di: Aniva, Leni, et al.
Pubblicazione: (2024)
From Debate to Equilibrium: Belief-Driven Multi-Agent LLM Reasoning via Bayesian Nash Equilibrium
di: Yi, Xie, et al.
Pubblicazione: (2025)
di: Yi, Xie, et al.
Pubblicazione: (2025)
Envisioning Outlier Exposure by Large Language Models for Out-of-Distribution Detection
di: Cao, Chentao, et al.
Pubblicazione: (2024)
di: Cao, Chentao, et al.
Pubblicazione: (2024)
Is Pre-training Truly Better Than Meta-Learning?
di: Miranda, Brando, et al.
Pubblicazione: (2023)
di: Miranda, Brando, et al.
Pubblicazione: (2023)
CURE: Cultural Understanding and Reasoning Evaluation - A Framework for "Thick" Culture Alignment Evaluation in LLMs
di: Vo, Truong, et al.
Pubblicazione: (2025)
di: Vo, Truong, et al.
Pubblicazione: (2025)
Less is More: One-shot Subgraph Reasoning on Large-scale Knowledge Graphs
di: Zhou, Zhanke, et al.
Pubblicazione: (2024)
di: Zhou, Zhanke, et al.
Pubblicazione: (2024)
Neural Atoms: Propagating Long-range Interaction in Molecular Graphs through Efficient Communication Channel
di: Li, Xuan, et al.
Pubblicazione: (2023)
di: Li, Xuan, et al.
Pubblicazione: (2023)
Noisy Test-Time Adaptation in Vision-Language Models
di: Cao, Chentao, et al.
Pubblicazione: (2025)
di: Cao, Chentao, et al.
Pubblicazione: (2025)
Putnam-AXIOM: A Functional and Static Benchmark for Measuring Higher Level Mathematical Reasoning in LLMs
di: Gulati, Aryan, et al.
Pubblicazione: (2025)
di: Gulati, Aryan, et al.
Pubblicazione: (2025)
ZIP-FIT: Embedding-Free Data Selection via Compression-Based Alignment
di: Obbad, Elyas, et al.
Pubblicazione: (2024)
di: Obbad, Elyas, et al.
Pubblicazione: (2024)
Beyond Scale: The Diversity Coefficient as a Data Quality Metric for Variability in Natural Language Data
di: Miranda, Brando, et al.
Pubblicazione: (2023)
di: Miranda, Brando, et al.
Pubblicazione: (2023)
Lean-ing on Quality: How High-Quality Data Beats Diverse Multilingual Data in AutoFormalization
di: Chan, Willy, et al.
Pubblicazione: (2025)
di: Chan, Willy, et al.
Pubblicazione: (2025)
Evaluating the Robustness of Chinchilla Compute-Optimal Scaling
di: Schaeffer, Rylan, et al.
Pubblicazione: (2025)
di: Schaeffer, Rylan, et al.
Pubblicazione: (2025)
Co-rewarding: Stable Self-supervised RL for Eliciting Reasoning in Large Language Models
di: Zhang, Zizhuo, et al.
Pubblicazione: (2025)
di: Zhang, Zizhuo, et al.
Pubblicazione: (2025)
Causally Inspired Regularization Enables Domain General Representations
di: Salaudeen, Olawale, et al.
Pubblicazione: (2024)
di: Salaudeen, Olawale, et al.
Pubblicazione: (2024)
Let's Measure Information Step-by-Step: AI-Based Evaluation Beyond Vibes
di: Robertson, Zachary, et al.
Pubblicazione: (2025)
di: Robertson, Zachary, et al.
Pubblicazione: (2025)
Reasoning Models Don't Just Think Longer, They Move Differently
di: Gjølbye, Anders, et al.
Pubblicazione: (2026)
di: Gjølbye, Anders, et al.
Pubblicazione: (2026)
SATBench: Benchmarking LLMs' Logical Reasoning via Automated Puzzle Generation from SAT Formulas
di: Wei, Anjiang, et al.
Pubblicazione: (2025)
di: Wei, Anjiang, et al.
Pubblicazione: (2025)
In-Situ Behavioral Evaluation for LLM Fairness, Not Standardized-Test Scores
di: Tang, Zeyu, et al.
Pubblicazione: (2026)
di: Tang, Zeyu, et al.
Pubblicazione: (2026)
Quantifying the Importance of Data Alignment in Downstream Model Performance
di: Chawla, Krrish, et al.
Pubblicazione: (2025)
di: Chawla, Krrish, et al.
Pubblicazione: (2025)
Model Inversion Attacks: A Survey of Approaches and Countermeasures
di: Zhou, Zhanke, et al.
Pubblicazione: (2024)
di: Zhou, Zhanke, et al.
Pubblicazione: (2024)
RewardFlow: Topology-Aware Reward Propagation on State Graphs for Agentic RL with Large Language Models
di: Feng, Xiao, et al.
Pubblicazione: (2026)
di: Feng, Xiao, et al.
Pubblicazione: (2026)
A Framework for Objective-Driven Dynamical Stochastic Fields
di: Zhang, Yibo Jacky, et al.
Pubblicazione: (2025)
di: Zhang, Yibo Jacky, et al.
Pubblicazione: (2025)
Decoupling the Class Label and the Target Concept in Machine Unlearning
di: Zhu, Jianing, et al.
Pubblicazione: (2024)
di: Zhu, Jianing, et al.
Pubblicazione: (2024)
Reliable and Efficient Amortized Model-based Evaluation
di: Truong, Sang, et al.
Pubblicazione: (2025)
di: Truong, Sang, et al.
Pubblicazione: (2025)
Towards Effective Evaluations and Comparisons for LLM Unlearning Methods
di: Wang, Qizhou, et al.
Pubblicazione: (2024)
di: Wang, Qizhou, et al.
Pubblicazione: (2024)
BrokenBind: Universal Modality Exploration beyond Dataset Boundaries
di: Huang, Zhuo, et al.
Pubblicazione: (2026)
di: Huang, Zhuo, et al.
Pubblicazione: (2026)
In-Context Learning of Energy Functions
di: Schaeffer, Rylan, et al.
Pubblicazione: (2024)
di: Schaeffer, Rylan, et al.
Pubblicazione: (2024)
Discovering Implicit Large Language Model Alignment Objectives
di: Chen, Edward, et al.
Pubblicazione: (2026)
di: Chen, Edward, et al.
Pubblicazione: (2026)
High-Dimensional Markov-switching Ordinary Differential Processes
di: Tsai, Katherine, et al.
Pubblicazione: (2024)
di: Tsai, Katherine, et al.
Pubblicazione: (2024)
Distributional Machine Unlearning via Selective Data Removal
di: Allouah, Youssef, et al.
Pubblicazione: (2025)
di: Allouah, Youssef, et al.
Pubblicazione: (2025)
SCENEBench: An Audio Understanding Benchmark Grounded in Assistive and Industrial Use Cases
di: Iyer, Laya, et al.
Pubblicazione: (2026)
di: Iyer, Laya, et al.
Pubblicazione: (2026)
HiFA: High-fidelity Text-to-3D Generation with Advanced Diffusion Guidance
di: Zhu, Junzhe, et al.
Pubblicazione: (2023)
di: Zhu, Junzhe, et al.
Pubblicazione: (2023)
Why Has Predicting Downstream Capabilities of Frontier AI Models with Scale Remained Elusive?
di: Schaeffer, Rylan, et al.
Pubblicazione: (2024)
di: Schaeffer, Rylan, et al.
Pubblicazione: (2024)
BadLabel: A Robust Perspective on Evaluating and Enhancing Label-noise Learning
di: Zhang, Jingfeng, et al.
Pubblicazione: (2023)
di: Zhang, Jingfeng, et al.
Pubblicazione: (2023)
Documenti analoghi
-
From Passive to Active Reasoning: Can Large Language Models Ask the Right Questions under Incomplete Information?
di: Zhou, Zhanke, et al.
Pubblicazione: (2025) -
Reference-guided Policy Optimization for Molecular Optimization via LLM Reasoning
di: Li, Xuan, et al.
Pubblicazione: (2026) -
DeepInception: Hypnotize Large Language Model to Be Jailbreaker
di: Li, Xuan, et al.
Pubblicazione: (2023) -
Pretraining Scaling Laws for Generative Evaluations of Language Models
di: Schaeffer, Rylan, et al.
Pubblicazione: (2025) -
Landscape of Thoughts: Visualizing the Reasoning Process of Large Language Models
di: Zhou, Zhanke, et al.
Pubblicazione: (2025)