Less is More: Recursive Reasoning with Tiny Networks
Fuente:
arXiv
Salvato in:
| Autore principale: | Jolicoeur-Martineau, Alexia |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Probabilistic Tiny Recursive Model
di: Sghaier, Amin, et al.
Pubblicazione: (2026)
di: Sghaier, Amin, et al.
Pubblicazione: (2026)
Multi-Agent Game Generation and Evaluation via Audio-Visual Recordings
di: Jolicoeur-Martineau, Alexia
Pubblicazione: (2025)
di: Jolicoeur-Martineau, Alexia
Pubblicazione: (2025)
Beyond FVD: Enhanced Evaluation Metrics for Video Generation Quality
di: Luo, Ge Ya, et al.
Pubblicazione: (2024)
di: Luo, Ge Ya, et al.
Pubblicazione: (2024)
Understanding Adam Requires Better Rotation Dependent Assumptions
di: Zhang, Tianyue H., et al.
Pubblicazione: (2024)
di: Zhang, Tianyue H., et al.
Pubblicazione: (2024)
Test-time Adaptation of Tiny Recursive Models
di: McGovern, Ronan Killian
Pubblicazione: (2025)
di: McGovern, Ronan Killian
Pubblicazione: (2025)
Generating and Imputing Tabular Data via Diffusion and Flow-based Gradient-Boosted Trees
di: Jolicoeur-Martineau, Alexia, et al.
Pubblicazione: (2023)
di: Jolicoeur-Martineau, Alexia, et al.
Pubblicazione: (2023)
LoGAH: Predicting 774-Million-Parameter Transformers using Graph HyperNetworks with 1/100 Parameters
di: Zhou, Xinyu, et al.
Pubblicazione: (2024)
di: Zhou, Xinyu, et al.
Pubblicazione: (2024)
Generating Tabular Data Using Heterogeneous Sequential Feature Forest Flow Matching
di: Akazan, Ange-Clément, et al.
Pubblicazione: (2024)
di: Akazan, Ange-Clément, et al.
Pubblicazione: (2024)
Less Is More -- On the Importance of Sparsification for Transformers and Graph Neural Networks for TSP
di: Lischka, Attila, et al.
Pubblicazione: (2024)
di: Lischka, Attila, et al.
Pubblicazione: (2024)
Transformer Multivariate Forecasting: Less is More?
di: Xu, Jingjing, et al.
Pubblicazione: (2023)
di: Xu, Jingjing, et al.
Pubblicazione: (2023)
On the Sample Efficiency of Inverse Dynamics Models for Semi-Supervised Imitation Learning
di: Morin, Sacha, et al.
Pubblicazione: (2026)
di: Morin, Sacha, et al.
Pubblicazione: (2026)
Less is More: Undertraining Experts Improves Model Upcycling
di: Horoi, Stefan, et al.
Pubblicazione: (2025)
di: Horoi, Stefan, et al.
Pubblicazione: (2025)
Less is More: on the Over-Globalizing Problem in Graph Transformers
di: Xing, Yujie, et al.
Pubblicazione: (2024)
di: Xing, Yujie, et al.
Pubblicazione: (2024)
Less Noise, More Voice: Reinforcement Learning for Reasoning via Instruction Purification
di: Guo, Yiju, et al.
Pubblicazione: (2026)
di: Guo, Yiju, et al.
Pubblicazione: (2026)
Tiny Recursive Models on ARC-AGI-1: Inductive Biases, Identity Conditioning, and Test-Time Compute
di: Roye-Azar, Antonio, et al.
Pubblicazione: (2025)
di: Roye-Azar, Antonio, et al.
Pubblicazione: (2025)
Accelerating Training Speed of Tiny Recursive Models with Curriculum Guided Adaptive Recursion
di: Qasim, Kaleem Ullah, et al.
Pubblicazione: (2025)
di: Qasim, Kaleem Ullah, et al.
Pubblicazione: (2025)
LIMR: Less is More for RL Scaling
di: Li, Xuefeng, et al.
Pubblicazione: (2025)
di: Li, Xuefeng, et al.
Pubblicazione: (2025)
Interaction Locality in Hierarchical Recursive Reasoning
di: Miyanishi, Yosuke, et al.
Pubblicazione: (2026)
di: Miyanishi, Yosuke, et al.
Pubblicazione: (2026)
Recursive Inference Machines for Neural Reasoning
di: Komisarczyk, Mieszko, et al.
Pubblicazione: (2026)
di: Komisarczyk, Mieszko, et al.
Pubblicazione: (2026)
Less is More: One-shot Subgraph Reasoning on Large-scale Knowledge Graphs
di: Zhou, Zhanke, et al.
Pubblicazione: (2024)
di: Zhou, Zhanke, et al.
Pubblicazione: (2024)
Draft Less, Retrieve More: Hybrid Tree Construction for Speculative Decoding
di: Shen, Yuhao, et al.
Pubblicazione: (2026)
di: Shen, Yuhao, et al.
Pubblicazione: (2026)
Less is More: Pseudo-Label Filtering for Continual Test-Time Adaptation
di: Tan, Jiayao, et al.
Pubblicazione: (2024)
di: Tan, Jiayao, et al.
Pubblicazione: (2024)
Supernova: Achieving More with Less in Transformer Architectures
di: Tanase, Andrei-Valentin, et al.
Pubblicazione: (2025)
di: Tanase, Andrei-Valentin, et al.
Pubblicazione: (2025)
No More, No Less: Task Alignment in Terminal Agents
di: Mavali, Sina, et al.
Pubblicazione: (2026)
di: Mavali, Sina, et al.
Pubblicazione: (2026)
Less is More: Multimodal Region Representation via Pairwise Inter-view Learning
di: Namgung, Min, et al.
Pubblicazione: (2025)
di: Namgung, Min, et al.
Pubblicazione: (2025)
Forget Less, Generalize More: Unifying Temporal and Structural Adaptation for Dynamic Graphs
di: Chang, Qian, et al.
Pubblicazione: (2026)
di: Chang, Qian, et al.
Pubblicazione: (2026)
Cut Less, Fold More: Model Compression through the Lens of Projection Geometry
di: Saukh, Olga, et al.
Pubblicazione: (2026)
di: Saukh, Olga, et al.
Pubblicazione: (2026)
Input-Time Scaling: Adding Noise and Irrelevance into Less-Is-More Drastically Improves Reasoning Performance and Efficiency
di: Huang, Rapheal, et al.
Pubblicazione: (2025)
di: Huang, Rapheal, et al.
Pubblicazione: (2025)
Recursive Decomposition with Dependencies for Generic Divide-and-Conquer Reasoning
di: Hernández-Gutiérrez, Sergio, et al.
Pubblicazione: (2025)
di: Hernández-Gutiérrez, Sergio, et al.
Pubblicazione: (2025)
Less is More for Improving Automatic Evaluation of Factual Consistency
di: Wang, Tong, et al.
Pubblicazione: (2024)
di: Wang, Tong, et al.
Pubblicazione: (2024)
Less is More: Unlocking Specialization of Time Series Foundation Models via Structured Pruning
di: Zhao, Lifan, et al.
Pubblicazione: (2025)
di: Zhao, Lifan, et al.
Pubblicazione: (2025)
When Less is More: 8-bit Quantization Improves Continual Learning in Large Language Models
di: Zhang, Michael S., et al.
Pubblicazione: (2025)
di: Zhang, Michael S., et al.
Pubblicazione: (2025)
Train Less, Learn More: Adaptive Efficient Rollout Optimization for Group-Based Reinforcement Learning
di: Zhang, Zhi, et al.
Pubblicazione: (2026)
di: Zhang, Zhi, et al.
Pubblicazione: (2026)
Any-Property-Conditional Molecule Generation with Self-Criticism using Spanning Trees
di: Jolicoeur-Martineau, Alexia, et al.
Pubblicazione: (2024)
di: Jolicoeur-Martineau, Alexia, et al.
Pubblicazione: (2024)
Tina: Tiny Reasoning Models via LoRA
di: Wang, Shangshang, et al.
Pubblicazione: (2025)
di: Wang, Shangshang, et al.
Pubblicazione: (2025)
When More is Less: Understanding Chain-of-Thought Length in LLMs
di: Wu, Yuyang, et al.
Pubblicazione: (2025)
di: Wu, Yuyang, et al.
Pubblicazione: (2025)
Less is More: Denoising Knowledge Graphs For Retrieval Augmented Generation
di: Zheng, Yilun, et al.
Pubblicazione: (2025)
di: Zheng, Yilun, et al.
Pubblicazione: (2025)
Less is More: Local Intrinsic Dimensions of Contextual Language Models
di: Ruppik, Benjamin Matthias, et al.
Pubblicazione: (2025)
di: Ruppik, Benjamin Matthias, et al.
Pubblicazione: (2025)
TinyGraph: Joint Feature and Node Condensation for Graph Neural Networks
di: Liu, Yezi, et al.
Pubblicazione: (2024)
di: Liu, Yezi, et al.
Pubblicazione: (2024)
Learning More with Less: A Dynamic Dual-Level Down-Sampling Framework for Efficient Policy Optimization
di: Wang, Chao, et al.
Pubblicazione: (2025)
di: Wang, Chao, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Probabilistic Tiny Recursive Model
di: Sghaier, Amin, et al.
Pubblicazione: (2026) -
Multi-Agent Game Generation and Evaluation via Audio-Visual Recordings
di: Jolicoeur-Martineau, Alexia
Pubblicazione: (2025) -
Beyond FVD: Enhanced Evaluation Metrics for Video Generation Quality
di: Luo, Ge Ya, et al.
Pubblicazione: (2024) -
Understanding Adam Requires Better Rotation Dependent Assumptions
di: Zhang, Tianyue H., et al.
Pubblicazione: (2024) -
Test-time Adaptation of Tiny Recursive Models
di: McGovern, Ronan Killian
Pubblicazione: (2025)