Tiny Autoregressive Recursive Models
Fuente:
arXiv
Guardado en:
| Autores principales: | Rauba, Paulius, Fanconi, Claudio, van der Schaar, Mihaela |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Deep Hierarchical Learning with Nested Subspace Networks for Large Language Models
por: Rauba, Paulius, et al.
Publicado: (2025)
por: Rauba, Paulius, et al.
Publicado: (2025)
Visualizing token importance for black-box language models
por: Rauba, Paulius, et al.
Publicado: (2025)
por: Rauba, Paulius, et al.
Publicado: (2025)
Quantifying perturbation impacts for large language models
por: Rauba, Paulius, et al.
Publicado: (2024)
por: Rauba, Paulius, et al.
Publicado: (2024)
No More, No Less: Least-Privilege Language Models
por: Rauba, Paulius, et al.
Publicado: (2026)
por: Rauba, Paulius, et al.
Publicado: (2026)
Context-Aware Testing: A New Paradigm for Model Testing with Large Language Models
por: Rauba, Paulius, et al.
Publicado: (2024)
por: Rauba, Paulius, et al.
Publicado: (2024)
Self-Healing Machine Learning: A Framework for Autonomous Adaptation in Real-World Environments
por: Rauba, Paulius, et al.
Publicado: (2024)
por: Rauba, Paulius, et al.
Publicado: (2024)
Few-shot Steerable Alignment: Adapting Rewards and LLM Policies with Neural Processes
por: Kobalczyk, Katarzyna, et al.
Publicado: (2024)
por: Kobalczyk, Katarzyna, et al.
Publicado: (2024)
Redefining Digital Health Interfaces with Large Language Models
por: Imrie, Fergus, et al.
Publicado: (2023)
por: Imrie, Fergus, et al.
Publicado: (2023)
Statistical Hypothesis Testing for Auditing Robustness in Language Models
por: Rauba, Paulius, et al.
Publicado: (2025)
por: Rauba, Paulius, et al.
Publicado: (2025)
Cascaded Language Models for Cost-effective Human-AI Decision-Making
por: Fanconi, Claudio, et al.
Publicado: (2025)
por: Fanconi, Claudio, et al.
Publicado: (2025)
Multi-Agent Systems Should be Treated as Principal-Agent Problems
por: Rauba, Paulius, et al.
Publicado: (2026)
por: Rauba, Paulius, et al.
Publicado: (2026)
Discovering Preference Optimization Algorithms with and for Large Language Models
por: Lu, Chris, et al.
Publicado: (2024)
por: Lu, Chris, et al.
Publicado: (2024)
Eliciting Numerical Predictive Distributions of LLMs Without Autoregression
por: Piskorz, Julianna, et al.
Publicado: (2026)
por: Piskorz, Julianna, et al.
Publicado: (2026)
Matchmaker: Self-Improving Large Language Model Programs for Schema Matching
por: Seedat, Nabeel, et al.
Publicado: (2024)
por: Seedat, Nabeel, et al.
Publicado: (2024)
Learning Reasoning Rewards from Expert Demonstrations with Inverse Reinforcement Learning
por: Fanconi, Claudio, et al.
Publicado: (2025)
por: Fanconi, Claudio, et al.
Publicado: (2025)
Why Tabular Foundation Models Should Be a Research Priority
por: van Breugel, Boris, et al.
Publicado: (2024)
por: van Breugel, Boris, et al.
Publicado: (2024)
Inverse-RLignment: Large Language Model Alignment from Demonstrations through Inverse Reinforcement Learning
por: Sun, Hao, et al.
Publicado: (2024)
por: Sun, Hao, et al.
Publicado: (2024)
Language Bottleneck Models for Qualitative Knowledge State Modeling
por: Berthon, Antonin, et al.
Publicado: (2025)
por: Berthon, Antonin, et al.
Publicado: (2025)
Shape Arithmetic Expressions: Advancing Scientific Discovery Beyond Closed-Form Equations
por: Kacprzyk, Krzysztof, et al.
Publicado: (2024)
por: Kacprzyk, Krzysztof, et al.
Publicado: (2024)
No Equations Needed: Learning System Dynamics Without Relying on Closed-Form ODEs
por: Kacprzyk, Krzysztof, et al.
Publicado: (2025)
por: Kacprzyk, Krzysztof, et al.
Publicado: (2025)
Towards Automated Knowledge Integration From Human-Interpretable Representations
por: Kobalczyk, Katarzyna, et al.
Publicado: (2024)
por: Kobalczyk, Katarzyna, et al.
Publicado: (2024)
On Error Propagation of Diffusion Models
por: Li, Yangming, et al.
Publicado: (2023)
por: Li, Yangming, et al.
Publicado: (2023)
Inverse Reinforcement Learning Meets Large Language Model Post-Training: Basics, Advances, and Opportunities
por: Sun, Hao, et al.
Publicado: (2025)
por: Sun, Hao, et al.
Publicado: (2025)
Active Timepoint Selection for Learning Measure-Valued Trajectories
por: Huynh, Nicolas, et al.
Publicado: (2026)
por: Huynh, Nicolas, et al.
Publicado: (2026)
Hyperparameter Trajectory Inference with Conditional Lagrangian Optimal Transport
por: Amad, Harry, et al.
Publicado: (2026)
por: Amad, Harry, et al.
Publicado: (2026)
Not All Explanations for Deep Learning Phenomena Are Equally Valuable
por: Jeffares, Alan, et al.
Publicado: (2025)
por: Jeffares, Alan, et al.
Publicado: (2025)
Preference Learning for AI Alignment: a Causal Perspective
por: Kobalczyk, Katarzyna, et al.
Publicado: (2025)
por: Kobalczyk, Katarzyna, et al.
Publicado: (2025)
Discovery of Hidden Miscalibration Regimes
por: Kobalczyk, Katarzyna, et al.
Publicado: (2026)
por: Kobalczyk, Katarzyna, et al.
Publicado: (2026)
Simulating Viva Voce Examinations to Evaluate Clinical Reasoning in Large Language Models
por: Chiu, Christopher, et al.
Publicado: (2025)
por: Chiu, Christopher, et al.
Publicado: (2025)
Risk-Sensitive Diffusion: Robustly Optimizing Diffusion Models with Noisy Samples
por: Li, Yangming, et al.
Publicado: (2024)
por: Li, Yangming, et al.
Publicado: (2024)
Deep Learning Through A Telescoping Lens: A Simple Model Provides Empirical Insights On Grokking, Gradient Boosting & Beyond
por: Jeffares, Alan, et al.
Publicado: (2024)
por: Jeffares, Alan, et al.
Publicado: (2024)
A Study of Posterior Stability for Time-Series Latent Diffusion
por: Li, Yangming, et al.
Publicado: (2024)
por: Li, Yangming, et al.
Publicado: (2024)
Automatically Learning Hybrid Digital Twins of Dynamical Systems
por: Holt, Samuel, et al.
Publicado: (2024)
por: Holt, Samuel, et al.
Publicado: (2024)
Decision Tree Induction Through LLMs via Semantically-Aware Evolution
por: Liu, Tennison, et al.
Publicado: (2025)
por: Liu, Tennison, et al.
Publicado: (2025)
Why do Random Forests Work? Understanding Tree Ensembles as Self-Regularizing Adaptive Smoothers
por: Curth, Alicia, et al.
Publicado: (2024)
por: Curth, Alicia, et al.
Publicado: (2024)
Dissecting Sample Hardness: A Fine-Grained Analysis of Hardness Characterization Methods for Data-Centric AI
por: Seedat, Nabeel, et al.
Publicado: (2024)
por: Seedat, Nabeel, et al.
Publicado: (2024)
Quantifying Aleatoric Uncertainty of the Treatment Effect: A Novel Orthogonal Learner
por: Melnychuk, Valentyn, et al.
Publicado: (2024)
por: Melnychuk, Valentyn, et al.
Publicado: (2024)
Adaptive Experiment Design with Synthetic Controls
por: Hüyük, Alihan, et al.
Publicado: (2024)
por: Hüyük, Alihan, et al.
Publicado: (2024)
What's the next frontier for Data-centric AI? Data Savvy Agents
por: Seedat, Nabeel, et al.
Publicado: (2025)
por: Seedat, Nabeel, et al.
Publicado: (2025)
Position: All Current Generative Fidelity and Diversity Metrics are Flawed
por: Räisä, Ossi, et al.
Publicado: (2025)
por: Räisä, Ossi, et al.
Publicado: (2025)
Ejemplares similares
-
Deep Hierarchical Learning with Nested Subspace Networks for Large Language Models
por: Rauba, Paulius, et al.
Publicado: (2025) -
Visualizing token importance for black-box language models
por: Rauba, Paulius, et al.
Publicado: (2025) -
Quantifying perturbation impacts for large language models
por: Rauba, Paulius, et al.
Publicado: (2024) -
No More, No Less: Least-Privilege Language Models
por: Rauba, Paulius, et al.
Publicado: (2026) -
Context-Aware Testing: A New Paradigm for Model Testing with Large Language Models
por: Rauba, Paulius, et al.
Publicado: (2024)