The Role of Generator Access in Autoregressive Post-Training
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Rege, Amit Kiran |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Data Attribution in Adaptive Learning
von: Rege, Amit Kiran
Veröffentlicht: (2026)
von: Rege, Amit Kiran
Veröffentlicht: (2026)
Multi-Agent Lipschitz Bandits
von: Chakraborty, Sourav, et al.
Veröffentlicht: (2026)
von: Chakraborty, Sourav, et al.
Veröffentlicht: (2026)
A Unified Framework for Locality in Scalable MARL
von: Chakraborty, Sourav, et al.
Veröffentlicht: (2026)
von: Chakraborty, Sourav, et al.
Veröffentlicht: (2026)
Flickering Multi-Armed Bandits
von: Chakraborty, Sourav, et al.
Veröffentlicht: (2026)
von: Chakraborty, Sourav, et al.
Veröffentlicht: (2026)
Incentivized Lipschitz Bandits
von: Chakraborty, Sourav, et al.
Veröffentlicht: (2025)
von: Chakraborty, Sourav, et al.
Veröffentlicht: (2025)
Where Did Your Model Learn That? Label-free Influence for Self-supervised Learning
von: Harilal, Nidhin, et al.
Veröffentlicht: (2024)
von: Harilal, Nidhin, et al.
Veröffentlicht: (2024)
Autoregressive Adversarial Post-Training for Real-Time Interactive Video Generation
von: Lin, Shanchuan, et al.
Veröffentlicht: (2025)
von: Lin, Shanchuan, et al.
Veröffentlicht: (2025)
Chunky Post-Training: Data Driven Failures of Generalization
von: Murray, Seoirse, et al.
Veröffentlicht: (2026)
von: Murray, Seoirse, et al.
Veröffentlicht: (2026)
Data Generation for Hardware-Friendly Post-Training Quantization
von: Dikstein, Lior, et al.
Veröffentlicht: (2024)
von: Dikstein, Lior, et al.
Veröffentlicht: (2024)
PAL: Pluralistic Alignment Framework for Learning from Heterogeneous Preferences
von: Chen, Daiwei, et al.
Veröffentlicht: (2024)
von: Chen, Daiwei, et al.
Veröffentlicht: (2024)
Decentralized Autoregressive Generation
von: Maschan, Stepan, et al.
Veröffentlicht: (2026)
von: Maschan, Stepan, et al.
Veröffentlicht: (2026)
Cognitive Fatigue in Autoregressive Transformers: Formalization and Measurement
von: Marwah, Riju, et al.
Veröffentlicht: (2026)
von: Marwah, Riju, et al.
Veröffentlicht: (2026)
Post-Training Augmentation Invariance
von: Eikenberry, Keenan, et al.
Veröffentlicht: (2025)
von: Eikenberry, Keenan, et al.
Veröffentlicht: (2025)
Improving Autoregressive Training with Dynamic Oracles
von: Yang, Jianing, et al.
Veröffentlicht: (2024)
von: Yang, Jianing, et al.
Veröffentlicht: (2024)
On Mesa-Optimization in Autoregressively Trained Transformers: Emergence and Capability
von: Zheng, Chenyu, et al.
Veröffentlicht: (2024)
von: Zheng, Chenyu, et al.
Veröffentlicht: (2024)
Training Dynamics Impact Post-Training Quantization Robustness
von: Catalan-Tatjer, Albert, et al.
Veröffentlicht: (2025)
von: Catalan-Tatjer, Albert, et al.
Veröffentlicht: (2025)
Unifying Autoregressive and Diffusion-Based Sequence Generation
von: Fathi, Nima, et al.
Veröffentlicht: (2025)
von: Fathi, Nima, et al.
Veröffentlicht: (2025)
On Powerful Ways to Generate: Autoregression, Diffusion, and Beyond
von: Yang, Chenxiao, et al.
Veröffentlicht: (2025)
von: Yang, Chenxiao, et al.
Veröffentlicht: (2025)
Accelerating Training of Autoregressive Video Generation Models via Local Optimization with Representation Continuity
von: Zhou, Yucheng, et al.
Veröffentlicht: (2026)
von: Zhou, Yucheng, et al.
Veröffentlicht: (2026)
Rethinking Training Dynamics in Scale-wise Autoregressive Generation
von: Zhou, Gengze, et al.
Veröffentlicht: (2025)
von: Zhou, Gengze, et al.
Veröffentlicht: (2025)
WILDCHAT-50M: A Deep Dive Into the Role of Synthetic Data in Post-Training
von: Feuer, Benjamin, et al.
Veröffentlicht: (2025)
von: Feuer, Benjamin, et al.
Veröffentlicht: (2025)
Convex Dataset Valuation for Post-Training
von: Zeng, Siqi, et al.
Veröffentlicht: (2026)
von: Zeng, Siqi, et al.
Veröffentlicht: (2026)
Steerable Scene Generation with Post Training and Inference-Time Search
von: Pfaff, Nicholas, et al.
Veröffentlicht: (2025)
von: Pfaff, Nicholas, et al.
Veröffentlicht: (2025)
TROJAN-GUARD: Hardware Trojans Detection Using GNN in RTL Designs
von: Thorat, Kiran, et al.
Veröffentlicht: (2025)
von: Thorat, Kiran, et al.
Veröffentlicht: (2025)
Compositional Generalization in Autoregressive Models via Logit Composition
von: Kumar, Aakash, et al.
Veröffentlicht: (2026)
von: Kumar, Aakash, et al.
Veröffentlicht: (2026)
ENMA: Tokenwise Autoregression for Generative Neural PDE Operators
von: Koupaï, Armand Kassaï, et al.
Veröffentlicht: (2025)
von: Koupaï, Armand Kassaï, et al.
Veröffentlicht: (2025)
Instruction-Guided Autoregressive Neural Network Parameter Generation
von: Bedionita, Soro, et al.
Veröffentlicht: (2025)
von: Bedionita, Soro, et al.
Veröffentlicht: (2025)
Active Exploration via Autoregressive Generation of Missing Data
von: Cai, Tiffany Tianhui, et al.
Veröffentlicht: (2024)
von: Cai, Tiffany Tianhui, et al.
Veröffentlicht: (2024)
Bidirectional Representations Augmented Autoregressive Biological Sequence Generation
von: Zhang, Xiang, et al.
Veröffentlicht: (2025)
von: Zhang, Xiang, et al.
Veröffentlicht: (2025)
Pard: Permutation-Invariant Autoregressive Diffusion for Graph Generation
von: Zhao, Lingxiao, et al.
Veröffentlicht: (2024)
von: Zhao, Lingxiao, et al.
Veröffentlicht: (2024)
Parallelizing Autoregressive Generation with Variational State Space Models
von: Lambrechts, Gaspard, et al.
Veröffentlicht: (2024)
von: Lambrechts, Gaspard, et al.
Veröffentlicht: (2024)
Step-Wise Refusal Dynamics in Autoregressive and Diffusion Language Models
von: Rahimi, Eliron, et al.
Veröffentlicht: (2026)
von: Rahimi, Eliron, et al.
Veröffentlicht: (2026)
Dr. Post-Training: A Data Regularization Perspective on LLM Post-Training
von: Hu, Pingbang, et al.
Veröffentlicht: (2026)
von: Hu, Pingbang, et al.
Veröffentlicht: (2026)
Train Separately, Merge Together: Modular Post-Training with Mixture-of-Experts
von: Morrison, Jacob, et al.
Veröffentlicht: (2026)
von: Morrison, Jacob, et al.
Veröffentlicht: (2026)
Materium: An Autoregressive Approach for Material Generation
von: Dobberstein, Niklas, et al.
Veröffentlicht: (2025)
von: Dobberstein, Niklas, et al.
Veröffentlicht: (2025)
Projected Autoregression: Autoregressive Language Generation in Continuous State Space
von: Naparstek, Oshri
Veröffentlicht: (2026)
von: Naparstek, Oshri
Veröffentlicht: (2026)
DQNC2S: DQN-based Cross-stream Crisis event Summarizer
von: Cambrin, Daniele Rege, et al.
Veröffentlicht: (2024)
von: Cambrin, Daniele Rege, et al.
Veröffentlicht: (2024)
Apriel-1.5-OpenReasoner: RL Post-Training for General-Purpose and Efficient Reasoning
von: Pardinas, Rafael, et al.
Veröffentlicht: (2026)
von: Pardinas, Rafael, et al.
Veröffentlicht: (2026)
Robust Post-Training for Generative Recommenders: Why Exponential Reward-Weighted SFT Outperforms RLHF
von: Chidambaram, Keertana, et al.
Veröffentlicht: (2026)
von: Chidambaram, Keertana, et al.
Veröffentlicht: (2026)
MAGE: Multi-scale Autoregressive Generation for Offline Reinforcement Learning
von: Lin, Chenxing, et al.
Veröffentlicht: (2026)
von: Lin, Chenxing, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Data Attribution in Adaptive Learning
von: Rege, Amit Kiran
Veröffentlicht: (2026) -
Multi-Agent Lipschitz Bandits
von: Chakraborty, Sourav, et al.
Veröffentlicht: (2026) -
A Unified Framework for Locality in Scalable MARL
von: Chakraborty, Sourav, et al.
Veröffentlicht: (2026) -
Flickering Multi-Armed Bandits
von: Chakraborty, Sourav, et al.
Veröffentlicht: (2026) -
Incentivized Lipschitz Bandits
von: Chakraborty, Sourav, et al.
Veröffentlicht: (2025)