Diffusion Controller: Framework, Algorithms and Parameterization
Fuente:
arXiv
Guardado en:
| Autores principales: | Yang, Tong, Ryu, Moonkyung, Hsu, Chih-Wei, Tennenholtz, Guy, Chi, Yuejie, Boutilier, Craig, Dai, Bo |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Preference Adaptive and Sequential Text-to-Image Generation
por: Nabati, Ofir, et al.
Publicado: (2024)
por: Nabati, Ofir, et al.
Publicado: (2024)
Embedding-Aligned Language Models
por: Tennenholtz, Guy, et al.
Publicado: (2024)
por: Tennenholtz, Guy, et al.
Publicado: (2024)
Synthetic Dialogue Generation for Interactive Conversational Elicitation & Recommendation (ICER)
por: Ryu, Moonkyung, et al.
Publicado: (2025)
por: Ryu, Moonkyung, et al.
Publicado: (2025)
Controllable User Simulation
por: Tennenholtz, Guy, et al.
Publicado: (2026)
por: Tennenholtz, Guy, et al.
Publicado: (2026)
Demystifying Embedding Spaces using Large Language Models
por: Tennenholtz, Guy, et al.
Publicado: (2023)
por: Tennenholtz, Guy, et al.
Publicado: (2023)
Descriptive History Representations: Learning Representations by Answering Questions
por: Tennenholtz, Guy, et al.
Publicado: (2025)
por: Tennenholtz, Guy, et al.
Publicado: (2025)
Inference-Aware Fine-Tuning for Best-of-N Sampling in Large Language Models
por: Chow, Yinlam, et al.
Publicado: (2024)
por: Chow, Yinlam, et al.
Publicado: (2024)
Asking Clarifying Questions for Preference Elicitation With Large Language Models
por: Montazeralghaem, Ali, et al.
Publicado: (2025)
por: Montazeralghaem, Ali, et al.
Publicado: (2025)
DynaMITE-RL: A Dynamic Model for Improved Temporal Meta-Reinforcement Learning
por: Liang, Anthony, et al.
Publicado: (2024)
por: Liang, Anthony, et al.
Publicado: (2024)
Incentivize without Bonus: Provably Efficient Model-based Online Multi-agent RL for Markov Games
por: Yang, Tong, et al.
Publicado: (2025)
por: Yang, Tong, et al.
Publicado: (2025)
Federated Natural Policy Gradient and Actor Critic Methods for Multi-task Reinforcement Learning
por: Yang, Tong, et al.
Publicado: (2023)
por: Yang, Tong, et al.
Publicado: (2023)
Representation-Driven Reinforcement Learning
por: Nabati, Ofir, et al.
Publicado: (2023)
por: Nabati, Ofir, et al.
Publicado: (2023)
Statistical and Algorithmic Foundations of Reinforcement Learning
por: Chi, Yuejie, et al.
Publicado: (2025)
por: Chi, Yuejie, et al.
Publicado: (2025)
Value-Incentivized Preference Optimization: A Unified Approach to Online and Offline RLHF
por: Cen, Shicong, et al.
Publicado: (2024)
por: Cen, Shicong, et al.
Publicado: (2024)
Agentic Transformers Provably Learn to Search via Reinforcement Learning
por: Yang, Tong, et al.
Publicado: (2026)
por: Yang, Tong, et al.
Publicado: (2026)
Beyond Expectations: Learning with Stochastic Dominance Made Practical
por: Cen, Shicong, et al.
Publicado: (2024)
por: Cen, Shicong, et al.
Publicado: (2024)
From Noise to Control: Parameterized Diffusion Policies
por: Zhang, Renhao, et al.
Publicado: (2026)
por: Zhang, Renhao, et al.
Publicado: (2026)
Multi-head Transformers Provably Learn Symbolic Multi-step Reasoning via Gradient Descent
por: Yang, Tong, et al.
Publicado: (2025)
por: Yang, Tong, et al.
Publicado: (2025)
Prompt-prompted Adaptive Structured Pruning for Efficient LLM Generation
por: Dong, Harry, et al.
Publicado: (2024)
por: Dong, Harry, et al.
Publicado: (2024)
Accelerating Convergence of Score-Based Diffusion Models, Provably
por: Li, Gen, et al.
Publicado: (2024)
por: Li, Gen, et al.
Publicado: (2024)
Get More with LESS: Synthesizing Recurrence with KV Cache Compression for Efficient LLM Inference
por: Dong, Harry, et al.
Publicado: (2024)
por: Dong, Harry, et al.
Publicado: (2024)
Preconditioning Benefits of Spectral Orthogonalization in Muon
por: Ma, Jianhao, et al.
Publicado: (2026)
por: Ma, Jianhao, et al.
Publicado: (2026)
Scalable LLM Reasoning Acceleration with Low-rank Distillation
por: Dong, Harry, et al.
Publicado: (2025)
por: Dong, Harry, et al.
Publicado: (2025)
Interpretable Neural Networks with Random Constructive Algorithm
por: Nan, Jing, et al.
Publicado: (2023)
por: Nan, Jing, et al.
Publicado: (2023)
Learning Discrete Concepts in Latent Hierarchical Models
por: Kong, Lingjing, et al.
Publicado: (2024)
por: Kong, Lingjing, et al.
Publicado: (2024)
DARE: Diffusion Language Model Activation Reuse for Efficient Inference
por: Frumkin, Natalia, et al.
Publicado: (2026)
por: Frumkin, Natalia, et al.
Publicado: (2026)
Transformers Provably Learn Chain-of-Thought Reasoning with Length Generalization
por: Huang, Yu, et al.
Publicado: (2025)
por: Huang, Yu, et al.
Publicado: (2025)
On the Design of KL-Regularized Policy Gradient Algorithms for LLM Reasoning
por: Zhang, Yifan, et al.
Publicado: (2025)
por: Zhang, Yifan, et al.
Publicado: (2025)
Exploration from a Primal-Dual Lens: Value-Incentivized Actor-Critic Methods for Sample-Efficient Online RL
por: Yang, Tong, et al.
Publicado: (2025)
por: Yang, Tong, et al.
Publicado: (2025)
Enhanced DACER Algorithm with High Diffusion Efficiency
por: Wang, Yinuo, et al.
Publicado: (2025)
por: Wang, Yinuo, et al.
Publicado: (2025)
The Implicit Curriculum: Learning Dynamics in RL with Verifiable Rewards
por: Huang, Yu, et al.
Publicado: (2026)
por: Huang, Yu, et al.
Publicado: (2026)
Policy Optimization for Personalized Interventions in Behavioral Health
por: Baek, Jackie, et al.
Publicado: (2023)
por: Baek, Jackie, et al.
Publicado: (2023)
Diffusion Spectral Representation for Reinforcement Learning
por: Shribak, Dmitry, et al.
Publicado: (2024)
por: Shribak, Dmitry, et al.
Publicado: (2024)
Diff-MN: Diffusion Parameterized MoE-NCDE for Continuous Time Series Generation with Irregular Observations
por: Zhang, Xu, et al.
Publicado: (2026)
por: Zhang, Xu, et al.
Publicado: (2026)
Comparative Analysis of Parameterized Action Actor-Critic Reinforcement Learning Algorithms for Web Search Match Plan Generation
por: Bapoo, Ubayd, et al.
Publicado: (2025)
por: Bapoo, Ubayd, et al.
Publicado: (2025)
Parameterized Projected Bellman Operator
por: Vincent, Théo, et al.
Publicado: (2023)
por: Vincent, Théo, et al.
Publicado: (2023)
Predicting Large-scale Urban Network Dynamics with Energy-informed Graph Neural Diffusion
por: Nie, Tong, et al.
Publicado: (2025)
por: Nie, Tong, et al.
Publicado: (2025)
Generalized Parallel Scaling with Interdependent Generations
por: Dong, Harry, et al.
Publicado: (2025)
por: Dong, Harry, et al.
Publicado: (2025)
Predicting Decisions of AI Agents from Limited Interaction through Text-Tabular Modeling
por: Shapira, Eilam, et al.
Publicado: (2026)
por: Shapira, Eilam, et al.
Publicado: (2026)
ControlTraj: Controllable Trajectory Generation with Topology-Constrained Diffusion Model
por: Zhu, Yuanshao, et al.
Publicado: (2024)
por: Zhu, Yuanshao, et al.
Publicado: (2024)
Ejemplares similares
-
Preference Adaptive and Sequential Text-to-Image Generation
por: Nabati, Ofir, et al.
Publicado: (2024) -
Embedding-Aligned Language Models
por: Tennenholtz, Guy, et al.
Publicado: (2024) -
Synthetic Dialogue Generation for Interactive Conversational Elicitation & Recommendation (ICER)
por: Ryu, Moonkyung, et al.
Publicado: (2025) -
Controllable User Simulation
por: Tennenholtz, Guy, et al.
Publicado: (2026) -
Demystifying Embedding Spaces using Large Language Models
por: Tennenholtz, Guy, et al.
Publicado: (2023)