Discrete Compositional Generation via General Soft Operators and Robust Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Jiralerspong, Marco, Derman, Esther, Vucetic, Danilo, Malkin, Nikolay, Sun, Bilun, Zhang, Tianyu, Bacon, Pierre-Luc, Gidel, Gauthier |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Expected flow networks in stochastic environments and two-player zero-sum games
by: Jiralerspong, Marco, et al.
Published: (2023)
by: Jiralerspong, Marco, et al.
Published: (2023)
General Causal Imputation via Synthetic Interventions
by: Jiralerspong, Marco, et al.
Published: (2024)
by: Jiralerspong, Marco, et al.
Published: (2024)
State Entropy Regularization for Robust Reinforcement Learning
by: Ashlag, Yonatan, et al.
Published: (2025)
by: Ashlag, Yonatan, et al.
Published: (2025)
On the Stability of Iterative Retraining of Generative Models on their own Data
by: Bertrand, Quentin, et al.
Published: (2023)
by: Bertrand, Quentin, et al.
Published: (2023)
Feature Likelihood Divergence: Evaluating the Generalization of Generative Models Using Samples
by: Jiralerspong, Marco, et al.
Published: (2023)
by: Jiralerspong, Marco, et al.
Published: (2023)
Reward Redistribution for CVaR MDPs using a Bellman Operator on L-infinity
by: Muni, Aneri, et al.
Published: (2026)
by: Muni, Aneri, et al.
Published: (2026)
Long-Horizon Model-Based Offline Reinforcement Learning Without Explicit Conservatism
by: Ni, Tianwei, et al.
Published: (2025)
by: Ni, Tianwei, et al.
Published: (2025)
Solving Hidden Monotone Variational Inequalities with Surrogate Losses
by: D'Orazio, Ryan, et al.
Published: (2024)
by: D'Orazio, Ryan, et al.
Published: (2024)
On Designing Diffusion Autoencoders for Efficient Generation and Representation Learning
by: Proszewska, Magdalena, et al.
Published: (2025)
by: Proszewska, Magdalena, et al.
Published: (2025)
The Three Regimes of Offline-to-Online Reinforcement Learning
by: Li, Lu, et al.
Published: (2025)
by: Li, Lu, et al.
Published: (2025)
Soft Robotic Devices for Mechanotherapy of the Upper and Lower Extremities
by: Trivoramai Jiralerspong, et al.
Published: (2024)
by: Trivoramai Jiralerspong, et al.
Published: (2024)
On Generalization for Generative Flow Networks
by: Krichel, Anas, et al.
Published: (2024)
by: Krichel, Anas, et al.
Published: (2024)
Self-Consuming Generative Models with Curated Data Provably Optimize Human Preferences
by: Ferbach, Damien, et al.
Published: (2024)
by: Ferbach, Damien, et al.
Published: (2024)
Soft Prompt Threats: Attacking Safety Alignment and Unlearning in Open-Source LLMs through the Embedding Space
by: Schwinn, Leo, et al.
Published: (2024)
by: Schwinn, Leo, et al.
Published: (2024)
Activation Steering for Aligned Open-ended Generation without Sacrificing Coherence
by: Herbster, Niklas, et al.
Published: (2026)
by: Herbster, Niklas, et al.
Published: (2026)
A Generative Approach to LLM Harmfulness Mitigation with Red Flag Tokens
by: Dobre, David, et al.
Published: (2025)
by: Dobre, David, et al.
Published: (2025)
Why Open Source? A Game-Theoretic Analysis of the AI Race
by: Mladenovic, Andjela, et al.
Published: (2026)
by: Mladenovic, Andjela, et al.
Published: (2026)
Performative Prediction with Neural Networks
by: Mofakhami, Mehrnaz, et al.
Published: (2023)
by: Mofakhami, Mehrnaz, et al.
Published: (2023)
Sarah Frank-Wolfe: Methods for Constrained Optimization with Best Rates and Practical Features
by: Beznosikov, Aleksandr, et al.
Published: (2023)
by: Beznosikov, Aleksandr, et al.
Published: (2023)
Delta-AI: Local objectives for amortized inference in sparse graphical models
by: Falet, Jean-Pierre, et al.
Published: (2023)
by: Falet, Jean-Pierre, et al.
Published: (2023)
Mol-MoE: Training Preference-Guided Routers for Molecule Generation
by: Calanzone, Diego, et al.
Published: (2025)
by: Calanzone, Diego, et al.
Published: (2025)
Proving Linear Mode Connectivity of Neural Networks via Optimal Transport
by: Ferbach, Damien, et al.
Published: (2023)
by: Ferbach, Damien, et al.
Published: (2023)
Learning diverse attacks on large language models for robust red-teaming and safety tuning
by: Lee, Seanie, et al.
Published: (2024)
by: Lee, Seanie, et al.
Published: (2024)
Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments
by: Luo, Ziyan, et al.
Published: (2025)
by: Luo, Ziyan, et al.
Published: (2025)
Machine learning and information theory concepts towards an AI Mathematician
by: Bengio, Yoshua, et al.
Published: (2024)
by: Bengio, Yoshua, et al.
Published: (2024)
Data-to-Energy Stochastic Dynamics
by: Tamogashev, Kirill, et al.
Published: (2025)
by: Tamogashev, Kirill, et al.
Published: (2025)
Discrete Probabilistic Inference as Control in Multi-path Environments
by: Deleu, Tristan, et al.
Published: (2024)
by: Deleu, Tristan, et al.
Published: (2024)
LLM-Safety Evaluations Lack Robustness
by: Beyer, Tim, et al.
Published: (2025)
by: Beyer, Tim, et al.
Published: (2025)
What Makes Value Learning Efficient in Residual Reinforcement Learning?
by: Ma, Guozheng, et al.
Published: (2026)
by: Ma, Guozheng, et al.
Published: (2026)
In-Context Learning Can Re-learn Forbidden Tasks
by: Xhonneux, Sophie, et al.
Published: (2024)
by: Xhonneux, Sophie, et al.
Published: (2024)
Network Sparsity Unlocks the Scaling Potential of Deep Reinforcement Learning
by: Ma, Guozheng, et al.
Published: (2025)
by: Ma, Guozheng, et al.
Published: (2025)
A Persuasive Approach to Combating Misinformation
by: Hossain, Safwan, et al.
Published: (2023)
by: Hossain, Safwan, et al.
Published: (2023)
Omega: Optimistic EMA Gradients
by: Ramirez, Juan, et al.
Published: (2023)
by: Ramirez, Juan, et al.
Published: (2023)
Self-Play Q-learners Can Provably Collude in the Iterated Prisoner's Dilemma
by: Bertrand, Quentin, et al.
Published: (2023)
by: Bertrand, Quentin, et al.
Published: (2023)
A Coin Flip for Safety: LLM Judges Fail to Reliably Measure Adversarial Robustness
by: Schwinn, Leo, et al.
Published: (2026)
by: Schwinn, Leo, et al.
Published: (2026)
Solving Non-Rectangular Reward-Robust MDPs via Frequency Regularization
by: Gadot, Uri, et al.
Published: (2023)
by: Gadot, Uri, et al.
Published: (2023)
A Complexity-Based Theory of Compositionality
by: Elmoznino, Eric, et al.
Published: (2024)
by: Elmoznino, Eric, et al.
Published: (2024)
Discrete, compositional, and symbolic representations through attractor dynamics
by: Nam, Andrew, et al.
Published: (2023)
by: Nam, Andrew, et al.
Published: (2023)
When is Momentum Extragradient Optimal? A Polynomial-Based Analysis
by: Kim, Junhyung Lyle, et al.
Published: (2022)
by: Kim, Junhyung Lyle, et al.
Published: (2022)
Plumbings of lens spaces and crepant resolutions of compound $A_n$ singularities
by: Xie, Bilun, et al.
Published: (2025)
by: Xie, Bilun, et al.
Published: (2025)
Similar Items
-
Expected flow networks in stochastic environments and two-player zero-sum games
by: Jiralerspong, Marco, et al.
Published: (2023) -
General Causal Imputation via Synthetic Interventions
by: Jiralerspong, Marco, et al.
Published: (2024) -
State Entropy Regularization for Robust Reinforcement Learning
by: Ashlag, Yonatan, et al.
Published: (2025) -
On the Stability of Iterative Retraining of Generative Models on their own Data
by: Bertrand, Quentin, et al.
Published: (2023) -
Feature Likelihood Divergence: Evaluating the Generalization of Generative Models Using Samples
by: Jiralerspong, Marco, et al.
Published: (2023)