Joint Learning of Energy-based Models and their Partition Function
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sander, Michael E., Roulet, Vincent, Liu, Tianlin, Blondel, Mathieu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Loss Functions and Operators Generated by f-Divergences
von: Roulet, Vincent, et al.
Veröffentlicht: (2025)
von: Roulet, Vincent, et al.
Veröffentlicht: (2025)
Autoregressive Language Models are Secretly Energy-Based Models: Insights into the Lookahead Capabilities of Next-Token Prediction
von: Blondel, Mathieu, et al.
Veröffentlicht: (2025)
von: Blondel, Mathieu, et al.
Veröffentlicht: (2025)
The Elements of Differentiable Programming
von: Blondel, Mathieu, et al.
Veröffentlicht: (2024)
von: Blondel, Mathieu, et al.
Veröffentlicht: (2024)
Stepping on the Edge: Curvature Aware Learning Rate Tuners
von: Roulet, Vincent, et al.
Veröffentlicht: (2024)
von: Roulet, Vincent, et al.
Veröffentlicht: (2024)
Routers in Vision Mixture of Experts: An Empirical Study
von: Liu, Tianlin, et al.
Veröffentlicht: (2024)
von: Liu, Tianlin, et al.
Veröffentlicht: (2024)
How do Transformers perform In-Context Autoregressive Learning?
von: Sander, Michael E., et al.
Veröffentlicht: (2024)
von: Sander, Michael E., et al.
Veröffentlicht: (2024)
Differentiable Knapsack and Top-k Operators via Dynamic Programming
von: Vivier-Ardisson, Germain, et al.
Veröffentlicht: (2026)
von: Vivier-Ardisson, Germain, et al.
Veröffentlicht: (2026)
Learning with Local Search MCMC Layers
von: Vivier-Ardisson, Germain, et al.
Veröffentlicht: (2025)
von: Vivier-Ardisson, Germain, et al.
Veröffentlicht: (2025)
Per-example gradients: a new frontier for understanding and improving optimizers
von: Roulet, Vincent, et al.
Veröffentlicht: (2025)
von: Roulet, Vincent, et al.
Veröffentlicht: (2025)
Learning with Fitzpatrick Losses
von: Rakotomandimby, Seta, et al.
Veröffentlicht: (2024)
von: Rakotomandimby, Seta, et al.
Veröffentlicht: (2024)
Decoding-time Realignment of Language Models
von: Liu, Tianlin, et al.
Veröffentlicht: (2024)
von: Liu, Tianlin, et al.
Veröffentlicht: (2024)
Regularized Large Neighborhood Search
von: Vivier-Ardisson, Germain, et al.
Veröffentlicht: (2026)
von: Vivier-Ardisson, Germain, et al.
Veröffentlicht: (2026)
On the Interplay Between Stepsize Tuning and Progressive Sharpening
von: Roulet, Vincent, et al.
Veröffentlicht: (2023)
von: Roulet, Vincent, et al.
Veröffentlicht: (2023)
Clustering in Deep Stochastic Transformers
von: Fedorov, Lev, et al.
Veröffentlicht: (2026)
von: Fedorov, Lev, et al.
Veröffentlicht: (2026)
Input Resolution Downsizing as a Compression Technique for Vision Deep Learning Systems
von: Morlier, Jeremy, et al.
Veröffentlicht: (2025)
von: Morlier, Jeremy, et al.
Veröffentlicht: (2025)
On Teacher Hacking in Language Model Distillation
von: Tiapkin, Daniil, et al.
Veröffentlicht: (2025)
von: Tiapkin, Daniil, et al.
Veröffentlicht: (2025)
In-Context Function Learning in Large Language Models
von: Akata, Elif, et al.
Veröffentlicht: (2026)
von: Akata, Elif, et al.
Veröffentlicht: (2026)
How far away are truly hyperparameter-free learning algorithms?
von: Kasimbeg, Priya, et al.
Veröffentlicht: (2025)
von: Kasimbeg, Priya, et al.
Veröffentlicht: (2025)
Joint Optimization of Model Partitioning and Resource Allocation for Anti-Jamming Collaborative Inference Systems
von: Wu, Mengru, et al.
Veröffentlicht: (2026)
von: Wu, Mengru, et al.
Veröffentlicht: (2026)
Towards Understanding the Universality of Transformers for Next-Token Prediction
von: Sander, Michael E., et al.
Veröffentlicht: (2024)
von: Sander, Michael E., et al.
Veröffentlicht: (2024)
Deep Reinforcement Learning based Triggering Function for Early Classifiers of Time Series
von: Renault, Aurélien, et al.
Veröffentlicht: (2025)
von: Renault, Aurélien, et al.
Veröffentlicht: (2025)
Nonparametric Estimation of Joint Entropy via Partitioned Sample-Spacing
von: Ho, Jungwoo, et al.
Veröffentlicht: (2025)
von: Ho, Jungwoo, et al.
Veröffentlicht: (2025)
Learning Data-Driven Uncertainty Set Partitions for Robust and Adaptive Energy Forecasting with Missing Data
von: Stratigakos, Akylas, et al.
Veröffentlicht: (2025)
von: Stratigakos, Akylas, et al.
Veröffentlicht: (2025)
Partition Function Estimation under Bounded f-Divergence
von: Block, Adam, et al.
Veröffentlicht: (2026)
von: Block, Adam, et al.
Veröffentlicht: (2026)
In-Context Learning of Energy Functions
von: Schaeffer, Rylan, et al.
Veröffentlicht: (2024)
von: Schaeffer, Rylan, et al.
Veröffentlicht: (2024)
Hierarchical Learning-based Graph Partition for Large-scale Vehicle Routing Problems
von: Pan, Yuxin, et al.
Veröffentlicht: (2025)
von: Pan, Yuxin, et al.
Veröffentlicht: (2025)
Latent Imitator: Generating Natural Individual Discriminatory Instances for Black-Box Fairness Testing
von: Xiao, Yisong, et al.
Veröffentlicht: (2023)
von: Xiao, Yisong, et al.
Veröffentlicht: (2023)
Function Approximation for Reinforcement Learning Controller for Energy from Spread Waves
von: Sarkar, Soumyendu, et al.
Veröffentlicht: (2024)
von: Sarkar, Soumyendu, et al.
Veröffentlicht: (2024)
Energy Generative Modeling: A Lyapunov-based Energy Matching Perspective
von: Wang, Yixuan, et al.
Veröffentlicht: (2026)
von: Wang, Yixuan, et al.
Veröffentlicht: (2026)
Learning Joint Models of Prediction and Optimization
von: Kotary, James, et al.
Veröffentlicht: (2024)
von: Kotary, James, et al.
Veröffentlicht: (2024)
Active-Passive Federated Learning for Vertically Partitioned Multi-view Data
von: Liu, Jiyuan, et al.
Veröffentlicht: (2024)
von: Liu, Jiyuan, et al.
Veröffentlicht: (2024)
Joint Partitioning and Placement of Foundation Models for Real-Time Edge AI
von: Djuhera, Aladin, et al.
Veröffentlicht: (2025)
von: Djuhera, Aladin, et al.
Veröffentlicht: (2025)
A Scale-Adaptive Framework for Joint Spatiotemporal Super-Resolution with Diffusion Models
von: Defez, Max, et al.
Veröffentlicht: (2026)
von: Defez, Max, et al.
Veröffentlicht: (2026)
Potential Energy based Mixture Model for Noisy Label Learning
von: Wang, Zijia, et al.
Veröffentlicht: (2024)
von: Wang, Zijia, et al.
Veröffentlicht: (2024)
Preference-based Conditional Treatment Effects and Policy Learning
von: Parnas, Dovid, et al.
Veröffentlicht: (2026)
von: Parnas, Dovid, et al.
Veröffentlicht: (2026)
Learning Centre Partitions from Summaries
von: Debaly, Zinsou Max, et al.
Veröffentlicht: (2025)
von: Debaly, Zinsou Max, et al.
Veröffentlicht: (2025)
Joint Optimization of Energy Consumption and Completion Time in Federated Learning
von: Zhou, Xinyu, et al.
Veröffentlicht: (2022)
von: Zhou, Xinyu, et al.
Veröffentlicht: (2022)
Partition Generative Modeling: Masked Modeling Without Masks
von: Deschenaux, Justin, et al.
Veröffentlicht: (2025)
von: Deschenaux, Justin, et al.
Veröffentlicht: (2025)
Implicit Diffusion: Efficient Optimization through Stochastic Sampling
von: Marion, Pierre, et al.
Veröffentlicht: (2024)
von: Marion, Pierre, et al.
Veröffentlicht: (2024)
Quantile Reward Policy Optimization: Alignment with Pointwise Regression and Exact Partition Functions
von: Matrenok, Simon, et al.
Veröffentlicht: (2025)
von: Matrenok, Simon, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Loss Functions and Operators Generated by f-Divergences
von: Roulet, Vincent, et al.
Veröffentlicht: (2025) -
Autoregressive Language Models are Secretly Energy-Based Models: Insights into the Lookahead Capabilities of Next-Token Prediction
von: Blondel, Mathieu, et al.
Veröffentlicht: (2025) -
The Elements of Differentiable Programming
von: Blondel, Mathieu, et al.
Veröffentlicht: (2024) -
Stepping on the Edge: Curvature Aware Learning Rate Tuners
von: Roulet, Vincent, et al.
Veröffentlicht: (2024) -
Routers in Vision Mixture of Experts: An Empirical Study
von: Liu, Tianlin, et al.
Veröffentlicht: (2024)