Learning Parametric Distributions from Samples and Preferences
Fuente:
arXiv
Salvato in:
| Autori principali: | Jourdan, Marc, Yüce, Gizem, Flammarion, Nicolas |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Learning In-context n-grams with Transformers: Sub-n-grams Are Near-stationary Points
di: Varre, Aditya, et al.
Pubblicazione: (2025)
di: Varre, Aditya, et al.
Pubblicazione: (2025)
Transformers Learn Latent Mixture Models In-Context via Mirror Descent
di: D'Angelo, Francesco, et al.
Pubblicazione: (2026)
di: D'Angelo, Francesco, et al.
Pubblicazione: (2026)
Pareto Set Identification With Posterior Sampling
di: Kone, Cyrille, et al.
Pubblicazione: (2024)
di: Kone, Cyrille, et al.
Pubblicazione: (2024)
Early alignment in two-layer networks training is a two-edged sword
di: Boursier, Etienne, et al.
Pubblicazione: (2024)
di: Boursier, Etienne, et al.
Pubblicazione: (2024)
Simplicity bias and optimization threshold in two-layer ReLU networks
di: Boursier, Etienne, et al.
Pubblicazione: (2024)
di: Boursier, Etienne, et al.
Pubblicazione: (2024)
Penalising the biases in norm regularisation enforces sparsity
di: Boursier, Etienne, et al.
Pubblicazione: (2023)
di: Boursier, Etienne, et al.
Pubblicazione: (2023)
(How) Learning Rates Regulate Catastrophic Overtraining
di: Rofin, Mark, et al.
Pubblicazione: (2026)
di: Rofin, Mark, et al.
Pubblicazione: (2026)
Learning Algorithms in the Limit
di: Papazov, Hristo, et al.
Pubblicazione: (2025)
di: Papazov, Hristo, et al.
Pubblicazione: (2025)
Does Refusal Training in LLMs Generalize to the Past Tense?
di: Andriushchenko, Maksym, et al.
Pubblicazione: (2024)
di: Andriushchenko, Maksym, et al.
Pubblicazione: (2024)
Exact Learning of Arithmetic with Differentiable Agents
di: Papazov, Hristo, et al.
Pubblicazione: (2025)
di: Papazov, Hristo, et al.
Pubblicazione: (2025)
Optimal Best Arm Identification under Differential Privacy
di: Jourdan, Marc, et al.
Pubblicazione: (2025)
di: Jourdan, Marc, et al.
Pubblicazione: (2025)
On the Out-of-Distribution Generalization of Reasoning in Multimodal LLMs for Simple Visual Planning Tasks
di: Neuhaus, Yannic, et al.
Pubblicazione: (2026)
di: Neuhaus, Yannic, et al.
Pubblicazione: (2026)
Why Do We Need Weight Decay in Modern Deep Learning?
di: D'Angelo, Francesco, et al.
Pubblicazione: (2023)
di: D'Angelo, Francesco, et al.
Pubblicazione: (2023)
Incremental Learning of Sparse Attention Patterns in Transformers
di: Yüksel, Oğuz Kaan, et al.
Pubblicazione: (2026)
di: Yüksel, Oğuz Kaan, et al.
Pubblicazione: (2026)
Gradient flow dynamics of shallow ReLU networks for square loss and orthogonal inputs
di: Boursier, Etienne, et al.
Pubblicazione: (2022)
di: Boursier, Etienne, et al.
Pubblicazione: (2022)
First-order ANIL provably learns representations despite overparametrization
di: Yüksel, Oğuz Kaan, et al.
Pubblicazione: (2023)
di: Yüksel, Oğuz Kaan, et al.
Pubblicazione: (2023)
Leveraging Continuous Time to Understand Momentum When Training Diagonal Linear Networks
di: Papazov, Hristo, et al.
Pubblicazione: (2024)
di: Papazov, Hristo, et al.
Pubblicazione: (2024)
Gradient Flow Polarizes Softmax Outputs towards Low-Entropy Solutions
di: Varre, Aditya, et al.
Pubblicazione: (2026)
di: Varre, Aditya, et al.
Pubblicazione: (2026)
Selective Induction Heads: How Transformers Select Causal Structures In Context
di: D'Angelo, Francesco, et al.
Pubblicazione: (2025)
di: D'Angelo, Francesco, et al.
Pubblicazione: (2025)
Is In-Context Learning Sufficient for Instruction Following in LLMs?
di: Zhao, Hao, et al.
Pubblicazione: (2024)
di: Zhao, Hao, et al.
Pubblicazione: (2024)
Jailbreaking Leading Safety-Aligned LLMs with Simple Adaptive Attacks
di: Andriushchenko, Maksym, et al.
Pubblicazione: (2024)
di: Andriushchenko, Maksym, et al.
Pubblicazione: (2024)
Implicit Bias of Mirror Flow on Separable Data
di: Pesme, Scott, et al.
Pubblicazione: (2024)
di: Pesme, Scott, et al.
Pubblicazione: (2024)
Long-Context Linear System Identification
di: Yüksel, Oğuz Kaan, et al.
Pubblicazione: (2024)
di: Yüksel, Oğuz Kaan, et al.
Pubblicazione: (2024)
An Anytime Algorithm for Good Arm Identification
di: Jourdan, Marc, et al.
Pubblicazione: (2023)
di: Jourdan, Marc, et al.
Pubblicazione: (2023)
Contextual Preference Distribution Learning
di: Hudson, Benjamin, et al.
Pubblicazione: (2026)
di: Hudson, Benjamin, et al.
Pubblicazione: (2026)
FuseLIP: Multimodal Embeddings via Early Fusion of Discrete Tokens
di: Schlarmann, Christian, et al.
Pubblicazione: (2025)
di: Schlarmann, Christian, et al.
Pubblicazione: (2025)
SIEVE: Sample-Efficient Parametric Learning from Natural Language
di: Asawa, Parth, et al.
Pubblicazione: (2026)
di: Asawa, Parth, et al.
Pubblicazione: (2026)
Best-Arm Identification in Unimodal Bandits
di: Poiani, Riccardo, et al.
Pubblicazione: (2024)
di: Poiani, Riccardo, et al.
Pubblicazione: (2024)
Speculative Sampling for Parametric Temporal Point Processes
di: Biloš, Marin, et al.
Pubblicazione: (2025)
di: Biloš, Marin, et al.
Pubblicazione: (2025)
Clear Preferences Leave Traces: Reference Model-Guided Sampling for Preference Learning
di: Diwan, Nirav, et al.
Pubblicazione: (2025)
di: Diwan, Nirav, et al.
Pubblicazione: (2025)
Leveraging Sparsity for Sample-Efficient Preference Learning: A Theoretical Perspective
di: Yao, Yunzhen, et al.
Pubblicazione: (2025)
di: Yao, Yunzhen, et al.
Pubblicazione: (2025)
Finite Sample Bounds for Non-Parametric Regression: Optimal Sample Efficiency and Space Complexity
di: Maran, Davide, et al.
Pubblicazione: (2024)
di: Maran, Davide, et al.
Pubblicazione: (2024)
Privacy Assessment of Federated Learning using Private Personalized Layers
di: Jourdan, Théo, et al.
Pubblicazione: (2021)
di: Jourdan, Théo, et al.
Pubblicazione: (2021)
HypeMARL: Multi-Agent Reinforcement Learning For High-Dimensional, Parametric, and Distributed Systems
di: Botteghi, Nicolò, et al.
Pubblicazione: (2025)
di: Botteghi, Nicolò, et al.
Pubblicazione: (2025)
Distributed Direct Preference Optimization
di: Jiang, Zhanhong
Pubblicazione: (2026)
di: Jiang, Zhanhong
Pubblicazione: (2026)
Graph Neural Networks with a Distribution of Parametrized Graphs
di: Lee, See Hian, et al.
Pubblicazione: (2023)
di: Lee, See Hian, et al.
Pubblicazione: (2023)
Preference as Reward, Maximum Preference Optimization with Importance Sampling
di: Jiang, Zaifan, et al.
Pubblicazione: (2023)
di: Jiang, Zaifan, et al.
Pubblicazione: (2023)
When Can Proxies Improve the Sample Complexity of Preference Learning?
di: Zhu, Yuchen, et al.
Pubblicazione: (2024)
di: Zhu, Yuchen, et al.
Pubblicazione: (2024)
Towards Unified Benchmark and Models for Multi-Modal Perceptual Metrics
di: Ghazanfari, Sara, et al.
Pubblicazione: (2024)
di: Ghazanfari, Sara, et al.
Pubblicazione: (2024)
Differentially Private Best-Arm Identification
di: Azize, Achraf, et al.
Pubblicazione: (2024)
di: Azize, Achraf, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Learning In-context n-grams with Transformers: Sub-n-grams Are Near-stationary Points
di: Varre, Aditya, et al.
Pubblicazione: (2025) -
Transformers Learn Latent Mixture Models In-Context via Mirror Descent
di: D'Angelo, Francesco, et al.
Pubblicazione: (2026) -
Pareto Set Identification With Posterior Sampling
di: Kone, Cyrille, et al.
Pubblicazione: (2024) -
Early alignment in two-layer networks training is a two-edged sword
di: Boursier, Etienne, et al.
Pubblicazione: (2024) -
Simplicity bias and optimization threshold in two-layer ReLU networks
di: Boursier, Etienne, et al.
Pubblicazione: (2024)