Sampling Data with Chains of Forward-Backward Diffusion Steps
Fuente:
arXiv
Salvato in:
| Autori principali: | Kang, Hyunmo, Levi, Noam Itzhak, Wegner, Corinna Elena, Korchinski, Daniel J., Wyart, Matthieu |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Probing the Latent Hierarchical Structure of Data via Diffusion Models
di: Sclocchi, Antonio, et al.
Pubblicazione: (2024)
di: Sclocchi, Antonio, et al.
Pubblicazione: (2024)
Learning curves theory for hierarchically compositional data with power-law distributed features
di: Cagnetta, Francesco, et al.
Pubblicazione: (2025)
di: Cagnetta, Francesco, et al.
Pubblicazione: (2025)
On the Emergence of Linear Analogies in Word Embeddings
di: Korchinski, Daniel J., et al.
Pubblicazione: (2025)
di: Korchinski, Daniel J., et al.
Pubblicazione: (2025)
Symmetry in language statistics shapes the geometry of model representations
di: Karkada, Dhruva, et al.
Pubblicazione: (2026)
di: Karkada, Dhruva, et al.
Pubblicazione: (2026)
How Deep Networks Learn Sparse and Hierarchical Data: the Sparse Random Hierarchy Model
di: Tomasini, Umberto, et al.
Pubblicazione: (2024)
di: Tomasini, Umberto, et al.
Pubblicazione: (2024)
On the different regimes of Stochastic Gradient Descent
di: Sclocchi, Antonio, et al.
Pubblicazione: (2023)
di: Sclocchi, Antonio, et al.
Pubblicazione: (2023)
Towards a theory of how the structure of language is acquired by deep neural networks
di: Cagnetta, Francesco, et al.
Pubblicazione: (2024)
di: Cagnetta, Francesco, et al.
Pubblicazione: (2024)
Microscopic description of the intermittent dynamics driving logarithmic creep
di: Korchinski, Daniel J., et al.
Pubblicazione: (2024)
di: Korchinski, Daniel J., et al.
Pubblicazione: (2024)
Scaling Laws and Representation Learning in Simple Hierarchical Languages: Transformers vs. Convolutional Architectures
di: Cagnetta, Francesco, et al.
Pubblicazione: (2025)
di: Cagnetta, Francesco, et al.
Pubblicazione: (2025)
A Phase Transition in Diffusion Models Reveals the Hierarchical Nature of Data
di: Sclocchi, Antonio, et al.
Pubblicazione: (2024)
di: Sclocchi, Antonio, et al.
Pubblicazione: (2024)
Deep networks learn to parse uniform-depth context-free languages from local statistics
di: Parley, Jack T., et al.
Pubblicazione: (2026)
di: Parley, Jack T., et al.
Pubblicazione: (2026)
Learning Shrinks the Hard Tail: Training-Dependent Inference Scaling in a Solvable Linear Model
di: Levi, Noam
Pubblicazione: (2026)
di: Levi, Noam
Pubblicazione: (2026)
Spectral Analysis of Representational Similarity with Limited Neurons
di: Kang, Hyunmo, et al.
Pubblicazione: (2025)
di: Kang, Hyunmo, et al.
Pubblicazione: (2025)
Grokking at the Edge of Linear Separability
di: Beck, Alon, et al.
Pubblicazione: (2024)
di: Beck, Alon, et al.
Pubblicazione: (2024)
Grokking in Linear Estimators -- A Solvable Model that Groks without Understanding
di: Levi, Noam, et al.
Pubblicazione: (2023)
di: Levi, Noam, et al.
Pubblicazione: (2023)
More Bang for the Buck: Improving the Inference of Large Language Models at a Fixed Budget using Reset and Discard (ReD)
di: Meir, Sagi, et al.
Pubblicazione: (2026)
di: Meir, Sagi, et al.
Pubblicazione: (2026)
The Underlying Scaling Laws and Universal Statistical Structure of Complex Datasets
di: Levi, Noam, et al.
Pubblicazione: (2023)
di: Levi, Noam, et al.
Pubblicazione: (2023)
A Generative Diffusion Model for Amorphous Materials
di: Yang, Kai, et al.
Pubblicazione: (2025)
di: Yang, Kai, et al.
Pubblicazione: (2025)
The Interplay of Data Structure and Imbalance in the Learning Dynamics of Diffusion Models
di: Nicoletti, Flavio, et al.
Pubblicazione: (2026)
di: Nicoletti, Flavio, et al.
Pubblicazione: (2026)
Convergence Acceleration of Markov Chain Monte Carlo-based Gradient Descent by Deep Unfolding
di: Hagiwara, Ryo, et al.
Pubblicazione: (2024)
di: Hagiwara, Ryo, et al.
Pubblicazione: (2024)
Diffusion Operator Geometry of Feedforward Representations
di: Reddy, Kanishka
Pubblicazione: (2026)
di: Reddy, Kanishka
Pubblicazione: (2026)
Dynamical Regimes of Multimodal Diffusion Models
di: Albrychiewicz, Emil, et al.
Pubblicazione: (2026)
di: Albrychiewicz, Emil, et al.
Pubblicazione: (2026)
Controlled Langevin Dynamics for Sampling of Feedforward Neural Networks Trained with Minibatches
di: Zambon, Alessandro, et al.
Pubblicazione: (2026)
di: Zambon, Alessandro, et al.
Pubblicazione: (2026)
EB-RANSAC: Random Sample Consensus based on Energy-Based Model
di: Yasuda, Muneki, et al.
Pubblicazione: (2026)
di: Yasuda, Muneki, et al.
Pubblicazione: (2026)
Emergence of Distortions in High-Dimensional Guided Diffusion Models
di: Ventura, Enrico, et al.
Pubblicazione: (2026)
di: Ventura, Enrico, et al.
Pubblicazione: (2026)
Short-range depinning in the presence of velocity-weakening
di: de Geus, Tom W. J., et al.
Pubblicazione: (2024)
di: de Geus, Tom W. J., et al.
Pubblicazione: (2024)
Dynamical heterogeneities of thermal creep in pinned interfaces
di: de Geus, Tom W. J., et al.
Pubblicazione: (2024)
di: de Geus, Tom W. J., et al.
Pubblicazione: (2024)
Theory of Speciation Transitions in Diffusion Models with General Class Structure
di: Achilli, Beatrice, et al.
Pubblicazione: (2026)
di: Achilli, Beatrice, et al.
Pubblicazione: (2026)
Why Diffusion Models Don't Memorize: The Role of Implicit Dynamical Regularization in Training
di: Bonnaire, Tony, et al.
Pubblicazione: (2025)
di: Bonnaire, Tony, et al.
Pubblicazione: (2025)
Regularization, early-stopping and dreaming: a Hopfield-like setup to address generalization and overfitting
di: Agliari, Elena, et al.
Pubblicazione: (2023)
di: Agliari, Elena, et al.
Pubblicazione: (2023)
Neural Scaling Laws Rooted in the Data Distribution
di: Brill, Ari
Pubblicazione: (2024)
di: Brill, Ari
Pubblicazione: (2024)
Demystifying Spectral Bias on Real-World Data
di: Lavie, Itay, et al.
Pubblicazione: (2024)
di: Lavie, Itay, et al.
Pubblicazione: (2024)
Modeling Structured Data Learning with Restricted Boltzmann Machines in the Teacher-Student Setting
di: Thériault, Robin, et al.
Pubblicazione: (2024)
di: Thériault, Robin, et al.
Pubblicazione: (2024)
Deep Linear Network Training Dynamics from Random Initialization: Data, Width, Depth, and Hyperparameter Transfer
di: Bordelon, Blake, et al.
Pubblicazione: (2025)
di: Bordelon, Blake, et al.
Pubblicazione: (2025)
Soft Quantization: Model Compression Via Weight Coupling
di: Bernstein, Daniel T., et al.
Pubblicazione: (2026)
di: Bernstein, Daniel T., et al.
Pubblicazione: (2026)
Stochastic Interpolants: A Unifying Framework for Flows and Diffusions
di: Albergo, Michael S., et al.
Pubblicazione: (2023)
di: Albergo, Michael S., et al.
Pubblicazione: (2023)
Sampling at intermediate temperatures is optimal for training large language models in protein structure prediction
di: Ghiringhelli, L., et al.
Pubblicazione: (2026)
di: Ghiringhelli, L., et al.
Pubblicazione: (2026)
Supervised Hebbian Learning
di: Alemanno, Francesco, et al.
Pubblicazione: (2022)
di: Alemanno, Francesco, et al.
Pubblicazione: (2022)
How does Chain of Thought decompose complex tasks?
di: Nadgir, Amrut, et al.
Pubblicazione: (2026)
di: Nadgir, Amrut, et al.
Pubblicazione: (2026)
Generative Inversion of Spectroscopic Data for Amorphous Structure Elucidation
di: Guo, Jiawei, et al.
Pubblicazione: (2026)
di: Guo, Jiawei, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Probing the Latent Hierarchical Structure of Data via Diffusion Models
di: Sclocchi, Antonio, et al.
Pubblicazione: (2024) -
Learning curves theory for hierarchically compositional data with power-law distributed features
di: Cagnetta, Francesco, et al.
Pubblicazione: (2025) -
On the Emergence of Linear Analogies in Word Embeddings
di: Korchinski, Daniel J., et al.
Pubblicazione: (2025) -
Symmetry in language statistics shapes the geometry of model representations
di: Karkada, Dhruva, et al.
Pubblicazione: (2026) -
How Deep Networks Learn Sparse and Hierarchical Data: the Sparse Random Hierarchy Model
di: Tomasini, Umberto, et al.
Pubblicazione: (2024)