Guided Star-Shaped Masked Diffusion
Fuente:
arXiv
Saved in:
| Main Authors: | Meshchaninov, Viacheslav, Shibaev, Egor, Makoian, Artem, Klimov, Ivan, Balagansky, Nikita, Gavrilov, Daniil, Alanov, Aibek, Vetrov, Dmitry |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Mechanistic Permutability: Match Features Across Layers
by: Balagansky, Nikita, et al.
Published: (2024)
by: Balagansky, Nikita, et al.
Published: (2024)
Steering LLM Reasoning Through Bias-Only Adaptation
by: Sinii, Viacheslav, et al.
Published: (2025)
by: Sinii, Viacheslav, et al.
Published: (2025)
You Do Not Fully Utilize Transformer's Representation Capacity
by: Gerasimov, Gleb, et al.
Published: (2025)
by: Gerasimov, Gleb, et al.
Published: (2025)
Analyze Feature Flow to Enhance Interpretation and Steering in Language Models
by: Laptev, Daniil, et al.
Published: (2025)
by: Laptev, Daniil, et al.
Published: (2025)
Diffusion Language Models Generation Can Be Halted Early
by: Vaina, Sofia Maria Lo Cicero, et al.
Published: (2023)
by: Vaina, Sofia Maria Lo Cicero, et al.
Published: (2023)
Diffusion on language model encodings for protein sequence generation
by: Meshchaninov, Viacheslav, et al.
Published: (2024)
by: Meshchaninov, Viacheslav, et al.
Published: (2024)
Next Embedding Prediction Makes World Models Stronger
by: Bredis, George, et al.
Published: (2026)
by: Bredis, George, et al.
Published: (2026)
Kronecker Factorization Improves Efficiency and Interpretability of Sparse Autoencoders
by: Kurochkin, Vadim, et al.
Published: (2025)
by: Kurochkin, Vadim, et al.
Published: (2025)
Teach Old SAEs New Domain Tricks with Boosting
by: Koriagin, Nikita, et al.
Published: (2025)
by: Koriagin, Nikita, et al.
Published: (2025)
Smoothie: Smoothing Diffusion on Token Embeddings for Text Generation
by: Shabalin, Alexander, et al.
Published: (2025)
by: Shabalin, Alexander, et al.
Published: (2025)
Small Vectors, Big Effects: A Mechanistic Study of RL-Induced Reasoning via Steering Vectors
by: Sinii, Viacheslav, et al.
Published: (2025)
by: Sinii, Viacheslav, et al.
Published: (2025)
Train One Sparse Autoencoder Across Multiple Sparsity Budgets to Preserve Interpretability and Accuracy
by: Balagansky, Nikita, et al.
Published: (2025)
by: Balagansky, Nikita, et al.
Published: (2025)
Cosmos: Compressed and Smooth Latent Space for Text Diffusion Modeling
by: Meshchaninov, Viacheslav, et al.
Published: (2025)
by: Meshchaninov, Viacheslav, et al.
Published: (2025)
SGD as Free Energy Minimization: A Thermodynamic View on Neural Network Training
by: Sadrtdinov, Ildus, et al.
Published: (2025)
by: Sadrtdinov, Ildus, et al.
Published: (2025)
Trust-Region Behavior Blending for On-Policy Distillation
by: Plyusov, Daniil, et al.
Published: (2026)
by: Plyusov, Daniil, et al.
Published: (2026)
Generative Flow Networks as Entropy-Regularized RL
by: Tiapkin, Daniil, et al.
Published: (2023)
by: Tiapkin, Daniil, et al.
Published: (2023)
How to Train Your Latent Diffusion Language Model Jointly With the Latent Space
by: Meshchaninov, Viacheslav, et al.
Published: (2026)
by: Meshchaninov, Viacheslav, et al.
Published: (2026)
Learn Your Reference Model for Real Good Alignment
by: Gorbatovski, Alexey, et al.
Published: (2024)
by: Gorbatovski, Alexey, et al.
Published: (2024)
The Devil is in the Details: StyleFeatureEditor for Detail-Rich StyleGAN Inversion and High Quality Image Editing
by: Bobkov, Denis, et al.
Published: (2024)
by: Bobkov, Denis, et al.
Published: (2024)
HairFastGAN: Realistic and Robust Hair Transfer with a Fast Encoder-Based Approach
by: Nikolaev, Maxim, et al.
Published: (2024)
by: Nikolaev, Maxim, et al.
Published: (2024)
VARAN: Variational Inference for Self-Supervised Speech Models Fine-Tuning on Downstream Tasks
by: Diatlova, Daria, et al.
Published: (2025)
by: Diatlova, Daria, et al.
Published: (2025)
Guide-and-Rescale: Self-Guidance Mechanism for Effective Tuning-Free Real Image Editing
by: Titov, Vadim, et al.
Published: (2024)
by: Titov, Vadim, et al.
Published: (2024)
Linear Transformers with Learnable Kernel Functions are Better In-Context Models
by: Aksenov, Yaroslav, et al.
Published: (2024)
by: Aksenov, Yaroslav, et al.
Published: (2024)
Improving GFlowNets with Monte Carlo Tree Search
by: Morozov, Nikita, et al.
Published: (2024)
by: Morozov, Nikita, et al.
Published: (2024)
Adaptive Destruction Processes for Diffusion Samplers
by: Gritsaev, Timofei, et al.
Published: (2025)
by: Gritsaev, Timofei, et al.
Published: (2025)
Why Gaussian Diffusion Models Fail on Discrete Data and How to Prevent It?
by: Shabalin, Alexander, et al.
Published: (2026)
by: Shabalin, Alexander, et al.
Published: (2026)
Can Stationary Distributions of Scale-Invariant Neural Networks Be Described by the Thermodynamics of an Ideal Gas?
by: Sadrtdinov, Ildus, et al.
Published: (2025)
by: Sadrtdinov, Ildus, et al.
Published: (2025)
The Differences Between Direct Alignment Algorithms are a Blur
by: Gorbatovski, Alexey, et al.
Published: (2025)
by: Gorbatovski, Alexey, et al.
Published: (2025)
Enhancing Vision-Language Model Training with Reinforcement Learning in Synthetic Worlds for Real-World Success
by: Bredis, George, et al.
Published: (2025)
by: Bredis, George, et al.
Published: (2025)
Neural Diffusion Models
by: Bartosh, Grigory, et al.
Published: (2023)
by: Bartosh, Grigory, et al.
Published: (2023)
F-GRPO: Don't Let Your Policy Learn the Obvious and Forget the Rare
by: Plyusov, Daniil, et al.
Published: (2026)
by: Plyusov, Daniil, et al.
Published: (2026)
OrthoFuse: Training-free Riemannian Fusion of Orthogonal Style-Concept Adapters for Diffusion Models
by: Aliev, Ali, et al.
Published: (2026)
by: Aliev, Ali, et al.
Published: (2026)
TEncDM: Understanding the Properties of the Diffusion Model in the Space of Language Model Encodings
by: Shabalin, Alexander, et al.
Published: (2024)
by: Shabalin, Alexander, et al.
Published: (2024)
Homeostasis and Sparsity in Transformer
by: Kotyuzanskiy, Leonid, et al.
Published: (2024)
by: Kotyuzanskiy, Leonid, et al.
Published: (2024)
Regularized Distribution Matching Distillation for One-step Unpaired Image-to-Image Translation
by: Rakitin, Denis, et al.
Published: (2024)
by: Rakitin, Denis, et al.
Published: (2024)
Stack Trace Deduplication: Faster, More Accurately, and in More Realistic Scenarios
by: Shibaev, Egor, et al.
Published: (2024)
by: Shibaev, Egor, et al.
Published: (2024)
Neural Flow Diffusion Models: Learnable Forward Process for Improved Diffusion Modelling
by: Bartosh, Grigory, et al.
Published: (2024)
by: Bartosh, Grigory, et al.
Published: (2024)
ESSA: Evolutionary Strategies for Scalable Alignment
by: Korotyshova, Daria, et al.
Published: (2025)
by: Korotyshova, Daria, et al.
Published: (2025)
Streaming Generation of Co-Speech Gestures via Accelerated Rolling Diffusion
by: Vu, Evgeniia, et al.
Published: (2025)
by: Vu, Evgeniia, et al.
Published: (2025)
TabDDPM: Modelling Tabular Data with Diffusion Models
by: Kotelnikov, Akim, et al.
Published: (2022)
by: Kotelnikov, Akim, et al.
Published: (2022)
Similar Items
-
Mechanistic Permutability: Match Features Across Layers
by: Balagansky, Nikita, et al.
Published: (2024) -
Steering LLM Reasoning Through Bias-Only Adaptation
by: Sinii, Viacheslav, et al.
Published: (2025) -
You Do Not Fully Utilize Transformer's Representation Capacity
by: Gerasimov, Gleb, et al.
Published: (2025) -
Analyze Feature Flow to Enhance Interpretation and Steering in Language Models
by: Laptev, Daniil, et al.
Published: (2025) -
Diffusion Language Models Generation Can Be Halted Early
by: Vaina, Sofia Maria Lo Cicero, et al.
Published: (2023)