Generalized Interpolating Discrete Diffusion
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | von Rütte, Dimitri, Fluri, Janis, Ding, Yuhui, Orvieto, Antonio, Schölkopf, Bernhard, Hofmann, Thomas |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Scaling Behavior of Discrete Diffusion Language Models
von: von Rütte, Dimitri, et al.
Veröffentlicht: (2025)
von: von Rütte, Dimitri, et al.
Veröffentlicht: (2025)
Identifying Intervenable and Interpretable Features via Orthogonality Regularization
von: Miller, Moritz, et al.
Veröffentlicht: (2026)
von: Miller, Moritz, et al.
Veröffentlicht: (2026)
A Language Model's Guide Through Latent Space
von: von Rütte, Dimitri, et al.
Veröffentlicht: (2024)
von: von Rütte, Dimitri, et al.
Veröffentlicht: (2024)
Improving Large Language Model Safety with Contrastive Representation Learning
von: Simko, Samuel, et al.
Veröffentlicht: (2025)
von: Simko, Samuel, et al.
Veröffentlicht: (2025)
Deriving Hyperparameter Scaling Laws via Modern Optimization Theory
von: Shulgin, Egor, et al.
Veröffentlicht: (2026)
von: Shulgin, Egor, et al.
Veröffentlicht: (2026)
Generalized Discrete Diffusion from Snapshots
von: Zekri, Oussama, et al.
Veröffentlicht: (2026)
von: Zekri, Oussama, et al.
Veröffentlicht: (2026)
Counterfactual reasoning: an analysis of in-context emergence
von: Miller, Moritz, et al.
Veröffentlicht: (2025)
von: Miller, Moritz, et al.
Veröffentlicht: (2025)
Planner and Executor: Collaboration between Discrete Diffusion And Autoregressive Models in Reasoning
von: Berrayana, Lina, et al.
Veröffentlicht: (2025)
von: Berrayana, Lina, et al.
Veröffentlicht: (2025)
MathGAP: Out-of-Distribution Evaluation on Problems with Arbitrarily Complex Proofs
von: Opedal, Andreas, et al.
Veröffentlicht: (2024)
von: Opedal, Andreas, et al.
Veröffentlicht: (2024)
Limits of Transformer Language Models on Learning to Compose Algorithms
von: Thomm, Jonathan, et al.
Veröffentlicht: (2024)
von: Thomm, Jonathan, et al.
Veröffentlicht: (2024)
Ultra-Fast Language Generation via Discrete Diffusion Divergence Instruct
von: Zheng, Haoyang, et al.
Veröffentlicht: (2025)
von: Zheng, Haoyang, et al.
Veröffentlicht: (2025)
Reparameterized LLM Training via Orthogonal Equivalence Transformation
von: Qiu, Zeju, et al.
Veröffentlicht: (2025)
von: Qiu, Zeju, et al.
Veröffentlicht: (2025)
Learning Beyond Pattern Matching? Assaying Mathematical Understanding in LLMs
von: Guo, Siyuan, et al.
Veröffentlicht: (2024)
von: Guo, Siyuan, et al.
Veröffentlicht: (2024)
Learning to Reason Efficiently with A* Post-Training
von: Opedal, Andreas, et al.
Veröffentlicht: (2026)
von: Opedal, Andreas, et al.
Veröffentlicht: (2026)
Continuous Diffusion Scales Competitively with Discrete Diffusion for Language
von: Yang, Zhihan, et al.
Veröffentlicht: (2026)
von: Yang, Zhihan, et al.
Veröffentlicht: (2026)
Orthogonal Finetuning Made Scalable
von: Qiu, Zeju, et al.
Veröffentlicht: (2025)
von: Qiu, Zeju, et al.
Veröffentlicht: (2025)
Unifying Continuous and Discrete Text Diffusion with Non-simultaneous Diffusion Processes
von: Li, Bocheng, et al.
Veröffentlicht: (2025)
von: Li, Bocheng, et al.
Veröffentlicht: (2025)
GRASP: Deterministic argument ranking in interaction graphs
von: Misra, Diganta, et al.
Veröffentlicht: (2026)
von: Misra, Diganta, et al.
Veröffentlicht: (2026)
Can Large Language Models Infer Causation from Correlation?
von: Jin, Zhijing, et al.
Veröffentlicht: (2023)
von: Jin, Zhijing, et al.
Veröffentlicht: (2023)
Analyzing the Role of Semantic Representations in the Era of Large Language Models
von: Jin, Zhijing, et al.
Veröffentlicht: (2024)
von: Jin, Zhijing, et al.
Veröffentlicht: (2024)
Do Language Models Exhibit the Same Cognitive Biases in Problem Solving as Human Learners?
von: Opedal, Andreas, et al.
Veröffentlicht: (2024)
von: Opedal, Andreas, et al.
Veröffentlicht: (2024)
Non-Markovian Discrete Diffusion with Causal Language Models
von: Zhang, Yangtian, et al.
Veröffentlicht: (2025)
von: Zhang, Yangtian, et al.
Veröffentlicht: (2025)
Fine-Tuning Discrete Diffusion Models with Policy Gradient Methods
von: Zekri, Oussama, et al.
Veröffentlicht: (2025)
von: Zekri, Oussama, et al.
Veröffentlicht: (2025)
Your Absorbing Discrete Diffusion Secretly Models the Bayesian Posterior
von: Doyle, Cooper
Veröffentlicht: (2025)
von: Doyle, Cooper
Veröffentlicht: (2025)
Quriosity: Analyzing Human Questioning Behavior and Causal Inquiry through Curiosity-Driven Queries
von: Ceraolo, Roberto, et al.
Veröffentlicht: (2024)
von: Ceraolo, Roberto, et al.
Veröffentlicht: (2024)
Soft-Masked Diffusion Language Models
von: Hersche, Michael, et al.
Veröffentlicht: (2025)
von: Hersche, Michael, et al.
Veröffentlicht: (2025)
Steering Without Breaking: Mechanistically Informed Interventions for Discrete Diffusion Language Models
von: Zhou, Hanhan, et al.
Veröffentlicht: (2026)
von: Zhou, Hanhan, et al.
Veröffentlicht: (2026)
Large Language Models for Detection of Life-Threatening Texts
von: Nguyen, Thanh Thi, et al.
Veröffentlicht: (2025)
von: Nguyen, Thanh Thi, et al.
Veröffentlicht: (2025)
MetaState: Persistent Working Memory Enhances Reasoning in Discrete Diffusion Language Models
von: Xia, Kejing, et al.
Veröffentlicht: (2026)
von: Xia, Kejing, et al.
Veröffentlicht: (2026)
Cultural Alien Sampler: Open-ended art generation balancing originality and coherence
von: Artiles, Alejandro H., et al.
Veröffentlicht: (2025)
von: Artiles, Alejandro H., et al.
Veröffentlicht: (2025)
Flexible-length Text Infilling for Discrete Diffusion Models
von: Zhang, Andrew, et al.
Veröffentlicht: (2025)
von: Zhang, Andrew, et al.
Veröffentlicht: (2025)
Most Likely Sequence Generation for $n$-Grams, Transformers, HMMs, and Markov Chains, by Using Rollout Algorithms
von: Li, Yuchao, et al.
Veröffentlicht: (2024)
von: Li, Yuchao, et al.
Veröffentlicht: (2024)
Think While You Generate: Discrete Diffusion with Planned Denoising
von: Liu, Sulin, et al.
Veröffentlicht: (2024)
von: Liu, Sulin, et al.
Veröffentlicht: (2024)
CLadder: Assessing Causal Reasoning in Language Models
von: Jin, Zhijing, et al.
Veröffentlicht: (2023)
von: Jin, Zhijing, et al.
Veröffentlicht: (2023)
Ten Words Only Still Help: Improving Black-Box AI-Generated Text Detection via Proxy-Guided Efficient Re-Sampling
von: Shi, Yuhui, et al.
Veröffentlicht: (2024)
von: Shi, Yuhui, et al.
Veröffentlicht: (2024)
Are Language Models Efficient Reasoners? A Perspective from Logic Programming
von: Opedal, Andreas, et al.
Veröffentlicht: (2025)
von: Opedal, Andreas, et al.
Veröffentlicht: (2025)
TAID: Temporally Adaptive Interpolated Distillation for Efficient Knowledge Transfer in Language Models
von: Shing, Makoto, et al.
Veröffentlicht: (2025)
von: Shing, Makoto, et al.
Veröffentlicht: (2025)
Discrete Markov Bridge
von: Li, Hengli, et al.
Veröffentlicht: (2025)
von: Li, Hengli, et al.
Veröffentlicht: (2025)
Scalable Non-Equivariant 3D Molecule Generation via Rotational Alignment
von: Ding, Yuhui, et al.
Veröffentlicht: (2025)
von: Ding, Yuhui, et al.
Veröffentlicht: (2025)
IAA: Inner-Adaptor Architecture Empowers Frozen Large Language Model with Multimodal Capabilities
von: Wang, Bin, et al.
Veröffentlicht: (2024)
von: Wang, Bin, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Scaling Behavior of Discrete Diffusion Language Models
von: von Rütte, Dimitri, et al.
Veröffentlicht: (2025) -
Identifying Intervenable and Interpretable Features via Orthogonality Regularization
von: Miller, Moritz, et al.
Veröffentlicht: (2026) -
A Language Model's Guide Through Latent Space
von: von Rütte, Dimitri, et al.
Veröffentlicht: (2024) -
Improving Large Language Model Safety with Contrastive Representation Learning
von: Simko, Samuel, et al.
Veröffentlicht: (2025) -
Deriving Hyperparameter Scaling Laws via Modern Optimization Theory
von: Shulgin, Egor, et al.
Veröffentlicht: (2026)