DiffuCoder: Understanding and Improving Masked Diffusion Models for Code Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Gong, Shansan, Zhang, Ruixiang, Zheng, Huangjie, Gu, Jiatao, Jaitly, Navdeep, Kong, Lingpeng, Zhang, Yizhe |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Embarrassingly Simple Self-Distillation Improves Code Generation
von: Zhang, Ruixiang, et al.
Veröffentlicht: (2026)
von: Zhang, Ruixiang, et al.
Veröffentlicht: (2026)
Continuously Augmented Discrete Diffusion model for Categorical Generative Modeling
von: Zheng, Huangjie, et al.
Veröffentlicht: (2025)
von: Zheng, Huangjie, et al.
Veröffentlicht: (2025)
Flexible Language Modeling in Continuous Space with Transformer-based Autoregressive Flows
von: Zhang, Ruixiang, et al.
Veröffentlicht: (2025)
von: Zhang, Ruixiang, et al.
Veröffentlicht: (2025)
PLANNER: Generating Diversified Paragraph via Latent Language Diffusion Model
von: Zhang, Yizhe, et al.
Veröffentlicht: (2023)
von: Zhang, Yizhe, et al.
Veröffentlicht: (2023)
SAGE: Steering Dialog Generation with Future-Aware State-Action Augmentation
von: Zhang, Yizhe, et al.
Veröffentlicht: (2025)
von: Zhang, Yizhe, et al.
Veröffentlicht: (2025)
How Far Are We from Intelligent Visual Deductive Reasoning?
von: Zhang, Yizhe, et al.
Veröffentlicht: (2024)
von: Zhang, Yizhe, et al.
Veröffentlicht: (2024)
Dream-Coder 7B: An Open Diffusion Language Model for Code
von: Xie, Zhihui, et al.
Veröffentlicht: (2025)
von: Xie, Zhihui, et al.
Veröffentlicht: (2025)
Improving GFlowNets for Text-to-Image Diffusion Alignment
von: Zhang, Dinghuai, et al.
Veröffentlicht: (2024)
von: Zhang, Dinghuai, et al.
Veröffentlicht: (2024)
What Makes the Preferred Thinking Direction for LLMs in Multiple-choice Questions?
von: Zhang, Yizhe, et al.
Veröffentlicht: (2025)
von: Zhang, Yizhe, et al.
Veröffentlicht: (2025)
Matryoshka Diffusion Models
von: Gu, Jiatao, et al.
Veröffentlicht: (2023)
von: Gu, Jiatao, et al.
Veröffentlicht: (2023)
Kaleido Diffusion: Improving Conditional Diffusion Models with Autoregressive Latent Modeling
von: Gu, Jiatao, et al.
Veröffentlicht: (2024)
von: Gu, Jiatao, et al.
Veröffentlicht: (2024)
Probing the Multi-turn Planning Capabilities of LLMs via 20 Question Games
von: Zhang, Yizhe, et al.
Veröffentlicht: (2023)
von: Zhang, Yizhe, et al.
Veröffentlicht: (2023)
ChipChat: Low-Latency Cascaded Conversational Agent in MLX
von: Likhomanenko, Tatiana, et al.
Veröffentlicht: (2025)
von: Likhomanenko, Tatiana, et al.
Veröffentlicht: (2025)
Divide-or-Conquer? Which Part Should You Distill Your LLM?
von: Wu, Zhuofeng, et al.
Veröffentlicht: (2024)
von: Wu, Zhuofeng, et al.
Veröffentlicht: (2024)
TypeScore: A Text Fidelity Metric for Text-to-Image Generative Models
von: Sampaio, Georgia Gabriela, et al.
Veröffentlicht: (2024)
von: Sampaio, Georgia Gabriela, et al.
Veröffentlicht: (2024)
Scaling Diffusion Language Models via Adaptation from Autoregressive Models
von: Gong, Shansan, et al.
Veröffentlicht: (2024)
von: Gong, Shansan, et al.
Veröffentlicht: (2024)
Normalizing Flows are Capable Generative Models
von: Zhai, Shuangfei, et al.
Veröffentlicht: (2024)
von: Zhai, Shuangfei, et al.
Veröffentlicht: (2024)
Training-Free Long-Context Scaling of Large Language Models
von: An, Chenxin, et al.
Veröffentlicht: (2024)
von: An, Chenxin, et al.
Veröffentlicht: (2024)
Beyond Autoregression: Discrete Diffusion for Complex Reasoning and Planning
von: Ye, Jiacheng, et al.
Veröffentlicht: (2024)
von: Ye, Jiacheng, et al.
Veröffentlicht: (2024)
DreamOn: Diffusion Language Models For Code Infilling Beyond Fixed-size Canvas
von: Wu, Zirui, et al.
Veröffentlicht: (2026)
von: Wu, Zirui, et al.
Veröffentlicht: (2026)
Rephrasing the Web: A Recipe for Compute and Data-Efficient Language Modeling
von: Maini, Pratyush, et al.
Veröffentlicht: (2024)
von: Maini, Pratyush, et al.
Veröffentlicht: (2024)
KGLens: Towards Efficient and Effective Knowledge Probing of Large Language Models with Knowledge Graphs
von: Zheng, Shangshang, et al.
Veröffentlicht: (2023)
von: Zheng, Shangshang, et al.
Veröffentlicht: (2023)
DiffuMask: Diffusion Language Model for Token-level Prompt Pruning
von: Zheng, Caleb, et al.
Veröffentlicht: (2026)
von: Zheng, Caleb, et al.
Veröffentlicht: (2026)
CLaRa: Bridging Retrieval and Generation with Continuous Latent Reasoning
von: He, Jie, et al.
Veröffentlicht: (2025)
von: He, Jie, et al.
Veröffentlicht: (2025)
Self-Infilling Code Generation
von: Zheng, Lin, et al.
Veröffentlicht: (2023)
von: Zheng, Lin, et al.
Veröffentlicht: (2023)
Why Does the Effective Context Length of LLMs Fall Short?
von: An, Chenxin, et al.
Veröffentlicht: (2024)
von: An, Chenxin, et al.
Veröffentlicht: (2024)
dMel: Speech Tokenization made Simple
von: Bai, Richard He, et al.
Veröffentlicht: (2024)
von: Bai, Richard He, et al.
Veröffentlicht: (2024)
Dream-VL & Dream-VLA: Open Vision-Language and Vision-Language-Action Models with Diffusion Language Model Backbone
von: Ye, Jiacheng, et al.
Veröffentlicht: (2025)
von: Ye, Jiacheng, et al.
Veröffentlicht: (2025)
LaDiR: Latent Diffusion Enhances LLMs for Text Reasoning
von: Kang, Haoqiang, et al.
Veröffentlicht: (2025)
von: Kang, Haoqiang, et al.
Veröffentlicht: (2025)
BBA: Bi-Modal Behavioral Alignment for Reasoning with Large Vision-Language Models
von: Zhao, Xueliang, et al.
Veröffentlicht: (2024)
von: Zhao, Xueliang, et al.
Veröffentlicht: (2024)
Diffusion of Thoughts: Chain-of-Thought Reasoning in Diffusion Language Models
von: Ye, Jiacheng, et al.
Veröffentlicht: (2024)
von: Ye, Jiacheng, et al.
Veröffentlicht: (2024)
DART: Denoising Autoregressive Transformer for Scalable Text-to-Image Generation
von: Gu, Jiatao, et al.
Veröffentlicht: (2024)
von: Gu, Jiatao, et al.
Veröffentlicht: (2024)
Omni-Router: Sharing Routing Decisions in Sparse Mixture-of-Experts for Speech Recognition
von: Gu, Zijin, et al.
Veröffentlicht: (2025)
von: Gu, Zijin, et al.
Veröffentlicht: (2025)
A Reparameterized Discrete Diffusion Model for Text Generation
von: Zheng, Lin, et al.
Veröffentlicht: (2023)
von: Zheng, Lin, et al.
Veröffentlicht: (2023)
Eliciting In-context Retrieval and Reasoning for Long-context Large Language Models
von: Qiu, Yifu, et al.
Veröffentlicht: (2025)
von: Qiu, Yifu, et al.
Veröffentlicht: (2025)
Training Software Engineering Agents and Verifiers with SWE-Gym
von: Pan, Jiayi, et al.
Veröffentlicht: (2024)
von: Pan, Jiayi, et al.
Veröffentlicht: (2024)
RefineCoder: Iterative Improving of Large Language Models via Adaptive Critique Refinement for Code Generation
von: Zhou, Changzhi, et al.
Veröffentlicht: (2025)
von: Zhou, Changzhi, et al.
Veröffentlicht: (2025)
DiffuGuard: How Intrinsic Safety is Lost and Found in Diffusion Large Language Models
von: Li, Zherui, et al.
Veröffentlicht: (2025)
von: Li, Zherui, et al.
Veröffentlicht: (2025)
StepCoder: Improve Code Generation with Reinforcement Learning from Compiler Feedback
von: Dou, Shihan, et al.
Veröffentlicht: (2024)
von: Dou, Shihan, et al.
Veröffentlicht: (2024)
SpeakStream: Streaming Text-to-Speech with Interleaved Data
von: Bai, Richard He, et al.
Veröffentlicht: (2025)
von: Bai, Richard He, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Embarrassingly Simple Self-Distillation Improves Code Generation
von: Zhang, Ruixiang, et al.
Veröffentlicht: (2026) -
Continuously Augmented Discrete Diffusion model for Categorical Generative Modeling
von: Zheng, Huangjie, et al.
Veröffentlicht: (2025) -
Flexible Language Modeling in Continuous Space with Transformer-based Autoregressive Flows
von: Zhang, Ruixiang, et al.
Veröffentlicht: (2025) -
PLANNER: Generating Diversified Paragraph via Latent Language Diffusion Model
von: Zhang, Yizhe, et al.
Veröffentlicht: (2023) -
SAGE: Steering Dialog Generation with Future-Aware State-Action Augmentation
von: Zhang, Yizhe, et al.
Veröffentlicht: (2025)