Balancing Understanding and Generation in Discrete Diffusion Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liu, Yue, Zhao, Yuzhong, Xie, Zheyong, Ye, Qixiang, Jiao, Jianbin, Hu, Yao, Cao, Shaosheng, Liu, Yunfan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
One Token Is Enough: Improving Diffusion Language Models with a Sink Token
von: Zhang, Zihou, et al.
Veröffentlicht: (2026)
von: Zhang, Zihou, et al.
Veröffentlicht: (2026)
EComStage: Stage-wise and Orientation-specific Benchmarking for Large Language Models in E-commerce
von: Zhao, Kaiyan, et al.
Veröffentlicht: (2026)
von: Zhao, Kaiyan, et al.
Veröffentlicht: (2026)
Benchmarking Machine Translation on Chinese Social Media Texts
von: Zhao, Kaiyan, et al.
Veröffentlicht: (2026)
von: Zhao, Kaiyan, et al.
Veröffentlicht: (2026)
VMamba: Visual State Space Model
von: Liu, Yue, et al.
Veröffentlicht: (2024)
von: Liu, Yue, et al.
Veröffentlicht: (2024)
Decouple Searching from Training: Scaling Data Mixing via Model Merging for Large Language Model Pre-training
von: Li, Shengrui, et al.
Veröffentlicht: (2026)
von: Li, Shengrui, et al.
Veröffentlicht: (2026)
Pet-Bench: Benchmarking the Abilities of Large Language Models as E-Pets in Social Network Services
von: Guo, Hongcheng, et al.
Veröffentlicht: (2025)
von: Guo, Hongcheng, et al.
Veröffentlicht: (2025)
CC-Diff: Enhancing Contextual Coherence in Remote Sensing Image Synthesis
von: Zhang, Mu, et al.
Veröffentlicht: (2024)
von: Zhang, Mu, et al.
Veröffentlicht: (2024)
SNS-Bench-VL: Benchmarking Multimodal Large Language Models in Social Networking Services
von: Guo, Hongcheng, et al.
Veröffentlicht: (2025)
von: Guo, Hongcheng, et al.
Veröffentlicht: (2025)
QP-OneModel: A Unified Generative LLM for Multi-Task Query Understanding in Xiaohongshu Search
von: Huang, Jianzhao, et al.
Veröffentlicht: (2026)
von: Huang, Jianzhao, et al.
Veröffentlicht: (2026)
Redefining Machine Translation on Social Network Services with Large Language Models
von: Guo, Hongcheng, et al.
Veröffentlicht: (2025)
von: Guo, Hongcheng, et al.
Veröffentlicht: (2025)
Geometric-Mean Policy Optimization
von: Zhao, Yuzhong, et al.
Veröffentlicht: (2025)
von: Zhao, Yuzhong, et al.
Veröffentlicht: (2025)
Constrained Code Generation with Discrete Diffusion
von: Shao, Lize, et al.
Veröffentlicht: (2026)
von: Shao, Lize, et al.
Veröffentlicht: (2026)
Self-supervised Feature-Gate Coupling for Dynamic Network Pruning
von: Shi, Mengnan, et al.
Veröffentlicht: (2021)
von: Shi, Mengnan, et al.
Veröffentlicht: (2021)
Adaptive Keyframe Sampling for Long Video Understanding
von: Tang, Xi, et al.
Veröffentlicht: (2025)
von: Tang, Xi, et al.
Veröffentlicht: (2025)
Interleaving Reasoning for Better Text-to-Image Generation
von: Huang, Wenxuan, et al.
Veröffentlicht: (2025)
von: Huang, Wenxuan, et al.
Veröffentlicht: (2025)
DIFFA-2: A Practical Diffusion Large Language Model for General Audio Understanding
von: Zhou, Jiaming, et al.
Veröffentlicht: (2026)
von: Zhou, Jiaming, et al.
Veröffentlicht: (2026)
Dynamic-LLaVA: Efficient Multimodal Large Language Models via Dynamic Vision-language Context Sparsification
von: Huang, Wenxuan, et al.
Veröffentlicht: (2024)
von: Huang, Wenxuan, et al.
Veröffentlicht: (2024)
Depth-guided Texture Diffusion for Image Semantic Segmentation
von: Sun, Wei, et al.
Veröffentlicht: (2024)
von: Sun, Wei, et al.
Veröffentlicht: (2024)
Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models
von: Huang, Wenxuan, et al.
Veröffentlicht: (2025)
von: Huang, Wenxuan, et al.
Veröffentlicht: (2025)
Cluster-Driven Expert Pruning for Mixture-of-Experts Large Language Models
von: Guo, Hongcheng, et al.
Veröffentlicht: (2025)
von: Guo, Hongcheng, et al.
Veröffentlicht: (2025)
MIRAGE: Exploring How Large Language Models Perform in Complex Social Interactive Environments
von: Cai, Yin, et al.
Veröffentlicht: (2025)
von: Cai, Yin, et al.
Veröffentlicht: (2025)
Discrete Diffusion for Reflective Vision-Language-Action Models in Autonomous Driving
von: Li, Pengxiang, et al.
Veröffentlicht: (2025)
von: Li, Pengxiang, et al.
Veröffentlicht: (2025)
Building Vision Models upon Heat Conduction
von: Wang, Zhaozhi, et al.
Veröffentlicht: (2024)
von: Wang, Zhaozhi, et al.
Veröffentlicht: (2024)
Retrieve-Plan-Generation: An Iterative Planning and Answering Framework for Knowledge-Intensive LLM Generation
von: Lyu, Yuanjie, et al.
Veröffentlicht: (2024)
von: Lyu, Yuanjie, et al.
Veröffentlicht: (2024)
Thinking with Images via Self-Calling Agent
von: Yang, Wenxi, et al.
Veröffentlicht: (2025)
von: Yang, Wenxi, et al.
Veröffentlicht: (2025)
Balanced Data Sampling for Language Model Training with Clustering
von: Shao, Yunfan, et al.
Veröffentlicht: (2024)
von: Shao, Yunfan, et al.
Veröffentlicht: (2024)
TableMaster: A Recipe to Advance Table Understanding with Language Models
von: Cao, Lang, et al.
Veröffentlicht: (2025)
von: Cao, Lang, et al.
Veröffentlicht: (2025)
In-Context Former: Lightning-fast Compressing Context for Large Language Model
von: Wang, Xiangfeng, et al.
Veröffentlicht: (2024)
von: Wang, Xiangfeng, et al.
Veröffentlicht: (2024)
A Reparameterized Discrete Diffusion Model for Text Generation
von: Zheng, Lin, et al.
Veröffentlicht: (2023)
von: Zheng, Lin, et al.
Veröffentlicht: (2023)
DynRefer: Delving into Region-level Multimodal Tasks via Dynamic Resolution
von: Zhao, Yuzhong, et al.
Veröffentlicht: (2024)
von: Zhao, Yuzhong, et al.
Veröffentlicht: (2024)
Targeted Remasking: Replacing Token Editing with Token-to-Mask Refinement in Discrete Diffusion Language Models
von: Yao, Lin
Veröffentlicht: (2026)
von: Yao, Lin
Veröffentlicht: (2026)
FedJudge: Federated Legal Large Language Model
von: Yue, Linan, et al.
Veröffentlicht: (2023)
von: Yue, Linan, et al.
Veröffentlicht: (2023)
Ultra-Fast Language Generation via Discrete Diffusion Divergence Instruct
von: Zheng, Haoyang, et al.
Veröffentlicht: (2025)
von: Zheng, Haoyang, et al.
Veröffentlicht: (2025)
Expandable Residual Approximation for Knowledge Distillation
von: Yan, Zhaoyi, et al.
Veröffentlicht: (2025)
von: Yan, Zhaoyi, et al.
Veröffentlicht: (2025)
Spatial Transform Decoupling for Oriented Object Detection
von: Yu, Hongtian, et al.
Veröffentlicht: (2023)
von: Yu, Hongtian, et al.
Veröffentlicht: (2023)
DiffListener: Discrete Diffusion Model for Listener Generation
von: Jung, Siyeol, et al.
Veröffentlicht: (2025)
von: Jung, Siyeol, et al.
Veröffentlicht: (2025)
LangFlow: Continuous Diffusion Rivals Discrete in Language Modeling
von: Chen, Yuxin, et al.
Veröffentlicht: (2026)
von: Chen, Yuxin, et al.
Veröffentlicht: (2026)
On Discrete Prompt Optimization for Diffusion Models
von: Wang, Ruochen, et al.
Veröffentlicht: (2024)
von: Wang, Ruochen, et al.
Veröffentlicht: (2024)
Timestep Embedding Tells: It's Time to Cache for Video Diffusion Model
von: Liu, Feng, et al.
Veröffentlicht: (2024)
von: Liu, Feng, et al.
Veröffentlicht: (2024)
Beyond Hard and Soft: Hybrid Context Compression for Balancing Local and Global Information Retention
von: Liao, Huanxuan, et al.
Veröffentlicht: (2025)
von: Liao, Huanxuan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
One Token Is Enough: Improving Diffusion Language Models with a Sink Token
von: Zhang, Zihou, et al.
Veröffentlicht: (2026) -
EComStage: Stage-wise and Orientation-specific Benchmarking for Large Language Models in E-commerce
von: Zhao, Kaiyan, et al.
Veröffentlicht: (2026) -
Benchmarking Machine Translation on Chinese Social Media Texts
von: Zhao, Kaiyan, et al.
Veröffentlicht: (2026) -
VMamba: Visual State Space Model
von: Liu, Yue, et al.
Veröffentlicht: (2024) -
Decouple Searching from Training: Scaling Data Mixing via Model Merging for Large Language Model Pre-training
von: Li, Shengrui, et al.
Veröffentlicht: (2026)