Block-R1: Rethinking the Role of Block Size in Multi-domain Reinforcement Learning for Diffusion Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jiang, Yan, Qiu, Ruihong, Huang, Zi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Break the Block: Dynamic-size Reasoning Blocks for Diffusion Large Language Models via Monotonic Entropy Descent with Reinforcement Learning
von: Jiang, Yan, et al.
Veröffentlicht: (2026)
von: Jiang, Yan, et al.
Veröffentlicht: (2026)
When to Commit? Towards Variable-Size Self-Contained Blocks for Discrete Diffusion Language Models
von: Wang, Danny, et al.
Veröffentlicht: (2026)
von: Wang, Danny, et al.
Veröffentlicht: (2026)
TRN-R1-Zero: Text-rich Network Reasoning via LLMs with Reinforcement Learning Only
von: Liu, Yilun, et al.
Veröffentlicht: (2026)
von: Liu, Yilun, et al.
Veröffentlicht: (2026)
Does Homophily Help in Robust Test-time Node Classification?
von: Jiang, Yan, et al.
Veröffentlicht: (2025)
von: Jiang, Yan, et al.
Veröffentlicht: (2025)
Text Meets Topology: Rethinking Out-of-distribution Detection in Text-Rich Networks
von: Wang, Danny, et al.
Veröffentlicht: (2025)
von: Wang, Danny, et al.
Veröffentlicht: (2025)
What Information Matters? Graph Out-of-Distribution Detection via Tri-Component Information Decomposition
von: Wang, Danny, et al.
Veröffentlicht: (2026)
von: Wang, Danny, et al.
Veröffentlicht: (2026)
GCondenser: Benchmarking Graph Condensation
von: Liu, Yilun, et al.
Veröffentlicht: (2024)
von: Liu, Yilun, et al.
Veröffentlicht: (2024)
GeoBlock: Inferring Block Granularity from Dependency Geometry in Diffusion Language Models
von: Wan, Lipeng, et al.
Veröffentlicht: (2026)
von: Wan, Lipeng, et al.
Veröffentlicht: (2026)
AdaBlock-dLLM: Semantic-Aware Diffusion LLM Inference via Adaptive Block Size
von: Lu, Guanxi, et al.
Veröffentlicht: (2025)
von: Lu, Guanxi, et al.
Veröffentlicht: (2025)
GOLD: Graph Out-of-Distribution Detection via Implicit Adversarial Latent Generation
von: Wang, Danny, et al.
Veröffentlicht: (2025)
von: Wang, Danny, et al.
Veröffentlicht: (2025)
PUMA: Efficient Continual Graph Learning for Node Classification with Graph Condensation
von: Liu, Yilun, et al.
Veröffentlicht: (2023)
von: Liu, Yilun, et al.
Veröffentlicht: (2023)
ParaBlock: Communication-Computation Parallel Block Coordinate Federated Learning for Large Language Models
von: Wang, Yujia, et al.
Veröffentlicht: (2025)
von: Wang, Yujia, et al.
Veröffentlicht: (2025)
Block Diffusion: Interpolating Between Autoregressive and Diffusion Language Models
von: Arriola, Marianne, et al.
Veröffentlicht: (2025)
von: Arriola, Marianne, et al.
Veröffentlicht: (2025)
Block Circulant Adapter for Large Language Models
von: Ding, Xinyu, et al.
Veröffentlicht: (2025)
von: Ding, Xinyu, et al.
Veröffentlicht: (2025)
BlockBatch: Multi-Scale Consensus Decoding for Efficient Diffusion Language Model Inference
von: Wu, Xiaoyou, et al.
Veröffentlicht: (2026)
von: Wu, Xiaoyou, et al.
Veröffentlicht: (2026)
From Tokens to Blocks: A Block-Diffusion Perspective on Molecular Generation
von: Yang, Qianwei, et al.
Veröffentlicht: (2026)
von: Yang, Qianwei, et al.
Veröffentlicht: (2026)
CBQ: Cross-Block Quantization for Large Language Models
von: Ding, Xin, et al.
Veröffentlicht: (2023)
von: Ding, Xin, et al.
Veröffentlicht: (2023)
DiffusionBlocks: Block-wise Neural Network Training via Diffusion Interpretation
von: Shing, Makoto, et al.
Veröffentlicht: (2025)
von: Shing, Makoto, et al.
Veröffentlicht: (2025)
TrajDLM: Topology-Aware Block Diffusion Language Model for Trajectory Generation
von: Wongso, Wilson, et al.
Veröffentlicht: (2026)
von: Wongso, Wilson, et al.
Veröffentlicht: (2026)
ALSA: Anchors in Logit Space for Out-of-Distribution Accuracy Estimation
von: Liu, Chenzhi, et al.
Veröffentlicht: (2025)
von: Liu, Chenzhi, et al.
Veröffentlicht: (2025)
Efficient Large Language Model Inference with Neural Block Linearization
von: Erdogan, Mete, et al.
Veröffentlicht: (2025)
von: Erdogan, Mete, et al.
Veröffentlicht: (2025)
LEGO: Language Model Building Blocks
von: Bhansali, Shrenik, et al.
Veröffentlicht: (2024)
von: Bhansali, Shrenik, et al.
Veröffentlicht: (2024)
SBGD: Improving Graph Diffusion Generative Model via Stochastic Block Diffusion
von: Su, Junwei, et al.
Veröffentlicht: (2025)
von: Su, Junwei, et al.
Veröffentlicht: (2025)
Blocked Gibbs meets Diffusion Transformers: Unsupervised Learning for Constraint Optimization
von: Xu, Yudong W., et al.
Veröffentlicht: (2026)
von: Xu, Yudong W., et al.
Veröffentlicht: (2026)
Heterogeneous Multi-agent Multi-armed Bandits on Stochastic Block Models
von: Xu, Mengfan, et al.
Veröffentlicht: (2025)
von: Xu, Mengfan, et al.
Veröffentlicht: (2025)
LoSA: Locality Aware Sparse Attention for Block-Wise Diffusion Language Models
von: Xi, Haocheng, et al.
Veröffentlicht: (2026)
von: Xi, Haocheng, et al.
Veröffentlicht: (2026)
Rethinking the Role of Temperature in Large Language Model Distillation
von: Luong, Hoang-Chau, et al.
Veröffentlicht: (2026)
von: Luong, Hoang-Chau, et al.
Veröffentlicht: (2026)
Block Flow: Learning Straight Flow on Data Blocks
von: Wang, Zibin, et al.
Veröffentlicht: (2025)
von: Wang, Zibin, et al.
Veröffentlicht: (2025)
Interactions Across Blocks in Post-Training Quantization of Large Language Models
von: Shabanovi, Khasmamad, et al.
Veröffentlicht: (2024)
von: Shabanovi, Khasmamad, et al.
Veröffentlicht: (2024)
Set Block Decoding is a Language Model Inference Accelerator
von: Gat, Itai, et al.
Veröffentlicht: (2025)
von: Gat, Itai, et al.
Veröffentlicht: (2025)
Large Language Model-Enhanced Reinforcement Learning for Diverse and Novel Recommendations
von: Woo, Jiin, et al.
Veröffentlicht: (2025)
von: Woo, Jiin, et al.
Veröffentlicht: (2025)
Efficient Reinforcement Learning with Large Language Model Priors
von: Yan, Xue, et al.
Veröffentlicht: (2024)
von: Yan, Xue, et al.
Veröffentlicht: (2024)
Review, Remask, Refine (R3): Process-Guided Block Diffusion for Text Generation
von: Mounier, Nikita, et al.
Veröffentlicht: (2025)
von: Mounier, Nikita, et al.
Veröffentlicht: (2025)
Blocking Bandits
von: Basu, Soumya, et al.
Veröffentlicht: (2019)
von: Basu, Soumya, et al.
Veröffentlicht: (2019)
PagedEviction: Structured Block-wise KV Cache Pruning for Efficient Large Language Model Inference
von: Chitty-Venkata, Krishna Teja, et al.
Veröffentlicht: (2025)
von: Chitty-Venkata, Krishna Teja, et al.
Veröffentlicht: (2025)
Consensus Knowledge Graph Learning via Multi-view Sparse Low Rank Block Model
von: Cai, Tianxi, et al.
Veröffentlicht: (2022)
von: Cai, Tianxi, et al.
Veröffentlicht: (2022)
AttentionLego: An Open-Source Building Block For Spatially-Scalable Large Language Model Accelerator With Processing-In-Memory Technology
von: Cong, Rongqing, et al.
Veröffentlicht: (2024)
von: Cong, Rongqing, et al.
Veröffentlicht: (2024)
Multi-task Learning for Heterogeneous Multi-source Block-Wise Missing Data
von: Sui, Yang, et al.
Veröffentlicht: (2025)
von: Sui, Yang, et al.
Veröffentlicht: (2025)
Accelerating Multi-Block Constrained Optimization Through Learning to Optimize
von: Liang, Ling, et al.
Veröffentlicht: (2024)
von: Liang, Ling, et al.
Veröffentlicht: (2024)
Guaranteed Conditional Diffusion: 3D Block-based Models for Scientific Data Compression
von: Lee, Jaemoon, et al.
Veröffentlicht: (2025)
von: Lee, Jaemoon, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Break the Block: Dynamic-size Reasoning Blocks for Diffusion Large Language Models via Monotonic Entropy Descent with Reinforcement Learning
von: Jiang, Yan, et al.
Veröffentlicht: (2026) -
When to Commit? Towards Variable-Size Self-Contained Blocks for Discrete Diffusion Language Models
von: Wang, Danny, et al.
Veröffentlicht: (2026) -
TRN-R1-Zero: Text-rich Network Reasoning via LLMs with Reinforcement Learning Only
von: Liu, Yilun, et al.
Veröffentlicht: (2026) -
Does Homophily Help in Robust Test-time Node Classification?
von: Jiang, Yan, et al.
Veröffentlicht: (2025) -
Text Meets Topology: Rethinking Out-of-distribution Detection in Text-Rich Networks
von: Wang, Danny, et al.
Veröffentlicht: (2025)