DiffusionBlocks: Block-wise Neural Network Training via Diffusion Interpretation
Fuente:
arXiv
Saved in:
| Main Authors: | Shing, Makoto, Koyama, Masanori, Akiba, Takuya |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
UnMaskFork: Test-Time Scaling for Masked Diffusion via Deterministic Action Branching
by: Misaki, Kou, et al.
Published: (2026)
by: Misaki, Kou, et al.
Published: (2026)
An Efficient Training Algorithm for Models with Block-wise Sparsity
by: Zhu, Ding, et al.
Published: (2025)
by: Zhu, Ding, et al.
Published: (2025)
AdaBlock-dLLM: Semantic-Aware Diffusion LLM Inference via Adaptive Block Size
by: Lu, Guanxi, et al.
Published: (2025)
by: Lu, Guanxi, et al.
Published: (2025)
Block Diffusion: Interpolating Between Autoregressive and Diffusion Language Models
by: Arriola, Marianne, et al.
Published: (2025)
by: Arriola, Marianne, et al.
Published: (2025)
TAID: Temporally Adaptive Interpolated Distillation for Efficient Knowledge Transfer in Language Models
by: Shing, Makoto, et al.
Published: (2025)
by: Shing, Makoto, et al.
Published: (2025)
Efficient ANN-Guided Distillation: Aligning Rate-based Features of Spiking Neural Networks through Hybrid Block-wise Replacement
by: Yang, Shu, et al.
Published: (2025)
by: Yang, Shu, et al.
Published: (2025)
Realizing Unaligned Block-wise Pruning for DNN Acceleration on Mobile Devices
by: Lee, Hayun, et al.
Published: (2024)
by: Lee, Hayun, et al.
Published: (2024)
Cognitive Chunking for Soft Prompts: Accelerating Compressor Learning via Block-wise Causal Masking
by: Liu, Guojie, et al.
Published: (2026)
by: Liu, Guojie, et al.
Published: (2026)
GUDA: Counterfactual Group-wise Training Data Attribution for Diffusion Models via Unlearning
by: Murata, Naoki, et al.
Published: (2026)
by: Murata, Naoki, et al.
Published: (2026)
Differentially Private Block-wise Gradient Shuffle for Deep Learning
by: Zagardo, David
Published: (2024)
by: Zagardo, David
Published: (2024)
DepCap: Adaptive Block-Wise Parallel Decoding for Efficient Diffusion LM Inference
by: Xia, Xiang, et al.
Published: (2026)
by: Xia, Xiang, et al.
Published: (2026)
BlockBatch: Multi-Scale Consensus Decoding for Efficient Diffusion Language Model Inference
by: Wu, Xiaoyou, et al.
Published: (2026)
by: Wu, Xiaoyou, et al.
Published: (2026)
Pushing the Limits of Block Rotations in Post-Training Quantization
by: Sanjeet, Sai, et al.
Published: (2026)
by: Sanjeet, Sai, et al.
Published: (2026)
BLAST: Block-Level Adaptive Structured Matrices for Efficient Deep Neural Network Inference
by: Lee, Changwoo, et al.
Published: (2024)
by: Lee, Changwoo, et al.
Published: (2024)
Training Neural Networks for Modularity aids Interpretability
by: Golechha, Satvik, et al.
Published: (2024)
by: Golechha, Satvik, et al.
Published: (2024)
Recurrent Stochastic Configuration Networks with Incremental Blocks
by: Dang, Gang, et al.
Published: (2024)
by: Dang, Gang, et al.
Published: (2024)
FOAM: Blocked State Folding for Memory-Efficient LLM Training
by: Wen, Ziqing, et al.
Published: (2025)
by: Wen, Ziqing, et al.
Published: (2025)
BEND: Bagging Deep Learning Training Based on Efficient Neural Network Diffusion
by: Wei, Jia, et al.
Published: (2024)
by: Wei, Jia, et al.
Published: (2024)
Unifying Block-wise PTQ and Distillation-based QAT for Progressive Quantization toward 2-bit Instruction-Tuned LLMs
by: Lee, Jung Hyun, et al.
Published: (2025)
by: Lee, Jung Hyun, et al.
Published: (2025)
Efficient Large Language Model Inference with Neural Block Linearization
by: Erdogan, Mete, et al.
Published: (2025)
by: Erdogan, Mete, et al.
Published: (2025)
Exploiting Block Coordinate Descent for Cost-Effective LLM Model Training
by: Liu, Zeyu, et al.
Published: (2025)
by: Liu, Zeyu, et al.
Published: (2025)
An Active Diffusion Neural Network for Graphs
by: Jiang, Mengying
Published: (2025)
by: Jiang, Mengying
Published: (2025)
Regularizing Neural Network Training via Identity-wise Discriminative Feature Suppression
by: Chapman, Avraham, et al.
Published: (2022)
by: Chapman, Avraham, et al.
Published: (2022)
Block-wise Adaptive Caching for Accelerating Diffusion Policy
by: Ji, Kangye, et al.
Published: (2025)
by: Ji, Kangye, et al.
Published: (2025)
Diffusion-TS: Interpretable Diffusion for General Time Series Generation
by: Yuan, Xinyu, et al.
Published: (2024)
by: Yuan, Xinyu, et al.
Published: (2024)
Block-Based Double Decoders
by: Labovich, Asher, et al.
Published: (2026)
by: Labovich, Asher, et al.
Published: (2026)
Diffusion-Based Neural Network Weights Generation
by: Soro, Bedionita, et al.
Published: (2024)
by: Soro, Bedionita, et al.
Published: (2024)
Structurally Flexible Neural Networks: Evolving the Building Blocks for General Agents
by: Pedersen, Joachim Winther, et al.
Published: (2024)
by: Pedersen, Joachim Winther, et al.
Published: (2024)
Interpretable Diffusion via Information Decomposition
by: Kong, Xianghao, et al.
Published: (2023)
by: Kong, Xianghao, et al.
Published: (2023)
ECHO: Efficient Chest X-ray Report Generation with One-step Block Diffusion
by: Chen, Lifeng, et al.
Published: (2026)
by: Chen, Lifeng, et al.
Published: (2026)
LNN-PINN: A Unified Physics-Only Training Framework with Liquid Residual Blocks
by: Tao, Ze, et al.
Published: (2025)
by: Tao, Ze, et al.
Published: (2025)
Neural Diffusion Processes for Physically Interpretable Survival Prediction
by: Cristofoletto, Alessio, et al.
Published: (2025)
by: Cristofoletto, Alessio, et al.
Published: (2025)
Structure-based RNA Design by Step-wise Optimization of Latent Diffusion Model
by: Si, Qi, et al.
Published: (2026)
by: Si, Qi, et al.
Published: (2026)
Graph Neural Diffusion Networks for Semi-supervised Learning
by: Ye, Wei, et al.
Published: (2022)
by: Ye, Wei, et al.
Published: (2022)
From Uniform to Adaptive: General Skip-Block Mechanisms for Efficient PDE Neural Operators
by: Liu, Lei, et al.
Published: (2025)
by: Liu, Lei, et al.
Published: (2025)
ReplaceMe: Network Simplification via Depth Pruning and Transformer Block Linearization
by: Shopkhoev, Dmitriy, et al.
Published: (2025)
by: Shopkhoev, Dmitriy, et al.
Published: (2025)
Training Long-Context LLMs Efficiently via Chunk-wise Optimization
by: Li, Wenhao, et al.
Published: (2025)
by: Li, Wenhao, et al.
Published: (2025)
Graph Neural Diffusion via Generalized Opinion Dynamics
by: Hevapathige, Asela, et al.
Published: (2025)
by: Hevapathige, Asela, et al.
Published: (2025)
Interactions Across Blocks in Post-Training Quantization of Large Language Models
by: Shabanovi, Khasmamad, et al.
Published: (2024)
by: Shabanovi, Khasmamad, et al.
Published: (2024)
BlockGPT: Spatio-Temporal Modelling of Rainfall via Frame-Level Autoregression
by: Meo, Cristian, et al.
Published: (2025)
by: Meo, Cristian, et al.
Published: (2025)
Similar Items
-
UnMaskFork: Test-Time Scaling for Masked Diffusion via Deterministic Action Branching
by: Misaki, Kou, et al.
Published: (2026) -
An Efficient Training Algorithm for Models with Block-wise Sparsity
by: Zhu, Ding, et al.
Published: (2025) -
AdaBlock-dLLM: Semantic-Aware Diffusion LLM Inference via Adaptive Block Size
by: Lu, Guanxi, et al.
Published: (2025) -
Block Diffusion: Interpolating Between Autoregressive and Diffusion Language Models
by: Arriola, Marianne, et al.
Published: (2025) -
TAID: Temporally Adaptive Interpolated Distillation for Efficient Knowledge Transfer in Language Models
by: Shing, Makoto, et al.
Published: (2025)