BLAST: Block-Level Adaptive Structured Matrices for Efficient Deep Neural Network Inference
Fuente:
arXiv
Saved in:
| Main Authors: | Lee, Changwoo, Kwon, Soo Min, Qu, Qing, Kim, Hun-Seok |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Differentiable Learning of Generalized Structured Matrices for Efficient Deep Neural Networks
by: Lee, Changwoo, et al.
Published: (2023)
by: Lee, Changwoo, et al.
Published: (2023)
Memory-Efficient Acceleration of Block Low-Rank Foundation Models on Resource Constrained GPUs
by: Abillama, Pierre, et al.
Published: (2025)
by: Abillama, Pierre, et al.
Published: (2025)
Efficient Large Language Model Inference with Neural Block Linearization
by: Erdogan, Mete, et al.
Published: (2025)
by: Erdogan, Mete, et al.
Published: (2025)
DepCap: Adaptive Block-Wise Parallel Decoding for Efficient Diffusion LM Inference
by: Xia, Xiang, et al.
Published: (2026)
by: Xia, Xiang, et al.
Published: (2026)
Explicit Feature Interaction-aware Graph Neural Networks
by: Kim, Minkyu, et al.
Published: (2022)
by: Kim, Minkyu, et al.
Published: (2022)
BugSweeper: Function-Level Detection of Smart Contract Vulnerabilities Using Graph Neural Networks
by: Lee, Uisang, et al.
Published: (2025)
by: Lee, Uisang, et al.
Published: (2025)
LRQ: Optimizing Post-Training Quantization for Large Language Models by Learning Low-Rank Weight-Scaling Matrices
by: Lee, Jung Hyun, et al.
Published: (2024)
by: Lee, Jung Hyun, et al.
Published: (2024)
RL-SPH: Learning to Achieve Feasible Solutions for Integer Linear Programs
by: Lee, Tae-Hoon, et al.
Published: (2024)
by: Lee, Tae-Hoon, et al.
Published: (2024)
From Uniform to Adaptive: General Skip-Block Mechanisms for Efficient PDE Neural Operators
by: Liu, Lei, et al.
Published: (2025)
by: Liu, Lei, et al.
Published: (2025)
ICaRus: Identical Cache Reuse for Efficient Multi Model Inference
by: Woo, Sunghyeon, et al.
Published: (2026)
by: Woo, Sunghyeon, et al.
Published: (2026)
From Algorithm to Hardware: A Survey on Efficient and Safe Deployment of Deep Neural Networks
by: Geng, Xue, et al.
Published: (2024)
by: Geng, Xue, et al.
Published: (2024)
Attention-Based Neural-Augmented Kalman Filter for Legged Robot State Estimation
by: Lee, Seokju, et al.
Published: (2026)
by: Lee, Seokju, et al.
Published: (2026)
AdaBlock-dLLM: Semantic-Aware Diffusion LLM Inference via Adaptive Block Size
by: Lu, Guanxi, et al.
Published: (2025)
by: Lu, Guanxi, et al.
Published: (2025)
From Accuracy to Readiness: Metrics and Benchmarks for Human-AI Decision-Making
by: Lee, Min Hun
Published: (2026)
by: Lee, Min Hun
Published: (2026)
Shortcut Features as Top Eigenfunctions of NTK: A Linear Neural Network Case and More
by: Lim, Jinwoo, et al.
Published: (2026)
by: Lim, Jinwoo, et al.
Published: (2026)
SMMF: Square-Matricized Momentum Factorization for Memory-Efficient Optimization
by: Park, Kwangryeol, et al.
Published: (2024)
by: Park, Kwangryeol, et al.
Published: (2024)
Bridging Neural Networks and Dynamic Time Warping for Adaptive Time Series Classification
by: Qu, Jintao, et al.
Published: (2025)
by: Qu, Jintao, et al.
Published: (2025)
Graph-Level Label-Only Membership Inference Attack against Graph Neural Networks
by: Dai, Jiazhu, et al.
Published: (2025)
by: Dai, Jiazhu, et al.
Published: (2025)
DiffusionBlocks: Block-wise Neural Network Training via Diffusion Interpretation
by: Shing, Makoto, et al.
Published: (2025)
by: Shing, Makoto, et al.
Published: (2025)
Multi-LLM Adaptive Conformal Inference for Reliable LLM Responses
by: Noh, Kangjun, et al.
Published: (2026)
by: Noh, Kangjun, et al.
Published: (2026)
Depth-Adaptive Graph Neural Networks via Learnable Bakry-'Emery Curvature
by: Hevapathige, Asela, et al.
Published: (2025)
by: Hevapathige, Asela, et al.
Published: (2025)
Quantum Circuit Structure Optimization for Quantum Reinforcement Learning
by: Son, Seok Bin, et al.
Published: (2025)
by: Son, Seok Bin, et al.
Published: (2025)
Anomaly Detection with Adaptive and Aggressive Rejection for Contaminated Training Data
by: Lee, Jungi, et al.
Published: (2025)
by: Lee, Jungi, et al.
Published: (2025)
Think in Blocks: Adaptive Reasoning from Direct Response to Deep Reasoning
by: Zhu, Yekun, et al.
Published: (2025)
by: Zhu, Yekun, et al.
Published: (2025)
Mitigating Overfitting in Graph Neural Networks via Feature and Hyperplane Perturbation
by: Choi, Yoonhyuk, et al.
Published: (2022)
by: Choi, Yoonhyuk, et al.
Published: (2022)
Eau De $Q$-Network: Adaptive Distillation of Neural Networks in Deep Reinforcement Learning
by: Vincent, Théo, et al.
Published: (2025)
by: Vincent, Théo, et al.
Published: (2025)
Permutation-Invariant Graph Partitioning:How Graph Neural Networks Capture Structural Interactions?
by: Hevapathige, Asela, et al.
Published: (2023)
by: Hevapathige, Asela, et al.
Published: (2023)
Exploring a Multimodal Fusion-based Deep Learning Network for Detecting Facial Palsy
by: Oo, Heng Yim Nicole, et al.
Published: (2024)
by: Oo, Heng Yim Nicole, et al.
Published: (2024)
Self-Abstraction Learning for Effective and Stable Training of Deep Neural Networks
by: Cho, Wonyong, et al.
Published: (2026)
by: Cho, Wonyong, et al.
Published: (2026)
Adaptive Inference-Time Scaling via Cyclic Diffusion Search
by: Lee, Gyubin, et al.
Published: (2025)
by: Lee, Gyubin, et al.
Published: (2025)
FEATHer: Fourier-Efficient Adaptive Temporal Hierarchy Forecaster for Time-Series Forecasting
by: Lee, Jaehoon, et al.
Published: (2026)
by: Lee, Jaehoon, et al.
Published: (2026)
NeuZip: Memory-Efficient Training and Inference with Dynamic Compression of Neural Networks
by: Hao, Yongchang, et al.
Published: (2024)
by: Hao, Yongchang, et al.
Published: (2024)
Adaptive Data Harvesting for Efficient Neural Network Learning with Universal Constraints
by: Kang, Siteng, et al.
Published: (2026)
by: Kang, Siteng, et al.
Published: (2026)
Better Not to Propagate: Understanding Edge Uncertainty and Over-smoothing in Signed Graph Neural Networks
by: Choi, Yoonhyuk, et al.
Published: (2024)
by: Choi, Yoonhyuk, et al.
Published: (2024)
AnyBCQ: Hardware Efficient Flexible Binary-Coded Quantization for Multi-Precision LLMs
by: Park, Gunho, et al.
Published: (2025)
by: Park, Gunho, et al.
Published: (2025)
BlockBatch: Multi-Scale Consensus Decoding for Efficient Diffusion Language Model Inference
by: Wu, Xiaoyou, et al.
Published: (2026)
by: Wu, Xiaoyou, et al.
Published: (2026)
TransPL: VQ-Code Transition Matrices for Pseudo-Labeling of Time Series Unsupervised Domain Adaptation
by: Kim, Jaeho, et al.
Published: (2025)
by: Kim, Jaeho, et al.
Published: (2025)
Multi-View Graph Convolution Network for Internal Talent Recommendation Based on Enterprise Emails
by: Kim, Soo Hyun, et al.
Published: (2025)
by: Kim, Soo Hyun, et al.
Published: (2025)
Controllable 3D Molecular Generation for Structure-Based Drug Design Through Bayesian Flow Networks and Gradient Integration
by: Choi, Seungyeon, et al.
Published: (2025)
by: Choi, Seungyeon, et al.
Published: (2025)
The Confusion is Real: GRAPHIC -- A Network Science Approach to Confusion Matrices in Deep Learning
by: Fröhlich, Johanna S., et al.
Published: (2026)
by: Fröhlich, Johanna S., et al.
Published: (2026)
Similar Items
-
Differentiable Learning of Generalized Structured Matrices for Efficient Deep Neural Networks
by: Lee, Changwoo, et al.
Published: (2023) -
Memory-Efficient Acceleration of Block Low-Rank Foundation Models on Resource Constrained GPUs
by: Abillama, Pierre, et al.
Published: (2025) -
Efficient Large Language Model Inference with Neural Block Linearization
by: Erdogan, Mete, et al.
Published: (2025) -
DepCap: Adaptive Block-Wise Parallel Decoding for Efficient Diffusion LM Inference
by: Xia, Xiang, et al.
Published: (2026) -
Explicit Feature Interaction-aware Graph Neural Networks
by: Kim, Minkyu, et al.
Published: (2022)