Tiled Bit Networks: Sub-Bit Neural Network Compression Through Reuse of Learnable Binary Vectors
Fuente:
arXiv
Saved in:
| Main Authors: | Gorbett, Matt, Shirazi, Hossein, Ray, Indrakshi |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Label-Free Reinforcement Learning via Cross-Model Entropy
by: Gorbett, Matt, et al.
Published: (2026)
by: Gorbett, Matt, et al.
Published: (2026)
Cross-Silo Federated Learning Across Divergent Domains with Iterative Parameter Alignment
by: Gorbett, Matt, et al.
Published: (2023)
by: Gorbett, Matt, et al.
Published: (2023)
BTC-LLM: Efficient Sub-1-Bit LLM Quantization via Learnable Transformation and Binary Codebook
by: Gu, Hao, et al.
Published: (2025)
by: Gu, Hao, et al.
Published: (2025)
A&B BNN: Add&Bit-Operation-Only Hardware-Friendly Binary Neural Network
by: Ma, Ruichen, et al.
Published: (2024)
by: Ma, Ruichen, et al.
Published: (2024)
Learning Grouped Lattice Vector Quantizers for Low-Bit LLM Compression
by: Zhang, Xi, et al.
Published: (2025)
by: Zhang, Xi, et al.
Published: (2025)
SplitQuant: Layer Splitting for Low-Bit Neural Network Quantization
by: Song, Jaewoo, et al.
Published: (2025)
by: Song, Jaewoo, et al.
Published: (2025)
Diagonal-Tiled Mixed-Precision Attention for Efficient Low-Bit MXFP Inference
by: Ding, Yifu, et al.
Published: (2026)
by: Ding, Yifu, et al.
Published: (2026)
Verification of Bit-Flip Attacks against Quantized Neural Networks
by: Zhang, Yedi, et al.
Published: (2025)
by: Zhang, Yedi, et al.
Published: (2025)
LittleBit-2: Maximizing the Spectral Energy Gain in Sub-1-Bit LLMs via Latent Geometry Alignment
by: Lee, Banseok, et al.
Published: (2026)
by: Lee, Banseok, et al.
Published: (2026)
FP=xINT:Representing Neural Networks via Low-Bit Series Basis Functions
by: Zhang, Boyang, et al.
Published: (2024)
by: Zhang, Boyang, et al.
Published: (2024)
Unlocking the Theory Behind Scaling 1-Bit Neural Networks
by: Daliri, Majid, et al.
Published: (2024)
by: Daliri, Majid, et al.
Published: (2024)
From Arithmetic to Logic: The Resilience of Logic and Lookup-Based Neural Networks Under Parameter Bit-Flips
by: Bacellar, Alan T. L., et al.
Published: (2026)
by: Bacellar, Alan T. L., et al.
Published: (2026)
Point Cloud Compression with Bits-back Coding
by: Hieu, Nguyen Quang, et al.
Published: (2024)
by: Hieu, Nguyen Quang, et al.
Published: (2024)
Q-Palette: Fractional-Bit Quantizers Toward Optimal Bit Allocation for Efficient LLM Deployment
by: Lee, Deokjae, et al.
Published: (2025)
by: Lee, Deokjae, et al.
Published: (2025)
UltraSketchLLM: Saliency-Driven Sketching for Ultra-Low Bit LLM Compression
by: Zou, Sunan, et al.
Published: (2025)
by: Zou, Sunan, et al.
Published: (2025)
BitsMoE: Efficient Spectral Energy-Guided Bit Allocation for MoE LLM Quantization
by: Zhao, Jiayu, et al.
Published: (2026)
by: Zhao, Jiayu, et al.
Published: (2026)
LittleBit: Ultra Low-Bit Quantization via Latent Factorization
by: Lee, Banseok, et al.
Published: (2025)
by: Lee, Banseok, et al.
Published: (2025)
Towards Cheaper Inference in Deep Networks with Lower Bit-Width Accumulators
by: Blumenfeld, Yaniv, et al.
Published: (2024)
by: Blumenfeld, Yaniv, et al.
Published: (2024)
Normalized Architectures are Natively 4-Bit
by: Fishman, Maxim, et al.
Published: (2026)
by: Fishman, Maxim, et al.
Published: (2026)
Attacking Graph Neural Networks with Bit Flips: Weisfeiler and Lehman Go Indifferent
by: Kummer, Lorenz, et al.
Published: (2023)
by: Kummer, Lorenz, et al.
Published: (2023)
Sign Lock-In: Randomly Initialized Weight Signs Persist and Bottleneck Sub-Bit Model Compression
by: Sakai, Akira, et al.
Published: (2026)
by: Sakai, Akira, et al.
Published: (2026)
On the Probabilistic Learnability of Compact Neural Network Preimage Bounds
by: Marzari, Luca, et al.
Published: (2025)
by: Marzari, Luca, et al.
Published: (2025)
NSNQuant: A Double Normalization Approach for Calibration-Free Low-Bit Vector Quantization of KV Cache
by: Son, Donghyun, et al.
Published: (2025)
by: Son, Donghyun, et al.
Published: (2025)
More Than Bits: Multi-Envelope Double Binary Factorization for Extreme Quantization
by: Ichikawa, Yuma, et al.
Published: (2025)
by: Ichikawa, Yuma, et al.
Published: (2025)
PATCH: Learnable Tile-level Hybrid Sparsity for LLMs
by: Hourri, Younes, et al.
Published: (2025)
by: Hourri, Younes, et al.
Published: (2025)
Human-aligned Chess with a Bit of Search
by: Zhang, Yiming, et al.
Published: (2024)
by: Zhang, Yiming, et al.
Published: (2024)
Boosting Graph Neural Network Expressivity with Learnable Lanczos Constraints
by: Azizi, Niloofar, et al.
Published: (2024)
by: Azizi, Niloofar, et al.
Published: (2024)
Towards Efficient and Accurate Spiking Neural Networks via Adaptive Bit Allocation
by: Yao, Xingting, et al.
Published: (2025)
by: Yao, Xingting, et al.
Published: (2025)
Müntz-Szász Networks: Neural Architectures with Learnable Power-Law Bases
by: N'guessan, Gnankan Landry Regis
Published: (2025)
by: N'guessan, Gnankan Landry Regis
Published: (2025)
Physics-Informed Neural Networks with Learnable Loss Balancing and Transfer Learning
by: Pirayeshshirazinezhad, Reza
Published: (2026)
by: Pirayeshshirazinezhad, Reza
Published: (2026)
Maximal Brain Damage Without Data or Optimization: Disrupting Neural Networks via Sign-Bit Flips
by: Galil, Ido, et al.
Published: (2025)
by: Galil, Ido, et al.
Published: (2025)
Self-Compressing Neural Networks
by: Cséfalvay, Szabolcs, et al.
Published: (2023)
by: Cséfalvay, Szabolcs, et al.
Published: (2023)
To be Continuous, or to be Discrete, Those are Bits of Questions
by: Wang, Yiran, et al.
Published: (2024)
by: Wang, Yiran, et al.
Published: (2024)
Skill Reuse as Compression in Agentic RL
by: Xu, Zhikun, et al.
Published: (2026)
by: Xu, Zhikun, et al.
Published: (2026)
Depth-Adaptive Graph Neural Networks via Learnable Bakry-'Emery Curvature
by: Hevapathige, Asela, et al.
Published: (2025)
by: Hevapathige, Asela, et al.
Published: (2025)
AdaQAT: Adaptive Bit-Width Quantization-Aware Training
by: Gernigon, Cédric, et al.
Published: (2024)
by: Gernigon, Cédric, et al.
Published: (2024)
On Exact Bit-level Reversible Transformers Without Changing Architectures
by: Zhang, Guoqiang, et al.
Published: (2024)
by: Zhang, Guoqiang, et al.
Published: (2024)
Safe Transformer: An Explicit Safety Bit For Interpretable And Controllable Alignment
by: Feng, Jingyuan, et al.
Published: (2026)
by: Feng, Jingyuan, et al.
Published: (2026)
Attn-QAT: 4-Bit Attention With Quantization-Aware Training
by: Zhang, Peiyuan, et al.
Published: (2026)
by: Zhang, Peiyuan, et al.
Published: (2026)
PuzzleMoE: Efficient Compression of Large Mixture-of-Experts Models via Sparse Expert Merging and Bit-packed inference
by: Zhao, Yushu, et al.
Published: (2025)
by: Zhao, Yushu, et al.
Published: (2025)
Similar Items
-
Label-Free Reinforcement Learning via Cross-Model Entropy
by: Gorbett, Matt, et al.
Published: (2026) -
Cross-Silo Federated Learning Across Divergent Domains with Iterative Parameter Alignment
by: Gorbett, Matt, et al.
Published: (2023) -
BTC-LLM: Efficient Sub-1-Bit LLM Quantization via Learnable Transformation and Binary Codebook
by: Gu, Hao, et al.
Published: (2025) -
A&B BNN: Add&Bit-Operation-Only Hardware-Friendly Binary Neural Network
by: Ma, Ruichen, et al.
Published: (2024) -
Learning Grouped Lattice Vector Quantizers for Low-Bit LLM Compression
by: Zhang, Xi, et al.
Published: (2025)