Adaptive Block-Scaled Data Types
Fuente:
arXiv
Saved in:
| Main Authors: | Cook, Jack, Lee, Hyemin S., Le, Kathryn, Guo, Junxian, Traverso, Giovanni, Chandrakasan, Anantha P., Han, Song |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Four Over Six: More Accurate NVFP4 Quantization with Adaptive Block Scaling
by: Cook, Jack, et al.
Published: (2025)
by: Cook, Jack, et al.
Published: (2025)
Optimizing Mixture of Block Attention
by: Xiao, Guangxuan, et al.
Published: (2025)
by: Xiao, Guangxuan, et al.
Published: (2025)
XAttention: Block Sparse Attention with Antidiagonal Scoring
by: Xu, Ruyi, et al.
Published: (2025)
by: Xu, Ruyi, et al.
Published: (2025)
ARGUS: Adaptive Rotation-Invariant Geometric Unsupervised System
by: Sharma, Anantha
Published: (2026)
by: Sharma, Anantha
Published: (2026)
Embedding Retrofitting: Data Engineering for better RAG
by: Sharma, Anantha
Published: (2026)
by: Sharma, Anantha
Published: (2026)
CORDIAL: Can Multimodal Large Language Models Effectively Understand Coherence Relationships?
by: Ramakrishnan, Aashish Anantha, et al.
Published: (2025)
by: Ramakrishnan, Aashish Anantha, et al.
Published: (2025)
Divide and Translate: Compositional First-Order Logic Translation and Verification for Complex Logical Reasoning
by: Ryu, Hyun, et al.
Published: (2024)
by: Ryu, Hyun, et al.
Published: (2024)
PGB: One-Shot Pruning for BERT via Weight Grouping and Permutation
by: Lim, Hyemin, et al.
Published: (2025)
by: Lim, Hyemin, et al.
Published: (2025)
IRONIC: Coherence-Aware Reasoning Chains for Multi-Modal Sarcasm Detection
by: Ramakrishnan, Aashish Anantha, et al.
Published: (2025)
by: Ramakrishnan, Aashish Anantha, et al.
Published: (2025)
RONA: Pragmatically Diverse Image Captioning with Coherence Relations
by: Ramakrishnan, Aashish Anantha, et al.
Published: (2025)
by: Ramakrishnan, Aashish Anantha, et al.
Published: (2025)
TableGuard -- Securing Structured & Unstructured Data
by: Sharma, Anantha, et al.
Published: (2024)
by: Sharma, Anantha, et al.
Published: (2024)
K/DA: Automated Data Generation Pipeline for Detoxifying Implicitly Offensive Language in Korean
by: Jeon, Minkyeong, et al.
Published: (2025)
by: Jeon, Minkyeong, et al.
Published: (2025)
Advancing Block Diffusion Language Models for Test-Time Scaling
by: Lu, Yi, et al.
Published: (2026)
by: Lu, Yi, et al.
Published: (2026)
DFlare: Scaling Up Draft Capacity for Block Diffusion Speculative Decoding
by: Zhang, Jiebin, et al.
Published: (2026)
by: Zhang, Jiebin, et al.
Published: (2026)
DuoAttention: Efficient Long-Context LLM Inference with Retrieval and Streaming Heads
by: Xiao, Guangxuan, et al.
Published: (2024)
by: Xiao, Guangxuan, et al.
Published: (2024)
Semantic Anchors in In-Context Learning: Why Small LLMs Cannot Flip Their Labels
by: Kumar, Anantha Padmanaban Krishna
Published: (2025)
by: Kumar, Anantha Padmanaban Krishna
Published: (2025)
Optimizing the Privacy-Utility Balance using Synthetic Data and Configurable Perturbation Pipelines
by: Sharma, Anantha, et al.
Published: (2025)
by: Sharma, Anantha, et al.
Published: (2025)
Subgraph-Aware Training of Language Models for Knowledge Graph Completion Using Structure-Aware Contrastive Learning
by: Ko, Youmin, et al.
Published: (2024)
by: Ko, Youmin, et al.
Published: (2024)
ANNA: Abstractive Text-to-Image Synthesis with Filtered News Captions
by: Ramakrishnan, Aashish Anantha, et al.
Published: (2023)
by: Ramakrishnan, Aashish Anantha, et al.
Published: (2023)
Applying RLAIF for Code Generation with API-usage in Lightweight LLMs
by: Dutta, Sujan, et al.
Published: (2024)
by: Dutta, Sujan, et al.
Published: (2024)
Predictive Data Selection: The Data That Predicts Is the Data That Teaches
by: Shum, Kashun, et al.
Published: (2025)
by: Shum, Kashun, et al.
Published: (2025)
No More Distractions: an Adaptive Up-Sampling Algorithm to Reduce Data Artifacts
by: Chen, Han
Published: (2024)
by: Chen, Han
Published: (2024)
From Intentions to Techniques: A Comprehensive Taxonomy and Challenges in Text Watermarking for Large Language Models
by: Lalai, Harsh Nishant, et al.
Published: (2024)
by: Lalai, Harsh Nishant, et al.
Published: (2024)
EnergyLens: Predictive Energy-Aware Exploration for Multi-GPU LLM Inference Optimization
by: Song, Zhiye, et al.
Published: (2026)
by: Song, Zhiye, et al.
Published: (2026)
EnergAIzer: Fast and Accurate GPU Power Estimation Framework for AI Workloads
by: Lee, Kyungmi, et al.
Published: (2026)
by: Lee, Kyungmi, et al.
Published: (2026)
Triplet-Block Diffusion RWKV
by: Lin, Ke, et al.
Published: (2026)
by: Lin, Ke, et al.
Published: (2026)
SynLogic: Synthesizing Verifiable Reasoning Data at Scale for Learning Logical Reasoning and Beyond
by: Liu, Junteng, et al.
Published: (2025)
by: Liu, Junteng, et al.
Published: (2025)
Towards Universal Dense Blocking for Entity Resolution
by: Wang, Tianshu, et al.
Published: (2024)
by: Wang, Tianshu, et al.
Published: (2024)
CAMPHOR: Collaborative Agents for Multi-input Planning and High-Order Reasoning On Device
by: Fu, Yicheng, et al.
Published: (2024)
by: Fu, Yicheng, et al.
Published: (2024)
Techniques to Improve Q&A Accuracy with Transformer-based models on Large Complex Documents
by: Liao, Chejui, et al.
Published: (2020)
by: Liao, Chejui, et al.
Published: (2020)
Swordsman: Entropy-Driven Adaptive Block Partition for Efficient Diffusion Language Models
by: Zhang, Yu, et al.
Published: (2026)
by: Zhang, Yu, et al.
Published: (2026)
SABlock: Semantic-Aware KV Cache Eviction with Adaptive Compression Block Size
by: Chen, Jinhan, et al.
Published: (2025)
by: Chen, Jinhan, et al.
Published: (2025)
Quantifying Data Contamination in Psychometric Evaluations of LLMs
by: Han, Jongwook, et al.
Published: (2025)
by: Han, Jongwook, et al.
Published: (2025)
ScalingFilter: Assessing Data Quality through Inverse Utilization of Scaling Laws
by: Li, Ruihang, et al.
Published: (2024)
by: Li, Ruihang, et al.
Published: (2024)
LEDRO: LLM-Enhanced Design Space Reduction and Optimization for Analog Circuits
by: Kochar, Dimple Vijay, et al.
Published: (2024)
by: Kochar, Dimple Vijay, et al.
Published: (2024)
Fast-dLLM v2: Efficient Block-Diffusion LLM
by: Wu, Chengyue, et al.
Published: (2025)
by: Wu, Chengyue, et al.
Published: (2025)
What Makes Good Data for Alignment? A Comprehensive Study of Automatic Data Selection in Instruction Tuning
by: Liu, Wei, et al.
Published: (2023)
by: Liu, Wei, et al.
Published: (2023)
Scaling Multi-Hop Training Data via Graph-Constrained Path Selection
by: Chen, Pengyu, et al.
Published: (2026)
by: Chen, Pengyu, et al.
Published: (2026)
On the Universal Truthfulness Hyperplane Inside LLMs
by: Liu, Junteng, et al.
Published: (2024)
by: Liu, Junteng, et al.
Published: (2024)
High-Dimensional Interlingual Representations of Large Language Models
by: Wilie, Bryan, et al.
Published: (2025)
by: Wilie, Bryan, et al.
Published: (2025)
Similar Items
-
Four Over Six: More Accurate NVFP4 Quantization with Adaptive Block Scaling
by: Cook, Jack, et al.
Published: (2025) -
Optimizing Mixture of Block Attention
by: Xiao, Guangxuan, et al.
Published: (2025) -
XAttention: Block Sparse Attention with Antidiagonal Scoring
by: Xu, Ruyi, et al.
Published: (2025) -
ARGUS: Adaptive Rotation-Invariant Geometric Unsupervised System
by: Sharma, Anantha
Published: (2026) -
Embedding Retrofitting: Data Engineering for better RAG
by: Sharma, Anantha
Published: (2026)