Adaptive Block-Scaled Data Types
Fuente:
arXiv
Guardado en:
| Autores principales: | Cook, Jack, Lee, Hyemin S., Le, Kathryn, Guo, Junxian, Traverso, Giovanni, Chandrakasan, Anantha P., Han, Song |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Four Over Six: More Accurate NVFP4 Quantization with Adaptive Block Scaling
por: Cook, Jack, et al.
Publicado: (2025)
por: Cook, Jack, et al.
Publicado: (2025)
Optimizing Mixture of Block Attention
por: Xiao, Guangxuan, et al.
Publicado: (2025)
por: Xiao, Guangxuan, et al.
Publicado: (2025)
XAttention: Block Sparse Attention with Antidiagonal Scoring
por: Xu, Ruyi, et al.
Publicado: (2025)
por: Xu, Ruyi, et al.
Publicado: (2025)
ARGUS: Adaptive Rotation-Invariant Geometric Unsupervised System
por: Sharma, Anantha
Publicado: (2026)
por: Sharma, Anantha
Publicado: (2026)
Embedding Retrofitting: Data Engineering for better RAG
por: Sharma, Anantha
Publicado: (2026)
por: Sharma, Anantha
Publicado: (2026)
CORDIAL: Can Multimodal Large Language Models Effectively Understand Coherence Relationships?
por: Ramakrishnan, Aashish Anantha, et al.
Publicado: (2025)
por: Ramakrishnan, Aashish Anantha, et al.
Publicado: (2025)
Divide and Translate: Compositional First-Order Logic Translation and Verification for Complex Logical Reasoning
por: Ryu, Hyun, et al.
Publicado: (2024)
por: Ryu, Hyun, et al.
Publicado: (2024)
PGB: One-Shot Pruning for BERT via Weight Grouping and Permutation
por: Lim, Hyemin, et al.
Publicado: (2025)
por: Lim, Hyemin, et al.
Publicado: (2025)
IRONIC: Coherence-Aware Reasoning Chains for Multi-Modal Sarcasm Detection
por: Ramakrishnan, Aashish Anantha, et al.
Publicado: (2025)
por: Ramakrishnan, Aashish Anantha, et al.
Publicado: (2025)
RONA: Pragmatically Diverse Image Captioning with Coherence Relations
por: Ramakrishnan, Aashish Anantha, et al.
Publicado: (2025)
por: Ramakrishnan, Aashish Anantha, et al.
Publicado: (2025)
TableGuard -- Securing Structured & Unstructured Data
por: Sharma, Anantha, et al.
Publicado: (2024)
por: Sharma, Anantha, et al.
Publicado: (2024)
K/DA: Automated Data Generation Pipeline for Detoxifying Implicitly Offensive Language in Korean
por: Jeon, Minkyeong, et al.
Publicado: (2025)
por: Jeon, Minkyeong, et al.
Publicado: (2025)
Advancing Block Diffusion Language Models for Test-Time Scaling
por: Lu, Yi, et al.
Publicado: (2026)
por: Lu, Yi, et al.
Publicado: (2026)
DFlare: Scaling Up Draft Capacity for Block Diffusion Speculative Decoding
por: Zhang, Jiebin, et al.
Publicado: (2026)
por: Zhang, Jiebin, et al.
Publicado: (2026)
DuoAttention: Efficient Long-Context LLM Inference with Retrieval and Streaming Heads
por: Xiao, Guangxuan, et al.
Publicado: (2024)
por: Xiao, Guangxuan, et al.
Publicado: (2024)
Semantic Anchors in In-Context Learning: Why Small LLMs Cannot Flip Their Labels
por: Kumar, Anantha Padmanaban Krishna
Publicado: (2025)
por: Kumar, Anantha Padmanaban Krishna
Publicado: (2025)
Optimizing the Privacy-Utility Balance using Synthetic Data and Configurable Perturbation Pipelines
por: Sharma, Anantha, et al.
Publicado: (2025)
por: Sharma, Anantha, et al.
Publicado: (2025)
Subgraph-Aware Training of Language Models for Knowledge Graph Completion Using Structure-Aware Contrastive Learning
por: Ko, Youmin, et al.
Publicado: (2024)
por: Ko, Youmin, et al.
Publicado: (2024)
ANNA: Abstractive Text-to-Image Synthesis with Filtered News Captions
por: Ramakrishnan, Aashish Anantha, et al.
Publicado: (2023)
por: Ramakrishnan, Aashish Anantha, et al.
Publicado: (2023)
Applying RLAIF for Code Generation with API-usage in Lightweight LLMs
por: Dutta, Sujan, et al.
Publicado: (2024)
por: Dutta, Sujan, et al.
Publicado: (2024)
Predictive Data Selection: The Data That Predicts Is the Data That Teaches
por: Shum, Kashun, et al.
Publicado: (2025)
por: Shum, Kashun, et al.
Publicado: (2025)
No More Distractions: an Adaptive Up-Sampling Algorithm to Reduce Data Artifacts
por: Chen, Han
Publicado: (2024)
por: Chen, Han
Publicado: (2024)
From Intentions to Techniques: A Comprehensive Taxonomy and Challenges in Text Watermarking for Large Language Models
por: Lalai, Harsh Nishant, et al.
Publicado: (2024)
por: Lalai, Harsh Nishant, et al.
Publicado: (2024)
EnergyLens: Predictive Energy-Aware Exploration for Multi-GPU LLM Inference Optimization
por: Song, Zhiye, et al.
Publicado: (2026)
por: Song, Zhiye, et al.
Publicado: (2026)
EnergAIzer: Fast and Accurate GPU Power Estimation Framework for AI Workloads
por: Lee, Kyungmi, et al.
Publicado: (2026)
por: Lee, Kyungmi, et al.
Publicado: (2026)
Triplet-Block Diffusion RWKV
por: Lin, Ke, et al.
Publicado: (2026)
por: Lin, Ke, et al.
Publicado: (2026)
SynLogic: Synthesizing Verifiable Reasoning Data at Scale for Learning Logical Reasoning and Beyond
por: Liu, Junteng, et al.
Publicado: (2025)
por: Liu, Junteng, et al.
Publicado: (2025)
Towards Universal Dense Blocking for Entity Resolution
por: Wang, Tianshu, et al.
Publicado: (2024)
por: Wang, Tianshu, et al.
Publicado: (2024)
CAMPHOR: Collaborative Agents for Multi-input Planning and High-Order Reasoning On Device
por: Fu, Yicheng, et al.
Publicado: (2024)
por: Fu, Yicheng, et al.
Publicado: (2024)
Techniques to Improve Q&A Accuracy with Transformer-based models on Large Complex Documents
por: Liao, Chejui, et al.
Publicado: (2020)
por: Liao, Chejui, et al.
Publicado: (2020)
Swordsman: Entropy-Driven Adaptive Block Partition for Efficient Diffusion Language Models
por: Zhang, Yu, et al.
Publicado: (2026)
por: Zhang, Yu, et al.
Publicado: (2026)
SABlock: Semantic-Aware KV Cache Eviction with Adaptive Compression Block Size
por: Chen, Jinhan, et al.
Publicado: (2025)
por: Chen, Jinhan, et al.
Publicado: (2025)
Quantifying Data Contamination in Psychometric Evaluations of LLMs
por: Han, Jongwook, et al.
Publicado: (2025)
por: Han, Jongwook, et al.
Publicado: (2025)
ScalingFilter: Assessing Data Quality through Inverse Utilization of Scaling Laws
por: Li, Ruihang, et al.
Publicado: (2024)
por: Li, Ruihang, et al.
Publicado: (2024)
LEDRO: LLM-Enhanced Design Space Reduction and Optimization for Analog Circuits
por: Kochar, Dimple Vijay, et al.
Publicado: (2024)
por: Kochar, Dimple Vijay, et al.
Publicado: (2024)
Fast-dLLM v2: Efficient Block-Diffusion LLM
por: Wu, Chengyue, et al.
Publicado: (2025)
por: Wu, Chengyue, et al.
Publicado: (2025)
What Makes Good Data for Alignment? A Comprehensive Study of Automatic Data Selection in Instruction Tuning
por: Liu, Wei, et al.
Publicado: (2023)
por: Liu, Wei, et al.
Publicado: (2023)
Scaling Multi-Hop Training Data via Graph-Constrained Path Selection
por: Chen, Pengyu, et al.
Publicado: (2026)
por: Chen, Pengyu, et al.
Publicado: (2026)
On the Universal Truthfulness Hyperplane Inside LLMs
por: Liu, Junteng, et al.
Publicado: (2024)
por: Liu, Junteng, et al.
Publicado: (2024)
High-Dimensional Interlingual Representations of Large Language Models
por: Wilie, Bryan, et al.
Publicado: (2025)
por: Wilie, Bryan, et al.
Publicado: (2025)
Ejemplares similares
-
Four Over Six: More Accurate NVFP4 Quantization with Adaptive Block Scaling
por: Cook, Jack, et al.
Publicado: (2025) -
Optimizing Mixture of Block Attention
por: Xiao, Guangxuan, et al.
Publicado: (2025) -
XAttention: Block Sparse Attention with Antidiagonal Scoring
por: Xu, Ruyi, et al.
Publicado: (2025) -
ARGUS: Adaptive Rotation-Invariant Geometric Unsupervised System
por: Sharma, Anantha
Publicado: (2026) -
Embedding Retrofitting: Data Engineering for better RAG
por: Sharma, Anantha
Publicado: (2026)