You Had One Job: Per-Task Quantization Using LLMs' Hidden Representations
Fuente:
arXiv
Saved in:
| Main Authors: | LeVi, Amit, Lapid, Raz, Himelstein, Rom, Baskin, Chaim, Ziv, Ravid Shwartz, Mendelson, Avi |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Silent Tokens, Loud Effects: Padding in LLMs
by: Himelstein, Rom, et al.
Published: (2025)
by: Himelstein, Rom, et al.
Published: (2025)
Step-Wise Refusal Dynamics in Autoregressive and Diffusion Language Models
by: Rahimi, Eliron, et al.
Published: (2026)
by: Rahimi, Eliron, et al.
Published: (2026)
Silenced Biases: The Dark Side LLMs Learned to Refuse
by: Himelstein, Rom, et al.
Published: (2025)
by: Himelstein, Rom, et al.
Published: (2025)
Jailbreak Attack Initializations as Extractors of Compliance Directions
by: Levi, Amit, et al.
Published: (2025)
by: Levi, Amit, et al.
Published: (2025)
Extracting Recurring Vulnerabilities from Black-Box LLM-Generated Software
by: Kordonsky, Tomer, et al.
Published: (2026)
by: Kordonsky, Tomer, et al.
Published: (2026)
AMED: Automatic Mixed-Precision Quantization for Edge Devices
by: Kimhi, Moshe, et al.
Published: (2022)
by: Kimhi, Moshe, et al.
Published: (2022)
Sparse patches adversarial attacks via extrapolating point-wise information
by: Nemcovsky, Yaniv, et al.
Published: (2024)
by: Nemcovsky, Yaniv, et al.
Published: (2024)
$\mathbf{R}^3$: Reconstruction, Raw, and Rain: Deraining Directly in the Bayer Domain
by: Rothschild, Nate, et al.
Published: (2025)
by: Rothschild, Nate, et al.
Published: (2025)
Hysteresis Activation Function for Efficient Inference
by: Kimhi, Moshe, et al.
Published: (2024)
by: Kimhi, Moshe, et al.
Published: (2024)
Video Representation Learning with Joint-Embedding Predictive Architectures
by: Drozdov, Katrina, et al.
Published: (2024)
by: Drozdov, Katrina, et al.
Published: (2024)
Learning to Compress: Local Rank and Information Compression in Deep Neural Networks
by: Patel, Niket, et al.
Published: (2024)
by: Patel, Niket, et al.
Published: (2024)
Breaking Audio Large Language Models by Attacking Only the Encoder: A Universal Targeted Latent-Space Audio Attack
by: Ziv, Roee, et al.
Published: (2025)
by: Ziv, Roee, et al.
Published: (2025)
Layer by Layer: Uncovering Hidden Representations in Language Models
by: Skean, Oscar, et al.
Published: (2025)
by: Skean, Oscar, et al.
Published: (2025)
Does Representation Matter? Exploring Intermediate Layers in Large Language Models
by: Skean, Oscar, et al.
Published: (2024)
by: Skean, Oscar, et al.
Published: (2024)
Variance-Covariance Regularization Improves Representation Learning
by: Zhu, Jiachen, et al.
Published: (2023)
by: Zhu, Jiachen, et al.
Published: (2023)
Representing LLMs in Prompt Semantic Task Space
by: Kashani, Idan, et al.
Published: (2025)
by: Kashani, Idan, et al.
Published: (2025)
From Tokens to Thoughts: How LLMs and Humans Trade Compression for Meaning
by: Shani, Chen, et al.
Published: (2025)
by: Shani, Chen, et al.
Published: (2025)
When Attention Collapses: How Degenerate Layers in LLMs Enable Smaller, Stronger Models
by: Sanyal, Sunny, et al.
Published: (2024)
by: Sanyal, Sunny, et al.
Published: (2024)
On the Robustness of Diffusion-Based Image Compression to Bit-Flip Errors
by: Vaisman, Amit, et al.
Published: (2026)
by: Vaisman, Amit, et al.
Published: (2026)
AI Must Embrace Specialization via Superhuman Adaptable Intelligence
by: Goldfeder, Judah, et al.
Published: (2026)
by: Goldfeder, Judah, et al.
Published: (2026)
The Entropy Enigma: Success and Failure of Entropy Minimization
by: Press, Ori, et al.
Published: (2024)
by: Press, Ori, et al.
Published: (2024)
Leveraging Temporal Graph Networks Using Module Decoupling
by: Feldman, Or, et al.
Published: (2023)
by: Feldman, Or, et al.
Published: (2023)
Antislop: A Comprehensive Framework for Identifying and Eliminating Repetitive Patterns in Language Models
by: Paech, Samuel, et al.
Published: (2025)
by: Paech, Samuel, et al.
Published: (2025)
Backdoors in Conditional Diffusion: Threats to Responsible Synthetic Data Pipelines
by: Lapid, Raz, et al.
Published: (2025)
by: Lapid, Raz, et al.
Published: (2025)
The Illusion of Progress: Re-evaluating Hallucination Detection in LLMs
by: Janiak, Denis, et al.
Published: (2025)
by: Janiak, Denis, et al.
Published: (2025)
On Training in Imagination
by: Timor, Nadav, et al.
Published: (2026)
by: Timor, Nadav, et al.
Published: (2026)
JEPA as a Neural Tokenizer: Learning Robust Speech Representations with Density Adaptive Attention
by: Ioannides, Georgios, et al.
Published: (2025)
by: Ioannides, Georgios, et al.
Published: (2025)
Seq-VCR: Preventing Collapse in Intermediate Transformer Representations for Enhanced Reasoning
by: Arefin, Md Rifat, et al.
Published: (2024)
by: Arefin, Md Rifat, et al.
Published: (2024)
REMIND: Input Loss Landscapes Reveal Residual Memorization in Post-Unlearning LLMs
by: Cohen, Liran, et al.
Published: (2025)
by: Cohen, Liran, et al.
Published: (2025)
Latent Transfer Attack: Adversarial Examples via Generative Latent Spaces
by: Shaar, Eitan, et al.
Published: (2026)
by: Shaar, Eitan, et al.
Published: (2026)
Attention Sinks and Compression Valleys in LLMs are Two Sides of the Same Coin
by: Queipo-de-Llano, Enrique, et al.
Published: (2025)
by: Queipo-de-Llano, Enrique, et al.
Published: (2025)
Soft Clustering Anchors for Self-Supervised Speech Representation Learning in Joint Embedding Prediction Architectures
by: Ioannides, Georgios, et al.
Published: (2026)
by: Ioannides, Georgios, et al.
Published: (2026)
An Information-Theoretic Perspective on Variance-Invariance-Covariance Regularization
by: Shwartz-Ziv, Ravid, et al.
Published: (2023)
by: Shwartz-Ziv, Ravid, et al.
Published: (2023)
Rate-In: Information-Driven Adaptive Dropout Rates for Improved Inference-Time Uncertainty Estimation
by: Zeevi, Tal, et al.
Published: (2024)
by: Zeevi, Tal, et al.
Published: (2024)
Domain Restriction via Multi SAE Layer Transitions
by: Shaheen, Elias, et al.
Published: (2026)
by: Shaheen, Elias, et al.
Published: (2026)
Pulling Back the Curtain: Unsupervised Adversarial Detection via Contrastive Auxiliary Networks
by: Mizrahi, Eylon, et al.
Published: (2025)
by: Mizrahi, Eylon, et al.
Published: (2025)
Patch of Invisibility: Naturalistic Physical Black-Box Adversarial Attacks on Object Detectors
by: Lapid, Raz, et al.
Published: (2023)
by: Lapid, Raz, et al.
Published: (2023)
Open Sesame! Universal Black Box Jailbreaking of Large Language Models
by: Lapid, Raz, et al.
Published: (2023)
by: Lapid, Raz, et al.
Published: (2023)
On the Robustness of Kolmogorov-Arnold Networks: An Adversarial Perspective
by: Alter, Tal, et al.
Published: (2024)
by: Alter, Tal, et al.
Published: (2024)
Fortify the Guardian, Not the Treasure: Resilient Adversarial Detectors
by: Lapid, Raz, et al.
Published: (2024)
by: Lapid, Raz, et al.
Published: (2024)
Similar Items
-
Silent Tokens, Loud Effects: Padding in LLMs
by: Himelstein, Rom, et al.
Published: (2025) -
Step-Wise Refusal Dynamics in Autoregressive and Diffusion Language Models
by: Rahimi, Eliron, et al.
Published: (2026) -
Silenced Biases: The Dark Side LLMs Learned to Refuse
by: Himelstein, Rom, et al.
Published: (2025) -
Jailbreak Attack Initializations as Extractors of Compliance Directions
by: Levi, Amit, et al.
Published: (2025) -
Extracting Recurring Vulnerabilities from Black-Box LLM-Generated Software
by: Kordonsky, Tomer, et al.
Published: (2026)