Position: Enough of Scaling LLMs! Lets Focus on Downscaling
Fuente:
arXiv
Saved in:
| Main Authors: | Goel, Yash, Sengupta, Ayan, Chakraborty, Tanmoy |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
How to Upscale Neural Networks with Scaling Law? A Survey and Practical Guidelines
by: Sengupta, Ayan, et al.
Published: (2025)
by: Sengupta, Ayan, et al.
Published: (2025)
Value-Guided KV Compression for LLMs via Approximated CUR Decomposition
by: Sengupta, Ayan, et al.
Published: (2025)
by: Sengupta, Ayan, et al.
Published: (2025)
Understanding the Physics of Key-Value Cache Compression for LLMs through Attention Dynamics
by: Ananthanarayanan, Samhruth, et al.
Published: (2026)
by: Ananthanarayanan, Samhruth, et al.
Published: (2026)
The Art of Scaling Test-Time Compute for Large Language Models
by: Agarwal, Aradhye, et al.
Published: (2025)
by: Agarwal, Aradhye, et al.
Published: (2025)
First Finish Search: Efficient Test-Time Scaling in Large Language Models
by: Agarwal, Aradhye, et al.
Published: (2025)
by: Agarwal, Aradhye, et al.
Published: (2025)
Compression Laws for Large Language Models
by: Sengupta, Ayan, et al.
Published: (2025)
by: Sengupta, Ayan, et al.
Published: (2025)
You Only Prune Once: Designing Calibration-Free Model Compression With Policy Learning
by: Sengupta, Ayan, et al.
Published: (2025)
by: Sengupta, Ayan, et al.
Published: (2025)
HIDE and Seek: Detecting Hallucinations in Language Models via Decoupled Representations
by: Chatterjee, Anwoy, et al.
Published: (2025)
by: Chatterjee, Anwoy, et al.
Published: (2025)
On the Generalization vs Fidelity Paradox in Knowledge Distillation
by: Ramesh, Suhas Kamasetty, et al.
Published: (2025)
by: Ramesh, Suhas Kamasetty, et al.
Published: (2025)
Persona-aware Generative Model for Code-mixed Language
by: Sengupta, Ayan, et al.
Published: (2023)
by: Sengupta, Ayan, et al.
Published: (2023)
From Images to Words: Efficient Cross-Modal Knowledge Distillation to Language Models from Black-box Teachers
by: Sengupta, Ayan, et al.
Published: (2026)
by: Sengupta, Ayan, et al.
Published: (2026)
Step-by-Step Unmasking for Parameter-Efficient Fine-tuning of Large Language Models
by: Agarwal, Aradhye, et al.
Published: (2024)
by: Agarwal, Aradhye, et al.
Published: (2024)
Robust and Efficient Fine-tuning of LLMs with Bayesian Reparameterization of Low-Rank Adaptation
by: Sengupta, Ayan, et al.
Published: (2024)
by: Sengupta, Ayan, et al.
Published: (2024)
Multilingual Test-Time Scaling via Initial Thought Transfer
by: Bajpai, Prasoon, et al.
Published: (2025)
by: Bajpai, Prasoon, et al.
Published: (2025)
Multilingual LLMs Struggle to Link Orthography and Semantics in Bilingual Word Processing
by: Tanwar, Eshaan, et al.
Published: (2025)
by: Tanwar, Eshaan, et al.
Published: (2025)
Temporal Referential Consistency: Do LLMs Favor Sequences Over Absolute Time References?
by: Bajpai, Ashutosh, et al.
Published: (2025)
by: Bajpai, Ashutosh, et al.
Published: (2025)
Understanding the Effects of Domain Finetuning on LLMs
by: Tanwar, Eshaan, et al.
Published: (2025)
by: Tanwar, Eshaan, et al.
Published: (2025)
Can LLMs replace Neil deGrasse Tyson? Evaluating the Reliability of LLMs as Science Communicators
by: Bajpai, Prasoon, et al.
Published: (2024)
by: Bajpai, Prasoon, et al.
Published: (2024)
Multilingual LLMs Inherently Reward In-Language Time-Sensitive Semantic Alignment for Low-Resource Languages
by: Bajpai, Ashutosh, et al.
Published: (2024)
by: Bajpai, Ashutosh, et al.
Published: (2024)
Harmonizing Code-mixed Conversations: Personality-assisted Code-mixed Response Generation in Dialogues
by: Kumar, Shivani, et al.
Published: (2024)
by: Kumar, Shivani, et al.
Published: (2024)
Decoding Memes: Benchmarking Narrative Role Classification across Multilingual and Multimodal Models
by: Sharma, Shivam, et al.
Published: (2025)
by: Sharma, Shivam, et al.
Published: (2025)
Can LLMs reason over extended multilingual contexts? Towards long-context evaluation beyond retrieval and haystacks
by: Hengle, Amey, et al.
Published: (2025)
by: Hengle, Amey, et al.
Published: (2025)
Latent Performance Profiling of Large Language Models
by: Chakraborty, Tanmoy, et al.
Published: (2026)
by: Chakraborty, Tanmoy, et al.
Published: (2026)
Hate Personified: Investigating the role of LLMs in content moderation
by: Masud, Sarah, et al.
Published: (2024)
by: Masud, Sarah, et al.
Published: (2024)
Rank, Chunk and Expand: Lineage-Oriented Reasoning for Taxonomy Expansion
by: Mishra, Sahil, et al.
Published: (2025)
by: Mishra, Sahil, et al.
Published: (2025)
Innocence in the Crossfire: Roles of Skip Connections in Jailbreaking Visual Language Models
by: Nandi, Palash, et al.
Published: (2025)
by: Nandi, Palash, et al.
Published: (2025)
Counterspeech the ultimate shield! Multi-Conditioned Counterspeech Generation through Attributed Prefix Learning
by: Kumar, Aswini, et al.
Published: (2025)
by: Kumar, Aswini, et al.
Published: (2025)
Markovian ODE-guided scoring can assess the quality of offline reasoning traces in language models
by: Nandi, Arghodeep, et al.
Published: (2026)
by: Nandi, Arghodeep, et al.
Published: (2026)
Information Anxiety in Large Language Models
by: Bajpai, Prasoon, et al.
Published: (2024)
by: Bajpai, Prasoon, et al.
Published: (2024)
Exposing Long-Tail Safety Failures in Large Language Models through Efficient Diverse Response Sampling
by: Hajra, Suvadeep, et al.
Published: (2026)
by: Hajra, Suvadeep, et al.
Published: (2026)
SAFE-MEME: Structured Reasoning Framework for Robust Hate Speech Detection in Memes
by: Nandi, Palash, et al.
Published: (2024)
by: Nandi, Palash, et al.
Published: (2024)
Focal Inferential Infusion Coupled with Tractable Density Discrimination for Implicit Hate Detection
by: Masud, Sarah, et al.
Published: (2023)
by: Masud, Sarah, et al.
Published: (2023)
CSEval: Towards Automated, Multi-Dimensional, and Reference-Free Counterspeech Evaluation using Auto-Calibrated LLMs
by: Hengle, Amey, et al.
Published: (2025)
by: Hengle, Amey, et al.
Published: (2025)
coTherapist: A Behavior-Aligned Small Language Model to Support Mental Healthcare Experts
by: Adhikary, Prottay Kumar, et al.
Published: (2026)
by: Adhikary, Prottay Kumar, et al.
Published: (2026)
Let's Focus on Neuron: Neuron-Level Supervised Fine-tuning for Large Language Model
by: Xu, Haoyun, et al.
Published: (2024)
by: Xu, Haoyun, et al.
Published: (2024)
Agentic AutoSurvey: Let LLMs Survey LLMs
by: Liu, Yixin, et al.
Published: (2025)
by: Liu, Yixin, et al.
Published: (2025)
Towards Privacy-aware Mental Health AI Models: Advances, Challenges, and Opportunities
by: Mandal, Aishik, et al.
Published: (2025)
by: Mandal, Aishik, et al.
Published: (2025)
MAGneT: Coordinated Multi-Agent Generation of Synthetic Multi-Turn Mental Health Counseling Sessions
by: Mandal, Aishik, et al.
Published: (2025)
by: Mandal, Aishik, et al.
Published: (2025)
SABER: Uncovering Vulnerabilities in Safety Alignment via Cross-Layer Residual Connection
by: Joshi, Maithili, et al.
Published: (2025)
by: Joshi, Maithili, et al.
Published: (2025)
$\texttt{LM}^\texttt{2}$: A Simple Society of Language Models Solves Complex Reasoning
by: Juneja, Gurusha, et al.
Published: (2024)
by: Juneja, Gurusha, et al.
Published: (2024)
Similar Items
-
How to Upscale Neural Networks with Scaling Law? A Survey and Practical Guidelines
by: Sengupta, Ayan, et al.
Published: (2025) -
Value-Guided KV Compression for LLMs via Approximated CUR Decomposition
by: Sengupta, Ayan, et al.
Published: (2025) -
Understanding the Physics of Key-Value Cache Compression for LLMs through Attention Dynamics
by: Ananthanarayanan, Samhruth, et al.
Published: (2026) -
The Art of Scaling Test-Time Compute for Large Language Models
by: Agarwal, Aradhye, et al.
Published: (2025) -
First Finish Search: Efficient Test-Time Scaling in Large Language Models
by: Agarwal, Aradhye, et al.
Published: (2025)