LLM Compression: How Far Can We Go in Balancing Size and Performance?
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Sk, Sahil, Dhal, Debasish, Khosla, Sonal, Shahid, Sk, Shekhar, Sambit, Dhaka, Akash, Parida, Shantipriya, Prasad, Dilip K., Bojar, Ondřej |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Towards a More Inclusive AI: Progress and Perspectives in Large Language Model Training for the Sámi Language
par: Paul, Ronny, et autres
Publié: (2024)
par: Paul, Ronny, et autres
Publié: (2024)
Building pre-train LLM Dataset for the INDIC Languages: a case study on Hindi
par: Parida, Shantipriya, et autres
Publié: (2024)
par: Parida, Shantipriya, et autres
Publié: (2024)
Breaking the Degeneracy of Sense Codons – How Far Can We Go?
par: Clark A. Jones, et autres
Publié: (2024)
par: Clark A. Jones, et autres
Publié: (2024)
How Far Can We Go with Practical Function-Level Program Repair?
par: Xiang, Jiahong, et autres
Publié: (2024)
par: Xiang, Jiahong, et autres
Publié: (2024)
Probing Light Particles With Optically Trapped Sensors Through Nucleon Scattering
par: Dutta, Bhaskar, et autres
Publié: (2025)
par: Dutta, Bhaskar, et autres
Publié: (2025)
Long lived inert Higgs in fast expanding universe and its imprint on cosmic microwave background
par: Ghosh, Dilip Kumar, et autres
Publié: (2022)
par: Ghosh, Dilip Kumar, et autres
Publié: (2022)
Chapter ‘We Go as Far as We Can without Losing Sleep Over It’
par: Schmid, Evi | https://orcid.org/0000-0003-0364-8390
Publié: (2026)
par: Schmid, Evi | https://orcid.org/0000-0003-0364-8390
Publié: (2026)
How Far Can We Compress Instant-NGP-Based NeRF?
par: Chen, Yihang, et autres
Publié: (2024)
par: Chen, Yihang, et autres
Publié: (2024)
Information availability in different languages and various technological constraints related to multilinguism on the Internet
par: Khosla, Sonal, et autres
Publié: (2025)
par: Khosla, Sonal, et autres
Publié: (2025)
Runtime Failure Hunting for Physics Engine Based Software Systems: How Far Can We Go?
par: Li, Shuqing, et autres
Publié: (2025)
par: Li, Shuqing, et autres
Publié: (2025)
Agent Factories for High Level Synthesis: How Far Can General-Purpose Coding Agents Go in Hardware Optimization?
par: Bhandwaldar, Abhishek, et autres
Publié: (2026)
par: Bhandwaldar, Abhishek, et autres
Publié: (2026)
Finetuning LLMs for EvaCun 2025 token prediction shared task
par: Jon, Josef, et autres
Publié: (2025)
par: Jon, Josef, et autres
Publié: (2025)
Intrinsic vs. Extrinsic Evaluation of Czech Sentence Embeddings: Semantic Relevance Doesn't Help with MT Evaluation
par: Barančíková, Petra, et autres
Publié: (2025)
par: Barančíková, Petra, et autres
Publié: (2025)
Overview of the Sensemaking Task at the ELOQUENT 2025 Lab: LLMs as Teachers, Students and Evaluators
par: Šindelář, Pavel, et autres
Publié: (2025)
par: Šindelář, Pavel, et autres
Publié: (2025)
Understanding the role of FFNs in driving multilingual behaviour in LLMs
par: Bhattacharya, Sunit, et autres
Publié: (2024)
par: Bhattacharya, Sunit, et autres
Publié: (2024)
Long-Form End-to-End Speech Translation via Latent Alignment Segmentation
par: Polák, Peter, et autres
Publié: (2023)
par: Polák, Peter, et autres
Publié: (2023)
Quality and Quantity of Machine Translation References for Automatic Metrics
par: Zouhar, Vilém, et autres
Publié: (2024)
par: Zouhar, Vilém, et autres
Publié: (2024)
End-to-end Automatic Speech Recognition and Speech Translation: Integration of Speech Foundational Models and LLMs
par: Luu, Nam, et autres
Publié: (2025)
par: Luu, Nam, et autres
Publié: (2025)
Characterization of fractional Sobolev--Poincaré and (localized) Hardy inequalities
par: Sk, Firoj
Publié: (2022)
par: Sk, Firoj
Publié: (2022)
Blazar Boosted ALP and vector portal Dark matter confronting light mediator searches
par: Jeesun, Sk
Publié: (2025)
par: Jeesun, Sk
Publié: (2025)
Development of a Deep Learning-Driven Control Framework for Exoskeleton Robots
par: Hasan, Sk
Publié: (2022)
par: Hasan, Sk
Publié: (2022)
How Far Can You Go with Your Studies at UNED?
par: Cristina Gutiérrez-Carranza
Publié: (2020)
par: Cristina Gutiérrez-Carranza
Publié: (2020)
The $N_{\rm eff}$ at CMB challenges $U(1)_X$ light gauge boson scenarios
par: Ghosh, Dilip Kumar, et autres
Publié: (2024)
par: Ghosh, Dilip Kumar, et autres
Publié: (2024)
ALP and $Z^\prime$ boson at the Electron-Ion collider
par: Adhikary, Amit, et autres
Publié: (2026)
par: Adhikary, Amit, et autres
Publié: (2026)
Hubble Tension and Cosmological Imprints of $U(1)_X$ Gauge Symmetry: $U(1)_{B_3-3 L_i}$ as a case study
par: Ghosh, Dilip Kumar, et autres
Publié: (2023)
par: Ghosh, Dilip Kumar, et autres
Publié: (2023)
Hairy Black Holes: Non-existence of Short Hairs and Bound on Light Ring Size
par: Ghosh, Rajes, et autres
Publié: (2023)
par: Ghosh, Rajes, et autres
Publié: (2023)
Can Non-Relativistic Strings Propagate Without Geometric Baggage?
par: Nandi, Partha, et autres
Publié: (2025)
par: Nandi, Partha, et autres
Publié: (2025)
How "Real" is Your Real-Time Simultaneous Speech-to-Text Translation System?
par: Papi, Sara, et autres
Publié: (2024)
par: Papi, Sara, et autres
Publié: (2024)
Financial Named Entity Recognition: How Far Can LLM Go?
par: Lu, Yi-Te, et autres
Publié: (2025)
par: Lu, Yi-Te, et autres
Publié: (2025)
Implying Volatility: How Fast Can We Go?
par: Floc'h, Fabien Le, et autres
Publié: (2026)
par: Floc'h, Fabien Le, et autres
Publié: (2026)
Can AI Agents Generate Microservices? How Far are We?
par: Adnan, Bassam, et autres
Publié: (2026)
par: Adnan, Bassam, et autres
Publié: (2026)
Comparative Evaluation of Phase Change Materials and Fins in Battery Thermal Management During High Discharge
par: Sk Mohammad Shareef, et autres
Publié: (2025)
par: Sk Mohammad Shareef, et autres
Publié: (2025)
Multipacking on graphs and Euclidean metric space
par: Islam, Sk Samim
Publié: (2026)
par: Islam, Sk Samim
Publié: (2026)
“Midnight’s Untouchable Children” and their Struggle for Existence: A Study of Bangla Dalit Poetry of Manohar Mouli Biswas in English Translation
par: Md Humayun Sk
Publié: (2016)
par: Md Humayun Sk
Publié: (2016)
How Far Can In-Context Alignment Go? Exploring the State of In-Context Alignment
par: Huang, Heyan, et autres
Publié: (2024)
par: Huang, Heyan, et autres
Publié: (2024)
How Far Can We Go with Pixels Alone? A Pilot Study on Screen-Only Navigation in Commercial 3D ARPGs
par: Xu, Kaijie, et autres
Publié: (2026)
par: Xu, Kaijie, et autres
Publié: (2026)
How Far Can We Trust Chaos? Extending the Horizon of Predictability
par: Angelidis, Alexandros K., et autres
Publié: (2025)
par: Angelidis, Alexandros K., et autres
Publié: (2025)
Continuous Rating as Reliable Human Evaluation of Simultaneous Speech Translation
par: Javorský, Dávid, et autres
Publié: (2022)
par: Javorský, Dávid, et autres
Publié: (2022)
MEEDAV: A Synchronous Web Viewer for EEG, Eye-Tracking and Speech Data
par: Pijálek, Jan, et autres
Publié: (2026)
par: Pijálek, Jan, et autres
Publié: (2026)
Prompting LLMs: Length Control for Isometric Machine Translation
par: Javorský, Dávid, et autres
Publié: (2025)
par: Javorský, Dávid, et autres
Publié: (2025)
Documents similaires
-
Towards a More Inclusive AI: Progress and Perspectives in Large Language Model Training for the Sámi Language
par: Paul, Ronny, et autres
Publié: (2024) -
Building pre-train LLM Dataset for the INDIC Languages: a case study on Hindi
par: Parida, Shantipriya, et autres
Publié: (2024) -
Breaking the Degeneracy of Sense Codons – How Far Can We Go?
par: Clark A. Jones, et autres
Publié: (2024) -
How Far Can We Go with Practical Function-Level Program Repair?
par: Xiang, Jiahong, et autres
Publié: (2024) -
Probing Light Particles With Optically Trapped Sensors Through Nucleon Scattering
par: Dutta, Bhaskar, et autres
Publié: (2025)