Just CHOP: Embarrassingly Simple LLM Compression
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jha, Ananya Harsh, Sherborne, Tom, Walsh, Evan Pete, Groeneveld, Dirk, Strubell, Emma, Beltagy, Iz |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Scalable Data Ablation Approximations for Language Models through Modular Training and Merging
von: Na, Clara, et al.
Veröffentlicht: (2024)
von: Na, Clara, et al.
Veröffentlicht: (2024)
Source-Aware Training Enables Knowledge Attribution in Language Models
von: Khalifa, Muhammad, et al.
Veröffentlicht: (2024)
von: Khalifa, Muhammad, et al.
Veröffentlicht: (2024)
Paloma: A Benchmark for Evaluating Language Model Fit
von: Magnusson, Ian, et al.
Veröffentlicht: (2023)
von: Magnusson, Ian, et al.
Veröffentlicht: (2023)
An Embarrassingly Simple Defense Against LLM Abliteration Attacks
von: Shairah, Harethah Abu, et al.
Veröffentlicht: (2025)
von: Shairah, Harethah Abu, et al.
Veröffentlicht: (2025)
An Embarrassingly Simple Approach for LLM with Strong ASR Capacity
von: Ma, Ziyang, et al.
Veröffentlicht: (2024)
von: Ma, Ziyang, et al.
Veröffentlicht: (2024)
Embarrassingly Simple Self-Distillation Improves Code Generation
von: Zhang, Ruixiang, et al.
Veröffentlicht: (2026)
von: Zhang, Ruixiang, et al.
Veröffentlicht: (2026)
Embarrassingly Simple Unsupervised Aspect Based Sentiment Tuple Extraction
von: Scaria, Kevin, et al.
Veröffentlicht: (2024)
von: Scaria, Kevin, et al.
Veröffentlicht: (2024)
Dolma: an Open Corpus of Three Trillion Tokens for Language Model Pretraining Research
von: Soldaini, Luca, et al.
Veröffentlicht: (2024)
von: Soldaini, Luca, et al.
Veröffentlicht: (2024)
Compositional Generalisation for Explainable Hate Speech Detection
von: Calabrese, Agostina, et al.
Veröffentlicht: (2025)
von: Calabrese, Agostina, et al.
Veröffentlicht: (2025)
Gradient Localization Improves Lifelong Pretraining of Language Models
von: Fernandez, Jared, et al.
Veröffentlicht: (2024)
von: Fernandez, Jared, et al.
Veröffentlicht: (2024)
TESS: Text-to-Text Self-Conditioned Simplex Diffusion
von: Mahabadi, Rabeeh Karimi, et al.
Veröffentlicht: (2023)
von: Mahabadi, Rabeeh Karimi, et al.
Veröffentlicht: (2023)
TRAM: Bridging Trust Regions and Sharpness Aware Minimization
von: Sherborne, Tom, et al.
Veröffentlicht: (2023)
von: Sherborne, Tom, et al.
Veröffentlicht: (2023)
TouchTTS: An Embarrassingly Simple TTS Framework that Everyone Can Touch
von: Song, Xingchen, et al.
Veröffentlicht: (2024)
von: Song, Xingchen, et al.
Veröffentlicht: (2024)
OLMo: Accelerating the Science of Language Models
von: Groeneveld, Dirk, et al.
Veröffentlicht: (2024)
von: Groeneveld, Dirk, et al.
Veröffentlicht: (2024)
Beyond Text: Characterizing Domain Expert Needs in Document Research
von: Gururaja, Sireesh, et al.
Veröffentlicht: (2025)
von: Gururaja, Sireesh, et al.
Veröffentlicht: (2025)
Agents Explore but Agents Ignore: LLMs Lack Environmental Curiosity
von: Engländer, Leon, et al.
Veröffentlicht: (2026)
von: Engländer, Leon, et al.
Veröffentlicht: (2026)
FicSim: A Dataset for Multi-Faceted Semantic Similarity in Long-Form Fiction
von: Johnson, Natasha, et al.
Veröffentlicht: (2025)
von: Johnson, Natasha, et al.
Veröffentlicht: (2025)
Establishing Task Scaling Laws via Compute-Efficient Model Ladders
von: Bhagia, Akshita, et al.
Veröffentlicht: (2024)
von: Bhagia, Akshita, et al.
Veröffentlicht: (2024)
JustRL: Scaling a 1.5B LLM with a Simple RL Recipe
von: He, Bingxiang, et al.
Veröffentlicht: (2025)
von: He, Bingxiang, et al.
Veröffentlicht: (2025)
Stereotype or Personalization? User Identity Biases Chatbot Recommendations
von: Kantharuban, Anjali, et al.
Veröffentlicht: (2024)
von: Kantharuban, Anjali, et al.
Veröffentlicht: (2024)
To Build Our Future, We Must Know Our Past: Contextualizing Paradigm Shifts in Natural Language Processing
von: Gururaja, Sireesh, et al.
Veröffentlicht: (2023)
von: Gururaja, Sireesh, et al.
Veröffentlicht: (2023)
HalluSearch at SemEval-2025 Task 3: A Search-Enhanced RAG Pipeline for Hallucination Detection
von: Abdallah, Mohamed A., et al.
Veröffentlicht: (2025)
von: Abdallah, Mohamed A., et al.
Veröffentlicht: (2025)
Exploring Retrieval Augmented Generation in Arabic
von: El-Beltagy, Samhaa R., et al.
Veröffentlicht: (2024)
von: El-Beltagy, Samhaa R., et al.
Veröffentlicht: (2024)
Energy and Carbon Considerations of Fine-Tuning BERT
von: Wang, Xiaorong, et al.
Veröffentlicht: (2023)
von: Wang, Xiaorong, et al.
Veröffentlicht: (2023)
Embarrassed to observe: The effects of directive language in brand conversation
von: Andriuzzi, Andria, et al.
Veröffentlicht: (2025)
von: Andriuzzi, Andria, et al.
Veröffentlicht: (2025)
CHOP: Chunkwise Context-Preserving Framework for RAG on Multi Documents
von: Park, Hyunseok, et al.
Veröffentlicht: (2026)
von: Park, Hyunseok, et al.
Veröffentlicht: (2026)
Demo: TOSense -- What Did You Just Agree to?
von: Chen, Xinzhang, et al.
Veröffentlicht: (2025)
von: Chen, Xinzhang, et al.
Veröffentlicht: (2025)
Optical Context Compression Is Just (Bad) Autoencoding
von: Lee, Ivan Yee, et al.
Veröffentlicht: (2025)
von: Lee, Ivan Yee, et al.
Veröffentlicht: (2025)
Energy Considerations of Large Language Model Inference and Efficiency Optimizations
von: Fernandez, Jared, et al.
Veröffentlicht: (2025)
von: Fernandez, Jared, et al.
Veröffentlicht: (2025)
Kinetics: Rethinking Test-Time Scaling Laws
von: Sadhukhan, Ranajoy, et al.
Veröffentlicht: (2025)
von: Sadhukhan, Ranajoy, et al.
Veröffentlicht: (2025)
If You Can't Use Them, Recycle Them: Optimizing Merging at Scale Mitigates Performance Tradeoffs
von: Khalifa, Muhammad, et al.
Veröffentlicht: (2024)
von: Khalifa, Muhammad, et al.
Veröffentlicht: (2024)
SpreadsheetArena: Decomposing Preference in LLM Generation of Spreadsheet Workbooks
von: Kundurthy, Srivatsa, et al.
Veröffentlicht: (2026)
von: Kundurthy, Srivatsa, et al.
Veröffentlicht: (2026)
What's In My Big Data?
von: Elazar, Yanai, et al.
Veröffentlicht: (2023)
von: Elazar, Yanai, et al.
Veröffentlicht: (2023)
AboutMe: Using Self-Descriptions in Webpages to Document the Effects of English Pretraining Data Filters
von: Lucy, Li, et al.
Veröffentlicht: (2024)
von: Lucy, Li, et al.
Veröffentlicht: (2024)
Moonshine v2: Ergodic Streaming Encoder ASR for Latency-Critical Speech Applications
von: Kudlur, Manjunath, et al.
Veröffentlicht: (2026)
von: Kudlur, Manjunath, et al.
Veröffentlicht: (2026)
On Leakage of Code Generation Evaluation Datasets
von: Matton, Alexandre, et al.
Veröffentlicht: (2024)
von: Matton, Alexandre, et al.
Veröffentlicht: (2024)
Reason-KE++: Aligning the Process, Not Just the Outcome, for Faithful LLM Knowledge Editing
von: Wu, Yuchen, et al.
Veröffentlicht: (2025)
von: Wu, Yuchen, et al.
Veröffentlicht: (2025)
Flavors of Moonshine: Tiny Specialized ASR Models for Edge Devices
von: King, Evan, et al.
Veröffentlicht: (2025)
von: King, Evan, et al.
Veröffentlicht: (2025)
Compressing LLMs: The Truth is Rarely Pure and Never Simple
von: Jaiswal, Ajay, et al.
Veröffentlicht: (2023)
von: Jaiswal, Ajay, et al.
Veröffentlicht: (2023)
Reasoning Path Compression: Compressing Generation Trajectories for Efficient LLM Reasoning
von: Song, Jiwon, et al.
Veröffentlicht: (2025)
von: Song, Jiwon, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Scalable Data Ablation Approximations for Language Models through Modular Training and Merging
von: Na, Clara, et al.
Veröffentlicht: (2024) -
Source-Aware Training Enables Knowledge Attribution in Language Models
von: Khalifa, Muhammad, et al.
Veröffentlicht: (2024) -
Paloma: A Benchmark for Evaluating Language Model Fit
von: Magnusson, Ian, et al.
Veröffentlicht: (2023) -
An Embarrassingly Simple Defense Against LLM Abliteration Attacks
von: Shairah, Harethah Abu, et al.
Veröffentlicht: (2025) -
An Embarrassingly Simple Approach for LLM with Strong ASR Capacity
von: Ma, Ziyang, et al.
Veröffentlicht: (2024)