OpenELM: An Efficient Language Model Family with Open Training and Inference Framework
Fuente:
arXiv
Saved in:
| Main Authors: | Mehta, Sachin, Sekhavat, Mohammad Hossein, Cao, Qingqing, Horton, Maxwell, Jin, Yanzi, Sun, Chenfan, Mirzadeh, Iman, Najibi, Mahyar, Belenko, Dmitry, Zatloukal, Peter, Rastegari, Mohammad |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CatLIP: CLIP-level Visual Recognition Accuracy with 2.7x Faster Pre-training on Web-scale Image-Text Data
by: Mehta, Sachin, et al.
Published: (2024)
by: Mehta, Sachin, et al.
Published: (2024)
KV Prediction for Improved Time to First Token
by: Horton, Maxwell, et al.
Published: (2024)
by: Horton, Maxwell, et al.
Published: (2024)
Diffusion Models as Masked Audio-Video Learners
by: Nunez, Elvis, et al.
Published: (2023)
by: Nunez, Elvis, et al.
Published: (2023)
CtrlSynth: Controllable Image Text Synthesis for Data-Efficient Multimodal Learning
by: Cao, Qingqing, et al.
Published: (2024)
by: Cao, Qingqing, et al.
Published: (2024)
LazyLLM: Dynamic Token Pruning for Efficient Long Context LLM Inference
by: Fu, Qichen, et al.
Published: (2024)
by: Fu, Qichen, et al.
Published: (2024)
Bytes Are All You Need: Transformers Operating Directly On File Bytes
by: Horton, Maxwell, et al.
Published: (2023)
by: Horton, Maxwell, et al.
Published: (2023)
LLM in a flash: Efficient Large Language Model Inference with Limited Memory
by: Alizadeh, Keivan, et al.
Published: (2023)
by: Alizadeh, Keivan, et al.
Published: (2023)
SeedLM: Compressing LLM Weights into Seeds of Pseudo-Random Generators
by: Shafipour, Rasoul, et al.
Published: (2024)
by: Shafipour, Rasoul, et al.
Published: (2024)
Duo-LLM: A Framework for Studying Adaptive Computation in Large Language Models
by: Alizadeh, Keivan, et al.
Published: (2024)
by: Alizadeh, Keivan, et al.
Published: (2024)
LEAP: Learnable End-to-End Adaptive Pruning of Large Language Models
by: Mozaffari, Mohammad, et al.
Published: (2026)
by: Mozaffari, Mohammad, et al.
Published: (2026)
Superposition Prompting: Improving and Accelerating Retrieval-Augmented Generation
by: Merth, Thomas, et al.
Published: (2024)
by: Merth, Thomas, et al.
Published: (2024)
Efficient Vision-Language Models by Summarizing Visual Tokens into Compact Registers
by: Wen, Yuxin, et al.
Published: (2024)
by: Wen, Yuxin, et al.
Published: (2024)
Speculative Streaming: Fast LLM Inference without Auxiliary Models
by: Bhendawade, Nikhil, et al.
Published: (2024)
by: Bhendawade, Nikhil, et al.
Published: (2024)
Computational Bottlenecks of Training Small-scale Large Language Models
by: Ashkboos, Saleh, et al.
Published: (2024)
by: Ashkboos, Saleh, et al.
Published: (2024)
From Dense to Dynamic: Token-Difficulty Driven MoEfication of Pre-Trained LLMs
by: Nishu, Kumari, et al.
Published: (2025)
by: Nishu, Kumari, et al.
Published: (2025)
An Efficient and Streaming Audio Visual Active Speaker Detection System
by: Kundu, Arnav, et al.
Published: (2024)
by: Kundu, Arnav, et al.
Published: (2024)
The Illusion of Thinking: Understanding the Strengths and Limitations of Reasoning Models via the Lens of Problem Complexity
by: Shojaee, Parshin, et al.
Published: (2025)
by: Shojaee, Parshin, et al.
Published: (2025)
Knowledge Transfer from Vision Foundation Models for Efficient Training of Small Task-specific Models
by: Vemulapalli, Raviteja, et al.
Published: (2023)
by: Vemulapalli, Raviteja, et al.
Published: (2023)
KV-Runahead: Scalable Causal LLM Inference by Parallel Key-Value Cache Generation
by: Cho, Minsik, et al.
Published: (2024)
by: Cho, Minsik, et al.
Published: (2024)
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference
by: Bhendawade, Nikhil, et al.
Published: (2025)
by: Bhendawade, Nikhil, et al.
Published: (2025)
SALSA: Soup-based Alignment Learning for Stronger Adaptation in RLHF
by: Chegini, Atoosa, et al.
Published: (2024)
by: Chegini, Atoosa, et al.
Published: (2024)
The role of shape operator in gauge theories
by: Zatloukal, Vaclav, et al.
Published: (2022)
by: Zatloukal, Vaclav, et al.
Published: (2022)
Open Sourcing GPTs: Economics of Open Sourcing Advanced AI Models
by: Habibi, Mahyar
Published: (2025)
by: Habibi, Mahyar
Published: (2025)
DeFi Liquidation Risk Modeling Using Geometric Brownian Motion
by: Belenko, Timofei, et al.
Published: (2025)
by: Belenko, Timofei, et al.
Published: (2025)
Optical conductivity and orbital magnetization of Floquet vortex states
by: Ahmadabadi, Iman, et al.
Published: (2022)
by: Ahmadabadi, Iman, et al.
Published: (2022)
Achieving Fine‐Grained Microstructure in Low‐Alloy Steel: A Study on Static Recrystallization Using Experimental and Simulation Approaches
by: Mahdiyeh Baharvand, et al.
Published: (2025)
by: Mahdiyeh Baharvand, et al.
Published: (2025)
Open Horn Type Theory
by: Poernomo, Iman
Published: (2025)
by: Poernomo, Iman
Published: (2025)
QuantSpec: Self-Speculative Decoding with Hierarchical Quantized KV Cache
by: Tiwari, Rishabh, et al.
Published: (2025)
by: Tiwari, Rishabh, et al.
Published: (2025)
World-POI: Global Point-of-Interest Data Enriched from Foursquare and OpenStreetMap as Tabular and Graph Data
by: Amiri, Hossein, et al.
Published: (2025)
by: Amiri, Hossein, et al.
Published: (2025)
Global Longitudinal Strain May Be the One that Appropriately Identifies Candidates of ICD Implantation
by: Mohammad Hossein Nikoo, et al.
Published: (2024)
by: Mohammad Hossein Nikoo, et al.
Published: (2024)
SAM-CLIP: Merging Vision Foundation Models towards Semantic and Spatial Understanding
by: Wang, Haoxiang, et al.
Published: (2023)
by: Wang, Haoxiang, et al.
Published: (2023)
Cutaneous Leishmaniasis Mimicking a Nasal Tumor: A Case Report
by: Maryam Hekmat, et al.
Published: (2025)
by: Maryam Hekmat, et al.
Published: (2025)
Overriding Safety protections of Open-source Models
by: Kumar, Sachin
Published: (2024)
by: Kumar, Sachin
Published: (2024)
TiC-LM: A Web-Scale Benchmark for Time-Continual LLM Pretraining
by: Li, Jeffrey, et al.
Published: (2025)
by: Li, Jeffrey, et al.
Published: (2025)
Open Stamped Parts Dataset
by: Antiles, Sarah, et al.
Published: (2024)
by: Antiles, Sarah, et al.
Published: (2024)
APT: Adaptive Pruning and Tuning Pretrained Language Models for Efficient Training and Inference
by: Zhao, Bowen, et al.
Published: (2024)
by: Zhao, Bowen, et al.
Published: (2024)
An Open-Source Framework for Coupled Vehicle-Bridge Interaction Analysis Using OpenSees
by: Talebi-Kalaleh, Mohammad, et al.
Published: (2026)
by: Talebi-Kalaleh, Mohammad, et al.
Published: (2026)
Gray-Box Computed Torque Control for Differential-Drive Mobile Robot Tracking
by: Pishkhani, Arman Javan Sekhavat
Published: (2025)
by: Pishkhani, Arman Javan Sekhavat
Published: (2025)
Barriers for Learning in an Evolving World: Mathematical Understanding of Loss of Plasticity
by: Joudaki, Amir, et al.
Published: (2025)
by: Joudaki, Amir, et al.
Published: (2025)
Scaling Smart: Accelerating Large Language Model Pre-training with Small Model Initialization
by: Samragh, Mohammad, et al.
Published: (2024)
by: Samragh, Mohammad, et al.
Published: (2024)
Similar Items
-
CatLIP: CLIP-level Visual Recognition Accuracy with 2.7x Faster Pre-training on Web-scale Image-Text Data
by: Mehta, Sachin, et al.
Published: (2024) -
KV Prediction for Improved Time to First Token
by: Horton, Maxwell, et al.
Published: (2024) -
Diffusion Models as Masked Audio-Video Learners
by: Nunez, Elvis, et al.
Published: (2023) -
CtrlSynth: Controllable Image Text Synthesis for Data-Efficient Multimodal Learning
by: Cao, Qingqing, et al.
Published: (2024) -
LazyLLM: Dynamic Token Pruning for Efficient Long Context LLM Inference
by: Fu, Qichen, et al.
Published: (2024)