Gespeichert in:
| Hauptverfasser: | Sundaram, Jainaveen, Iyer, Ravi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2408.13402 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Bitnet.cpp: Efficient Edge Inference for Ternary LLMs
von: Wang, Jinheng, et al.
Veröffentlicht: (2025)
von: Wang, Jinheng, et al.
Veröffentlicht: (2025)
2 OLMo 2 Furious
von: OLMo, Team, et al.
Veröffentlicht: (2024)
von: OLMo, Team, et al.
Veröffentlicht: (2024)
OLMoE: Open Mixture-of-Experts Language Models
von: Muennighoff, Niklas, et al.
Veröffentlicht: (2024)
von: Muennighoff, Niklas, et al.
Veröffentlicht: (2024)
OLMoASR: Open Models and Data for Training Robust Speech Recognition Models
von: Ngo, Huong, et al.
Veröffentlicht: (2025)
von: Ngo, Huong, et al.
Veröffentlicht: (2025)
Generation of Human Comprehensible Access Control Policies from Audit Logs
von: Kumar, Gautam, et al.
Veröffentlicht: (2026)
von: Kumar, Gautam, et al.
Veröffentlicht: (2026)
Graph Persistence goes Spectral
von: Ji, Mattie, et al.
Veröffentlicht: (2025)
von: Ji, Mattie, et al.
Veröffentlicht: (2025)
TernaryLLM: Ternarized Large Language Model
von: Chen, Tianqi, et al.
Veröffentlicht: (2024)
von: Chen, Tianqi, et al.
Veröffentlicht: (2024)
Understanding the Effect of Noise in LLM Training Data with Algorithmic Chains of Thought
von: Havrilla, Alex, et al.
Veröffentlicht: (2024)
von: Havrilla, Alex, et al.
Veröffentlicht: (2024)
SMA: Submodular Modality Aligner For Data Efficient Multimodal Learning
von: Pham, Truong, et al.
Veröffentlicht: (2026)
von: Pham, Truong, et al.
Veröffentlicht: (2026)
TinyLLaVA: A Framework of Small-scale Large Multimodal Models
von: Zhou, Baichuan, et al.
Veröffentlicht: (2024)
von: Zhou, Baichuan, et al.
Veröffentlicht: (2024)
TinyLLaVA Factory: A Modularized Codebase for Small-scale Large Multimodal Models
von: Jia, Junlong, et al.
Veröffentlicht: (2024)
von: Jia, Junlong, et al.
Veröffentlicht: (2024)
State Contamination in Memory-Augmented LLM Agents
von: Wang, Yian, et al.
Veröffentlicht: (2026)
von: Wang, Yian, et al.
Veröffentlicht: (2026)
The Fourth State: Signed-Zero Ternary for Stable LLM Quantization (and More)
von: Uhlmann, Jeffrey
Veröffentlicht: (2025)
von: Uhlmann, Jeffrey
Veröffentlicht: (2025)
LLaPipe: LLM-Guided Reinforcement Learning for Automated Data Preparation Pipeline Construction
von: Chang, Jing, et al.
Veröffentlicht: (2025)
von: Chang, Jing, et al.
Veröffentlicht: (2025)
LogLLaMA: Transformer-based log anomaly detection with LLaMA
von: Yang, Zhuoyi, et al.
Veröffentlicht: (2025)
von: Yang, Zhuoyi, et al.
Veröffentlicht: (2025)
ECHO-LLaMA: Efficient Caching for High-Performance LLaMA Training
von: Dialameh, Maryam, et al.
Veröffentlicht: (2025)
von: Dialameh, Maryam, et al.
Veröffentlicht: (2025)
KaVa: Latent Reasoning via Compressed KV-Cache Distillation
von: Kuzina, Anna, et al.
Veröffentlicht: (2025)
von: Kuzina, Anna, et al.
Veröffentlicht: (2025)
Single-Stage Huffman Encoder for ML Compression
von: Agrawal, Aditya, et al.
Veröffentlicht: (2026)
von: Agrawal, Aditya, et al.
Veröffentlicht: (2026)
VaCDA: Variational Contrastive Alignment-based Scalable Human Activity Recognition
von: Khisa, Soham, et al.
Veröffentlicht: (2025)
von: Khisa, Soham, et al.
Veröffentlicht: (2025)
AstroLLaMA-Chat: Scaling AstroLLaMA with Conversational and Diverse Datasets
von: Perkowski, Ernest, et al.
Veröffentlicht: (2024)
von: Perkowski, Ernest, et al.
Veröffentlicht: (2024)
FairyFuse: Multiplication-Free LLM Inference on CPUs via Fused Ternary Kernels
von: Zuo, Fei, et al.
Veröffentlicht: (2026)
von: Zuo, Fei, et al.
Veröffentlicht: (2026)
MLLM-LLaVA-FL: Multimodal Large Language Model Assisted Federated Learning
von: Zhang, Jianyi, et al.
Veröffentlicht: (2024)
von: Zhang, Jianyi, et al.
Veröffentlicht: (2024)
HiDe-LLaVA: Hierarchical Decoupling for Continual Instruction Tuning of Multimodal Large Language Model
von: Guo, Haiyang, et al.
Veröffentlicht: (2025)
von: Guo, Haiyang, et al.
Veröffentlicht: (2025)
DeRDaVa: Deletion-Robust Data Valuation for Machine Learning
von: Tian, Xiao, et al.
Veröffentlicht: (2023)
von: Tian, Xiao, et al.
Veröffentlicht: (2023)
Better Together: Leveraging Unpaired Multimodal Data for Stronger Unimodal Models
von: Gupta, Sharut, et al.
Veröffentlicht: (2025)
von: Gupta, Sharut, et al.
Veröffentlicht: (2025)
Apparate: Rethinking Early Exits to Tame Latency-Throughput Tensions in ML Serving
von: Dai, Yinwei, et al.
Veröffentlicht: (2023)
von: Dai, Yinwei, et al.
Veröffentlicht: (2023)
Quad Length Codes for Lossless Compression of e4m3
von: Agrawal, Aditya, et al.
Veröffentlicht: (2026)
von: Agrawal, Aditya, et al.
Veröffentlicht: (2026)
Classification with a Network of Partially Informative Agents: Enabling Wise Crowds from Individually Myopic Classifiers
von: Yao, Tong, et al.
Veröffentlicht: (2024)
von: Yao, Tong, et al.
Veröffentlicht: (2024)
Transformer-based CoVaR: Systemic Risk in Textual Information
von: Chen, Junyu, et al.
Veröffentlicht: (2026)
von: Chen, Junyu, et al.
Veröffentlicht: (2026)
FiSH: Fair Spatial Hotspots
von: P, Deepak, et al.
Veröffentlicht: (2021)
von: P, Deepak, et al.
Veröffentlicht: (2021)
BaKlaVa -- Budgeted Allocation of KV cache for Long-context Inference
von: Gulhan, Ahmed Burak, et al.
Veröffentlicht: (2025)
von: Gulhan, Ahmed Burak, et al.
Veröffentlicht: (2025)
The Uniqueness of LLaMA3-70B Series with Per-Channel Quantization
von: Qin, Minghai
Veröffentlicht: (2024)
von: Qin, Minghai
Veröffentlicht: (2024)
VaPR -- Vision-language Preference alignment for Reasoning
von: Wadhawan, Rohan, et al.
Veröffentlicht: (2025)
von: Wadhawan, Rohan, et al.
Veröffentlicht: (2025)
MoTE: Mixture of Ternary Experts for Memory-efficient Large Multimodal Models
von: Wang, Hongyu, et al.
Veröffentlicht: (2025)
von: Wang, Hongyu, et al.
Veröffentlicht: (2025)
TeLLMe: An Energy-Efficient Ternary LLM Accelerator for Prefilling and Decoding on Edge FPGAs
von: Qiao, Ye, et al.
Veröffentlicht: (2025)
von: Qiao, Ye, et al.
Veröffentlicht: (2025)
LLaGA: Large Language and Graph Assistant
von: Chen, Runjin, et al.
Veröffentlicht: (2024)
von: Chen, Runjin, et al.
Veröffentlicht: (2024)
Dynamic Activation Pitfalls in LLaMA Models: An Empirical Study
von: Ma, Chi, et al.
Veröffentlicht: (2024)
von: Ma, Chi, et al.
Veröffentlicht: (2024)
Learning State-Space Models of Dynamic Systems from Arbitrary Data using Joint Embedding Predictive Architectures
von: Ulmen, Jonas, et al.
Veröffentlicht: (2025)
von: Ulmen, Jonas, et al.
Veröffentlicht: (2025)
DIVEBATCH: Accelerating Model Training Through Gradient-Diversity Aware Batch Size Adaptation
von: Chen, Yuen, et al.
Veröffentlicht: (2025)
von: Chen, Yuen, et al.
Veröffentlicht: (2025)
STENCIL: Submodular Mutual Information Based Weak Supervision for Cold-Start Active Learning
von: Beck, Nathan, et al.
Veröffentlicht: (2024)
von: Beck, Nathan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Bitnet.cpp: Efficient Edge Inference for Ternary LLMs
von: Wang, Jinheng, et al.
Veröffentlicht: (2025) -
2 OLMo 2 Furious
von: OLMo, Team, et al.
Veröffentlicht: (2024) -
OLMoE: Open Mixture-of-Experts Language Models
von: Muennighoff, Niklas, et al.
Veröffentlicht: (2024) -
OLMoASR: Open Models and Data for Training Robust Speech Recognition Models
von: Ngo, Huong, et al.
Veröffentlicht: (2025) -
Generation of Human Comprehensible Access Control Policies from Audit Logs
von: Kumar, Gautam, et al.
Veröffentlicht: (2026)