Optimization Strategies for Enhancing Resource Efficiency in Transformers & Large Language Models
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Wallace, Tom, Ezzati-Jivan, Naser, Ombuki-Berman, Beatrice |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Evaluating Large Language Models for Public Health Classification and Extraction Tasks
par: Harris, Joshua, et autres
Publié: (2024)
par: Harris, Joshua, et autres
Publié: (2024)
Review GIDE -- Restaurant Review Gastrointestinal Illness Detection and Extraction with Large Language Models
par: Laurence, Timothy, et autres
Publié: (2025)
par: Laurence, Timothy, et autres
Publié: (2025)
Cache-to-Cache: Direct Semantic Communication Between Large Language Models
par: Fu, Tianyu, et autres
Publié: (2025)
par: Fu, Tianyu, et autres
Publié: (2025)
Surfing the modeling of PoS taggers in low-resource scenarios
par: Ferro, Manuel Vilares, et autres
Publié: (2024)
par: Ferro, Manuel Vilares, et autres
Publié: (2024)
Pivot Language for Low-Resource Machine Translation
par: Talwar, Abhimanyu, et autres
Publié: (2025)
par: Talwar, Abhimanyu, et autres
Publié: (2025)
LawInstruct: A Resource for Studying Language Model Adaptation to the Legal Domain
par: Niklaus, Joel, et autres
Publié: (2024)
par: Niklaus, Joel, et autres
Publié: (2024)
A Flexible Large Language Models Guardrail Development Methodology Applied to Off-Topic Prompt Detection
par: Chua, Gabriel, et autres
Publié: (2024)
par: Chua, Gabriel, et autres
Publié: (2024)
Uncertainty Estimation and Quantification for LLMs: A Simple Supervised Approach
par: Liu, Linyu, et autres
Publié: (2024)
par: Liu, Linyu, et autres
Publié: (2024)
A Language Model-Driven Semi-Supervised Ensemble Framework for Illicit Market Detection Across Deep/Dark Web and Social Platforms
par: Yazdanjue, Navid, et autres
Publié: (2025)
par: Yazdanjue, Navid, et autres
Publié: (2025)
LEGAL-UQA: A Low-Resource Urdu-English Dataset for Legal Question Answering
par: Faisal, Faizan, et autres
Publié: (2024)
par: Faisal, Faizan, et autres
Publié: (2024)
CausalSent: Interpretable Sentiment Classification with RieszNet
par: Frees, Daniel, et autres
Publié: (2025)
par: Frees, Daniel, et autres
Publié: (2025)
Healthy LLMs? Benchmarking LLM Knowledge of UK Government Public Health Information
par: Harris, Joshua, et autres
Publié: (2025)
par: Harris, Joshua, et autres
Publié: (2025)
Does It Make Sense to Explain a Black Box With Another Black Box?
par: Delaunay, Julien, et autres
Publié: (2024)
par: Delaunay, Julien, et autres
Publié: (2024)
QuAILoRA: Quantization-Aware Initialization for LoRA
par: Lawton, Neal, et autres
Publié: (2024)
par: Lawton, Neal, et autres
Publié: (2024)
OptPO: Optimal Rollout Allocation for Test-time Policy Optimization
par: Wang, Youkang, et autres
Publié: (2025)
par: Wang, Youkang, et autres
Publié: (2025)
PolyTruth: Multilingual Disinformation Detection using Transformer-Based Language Models
par: Gouliev, Zaur, et autres
Publié: (2025)
par: Gouliev, Zaur, et autres
Publié: (2025)
ENIGMA: The Geometry of Reasoning and Alignment in Large-Language Models
par: Seneque, Gareth, et autres
Publié: (2025)
par: Seneque, Gareth, et autres
Publié: (2025)
Fine-tuning Large Language Models for Entity Matching
par: Steiner, Aaron, et autres
Publié: (2024)
par: Steiner, Aaron, et autres
Publié: (2024)
Multi-Model Synthetic Training for Mission-Critical Small Language Models
par: Platt, Nolan, et autres
Publié: (2025)
par: Platt, Nolan, et autres
Publié: (2025)
Recent Advances in Named Entity Recognition: A Comprehensive Survey and Comparative Study
par: Keraghel, Imed, et autres
Publié: (2024)
par: Keraghel, Imed, et autres
Publié: (2024)
Parameter-Efficient Transformer Embeddings
par: Ndubuaku, Henry, et autres
Publié: (2025)
par: Ndubuaku, Henry, et autres
Publié: (2025)
Large Language Models versus Classical Machine Learning: Performance in COVID-19 Mortality Prediction Using High-Dimensional Tabular Data
par: Ghaffarzadeh-Esfahani, Mohammadreza, et autres
Publié: (2024)
par: Ghaffarzadeh-Esfahani, Mohammadreza, et autres
Publié: (2024)
ABC Align: Large Language Model Alignment for Safety & Accuracy
par: Seneque, Gareth, et autres
Publié: (2024)
par: Seneque, Gareth, et autres
Publié: (2024)
LLM-Assisted Crisis Management: Building Advanced LLM Platforms for Effective Emergency Response and Public Collaboration
par: Otal, Hakan T., et autres
Publié: (2024)
par: Otal, Hakan T., et autres
Publié: (2024)
HEFT: A Coarse-to-Fine Hierarchy for Enhancing the Efficiency and Accuracy of Language Model Reasoning
par: Hill, Brennen
Publié: (2025)
par: Hill, Brennen
Publié: (2025)
Lightweight Transformers for Clinical Natural Language Processing
par: Rohanian, Omid, et autres
Publié: (2023)
par: Rohanian, Omid, et autres
Publié: (2023)
DYNAMAX: Dynamic computing for Transformers and Mamba based architectures
par: Nogales, Miguel, et autres
Publié: (2025)
par: Nogales, Miguel, et autres
Publié: (2025)
Mechanistic Analysis of Catastrophic Forgetting in Large Language Models During Continual Fine-tuning
par: Imanov, Olaf Yunus Laitinen
Publié: (2026)
par: Imanov, Olaf Yunus Laitinen
Publié: (2026)
TRUE: A Trustworthy Unified Explanation Framework for Large Language Model Reasoning
par: Yang, Yujiao
Publié: (2026)
par: Yang, Yujiao
Publié: (2026)
Fairness Certification for Natural Language Processing and Large Language Models
par: Freiberger, Vincent, et autres
Publié: (2024)
par: Freiberger, Vincent, et autres
Publié: (2024)
Template-Based Probes Are Imperfect Lenses for Counterfactual Bias Evaluation in LLMs
par: Kohankhaki, Farnaz, et autres
Publié: (2024)
par: Kohankhaki, Farnaz, et autres
Publié: (2024)
Knowledge Editing for Large Language Model with Knowledge Neuronal Ensemble
par: Li, Yongchang, et autres
Publié: (2024)
par: Li, Yongchang, et autres
Publié: (2024)
A Comprehensive Survey of Small Language Models in the Era of Large Language Models: Techniques, Enhancements, Applications, Collaboration with LLMs, and Trustworthiness
par: Wang, Fali, et autres
Publié: (2024)
par: Wang, Fali, et autres
Publié: (2024)
Clustering in pure-attention hardmax transformers and its role in sentiment analysis
par: Alcalde, Albert, et autres
Publié: (2024)
par: Alcalde, Albert, et autres
Publié: (2024)
Semantic Needles in Document Haystacks: Sensitivity Testing of LLM-as-a-Judge Similarity Scoring
par: Aksoy, Sinan G., et autres
Publié: (2026)
par: Aksoy, Sinan G., et autres
Publié: (2026)
A Survey on Hypothesis Generation for Scientific Discovery in the Era of Large Language Models
par: Alkan, Atilla Kaan, et autres
Publié: (2025)
par: Alkan, Atilla Kaan, et autres
Publié: (2025)
Fishing for Magikarp: Automatically Detecting Under-trained Tokens in Large Language Models
par: Land, Sander, et autres
Publié: (2024)
par: Land, Sander, et autres
Publié: (2024)
On The Role of Reasoning in the Identification of Subtle Stereotypes in Natural Language
par: Tian, Jacob-Junqi, et autres
Publié: (2023)
par: Tian, Jacob-Junqi, et autres
Publié: (2023)
VisGraphVar: A Benchmark Generator for Assessing Variability in Graph Analysis Using Large Vision-Language Models
par: Sartori, Camilo Chacón, et autres
Publié: (2024)
par: Sartori, Camilo Chacón, et autres
Publié: (2024)
Anonymity at Risk? Assessing Re-Identification Capabilities of Large Language Models
par: Nyffenegger, Alex, et autres
Publié: (2023)
par: Nyffenegger, Alex, et autres
Publié: (2023)
Documents similaires
-
Evaluating Large Language Models for Public Health Classification and Extraction Tasks
par: Harris, Joshua, et autres
Publié: (2024) -
Review GIDE -- Restaurant Review Gastrointestinal Illness Detection and Extraction with Large Language Models
par: Laurence, Timothy, et autres
Publié: (2025) -
Cache-to-Cache: Direct Semantic Communication Between Large Language Models
par: Fu, Tianyu, et autres
Publié: (2025) -
Surfing the modeling of PoS taggers in low-resource scenarios
par: Ferro, Manuel Vilares, et autres
Publié: (2024) -
Pivot Language for Low-Resource Machine Translation
par: Talwar, Abhimanyu, et autres
Publié: (2025)