Fragile Mastery: Are Domain-Specific Trade-Offs Undermining On-Device Language Models?
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jha, Basab, Paudel, Firoj |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
The Rosetta Paradox: Domain-Specific Performance Inversions in Large Language Models
von: Jha, Basab, et al.
Veröffentlicht: (2024)
von: Jha, Basab, et al.
Veröffentlicht: (2024)
Thinking About Thinking: SAGE-nano's Inverse Reasoning for Self-Aware Language Models
von: Jha, Basab, et al.
Veröffentlicht: (2025)
von: Jha, Basab, et al.
Veröffentlicht: (2025)
SAGE-32B: Agentic Reasoning via Iterative Distillation
von: Jha, Basab, et al.
Veröffentlicht: (2026)
von: Jha, Basab, et al.
Veröffentlicht: (2026)
Exploring Safety-Utility Trade-Offs in Personalized Language Models
von: Vijjini, Anvesh Rao, et al.
Veröffentlicht: (2024)
von: Vijjini, Anvesh Rao, et al.
Veröffentlicht: (2024)
SAGE Celer 2.6 Technical Card
von: SAGEA Research Team, et al.
Veröffentlicht: (2026)
von: SAGEA Research Team, et al.
Veröffentlicht: (2026)
Green AI: Exploring Carbon Footprints, Mitigation Strategies, and Trade Offs in Large Language Model Training
von: Liu, Vivian, et al.
Veröffentlicht: (2024)
von: Liu, Vivian, et al.
Veröffentlicht: (2024)
Chain-of-Rank: Enhancing Large Language Models for Domain-Specific RAG in Edge Device
von: Lee, Juntae, et al.
Veröffentlicht: (2025)
von: Lee, Juntae, et al.
Veröffentlicht: (2025)
Exploring the Trade-Offs: Quantization Methods, Task Difficulty, and Model Size in Large Language Models From Edge to Giant
von: Lee, Jemin, et al.
Veröffentlicht: (2024)
von: Lee, Jemin, et al.
Veröffentlicht: (2024)
MKGL: Mastery of a Three-Word Language
von: Guo, Lingbing, et al.
Veröffentlicht: (2024)
von: Guo, Lingbing, et al.
Veröffentlicht: (2024)
Engagement Undermines Safety: How Stereotypes and Toxicity Shape Humor in Language Models
von: Dogra, Atharvan, et al.
Veröffentlicht: (2025)
von: Dogra, Atharvan, et al.
Veröffentlicht: (2025)
Compromesso! Italian Many-Shot Jailbreaks Undermine the Safety of Large Language Models
von: Pernisi, Fabio, et al.
Veröffentlicht: (2024)
von: Pernisi, Fabio, et al.
Veröffentlicht: (2024)
The Rise of Darkness: Safety-Utility Trade-Offs in Role-Playing Dialogue Agents
von: Tang, Yihong, et al.
Veröffentlicht: (2025)
von: Tang, Yihong, et al.
Veröffentlicht: (2025)
A Deep Dive into the Trade-Offs of Parameter-Efficient Preference Alignment Techniques
von: Thakkar, Megh, et al.
Veröffentlicht: (2024)
von: Thakkar, Megh, et al.
Veröffentlicht: (2024)
An Industrial-Scale Insurance LLM Achieving Verifiable Domain Mastery and Hallucination Control without Competence Trade-offs
von: Zhu, Qian, et al.
Veröffentlicht: (2026)
von: Zhu, Qian, et al.
Veröffentlicht: (2026)
From Correction to Mastery: Reinforced Distillation of Large Language Model Agents
von: Lyu, Yuanjie, et al.
Veröffentlicht: (2025)
von: Lyu, Yuanjie, et al.
Veröffentlicht: (2025)
Natural Context Drift Undermines the Natural Language Understanding of Large Language Models
von: Wu, Yulong, et al.
Veröffentlicht: (2025)
von: Wu, Yulong, et al.
Veröffentlicht: (2025)
On the Limits of Language Generation: Trade-Offs Between Hallucination and Mode Collapse
von: Kalavasis, Alkis, et al.
Veröffentlicht: (2024)
von: Kalavasis, Alkis, et al.
Veröffentlicht: (2024)
The Landscape of Arabic Large Language Models (ALLMs): A New Era for Arabic Language Technology
von: Al-Khalifa, Shahad, et al.
Veröffentlicht: (2025)
von: Al-Khalifa, Shahad, et al.
Veröffentlicht: (2025)
Seamless Language Expansion: Enhancing Multilingual Mastery in Self-Supervised Models
von: Xu, Jing, et al.
Veröffentlicht: (2024)
von: Xu, Jing, et al.
Veröffentlicht: (2024)
Uncertainty is Fragile: Manipulating Uncertainty in Large Language Models
von: Zeng, Qingcheng, et al.
Veröffentlicht: (2024)
von: Zeng, Qingcheng, et al.
Veröffentlicht: (2024)
Understanding Fairness-Accuracy Trade-offs in Machine Learning Models: Does Promoting Fairness Undermine Performance?
von: Liu, Junhua, et al.
Veröffentlicht: (2024)
von: Liu, Junhua, et al.
Veröffentlicht: (2024)
Biased or Flawed? Mitigating Stereotypes in Generative Language Models by Addressing Task-Specific Flaws
von: Jha, Akshita, et al.
Veröffentlicht: (2024)
von: Jha, Akshita, et al.
Veröffentlicht: (2024)
Semantic Mastery: Enhancing LLMs with Advanced Natural Language Understanding
von: Hariharan, Mohanakrishnan
Veröffentlicht: (2025)
von: Hariharan, Mohanakrishnan
Veröffentlicht: (2025)
Once Correct, Still Wrong: Counterfactual Hallucination in Multilingual Vision-Language Models
von: Mousi, Basel, et al.
Veröffentlicht: (2026)
von: Mousi, Basel, et al.
Veröffentlicht: (2026)
Cross-Domain Content Generation with Domain-Specific Small Language Models
von: Maloo, Ankit, et al.
Veröffentlicht: (2024)
von: Maloo, Ankit, et al.
Veröffentlicht: (2024)
Step-by-Step Mastery: Enhancing Soft Constraint Following Ability of Large Language Models
von: Ren, Qingyu, et al.
Veröffentlicht: (2025)
von: Ren, Qingyu, et al.
Veröffentlicht: (2025)
Language Models are Surprisingly Fragile to Drug Names in Biomedical Benchmarks
von: Gallifant, Jack, et al.
Veröffentlicht: (2024)
von: Gallifant, Jack, et al.
Veröffentlicht: (2024)
DeepThink: Aligning Language Models with Domain-Specific User Intents
von: Li, Yang, et al.
Veröffentlicht: (2025)
von: Li, Yang, et al.
Veröffentlicht: (2025)
HUKUKBERT: Domain-Specific Language Model for Turkish Law
von: Öztürk, Mehmet Utku, et al.
Veröffentlicht: (2026)
von: Öztürk, Mehmet Utku, et al.
Veröffentlicht: (2026)
RAFT: Adapting Language Model to Domain Specific RAG
von: Zhang, Tianjun, et al.
Veröffentlicht: (2024)
von: Zhang, Tianjun, et al.
Veröffentlicht: (2024)
BLADE: Enhancing Black-box Large Language Models with Small Domain-Specific Models
von: Li, Haitao, et al.
Veröffentlicht: (2024)
von: Li, Haitao, et al.
Veröffentlicht: (2024)
Injecting Domain-Specific Knowledge into Large Language Models: A Comprehensive Survey
von: Song, Zirui, et al.
Veröffentlicht: (2025)
von: Song, Zirui, et al.
Veröffentlicht: (2025)
MiningGPT -- A Domain-Specific Large Language Model for the Mining Industry
von: Demartini, Kurukulasooriya Fernando ana Gianluca
Veröffentlicht: (2024)
von: Demartini, Kurukulasooriya Fernando ana Gianluca
Veröffentlicht: (2024)
Efficient Continual Pre-training for Building Domain Specific Large Language Models
von: Xie, Yong, et al.
Veröffentlicht: (2023)
von: Xie, Yong, et al.
Veröffentlicht: (2023)
Exploiting Domain-Specific Parallel Data on Multilingual Language Models for Low-resource Language Translation
von: Ranathungaa, Surangika, et al.
Veröffentlicht: (2024)
von: Ranathungaa, Surangika, et al.
Veröffentlicht: (2024)
Large Language Models for Propaganda Span Annotation
von: Hasanain, Maram, et al.
Veröffentlicht: (2023)
von: Hasanain, Maram, et al.
Veröffentlicht: (2023)
The Reliability Paradox: Exploring How Shortcut Learning Undermines Language Model Calibration
von: Bihani, Geetanjali, et al.
Veröffentlicht: (2024)
von: Bihani, Geetanjali, et al.
Veröffentlicht: (2024)
Fine-Tuned Language Models for Domain-Specific Summarization and Tagging
von: Wang, Jun, et al.
Veröffentlicht: (2025)
von: Wang, Jun, et al.
Veröffentlicht: (2025)
Beyond Behavioural Trade-Offs: Mechanistic Tracing of Pain-Pleasure Decisions in an LLM
von: Bianco, Francesca, et al.
Veröffentlicht: (2026)
von: Bianco, Francesca, et al.
Veröffentlicht: (2026)
Retrieval-Enhanced Mutation Mastery: Augmenting Zero-Shot Prediction of Protein Language Model
von: Tan, Yang, et al.
Veröffentlicht: (2024)
von: Tan, Yang, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
The Rosetta Paradox: Domain-Specific Performance Inversions in Large Language Models
von: Jha, Basab, et al.
Veröffentlicht: (2024) -
Thinking About Thinking: SAGE-nano's Inverse Reasoning for Self-Aware Language Models
von: Jha, Basab, et al.
Veröffentlicht: (2025) -
SAGE-32B: Agentic Reasoning via Iterative Distillation
von: Jha, Basab, et al.
Veröffentlicht: (2026) -
Exploring Safety-Utility Trade-Offs in Personalized Language Models
von: Vijjini, Anvesh Rao, et al.
Veröffentlicht: (2024) -
SAGE Celer 2.6 Technical Card
von: SAGEA Research Team, et al.
Veröffentlicht: (2026)