Tag-LLM: Repurposing General-Purpose LLMs for Specialized Domains
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Shen, Junhong, Tenenholtz, Neil, Hall, James Brian, Alvarez-Melis, David, Fusi, Nicolo |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Adapting Language Models via Token Translation
von: Feng, Zhili, et al.
Veröffentlicht: (2024)
von: Feng, Zhili, et al.
Veröffentlicht: (2024)
IGDA: Interactive Graph Discovery through Large Language Model Agents
von: Havrilla, Alex, et al.
Veröffentlicht: (2025)
von: Havrilla, Alex, et al.
Veröffentlicht: (2025)
Activation Oracles: Training and Evaluating LLMs as General-Purpose Activation Explainers
von: Karvonen, Adam, et al.
Veröffentlicht: (2025)
von: Karvonen, Adam, et al.
Veröffentlicht: (2025)
LLMs as Scalable, General-Purpose Simulators For Evolving Digital Agent Training
von: Wang, Yiming, et al.
Veröffentlicht: (2025)
von: Wang, Yiming, et al.
Veröffentlicht: (2025)
TrimLLM: Progressive Layer Dropping for Domain-Specific LLMs
von: Hu, Lanxiang, et al.
Veröffentlicht: (2024)
von: Hu, Lanxiang, et al.
Veröffentlicht: (2024)
Are complicated loss functions necessary for teaching LLMs to reason?
von: Carrino, Gabriele, et al.
Veröffentlicht: (2026)
von: Carrino, Gabriele, et al.
Veröffentlicht: (2026)
CodePDE: An Inference Framework for LLM-driven PDE Solver Generation
von: Li, Shanda, et al.
Veröffentlicht: (2025)
von: Li, Shanda, et al.
Veröffentlicht: (2025)
Diversity as a Reward: Fine-Tuning LLMs on a Mixture of Domain-Undetermined Data
von: Ling, Zhenqing, et al.
Veröffentlicht: (2025)
von: Ling, Zhenqing, et al.
Veröffentlicht: (2025)
Can Interpretation Predict Behavior on Unseen Data?
von: Li, Victoria R., et al.
Veröffentlicht: (2025)
von: Li, Victoria R., et al.
Veröffentlicht: (2025)
Towards Specialized Generalists: A Multi-Task MoE-LoRA Framework for Domain-Specific LLM Adaptation
von: Yang, Yuxin, et al.
Veröffentlicht: (2026)
von: Yang, Yuxin, et al.
Veröffentlicht: (2026)
Advancing General-Purpose Reasoning Models with Modular Gradient Surgery
von: Cai, Min, et al.
Veröffentlicht: (2026)
von: Cai, Min, et al.
Veröffentlicht: (2026)
TopicTag: Automatic Annotation of NMF Topic Models Using Chain of Thought and Prompt Tuning with LLMs
von: Wanna, Selma, et al.
Veröffentlicht: (2024)
von: Wanna, Selma, et al.
Veröffentlicht: (2024)
An Empirical Study on Context Length for Open-Domain Dialog Generation
von: Shen, Xinyi, et al.
Veröffentlicht: (2024)
von: Shen, Xinyi, et al.
Veröffentlicht: (2024)
Can General-Purpose Large Language Models Generalize to English-Thai Machine Translation ?
von: Chiaranaipanich, Jirat, et al.
Veröffentlicht: (2024)
von: Chiaranaipanich, Jirat, et al.
Veröffentlicht: (2024)
CRAFT: Customizing LLMs by Creating and Retrieving from Specialized Toolsets
von: Yuan, Lifan, et al.
Veröffentlicht: (2023)
von: Yuan, Lifan, et al.
Veröffentlicht: (2023)
FBI-LLM: Scaling Up Fully Binarized LLMs from Scratch via Autoregressive Distillation
von: Ma, Liqun, et al.
Veröffentlicht: (2024)
von: Ma, Liqun, et al.
Veröffentlicht: (2024)
Nemotron-Cascade: Scaling Cascaded Reinforcement Learning for General-Purpose Reasoning Models
von: Wang, Boxin, et al.
Veröffentlicht: (2025)
von: Wang, Boxin, et al.
Veröffentlicht: (2025)
Learn from Weaknesses: Automated Domain Specialization for Small Computer-Use Agents
von: Kim, Suji, et al.
Veröffentlicht: (2026)
von: Kim, Suji, et al.
Veröffentlicht: (2026)
InversionView: A General-Purpose Method for Reading Information from Neural Activations
von: Huang, Xinting, et al.
Veröffentlicht: (2024)
von: Huang, Xinting, et al.
Veröffentlicht: (2024)
Draft-Conditioned Constrained Decoding for Structured Generation in LLMs
von: Reddy, Avinash, et al.
Veröffentlicht: (2026)
von: Reddy, Avinash, et al.
Veröffentlicht: (2026)
Agentic Adversarial QA for Improving Domain-Specific LLMs
von: Grari, Vincent, et al.
Veröffentlicht: (2026)
von: Grari, Vincent, et al.
Veröffentlicht: (2026)
CreativityBench: Evaluating Agent Creative Reasoning via Affordance-Based Tool Repurposing
von: Qian, Cheng, et al.
Veröffentlicht: (2026)
von: Qian, Cheng, et al.
Veröffentlicht: (2026)
Dynamic Expert Specialization: Towards Catastrophic Forgetting-Free Multi-Domain MoE Adaptation
von: Li, Junzhuo, et al.
Veröffentlicht: (2025)
von: Li, Junzhuo, et al.
Veröffentlicht: (2025)
DrKGC: Dynamic Subgraph Retrieval-Augmented LLMs for Knowledge Graph Completion across General and Biomedical Domains
von: Xiao, Yongkang, et al.
Veröffentlicht: (2025)
von: Xiao, Yongkang, et al.
Veröffentlicht: (2025)
Decomposing Elements of Problem Solving: What "Math" Does RL Teach?
von: Qin, Tian, et al.
Veröffentlicht: (2025)
von: Qin, Tian, et al.
Veröffentlicht: (2025)
SyGra: A Unified Graph-Based Framework for Scalable Generation, Quality Tagging, and Management of Synthetic Data
von: Pradhan, Bidyapati, et al.
Veröffentlicht: (2025)
von: Pradhan, Bidyapati, et al.
Veröffentlicht: (2025)
SemEval-2025 Task 5: LLMs4Subjects -- LLM-based Automated Subject Tagging for a National Technical Library's Open-Access Catalog
von: D'Souza, Jennifer, et al.
Veröffentlicht: (2025)
von: D'Souza, Jennifer, et al.
Veröffentlicht: (2025)
Attention Sinks: A 'Catch, Tag, Release' Mechanism for Embeddings
von: Zhang, Stephen, et al.
Veröffentlicht: (2025)
von: Zhang, Stephen, et al.
Veröffentlicht: (2025)
Empowering Small-Scale Knowledge Graphs: A Strategy of Leveraging General-Purpose Knowledge Graphs for Enriched Embeddings
von: Sawczyn, Albert, et al.
Veröffentlicht: (2024)
von: Sawczyn, Albert, et al.
Veröffentlicht: (2024)
On Evaluating LLM Alignment by Evaluating LLMs as Judges
von: Liu, Yixin, et al.
Veröffentlicht: (2025)
von: Liu, Yixin, et al.
Veröffentlicht: (2025)
Dr.LLM: Dynamic Layer Routing in LLMs
von: Heakl, Ahmed, et al.
Veröffentlicht: (2025)
von: Heakl, Ahmed, et al.
Veröffentlicht: (2025)
References Improve LLM Alignment in Non-Verifiable Domains
von: Shi, Kejian, et al.
Veröffentlicht: (2026)
von: Shi, Kejian, et al.
Veröffentlicht: (2026)
Benchmarking for Domain-Specific LLMs: A Case Study on Academia and Beyond
von: Chen, Rubing, et al.
Veröffentlicht: (2025)
von: Chen, Rubing, et al.
Veröffentlicht: (2025)
Constructing Synthetic Instruction Datasets for Improving Reasoning in Domain-Specific LLMs: A Case Study in the Japanese Financial Domain
von: Okochi, Yuma, et al.
Veröffentlicht: (2026)
von: Okochi, Yuma, et al.
Veröffentlicht: (2026)
SimRAG: Self-Improving Retrieval-Augmented Generation for Adapting Large Language Models to Specialized Domains
von: Xu, Ran, et al.
Veröffentlicht: (2024)
von: Xu, Ran, et al.
Veröffentlicht: (2024)
RouteLLM: Learning to Route LLMs with Preference Data
von: Ong, Isaac, et al.
Veröffentlicht: (2024)
von: Ong, Isaac, et al.
Veröffentlicht: (2024)
DB-LLM: Accurate Dual-Binarization for Efficient LLMs
von: Chen, Hong, et al.
Veröffentlicht: (2024)
von: Chen, Hong, et al.
Veröffentlicht: (2024)
A Comparative Study of Specialized LLMs as Dense Retrievers
von: Zhang, Hengran, et al.
Veröffentlicht: (2025)
von: Zhang, Hengran, et al.
Veröffentlicht: (2025)
Synthetic vs. Gold: The Role of LLM Generated Labels and Data in Cyberbullying Detection
von: Kazemi, Arefeh, et al.
Veröffentlicht: (2025)
von: Kazemi, Arefeh, et al.
Veröffentlicht: (2025)
Domain-level metacognitive monitoring in frontier LLMs: A 33-model atlas
von: Cacioli, Jon-Paul
Veröffentlicht: (2026)
von: Cacioli, Jon-Paul
Veröffentlicht: (2026)
Ähnliche Einträge
-
Adapting Language Models via Token Translation
von: Feng, Zhili, et al.
Veröffentlicht: (2024) -
IGDA: Interactive Graph Discovery through Large Language Model Agents
von: Havrilla, Alex, et al.
Veröffentlicht: (2025) -
Activation Oracles: Training and Evaluating LLMs as General-Purpose Activation Explainers
von: Karvonen, Adam, et al.
Veröffentlicht: (2025) -
LLMs as Scalable, General-Purpose Simulators For Evolving Digital Agent Training
von: Wang, Yiming, et al.
Veröffentlicht: (2025) -
TrimLLM: Progressive Layer Dropping for Domain-Specific LLMs
von: Hu, Lanxiang, et al.
Veröffentlicht: (2024)