PANDA: Preference Adaptation for Enhancing Domain-Specific Abilities of LLMs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liu, An, Yang, Zonghan, Zhang, Zhenhe, Hu, Qingyuan, Li, Peng, Yan, Ming, Zhang, Ji, Huang, Fei, Liu, Yang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ReAct Meets ActRe: When Language Agents Enjoy Training Data Autonomy
von: Yang, Zonghan, et al.
Veröffentlicht: (2024)
von: Yang, Zonghan, et al.
Veröffentlicht: (2024)
AIGS: Generating Science from AI-Powered Automated Falsification
von: Liu, Zijun, et al.
Veröffentlicht: (2024)
von: Liu, Zijun, et al.
Veröffentlicht: (2024)
Towards Unified Alignment Between Agents, Humans, and Environment
von: Yang, Zonghan, et al.
Veröffentlicht: (2024)
von: Yang, Zonghan, et al.
Veröffentlicht: (2024)
Enabling Weak LLMs to Judge Response Reliability via Meta Ranking
von: Liu, Zijun, et al.
Veröffentlicht: (2024)
von: Liu, Zijun, et al.
Veröffentlicht: (2024)
On the Generalization and Adaptation Ability of Machine-Generated Text Detectors in Academic Writing
von: Liu, Yule, et al.
Veröffentlicht: (2024)
von: Liu, Yule, et al.
Veröffentlicht: (2024)
SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization
von: Sun, Huashan, et al.
Veröffentlicht: (2025)
von: Sun, Huashan, et al.
Veröffentlicht: (2025)
Enhancing Reasoning Abilities of Small LLMs with Cognitive Alignment
von: Cai, Wenrui, et al.
Veröffentlicht: (2025)
von: Cai, Wenrui, et al.
Veröffentlicht: (2025)
SLearnLLM: A Self-Learning Framework for Efficient Domain-Specific Adaptation of Large Language Models
von: Liu, Xiang, et al.
Veröffentlicht: (2025)
von: Liu, Xiang, et al.
Veröffentlicht: (2025)
Enhancing the Geometric Problem-Solving Ability of Multimodal LLMs via Symbolic-Neural Integration
von: Pan, Yicheng, et al.
Veröffentlicht: (2025)
von: Pan, Yicheng, et al.
Veröffentlicht: (2025)
Domain-Specific Data Generation Framework for RAG Adaptation
von: Tian, Chris Xing, et al.
Veröffentlicht: (2025)
von: Tian, Chris Xing, et al.
Veröffentlicht: (2025)
Knowledgeable Preference Alignment for LLMs in Domain-specific Question Answering
von: Zhang, Yichi, et al.
Veröffentlicht: (2023)
von: Zhang, Yichi, et al.
Veröffentlicht: (2023)
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning
von: Fei, Zhaoye, et al.
Veröffentlicht: (2025)
von: Fei, Zhaoye, et al.
Veröffentlicht: (2025)
TrimLLM: Progressive Layer Dropping for Domain-Specific LLMs
von: Hu, Lanxiang, et al.
Veröffentlicht: (2024)
von: Hu, Lanxiang, et al.
Veröffentlicht: (2024)
An Empirical Investigation of Domain Adaptation Ability for Chinese Spelling Check Models
von: Wang, Xi, et al.
Veröffentlicht: (2024)
von: Wang, Xi, et al.
Veröffentlicht: (2024)
Enhancing the Medical Context-Awareness Ability of LLMs via Multifaceted Self-Refinement Learning
von: Zhou, Yuxuan, et al.
Veröffentlicht: (2025)
von: Zhou, Yuxuan, et al.
Veröffentlicht: (2025)
IOPO: Empowering LLMs with Complex Instruction Following via Input-Output Preference Optimization
von: Zhang, Xinghua, et al.
Veröffentlicht: (2024)
von: Zhang, Xinghua, et al.
Veröffentlicht: (2024)
Penrose Tiled Low-Rank Compression and Section-Wise Q&A Fine-Tuning: A General Framework for Domain-Specific Large Language Model Adaptation
von: Kuo, Chuan-Wei, et al.
Veröffentlicht: (2025)
von: Kuo, Chuan-Wei, et al.
Veröffentlicht: (2025)
Pi-SQL: Enhancing Text-to-SQL with Fine-Grained Guidance from Pivot Programming Languages
von: chi, Yongdong, et al.
Veröffentlicht: (2025)
von: chi, Yongdong, et al.
Veröffentlicht: (2025)
Small LLMs Are Weak Tool Learners: A Multi-LLM Agent
von: Shen, Weizhou, et al.
Veröffentlicht: (2024)
von: Shen, Weizhou, et al.
Veröffentlicht: (2024)
StealthGraph: Exposing Domain-Specific Risks in LLMs through Knowledge-Graph-Guided Harmful Prompt Generation
von: Zheng, Huawei, et al.
Veröffentlicht: (2026)
von: Zheng, Huawei, et al.
Veröffentlicht: (2026)
DO-RAG: A Domain-Specific QA Framework Using Knowledge Graph-Enhanced Retrieval-Augmented Generation
von: Opoku, David Osei, et al.
Veröffentlicht: (2025)
von: Opoku, David Osei, et al.
Veröffentlicht: (2025)
RoleLLM: Benchmarking, Eliciting, and Enhancing Role-Playing Abilities of Large Language Models
von: Wang, Zekun Moore, et al.
Veröffentlicht: (2023)
von: Wang, Zekun Moore, et al.
Veröffentlicht: (2023)
Domain Adaptation of LLMs for Process Data
von: Oyamada, Rafael Seidi, et al.
Veröffentlicht: (2025)
von: Oyamada, Rafael Seidi, et al.
Veröffentlicht: (2025)
Evaluation Ethics of LLMs in Legal Domain
von: Zhang, Ruizhe, et al.
Veröffentlicht: (2024)
von: Zhang, Ruizhe, et al.
Veröffentlicht: (2024)
PharmaGPT: Domain-Specific Large Language Models for Bio-Pharmaceutical and Chemistry
von: Chen, Linqing, et al.
Veröffentlicht: (2024)
von: Chen, Linqing, et al.
Veröffentlicht: (2024)
What External Knowledge is Preferred by LLMs? Characterizing and Exploring Chain of Evidence in Imperfect Context for Multi-Hop QA
von: Chang, Zhiyuan, et al.
Veröffentlicht: (2024)
von: Chang, Zhiyuan, et al.
Veröffentlicht: (2024)
Ranking LLMs by compression
von: Guo, Peijia, et al.
Veröffentlicht: (2024)
von: Guo, Peijia, et al.
Veröffentlicht: (2024)
Toward Robust Multilingual Adaptation of LLMs for Low-Resource Languages
von: Li, Haolin, et al.
Veröffentlicht: (2025)
von: Li, Haolin, et al.
Veröffentlicht: (2025)
KBAlign: Efficient Self Adaptation on Specific Knowledge Bases
von: Zeng, Zheni, et al.
Veröffentlicht: (2024)
von: Zeng, Zheni, et al.
Veröffentlicht: (2024)
TactfulToM: Do LLMs Have the Theory of Mind Ability to Understand White Lies?
von: Liu, Yiwei, et al.
Veröffentlicht: (2025)
von: Liu, Yiwei, et al.
Veröffentlicht: (2025)
Domain Knowledge-Enhanced LLMs for Fraud and Concept Drift Detection
von: Şenol, Ali, et al.
Veröffentlicht: (2025)
von: Şenol, Ali, et al.
Veröffentlicht: (2025)
All-in-One Tuning and Structural Pruning for Domain-Specific LLMs
von: Lu, Lei, et al.
Veröffentlicht: (2024)
von: Lu, Lei, et al.
Veröffentlicht: (2024)
MATH-Perturb: Benchmarking LLMs' Math Reasoning Abilities against Hard Perturbations
von: Huang, Kaixuan, et al.
Veröffentlicht: (2025)
von: Huang, Kaixuan, et al.
Veröffentlicht: (2025)
Enhancing Multi-Step Reasoning Abilities of Language Models through Direct Q-Function Optimization
von: Ji, Kaixuan, et al.
Veröffentlicht: (2024)
von: Ji, Kaixuan, et al.
Veröffentlicht: (2024)
LLMs to Support a Domain Specific Knowledge Assistant
von: Lovin, Maria-Flavia
Veröffentlicht: (2025)
von: Lovin, Maria-Flavia
Veröffentlicht: (2025)
Discrete Tokenization for Multimodal LLMs: A Comprehensive Survey
von: Li, Jindong, et al.
Veröffentlicht: (2025)
von: Li, Jindong, et al.
Veröffentlicht: (2025)
Towards Specialized Generalists: A Multi-Task MoE-LoRA Framework for Domain-Specific LLM Adaptation
von: Yang, Yuxin, et al.
Veröffentlicht: (2026)
von: Yang, Yuxin, et al.
Veröffentlicht: (2026)
Chain-of-Rank: Enhancing Large Language Models for Domain-Specific RAG in Edge Device
von: Lee, Juntae, et al.
Veröffentlicht: (2025)
von: Lee, Juntae, et al.
Veröffentlicht: (2025)
LMM-R1: Empowering 3B LMMs with Strong Reasoning Abilities Through Two-Stage Rule-Based RL
von: Peng, Yingzhe, et al.
Veröffentlicht: (2025)
von: Peng, Yingzhe, et al.
Veröffentlicht: (2025)
An Empirical Study of Data Ability Boundary in LLMs' Math Reasoning
von: Chen, Zui, et al.
Veröffentlicht: (2024)
von: Chen, Zui, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
ReAct Meets ActRe: When Language Agents Enjoy Training Data Autonomy
von: Yang, Zonghan, et al.
Veröffentlicht: (2024) -
AIGS: Generating Science from AI-Powered Automated Falsification
von: Liu, Zijun, et al.
Veröffentlicht: (2024) -
Towards Unified Alignment Between Agents, Humans, and Environment
von: Yang, Zonghan, et al.
Veröffentlicht: (2024) -
Enabling Weak LLMs to Judge Response Reliability via Meta Ranking
von: Liu, Zijun, et al.
Veröffentlicht: (2024) -
On the Generalization and Adaptation Ability of Machine-Generated Text Detectors in Academic Writing
von: Liu, Yule, et al.
Veröffentlicht: (2024)