APT: Improving Specialist LLM Performance with Weakness Case Acquisition and Iterative Preference Training
Fuente:
arXiv
Saved in:
| Main Authors: | Rao, Jun, Lin, Zepeng, Liu, Xuebo, Ke, Xiaopeng, Lian, Lian, Jin, Dong, Cheng, Shengjun, Yu, Jun, Zhang, Min |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SeaPO: Strategic Error Amplification for Robust Preference Optimization of Large Language Models
by: Rao, Jun, et al.
Published: (2025)
by: Rao, Jun, et al.
Published: (2025)
CommonIT: Commonality-Aware Instruction Tuning for Large Language Models via Data Partitions
by: Rao, Jun, et al.
Published: (2024)
by: Rao, Jun, et al.
Published: (2024)
AQuilt: Weaving Logic and Self-Inspection into Low-Cost, High-Relevance Data Synthesis for Specialist LLMs
by: Ke, Xiaopeng, et al.
Published: (2025)
by: Ke, Xiaopeng, et al.
Published: (2025)
Dynamic Sampling that Adapts: Self-Aware Iterative Data Persistent Optimization for Mathematical Reasoning
by: Rao, Jun, et al.
Published: (2025)
by: Rao, Jun, et al.
Published: (2025)
Exploring and Enhancing the Transfer of Distribution in Knowledge Distillation for Autoregressive Language Models
by: Rao, Jun, et al.
Published: (2024)
by: Rao, Jun, et al.
Published: (2024)
APT-CGLP: Advanced Persistent Threat Hunting via Contrastive Graph-Language Pre-Training
by: Qiu, Xuebo, et al.
Published: (2025)
by: Qiu, Xuebo, et al.
Published: (2025)
Efficient Exploration for Iterative Nash Preference Optimization
by: Nan, Tianlong, et al.
Published: (2026)
by: Nan, Tianlong, et al.
Published: (2026)
REA-RL: Reflection-Aware Online Reinforcement Learning for Efficient Reasoning
by: Deng, Hexuan, et al.
Published: (2025)
by: Deng, Hexuan, et al.
Published: (2025)
Grid-Orch: An LLM-Powered Orchestrator for Distribution Grid Simulation and Analytics
by: Liu, Boming, et al.
Published: (2026)
by: Liu, Boming, et al.
Published: (2026)
AIPO: Improving Training Objective for Iterative Preference Optimization
by: Shen, Yaojie, et al.
Published: (2024)
by: Shen, Yaojie, et al.
Published: (2024)
ProHunter: A Comprehensive APT Hunting System Based on Whole-System Provenance
by: Qiu, Xuebo, et al.
Published: (2026)
by: Qiu, Xuebo, et al.
Published: (2026)
Table-LLM-Specialist: Language Model Specialists for Tables using Iterative Generator-Validator Fine-tuning
by: Xing, Junjie, et al.
Published: (2024)
by: Xing, Junjie, et al.
Published: (2024)
APT: Architectural Planning and Text-to-Blueprint Construction Using Large Language Models for Open-World Agents
by: Chen, Jun Yu, et al.
Published: (2024)
by: Chen, Jun Yu, et al.
Published: (2024)
Weakly Approximating Knapsack in Subquadratic Time
by: Chen, Lin, et al.
Published: (2025)
by: Chen, Lin, et al.
Published: (2025)
TREC: APT Tactic / Technique Recognition via Few-Shot Provenance Subgraph Learning
by: Lv, Mingqi, et al.
Published: (2024)
by: Lv, Mingqi, et al.
Published: (2024)
Improving Attributed Text Generation of Large Language Models via Preference Learning
by: Li, Dongfang, et al.
Published: (2024)
by: Li, Dongfang, et al.
Published: (2024)
Intelligent Channel Allocation for IEEE 802.11be Multi-Link Operation: When MAB Meets LLM
by: Lian, Shumin, et al.
Published: (2025)
by: Lian, Shumin, et al.
Published: (2025)
Comparative Analysis and Parametric Tuning of PPO, GRPO, and DAPO for LLM Reasoning Enhancement
by: Lian, Yongsheng
Published: (2025)
by: Lian, Yongsheng
Published: (2025)
Editorial for “Acquisition Efficiency and Technical Repeatability of Dual‐Frequency 3D Vector MR Elastography of the Liver”
by: Yufan Qian, et al.
Published: (2024)
by: Yufan Qian, et al.
Published: (2024)
Nezha‐SeaDart: A tail‐sitting fixed‐wing vertical takeoff and landing hybrid aerial underwater vehicle
by: Yufei Jin, et al.
Published: (2024)
by: Yufei Jin, et al.
Published: (2024)
Evidence of validity of the Attentional Performance Test (APT)
by: Jonatas R. Bessa
Published: (2021)
by: Jonatas R. Bessa
Published: (2021)
Bosonic Quantum Breakdown Hubbard Model
by: Hu, Yu-Min, et al.
Published: (2024)
by: Hu, Yu-Min, et al.
Published: (2024)
From the Quantum Breakdown Model to the Lattice Gauge Theory
by: Hu, Yu-Min, et al.
Published: (2024)
by: Hu, Yu-Min, et al.
Published: (2024)
Abdominal Undulation with Compliant Mechanism Improves Flight Performance of Biomimetic Robotic Butterfly
by: Lian, Xuyi, et al.
Published: (2025)
by: Lian, Xuyi, et al.
Published: (2025)
WaferSAGE: Large Language Model-Powered Wafer Defect Analysis via Synthetic Data Generation and Rubric-Guided Reinforcement Learning
by: Xu, Ke, et al.
Published: (2026)
by: Xu, Ke, et al.
Published: (2026)
The novel HLA‐C*01:274 allele, identified by Sanger dideoxy nucleotide sequencing in a Chinese individual
by: Xue Lian, et al.
Published: (2024)
by: Xue Lian, et al.
Published: (2024)
Distributed Iterative Hard Thresholding for Variable Selection in Tobit Models
by: Yang, Changxin, et al.
Published: (2024)
by: Yang, Changxin, et al.
Published: (2024)
Hindsight Preference Learning for Offline Preference-based Reinforcement Learning
by: Gao, Chen-Xiao, et al.
Published: (2024)
by: Gao, Chen-Xiao, et al.
Published: (2024)
SHIELD: APT Detection and Intelligent Explanation Using LLM
by: Gandhi, Parth Atulbhai, et al.
Published: (2025)
by: Gandhi, Parth Atulbhai, et al.
Published: (2025)
SuperOffload: Unleashing the Power of Large-Scale LLM Training on Superchips
by: Lian, Xinyu, et al.
Published: (2025)
by: Lian, Xinyu, et al.
Published: (2025)
Multimodal Contrastive Learning for 3D Object Classification and Part‐Segmentation by Leveraging V‐LLM and CNNs
by: Jiaxin Jiang, et al.
Published: (2025)
by: Jiaxin Jiang, et al.
Published: (2025)
High‐Performance Nacre‐Inspired 2D Carbon‐Based Nanocomposites
by: Yuchen Li, et al.
Published: (2025)
by: Yuchen Li, et al.
Published: (2025)
CFMD: Dynamic Cross-layer Feature Fusion for Salient Object Detection
by: Lian, Jin, et al.
Published: (2025)
by: Lian, Jin, et al.
Published: (2025)
Primary Mesenteric Malignant Yolk Sac Tumor: A Case Report
by: Lian‐di Liu, et al.
Published: (2025)
by: Lian‐di Liu, et al.
Published: (2025)
Electric circuit analog of Landau-Zener tunneling using time-varying elements
by: Cheng, Enhong, et al.
Published: (2025)
by: Cheng, Enhong, et al.
Published: (2025)
Clinical and laboratory characteristics of salivary gland ultrasonography‐positive patients with primary Sjögren's syndrome
by: Yan Yang, et al.
Published: (2024)
by: Yan Yang, et al.
Published: (2024)
WESE: Weak Exploration to Strong Exploitation for LLM Agents
by: Huang, Xu, et al.
Published: (2024)
by: Huang, Xu, et al.
Published: (2024)
APT: Adaptive Personalized Training for Diffusion Models with Limited Data
by: Chae, JungWoo, et al.
Published: (2025)
by: Chae, JungWoo, et al.
Published: (2025)
Acute Intermittent Porphyria With Epilepsy as the Initial Symptom and Posterior Reversible Encephalopathy Syndrome: A Case Report
by: Wei Li, et al.
Published: (2025)
by: Wei Li, et al.
Published: (2025)
ScalSelect: Scalable Training-Free Multimodal Data Selection for Efficient Visual Instruction Tuning
by: Wu, Changti, et al.
Published: (2026)
by: Wu, Changti, et al.
Published: (2026)
Similar Items
-
SeaPO: Strategic Error Amplification for Robust Preference Optimization of Large Language Models
by: Rao, Jun, et al.
Published: (2025) -
CommonIT: Commonality-Aware Instruction Tuning for Large Language Models via Data Partitions
by: Rao, Jun, et al.
Published: (2024) -
AQuilt: Weaving Logic and Self-Inspection into Low-Cost, High-Relevance Data Synthesis for Specialist LLMs
by: Ke, Xiaopeng, et al.
Published: (2025) -
Dynamic Sampling that Adapts: Self-Aware Iterative Data Persistent Optimization for Mathematical Reasoning
by: Rao, Jun, et al.
Published: (2025) -
Exploring and Enhancing the Transfer of Distribution in Knowledge Distillation for Autoregressive Language Models
by: Rao, Jun, et al.
Published: (2024)