DecIF: Improving Instruction-Following through Meta-Decomposition
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Hui, Tingfeng, Zhu, Pengyu, Ping, Bowen, Tang, Ling, Dong, Guanting, Zhang, Yaqi, Su, Sen |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Smaller Language Models Are Better Instruction Evolvers
par: Hui, Tingfeng, et autres
Publié: (2024)
par: Hui, Tingfeng, et autres
Publié: (2024)
Upcycling Instruction Tuning from Dense to Mixture-of-Experts via Parameter Merging
par: Hui, Tingfeng, et autres
Publié: (2024)
par: Hui, Tingfeng, et autres
Publié: (2024)
STT-Arena: A More Realistic Environment for Tool-Using with Spatio-Temporal Dynamics
par: Hui, Tingfeng, et autres
Publié: (2026)
par: Hui, Tingfeng, et autres
Publié: (2026)
Noise-BERT: A Unified Perturbation-Robust Framework with Noise Alignment Pre-training for Noisy Slot Filling Task
par: Zhao, Jinxu, et autres
Publié: (2024)
par: Zhao, Jinxu, et autres
Publié: (2024)
Toward General Instruction-Following Alignment for Retrieval-Augmented Generation
par: Dong, Guanting, et autres
Publié: (2024)
par: Dong, Guanting, et autres
Publié: (2024)
Self-play with Execution Feedback: Improving Instruction-following Capabilities of Large Language Models
par: Dong, Guanting, et autres
Publié: (2024)
par: Dong, Guanting, et autres
Publié: (2024)
LongR: Unleashing Long-Context Reasoning via Reinforcement Learning with Dense Utility Rewards
par: Ping, Bowen, et autres
Publié: (2026)
par: Ping, Bowen, et autres
Publié: (2026)
MUFFIN: Curating Multi-Faceted Instructions for Improving Instruction-Following
par: Lou, Renze, et autres
Publié: (2023)
par: Lou, Renze, et autres
Publié: (2023)
LongAct: Harnessing Intrinsic Activation Patterns for Long-Context Reinforcement Learning
par: Ping, Bowen, et autres
Publié: (2026)
par: Ping, Bowen, et autres
Publié: (2026)
LARFT: Closing the Cognition-Action Gap for Length Instruction Following in Large Language Models
par: Zhang, Wei, et autres
Publié: (2026)
par: Zhang, Wei, et autres
Publié: (2026)
Following Length Constraints in Instructions
par: Yuan, Weizhe, et autres
Publié: (2024)
par: Yuan, Weizhe, et autres
Publié: (2024)
Improving Instruction-Following in Language Models through Activation Steering
par: Stolfo, Alessandro, et autres
Publié: (2024)
par: Stolfo, Alessandro, et autres
Publié: (2024)
MuSC: Improving Complex Instruction Following with Multi-granularity Self-Contrastive Training
par: Huang, Hui, et autres
Publié: (2025)
par: Huang, Hui, et autres
Publié: (2025)
LIFEBench: Evaluating Length Instruction Following in Large Language Models
par: Zhang, Wei, et autres
Publié: (2025)
par: Zhang, Wei, et autres
Publié: (2025)
M3PO: Multimodal-Model-Guided Preference Optimization for Visual Instruction Following
par: Gao, Ruirui, et autres
Publié: (2025)
par: Gao, Ruirui, et autres
Publié: (2025)
MedINST: Meta Dataset of Biomedical Instructions
par: Han, Wenhan, et autres
Publié: (2024)
par: Han, Wenhan, et autres
Publié: (2024)
CoachLM: Automatic Instruction Revisions Improve the Data Quality in LLM Instruction Tuning
par: Liu, Yilun, et autres
Publié: (2023)
par: Liu, Yilun, et autres
Publié: (2023)
Improving Instruction Following in Language Models through Proxy-Based Uncertainty Estimation
par: Lee, JoonHo, et autres
Publié: (2024)
par: Lee, JoonHo, et autres
Publié: (2024)
IF-RewardBench: Benchmarking Judge Models for Instruction-Following Evaluation
par: Wen, Bosi, et autres
Publié: (2026)
par: Wen, Bosi, et autres
Publié: (2026)
Demonstrating ViviDoc: Generating Interactive Documents through Human-Agent Collaboration
par: Tang, Yinghao, et autres
Publié: (2026)
par: Tang, Yinghao, et autres
Publié: (2026)
Learning Instruction-Following Policies through Open-Ended Instruction Relabeling with Large Language Models
par: Zhang, Zhicheng, et autres
Publié: (2025)
par: Zhang, Zhicheng, et autres
Publié: (2025)
Instruction Following without Instruction Tuning
par: Hewitt, John, et autres
Publié: (2024)
par: Hewitt, John, et autres
Publié: (2024)
SPaR: Self-Play with Tree-Search Refinement to Improve Instruction-Following in Large Language Models
par: Cheng, Jiale, et autres
Publié: (2024)
par: Cheng, Jiale, et autres
Publié: (2024)
DotaMath: Decomposition of Thought with Code Assistance and Self-correction for Mathematical Reasoning
par: Li, Chengpeng, et autres
Publié: (2024)
par: Li, Chengpeng, et autres
Publié: (2024)
CoGenesis: A Framework Collaborating Large and Small Language Models for Secure Context-Aware Instruction Following
par: Zhang, Kaiyan, et autres
Publié: (2024)
par: Zhang, Kaiyan, et autres
Publié: (2024)
MDCure: A Scalable Pipeline for Multi-Document Instruction-Following
par: Liu, Gabrielle Kaili-May, et autres
Publié: (2024)
par: Liu, Gabrielle Kaili-May, et autres
Publié: (2024)
FollowTable: A Benchmark for Instruction-Following Table Retrieval
par: Jin, Rihui, et autres
Publié: (2026)
par: Jin, Rihui, et autres
Publié: (2026)
DecMetrics: Structured Claim Decomposition Scoring for Factually Consistent LLM Outputs
par: Huang, Minghui
Publié: (2025)
par: Huang, Minghui
Publié: (2025)
IHEval: Evaluating Language Models on Following the Instruction Hierarchy
par: Zhang, Zhihan, et autres
Publié: (2025)
par: Zhang, Zhihan, et autres
Publié: (2025)
GraphWiz: An Instruction-Following Language Model for Graph Problems
par: Chen, Nuo, et autres
Publié: (2024)
par: Chen, Nuo, et autres
Publié: (2024)
MSI-Agent: Incorporating Multi-Scale Insight into Embodied Agents for Superior Planning and Decision-Making
par: Fu, Dayuan, et autres
Publié: (2024)
par: Fu, Dayuan, et autres
Publié: (2024)
IF-CRITIC: Towards a Fine-Grained LLM Critic for Instruction-Following Evaluation
par: Wen, Bosi, et autres
Publié: (2025)
par: Wen, Bosi, et autres
Publié: (2025)
Conifer: Improving Complex Constrained Instruction-Following Ability of Large Language Models
par: Sun, Haoran, et autres
Publié: (2024)
par: Sun, Haoran, et autres
Publié: (2024)
IFDECORATOR: Wrapping Instruction Following Reinforcement Learning with Verifiable Rewards
par: Guo, Xu, et autres
Publié: (2025)
par: Guo, Xu, et autres
Publié: (2025)
MetaAlign: Align Large Language Models with Diverse Preferences during Inference Time
par: Zhang, Mozhi, et autres
Publié: (2024)
par: Zhang, Mozhi, et autres
Publié: (2024)
Hierarchical Alignment: Enforcing Hierarchical Instruction-Following in LLMs through Logical Consistency
par: Yang, Shu, et autres
Publié: (2026)
par: Yang, Shu, et autres
Publié: (2026)
TAG-INSTRUCT: Controlled Instruction Complexity Enhancement through Structure-based Augmentation
par: Zhu, He, et autres
Publié: (2025)
par: Zhu, He, et autres
Publié: (2025)
In-Context Examples Matter: Improving Emotion Recognition in Conversation with Instruction Tuning
par: Ma, Hui, et autres
Publié: (2025)
par: Ma, Hui, et autres
Publié: (2025)
Generalizing Verifiable Instruction Following
par: Pyatkin, Valentina, et autres
Publié: (2025)
par: Pyatkin, Valentina, et autres
Publié: (2025)
Revisiting the Reliability of Language Models in Instruction-Following
par: Dong, Jianshuo, et autres
Publié: (2025)
par: Dong, Jianshuo, et autres
Publié: (2025)
Documents similaires
-
Smaller Language Models Are Better Instruction Evolvers
par: Hui, Tingfeng, et autres
Publié: (2024) -
Upcycling Instruction Tuning from Dense to Mixture-of-Experts via Parameter Merging
par: Hui, Tingfeng, et autres
Publié: (2024) -
STT-Arena: A More Realistic Environment for Tool-Using with Spatio-Temporal Dynamics
par: Hui, Tingfeng, et autres
Publié: (2026) -
Noise-BERT: A Unified Perturbation-Robust Framework with Noise Alignment Pre-training for Noisy Slot Filling Task
par: Zhao, Jinxu, et autres
Publié: (2024) -
Toward General Instruction-Following Alignment for Retrieval-Augmented Generation
par: Dong, Guanting, et autres
Publié: (2024)