DecIF: Improving Instruction-Following through Meta-Decomposition
Fuente:
arXiv
Salvato in:
| Autori principali: | Hui, Tingfeng, Zhu, Pengyu, Ping, Bowen, Tang, Ling, Dong, Guanting, Zhang, Yaqi, Su, Sen |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Smaller Language Models Are Better Instruction Evolvers
di: Hui, Tingfeng, et al.
Pubblicazione: (2024)
di: Hui, Tingfeng, et al.
Pubblicazione: (2024)
Upcycling Instruction Tuning from Dense to Mixture-of-Experts via Parameter Merging
di: Hui, Tingfeng, et al.
Pubblicazione: (2024)
di: Hui, Tingfeng, et al.
Pubblicazione: (2024)
STT-Arena: A More Realistic Environment for Tool-Using with Spatio-Temporal Dynamics
di: Hui, Tingfeng, et al.
Pubblicazione: (2026)
di: Hui, Tingfeng, et al.
Pubblicazione: (2026)
Noise-BERT: A Unified Perturbation-Robust Framework with Noise Alignment Pre-training for Noisy Slot Filling Task
di: Zhao, Jinxu, et al.
Pubblicazione: (2024)
di: Zhao, Jinxu, et al.
Pubblicazione: (2024)
Toward General Instruction-Following Alignment for Retrieval-Augmented Generation
di: Dong, Guanting, et al.
Pubblicazione: (2024)
di: Dong, Guanting, et al.
Pubblicazione: (2024)
Self-play with Execution Feedback: Improving Instruction-following Capabilities of Large Language Models
di: Dong, Guanting, et al.
Pubblicazione: (2024)
di: Dong, Guanting, et al.
Pubblicazione: (2024)
LongR: Unleashing Long-Context Reasoning via Reinforcement Learning with Dense Utility Rewards
di: Ping, Bowen, et al.
Pubblicazione: (2026)
di: Ping, Bowen, et al.
Pubblicazione: (2026)
MUFFIN: Curating Multi-Faceted Instructions for Improving Instruction-Following
di: Lou, Renze, et al.
Pubblicazione: (2023)
di: Lou, Renze, et al.
Pubblicazione: (2023)
LongAct: Harnessing Intrinsic Activation Patterns for Long-Context Reinforcement Learning
di: Ping, Bowen, et al.
Pubblicazione: (2026)
di: Ping, Bowen, et al.
Pubblicazione: (2026)
LARFT: Closing the Cognition-Action Gap for Length Instruction Following in Large Language Models
di: Zhang, Wei, et al.
Pubblicazione: (2026)
di: Zhang, Wei, et al.
Pubblicazione: (2026)
Following Length Constraints in Instructions
di: Yuan, Weizhe, et al.
Pubblicazione: (2024)
di: Yuan, Weizhe, et al.
Pubblicazione: (2024)
Improving Instruction-Following in Language Models through Activation Steering
di: Stolfo, Alessandro, et al.
Pubblicazione: (2024)
di: Stolfo, Alessandro, et al.
Pubblicazione: (2024)
MuSC: Improving Complex Instruction Following with Multi-granularity Self-Contrastive Training
di: Huang, Hui, et al.
Pubblicazione: (2025)
di: Huang, Hui, et al.
Pubblicazione: (2025)
LIFEBench: Evaluating Length Instruction Following in Large Language Models
di: Zhang, Wei, et al.
Pubblicazione: (2025)
di: Zhang, Wei, et al.
Pubblicazione: (2025)
M3PO: Multimodal-Model-Guided Preference Optimization for Visual Instruction Following
di: Gao, Ruirui, et al.
Pubblicazione: (2025)
di: Gao, Ruirui, et al.
Pubblicazione: (2025)
MedINST: Meta Dataset of Biomedical Instructions
di: Han, Wenhan, et al.
Pubblicazione: (2024)
di: Han, Wenhan, et al.
Pubblicazione: (2024)
CoachLM: Automatic Instruction Revisions Improve the Data Quality in LLM Instruction Tuning
di: Liu, Yilun, et al.
Pubblicazione: (2023)
di: Liu, Yilun, et al.
Pubblicazione: (2023)
Improving Instruction Following in Language Models through Proxy-Based Uncertainty Estimation
di: Lee, JoonHo, et al.
Pubblicazione: (2024)
di: Lee, JoonHo, et al.
Pubblicazione: (2024)
IF-RewardBench: Benchmarking Judge Models for Instruction-Following Evaluation
di: Wen, Bosi, et al.
Pubblicazione: (2026)
di: Wen, Bosi, et al.
Pubblicazione: (2026)
Demonstrating ViviDoc: Generating Interactive Documents through Human-Agent Collaboration
di: Tang, Yinghao, et al.
Pubblicazione: (2026)
di: Tang, Yinghao, et al.
Pubblicazione: (2026)
Learning Instruction-Following Policies through Open-Ended Instruction Relabeling with Large Language Models
di: Zhang, Zhicheng, et al.
Pubblicazione: (2025)
di: Zhang, Zhicheng, et al.
Pubblicazione: (2025)
Instruction Following without Instruction Tuning
di: Hewitt, John, et al.
Pubblicazione: (2024)
di: Hewitt, John, et al.
Pubblicazione: (2024)
SPaR: Self-Play with Tree-Search Refinement to Improve Instruction-Following in Large Language Models
di: Cheng, Jiale, et al.
Pubblicazione: (2024)
di: Cheng, Jiale, et al.
Pubblicazione: (2024)
DotaMath: Decomposition of Thought with Code Assistance and Self-correction for Mathematical Reasoning
di: Li, Chengpeng, et al.
Pubblicazione: (2024)
di: Li, Chengpeng, et al.
Pubblicazione: (2024)
CoGenesis: A Framework Collaborating Large and Small Language Models for Secure Context-Aware Instruction Following
di: Zhang, Kaiyan, et al.
Pubblicazione: (2024)
di: Zhang, Kaiyan, et al.
Pubblicazione: (2024)
MDCure: A Scalable Pipeline for Multi-Document Instruction-Following
di: Liu, Gabrielle Kaili-May, et al.
Pubblicazione: (2024)
di: Liu, Gabrielle Kaili-May, et al.
Pubblicazione: (2024)
FollowTable: A Benchmark for Instruction-Following Table Retrieval
di: Jin, Rihui, et al.
Pubblicazione: (2026)
di: Jin, Rihui, et al.
Pubblicazione: (2026)
DecMetrics: Structured Claim Decomposition Scoring for Factually Consistent LLM Outputs
di: Huang, Minghui
Pubblicazione: (2025)
di: Huang, Minghui
Pubblicazione: (2025)
IHEval: Evaluating Language Models on Following the Instruction Hierarchy
di: Zhang, Zhihan, et al.
Pubblicazione: (2025)
di: Zhang, Zhihan, et al.
Pubblicazione: (2025)
GraphWiz: An Instruction-Following Language Model for Graph Problems
di: Chen, Nuo, et al.
Pubblicazione: (2024)
di: Chen, Nuo, et al.
Pubblicazione: (2024)
MSI-Agent: Incorporating Multi-Scale Insight into Embodied Agents for Superior Planning and Decision-Making
di: Fu, Dayuan, et al.
Pubblicazione: (2024)
di: Fu, Dayuan, et al.
Pubblicazione: (2024)
IF-CRITIC: Towards a Fine-Grained LLM Critic for Instruction-Following Evaluation
di: Wen, Bosi, et al.
Pubblicazione: (2025)
di: Wen, Bosi, et al.
Pubblicazione: (2025)
Conifer: Improving Complex Constrained Instruction-Following Ability of Large Language Models
di: Sun, Haoran, et al.
Pubblicazione: (2024)
di: Sun, Haoran, et al.
Pubblicazione: (2024)
IFDECORATOR: Wrapping Instruction Following Reinforcement Learning with Verifiable Rewards
di: Guo, Xu, et al.
Pubblicazione: (2025)
di: Guo, Xu, et al.
Pubblicazione: (2025)
MetaAlign: Align Large Language Models with Diverse Preferences during Inference Time
di: Zhang, Mozhi, et al.
Pubblicazione: (2024)
di: Zhang, Mozhi, et al.
Pubblicazione: (2024)
Hierarchical Alignment: Enforcing Hierarchical Instruction-Following in LLMs through Logical Consistency
di: Yang, Shu, et al.
Pubblicazione: (2026)
di: Yang, Shu, et al.
Pubblicazione: (2026)
TAG-INSTRUCT: Controlled Instruction Complexity Enhancement through Structure-based Augmentation
di: Zhu, He, et al.
Pubblicazione: (2025)
di: Zhu, He, et al.
Pubblicazione: (2025)
In-Context Examples Matter: Improving Emotion Recognition in Conversation with Instruction Tuning
di: Ma, Hui, et al.
Pubblicazione: (2025)
di: Ma, Hui, et al.
Pubblicazione: (2025)
Generalizing Verifiable Instruction Following
di: Pyatkin, Valentina, et al.
Pubblicazione: (2025)
di: Pyatkin, Valentina, et al.
Pubblicazione: (2025)
Revisiting the Reliability of Language Models in Instruction-Following
di: Dong, Jianshuo, et al.
Pubblicazione: (2025)
di: Dong, Jianshuo, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Smaller Language Models Are Better Instruction Evolvers
di: Hui, Tingfeng, et al.
Pubblicazione: (2024) -
Upcycling Instruction Tuning from Dense to Mixture-of-Experts via Parameter Merging
di: Hui, Tingfeng, et al.
Pubblicazione: (2024) -
STT-Arena: A More Realistic Environment for Tool-Using with Spatio-Temporal Dynamics
di: Hui, Tingfeng, et al.
Pubblicazione: (2026) -
Noise-BERT: A Unified Perturbation-Robust Framework with Noise Alignment Pre-training for Noisy Slot Filling Task
di: Zhao, Jinxu, et al.
Pubblicazione: (2024) -
Toward General Instruction-Following Alignment for Retrieval-Augmented Generation
di: Dong, Guanting, et al.
Pubblicazione: (2024)