MuSC: Improving Complex Instruction Following with Multi-granularity Self-Contrastive Training
Fuente:
arXiv
Guardado en:
| Autores principales: | Huang, Hui, Liu, Jiaheng, He, Yancheng, Li, Shilong, Xu, Bing, Zhu, Conghui, Yang, Muyun, Zhao, Tiejun |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Long-form RewardBench: Evaluating Reward Models for Long-form Generation
por: Huang, Hui, et al.
Publicado: (2026)
por: Huang, Hui, et al.
Publicado: (2026)
AIR: Complex Instruction Generation via Automatic Iterative Refinement
por: Liu, Wei, et al.
Publicado: (2025)
por: Liu, Wei, et al.
Publicado: (2025)
Toward Robust LLM-Based Judges: Taxonomic Bias Evaluation and Debiasing Optimization
por: Zhou, Hongli, et al.
Publicado: (2026)
por: Zhou, Hongli, et al.
Publicado: (2026)
Mitigating the Bias of Large Language Model Evaluation
por: Zhou, Hongli, et al.
Publicado: (2024)
por: Zhou, Hongli, et al.
Publicado: (2024)
Self-Evaluation of Large Language Model based on Glass-box Features
por: Huang, Hui, et al.
Publicado: (2024)
por: Huang, Hui, et al.
Publicado: (2024)
LoRA-drop: Efficient LoRA Parameter Pruning based on Output Evaluation
por: Zhou, Hongyun, et al.
Publicado: (2024)
por: Zhou, Hongyun, et al.
Publicado: (2024)
DuplexMamba: Enhancing Real-time Speech Conversations with Duplex and Streaming Capabilities
por: Lu, Xiangyu, et al.
Publicado: (2025)
por: Lu, Xiangyu, et al.
Publicado: (2025)
An Empirical Study of LLM-as-a-Judge for LLM Evaluation: Fine-tuned Judge Model is not a General Substitute for GPT-4
por: Huang, Hui, et al.
Publicado: (2024)
por: Huang, Hui, et al.
Publicado: (2024)
Lost in Benchmarks? Rethinking Large Language Model Benchmarking with Item Response Theory
por: Zhou, Hongli, et al.
Publicado: (2025)
por: Zhou, Hongli, et al.
Publicado: (2025)
PMoL: Parameter Efficient MoE for Preference Mixing of LLM Alignment
por: Liu, Dongxu, et al.
Publicado: (2024)
por: Liu, Dongxu, et al.
Publicado: (2024)
Beyond Token-Level Policy Gradients for Complex Reasoning with Large Language Models
por: Xu, Mufan, et al.
Publicado: (2026)
por: Xu, Mufan, et al.
Publicado: (2026)
MUFFIN: Curating Multi-Faceted Instructions for Improving Instruction-Following
por: Lou, Renze, et al.
Publicado: (2023)
por: Lou, Renze, et al.
Publicado: (2023)
RECAST: Expanding the Boundaries of LLMs' Complex Instruction Following with Multi-Constraint Data
por: Guo, Zhengkang, et al.
Publicado: (2025)
por: Guo, Zhengkang, et al.
Publicado: (2025)
Thinking with Comics: Enhancing Multimodal Reasoning through Structured Visual Storytelling
por: Chen, Andong, et al.
Publicado: (2026)
por: Chen, Andong, et al.
Publicado: (2026)
LLM-based Discriminative Reasoning for Knowledge Graph Question Answering
por: Xu, Mufan, et al.
Publicado: (2024)
por: Xu, Mufan, et al.
Publicado: (2024)
DecIF: Improving Instruction-Following through Meta-Decomposition
por: Hui, Tingfeng, et al.
Publicado: (2025)
por: Hui, Tingfeng, et al.
Publicado: (2025)
From Complex to Simple: Enhancing Multi-Constraint Complex Instruction Following Ability of Large Language Models
por: He, Qianyu, et al.
Publicado: (2024)
por: He, Qianyu, et al.
Publicado: (2024)
2D-DPO: Scaling Direct Preference Optimization with 2-Dimensional Supervision
por: Li, Shilong, et al.
Publicado: (2024)
por: Li, Shilong, et al.
Publicado: (2024)
Large Language Models for Classical Chinese Poetry Translation: Benchmarking, Evaluating, and Improving
por: Chen, Andong, et al.
Publicado: (2024)
por: Chen, Andong, et al.
Publicado: (2024)
DiVA: Fine-grained Factuality Verification with Agentic-Discriminative Verifier
por: Huang, Hui, et al.
Publicado: (2026)
por: Huang, Hui, et al.
Publicado: (2026)
Speculative Decoding Meets Quantization: Compatibility Evaluation and Hierarchical Framework Design
por: Zhang, Yudi, et al.
Publicado: (2025)
por: Zhang, Yudi, et al.
Publicado: (2025)
User-Aware Active Knowledge Acquisition for Emotional Support Dialogue
por: Xu, Mufan, et al.
Publicado: (2026)
por: Xu, Mufan, et al.
Publicado: (2026)
Constraint Back-translation Improves Complex Instruction Following of Large Language Models
por: Qi, Yunjia, et al.
Publicado: (2024)
por: Qi, Yunjia, et al.
Publicado: (2024)
Conifer: Improving Complex Constrained Instruction-Following Ability of Large Language Models
por: Sun, Haoran, et al.
Publicado: (2024)
por: Sun, Haoran, et al.
Publicado: (2024)
Dual Instruction Tuning with Large Language Models for Mathematical Reasoning
por: Zhou, Yongwei, et al.
Publicado: (2024)
por: Zhou, Yongwei, et al.
Publicado: (2024)
Culture In a Frame: C$^3$B as a Comic-Based Benchmark for Multimodal Culturally Awareness
por: Song, Yuchen, et al.
Publicado: (2025)
por: Song, Yuchen, et al.
Publicado: (2025)
Evaluating o1-Like LLMs: Unlocking Reasoning for Translation through Comprehensive Analysis
por: Chen, Andong, et al.
Publicado: (2025)
por: Chen, Andong, et al.
Publicado: (2025)
Reverse Preference Optimization for Complex Instruction Following
por: Huang, Xiang, et al.
Publicado: (2025)
por: Huang, Xiang, et al.
Publicado: (2025)
Make Imagination Clearer! Stable Diffusion-based Visual Imagination for Multimodal Machine Translation
por: Chen, Andong, et al.
Publicado: (2024)
por: Chen, Andong, et al.
Publicado: (2024)
Benchmarking Complex Instruction-Following with Multiple Constraints Composition
por: Wen, Bosi, et al.
Publicado: (2024)
por: Wen, Bosi, et al.
Publicado: (2024)
Beyond Static Bias: Adaptive Multi-Fidelity Bandits with Improving Proxies
por: Lu, Muyun, et al.
Publicado: (2026)
por: Lu, Muyun, et al.
Publicado: (2026)
SEIF: Self-Evolving Reinforcement Learning for Instruction Following
por: Ren, Qingyu, et al.
Publicado: (2026)
por: Ren, Qingyu, et al.
Publicado: (2026)
Video2Layout: Recall and Reconstruct Metric-Grounded Cognitive Map for Spatial Reasoning
por: Huang, Yibin, et al.
Publicado: (2025)
por: Huang, Yibin, et al.
Publicado: (2025)
MuCo: Multi-turn Contrastive Learning for Multimodal Embedding Model
por: Gu, Geonmo, et al.
Publicado: (2026)
por: Gu, Geonmo, et al.
Publicado: (2026)
Multi-IF: Benchmarking LLMs on Multi-Turn and Multilingual Instructions Following
por: He, Yun, et al.
Publicado: (2024)
por: He, Yun, et al.
Publicado: (2024)
Light-IF: Endowing LLMs with Generalizable Reasoning via Preview and Self-Checking for Complex Instruction Following
por: Wang, Chenyang, et al.
Publicado: (2025)
por: Wang, Chenyang, et al.
Publicado: (2025)
Sparse Activation Editing for Reliable Instruction Following in Narratives
por: Zhao, Runcong, et al.
Publicado: (2025)
por: Zhao, Runcong, et al.
Publicado: (2025)
Dual-View Training for Instruction-Following Information Retrieval
por: Zeng, Qingcheng, et al.
Publicado: (2026)
por: Zeng, Qingcheng, et al.
Publicado: (2026)
Training with Pseudo-Code for Instruction Following
por: Kumar, Prince, et al.
Publicado: (2025)
por: Kumar, Prince, et al.
Publicado: (2025)
From Perception to Reasoning: Deep Thinking Empowers Multimodal Large Language Models
por: Zhu, Wenxin, et al.
Publicado: (2025)
por: Zhu, Wenxin, et al.
Publicado: (2025)
Ejemplares similares
-
Long-form RewardBench: Evaluating Reward Models for Long-form Generation
por: Huang, Hui, et al.
Publicado: (2026) -
AIR: Complex Instruction Generation via Automatic Iterative Refinement
por: Liu, Wei, et al.
Publicado: (2025) -
Toward Robust LLM-Based Judges: Taxonomic Bias Evaluation and Debiasing Optimization
por: Zhou, Hongli, et al.
Publicado: (2026) -
Mitigating the Bias of Large Language Model Evaluation
por: Zhou, Hongli, et al.
Publicado: (2024) -
Self-Evaluation of Large Language Model based on Glass-box Features
por: Huang, Hui, et al.
Publicado: (2024)