OrderChain: Towards General Instruct-Tuning for Stimulating the Ordinal Understanding Ability of MLLM
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Jinhong, Tong, Shuo, liu, Jian, Tang, Dongqi, Wang, Weiqiang, Li, Wentong, Xu, Hongxia, Chen, Danny, Chen, Jintai, Wu, Jian |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
STORM: Benchmarking Visual Rating of MLLMs with a Comprehensive Ordinal Regression Dataset
von: Wang, Jinhong, et al.
Veröffentlicht: (2025)
von: Wang, Jinhong, et al.
Veröffentlicht: (2025)
A Survey on Ordinal Regression: Applications, Advances and Prospects
von: Wang, Jinhong, et al.
Veröffentlicht: (2025)
von: Wang, Jinhong, et al.
Veröffentlicht: (2025)
Scalable Autoregressive Monocular Depth Estimation
von: Wang, Jinhong, et al.
Veröffentlicht: (2024)
von: Wang, Jinhong, et al.
Veröffentlicht: (2024)
LKM-UNet: Large Kernel Vision Mamba UNet for Medical Image Segmentation
von: Wang, Jinhong, et al.
Veröffentlicht: (2024)
von: Wang, Jinhong, et al.
Veröffentlicht: (2024)
Multi-rater Prompting for Ambiguous Medical Image Segmentation
von: Wang, Jinhong, et al.
Veröffentlicht: (2024)
von: Wang, Jinhong, et al.
Veröffentlicht: (2024)
Dual-level Fuzzy Learning with Patch Guidance for Image Ordinal Regression
von: Dong, Chunlai, et al.
Veröffentlicht: (2025)
von: Dong, Chunlai, et al.
Veröffentlicht: (2025)
Osprey: Pixel Understanding with Visual Instruction Tuning
von: Yuan, Yuqian, et al.
Veröffentlicht: (2023)
von: Yuan, Yuqian, et al.
Veröffentlicht: (2023)
DetToolChain: A New Prompting Paradigm to Unleash Detection Ability of MLLM
von: Wu, Yixuan, et al.
Veröffentlicht: (2024)
von: Wu, Yixuan, et al.
Veröffentlicht: (2024)
PoCo: A Self-Supervised Approach via Polar Transformation Based Progressive Contrastive Learning for Ophthalmic Disease Diagnosis
von: Wang, Jinhong, et al.
Veröffentlicht: (2024)
von: Wang, Jinhong, et al.
Veröffentlicht: (2024)
Team up GBDTs and DNNs: Advancing Efficient and Effective Tabular Prediction with Tree-hybrid MLPs
von: Yan, Jiahuan, et al.
Veröffentlicht: (2024)
von: Yan, Jiahuan, et al.
Veröffentlicht: (2024)
Towards Clinical Practice in CT-Based Pulmonary Disease Screening: An Efficient and Reliable Framework
von: Shao, Qian, et al.
Veröffentlicht: (2024)
von: Shao, Qian, et al.
Veröffentlicht: (2024)
Making Pre-trained Language Models Great on Tabular Prediction
von: Yan, Jiahuan, et al.
Veröffentlicht: (2024)
von: Yan, Jiahuan, et al.
Veröffentlicht: (2024)
InstructX: Towards Unified Visual Editing with MLLM Guidance
von: Mou, Chong, et al.
Veröffentlicht: (2025)
von: Mou, Chong, et al.
Veröffentlicht: (2025)
OmniZip: Audio-Guided Dynamic Token Compression for Fast Omnimodal Large Language Models
von: Tao, Keda, et al.
Veröffentlicht: (2025)
von: Tao, Keda, et al.
Veröffentlicht: (2025)
Unraveling Babel: Exploring Multilingual Activation Patterns of LLMs and Their Applications
von: Liu, Weize, et al.
Veröffentlicht: (2024)
von: Liu, Weize, et al.
Veröffentlicht: (2024)
Uncertainty-Instructed Structure Injection for Generalizable HD Map Construction
von: Liu, Xiaolu, et al.
Veröffentlicht: (2025)
von: Liu, Xiaolu, et al.
Veröffentlicht: (2025)
ExcelFormer: A neural network surpassing GBDTs on tabular data
von: Chen, Jintai, et al.
Veröffentlicht: (2023)
von: Chen, Jintai, et al.
Veröffentlicht: (2023)
CurvZO: Adaptive Curvature-Guided Sparse Zeroth-Order Optimization for Efficient LLM Fine-Tuning
von: Wang, Shuo, et al.
Veröffentlicht: (2026)
von: Wang, Shuo, et al.
Veröffentlicht: (2026)
InstructVLA: Vision-Language-Action Instruction Tuning from Understanding to Manipulation
von: Yang, Shuai, et al.
Veröffentlicht: (2025)
von: Yang, Shuai, et al.
Veröffentlicht: (2025)
InstructSing: High-Fidelity Singing Voice Generation via Instructing Yourself
von: Zeng, Chang, et al.
Veröffentlicht: (2024)
von: Zeng, Chang, et al.
Veröffentlicht: (2024)
TeleOR: Real-time Telemedicine System for Full-Scene Operating Room
von: Wu, Yixuan, et al.
Veröffentlicht: (2024)
von: Wu, Yixuan, et al.
Veröffentlicht: (2024)
TokenPacker: Efficient Visual Projector for Multimodal LLM
von: Li, Wentong, et al.
Veröffentlicht: (2024)
von: Li, Wentong, et al.
Veröffentlicht: (2024)
VisionTrim: Unified Vision Token Compression for Training-Free MLLM Acceleration
von: Yu, Hanxun, et al.
Veröffentlicht: (2026)
von: Yu, Hanxun, et al.
Veröffentlicht: (2026)
Inst3D-LMM: Instance-Aware 3D Scene Understanding with Multi-modal Instruction Tuning
von: Yu, Hanxun, et al.
Veröffentlicht: (2025)
von: Yu, Hanxun, et al.
Veröffentlicht: (2025)
Instruct Once, Chat Consistently in Multiple Rounds: An Efficient Tuning Framework for Dialogue
von: Wang, Jian, et al.
Veröffentlicht: (2024)
von: Wang, Jian, et al.
Veröffentlicht: (2024)
Verification of Electromagnetic Fully-kinetic Symplectic Particle-in-cell Method in Microinstabilities Simulation of
von: Xiao, Jianyuan, et al.
Veröffentlicht: (2025)
von: Xiao, Jianyuan, et al.
Veröffentlicht: (2025)
Decision-Level Ordinal Modeling for Multimodal Essay Scoring with Large Language Models
von: Zhang, Han, et al.
Veröffentlicht: (2026)
von: Zhang, Han, et al.
Veröffentlicht: (2026)
Scientists' First Exam: Probing Cognitive Abilities of MLLM via Perception, Understanding, and Reasoning
von: Zhou, Yuhao, et al.
Veröffentlicht: (2025)
von: Zhou, Yuhao, et al.
Veröffentlicht: (2025)
AbsInstruct: Eliciting Abstraction Ability from LLMs through Explanation Tuning with Plausibility Estimation
von: Wang, Zhaowei, et al.
Veröffentlicht: (2024)
von: Wang, Zhaowei, et al.
Veröffentlicht: (2024)
Med-Scout: Curing MLLMs' Geometric Blindness in Medical Perception via Geometry-Aware RL Post-Training
von: Liu, Anglin, et al.
Veröffentlicht: (2026)
von: Liu, Anglin, et al.
Veröffentlicht: (2026)
Decomposing the Basic Abilities of Large Language Models: Mitigating Cross-Task Interference in Multi-Task Instruct-Tuning
von: Wang, Bing, et al.
Veröffentlicht: (2026)
von: Wang, Bing, et al.
Veröffentlicht: (2026)
Utilizing Training Data to Improve LLM Reasoning for Tabular Understanding
von: Gao, Chufan, et al.
Veröffentlicht: (2025)
von: Gao, Chufan, et al.
Veröffentlicht: (2025)
The Principle of Generative Order: A Variational Unification of Sampling Theory, Holography, and Dream Construction
von: jiazheng liu
Veröffentlicht: (2026)
von: jiazheng liu
Veröffentlicht: (2026)
MM-DADM: Multimodal Drug-Aware Diffusion Model for Virtual Clinical Trials
von: Shao, Qian, et al.
Veröffentlicht: (2025)
von: Shao, Qian, et al.
Veröffentlicht: (2025)
Learning What Matters: Dynamic Dimension Selection and Aggregation for Interpretable Vision-Language Reward Modeling
von: Chen, Qiyuan, et al.
Veröffentlicht: (2026)
von: Chen, Qiyuan, et al.
Veröffentlicht: (2026)
Curing Semantic Drift: A Dynamic Approach to Grounding Generation in Large Vision-Language Models
von: Chen, Jiahe, et al.
Veröffentlicht: (2025)
von: Chen, Jiahe, et al.
Veröffentlicht: (2025)
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation
von: Chen, Qiyuan, et al.
Veröffentlicht: (2025)
von: Chen, Qiyuan, et al.
Veröffentlicht: (2025)
Closed-Form Expressions for Odd-Order Riemann Zeta Functions and Their Rigorous Proofs (Part II)
von: liu, shifa
Veröffentlicht: (2026)
von: liu, shifa
Veröffentlicht: (2026)
R-Tuning: Instructing Large Language Models to Say `I Don't Know'
von: Zhang, Hanning, et al.
Veröffentlicht: (2023)
von: Zhang, Hanning, et al.
Veröffentlicht: (2023)
Closed-Form Expressions, Best Approximation Theory, and Arithmetic Properties of Dirichlet L-Functions of Real Order
von: liu, shifa
Veröffentlicht: (2026)
von: liu, shifa
Veröffentlicht: (2026)
Ähnliche Einträge
-
STORM: Benchmarking Visual Rating of MLLMs with a Comprehensive Ordinal Regression Dataset
von: Wang, Jinhong, et al.
Veröffentlicht: (2025) -
A Survey on Ordinal Regression: Applications, Advances and Prospects
von: Wang, Jinhong, et al.
Veröffentlicht: (2025) -
Scalable Autoregressive Monocular Depth Estimation
von: Wang, Jinhong, et al.
Veröffentlicht: (2024) -
LKM-UNet: Large Kernel Vision Mamba UNet for Medical Image Segmentation
von: Wang, Jinhong, et al.
Veröffentlicht: (2024) -
Multi-rater Prompting for Ambiguous Medical Image Segmentation
von: Wang, Jinhong, et al.
Veröffentlicht: (2024)