APEX: Learning Adaptive Priorities for Multi-Objective Alignment in Vision-Language Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Dongliang, Zhuang, Xinlin, Xu, Junjie, Xie, Luojian, Wang, Zehui, Zhuang, Jiaxi, Yang, Haolin, Dou, Liang, He, Xiao, Wu, Xingjiao, Qian, Ying |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ClinCoT: Clinical-Aware Visual Chain-of-Thought for Medical Vision Language Models
by: Liu, Xiwei, et al.
Published: (2026)
by: Liu, Xiwei, et al.
Published: (2026)
Multi-Type Preference Learning: Empowering Preference-Based Reinforcement Learning with Equal Preferences
by: Liu, Ziang, et al.
Published: (2024)
by: Liu, Ziang, et al.
Published: (2024)
Copy-Augmented Representation for Structure Invariant Template-Free Retrosynthesis
by: Zhuang, Jiaxi, et al.
Published: (2025)
by: Zhuang, Jiaxi, et al.
Published: (2025)
Towards Robust Visual Continual Learning with Multi-Prototype Supervision
by: Liu, Xiwei, et al.
Published: (2025)
by: Liu, Xiwei, et al.
Published: (2025)
Dynamic Multimodal Activation Steering for Hallucination Mitigation in Large Vision-Language Models
by: Yin, Jianghao, et al.
Published: (2026)
by: Yin, Jianghao, et al.
Published: (2026)
Pareto Multi-Objective Alignment for Language Models
by: He, Qiang, et al.
Published: (2025)
by: He, Qiang, et al.
Published: (2025)
MOSAIC: Multi-Objective Slice-Aware Iterative Curation for Alignment
by: Dou, Yipu, et al.
Published: (2026)
by: Dou, Yipu, et al.
Published: (2026)
Meta-rater: A Multi-dimensional Data Selection Method for Pre-training Language Models
by: Zhuang, Xinlin, et al.
Published: (2025)
by: Zhuang, Xinlin, et al.
Published: (2025)
Retro3D: A 3D-aware Template-free Method for Enhancing Retrosynthesis via Molecular Conformer Information
by: Zhuang, Jiaxi, et al.
Published: (2025)
by: Zhuang, Jiaxi, et al.
Published: (2025)
scAGC: Learning Adaptive Cell Graphs with Contrastive Guidance for Single-Cell Clustering
by: Li, Huifa, et al.
Published: (2025)
by: Li, Huifa, et al.
Published: (2025)
Gradient-Adaptive Policy Optimization: Towards Multi-Objective Alignment of Large Language Models
by: Li, Chengao, et al.
Published: (2025)
by: Li, Chengao, et al.
Published: (2025)
FairMonitor: A Dual-framework for Detecting Stereotypes and Biases in Large Language Models
by: Bai, Yanhong, et al.
Published: (2024)
by: Bai, Yanhong, et al.
Published: (2024)
AMO: Adaptive Muon Orthogonalization
by: Zhuang, Xinlin, et al.
Published: (2026)
by: Zhuang, Xinlin, et al.
Published: (2026)
CHIPS: Efficient CLIP Adaptation via Curvature-aware Hybrid Influence-based Data Selection
by: Zhuang, Xinlin, et al.
Published: (2025)
by: Zhuang, Xinlin, et al.
Published: (2025)
APEX$^2$: Adaptive and Extreme Summarization for Personalized Knowledge Graphs
by: Li, Zihao, et al.
Published: (2024)
by: Li, Zihao, et al.
Published: (2024)
Cross-Subject EEG Emotion Recognition Based on Temporal Asynchronous Alignment Contrastive Learning
by: Xie, Ying, et al.
Published: (2026)
by: Xie, Ying, et al.
Published: (2026)
Topic Over Source: The Key to Effective Data Mixing for Language Models Pre-training
by: Peng, Jiahui, et al.
Published: (2025)
by: Peng, Jiahui, et al.
Published: (2025)
Towards Efficient Medical Reasoning with Minimal Fine-Tuning Data
by: Zhuang, Xinlin, et al.
Published: (2025)
by: Zhuang, Xinlin, et al.
Published: (2025)
PanoGen++: Domain-Adapted Text-Guided Panoramic Environment Generation for Vision-and-Language Navigation
by: Wang, Sen, et al.
Published: (2025)
by: Wang, Sen, et al.
Published: (2025)
MindScope: Exploring cognitive biases in large language models through Multi-Agent Systems
by: Xie, Zhentao, et al.
Published: (2024)
by: Xie, Zhentao, et al.
Published: (2024)
MetaAligner: Towards Generalizable Multi-Objective Alignment of Language Models
by: Yang, Kailai, et al.
Published: (2024)
by: Yang, Kailai, et al.
Published: (2024)
Magnet: We Never Know How Text-to-Image Diffusion Models Work, Until We Learn How Vision-Language Models Function
by: Zhuang, Chenyi, et al.
Published: (2024)
by: Zhuang, Chenyi, et al.
Published: (2024)
Unveiling Deep Semantic Uncertainty Perception for Language-Anchored Multi-modal Vision-Brain Alignment
by: Feng, Zehui, et al.
Published: (2025)
by: Feng, Zehui, et al.
Published: (2025)
Draft-Refine-Optimize: Self-Evolved Learning for Natural Language to MongoDB Query Generation
by: Ye, Mingwei, et al.
Published: (2026)
by: Ye, Mingwei, et al.
Published: (2026)
Aesthetic Matters in Music Perception for Image Stylization: A Emotion-driven Music-to-Visual Manipulation
by: Xu, Junjie, et al.
Published: (2025)
by: Xu, Junjie, et al.
Published: (2025)
ST-Booster: An Iterative SpatioTemporal Perception Booster for Vision-and-Language Navigation in Continuous Environments
by: Yue, Lu, et al.
Published: (2025)
by: Yue, Lu, et al.
Published: (2025)
Generating Vision-Language Navigation Instructions Incorporated Fine-Grained Alignment Annotations
by: Cui, Yibo, et al.
Published: (2025)
by: Cui, Yibo, et al.
Published: (2025)
Lookahead: An Inference Acceleration Framework for Large Language Model with Lossless Generation Accuracy
by: Zhao, Yao, et al.
Published: (2023)
by: Zhao, Yao, et al.
Published: (2023)
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models
by: Chen, Yiteng, et al.
Published: (2025)
by: Chen, Yiteng, et al.
Published: (2025)
Multi-Objective Alignment of Language Models for Personalized Psychotherapy
by: Beikzadeh, Mehrab, et al.
Published: (2026)
by: Beikzadeh, Mehrab, et al.
Published: (2026)
PiCo: Enhancing Text-Image Alignment with Improved Noise Selection and Precise Mask Control in Diffusion Models
by: Xie, Chang, et al.
Published: (2025)
by: Xie, Chang, et al.
Published: (2025)
Data-Adaptive Graph Framelets with Generalized Vanishing Moments for Graph Machine Learning
by: Zheng, Ruigang, et al.
Published: (2023)
by: Zheng, Ruigang, et al.
Published: (2023)
APEX-Agents
by: Vidgen, Bertie, et al.
Published: (2026)
by: Vidgen, Bertie, et al.
Published: (2026)
APEX-SWE
by: Kottamasu, Abhi, et al.
Published: (2026)
by: Kottamasu, Abhi, et al.
Published: (2026)
Adaptive Duration Model for Text Speech Alignment
by: Cao, Junjie
Published: (2025)
by: Cao, Junjie
Published: (2025)
Flatness Guided Test-Time Adaptation for Vision-Language Models
by: Li, Aodi, et al.
Published: (2025)
by: Li, Aodi, et al.
Published: (2025)
Are Large Vision Language Models Good Game Players?
by: Wang, Xinyu, et al.
Published: (2025)
by: Wang, Xinyu, et al.
Published: (2025)
Context-Adaptive Multi-Prompt Embedding with Large Language Models for Vision-Language Alignment
by: Kim, Dahun, et al.
Published: (2025)
by: Kim, Dahun, et al.
Published: (2025)
Figure 1 from: Wu X-Q, Zhuang W-Y, Zeng Z-Q (2026) Two new species of Pseudocosmospora (Hypocreales) revealed through morphological and phylogenetic analyses. MycoKeys 128: 285-300. https://doi.org/10.3897/mycokeys.128.180941
by: Wu, Xiao-Qian, et al.
Published: (2026)
by: Wu, Xiao-Qian, et al.
Published: (2026)
Figure 5 from: Wu X-Q, Zhuang W-Y, Zeng Z-Q (2026) Two new species of Pseudocosmospora (Hypocreales) revealed through morphological and phylogenetic analyses. MycoKeys 128: 285-300. https://doi.org/10.3897/mycokeys.128.180941
by: Wu, Xiao-Qian, et al.
Published: (2026)
by: Wu, Xiao-Qian, et al.
Published: (2026)
Similar Items
-
ClinCoT: Clinical-Aware Visual Chain-of-Thought for Medical Vision Language Models
by: Liu, Xiwei, et al.
Published: (2026) -
Multi-Type Preference Learning: Empowering Preference-Based Reinforcement Learning with Equal Preferences
by: Liu, Ziang, et al.
Published: (2024) -
Copy-Augmented Representation for Structure Invariant Template-Free Retrosynthesis
by: Zhuang, Jiaxi, et al.
Published: (2025) -
Towards Robust Visual Continual Learning with Multi-Prototype Supervision
by: Liu, Xiwei, et al.
Published: (2025) -
Dynamic Multimodal Activation Steering for Hallucination Mitigation in Large Vision-Language Models
by: Yin, Jianghao, et al.
Published: (2026)