Turning Drift into Constraint: Robust Reasoning Alignment in Non-Stationary Multi-Stream Environments
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yang, Xiaoyu, Yu, En, Duan, Wei, Lu, Jie |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Resilient Contrastive Pre-training under Non-Stationary Drift
von: Yang, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Yang, Xiaoyu, et al.
Veröffentlicht: (2025)
Walking the Tightrope: Disentangling Beneficial and Detrimental Drifts in Non-Stationary Custom-Tuning
von: Yang, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Yang, Xiaoyu, et al.
Veröffentlicht: (2025)
Autonomous Drift Learning in Data Streams: A Unified Perspective
von: Yang, Xiaoyu, et al.
Veröffentlicht: (2026)
von: Yang, Xiaoyu, et al.
Veröffentlicht: (2026)
Towards Robust Endogenous Reasoning: Unifying Drift Adaptation in Non-Stationary Tuning
von: Yang, Xiaoyu, et al.
Veröffentlicht: (2026)
von: Yang, Xiaoyu, et al.
Veröffentlicht: (2026)
Adapting Multi-modal Large Language Model to Concept Drift From Pre-training Onwards
von: Yang, Xiaoyu, et al.
Veröffentlicht: (2024)
von: Yang, Xiaoyu, et al.
Veröffentlicht: (2024)
Robust Multi-View Learning via Representation Fusion of Sample-Level Attention and Alignment of Simulated Perturbation
von: Xu, Jie, et al.
Veröffentlicht: (2025)
von: Xu, Jie, et al.
Veröffentlicht: (2025)
Bespoke Non-Stationary Solvers for Fast Sampling of Diffusion and Flow Models
von: Shaul, Neta, et al.
Veröffentlicht: (2024)
von: Shaul, Neta, et al.
Veröffentlicht: (2024)
MMDU: A Multi-Turn Multi-Image Dialog Understanding Benchmark and Instruction-Tuning Dataset for LVLMs
von: Liu, Ziyu, et al.
Veröffentlicht: (2024)
von: Liu, Ziyu, et al.
Veröffentlicht: (2024)
When Thinking Drifts: Evidential Grounding for Robust Video Reasoning
von: Luo, Mi, et al.
Veröffentlicht: (2025)
von: Luo, Mi, et al.
Veröffentlicht: (2025)
Towards Multi-dimensional Explanation Alignment for Medical Classification
von: Hu, Lijie, et al.
Veröffentlicht: (2024)
von: Hu, Lijie, et al.
Veröffentlicht: (2024)
Streaming 4D Visual Geometry Transformer
von: Zhuo, Dong, et al.
Veröffentlicht: (2025)
von: Zhuo, Dong, et al.
Veröffentlicht: (2025)
Point3R: Streaming 3D Reconstruction with Explicit Spatial Pointer Memory
von: Wu, Yuqi, et al.
Veröffentlicht: (2025)
von: Wu, Yuqi, et al.
Veröffentlicht: (2025)
GaussianWorld: Gaussian World Model for Streaming 3D Occupancy Prediction
von: Zuo, Sicheng, et al.
Veröffentlicht: (2024)
von: Zuo, Sicheng, et al.
Veröffentlicht: (2024)
SPHINX: A Synthetic Environment for Visual Perception and Reasoning
von: Alam, Md Tanvirul, et al.
Veröffentlicht: (2025)
von: Alam, Md Tanvirul, et al.
Veröffentlicht: (2025)
Think While Watching: Online Streaming Segment-Level Memory for Multi-Turn Video Reasoning in Multimodal Large Language Models
von: Wang, Lu, et al.
Veröffentlicht: (2026)
von: Wang, Lu, et al.
Veröffentlicht: (2026)
LINA: Learning INterventions Adaptively for Physical Alignment and Generalization in Diffusion Models
von: Yu, Shu, et al.
Veröffentlicht: (2025)
von: Yu, Shu, et al.
Veröffentlicht: (2025)
Learning Visual Abstract Reasoning through Dual-Stream Networks
von: Zhao, Kai, et al.
Veröffentlicht: (2024)
von: Zhao, Kai, et al.
Veröffentlicht: (2024)
Mitigating Object and Action Hallucinations in Multimodal LLMs via Self-Augmented Contrastive Alignment
von: Chang, Kai-Po, et al.
Veröffentlicht: (2025)
von: Chang, Kai-Po, et al.
Veröffentlicht: (2025)
EdiVal-Agent: An Object-Centric Framework for Automated, Fine-Grained Evaluation of Multi-Turn Editing
von: Chen, Tianyu, et al.
Veröffentlicht: (2025)
von: Chen, Tianyu, et al.
Veröffentlicht: (2025)
Streaming Video Understanding and Multi-round Interaction with Memory-enhanced Knowledge
von: Xiong, Haomiao, et al.
Veröffentlicht: (2025)
von: Xiong, Haomiao, et al.
Veröffentlicht: (2025)
Proactive Agents for Multi-Turn Text-to-Image Generation Under Uncertainty
von: Hahn, Meera, et al.
Veröffentlicht: (2024)
von: Hahn, Meera, et al.
Veröffentlicht: (2024)
IMAGAgent: Orchestrating Multi-Turn Image Editing via Constraint-Aware Planning and Reflection
von: Shen, Fei, et al.
Veröffentlicht: (2026)
von: Shen, Fei, et al.
Veröffentlicht: (2026)
Lookahead Drifting Model
von: Zhang, Guoqiang, et al.
Veröffentlicht: (2026)
von: Zhang, Guoqiang, et al.
Veröffentlicht: (2026)
VGAS: Value-Guided Action-Chunk Selection for Few-Shot Vision-Language-Action Adaptation
von: Xu, Changhua, et al.
Veröffentlicht: (2026)
von: Xu, Changhua, et al.
Veröffentlicht: (2026)
Adaptive Parameter Optimization for Robust Remote Photoplethysmography
von: Morales, Cecilia G., et al.
Veröffentlicht: (2025)
von: Morales, Cecilia G., et al.
Veröffentlicht: (2025)
Force Matching with Relativistic Constraints: A Physics-Inspired Approach to Stable and Efficient Generative Modeling
von: Cao, Yang, et al.
Veröffentlicht: (2025)
von: Cao, Yang, et al.
Veröffentlicht: (2025)
MLVTG: Mamba-Based Feature Alignment and LLM-Driven Purification for Multi-Modal Video Temporal Grounding
von: Zhu, Zhiyi, et al.
Veröffentlicht: (2025)
von: Zhu, Zhiyi, et al.
Veröffentlicht: (2025)
ThinkAct: Vision-Language-Action Reasoning via Reinforced Visual Latent Planning
von: Huang, Chi-Pin, et al.
Veröffentlicht: (2025)
von: Huang, Chi-Pin, et al.
Veröffentlicht: (2025)
Imputation-free and Alignment-free: Incomplete Multi-view Clustering Driven by Consensus Semantic Learning
von: Dai, Yuzhuo, et al.
Veröffentlicht: (2025)
von: Dai, Yuzhuo, et al.
Veröffentlicht: (2025)
UniGame: Turning a Unified Multimodal Model Into Its Own Adversary
von: Su, Zhaolong, et al.
Veröffentlicht: (2025)
von: Su, Zhaolong, et al.
Veröffentlicht: (2025)
Fast-ThinkAct: Efficient Vision-Language-Action Reasoning via Verbalizable Latent Planning
von: Huang, Chi-Pin, et al.
Veröffentlicht: (2026)
von: Huang, Chi-Pin, et al.
Veröffentlicht: (2026)
Logit Calibration and Feature Contrast for Robust Federated Learning on Non-IID Data
von: Qiao, Yu, et al.
Veröffentlicht: (2024)
von: Qiao, Yu, et al.
Veröffentlicht: (2024)
Reinforcing VLMs to Use Tools for Detailed Visual Reasoning Under Resource Constraints
von: Kumar, Sunil, et al.
Veröffentlicht: (2025)
von: Kumar, Sunil, et al.
Veröffentlicht: (2025)
Data Stream Sampling with Fuzzy Task Boundaries and Noisy Labels
von: Chen, Yu-Hsi
Veröffentlicht: (2024)
von: Chen, Yu-Hsi
Veröffentlicht: (2024)
Contrastive Continual Multi-view Clustering with Filtered Structural Fusion
von: Wan, Xinhang, et al.
Veröffentlicht: (2023)
von: Wan, Xinhang, et al.
Veröffentlicht: (2023)
Memory-V2V: Memory-Augmented Video-to-Video Diffusion for Consistent Multi-Turn Editing
von: Lee, Dohun, et al.
Veröffentlicht: (2026)
von: Lee, Dohun, et al.
Veröffentlicht: (2026)
ORCA: An Agentic Reasoning Framework for Hallucination and Adversarial Robustness in Vision-Language Models
von: Yu, Chung-En Johnny, et al.
Veröffentlicht: (2025)
von: Yu, Chung-En Johnny, et al.
Veröffentlicht: (2025)
CARE: Contrastive Alignment for ADL Recognition from Event-Triggered Sensor Streams
von: Zhao, Junhao, et al.
Veröffentlicht: (2025)
von: Zhao, Junhao, et al.
Veröffentlicht: (2025)
TA-Prompting: Enhancing Video Large Language Models for Dense Video Captioning via Temporal Anchors
von: Cheng, Wei-Yuan, et al.
Veröffentlicht: (2026)
von: Cheng, Wei-Yuan, et al.
Veröffentlicht: (2026)
NukesFormers: Unpaired Hyperspectral Image Generation with Non-Uniform Domain Alignment
von: Li, Jiaojiao, et al.
Veröffentlicht: (2025)
von: Li, Jiaojiao, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Resilient Contrastive Pre-training under Non-Stationary Drift
von: Yang, Xiaoyu, et al.
Veröffentlicht: (2025) -
Walking the Tightrope: Disentangling Beneficial and Detrimental Drifts in Non-Stationary Custom-Tuning
von: Yang, Xiaoyu, et al.
Veröffentlicht: (2025) -
Autonomous Drift Learning in Data Streams: A Unified Perspective
von: Yang, Xiaoyu, et al.
Veröffentlicht: (2026) -
Towards Robust Endogenous Reasoning: Unifying Drift Adaptation in Non-Stationary Tuning
von: Yang, Xiaoyu, et al.
Veröffentlicht: (2026) -
Adapting Multi-modal Large Language Model to Concept Drift From Pre-training Onwards
von: Yang, Xiaoyu, et al.
Veröffentlicht: (2024)