Salvato in:
| Autori principali: | Wan, Cong, Guo, Zeyu, Li, Jiangyang, Dong, SongLin, Bai, Yifan, Peng, Lin, Ma, Zhiheng, Gong, Yihong |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2603.00461 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Trajectory-Diversity-Driven Robust Vision-and-Language Navigation
di: Li, Jiangyang, et al.
Pubblicazione: (2026)
di: Li, Jiangyang, et al.
Pubblicazione: (2026)
Continuous Expert Assembly: Instance-Conditioned Low-Rank Residuals for All-in-One Image Restoration
di: He, Haisen, et al.
Pubblicazione: (2026)
di: He, Haisen, et al.
Pubblicazione: (2026)
VDC-Agent: When Video Detailed Captioners Evolve Themselves via Agentic Self-Reflection
di: Wang, Qiang, et al.
Pubblicazione: (2025)
di: Wang, Qiang, et al.
Pubblicazione: (2025)
DualCP: Rehearsal-Free Domain-Incremental Learning via Dual-Level Concept Prototype
di: Wang, Qiang, et al.
Pubblicazione: (2025)
di: Wang, Qiang, et al.
Pubblicazione: (2025)
P2L-CA: An Effective Parameter Tuning Framework for Rehearsal-Free Multi-Label Class-Incremental Learning
di: Dong, Songlin, et al.
Pubblicazione: (2026)
di: Dong, Songlin, et al.
Pubblicazione: (2026)
Unleashing the Potential of All Test Samples: Mean-Shift Guided Test-Time Adaptation
di: Han, Jizhou, et al.
Pubblicazione: (2025)
di: Han, Jizhou, et al.
Pubblicazione: (2025)
Learning Like Humans: Analogical Concept Learning for Generalized Category Discovery
di: Han, Jizhou, et al.
Pubblicazione: (2026)
di: Han, Jizhou, et al.
Pubblicazione: (2026)
Retrieve-then-Steer: Online Success Memory for Test-Time Adaptation of Generative VLAs
di: Zhao, Jianchao, et al.
Pubblicazione: (2026)
di: Zhao, Jianchao, et al.
Pubblicazione: (2026)
ProSR: Process-Shaped Spatial Reasoning for Reliable Chain-of-Thought in VLMs
di: Li, Jiangyang, et al.
Pubblicazione: (2026)
di: Li, Jiangyang, et al.
Pubblicazione: (2026)
GOAL: Geometrically Optimal Alignment for Continual Generalized Category Discovery
di: Han, Jizhou, et al.
Pubblicazione: (2026)
di: Han, Jizhou, et al.
Pubblicazione: (2026)
Consistent Supervised-Unsupervised Alignment for Generalized Category Discovery
di: Han, Jizhou, et al.
Pubblicazione: (2025)
di: Han, Jizhou, et al.
Pubblicazione: (2025)
Beyond World-Frame Action Heads: Motion-Centric Action Frames for Vision-Language-Action Models
di: Yang, Huoren, et al.
Pubblicazione: (2026)
di: Yang, Huoren, et al.
Pubblicazione: (2026)
Beyond CLIP Generalization: Against Forward&Backward Forgetting Adapter for Continual Learning of Vision-Language Models
di: Dong, Songlin, et al.
Pubblicazione: (2025)
di: Dong, Songlin, et al.
Pubblicazione: (2025)
Prompt-Agnostic Adversarial Perturbation for Customized Diffusion Models
di: Wan, Cong, et al.
Pubblicazione: (2024)
di: Wan, Cong, et al.
Pubblicazione: (2024)
ReMoMask: Retrieval-Augmented Masked Motion Generation
di: Li, Zhengdao, et al.
Pubblicazione: (2025)
di: Li, Zhengdao, et al.
Pubblicazione: (2025)
Few-shot Online Anomaly Detection and Segmentation
di: Wei, Shenxing, et al.
Pubblicazione: (2024)
di: Wei, Shenxing, et al.
Pubblicazione: (2024)
Shared & Domain Self-Adaptive Experts with Frequency-Aware Discrimination for Continual Test-Time Adaptation
di: Zhao, JianChao, et al.
Pubblicazione: (2025)
di: Zhao, JianChao, et al.
Pubblicazione: (2025)
ARTrackV2: Prompting Autoregressive Tracker Where to Look and How to Describe
di: Bai, Yifan, et al.
Pubblicazione: (2023)
di: Bai, Yifan, et al.
Pubblicazione: (2023)
MoRL: Reinforced Reasoning for Unified Motion Understanding and Generation
di: Wang, Hongpeng, et al.
Pubblicazione: (2026)
di: Wang, Hongpeng, et al.
Pubblicazione: (2026)
Curriculum Dataset Distillation
di: Ma, Zhiheng, et al.
Pubblicazione: (2024)
di: Ma, Zhiheng, et al.
Pubblicazione: (2024)
Grid: Omni Visual Generation
di: Wan, Cong, et al.
Pubblicazione: (2024)
di: Wan, Cong, et al.
Pubblicazione: (2024)
CoGenAV: Versatile Audio-Visual Representation Learning via Contrastive-Generative Synchronization
di: Bai, Detao, et al.
Pubblicazione: (2025)
di: Bai, Detao, et al.
Pubblicazione: (2025)
MoVieS: Motion-Aware 4D Dynamic View Synthesis in One Second
di: Lin, Chenguo, et al.
Pubblicazione: (2025)
di: Lin, Chenguo, et al.
Pubblicazione: (2025)
MSP-ReID: Hairstyle-Robust Cloth-Changing Person Re-Identification
di: He, Xiangyang, et al.
Pubblicazione: (2026)
di: He, Xiangyang, et al.
Pubblicazione: (2026)
Projecting Points to Axes: Oriented Object Detection via Point-Axis Representation
di: Zhao, Zeyang, et al.
Pubblicazione: (2024)
di: Zhao, Zeyang, et al.
Pubblicazione: (2024)
ConMo: Controllable Motion Disentanglement and Recomposition for Zero-Shot Motion Transfer
di: Gao, Jiayi, et al.
Pubblicazione: (2025)
di: Gao, Jiayi, et al.
Pubblicazione: (2025)
IOTA: Corrective Knowledge-Guided Prompt Learning via Black-White Box Framework
di: Wang, Shaokun, et al.
Pubblicazione: (2026)
di: Wang, Shaokun, et al.
Pubblicazione: (2026)
MoTiC: Momentum Tightness and Contrast for Few-Shot Class-Incremental Learning
di: He, Zeyu, et al.
Pubblicazione: (2025)
di: He, Zeyu, et al.
Pubblicazione: (2025)
Prototypical Contrastive Learning-based CLIP Fine-tuning for Object Re-identification
di: Li, Jiachen, et al.
Pubblicazione: (2023)
di: Li, Jiachen, et al.
Pubblicazione: (2023)
H-MoRe: Learning Human-centric Motion Representation for Action Analysis
di: Huang, Zhanbo, et al.
Pubblicazione: (2025)
di: Huang, Zhanbo, et al.
Pubblicazione: (2025)
SemGeoMo: Dynamic Contextual Human Motion Generation with Semantic and Geometric Guidance
di: Cong, Peishan, et al.
Pubblicazione: (2025)
di: Cong, Peishan, et al.
Pubblicazione: (2025)
MoReFun: Past-Movement Guided Motion Representation Learning for Future Motion Prediction and Understanding
di: Shi, Junyu, et al.
Pubblicazione: (2024)
di: Shi, Junyu, et al.
Pubblicazione: (2024)
MoReGen: Multi-Agent Motion-Reasoning Engine for Code-based Text-to-Video Synthesis
di: Bai, Xiangyu, et al.
Pubblicazione: (2025)
di: Bai, Xiangyu, et al.
Pubblicazione: (2025)
Neural Implicit Action Fields: From Discrete Waypoints to Continuous Functions for Vision-Language-Action Models
di: Liu, Haoyun, et al.
Pubblicazione: (2026)
di: Liu, Haoyun, et al.
Pubblicazione: (2026)
SafeMo: Linguistically Grounded Unlearning for Trustworthy Text-to-Motion Generation
di: Wang, Yiling, et al.
Pubblicazione: (2026)
di: Wang, Yiling, et al.
Pubblicazione: (2026)
Fine-Grained Spatiotemporal Motion Alignment for Contrastive Video Representation Learning
di: Zhu, Minghao, et al.
Pubblicazione: (2023)
di: Zhu, Minghao, et al.
Pubblicazione: (2023)
SiCL: Silhouette-Driven Contrastive Learning for Unsupervised Person Re-Identification with Clothes Change
di: Li, Mingkun, et al.
Pubblicazione: (2023)
di: Li, Mingkun, et al.
Pubblicazione: (2023)
Re$^2$MoGen: Open-Vocabulary Motion Generation via LLM Reasoning and Physics-Aware Refinement
di: Zheng, Jiakun, et al.
Pubblicazione: (2026)
di: Zheng, Jiakun, et al.
Pubblicazione: (2026)
Beyond Prompt Learning: Continual Adapter for Efficient Rehearsal-Free Continual Learning
di: Gao, Xinyuan, et al.
Pubblicazione: (2024)
di: Gao, Xinyuan, et al.
Pubblicazione: (2024)
Gramformer: Learning Crowd Counting via Graph-Modulated Transformer
di: Lin, Hui, et al.
Pubblicazione: (2024)
di: Lin, Hui, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Trajectory-Diversity-Driven Robust Vision-and-Language Navigation
di: Li, Jiangyang, et al.
Pubblicazione: (2026) -
Continuous Expert Assembly: Instance-Conditioned Low-Rank Residuals for All-in-One Image Restoration
di: He, Haisen, et al.
Pubblicazione: (2026) -
VDC-Agent: When Video Detailed Captioners Evolve Themselves via Agentic Self-Reflection
di: Wang, Qiang, et al.
Pubblicazione: (2025) -
DualCP: Rehearsal-Free Domain-Incremental Learning via Dual-Level Concept Prototype
di: Wang, Qiang, et al.
Pubblicazione: (2025) -
P2L-CA: An Effective Parameter Tuning Framework for Rehearsal-Free Multi-Label Class-Incremental Learning
di: Dong, Songlin, et al.
Pubblicazione: (2026)