Guardado en:
| Autores principales: | Liu, Wenxuan, Zhou, Zhuo, Jia, Xuemei, Yang, Siyuan, Huang, Wenxin, Zhong, Xian, Lin, Chia-Wen |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2504.20530 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
OccludeNet: A Causal Journey into Mixed-View Actor-Centric Video Action Recognition under Occlusions
por: Zhou, Guanyu, et al.
Publicado: (2024)
por: Zhou, Guanyu, et al.
Publicado: (2024)
Towards Low-latency Event-based Visual Recognition with Hybrid Step-wise Distillation Spiking Neural Networks
por: Zhong, Xian, et al.
Publicado: (2024)
por: Zhong, Xian, et al.
Publicado: (2024)
See What You Seek: Semantic Contextual Integration for Cloth-Changing Person Re-Identification
por: Han, Xiyu, et al.
Publicado: (2024)
por: Han, Xiyu, et al.
Publicado: (2024)
Enhancing Visual Question Answering with Multimodal LLMs via Chain-of-Question Guided Retrieval-Augmented Generation
por: Xu, Quanxing, et al.
Publicado: (2026)
por: Xu, Quanxing, et al.
Publicado: (2026)
QIRL: Boosting Visual Question Answering via Optimized Question-Image Relation Learning
por: Xu, Quanxing, et al.
Publicado: (2025)
por: Xu, Quanxing, et al.
Publicado: (2025)
Decoupled Contrastive Multi-View Clustering with High-Order Random Walks
por: Lu, Yiding, et al.
Publicado: (2023)
por: Lu, Yiding, et al.
Publicado: (2023)
SpikeDerain: Unveiling Clear Videos from Rainy Sequences Using Color Spike Streams
por: Liang, Hanwen, et al.
Publicado: (2025)
por: Liang, Hanwen, et al.
Publicado: (2025)
DualVLA: Building a Generalizable Embodied Agent via Partial Decoupling of Reasoning and Action
por: Fang, Zhen, et al.
Publicado: (2025)
por: Fang, Zhen, et al.
Publicado: (2025)
Heterogeneous Semantic Transfer for Multi-label Recognition with Partial Labels
por: Chen, Tianshui, et al.
Publicado: (2022)
por: Chen, Tianshui, et al.
Publicado: (2022)
Over-PINNs: Enhancing Physics-Informed Neural Networks via Higher-Order Partial Derivative Overdetermination of PDEs
por: Huo, Wenxuan, et al.
Publicado: (2025)
por: Huo, Wenxuan, et al.
Publicado: (2025)
MV-GMN: State Space Model for Multi-View Action Recognition
por: Lin, Yuhui, et al.
Publicado: (2025)
por: Lin, Yuhui, et al.
Publicado: (2025)
FALCON: Future-Aware Learning with Contextual Object-Centric Pretraining for UAV Action Recognition
por: Xian, Ruiqi, et al.
Publicado: (2024)
por: Xian, Ruiqi, et al.
Publicado: (2024)
UAV-OVO: Out-of-Viewpoint Generalization in UAV Action Recognition
por: Xia, Yu, et al.
Publicado: (2026)
por: Xia, Yu, et al.
Publicado: (2026)
MetaRA: Metamorphic Robustness Assessment for Multimodal Large Language Model-based Visual Question Answering Systems
por: Xu, Quanxing, et al.
Publicado: (2026)
por: Xu, Quanxing, et al.
Publicado: (2026)
PAD: Phase-Amplitude Decoupling Fusion for Multi-Modal Land Cover Classification
por: Zheng, Huiling, et al.
Publicado: (2025)
por: Zheng, Huiling, et al.
Publicado: (2025)
Improving the Transferability of Adversarial Attacks on Face Recognition with Diverse Parameters Augmentation
por: Zhou, Fengfan, et al.
Publicado: (2024)
por: Zhou, Fengfan, et al.
Publicado: (2024)
One-Shot Action Recognition via Multi-Scale Spatial-Temporal Skeleton Matching
por: Yang, Siyuan, et al.
Publicado: (2023)
por: Yang, Siyuan, et al.
Publicado: (2023)
Beyond Viewpoint: Robust 3D Object Recognition under Arbitrary Views through Joint Multi-Part Representation
por: Fan, Linlong, et al.
Publicado: (2024)
por: Fan, Linlong, et al.
Publicado: (2024)
Filter‐Less Polychromatic Recognition in One Photosensitive Oscillator via Dual‐Perception Feature Decoupling
por: Hailong Wang, et al.
Publicado: (2025)
por: Hailong Wang, et al.
Publicado: (2025)
Generalized Jersey Number Recognition Using Multi-task Learning With Orientation-guided Weight Refinement
por: Lin, Yung-Hui, et al.
Publicado: (2024)
por: Lin, Yung-Hui, et al.
Publicado: (2024)
Evidential Deep Partial Multi-View Classification With Discount Fusion
por: Huang, Haojian, et al.
Publicado: (2024)
por: Huang, Haojian, et al.
Publicado: (2024)
LessMimic: Long-Horizon Humanoid Interaction with Unified Distance Field Representations
por: Lin, Yutang, et al.
Publicado: (2026)
por: Lin, Yutang, et al.
Publicado: (2026)
Beyond Entangled Planning: Task-Decoupled Planning for Long-Horizon Agents
por: Li, Yunfan, et al.
Publicado: (2026)
por: Li, Yunfan, et al.
Publicado: (2026)
SkeFi: Cross-Modal Knowledge Transfer for Wireless Skeleton-Based Action Recognition
por: Huang, Shunyu, et al.
Publicado: (2026)
por: Huang, Shunyu, et al.
Publicado: (2026)
Agent-BRACE: Decoupling Beliefs from Actions in Long-Horizon Tasks via Verbalized State Uncertainty
por: Singh, Joykirat, et al.
Publicado: (2026)
por: Singh, Joykirat, et al.
Publicado: (2026)
Beyond Night Visibility: Adaptive Multi-Scale Fusion of Infrared and Visible Images
por: Pei, Shufan, et al.
Publicado: (2024)
por: Pei, Shufan, et al.
Publicado: (2024)
View-aware Cross-modal Distillation for Multi-view Action Recognition
por: Nguyen, Trung Thanh, et al.
Publicado: (2025)
por: Nguyen, Trung Thanh, et al.
Publicado: (2025)
Hypergraph-based Multi-View Action Recognition using Event Cameras
por: Gao, Yue, et al.
Publicado: (2024)
por: Gao, Yue, et al.
Publicado: (2024)
MAVR-Net: Robust Multi-View Learning for MAV Action Recognition with Cross-View Attention
por: Zhang, Nengbo, et al.
Publicado: (2025)
por: Zhang, Nengbo, et al.
Publicado: (2025)
Long-Horizon Streaming Video Generation via Hybrid Attention with Decoupled Distillation
por: Li, Ruibin, et al.
Publicado: (2026)
por: Li, Ruibin, et al.
Publicado: (2026)
Multi-View Active Sensing for Human-Robot Interaction via Hierarchically Connected Tree
por: Ying, Yuanjiong, et al.
Publicado: (2024)
por: Ying, Yuanjiong, et al.
Publicado: (2024)
3D Scene Change Modeling With Consistent Multi-View Aggregation
por: Zhou, Zirui, et al.
Publicado: (2025)
por: Zhou, Zirui, et al.
Publicado: (2025)
Multi-Domain Biometric Recognition using Body Embeddings
por: Nanduri, Anirudh, et al.
Publicado: (2025)
por: Nanduri, Anirudh, et al.
Publicado: (2025)
Task-Augmented Cross-View Imputation Network for Partial Multi-View Incomplete Multi-Label Classification
por: Zhao, Lian, et al.
Publicado: (2024)
por: Zhao, Lian, et al.
Publicado: (2024)
Trace-Focused Diffusion Policy for Multi-Modal Action Disambiguation in Long-Horizon Robotic Manipulation
por: Hu, Yuxuan, et al.
Publicado: (2026)
por: Hu, Yuxuan, et al.
Publicado: (2026)
MVAFormer: RGB-based Multi-View Spatio-Temporal Action Recognition with Transformer
por: Yamane, Taiga, et al.
Publicado: (2025)
por: Yamane, Taiga, et al.
Publicado: (2025)
WeatherCycle: Unpaired Multi-Weather Restoration via Color Space Decoupled Cycle Learning
por: Fang, Wenxuan, et al.
Publicado: (2025)
por: Fang, Wenxuan, et al.
Publicado: (2025)
Multi-Level Decoupled Relational Distillation for Heterogeneous Architectures
por: Yang, Yaoxin, et al.
Publicado: (2025)
por: Yang, Yaoxin, et al.
Publicado: (2025)
NavAgent: Multi-scale Urban Street View Fusion For UAV Embodied Vision-and-Language Navigation
por: Liu, Youzhi, et al.
Publicado: (2024)
por: Liu, Youzhi, et al.
Publicado: (2024)
Decoupled Action Expert: Confining Task Knowledge to the Conditioning Pathway
por: Zhou, Jian, et al.
Publicado: (2025)
por: Zhou, Jian, et al.
Publicado: (2025)
Ejemplares similares
-
OccludeNet: A Causal Journey into Mixed-View Actor-Centric Video Action Recognition under Occlusions
por: Zhou, Guanyu, et al.
Publicado: (2024) -
Towards Low-latency Event-based Visual Recognition with Hybrid Step-wise Distillation Spiking Neural Networks
por: Zhong, Xian, et al.
Publicado: (2024) -
See What You Seek: Semantic Contextual Integration for Cloth-Changing Person Re-Identification
por: Han, Xiyu, et al.
Publicado: (2024) -
Enhancing Visual Question Answering with Multimodal LLMs via Chain-of-Question Guided Retrieval-Augmented Generation
por: Xu, Quanxing, et al.
Publicado: (2026) -
QIRL: Boosting Visual Question Answering via Optimized Question-Image Relation Learning
por: Xu, Quanxing, et al.
Publicado: (2025)