Saved in:
| Main Authors: | Wang, Changwei, Chen, Shunpeng, Song, Yukun, Xu, Rongtao, Zhang, Zherui, Zhang, Jiguang, Yang, Haoran, Zhang, Yu, Fu, Kexue, Du, Shide, Xu, Zhiwei, Gao, Longxiang, Guo, Li, Xu, Shibiao |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2504.09881 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Region Matters: Efficient and Reliable Region-Aware Visual Place Recognition
by: Chen, Shunpeng, et al.
Published: (2026)
by: Chen, Shunpeng, et al.
Published: (2026)
SAGE: Spatial-visual Adaptive Graph Exploration for Efficient Visual Place Recognition
by: Chen, Shunpeng, et al.
Published: (2025)
by: Chen, Shunpeng, et al.
Published: (2025)
Image Recognition with Online Lightweight Vision Transformer: A Survey
by: Zhang, Zherui, et al.
Published: (2025)
by: Zhang, Zherui, et al.
Published: (2025)
Local Feature Matching Using Deep Learning: A Survey
by: Xu, Shibiao, et al.
Published: (2024)
by: Xu, Shibiao, et al.
Published: (2024)
DeliCIR: Deliberative Test-Time Evolutionary Hierarchical Multi-Agents for Composed Image Retrieval
by: Pei, Xingtian, et al.
Published: (2026)
by: Pei, Xingtian, et al.
Published: (2026)
SkinFormer: Learning Statistical Texture Representation with Transformer for Skin Lesion Segmentation
by: Xu, Rongtao, et al.
Published: (2024)
by: Xu, Rongtao, et al.
Published: (2024)
CurriFlow: Curriculum-Guided Depth Fusion with Optical Flow-Based Temporal Alignment for 3D Semantic Scene Completion
by: Lin, Jinzhou, et al.
Published: (2025)
by: Lin, Jinzhou, et al.
Published: (2025)
CAE-DFKD: Bridging the Transferability Gap in Data-Free Knowledge Distillation
by: Zhang, Zherui, et al.
Published: (2025)
by: Zhang, Zherui, et al.
Published: (2025)
SAMamba: Adaptive State Space Modeling with Hierarchical Vision for Infrared Small Target Detection
by: Xu, Wenhao, et al.
Published: (2025)
by: Xu, Wenhao, et al.
Published: (2025)
FDBPL: Faster Distillation-Based Prompt Learning for Region-Aware Vision-Language Models Adaptation
by: Zhang, Zherui, et al.
Published: (2025)
by: Zhang, Zherui, et al.
Published: (2025)
Generalization Boosted Adapter for Open-Vocabulary Segmentation
by: Xu, Wenhao, et al.
Published: (2024)
by: Xu, Wenhao, et al.
Published: (2024)
3D-MoRe: Unified Modal-Contextual Reasoning for Embodied Question Answering
by: Xu, Rongtao, et al.
Published: (2025)
by: Xu, Rongtao, et al.
Published: (2025)
Spectral Prompt Tuning:Unveiling Unseen Classes for Zero-Shot Semantic Segmentation
by: Xu, Wenhao, et al.
Published: (2023)
by: Xu, Wenhao, et al.
Published: (2023)
HCF-Net: Hierarchical Context Fusion Network for Infrared Small Object Detection
by: Xu, Shibiao, et al.
Published: (2024)
by: Xu, Shibiao, et al.
Published: (2024)
Multimodal Fusion and Vision-Language Models: A Survey for Robot Vision
by: Han, Xiaofeng, et al.
Published: (2025)
by: Han, Xiaofeng, et al.
Published: (2025)
PSTNet: Enhanced Polyp Segmentation with Multi-scale Alignment and Frequency Domain Integration
by: Xu, Wenhao, et al.
Published: (2024)
by: Xu, Wenhao, et al.
Published: (2024)
Segment Anything Model is a Good Teacher for Local Feature Learning
by: Wu, Jingqian, et al.
Published: (2023)
by: Wu, Jingqian, et al.
Published: (2023)
Advances in Embodied Navigation Using Large Language Models: A Survey
by: Lin, Jinzhou, et al.
Published: (2023)
by: Lin, Jinzhou, et al.
Published: (2023)
Vision Also You Need: Navigating Out-of-Distribution Detection with Multimodal Large Language Model
by: Xu, Haoran, et al.
Published: (2026)
by: Xu, Haoran, et al.
Published: (2026)
Key‐point‐guided adaptive convolution and instance normalization for continuous transitive face reenactment of any person
by: Shibiao Xu, et al.
Published: (2024)
by: Shibiao Xu, et al.
Published: (2024)
Complementary Information Guided Occupancy Prediction via Multi-Level Representation Fusion
by: Xu, Rongtao, et al.
Published: (2025)
by: Xu, Rongtao, et al.
Published: (2025)
Adaptive Visual Autoregressive Acceleration via Dual-Linkage Entropy Analysis
by: Zhang, Yu, et al.
Published: (2026)
by: Zhang, Yu, et al.
Published: (2026)
DisPlace: Discriminative Place Projections for Multi-Reference Visual Place Recognition
by: Rajani, Dhyey Manish, et al.
Published: (2026)
by: Rajani, Dhyey Manish, et al.
Published: (2026)
Marine Saliency Segmenter: Object-Focused Conditional Diffusion with Region-Level Semantic Knowledge Distillation
by: Chang, Laibin, et al.
Published: (2025)
by: Chang, Laibin, et al.
Published: (2025)
Explicit Temporal-Semantic Modeling for Dense Video Captioning via Context-Aware Cross-Modal Interaction
by: Jia, Mingda, et al.
Published: (2025)
by: Jia, Mingda, et al.
Published: (2025)
LaplacianFormer:Rethinking Linear Attention with Laplacian Kernel
by: Feng, Zhe, et al.
Published: (2026)
by: Feng, Zhe, et al.
Published: (2026)
Synchronization of Complex Dynamical Networks via Event-Triggered Pinning Impulses
by: Zhang, Kexue
Published: (2024)
by: Zhang, Kexue
Published: (2024)
Visual Para-Thinker: Divide-and-Conquer Reasoning for Visual Comprehension
by: Xu, Haoran, et al.
Published: (2026)
by: Xu, Haoran, et al.
Published: (2026)
OpenViewer: Openness-Aware Multi-View Learning
by: Du, Shide, et al.
Published: (2024)
by: Du, Shide, et al.
Published: (2024)
HMR-1: Hierarchical Massage Robot with Vision-Language-Model for Embodied Healthcare
by: Xu, Rongtao, et al.
Published: (2026)
by: Xu, Rongtao, et al.
Published: (2026)
Graph of Verification: Structured Verification of LLM Reasoning with Directed Acyclic Graphs
by: Fang, Jiwei, et al.
Published: (2025)
by: Fang, Jiwei, et al.
Published: (2025)
OptiCorNet: Optimizing Sequence-Based Context Correlation for Visual Place Recognition
by: Li, Zhenyu, et al.
Published: (2025)
by: Li, Zhenyu, et al.
Published: (2025)
QAGait: Revisit Gait Recognition from a Quality Perspective
by: Wang, Zengbin, et al.
Published: (2024)
by: Wang, Zengbin, et al.
Published: (2024)
RevTogether: Supporting Science Story Revision with Multiple AI Agents
by: Zhang, Yu, et al.
Published: (2025)
by: Zhang, Yu, et al.
Published: (2025)
LargeMvC-Net: Anchor-based Deep Unfolding Network for Large-scale Multi-view Clustering
by: Du, Shide, et al.
Published: (2025)
by: Du, Shide, et al.
Published: (2025)
Register assisted aggregation for Visual Place Recognition
by: Yu, Xuan, et al.
Published: (2024)
by: Yu, Xuan, et al.
Published: (2024)
A Comprehensive Investigation on Speaker Augmentation for Speaker Recognition
by: Zhou, Zhenyu, et al.
Published: (2024)
by: Zhou, Zhenyu, et al.
Published: (2024)
Enhancing Multi-view Open-set Learning via Ambiguity Uncertainty Calibration and View-wise Debiasing
by: Fang, Zihan, et al.
Published: (2025)
by: Fang, Zihan, et al.
Published: (2025)
Hierarchical Augmentation and Distillation for Class Incremental Audio-Visual Video Recognition
by: Zuo, Yukun, et al.
Published: (2024)
by: Zuo, Yukun, et al.
Published: (2024)
An Optimal Discriminator Weighted Imitation Perspective for Reinforcement Learning
by: Xu, Haoran, et al.
Published: (2025)
by: Xu, Haoran, et al.
Published: (2025)
Similar Items
-
Region Matters: Efficient and Reliable Region-Aware Visual Place Recognition
by: Chen, Shunpeng, et al.
Published: (2026) -
SAGE: Spatial-visual Adaptive Graph Exploration for Efficient Visual Place Recognition
by: Chen, Shunpeng, et al.
Published: (2025) -
Image Recognition with Online Lightweight Vision Transformer: A Survey
by: Zhang, Zherui, et al.
Published: (2025) -
Local Feature Matching Using Deep Learning: A Survey
by: Xu, Shibiao, et al.
Published: (2024) -
DeliCIR: Deliberative Test-Time Evolutionary Hierarchical Multi-Agents for Composed Image Retrieval
by: Pei, Xingtian, et al.
Published: (2026)