DualCross: Cross-Modality Cross-Domain Adaptation for Monocular BEV Perception
Fuente:
arXiv
Saved in:
| Main Authors: | Man, Yunze, Gui, Liang-Yan, Wang, Yu-Xiong |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
REALM: An RGB and Event Aligned Latent Manifold for Cross-Modal Perception
by: Polizzi, Vincenzo, et al.
Published: (2026)
by: Polizzi, Vincenzo, et al.
Published: (2026)
Residual Cross-Modal Fusion Networks for Audio-Visual Navigation
by: Wang, Yi, et al.
Published: (2026)
by: Wang, Yi, et al.
Published: (2026)
CoCMT: Communication-Efficient Cross-Modal Transformer for Collaborative Perception
by: Wang, Rujia, et al.
Published: (2025)
by: Wang, Rujia, et al.
Published: (2025)
Lexicon3D: Probing Visual Foundation Models for Complex 3D Scene Understanding
by: Man, Yunze, et al.
Published: (2024)
by: Man, Yunze, et al.
Published: (2024)
Single-Frame Point-Pixel Registration via Supervised Cross-Modal Feature Matching
by: Han, Yu, et al.
Published: (2025)
by: Han, Yu, et al.
Published: (2025)
UrbanCross: Enhancing Satellite Image-Text Retrieval with Cross-Domain Adaptation
by: Zhong, Siru, et al.
Published: (2024)
by: Zhong, Siru, et al.
Published: (2024)
Situational Awareness Matters in 3D Vision Language Reasoning
by: Man, Yunze, et al.
Published: (2024)
by: Man, Yunze, et al.
Published: (2024)
Cross-Modal Instructions for Robot Motion Generation
by: Barron, William, et al.
Published: (2025)
by: Barron, William, et al.
Published: (2025)
Online,Target-Free LiDAR-Camera Extrinsic Calibration via Cross-Modal Mask Matching
by: Huang, Zhiwei, et al.
Published: (2024)
by: Huang, Zhiwei, et al.
Published: (2024)
OnlineHOI: Towards Online Human-Object Interaction Generation and Perception
by: Ji, Yihong, et al.
Published: (2025)
by: Ji, Yihong, et al.
Published: (2025)
Unsupervised Domain Adaptation via Similarity-based Prototypes for Cross-Modality Segmentation
by: Ye, Ziyu, et al.
Published: (2025)
by: Ye, Ziyu, et al.
Published: (2025)
RAAP: Retrieval-Augmented Affordance Prediction with Cross-Image Action Alignment
by: Zhuang, Qiyuan, et al.
Published: (2026)
by: Zhuang, Qiyuan, et al.
Published: (2026)
VLA-Pro: Cross-Task Procedural Memory Transfer for Vision-Language-Action Models
by: Si, Shengyu, et al.
Published: (2026)
by: Si, Shengyu, et al.
Published: (2026)
TempBEV: Improving Learned BEV Encoders with Combined Image and BEV Space Temporal Aggregation
by: Monninger, Thomas, et al.
Published: (2024)
by: Monninger, Thomas, et al.
Published: (2024)
CleverDistiller: Simple and Spatially Consistent Cross-modal Distillation
by: Govindarajan, Hariprasath, et al.
Published: (2025)
by: Govindarajan, Hariprasath, et al.
Published: (2025)
X-Diffusion: Training Diffusion Policies on Cross-Embodiment Human Demonstrations
by: Pace, Maximus A., et al.
Published: (2025)
by: Pace, Maximus A., et al.
Published: (2025)
LetsMap: Unsupervised Representation Learning for Semantic BEV Mapping
by: Gosala, Nikhil, et al.
Published: (2024)
by: Gosala, Nikhil, et al.
Published: (2024)
X-VLA: Soft-Prompted Transformer as Scalable Cross-Embodiment Vision-Language-Action Model
by: Zheng, Jinliang, et al.
Published: (2025)
by: Zheng, Jinliang, et al.
Published: (2025)
Cross from Left to Right Brain: Adaptive Text Dreamer for Vision-and-Language Navigation
by: Zhang, Pingrui, et al.
Published: (2025)
by: Zhang, Pingrui, et al.
Published: (2025)
DenseMTL: Cross-task Attention Mechanism for Dense Multi-task Learning
by: Lopes, Ivan, et al.
Published: (2022)
by: Lopes, Ivan, et al.
Published: (2022)
Accelerating Transformer-Based Monocular SLAM via Geometric Utility Scoring
by: Xiong, Xinmiao, et al.
Published: (2026)
by: Xiong, Xinmiao, et al.
Published: (2026)
PaintScene4D: Consistent 4D Scene Generation from Text Prompts
by: Gupta, Vinayak, et al.
Published: (2024)
by: Gupta, Vinayak, et al.
Published: (2024)
See, Act, Adapt: Active Perception for Unsupervised Cross-Domain Visual Adaptation via Personalized VLM-Guided Agent
by: Tang, Tianci, et al.
Published: (2026)
by: Tang, Tianci, et al.
Published: (2026)
Weather-Robust Cross-View Geo-Localization via Prototype-Based Semantic Part Discovery
by: Tran, Chi-Nguyen, et al.
Published: (2026)
by: Tran, Chi-Nguyen, et al.
Published: (2026)
ROCKET-2: Steering Visuomotor Policy via Cross-View Goal Alignment
by: Cai, Shaofei, et al.
Published: (2025)
by: Cai, Shaofei, et al.
Published: (2025)
Quantifying Context Bias in Domain Adaptation for Object Detection
by: Son, Hojun, et al.
Published: (2024)
by: Son, Hojun, et al.
Published: (2024)
Training-Free Dual Hyperbolic Adapters for Better Cross-Modal Reasoning
by: Zhang, Yi, et al.
Published: (2025)
by: Zhang, Yi, et al.
Published: (2025)
Turning Adaptation into Assets: Cross-Domain Bridging for Online Vision-Language Navigation
by: Hu, Zixuan, et al.
Published: (2026)
by: Hu, Zixuan, et al.
Published: (2026)
Beyond Cross-Modal Alignment: Measuring and Leveraging Modality Gap in Vision-Language Models
by: Yan, Hanqi, et al.
Published: (2025)
by: Yan, Hanqi, et al.
Published: (2025)
Towards Dense and Accurate Radar Perception Via Efficient Cross-Modal Diffusion Model
by: Zhang, Ruibin, et al.
Published: (2024)
by: Zhang, Ruibin, et al.
Published: (2024)
Monocular Visual Place Recognition in LiDAR Maps via Cross-Modal State Space Model and Multi-View Matching
by: Yao, Gongxin, et al.
Published: (2024)
by: Yao, Gongxin, et al.
Published: (2024)
Boosting Cross-spectral Unsupervised Domain Adaptation for Thermal Semantic Segmentation
by: Kwon, Seokjun, et al.
Published: (2025)
by: Kwon, Seokjun, et al.
Published: (2025)
Flat'n'Fold: A Diverse Multi-Modal Dataset for Garment Perception and Manipulation
by: Zhuang, Lipeng, et al.
Published: (2024)
by: Zhuang, Lipeng, et al.
Published: (2024)
BEVal: A Cross-dataset Evaluation Study of BEV Segmentation Models for Autonomous Driving
by: Diaz-Zapata, Manuel Alejandro, et al.
Published: (2024)
by: Diaz-Zapata, Manuel Alejandro, et al.
Published: (2024)
Monocular Reconstruction of Neural Tactile Fields
by: Mantripragada, Pavan, et al.
Published: (2026)
by: Mantripragada, Pavan, et al.
Published: (2026)
Pseudo Label Refinery for Unsupervised Domain Adaptation on Cross-dataset 3D Object Detection
by: Zhang, Zhanwei, et al.
Published: (2024)
by: Zhang, Zhanwei, et al.
Published: (2024)
MCRL4OR: Multimodal Contrastive Representation Learning for Off-Road Environmental Perception
by: Yang, Yi, et al.
Published: (2025)
by: Yang, Yi, et al.
Published: (2025)
Cross-Source Supervision for Bone Infection Segmentation in Dual-Modality PET-CT
by: Yang, Zonglin, et al.
Published: (2026)
by: Yang, Zonglin, et al.
Published: (2026)
Enhancing Multimodal Unified Representations for Cross Modal Generalization
by: Huang, Hai, et al.
Published: (2024)
by: Huang, Hai, et al.
Published: (2024)
Visual Sync: Multi-Camera Synchronization via Cross-View Object Motion
by: Liu, Shaowei, et al.
Published: (2025)
by: Liu, Shaowei, et al.
Published: (2025)
Similar Items
-
REALM: An RGB and Event Aligned Latent Manifold for Cross-Modal Perception
by: Polizzi, Vincenzo, et al.
Published: (2026) -
Residual Cross-Modal Fusion Networks for Audio-Visual Navigation
by: Wang, Yi, et al.
Published: (2026) -
CoCMT: Communication-Efficient Cross-Modal Transformer for Collaborative Perception
by: Wang, Rujia, et al.
Published: (2025) -
Lexicon3D: Probing Visual Foundation Models for Complex 3D Scene Understanding
by: Man, Yunze, et al.
Published: (2024) -
Single-Frame Point-Pixel Registration via Supervised Cross-Modal Feature Matching
by: Han, Yu, et al.
Published: (2025)