ViewSAM: Learning View-aware Cross-modal Semantics for Weakly Supervised Cross-view Referring Multi-Object Tracking
Fuente:
arXiv
Salvato in:
| Autori principali: | Ge, Jiawei, Zhang, Xintian, Cao, Jiuxin, Liu, Bo, Deuser, Fabian, Liu, Chang, Wenkang, Gong, Li, Siyou, Shao, Juexi, Wu, Wenqing, Feng, Chen, Patras, Ioannis |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Debate-Enhanced Pseudo Labeling and Frequency-Aware Progressive Debiasing for Weakly-Supervised Camouflaged Object Detection with Scribble Annotations
di: Ge, Jiawei, et al.
Pubblicazione: (2025)
di: Ge, Jiawei, et al.
Pubblicazione: (2025)
Cross-View Referring Multi-Object Tracking
di: Chen, Sijia, et al.
Pubblicazione: (2024)
di: Chen, Sijia, et al.
Pubblicazione: (2024)
View-aware Cross-modal Distillation for Multi-view Action Recognition
di: Nguyen, Trung Thanh, et al.
Pubblicazione: (2025)
di: Nguyen, Trung Thanh, et al.
Pubblicazione: (2025)
ViewSparsifier: Killing Redundancy in Multi-View Plant Phenotyping
di: Kampa, Robin-Nico, et al.
Pubblicazione: (2025)
di: Kampa, Robin-Nico, et al.
Pubblicazione: (2025)
Cross-View Geolocalization and Disaster Mapping with Street-View and VHR Satellite Imagery: A Case Study of Hurricane IAN
di: Li, Hao, et al.
Pubblicazione: (2024)
di: Li, Hao, et al.
Pubblicazione: (2024)
GlazyBench: A Benchmark for Ceramic Glaze Property Prediction and Image Generation
di: Zhai, Ziyu, et al.
Pubblicazione: (2026)
di: Zhai, Ziyu, et al.
Pubblicazione: (2026)
V$^{2}$-SAM: Marrying SAM2 with Multi-Prompt Experts for Cross-View Object Correspondence
di: Pan, Jiancheng, et al.
Pubblicazione: (2025)
di: Pan, Jiancheng, et al.
Pubblicazione: (2025)
VAGeo: View-specific Attention for Cross-View Object Geo-Localization
di: Li, Zhongyang, et al.
Pubblicazione: (2025)
di: Li, Zhongyang, et al.
Pubblicazione: (2025)
Making Dialogue Grounding Data Rich: A Three-Tier Data Synthesis Framework for Generalized Referring Expression Comprehension
di: Shao, Juexi, et al.
Pubblicazione: (2025)
di: Shao, Juexi, et al.
Pubblicazione: (2025)
NovBench: Evaluating Large Language Models on Academic Paper Novelty Assessment
di: Wu, Wenqing, et al.
Pubblicazione: (2026)
di: Wu, Wenqing, et al.
Pubblicazione: (2026)
Beyond Visual Cues: Synchronously Exploring Target-Centric Semantics for Vision-Language Tracking
di: Ge, Jiawei, et al.
Pubblicazione: (2023)
di: Ge, Jiawei, et al.
Pubblicazione: (2023)
Self-Supervised Facial Representation Learning with Facial Region Awareness
di: Gao, Zheng, et al.
Pubblicazione: (2024)
di: Gao, Zheng, et al.
Pubblicazione: (2024)
GeoDistill: Geometry-Guided Self-Distillation for Weakly Supervised Cross-View Localization
di: Tong, Shaowen, et al.
Pubblicazione: (2025)
di: Tong, Shaowen, et al.
Pubblicazione: (2025)
Recurrent Cross-View Object Geo-Localization
di: Zhang, Xiaohan, et al.
Pubblicazione: (2025)
di: Zhang, Xiaohan, et al.
Pubblicazione: (2025)
RefSAM3D: Adapting SAM with Cross-modal Reference for 3D Medical Image Segmentation
di: Gao, Xiang, et al.
Pubblicazione: (2024)
di: Gao, Xiang, et al.
Pubblicazione: (2024)
SAM-Driven Weakly Supervised Nodule Segmentation with Uncertainty-Aware Cross Teaching
di: Zhao, Xingyue, et al.
Pubblicazione: (2024)
di: Zhao, Xingyue, et al.
Pubblicazione: (2024)
CrossView-GS: Cross-view Gaussian Splatting For Large-scale Scene Reconstruction
di: Zhang, Chenhao, et al.
Pubblicazione: (2025)
di: Zhang, Chenhao, et al.
Pubblicazione: (2025)
Towards Generative Location Awareness for Disaster Response: A Probabilistic Cross-view Geolocalization Approach
di: Li, Hao, et al.
Pubblicazione: (2025)
di: Li, Hao, et al.
Pubblicazione: (2025)
Weakly Supervised Camouflaged Object Detection Based on the SAM Model and Mask Guidance
di: Li, Xia, et al.
Pubblicazione: (2026)
di: Li, Xia, et al.
Pubblicazione: (2026)
DOMR: Establishing Cross-View Segmentation via Dense Object Matching
di: Liao, Jitong, et al.
Pubblicazione: (2025)
di: Liao, Jitong, et al.
Pubblicazione: (2025)
VidCtx: Context-aware Video Question Answering with Image Models
di: Goulas, Andreas, et al.
Pubblicazione: (2024)
di: Goulas, Andreas, et al.
Pubblicazione: (2024)
Effects of Repeated Viewing on Attention to Bilingual Subtitles and L2 Incidental Vocabulary Learning: An Eye‐Tracking Study
di: Wenqing Zhong, et al.
Pubblicazione: (2026)
di: Wenqing Zhong, et al.
Pubblicazione: (2026)
Self-Supervised Bird's Eye View Motion Prediction with Cross-Modality Signals
di: Fang, Shaoheng, et al.
Pubblicazione: (2024)
di: Fang, Shaoheng, et al.
Pubblicazione: (2024)
CrossView Suite: Harnessing Cross-view Spatial Intelligence of MLLMs with Dataset, Model and Benchmark
di: Wang, Wei, et al.
Pubblicazione: (2026)
di: Wang, Wei, et al.
Pubblicazione: (2026)
Visual Sync: Multi-Camera Synchronization via Cross-View Object Motion
di: Liu, Shaowei, et al.
Pubblicazione: (2025)
di: Liu, Shaowei, et al.
Pubblicazione: (2025)
CrossViewDiff: A Cross-View Diffusion Model for Satellite-to-Street View Synthesis
di: Li, Weijia, et al.
Pubblicazione: (2024)
di: Li, Weijia, et al.
Pubblicazione: (2024)
Multi-Grained Cross-modal Alignment for Learning Open-vocabulary Semantic Segmentation from Text Supervision
di: Liu, Yajie, et al.
Pubblicazione: (2024)
di: Liu, Yajie, et al.
Pubblicazione: (2024)
IMDPrompter: Adapting SAM to Image Manipulation Detection by Cross-View Automated Prompt Learning
di: Zhang, Quan, et al.
Pubblicazione: (2025)
di: Zhang, Quan, et al.
Pubblicazione: (2025)
SAM-COD: SAM-guided Unified Framework for Weakly-Supervised Camouflaged Object Detection
di: Chen, Huafeng, et al.
Pubblicazione: (2024)
di: Chen, Huafeng, et al.
Pubblicazione: (2024)
MOCHA: Multi-modal Objects-aware Cross-arcHitecture Alignment
di: Camuffo, Elena, et al.
Pubblicazione: (2025)
di: Camuffo, Elena, et al.
Pubblicazione: (2025)
ICG-MVSNet: Learning Intra-view and Cross-view Relationships for Guidance in Multi-View Stereo
di: Hu, Yuxi, et al.
Pubblicazione: (2025)
di: Hu, Yuxi, et al.
Pubblicazione: (2025)
Cross-modal Offset-guided Dynamic Alignment and Fusion for Weakly Aligned UAV Object Detection
di: Zongzhen, Liu, et al.
Pubblicazione: (2025)
di: Zongzhen, Liu, et al.
Pubblicazione: (2025)
Active View Selector: Fast and Accurate Active View Selection with Cross Reference Image Quality Assessment
di: Wang, Zirui, et al.
Pubblicazione: (2025)
di: Wang, Zirui, et al.
Pubblicazione: (2025)
Geometry-guided Cross-view Diffusion for One-to-many Cross-view Image Synthesis
di: Lin, Tao Jun, et al.
Pubblicazione: (2024)
di: Lin, Tao Jun, et al.
Pubblicazione: (2024)
ViewBridge:Revisiting Cross-View Localization from Image Matching
di: Xia, Panwang, et al.
Pubblicazione: (2025)
di: Xia, Panwang, et al.
Pubblicazione: (2025)
MVAT: Multi-View Aware Teacher for Weakly Supervised 3D Object Detection
di: Lahlali, Saad, et al.
Pubblicazione: (2025)
di: Lahlali, Saad, et al.
Pubblicazione: (2025)
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos
di: Majumder, Sagnik, et al.
Pubblicazione: (2024)
di: Majumder, Sagnik, et al.
Pubblicazione: (2024)
Cross-View Open-Vocabulary Object Detection in Aerial Imagery
di: Kini, Jyoti, et al.
Pubblicazione: (2025)
di: Kini, Jyoti, et al.
Pubblicazione: (2025)
Multi-View Industrial Anomaly Detection with Epipolar Constrained Cross-View Fusion
di: Liu, Yifan, et al.
Pubblicazione: (2025)
di: Liu, Yifan, et al.
Pubblicazione: (2025)
Pro2SAM: Mask Prompt to SAM with Grid Points for Weakly Supervised Object Localization
di: Yang, Xi, et al.
Pubblicazione: (2025)
di: Yang, Xi, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Debate-Enhanced Pseudo Labeling and Frequency-Aware Progressive Debiasing for Weakly-Supervised Camouflaged Object Detection with Scribble Annotations
di: Ge, Jiawei, et al.
Pubblicazione: (2025) -
Cross-View Referring Multi-Object Tracking
di: Chen, Sijia, et al.
Pubblicazione: (2024) -
View-aware Cross-modal Distillation for Multi-view Action Recognition
di: Nguyen, Trung Thanh, et al.
Pubblicazione: (2025) -
ViewSparsifier: Killing Redundancy in Multi-View Plant Phenotyping
di: Kampa, Robin-Nico, et al.
Pubblicazione: (2025) -
Cross-View Geolocalization and Disaster Mapping with Street-View and VHR Satellite Imagery: A Case Study of Hurricane IAN
di: Li, Hao, et al.
Pubblicazione: (2024)