MOD-UV: Learning Mobile Object Detectors from Unlabeled Videos
Fuente:
arXiv
Saved in:
| Main Authors: | Sun, Yihong, Hariharan, Bharath |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Tracking and Understanding Object Transformations
by: Sun, Yihong, et al.
Published: (2025)
by: Sun, Yihong, et al.
Published: (2025)
Live Interactive Training for Video Segmentation
by: Yang, Xinyu, et al.
Published: (2026)
by: Yang, Xinyu, et al.
Published: (2026)
Video Creation by Demonstration
by: Sun, Yihong, et al.
Published: (2024)
by: Sun, Yihong, et al.
Published: (2024)
Learning Feature Descriptors using Camera Pose Supervision
by: Wang, Qianqian, et al.
Published: (2020)
by: Wang, Qianqian, et al.
Published: (2020)
MovieRecapsQA: A Multimodal Open-Ended Video Question-Answering Benchmark
by: Shaar, Shaden, et al.
Published: (2026)
by: Shaar, Shaden, et al.
Published: (2026)
Composing People Together: Iterative Pose-Image Generation for Multi-Person Interaction Scenes
by: Peng, Wenxuan, et al.
Published: (2026)
by: Peng, Wenxuan, et al.
Published: (2026)
VideoWorld: Exploring Knowledge Learning from Unlabeled Videos
by: Ren, Zhongwei, et al.
Published: (2025)
by: Ren, Zhongwei, et al.
Published: (2025)
Learning 3D Perception from Others' Predictions
by: Yoo, Jinsu, et al.
Published: (2024)
by: Yoo, Jinsu, et al.
Published: (2024)
ObjectCarver: Semi-automatic segmentation, reconstruction and separation of 3D objects
by: Hassena, Gemmechu, et al.
Published: (2024)
by: Hassena, Gemmechu, et al.
Published: (2024)
MOD-CL: Multi-label Object Detection with Constrained Loss
by: Moriyama, Sota, et al.
Published: (2024)
by: Moriyama, Sota, et al.
Published: (2024)
Ponymation: Learning Articulated 3D Animal Motions from Unlabeled Online Videos
by: Sun, Keqiang, et al.
Published: (2023)
by: Sun, Keqiang, et al.
Published: (2023)
Pre-Training LiDAR-Based 3D Object Detectors Through Colorization
by: Pan, Tai-Yu, et al.
Published: (2023)
by: Pan, Tai-Yu, et al.
Published: (2023)
Lost and Found: Overcoming Detector Failures in Online Multi-Object Tracking
by: Vaquero, Lorenzo, et al.
Published: (2024)
by: Vaquero, Lorenzo, et al.
Published: (2024)
C3Po: Cross-View Cross-Modality Correspondence by Pointmap Prediction
by: Huang, Kuan Wei, et al.
Published: (2025)
by: Huang, Kuan Wei, et al.
Published: (2025)
KFC-W: Generating 3D-Consistent Videos from Unposed Internet Photos
by: Chou, Gene, et al.
Published: (2024)
by: Chou, Gene, et al.
Published: (2024)
Better Monocular 3D Detectors with LiDAR from the Past
by: You, Yurong, et al.
Published: (2024)
by: You, Yurong, et al.
Published: (2024)
Sports Re-ID: Improving Re-Identification Of Players In Broadcast Videos Of Team Sports
by: Comandur, Bharath
Published: (2022)
by: Comandur, Bharath
Published: (2022)
Switch-a-View: View Selection Learned from Unlabeled In-the-wild Videos
by: Majumder, Sagnik, et al.
Published: (2024)
by: Majumder, Sagnik, et al.
Published: (2024)
Color Bind: Exploring Color Perception in Text-to-Image Models
by: Shomer-Chai, Shay, et al.
Published: (2025)
by: Shomer-Chai, Shay, et al.
Published: (2025)
FlashDepth: Real-time Streaming Video Depth Estimation at 2K Resolution
by: Chou, Gene, et al.
Published: (2025)
by: Chou, Gene, et al.
Published: (2025)
xMOD: Cross-Modal Distillation for 2D/3D Multi-Object Discovery from 2D motion
by: Lahlali, Saad, et al.
Published: (2025)
by: Lahlali, Saad, et al.
Published: (2025)
Vid-Morp: Video Moment Retrieval Pretraining from Unlabeled Videos in the Wild
by: Bao, Peijun, et al.
Published: (2024)
by: Bao, Peijun, et al.
Published: (2024)
On the Feasibility and Opportunity of Autoregressive 3D Object Detection
by: Huang, Zanming, et al.
Published: (2026)
by: Huang, Zanming, et al.
Published: (2026)
Scale-Aware Recognition in Satellite Images under Resource Constraints
by: Revankar, Shreelekha, et al.
Published: (2024)
by: Revankar, Shreelekha, et al.
Published: (2024)
MONITRS: Multimodal Observations of Natural Incidents Through Remote Sensing
by: Revankar, Shreelekha, et al.
Published: (2025)
by: Revankar, Shreelekha, et al.
Published: (2025)
SOOD++: Leveraging Unlabeled Data to Boost Oriented Object Detection
by: Liang, Dingkang, et al.
Published: (2024)
by: Liang, Dingkang, et al.
Published: (2024)
Adaptive Pseudo Label Selection for Individual Unlabeled Data by Positive and Unlabeled Learning
by: Yamane, Takehiro, et al.
Published: (2025)
by: Yamane, Takehiro, et al.
Published: (2025)
CityRAG: Stepping Into a City via Spatially-Grounded Video Generation
by: Chou, Gene, et al.
Published: (2026)
by: Chou, Gene, et al.
Published: (2026)
$Δ$ynamics: Language-Based Representation for Inferring Rigid-Body Dynamics From Videos
by: Kao, Chia-Hsiang, et al.
Published: (2026)
by: Kao, Chia-Hsiang, et al.
Published: (2026)
Advancing Human Action Recognition with Foundation Models trained on Unlabeled Public Videos
by: Qian, Yang, et al.
Published: (2024)
by: Qian, Yang, et al.
Published: (2024)
Fine-Tuning Video-Text Contrastive Model for Primate Behavior Retrieval from Unlabeled Raw Videos
by: Santo, Giulio Cesare Mastrocinque, et al.
Published: (2025)
by: Santo, Giulio Cesare Mastrocinque, et al.
Published: (2025)
DiSciPLE: Learning Interpretable Programs for Scientific Visual Discovery
by: Mall, Utkarsh, et al.
Published: (2025)
by: Mall, Utkarsh, et al.
Published: (2025)
Towards Unstructured Unlabeled Optical Mocap: A Video Helps!
by: Milef, Nicholas, et al.
Published: (2024)
by: Milef, Nicholas, et al.
Published: (2024)
Task Integration Distillation for Object Detectors
by: Su, Hai, et al.
Published: (2024)
by: Su, Hai, et al.
Published: (2024)
DiffuBox: Refining 3D Object Detection with Point Diffusion
by: Chen, Xiangyu, et al.
Published: (2024)
by: Chen, Xiangyu, et al.
Published: (2024)
Automatic Retrieval of Specific Cows from Unlabeled Videos
by: Lyu, Jiawen, et al.
Published: (2025)
by: Lyu, Jiawen, et al.
Published: (2025)
On Calibration of Object Detectors: Pitfalls, Evaluation and Baselines
by: Kuzucu, Selim, et al.
Published: (2024)
by: Kuzucu, Selim, et al.
Published: (2024)
Domain Adaptation for Large-Vocabulary Object Detectors
by: Jiang, Kai, et al.
Published: (2024)
by: Jiang, Kai, et al.
Published: (2024)
Calibrating Probabilistic Object Detectors with Annotator Disagreement
by: Tan, Zhi Qin, et al.
Published: (2026)
by: Tan, Zhi Qin, et al.
Published: (2026)
DiffMOD: Progressive Diffusion Point Denoising for Moving Object Detection in Remote Sensing
by: Zhang, Jinyue, et al.
Published: (2025)
by: Zhang, Jinyue, et al.
Published: (2025)
Similar Items
-
Tracking and Understanding Object Transformations
by: Sun, Yihong, et al.
Published: (2025) -
Live Interactive Training for Video Segmentation
by: Yang, Xinyu, et al.
Published: (2026) -
Video Creation by Demonstration
by: Sun, Yihong, et al.
Published: (2024) -
Learning Feature Descriptors using Camera Pose Supervision
by: Wang, Qianqian, et al.
Published: (2020) -
MovieRecapsQA: A Multimodal Open-Ended Video Question-Answering Benchmark
by: Shaar, Shaden, et al.
Published: (2026)