PLOT: Pseudo-Labeling via Video Object Tracking for Scalable Monocular 3D Object Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Lee, Seokyeong, Aung, Sithu, Choi, Junyong, Kim, Seungryong, Kim, Ig-Jae, Cho, Junghyun |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
VideoFrom3D: 3D Scene Video Generation via Complementary Image and Video Diffusion Models
by: Kim, Geonung, et al.
Published: (2025)
by: Kim, Geonung, et al.
Published: (2025)
ASMR: Adaptive Skeleton‐Mesh Rigging and Skinning via 2D Generative Prior
by: Seokhyeon Hong, et al.
Published: (2025)
by: Seokhyeon Hong, et al.
Published: (2025)
ASMR: Adaptive Skeleton-Mesh Rigging and Skinning via 2D Generative Prior
by: Hong, Seokhyeon, et al.
Published: (2025)
by: Hong, Seokhyeon, et al.
Published: (2025)
StyleMM: Stylized 3D Morphable Face Model via Text‐Driven Aligned Image Translation
by: Seungmi Lee, et al.
Published: (2025)
by: Seungmi Lee, et al.
Published: (2025)
Skinned Motion Retargeting with Spatially Adaptive Interaction Guidance
by: Choi, Soojin, et al.
Published: (2026)
by: Choi, Soojin, et al.
Published: (2026)
Monocular Human-Object Reconstruction in the Wild
by: Huo, Chaofan, et al.
Published: (2024)
by: Huo, Chaofan, et al.
Published: (2024)
ObjectMover: Generative Object Movement with Video Prior
by: Yu, Xin, et al.
Published: (2025)
by: Yu, Xin, et al.
Published: (2025)
ComVi: Context-Aware Optimized Comment Display in Video Playback
by: Kim, Minsun, et al.
Published: (2026)
by: Kim, Minsun, et al.
Published: (2026)
Multi-View Pedestrian Occupancy Prediction with a Novel Synthetic Dataset
by: Aung, Sithu, et al.
Published: (2024)
by: Aung, Sithu, et al.
Published: (2024)
Implicit Modeling for 3D-printed Multi-material Computational Object Design via Python
by: Wade, Charles, et al.
Published: (2025)
by: Wade, Charles, et al.
Published: (2025)
Style Brush: Guided Style Transfer for 3D Objects
by: Kovács, Áron Samuel, et al.
Published: (2025)
by: Kovács, Áron Samuel, et al.
Published: (2025)
Learning Generalizable Hand-Object Tracking from Synthetic Demonstrations
by: Wang, Yinhuai, et al.
Published: (2025)
by: Wang, Yinhuai, et al.
Published: (2025)
Neural Clustering for Prefractured Mesh Generation in Real-time Object Destruction
by: Kim, Seunghwan, et al.
Published: (2025)
by: Kim, Seunghwan, et al.
Published: (2025)
Deep Learning Based Facial Retargeting Using Local Patches
by: Choi, Yeonsoo, et al.
Published: (2026)
by: Choi, Yeonsoo, et al.
Published: (2026)
MATT-GS: Masked Attention-based 3DGS for Robot Perception and Object Detection
by: Lee, Jee Won, et al.
Published: (2025)
by: Lee, Jee Won, et al.
Published: (2025)
SALAD: Skeleton-aware Latent Diffusion for Text-driven Motion Generation and Editing
by: Hong, Seokhyeon, et al.
Published: (2025)
by: Hong, Seokhyeon, et al.
Published: (2025)
Instant3dit: Multiview Inpainting for Fast Editing of 3D Objects
by: Barda, Amir, et al.
Published: (2024)
by: Barda, Amir, et al.
Published: (2024)
An Object is Worth 64x64 Pixels: Generating 3D Object via Image Diffusion
by: Yan, Xingguang, et al.
Published: (2024)
by: Yan, Xingguang, et al.
Published: (2024)
3DPhysVideo: Consistency-Guided Flow SDE for Video Generation via 3D Scene Reconstruction and Physical Simulation
by: Kim, Hwidong, et al.
Published: (2026)
by: Kim, Hwidong, et al.
Published: (2026)
StackFLOW: Monocular Human-Object Reconstruction by Stacked Normalizing Flow with Offset
by: Huo, Chaofan, et al.
Published: (2024)
by: Huo, Chaofan, et al.
Published: (2024)
BoxFusion: Reconstruction‐Free Open‐Vocabulary 3D Object Detection via Real‐Time Multi‐View Box Fusion
by: Yuqing Lan, et al.
Published: (2025)
by: Yuqing Lan, et al.
Published: (2025)
Dual Prototype Attention for Unsupervised Video Object Segmentation
by: Cho, Suhwan, et al.
Published: (2022)
by: Cho, Suhwan, et al.
Published: (2022)
Efficient Quantum Circuit Encoding of Object Information in 2D Ray Casting
by: Lee, Seungjae, et al.
Published: (2024)
by: Lee, Seungjae, et al.
Published: (2024)
Learning Explicit Contact for Implicit Reconstruction of Hand-held Objects from Monocular Images
by: Hu, Junxing, et al.
Published: (2023)
by: Hu, Junxing, et al.
Published: (2023)
SceneMotifCoder: Example-driven Visual Program Learning for Generating 3D Object Arrangements
by: Tam, Hou In Ivan, et al.
Published: (2024)
by: Tam, Hou In Ivan, et al.
Published: (2024)
VR-Pipe: Streamlining Hardware Graphics Pipeline for Volume Rendering
by: Lee, Junseo, et al.
Published: (2025)
by: Lee, Junseo, et al.
Published: (2025)
MatMart: Material Reconstruction of 3D Objects via Diffusion
by: Wu, Xiuchao, et al.
Published: (2025)
by: Wu, Xiuchao, et al.
Published: (2025)
SOD‐diffusion: Salient Object Detection via Diffusion‐Based Image Generators
by: Shuo Zhang, et al.
Published: (2024)
by: Shuo Zhang, et al.
Published: (2024)
DC-VSR: Spatially and Temporally Consistent Video Super-Resolution with Video Diffusion Prior
by: Han, Janghyeok, et al.
Published: (2025)
by: Han, Janghyeok, et al.
Published: (2025)
Holographic Parallax Improves 3D Perceptual Realism
by: Kim, Dongyeon, et al.
Published: (2024)
by: Kim, Dongyeon, et al.
Published: (2024)
D$^3$-Human: Dynamic Disentangled Digital Human from Monocular Video
by: Chen, Honghu, et al.
Published: (2025)
by: Chen, Honghu, et al.
Published: (2025)
From Geometry to Culture: An Iterative VLM Layout Framework for Placing Objects in Complex 3D Scene Contexts
by: Asano, Yuto, et al.
Published: (2025)
by: Asano, Yuto, et al.
Published: (2025)
THGS: Lifelike Talking Human Avatar Synthesis From Monocular Video Via 3D Gaussian Splatting
by: Chuang Chen, et al.
Published: (2025)
by: Chuang Chen, et al.
Published: (2025)
VoxDet: Rethinking 3D Semantic Occupancy Prediction as Dense Object Detection
by: Li, Wuyang, et al.
Published: (2025)
by: Li, Wuyang, et al.
Published: (2025)
Material Anything: Generating Materials for Any 3D Object via Diffusion
by: Huang, Xin, et al.
Published: (2024)
by: Huang, Xin, et al.
Published: (2024)
Inter3D: A Benchmark and Strong Baseline for Human-Interactive 3D Object Reconstruction
by: Chen, Gan, et al.
Published: (2025)
by: Chen, Gan, et al.
Published: (2025)
ProjectiveShading: Inserting 3D Objects into Indoor Images with Complex Shadows
by: J. Luo†, et al.
Published: (2026)
by: J. Luo†, et al.
Published: (2026)
GaussianObject: High-Quality 3D Object Reconstruction from Four Views with Gaussian Splatting
by: Yang, Chen, et al.
Published: (2024)
by: Yang, Chen, et al.
Published: (2024)
GRTX: Efficient Ray Tracing for 3D Gaussian-Based Rendering
by: Lee, Junseo, et al.
Published: (2026)
by: Lee, Junseo, et al.
Published: (2026)
MotionDuet: Dual-Conditioned 3D Human Motion Generation with Video-Regularized Text Learning
by: Zhang, Yi-Yang, et al.
Published: (2025)
by: Zhang, Yi-Yang, et al.
Published: (2025)
Similar Items
-
VideoFrom3D: 3D Scene Video Generation via Complementary Image and Video Diffusion Models
by: Kim, Geonung, et al.
Published: (2025) -
ASMR: Adaptive Skeleton‐Mesh Rigging and Skinning via 2D Generative Prior
by: Seokhyeon Hong, et al.
Published: (2025) -
ASMR: Adaptive Skeleton-Mesh Rigging and Skinning via 2D Generative Prior
by: Hong, Seokhyeon, et al.
Published: (2025) -
StyleMM: Stylized 3D Morphable Face Model via Text‐Driven Aligned Image Translation
by: Seungmi Lee, et al.
Published: (2025) -
Skinned Motion Retargeting with Spatially Adaptive Interaction Guidance
by: Choi, Soojin, et al.
Published: (2026)