PTT: Point-Trajectory Transformer for Efficient Temporal 3D Object Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Huang, Kuan-Chih, Lyu, Weijie, Yang, Ming-Hsuan, Tsai, Yi-Hsuan |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Weakly Supervised 3D Object Detection via Multi-Level Visual Guidance
by: Huang, Kuan-Chih, et al.
Published: (2023)
by: Huang, Kuan-Chih, et al.
Published: (2023)
Gaga: Group Any Gaussians via 3D-aware Memory Bank
by: Lyu, Weijie, et al.
Published: (2024)
by: Lyu, Weijie, et al.
Published: (2024)
FaceLift: Learning Generalizable Single Image 3D Face Reconstruction from Synthetic Heads
by: Lyu, Weijie, et al.
Published: (2024)
by: Lyu, Weijie, et al.
Published: (2024)
Reason3D: Searching and Reasoning 3D Segmentation via Large Language Model
by: Huang, Kuan-Chih, et al.
Published: (2024)
by: Huang, Kuan-Chih, et al.
Published: (2024)
FaceCam: Portrait Video Camera Control via Scale-Aware Conditioning
by: Lyu, Weijie, et al.
Published: (2026)
by: Lyu, Weijie, et al.
Published: (2026)
Chat-Edit-3D: Interactive 3D Scene Editing via Text Prompts
by: Fang, Shuangkang, et al.
Published: (2024)
by: Fang, Shuangkang, et al.
Published: (2024)
Edit3r: Instant 3D Scene Editing from Sparse Unposed Images
by: Liu, Jiageng, et al.
Published: (2025)
by: Liu, Jiageng, et al.
Published: (2025)
InstaInpaint: Instant 3D-Scene Inpainting with Masked Large Reconstruction Model
by: You, Junqi, et al.
Published: (2025)
by: You, Junqi, et al.
Published: (2025)
Ranking-aware adapter for text-driven image ordering with CLIP
by: Yu, Wei-Hsiang, et al.
Published: (2024)
by: Yu, Wei-Hsiang, et al.
Published: (2024)
OpenAD: Open-World Autonomous Driving Benchmark for 3D Object Detection
by: Xia, Zhongyu, et al.
Published: (2024)
by: Xia, Zhongyu, et al.
Published: (2024)
Efficiently Disentangling CLIP for Multi-Object Perception
by: Rawlekar, Samyak, et al.
Published: (2025)
by: Rawlekar, Samyak, et al.
Published: (2025)
MeshLLM: Empowering Large Language Models to Progressively Understand and Generate 3D Mesh
by: Fang, Shuangkang, et al.
Published: (2025)
by: Fang, Shuangkang, et al.
Published: (2025)
EA3D: Online Open-World 3D Object Extraction from Streaming Videos
by: Zhou, Xiaoyu, et al.
Published: (2025)
by: Zhou, Xiaoyu, et al.
Published: (2025)
Safety-Aligned 3D Object Detection: Single-Vehicle, Cooperative, and End-to-End Perspectives
by: Liao, Brian Hsuan-Cheng, et al.
Published: (2026)
by: Liao, Brian Hsuan-Cheng, et al.
Published: (2026)
Distribution Discrepancy and Feature Heterogeneity for Active 3D Object Detection
by: Chen, Huang-Yu, et al.
Published: (2024)
by: Chen, Huang-Yu, et al.
Published: (2024)
Self-Attention with State-Object Weighted Combination for Compositional Zero Shot Learning
by: Chang, Cheng-Hong, et al.
Published: (2025)
by: Chang, Cheng-Hong, et al.
Published: (2025)
Layout-your-3D: Controllable and Precise 3D Generation with 2D Blueprint
by: Zhou, Junwei, et al.
Published: (2024)
by: Zhou, Junwei, et al.
Published: (2024)
Text-Driven Image Editing via Learnable Regions
by: Lin, Yuanze, et al.
Published: (2023)
by: Lin, Yuanze, et al.
Published: (2023)
RobuRCDet: Enhancing Robustness of Radar-Camera Fusion in Bird's Eye View for 3D Object Detection
by: Yue, Jingtong, et al.
Published: (2025)
by: Yue, Jingtong, et al.
Published: (2025)
Collaborative Temporal Consistency Learning for Point-supervised Natural Language Video Localization
by: Tao, Zhuo, et al.
Published: (2025)
by: Tao, Zhuo, et al.
Published: (2025)
Confronting Ambiguity in 6D Object Pose Estimation via Score-Based Diffusion on SE(3)
by: Hsiao, Tsu-Ching, et al.
Published: (2023)
by: Hsiao, Tsu-Ching, et al.
Published: (2023)
Tex4D: Zero-shot 4D Scene Texturing with Video Diffusion Models
by: Bao, Jingzhi, et al.
Published: (2024)
by: Bao, Jingzhi, et al.
Published: (2024)
Spatial-Temporal Multi-level Association for Video Object Segmentation
by: Miao, Deshui, et al.
Published: (2024)
by: Miao, Deshui, et al.
Published: (2024)
Efficient Concertormer for Image Deblurring and Beyond
by: Kuo, Pin-Hung, et al.
Published: (2024)
by: Kuo, Pin-Hung, et al.
Published: (2024)
IllumiCraft: Unified Geometry and Illumination Diffusion for Controllable Video Generation
by: Lin, Yuanze, et al.
Published: (2025)
by: Lin, Yuanze, et al.
Published: (2025)
Toward Real-world BEV Perception: Depth Uncertainty Estimation via Gaussian Splatting
by: Lu, Shu-Wei, et al.
Published: (2025)
by: Lu, Shu-Wei, et al.
Published: (2025)
UNETR++: Delving into Efficient and Accurate 3D Medical Image Segmentation
by: Shaker, Abdelrahman, et al.
Published: (2022)
by: Shaker, Abdelrahman, et al.
Published: (2022)
Restage4D: Reanimating Deformable 3D Reconstruction from a Single Video
by: He, Jixuan, et al.
Published: (2025)
by: He, Jixuan, et al.
Published: (2025)
PVTransformer: Point-to-Voxel Transformer for Scalable 3D Object Detection
by: Leng, Zhaoqi, et al.
Published: (2024)
by: Leng, Zhaoqi, et al.
Published: (2024)
Efficient Video Object Segmentation via Modulated Cross-Attention Memory
by: Shaker, Abdelrahman, et al.
Published: (2024)
by: Shaker, Abdelrahman, et al.
Published: (2024)
Voxel or Pillar: Exploring Efficient Point Cloud Representation for 3D Object Detection
by: Huang, Yuhao, et al.
Published: (2023)
by: Huang, Yuhao, et al.
Published: (2023)
HENet++: Hybrid Encoding and Multi-task Learning for 3D Perception and End-to-end Autonomous Driving
by: Xia, Zhongyu, et al.
Published: (2025)
by: Xia, Zhongyu, et al.
Published: (2025)
CoCo4D: Comprehensive and Complex 4D Scene Generation
by: Zhou, Junwei, et al.
Published: (2025)
by: Zhou, Junwei, et al.
Published: (2025)
Learning Knowledge-based Prompts for Robust 3D Mask Presentation Attack Detection
by: Jiang, Fangling, et al.
Published: (2025)
by: Jiang, Fangling, et al.
Published: (2025)
DreamScene4D: Dynamic Multi-Object Scene Generation from Monocular Videos
by: Chu, Wen-Hsuan, et al.
Published: (2024)
by: Chu, Wen-Hsuan, et al.
Published: (2024)
TrajSSL: Trajectory-Enhanced Semi-Supervised 3D Object Detection
by: Jacobson, Philip, et al.
Published: (2024)
by: Jacobson, Philip, et al.
Published: (2024)
USC: Uncompromising Spatial Constraints for Safety-Oriented 3D Object Detectors in Autonomous Driving
by: Liao, Brian Hsuan-Cheng, et al.
Published: (2022)
by: Liao, Brian Hsuan-Cheng, et al.
Published: (2022)
Auto-Vocabulary 3D Object Detection
by: Zhang, Haomeng, et al.
Published: (2025)
by: Zhang, Haomeng, et al.
Published: (2025)
Video Prediction Transformers without Recurrence or Convolution
by: Tang, Yujin, et al.
Published: (2024)
by: Tang, Yujin, et al.
Published: (2024)
GRACE: Graph-Regularized Attentive Convolutional Entanglement with Laplacian Smoothing for Robust DeepFake Video Detection
by: Hsu, Chih-Chung, et al.
Published: (2024)
by: Hsu, Chih-Chung, et al.
Published: (2024)
Similar Items
-
Weakly Supervised 3D Object Detection via Multi-Level Visual Guidance
by: Huang, Kuan-Chih, et al.
Published: (2023) -
Gaga: Group Any Gaussians via 3D-aware Memory Bank
by: Lyu, Weijie, et al.
Published: (2024) -
FaceLift: Learning Generalizable Single Image 3D Face Reconstruction from Synthetic Heads
by: Lyu, Weijie, et al.
Published: (2024) -
Reason3D: Searching and Reasoning 3D Segmentation via Large Language Model
by: Huang, Kuan-Chih, et al.
Published: (2024) -
FaceCam: Portrait Video Camera Control via Scale-Aware Conditioning
by: Lyu, Weijie, et al.
Published: (2026)