Few-View Object Reconstruction with Unknown Categories and Camera Poses
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jiang, Hanwen, Jiang, Zhenyu, Grauman, Kristen, Zhu, Yuke |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2022
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Audio-Visual Camera Pose Estimation with Passive Scene Sounds and In-the-Wild Video
von: Adebi, Daniel, et al.
Veröffentlicht: (2025)
von: Adebi, Daniel, et al.
Veröffentlicht: (2025)
Generic Objects as Pose Probes for Few-shot View Synthesis
von: Gao, Zhirui, et al.
Veröffentlicht: (2024)
von: Gao, Zhirui, et al.
Veröffentlicht: (2024)
Learning Object State Changes in Videos: An Open-World Perspective
von: Xue, Zihui, et al.
Veröffentlicht: (2023)
von: Xue, Zihui, et al.
Veröffentlicht: (2023)
HieraMamba: Video Temporal Grounding via Hierarchical Anchor-Mamba Pooling
von: An, Joungbin, et al.
Veröffentlicht: (2025)
von: An, Joungbin, et al.
Veröffentlicht: (2025)
HOI-Swap: Swapping Objects in Videos with Hand-Object Interaction Awareness
von: Xue, Zihui, et al.
Veröffentlicht: (2024)
von: Xue, Zihui, et al.
Veröffentlicht: (2024)
Switch-a-View: View Selection Learned from Unlabeled In-the-wild Videos
von: Majumder, Sagnik, et al.
Veröffentlicht: (2024)
von: Majumder, Sagnik, et al.
Veröffentlicht: (2024)
Seeing without Pixels: Perception from Camera Trajectories
von: Xue, Zihui, et al.
Veröffentlicht: (2025)
von: Xue, Zihui, et al.
Veröffentlicht: (2025)
ExpertEdit: Learning Skill-Aware Motion Editing from Expert Videos
von: Somayazulu, Arjun, et al.
Veröffentlicht: (2026)
von: Somayazulu, Arjun, et al.
Veröffentlicht: (2026)
Learning Skill-Attributes for Transferable Assessment in Video
von: Ashutosh, Kumar, et al.
Veröffentlicht: (2025)
von: Ashutosh, Kumar, et al.
Veröffentlicht: (2025)
ViewBridge: Curriculum Knowledge Distillation for Activity View-Invariance Under Extreme Viewpoint Changes
von: Somayazulu, Arjun, et al.
Veröffentlicht: (2025)
von: Somayazulu, Arjun, et al.
Veröffentlicht: (2025)
MZEN: Multi-Zoom Enhanced NeRF for 3-D Reconstruction with Unknown Camera Poses
von: Park, Jong-Ik, et al.
Veröffentlicht: (2025)
von: Park, Jong-Ik, et al.
Veröffentlicht: (2025)
Category-level Object Detection, Pose Estimation and Reconstruction from Stereo Images
von: Zhang, Chuanrui, et al.
Veröffentlicht: (2024)
von: Zhang, Chuanrui, et al.
Veröffentlicht: (2024)
UniversalVTG: A Universal and Lightweight Foundation Model for Video Temporal Grounding
von: An, Joungbin, et al.
Veröffentlicht: (2026)
von: An, Joungbin, et al.
Veröffentlicht: (2026)
Dense Dynamic Scene Reconstruction and Camera Pose Estimation from Multi-View Videos
von: Sun, Shuo, et al.
Veröffentlicht: (2026)
von: Sun, Shuo, et al.
Veröffentlicht: (2026)
SPOC: Spatially-Progressing Object State Change Segmentation in Video
von: Mandikal, Priyanka, et al.
Veröffentlicht: (2025)
von: Mandikal, Priyanka, et al.
Veröffentlicht: (2025)
Beyond 'Templates': Category-Agnostic Object Pose, Size, and Shape Estimation from a Single View
von: Zhang, Jinyu, et al.
Veröffentlicht: (2025)
von: Zhang, Jinyu, et al.
Veröffentlicht: (2025)
CleanPose: Category-Level Object Pose Estimation via Causal Learning and Knowledge Distillation
von: Lin, Xiao, et al.
Veröffentlicht: (2025)
von: Lin, Xiao, et al.
Veröffentlicht: (2025)
FIction: 4D Future Interaction Prediction from Video
von: Ashutosh, Kumar, et al.
Veröffentlicht: (2024)
von: Ashutosh, Kumar, et al.
Veröffentlicht: (2024)
Seeing the Arrow of Time in Large Multimodal Models
von: Xue, Zihui, et al.
Veröffentlicht: (2025)
von: Xue, Zihui, et al.
Veröffentlicht: (2025)
Don't Let the Video Speak: Audio-Contrastive Preference Optimization for Audio-Visual Language Models
von: Baid, Ami, et al.
Veröffentlicht: (2026)
von: Baid, Ami, et al.
Veröffentlicht: (2026)
A Construct-Optimize Approach to Sparse View Synthesis without Camera Pose
von: Jiang, Kaiwen, et al.
Veröffentlicht: (2024)
von: Jiang, Kaiwen, et al.
Veröffentlicht: (2024)
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos
von: Majumder, Sagnik, et al.
Veröffentlicht: (2024)
von: Majumder, Sagnik, et al.
Veröffentlicht: (2024)
Learning a Category-level Object Pose Estimator without Pose Annotations
von: Tian, Fengrui, et al.
Veröffentlicht: (2024)
von: Tian, Fengrui, et al.
Veröffentlicht: (2024)
GCE-Pose: Global Context Enhancement for Category-level Object Pose Estimation
von: Li, Weihang, et al.
Veröffentlicht: (2025)
von: Li, Weihang, et al.
Veröffentlicht: (2025)
Learning Unknowns from Unknowns: Diversified Negative Prototypes Generator for Few-Shot Open-Set Recognition
von: Zhang, Zhenyu, et al.
Veröffentlicht: (2024)
von: Zhang, Zhenyu, et al.
Veröffentlicht: (2024)
Indoor 3D Reconstruction with an Unknown Camera-Projector Pair
von: Qi, Zhaoshuai, et al.
Veröffentlicht: (2024)
von: Qi, Zhaoshuai, et al.
Veröffentlicht: (2024)
SkillSight: Efficient First-Person Skill Assessment with Gaze
von: Wu, Chi Hsuan, et al.
Veröffentlicht: (2025)
von: Wu, Chi Hsuan, et al.
Veröffentlicht: (2025)
EgoExo-WM: Unlocking Exo Video for Ego World Models
von: Tran, Danny, et al.
Veröffentlicht: (2026)
von: Tran, Danny, et al.
Veröffentlicht: (2026)
Progress-Aware Video Frame Captioning
von: Xue, Zihui, et al.
Veröffentlicht: (2024)
von: Xue, Zihui, et al.
Veröffentlicht: (2024)
Stitch-a-Demo: Video Demonstrations from Multistep Descriptions
von: Wu, Chi Hsuan, et al.
Veröffentlicht: (2025)
von: Wu, Chi Hsuan, et al.
Veröffentlicht: (2025)
SportSkills: Physical Skill Learning from Sports Instructional Videos
von: Ashutosh, Kumar, et al.
Veröffentlicht: (2026)
von: Ashutosh, Kumar, et al.
Veröffentlicht: (2026)
Marginalized Bundle Adjustment: Multi-View Camera Pose from Monocular Depth Estimates
von: Zhu, Shengjie, et al.
Veröffentlicht: (2026)
von: Zhu, Shengjie, et al.
Veröffentlicht: (2026)
Free-Moving Object Reconstruction and Pose Estimation with Virtual Camera
von: Shi, Haixin, et al.
Veröffentlicht: (2024)
von: Shi, Haixin, et al.
Veröffentlicht: (2024)
TSM-Pose: Topology-Aware Learning with Semantic Mamba for Category-Level Object Pose Estimation
von: Liu, Jinshuo, et al.
Veröffentlicht: (2026)
von: Liu, Jinshuo, et al.
Veröffentlicht: (2026)
CLIPose: Category-Level Object Pose Estimation with Pre-trained Vision-Language Knowledge
von: Lin, Xiao, et al.
Veröffentlicht: (2024)
von: Lin, Xiao, et al.
Veröffentlicht: (2024)
Real3D: Scaling Up Large Reconstruction Models with Real-World Images
von: Jiang, Hanwen, et al.
Veröffentlicht: (2024)
von: Jiang, Hanwen, et al.
Veröffentlicht: (2024)
ComPose: A Unified Completion-Pose Framework for Robust Category-Level Object Pose Estimation
von: Ren, Huan, et al.
Veröffentlicht: (2026)
von: Ren, Huan, et al.
Veröffentlicht: (2026)
THE-Pose: Topological Prior with Hybrid Graph Fusion for Estimating Category-Level 6D Object Pose
von: Lee, Eunho, et al.
Veröffentlicht: (2025)
von: Lee, Eunho, et al.
Veröffentlicht: (2025)
LaPose: Laplacian Mixture Shape Modeling for RGB-Based Category-Level Object Pose Estimation
von: Zhang, Ruida, et al.
Veröffentlicht: (2024)
von: Zhang, Ruida, et al.
Veröffentlicht: (2024)
PosePilot: Steering Camera Pose for Generative World Models with Self-supervised Depth
von: Jin, Bu, et al.
Veröffentlicht: (2025)
von: Jin, Bu, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Audio-Visual Camera Pose Estimation with Passive Scene Sounds and In-the-Wild Video
von: Adebi, Daniel, et al.
Veröffentlicht: (2025) -
Generic Objects as Pose Probes for Few-shot View Synthesis
von: Gao, Zhirui, et al.
Veröffentlicht: (2024) -
Learning Object State Changes in Videos: An Open-World Perspective
von: Xue, Zihui, et al.
Veröffentlicht: (2023) -
HieraMamba: Video Temporal Grounding via Hierarchical Anchor-Mamba Pooling
von: An, Joungbin, et al.
Veröffentlicht: (2025) -
HOI-Swap: Swapping Objects in Videos with Hand-Object Interaction Awareness
von: Xue, Zihui, et al.
Veröffentlicht: (2024)