Enhancing 2D Representation Learning with a 3D Prior
Fuente:
arXiv
Saved in:
| Main Authors: | Aygün, Mehmet, Dhar, Prithviraj, Yan, Zhicheng, Mac Aodha, Oisin, Ranjan, Rakesh |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SAOR: Single-View Articulated Object Reconstruction
by: Aygün, Mehmet, et al.
Published: (2023)
by: Aygün, Mehmet, et al.
Published: (2023)
DepthCues: Evaluating Monocular Depth Perception in Large Vision Models
by: Danier, Duolikun, et al.
Published: (2024)
by: Danier, Duolikun, et al.
Published: (2024)
Interpretable Text-Guided Image Clustering via Iterative Search
by: Zhao, Bingchen, et al.
Published: (2025)
by: Zhao, Bingchen, et al.
Published: (2025)
View-Consistent Diffusion Representations for 3D-Consistent Video Generation
by: Danier, Duolikun, et al.
Published: (2025)
by: Danier, Duolikun, et al.
Published: (2025)
Representational Similarity via Interpretable Visual Concepts
by: Kondapaneni, Neehar, et al.
Published: (2025)
by: Kondapaneni, Neehar, et al.
Published: (2025)
Improving Semantic Correspondence with Viewpoint-Guided Spherical Maps
by: Mariotti, Octave, et al.
Published: (2023)
by: Mariotti, Octave, et al.
Published: (2023)
Representational Difference Explanations
by: Kondapaneni, Neehar, et al.
Published: (2025)
by: Kondapaneni, Neehar, et al.
Published: (2025)
VesselSDF: Distance Field Priors for Vascular Network Reconstruction
by: Esposito, Salvatore, et al.
Published: (2025)
by: Esposito, Salvatore, et al.
Published: (2025)
CrossSDF: 3D Reconstruction of Thin Structures From Cross-Sections
by: Walker, Thomas, et al.
Published: (2024)
by: Walker, Thomas, et al.
Published: (2024)
Less is More: Discovering Concise Network Explanations
by: Kondapaneni, Neehar, et al.
Published: (2024)
by: Kondapaneni, Neehar, et al.
Published: (2024)
Labeled Data Selection for Category Discovery
by: Zhao, Bingchen, et al.
Published: (2024)
by: Zhao, Bingchen, et al.
Published: (2024)
AirPlanes: Accurate Plane Estimation via 3D-Consistent Embeddings
by: Watson, Jamie, et al.
Published: (2024)
by: Watson, Jamie, et al.
Published: (2024)
BiMotion: B-spline Motion for Text-guided Dynamic 3D Character Generation
by: Wang, Miaowei, et al.
Published: (2026)
by: Wang, Miaowei, et al.
Published: (2026)
WildSAT: Learning Satellite Image Representations from Wildlife Observations
by: Daroya, Rangel, et al.
Published: (2024)
by: Daroya, Rangel, et al.
Published: (2024)
Jamais Vu: Exposing the Generalization Gap in Supervised Semantic Correspondence
by: Mariotti, Octave, et al.
Published: (2025)
by: Mariotti, Octave, et al.
Published: (2025)
The Temporal Trap: Entanglement in Pre-Trained Visual Representations for Visuomotor Policy Learning
by: Tsagkas, Nikolaos, et al.
Published: (2025)
by: Tsagkas, Nikolaos, et al.
Published: (2025)
MotionPhysics: Learnable Motion Distillation for Text-Guided Simulation
by: Wang, Miaowei, et al.
Published: (2026)
by: Wang, Miaowei, et al.
Published: (2026)
CleverBirds: A Multiple-Choice Benchmark for Fine-grained Human Knowledge Tracing
by: Bossemeyer, Leonie, et al.
Published: (2025)
by: Bossemeyer, Leonie, et al.
Published: (2025)
Generating Binary Species Range Maps
by: Dorm, Filip, et al.
Published: (2024)
by: Dorm, Filip, et al.
Published: (2024)
RI3D: Few-Shot Gaussian Splatting With Repair and Inpainting Diffusion Priors
by: Paliwal, Avinash, et al.
Published: (2025)
by: Paliwal, Avinash, et al.
Published: (2025)
Sample-efficient Integration of New Modalities into Large Language Models
by: İnce, Osman Batur, et al.
Published: (2025)
by: İnce, Osman Batur, et al.
Published: (2025)
Attentive Feature Aggregation or: How Policies Learn to Stop Worrying about Robustness and Attend to Task-Relevant Visual Cues
by: Tsagkas, Nikolaos, et al.
Published: (2025)
by: Tsagkas, Nikolaos, et al.
Published: (2025)
Click to Grasp: Zero-Shot Precise Manipulation via Visual Diffusion Descriptors
by: Tsagkas, Nikolaos, et al.
Published: (2024)
by: Tsagkas, Nikolaos, et al.
Published: (2024)
Encoding Semantic Priors into the Weights of Implicit Neural Representation
by: Cai, Zhicheng, et al.
Published: (2024)
by: Cai, Zhicheng, et al.
Published: (2024)
Garment3DGen: 3D Garment Stylization and Texture Generation
by: Sarafianos, Nikolaos, et al.
Published: (2024)
by: Sarafianos, Nikolaos, et al.
Published: (2024)
Enhancing Monocular 3D Hand Reconstruction with Learned Texture Priors
by: Karvounas, Giorgos, et al.
Published: (2025)
by: Karvounas, Giorgos, et al.
Published: (2025)
Enhancing 3D Lane Detection and Topology Reasoning with 2D Lane Priors
by: Li, Han, et al.
Published: (2024)
by: Li, Han, et al.
Published: (2024)
ComPC: Completing a 3D Point Cloud with 2D Diffusion Priors
by: Huang, Tianxin, et al.
Published: (2024)
by: Huang, Tianxin, et al.
Published: (2024)
Vision Learners Meet Web Image-Text Pairs
by: Zhao, Bingchen, et al.
Published: (2023)
by: Zhao, Bingchen, et al.
Published: (2023)
DreamDissector: Learning Disentangled Text-to-3D Generation from 2D Diffusion Priors
by: Yan, Zizheng, et al.
Published: (2024)
by: Yan, Zizheng, et al.
Published: (2024)
MV-DUSt3R+: Single-Stage Scene Reconstruction from Sparse Views In 2 Seconds
by: Tang, Zhenggang, et al.
Published: (2024)
by: Tang, Zhenggang, et al.
Published: (2024)
Enhancing LiDAR Point Features with Foundation Model Priors for 3D Object Detection
by: Mo, Yujian, et al.
Published: (2025)
by: Mo, Yujian, et al.
Published: (2025)
WaSt-3D: Wasserstein-2 Distance for Scene-to-Scene Stylization on 3D Gaussians
by: Kotovenko, Dmytro, et al.
Published: (2024)
by: Kotovenko, Dmytro, et al.
Published: (2024)
MVSAnywhere: Zero-Shot Multi-View Stereo
by: Izquierdo, Sergio, et al.
Published: (2025)
by: Izquierdo, Sergio, et al.
Published: (2025)
Sparse Input View Synthesis: 3D Representations and Reliable Priors
by: Somraj, Nagabhushan
Published: (2024)
by: Somraj, Nagabhushan
Published: (2024)
Shoot-Bounce-3D: Single-Shot Occlusion-Aware 3D from Lidar by Decomposing Two-Bounce Light
by: Klinghoffer, Tzofi, et al.
Published: (2025)
by: Klinghoffer, Tzofi, et al.
Published: (2025)
AssetGen: Deployable 3D Asset Generation at Interactive Speed
by: Wang, Dilin, et al.
Published: (2026)
by: Wang, Dilin, et al.
Published: (2026)
DeTurb: Atmospheric Turbulence Mitigation with Deformable 3D Convolutions and 3D Swin Transformers
by: Zou, Zhicheng, et al.
Published: (2024)
by: Zou, Zhicheng, et al.
Published: (2024)
Illusion3D: 3D Multiview Illusion with 2D Diffusion Priors
by: Feng, Yue, et al.
Published: (2024)
by: Feng, Yue, et al.
Published: (2024)
3D Mesh Editing using Masked LRMs
by: Gao, Will, et al.
Published: (2024)
by: Gao, Will, et al.
Published: (2024)
Similar Items
-
SAOR: Single-View Articulated Object Reconstruction
by: Aygün, Mehmet, et al.
Published: (2023) -
DepthCues: Evaluating Monocular Depth Perception in Large Vision Models
by: Danier, Duolikun, et al.
Published: (2024) -
Interpretable Text-Guided Image Clustering via Iterative Search
by: Zhao, Bingchen, et al.
Published: (2025) -
View-Consistent Diffusion Representations for 3D-Consistent Video Generation
by: Danier, Duolikun, et al.
Published: (2025) -
Representational Similarity via Interpretable Visual Concepts
by: Kondapaneni, Neehar, et al.
Published: (2025)