Saved in:
| Main Authors: | Sun, Fan-Yun, Tremblay, Jonathan, Blukis, Valts, Lin, Kevin, Xu, Danfei, Ivanovic, Boris, Karkus, Peter, Birchfield, Stan, Fox, Dieter, Zhang, Ruohan, Li, Yunzhu, Wu, Jiajun, Pavone, Marco, Haber, Nick |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2304.00673 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Neural Implicit Representation for Building Digital Twins of Unknown Articulated Objects
by: Weng, Yijia, et al.
Published: (2024)
by: Weng, Yijia, et al.
Published: (2024)
GRS: Generating Robotic Simulation Tasks from Real-World Images
by: Zook, Alex, et al.
Published: (2024)
by: Zook, Alex, et al.
Published: (2024)
OG-VLA: Orthographic Image Generation for 3D-Aware Vision-Language Action Model
by: Singh, Ishika, et al.
Published: (2025)
by: Singh, Ishika, et al.
Published: (2025)
RoboSpatial: Teaching Spatial Understanding to 2D and 3D Vision-Language Models for Robotics
by: Song, Chan Hee, et al.
Published: (2024)
by: Song, Chan Hee, et al.
Published: (2024)
BOP-ASK: Object-Interaction Reasoning for Vision-Language Models
by: Bhat, Vineet, et al.
Published: (2025)
by: Bhat, Vineet, et al.
Published: (2025)
3D-Generalist: Self-Improving Vision-Language-Action Models for Crafting 3D Worlds
by: Sun, Fan-Yun, et al.
Published: (2025)
by: Sun, Fan-Yun, et al.
Published: (2025)
Robot Policy Evaluation for Sim-to-Real Transfer: A Benchmarking Perspective
by: Yang, Xuning, et al.
Published: (2025)
by: Yang, Xuning, et al.
Published: (2025)
NeRFDeformer: NeRF Transformation from a Single View via 3D Scene Flows
by: Tang, Zhenggang, et al.
Published: (2024)
by: Tang, Zhenggang, et al.
Published: (2024)
DTPP: Differentiable Joint Conditional Prediction and Cost Evaluation for Tree Policy Planning in Autonomous Driving
by: Huang, Zhiyu, et al.
Published: (2023)
by: Huang, Zhiyu, et al.
Published: (2023)
3D-MVP: 3D Multiview Pretraining for Robotic Manipulation
by: Qian, Shengyi, et al.
Published: (2024)
by: Qian, Shengyi, et al.
Published: (2024)
RVT-2: Learning Precise Manipulation from Few Demonstrations
by: Goyal, Ankit, et al.
Published: (2024)
by: Goyal, Ankit, et al.
Published: (2024)
DreamDrive: Generative 4D Scene Modeling from Street View Images
by: Mao, Jiageng, et al.
Published: (2024)
by: Mao, Jiageng, et al.
Published: (2024)
Closed-Loop Supervised Fine-Tuning of Tokenized Traffic Models
by: Zhang, Zhejun, et al.
Published: (2024)
by: Zhang, Zhejun, et al.
Published: (2024)
RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
by: Yuan, Wentao, et al.
Published: (2024)
by: Yuan, Wentao, et al.
Published: (2024)
Snap-it, Tap-it, Splat-it: Tactile-Informed 3D Gaussian Splatting for Reconstructing Challenging Surfaces
by: Comi, Mauro, et al.
Published: (2024)
by: Comi, Mauro, et al.
Published: (2024)
LoRA3D: Low-Rank Self-Calibration of 3D Geometric Foundation Models
by: Lu, Ziqi, et al.
Published: (2024)
by: Lu, Ziqi, et al.
Published: (2024)
System-Level Analysis of Module Uncertainty Quantification in the Autonomy Pipeline
by: Deglurkar, Sampada, et al.
Published: (2024)
by: Deglurkar, Sampada, et al.
Published: (2024)
RoaD: Rollouts as Demonstrations for Closed-Loop Supervised Fine-Tuning of Autonomous Driving Policies
by: Garcia-Cobo, Guillermo, et al.
Published: (2025)
by: Garcia-Cobo, Guillermo, et al.
Published: (2025)
STORM: Spatio-Temporal Reconstruction Model for Large-Scale Outdoor Scenes
by: Yang, Jiawei, et al.
Published: (2024)
by: Yang, Jiawei, et al.
Published: (2024)
VLA-0: Building State-of-the-Art VLAs with Zero Modification
by: Goyal, Ankit, et al.
Published: (2025)
by: Goyal, Ankit, et al.
Published: (2025)
Extrapolated Urban View Synthesis Benchmark
by: Han, Xiangyu, et al.
Published: (2024)
by: Han, Xiangyu, et al.
Published: (2024)
FoundationPose: Unified 6D Pose Estimation and Tracking of Novel Objects
by: Wen, Bowen, et al.
Published: (2023)
by: Wen, Bowen, et al.
Published: (2023)
cuRoboV2: Dynamics-Aware Motion Generation with Depth-Fused Distance Fields for High-DoF Robots
by: Sundaralingam, Balakumar, et al.
Published: (2026)
by: Sundaralingam, Balakumar, et al.
Published: (2026)
Fast-FoundationStereo: Real-Time Zero-Shot Stereo Matching
by: Wen, Bowen, et al.
Published: (2025)
by: Wen, Bowen, et al.
Published: (2025)
RoboLab: A High-Fidelity Simulation Benchmark for Analysis of Task Generalist Policies
by: Yang, Xuning, et al.
Published: (2026)
by: Yang, Xuning, et al.
Published: (2026)
DT-NVS: Diffusion Transformers for Novel View Synthesis
by: Jang, Wonbong, et al.
Published: (2025)
by: Jang, Wonbong, et al.
Published: (2025)
DistillNeRF: Perceiving 3D Scenes from Single-Glance Images by Distilling Neural Fields and Foundation Model Features
by: Wang, Letian, et al.
Published: (2024)
by: Wang, Letian, et al.
Published: (2024)
FactorSim: Generative Simulation via Factorized Representation
by: Sun, Fan-Yun, et al.
Published: (2024)
by: Sun, Fan-Yun, et al.
Published: (2024)
LoRD: Adapting Differentiable Driving Policies to Distribution Shifts
by: Diehl, Christopher, et al.
Published: (2024)
by: Diehl, Christopher, et al.
Published: (2024)
InstantSplat: Sparse-view Gaussian Splatting in Seconds
by: Fan, Zhiwen, et al.
Published: (2024)
by: Fan, Zhiwen, et al.
Published: (2024)
Producing and Leveraging Online Map Uncertainty in Trajectory Prediction
by: Gu, Xunjiang, et al.
Published: (2024)
by: Gu, Xunjiang, et al.
Published: (2024)
Accelerating Online Mapping and Behavior Prediction via Direct BEV Feature Attention
by: Gu, Xunjiang, et al.
Published: (2024)
by: Gu, Xunjiang, et al.
Published: (2024)
Sample-Efficient Safety Assurances using Conformal Prediction
by: Luo, Rachel, et al.
Published: (2021)
by: Luo, Rachel, et al.
Published: (2021)
Tokenize the World into Object-level Knowledge to Address Long-tail Events in Autonomous Driving
by: Tian, Ran, et al.
Published: (2024)
by: Tian, Ran, et al.
Published: (2024)
Differentiable Room Acoustic Rendering with Multi-View Vision Priors
by: Jin, Derong, et al.
Published: (2025)
by: Jin, Derong, et al.
Published: (2025)
RaySt3R: Predicting Novel Depth Maps for Zero-Shot Object Completion
by: Duisterhof, Bardienus P., et al.
Published: (2025)
by: Duisterhof, Bardienus P., et al.
Published: (2025)
Reconstruction and Simulation of Elastic Objects with Spring-Mass 3D Gaussians
by: Zhong, Licheng, et al.
Published: (2024)
by: Zhong, Licheng, et al.
Published: (2024)
Learning Multiple Initial Solutions to Optimization Problems
by: Sharony, Elad, et al.
Published: (2024)
by: Sharony, Elad, et al.
Published: (2024)
Efficient Multi-Camera Tokenization with Triplanes for End-to-End Driving
by: Ivanovic, Boris, et al.
Published: (2025)
by: Ivanovic, Boris, et al.
Published: (2025)
MVP: Multiple View Prediction Improves GUI Grounding
by: Zhang, Yunzhu, et al.
Published: (2025)
by: Zhang, Yunzhu, et al.
Published: (2025)
Similar Items
-
Neural Implicit Representation for Building Digital Twins of Unknown Articulated Objects
by: Weng, Yijia, et al.
Published: (2024) -
GRS: Generating Robotic Simulation Tasks from Real-World Images
by: Zook, Alex, et al.
Published: (2024) -
OG-VLA: Orthographic Image Generation for 3D-Aware Vision-Language Action Model
by: Singh, Ishika, et al.
Published: (2025) -
RoboSpatial: Teaching Spatial Understanding to 2D and 3D Vision-Language Models for Robotics
by: Song, Chan Hee, et al.
Published: (2024) -
BOP-ASK: Object-Interaction Reasoning for Vision-Language Models
by: Bhat, Vineet, et al.
Published: (2025)