Pixie: Fast and Generalizable Supervised Learning of 3D Physics from Pixels
Fuente:
arXiv
Saved in:
| Main Authors: | Le, Long, Lucas, Ryan, Wang, Chen, Chen, Chuhao, Jayaraman, Dinesh, Eaton, Eric, Liu, Lingjie |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Vid2Sim: Generalizable, Video-based Reconstruction of Appearance, Geometry and Physics for Mesh-free Simulation
by: Chen, Chuhao, et al.
Published: (2025)
by: Chen, Chuhao, et al.
Published: (2025)
PhysCtrl: Generative Physics for Controllable and Physics-Grounded Video Generation
by: Wang, Chen, et al.
Published: (2025)
by: Wang, Chen, et al.
Published: (2025)
TLControl: Trajectory and Language Control for Human Motion Synthesis
by: Wan, Weilin, et al.
Published: (2023)
by: Wan, Weilin, et al.
Published: (2023)
GECO: Generative Image-to-3D within a SECOnd
by: Wang, Chen, et al.
Published: (2024)
by: Wang, Chen, et al.
Published: (2024)
Articulate-Anything: Automatic Modeling of Articulated Objects via a Vision-Language Foundation Model
by: Le, Long, et al.
Published: (2024)
by: Le, Long, et al.
Published: (2024)
Improving Pixel Embedding Learning through Intermediate Distance Regression Supervision for Instance Segmentation
by: Wu, Yuli, et al.
Published: (2020)
by: Wu, Yuli, et al.
Published: (2020)
Learning Generalizable 3D Manipulation With 10 Demonstrations
by: Ren, Yu, et al.
Published: (2024)
by: Ren, Yu, et al.
Published: (2024)
FreeArt3D: Training-Free Articulated Object Generation using 3D Diffusion
by: Chen, Chuhao, et al.
Published: (2025)
by: Chen, Chuhao, et al.
Published: (2025)
SG-Reg: Generalizable and Efficient Scene Graph Registration
by: Liu, Chuhao, et al.
Published: (2025)
by: Liu, Chuhao, et al.
Published: (2025)
latentSplat: Autoencoding Variational Gaussians for Fast Generalizable 3D Reconstruction
by: Wewer, Christopher, et al.
Published: (2024)
by: Wewer, Christopher, et al.
Published: (2024)
Learning Generalizable 3D Medical Image Representations from Mask-Guided Self-Supervision
by: Gao, Yunhe, et al.
Published: (2026)
by: Gao, Yunhe, et al.
Published: (2026)
DIMO: Diverse 3D Motion Generation for Arbitrary Objects
by: Mou, Linzhan, et al.
Published: (2025)
by: Mou, Linzhan, et al.
Published: (2025)
ZeroFlow: Scalable Scene Flow via Distillation
by: Vedder, Kyle, et al.
Published: (2023)
by: Vedder, Kyle, et al.
Published: (2023)
GGHead: Fast and Generalizable 3D Gaussian Heads
by: Kirschstein, Tobias, et al.
Published: (2024)
by: Kirschstein, Tobias, et al.
Published: (2024)
Self-Supervised and Generalizable Tokenization for CLIP-Based 3D Understanding
by: Mei, Guofeng, et al.
Published: (2025)
by: Mei, Guofeng, et al.
Published: (2025)
GPS-Gaussian: Generalizable Pixel-wise 3D Gaussian Splatting for Real-time Human Novel View Synthesis
by: Zheng, Shunyuan, et al.
Published: (2023)
by: Zheng, Shunyuan, et al.
Published: (2023)
Recasting Generic Pretrained Vision Transformers As Object-Centric Scene Encoders For Manipulation Policies
by: Qian, Jianing, et al.
Published: (2024)
by: Qian, Jianing, et al.
Published: (2024)
Can Transformers Capture Spatial Relations between Objects?
by: Wen, Chuan, et al.
Published: (2024)
by: Wen, Chuan, et al.
Published: (2024)
Pyramid Pixel Context Adaption Network for Medical Image Classification with Supervised Contrastive Learning
by: Zhang, Xiaoqing, et al.
Published: (2023)
by: Zhang, Xiaoqing, et al.
Published: (2023)
Zero-1-to-G: Taming Pretrained 2D Diffusion Model for Direct 3D Generation
by: Meng, Xuyi, et al.
Published: (2025)
by: Meng, Xuyi, et al.
Published: (2025)
TRAM: Global Trajectory and Motion of 3D Humans from in-the-wild Videos
by: Wang, Yufu, et al.
Published: (2024)
by: Wang, Yufu, et al.
Published: (2024)
Generalizable Pancreas Segmentation via a Dual Self-Supervised Learning Framework
by: Li, Jun, et al.
Published: (2025)
by: Li, Jun, et al.
Published: (2025)
Track Everything Everywhere Fast and Robustly
by: Song, Yunzhou, et al.
Published: (2024)
by: Song, Yunzhou, et al.
Published: (2024)
Pixel to Gaussian: Ultra-Fast Continuous Super-Resolution with 2D Gaussian Modeling
by: Peng, Long, et al.
Published: (2025)
by: Peng, Long, et al.
Published: (2025)
GPS-Gaussian+: Generalizable Pixel-wise 3D Gaussian Splatting for Real-Time Human-Scene Rendering from Sparse Views
by: Zhou, Boyao, et al.
Published: (2024)
by: Zhou, Boyao, et al.
Published: (2024)
PiD: Fast and High-Resolution Latent Decoding with Pixel Diffusion
by: Lu, Yifan, et al.
Published: (2026)
by: Lu, Yifan, et al.
Published: (2026)
tttLRM: Test-Time Training for Long Context and Autoregressive 3D Reconstruction
by: Wang, Chen, et al.
Published: (2026)
by: Wang, Chen, et al.
Published: (2026)
Pixal3D: Pixel-Aligned 3D Generation from Images
by: Li, Dong-Yang, et al.
Published: (2026)
by: Li, Dong-Yang, et al.
Published: (2026)
Frequency Prior Guided Matching: A Data Augmentation Approach for Generalizable Semi-Supervised Polyp Segmentation
by: Xi, Haoran, et al.
Published: (2025)
by: Xi, Haoran, et al.
Published: (2025)
PiLoT: Neural Pixel-to-3D Registration for UAV-based Ego and Target Geo-localization
by: Cheng, Xiaoya, et al.
Published: (2026)
by: Cheng, Xiaoya, et al.
Published: (2026)
In Pursuit of Pixel Supervision for Visual Pre-training
by: Yang, Lihe, et al.
Published: (2025)
by: Yang, Lihe, et al.
Published: (2025)
PhysHMR: Learning Humanoid Control Policies from Vision for Physically Plausible Human Motion Reconstruction
by: Feng, Qiao, et al.
Published: (2025)
by: Feng, Qiao, et al.
Published: (2025)
Supervised Representation Learning towards Generalizable Assembly State Recognition
by: Schoonbeek, Tim J., et al.
Published: (2024)
by: Schoonbeek, Tim J., et al.
Published: (2024)
PixelGaussian: Generalizable 3D Gaussian Reconstruction from Arbitrary Views
by: Fei, Xin, et al.
Published: (2024)
by: Fei, Xin, et al.
Published: (2024)
Plug-and-Play DISep: Separating Dense Instances for Scene-to-Pixel Weakly-Supervised Change Detection in High-Resolution Remote Sensing Images
by: Zhao, Zhenghui, et al.
Published: (2025)
by: Zhao, Zhenghui, et al.
Published: (2025)
GARF: Learning Generalizable 3D Reassembly for Real-World Fractures
by: Li, Sihang, et al.
Published: (2025)
by: Li, Sihang, et al.
Published: (2025)
Speedy-Splat: Fast 3D Gaussian Splatting with Sparse Pixels and Sparse Primitives
by: Hanson, Alex, et al.
Published: (2024)
by: Hanson, Alex, et al.
Published: (2024)
Beyond Gaussians: Fast and High-Fidelity 3D Splatting with Linear Kernels
by: Chen, Haodong, et al.
Published: (2024)
by: Chen, Haodong, et al.
Published: (2024)
FreeSplat++: Generalizable 3D Gaussian Splatting for Efficient Indoor Scene Reconstruction
by: Wang, Yunsong, et al.
Published: (2025)
by: Wang, Yunsong, et al.
Published: (2025)
Learning Efficient and Generalizable Human Representation with Human Gaussian Model
by: Liu, Yifan, et al.
Published: (2025)
by: Liu, Yifan, et al.
Published: (2025)
Similar Items
-
Vid2Sim: Generalizable, Video-based Reconstruction of Appearance, Geometry and Physics for Mesh-free Simulation
by: Chen, Chuhao, et al.
Published: (2025) -
PhysCtrl: Generative Physics for Controllable and Physics-Grounded Video Generation
by: Wang, Chen, et al.
Published: (2025) -
TLControl: Trajectory and Language Control for Human Motion Synthesis
by: Wan, Weilin, et al.
Published: (2023) -
GECO: Generative Image-to-3D within a SECOnd
by: Wang, Chen, et al.
Published: (2024) -
Articulate-Anything: Automatic Modeling of Articulated Objects via a Vision-Language Foundation Model
by: Le, Long, et al.
Published: (2024)