Next-Scale Autoregressive Models are Zero-Shot Single-Image Object View Synthesizers
Fuente:
arXiv
Saved in:
| Main Authors: | Yuan, Shiran, Zhao, Hao |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ZeroNVS: Zero-Shot 360-Degree View Synthesis from a Single Image
by: Sargent, Kyle, et al.
Published: (2023)
by: Sargent, Kyle, et al.
Published: (2023)
InteractEdit: Zero-Shot Editing of Human-Object Interactions in Images
by: Hoe, Jiun Tian, et al.
Published: (2025)
by: Hoe, Jiun Tian, et al.
Published: (2025)
ZeroScene: A Zero-Shot Framework for 3D Scene Generation from a Single Image and Controllable Texture Editing
by: Tang, Xiang, et al.
Published: (2025)
by: Tang, Xiang, et al.
Published: (2025)
Copy-Trasform-Paste: Zero-Shot Object-Object Alignment Guided by Vision-Language and Geometric Constraints
by: Gatenyo, Rotem, et al.
Published: (2026)
by: Gatenyo, Rotem, et al.
Published: (2026)
ACT-R: Adaptive Camera Trajectories for Single View 3D Reconstruction
by: Wang, Yizhi, et al.
Published: (2025)
by: Wang, Yizhi, et al.
Published: (2025)
ReShader: View-Dependent Highlights for Single Image View-Synthesis
by: Paliwal, Avinash, et al.
Published: (2023)
by: Paliwal, Avinash, et al.
Published: (2023)
HOI-PAGE: Zero-Shot Human-Object Interaction Generation with Part Affordance Guidance
by: Li, Lei, et al.
Published: (2025)
by: Li, Lei, et al.
Published: (2025)
In-2-4D: Inbetweening from Two Single-View Images to 4D Generation
by: Nag, Sauradip, et al.
Published: (2025)
by: Nag, Sauradip, et al.
Published: (2025)
ARMesh: Autoregressive Mesh Generation via Next-Level-of-Detail Prediction
by: Lei, Jiabao, et al.
Published: (2025)
by: Lei, Jiabao, et al.
Published: (2025)
Zero-Shot Human-Object Interaction Synthesis with Multimodal Priors
by: Lou, Yuke, et al.
Published: (2025)
by: Lou, Yuke, et al.
Published: (2025)
One Shot, One Talk: Whole-body Talking Avatar from a Single Image
by: Xiang, Jun, et al.
Published: (2024)
by: Xiang, Jun, et al.
Published: (2024)
PixARMesh: Autoregressive Mesh-Native Single-View Scene Reconstruction
by: Zhang, Xiang, et al.
Published: (2026)
by: Zhang, Xiang, et al.
Published: (2026)
Self-Ensembling Gaussian Splatting for Few-Shot Novel View Synthesis
by: Zhao, Chen, et al.
Published: (2024)
by: Zhao, Chen, et al.
Published: (2024)
Pippo: High-Resolution Multi-View Humans from a Single Image
by: Kant, Yash, et al.
Published: (2025)
by: Kant, Yash, et al.
Published: (2025)
Capture, Canonicalize, Splat: Zero-Shot 3D Gaussian Avatars from Unstructured Phone Images
by: Garbin, Emanuel, et al.
Published: (2025)
by: Garbin, Emanuel, et al.
Published: (2025)
Streetscapes: Large-scale Consistent Street View Generation Using Autoregressive Video Diffusion
by: Deng, Boyang, et al.
Published: (2024)
by: Deng, Boyang, et al.
Published: (2024)
Cascade-Zero123: One Image to Highly Consistent 3D with Self-Prompted Nearby Views
by: Chen, Yabo, et al.
Published: (2023)
by: Chen, Yabo, et al.
Published: (2023)
ZeroHSI: Zero-Shot 4D Human-Scene Interaction by Video Generation
by: Li, Hongjie, et al.
Published: (2024)
by: Li, Hongjie, et al.
Published: (2024)
Post-mastoidectomy Surface Multi-View Synthesis from a Single Microscopy Image
by: Zhang, Yike, et al.
Published: (2024)
by: Zhang, Yike, et al.
Published: (2024)
Pano2Room: Novel View Synthesis from a Single Indoor Panorama
by: Pu, Guo, et al.
Published: (2024)
by: Pu, Guo, et al.
Published: (2024)
LayerPeeler: Autoregressive Peeling for Layer-wise Image Vectorization
by: Wu, Ronghuan, et al.
Published: (2025)
by: Wu, Ronghuan, et al.
Published: (2025)
Split&Splat: Zero-Shot Panoptic Segmentation via Explicit Instance Modeling and 3D Gaussian Splatting
by: Monchieri, Leonardo, et al.
Published: (2026)
by: Monchieri, Leonardo, et al.
Published: (2026)
SyncLight: Single-Edit Multi-View Relighting
by: Serrano-Lozano, David, et al.
Published: (2026)
by: Serrano-Lozano, David, et al.
Published: (2026)
GaussianObject: High-Quality 3D Object Reconstruction from Four Views with Gaussian Splatting
by: Yang, Chen, et al.
Published: (2024)
by: Yang, Chen, et al.
Published: (2024)
SPAR3D: Stable Point-Aware Reconstruction of 3D Objects from Single Images
by: Huang, Zixuan, et al.
Published: (2025)
by: Huang, Zixuan, et al.
Published: (2025)
Neural Appearance Modeling From Single Images
by: Idema, Jay, et al.
Published: (2024)
by: Idema, Jay, et al.
Published: (2024)
Scaling Transformer-Based Novel View Synthesis Models with Token Disentanglement and Synthetic Data
by: Nair, Nithin Gopalakrishnan, et al.
Published: (2025)
by: Nair, Nithin Gopalakrishnan, et al.
Published: (2025)
Part123: Part-aware 3D Reconstruction from a Single-view Image
by: Liu, Anran, et al.
Published: (2024)
by: Liu, Anran, et al.
Published: (2024)
DartControl: A Diffusion-Based Autoregressive Motion Model for Real-Time Text-Driven Motion Control
by: Zhao, Kaifeng, et al.
Published: (2024)
by: Zhao, Kaifeng, et al.
Published: (2024)
Blended-NeRF: Zero-Shot Object Generation and Blending in Existing Neural Radiance Fields
by: Gordon, Ori, et al.
Published: (2023)
by: Gordon, Ori, et al.
Published: (2023)
NEMTO: Neural Environment Matting for Novel View and Relighting Synthesis of Transparent Objects
by: Wang, Dongqing, et al.
Published: (2023)
by: Wang, Dongqing, et al.
Published: (2023)
ZS-SRT: An Efficient Zero-Shot Super-Resolution Training Method for Neural Radiance Fields
by: Feng, Xiang, et al.
Published: (2023)
by: Feng, Xiang, et al.
Published: (2023)
SuperVoxelGPT: Adaptive and Ordered 3D Tokenization for Autoregressive Shape Generation
by: Li, Yuan, et al.
Published: (2026)
by: Li, Yuan, et al.
Published: (2026)
TelePhysics: Physics-Grounded Multi-Object Scene Generation from a Single Image with Real-Time Interaction
by: Zhang, Xin, et al.
Published: (2026)
by: Zhang, Xin, et al.
Published: (2026)
MotionDreamer: Exploring Semantic Video Diffusion features for Zero-Shot 3D Mesh Animation
by: Uzolas, Lukas, et al.
Published: (2024)
by: Uzolas, Lukas, et al.
Published: (2024)
QuadLink: Autoregressive Quad-Dominant Mesh Generation via Point-Relation Learning
by: Zhang, Yiheng, et al.
Published: (2026)
by: Zhang, Yiheng, et al.
Published: (2026)
3D View Optimization for Improving Image Aesthetics
by: Uchida, Taichi, et al.
Published: (2024)
by: Uchida, Taichi, et al.
Published: (2024)
Real-Time Position-Aware View Synthesis from Single-View Input
by: Gond, Manu, et al.
Published: (2024)
by: Gond, Manu, et al.
Published: (2024)
FSFSplatter: Build Surface and Novel Views with Sparse-Views within 2min
by: Zhao, Yibin, et al.
Published: (2025)
by: Zhao, Yibin, et al.
Published: (2025)
High-Fidelity Single-Image Head Modeling with Industry-Grade Topology
by: Wang, Yunmu, et al.
Published: (2026)
by: Wang, Yunmu, et al.
Published: (2026)
Similar Items
-
ZeroNVS: Zero-Shot 360-Degree View Synthesis from a Single Image
by: Sargent, Kyle, et al.
Published: (2023) -
InteractEdit: Zero-Shot Editing of Human-Object Interactions in Images
by: Hoe, Jiun Tian, et al.
Published: (2025) -
ZeroScene: A Zero-Shot Framework for 3D Scene Generation from a Single Image and Controllable Texture Editing
by: Tang, Xiang, et al.
Published: (2025) -
Copy-Trasform-Paste: Zero-Shot Object-Object Alignment Guided by Vision-Language and Geometric Constraints
by: Gatenyo, Rotem, et al.
Published: (2026) -
ACT-R: Adaptive Camera Trajectories for Single View 3D Reconstruction
by: Wang, Yizhi, et al.
Published: (2025)