SceneScript: Reconstructing Scenes With An Autoregressive Structured Language Model
Fuente:
arXiv
Saved in:
| Main Authors: | Avetisyan, Armen, Xie, Christopher, Howard-Jenkins, Henry, Yang, Tsun-Yi, Aroudj, Samir, Patra, Suvam, Zhang, Fuyang, Frost, Duncan, Holland, Luke, Orme, Campbell, Engel, Jakob, Miller, Edward, Newcombe, Richard, Balntas, Vasileios |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Human-in-the-Loop Local Corrections of 3D Scene Layouts via Infilling
by: Xie, Christopher, et al.
Published: (2025)
by: Xie, Christopher, et al.
Published: (2025)
JRM: Joint Reconstruction Model for Multiple Objects without Alignment
by: Wu, Qirui, et al.
Published: (2026)
by: Wu, Qirui, et al.
Published: (2026)
ShapeR: Robust Conditional 3D Shape Generation from Casual Captures
by: Siddiqui, Yawar, et al.
Published: (2026)
by: Siddiqui, Yawar, et al.
Published: (2026)
Fast SceneScript: Fast and Accurate Language-Based 3D Scene Understanding via Multi-Token Prediction
by: Yin, Ruihong, et al.
Published: (2025)
by: Yin, Ruihong, et al.
Published: (2025)
VertexRegen: Mesh Generation with Continuous Level of Detail
by: Zhang, Xiang, et al.
Published: (2025)
by: Zhang, Xiang, et al.
Published: (2025)
Photoreal Scene Reconstruction from an Egocentric Device
by: Lv, Zhaoyang, et al.
Published: (2025)
by: Lv, Zhaoyang, et al.
Published: (2025)
ReplaceAnything3D:Text-Guided 3D Scene Editing with Compositional Neural Radiance Fields
by: Bartrum, Edward, et al.
Published: (2024)
by: Bartrum, Edward, et al.
Published: (2024)
ReScene4D: Temporally Consistent Semantic Instance Segmentation of Evolving Indoor 3D Scenes
by: Steiner, Emily, et al.
Published: (2026)
by: Steiner, Emily, et al.
Published: (2026)
Enhanced detection of time-dependent dielectric structure: Rayleigh's limit and quantum vacuum
by: Mkrtchian, Vanik E., et al.
Published: (2024)
by: Mkrtchian, Vanik E., et al.
Published: (2024)
Select and Summarize: Scene Saliency for Movie Script Summarization
by: Saxena, Rohit, et al.
Published: (2024)
by: Saxena, Rohit, et al.
Published: (2024)
Read the Scene, Not the Script: Outcome-Aware Safety for LLMs
by: Wu, Rui, et al.
Published: (2025)
by: Wu, Rui, et al.
Published: (2025)
Sonata: Self-Supervised Learning of Reliable Point Representations
by: Wu, Xiaoyang, et al.
Published: (2025)
by: Wu, Xiaoyang, et al.
Published: (2025)
NymeriaPlus: Enriching Nymeria Dataset with Additional Annotations and Data
by: DeTone, Daniel, et al.
Published: (2026)
by: DeTone, Daniel, et al.
Published: (2026)
Librarians behind the Scene before the Year Starts
by: Howard, Sue
Published: (2004)
by: Howard, Sue
Published: (2004)
ART: Articulated Reconstruction Transformer
by: Li, Zizhang, et al.
Published: (2025)
by: Li, Zizhang, et al.
Published: (2025)
SATURN: Autoregressive Image Generation Guided by Scene Graphs
by: Vo, Thanh-Nhan, et al.
Published: (2025)
by: Vo, Thanh-Nhan, et al.
Published: (2025)
SIMS: Simulating Stylized Human-Scene Interactions with Retrieval-Augmented Script Generation
by: Wang, Wenjia, et al.
Published: (2024)
by: Wang, Wenjia, et al.
Published: (2024)
LaGen: Towards Autoregressive LiDAR Scene Generation
by: Zhou, Sizhuo, et al.
Published: (2025)
by: Zhou, Sizhuo, et al.
Published: (2025)
The NCPL [Natrona County, Wyoming Public Library] Scene 1980; Script for Video Tape.
Published: (1974)
Published: (1974)
LAMP: Localization Aware Multi-camera People Tracking in Metric 3D World
by: Yang, Nan, et al.
Published: (2026)
by: Yang, Nan, et al.
Published: (2026)
Exponential decay estimates for the resolvent kernel on a Riemannian manifold
by: Avetisyan, Zhirayr
Published: (2025)
by: Avetisyan, Zhirayr
Published: (2025)
PixARMesh: Autoregressive Mesh-Native Single-View Scene Reconstruction
by: Zhang, Xiang, et al.
Published: (2026)
by: Zhang, Xiang, et al.
Published: (2026)
GaussianGPT: Towards Autoregressive 3D Gaussian Scene Generation
by: von Lützow, Nicolas, et al.
Published: (2026)
by: von Lützow, Nicolas, et al.
Published: (2026)
Clockwork Neutrinogenesis: Baryogenesis from theory space
by: Maharana, Suvam, et al.
Published: (2023)
by: Maharana, Suvam, et al.
Published: (2023)
Skeletal Editing through Molecular Recombination of 2H‐Indazoles to Azo‐Linked‐Quinazolinones
by: Suvam Bhattacharjee, et al.
Published: (2024)
by: Suvam Bhattacharjee, et al.
Published: (2024)
TimeChat-Captioner: Scripting Multi-Scene Videos with Time-Aware and Structural Audio-Visual Captions
by: Yao, Linli, et al.
Published: (2026)
by: Yao, Linli, et al.
Published: (2026)
SceneDecorator: Towards Scene-Oriented Story Generation with Scene Planning and Scene Consistency
by: Song, Quanjian, et al.
Published: (2025)
by: Song, Quanjian, et al.
Published: (2025)
PhyScensis: Physics-Augmented LLM Agents for Complex Physical Scene Arrangement
by: Wang, Yian, et al.
Published: (2026)
by: Wang, Yian, et al.
Published: (2026)
Beyond Scanpaths: Graph-Based Gaze Simulation in Dynamic Scenes
by: Palmer, Luke, et al.
Published: (2026)
by: Palmer, Luke, et al.
Published: (2026)
SD-VSum: A Method and Dataset for Script-Driven Video Summarization
by: Mylonas, Manolis, et al.
Published: (2025)
by: Mylonas, Manolis, et al.
Published: (2025)
Ground4D: Spatially-Grounded Feedforward 4D Reconstruction for Unstructured Off-Road Scenes
by: Wang, Shuo, et al.
Published: (2026)
by: Wang, Shuo, et al.
Published: (2026)
ReSpace: Text-Driven Autoregressive 3D Indoor Scene Synthesis and Editing
by: Bucher, Martin JJ., et al.
Published: (2025)
by: Bucher, Martin JJ., et al.
Published: (2025)
DreamForge: Motion-Aware Autoregressive Video Generation for Multi-View Driving Scenes
by: Mei, Jianbiao, et al.
Published: (2024)
by: Mei, Jianbiao, et al.
Published: (2024)
HAAP: Vision-context Hierarchical Attention Autoregressive with Adaptive Permutation for Scene Text Recognition
by: Chen, Honghui, et al.
Published: (2024)
by: Chen, Honghui, et al.
Published: (2024)
Architect: Generating Vivid and Interactive 3D Scenes with Hierarchical 2D Inpainting
by: Wang, Yian, et al.
Published: (2024)
by: Wang, Yian, et al.
Published: (2024)
Open-Vocabulary vs Supervised Learning Methods for Post-Disaster Visual Scene Understanding
by: Michailidou, Anna, et al.
Published: (2026)
by: Michailidou, Anna, et al.
Published: (2026)
Symbolic Graph Inference for Compound Scene Understanding
by: Aryan, FNU, et al.
Published: (2024)
by: Aryan, FNU, et al.
Published: (2024)
SceneAlign: Aligning Multimodal Reasoning to Scene Graphs in Complex Visual Scenes
by: Wang, Chuhan, et al.
Published: (2026)
by: Wang, Chuhan, et al.
Published: (2026)
Interactive Augmented Reality-enabled Outdoor Scene Visualization For Enhanced Real-time Disaster Response
by: Apostolakis, Dimitrios, et al.
Published: (2026)
by: Apostolakis, Dimitrios, et al.
Published: (2026)
Boxer: Robust Lifting of Open-World 2D Bounding Boxes to 3D
by: DeTone, Daniel, et al.
Published: (2026)
by: DeTone, Daniel, et al.
Published: (2026)
Similar Items
-
Human-in-the-Loop Local Corrections of 3D Scene Layouts via Infilling
by: Xie, Christopher, et al.
Published: (2025) -
JRM: Joint Reconstruction Model for Multiple Objects without Alignment
by: Wu, Qirui, et al.
Published: (2026) -
ShapeR: Robust Conditional 3D Shape Generation from Casual Captures
by: Siddiqui, Yawar, et al.
Published: (2026) -
Fast SceneScript: Fast and Accurate Language-Based 3D Scene Understanding via Multi-Token Prediction
by: Yin, Ruihong, et al.
Published: (2025) -
VertexRegen: Mesh Generation with Continuous Level of Detail
by: Zhang, Xiang, et al.
Published: (2025)