MesaTask: Towards Task-Driven Tabletop Scene Generation via 3D Spatial Reasoning
Fuente:
arXiv
Saved in:
| Main Authors: | Hao, Jinkun, Liang, Naifu, Luo, Zhen, Xu, Xudong, Zhong, Weipeng, Yi, Ran, Jin, Yichen, Lyu, Zhaoyang, Zheng, Feng, Ma, Lizhuang, Pang, Jiangmiao |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
STABLE: Simulation-Ready Tabletop Layout Generation via a Semantics-Physics Dual System
by: Luo, Zhen, et al.
Published: (2026)
by: Luo, Zhen, et al.
Published: (2026)
InternScenes: A Large-scale Simulatable Indoor Scene Dataset with Realistic Layouts
by: Zhong, Weipeng, et al.
Published: (2025)
by: Zhong, Weipeng, et al.
Published: (2025)
EgoSim: Egocentric World Simulator for Embodied Interaction Generation
by: Hao, Jinkun, et al.
Published: (2026)
by: Hao, Jinkun, et al.
Published: (2026)
Multi-Object Tracking by Hierarchical Visual Representations
by: Cao, Jinkun, et al.
Published: (2024)
by: Cao, Jinkun, et al.
Published: (2024)
Pair2Scene: Learning Local Object Relations for Procedural Scene Generation
by: Ran, Xingjian, et al.
Published: (2026)
by: Ran, Xingjian, et al.
Published: (2026)
RoboVIP: Multi-View Video Generation with Visual Identity Prompting Augments Robot Manipulation
by: Wang, Boyang, et al.
Published: (2026)
by: Wang, Boyang, et al.
Published: (2026)
EmbodiedVSR: Dynamic Scene Graph-Guided Chain-of-Thought Reasoning for Visual Spatial Tasks
by: Zhang, Yi, et al.
Published: (2025)
by: Zhang, Yi, et al.
Published: (2025)
Code-as-Room: Generating 3D Rooms from Top-Down View Images via Agentic Code Synthesis
by: Yang, Yixuan, et al.
Published: (2026)
by: Yang, Yixuan, et al.
Published: (2026)
Unified Human-Scene Interaction via Prompted Chain-of-Contacts
by: Xiao, Zeqi, et al.
Published: (2023)
by: Xiao, Zeqi, et al.
Published: (2023)
MGF: Mixed Gaussian Flow for Diverse Trajectory Prediction
by: Chen, Jiahe, et al.
Published: (2024)
by: Chen, Jiahe, et al.
Published: (2024)
High-Performance Dual-Arm Task and Motion Planning for Tabletop Rearrangement
by: Zhang, Duo, et al.
Published: (2025)
by: Zhang, Duo, et al.
Published: (2025)
TaskGround: Structured Executable Task Inference for Full-Scene Household Reasoning
by: Feng, ZhiYuan, et al.
Published: (2026)
by: Feng, ZhiYuan, et al.
Published: (2026)
SceneFunRI: Reasoning the Invisible for Task-Driven Functional Object Localization
by: Chen, Posheng, et al.
Published: (2026)
by: Chen, Posheng, et al.
Published: (2026)
Reconstructing Topology-Consistent Face Mesh by Volume Rendering from Multi-View Images
by: Wang, Yating, et al.
Published: (2024)
by: Wang, Yating, et al.
Published: (2024)
Delay-Optimal Forwarding and Computation Offloading for Service Chain Tasks
by: Zhang, Jinkun, et al.
Published: (2024)
by: Zhang, Jinkun, et al.
Published: (2024)
PhyScene3D: Physically Consistent Interactive 3D Tabletop Scene Generation
by: Chen, Weixing, et al.
Published: (2026)
by: Chen, Weixing, et al.
Published: (2026)
V-PRISM: Probabilistic Mapping of Unknown Tabletop Scenes
by: Wright, Herbert, et al.
Published: (2024)
by: Wright, Herbert, et al.
Published: (2024)
TabletopGen: Instance-Level Interactive 3D Tabletop Scene Generation from Text or Single Image
by: Wang, Ziqian, et al.
Published: (2025)
by: Wang, Ziqian, et al.
Published: (2025)
Infinite Mobility: Scalable High-Fidelity Synthesis of Articulated Objects via Procedural Generation
by: Lian, Xinyu, et al.
Published: (2025)
by: Lian, Xinyu, et al.
Published: (2025)
Humanoid Goalkeeper: Learning from Position Conditioned Task-Motion Constraints
by: Ren, Junli, et al.
Published: (2025)
by: Ren, Junli, et al.
Published: (2025)
LHManip: A Dataset for Long-Horizon Language-Grounded Manipulation Tasks in Cluttered Tabletop Environments
by: Ceola, Federico, et al.
Published: (2023)
by: Ceola, Federico, et al.
Published: (2023)
MetaGPT: Merging Large Language Models Using Model Exclusive Task Arithmetic
by: Zhou, Yuyan, et al.
Published: (2024)
by: Zhou, Yuyan, et al.
Published: (2024)
Towards Terrain-Aware Task-Driven 3D Scene Graph Generation in Outdoor Environments
by: Samuelson, Chad R, et al.
Published: (2025)
by: Samuelson, Chad R, et al.
Published: (2025)
Enhancing Vision-Language Models for Autonomous Driving through Task-Specific Prompting and Spatial Reasoning
by: Wu, Aodi, et al.
Published: (2025)
by: Wu, Aodi, et al.
Published: (2025)
G$^2$VLM: Geometry Grounded Vision Language Model with Unified 3D Reconstruction and Spatial Reasoning
by: Hu, Wenbo, et al.
Published: (2025)
by: Hu, Wenbo, et al.
Published: (2025)
Chat-Scene++: Exploiting Context-Rich Object Identification for 3D LLM
by: Huang, Haifeng, et al.
Published: (2026)
by: Huang, Haifeng, et al.
Published: (2026)
SIRI-Bench: Challenging VLMs' Spatial Intelligence through Complex Reasoning Tasks
by: Song, Zijian, et al.
Published: (2025)
by: Song, Zijian, et al.
Published: (2025)
Evaluation Hallucination in Multi-Round Incomplete Information Lateral-Driven Reasoning Tasks
by: Dong, Wenhan, et al.
Published: (2025)
by: Dong, Wenhan, et al.
Published: (2025)
MeshCoder: LLM-Powered Structured Mesh Code Generation from Point Clouds
by: Dai, Bingquan, et al.
Published: (2025)
by: Dai, Bingquan, et al.
Published: (2025)
Towards Personal Data Sharing Autonomy:A Task-driven Data Capsule Sharing System
by: Lyu, Qiuyun, et al.
Published: (2024)
by: Lyu, Qiuyun, et al.
Published: (2024)
Towards Latency-Aware 3D Streaming Perception for Autonomous Driving
by: Peng, Jiaqi, et al.
Published: (2025)
by: Peng, Jiaqi, et al.
Published: (2025)
LOGIGEN: Logic-Driven Generation of Verifiable Agentic Tasks
by: Zeng, Yucheng, et al.
Published: (2026)
by: Zeng, Yucheng, et al.
Published: (2026)
Relationship-Aware Hierarchical 3D Scene Graph for Task Reasoning
by: Puigjaner, Albert Gassol, et al.
Published: (2026)
by: Puigjaner, Albert Gassol, et al.
Published: (2026)
ReasonIR: Training Retrievers for Reasoning Tasks
by: Shao, Rulin, et al.
Published: (2025)
by: Shao, Rulin, et al.
Published: (2025)
Ensemble Successor Representations for Task Generalization in Offline-to-Online Reinforcement Learning
by: Wang, Changhong, et al.
Published: (2024)
by: Wang, Changhong, et al.
Published: (2024)
FoREST: Frame of Reference Evaluation in Spatial Reasoning Tasks
by: Premsri, Tanawan, et al.
Published: (2025)
by: Premsri, Tanawan, et al.
Published: (2025)
Spatial propagation in a delayed spruce budworm diffusive model
by: Lizhuang Huang, et al.
Published: (2024)
by: Lizhuang Huang, et al.
Published: (2024)
Omnigrasp: Grasping Diverse Objects with Simulated Humanoids
by: Luo, Zhengyi, et al.
Published: (2024)
by: Luo, Zhengyi, et al.
Published: (2024)
From Physical Degradation Models to Task-Aware All-in-One Image Restoration
by: Gao, Hu, et al.
Published: (2026)
by: Gao, Hu, et al.
Published: (2026)
TDA-RC: Task-Driven Alignment for Knowledge-Based Reasoning Chains in Large Language Models
by: Zhang, Jiaquan, et al.
Published: (2026)
by: Zhang, Jiaquan, et al.
Published: (2026)
Similar Items
-
STABLE: Simulation-Ready Tabletop Layout Generation via a Semantics-Physics Dual System
by: Luo, Zhen, et al.
Published: (2026) -
InternScenes: A Large-scale Simulatable Indoor Scene Dataset with Realistic Layouts
by: Zhong, Weipeng, et al.
Published: (2025) -
EgoSim: Egocentric World Simulator for Embodied Interaction Generation
by: Hao, Jinkun, et al.
Published: (2026) -
Multi-Object Tracking by Hierarchical Visual Representations
by: Cao, Jinkun, et al.
Published: (2024) -
Pair2Scene: Learning Local Object Relations for Procedural Scene Generation
by: Ran, Xingjian, et al.
Published: (2026)