RAISECity: A Multimodal Agent Framework for Reality-Aligned 3D World Generation at City-Scale
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Shengyuan, Zheng, Zhiheng, Shang, Yu, He, Lixuan, Yu, Yangcheng, Hangyu, Fan, Feng, Jie, Liao, Qingmin, Li, Yong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
UrbanWorld: An Urban World Model for 3D City Generation
von: Shang, Yu, et al.
Veröffentlicht: (2024)
von: Shang, Yu, et al.
Veröffentlicht: (2024)
Mem4Nav: Boosting Vision-and-Language Navigation in Urban Environments with a Hierarchical Spatial-Cognition Long-Short Memory System
von: He, Lixuan, et al.
Veröffentlicht: (2025)
von: He, Lixuan, et al.
Veröffentlicht: (2025)
SpatialAct: Probing Spatial Reasoning-to-Action Capabilities of VLM Agents in 3D Scenes
von: Liu, Tianhui, et al.
Veröffentlicht: (2026)
von: Liu, Tianhui, et al.
Veröffentlicht: (2026)
AMFT: Aligning LLM Reasoners by Meta-Learning the Optimal Imitation-Exploration Balance
von: He, Lixuan, et al.
Veröffentlicht: (2025)
von: He, Lixuan, et al.
Veröffentlicht: (2025)
MoWM: Mixture-of-World-Models for Embodied Planning via Latent-to-Pixel Feature Modulation
von: Yu, Yangcheng, et al.
Veröffentlicht: (2025)
von: Yu, Yangcheng, et al.
Veröffentlicht: (2025)
MASC: Boosting Autoregressive Image Generation with a Manifold-Aligned Semantic Clustering
von: He, Lixuan, et al.
Veröffentlicht: (2025)
von: He, Lixuan, et al.
Veröffentlicht: (2025)
AgentExpt: Automating AI Experiment Design with LLM-based Resource Retrieval Agent
von: Li, Yu, et al.
Veröffentlicht: (2025)
von: Li, Yu, et al.
Veröffentlicht: (2025)
OpenCity: A Scalable Platform to Simulate Urban Activities with Massive LLM Agents
von: Yan, Yuwei, et al.
Veröffentlicht: (2024)
von: Yan, Yuwei, et al.
Veröffentlicht: (2024)
Mitigating Geospatial Knowledge Hallucination in Large Language Models: Benchmarking and Dynamic Factuality Aligning
von: Wang, Shengyuan, et al.
Veröffentlicht: (2025)
von: Wang, Shengyuan, et al.
Veröffentlicht: (2025)
Nebula: Enable City-Scale 3D Gaussian Splatting in Virtual Reality via Collaborative Rendering and Accelerated Stereo Rasterization
von: Zhu, He, et al.
Veröffentlicht: (2025)
von: Zhu, He, et al.
Veröffentlicht: (2025)
EconAgent: Large Language Model-Empowered Agents for Simulating Macroeconomic Activities
von: Li, Nian, et al.
Veröffentlicht: (2023)
von: Li, Nian, et al.
Veröffentlicht: (2023)
Learning from Suboptimal Data in Continuous Control via Auto-Regressive Soft Q-Network
von: Liu, Jijia, et al.
Veröffentlicht: (2025)
von: Liu, Jijia, et al.
Veröffentlicht: (2025)
An Infinite Family of Primitive Heron Triangles with Two Sides as Perfect Squares
von: Li, Yangcheng
Veröffentlicht: (2026)
von: Li, Yangcheng
Veröffentlicht: (2026)
A new perspective of arithmetic billiards
von: Li, Yangcheng
Veröffentlicht: (2023)
von: Li, Yangcheng
Veröffentlicht: (2023)
CityEQA: A Hierarchical LLM Agent on Embodied Question Answering Benchmark in City Space
von: Zhao, Yong, et al.
Veröffentlicht: (2025)
von: Zhao, Yong, et al.
Veröffentlicht: (2025)
AgentSwift: Efficient LLM Agent Design via Value-guided Hierarchical Search
von: Li, Yu, et al.
Veröffentlicht: (2025)
von: Li, Yu, et al.
Veröffentlicht: (2025)
UV-SAM: Adapting Segment Anything Model for Urban Village Identification
von: Zhang, Xin, et al.
Veröffentlicht: (2024)
von: Zhang, Xin, et al.
Veröffentlicht: (2024)
Multiple Weaks Win Single Strong: Large Language Models Ensemble Weak Reinforcement Learning Agents into a Supreme One
von: Song, Yiwen, et al.
Veröffentlicht: (2025)
von: Song, Yiwen, et al.
Veröffentlicht: (2025)
Efficient Feature Aggregation and Scale-Aware Regression for Monocular 3D Object Detection
von: Wang, Yifan, et al.
Veröffentlicht: (2024)
von: Wang, Yifan, et al.
Veröffentlicht: (2024)
LLM-Powered Hierarchical Language Agent for Real-time Human-AI Coordination
von: Liu, Jijia, et al.
Veröffentlicht: (2023)
von: Liu, Jijia, et al.
Veröffentlicht: (2023)
AgentSociety Challenge: Designing LLM Agents for User Modeling and Recommendation on Web Platforms
von: Yan, Yuwei, et al.
Veröffentlicht: (2025)
von: Yan, Yuwei, et al.
Veröffentlicht: (2025)
ChemLabs on ChemO: A Multi-Agent System for Multimodal Reasoning on IChO 2025
von: Xu, Qiang, et al.
Veröffentlicht: (2025)
von: Xu, Qiang, et al.
Veröffentlicht: (2025)
A Survey of Large Language Model-Powered Spatial Intelligence Across Scales: Advances in Embodied Agents, Smart Cities, and Earth Science
von: Feng, Jie, et al.
Veröffentlicht: (2025)
von: Feng, Jie, et al.
Veröffentlicht: (2025)
PhysiAgent: An Embodied Agent Framework in Physical World
von: Wang, Zhihao, et al.
Veröffentlicht: (2025)
von: Wang, Zhihao, et al.
Veröffentlicht: (2025)
EmbodiedCity: A Benchmark Platform for Embodied Agent in Real-world City Environment
von: Gao, Chen, et al.
Veröffentlicht: (2024)
von: Gao, Chen, et al.
Veröffentlicht: (2024)
AlignRec: Aligning and Training in Multimodal Recommendations
von: Liu, Yifan, et al.
Veröffentlicht: (2024)
von: Liu, Yifan, et al.
Veröffentlicht: (2024)
Predicting the Energy Landscape of Stochastic Dynamical System via Physics-informed Self-supervised Learning
von: Li, Ruikun, et al.
Veröffentlicht: (2025)
von: Li, Ruikun, et al.
Veröffentlicht: (2025)
GROW: Aligning GRPO with State-Action Modeling for Open-World VLM Agents
von: Wu, Xiongbin, et al.
Veröffentlicht: (2026)
von: Wu, Xiongbin, et al.
Veröffentlicht: (2026)
Self‐Reinforcing Carbon Nanotube Framework with Aligned Network Structure for Flexible and Compact Energy Storage
von: Mingyu Ye, et al.
Veröffentlicht: (2025)
von: Mingyu Ye, et al.
Veröffentlicht: (2025)
Learning Dual-Level Deformable Implicit Representation for Real-World Scale Arbitrary Super-Resolution
von: Li, Zhiheng, et al.
Veröffentlicht: (2024)
von: Li, Zhiheng, et al.
Veröffentlicht: (2024)
TrajAgent: An LLM-Agent Framework for Trajectory Modeling via Large-and-Small Model Collaboration
von: Du, Yuwei, et al.
Veröffentlicht: (2024)
von: Du, Yuwei, et al.
Veröffentlicht: (2024)
$(G,F)$-points on $\mathbb{Q}$-algebraic varieties
von: Li, Yangcheng, et al.
Veröffentlicht: (2025)
von: Li, Yangcheng, et al.
Veröffentlicht: (2025)
Improving locomotive syndrome risk level through community‐led activities to establish walking habits
von: Daisuke Matsushita, et al.
Veröffentlicht: (2024)
von: Daisuke Matsushita, et al.
Veröffentlicht: (2024)
CAMS: A CityGPT-Powered Agentic Framework for Urban Human Mobility Simulation
von: Du, Yuwei, et al.
Veröffentlicht: (2025)
von: Du, Yuwei, et al.
Veröffentlicht: (2025)
SeqTrack3D: Exploring Sequence Information for Robust 3D Point Cloud Tracking
von: Lin, Yu, et al.
Veröffentlicht: (2024)
von: Lin, Yu, et al.
Veröffentlicht: (2024)
KeyWorld: Key Frame Reasoning Enables Effective and Efficient World Models
von: Li, Sibo, et al.
Veröffentlicht: (2025)
von: Li, Sibo, et al.
Veröffentlicht: (2025)
The Evolution of Eco-routing under Population Growth: Evidence from Six U.S. Cities
von: Shi, Zhiheng, et al.
Veröffentlicht: (2026)
von: Shi, Zhiheng, et al.
Veröffentlicht: (2026)
xbench: Tracking Agents Productivity Scaling with Profession-Aligned Real-World Evaluations
von: Chen, Kaiyuan, et al.
Veröffentlicht: (2025)
von: Chen, Kaiyuan, et al.
Veröffentlicht: (2025)
Fusing Cross-Domain Knowledge from Multimodal Data to Solve Problems in the Physical World
von: Zheng, Yu
Veröffentlicht: (2025)
von: Zheng, Yu
Veröffentlicht: (2025)
Carboxyl‐Decorated UiO‐66 Supporting Pd Nanoparticles for Efficient Room‐Temperature Hydrodeoxygenation of Lignin Derivatives
von: Ruixue Yangcheng, et al.
Veröffentlicht: (2024)
von: Ruixue Yangcheng, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
UrbanWorld: An Urban World Model for 3D City Generation
von: Shang, Yu, et al.
Veröffentlicht: (2024) -
Mem4Nav: Boosting Vision-and-Language Navigation in Urban Environments with a Hierarchical Spatial-Cognition Long-Short Memory System
von: He, Lixuan, et al.
Veröffentlicht: (2025) -
SpatialAct: Probing Spatial Reasoning-to-Action Capabilities of VLM Agents in 3D Scenes
von: Liu, Tianhui, et al.
Veröffentlicht: (2026) -
AMFT: Aligning LLM Reasoners by Meta-Learning the Optimal Imitation-Exploration Balance
von: He, Lixuan, et al.
Veröffentlicht: (2025) -
MoWM: Mixture-of-World-Models for Embodied Planning via Latent-to-Pixel Feature Modulation
von: Yu, Yangcheng, et al.
Veröffentlicht: (2025)