Saved in:
| Main Authors: | Wei, Yuxi, Wang, Jingbo, Du, Yuwen, Wang, Dingju, Pan, Liang, Xu, Chenxin, Feng, Yao, Dai, Bo, Chen, Siheng |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2412.08685 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Language-Driven Interactive Traffic Trajectory Generation
by: Xia, Junkai, et al.
Published: (2024)
by: Xia, Junkai, et al.
Published: (2024)
Editable Scene Simulation for Autonomous Driving via Collaborative LLM-Agents
by: Wei, Yuxi, et al.
Published: (2024)
by: Wei, Yuxi, et al.
Published: (2024)
HoloDrive: Holistic 2D-3D Multi-Modal Street Scene Generation for Autonomous Driving
by: Wu, Zehuan, et al.
Published: (2024)
by: Wu, Zehuan, et al.
Published: (2024)
Self-Localized Collaborative Perception
by: Ni, Zhenyang, et al.
Published: (2024)
by: Ni, Zhenyang, et al.
Published: (2024)
Robust Collaborative Perception without External Localization and Clock Devices
by: Lei, Zixing, et al.
Published: (2024)
by: Lei, Zixing, et al.
Published: (2024)
Unveiling the Impact of Data and Model Scaling on High-Level Control for Humanoid Robots
by: Wei, Yuxi, et al.
Published: (2025)
by: Wei, Yuxi, et al.
Published: (2025)
TokenHSI: Unified Synthesis of Physical Human-Scene Interactions through Task Tokenization
by: Pan, Liang, et al.
Published: (2025)
by: Pan, Liang, et al.
Published: (2025)
ELoG-GS: Dual-Branch Gaussian Splatting with Luminance-Guided Enhancement for Extreme Low-light 3D Reconstruction
by: Liu, Yuhao, et al.
Published: (2026)
by: Liu, Yuhao, et al.
Published: (2026)
Decentralized and Lifelong-Adaptive Multi-Agent Collaborative Learning
by: Tang, Shuo, et al.
Published: (2024)
by: Tang, Shuo, et al.
Published: (2024)
From Complex Dynamics to DynFormer: Rethinking Transformers for PDEs
by: Lai, Pengyu, et al.
Published: (2026)
by: Lai, Pengyu, et al.
Published: (2026)
Ethical-Lens: Curbing Malicious Usages of Open-Source Text-to-Image Models
by: Cai, Yuzhu, et al.
Published: (2024)
by: Cai, Yuzhu, et al.
Published: (2024)
DynOPETs: A Versatile Benchmark for Dynamic Object Pose Estimation and Tracking in Moving Camera Scenarios
by: Meng, Xiangting, et al.
Published: (2025)
by: Meng, Xiangting, et al.
Published: (2025)
ChatBEV: A Visual Language Model that Understands BEV Maps
by: Xu, Qingyao, et al.
Published: (2025)
by: Xu, Qingyao, et al.
Published: (2025)
LongSeeker: Elastic Context Orchestration for Long-Horizon Search Agents
by: Lu, Yijun, et al.
Published: (2026)
by: Lu, Yijun, et al.
Published: (2026)
DynHD: Hallucination Detection for Diffusion Large Language Models via Denoising Dynamics Deviation Learning
by: Qian, Yanyu, et al.
Published: (2026)
by: Qian, Yanyu, et al.
Published: (2026)
SemGrasp: Semantic Grasp Generation via Language Aligned Discretization
by: Li, Kailin, et al.
Published: (2024)
by: Li, Kailin, et al.
Published: (2024)
BrowseMaster: Towards Scalable Web Browsing via Tool-Augmented Programmatic Agent Pair
by: Pang, Xianghe, et al.
Published: (2025)
by: Pang, Xianghe, et al.
Published: (2025)
BuilDyn: Excitation-Driven Data Generation for Building Thermal Dynamics Modeling and Control
by: Koch, Felix, et al.
Published: (2026)
by: Koch, Felix, et al.
Published: (2026)
Synthesizing Physically Plausible Human Motions in 3D Scenes
by: Pan, Liang, et al.
Published: (2023)
by: Pan, Liang, et al.
Published: (2023)
OregairuChar: A Benchmark Dataset for Character Appearance Frequency Analysis in My Teen Romantic Comedy SNAFU
by: Sun, Qi, et al.
Published: (2025)
by: Sun, Qi, et al.
Published: (2025)
DynORecon: Dynamic Object Reconstruction for Navigation
by: Wang, Yiduo, et al.
Published: (2024)
by: Wang, Yiduo, et al.
Published: (2024)
Street Gaussians: Modeling Dynamic Urban Scenes with Gaussian Splatting
by: Yan, Yunzhi, et al.
Published: (2024)
by: Yan, Yunzhi, et al.
Published: (2024)
From Image Generation to Infrastructure Design: a Multi-agent Pipeline for Street Design Generation
by: Wang, Chenguang, et al.
Published: (2025)
by: Wang, Chenguang, et al.
Published: (2025)
Physical Backdoor Attack can Jeopardize Driving with Vision-Large-Language Models
by: Ni, Zhenyang, et al.
Published: (2024)
by: Ni, Zhenyang, et al.
Published: (2024)
DynFlowDrive: Flow-Based Dynamic World Modeling for Autonomous Driving
by: Liu, Xiaolu, et al.
Published: (2026)
by: Liu, Xiaolu, et al.
Published: (2026)
Controllable Text-to-Motion Generation via Modular Body-Part Phase Control
by: Dai, Minyue, et al.
Published: (2026)
by: Dai, Minyue, et al.
Published: (2026)
Text2NeRF: Text-Driven 3D Scene Generation with Neural Radiance Fields
by: Zhang, Jingbo, et al.
Published: (2023)
by: Zhang, Jingbo, et al.
Published: (2023)
Collaborative Uncertainty Benefits Multi-Agent Multi-Modal Trajectory Forecasting
by: Tang, Bohan, et al.
Published: (2022)
by: Tang, Bohan, et al.
Published: (2022)
Unified Human-Scene Interaction via Prompted Chain-of-Contacts
by: Xiao, Zeqi, et al.
Published: (2023)
by: Xiao, Zeqi, et al.
Published: (2023)
RoomTex: Texturing Compositional Indoor Scenes via Iterative Inpainting
by: Wang, Qi, et al.
Published: (2024)
by: Wang, Qi, et al.
Published: (2024)
InterControl: Zero-shot Human Interaction Generation by Controlling Every Joint
by: Wang, Zhenzhi, et al.
Published: (2023)
by: Wang, Zhenzhi, et al.
Published: (2023)
Self-Supervised Bird's Eye View Motion Prediction with Cross-Modality Signals
by: Fang, Shaoheng, et al.
Published: (2024)
by: Fang, Shaoheng, et al.
Published: (2024)
RoomCraft: Controllable and Complete 3D Indoor Scene Generation
by: Zhou, Mengqi, et al.
Published: (2025)
by: Zhou, Mengqi, et al.
Published: (2025)
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding
by: Zhang, Peng, et al.
Published: (2026)
by: Zhang, Peng, et al.
Published: (2026)
SIMS: Simulating Stylized Human-Scene Interactions with Retrieval-Augmented Script Generation
by: Wang, Wenjia, et al.
Published: (2024)
by: Wang, Wenjia, et al.
Published: (2024)
Image Patch-Matching with Graph-Based Learning in Street Scenes
by: She, Rui, et al.
Published: (2023)
by: She, Rui, et al.
Published: (2023)
DynST: Dynamic Sparse Training for Resource-Constrained Spatio-Temporal Forecasting
by: Wu, Hao, et al.
Published: (2024)
by: Wu, Hao, et al.
Published: (2024)
DynSplit-KV: Dynamic Semantic Splitting for KVCache Compression in Efficient Long-Context LLM Inference
by: Ye, Jiancai, et al.
Published: (2026)
by: Ye, Jiancai, et al.
Published: (2026)
DynCIM: Dynamic Curriculum for Imbalanced Multimodal Learning
by: Qian, Chengxuan, et al.
Published: (2025)
by: Qian, Chengxuan, et al.
Published: (2025)
InterDyn: Controllable Interactive Dynamics with Video Diffusion Models
by: Akkerman, Rick, et al.
Published: (2024)
by: Akkerman, Rick, et al.
Published: (2024)
Similar Items
-
Language-Driven Interactive Traffic Trajectory Generation
by: Xia, Junkai, et al.
Published: (2024) -
Editable Scene Simulation for Autonomous Driving via Collaborative LLM-Agents
by: Wei, Yuxi, et al.
Published: (2024) -
HoloDrive: Holistic 2D-3D Multi-Modal Street Scene Generation for Autonomous Driving
by: Wu, Zehuan, et al.
Published: (2024) -
Self-Localized Collaborative Perception
by: Ni, Zhenyang, et al.
Published: (2024) -
Robust Collaborative Perception without External Localization and Clock Devices
by: Lei, Zixing, et al.
Published: (2024)