Saved in:
| Main Authors: | Gao, Xiangjun, Zhang, Zhensong, Chen, Dave Zhenyu, Xu, Songcen, Quan, Long, Pérez-Pellitero, Eduardo, Jang, Youngkyoon |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2601.11442 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CoMapGS: Covisibility Map-based Gaussian Splatting for Sparse Novel View Synthesis
by: Jang, Youngkyoon, et al.
Published: (2025)
by: Jang, Youngkyoon, et al.
Published: (2025)
ICo3D: An Interactive Conversational 3D Virtual Human
by: Shaw, Richard, et al.
Published: (2026)
by: Shaw, Richard, et al.
Published: (2026)
SA-ResGS: Self-Augmented Residual 3D Gaussian Splatting for Next Best View Selection
by: Jun-Seong, Kim, et al.
Published: (2026)
by: Jun-Seong, Kim, et al.
Published: (2026)
Off The Grid: Detection of Primitives for Feed-Forward 3D Gaussian Splatting
by: Moreau, Arthur, et al.
Published: (2025)
by: Moreau, Arthur, et al.
Published: (2025)
Video2Layout: Recall and Reconstruct Metric-Grounded Cognitive Map for Spatial Reasoning
by: Huang, Yibin, et al.
Published: (2025)
by: Huang, Yibin, et al.
Published: (2025)
Better Together: Unified Motion Capture and 3D Avatar Reconstruction
by: Moreau, Arthur, et al.
Published: (2025)
by: Moreau, Arthur, et al.
Published: (2025)
Charge: A Comprehensive Novel View Synthesis Benchmark and Dataset to Bind Them All
by: Nazarczuk, Michal, et al.
Published: (2025)
by: Nazarczuk, Michal, et al.
Published: (2025)
SCRREAM : SCan, Register, REnder And Map:A Framework for Annotating Accurate and Dense 3D Indoor Scenes with a Benchmark
by: Jung, HyunJun, et al.
Published: (2024)
by: Jung, HyunJun, et al.
Published: (2024)
ViDAR: Video Diffusion-Aware 4D Reconstruction From Monocular Inputs
by: Nazarczuk, Michal, et al.
Published: (2025)
by: Nazarczuk, Michal, et al.
Published: (2025)
CHROMA: Consistent Harmonization of Multi-View Appearance via Bilateral Grid Prediction
by: Shin, Jisu, et al.
Published: (2025)
by: Shin, Jisu, et al.
Published: (2025)
Diffusion-Based Makeup Transfer with Facial Region-Aware Makeup Features
by: Gao, Zheng, et al.
Published: (2026)
by: Gao, Zheng, et al.
Published: (2026)
GASPACHO: Gaussian Splatting for Controllable Humans and Objects
by: Mir, Aymen, et al.
Published: (2025)
by: Mir, Aymen, et al.
Published: (2025)
SCENIC: Scene-aware Semantic Navigation with Instruction-guided Control
by: Zhang, Xiaohan, et al.
Published: (2024)
by: Zhang, Xiaohan, et al.
Published: (2024)
SpatialReasoner: Towards Explicit and Generalizable 3D Spatial Reasoning
by: Ma, Wufei, et al.
Published: (2025)
by: Ma, Wufei, et al.
Published: (2025)
Complex-Valued 2D Gaussian Representation for Computer-Generated Holography
by: Zhan, Yicheng, et al.
Published: (2025)
by: Zhan, Yicheng, et al.
Published: (2025)
CogniMap3D: Cognitive 3D Mapping and Rapid Retrieval
by: Wang, Feiran, et al.
Published: (2026)
by: Wang, Feiran, et al.
Published: (2026)
MetaSpatial: Reinforcing 3D Spatial Reasoning in VLMs for the Metaverse
by: Pan, Zhenyu, et al.
Published: (2025)
by: Pan, Zhenyu, et al.
Published: (2025)
GRVS: a Generalizable and Recurrent Approach to Monocular Dynamic View Synthesis
by: Tanay, Thomas, et al.
Published: (2026)
by: Tanay, Thomas, et al.
Published: (2026)
Co-Speech Gesture Video Generation via Motion-Decoupled Diffusion Model
by: He, Xu, et al.
Published: (2024)
by: He, Xu, et al.
Published: (2024)
Spatial Chain-of-Thought: Bridging Understanding and Generation Models for Spatial Reasoning Generation
by: Chen, Wei, et al.
Published: (2026)
by: Chen, Wei, et al.
Published: (2026)
Color When It Counts: Grayscale-Guided Online Triggering for Always-On Streaming Video Sensing
by: Cai, Weitong, et al.
Published: (2026)
by: Cai, Weitong, et al.
Published: (2026)
Cog3DMap: Multi-View Vision-Language Reasoning with 3D Cognitive Maps
by: Gwak, Chanyoung, et al.
Published: (2026)
by: Gwak, Chanyoung, et al.
Published: (2026)
Semantics-aware Motion Retargeting with Vision-Language Models
by: Zhang, Haodong, et al.
Published: (2023)
by: Zhang, Haodong, et al.
Published: (2023)
SpatialReward: Bridging the Perception Gap in Online RL for Image Editing via Explicit Spatial Reasoning
by: Long, Yancheng, et al.
Published: (2026)
by: Long, Yancheng, et al.
Published: (2026)
Reasmory: 3D Reconstruction as Explicit Memory for VLMs Spatial Reasoning
by: He, Jixuan, et al.
Published: (2026)
by: He, Jixuan, et al.
Published: (2026)
Reality's Canvas, Language's Brush: Crafting 3D Avatars from Monocular Video
by: Rao, Yuchen, et al.
Published: (2023)
by: Rao, Yuchen, et al.
Published: (2023)
HeadGaS: Real-Time Animatable Head Avatars via 3D Gaussian Splatting
by: Dhamo, Helisa, et al.
Published: (2023)
by: Dhamo, Helisa, et al.
Published: (2023)
LASAR: Towards Spatio-temporal Reasoning with Latent Cognitive Map
by: Tang, Jinzhou, et al.
Published: (2026)
by: Tang, Jinzhou, et al.
Published: (2026)
XGrid-Mapping: Explicit Implicit Hybrid Grid Submaps for Efficient Incremental Neural LiDAR Mapping
by: Song, Zeqing, et al.
Published: (2025)
by: Song, Zeqing, et al.
Published: (2025)
Tag Map: A Text-Based Map for Spatial Reasoning and Navigation with Large Language Models
by: Zhang, Mike, et al.
Published: (2024)
by: Zhang, Mike, et al.
Published: (2024)
Are you Struggling? Dataset and Baselines for Struggle Determination in Assembly Videos
by: Feng, Shijia, et al.
Published: (2024)
by: Feng, Shijia, et al.
Published: (2024)
FreeScale: Scaling 3D Scenes via Certainty-Aware Free-View Generation
by: Jiang, Chenhan, et al.
Published: (2026)
by: Jiang, Chenhan, et al.
Published: (2026)
SpaceMind++: Toward Allocentric Cognitive Maps for Spatially Grounded Video MLLMs
by: Gu, Bo, et al.
Published: (2026)
by: Gu, Bo, et al.
Published: (2026)
MapAnything: Universal Feed-Forward Metric 3D Reconstruction
by: Keetha, Nikhil, et al.
Published: (2025)
by: Keetha, Nikhil, et al.
Published: (2025)
HeightMapNet: Explicit Height Modeling for End-to-End HD Map Learning
by: Qiu, Wenzhao, et al.
Published: (2024)
by: Qiu, Wenzhao, et al.
Published: (2024)
GAP-MLLM: Geometry-Aligned Pre-training for Activating 3D Spatial Perception in Multimodal Large Language Models
by: Zhang, Jiaxin, et al.
Published: (2026)
by: Zhang, Jiaxin, et al.
Published: (2026)
TextField3D: Towards Enhancing Open-Vocabulary 3D Generation with Noisy Text Fields
by: Huang, Tianyu, et al.
Published: (2023)
by: Huang, Tianyu, et al.
Published: (2023)
DreamControl: Control-Based Text-to-3D Generation with 3D Self-Prior
by: Huang, Tianyu, et al.
Published: (2023)
by: Huang, Tianyu, et al.
Published: (2023)
HiFi-123: Towards High-fidelity One Image to 3D Content Generation
by: Yu, Wangbo, et al.
Published: (2023)
by: Yu, Wangbo, et al.
Published: (2023)
Map Feature Perception Metric for Map Generation Quality Assessment and Loss Optimization
by: Sun, Chenxing, et al.
Published: (2025)
by: Sun, Chenxing, et al.
Published: (2025)
Similar Items
-
CoMapGS: Covisibility Map-based Gaussian Splatting for Sparse Novel View Synthesis
by: Jang, Youngkyoon, et al.
Published: (2025) -
ICo3D: An Interactive Conversational 3D Virtual Human
by: Shaw, Richard, et al.
Published: (2026) -
SA-ResGS: Self-Augmented Residual 3D Gaussian Splatting for Next Best View Selection
by: Jun-Seong, Kim, et al.
Published: (2026) -
Off The Grid: Detection of Primitives for Feed-Forward 3D Gaussian Splatting
by: Moreau, Arthur, et al.
Published: (2025) -
Video2Layout: Recall and Reconstruct Metric-Grounded Cognitive Map for Spatial Reasoning
by: Huang, Yibin, et al.
Published: (2025)