Thinking with Novel Views: A Systematic Analysis of Generative-Augmented Spatial Intelligence
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Yanbing, Wang, Bo, Liu, Jianhui, Jiang, Nan, Jiang, Jiaxiu, Sun, Haoze, Yang, Yijun, Zheng, Shenghe, Song, Lin, Huang, Haoyang, Duan, Nan, Li, Wenbo |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
OpenSpatial: A Principled Data Engine for Empowering Spatial Intelligence
by: Liu, Jianhui, et al.
Published: (2026)
by: Liu, Jianhui, et al.
Published: (2026)
TextLDM: Language Modeling with Continuous Latent Diffusion
by: Jiang, Jiaxiu, et al.
Published: (2026)
by: Jiang, Jiaxiu, et al.
Published: (2026)
JoyAI-Image: Awaking Spatial Intelligence in Unified Multimodal Understanding and Generation
by: Song, Lin, et al.
Published: (2026)
by: Song, Lin, et al.
Published: (2026)
V-Bridge: Bridging Video Generative Priors to Versatile Few-shot Image Restoration
by: Zheng, Shenghe, et al.
Published: (2026)
by: Zheng, Shenghe, et al.
Published: (2026)
SpatialEdit: Benchmarking Fine-Grained Image Spatial Editing
by: Xiao, Yicheng, et al.
Published: (2026)
by: Xiao, Yicheng, et al.
Published: (2026)
Frame-Level Captions for Long Video Generation with Complex Multi Scenes
by: Zheng, Guangcong, et al.
Published: (2025)
by: Zheng, Guangcong, et al.
Published: (2025)
A two-state Kalman estimator for atomic gravimetry
by: Jiang, Bo-Nan
Published: (2023)
by: Jiang, Bo-Nan
Published: (2023)
Development of a persuasive User Experience Research (UXR) Point of View for Explainable Artificial Intelligence (XAI)
by: Naiseh, Mohammad, et al.
Published: (2025)
by: Naiseh, Mohammad, et al.
Published: (2025)
Generative Pre-trained Autoregressive Diffusion Transformer
by: Zhang, Yuan, et al.
Published: (2025)
by: Zhang, Yuan, et al.
Published: (2025)
A Systematic Post-Train Framework for Video Generation
by: Xue, Zeyue, et al.
Published: (2026)
by: Xue, Zeyue, et al.
Published: (2026)
Thinking Augmented Pre-training
by: Wang, Liang, et al.
Published: (2025)
by: Wang, Liang, et al.
Published: (2025)
WiS Platform: Enhancing Evaluation of LLM-Based Multi-Agent Systems Through Game-Based Analysis
by: Hu, Chengwei, et al.
Published: (2024)
by: Hu, Chengwei, et al.
Published: (2024)
Towards Effective Next POI Prediction: Spatial and Semantic Augmentation with Remote Sensing Data
by: Jiang, Nan, et al.
Published: (2024)
by: Jiang, Nan, et al.
Published: (2024)
DINO-MVR: Multi-View Readout of Frozen DINOv3 for Annotation-Efficient Medical Segmentation
by: Jiang, Wei, et al.
Published: (2026)
by: Jiang, Wei, et al.
Published: (2026)
STORYANCHORS: Generating Consistent Multi-Scene Story Frames for Long-Form Narratives
by: Wang, Bo, et al.
Published: (2025)
by: Wang, Bo, et al.
Published: (2025)
A Meta‐Analysis of the Impact of Generative Artificial Intelligence on Learning Outcomes
by: Nan Ma, et al.
Published: (2025)
by: Nan Ma, et al.
Published: (2025)
ViewFusion: Structured Spatial Thinking Chains for Multi-View Reasoning
by: Tao, Xingjian, et al.
Published: (2026)
by: Tao, Xingjian, et al.
Published: (2026)
The Correlational Essence of Dimensions in the URUT Framework: A Unified Cosmological Framework from Zero-Dimensional Singularity to Three-Dimensional Entities
by: JiangNan
Published: (2026)
by: JiangNan
Published: (2026)
A Note on Loss Functions and Error Compounding in Model-based Reinforcement Learning
by: Jiang, Nan
Published: (2024)
by: Jiang, Nan
Published: (2024)
Selecting Belief-State Approximations in Simulators with Latent States
by: Jiang, Nan
Published: (2025)
by: Jiang, Nan
Published: (2025)
FREE-Merging: Fourier Transform for Efficient Model Merging
by: Zheng, Shenghe, et al.
Published: (2024)
by: Zheng, Shenghe, et al.
Published: (2024)
IntraMix: Intra-Class Mixup Generation for Accurate Labels and Neighbors
by: Zheng, Shenghe, et al.
Published: (2024)
by: Zheng, Shenghe, et al.
Published: (2024)
Test-time Scaling over Perception: Resolving the Grounding Paradox in Thinking with Images
by: Jiang, Zheng, et al.
Published: (2026)
by: Jiang, Zheng, et al.
Published: (2026)
Enhanced Quantum Entanglement Detection of General Two Qubits Systems Based on Modified CNN‐BiLSTM Model
by: Qian Sun, et al.
Published: (2024)
by: Qian Sun, et al.
Published: (2024)
LoViC: Efficient Long Video Generation with Context Compression
by: Jiang, Jiaxiu, et al.
Published: (2025)
by: Jiang, Jiaxiu, et al.
Published: (2025)
MC$^2$: Multi-concept Guidance for Customized Multi-concept Generation
by: Jiang, Jiaxiu, et al.
Published: (2024)
by: Jiang, Jiaxiu, et al.
Published: (2024)
Can Multimodal Large Language Model Think Analogically?
by: Guo, Diandian, et al.
Published: (2024)
by: Guo, Diandian, et al.
Published: (2024)
Front Cover: Enhanced Quantum Entanglement Detection of General Two Qubits Systems Based on Modified CNN‐BiLSTM Model (Adv. Quantum Technol. 1/2025)
by: Qian Sun, et al.
Published: (2025)
by: Qian Sun, et al.
Published: (2025)
Ultra-Fast Language Generation via Discrete Diffusion Divergence Instruct
by: Zheng, Haoyang, et al.
Published: (2025)
by: Zheng, Haoyang, et al.
Published: (2025)
The Spatial Analysis of the Impact of Economic Uncertainty on Household Consumption in Taiwan
by: Jiun‐Nan Pan
Published: (2025)
by: Jiun‐Nan Pan
Published: (2025)
Embodied3DBench: Benchmarking Low-Level Embodied Spatial Intelligence of Vision Language Models
by: Zhang, Jiyao, et al.
Published: (2026)
by: Zhang, Jiyao, et al.
Published: (2026)
A Unifying View of Coverage in Linear Off-Policy Evaluation
by: Amortila, Philip, et al.
Published: (2026)
by: Amortila, Philip, et al.
Published: (2026)
From Covalent to Supramolecular: Re‐Thinking Lysosome‐Targeting Chimeras Through Modular and Tunable Proximity Control
by: Nan Wang, et al.
Published: (2026)
by: Nan Wang, et al.
Published: (2026)
View-Consistent 3D Scene Editing via Dual-Path Structural Correspondense and Semantic Continuity
by: Li, Pufan, et al.
Published: (2026)
by: Li, Pufan, et al.
Published: (2026)
FSFSplatter: Build Surface and Novel Views with Sparse-Views within 2min
by: Zhao, Yibin, et al.
Published: (2025)
by: Zhao, Yibin, et al.
Published: (2025)
Intelligent design of artificial biocatalyst for biomedical diseases
by: Lijie Zhang, et al.
Published: (2026)
by: Lijie Zhang, et al.
Published: (2026)
SQS: Bayesian DNN Compression through Sparse Quantized Sub-distributions
by: Wang, Ziyi, et al.
Published: (2025)
by: Wang, Ziyi, et al.
Published: (2025)
SubFLOT: Submodel Extraction for Efficient and Personalized Federated Learning via Optimal Transport
by: Jiang, Zheng, et al.
Published: (2026)
by: Jiang, Zheng, et al.
Published: (2026)
Vertical Structure of Local Disk Galaxies revealed by DESI Imaging Data
by: Chen, Nan, et al.
Published: (2026)
by: Chen, Nan, et al.
Published: (2026)
MVLLaVA: An Intelligent Agent for Unified and Flexible Novel View Synthesis
by: Jiang, Hanyu, et al.
Published: (2024)
by: Jiang, Hanyu, et al.
Published: (2024)
Similar Items
-
OpenSpatial: A Principled Data Engine for Empowering Spatial Intelligence
by: Liu, Jianhui, et al.
Published: (2026) -
TextLDM: Language Modeling with Continuous Latent Diffusion
by: Jiang, Jiaxiu, et al.
Published: (2026) -
JoyAI-Image: Awaking Spatial Intelligence in Unified Multimodal Understanding and Generation
by: Song, Lin, et al.
Published: (2026) -
V-Bridge: Bridging Video Generative Priors to Versatile Few-shot Image Restoration
by: Zheng, Shenghe, et al.
Published: (2026) -
SpatialEdit: Benchmarking Fine-Grained Image Spatial Editing
by: Xiao, Yicheng, et al.
Published: (2026)