SpatialMosaic: A Multiview VLM Dataset for Partial Visibility
Fuente:
arXiv
Saved in:
| Main Authors: | Lee, Kanghee, Lee, Injae, Kwak, Minseok, Hong, Jungi, Ryu, Kwonyoung, Park, Jaesik |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
OpenBox: Annotate Any Bounding Boxes in 3D
by: Lee, In-Jae, et al.
Published: (2025)
by: Lee, In-Jae, et al.
Published: (2025)
BUFFER-X: Towards Zero-Shot Point Cloud Registration in Diverse Scenes
by: Seo, Minkyun, et al.
Published: (2025)
by: Seo, Minkyun, et al.
Published: (2025)
3D Geometric Shape Assembly via Efficient Point Cloud Matching
by: Lee, Nahyuk, et al.
Published: (2024)
by: Lee, Nahyuk, et al.
Published: (2024)
Robust Multimodal 3D Object Detection via Modality-Agnostic Decoding and Proximity-based Modality Ensemble
by: Cha, Juhan, et al.
Published: (2024)
by: Cha, Juhan, et al.
Published: (2024)
Distribution Matching Distillation without Fake Score Network
by: Kim, Youngjoong, et al.
Published: (2026)
by: Kim, Youngjoong, et al.
Published: (2026)
F4Splat: Feed-Forward Predictive Densification for Feed-Forward 3D Gaussian Splatting
by: Kim, Injae, et al.
Published: (2026)
by: Kim, Injae, et al.
Published: (2026)
Multiview Geometric Regularization of Gaussian Splatting for Accurate Radiance Fields
by: Kim, Jungeon, et al.
Published: (2025)
by: Kim, Jungeon, et al.
Published: (2025)
PointFix: Learning to Fix Domain Bias for Robust Online Stereo Adaptation
by: Kim, Kwonyoung, et al.
Published: (2022)
by: Kim, Kwonyoung, et al.
Published: (2022)
GRTX: Efficient Ray Tracing for 3D Gaussian-Based Rendering
by: Lee, Junseo, et al.
Published: (2026)
by: Lee, Junseo, et al.
Published: (2026)
Extending CLIP's Image-Text Alignment to Referring Image Segmentation
by: Kim, Seoyeon, et al.
Published: (2023)
by: Kim, Seoyeon, et al.
Published: (2023)
Mosaic3D: Foundation Dataset and Model for Open-Vocabulary 3D Segmentation
by: Lee, Junha, et al.
Published: (2025)
by: Lee, Junha, et al.
Published: (2025)
ParCo-SDF: Learning Prior-Free Partial-to-Complete Signed Distance Fields of Deformable Objects
by: Hwang, Deokmin, et al.
Published: (2026)
by: Hwang, Deokmin, et al.
Published: (2026)
CF3: Compact and Fast 3D Feature Fields
by: Lee, Hyunjoon, et al.
Published: (2025)
by: Lee, Hyunjoon, et al.
Published: (2025)
Deep Cost Ray Fusion for Sparse Depth Video Completion
by: Kim, Jungeon, et al.
Published: (2024)
by: Kim, Jungeon, et al.
Published: (2024)
SPAD : Spatially Aware Multiview Diffusers
by: Kant, Yash, et al.
Published: (2024)
by: Kant, Yash, et al.
Published: (2024)
Improving Editability in Image Generation with Layer-wise Memory
by: Kim, Daneul, et al.
Published: (2025)
by: Kim, Daneul, et al.
Published: (2025)
Recovering Dynamic 3D Sketches from Videos
by: Lee, Jaeah, et al.
Published: (2025)
by: Lee, Jaeah, et al.
Published: (2025)
3Doodle: Compact Abstraction of Objects with 3D Strokes
by: Choi, Changwoon, et al.
Published: (2024)
by: Choi, Changwoon, et al.
Published: (2024)
MASH-VLM: Mitigating Action-Scene Hallucination in Video-LLMs through Disentangled Spatial-Temporal Representations
by: Bae, Kyungho, et al.
Published: (2025)
by: Bae, Kyungho, et al.
Published: (2025)
Designing Concise ConvNets with Columnar Stages
by: Kumar, Ashish, et al.
Published: (2024)
by: Kumar, Ashish, et al.
Published: (2024)
Cross Resolution Encoding-Decoding For Detection Transformers
by: Kumar, Ashish, et al.
Published: (2024)
by: Kumar, Ashish, et al.
Published: (2024)
360 in the Wild: Dataset for Depth Prediction and View Synthesis
by: Park, Kibaek, et al.
Published: (2024)
by: Park, Kibaek, et al.
Published: (2024)
Group-wise Scaling and Orthogonal Decomposition for Domain-Invariant Feature Extraction in Face Anti-Spoofing
by: Jung, Seungjin, et al.
Published: (2025)
by: Jung, Seungjin, et al.
Published: (2025)
Anomaly Detection by Effectively Leveraging Synthetic Images
by: Kang, Sungho, et al.
Published: (2025)
by: Kang, Sungho, et al.
Published: (2025)
TRACE: Your Diffusion Model is Secretly an Instance Edge Detector
by: Jo, Sanghyun, et al.
Published: (2025)
by: Jo, Sanghyun, et al.
Published: (2025)
Ego-1K -- A Large-Scale Multiview Video Dataset for Egocentric Vision
by: Lee, Jae Yong, et al.
Published: (2026)
by: Lee, Jae Yong, et al.
Published: (2026)
Why Far Looks Up: Probing Spatial Representation in Vision-Language Models
by: Min, Cheolhong, et al.
Published: (2026)
by: Min, Cheolhong, et al.
Published: (2026)
Holistic Order Prediction in Natural Scenes
by: Musacchio, Pierre, et al.
Published: (2025)
by: Musacchio, Pierre, et al.
Published: (2025)
Retrieve What's Missing: Coverage-Maximizing Retrieval for Consistent Long Video Generation
by: Joo, Minseok, et al.
Published: (2026)
by: Joo, Minseok, et al.
Published: (2026)
Finding NeMo: Negative-mined Mosaic Augmentation for Referring Image Segmentation
by: Ha, Seongsu, et al.
Published: (2024)
by: Ha, Seongsu, et al.
Published: (2024)
Blockwise Flow Matching: Improving Flow Matching Models For Efficient High-Quality Generation
by: Park, Dogyun, et al.
Published: (2025)
by: Park, Dogyun, et al.
Published: (2025)
Syn4D: A Multiview Synthetic 4D Dataset
by: Jiang, Zeren, et al.
Published: (2026)
by: Jiang, Zeren, et al.
Published: (2026)
Alignment Scores: Robust Metrics for Multiview Pose Accuracy Evaluation
by: Lee, Seong Hun, et al.
Published: (2024)
by: Lee, Seong Hun, et al.
Published: (2024)
Efficient multi-view training for 3D Gaussian Splatting
by: Choi, Minhyuk, et al.
Published: (2025)
by: Choi, Minhyuk, et al.
Published: (2025)
Leveraging Learned Image Prior for 3D Gaussian Compression
by: Shin, Seungjoo, et al.
Published: (2025)
by: Shin, Seungjoo, et al.
Published: (2025)
Locality-aware Gaussian Compression for Fast and High-quality Rendering
by: Shin, Seungjoo, et al.
Published: (2025)
by: Shin, Seungjoo, et al.
Published: (2025)
Metropolis-Hastings Sampling for 3D Gaussian Reconstruction
by: Kim, Hyunjin, et al.
Published: (2025)
by: Kim, Hyunjin, et al.
Published: (2025)
InstantDrag: Improving Interactivity in Drag-based Image Editing
by: Shin, Joonghyuk, et al.
Published: (2024)
by: Shin, Joonghyuk, et al.
Published: (2024)
Optimized Minimal 4D Gaussian Splatting
by: Lee, Minseo, et al.
Published: (2025)
by: Lee, Minseo, et al.
Published: (2025)
Faster Parameter-Efficient Tuning with Token Redundancy Reduction
by: Kim, Kwonyoung, et al.
Published: (2025)
by: Kim, Kwonyoung, et al.
Published: (2025)
Similar Items
-
OpenBox: Annotate Any Bounding Boxes in 3D
by: Lee, In-Jae, et al.
Published: (2025) -
BUFFER-X: Towards Zero-Shot Point Cloud Registration in Diverse Scenes
by: Seo, Minkyun, et al.
Published: (2025) -
3D Geometric Shape Assembly via Efficient Point Cloud Matching
by: Lee, Nahyuk, et al.
Published: (2024) -
Robust Multimodal 3D Object Detection via Modality-Agnostic Decoding and Proximity-based Modality Ensemble
by: Cha, Juhan, et al.
Published: (2024) -
Distribution Matching Distillation without Fake Score Network
by: Kim, Youngjoong, et al.
Published: (2026)