Why Far Looks Up: Probing Spatial Representation in Vision-Language Models
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Min, Cheolhong, Jung, Jaeyun, Lee, Daeun, Jeon, Hyeonseong, Su, Yu, Tremblay, Jonathan, Song, Chan Hee, Park, Jaesik |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Tree-Guided Diffusion Planner
par: Jeon, Hyeonseong, et autres
Publié: (2025)
par: Jeon, Hyeonseong, et autres
Publié: (2025)
RoboSpatial: Teaching Spatial Understanding to 2D and 3D Vision-Language Models for Robotics
par: Song, Chan Hee, et autres
Publié: (2024)
par: Song, Chan Hee, et autres
Publié: (2024)
Keep it SymPL: Symbolic Projective Layout for Allocentric Spatial Reasoning in Vision-Language Models
par: Jang, Jaeyun, et autres
Publié: (2026)
par: Jang, Jaeyun, et autres
Publié: (2026)
Your feelings are reasonable: Emotional validation promotes persistence among preschoolers
par: Jeewon Jeon, et autres
Publié: (2024)
par: Jeewon Jeon, et autres
Publié: (2024)
Inconsistency-Aware Minimization: Improving Generalization with Unlabeled Data
par: Kim, Hee-Sung, et autres
Publié: (2026)
par: Kim, Hee-Sung, et autres
Publié: (2026)
Learning to Continually Learn with the Bayesian Principle
par: Lee, Soochan, et autres
Publié: (2024)
par: Lee, Soochan, et autres
Publié: (2024)
CF3: Compact and Fast 3D Feature Fields
par: Lee, Hyunjoon, et autres
Publié: (2025)
par: Lee, Hyunjoon, et autres
Publié: (2025)
Blind to Position, Biased in Language: Probing Mid-Layer Representational Bias in Vision-Language Encoders for Zero-Shot Language-Grounded Spatial Understanding
par: An, Na Min, et autres
Publié: (2025)
par: An, Na Min, et autres
Publié: (2025)
IM-LUT: Interpolation Mixing Look-Up Tables for Image Super-Resolution
par: Park, Sejin, et autres
Publié: (2025)
par: Park, Sejin, et autres
Publié: (2025)
Deontological Keyword Bias: The Impact of Modal Expressions on Normative Judgments of Language Models
par: Park, Bumjin, et autres
Publié: (2025)
par: Park, Bumjin, et autres
Publié: (2025)
Memorizing Documents with Guidance in Large Language Models
par: Park, Bumjin, et autres
Publié: (2024)
par: Park, Bumjin, et autres
Publié: (2024)
Identifying the Source of Generation for Large Language Models
par: Park, Bumjin, et autres
Publié: (2024)
par: Park, Bumjin, et autres
Publié: (2024)
Revealing Multi-View Hallucination in Large Vision-Language Models
par: Park, Wooje, et autres
Publié: (2026)
par: Park, Wooje, et autres
Publié: (2026)
Metropolis-Hastings Sampling for 3D Gaussian Reconstruction
par: Kim, Hyunjin, et autres
Publié: (2025)
par: Kim, Hyunjin, et autres
Publié: (2025)
Video Color Grading via Look-Up Table Generation
par: Shin, Seunghyun, et autres
Publié: (2025)
par: Shin, Seunghyun, et autres
Publié: (2025)
SVRecon: Sparse Voxel Rasterization for Surface Reconstruction
par: Oh, Seunghun, et autres
Publié: (2025)
par: Oh, Seunghun, et autres
Publié: (2025)
Improving the Machine Learning Stock Trading System: An N‐Period Volatility Labeling and Instance Selection Technique
par: Young Hun Song, et autres
Publié: (2024)
par: Young Hun Song, et autres
Publié: (2024)
SERQ: Saliency-Aware Low-Rank Error Reconstruction for LLM Quantization
par: Park, Yeonsik, et autres
Publié: (2026)
par: Park, Yeonsik, et autres
Publié: (2026)
Probing Network Decisions: Capturing Uncertainties and Unveiling Vulnerabilities Without Label Information
par: Joung, Youngju, et autres
Publié: (2025)
par: Joung, Youngju, et autres
Publié: (2025)
Pre-trained Vision and Language Transformers Are Few-Shot Incremental Learners
par: Park, Keon-Hee, et autres
Publié: (2024)
par: Park, Keon-Hee, et autres
Publié: (2024)
VL-TGS: Trajectory Generation and Selection using Vision Language Models in Mapless Outdoor Environments
par: Song, Daeun, et autres
Publié: (2024)
par: Song, Daeun, et autres
Publié: (2024)
SpatialMosaic: A Multiview VLM Dataset for Partial Visibility
par: Lee, Kanghee, et autres
Publié: (2025)
par: Lee, Kanghee, et autres
Publié: (2025)
Development of Marine‐Degradable Poly(Ester Amide)s with Strong, Up‐Scalable, and Up‐Cyclable Performance
par: Sung Bae Park, et autres
Publié: (2024)
par: Sung Bae Park, et autres
Publié: (2024)
Analyzing Process Data from Computer-Based Assessments: A Tutorial on Preprocessing, Feature Extraction, and Model-Based Inference
par: Hwangbo, Daeun, et autres
Publié: (2026)
par: Hwangbo, Daeun, et autres
Publié: (2026)
AutoSpatial: Visual-Language Reasoning for Social Robot Navigation through Efficient Spatial Reasoning Learning
par: Kong, Yangzhe, et autres
Publié: (2025)
par: Kong, Yangzhe, et autres
Publié: (2025)
Unifying reciprocal and real space atomic dynamics in dilute gases
par: Moon, Jaeyun
Publié: (2026)
par: Moon, Jaeyun
Publié: (2026)
Microscopic view of materials properties of liquids: An atomic scale perspective
par: Moon, Jaeyun
Publié: (2026)
par: Moon, Jaeyun
Publié: (2026)
Psychological capital and individual performance via social capital
par: Jaeyun Jeong
Publié: (2025)
par: Jaeyun Jeong
Publié: (2025)
Recovering Dynamic 3D Sketches from Videos
par: Lee, Jaeah, et autres
Publié: (2025)
par: Lee, Jaeah, et autres
Publié: (2025)
3Doodle: Compact Abstraction of Objects with 3D Strokes
par: Choi, Changwoon, et autres
Publié: (2024)
par: Choi, Changwoon, et autres
Publié: (2024)
Towards Time Series Generation Conditioned on Unstructured Natural Language
par: Woo, Jaeyun, et autres
Publié: (2025)
par: Woo, Jaeyun, et autres
Publié: (2025)
Granular Concept Circuits: Toward a Fine-Grained Circuit Discovery for Concept Representations
par: Kwon, Dahee, et autres
Publié: (2025)
par: Kwon, Dahee, et autres
Publié: (2025)
Holistic Order Prediction in Natural Scenes
par: Musacchio, Pierre, et autres
Publié: (2025)
par: Musacchio, Pierre, et autres
Publié: (2025)
Improving Editability in Image Generation with Layer-wise Memory
par: Kim, Daneul, et autres
Publié: (2025)
par: Kim, Daneul, et autres
Publié: (2025)
Distribution Matching Distillation without Fake Score Network
par: Kim, Youngjoong, et autres
Publié: (2026)
par: Kim, Youngjoong, et autres
Publié: (2026)
Spatial Evaluation of Chironomus plumosus Distribution Around Agricultural Processing Facilities in Response to Climate Change
par: Tae‐Hyeon Kim, et autres
Publié: (2025)
par: Tae‐Hyeon Kim, et autres
Publié: (2025)
Development of Marine‐Degradable Poly(Ester Amide)s with Strong, Up‐Scalable, and Up‐Cyclable Performance (Adv. Mater. 9/2025)
par: Sung Bae Park, et autres
Publié: (2025)
par: Sung Bae Park, et autres
Publié: (2025)
A Dual-Prompting for Interpretable Mental Health Language Models
par: Jeon, Hyolim, et autres
Publié: (2024)
par: Jeon, Hyolim, et autres
Publié: (2024)
Targetless LiDAR-Camera Calibration with Neural Gaussian Splatting
par: Jung, Haebeom, et autres
Publié: (2025)
par: Jung, Haebeom, et autres
Publié: (2025)
FIGLUT: An Energy-Efficient Accelerator Design for FP-INT GEMM Using Look-Up Tables
par: Park, Gunho, et autres
Publié: (2025)
par: Park, Gunho, et autres
Publié: (2025)
Documents similaires
-
Tree-Guided Diffusion Planner
par: Jeon, Hyeonseong, et autres
Publié: (2025) -
RoboSpatial: Teaching Spatial Understanding to 2D and 3D Vision-Language Models for Robotics
par: Song, Chan Hee, et autres
Publié: (2024) -
Keep it SymPL: Symbolic Projective Layout for Allocentric Spatial Reasoning in Vision-Language Models
par: Jang, Jaeyun, et autres
Publié: (2026) -
Your feelings are reasonable: Emotional validation promotes persistence among preschoolers
par: Jeewon Jeon, et autres
Publié: (2024) -
Inconsistency-Aware Minimization: Improving Generalization with Unlabeled Data
par: Kim, Hee-Sung, et autres
Publié: (2026)