Spatial Reasoning in Foundation Models: Benchmarking Object-Centric Spatial Understanding
Fuente:
arXiv
Saved in:
| Main Authors: | Mirjalili, Vahid, Giahi, Ramin, Kollipara, Sriram, Kekuda, Akshay, Yao, Kehui, Zhao, Kai, Xu, Jianpeng, Nag, Kaushiki, Subramaniam, Sinduja, Biswas, Topojoy, Korpeoglu, Evren, Achan, Kannan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
VL-CLIP: Enhancing Multimodal Recommendations via Visual Grounding and LLM-Augmented CLIP Embeddings
by: Giahi, Ramin, et al.
Published: (2025)
by: Giahi, Ramin, et al.
Published: (2025)
CASE: Cadence-Aware Set Encoding for Large-Scale Next Basket Repurchase Recommendation
by: Cao, Yanan, et al.
Published: (2026)
by: Cao, Yanan, et al.
Published: (2026)
LayoutAgent: A Vision-Language Agent Guided Compositional Diffusion for Spatial Layout Planning
by: Fan, Zezhong, et al.
Published: (2025)
by: Fan, Zezhong, et al.
Published: (2025)
CAL-RAG: Retrieval-Augmented Multi-Agent Generation for Content-Aware Layout Design
by: Forouzandehmehr, Najmeh, et al.
Published: (2025)
by: Forouzandehmehr, Najmeh, et al.
Published: (2025)
Grocery to General Merchandise: A Cross-Pollination Recommender using LLMs and Real-Time Cart Context
by: Kekuda, Akshay, et al.
Published: (2025)
by: Kekuda, Akshay, et al.
Published: (2025)
Decoding Style: Efficient Fine-Tuning of LLMs for Image-Guided Outfit Recommendation with Preference
by: Forouzandehmehr, Najmeh, et al.
Published: (2024)
by: Forouzandehmehr, Najmeh, et al.
Published: (2024)
Prompt Optimizer of Text-to-Image Diffusion Models for Abstract Concept Understanding
by: Fan, Zezhong, et al.
Published: (2024)
by: Fan, Zezhong, et al.
Published: (2024)
Segment and Matte Anything in a Unified Model
by: Fan, Zezhong, et al.
Published: (2026)
by: Fan, Zezhong, et al.
Published: (2026)
S2SRec2: Set-to-Set Recommendation for Basket Completion with Recipe
by: Cao, Yanan, et al.
Published: (2025)
by: Cao, Yanan, et al.
Published: (2025)
LLM-Ensemble: Optimal Large Language Model Ensemble Method for E-commerce Product Attribute Value Extraction
by: Fang, Chenhao, et al.
Published: (2024)
by: Fang, Chenhao, et al.
Published: (2024)
Is More Context Always Better? Examining LLM Reasoning Capability for Time Interval Prediction
by: Cao, Yanan, et al.
Published: (2026)
by: Cao, Yanan, et al.
Published: (2026)
Character-based Outfit Generation with Vision-augmented Style Extraction via LLMs
by: Forouzandehmehr, Najmeh, et al.
Published: (2024)
by: Forouzandehmehr, Najmeh, et al.
Published: (2024)
CARTS: Collaborative Agents for Recommendation Textual Summarization
by: Chen, Jiao, et al.
Published: (2025)
by: Chen, Jiao, et al.
Published: (2025)
No-Human in the Loop: Agentic Evaluation at Scale for Recommendation
by: Zhang, Tao, et al.
Published: (2025)
by: Zhang, Tao, et al.
Published: (2025)
Leveraging User-Generated Reviews for Recommender Systems with Dynamic Headers
by: Vashishtha, Shanu, et al.
Published: (2024)
by: Vashishtha, Shanu, et al.
Published: (2024)
Campaign-2-PT-RAG: LLM-Guided Semantic Product Type Attribution for Scalable Campaign Ranking
by: Che, Yiming, et al.
Published: (2026)
by: Che, Yiming, et al.
Published: (2026)
LLM-HYPER: Generative CTR Modeling for Cold-Start Ad Personalization via LLM-Based Hypernetworks
by: Ma, Luyi, et al.
Published: (2026)
by: Ma, Luyi, et al.
Published: (2026)
Chaining text-to-image and large language model: A novel approach for generating personalized e-commerce banners
by: Vashishtha, Shanu, et al.
Published: (2024)
by: Vashishtha, Shanu, et al.
Published: (2024)
Triple Modality Fusion: Aligning Visual, Textual, and Graph Data with Large Language Models for Multi-Behavior Recommendations
by: Ma, Luyi, et al.
Published: (2024)
by: Ma, Luyi, et al.
Published: (2024)
GRACE: Generative Recommendation via Journey-Aware Sparse Attention on Chain-of-Thought Tokenization
by: Ma, Luyi, et al.
Published: (2025)
by: Ma, Luyi, et al.
Published: (2025)
CRAB: Codebook Rebalancing for Bias Mitigation in Generative Recommendation
by: Fan, Zezhong, et al.
Published: (2026)
by: Fan, Zezhong, et al.
Published: (2026)
Latent Customer Segmentation and Value-Based Recommendation Leveraging a Two-Stage Model with Missing Labels
by: Gopalakrishnan, Keerthi, et al.
Published: (2026)
by: Gopalakrishnan, Keerthi, et al.
Published: (2026)
Personalized Product Search Ranking: A Multi-Task Learning Approach with Tabular and Non-Tabular Data
by: Morishetti, Lalitesh, et al.
Published: (2025)
by: Morishetti, Lalitesh, et al.
Published: (2025)
Embedding based retrieval for long tail search queries in ecommerce
by: Kekuda, Akshay, et al.
Published: (2025)
by: Kekuda, Akshay, et al.
Published: (2025)
Leveraging Foundation Models for Enhancing Robot Perception and Action
by: Mirjalili, Reihaneh
Published: (2025)
by: Mirjalili, Reihaneh
Published: (2025)
CausalSpatial: A Benchmark for Object-Centric Causal Spatial Reasoning
by: Ma, Wenxin, et al.
Published: (2026)
by: Ma, Wenxin, et al.
Published: (2026)
Dynamic Decision Making in Engineering System Design: A Deep Q-Learning Approach
by: Giahi, Ramin, et al.
Published: (2023)
by: Giahi, Ramin, et al.
Published: (2023)
Improving Sequential Recommender Systems with Online and In-store User Behavior
by: Ma, Luyi, et al.
Published: (2024)
by: Ma, Luyi, et al.
Published: (2024)
Twelve feminist lessons of war. By CynthiaEnloe. Oakland: University of California Press, 2023. 224 pages. $18.95 (paperback). ISBN: 978‐0520397675
by: Sinduja Raja
Published: (2024)
by: Sinduja Raja
Published: (2024)
PinpointQA: A Dataset and Benchmark for Small Object-Centric Spatial Understanding in Indoor Videos
by: Zhou, Zhiyu, et al.
Published: (2026)
by: Zhou, Zhiyu, et al.
Published: (2026)
Video Spatial Reasoning with Object-Centric 3D Rollout
by: Tang, Haoran, et al.
Published: (2025)
by: Tang, Haoran, et al.
Published: (2025)
Platelet Inventory Management with Approximate Dynamic Programming
by: Abouee-Mehrizi, Hossein, et al.
Published: (2023)
by: Abouee-Mehrizi, Hossein, et al.
Published: (2023)
A Combined Experimental and Numerical Study on Copper Oxide Thin Films: Oxygen Flow‐Driven Phase Tuning and Solar Cell Efficiency
by: Jagadish K. A., et al.
Published: (2025)
by: Jagadish K. A., et al.
Published: (2025)
Learning Object-Centric Spatial Reasoning for Sequential Manipulation in Cluttered Environments
by: Eze, Chrisantus, et al.
Published: (2026)
by: Eze, Chrisantus, et al.
Published: (2026)
GTA: Guided Transfer of Spatial Attention from Object-Centric Representations
by: Seo, SeokHyun, et al.
Published: (2024)
by: Seo, SeokHyun, et al.
Published: (2024)
Full Network Nonlocality Based Security In Quantum Key Distribution
by: Mukherjee, Kaushiki
Published: (2026)
by: Mukherjee, Kaushiki
Published: (2026)
Benchmarking Pathology Foundation Models for Spatial Domain Understanding
by: Zhao, Bokai, et al.
Published: (2026)
by: Zhao, Bokai, et al.
Published: (2026)
Comportamiento de especies arbóreas en dos arboretos del Instituto de Ciencia Animal
by: G. Achan
Published: (2011)
by: G. Achan
Published: (2011)
Bounding-Focused Discretization Methods for the Global Optimization of Nonconvex Semi-Infinite Programs
by: Turan, Evren M., et al.
Published: (2023)
by: Turan, Evren M., et al.
Published: (2023)
SpatialMath: Spatial Comprehension-Infused Symbolic Reasoning for Mathematical Problem-Solving
by: Bajpai, Ashutosh, et al.
Published: (2026)
by: Bajpai, Ashutosh, et al.
Published: (2026)
Similar Items
-
VL-CLIP: Enhancing Multimodal Recommendations via Visual Grounding and LLM-Augmented CLIP Embeddings
by: Giahi, Ramin, et al.
Published: (2025) -
CASE: Cadence-Aware Set Encoding for Large-Scale Next Basket Repurchase Recommendation
by: Cao, Yanan, et al.
Published: (2026) -
LayoutAgent: A Vision-Language Agent Guided Compositional Diffusion for Spatial Layout Planning
by: Fan, Zezhong, et al.
Published: (2025) -
CAL-RAG: Retrieval-Augmented Multi-Agent Generation for Content-Aware Layout Design
by: Forouzandehmehr, Najmeh, et al.
Published: (2025) -
Grocery to General Merchandise: A Cross-Pollination Recommender using LLMs and Real-Time Cart Context
by: Kekuda, Akshay, et al.
Published: (2025)