Analyzing the Shopping Journey: Computing Shelf Browsing Visits in a Physical Retail Store
Fuente:
arXiv
Guardado en:
| Autores principales: | Morales, Luis Yoichi, Zanlungo, Francesco, Woollard, David M. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
MARLIN: A Cloud Integrated Robotic Solution to Support Intralogistics in Retail
por: Mronga, Dennis, et al.
Publicado: (2024)
por: Mronga, Dennis, et al.
Publicado: (2024)
MindJourney: Test-Time Scaling with World Models for Spatial Reasoning
por: Yang, Yuncong, et al.
Publicado: (2025)
por: Yang, Yuncong, et al.
Publicado: (2025)
PRISM: A Multi-View Multi-Capability Retail Video Dataset for Embodied Vision-Language Models
por: Rouhi, Amirreza, et al.
Publicado: (2026)
por: Rouhi, Amirreza, et al.
Publicado: (2026)
Break Out the Silverware -- Semantic Understanding of Stored Household Items
por: Levi-Richter, Michaela, et al.
Publicado: (2025)
por: Levi-Richter, Michaela, et al.
Publicado: (2025)
MSC-Bench: Benchmarking and Analyzing Multi-Sensor Corruption for Driving Perception
por: Hao, Xiaoshuai, et al.
Publicado: (2025)
por: Hao, Xiaoshuai, et al.
Publicado: (2025)
VisBrowse-Bench: Benchmarking Visual-Native Search for Multimodal Browsing Agents
por: Zhang, Zhengbo, et al.
Publicado: (2026)
por: Zhang, Zhengbo, et al.
Publicado: (2026)
Enhancing Vision-Language Models with Scene Graphs for Traffic Accident Understanding
por: Lohner, Aaron, et al.
Publicado: (2024)
por: Lohner, Aaron, et al.
Publicado: (2024)
4th Workshop on Maritime Computer Vision (MaCVi): Challenge Overview
por: Kiefer, Benjamin, et al.
Publicado: (2026)
por: Kiefer, Benjamin, et al.
Publicado: (2026)
A Systematic Literature Review on Deep Learning-based Depth Estimation in Computer Vision
por: Rohan, Ali, et al.
Publicado: (2025)
por: Rohan, Ali, et al.
Publicado: (2025)
A Systematic Literature Review of Computer Vision Applications in Robotized Wire Harness Assembly
por: Wang, Hao, et al.
Publicado: (2023)
por: Wang, Hao, et al.
Publicado: (2023)
Overview of Computer Vision Techniques in Robotized Wire Harness Assembly: Current State and Future Opportunities
por: Wang, Hao, et al.
Publicado: (2023)
por: Wang, Hao, et al.
Publicado: (2023)
Robot Learning from a Physical World Model
por: Mao, Jiageng, et al.
Publicado: (2025)
por: Mao, Jiageng, et al.
Publicado: (2025)
Physically Grounded Vision-Language Models for Robotic Manipulation
por: Gao, Jensen, et al.
Publicado: (2023)
por: Gao, Jensen, et al.
Publicado: (2023)
Robotic Manipulation by Imitating Generated Videos Without Physical Demonstrations
por: Patel, Shivansh, et al.
Publicado: (2025)
por: Patel, Shivansh, et al.
Publicado: (2025)
GaussianProperty: Integrating Physical Properties to 3D Gaussians with LMMs
por: Xu, Xinli, et al.
Publicado: (2024)
por: Xu, Xinli, et al.
Publicado: (2024)
Digital Gene: Learning about the Physical World through Analytic Concepts
por: Sun, Jianhua, et al.
Publicado: (2025)
por: Sun, Jianhua, et al.
Publicado: (2025)
Goal Force: Teaching Video Models To Accomplish Physics-Conditioned Goals
por: Gillman, Nate, et al.
Publicado: (2026)
por: Gillman, Nate, et al.
Publicado: (2026)
ContactGaussian-WM: Learning Physics-Grounded World Model from Videos
por: Wang, Meizhong, et al.
Publicado: (2026)
por: Wang, Meizhong, et al.
Publicado: (2026)
DexGrasp Anything: Towards Universal Robotic Dexterous Grasping with Physics Awareness
por: Zhong, Yiming, et al.
Publicado: (2025)
por: Zhong, Yiming, et al.
Publicado: (2025)
PhysTwin: Physics-Informed Reconstruction and Simulation of Deformable Objects from Videos
por: Jiang, Hanxiao, et al.
Publicado: (2025)
por: Jiang, Hanxiao, et al.
Publicado: (2025)
A Physical Agentic Loop for Language-Guided Grasping with Execution-State Monitoring
por: Wang, Wenze, et al.
Publicado: (2026)
por: Wang, Wenze, et al.
Publicado: (2026)
PhysHanDI: Physics-Based Reconstruction of Hand-Deformable Object Interactions
por: Lee, Jihyun, et al.
Publicado: (2026)
por: Lee, Jihyun, et al.
Publicado: (2026)
Free-form language-based robotic reasoning and grasping
por: Jiao, Runyu, et al.
Publicado: (2025)
por: Jiao, Runyu, et al.
Publicado: (2025)
Obstruction reasoning for robotic grasping
por: Jiao, Runyu, et al.
Publicado: (2025)
por: Jiao, Runyu, et al.
Publicado: (2025)
MM-BrowseComp: A Comprehensive Benchmark for Multimodal Browsing Agents
por: Li, Shilong, et al.
Publicado: (2025)
por: Li, Shilong, et al.
Publicado: (2025)
Continuous Vision-Language-Action Co-Learning with Semantic-Physical Alignment for Behavioral Cloning
por: Qi, Xiuxiu, et al.
Publicado: (2025)
por: Qi, Xiuxiu, et al.
Publicado: (2025)
SIM1: Physics-Aligned Simulator as Zero-Shot Data Scaler in Deformable Worlds
por: Zhou, Yunsong, et al.
Publicado: (2026)
por: Zhou, Yunsong, et al.
Publicado: (2026)
PhyGile: Physics-Prefix Guided Motion Generation for Agile General Humanoid Motion Tracking
por: Bao, Jiacheng, et al.
Publicado: (2026)
por: Bao, Jiacheng, et al.
Publicado: (2026)
PhyBlock: A Progressive Benchmark for Physical Understanding and Planning via 3D Block Assembly
por: Ma, Liang, et al.
Publicado: (2025)
por: Ma, Liang, et al.
Publicado: (2025)
Multi-Modal World Model for Physical Robot Interactions: Simultaneous Visual and Tactile Predictions for Enhanced Accuracy
por: Mandil, Willow, et al.
Publicado: (2023)
por: Mandil, Willow, et al.
Publicado: (2023)
Phys4D: Fine-Grained Physics-Consistent 4D Modeling from Video Diffusion
por: Lu, Haoran, et al.
Publicado: (2026)
por: Lu, Haoran, et al.
Publicado: (2026)
PhysWorld: From Real Videos to World Models of Deformable Objects via Physics-Aware Demonstration Synthesis
por: Yang, Yu, et al.
Publicado: (2025)
por: Yang, Yu, et al.
Publicado: (2025)
DINO-VO: A Feature-based Visual Odometry Leveraging a Visual Foundation Model
por: Azhari, Maulana Bisyir, et al.
Publicado: (2025)
por: Azhari, Maulana Bisyir, et al.
Publicado: (2025)
FlowBot3D: Learning 3D Articulation Flow to Manipulate Articulated Objects
por: Eisner, Ben, et al.
Publicado: (2022)
por: Eisner, Ben, et al.
Publicado: (2022)
H2R-Grounder: A Paired-Data-Free Paradigm for Translating Human Interaction Videos into Physically Grounded Robot Videos
por: Ci, Hai, et al.
Publicado: (2025)
por: Ci, Hai, et al.
Publicado: (2025)
REALM: An RGB and Event Aligned Latent Manifold for Cross-Modal Perception
por: Polizzi, Vincenzo, et al.
Publicado: (2026)
por: Polizzi, Vincenzo, et al.
Publicado: (2026)
Hierarchical place recognition with omnidirectional images and curriculum learning-based loss functions
por: Alfaro, Marcos, et al.
Publicado: (2024)
por: Alfaro, Marcos, et al.
Publicado: (2024)
Evaluation of Large Language Models for Anomaly Detection in Autonomous Vehicles
por: Loukas, Petros, et al.
Publicado: (2025)
por: Loukas, Petros, et al.
Publicado: (2025)
Reg-NF: Efficient Registration of Implicit Surfaces within Neural Fields
por: Hausler, Stephen, et al.
Publicado: (2024)
por: Hausler, Stephen, et al.
Publicado: (2024)
Robusto-1 Dataset: Comparing Humans and VLMs on real out-of-distribution Autonomous Driving VQA from Peru
por: Cusipuma, Dunant, et al.
Publicado: (2025)
por: Cusipuma, Dunant, et al.
Publicado: (2025)
Ejemplares similares
-
MARLIN: A Cloud Integrated Robotic Solution to Support Intralogistics in Retail
por: Mronga, Dennis, et al.
Publicado: (2024) -
MindJourney: Test-Time Scaling with World Models for Spatial Reasoning
por: Yang, Yuncong, et al.
Publicado: (2025) -
PRISM: A Multi-View Multi-Capability Retail Video Dataset for Embodied Vision-Language Models
por: Rouhi, Amirreza, et al.
Publicado: (2026) -
Break Out the Silverware -- Semantic Understanding of Stored Household Items
por: Levi-Richter, Michaela, et al.
Publicado: (2025) -
MSC-Bench: Benchmarking and Analyzing Multi-Sensor Corruption for Driving Perception
por: Hao, Xiaoshuai, et al.
Publicado: (2025)