OpenGVL -- Benchmarking Visual Temporal Progress for Data Curation
Fuente:
arXiv
Saved in:
| Main Authors: | Budzianowski, Paweł, Wiśnios, Emilia, Tyrolski, Michał, Góral, Gracjan, Kulakov, Igor, Petrenko, Viktor, Walas, Krzysztof |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Wait, that's not an option: LLMs Robustness with Incorrect Multiple-Choice Options
by: Góral, Gracjan, et al.
Published: (2024)
by: Góral, Gracjan, et al.
Published: (2024)
What Matters in Hierarchical Search for Combinatorial Reasoning Problems?
by: Zawalski, Michał, et al.
Published: (2024)
by: Zawalski, Michał, et al.
Published: (2024)
Seeing Through Their Eyes: Evaluating Visual Perspective Taking in Vision Language Models
by: Góral, Gracjan, et al.
Published: (2024)
by: Góral, Gracjan, et al.
Published: (2024)
EdgeVLA: Efficient Vision-Language-Action Models
by: Budzianowski, Paweł, et al.
Published: (2025)
by: Budzianowski, Paweł, et al.
Published: (2025)
Learning dynamics models for velocity estimation in autonomous racing
by: Węgrzynowski, Jan, et al.
Published: (2024)
by: Węgrzynowski, Jan, et al.
Published: (2024)
On learning racing policies with reinforcement learning
by: Czechmanowski, Grzegorz, et al.
Published: (2025)
by: Czechmanowski, Grzegorz, et al.
Published: (2025)
Beyond Constant Parameters: Hyper Prediction Models and HyperMPC
by: Węgrzynowski, Jan, et al.
Published: (2025)
by: Węgrzynowski, Jan, et al.
Published: (2025)
Evaluation of an Actuated Spine in Agile Quadruped Locomotion
by: Bohlinger, Nico, et al.
Published: (2026)
by: Bohlinger, Nico, et al.
Published: (2026)
Semantic-Drive: Democratizing Long-Tail Data Curation via Open-Vocabulary Grounding and Neuro-Symbolic VLM Consensus
by: Guillen-Perez, Antonio
Published: (2025)
by: Guillen-Perez, Antonio
Published: (2025)
Beyond Recognition: Evaluating Visual Perspective Taking in Vision Language Models
by: Góral, Gracjan, et al.
Published: (2025)
by: Góral, Gracjan, et al.
Published: (2025)
Behind Closed Words: Creating and Investigating the forePLay Annotated Dataset for Polish Erotic Discourse
by: Kołos, Anna, et al.
Published: (2024)
by: Kołos, Anna, et al.
Published: (2024)
Depth-Wise Activation Steering for Honest Language Models
by: Góral, Gracjan, et al.
Published: (2025)
by: Góral, Gracjan, et al.
Published: (2025)
Bridging the gap between Learning-to-plan, Motion Primitives and Safe Reinforcement Learning
by: Kicki, Piotr, et al.
Published: (2024)
by: Kicki, Piotr, et al.
Published: (2024)
BAN-PL: a Novel Polish Dataset of Banned Harmful and Offensive Content from Wykop.pl web service
by: Kołos, Anna, et al.
Published: (2023)
by: Kołos, Anna, et al.
Published: (2023)
Pheme: Efficient and Conversational Speech Generation
by: Budzianowski, Paweł, et al.
Published: (2024)
by: Budzianowski, Paweł, et al.
Published: (2024)
The Design of Informative Take-Over Requests for Semi-Autonomous Cyber-Physical Systems: Combining Spoken Language and Visual Icons in a Drone-Controller Setting
by: Gundappa, Ashwini, et al.
Published: (2024)
by: Gundappa, Ashwini, et al.
Published: (2024)
HyperGVL: Benchmarking and Improving Large Vision-Language Models in Hypergraph Understanding and Reasoning
by: Wei, Yanbin, et al.
Published: (2026)
by: Wei, Yanbin, et al.
Published: (2026)
OpenFMNav: Towards Open-Set Zero-Shot Object Navigation via Vision-Language Foundation Models
by: Kuang, Yuxuan, et al.
Published: (2024)
by: Kuang, Yuxuan, et al.
Published: (2024)
One Policy to Run Them All: an End-to-end Learning Approach to Multi-Embodiment Locomotion
by: Bohlinger, Nico, et al.
Published: (2024)
by: Bohlinger, Nico, et al.
Published: (2024)
Open-Loop Planning, Closed-Loop Verification: Speculative Verification for VLA
by: Wang, Zihua, et al.
Published: (2026)
by: Wang, Zihua, et al.
Published: (2026)
Benchmarking Local Language Models for Social Robots using Edge Devices
by: Lamouille, Dorian, et al.
Published: (2026)
by: Lamouille, Dorian, et al.
Published: (2026)
Robust Localization, Mapping, and Navigation for Quadruped Robots
by: Aditya, Dyuman, et al.
Published: (2025)
by: Aditya, Dyuman, et al.
Published: (2025)
Robot Data Curation with Mutual Information Estimators
by: Hejna, Joey, et al.
Published: (2025)
by: Hejna, Joey, et al.
Published: (2025)
Visual Cues Enhance Predictive Turn-Taking for Two-Party Human Interaction
by: Russell, Sam O'Connor, et al.
Published: (2025)
by: Russell, Sam O'Connor, et al.
Published: (2025)
PSALM-V: Automating Symbolic Planning in Interactive Visual Environments with Large Language Models
by: Zhu, Wang Bill, et al.
Published: (2025)
by: Zhu, Wang Bill, et al.
Published: (2025)
P-RAG: Progressive Retrieval Augmented Generation For Planning on Embodied Everyday Task
by: Xu, Weiye, et al.
Published: (2024)
by: Xu, Weiye, et al.
Published: (2024)
Polaris: Open-ended Interactive Robotic Manipulation via Syn2Real Visual Grounding and Large Language Models
by: Wang, Tianyu, et al.
Published: (2024)
by: Wang, Tianyu, et al.
Published: (2024)
VLURes: Benchmarking VLM Visual and Linguistic Understanding in Low-Resource Languages
by: Atuhurra, Jesse, et al.
Published: (2025)
by: Atuhurra, Jesse, et al.
Published: (2025)
Scene-Graph ViT: End-to-End Open-Vocabulary Visual Relationship Detection
by: Salzmann, Tim, et al.
Published: (2024)
by: Salzmann, Tim, et al.
Published: (2024)
Robotouille: An Asynchronous Planning Benchmark for LLM Agents
by: Gonzalez-Pumariega, Gonzalo, et al.
Published: (2025)
by: Gonzalez-Pumariega, Gonzalo, et al.
Published: (2025)
Planning with Linear Temporal Logic Specifications: Handling Quantifiable and Unquantifiable Uncertainty
by: Yu, Pian, et al.
Published: (2025)
by: Yu, Pian, et al.
Published: (2025)
Neuro-Symbolic Generation of Explanations for Robot Policies with Weighted Signal Temporal Logic
by: Yuasa, Mikihisa, et al.
Published: (2025)
by: Yuasa, Mikihisa, et al.
Published: (2025)
Compiling OpenSCENARIO 2.1 for Scenario-Based Testing in CARLA
by: Gamage, Thoshitha, et al.
Published: (2026)
by: Gamage, Thoshitha, et al.
Published: (2026)
TimeSpot: Benchmarking Geo-Temporal Understanding in Vision-Language Models in Real-World Settings
by: Wasi, Azmine Toushik, et al.
Published: (2026)
by: Wasi, Azmine Toushik, et al.
Published: (2026)
RoboTwin: Dual-Arm Robot Benchmark with Generative Digital Twins
by: Mu, Yao, et al.
Published: (2025)
by: Mu, Yao, et al.
Published: (2025)
InCoRo: In-Context Learning for Robotics Control with Feedback Loops
by: Zhu, Jiaqiang Ye, et al.
Published: (2024)
by: Zhu, Jiaqiang Ye, et al.
Published: (2024)
LTL-D*: Incrementally Optimal Replanning for Feasible and Infeasible Tasks in Linear Temporal Logic Specifications
by: Ren, Jiming, et al.
Published: (2024)
by: Ren, Jiming, et al.
Published: (2024)
ReMEmbR: Building and Reasoning Over Long-Horizon Spatio-Temporal Memory for Robot Navigation
by: Anwar, Abrar, et al.
Published: (2024)
by: Anwar, Abrar, et al.
Published: (2024)
An Iterative Approach for Heterogeneous Multi-Agent Route Planning with Resource Transportation Uncertainty and Temporal Logic Goals
by: Cardona, Gustavo A., et al.
Published: (2025)
by: Cardona, Gustavo A., et al.
Published: (2025)
When Robots Say No: Temporal Trust Recovery Through Explanation
by: Webb, Nicola, et al.
Published: (2025)
by: Webb, Nicola, et al.
Published: (2025)
Similar Items
-
Wait, that's not an option: LLMs Robustness with Incorrect Multiple-Choice Options
by: Góral, Gracjan, et al.
Published: (2024) -
What Matters in Hierarchical Search for Combinatorial Reasoning Problems?
by: Zawalski, Michał, et al.
Published: (2024) -
Seeing Through Their Eyes: Evaluating Visual Perspective Taking in Vision Language Models
by: Góral, Gracjan, et al.
Published: (2024) -
EdgeVLA: Efficient Vision-Language-Action Models
by: Budzianowski, Paweł, et al.
Published: (2025) -
Learning dynamics models for velocity estimation in autonomous racing
by: Węgrzynowski, Jan, et al.
Published: (2024)