AdaGlimpse: Active Visual Exploration with Arbitrary Glimpse Position and Scale
Fuente:
arXiv
Saved in:
| Main Authors: | Pardyl, Adam, Wronka, Michał, Wołczyk, Maciej, Adamczewski, Kamil, Trzciński, Tomasz, Zieliński, Bartosz |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Beyond Grids: Exploring Elastic Input Sampling for Vision Transformers
by: Pardyl, Adam, et al.
Published: (2023)
by: Pardyl, Adam, et al.
Published: (2023)
FlySearch: Exploring how vision-language models explore
by: Pardyl, Adam, et al.
Published: (2025)
by: Pardyl, Adam, et al.
Published: (2025)
SpaRRTa: A Synthetic Benchmark for Evaluating Spatial Intelligence in Visual Foundation Models
by: Kargin, Turhan Can, et al.
Published: (2026)
by: Kargin, Turhan Can, et al.
Published: (2026)
A Glimpse to Compress: Dynamic Visual Token Pruning for Large Vision-Language Models
by: Zeng, Quan-Sheng, et al.
Published: (2025)
by: Zeng, Quan-Sheng, et al.
Published: (2025)
Mind the GAP: Glimpse-based Active Perception improves generalization and sample efficiency of visual reasoning
by: Kolner, Oleh, et al.
Published: (2024)
by: Kolner, Oleh, et al.
Published: (2024)
Seeing Through Their Eyes: Evaluating Visual Perspective Taking in Vision Language Models
by: Góral, Gracjan, et al.
Published: (2024)
by: Góral, Gracjan, et al.
Published: (2024)
Beyond Recognition: Evaluating Visual Perspective Taking in Vision Language Models
by: Góral, Gracjan, et al.
Published: (2025)
by: Góral, Gracjan, et al.
Published: (2025)
Chain-of-Glimpse: Search-Guided Progressive Object-Grounded Reasoning for Video Understanding
by: Wu, Zhixuan, et al.
Published: (2026)
by: Wu, Zhixuan, et al.
Published: (2026)
Unifying Deep Stochastic Processes for Image Enhancement
by: Kozłowski, Wojciech, et al.
Published: (2026)
by: Kozłowski, Wojciech, et al.
Published: (2026)
Sparrow: Text-Anchored Window Attention with Visual-Semantic Glimpsing for Speculative Decoding in Video LLMs
by: Zhang, Libo, et al.
Published: (2026)
by: Zhang, Libo, et al.
Published: (2026)
Glimpse: Generalized Locality for Scalable and Robust CT
by: Khorashadizadeh, AmirEhsan, et al.
Published: (2024)
by: Khorashadizadeh, AmirEhsan, et al.
Published: (2024)
Divide and not forget: Ensemble of selectively trained experts in Continual Learning
by: Rypeść, Grzegorz, et al.
Published: (2024)
by: Rypeść, Grzegorz, et al.
Published: (2024)
Action Anticipation at a Glimpse: To What Extent Can Multimodal Cues Replace Video?
by: Benavent-Lledo, Manuel, et al.
Published: (2025)
by: Benavent-Lledo, Manuel, et al.
Published: (2025)
GLIMPSE: Do Large Vision-Language Models Truly Think With Videos or Just Glimpse at Them?
by: Zhou, Yiyang, et al.
Published: (2025)
by: Zhou, Yiyang, et al.
Published: (2025)
Guess The Unseen: Dynamic 3D Scene Reconstruction from Partial 2D Glimpses
by: Lee, Inhee, et al.
Published: (2024)
by: Lee, Inhee, et al.
Published: (2024)
TORE: Token Recycling in Vision Transformers for Efficient Active Visual Exploration
by: Olszewski, Jan, et al.
Published: (2023)
by: Olszewski, Jan, et al.
Published: (2023)
Unmasking the Uniqueness: A Glimpse into Age-Invariant Face Recognition of Indigenous African Faces
by: Ajewole, Fakunle, et al.
Published: (2024)
by: Ajewole, Fakunle, et al.
Published: (2024)
Novel View Synthesis from A Few Glimpses via Test-Time Natural Video Completion
by: Xu, Yan, et al.
Published: (2025)
by: Xu, Yan, et al.
Published: (2025)
Conditioned Activation Transport for T2I Safety Steering
by: Chrabąszcz, Maciej, et al.
Published: (2026)
by: Chrabąszcz, Maciej, et al.
Published: (2026)
Neural Bloom: A Deep Learning Approach to Real-Time Lighting
by: Karp, Rafal, et al.
Published: (2025)
by: Karp, Rafal, et al.
Published: (2025)
Modeling 3D Surface Manifolds with a Locally Conditioned Atlas
by: Spurek, Przemysław, et al.
Published: (2021)
by: Spurek, Przemysław, et al.
Published: (2021)
Shapley Pruning for Neural Network Compression
by: Adamczewski, Kamil, et al.
Published: (2024)
by: Adamczewski, Kamil, et al.
Published: (2024)
AI-Driven Rapid Identification of Bacterial and Fungal Pathogens in Blood Smears of Septic Patients
by: Sroka-Oleksiak, Agnieszka, et al.
Published: (2025)
by: Sroka-Oleksiak, Agnieszka, et al.
Published: (2025)
ExpertSim: Fast Particle Detector Simulation Using Mixture-of-Generative-Experts
by: Będkowski, Patryk, et al.
Published: (2025)
by: Będkowski, Patryk, et al.
Published: (2025)
Points2NeRF: Generating Neural Radiance Fields from 3D point cloud
by: Zimny, Dominik, et al.
Published: (2022)
by: Zimny, Dominik, et al.
Published: (2022)
LumiGauss: Relightable Gaussian Splatting in the Wild
by: Kaleta, Joanna, et al.
Published: (2024)
by: Kaleta, Joanna, et al.
Published: (2024)
AdaVLN: Towards Visual Language Navigation in Continuous Indoor Environments with Moving Humans
by: Loh, Dillon, et al.
Published: (2024)
by: Loh, Dillon, et al.
Published: (2024)
Self-Supervised Event Representations: Towards Accurate, Real-Time Perception on SoC FPGAs
by: Jeziorek, Kamil, et al.
Published: (2025)
by: Jeziorek, Kamil, et al.
Published: (2025)
Precise Parameter Localization for Textual Generation in Diffusion Models
by: Staniszewski, Łukasz, et al.
Published: (2025)
by: Staniszewski, Łukasz, et al.
Published: (2025)
Revisiting Supervision for Continual Representation Learning
by: Marczak, Daniel, et al.
Published: (2023)
by: Marczak, Daniel, et al.
Published: (2023)
Task-recency bias strikes back: Adapting covariances in Exemplar-Free Class Incremental Learning
by: Rypeść, Grzegorz, et al.
Published: (2024)
by: Rypeść, Grzegorz, et al.
Published: (2024)
Realistic Evaluation of Test-Time Adaptation Algorithms: Unsupervised Hyperparameter Selection
by: Cygert, Sebastian, et al.
Published: (2024)
by: Cygert, Sebastian, et al.
Published: (2024)
Agentic Discovery with Active Hypothesis Exploration for Visual Recognition
by: Koo, Jaywon, et al.
Published: (2026)
by: Koo, Jaywon, et al.
Published: (2026)
AdaViPro: Region-based Adaptive Visual Prompt for Large-Scale Models Adapting
by: Yang, Mengyu, et al.
Published: (2024)
by: Yang, Mengyu, et al.
Published: (2024)
Parameter-Efficient Interventions for Enhanced Model Merging
by: Osial, Marcin, et al.
Published: (2024)
by: Osial, Marcin, et al.
Published: (2024)
Task-driven real-world super-resolution of document scans
by: Zyrek, Maciej, et al.
Published: (2025)
by: Zyrek, Maciej, et al.
Published: (2025)
AdaSCALE: Adaptive Scaling for OOD Detection
by: Regmi, Sudarshan
Published: (2025)
by: Regmi, Sudarshan
Published: (2025)
CLIP-DINOiser: Teaching CLIP a few DINO tricks for open-vocabulary semantic segmentation
by: Wysoczańska, Monika, et al.
Published: (2023)
by: Wysoczańska, Monika, et al.
Published: (2023)
LumiMotion: Improving Gaussian Relighting with Scene Dynamics
by: Kaleta, Joanna, et al.
Published: (2026)
by: Kaleta, Joanna, et al.
Published: (2026)
Learning from Noise: Enhancing DNNs for Event-Based Vision through Controlled Noise Injection
by: Kowalczyk, Marcin, et al.
Published: (2025)
by: Kowalczyk, Marcin, et al.
Published: (2025)
Similar Items
-
Beyond Grids: Exploring Elastic Input Sampling for Vision Transformers
by: Pardyl, Adam, et al.
Published: (2023) -
FlySearch: Exploring how vision-language models explore
by: Pardyl, Adam, et al.
Published: (2025) -
SpaRRTa: A Synthetic Benchmark for Evaluating Spatial Intelligence in Visual Foundation Models
by: Kargin, Turhan Can, et al.
Published: (2026) -
A Glimpse to Compress: Dynamic Visual Token Pruning for Large Vision-Language Models
by: Zeng, Quan-Sheng, et al.
Published: (2025) -
Mind the GAP: Glimpse-based Active Perception improves generalization and sample efficiency of visual reasoning
by: Kolner, Oleh, et al.
Published: (2024)