Salvato in:
| Autori principali: | Fuller, Anthony, Yassin, Yousef, Wen, Junfeng, Kyrollos, Daniel G., Ibrahim, Tarek, Green, James R., Shelhamer, Evan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2505.18051 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
LookWhen? Fast Video Recognition by Learning When, Where, and What to Compute
di: Salamatian, Ali, et al.
Pubblicazione: (2026)
di: Salamatian, Ali, et al.
Pubblicazione: (2026)
LookHere: Vision Transformers with Directed Attention Generalize and Extrapolate
di: Fuller, Anthony, et al.
Pubblicazione: (2024)
di: Fuller, Anthony, et al.
Pubblicazione: (2024)
Thicker and Quicker: A Jumbo Token for Fast Plain Vision Transformers
di: Fuller, Anthony, et al.
Pubblicazione: (2025)
di: Fuller, Anthony, et al.
Pubblicazione: (2025)
Self-Soupervision: Cooking Model Soups without Labels
di: Fuller, Anthony, et al.
Pubblicazione: (2026)
di: Fuller, Anthony, et al.
Pubblicazione: (2026)
LookSharp: Attention Entropy Minimization for Test-Time Adaptation
di: Mali, Yash, et al.
Pubblicazione: (2025)
di: Mali, Yash, et al.
Pubblicazione: (2025)
Self-Distillation of Hidden Layers for Self-Supervised Representation Learning
di: Lowe, Scott C., et al.
Pubblicazione: (2026)
di: Lowe, Scott C., et al.
Pubblicazione: (2026)
A Closer Look at In-Distribution vs. Out-of-Distribution Accuracy for Open-Set Test-time Adaptation
di: Li, Zefeng, et al.
Pubblicazione: (2026)
di: Li, Zefeng, et al.
Pubblicazione: (2026)
Pay Attention to Where You Looked
di: Berian, Alex, et al.
Pubblicazione: (2026)
di: Berian, Alex, et al.
Pubblicazione: (2026)
GridPrune: From "Where to Look" to "What to Select" in Visual Token Pruning for MLLMs
di: Duan, Yuxiang, et al.
Pubblicazione: (2025)
di: Duan, Yuxiang, et al.
Pubblicazione: (2025)
Learning Where to Look: Self-supervised Viewpoint Selection for Active Localization using Geometrical Information
di: Di Giammarino, Luca, et al.
Pubblicazione: (2024)
di: Di Giammarino, Luca, et al.
Pubblicazione: (2024)
What ZTF Saw Where Rubin Looked: Anomaly Hunting in DR23
di: Pruzhinskaya, Maria V., et al.
Pubblicazione: (2025)
di: Pruzhinskaya, Maria V., et al.
Pubblicazione: (2025)
LookWise: Knowing When and Where to Look for Fine-Grained Visual Reasoning in Multimodal Large Language Models
di: Shen, Yuxiang, et al.
Pubblicazione: (2026)
di: Shen, Yuxiang, et al.
Pubblicazione: (2026)
Look Where It Matters: High-Resolution Crops Retrieval for Efficient VLMs
di: Shabtay, Nimrod, et al.
Pubblicazione: (2026)
di: Shabtay, Nimrod, et al.
Pubblicazione: (2026)
Why We Look Where We Look: Emergent Human-like Fixations of a Foveated Visual Language Model Maximizing Scene Understanding
di: Murlidaran, Shravan, et al.
Pubblicazione: (2026)
di: Murlidaran, Shravan, et al.
Pubblicazione: (2026)
Canadian Libraries and Librarianship: Where to Start Looking for the Information.
di: Anderson, Beryl L.
Pubblicazione: (1982)
di: Anderson, Beryl L.
Pubblicazione: (1982)
Where do Large Vision-Language Models Look at when Answering Questions?
di: Xing, Xiaoying, et al.
Pubblicazione: (2025)
di: Xing, Xiaoying, et al.
Pubblicazione: (2025)
Where to Look Next: Learning Viewpoint Recommendations for Informative Trajectory Planning
di: Lodel, Max, et al.
Pubblicazione: (2022)
di: Lodel, Max, et al.
Pubblicazione: (2022)
Beyond Where to Look: Trajectory-Guided Reinforcement Learning for Multimodal RLVR
di: Lu, Jinda, et al.
Pubblicazione: (2026)
di: Lu, Jinda, et al.
Pubblicazione: (2026)
VANP: Learning Where to See for Navigation with Self-Supervised Vision-Action Pre-Training
di: Nazeri, Mohammad, et al.
Pubblicazione: (2024)
di: Nazeri, Mohammad, et al.
Pubblicazione: (2024)
The CD-ROM Journal Literature: Where Do You Look?
di: Schwartz, Candy
Pubblicazione: (1992)
di: Schwartz, Candy
Pubblicazione: (1992)
What They Saw, Not Just Where They Looked: Semantic Scanpath Similarity via VLMs and NLP metric
di: Kerkouri, Mohamed Amine, et al.
Pubblicazione: (2026)
di: Kerkouri, Mohamed Amine, et al.
Pubblicazione: (2026)
Learning Where to Look: UCB-Driven Controlled Sensing for Quickest Change Detection
di: Huang, Yu-Han, et al.
Pubblicazione: (2026)
di: Huang, Yu-Han, et al.
Pubblicazione: (2026)
MLLMs Know Where to Look: Training-free Perception of Small Visual Details with Multimodal LLMs
di: Zhang, Jiarui, et al.
Pubblicazione: (2025)
di: Zhang, Jiarui, et al.
Pubblicazione: (2025)
IE-002 Corrected: The Projection Asymmetry Principle - What You Can Read Depends on Where You Stand When You Look
di: Dunn, James E.
Pubblicazione: (2026)
di: Dunn, James E.
Pubblicazione: (2026)
IE-002 Corrected: The Projection Asymmetry Principle - What You Can Read Depends on Where You Stand When You Look
di: Dunn, James E.
Pubblicazione: (2026)
di: Dunn, James E.
Pubblicazione: (2026)
Galileo: Learning Global & Local Features of Many Remote Sensing Modalities
di: Tseng, Gabriel, et al.
Pubblicazione: (2025)
di: Tseng, Gabriel, et al.
Pubblicazione: (2025)
Decoupling What to Count and Where to See for Referring Expression Counting
di: Zou, Yuda, et al.
Pubblicazione: (2025)
di: Zou, Yuda, et al.
Pubblicazione: (2025)
MAGE: All-[MASK] Block Already Knows Where to Look in Diffusion LLM
di: Kwon, Omin, et al.
Pubblicazione: (2026)
di: Kwon, Omin, et al.
Pubblicazione: (2026)
ARTrackV2: Prompting Autoregressive Tracker Where to Look and How to Describe
di: Bai, Yifan, et al.
Pubblicazione: (2023)
di: Bai, Yifan, et al.
Pubblicazione: (2023)
Tell Model Where to Look: Mitigating Hallucinations in MLLMs by Vision-Guided Attention
di: Zhao, Jianfei, et al.
Pubblicazione: (2025)
di: Zhao, Jianfei, et al.
Pubblicazione: (2025)
Where to Look It Up. Saturday Review's Semi-Annual Reference Book Roundup
di: Glixon, David M.
Pubblicazione: (1969)
di: Glixon, David M.
Pubblicazione: (1969)
Learning to Read Where to Look: Disease-Aware Vision-Language Pretraining for 3D CT
di: Ging, Simon, et al.
Pubblicazione: (2026)
di: Ging, Simon, et al.
Pubblicazione: (2026)
Overdue: I'll Never Forget Where I Looked for Whatsisname in the Card Catalog
di: Purucker, Mary I., et al.
Pubblicazione: (1972)
di: Purucker, Mary I., et al.
Pubblicazione: (1972)
iCub Knows Where You Look: Exploiting Social Cues for Interactive Object Detection Learning
di: Lombardi, Maria, et al.
Pubblicazione: (2022)
di: Lombardi, Maria, et al.
Pubblicazione: (2022)
Components of a Quality School Library Program: What Do They Look Like? How Do We Assess Where We Are?
di: Brown, Gerald R.
Pubblicazione: (2003)
di: Brown, Gerald R.
Pubblicazione: (2003)
Where Did It All Go Wrong? A Hierarchical Look into Multi-Agent Error Attribution
di: Banerjee, Adi, et al.
Pubblicazione: (2025)
di: Banerjee, Adi, et al.
Pubblicazione: (2025)
Where to Look: Can Foundation Models Reach a Target Viewpoint Through Active Exploration?
di: Li, Liyang, et al.
Pubblicazione: (2026)
di: Li, Liyang, et al.
Pubblicazione: (2026)
Look over there. Where? A compositional approach to the modeling of public opinion on the most important problem
di: Steven Jokinsky, et al.
Pubblicazione: (2024)
di: Steven Jokinsky, et al.
Pubblicazione: (2024)
Learning to Perceive "Where": Spatial Pretext Tasks for Robust Self-Supervised Learning
di: Shen, Yang, et al.
Pubblicazione: (2026)
di: Shen, Yang, et al.
Pubblicazione: (2026)
What the Consumer Looks For.
di: Ratcliffe, F. W.
Pubblicazione: (1987)
di: Ratcliffe, F. W.
Pubblicazione: (1987)
Documenti analoghi
-
LookWhen? Fast Video Recognition by Learning When, Where, and What to Compute
di: Salamatian, Ali, et al.
Pubblicazione: (2026) -
LookHere: Vision Transformers with Directed Attention Generalize and Extrapolate
di: Fuller, Anthony, et al.
Pubblicazione: (2024) -
Thicker and Quicker: A Jumbo Token for Fast Plain Vision Transformers
di: Fuller, Anthony, et al.
Pubblicazione: (2025) -
Self-Soupervision: Cooking Model Soups without Labels
di: Fuller, Anthony, et al.
Pubblicazione: (2026) -
LookSharp: Attention Entropy Minimization for Test-Time Adaptation
di: Mali, Yash, et al.
Pubblicazione: (2025)