GRAZE: Grounded Refinement and Motion-Aware Zero-Shot Event Localization
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Zaidi, Syed Ahsan Masud, Shamir, Lior, Hsu, William, Dietrich, Scott, Zaidi, Talha |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
ViTs for Action Classification in Videos: An Approach to Risky Tackle Detection in American Football Practice Videos
par: Zaidi, Syed Ahsan Masud, et autres
Publié: (2026)
par: Zaidi, Syed Ahsan Masud, et autres
Publié: (2026)
CornViT: A Multi-Stage Convolutional Vision Transformer Framework for Hierarchical Corn Kernel Analysis
par: Erukude, Sai Teja, et autres
Publié: (2025)
par: Erukude, Sai Teja, et autres
Publié: (2025)
Improving Adversarial Robustness Through Adaptive Learning-Driven Multi-Teacher Knowledge Distillation
par: Ullah, Hayat, et autres
Publié: (2025)
par: Ullah, Hayat, et autres
Publié: (2025)
DynaPURLS: Dynamic Refinement of Part-Aware Representations for Skeleton-Based Zero-Shot Action Recognition
par: Zhu, Jingmin, et autres
Publié: (2025)
par: Zhu, Jingmin, et autres
Publié: (2025)
Next-Generation License Plate Detection and Recognition System using YOLOv8
par: Amin, Arslan, et autres
Publié: (2025)
par: Amin, Arslan, et autres
Publié: (2025)
Semantic Segmentation Refiner for Ultrasound Applications with Zero-Shot Foundation Models
par: Indelman, Hedda Cohen, et autres
Publié: (2024)
par: Indelman, Hedda Cohen, et autres
Publié: (2024)
Identifying Bias in Deep Neural Networks Using Image Transforms
par: Erukude, Sai Teja, et autres
Publié: (2024)
par: Erukude, Sai Teja, et autres
Publié: (2024)
MotionShop: Zero-Shot Motion Transfer in Video Diffusion Models with Mixture of Score Guidance
par: Yesiltepe, Hidir, et autres
Publié: (2024)
par: Yesiltepe, Hidir, et autres
Publié: (2024)
Fruit Classification System with Deep Learning and Neural Architecture Search
par: Dewi, Christine, et autres
Publié: (2024)
par: Dewi, Christine, et autres
Publié: (2024)
RARE: Refine Any Registration of Pairwise Point Clouds via Zero-Shot Learning
par: Zheng, Chengyu, et autres
Publié: (2025)
par: Zheng, Chengyu, et autres
Publié: (2025)
SPATIOROUTE: Dynamic Prompt Routing for Zero-Shot Spatial Reasoning
par: Chunhachatrachai, Pawat, et autres
Publié: (2026)
par: Chunhachatrachai, Pawat, et autres
Publié: (2026)
Zero-Shot Monocular Motion Segmentation in the Wild by Combining Deep Learning with Geometric Motion Model Fusion
par: Huang, Yuxiang, et autres
Publié: (2024)
par: Huang, Yuxiang, et autres
Publié: (2024)
VTG-GPT: Tuning-Free Zero-Shot Video Temporal Grounding with GPT
par: Xu, Yifang, et autres
Publié: (2024)
par: Xu, Yifang, et autres
Publié: (2024)
Context-Aware Pseudo-Label Scoring for Zero-Shot Video Summarization
par: Wu, Yuanli, et autres
Publié: (2025)
par: Wu, Yuanli, et autres
Publié: (2025)
Zero-Shot Industrial Anomaly Segmentation with Image-Aware Prompt Generation
par: Park, SoYoung, et autres
Publié: (2025)
par: Park, SoYoung, et autres
Publié: (2025)
SalientFusion: Context-Aware Compositional Zero-Shot Food Recognition
par: Song, Jiajun, et autres
Publié: (2025)
par: Song, Jiajun, et autres
Publié: (2025)
Zero-Shot Temporal Interaction Localization for Egocentric Videos
par: Zhang, Erhang, et autres
Publié: (2025)
par: Zhang, Erhang, et autres
Publié: (2025)
Zero-shot Vision-Language Reranking for Cross-View Geolocalization
par: Erzurumlu, Yunus Talha, et autres
Publié: (2026)
par: Erzurumlu, Yunus Talha, et autres
Publié: (2026)
Zero-TIG: Temporal Consistency-Aware Zero-Shot Illumination-Guided Low-light Video Enhancement
par: Li, Yini, et autres
Publié: (2025)
par: Li, Yini, et autres
Publié: (2025)
Generating Adversarial Events: A Motion-Aware Point Cloud Framework
par: Ren, Hongwei, et autres
Publié: (2026)
par: Ren, Hongwei, et autres
Publié: (2026)
DreamMotion: Space-Time Self-Similar Score Distillation for Zero-Shot Video Editing
par: Jeong, Hyeonho, et autres
Publié: (2024)
par: Jeong, Hyeonho, et autres
Publié: (2024)
ZSPAPrune: Zero-Shot Prompt-Aware Token Pruning for Vision-Language Models
par: Zhang, Pu, et autres
Publié: (2025)
par: Zhang, Pu, et autres
Publié: (2025)
Context-Aware Weakly Supervised Image Manipulation Localization with SAM Refinement
par: Wang, Xinghao, et autres
Publié: (2025)
par: Wang, Xinghao, et autres
Publié: (2025)
RSVG-ZeroOV: Exploring a Training-Free Framework for Zero-Shot Open-Vocabulary Visual Grounding in Remote Sensing Images
par: Li, Ke, et autres
Publié: (2025)
par: Li, Ke, et autres
Publié: (2025)
MotionCraft: Physics-based Zero-Shot Video Generation
par: Aira, Luca Savant, et autres
Publié: (2024)
par: Aira, Luca Savant, et autres
Publié: (2024)
Transductive Zero-Shot and Few-Shot CLIP
par: Martin, Ségolène, et autres
Publié: (2024)
par: Martin, Ségolène, et autres
Publié: (2024)
SAModified: A Foundation Model-Based Zero-Shot Approach for Refining Noisy Land-Use Land-Cover Maps
par: Pekhale, Sparsh, et autres
Publié: (2024)
par: Pekhale, Sparsh, et autres
Publié: (2024)
Can Large Language Models Grasp Event Signals? Exploring Pure Zero-Shot Event-based Recognition
par: Yu, Zongyou, et autres
Publié: (2024)
par: Yu, Zongyou, et autres
Publié: (2024)
A Guideline-Aware AI Agent for Zero-Shot Target Volume Auto-Delineation
par: Kim, Yoon Jo, et autres
Publié: (2026)
par: Kim, Yoon Jo, et autres
Publié: (2026)
Refining Pre-Trained Motion Models
par: Sun, Xinglong, et autres
Publié: (2024)
par: Sun, Xinglong, et autres
Publié: (2024)
Think, Act, Build: An Agentic Framework with Vision Language Models for Zero-Shot 3D Visual Grounding
par: Wang, Haibo, et autres
Publié: (2026)
par: Wang, Haibo, et autres
Publié: (2026)
Just Zoom In: Cross-View Geo-Localization via Autoregressive Zooming
par: Erzurumlu, Yunus Talha, et autres
Publié: (2026)
par: Erzurumlu, Yunus Talha, et autres
Publié: (2026)
TIMA: Text-Image Mutual Awareness for Balancing Zero-Shot Adversarial Robustness and Generalization Ability
par: Ma, Fengji, et autres
Publié: (2024)
par: Ma, Fengji, et autres
Publié: (2024)
Hybrid Global-Local Representation with Augmented Spatial Guidance for Zero-Shot Referring Image Segmentation
par: Liu, Ting, et autres
Publié: (2025)
par: Liu, Ting, et autres
Publié: (2025)
GVDepth: Zero-Shot Monocular Depth Estimation for Ground Vehicles based on Probabilistic Cue Fusion
par: Koledić, Karlo, et autres
Publié: (2024)
par: Koledić, Karlo, et autres
Publié: (2024)
SurgAtt-Tracker: Online Surgical Attention Tracking via Temporal Proposal Reranking and Motion-Aware Refinement
par: Zhou, Rulin, et autres
Publié: (2026)
par: Zhou, Rulin, et autres
Publié: (2026)
PoseTalk: Text-and-Audio-based Pose Control and Motion Refinement for One-Shot Talking Head Generation
par: Ling, Jun, et autres
Publié: (2024)
par: Ling, Jun, et autres
Publié: (2024)
Binary Verification for Zero-Shot Vision
par: Hu, Rongbin, et autres
Publié: (2025)
par: Hu, Rongbin, et autres
Publié: (2025)
Zero-Shot Refinement of Buildings' Segmentation Models using SAM
par: Mayladan, Ali, et autres
Publié: (2023)
par: Mayladan, Ali, et autres
Publié: (2023)
Mocap Anywhere: Towards Pairwise-Distance based Motion Capture in the Wild (for the Wild)
par: Abramovich, Ofir, et autres
Publié: (2026)
par: Abramovich, Ofir, et autres
Publié: (2026)
Documents similaires
-
ViTs for Action Classification in Videos: An Approach to Risky Tackle Detection in American Football Practice Videos
par: Zaidi, Syed Ahsan Masud, et autres
Publié: (2026) -
CornViT: A Multi-Stage Convolutional Vision Transformer Framework for Hierarchical Corn Kernel Analysis
par: Erukude, Sai Teja, et autres
Publié: (2025) -
Improving Adversarial Robustness Through Adaptive Learning-Driven Multi-Teacher Knowledge Distillation
par: Ullah, Hayat, et autres
Publié: (2025) -
DynaPURLS: Dynamic Refinement of Part-Aware Representations for Skeleton-Based Zero-Shot Action Recognition
par: Zhu, Jingmin, et autres
Publié: (2025) -
Next-Generation License Plate Detection and Recognition System using YOLOv8
par: Amin, Arslan, et autres
Publié: (2025)