ShotFinder: Imagination-Driven Open-Domain Video Shot Retrieval via Web Search
Fuente:
arXiv
Salvato in:
| Autori principali: | Yu, Tao, Jin, Haopeng, Wang, Hao, Chai, Shenghua, Yang, Yujia, Gong, Junhao, Guo, Jiaming, Zhang, Minghui, Chen, Xinlong, Zhang, Zhenghao, Zhou, Yuxuan, Xiong, Yufei, Zhang, Shanbin, Yang, Jiabing, Yi, Hongzhu, Wang, Xinming, Zhong, Cheng, Ma, Xiao, Zhang, Zhang, Huang, Yan, Wang, Liang |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Beyond Closed-Pool Video Retrieval: A Benchmark and Agent Framework for Real-World Video Search and Moment Localization
di: Yu, Tao, et al.
Pubblicazione: (2026)
di: Yu, Tao, et al.
Pubblicazione: (2026)
Omni-DeepSearch: A Benchmark for Audio-Driven Omni-Modal Deep Search
di: Yu, Tao, et al.
Pubblicazione: (2026)
di: Yu, Tao, et al.
Pubblicazione: (2026)
PaperX: A Unified Framework for Multimodal Academic Presentation Generation with Scholar DAG
di: Yu, Tao, et al.
Pubblicazione: (2026)
di: Yu, Tao, et al.
Pubblicazione: (2026)
BrowserAgent: Building Web Agents with Human-Inspired Web Browsing Actions
di: Yu, Tao, et al.
Pubblicazione: (2025)
di: Yu, Tao, et al.
Pubblicazione: (2025)
Beyond the All-in-One Agent: Benchmarking Role-Specialized Multi-Agent Collaboration in Enterprise Workflows
di: Yu, Tao, et al.
Pubblicazione: (2026)
di: Yu, Tao, et al.
Pubblicazione: (2026)
FSUNav: A Cerebrum-Cerebellum Architecture for Fast, Safe, and Universal Zero-Shot Goal-Oriented Navigation
di: Tan, Mingao, et al.
Pubblicazione: (2026)
di: Tan, Mingao, et al.
Pubblicazione: (2026)
Learning by Imagining: Debiased Feature Augmentation for Compositional Zero-Shot Learning
di: Zhang, Haozhe, et al.
Pubblicazione: (2025)
di: Zhang, Haozhe, et al.
Pubblicazione: (2025)
EndoFinder: Online Image Retrieval for Explainable Colorectal Polyp Diagnosis
di: Yang, Ruijie, et al.
Pubblicazione: (2024)
di: Yang, Ruijie, et al.
Pubblicazione: (2024)
Pseudo-label Based Domain Adaptation for Zero-Shot Text Steganalysis
di: Luo, Yufei, et al.
Pubblicazione: (2024)
di: Luo, Yufei, et al.
Pubblicazione: (2024)
Fine-Grained Prototypes Distillation for Few-Shot Object Detection
di: Wang, Zichen, et al.
Pubblicazione: (2024)
di: Wang, Zichen, et al.
Pubblicazione: (2024)
ObjectFinder: An Open-Vocabulary Assistive System for Interactive Object Search by Blind People
di: Liu, Ruiping, et al.
Pubblicazione: (2024)
di: Liu, Ruiping, et al.
Pubblicazione: (2024)
PhotoDoodle: Learning Artistic Image Editing from Few-Shot Pairwise Data
di: Huang, Shijie, et al.
Pubblicazione: (2025)
di: Huang, Shijie, et al.
Pubblicazione: (2025)
EndoFinder: Online Lesion Retrieval for Explainable Colorectal Polyp Diagnosis Leveraging Latent Scene Representations
di: Yang, Ruijie, et al.
Pubblicazione: (2025)
di: Yang, Ruijie, et al.
Pubblicazione: (2025)
Breaking Robustness Barriers in Cognitive Diagnosis: A One-Shot Neural Architecture Search Perspective
di: Wang, Ziwen, et al.
Pubblicazione: (2026)
di: Wang, Ziwen, et al.
Pubblicazione: (2026)
SynPo: Boosting Training-Free Few-Shot Medical Segmentation via High-Quality Negative Prompts
di: Liu, Yufei, et al.
Pubblicazione: (2025)
di: Liu, Yufei, et al.
Pubblicazione: (2025)
TrajRAG: Retrieving Geometric-Semantic Experience for Zero-Shot Object Navigation
di: Wang, Yiyao, et al.
Pubblicazione: (2026)
di: Wang, Yiyao, et al.
Pubblicazione: (2026)
RPO:Reinforcement Fine-Tuning with Partial Reasoning Optimization
di: Yi, Hongzhu, et al.
Pubblicazione: (2026)
di: Yi, Hongzhu, et al.
Pubblicazione: (2026)
Transvalvular Precision: Digital Cholangioscopy‐Guided SEMS Deployment for Malignant Ileocecal Obstruction
di: Shanbin Wu, et al.
Pubblicazione: (2025)
di: Shanbin Wu, et al.
Pubblicazione: (2025)
Unveiling the Magic: Investigating Attention Distillation in Retrieval-augmented Generation
di: Li, Zizhong, et al.
Pubblicazione: (2024)
di: Li, Zizhong, et al.
Pubblicazione: (2024)
MPCI-Bench: A Benchmark for Multimodal Pairwise Contextual Integrity Evaluation of Language Model Agents
di: Wang, Shouju, et al.
Pubblicazione: (2026)
di: Wang, Shouju, et al.
Pubblicazione: (2026)
DreamNav: A Trajectory-Based Imaginative Framework for Zero-Shot Vision-and-Language Navigation
di: Wang, Yunheng, et al.
Pubblicazione: (2025)
di: Wang, Yunheng, et al.
Pubblicazione: (2025)
Adapting Foundation Models for Few-Shot Medical Image Segmentation: Actively and Sequentially
di: Yang, Jingyun, et al.
Pubblicazione: (2025)
di: Yang, Jingyun, et al.
Pubblicazione: (2025)
ShotVerse: Advancing Cinematic Camera Control for Text-Driven Multi-Shot Video Creation
di: Yang, Songlin, et al.
Pubblicazione: (2026)
di: Yang, Songlin, et al.
Pubblicazione: (2026)
Token-Level Precise Attack on RAG: Searching for the Best Alternatives to Mislead Generation
di: Li, Zizhong, et al.
Pubblicazione: (2025)
di: Li, Zizhong, et al.
Pubblicazione: (2025)
Intermediate Distillation: Data-Efficient Distillation from Black-Box LLMs for Information Retrieval
di: Li, Zizhong, et al.
Pubblicazione: (2024)
di: Li, Zizhong, et al.
Pubblicazione: (2024)
OmniShotCut: Holistic Relational Shot Boundary Detection with Shot-Query Transformer
di: Wang, Boyang, et al.
Pubblicazione: (2026)
di: Wang, Boyang, et al.
Pubblicazione: (2026)
Solving Zero-Shot 3D Visual Grounding as Constraint Satisfaction Problems
di: Yuan, Qihao, et al.
Pubblicazione: (2024)
di: Yuan, Qihao, et al.
Pubblicazione: (2024)
MRAD: Zero-Shot Anomaly Detection with Memory-Driven Retrieval
di: Xu, Chaoran, et al.
Pubblicazione: (2026)
di: Xu, Chaoran, et al.
Pubblicazione: (2026)
Generative Editing in the Joint Vision-Language Space for Zero-Shot Composed Image Retrieval
di: Wang, Xin, et al.
Pubblicazione: (2025)
di: Wang, Xin, et al.
Pubblicazione: (2025)
Omni IIE Bench: Benchmarking the Practical Capabilities of Image Editing Models
di: Yang, Yujia, et al.
Pubblicazione: (2026)
di: Yang, Yujia, et al.
Pubblicazione: (2026)
In-Context Learning for Few-Shot Nested Named Entity Recognition
di: Zhang, Meishan, et al.
Pubblicazione: (2024)
di: Zhang, Meishan, et al.
Pubblicazione: (2024)
The Universal Personalizer: Few-Shot Dysarthric Speech Recognition via Meta-Learning
di: Agarwal, Dhruuv, et al.
Pubblicazione: (2025)
di: Agarwal, Dhruuv, et al.
Pubblicazione: (2025)
ZeroSim: Zero-Shot Analog Circuit Evaluation with Unified Transformer Embeddings
di: Yang, Xiaomeng, et al.
Pubblicazione: (2025)
di: Yang, Xiaomeng, et al.
Pubblicazione: (2025)
Otter: Mitigating Background Distractions of Wide-Angle Few-Shot Action Recognition with Enhanced RWKV
di: Huang, Wenbo, et al.
Pubblicazione: (2025)
di: Huang, Wenbo, et al.
Pubblicazione: (2025)
Commonality in Few: Few-Shot Multimodal Anomaly Detection via Hypergraph-Enhanced Memory
di: Lin, Yuxuan, et al.
Pubblicazione: (2025)
di: Lin, Yuxuan, et al.
Pubblicazione: (2025)
WISER: Wider Search, Deeper Thinking, and Adaptive Fusion for Training-Free Zero-Shot Composed Image Retrieval
di: Wang, Tianyue, et al.
Pubblicazione: (2026)
di: Wang, Tianyue, et al.
Pubblicazione: (2026)
Node-Time Conditional Prompt Learning In Dynamic Graphs
di: Yu, Xingtong, et al.
Pubblicazione: (2024)
di: Yu, Xingtong, et al.
Pubblicazione: (2024)
Multi-Floor Zero-Shot Object Navigation Policy
di: Zhang, Lingfeng, et al.
Pubblicazione: (2024)
di: Zhang, Lingfeng, et al.
Pubblicazione: (2024)
GATHER: Convergence-Centric Hyper-Entity Retrieval for Zero-Shot Cell-Type Annotation
di: Zhang, Zhonghui, et al.
Pubblicazione: (2026)
di: Zhang, Zhonghui, et al.
Pubblicazione: (2026)
Matcher: Segment Anything with One Shot Using All-Purpose Feature Matching
di: Liu, Yang, et al.
Pubblicazione: (2023)
di: Liu, Yang, et al.
Pubblicazione: (2023)
Documenti analoghi
-
Beyond Closed-Pool Video Retrieval: A Benchmark and Agent Framework for Real-World Video Search and Moment Localization
di: Yu, Tao, et al.
Pubblicazione: (2026) -
Omni-DeepSearch: A Benchmark for Audio-Driven Omni-Modal Deep Search
di: Yu, Tao, et al.
Pubblicazione: (2026) -
PaperX: A Unified Framework for Multimodal Academic Presentation Generation with Scholar DAG
di: Yu, Tao, et al.
Pubblicazione: (2026) -
BrowserAgent: Building Web Agents with Human-Inspired Web Browsing Actions
di: Yu, Tao, et al.
Pubblicazione: (2025) -
Beyond the All-in-One Agent: Benchmarking Role-Specialized Multi-Agent Collaboration in Enterprise Workflows
di: Yu, Tao, et al.
Pubblicazione: (2026)