PREGEN: Uncovering Latent Thoughts in Composed Video Retrieval
Fuente:
arXiv
Saved in:
| Main Authors: | Serussi, Gabriele, Vainshtein, David, Kouchly, Jonathan, Di Castro, Dotan, Baskin, Chaim |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Robot Instance Segmentation with Few Annotations for Grasping
by: Kimhi, Moshe, et al.
Published: (2024)
by: Kimhi, Moshe, et al.
Published: (2024)
HiMu: Hierarchical Multimodal Frame Selection for Long Video Question Answering
by: Ben-Ami, Dan, et al.
Published: (2026)
by: Ben-Ami, Dan, et al.
Published: (2026)
HERBench: A Benchmark for Multi-Evidence Integration in Video Question Answering
by: Ben-Ami, Dan, et al.
Published: (2025)
by: Ben-Ami, Dan, et al.
Published: (2025)
Leveraging Latents for Efficient Thermography Classification and Segmentation
by: Shor, Tamir, et al.
Published: (2024)
by: Shor, Tamir, et al.
Published: (2024)
Cross-Instance Gaussian Splatting Registration via Geometry-Aware Feature-Guided Alignment
by: Amoyal, Roy, et al.
Published: (2026)
by: Amoyal, Roy, et al.
Published: (2026)
Conceptual Learning via Embedding Approximations for Reinforcing Interpretability and Transparency
by: Dikter, Maor, et al.
Published: (2024)
by: Dikter, Maor, et al.
Published: (2024)
Sparse patches adversarial attacks via extrapolating point-wise information
by: Nemcovsky, Yaniv, et al.
Published: (2024)
by: Nemcovsky, Yaniv, et al.
Published: (2024)
$\mathbf{R}^3$: Reconstruction, Raw, and Rain: Deraining Directly in the Bayer Domain
by: Rothschild, Nate, et al.
Published: (2025)
by: Rothschild, Nate, et al.
Published: (2025)
Single Image Test-Time Adaptation for Segmentation
by: Janouskova, Klara, et al.
Published: (2023)
by: Janouskova, Klara, et al.
Published: (2023)
TEAM PILOT -- Learned Feasible Extendable Set of Dynamic MRI Acquisition Trajectories
by: Shor, Tamir, et al.
Published: (2024)
by: Shor, Tamir, et al.
Published: (2024)
Efficient Discriminative Joint Encoders for Large Scale Vision-Language Reranking
by: Taraday, Mitchell Keren, et al.
Published: (2025)
by: Taraday, Mitchell Keren, et al.
Published: (2025)
Depth Over RGB: Automatic Evaluation of Open Surgery Skills Using Depth Camera
by: Zuckerman, Ido, et al.
Published: (2024)
by: Zuckerman, Ido, et al.
Published: (2024)
Class-Conditioned Transformation for Enhanced Robust Image Classification
by: Blau, Tsachi, et al.
Published: (2023)
by: Blau, Tsachi, et al.
Published: (2023)
Semi-Supervised Semantic Segmentation via Marginal Contextual Information
by: Kimhi, Moshe, et al.
Published: (2023)
by: Kimhi, Moshe, et al.
Published: (2023)
CoVR-R:Reason-Aware Composed Video Retrieval
by: Thawakar, Omkar, et al.
Published: (2026)
by: Thawakar, Omkar, et al.
Published: (2026)
Composed Video Retrieval via Enriched Context and Discriminative Embeddings
by: Thawakar, Omkar, et al.
Published: (2024)
by: Thawakar, Omkar, et al.
Published: (2024)
Beyond Simple Edits: Composed Video Retrieval with Dense Modifications
by: Thawakar, Omkar, et al.
Published: (2025)
by: Thawakar, Omkar, et al.
Published: (2025)
From Play to Replay: Composed Video Retrieval for Temporally Fine-Grained Videos
by: Gupta, Animesh, et al.
Published: (2025)
by: Gupta, Animesh, et al.
Published: (2025)
Composed Object Retrieval: Object-level Retrieval via Composed Expressions
by: Wang, Tong, et al.
Published: (2025)
by: Wang, Tong, et al.
Published: (2025)
Reason-before-Retrieve: One-Stage Reflective Chain-of-Thoughts for Training-Free Zero-Shot Composed Image Retrieval
by: Tang, Yuanmin, et al.
Published: (2024)
by: Tang, Yuanmin, et al.
Published: (2024)
Look Where It Matters: High-Resolution Crops Retrieval for Efficient VLMs
by: Shabtay, Nimrod, et al.
Published: (2026)
by: Shabtay, Nimrod, et al.
Published: (2026)
CompoDiff: Versatile Composed Image Retrieval With Latent Diffusion
by: Gu, Geonmo, et al.
Published: (2023)
by: Gu, Geonmo, et al.
Published: (2023)
CIR-CoT: Towards Interpretable Composed Image Retrieval via End-to-End Chain-of-Thought Reasoning
by: Lin, Weihuang, et al.
Published: (2025)
by: Lin, Weihuang, et al.
Published: (2025)
EgoCVR: An Egocentric Benchmark for Fine-Grained Composed Video Retrieval
by: Hummel, Thomas, et al.
Published: (2024)
by: Hummel, Thomas, et al.
Published: (2024)
CoVR-2: Automatic Data Construction for Composed Video Retrieval
by: Ventura, Lucas, et al.
Published: (2023)
by: Ventura, Lucas, et al.
Published: (2023)
Noisy Annotations in Semantic Segmentation
by: Kimhi, Moshe, et al.
Published: (2024)
by: Kimhi, Moshe, et al.
Published: (2024)
DiTASK: Multi-Task Fine-Tuning with Diffeomorphic Transformations
by: Mantri, Krishna Sri Ipsit, et al.
Published: (2025)
by: Mantri, Krishna Sri Ipsit, et al.
Published: (2025)
ISCUTE: Instance Segmentation of Cables Using Text Embedding
by: Kozlovsky, Shir, et al.
Published: (2024)
by: Kozlovsky, Shir, et al.
Published: (2024)
Towards General Modality Translation with Contrastive and Predictive Latent Diffusion Bridge
by: Berman, Nimrod, et al.
Published: (2025)
by: Berman, Nimrod, et al.
Published: (2025)
MCoT-MVS: Multi-level Vision Selection by Multi-modal Chain-of-Thought Reasoning for Composed Image Retrieval
by: Ge, Xuri, et al.
Published: (2026)
by: Ge, Xuri, et al.
Published: (2026)
CoVA: Text-Guided Composed Video Retrieval for Audio-Visual Content
by: Han, Gyuwon, et al.
Published: (2026)
by: Han, Gyuwon, et al.
Published: (2026)
T1-PILOT: Optimized Trajectories for T1 Mapping Acceleration
by: Shor, Tamir, et al.
Published: (2025)
by: Shor, Tamir, et al.
Published: (2025)
MCoT-RE: Multi-Faceted Chain-of-Thought and Re-Ranking for Training-Free Zero-Shot Composed Image Retrieval
by: Park, Jeong-Woo, et al.
Published: (2025)
by: Park, Jeong-Woo, et al.
Published: (2025)
WAVECLIP: Wavelet Tokenization for Adaptive-Resolution CLIP
by: Kimhi, Moshe, et al.
Published: (2025)
by: Kimhi, Moshe, et al.
Published: (2025)
HUD: Hierarchical Uncertainty-Aware Disambiguation Network for Composed Video Retrieval
by: Chen, Zhiwei, et al.
Published: (2025)
by: Chen, Zhiwei, et al.
Published: (2025)
R^3: Composed Video Retrieval via Reasoning-Guided Recalling and Re-ranking
by: Li, Zixu, et al.
Published: (2026)
by: Li, Zixu, et al.
Published: (2026)
Instance-Level Composed Image Retrieval
by: Psomas, Bill, et al.
Published: (2025)
by: Psomas, Bill, et al.
Published: (2025)
Zero Shot Composed Image Retrieval
by: Kakarla, Santhosh, et al.
Published: (2025)
by: Kakarla, Santhosh, et al.
Published: (2025)
Composed Image Retrieval for Remote Sensing
by: Psomas, Bill, et al.
Published: (2024)
by: Psomas, Bill, et al.
Published: (2024)
Structured Diffusion Bridges: Inductive Bias for Denoising Diffusion Bridges
by: Kosman, Eitan, et al.
Published: (2026)
by: Kosman, Eitan, et al.
Published: (2026)
Similar Items
-
Robot Instance Segmentation with Few Annotations for Grasping
by: Kimhi, Moshe, et al.
Published: (2024) -
HiMu: Hierarchical Multimodal Frame Selection for Long Video Question Answering
by: Ben-Ami, Dan, et al.
Published: (2026) -
HERBench: A Benchmark for Multi-Evidence Integration in Video Question Answering
by: Ben-Ami, Dan, et al.
Published: (2025) -
Leveraging Latents for Efficient Thermography Classification and Segmentation
by: Shor, Tamir, et al.
Published: (2024) -
Cross-Instance Gaussian Splatting Registration via Geometry-Aware Feature-Guided Alignment
by: Amoyal, Roy, et al.
Published: (2026)