Reason-Then-Retrieve for CoVR-R with Structured Edit Prompts and Dense-Sparse Fusion
Fuente:
arXiv
Salvato in:
| Autori principali: | Liu, DongQing, Qi, MengShi, Ji, HongWei |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
CoVR-R:Reason-Aware Composed Video Retrieval
di: Thawakar, Omkar, et al.
Pubblicazione: (2026)
di: Thawakar, Omkar, et al.
Pubblicazione: (2026)
CoVR-2: Automatic Data Construction for Composed Video Retrieval
di: Ventura, Lucas, et al.
Pubblicazione: (2023)
di: Ventura, Lucas, et al.
Pubblicazione: (2023)
2DMCG:2DMambawith Change Flow Guidance for Change Detection in Remote Sensing
di: Kaung, JunYao, et al.
Pubblicazione: (2025)
di: Kaung, JunYao, et al.
Pubblicazione: (2025)
Beyond Simple Edits: Composed Video Retrieval with Dense Modifications
di: Thawakar, Omkar, et al.
Pubblicazione: (2025)
di: Thawakar, Omkar, et al.
Pubblicazione: (2025)
ProEdit: Inversion-based Editing From Prompts Done Right
di: Ouyang, Zhi, et al.
Pubblicazione: (2025)
di: Ouyang, Zhi, et al.
Pubblicazione: (2025)
CogniEdit: Dense Gradient Flow Optimization for Fine-Grained Image Editing
di: Li, Yan, et al.
Pubblicazione: (2025)
di: Li, Yan, et al.
Pubblicazione: (2025)
Sprint: Sparse-Dense Residual Fusion for Efficient Diffusion Transformers
di: Park, Dogyun, et al.
Pubblicazione: (2025)
di: Park, Dogyun, et al.
Pubblicazione: (2025)
Exploring Sparse Visual Prompt for Domain Adaptive Dense Prediction
di: Yang, Senqiao, et al.
Pubblicazione: (2023)
di: Yang, Senqiao, et al.
Pubblicazione: (2023)
Orchestrate Latent Expertise: Advancing Online Continual Learning with Multi-Level Supervision and Reverse Self-Distillation
di: Yan, HongWei, et al.
Pubblicazione: (2024)
di: Yan, HongWei, et al.
Pubblicazione: (2024)
Edit2Perceive: Image Editing Diffusion Models Are Strong Dense Perceivers
di: Shi, Yiqing, et al.
Pubblicazione: (2025)
di: Shi, Yiqing, et al.
Pubblicazione: (2025)
From Sparse to Dense: Spatio-Temporal Fusion for Multi-View 3D Human Pose Estimation with DenseWarper
di: Li, Ling, et al.
Pubblicazione: (2026)
di: Li, Ling, et al.
Pubblicazione: (2026)
EditTransfer++: Toward Faithful and Efficient Visual-Prompt-Guided Image Editing
di: Chen, Lan, et al.
Pubblicazione: (2026)
di: Chen, Lan, et al.
Pubblicazione: (2026)
FlowR: Flowing from Sparse to Dense 3D Reconstructions
di: Fischer, Tobias, et al.
Pubblicazione: (2025)
di: Fischer, Tobias, et al.
Pubblicazione: (2025)
ReasonEdit: Towards Reasoning-Enhanced Image Editing Models
di: Yin, Fukun, et al.
Pubblicazione: (2025)
di: Yin, Fukun, et al.
Pubblicazione: (2025)
PhysEditBench: A Protocol-Conditioned Benchmark for Dense Physical-Map Prediction with Image Editors
di: Yang, Jiaxin, et al.
Pubblicazione: (2026)
di: Yang, Jiaxin, et al.
Pubblicazione: (2026)
SparseFusion: Efficient Sparse Multi-Modal Fusion Framework for Long-Range 3D Perception
di: Li, Yiheng, et al.
Pubblicazione: (2024)
di: Li, Yiheng, et al.
Pubblicazione: (2024)
CIR-CoT: Towards Interpretable Composed Image Retrieval via End-to-End Chain-of-Thought Reasoning
di: Lin, Weihuang, et al.
Pubblicazione: (2025)
di: Lin, Weihuang, et al.
Pubblicazione: (2025)
DesignEdit: Multi-Layered Latent Decomposition and Fusion for Unified & Accurate Image Editing
di: Jia, Yueru, et al.
Pubblicazione: (2024)
di: Jia, Yueru, et al.
Pubblicazione: (2024)
EditSleuth: A Dataset of Grounded Reasoning Chains for Image-Edit Forensics
di: Nguyen, Van-Loc, et al.
Pubblicazione: (2026)
di: Nguyen, Van-Loc, et al.
Pubblicazione: (2026)
FusionEdit: Semantic Fusion and Attention Modulation for Training-Free Image Editing
di: Lai, Yongwen, et al.
Pubblicazione: (2026)
di: Lai, Yongwen, et al.
Pubblicazione: (2026)
OG-Mapping: Octree-based Structured 3D Gaussians for Online Dense Mapping
di: Wang, Meng, et al.
Pubblicazione: (2024)
di: Wang, Meng, et al.
Pubblicazione: (2024)
DenseFormer: Learning Dense Depth Map from Sparse Depth and Image via Conditional Diffusion Model
di: Yuan, Ming, et al.
Pubblicazione: (2025)
di: Yuan, Ming, et al.
Pubblicazione: (2025)
Dense360: Dense Understanding from Omnidirectional Panoramas
di: Zhou, Yikang, et al.
Pubblicazione: (2025)
di: Zhou, Yikang, et al.
Pubblicazione: (2025)
ThinkRL-Edit: Thinking in Reinforcement Learning for Reasoning-Centric Image Editing
di: Li, Hengjia, et al.
Pubblicazione: (2026)
di: Li, Hengjia, et al.
Pubblicazione: (2026)
FGS-SLAM: Fourier-based Gaussian Splatting for Real-time SLAM with Sparse and Dense Map Fusion
di: Xu, Yansong, et al.
Pubblicazione: (2025)
di: Xu, Yansong, et al.
Pubblicazione: (2025)
Edit-Compass & EditReward-Compass: A Unified Benchmark for Image Editing and Reward Modeling
di: Bai, Xuehai, et al.
Pubblicazione: (2026)
di: Bai, Xuehai, et al.
Pubblicazione: (2026)
LoVR: A Benchmark for Long Video Retrieval in Multimodal Contexts
di: Cai, Qifeng, et al.
Pubblicazione: (2025)
di: Cai, Qifeng, et al.
Pubblicazione: (2025)
Retrv-R1: A Reasoning-Driven MLLM Framework for Universal and Efficient Multimodal Retrieval
di: Zhu, Lanyun, et al.
Pubblicazione: (2025)
di: Zhu, Lanyun, et al.
Pubblicazione: (2025)
S2D: Sparse to Dense Lifting for 3D Reconstruction with Minimal Inputs
di: Ji, Yuzhou, et al.
Pubblicazione: (2026)
di: Ji, Yuzhou, et al.
Pubblicazione: (2026)
DenseGRPO: From Sparse to Dense Reward for Flow Matching Model Alignment
di: Deng, Haoyou, et al.
Pubblicazione: (2026)
di: Deng, Haoyou, et al.
Pubblicazione: (2026)
FineEdit: Fine-Grained Image Edit with Bounding Box Guidance
di: Xu, Haohang, et al.
Pubblicazione: (2026)
di: Xu, Haohang, et al.
Pubblicazione: (2026)
Reasoning to Edit: Hypothetical Instruction-Based Image Editing with Visual Reasoning
di: He, Qingdong, et al.
Pubblicazione: (2025)
di: He, Qingdong, et al.
Pubblicazione: (2025)
FlexiEdit: Frequency-Aware Latent Refinement for Enhanced Non-Rigid Editing
di: Koo, Gwanhyeong, et al.
Pubblicazione: (2024)
di: Koo, Gwanhyeong, et al.
Pubblicazione: (2024)
Sparse Beats Dense: Rethinking Supervision in Radar-Camera Depth Completion
di: Li, Huadong, et al.
Pubblicazione: (2023)
di: Li, Huadong, et al.
Pubblicazione: (2023)
Adaptive Dense Evidence Refinement for Video Relational Reasoning for VRR-QA Challenge
di: Sun, Yuyang, et al.
Pubblicazione: (2026)
di: Sun, Yuyang, et al.
Pubblicazione: (2026)
SD4R: Sparse-to-Dense Learning for 3D Object Detection with 4D Radar
di: Bai, Xiaokai, et al.
Pubblicazione: (2026)
di: Bai, Xiaokai, et al.
Pubblicazione: (2026)
Follow the Saliency: Supervised Saliency for Retrieval-augmented Dense Video Captioning
di: Choi, Seung hee, et al.
Pubblicazione: (2026)
di: Choi, Seung hee, et al.
Pubblicazione: (2026)
Prompt to Restore, Restore to Prompt: Cyclic Prompting for Universal Adverse Weather Removal
di: Liao, Rongxin, et al.
Pubblicazione: (2025)
di: Liao, Rongxin, et al.
Pubblicazione: (2025)
Towards 3D VR-Sketch to 3D Shape Retrieval
di: Luo, Ling, et al.
Pubblicazione: (2022)
di: Luo, Ling, et al.
Pubblicazione: (2022)
CogCoM: A Visual Language Model with Chain-of-Manipulations Reasoning
di: Qi, Ji, et al.
Pubblicazione: (2024)
di: Qi, Ji, et al.
Pubblicazione: (2024)
Documenti analoghi
-
CoVR-R:Reason-Aware Composed Video Retrieval
di: Thawakar, Omkar, et al.
Pubblicazione: (2026) -
CoVR-2: Automatic Data Construction for Composed Video Retrieval
di: Ventura, Lucas, et al.
Pubblicazione: (2023) -
2DMCG:2DMambawith Change Flow Guidance for Change Detection in Remote Sensing
di: Kaung, JunYao, et al.
Pubblicazione: (2025) -
Beyond Simple Edits: Composed Video Retrieval with Dense Modifications
di: Thawakar, Omkar, et al.
Pubblicazione: (2025) -
ProEdit: Inversion-based Editing From Prompts Done Right
di: Ouyang, Zhi, et al.
Pubblicazione: (2025)