Mastering Negation: Boosting Grounding Models via Grouped Opposition-Based Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Yang, Zesheng, Jiang, Xi, Hu, Bingzhang, Guan, Weili, Cong, Runmin, Qi, Guo-Jun, Zheng, Feng |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AnomalyClaw: A Universal Visual Anomaly Detection Agent via Tool-Grounded Refutation
by: Jiang, Xi, et al.
Published: (2026)
by: Jiang, Xi, et al.
Published: (2026)
Object-Shot Enhanced Grounding Network for Egocentric Video
by: Feng, Yisen, et al.
Published: (2025)
by: Feng, Yisen, et al.
Published: (2025)
AutoIAD: Manager-Driven Multi-Agent Collaboration for Automated Industrial Anomaly Detection
by: Ji, Dongwei, et al.
Published: (2025)
by: Ji, Dongwei, et al.
Published: (2025)
Pix2Fact: When Vision Is Not Enough -- Benchmarking Fine-Grained VQA with Web Verification on High-Resolution Real-World Scenes
by: Jiang, Yifan, et al.
Published: (2026)
by: Jiang, Yifan, et al.
Published: (2026)
Empowering DINO Representations for Underwater Instance Segmentation via Aligner and Prompter
by: Chen, Zhiyang, et al.
Published: (2025)
by: Chen, Zhiyang, et al.
Published: (2025)
Understanding Model Reprogramming for CLIP via Decoupling Visual Prompts
by: Cai, Chengyi, et al.
Published: (2025)
by: Cai, Chengyi, et al.
Published: (2025)
FastBoost: Progressive Attention with Dynamic Scaling for Efficient Deep Learning
by: Yuan, JunXi
Published: (2025)
by: Yuan, JunXi
Published: (2025)
From Sight to Insight: Unleashing Eye-Tracking in Weakly Supervised Video Salient Object Detection
by: Qin, Qi, et al.
Published: (2025)
by: Qin, Qi, et al.
Published: (2025)
SDDNet: Style-guided Dual-layer Disentanglement Network for Shadow Detection
by: Cong, Runmin, et al.
Published: (2023)
by: Cong, Runmin, et al.
Published: (2023)
A Mutual Learning Method for Salient Object Detection with intertwined Multi-Supervision--Revised
by: Wu, Runmin, et al.
Published: (2025)
by: Wu, Runmin, et al.
Published: (2025)
Video Object Segmentation via SAM 2: The 4th Solution for LSVOS Challenge VOS Track
by: Pan, Feiyu, et al.
Published: (2024)
by: Pan, Feiyu, et al.
Published: (2024)
UIS-Mamba: Exploring Mamba for Underwater Instance Segmentation via Dynamic Tree Scan and Hidden State Weaken
by: Cong, Runmin, et al.
Published: (2025)
by: Cong, Runmin, et al.
Published: (2025)
IOTA: Corrective Knowledge-Guided Prompt Learning via Black-White Box Framework
by: Wang, Shaokun, et al.
Published: (2026)
by: Wang, Shaokun, et al.
Published: (2026)
PhysMaster: Mastering Physical Representation for Video Generation via Reinforcement Learning
by: Ji, Sihui, et al.
Published: (2025)
by: Ji, Sihui, et al.
Published: (2025)
Point-aware Interaction and CNN-induced Refinement Network for RGB-D Salient Object Detection
by: Cong, Runmin, et al.
Published: (2023)
by: Cong, Runmin, et al.
Published: (2023)
Semantic Concentration for Self-Supervised Dense Representations Learning
by: Wen, Peisong, et al.
Published: (2025)
by: Wen, Peisong, et al.
Published: (2025)
GraSP-VL: Length as a Semantic Granularity Interface for Vision-Language Representations
by: Li, Zesheng, et al.
Published: (2026)
by: Li, Zesheng, et al.
Published: (2026)
Bayesian-guided Label Mapping for Visual Reprogramming
by: Cai, Chengyi, et al.
Published: (2024)
by: Cai, Chengyi, et al.
Published: (2024)
Attribute-based Visual Reprogramming for Vision-Language Models
by: Cai, Chengyi, et al.
Published: (2025)
by: Cai, Chengyi, et al.
Published: (2025)
Sample-specific Masks for Visual Reprogramming-based Prompting
by: Cai, Chengyi, et al.
Published: (2024)
by: Cai, Chengyi, et al.
Published: (2024)
LTRL: Boosting Long-tail Recognition via Reflective Learning
by: Zhao, Qihao, et al.
Published: (2024)
by: Zhao, Qihao, et al.
Published: (2024)
Unleashing Correlation and Continuity for Hyperspectral Reconstruction from RGB Images
by: Feng, Fuxiang, et al.
Published: (2025)
by: Feng, Fuxiang, et al.
Published: (2025)
R^3: Composed Video Retrieval via Reasoning-Guided Recalling and Re-ranking
by: Li, Zixu, et al.
Published: (2026)
by: Li, Zixu, et al.
Published: (2026)
Boosting Active Learning with Knowledge Transfer
by: Wang, Tianyang, et al.
Published: (2025)
by: Wang, Tianyang, et al.
Published: (2025)
Learning Hierarchical Color Guidance for Depth Map Super-Resolution
by: Cong, Runmin, et al.
Published: (2024)
by: Cong, Runmin, et al.
Published: (2024)
Learning Spatio-Temporal Patterns of Polar Ice Layers With Physics-Informed Graph Neural Network
by: Liu, Zesheng, et al.
Published: (2024)
by: Liu, Zesheng, et al.
Published: (2024)
Boosting Consistency in Story Visualization with Rich-Contextual Conditional Diffusion Models
by: Shen, Fei, et al.
Published: (2024)
by: Shen, Fei, et al.
Published: (2024)
Taming Real-World Space-Time Video Super-Resolution with One-Step Diffusion
by: Wei, Shuoyan, et al.
Published: (2026)
by: Wei, Shuoyan, et al.
Published: (2026)
Divide-and-Conquer Decoupled Network for Cross-Domain Few-Shot Segmentation
by: Cong, Runmin, et al.
Published: (2025)
by: Cong, Runmin, et al.
Published: (2025)
FashionLens: Toward Versatile Fashion Image Retrieval via Task-Adaptive Learning
by: Wen, Haokun, et al.
Published: (2026)
by: Wen, Haokun, et al.
Published: (2026)
GA2-CLIP: Generic Attribute Anchor for Efficient Prompt Tuningin Video-Language Models
by: Wang, Bin, et al.
Published: (2025)
by: Wang, Bin, et al.
Published: (2025)
Improving Generalized Visual Grounding with Instance-aware Joint Learning
by: Dai, Ming, et al.
Published: (2025)
by: Dai, Ming, et al.
Published: (2025)
Boosting Temporal Sentence Grounding via Causal Inference
by: Tang, Kefan, et al.
Published: (2025)
by: Tang, Kefan, et al.
Published: (2025)
Boosting Multi-View Indoor 3D Object Detection via Adaptive 3D Volume Construction
by: Zhang, Runmin, et al.
Published: (2025)
by: Zhang, Runmin, et al.
Published: (2025)
IGFuse: Interactive 3D Gaussian Scene Reconstruction via Multi-Scans Fusion
by: Hu, Wenhao, et al.
Published: (2025)
by: Hu, Wenhao, et al.
Published: (2025)
AbductiveMLLM: Boosting Visual Abductive Reasoning Within MLLMs
by: Chang, Boyu, et al.
Published: (2026)
by: Chang, Boyu, et al.
Published: (2026)
UNINEXT-Cutie: The 1st Solution for LSVOS Challenge RVOS Track
by: Fang, Hao, et al.
Published: (2024)
by: Fang, Hao, et al.
Published: (2024)
The 1st Solution for 4th PVUW MeViS Challenge: Unleashing the Potential of Large Multimodal Models for Referring Video Segmentation
by: Fang, Hao, et al.
Published: (2025)
by: Fang, Hao, et al.
Published: (2025)
Beyond Global Scanning: Adaptive Visual State Space Modeling for Salient Object Detection in Optical Remote Sensing Images
by: Ren, Mengyu, et al.
Published: (2025)
by: Ren, Mengyu, et al.
Published: (2025)
Context Consistency Learning via Sentence Removal for Semi-Supervised Video Paragraph Grounding
by: Zhong, Yaokun, et al.
Published: (2025)
by: Zhong, Yaokun, et al.
Published: (2025)
Similar Items
-
AnomalyClaw: A Universal Visual Anomaly Detection Agent via Tool-Grounded Refutation
by: Jiang, Xi, et al.
Published: (2026) -
Object-Shot Enhanced Grounding Network for Egocentric Video
by: Feng, Yisen, et al.
Published: (2025) -
AutoIAD: Manager-Driven Multi-Agent Collaboration for Automated Industrial Anomaly Detection
by: Ji, Dongwei, et al.
Published: (2025) -
Pix2Fact: When Vision Is Not Enough -- Benchmarking Fine-Grained VQA with Web Verification on High-Resolution Real-World Scenes
by: Jiang, Yifan, et al.
Published: (2026) -
Empowering DINO Representations for Underwater Instance Segmentation via Aligner and Prompter
by: Chen, Zhiyang, et al.
Published: (2025)