C3-OWD: A Curriculum Cross-modal Contrastive Learning Framework for Open-World Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Siheng, Li, Zhengdao, Li, Yanshu, Xiao, Canran, Zhan, Haibo, Yao, Zhengtao, Zhang, Xuzhi, Kang, Jiale, Li, Linshan, Liu, Weiming, Dong, Zhikang, Shen, Jifeng, Dong, Junhao, Sun, Qiang, Koniusz, Piotr |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DeCo-DETR: Decoupled Cognition DETR for efficient Open-Vocabulary Object Detection
by: Wang, Siheng, et al.
Published: (2026)
by: Wang, Siheng, et al.
Published: (2026)
JEPA-T: Joint-Embedding Predictive Architecture with Text Fusion for Image Generation
by: Wan, Siheng, et al.
Published: (2025)
by: Wan, Siheng, et al.
Published: (2025)
ICED: Concept-level Machine Unlearning via Interpretable Concept Decomposition
by: Lin, Shen, et al.
Published: (2026)
by: Lin, Shen, et al.
Published: (2026)
ReMoMask: Retrieval-Augmented Masked Motion Generation
by: Li, Zhengdao, et al.
Published: (2025)
by: Li, Zhengdao, et al.
Published: (2025)
Hierarchically Robust Zero-shot Vision-language Models
by: Dong, Junhao, et al.
Published: (2026)
by: Dong, Junhao, et al.
Published: (2026)
OpenKD: Opening Prompt Diversity for Zero- and Few-shot Keypoint Detection
by: Lu, Changsheng, et al.
Published: (2024)
by: Lu, Changsheng, et al.
Published: (2024)
Adaptive Multi-head Contrastive Learning
by: Wang, Lei, et al.
Published: (2023)
by: Wang, Lei, et al.
Published: (2023)
ChromouVQA: Benchmarking Vision-Language Models under Chromatic Camouflaged Images
by: Zhang, Yunfei, et al.
Published: (2025)
by: Zhang, Yunfei, et al.
Published: (2025)
CHAIN: Enhancing Generalization in Data-Efficient GANs via lipsCHitz continuity constrAIned Normalization
by: Ni, Yao, et al.
Published: (2024)
by: Ni, Yao, et al.
Published: (2024)
Feature Hallucination for Self-supervised Action Recognition
by: Wang, Lei, et al.
Published: (2025)
by: Wang, Lei, et al.
Published: (2025)
MMT-ARD: Multimodal Multi-Teacher Adversarial Distillation for Robust Vision-Language Models
by: Li, Yuqi, et al.
Published: (2025)
by: Li, Yuqi, et al.
Published: (2025)
Advancing Multimodal In-Context Learning in Large Vision-Language Models with Task-aware Demonstrations
by: Li, Yanshu
Published: (2025)
by: Li, Yanshu
Published: (2025)
Can Green Technology and Renewable Energy Drive Carbon Neutrality in CAREC Nations? Evidence From Financial Development and Sustainable Growth
by: Zhengdao Li, et al.
Published: (2025)
by: Zhengdao Li, et al.
Published: (2025)
Token Entropy Regularization for Multi-modal Antenna Affiliation Identification
by: Chen, Dong, et al.
Published: (2026)
by: Chen, Dong, et al.
Published: (2026)
MM-StanceDet: Retrieval-Augmented Multi-modal Multi-agent Stance Detection
by: Lu, Weihai, et al.
Published: (2026)
by: Lu, Weihai, et al.
Published: (2026)
Uncertainty-DTW for Sequences and Visual Tokens
by: Wang, Lei, et al.
Published: (2026)
by: Wang, Lei, et al.
Published: (2026)
PACE: Marrying generalization in PArameter-efficient fine-tuning with Consistency rEgularization
by: Ni, Yao, et al.
Published: (2024)
by: Ni, Yao, et al.
Published: (2024)
Video Understanding by Design: How Datasets Shape Architectures and Insights
by: Wang, Lei, et al.
Published: (2025)
by: Wang, Lei, et al.
Published: (2025)
Graph Self-Supervised Learning with Learnable Structural and Positional Encodings
by: Wijesinghe, Asiri, et al.
Published: (2025)
by: Wijesinghe, Asiri, et al.
Published: (2025)
Multispectral State-Space Feature Fusion: Bridging Shared and Cross-Parametric Interactions for Object Detection
by: Shen, Jifeng, et al.
Published: (2025)
by: Shen, Jifeng, et al.
Published: (2025)
GLIP-OOD: Zero-Shot Graph OOD Detection with Graph Foundation Model
by: Xu, Haoyan, et al.
Published: (2025)
by: Xu, Haoyan, et al.
Published: (2025)
On the Interpolation Effect of Score Smoothing in Diffusion Models
by: Chen, Zhengdao
Published: (2025)
by: Chen, Zhengdao
Published: (2025)
Neural Hilbert Ladders: Multi-Layer Neural Networks in Function Space
by: Chen, Zhengdao
Published: (2023)
by: Chen, Zhengdao
Published: (2023)
Inductive Graph Few-shot Class Incremental Learning
by: Li, Yayong, et al.
Published: (2024)
by: Li, Yayong, et al.
Published: (2024)
SDCD: Structure-Disrupted Contrastive Decoding for Mitigating Hallucinations in Large Vision-Language Models
by: Xia, Yuxuan, et al.
Published: (2026)
by: Xia, Yuxuan, et al.
Published: (2026)
Logistic-aided Huber M-estimator for robust GNSS positioning
by: Li, Zhengdao, et al.
Published: (2026)
by: Li, Zhengdao, et al.
Published: (2026)
Understanding and Mitigating Hyperbolic Dimensional Collapse in Graph Contrastive Learning
by: Zhang, Yifei, et al.
Published: (2023)
by: Zhang, Yifei, et al.
Published: (2023)
CCPrefix: Counterfactual Contrastive Prefix-Tuning for Many-Class Classification
by: Li, Yang, et al.
Published: (2022)
by: Li, Yang, et al.
Published: (2022)
CP-PINNs: Data-Driven Changepoints Detection in PDEs Using Online Optimized Physics-Informed Neural Networks
by: Dong, Zhikang, et al.
Published: (2022)
by: Dong, Zhikang, et al.
Published: (2022)
WIMLE: Uncertainty-Aware World Models with IMLE for Sample-Efficient Continuous Control
by: Aghabozorgi, Mehran, et al.
Published: (2026)
by: Aghabozorgi, Mehran, et al.
Published: (2026)
AMMKD: Adaptive Multimodal Multi-teacher Distillation for Lightweight Vision-Language Models
by: Li, Yuqi, et al.
Published: (2025)
by: Li, Yuqi, et al.
Published: (2025)
Towards Open Federated Learning Platforms: Survey and Vision from Technical and Legal Perspectives
by: Duan, Moming, et al.
Published: (2023)
by: Duan, Moming, et al.
Published: (2023)
Swift Sampler: Efficient Learning of Sampler by 10 Parameters
by: Yao, Jiawei, et al.
Published: (2024)
by: Yao, Jiawei, et al.
Published: (2024)
Commentary on ‘Rationalising Hepatocellular Carcinoma Screening in Chronic Hepatitis B: Evaluation of Predictive Scores in an Australian Cohort’
by: Jiawei Hu, et al.
Published: (2025)
by: Jiawei Hu, et al.
Published: (2025)
OWMM-Agent: Open World Mobile Manipulation With Multi-modal Agentic Data Synthesis
by: Chen, Junting, et al.
Published: (2025)
by: Chen, Junting, et al.
Published: (2025)
Subspace Kernel Learning on Tensor Sequences
by: Wang, Lei, et al.
Published: (2026)
by: Wang, Lei, et al.
Published: (2026)
Motion meets Attention: Video Motion Prompts
by: Chen, Qixiang, et al.
Published: (2024)
by: Chen, Qixiang, et al.
Published: (2024)
Graph Your Own Prompt
by: Ding, Xi, et al.
Published: (2025)
by: Ding, Xi, et al.
Published: (2025)
Focus on Background: Exploring SAM's Potential in Few-shot Medical Image Segmentation with Background-centric Prompting
by: Bo, Yuntian, et al.
Published: (2026)
by: Bo, Yuntian, et al.
Published: (2026)
Learning Time in Static Classifiers
by: Ding, Xi, et al.
Published: (2025)
by: Ding, Xi, et al.
Published: (2025)
Similar Items
-
DeCo-DETR: Decoupled Cognition DETR for efficient Open-Vocabulary Object Detection
by: Wang, Siheng, et al.
Published: (2026) -
JEPA-T: Joint-Embedding Predictive Architecture with Text Fusion for Image Generation
by: Wan, Siheng, et al.
Published: (2025) -
ICED: Concept-level Machine Unlearning via Interpretable Concept Decomposition
by: Lin, Shen, et al.
Published: (2026) -
ReMoMask: Retrieval-Augmented Masked Motion Generation
by: Li, Zhengdao, et al.
Published: (2025) -
Hierarchically Robust Zero-shot Vision-language Models
by: Dong, Junhao, et al.
Published: (2026)