Contrast-Unity for Partially-Supervised Temporal Sentence Grounding
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Haicheng, Ju, Chen, Lin, Weixiong, Ma, Chaofan, Xiao, Shuai, Zhang, Ya, Wang, Yanfeng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Advancing Myopia To Holism: Fully Contrastive Language-Image Pre-training
von: Wang, Haicheng, et al.
Veröffentlicht: (2024)
von: Wang, Haicheng, et al.
Veröffentlicht: (2024)
Multi-Sentence Grounding for Long-term Instructional Video
von: Li, Zeqian, et al.
Veröffentlicht: (2023)
von: Li, Zeqian, et al.
Veröffentlicht: (2023)
Weakly Supervised Temporal Sentence Grounding via Positive Sample Mining
von: Dong, Lu, et al.
Veröffentlicht: (2025)
von: Dong, Lu, et al.
Veröffentlicht: (2025)
Multi-Modal Prototypes for Open-World Semantic Segmentation
von: Yang, Yuhuan, et al.
Veröffentlicht: (2023)
von: Yang, Yuhuan, et al.
Veröffentlicht: (2023)
AttrSeg: Open-Vocabulary Semantic Segmentation via Attribute Decomposition-Aggregation
von: Ma, Chaofan, et al.
Veröffentlicht: (2023)
von: Ma, Chaofan, et al.
Veröffentlicht: (2023)
PMC-VQA: Visual Instruction Tuning for Medical Visual Question Answering
von: Zhang, Xiaoman, et al.
Veröffentlicht: (2023)
von: Zhang, Xiaoman, et al.
Veröffentlicht: (2023)
Length Matters: Length-Aware Transformer for Temporal Sentence Grounding
von: Wang, Yifan, et al.
Veröffentlicht: (2025)
von: Wang, Yifan, et al.
Veröffentlicht: (2025)
MoMa: Modulating Mamba for Adapting Image Foundation Models to Video Recognition
von: Yang, Yuhuan, et al.
Veröffentlicht: (2025)
von: Yang, Yuhuan, et al.
Veröffentlicht: (2025)
Diversifying Query: Region-Guided Transformer for Temporal Sentence Grounding
von: Sun, Xiaolong, et al.
Veröffentlicht: (2024)
von: Sun, Xiaolong, et al.
Veröffentlicht: (2024)
Sim-DETR: Unlock DETR for Temporal Sentence Grounding
von: Tang, Jiajin, et al.
Veröffentlicht: (2025)
von: Tang, Jiajin, et al.
Veröffentlicht: (2025)
GenMask: Adapting DiT for Segmentation via Direct Mask Generation
von: Yang, Yuhuan, et al.
Veröffentlicht: (2026)
von: Yang, Yuhuan, et al.
Veröffentlicht: (2026)
DENOISER: Rethinking the Robustness for Open-Vocabulary Action Recognition
von: Cheng, Haozhe, et al.
Veröffentlicht: (2024)
von: Cheng, Haozhe, et al.
Veröffentlicht: (2024)
ReMamber: Referring Image Segmentation with Mamba Twister
von: Yang, Yuhuan, et al.
Veröffentlicht: (2024)
von: Yang, Yuhuan, et al.
Veröffentlicht: (2024)
Bias-Conflict Sample Synthesis and Adversarial Removal Debias Strategy for Temporal Sentence Grounding in Video
von: Qi, Zhaobo, et al.
Veröffentlicht: (2024)
von: Qi, Zhaobo, et al.
Veröffentlicht: (2024)
FOLDER: Accelerating Multi-modal Large Language Models with Enhanced Performance
von: Wang, Haicheng, et al.
Veröffentlicht: (2025)
von: Wang, Haicheng, et al.
Veröffentlicht: (2025)
Boosting Temporal Sentence Grounding via Causal Inference
von: Tang, Kefan, et al.
Veröffentlicht: (2025)
von: Tang, Kefan, et al.
Veröffentlicht: (2025)
Multi-Pair Temporal Sentence Grounding via Multi-Thread Knowledge Transfer Network
von: Fang, Xiang, et al.
Veröffentlicht: (2024)
von: Fang, Xiang, et al.
Veröffentlicht: (2024)
Squeeze Out Tokens from Sample for Finer-Grained Data Governance
von: Lin, Weixiong, et al.
Veröffentlicht: (2025)
von: Lin, Weixiong, et al.
Veröffentlicht: (2025)
Efficient Temporal Sentence Grounding in Videos with Multi-Teacher Knowledge Distillation
von: Liang, Renjie, et al.
Veröffentlicht: (2023)
von: Liang, Renjie, et al.
Veröffentlicht: (2023)
SaFiRe: Saccade-Fixation Reiteration with Mamba for Referring Image Segmentation
von: Mao, Zhenjie, et al.
Veröffentlicht: (2025)
von: Mao, Zhenjie, et al.
Veröffentlicht: (2025)
HERO: Hierarchical Embedding-Refinement for Open-Vocabulary Temporal Sentence Grounding in Videos
von: Han, Tingting, et al.
Veröffentlicht: (2026)
von: Han, Tingting, et al.
Veröffentlicht: (2026)
Context Consistency Learning via Sentence Removal for Semi-Supervised Video Paragraph Grounding
von: Zhong, Yaokun, et al.
Veröffentlicht: (2025)
von: Zhong, Yaokun, et al.
Veröffentlicht: (2025)
Empower Words: DualGround for Structured Phrase and Sentence-Level Temporal Grounding
von: Kang, Minseok, et al.
Veröffentlicht: (2025)
von: Kang, Minseok, et al.
Veröffentlicht: (2025)
Universal Video Temporal Grounding with Generative Multi-modal Large Language Models
von: Li, Zeqian, et al.
Veröffentlicht: (2025)
von: Li, Zeqian, et al.
Veröffentlicht: (2025)
Let Me Finish My Sentence: Video Temporal Grounding with Holistic Text Understanding
von: Woo, Jongbhin, et al.
Veröffentlicht: (2024)
von: Woo, Jongbhin, et al.
Veröffentlicht: (2024)
Turbo: Informativity-Driven Acceleration Plug-In for Vision-Language Large Models
von: Ju, Chen, et al.
Veröffentlicht: (2024)
von: Ju, Chen, et al.
Veröffentlicht: (2024)
Contrastive Prompt Clustering for Weakly Supervised Semantic Segmentation
von: Wu, Wangyu, et al.
Veröffentlicht: (2025)
von: Wu, Wangyu, et al.
Veröffentlicht: (2025)
Region-aware Distribution Contrast: A Novel Approach to Multi-Task Partially Supervised Learning
von: Li, Meixuan, et al.
Veröffentlicht: (2024)
von: Li, Meixuan, et al.
Veröffentlicht: (2024)
RadGenome-Chest CT: A Grounded Vision-Language Dataset for Chest CT Analysis
von: Zhang, Xiaoman, et al.
Veröffentlicht: (2024)
von: Zhang, Xiaoman, et al.
Veröffentlicht: (2024)
Anomaly Detection in Electrocardiograms: Advancing Clinical Diagnosis Through Self-Supervised Learning
von: Jiang, Aofan, et al.
Veröffentlicht: (2024)
von: Jiang, Aofan, et al.
Veröffentlicht: (2024)
Improving Human Image Animation via Semantic Representation Alignment
von: Liu, Chang, et al.
Veröffentlicht: (2026)
von: Liu, Chang, et al.
Veröffentlicht: (2026)
Hierarchical Local-Global Transformer for Temporal Sentence Grounding
von: Fang, Xiang, et al.
Veröffentlicht: (2022)
von: Fang, Xiang, et al.
Veröffentlicht: (2022)
Multi-Scale Contrastive Learning for Video Temporal Grounding
von: Nguyen, Thong Thanh, et al.
Veröffentlicht: (2024)
von: Nguyen, Thong Thanh, et al.
Veröffentlicht: (2024)
From Priors to Perception: Grounding Video-LLMs in Physical Reality
von: Zhao, Zicheng, et al.
Veröffentlicht: (2026)
von: Zhao, Zicheng, et al.
Veröffentlicht: (2026)
Backdooring Self-Supervised Contrastive Learning by Noisy Alignment
von: Chen, Tuo, et al.
Veröffentlicht: (2025)
von: Chen, Tuo, et al.
Veröffentlicht: (2025)
Unsupervised Domain Adaptation via Similarity-based Prototypes for Cross-Modality Segmentation
von: Ye, Ziyu, et al.
Veröffentlicht: (2025)
von: Ye, Ziyu, et al.
Veröffentlicht: (2025)
Adaptive Patch Contrast for Weakly Supervised Semantic Segmentation
von: Wu, Wangyu, et al.
Veröffentlicht: (2024)
von: Wu, Wangyu, et al.
Veröffentlicht: (2024)
AMD: Adaptive Momentum and Decoupled Contrastive Learning Framework for Robust Long-Tail Trajectory Prediction
von: Rao, Bin, et al.
Veröffentlicht: (2025)
von: Rao, Bin, et al.
Veröffentlicht: (2025)
STPro: Spatial and Temporal Progressive Learning for Weakly Supervised Spatio-Temporal Grounding
von: Garg, Aaryan, et al.
Veröffentlicht: (2025)
von: Garg, Aaryan, et al.
Veröffentlicht: (2025)
Demographic-Aware Self-Supervised Anomaly Detection Pretraining for Equitable Rare Cardiac Diagnosis
von: Huang, Chaoqin, et al.
Veröffentlicht: (2026)
von: Huang, Chaoqin, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Advancing Myopia To Holism: Fully Contrastive Language-Image Pre-training
von: Wang, Haicheng, et al.
Veröffentlicht: (2024) -
Multi-Sentence Grounding for Long-term Instructional Video
von: Li, Zeqian, et al.
Veröffentlicht: (2023) -
Weakly Supervised Temporal Sentence Grounding via Positive Sample Mining
von: Dong, Lu, et al.
Veröffentlicht: (2025) -
Multi-Modal Prototypes for Open-World Semantic Segmentation
von: Yang, Yuhuan, et al.
Veröffentlicht: (2023) -
AttrSeg: Open-Vocabulary Semantic Segmentation via Attribute Decomposition-Aggregation
von: Ma, Chaofan, et al.
Veröffentlicht: (2023)