LVOS: A Benchmark for Large-scale Long-term Video Object Segmentation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Hong, Lingyi, Liu, Zhongying, Chen, Wenchao, Tan, Chenzhi, Feng, Yuang, Zhou, Xinyu, Guo, Pinxue, Li, Jinglun, Chen, Zhaoyu, Gao, Shuyong, Zhang, Wei, Zhang, Wenqiang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ClickVOS: Click Video Object Segmentation
von: Guo, Pinxue, et al.
Veröffentlicht: (2024)
von: Guo, Pinxue, et al.
Veröffentlicht: (2024)
X-Prompt: Multi-modal Visual Prompt for Video Object Segmentation
von: Guo, Pinxue, et al.
Veröffentlicht: (2024)
von: Guo, Pinxue, et al.
Veröffentlicht: (2024)
General Compression Framework for Efficient Transformer Object Tracking
von: Hong, Lingyi, et al.
Veröffentlicht: (2024)
von: Hong, Lingyi, et al.
Veröffentlicht: (2024)
Scoring, Remember, and Reference: Catching Camouflaged Objects in Videos
von: Feng, Yuang, et al.
Veröffentlicht: (2025)
von: Feng, Yuang, et al.
Veröffentlicht: (2025)
Reading Relevant Feature from Global Representation Memory for Visual Object Tracking
von: Zhou, Xinyu, et al.
Veröffentlicht: (2024)
von: Zhou, Xinyu, et al.
Veröffentlicht: (2024)
MSVCOD:A Large-Scale Multi-Scene Dataset for Video Camouflage Object Detection
von: Gao, Shuyong, et al.
Veröffentlicht: (2025)
von: Gao, Shuyong, et al.
Veröffentlicht: (2025)
TagOOD: A Novel Approach to Out-of-Distribution Detection via Vision-Language Representations and Class Center Learning
von: Li, Jinglun, et al.
Veröffentlicht: (2024)
von: Li, Jinglun, et al.
Veröffentlicht: (2024)
DeTrack: In-model Latent Denoising Learning for Visual Object Tracking
von: Zhou, Xinyu, et al.
Veröffentlicht: (2025)
von: Zhou, Xinyu, et al.
Veröffentlicht: (2025)
VideoPure: Diffusion-based Adversarial Purification for Video Recognition
von: Jiang, Kaixun, et al.
Veröffentlicht: (2025)
von: Jiang, Kaixun, et al.
Veröffentlicht: (2025)
OneTracker: Unifying Visual Object Tracking with Foundation Models and Efficient Tuning
von: Hong, Lingyi, et al.
Veröffentlicht: (2024)
von: Hong, Lingyi, et al.
Veröffentlicht: (2024)
OneVOS: Unifying Video Object Segmentation with All-in-One Transformer Framework
von: Li, Wanyun, et al.
Veröffentlicht: (2024)
von: Li, Wanyun, et al.
Veröffentlicht: (2024)
Unified Multimodal Visual Tracking with Dual Mixture-of-Experts
von: Hong, Lingyi, et al.
Veröffentlicht: (2026)
von: Hong, Lingyi, et al.
Veröffentlicht: (2026)
Improving Adversarial Transferability with Neighbourhood Gradient Information
von: Guo, Haijing, et al.
Veröffentlicht: (2024)
von: Guo, Haijing, et al.
Veröffentlicht: (2024)
VLA-Hijack: A Transferable Patch Attack against Vision-Language-Action Models via Visual Proprioception Hijacking
von: Fu, Jiyuan, et al.
Veröffentlicht: (2026)
von: Fu, Jiyuan, et al.
Veröffentlicht: (2026)
Seeing is Believing: Rich-Context Hallucination Detection for MLLMs via Backward Visual Grounding
von: Guo, Pinxue, et al.
Veröffentlicht: (2025)
von: Guo, Pinxue, et al.
Veröffentlicht: (2025)
Dynamic Semantic-Aware Correlation Modeling for UAV Tracking
von: Zhou, Xinyu, et al.
Veröffentlicht: (2025)
von: Zhou, Xinyu, et al.
Veröffentlicht: (2025)
OpenVIS: Open-vocabulary Video Instance Segmentation
von: Guo, Pinxue, et al.
Veröffentlicht: (2023)
von: Guo, Pinxue, et al.
Veröffentlicht: (2023)
PanoVOS: Bridging Non-panoramic and Panoramic Views with Transformer for Video Segmentation
von: Yan, Shilin, et al.
Veröffentlicht: (2023)
von: Yan, Shilin, et al.
Veröffentlicht: (2023)
Boosting the Transferability of Adversarial Attacks with Global Momentum Initialization
von: Wang, Jiafeng, et al.
Veröffentlicht: (2022)
von: Wang, Jiafeng, et al.
Veröffentlicht: (2022)
LingoLoop Attack: Trapping MLLMs via Linguistic Context and State Entrapment into Endless Loops
von: Fu, Jiyuan, et al.
Veröffentlicht: (2025)
von: Fu, Jiyuan, et al.
Veröffentlicht: (2025)
RSAgent: Learning to Reason and Act for Text-Guided Segmentation via Multi-Turn Tool Invocations
von: He, Xingqi, et al.
Veröffentlicht: (2025)
von: He, Xingqi, et al.
Veröffentlicht: (2025)
PG-Attack: A Precision-Guided Adversarial Attack Framework Against Vision Foundation Models for Autonomous Driving
von: Fu, Jiyuan, et al.
Veröffentlicht: (2024)
von: Fu, Jiyuan, et al.
Veröffentlicht: (2024)
Enhancing Diffusion-based Unrestricted Adversarial Attacks via Adversary Preferences Alignment
von: Jiang, Kaixun, et al.
Veröffentlicht: (2025)
von: Jiang, Kaixun, et al.
Veröffentlicht: (2025)
Hierarchical Visual Categories Modeling: A Joint Representation Learning and Density Estimation Framework for Out-of-Distribution Detection
von: Li, Jinglun, et al.
Veröffentlicht: (2024)
von: Li, Jinglun, et al.
Veröffentlicht: (2024)
VideoSAM: Open-World Video Segmentation
von: Guo, Pinxue, et al.
Veröffentlicht: (2024)
von: Guo, Pinxue, et al.
Veröffentlicht: (2024)
P3S-Diffusion:A Selective Subject-driven Generation Framework via Point Supervision
von: Hu, Junjie, et al.
Veröffentlicht: (2024)
von: Hu, Junjie, et al.
Veröffentlicht: (2024)
AnimatePainter: A Self-Supervised Rendering Framework for Reconstructing Painting Process
von: Hu, Junjie, et al.
Veröffentlicht: (2025)
von: Hu, Junjie, et al.
Veröffentlicht: (2025)
Improving Adversarial Transferability of Vision-Language Pre-training Models through Collaborative Multimodal Interaction
von: Fu, Jiyuan, et al.
Veröffentlicht: (2024)
von: Fu, Jiyuan, et al.
Veröffentlicht: (2024)
Long-RVOS: A Comprehensive Benchmark for Long-term Referring Video Object Segmentation
von: Liang, Tianming, et al.
Veröffentlicht: (2025)
von: Liang, Tianming, et al.
Veröffentlicht: (2025)
LSVOS Challenge Report: Large-scale Complex and Long Video Object Segmentation
von: Ding, Henghui, et al.
Veröffentlicht: (2024)
von: Ding, Henghui, et al.
Veröffentlicht: (2024)
Synthesizing Near-Boundary OOD Samples for Out-of-Distribution Detection
von: Li, Jinglun, et al.
Veröffentlicht: (2025)
von: Li, Jinglun, et al.
Veröffentlicht: (2025)
A Holistically Point-guided Text Framework for Weakly-Supervised Camouflaged Object Detection
von: Mok, Tsui Qin, et al.
Veröffentlicht: (2025)
von: Mok, Tsui Qin, et al.
Veröffentlicht: (2025)
LTCF-Net: A Transformer-Enhanced Dual-Channel Fourier Framework for Low-Light Image Restoration
von: Zhang, Gaojing, et al.
Veröffentlicht: (2024)
von: Zhang, Gaojing, et al.
Veröffentlicht: (2024)
Delving into Decision-based Black-box Attacks on Semantic Segmentation
von: Chen, Zhaoyu, et al.
Veröffentlicht: (2024)
von: Chen, Zhaoyu, et al.
Veröffentlicht: (2024)
Supervised Learning Model for Key Frame Identification from Cow Teat Videos
von: Wang, Minghao, et al.
Veröffentlicht: (2024)
von: Wang, Minghao, et al.
Veröffentlicht: (2024)
HLV-1K: A Large-scale Hour-Long Video Benchmark for Time-Specific Long Video Understanding
von: Zou, Heqing, et al.
Veröffentlicht: (2025)
von: Zou, Heqing, et al.
Veröffentlicht: (2025)
One-shot Training for Video Object Segmentation
von: Chen, Baiyu, et al.
Veröffentlicht: (2024)
von: Chen, Baiyu, et al.
Veröffentlicht: (2024)
Boosting Salient Object Detection with Knowledge Distillated from Large Foundation Models
von: He, Miaoyang, et al.
Veröffentlicht: (2025)
von: He, Miaoyang, et al.
Veröffentlicht: (2025)
Enhancing Object Discovery for Unsupervised Instance Segmentation and Object Detection
von: Feng, Xingyu, et al.
Veröffentlicht: (2025)
von: Feng, Xingyu, et al.
Veröffentlicht: (2025)
Seeing the Scene Matters: Revealing Forgetting in Video Understanding Models with a Scene-Aware Long-Video Benchmark
von: Chen, Seng Nam, et al.
Veröffentlicht: (2026)
von: Chen, Seng Nam, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
ClickVOS: Click Video Object Segmentation
von: Guo, Pinxue, et al.
Veröffentlicht: (2024) -
X-Prompt: Multi-modal Visual Prompt for Video Object Segmentation
von: Guo, Pinxue, et al.
Veröffentlicht: (2024) -
General Compression Framework for Efficient Transformer Object Tracking
von: Hong, Lingyi, et al.
Veröffentlicht: (2024) -
Scoring, Remember, and Reference: Catching Camouflaged Objects in Videos
von: Feng, Yuang, et al.
Veröffentlicht: (2025) -
Reading Relevant Feature from Global Representation Memory for Visual Object Tracking
von: Zhou, Xinyu, et al.
Veröffentlicht: (2024)