Gespeichert in:
| Hauptverfasser: | Zhang, Chunhui, Cui, Yawen, Lin, Weilin, Huang, Guanjie, Rong, Yan, Liu, Li, Shan, Shiguang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2408.08315 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Comprehensive Survey on Human Video Generation: Challenges, Methods, and Insights
von: Lei, Wentao, et al.
Veröffentlicht: (2024)
von: Lei, Wentao, et al.
Veröffentlicht: (2024)
WebUOT-1M: Advancing Deep Underwater Object Tracking with A Million-Scale Benchmark
von: Zhang, Chunhui, et al.
Veröffentlicht: (2024)
von: Zhang, Chunhui, et al.
Veröffentlicht: (2024)
RefSAM: Efficiently Adapting Segmenting Anything Model for Referring Video Object Segmentation
von: Li, Yonglin, et al.
Veröffentlicht: (2023)
von: Li, Yonglin, et al.
Veröffentlicht: (2023)
LensWalk: Agentic Video Understanding by Planning How You See in Videos
von: Li, Keliang, et al.
Veröffentlicht: (2026)
von: Li, Keliang, et al.
Veröffentlicht: (2026)
SAMAug: Point Prompt Augmentation for Segment Anything Model
von: Dai, Haixing, et al.
Veröffentlicht: (2023)
von: Dai, Haixing, et al.
Veröffentlicht: (2023)
X-SAM: From Segment Anything to Any Segmentation
von: Wang, Hao, et al.
Veröffentlicht: (2025)
von: Wang, Hao, et al.
Veröffentlicht: (2025)
INFACT: A Diagnostic Benchmark for Induced Faithfulness and Factuality Hallucinations in Video-LLMs
von: Yang, Junqi, et al.
Veröffentlicht: (2026)
von: Yang, Junqi, et al.
Veröffentlicht: (2026)
CamSAM2: Segment Anything Accurately in Camouflaged Videos
von: Zhou, Yuli, et al.
Veröffentlicht: (2025)
von: Zhou, Yuli, et al.
Veröffentlicht: (2025)
Underwater Camouflaged Object Tracking Meets Vision-Language SAM2
von: Zhang, Chunhui, et al.
Veröffentlicht: (2024)
von: Zhang, Chunhui, et al.
Veröffentlicht: (2024)
Describe Anything: Detailed Localized Image and Video Captioning
von: Lian, Long, et al.
Veröffentlicht: (2025)
von: Lian, Long, et al.
Veröffentlicht: (2025)
Measuring the Measurers: Quality Evaluation of Hallucination Benchmarks for Large Vision-Language Models
von: Yan, Bei, et al.
Veröffentlicht: (2024)
von: Yan, Bei, et al.
Veröffentlicht: (2024)
MM-MoralBench: A MultiModal Moral Evaluation Benchmark for Large Vision-Language Models
von: Yan, Bei, et al.
Veröffentlicht: (2024)
von: Yan, Bei, et al.
Veröffentlicht: (2024)
Segment-Anything Models Achieve Zero-shot Robustness in Autonomous Driving
von: Yan, Jun, et al.
Veröffentlicht: (2024)
von: Yan, Jun, et al.
Veröffentlicht: (2024)
SAM 3: Segment Anything with Concepts
von: Carion, Nicolas, et al.
Veröffentlicht: (2025)
von: Carion, Nicolas, et al.
Veröffentlicht: (2025)
SAM 2: Segment Anything in Images and Videos
von: Ravi, Nikhila, et al.
Veröffentlicht: (2024)
von: Ravi, Nikhila, et al.
Veröffentlicht: (2024)
LENS: Learning to Segment Anything with Unified Reinforced Reasoning
von: Zhu, Lianghui, et al.
Veröffentlicht: (2025)
von: Zhu, Lianghui, et al.
Veröffentlicht: (2025)
Trigger without Trace: Towards Stealthy Backdoor Attack on Text-to-Image Diffusion Models
von: Zhang, Jie, et al.
Veröffentlicht: (2025)
von: Zhang, Jie, et al.
Veröffentlicht: (2025)
WeakSAM: Segment Anything Meets Weakly-supervised Instance-level Recognition
von: Zhu, Lianghui, et al.
Veröffentlicht: (2024)
von: Zhu, Lianghui, et al.
Veröffentlicht: (2024)
Video Depth Anything: Consistent Depth Estimation for Super-Long Videos
von: Chen, Sili, et al.
Veröffentlicht: (2025)
von: Chen, Sili, et al.
Veröffentlicht: (2025)
Segment Anything in Pathology Images with Natural Language
von: Chen, Zhixuan, et al.
Veröffentlicht: (2025)
von: Chen, Zhixuan, et al.
Veröffentlicht: (2025)
Tracking and Segmenting Anything in Any Modality
von: Zhang, Tianlu, et al.
Veröffentlicht: (2025)
von: Zhang, Tianlu, et al.
Veröffentlicht: (2025)
Segment and Matte Anything in a Unified Model
von: Fan, Zezhong, et al.
Veröffentlicht: (2026)
von: Fan, Zezhong, et al.
Veröffentlicht: (2026)
ALPS: An Auto-Labeling and Pre-training Scheme for Remote Sensing Segmentation With Segment Anything Model
von: Zhang, Song, et al.
Veröffentlicht: (2024)
von: Zhang, Song, et al.
Veröffentlicht: (2024)
Compress Any Segment Anything Model (SAM)
von: Fan, Juntong, et al.
Veröffentlicht: (2025)
von: Fan, Juntong, et al.
Veröffentlicht: (2025)
Promoting Segment Anything Model towards Highly Accurate Dichotomous Image Segmentation
von: Liu, Xianjie, et al.
Veröffentlicht: (2023)
von: Liu, Xianjie, et al.
Veröffentlicht: (2023)
MedSAM3: Delving into Segment Anything with Medical Concepts
von: Liu, Anglin, et al.
Veröffentlicht: (2025)
von: Liu, Anglin, et al.
Veröffentlicht: (2025)
Customize Segment Anything Model for Multi-Modal Semantic Segmentation with Mixture of LoRA Experts
von: Zhu, Chenyang, et al.
Veröffentlicht: (2024)
von: Zhu, Chenyang, et al.
Veröffentlicht: (2024)
MetricAnything: Scaling Metric Depth Pretraining with Noisy Heterogeneous Sources
von: Ma, Baorui, et al.
Veröffentlicht: (2026)
von: Ma, Baorui, et al.
Veröffentlicht: (2026)
UniPose: A Unified Multimodal Framework for Human Pose Comprehension, Generation and Editing
von: Li, Yiheng, et al.
Veröffentlicht: (2024)
von: Li, Yiheng, et al.
Veröffentlicht: (2024)
Semantic Localization Guiding Segment Anything Model For Reference Remote Sensing Image Segmentation
von: Li, Shuyang, et al.
Veröffentlicht: (2025)
von: Li, Shuyang, et al.
Veröffentlicht: (2025)
Pose-Robust Calibration Strategy for Point-of-Gaze Estimation on Mobile Phones
von: Zhao, Yujie, et al.
Veröffentlicht: (2025)
von: Zhao, Yujie, et al.
Veröffentlicht: (2025)
RASA: Replace Anyone, Say Anything -- A Training-Free Framework for Audio-Driven and Universal Portrait Video Editing
von: Pan, Tianrui, et al.
Veröffentlicht: (2025)
von: Pan, Tianrui, et al.
Veröffentlicht: (2025)
Prior-free Balanced Replay: Uncertainty-guided Reservoir Sampling for Long-Tailed Continual Learning
von: Liu, Lei, et al.
Veröffentlicht: (2024)
von: Liu, Lei, et al.
Veröffentlicht: (2024)
SAM2 for Image and Video Segmentation: A Comprehensive Survey
von: Jiaxing, Zhang, et al.
Veröffentlicht: (2025)
von: Jiaxing, Zhang, et al.
Veröffentlicht: (2025)
AnyMo: Scaling Any-Modality Conditional Motion Generation with Masked Modeling
von: Li, Yiheng, et al.
Veröffentlicht: (2026)
von: Li, Yiheng, et al.
Veröffentlicht: (2026)
Annolid: Annotate, Segment, and Track Anything You Need
von: Yang, Chen, et al.
Veröffentlicht: (2024)
von: Yang, Chen, et al.
Veröffentlicht: (2024)
Adapting Segment Anything Model to Melanoma Segmentation in Microscopy Slide Images
von: Liu, Qingyuan, et al.
Veröffentlicht: (2024)
von: Liu, Qingyuan, et al.
Veröffentlicht: (2024)
Efficient Quantization-Aware Training on Segment Anything Model in Medical Images and Its Deployment
von: Lu, Haisheng, et al.
Veröffentlicht: (2024)
von: Lu, Haisheng, et al.
Veröffentlicht: (2024)
Composition Vision-Language Understanding via Segment and Depth Anything Model
von: Huo, Mingxiao, et al.
Veröffentlicht: (2024)
von: Huo, Mingxiao, et al.
Veröffentlicht: (2024)
SparseSAM: Structured Sparsification of Activations in Segment Anything Models
von: Tran, Hoai-Chau, et al.
Veröffentlicht: (2026)
von: Tran, Hoai-Chau, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
A Comprehensive Survey on Human Video Generation: Challenges, Methods, and Insights
von: Lei, Wentao, et al.
Veröffentlicht: (2024) -
WebUOT-1M: Advancing Deep Underwater Object Tracking with A Million-Scale Benchmark
von: Zhang, Chunhui, et al.
Veröffentlicht: (2024) -
RefSAM: Efficiently Adapting Segmenting Anything Model for Referring Video Object Segmentation
von: Li, Yonglin, et al.
Veröffentlicht: (2023) -
LensWalk: Agentic Video Understanding by Planning How You See in Videos
von: Li, Keliang, et al.
Veröffentlicht: (2026) -
SAMAug: Point Prompt Augmentation for Segment Anything Model
von: Dai, Haixing, et al.
Veröffentlicht: (2023)