The 1st Solution for 7th LSVOS RVOS Track: SaSaSa2VA
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Niu, Quanzhu, Gong, Dengxian, Chen, Shihao, Zhang, Tao, Zhou, Yikang, Yuan, Haobo, Qi, Lu, Li, Xiangtai, Ji, Shunping |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
SaSaSaSa2VA: 2nd Place of the 5th PVUW MeViS-Text Track
par: Gong, Dengxian, et autres
Publié: (2026)
par: Gong, Dengxian, et autres
Publié: (2026)
Enhancing Sa2VA for Referent Video Object Segmentation: 2nd Solution for 7th LSVOS RVOS Track
par: Hong, Ran, et autres
Publié: (2025)
par: Hong, Ran, et autres
Publié: (2025)
2nd of the 5th PVUW MeViS-Audio Track: ASR-SaSaSa2VA
par: Wang, Zhiyu, et autres
Publié: (2026)
par: Wang, Zhiyu, et autres
Publié: (2026)
Sa2VA: Marrying SAM2 with LLaVA for Dense Grounded Understanding of Images and Videos
par: Yuan, Haobo, et autres
Publié: (2025)
par: Yuan, Haobo, et autres
Publié: (2025)
4th PVUW MeViS 3rd Place Report: Sa2VA
par: Yuan, Haobo, et autres
Publié: (2025)
par: Yuan, Haobo, et autres
Publié: (2025)
UNINEXT-Cutie: The 1st Solution for LSVOS Challenge RVOS Track
par: Fang, Hao, et autres
Publié: (2024)
par: Fang, Hao, et autres
Publié: (2024)
Beyond Appearance: Geometric Cues for Robust Video Instance Segmentation
par: Niu, Quanzhu, et autres
Publié: (2025)
par: Niu, Quanzhu, et autres
Publié: (2025)
Sa2VA-i: Improving Sa2VA Results with Consistent Training and Inference
par: Nekrasov, Alexey, et autres
Publié: (2025)
par: Nekrasov, Alexey, et autres
Publié: (2025)
DeH4R: A Decoupled and Hybrid Method for Road Network Graph Extraction
par: Gong, Dengxian, et autres
Publié: (2025)
par: Gong, Dengxian, et autres
Publié: (2025)
SeBaSa
Publié: (2022)
Publié: (2022)
The Sa'dan-Toraja
par: Nooy-Palm, H.
Publié: (2016)
par: Nooy-Palm, H.
Publié: (2016)
The Sa'dan-Toraja
par: Nooy-Palm, Hetty
Publié: (2022)
par: Nooy-Palm, Hetty
Publié: (2022)
The Instance-centric Transformer for the RVOS Track of LSVOS Challenge: 3rd Place Solution
par: Cao, Bin, et autres
Publié: (2024)
par: Cao, Bin, et autres
Publié: (2024)
Contribution to study on the seaweeds in Truong Sa archipelago (Truong Sa Lon and Nam Yet islands)
par: Pham, Huu Tri
Publié: (1996)
par: Pham, Huu Tri
Publié: (1996)
The 2nd Solution for LSVOS Challenge RVOS Track: Spatial-temporal Refinement for Consistent Semantic Segmentation
par: Tran, Tuyen
Publié: (2024)
par: Tran, Tuyen
Publié: (2024)
Dense360: Dense Understanding from Omnidirectional Panoramas
par: Zhou, Yikang, et autres
Publié: (2025)
par: Zhou, Yikang, et autres
Publié: (2025)
The Sa'dan Toradja Chant for the Deceased
par: van der Veen, H.
Publié: (2016)
par: van der Veen, H.
Publié: (2016)
The Merok Feast of the Sa'dan Toradja
par: van der Veen, H.
Publié: (2016)
par: van der Veen, H.
Publié: (2016)
Leksyon Na Didyital Ang Porma: Sagot Sa Pagpapatuloy Ng Pag-Aaral Sa Asignaturang Filipino
par: Saycon, Maria Rosalie
Publié: (2024)
par: Saycon, Maria Rosalie
Publié: (2024)
DVIS-DAQ: Improving Video Segmentation via Dynamic Anchor Queries
par: Zhou, Yikang, et autres
Publié: (2024)
par: Zhou, Yikang, et autres
Publié: (2024)
SaD: A Scenario-Aware Discriminator for Speech Enhancement
par: Yuan, Xihao, et autres
Publié: (2025)
par: Yuan, Xihao, et autres
Publié: (2025)
Phloeosinus aubei N C S Si Sa
par: Colonnelli, Enzo
Publié: (2003)
par: Colonnelli, Enzo
Publié: (2003)
Sa Pagitan ng Pa(g)hinga
par: Austria, Jeferson A.
Publié: (2026)
par: Austria, Jeferson A.
Publié: (2026)
PaSa: An LLM Agent for Comprehensive Academic Paper Search
par: He, Yichen, et autres
Publié: (2025)
par: He, Yichen, et autres
Publié: (2025)
MoSa: Motion Generation with Scalable Autoregressive Modeling
par: Liu, Mengyuan, et autres
Publié: (2025)
par: Liu, Mengyuan, et autres
Publié: (2025)
Sa Likod Ng Kanyang Tahimik Na Ngiti
par: Torres, Ellaine Mae R.
Publié: (2026)
par: Torres, Ellaine Mae R.
Publié: (2026)
Paths and Rivers; Sa’dan Toraja Society in Transformation
par: Waterson, Roxana
Publié: (2011)
par: Waterson, Roxana
Publié: (2011)
UniSaT: Unified-Objective Belief Model and Planner to Search for and Track Multiple Objects
par: Santos, Leonardo, et autres
Publié: (2024)
par: Santos, Leonardo, et autres
Publié: (2024)
LSVOS 2025 Challenge Report: Recent Advances in Complex Video Object Segmentation
par: Liu, Chang, et autres
Publié: (2025)
par: Liu, Chang, et autres
Publié: (2025)
SAMTok: Representing Any Mask with Two Words
par: Zhou, Yikang, et autres
Publié: (2026)
par: Zhou, Yikang, et autres
Publié: (2026)
Are They the Same? Exploring Visual Correspondence Shortcomings of Multimodal LLMs
par: Zhou, Yikang, et autres
Publié: (2025)
par: Zhou, Yikang, et autres
Publié: (2025)
Point Cloud Mamba: Point Cloud Learning via State Space Model
par: Zhang, Tao, et autres
Publié: (2024)
par: Zhang, Tao, et autres
Publié: (2024)
SaMOSA: Sandbox for Malware Orchestration and Side-Channel Analysis
par: Udeshi, Meet, et autres
Publié: (2025)
par: Udeshi, Meet, et autres
Publié: (2025)
SaTor: Exploring Satellite Routing in Tor to Reduce Latency
par: Li, Haozhi, et autres
Publié: (2024)
par: Li, Haozhi, et autres
Publié: (2024)
SaGE: Evaluating Moral Consistency in Large Language Models
par: Bonagiri, Vamshi Krishna, et autres
Publié: (2024)
par: Bonagiri, Vamshi Krishna, et autres
Publié: (2024)
SaFARi: State-Space Models for Frame-Agnostic Representation
par: Babaei, Hossein, et autres
Publié: (2025)
par: Babaei, Hossein, et autres
Publié: (2025)
Preliminary study on composition and distribution of Gastropods of Truong Sa archipelago
par: Lang, Van Keng
Publié: (1996)
par: Lang, Van Keng
Publié: (1996)
Discriminative Spatial-Semantic VOS Solution: 1st Place Solution for 6th LSVOS
par: Miao, Deshui, et autres
Publié: (2024)
par: Miao, Deshui, et autres
Publié: (2024)
OMG-LLaVA: Bridging Image-level, Object-level, Pixel-level Reasoning and Understanding
par: Zhang, Tao, et autres
Publié: (2024)
par: Zhang, Tao, et autres
Publié: (2024)
SaRO: Enhancing LLM Safety through Reasoning-based Alignment
par: Mou, Yutao, et autres
Publié: (2025)
par: Mou, Yutao, et autres
Publié: (2025)
Documents similaires
-
SaSaSaSa2VA: 2nd Place of the 5th PVUW MeViS-Text Track
par: Gong, Dengxian, et autres
Publié: (2026) -
Enhancing Sa2VA for Referent Video Object Segmentation: 2nd Solution for 7th LSVOS RVOS Track
par: Hong, Ran, et autres
Publié: (2025) -
2nd of the 5th PVUW MeViS-Audio Track: ASR-SaSaSa2VA
par: Wang, Zhiyu, et autres
Publié: (2026) -
Sa2VA: Marrying SAM2 with LLaVA for Dense Grounded Understanding of Images and Videos
par: Yuan, Haobo, et autres
Publié: (2025) -
4th PVUW MeViS 3rd Place Report: Sa2VA
par: Yuan, Haobo, et autres
Publié: (2025)