Vision-Aware Text Features in Referring Image Segmentation: From Object Understanding to Context Understanding
Fuente:
arXiv
Salvato in:
| Autori principali: | Nguyen-Truong, Hai, Nguyen, E-Ro, Vu, Tuan-Anh, Tran, Minh-Triet, Hua, Binh-Son, Yeung, Sai-Kit |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Power of Boundary and Reflection: Semantic Transparent Object Segmentation using Pyramid Vision Transformer with Transparent Cues
di: Vu, Tuan-Anh, et al.
Pubblicazione: (2025)
di: Vu, Tuan-Anh, et al.
Pubblicazione: (2025)
Self-supervised Video Object Segmentation with Distillation Learning of Deformable Attention
di: Truong, Quang-Trung, et al.
Pubblicazione: (2024)
di: Truong, Quang-Trung, et al.
Pubblicazione: (2024)
Catch Me If You Can Describe Me: Open-Vocabulary Camouflaged Instance Segmentation with Diffusion
di: Vu, Tuan-Anh, et al.
Pubblicazione: (2023)
di: Vu, Tuan-Anh, et al.
Pubblicazione: (2023)
ModeDreamer: Mode Guiding Score Distillation for Text-to-3D Generation using Reference Image Prompts
di: Tran, Uy Dieu, et al.
Pubblicazione: (2024)
di: Tran, Uy Dieu, et al.
Pubblicazione: (2024)
Language-driven Object Fusion into Neural Radiance Fields with Pose-Conditioned Dataset Updates
di: Shum, Ka Chun, et al.
Pubblicazione: (2023)
di: Shum, Ka Chun, et al.
Pubblicazione: (2023)
AUTV: Creating Underwater Video Datasets with Pixel-wise Annotations
di: Truong, Quang Trung, et al.
Pubblicazione: (2025)
di: Truong, Quang Trung, et al.
Pubblicazione: (2025)
Color Alignment in Diffusion
di: Shum, Ka Chun, et al.
Pubblicazione: (2025)
di: Shum, Ka Chun, et al.
Pubblicazione: (2025)
DiverseDream: Diverse Text-to-3D Synthesis with Augmented Text Embedding
di: Tran, Uy Dieu, et al.
Pubblicazione: (2023)
di: Tran, Uy Dieu, et al.
Pubblicazione: (2023)
LiftRefine: Progressively Refined View Synthesis from 3D Lifting with Volume-Triplane Representations
di: Do, Tung, et al.
Pubblicazione: (2024)
di: Do, Tung, et al.
Pubblicazione: (2024)
STER-VLM: Spatio-Temporal With Enhanced Reference Vision-Language Models
di: Nguyen-Nhu, Tinh-Anh, et al.
Pubblicazione: (2025)
di: Nguyen-Nhu, Tinh-Anh, et al.
Pubblicazione: (2025)
Interactive Interface For Semantic Segmentation Dataset Synthesis
di: Tran, Ngoc-Do, et al.
Pubblicazione: (2025)
di: Tran, Ngoc-Do, et al.
Pubblicazione: (2025)
Understanding SAM's Robustness to Noisy Labels through Gradient Down-weighting
di: Luong, Hoang-Chau, et al.
Pubblicazione: (2024)
di: Luong, Hoang-Chau, et al.
Pubblicazione: (2024)
Remarks on Semilinear σ‐Evolution Equations With Critical Damping and Critical Nonlinearity of Derivative Type
di: Tuan Anh Dao, et al.
Pubblicazione: (2025)
di: Tuan Anh Dao, et al.
Pubblicazione: (2025)
Link prediction Graph Neural Networks for structure recognition of Handwritten Mathematical Expressions
di: Nguyen, Cuong Tuan, et al.
Pubblicazione: (2025)
di: Nguyen, Cuong Tuan, et al.
Pubblicazione: (2025)
Automated Image Recognition Framework
di: Nguyen, Quang-Binh, et al.
Pubblicazione: (2025)
di: Nguyen, Quang-Binh, et al.
Pubblicazione: (2025)
SAMURAI: Shape-Aware Multimodal Retrieval for 3D Object Identification
di: Vo, Dinh-Khoi, et al.
Pubblicazione: (2025)
di: Vo, Dinh-Khoi, et al.
Pubblicazione: (2025)
CSD-VAR: Content-Style Decomposition in Visual Autoregressive Models
di: Nguyen, Quang-Binh, et al.
Pubblicazione: (2025)
di: Nguyen, Quang-Binh, et al.
Pubblicazione: (2025)
TF-SASM: Training-free Spatial-aware Sparse Memory for Multi-object Tracking
di: Nguyen-Quang, Thuc, et al.
Pubblicazione: (2024)
di: Nguyen-Quang, Thuc, et al.
Pubblicazione: (2024)
The Art of Camouflage: Few-Shot Learning for Animal Detection and Segmentation
di: Nguyen, Thanh-Danh, et al.
Pubblicazione: (2023)
di: Nguyen, Thanh-Danh, et al.
Pubblicazione: (2023)
Curriculum Demonstration Selection for In-Context Learning
di: Vu, Duc Anh, et al.
Pubblicazione: (2024)
di: Vu, Duc Anh, et al.
Pubblicazione: (2024)
Robust adaptive fuzzy sliding mode control for trajectory tracking for of cylindrical manipulator
di: Pham, Van Cuong, et al.
Pubblicazione: (2025)
di: Pham, Van Cuong, et al.
Pubblicazione: (2025)
SDPA++: A General Framework for Self-Supervised Denoising with Patch Aggregation
di: Nguyen, Huy Minh Nhat, et al.
Pubblicazione: (2025)
di: Nguyen, Huy Minh Nhat, et al.
Pubblicazione: (2025)
A Vietnamese Dataset for Text Segmentation and Multiple Choices Reading Comprehension
di: Hai, Toan Nguyen, et al.
Pubblicazione: (2025)
di: Hai, Toan Nguyen, et al.
Pubblicazione: (2025)
Flexible Genetic Algorithm for Quantum Support Vector Machines
di: Duc, Nguyen Minh, et al.
Pubblicazione: (2025)
di: Duc, Nguyen Minh, et al.
Pubblicazione: (2025)
RoboDesign1M: A Large-scale Dataset for Robot Design Understanding
di: Le, Tri, et al.
Pubblicazione: (2025)
di: Le, Tri, et al.
Pubblicazione: (2025)
VietMEAgent: Culturally-Aware Few-Shot Multimodal Explanation for Vietnamese Visual Question Answering
di: Nguyen, Hai-Dang, et al.
Pubblicazione: (2025)
di: Nguyen, Hai-Dang, et al.
Pubblicazione: (2025)
Online Trajectory Replanner for Dynamically Grasping Irregular Objects
di: Vu, Minh Nhat, et al.
Pubblicazione: (2025)
di: Vu, Minh Nhat, et al.
Pubblicazione: (2025)
Moment Sum-of-Squares Hierarchy for Gromov Wasserstein: Continuous Extensions and Sample Complexity
di: Tran, Hoang Anh, et al.
Pubblicazione: (2025)
di: Tran, Hoang Anh, et al.
Pubblicazione: (2025)
Sum-of-Squares Hierarchy for the Gromov Wasserstein Problem
di: Tran, Hoang Anh, et al.
Pubblicazione: (2025)
di: Tran, Hoang Anh, et al.
Pubblicazione: (2025)
Enhancing Visual Feature Attribution via Weighted Integrated Gradients
di: Tuan, Kien Tran Duc, et al.
Pubblicazione: (2025)
di: Tuan, Kien Tran Duc, et al.
Pubblicazione: (2025)
ShowFlow: From Robust Single Concept to Condition-Free Multi-Concept Generation
di: Hoang, Trong-Vu, et al.
Pubblicazione: (2025)
di: Hoang, Trong-Vu, et al.
Pubblicazione: (2025)
SAM-EG: Segment Anything Model with Egde Guidance framework for efficient Polyp Segmentation
di: Trinh, Quoc-Huy, et al.
Pubblicazione: (2024)
di: Trinh, Quoc-Huy, et al.
Pubblicazione: (2024)
An Efficient Approach for Machine Translation on Low-resource Languages: A Case Study in Vietnamese-Chinese
di: Son, Tran Ngoc, et al.
Pubblicazione: (2025)
di: Son, Tran Ngoc, et al.
Pubblicazione: (2025)
Stereochemical assignment of four monoterpene glucoside derivatives from Turpinia montana Kurz by NMR study combined with CD spectroscopy
di: Le Thanh Huong, et al.
Pubblicazione: (2024)
di: Le Thanh Huong, et al.
Pubblicazione: (2024)
Pacidesmus tuachua Nguyen, Eguchi and Vu 2024
di: Nguyen, Anh D., et al.
Pubblicazione: (2025)
di: Nguyen, Anh D., et al.
Pubblicazione: (2025)
Orthomorpha scabra subsp. scabra Jeekel 1964
di: Nguyen, Anh D., et al.
Pubblicazione: (2025)
di: Nguyen, Anh D., et al.
Pubblicazione: (2025)
Sphaerobelum clavigerum Verhoeff 1924
di: Nguyen, Anh D., et al.
Pubblicazione: (2025)
di: Nguyen, Anh D., et al.
Pubblicazione: (2025)
Tylopus procurvus Golovatch 1984
di: Nguyen, Anh D., et al.
Pubblicazione: (2025)
di: Nguyen, Anh D., et al.
Pubblicazione: (2025)
Cambalopsidae Cook 1895
di: Nguyen, Anh D., et al.
Pubblicazione: (2025)
di: Nguyen, Anh D., et al.
Pubblicazione: (2025)
Integrating Image Features with Convolutional Sequence-to-sequence Network for Multilingual Visual Question Answering
di: Thai, Triet Minh, et al.
Pubblicazione: (2023)
di: Thai, Triet Minh, et al.
Pubblicazione: (2023)
Documenti analoghi
-
Power of Boundary and Reflection: Semantic Transparent Object Segmentation using Pyramid Vision Transformer with Transparent Cues
di: Vu, Tuan-Anh, et al.
Pubblicazione: (2025) -
Self-supervised Video Object Segmentation with Distillation Learning of Deformable Attention
di: Truong, Quang-Trung, et al.
Pubblicazione: (2024) -
Catch Me If You Can Describe Me: Open-Vocabulary Camouflaged Instance Segmentation with Diffusion
di: Vu, Tuan-Anh, et al.
Pubblicazione: (2023) -
ModeDreamer: Mode Guiding Score Distillation for Text-to-3D Generation using Reference Image Prompts
di: Tran, Uy Dieu, et al.
Pubblicazione: (2024) -
Language-driven Object Fusion into Neural Radiance Fields with Pose-Conditioned Dataset Updates
di: Shum, Ka Chun, et al.
Pubblicazione: (2023)