Domain Generalization through Spatial Relation Induction over Visual Primitives
Fuente:
arXiv
Salvato in:
| Autori principali: | Nguyen, Dat, Nguyen, Duc-Duy |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
VirDA: Reusing Backbone for Unsupervised Domain Adaptation with Visual Reprogramming
di: Nguyen, Duy, et al.
Pubblicazione: (2025)
di: Nguyen, Duy, et al.
Pubblicazione: (2025)
MoBind: Motion Binding for Fine-Grained IMU-Video Pose Alignment
di: Nguyen, Duc Duy, et al.
Pubblicazione: (2026)
di: Nguyen, Duc Duy, et al.
Pubblicazione: (2026)
CutPaste&Find: Efficient Multimodal Hallucination Detector with Visual-aid Knowledge Base
di: Nguyen, Cong-Duy, et al.
Pubblicazione: (2025)
di: Nguyen, Cong-Duy, et al.
Pubblicazione: (2025)
MMAP: A Multi-Magnification and Prototype-Aware Architecture for Predicting Spatial Gene Expression
di: Nguyen, Hai Dang, et al.
Pubblicazione: (2025)
di: Nguyen, Hai Dang, et al.
Pubblicazione: (2025)
Count What You Want: Exemplar Identification and Few-shot Counting of Human Actions in the Wild
di: Huang, Yifeng, et al.
Pubblicazione: (2023)
di: Huang, Yifeng, et al.
Pubblicazione: (2023)
EA-Swin: An Embedding-Agnostic Swin Transformer for AI-Generated Video Detection
di: Mai, Hung, et al.
Pubblicazione: (2026)
di: Mai, Hung, et al.
Pubblicazione: (2026)
AutoViVQA: A Large-Scale Automatically Constructed Dataset for Vietnamese Visual Question Answering
di: Tuong, Nguyen Anh, et al.
Pubblicazione: (2026)
di: Tuong, Nguyen Anh, et al.
Pubblicazione: (2026)
Insect-Foundation: A Foundation Model and Large-scale 1M Dataset for Visual Insect Understanding
di: Nguyen, Hoang-Quan, et al.
Pubblicazione: (2023)
di: Nguyen, Hoang-Quan, et al.
Pubblicazione: (2023)
A Dual-Module Denoising Approach with Curriculum Learning for Enhancing Multimodal Aspect-Based Sentiment Analysis
di: Van Doan, Nguyen, et al.
Pubblicazione: (2024)
di: Van Doan, Nguyen, et al.
Pubblicazione: (2024)
Semi-supervised 3D Semantic Scene Completion with 2D Vision Foundation Model Guidance
di: Pham, Duc-Hai, et al.
Pubblicazione: (2024)
di: Pham, Duc-Hai, et al.
Pubblicazione: (2024)
A model-agnostic active learning approach for animal detection from camera traps
di: Nguyen, Thi Thu Thuy, et al.
Pubblicazione: (2025)
di: Nguyen, Thi Thu Thuy, et al.
Pubblicazione: (2025)
The Art of Camouflage: Few-Shot Learning for Animal Detection and Segmentation
di: Nguyen, Thanh-Danh, et al.
Pubblicazione: (2023)
di: Nguyen, Thanh-Danh, et al.
Pubblicazione: (2023)
GenKOL: Modular Generative AI Framework For Scalable Virtual KOL Generation
di: To, Tan-Hiep, et al.
Pubblicazione: (2025)
di: To, Tan-Hiep, et al.
Pubblicazione: (2025)
Learning to Stop Overthinking at Test Time
di: Bao, Hieu Tran, et al.
Pubblicazione: (2025)
di: Bao, Hieu Tran, et al.
Pubblicazione: (2025)
UniSemAlign: Text-Prototype Alignment with a Foundation Encoder for Semi-Supervised Histopathology Segmentation
di: Thai, Le-Van, et al.
Pubblicazione: (2026)
di: Thai, Le-Van, et al.
Pubblicazione: (2026)
Unsupervised Domain Adaptation with SAM-RefiSeR for Enhanced Brain Tumor Segmentation
di: Imans, Dillan, et al.
Pubblicazione: (2026)
di: Imans, Dillan, et al.
Pubblicazione: (2026)
h-Edit: Effective and Flexible Diffusion-Based Editing via Doob's h-Transform
di: Nguyen, Toan, et al.
Pubblicazione: (2025)
di: Nguyen, Toan, et al.
Pubblicazione: (2025)
SHREC 2025: Retrieval of Optimal Objects for Multi-modal Enhanced Language and Spatial Assistance (ROOMELSA)
di: Nguyen, Trong-Thuan, et al.
Pubblicazione: (2025)
di: Nguyen, Trong-Thuan, et al.
Pubblicazione: (2025)
Adaptive Cache Enhancement for Test-Time Adaptation of Vision-Language Models
di: Nguyen, Khanh-Binh, et al.
Pubblicazione: (2025)
di: Nguyen, Khanh-Binh, et al.
Pubblicazione: (2025)
Frequency Adapter with SAM for Generalized Medical Image Segmentation
di: Bui, Phuoc-Nguyen, et al.
Pubblicazione: (2026)
di: Bui, Phuoc-Nguyen, et al.
Pubblicazione: (2026)
Towards multi-modal forgery representation learning for AI-generated video detection and localization
di: Le, Dat, et al.
Pubblicazione: (2026)
di: Le, Dat, et al.
Pubblicazione: (2026)
Multi-view Action Recognition via Directed Gromov-Wasserstein Discrepancy
di: Nguyen, Hoang-Quan, et al.
Pubblicazione: (2024)
di: Nguyen, Hoang-Quan, et al.
Pubblicazione: (2024)
FakeFormer: Efficient Vulnerability-Driven Transformers for Generalisable Deepfake Detection
di: Nguyen, Dat, et al.
Pubblicazione: (2024)
di: Nguyen, Dat, et al.
Pubblicazione: (2024)
Audio-3DVG: Unified Audio -- Point Cloud Fusion for 3D Visual Grounding
di: Cao-Dinh, Duc, et al.
Pubblicazione: (2025)
di: Cao-Dinh, Duc, et al.
Pubblicazione: (2025)
Representation Learning with Semantic-aware Instance and Sparse Token Alignments
di: Bui, Phuoc-Nguyen, et al.
Pubblicazione: (2026)
di: Bui, Phuoc-Nguyen, et al.
Pubblicazione: (2026)
Improving Imbalanced Multi-Label Chest X-Ray Diagnosis via CBAM-Enhanced CNN Backbones
di: Huu, Duy Nguyen, et al.
Pubblicazione: (2026)
di: Huu, Duy Nguyen, et al.
Pubblicazione: (2026)
Momentum-Anchored Multi-Scale Fusion Model for Long-Tailed Chest X-Ray Classification
di: Khuong, Duy Hoang, et al.
Pubblicazione: (2026)
di: Khuong, Duy Hoang, et al.
Pubblicazione: (2026)
S-Chain: Structured Visual Chain-of-Thought For Medicine
di: Le-Duc, Khai, et al.
Pubblicazione: (2025)
di: Le-Duc, Khai, et al.
Pubblicazione: (2025)
FVO: Fast Visual Odometry with Transformers
di: Yugay, Vlardimir, et al.
Pubblicazione: (2025)
di: Yugay, Vlardimir, et al.
Pubblicazione: (2025)
YOWOv3: An Efficient and Generalized Framework for Human Action Detection and Recognition
di: Dang, Duc Manh Nguyen, et al.
Pubblicazione: (2024)
di: Dang, Duc Manh Nguyen, et al.
Pubblicazione: (2024)
WIPES: Wavelet-based Visual Primitives
di: Zhang, Wenhao, et al.
Pubblicazione: (2025)
di: Zhang, Wenhao, et al.
Pubblicazione: (2025)
Stratified Domain Adaptation: A Progressive Self-Training Approach for Scene Text Recognition
di: Le, Kha Nhat, et al.
Pubblicazione: (2024)
di: Le, Kha Nhat, et al.
Pubblicazione: (2024)
Reperio-rPPG: Relational Temporal Graph Neural Networks for Periodicity Learning in Remote Physiological Measurement
di: Nguyen, Ba-Thinh, et al.
Pubblicazione: (2025)
di: Nguyen, Ba-Thinh, et al.
Pubblicazione: (2025)
Motion-aware Contrastive Learning for Temporal Panoptic Scene Graph Generation
di: Nguyen, Thong Thanh, et al.
Pubblicazione: (2024)
di: Nguyen, Thong Thanh, et al.
Pubblicazione: (2024)
Fast Online 3D Multi-Camera Multi-Object Tracking and Pose Estimation
di: Van Ma, Linh, et al.
Pubblicazione: (2026)
di: Van Ma, Linh, et al.
Pubblicazione: (2026)
Detection Fire in Camera RGB-NIR
di: Khai, Nguyen Truong, et al.
Pubblicazione: (2025)
di: Khai, Nguyen Truong, et al.
Pubblicazione: (2025)
Guiding Noisy Label Conditional Diffusion Models with Score-based Discriminator Correction
di: Cong, Dat Nguyen, et al.
Pubblicazione: (2025)
di: Cong, Dat Nguyen, et al.
Pubblicazione: (2025)
Dual Strategies for Test-Time Adaptation
di: Phuong, Nam Nguyen, et al.
Pubblicazione: (2026)
di: Phuong, Nam Nguyen, et al.
Pubblicazione: (2026)
RangeSAM: On the Potential of Visual Foundation Models for Range-View represented LiDAR segmentation
di: Kühn, Paul Julius, et al.
Pubblicazione: (2025)
di: Kühn, Paul Julius, et al.
Pubblicazione: (2025)
FlexEdit: Flexible and Controllable Diffusion-based Object-centric Image Editing
di: Nguyen, Trong-Tung, et al.
Pubblicazione: (2024)
di: Nguyen, Trong-Tung, et al.
Pubblicazione: (2024)
Documenti analoghi
-
VirDA: Reusing Backbone for Unsupervised Domain Adaptation with Visual Reprogramming
di: Nguyen, Duy, et al.
Pubblicazione: (2025) -
MoBind: Motion Binding for Fine-Grained IMU-Video Pose Alignment
di: Nguyen, Duc Duy, et al.
Pubblicazione: (2026) -
CutPaste&Find: Efficient Multimodal Hallucination Detector with Visual-aid Knowledge Base
di: Nguyen, Cong-Duy, et al.
Pubblicazione: (2025) -
MMAP: A Multi-Magnification and Prototype-Aware Architecture for Predicting Spatial Gene Expression
di: Nguyen, Hai Dang, et al.
Pubblicazione: (2025) -
Count What You Want: Exemplar Identification and Few-shot Counting of Human Actions in the Wild
di: Huang, Yifeng, et al.
Pubblicazione: (2023)