PAT: Pixel-wise Adaptive Training for Long-tailed Segmentation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Do, Khoi, Nguyen, Duong, Tran, Nguyen H., Nguyen, Viet Dung |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Phantasia: Context-Adaptive Backdoors in Vision Language Models
von: Tran, Nam Duong, et al.
Veröffentlicht: (2026)
von: Tran, Nam Duong, et al.
Veröffentlicht: (2026)
OE3DIS: Open-Ended 3D Point Cloud Instance Segmentation
von: Nguyen, Phuc D. A., et al.
Veröffentlicht: (2024)
von: Nguyen, Phuc D. A., et al.
Veröffentlicht: (2024)
CSD-VAR: Content-Style Decomposition in Visual Autoregressive Models
von: Nguyen, Quang-Binh, et al.
Veröffentlicht: (2025)
von: Nguyen, Quang-Binh, et al.
Veröffentlicht: (2025)
SwiftTry: Fast and Consistent Video Virtual Try-On with Diffusion Models
von: Nguyen, Hung, et al.
Veröffentlicht: (2024)
von: Nguyen, Hung, et al.
Veröffentlicht: (2024)
InverFill: One-Step Inversion for Enhanced Few-Step Diffusion Inpainting
von: Vu, Duc, et al.
Veröffentlicht: (2026)
von: Vu, Duc, et al.
Veröffentlicht: (2026)
VEIGAR: View-consistent Explicit Inpainting and Geometry Alignment for 3D object Removal
von: Do, Pham Khai Nguyen, et al.
Veröffentlicht: (2025)
von: Do, Pham Khai Nguyen, et al.
Veröffentlicht: (2025)
MasHeNe: A Benchmark for Head and Neck CT Mass Segmentation using Window-Enhanced Mamba with Frequency-Domain Integration
von: Dao, Thao Thi Phuong, et al.
Veröffentlicht: (2025)
von: Dao, Thao Thi Phuong, et al.
Veröffentlicht: (2025)
PANDORA: Pixel-wise Attention Dissolution and Latent Guidance for Zero-Shot Object Removal
von: Vo, Dinh-Khoi, et al.
Veröffentlicht: (2026)
von: Vo, Dinh-Khoi, et al.
Veröffentlicht: (2026)
Enhancing Multimodal Entity Linking with Jaccard Distance-based Conditional Contrastive Learning and Contextual Visual Augmentation
von: Nguyen, Cong-Duy, et al.
Veröffentlicht: (2025)
von: Nguyen, Cong-Duy, et al.
Veröffentlicht: (2025)
Comparing Deep Neural Network for Multi-Label ECG Diagnosis From Scanned ECG
von: Nguyen, Cuong V., et al.
Veröffentlicht: (2025)
von: Nguyen, Cuong V., et al.
Veröffentlicht: (2025)
ConstStyle: Robust Domain Generalization with Unified Style Transformation
von: Tran, Nam Duong, et al.
Veröffentlicht: (2025)
von: Tran, Nam Duong, et al.
Veröffentlicht: (2025)
SwiftBrush v2: Make Your One-step Diffusion Model Better Than Its Teacher
von: Dao, Trung, et al.
Veröffentlicht: (2024)
von: Dao, Trung, et al.
Veröffentlicht: (2024)
IGL-DT: Iterative Global-Local Feature Learning with Dual-Teacher Semantic Segmentation Framework under Limited Annotation Scheme
von: Tran, Dinh Dai Quan, et al.
Veröffentlicht: (2025)
von: Tran, Dinh Dai Quan, et al.
Veröffentlicht: (2025)
How Homogenizing the Channel-wise Magnitude Can Enhance EEG Classification Model?
von: Ngo, Huyen, et al.
Veröffentlicht: (2024)
von: Ngo, Huyen, et al.
Veröffentlicht: (2024)
Aleatoric Uncertainty Medical Image Segmentation Estimation via Flow Matching
von: Van Nguyen, Phi, et al.
Veröffentlicht: (2025)
von: Van Nguyen, Phi, et al.
Veröffentlicht: (2025)
Bidirectional Diffusion Bridge Models
von: Kieu, Duc, et al.
Veröffentlicht: (2025)
von: Kieu, Duc, et al.
Veröffentlicht: (2025)
SegMaFormer: A Hybrid State-Space and Transformer Model for Efficient Segmentation
von: Nguyen, Duy D., et al.
Veröffentlicht: (2026)
von: Nguyen, Duy D., et al.
Veröffentlicht: (2026)
AC-MAMBASEG: An adaptive convolution and Mamba-based architecture for enhanced skin lesion segmentation
von: Nguyen, Viet-Thanh, et al.
Veröffentlicht: (2024)
von: Nguyen, Viet-Thanh, et al.
Veröffentlicht: (2024)
Semi-Supervised Semantic Segmentation using Redesigned Self-Training for White Blood Cells
von: Luu, Vinh Quoc, et al.
Veröffentlicht: (2024)
von: Luu, Vinh Quoc, et al.
Veröffentlicht: (2024)
Not All Pixels Are Equal: Pixel-wise Meta-Learning for Medical Segmentation with Noisy Labels
von: Mu, Chenyu, et al.
Veröffentlicht: (2025)
von: Mu, Chenyu, et al.
Veröffentlicht: (2025)
SparseSAM: Structured Sparsification of Activations in Segment Anything Models
von: Tran, Hoai-Chau, et al.
Veröffentlicht: (2026)
von: Tran, Hoai-Chau, et al.
Veröffentlicht: (2026)
AUCSeg: AUC-oriented Pixel-level Long-tail Semantic Segmentation
von: Han, Boyu, et al.
Veröffentlicht: (2024)
von: Han, Boyu, et al.
Veröffentlicht: (2024)
V-Math: An Agentic Approach to the Vietnamese National High School Graduation Mathematics Exams
von: Nguyen, Duong Q., et al.
Veröffentlicht: (2025)
von: Nguyen, Duong Q., et al.
Veröffentlicht: (2025)
Robustness Evaluation of OCR-based Visual Document Understanding under Multi-Modal Adversarial Attacks
von: Tien, Dong Nguyen, et al.
Veröffentlicht: (2025)
von: Tien, Dong Nguyen, et al.
Veröffentlicht: (2025)
Federated Prompt-Tuning with Heterogeneous and Incomplete Multimodal Client Data
von: Phung, Thu Hang, et al.
Veröffentlicht: (2026)
von: Phung, Thu Hang, et al.
Veröffentlicht: (2026)
MedSteer: Counterfactual Endoscopic Synthesis via Training-Free Activation Steering
von: Pham, Trong-Thang, et al.
Veröffentlicht: (2026)
von: Pham, Trong-Thang, et al.
Veröffentlicht: (2026)
Supercharged One-step Text-to-Image Diffusion Models with Negative Prompts
von: Nguyen, Viet, et al.
Veröffentlicht: (2024)
von: Nguyen, Viet, et al.
Veröffentlicht: (2024)
HDC: Hierarchical Distillation for Multi-level Noisy Consistency in Semi-Supervised Fetal Ultrasound Segmentation
von: Le, Tran Quoc Khanh, et al.
Veröffentlicht: (2025)
von: Le, Tran Quoc Khanh, et al.
Veröffentlicht: (2025)
Training Deep Visual Networks Beyond Loss and Accuracy Through a Dynamical Systems Approach
von: La Quang, Hai, et al.
Veröffentlicht: (2026)
von: La Quang, Hai, et al.
Veröffentlicht: (2026)
ViCLIP-OT: The First Foundation Vision-Language Model for Vietnamese Image-Text Retrieval with Optimal Transport
von: Tran, Quoc-Khang, et al.
Veröffentlicht: (2026)
von: Tran, Quoc-Khang, et al.
Veröffentlicht: (2026)
Text-to-3D Generation using Jensen-Shannon Score Distillation
von: Do, Khoi, et al.
Veröffentlicht: (2025)
von: Do, Khoi, et al.
Veröffentlicht: (2025)
CGCE: Classifier-Guided Concept Erasure in Generative Models
von: Nguyen, Viet, et al.
Veröffentlicht: (2025)
von: Nguyen, Viet, et al.
Veröffentlicht: (2025)
Brain Tumor Segmentation in MRI Images with 3D U-Net and Contextual Transformer
von: Nguyen, Thien-Qua T., et al.
Veröffentlicht: (2024)
von: Nguyen, Thien-Qua T., et al.
Veröffentlicht: (2024)
Vision-Aware Text Features in Referring Image Segmentation: From Object Understanding to Context Understanding
von: Nguyen-Truong, Hai, et al.
Veröffentlicht: (2024)
von: Nguyen-Truong, Hai, et al.
Veröffentlicht: (2024)
SUGAR: A Sweeter Spot for Generative Unlearning of Many Identities
von: Nguyen, Dung Thuy, et al.
Veröffentlicht: (2025)
von: Nguyen, Dung Thuy, et al.
Veröffentlicht: (2025)
Any3DIS: Class-Agnostic 3D Instance Segmentation by 2D Mask Tracking
von: Nguyen, Phuc, et al.
Veröffentlicht: (2024)
von: Nguyen, Phuc, et al.
Veröffentlicht: (2024)
Anti-I2V: Safeguarding your photos from malicious image-to-video generation
von: Vu, Duc, et al.
Veröffentlicht: (2026)
von: Vu, Duc, et al.
Veröffentlicht: (2026)
Diffusion Model in Latent Space for Medical Image Segmentation Task
von: Ngoc, Huynh Trinh, et al.
Veröffentlicht: (2025)
von: Ngoc, Huynh Trinh, et al.
Veröffentlicht: (2025)
FurniMAS: Language-Guided Furniture Decoration using Multi-Agent System
von: Nguyen, Toan, et al.
Veröffentlicht: (2025)
von: Nguyen, Toan, et al.
Veröffentlicht: (2025)
Region-Grounded Report Generation for 3D Medical Imaging: A Fine-Grained Dataset and Graph-Enhanced Framework
von: Nguyen, Cong Huy, et al.
Veröffentlicht: (2026)
von: Nguyen, Cong Huy, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Phantasia: Context-Adaptive Backdoors in Vision Language Models
von: Tran, Nam Duong, et al.
Veröffentlicht: (2026) -
OE3DIS: Open-Ended 3D Point Cloud Instance Segmentation
von: Nguyen, Phuc D. A., et al.
Veröffentlicht: (2024) -
CSD-VAR: Content-Style Decomposition in Visual Autoregressive Models
von: Nguyen, Quang-Binh, et al.
Veröffentlicht: (2025) -
SwiftTry: Fast and Consistent Video Virtual Try-On with Diffusion Models
von: Nguyen, Hung, et al.
Veröffentlicht: (2024) -
InverFill: One-Step Inversion for Enhanced Few-Step Diffusion Inpainting
von: Vu, Duc, et al.
Veröffentlicht: (2026)