Fourier-Attentive Representation Learning: A Fourier-Guided Framework for Few-Shot Generalization in Vision-Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Pham, Hieu Dinh Trung, Nguyen, Huy Minh Nhat, Nguyen, Cuong Tuan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Link prediction Graph Neural Networks for structure recognition of Handwritten Mathematical Expressions
by: Nguyen, Cuong Tuan, et al.
Published: (2025)
by: Nguyen, Cuong Tuan, et al.
Published: (2025)
SDPA++: A General Framework for Self-Supervised Denoising with Patch Aggregation
by: Nguyen, Huy Minh Nhat, et al.
Published: (2025)
by: Nguyen, Huy Minh Nhat, et al.
Published: (2025)
Handling Supervision Scarcity in Chest X-ray Classification: Long-Tailed and Zero-Shot Learning
by: Pham, Ha-Hieu, et al.
Published: (2026)
by: Pham, Ha-Hieu, et al.
Published: (2026)
Predictive Spectral Calibration for Source-Free Test-Time Regression
by: Kiet, Nguyen Viet Tuan, et al.
Published: (2026)
by: Kiet, Nguyen Viet Tuan, et al.
Published: (2026)
VisionGuard: Synergistic Framework for Helmet Violation Detection
by: Nguyen, Lam-Huy, et al.
Published: (2025)
by: Nguyen, Lam-Huy, et al.
Published: (2025)
MGPATH: Vision-Language Model with Multi-Granular Prompt Learning for Few-Shot WSI Classification
by: Nguyen, Anh-Tien, et al.
Published: (2025)
by: Nguyen, Anh-Tien, et al.
Published: (2025)
Learning Disentangled Stain and Structural Representations for Semi-Supervised Histopathology Segmentation
by: Pham, Ha-Hieu, et al.
Published: (2025)
by: Pham, Ha-Hieu, et al.
Published: (2025)
Unlocking Compositional Generalization in Continual Few-Shot Learning
by: Nguyen-Lam, Phu-Quy, et al.
Published: (2026)
by: Nguyen-Lam, Phu-Quy, et al.
Published: (2026)
Toward a Vision-Language Foundation Model for Medical Data: Multimodal Dataset and Benchmarks for Vietnamese PET/CT Report Generation
by: Nguyen, Huu Tien, et al.
Published: (2025)
by: Nguyen, Huu Tien, et al.
Published: (2025)
Med-StepBench: A Hierarchical Reasoning Framework for Evaluating Hallucinations in Medical Vision-Language Models
by: Nguyen, Minh Khoi, et al.
Published: (2026)
by: Nguyen, Minh Khoi, et al.
Published: (2026)
GraspMamba: A Mamba-based Language-driven Grasp Detection Framework with Hierarchical Feature Learning
by: Nguyen, Huy Hoang, et al.
Published: (2024)
by: Nguyen, Huy Hoang, et al.
Published: (2024)
Variational Autoencoder for Anomaly Detection: A Comparative Study
by: Nguyen, Huy Hoang, et al.
Published: (2024)
by: Nguyen, Huy Hoang, et al.
Published: (2024)
Language-driven Grasp Detection
by: Vuong, An Dinh, et al.
Published: (2024)
by: Vuong, An Dinh, et al.
Published: (2024)
The Art of Camouflage: Few-Shot Learning for Animal Detection and Segmentation
by: Nguyen, Thanh-Danh, et al.
Published: (2023)
by: Nguyen, Thanh-Danh, et al.
Published: (2023)
Model and Feature Diversity for Bayesian Neural Networks in Mutual Learning
by: Pham, Cuong, et al.
Published: (2024)
by: Pham, Cuong, et al.
Published: (2024)
Comparing Deep Neural Network for Multi-Label ECG Diagnosis From Scanned ECG
by: Nguyen, Cuong V., et al.
Published: (2025)
by: Nguyen, Cuong V., et al.
Published: (2025)
Bridging Classification and Segmentation in Osteosarcoma Assessment via Foundation and Discrete Diffusion Models
by: Nguyen, Manh Duong, et al.
Published: (2025)
by: Nguyen, Manh Duong, et al.
Published: (2025)
Evaluating Precise Geolocation Inference Capabilities of Vision Language Models
by: Jay, Neel, et al.
Published: (2025)
by: Jay, Neel, et al.
Published: (2025)
VietMEAgent: Culturally-Aware Few-Shot Multimodal Explanation for Vietnamese Visual Question Answering
by: Nguyen, Hai-Dang, et al.
Published: (2025)
by: Nguyen, Hai-Dang, et al.
Published: (2025)
CamoFA: A Learnable Fourier-based Augmentation for Camouflage Segmentation
by: Le, Minh-Quan, et al.
Published: (2023)
by: Le, Minh-Quan, et al.
Published: (2023)
NeIn: Telling What You Don't Want
by: Bui, Nhat-Tan, et al.
Published: (2024)
by: Bui, Nhat-Tan, et al.
Published: (2024)
Overthinking Causes Hallucination: Tracing Confounder Propagation in Vision Language Models
by: Shoby, Abin, et al.
Published: (2026)
by: Shoby, Abin, et al.
Published: (2026)
Count What You Want: Exemplar Identification and Few-shot Counting of Human Actions in the Wild
by: Huang, Yifeng, et al.
Published: (2023)
by: Huang, Yifeng, et al.
Published: (2023)
Deep-Wide Learning Assistance for Insect Pest Classification
by: Nguyen, Toan, et al.
Published: (2024)
by: Nguyen, Toan, et al.
Published: (2024)
CT to PET Translation: A Large-scale Dataset and Domain-Knowledge-Guided Diffusion Approach
by: Nguyen, Dac Thai, et al.
Published: (2024)
by: Nguyen, Dac Thai, et al.
Published: (2024)
MaskDiff: Modeling Mask Distribution with Diffusion Probabilistic Model for Few-Shot Instance Segmentation
by: Le, Minh-Quan, et al.
Published: (2023)
by: Le, Minh-Quan, et al.
Published: (2023)
SwiftPie: Lightning-fast Subject-driven Image Personalization via One step Diffusion
by: Duong, Huy, et al.
Published: (2026)
by: Duong, Huy, et al.
Published: (2026)
ReCap: Event-Aware Image Captioning with Article Retrieval and Semantic Gaussian Normalization
by: Nguyen, Thinh-Phuc, et al.
Published: (2025)
by: Nguyen, Thinh-Phuc, et al.
Published: (2025)
LoG-VMamba: Local-Global Vision Mamba for Medical Image Segmentation
by: Dang, Trung Dinh Quoc, et al.
Published: (2024)
by: Dang, Trung Dinh Quoc, et al.
Published: (2024)
Provably Improving Generalization of Few-Shot Models with Synthetic Data
by: Nguyen, Lan-Cuong, et al.
Published: (2025)
by: Nguyen, Lan-Cuong, et al.
Published: (2025)
Digital FAST: An AI-Driven Multimodal Framework for Rapid and Early Stroke Screening
by: Hoang, Ngoc-Khai, et al.
Published: (2026)
by: Hoang, Ngoc-Khai, et al.
Published: (2026)
Towards Efficient and Robust Moment Retrieval System: A Unified Framework for Multi-Granularity Models and Temporal Reranking
by: Tran, Huu-Loc, et al.
Published: (2025)
by: Tran, Huu-Loc, et al.
Published: (2025)
Supercharged One-step Text-to-Image Diffusion Models with Negative Prompts
by: Nguyen, Viet, et al.
Published: (2024)
by: Nguyen, Viet, et al.
Published: (2024)
PANDORA: Pixel-wise Attention Dissolution and Latent Guidance for Zero-Shot Object Removal
by: Vo, Dinh-Khoi, et al.
Published: (2026)
by: Vo, Dinh-Khoi, et al.
Published: (2026)
VinDr-CXR-VQA: A Visual Question Answering Dataset for Explainable Chest X-Ray Analysis with Multi-Task Learning
by: Nguyen, Dang H., et al.
Published: (2025)
by: Nguyen, Dang H., et al.
Published: (2025)
Synergizing Deep Learning and Biological Heuristics for Extreme Long-Tail White Blood Cell Classification
by: Nguyen, Duc T., et al.
Published: (2026)
by: Nguyen, Duc T., et al.
Published: (2026)
DoRAN: Stabilizing Weight-Decomposed Low-Rank Adaptation via Noise Injection and Auxiliary Networks
by: Diep, Nghiem T., et al.
Published: (2025)
by: Diep, Nghiem T., et al.
Published: (2025)
MELEP: A Novel Predictive Measure of Transferability in Multi-Label ECG Diagnosis
by: Nguyen, Cuong V., et al.
Published: (2023)
by: Nguyen, Cuong V., et al.
Published: (2023)
Fusionista2.0: Efficiency Retrieval System for Large-Scale Datasets
by: Le, Huy M., et al.
Published: (2025)
by: Le, Huy M., et al.
Published: (2025)
Any3DIS: Class-Agnostic 3D Instance Segmentation by 2D Mask Tracking
by: Nguyen, Phuc, et al.
Published: (2024)
by: Nguyen, Phuc, et al.
Published: (2024)
Similar Items
-
Link prediction Graph Neural Networks for structure recognition of Handwritten Mathematical Expressions
by: Nguyen, Cuong Tuan, et al.
Published: (2025) -
SDPA++: A General Framework for Self-Supervised Denoising with Patch Aggregation
by: Nguyen, Huy Minh Nhat, et al.
Published: (2025) -
Handling Supervision Scarcity in Chest X-ray Classification: Long-Tailed and Zero-Shot Learning
by: Pham, Ha-Hieu, et al.
Published: (2026) -
Predictive Spectral Calibration for Source-Free Test-Time Regression
by: Kiet, Nguyen Viet Tuan, et al.
Published: (2026) -
VisionGuard: Synergistic Framework for Helmet Violation Detection
by: Nguyen, Lam-Huy, et al.
Published: (2025)