F-ViTA: Foundation Model Guided Visible to Thermal Translation
Fuente:
arXiv
Guardado en:
| Autores principales: | Paranjape, Jay N., de Melo, Celso, Patel, Vishal M. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
A Mamba-based Siamese Network for Remote Sensing Change Detection
por: Paranjape, Jay N., et al.
Publicado: (2024)
por: Paranjape, Jay N., et al.
Publicado: (2024)
ViTA-Seg: Vision Transformer for Amodal Segmentation in Robotics
por: Caramia, Donato, et al.
Publicado: (2025)
por: Caramia, Donato, et al.
Publicado: (2025)
Referring Change Detection in Remote Sensing Imagery
por: Korkmaz, Yilmaz, et al.
Publicado: (2025)
por: Korkmaz, Yilmaz, et al.
Publicado: (2025)
Thermo-VL: Extending Vision-Language Models to Thermal Infrared Perception
por: Thushara, Rusiru, et al.
Publicado: (2026)
por: Thushara, Rusiru, et al.
Publicado: (2026)
ViTA-PAR: Visual and Textual Attribute Alignment with Attribute Prompting for Pedestrian Attribute Recognition
por: Park, Minjeong, et al.
Publicado: (2025)
por: Park, Minjeong, et al.
Publicado: (2025)
S-SAM: SVD-based Fine-Tuning of Segment Anything Model for Medical Image Segmentation
por: Paranjape, Jay N., et al.
Publicado: (2024)
por: Paranjape, Jay N., et al.
Publicado: (2024)
Blackbox Adaptation for Medical Image Segmentation
por: Paranjape, Jay N., et al.
Publicado: (2024)
por: Paranjape, Jay N., et al.
Publicado: (2024)
Federated Black-Box Adaptation for Semantic Segmentation
por: Paranjape, Jay N., et al.
Publicado: (2024)
por: Paranjape, Jay N., et al.
Publicado: (2024)
GenDeg: Diffusion-based Degradation Synthesis for Generalizable All-In-One Image Restoration
por: Rajagopalan, Sudarshan, et al.
Publicado: (2024)
por: Rajagopalan, Sudarshan, et al.
Publicado: (2024)
Zero-Shot Scene Understanding for Automatic Target Recognition Using Large Vision-Language Models
por: Ranasinghe, Yasiru, et al.
Publicado: (2025)
por: Ranasinghe, Yasiru, et al.
Publicado: (2025)
FreeViS: Training-free Video Stylization with Inconsistent References
por: Xu, Jiacong, et al.
Publicado: (2025)
por: Xu, Jiacong, et al.
Publicado: (2025)
Thermal-Det: Language-Guided Cross-Modal Distillation for Open-Vocabulary Thermal Object Detection
por: Ranasinghe, Yasiru, et al.
Publicado: (2026)
por: Ranasinghe, Yasiru, et al.
Publicado: (2026)
Certainty and Uncertainty Guided Active Domain Adaptation
por: Safaei, Bardia, et al.
Publicado: (2025)
por: Safaei, Bardia, et al.
Publicado: (2025)
ViLReF: An Expert Knowledge Enabled Vision-Language Retinal Foundation Model
por: Yang, Shengzhu, et al.
Publicado: (2024)
por: Yang, Shengzhu, et al.
Publicado: (2024)
CGCE: Classifier-Guided Concept Erasure in Generative Models
por: Nguyen, Viet, et al.
Publicado: (2025)
por: Nguyen, Viet, et al.
Publicado: (2025)
Active Learning for Vision-Language Models
por: Safaei, Bardia, et al.
Publicado: (2024)
por: Safaei, Bardia, et al.
Publicado: (2024)
Silhouette-based Gait Foundation Model
por: Ye, Dingqiang, et al.
Publicado: (2025)
por: Ye, Dingqiang, et al.
Publicado: (2025)
ModelMix: A New Model-Mixup Strategy to Minimize Vicinal Risk across Tasks for Few-scribble based Cardiac Segmentation
por: Zhang, Ke, et al.
Publicado: (2024)
por: Zhang, Ke, et al.
Publicado: (2024)
RemoteVAR: Autoregressive Visual Modeling for Remote Sensing Change Detection
por: Korkmaz, Yilmaz, et al.
Publicado: (2026)
por: Korkmaz, Yilmaz, et al.
Publicado: (2026)
I2I-Galip: Unsupervised Medical Image Translation Using Generative Adversarial CLIP
por: Korkmaz, Yilmaz, et al.
Publicado: (2024)
por: Korkmaz, Yilmaz, et al.
Publicado: (2024)
CanViT: Toward Active-Vision Foundation Models
por: Berreby, Yohaï-Eliel, et al.
Publicado: (2026)
por: Berreby, Yohaï-Eliel, et al.
Publicado: (2026)
Attention Prompt Tuning: Parameter-efficient Adaptation of Pre-trained Models for Spatiotemporal Modeling
por: Bandara, Wele Gedara Chaminda, et al.
Publicado: (2024)
por: Bandara, Wele Gedara Chaminda, et al.
Publicado: (2024)
ViLCo-Bench: VIdeo Language COntinual learning Benchmark
por: Tang, Tianqi, et al.
Publicado: (2024)
por: Tang, Tianqi, et al.
Publicado: (2024)
MedCL: Learning Consistent Anatomy Distribution for Scribble-supervised Medical Image Segmentation
por: Zhang, Ke, et al.
Publicado: (2025)
por: Zhang, Ke, et al.
Publicado: (2025)
AWRaCLe: All-Weather Image Restoration using Visual In-Context Learning
por: Rajagopalan, Sudarshan, et al.
Publicado: (2024)
por: Rajagopalan, Sudarshan, et al.
Publicado: (2024)
Not All Tokens Need 40 Steps: Heterogeneous Step Allocation in Diffusion Transformers for Efficient Video Generation
por: Chu, Ernie, et al.
Publicado: (2026)
por: Chu, Ernie, et al.
Publicado: (2026)
Low-rank Adaptation-based All-Weather Removal for Autonomous Navigation
por: Rajagopalan, Sudarshan, et al.
Publicado: (2024)
por: Rajagopalan, Sudarshan, et al.
Publicado: (2024)
Hyp-OC: Hyperbolic One Class Classification for Face Anti-Spoofing
por: Narayan, Kartik, et al.
Publicado: (2024)
por: Narayan, Kartik, et al.
Publicado: (2024)
Implicit Neural Representations: A Signal Processing Perspective
por: Jayasundara, Dhananjaya, et al.
Publicado: (2026)
por: Jayasundara, Dhananjaya, et al.
Publicado: (2026)
Leveraging Thermal Modality to Enhance Reconstruction in Low-Light Conditions
por: Xu, Jiacong, et al.
Publicado: (2024)
por: Xu, Jiacong, et al.
Publicado: (2024)
Your Pre-trained Diffusion Model Secretly Knows Restoration
por: Rajagopalan, Sudarshan, et al.
Publicado: (2026)
por: Rajagopalan, Sudarshan, et al.
Publicado: (2026)
Face-to-Face: A Video Dataset for Multi-Person Interaction Modeling
por: Chu, Ernie, et al.
Publicado: (2026)
por: Chu, Ernie, et al.
Publicado: (2026)
UNIV: Unified Foundation Model for Infrared and Visible Modalities
por: Mao, Fangyuan, et al.
Publicado: (2025)
por: Mao, Fangyuan, et al.
Publicado: (2025)
Frame by Familiar Frame: Understanding Replication in Video Diffusion Models
por: Rahman, Aimon, et al.
Publicado: (2024)
por: Rahman, Aimon, et al.
Publicado: (2024)
Dreamguider: Improved Training free Diffusion-based Conditional Generation
por: Nair, Nithin Gopalakrishnan, et al.
Publicado: (2024)
por: Nair, Nithin Gopalakrishnan, et al.
Publicado: (2024)
Latent Feature-Guided Diffusion Models for Shadow Removal
por: Mei, Kangfu, et al.
Publicado: (2023)
por: Mei, Kangfu, et al.
Publicado: (2023)
Supervised Image Translation from Visible to Infrared Domain for Object Detection
por: Anand, Prahlad, et al.
Publicado: (2024)
por: Anand, Prahlad, et al.
Publicado: (2024)
MambaRecon: MRI Reconstruction with Structured State Space Models
por: Korkmaz, Yilmaz, et al.
Publicado: (2024)
por: Korkmaz, Yilmaz, et al.
Publicado: (2024)
ViT-Split: Unleashing the Power of Vision Foundation Models via Efficient Splitting Heads
por: Li, Yifan, et al.
Publicado: (2025)
por: Li, Yifan, et al.
Publicado: (2025)
FaceXBench: Evaluating Multimodal LLMs on Face Understanding
por: Narayan, Kartik, et al.
Publicado: (2025)
por: Narayan, Kartik, et al.
Publicado: (2025)
Ejemplares similares
-
A Mamba-based Siamese Network for Remote Sensing Change Detection
por: Paranjape, Jay N., et al.
Publicado: (2024) -
ViTA-Seg: Vision Transformer for Amodal Segmentation in Robotics
por: Caramia, Donato, et al.
Publicado: (2025) -
Referring Change Detection in Remote Sensing Imagery
por: Korkmaz, Yilmaz, et al.
Publicado: (2025) -
Thermo-VL: Extending Vision-Language Models to Thermal Infrared Perception
por: Thushara, Rusiru, et al.
Publicado: (2026) -
ViTA-PAR: Visual and Textual Attribute Alignment with Attribute Prompting for Pedestrian Attribute Recognition
por: Park, Minjeong, et al.
Publicado: (2025)