Enregistré dans:
| Auteurs principaux: | Shipard, Jordan, Wiliem, Arnold, Thanh, Kien Nguyen, Xiang, Wei, Fookes, Clinton |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | https://arxiv.org/abs/2401.11633 |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
OmniGCD: Abstracting Generalized Category Discovery for Modality Agnosticism
par: Shipard, Jordan, et autres
Publié: (2026)
par: Shipard, Jordan, et autres
Publié: (2026)
Filling the Gaps: A Multitask Hybrid Multiscale Generative Framework for Missing Modality in Remote Sensing Semantic Segmentation
par: Kieu, Nhi, et autres
Publié: (2025)
par: Kieu, Nhi, et autres
Publié: (2025)
DIS2: Disentanglement Meets Distillation with Classwise Attention for Robust Remote Sensing Segmentation under Missing Modalities
par: Kieu, Nhi, et autres
Publié: (2026)
par: Kieu, Nhi, et autres
Publié: (2026)
Physics-Informed Computer Vision: A Review and Perspectives
par: Banerjee, Chayan, et autres
Publié: (2023)
par: Banerjee, Chayan, et autres
Publié: (2023)
AG-ReID.v2: Bridging Aerial and Ground Views for Person Re-identification
par: Nguyen, Huy, et autres
Publié: (2024)
par: Nguyen, Huy, et autres
Publié: (2024)
Revisiting the Role of Texture in 3D Person Re-identification
par: Nguyen, Huy, et autres
Publié: (2024)
par: Nguyen, Huy, et autres
Publié: (2024)
PDV: Prompt Directional Vectors for Zero-shot Composed Image Retrieval
par: Tursun, Osman, et autres
Publié: (2025)
par: Tursun, Osman, et autres
Publié: (2025)
Ilov3Splat: Instance-Level Open-Vocabulary 3D Scene Understanding in Gaussian Splatting
par: Nguyen, Binh Long, et autres
Publié: (2026)
par: Nguyen, Binh Long, et autres
Publié: (2026)
Size and Smoothness Aware Adaptive Focal Loss for Small Tumor Segmentation
par: Islam, Md Rakibul, et autres
Publié: (2024)
par: Islam, Md Rakibul, et autres
Publié: (2024)
PINNs for Medical Image Analysis: A Survey
par: Banerjee, Chayan, et autres
Publié: (2024)
par: Banerjee, Chayan, et autres
Publié: (2024)
AG-VPReID.VIR: Bridging Aerial and Ground Platforms for Video-based Visible-Infrared Person Re-ID
par: Nguyen, Huy, et autres
Publié: (2025)
par: Nguyen, Huy, et autres
Publié: (2025)
AG-VPReID: A Challenging Large-Scale Benchmark for Aerial-Ground Video-based Person Re-Identification
par: Nguyen, Huy, et autres
Publié: (2025)
par: Nguyen, Huy, et autres
Publié: (2025)
Mining--Gym: A Configurable RL Benchmarking Environment for Truck Dispatch Scheduling
par: Banerjee, Chayan, et autres
Publié: (2025)
par: Banerjee, Chayan, et autres
Publié: (2025)
Person Recognition in Aerial Surveillance: A Decade Survey
par: Nguyen, Kien, et autres
Publié: (2025)
par: Nguyen, Kien, et autres
Publié: (2025)
Physics-Informed Operator Learning for Hemodynamic Modeling
par: Chappell, Ryan, et autres
Publié: (2025)
par: Chappell, Ryan, et autres
Publié: (2025)
A Survey on Physics Informed Reinforcement Learning: Review and Open Problems
par: Banerjee, Chayan, et autres
Publié: (2023)
par: Banerjee, Chayan, et autres
Publié: (2023)
Uncertainty in Real-Time Semantic Segmentation on Embedded Systems
par: Goan, Ethan, et autres
Publié: (2022)
par: Goan, Ethan, et autres
Publié: (2022)
MTReD: 3D Reconstruction Dataset for Fly-over Videos of Maritime Domain
par: Yong, Rui Yi, et autres
Publié: (2025)
par: Yong, Rui Yi, et autres
Publié: (2025)
DeepForgeSeal: Latent Space-Driven Semi-Fragile Watermarking for Deepfake Detection Using Multi-Agent Adversarial Reinforcement Learning
par: Fernando, Tharindu, et autres
Publié: (2025)
par: Fernando, Tharindu, et autres
Publié: (2025)
CLIP-Decoder : ZeroShot Multilabel Classification using Multimodal CLIP Aligned Representation
par: Ali, Muhammad, et autres
Publié: (2024)
par: Ali, Muhammad, et autres
Publié: (2024)
BackdoorIDS: Zero-shot Backdoor Detection for Pretrained Vision Encoder
par: Huang, Siquan, et autres
Publié: (2026)
par: Huang, Siquan, et autres
Publié: (2026)
An Adversarial Approach to Register Extreme Resolution Tissue Cleared 3D Brain Images
par: Naziba, Abdullah, et autres
Publié: (2025)
par: Naziba, Abdullah, et autres
Publié: (2025)
Zero-Shot Vision Encoder Grafting via LLM Surrogates
par: Yue, Kaiyu, et autres
Publié: (2025)
par: Yue, Kaiyu, et autres
Publié: (2025)
GenCLIP: Generalizing CLIP Prompts for Zero-shot Anomaly Detection
par: Kim, Donghyeong, et autres
Publié: (2025)
par: Kim, Donghyeong, et autres
Publié: (2025)
Transductive Zero-Shot and Few-Shot CLIP
par: Martin, Ségolène, et autres
Publié: (2024)
par: Martin, Ségolène, et autres
Publié: (2024)
Leveraging CLIP Encoder for Multimodal Emotion Recognition
par: Song, Yehun, et autres
Publié: (2025)
par: Song, Yehun, et autres
Publié: (2025)
Physics-Guided Attention in a Lightweight TCN for Efficient WiFi CSI-Based Human Activity Recognition
par: Ranasingha, Chinthaka, et autres
Publié: (2026)
par: Ranasingha, Chinthaka, et autres
Publié: (2026)
Dual-Stream Alignment for Action Segmentation
par: Gammulle, Harshala, et autres
Publié: (2025)
par: Gammulle, Harshala, et autres
Publié: (2025)
Proto-CLIP: Vision-Language Prototypical Network for Few-Shot Learning
par: P, Jishnu Jaykumar, et autres
Publié: (2023)
par: P, Jishnu Jaykumar, et autres
Publié: (2023)
HAC: Parameter-Efficient Hyperbolic Adaptation of CLIP for Zero-Shot VQA
par: Dibitonto, Francesco, et autres
Publié: (2026)
par: Dibitonto, Francesco, et autres
Publié: (2026)
Cross-Branch Orthogonality for Improved Generalization in Face Deepfake Detection
par: Fernando, Tharindu, et autres
Publié: (2025)
par: Fernando, Tharindu, et autres
Publié: (2025)
Online Zero-Shot Classification with CLIP
par: Qian, Qi, et autres
Publié: (2024)
par: Qian, Qi, et autres
Publié: (2024)
CLIP-driven Zero-shot Learning with Ambiguous Labels
par: Fan, Jinfu, et autres
Publié: (2026)
par: Fan, Jinfu, et autres
Publié: (2026)
Explaining CLIP Zero-shot Predictions Through Concepts
par: Ozdemir, Onat, et autres
Publié: (2026)
par: Ozdemir, Onat, et autres
Publié: (2026)
SPECIAL: Zero-shot Hyperspectral Image Classification With CLIP
par: Pang, Li, et autres
Publié: (2025)
par: Pang, Li, et autres
Publié: (2025)
FrameDiT: Diffusion Transformer with Matrix Attention for Efficient Video Generation
par: Le, Minh Khoa, et autres
Publié: (2026)
par: Le, Minh Khoa, et autres
Publié: (2026)
AI, Entrepreneurs, and Privacy: Deep Learning Outperforms Humans in Detecting Entrepreneurs from Image Data
par: Obschonka, Martin, et autres
Publié: (2024)
par: Obschonka, Martin, et autres
Publié: (2024)
Uncertainty Driven Bottleneck Attention U-net for Organ at Risk Segmentation
par: Nazib, Abdullah, et autres
Publié: (2023)
par: Nazib, Abdullah, et autres
Publié: (2023)
SpectralZoom: Efficient Segmentation with an Adaptive Hyperspectral Camera
par: Arnold, Jackson, et autres
Publié: (2024)
par: Arnold, Jackson, et autres
Publié: (2024)
MV-CLIP: Multi-View CLIP for Zero-shot 3D Shape Recognition
par: Song, Dan, et autres
Publié: (2023)
par: Song, Dan, et autres
Publié: (2023)
Documents similaires
-
OmniGCD: Abstracting Generalized Category Discovery for Modality Agnosticism
par: Shipard, Jordan, et autres
Publié: (2026) -
Filling the Gaps: A Multitask Hybrid Multiscale Generative Framework for Missing Modality in Remote Sensing Semantic Segmentation
par: Kieu, Nhi, et autres
Publié: (2025) -
DIS2: Disentanglement Meets Distillation with Classwise Attention for Robust Remote Sensing Segmentation under Missing Modalities
par: Kieu, Nhi, et autres
Publié: (2026) -
Physics-Informed Computer Vision: A Review and Perspectives
par: Banerjee, Chayan, et autres
Publié: (2023) -
AG-ReID.v2: Bridging Aerial and Ground Views for Person Re-identification
par: Nguyen, Huy, et autres
Publié: (2024)