Towards a Large Language-Vision Question Answering Model for MSTAR Automatic Target Recognition
Fuente:
arXiv
Guardado en:
| Autores principales: | Ramirez, David F., Overman, Tim L., Jaskie, Kristen, Kleine, Marv, Spanias, Andreas |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
SAR-RAG: ATR Visual Question Answering by Semantic Search, Retrieval, and MLLM Generation
por: Ramirez, David F., et al.
Publicado: (2026)
por: Ramirez, David F., et al.
Publicado: (2026)
Geospatial-Temporal Sensemaking of Remote Sensing Activity Detections with Multimodal Large Language Model
por: Ramirez, David F., et al.
Publicado: (2026)
por: Ramirez, David F., et al.
Publicado: (2026)
Infrared Computer Vision for Utility-Scale Photovoltaic Array Inspection
por: Ramirez, David F., et al.
Publicado: (2024)
por: Ramirez, David F., et al.
Publicado: (2024)
The OCR-PT-CT Project: Semi-Automatic Recognition of Ancient Egyptian Hieroglyphs Based on Metric Learning
por: Fuentes-Jimenez, David, et al.
Publicado: (2025)
por: Fuentes-Jimenez, David, et al.
Publicado: (2025)
Generalized User-Oriented Image Semantic Coding Empowered by Large Vision-Language Model
por: Huang, Sin-Yu, et al.
Publicado: (2025)
por: Huang, Sin-Yu, et al.
Publicado: (2025)
A Method for Target Detection Based on Mmw Radar and Vision Fusion
por: Zong, Ming, et al.
Publicado: (2024)
por: Zong, Ming, et al.
Publicado: (2024)
Surgical-LVLM: Learning to Adapt Large Vision-Language Model for Grounded Visual Question Answering in Robotic Surgery
por: Wang, Guankun, et al.
Publicado: (2024)
por: Wang, Guankun, et al.
Publicado: (2024)
Black Sheep in the Herd: Playing with Spuriously Correlated Attributes for Vision-Language Recognition
por: Tian, Xinyu, et al.
Publicado: (2025)
por: Tian, Xinyu, et al.
Publicado: (2025)
LDSF: Lightweight Dual-Stream Framework for SAR Target Recognition by Coupling Local Electromagnetic Scattering Features and Global Visual Features
por: Xiong, Xuying, et al.
Publicado: (2024)
por: Xiong, Xuying, et al.
Publicado: (2024)
Cochlear Implantation of Slim Pre-curved Arrays using Automatic Pre-operative Insertion Plans
por: Tawfik, Kareem O., et al.
Publicado: (2024)
por: Tawfik, Kareem O., et al.
Publicado: (2024)
An approach to Physics Based Single Photon Vision System Design
por: Lu, Yizhou, et al.
Publicado: (2023)
por: Lu, Yizhou, et al.
Publicado: (2023)
Unifying Image Processing as Visual Prompting Question Answering
por: Liu, Yihao, et al.
Publicado: (2023)
por: Liu, Yihao, et al.
Publicado: (2023)
RATNUS: Rapid, Automatic Thalamic Nuclei Segmentation using Multimodal MRI inputs
por: Feng, Anqi, et al.
Publicado: (2024)
por: Feng, Anqi, et al.
Publicado: (2024)
Towards robust quantitative photoacoustic tomography via learned iterative methods
por: Manninen, Anssi, et al.
Publicado: (2025)
por: Manninen, Anssi, et al.
Publicado: (2025)
Benchmarking Deep Learning Classifiers for SAR Automatic Target Recognition
por: Fein-Ashley, Jacob, et al.
Publicado: (2023)
por: Fein-Ashley, Jacob, et al.
Publicado: (2023)
CONRep: Uncertainty-Aware Vision-Language Report Drafting Using Conformal Prediction
por: Elyassirad, Danial, et al.
Publicado: (2026)
por: Elyassirad, Danial, et al.
Publicado: (2026)
Goal-Oriented Semantic Communication for Wireless Visual Question Answering
por: Liu, Sige, et al.
Publicado: (2024)
por: Liu, Sige, et al.
Publicado: (2024)
Context-Aware Vision Language Foundation Models for Ocular Disease Screening in Retinal Images
por: Berger, Lucie, et al.
Publicado: (2025)
por: Berger, Lucie, et al.
Publicado: (2025)
Compressed Learning for Nanosurface Deficiency Recognition Using Angle-resolved Scatterometry Data
por: Abdollahpour, Mehdi, et al.
Publicado: (2025)
por: Abdollahpour, Mehdi, et al.
Publicado: (2025)
Automatic Multi-Class Cardiovascular Magnetic Resonance Image Quality Assessment using Unsupervised Domain Adaptation in Spatial and Frequency Domains
por: Nabavi, Shahabedin, et al.
Publicado: (2021)
por: Nabavi, Shahabedin, et al.
Publicado: (2021)
Automatic Tuning of Denoising Algorithms Parameters Without Ground Truth
por: Floquet, Arthur, et al.
Publicado: (2024)
por: Floquet, Arthur, et al.
Publicado: (2024)
HERBench: A Benchmark for Multi-Evidence Integration in Video Question Answering
por: Ben-Ami, Dan, et al.
Publicado: (2025)
por: Ben-Ami, Dan, et al.
Publicado: (2025)
SPICER: Self-Supervised Learning for MRI with Automatic Coil Sensitivity Estimation and Reconstruction
por: Hu, Yuyang, et al.
Publicado: (2022)
por: Hu, Yuyang, et al.
Publicado: (2022)
LiNUS: Lightweight Automatic Segmentation of Deep Brain Nuclei for Real-Time DBS Surgery
por: Zhang, Shuo, et al.
Publicado: (2026)
por: Zhang, Shuo, et al.
Publicado: (2026)
Ultra-fast Real-time Target Recognition Using a Shift, Scale, and Rotation Invariant Hybrid Opto-electronic Joint Transform Correlator
por: Shen, Xi, et al.
Publicado: (2025)
por: Shen, Xi, et al.
Publicado: (2025)
Towards robust radiomics and radiogenomics predictive models for brain tumor characterization
por: Nadeem, Maria, et al.
Publicado: (2024)
por: Nadeem, Maria, et al.
Publicado: (2024)
Beyond Pixel Agreement: Large Language Models as Clinical Guardrails for Reliable Medical Image Segmentation
por: Sheng, Jiaxi, et al.
Publicado: (2025)
por: Sheng, Jiaxi, et al.
Publicado: (2025)
Automatically Generating Narrative-Style Radiology Reports from Volumetric CT Images; a Proof of Concept
por: Borghouts, Marijn
Publicado: (2024)
por: Borghouts, Marijn
Publicado: (2024)
Diver-Robot Communication Dataset for Underwater Hand Gesture Recognition
por: Kvasić, Igor, et al.
Publicado: (2025)
por: Kvasić, Igor, et al.
Publicado: (2025)
Nine Years of Pediatric Iris Recognition: Evidence for Biometric Permanence
por: Venkataswamy, Naveenkumar G, et al.
Publicado: (2025)
por: Venkataswamy, Naveenkumar G, et al.
Publicado: (2025)
Solution for Authenticity Identification of Typical Target Remote Sensing Images
por: Lin, Yipeng, et al.
Publicado: (2024)
por: Lin, Yipeng, et al.
Publicado: (2024)
TRUST -- Transformer-Driven U-Net for Sparse Target Recovery
por: An, Di, et al.
Publicado: (2025)
por: An, Di, et al.
Publicado: (2025)
A BTR-Based Approach for Detection of Infrared Small Targets
por: Li, Ke-Xin
Publicado: (2025)
por: Li, Ke-Xin
Publicado: (2025)
Two-Stage nnU-Net for Automatic Multi-class Bi-Atrial Segmentation from LGE-MRIs
por: On, Y., et al.
Publicado: (2025)
por: On, Y., et al.
Publicado: (2025)
Entropy Coding for Non-Rectangular Transform Blocks using Partitioned DCT Dictionaries for AV1
por: Das, Priyanka, et al.
Publicado: (2025)
por: Das, Priyanka, et al.
Publicado: (2025)
Shadow Augmentation for Handwashing Action Recognition: from Synthetic to Real Datasets
por: Ju, Shengtai, et al.
Publicado: (2024)
por: Ju, Shengtai, et al.
Publicado: (2024)
Automatic Detection and Classification of Corona Infection (COVID-19) from X-ray Images Using Convolution Neural Network
por: Patel, Kinjal A, et al.
Publicado: (2024)
por: Patel, Kinjal A, et al.
Publicado: (2024)
Local Patch Network with Global Attention for Infrared Small Target Detection
por: Chen, Fang, et al.
Publicado: (2021)
por: Chen, Fang, et al.
Publicado: (2021)
MaDiNet: Mamba Diffusion Network for SAR Target Detection
por: Zhou, Jie, et al.
Publicado: (2024)
por: Zhou, Jie, et al.
Publicado: (2024)
Twofold Structured Features-Based Siamese Network for Infrared Target Tracking
por: Yan, Wei-Jie, et al.
Publicado: (2023)
por: Yan, Wei-Jie, et al.
Publicado: (2023)
Ejemplares similares
-
SAR-RAG: ATR Visual Question Answering by Semantic Search, Retrieval, and MLLM Generation
por: Ramirez, David F., et al.
Publicado: (2026) -
Geospatial-Temporal Sensemaking of Remote Sensing Activity Detections with Multimodal Large Language Model
por: Ramirez, David F., et al.
Publicado: (2026) -
Infrared Computer Vision for Utility-Scale Photovoltaic Array Inspection
por: Ramirez, David F., et al.
Publicado: (2024) -
The OCR-PT-CT Project: Semi-Automatic Recognition of Ancient Egyptian Hieroglyphs Based on Metric Learning
por: Fuentes-Jimenez, David, et al.
Publicado: (2025) -
Generalized User-Oriented Image Semantic Coding Empowered by Large Vision-Language Model
por: Huang, Sin-Yu, et al.
Publicado: (2025)