Computer vision tasks for intelligent aerospace missions: An overview
Fuente:
arXiv
Salvato in:
| Autori principali: | Chen, Huilin, Sun, Qiyu, Li, Fangfei, Tang, Yang |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Applying Deep Neural Networks to automate visual verification of manual bracket installations in aerospace
di: Oyekan, John, et al.
Pubblicazione: (2024)
di: Oyekan, John, et al.
Pubblicazione: (2024)
Thinker: A vision-language foundation model for embodied intelligence
di: Pan, Baiyu, et al.
Pubblicazione: (2026)
di: Pan, Baiyu, et al.
Pubblicazione: (2026)
Robust Single-shot Structured Light 3D Imaging via Neural Feature Decoding
di: Li, Jiaheng, et al.
Pubblicazione: (2025)
di: Li, Jiaheng, et al.
Pubblicazione: (2025)
ProFound: A moderate-sized vision foundation model for multi-task prostate imaging
di: Wang, Yipei, et al.
Pubblicazione: (2026)
di: Wang, Yipei, et al.
Pubblicazione: (2026)
Adversarial Examples in Environment Perception for Automated Driving (Review)
di: Yan, Jun, et al.
Pubblicazione: (2025)
di: Yan, Jun, et al.
Pubblicazione: (2025)
HSFusion: A high-level vision task-driven infrared and visible image fusion network via semantic and geometric domain transformation
di: Jiang, Chengjie, et al.
Pubblicazione: (2024)
di: Jiang, Chengjie, et al.
Pubblicazione: (2024)
UniPINN: A Unified PINN Framework for Multi-task Learning of Diverse Navier-Stokes Equations
di: Sun, Dengdi, et al.
Pubblicazione: (2026)
di: Sun, Dengdi, et al.
Pubblicazione: (2026)
Utilizing the Mean Teacher with Supcontrast Loss for Wafer Pattern Recognition
di: Wei, Qiyu, et al.
Pubblicazione: (2024)
di: Wei, Qiyu, et al.
Pubblicazione: (2024)
Bridging Cross-task Protocol Inconsistency for Distillation in Dense Object Detection
di: Yang, Longrong, et al.
Pubblicazione: (2023)
di: Yang, Longrong, et al.
Pubblicazione: (2023)
Computer vision-based estimation of invertebrate biomass
di: Impiö, Mikko, et al.
Pubblicazione: (2026)
di: Impiö, Mikko, et al.
Pubblicazione: (2026)
Adaptive Dual-Constrained Line Aggregation for Robust Generic and Wireframe Line Segment Detection
di: Liu, Chenguang, et al.
Pubblicazione: (2025)
di: Liu, Chenguang, et al.
Pubblicazione: (2025)
A Unified Anomaly Synthesis Strategy with Gradient Ascent for Industrial Anomaly Detection and Localization
di: Chen, Qiyu, et al.
Pubblicazione: (2024)
di: Chen, Qiyu, et al.
Pubblicazione: (2024)
RoMa: Robust Dense Feature Matching
di: Edstedt, Johan, et al.
Pubblicazione: (2023)
di: Edstedt, Johan, et al.
Pubblicazione: (2023)
Bias-constrained multimodal intelligence for equitable and reliable clinical AI
di: Li, Cheng, et al.
Pubblicazione: (2026)
di: Li, Cheng, et al.
Pubblicazione: (2026)
POPCat: Propagation of particles for complex annotation tasks
di: Yang, Adam Srebrnjak, et al.
Pubblicazione: (2024)
di: Yang, Adam Srebrnjak, et al.
Pubblicazione: (2024)
ViSTa Dataset: Do vision-language models understand sequential tasks?
di: Wybitul, Evžen, et al.
Pubblicazione: (2024)
di: Wybitul, Evžen, et al.
Pubblicazione: (2024)
Representation geometry shapes task performance in vision-language modeling for CT enterography
di: Minoccheri, Cristian, et al.
Pubblicazione: (2026)
di: Minoccheri, Cristian, et al.
Pubblicazione: (2026)
OmniCamera: A Unified Framework for Multi-task Video Generation with Arbitrary Camera Control
di: Wang, Yukun, et al.
Pubblicazione: (2026)
di: Wang, Yukun, et al.
Pubblicazione: (2026)
Beyond conventional vision: RGB-event fusion for robust object detection in dynamic traffic scenarios
di: Liu, Zhanwen, et al.
Pubblicazione: (2025)
di: Liu, Zhanwen, et al.
Pubblicazione: (2025)
Fine-tuning vision foundation model for crack segmentation in civil infrastructures
di: Ge, Kang, et al.
Pubblicazione: (2023)
di: Ge, Kang, et al.
Pubblicazione: (2023)
MT-Depth: Multi-task Instance feature analysis for the Depth Completion
di: Nizamani, Abdul Haseeb, et al.
Pubblicazione: (2025)
di: Nizamani, Abdul Haseeb, et al.
Pubblicazione: (2025)
Efficient RGB-D Scene Understanding via Multi-task Adaptive Learning and Cross-dimensional Feature Guidance
di: Sun, Guodong, et al.
Pubblicazione: (2026)
di: Sun, Guodong, et al.
Pubblicazione: (2026)
VMAD: Visual-enhanced Multimodal Large Language Model for Zero-Shot Anomaly Detection
di: Deng, Huilin, et al.
Pubblicazione: (2024)
di: Deng, Huilin, et al.
Pubblicazione: (2024)
MRAD: Zero-Shot Anomaly Detection with Memory-Driven Retrieval
di: Xu, Chaoran, et al.
Pubblicazione: (2026)
di: Xu, Chaoran, et al.
Pubblicazione: (2026)
Co-Training Vision Language Models for Remote Sensing Multi-task Learning
di: Li, Qingyun, et al.
Pubblicazione: (2025)
di: Li, Qingyun, et al.
Pubblicazione: (2025)
LiMT: A Multi-task Liver Image Benchmark Dataset
di: Liu, Zhe, et al.
Pubblicazione: (2025)
di: Liu, Zhe, et al.
Pubblicazione: (2025)
Openfly: A comprehensive platform for aerial vision-language navigation
di: Gao, Yunpeng, et al.
Pubblicazione: (2025)
di: Gao, Yunpeng, et al.
Pubblicazione: (2025)
PhysMLE: Generalizable and Priors-Inclusive Multi-task Remote Physiological Measurement
di: Wang, Jiyao, et al.
Pubblicazione: (2024)
di: Wang, Jiyao, et al.
Pubblicazione: (2024)
Learning Physical Dynamics for Object-centric Visual Prediction
di: Xu, Huilin, et al.
Pubblicazione: (2024)
di: Xu, Huilin, et al.
Pubblicazione: (2024)
Dynamic Proxy Domain Generalizes the Crowd Localization by Better Binary Segmentation
di: Gao, Junyu, et al.
Pubblicazione: (2024)
di: Gao, Junyu, et al.
Pubblicazione: (2024)
SOTA: Self-adaptive Optimal Transport for Zero-Shot Classification with Multiple Foundation Models
di: Hu, Zhanxuan, et al.
Pubblicazione: (2025)
di: Hu, Zhanxuan, et al.
Pubblicazione: (2025)
OccScene: Semantic Occupancy-based Cross-task Mutual Learning for 3D Scene Generation
di: Li, Bohan, et al.
Pubblicazione: (2024)
di: Li, Bohan, et al.
Pubblicazione: (2024)
MMHMER:Multi-viewer and Multi-task for Handwritten Mathematical Expression Recognition
di: Chen, Kehua, et al.
Pubblicazione: (2025)
di: Chen, Kehua, et al.
Pubblicazione: (2025)
4D-Rotor Gaussian Splatting: Towards Efficient Novel View Synthesis for Dynamic Scenes
di: Duan, Yuanxing, et al.
Pubblicazione: (2024)
di: Duan, Yuanxing, et al.
Pubblicazione: (2024)
DSPNet: Dual-vision Scene Perception for Robust 3D Question Answering
di: Luo, Jingzhou, et al.
Pubblicazione: (2025)
di: Luo, Jingzhou, et al.
Pubblicazione: (2025)
WS-DETR: Robust Water Surface Object Detection through Vision-Radar Fusion with Detection Transformer
di: Yin, Huilin, et al.
Pubblicazione: (2025)
di: Yin, Huilin, et al.
Pubblicazione: (2025)
Flatten: Video Action Recognition is an Image Classification task
di: Chen, Junlin, et al.
Pubblicazione: (2024)
di: Chen, Junlin, et al.
Pubblicazione: (2024)
TABLET: Table Structure Recognition using Encoder-only Transformers
di: Hou, Qiyu, et al.
Pubblicazione: (2025)
di: Hou, Qiyu, et al.
Pubblicazione: (2025)
An aerial color image anomaly dataset for search missions in complex forested terrain
di: Nathan, Rakesh John Amala Arokia, et al.
Pubblicazione: (2025)
di: Nathan, Rakesh John Amala Arokia, et al.
Pubblicazione: (2025)
SGIA: Enhancing Fine-Grained Visual Classification with Sequence Generative Image Augmentation
di: Liao, Qiyu, et al.
Pubblicazione: (2024)
di: Liao, Qiyu, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Applying Deep Neural Networks to automate visual verification of manual bracket installations in aerospace
di: Oyekan, John, et al.
Pubblicazione: (2024) -
Thinker: A vision-language foundation model for embodied intelligence
di: Pan, Baiyu, et al.
Pubblicazione: (2026) -
Robust Single-shot Structured Light 3D Imaging via Neural Feature Decoding
di: Li, Jiaheng, et al.
Pubblicazione: (2025) -
ProFound: A moderate-sized vision foundation model for multi-task prostate imaging
di: Wang, Yipei, et al.
Pubblicazione: (2026) -
Adversarial Examples in Environment Perception for Automated Driving (Review)
di: Yan, Jun, et al.
Pubblicazione: (2025)