Salvato in:
| Autori principali: | Dang, Bo, Zhao, Wenchao, Li, Yufeng, Ma, Danqing, Yu, Qixuan, Zhu, Elly Yijun |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2405.05983 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Research on Brain Tumor Classification Method Based on Improved ResNet34 Network
di: Li, Yufeng, et al.
Pubblicazione: (2025)
di: Li, Yufeng, et al.
Pubblicazione: (2025)
Fostc3net:A Lightweight YOLOv5 Based On the Network Structure Optimization
di: Ma, Danqing, et al.
Pubblicazione: (2024)
di: Ma, Danqing, et al.
Pubblicazione: (2024)
Real-Time Object Detection in Occluded Environment with Background Cluttering Effects Using Deep Learning
di: Aamir, Syed Muhammad, et al.
Pubblicazione: (2024)
di: Aamir, Syed Muhammad, et al.
Pubblicazione: (2024)
MedEyes: Learning Dynamic Visual Focus for Medical Progressive Diagnosis
di: Zhu, Chunzheng, et al.
Pubblicazione: (2025)
di: Zhu, Chunzheng, et al.
Pubblicazione: (2025)
NaVIP: An Image-Centric Indoor Navigation Solution for Visually Impaired People
di: Yu, Jun, et al.
Pubblicazione: (2024)
di: Yu, Jun, et al.
Pubblicazione: (2024)
WalkVLM:Aid Visually Impaired People Walking by Vision Language Model
di: Yuan, Zhiqiang, et al.
Pubblicazione: (2024)
di: Yuan, Zhiqiang, et al.
Pubblicazione: (2024)
Restoring Real-World Images with an Internal Detail Enhancement Diffusion Model
di: Xiao, Peng, et al.
Pubblicazione: (2025)
di: Xiao, Peng, et al.
Pubblicazione: (2025)
Near--Real-Time Conflict-Related Fire Detection in Sudan Using Unsupervised Deep Learning
di: Atwal, Kuldip Singh, et al.
Pubblicazione: (2025)
di: Atwal, Kuldip Singh, et al.
Pubblicazione: (2025)
Preserving Cross-Modal Stability for Visual Unlearning in Multimodal Scenarios
di: Li, Jinghan Xu Yuyang Zhang Qixuan Cai Jiancheng Chen Keqiu
Pubblicazione: (2025)
di: Li, Jinghan Xu Yuyang Zhang Qixuan Cai Jiancheng Chen Keqiu
Pubblicazione: (2025)
Real-Time Currency Detection and Voice Feedback for Visually Impaired Individuals
di: Shreya, Saraf Anzum, et al.
Pubblicazione: (2025)
di: Shreya, Saraf Anzum, et al.
Pubblicazione: (2025)
VIALM: A Survey and Benchmark of Visually Impaired Assistance with Large Models
di: Zhao, Yi, et al.
Pubblicazione: (2024)
di: Zhao, Yi, et al.
Pubblicazione: (2024)
A Forward and Backward Compatible Framework for Few-shot Class-incremental Pill Recognition
di: Zhang, Jinghua, et al.
Pubblicazione: (2023)
di: Zhang, Jinghua, et al.
Pubblicazione: (2023)
Boosting Fine-Grained Visual Anomaly Detection with Coarse-Knowledge-Aware Adversarial Learning
di: Fang, Qingqing, et al.
Pubblicazione: (2024)
di: Fang, Qingqing, et al.
Pubblicazione: (2024)
Probing Visual Planning in Image Editing Models
di: Zhou, Zhimu, et al.
Pubblicazione: (2026)
di: Zhou, Zhimu, et al.
Pubblicazione: (2026)
MedSynapse-V: Bridging Visual Perception and Clinical Intuition via Latent Memory Evolution
di: Zhu, Chunzheng, et al.
Pubblicazione: (2026)
di: Zhu, Chunzheng, et al.
Pubblicazione: (2026)
Real-Time Roadway Obstacle Detection for Electric Scooters Using Deep Learning and Multi-Sensor Fusion
di: Zheng, Zeyang, et al.
Pubblicazione: (2025)
di: Zheng, Zeyang, et al.
Pubblicazione: (2025)
Diffusion Curriculum: Synthetic-to-Real Data Curriculum via Image-Guided Diffusion
di: Liang, Yijun, et al.
Pubblicazione: (2024)
di: Liang, Yijun, et al.
Pubblicazione: (2024)
GPT-based Textile Pilling Classification Using 3D Point Cloud Data
di: Lu, Yu, et al.
Pubblicazione: (2024)
di: Lu, Yu, et al.
Pubblicazione: (2024)
Real-Time Visual Attribution Streaming in Thinking Model
di: Kang, Seil, et al.
Pubblicazione: (2026)
di: Kang, Seil, et al.
Pubblicazione: (2026)
Deep Learning-Powered Visual SLAM Aimed at Assisting Visually Impaired Navigation
di: Bamdad, Marziyeh, et al.
Pubblicazione: (2025)
di: Bamdad, Marziyeh, et al.
Pubblicazione: (2025)
Real Time Deep Learning Weapon Detection Techniques for Mitigating Lone Wolf Attacks
di: Akhila, Kambhatla, et al.
Pubblicazione: (2024)
di: Akhila, Kambhatla, et al.
Pubblicazione: (2024)
Exploring Task-Level Optimal Prompts for Visual In-Context Learning
di: Zhu, Yan, et al.
Pubblicazione: (2025)
di: Zhu, Yan, et al.
Pubblicazione: (2025)
Evaluating Few-Shot Pill Recognition Under Visual Domain Shift
di: Chu, W. I., et al.
Pubblicazione: (2026)
di: Chu, W. I., et al.
Pubblicazione: (2026)
Improving the Spatial Resolution of GONG Solar Images to GST Quality Using Deep Learning
di: Li, Chenyang, et al.
Pubblicazione: (2025)
di: Li, Chenyang, et al.
Pubblicazione: (2025)
A Comprehensive Survey of Data Augmentation in Visual Reinforcement Learning
di: Ma, Guozheng, et al.
Pubblicazione: (2022)
di: Ma, Guozheng, et al.
Pubblicazione: (2022)
Breaking the SFT Plateau: Multimodal Structured Reinforcement Learning for Chart-to-Code Generation
di: Chen, Lei, et al.
Pubblicazione: (2025)
di: Chen, Lei, et al.
Pubblicazione: (2025)
An Overall Real-Time Mechanism for Classification and Quality Evaluation of Rice
di: Xia, Wanke, et al.
Pubblicazione: (2025)
di: Xia, Wanke, et al.
Pubblicazione: (2025)
Real-Time Fusion of Visual and Chart Data for Enhanced Maritime Vision
di: Kreis, Marten, et al.
Pubblicazione: (2025)
di: Kreis, Marten, et al.
Pubblicazione: (2025)
KALAHash: Knowledge-Anchored Low-Resource Adaptation for Deep Hashing
di: Zhao, Shu, et al.
Pubblicazione: (2024)
di: Zhao, Shu, et al.
Pubblicazione: (2024)
edgeVLM: Cloud-edge Collaborative Real-time VLM based on Context Transfer
di: Qian, Chen, et al.
Pubblicazione: (2025)
di: Qian, Chen, et al.
Pubblicazione: (2025)
PhiP-G: Physics-Guided Text-to-3D Compositional Scene Generation
di: Li, Qixuan, et al.
Pubblicazione: (2025)
di: Li, Qixuan, et al.
Pubblicazione: (2025)
REMOTE: Real-time Ego-motion Tracking for Various Endoscopes via Multimodal Visual Feature Learning
di: Shao, Liangjing, et al.
Pubblicazione: (2025)
di: Shao, Liangjing, et al.
Pubblicazione: (2025)
DART: Depth-Enhanced Accurate and Real-Time Background Matting
di: Li, Hanxi, et al.
Pubblicazione: (2024)
di: Li, Hanxi, et al.
Pubblicazione: (2024)
Transformer-Based Classification Outcome Prediction for Multimodal Stroke Treatment
di: Ma, Danqing, et al.
Pubblicazione: (2024)
di: Ma, Danqing, et al.
Pubblicazione: (2024)
A Large Vision-Language Model based Environment Perception System for Visually Impaired People
di: Chen, Zezhou, et al.
Pubblicazione: (2025)
di: Chen, Zezhou, et al.
Pubblicazione: (2025)
MelNet: A Real-Time Deep Learning Algorithm for Object Detection
di: Azadvatan, Yashar, et al.
Pubblicazione: (2024)
di: Azadvatan, Yashar, et al.
Pubblicazione: (2024)
Persistent Visual Memory: Sustaining Perception for Deep Generation in LVLMs
di: Huang, Siyuan, et al.
Pubblicazione: (2026)
di: Huang, Siyuan, et al.
Pubblicazione: (2026)
Real-Time Object Tracking with On-Device Deep Learning for Adaptive Beamforming in Dynamic Acoustic Environments
di: Ortigoso-Narro, Jorge, et al.
Pubblicazione: (2025)
di: Ortigoso-Narro, Jorge, et al.
Pubblicazione: (2025)
HoliTracer: Holistic Vectorization of Geographic Objects from Large-Size Remote Sensing Imagery
di: Wang, Yu, et al.
Pubblicazione: (2025)
di: Wang, Yu, et al.
Pubblicazione: (2025)
A Deep Learning Framework for Real-Time Image Processing in Medical Diagnostics: Enhancing Accuracy and Speed in Clinical Applications
di: Filvantorkaman, Melika, et al.
Pubblicazione: (2025)
di: Filvantorkaman, Melika, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Research on Brain Tumor Classification Method Based on Improved ResNet34 Network
di: Li, Yufeng, et al.
Pubblicazione: (2025) -
Fostc3net:A Lightweight YOLOv5 Based On the Network Structure Optimization
di: Ma, Danqing, et al.
Pubblicazione: (2024) -
Real-Time Object Detection in Occluded Environment with Background Cluttering Effects Using Deep Learning
di: Aamir, Syed Muhammad, et al.
Pubblicazione: (2024) -
MedEyes: Learning Dynamic Visual Focus for Medical Progressive Diagnosis
di: Zhu, Chunzheng, et al.
Pubblicazione: (2025) -
NaVIP: An Image-Centric Indoor Navigation Solution for Visually Impaired People
di: Yu, Jun, et al.
Pubblicazione: (2024)