Saved in:
| Main Authors: | Russo, Ciro, Bria, Alessandro, Marrocco, Claudio |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2309.12876 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SynSur: An end-to-end generative pipeline for synthetic industrial surface defect generation and detection
by: Kühn, Paul Julius, et al.
Published: (2026)
by: Kühn, Paul Julius, et al.
Published: (2026)
Stability Plasticity Decoupled Fine-tuning For Few-shot end-to-end Object Detection
by: Yin, Yuantao, et al.
Published: (2024)
by: Yin, Yuantao, et al.
Published: (2024)
RoofSeg: An edge-aware transformer-based network for end-to-end roof plane segmentation
by: You, Siyuan, et al.
Published: (2025)
by: You, Siyuan, et al.
Published: (2025)
ArtBrain: An Explainable end-to-end Toolkit for Classification and Attribution of AI-Generated Art and Style
by: Silva, Ravidu Suien Rammuni, et al.
Published: (2024)
by: Silva, Ravidu Suien Rammuni, et al.
Published: (2024)
Validation of an AI-based end-to-end model for prostate pathology using long-term archived routine samples
by: Ji, Xiaoyi, et al.
Published: (2026)
by: Ji, Xiaoyi, et al.
Published: (2026)
EVC-MF: End-to-end Video Captioning Network with Multi-scale Features
by: Niu, Tian-Zi, et al.
Published: (2024)
by: Niu, Tian-Zi, et al.
Published: (2024)
VidFormer: A novel end-to-end framework fused by 3DCNN and Transformer for Video-based Remote Physiological Measurement
by: Li, Jiachen, et al.
Published: (2025)
by: Li, Jiachen, et al.
Published: (2025)
Full end-to-end diagnostic workflow automation of 3D OCT via foundation model-driven AI for retinal diseases
by: Zhang, Jinze, et al.
Published: (2026)
by: Zhang, Jinze, et al.
Published: (2026)
SToRM: Supervised Token Reduction for Multi-modal LLMs toward efficient end-to-end autonomous driving
by: Kim, Seo Hyun, et al.
Published: (2026)
by: Kim, Seo Hyun, et al.
Published: (2026)
SGTR+: End-to-end Scene Graph Generation with Transformer
by: Li, Rongjie, et al.
Published: (2024)
by: Li, Rongjie, et al.
Published: (2024)
Handling Missing Modalities in Multimodal Survival Prediction for Non-Small Cell Lung Cancer
by: Ruffini, Filippo, et al.
Published: (2026)
by: Ruffini, Filippo, et al.
Published: (2026)
Module-wise Adaptive Adversarial Training for End-to-end Autonomous Driving
by: Zhang, Tianyuan, et al.
Published: (2024)
by: Zhang, Tianyuan, et al.
Published: (2024)
ParkingE2E: Camera-based End-to-end Parking Network, from Images to Planning
by: Li, Changze, et al.
Published: (2024)
by: Li, Changze, et al.
Published: (2024)
Combating the effects of speed and delays in end-to-end self-driving
by: Tampuu, Ardi, et al.
Published: (2023)
by: Tampuu, Ardi, et al.
Published: (2023)
End-to-end Autonomous Driving: Challenges and Frontiers
by: Chen, Li, et al.
Published: (2023)
by: Chen, Li, et al.
Published: (2023)
AI-driven Automation of End-to-end Assessment of Suturing Expertise
by: Deo, Atharva, et al.
Published: (2025)
by: Deo, Atharva, et al.
Published: (2025)
Path-RAG: Knowledge-Guided Key Region Retrieval for Open-ended Pathology Visual Question Answering
by: Naeem, Awais, et al.
Published: (2024)
by: Naeem, Awais, et al.
Published: (2024)
Codebook-enabled Generative End-to-end Semantic Communication Powered by Transformer
by: Ye, Peigen, et al.
Published: (2024)
by: Ye, Peigen, et al.
Published: (2024)
How Well Can Vison-Language Models Understand Humans' Intention? An Open-ended Theory of Mind Question Evaluation Benchmark
by: Wen, Ximing, et al.
Published: (2025)
by: Wen, Ximing, et al.
Published: (2025)
Shedding Light on Depth: Explainability Assessment in Monocular Depth Estimation
by: Cirillo, Lorenzo, et al.
Published: (2025)
by: Cirillo, Lorenzo, et al.
Published: (2025)
Point2RBox: Combine Knowledge from Synthetic Visual Patterns for End-to-end Oriented Object Detection with Single Point Supervision
by: Yu, Yi, et al.
Published: (2023)
by: Yu, Yi, et al.
Published: (2023)
Latent Space Analysis for Melanoma Prevention
by: Listone, Ciro, et al.
Published: (2025)
by: Listone, Ciro, et al.
Published: (2025)
End-to-end Listen, Look, Speak and Act
by: Wang, Siyin, et al.
Published: (2025)
by: Wang, Siyin, et al.
Published: (2025)
End-to-end Semantic-centric Video-based Multimodal Affective Computing
by: Lin, Ronghao, et al.
Published: (2024)
by: Lin, Ronghao, et al.
Published: (2024)
EARTalking: End-to-end GPT-style Autoregressive Talking Head Synthesis with Frame-wise Control
by: Weng, Yuzhe, et al.
Published: (2026)
by: Weng, Yuzhe, et al.
Published: (2026)
Advancing vision-language models in front-end development via data synthesis
by: Ge, Tong, et al.
Published: (2025)
by: Ge, Tong, et al.
Published: (2025)
LiDAR-based End-to-end Temporal Perception for Vehicle-Infrastructure Cooperation
by: Yang, Zhenwei, et al.
Published: (2024)
by: Yang, Zhenwei, et al.
Published: (2024)
End-to-end Topographic Auditory Models Replicate Signatures of Human Auditory Cortex
by: Al-Tahan, Haider, et al.
Published: (2025)
by: Al-Tahan, Haider, et al.
Published: (2025)
CLAP: Contrastive Latent-space Prompt Optimization for End-to-end Autonomous Driving
by: Zhu, Ruiyang, et al.
Published: (2026)
by: Zhu, Ruiyang, et al.
Published: (2026)
A multimodal deep learning architecture for smoking detection with a small data approach
by: Lakatos, Robert, et al.
Published: (2023)
by: Lakatos, Robert, et al.
Published: (2023)
Point2Graph: An End-to-end Point Cloud-based 3D Open-Vocabulary Scene Graph for Robot Navigation
by: Xu, Yifan, et al.
Published: (2024)
by: Xu, Yifan, et al.
Published: (2024)
PolyR-CNN: R-CNN for end-to-end polygonal building outline extraction
by: Jiao, Weiqin, et al.
Published: (2024)
by: Jiao, Weiqin, et al.
Published: (2024)
AI's Blind Spots: Geographic Knowledge and Diversity Deficit in Generated Urban Scenario
by: Beneduce, Ciro, et al.
Published: (2025)
by: Beneduce, Ciro, et al.
Published: (2025)
Adversarial Magnification to Deceive Deepfake Detection through Super Resolution
by: Coccomini, Davide Alessandro, et al.
Published: (2024)
by: Coccomini, Davide Alessandro, et al.
Published: (2024)
ImagenWorld: Stress-Testing Image Generation Models with Explainable Human Evaluation on Open-ended Real-World Tasks
by: Sani, Samin Mahdizadeh, et al.
Published: (2026)
by: Sani, Samin Mahdizadeh, et al.
Published: (2026)
Deepfake Detection without Deepfakes: Generalization via Synthetic Frequency Patterns Injection
by: Coccomini, Davide Alessandro, et al.
Published: (2024)
by: Coccomini, Davide Alessandro, et al.
Published: (2024)
Benchmarking Jetson Edge Devices with an End-to-end Video-based Anomaly Detection System
by: Pham, Hoang Viet, et al.
Published: (2023)
by: Pham, Hoang Viet, et al.
Published: (2023)
Reasoning over the Behaviour of Objects in Video-Clips for Adverb-Type Recognition
by: Seshadri, Amrit Diggavi, et al.
Published: (2023)
by: Seshadri, Amrit Diggavi, et al.
Published: (2023)
Multi-style conversion for semantic segmentation of lesions in fundus images by adversarial attacks
by: Playout, Clément, et al.
Published: (2024)
by: Playout, Clément, et al.
Published: (2024)
AC-MAMBASEG: An adaptive convolution and Mamba-based architecture for enhanced skin lesion segmentation
by: Nguyen, Viet-Thanh, et al.
Published: (2024)
by: Nguyen, Viet-Thanh, et al.
Published: (2024)
Similar Items
-
SynSur: An end-to-end generative pipeline for synthetic industrial surface defect generation and detection
by: Kühn, Paul Julius, et al.
Published: (2026) -
Stability Plasticity Decoupled Fine-tuning For Few-shot end-to-end Object Detection
by: Yin, Yuantao, et al.
Published: (2024) -
RoofSeg: An edge-aware transformer-based network for end-to-end roof plane segmentation
by: You, Siyuan, et al.
Published: (2025) -
ArtBrain: An Explainable end-to-end Toolkit for Classification and Attribution of AI-Generated Art and Style
by: Silva, Ravidu Suien Rammuni, et al.
Published: (2024) -
Validation of an AI-based end-to-end model for prostate pathology using long-term archived routine samples
by: Ji, Xiaoyi, et al.
Published: (2026)