Guardado en:
| Autores principales: | Tareen, Shaharyar Ahmed Khan, Tareen, Filza Khan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2510.10764 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
DeepDetect: Learning All-in-One Dense Keypoints
por: Tareen, Shaharyar Ahmed Khan, et al.
Publicado: (2025)
por: Tareen, Shaharyar Ahmed Khan, et al.
Publicado: (2025)
SOLAR: Switchable Output Layer for Accuracy and Robustness in Once-for-All Training
por: Tareen, Shaharyar Ahmed Khan, et al.
Publicado: (2025)
por: Tareen, Shaharyar Ahmed Khan, et al.
Publicado: (2025)
Balancing Accuracy and Efficiency: CNN Fusion Models for Diabetic Retinopathy Screening
por: Islam, Md Rafid, et al.
Publicado: (2025)
por: Islam, Md Rafid, et al.
Publicado: (2025)
Comparative Analysis of Deep Learning Models for Perception in Autonomous Vehicles
por: Khan, Jalal
Publicado: (2025)
por: Khan, Jalal
Publicado: (2025)
Adapting Large Multimodal Models to Distribution Shifts: The Role of In-Context Learning
por: Zhou, Guanglin, et al.
Publicado: (2024)
por: Zhou, Guanglin, et al.
Publicado: (2024)
WorldCache: Content-Aware Caching for Accelerated Video World Models
por: Nawaz, Umair, et al.
Publicado: (2026)
por: Nawaz, Umair, et al.
Publicado: (2026)
PILOT: Policy-Informed Learned Optimization for Adaptive Deep Network Training
por: Altuuaim, Sattam, et al.
Publicado: (2026)
por: Altuuaim, Sattam, et al.
Publicado: (2026)
Visual Attention Methods in Deep Learning: An In-Depth Survey
por: Hassanin, Mohammed, et al.
Publicado: (2022)
por: Hassanin, Mohammed, et al.
Publicado: (2022)
AttentionDrop: A Novel Regularization Method for Transformer Models
por: Baig, Mirza Samad Ahmed, et al.
Publicado: (2025)
por: Baig, Mirza Samad Ahmed, et al.
Publicado: (2025)
Enhancing Novel Object Detection via Cooperative Foundational Models
por: Bharadwaj, Rohit, et al.
Publicado: (2023)
por: Bharadwaj, Rohit, et al.
Publicado: (2023)
How Well Does GPT-4V(ision) Adapt to Distribution Shifts? A Preliminary Investigation
por: Han, Zhongyi, et al.
Publicado: (2023)
por: Han, Zhongyi, et al.
Publicado: (2023)
Enhancing Breast Cancer Diagnosis in Mammography: Evaluation and Integration of Convolutional Neural Networks and Explainable AI
por: Ahmed, Maryam, et al.
Publicado: (2024)
por: Ahmed, Maryam, et al.
Publicado: (2024)
Ensemble Deep Learning and LLM-Assisted Reporting for Automated Skin Lesion Diagnosis
por: Khan, Sher, et al.
Publicado: (2025)
por: Khan, Sher, et al.
Publicado: (2025)
Deep Networks Always Grok and Here is Why
por: Humayun, Ahmed Imtiaz, et al.
Publicado: (2024)
por: Humayun, Ahmed Imtiaz, et al.
Publicado: (2024)
Adversarial robustness of VAEs through the lens of local geometry
por: Khan, Asif, et al.
Publicado: (2022)
por: Khan, Asif, et al.
Publicado: (2022)
IKD+: Reliable Low Complexity Deep Models For Retinopathy Classification
por: Brahmavar, Shreyas Bhat, et al.
Publicado: (2023)
por: Brahmavar, Shreyas Bhat, et al.
Publicado: (2023)
Explaining Recovery Trajectories of Older Adults Post Lower-Limb Fracture Using Modality-wise Multiview Clustering and Large Language Models
por: Khan, Shehroz S., et al.
Publicado: (2025)
por: Khan, Shehroz S., et al.
Publicado: (2025)
Adapting Vision-Language Models for Evaluating World Models
por: Hendriksen, Mariya, et al.
Publicado: (2025)
por: Hendriksen, Mariya, et al.
Publicado: (2025)
Residual-SwinCA-Net: A Channel-Aware Integrated Residual CNN-Swin Transformer for Malignant Lesion Segmentation in BUSI
por: Naz, Saeeda, et al.
Publicado: (2025)
por: Naz, Saeeda, et al.
Publicado: (2025)
RS-CA-HSICT: A Residual and Spatial Channel Augmented CNN Transformer Framework for Monkeypox Detection
por: Iqbal, Rashid, et al.
Publicado: (2025)
por: Iqbal, Rashid, et al.
Publicado: (2025)
Training-Only Heterogeneous Image-Patch-Text Graph Supervision for Advancing Few-Shot Learning Adapters
por: Mohammad, Mohammed Rahman Sherif Khan, et al.
Publicado: (2026)
por: Mohammad, Mohammed Rahman Sherif Khan, et al.
Publicado: (2026)
TextPixs: Glyph-Conditioned Diffusion with Character-Aware Attention and OCR-Guided Supervision
por: Gillani, Syeda Anshrah, et al.
Publicado: (2025)
por: Gillani, Syeda Anshrah, et al.
Publicado: (2025)
PolyPath: Adapting a Large Multimodal Model for Multi-slide Pathology Report Generation
por: Ahmed, Faruk, et al.
Publicado: (2025)
por: Ahmed, Faruk, et al.
Publicado: (2025)
PAVE: Patching and Adapting Video Large Language Models
por: Liu, Zhuoming, et al.
Publicado: (2025)
por: Liu, Zhuoming, et al.
Publicado: (2025)
TFT-ACB-XML: Decision-Level Integration of Customized Temporal Fusion Transformer and Attention-BiLSTM with XGBoost Meta-Learner for BTC Price Forecasting
por: Din, Raiz Ud, et al.
Publicado: (2026)
por: Din, Raiz Ud, et al.
Publicado: (2026)
Adapt then Unlearn: Exploring Parameter Space Semantics for Unlearning in Generative Adversarial Networks
por: Tiwary, Piyush, et al.
Publicado: (2023)
por: Tiwary, Piyush, et al.
Publicado: (2023)
ECoDepth: Effective Conditioning of Diffusion Models for Monocular Depth Estimation
por: Patni, Suraj, et al.
Publicado: (2024)
por: Patni, Suraj, et al.
Publicado: (2024)
Impact of Tuning Parameters in Deep Convolutional Neural Network Using a Crack Image Dataset
por: Zabin, Mahe, et al.
Publicado: (2025)
por: Zabin, Mahe, et al.
Publicado: (2025)
Evaluation and Analysis of Deep Neural Transformers and Convolutional Neural Networks on Modern Remote Sensing Datasets
por: Hurt, J. Alex, et al.
Publicado: (2025)
por: Hurt, J. Alex, et al.
Publicado: (2025)
Universal Bovine Identification via Depth Data and Deep Metric Learning
por: Sharma, Asheesh, et al.
Publicado: (2024)
por: Sharma, Asheesh, et al.
Publicado: (2024)
AIN: The Arabic INclusive Large Multimodal Model
por: Heakl, Ahmed, et al.
Publicado: (2025)
por: Heakl, Ahmed, et al.
Publicado: (2025)
A Comprehensive Survey on Architectural Advances in Deep CNNs: Challenges, Applications, and Emerging Research Directions
por: Khan, Saddam Hussain, et al.
Publicado: (2025)
por: Khan, Saddam Hussain, et al.
Publicado: (2025)
Learning Regional Monsoon Patterns with a Multimodal Attention U-Net
por: Mazumder, Swaib Ilias, et al.
Publicado: (2025)
por: Mazumder, Swaib Ilias, et al.
Publicado: (2025)
DAUNet: A Lightweight UNet Variant with Deformable Convolutions and Parameter-Free Attention for Medical Image Segmentation
por: Munir, Adnan, et al.
Publicado: (2025)
por: Munir, Adnan, et al.
Publicado: (2025)
SemSegDepth: A Combined Model for Semantic Segmentation and Depth Completion
por: Lagos, Juan Pablo, et al.
Publicado: (2022)
por: Lagos, Juan Pablo, et al.
Publicado: (2022)
Promptception: How Sensitive Are Large Multimodal Models to Prompts?
por: Ismithdeen, Mohamed Insaf, et al.
Publicado: (2025)
por: Ismithdeen, Mohamed Insaf, et al.
Publicado: (2025)
Adapting Vision-Language Models Without Labels: A Comprehensive Survey
por: Dong, Hao, et al.
Publicado: (2025)
por: Dong, Hao, et al.
Publicado: (2025)
Adapting Segment Anything Model to Melanoma Segmentation in Microscopy Slide Images
por: Liu, Qingyuan, et al.
Publicado: (2024)
por: Liu, Qingyuan, et al.
Publicado: (2024)
Synthetic Data Generation Framework, Dataset, and Efficient Deep Model for Pedestrian Intention Prediction
por: Riaz, Muhammad Naveed, et al.
Publicado: (2024)
por: Riaz, Muhammad Naveed, et al.
Publicado: (2024)
BiDepth: A Bidirectional-Depth Neural Network for Spatio-Temporal Prediction
por: Ehsani, Sina, et al.
Publicado: (2025)
por: Ehsani, Sina, et al.
Publicado: (2025)
Ejemplares similares
-
DeepDetect: Learning All-in-One Dense Keypoints
por: Tareen, Shaharyar Ahmed Khan, et al.
Publicado: (2025) -
SOLAR: Switchable Output Layer for Accuracy and Robustness in Once-for-All Training
por: Tareen, Shaharyar Ahmed Khan, et al.
Publicado: (2025) -
Balancing Accuracy and Efficiency: CNN Fusion Models for Diabetic Retinopathy Screening
por: Islam, Md Rafid, et al.
Publicado: (2025) -
Comparative Analysis of Deep Learning Models for Perception in Autonomous Vehicles
por: Khan, Jalal
Publicado: (2025) -
Adapting Large Multimodal Models to Distribution Shifts: The Role of In-Context Learning
por: Zhou, Guanglin, et al.
Publicado: (2024)