Gespeichert in:
| Hauptverfasser: | Khalil, Ahmad, Khalil, Mahmoud, Ngom, Alioune |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2504.14429 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ResNetVLLM -- Multi-modal Vision LLM for the Video Understanding Task
von: Khalil, Ahmad, et al.
Veröffentlicht: (2025)
von: Khalil, Ahmad, et al.
Veröffentlicht: (2025)
Representation Learning with Adaptive Superpixel Coding
von: Khalil, Mahmoud, et al.
Veröffentlicht: (2025)
von: Khalil, Mahmoud, et al.
Veröffentlicht: (2025)
Expand VSR Benchmark for VLLM to Expertize in Spatial Rules
von: Xie, Peijin, et al.
Veröffentlicht: (2024)
von: Xie, Peijin, et al.
Veröffentlicht: (2024)
PhyVLLM: Physics-Guided Video Language Model with Motion-Appearance Disentanglement
von: Zhan, Yu-Wei, et al.
Veröffentlicht: (2025)
von: Zhan, Yu-Wei, et al.
Veröffentlicht: (2025)
MFI-ResNet: Efficient ResNet Architecture Optimization via MeanFlow Compression and Selective Incubation
von: Sun, Nuolin, et al.
Veröffentlicht: (2025)
von: Sun, Nuolin, et al.
Veröffentlicht: (2025)
Res2NetFuse: A Novel Res2Net-based Fusion Method for Infrared and Visible Images
von: Song, Xu, et al.
Veröffentlicht: (2021)
von: Song, Xu, et al.
Veröffentlicht: (2021)
B-VLLM: A Vision Large Language Model with Balanced Spatio-Temporal Tokens
von: Lu, Zhuqiang, et al.
Veröffentlicht: (2024)
von: Lu, Zhuqiang, et al.
Veröffentlicht: (2024)
VISion On Request: Enhanced VLLM efficiency with sparse, dynamically selected, vision-language interactions
von: Bulat, Adrian, et al.
Veröffentlicht: (2026)
von: Bulat, Adrian, et al.
Veröffentlicht: (2026)
Quotient Network -- A Network Similar to ResNet but Learning Quotients
von: Hui, Peng, et al.
Veröffentlicht: (2025)
von: Hui, Peng, et al.
Veröffentlicht: (2025)
Interpreting ResNet-based CLIP via Neuron-Attention Decomposition
von: Bu, Edmund, et al.
Veröffentlicht: (2025)
von: Bu, Edmund, et al.
Veröffentlicht: (2025)
Wavelet-based GAN Fingerprint Detection using ResNet50
von: Erukude, Sai Teja, et al.
Veröffentlicht: (2025)
von: Erukude, Sai Teja, et al.
Veröffentlicht: (2025)
ResNet: Enabling Deep Convolutional Neural Networks through Residual Learning
von: Liu, Xingyu, et al.
Veröffentlicht: (2025)
von: Liu, Xingyu, et al.
Veröffentlicht: (2025)
Multimodal Detection of Fake Reviews using BERT and ResNet-50
von: Veluru, Suhasnadh Reddy, et al.
Veröffentlicht: (2025)
von: Veluru, Suhasnadh Reddy, et al.
Veröffentlicht: (2025)
Research on Brain Tumor Classification Method Based on Improved ResNet34 Network
von: Li, Yufeng, et al.
Veröffentlicht: (2025)
von: Li, Yufeng, et al.
Veröffentlicht: (2025)
Transfer Learning for Wildlife Classification: Evaluating YOLOv8 against DenseNet, ResNet, and VGGNet on a Custom Dataset
von: Sharma, Subek, et al.
Veröffentlicht: (2024)
von: Sharma, Subek, et al.
Veröffentlicht: (2024)
XAI-Driven Skin Disease Classification: Leveraging GANs to Augment ResNet-50 Performance
von: Villanueva, Kim Gerard A., et al.
Veröffentlicht: (2025)
von: Villanueva, Kim Gerard A., et al.
Veröffentlicht: (2025)
Acute Lymphoblastic Leukemia Diagnosis Employing YOLOv11, YOLOv8, ResNet50, and Inception-ResNet-v2 Deep Learning Models
von: Awad, Alaa, et al.
Veröffentlicht: (2025)
von: Awad, Alaa, et al.
Veröffentlicht: (2025)
ResAF-Net: An Anchor-Free Attention-Based Network for Tree Detection and Agricultural Mapping in Palestine
von: Al-Qasem, Rabee
Veröffentlicht: (2026)
von: Al-Qasem, Rabee
Veröffentlicht: (2026)
Benchmarking ResNet Backbones in RT-DETR: Impact of Depth and Regularization under environmental conditions
von: Barboza, Pamela, et al.
Veröffentlicht: (2026)
von: Barboza, Pamela, et al.
Veröffentlicht: (2026)
Brain Tumor Classifiers Under Attack: Robustness of ResNet Variants Against Transferable FGSM and PGD Attacks
von: Deem, Ryan, et al.
Veröffentlicht: (2026)
von: Deem, Ryan, et al.
Veröffentlicht: (2026)
Multi-Phase Automated Segmentation of Dental Structures in CBCT Using a Lightweight Auto3DSeg and SegResNet Implementation
von: LaBella, Dominic, et al.
Veröffentlicht: (2025)
von: LaBella, Dominic, et al.
Veröffentlicht: (2025)
ResNet-34 with Lightweight Decoder for Accurate and Efficient Segmentation of Fetal Brain MRI
von: Rahman, Ashiqur, et al.
Veröffentlicht: (2026)
von: Rahman, Ashiqur, et al.
Veröffentlicht: (2026)
Hybrid ResNet-1D-BiGRU with Multi-Head Attention for Cyberattack Detection in Industrial IoT Environments
von: Gueriani, Afrah, et al.
Veröffentlicht: (2026)
von: Gueriani, Afrah, et al.
Veröffentlicht: (2026)
An Explainable Two Stage Deep Learning Framework for Pericoronitis Assessment in Panoramic Radiographs Using YOLOv8 and ResNet-50
von: George, Ajo Babu, et al.
Veröffentlicht: (2026)
von: George, Ajo Babu, et al.
Veröffentlicht: (2026)
Explainable AI: Comparative Analysis of Normal and Dilated ResNet Models for Fundus Disease Classification
von: Karthikayan, P. N., et al.
Veröffentlicht: (2024)
von: Karthikayan, P. N., et al.
Veröffentlicht: (2024)
Brain Ageing Prediction using Isolation Forest Technique and Residual Neural Network (ResNet)
von: Behzadi, Saadat, et al.
Veröffentlicht: (2024)
von: Behzadi, Saadat, et al.
Veröffentlicht: (2024)
SynA-ResNet: Spike-driven ResNet Achieved through OR Residual Connection
von: Shan, Yimeng, et al.
Veröffentlicht: (2023)
von: Shan, Yimeng, et al.
Veröffentlicht: (2023)
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense
von: Guo, Yangyang, et al.
Veröffentlicht: (2024)
von: Guo, Yangyang, et al.
Veröffentlicht: (2024)
PlantDiseaseNet-RT50: A Fine-tuned ResNet50 Architecture for High-Accuracy Plant Disease Detection Beyond Standard CNNs
von: Sagnika, Santwana, et al.
Veröffentlicht: (2025)
von: Sagnika, Santwana, et al.
Veröffentlicht: (2025)
FlexEdit: Marrying Free-Shape Masks to VLLM for Flexible Image Editing
von: Yuan, Tianshuo, et al.
Veröffentlicht: (2024)
von: Yuan, Tianshuo, et al.
Veröffentlicht: (2024)
Precise Shield: Explaining and Aligning VLLM Safety via Neuron-Level Guidance
von: Shi, Enyi, et al.
Veröffentlicht: (2026)
von: Shi, Enyi, et al.
Veröffentlicht: (2026)
Detection of pulmonary pathologies using convolutional neural networks, Data Augmentation, ResNet50 and Vision Transformers
von: Amador, Pablo Ramirez, et al.
Veröffentlicht: (2024)
von: Amador, Pablo Ramirez, et al.
Veröffentlicht: (2024)
A Neural Architecture Search Method using Auxiliary Evaluation Metric based on ResNet Architecture
von: Wang, Shang, et al.
Veröffentlicht: (2025)
von: Wang, Shang, et al.
Veröffentlicht: (2025)
ViGText: Deepfake Image Detection with Vision-Language Model Explanations and Graph Neural Networks
von: ALBarqawi, Ahmad, et al.
Veröffentlicht: (2025)
von: ALBarqawi, Ahmad, et al.
Veröffentlicht: (2025)
UMambaAdj: Advancing GTV Segmentation for Head and Neck Cancer in MRI-Guided RT with UMamba and nnU-Net ResEnc Planner
von: Ren, Jintao, et al.
Veröffentlicht: (2024)
von: Ren, Jintao, et al.
Veröffentlicht: (2024)
ResAdapt: Adaptive Resolution for Efficient Multimodal Reasoning
von: Liao, Huanxuan, et al.
Veröffentlicht: (2026)
von: Liao, Huanxuan, et al.
Veröffentlicht: (2026)
SpaceVLLM: Endowing Multimodal Large Language Model with Spatio-Temporal Video Grounding Capability
von: Wang, Jiankang, et al.
Veröffentlicht: (2025)
von: Wang, Jiankang, et al.
Veröffentlicht: (2025)
Lung Cancer Classification from CT Images Using ResNet
von: Adekunle, Olajumoke O., et al.
Veröffentlicht: (2025)
von: Adekunle, Olajumoke O., et al.
Veröffentlicht: (2025)
PushPull-Net: Inhibition-driven ResNet robust to image corruptions
von: Bennabhaktula, Guru Swaroop, et al.
Veröffentlicht: (2024)
von: Bennabhaktula, Guru Swaroop, et al.
Veröffentlicht: (2024)
GCA-ResUNet:Image segmentation in medical images using grouped coordinate attention
von: Ding, Jun, et al.
Veröffentlicht: (2025)
von: Ding, Jun, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
ResNetVLLM -- Multi-modal Vision LLM for the Video Understanding Task
von: Khalil, Ahmad, et al.
Veröffentlicht: (2025) -
Representation Learning with Adaptive Superpixel Coding
von: Khalil, Mahmoud, et al.
Veröffentlicht: (2025) -
Expand VSR Benchmark for VLLM to Expertize in Spatial Rules
von: Xie, Peijin, et al.
Veröffentlicht: (2024) -
PhyVLLM: Physics-Guided Video Language Model with Motion-Appearance Disentanglement
von: Zhan, Yu-Wei, et al.
Veröffentlicht: (2025) -
MFI-ResNet: Efficient ResNet Architecture Optimization via MeanFlow Compression and Selective Incubation
von: Sun, Nuolin, et al.
Veröffentlicht: (2025)