Banana Ripeness Level Classification using a Simple CNN Model Trained with Real and Synthetic Datasets
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chuquimarca, Luis, Vintimilla, Boris, Velastin, Sergio |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Learning Sign Language Representation using CNN LSTM, 3DCNN, CNN RNN LSTM and CCN TD
von: Louison, Nikita, et al.
Veröffentlicht: (2024)
von: Louison, Nikita, et al.
Veröffentlicht: (2024)
Evaluating Visual Mathematics in Multimodal LLMs: A Multilingual Benchmark Based on the Kangaroo Tests
von: Sáez, Arnau Igualde, et al.
Veröffentlicht: (2025)
von: Sáez, Arnau Igualde, et al.
Veröffentlicht: (2025)
Image-based Facial Rig Inversion
von: Yang, Tianxiang, et al.
Veröffentlicht: (2025)
von: Yang, Tianxiang, et al.
Veröffentlicht: (2025)
Combined Hyperbolic and Euclidean Soft Triple Loss Beyond the Single Space Deep Metric Learning
von: Saeki, Shozo, et al.
Veröffentlicht: (2025)
von: Saeki, Shozo, et al.
Veröffentlicht: (2025)
TowerVision: Understanding and Improving Multilinguality in Vision-Language Models
von: Viveiros, André G., et al.
Veröffentlicht: (2025)
von: Viveiros, André G., et al.
Veröffentlicht: (2025)
Unpacking Hateful Memes: Presupposed Context and False Claims
von: Cai, Weibin, et al.
Veröffentlicht: (2025)
von: Cai, Weibin, et al.
Veröffentlicht: (2025)
Does CLIP perceive art the same way we do?
von: Asperti, Andrea, et al.
Veröffentlicht: (2025)
von: Asperti, Andrea, et al.
Veröffentlicht: (2025)
Anonymization-Enhanced Privacy Protection for Mobile GUI Agents: Available but Invisible
von: Zhao, Lepeng, et al.
Veröffentlicht: (2026)
von: Zhao, Lepeng, et al.
Veröffentlicht: (2026)
Think, Act, Learn: A Framework for Autonomous Robotic Agents using Closed-Loop Large Language Models
von: Menon, Anjali R., et al.
Veröffentlicht: (2025)
von: Menon, Anjali R., et al.
Veröffentlicht: (2025)
IDOL: Instant Photorealistic 3D Human Creation from a Single Image
von: Zhuang, Yiyu, et al.
Veröffentlicht: (2024)
von: Zhuang, Yiyu, et al.
Veröffentlicht: (2024)
A Landmark-Aware Visual Navigation Dataset
von: Johnson, Faith, et al.
Veröffentlicht: (2024)
von: Johnson, Faith, et al.
Veröffentlicht: (2024)
Exploring the Capabilities of Large Language Model Encoders for Image-Text Retrieval in Chest X-rays
von: Ko, Hanbin, et al.
Veröffentlicht: (2025)
von: Ko, Hanbin, et al.
Veröffentlicht: (2025)
Dual-sensing driving detection model
von: K, Leon C. C., et al.
Veröffentlicht: (2025)
von: K, Leon C. C., et al.
Veröffentlicht: (2025)
DOD-SA: Infrared-Visible Decoupled Object Detection with Single-Modality Annotations
von: Jin, Hang, et al.
Veröffentlicht: (2025)
von: Jin, Hang, et al.
Veröffentlicht: (2025)
HuMoCon: Concept Discovery for Human Motion Understanding
von: Fang, Qihang, et al.
Veröffentlicht: (2025)
von: Fang, Qihang, et al.
Veröffentlicht: (2025)
Transforming faces into video stories -- VideoFace2.0
von: Brkljač, Branko, et al.
Veröffentlicht: (2025)
von: Brkljač, Branko, et al.
Veröffentlicht: (2025)
TableMoE: Neuro-Symbolic Routing for Structured Expert Reasoning in Multimodal Table Understanding
von: Zhang, Junwen, et al.
Veröffentlicht: (2025)
von: Zhang, Junwen, et al.
Veröffentlicht: (2025)
PathFormer: A Transformer with 3D Grid Constraints for Digital Twin Robot-Arm Trajectory Generation
von: Alanazi, Ahmed, et al.
Veröffentlicht: (2025)
von: Alanazi, Ahmed, et al.
Veröffentlicht: (2025)
Inducing Causal World Models in LLMs for Zero-Shot Physical Reasoning
von: Sharma, Aditya, et al.
Veröffentlicht: (2025)
von: Sharma, Aditya, et al.
Veröffentlicht: (2025)
Perception-Consistency Multimodal Large Language Models Reasoning via Caption-Regularized Policy Optimization
von: Tu, Songjun, et al.
Veröffentlicht: (2025)
von: Tu, Songjun, et al.
Veröffentlicht: (2025)
Classifying Healthy and Defective Fruits with a Multi-Input Architecture and CNN Models
von: Chuquimarca, Luis, et al.
Veröffentlicht: (2024)
von: Chuquimarca, Luis, et al.
Veröffentlicht: (2024)
Biomedical Visual Instruction Tuning with Clinician Preference Alignment
von: Cui, Hejie, et al.
Veröffentlicht: (2024)
von: Cui, Hejie, et al.
Veröffentlicht: (2024)
GLL: A Differentiable Graph Learning Layer for Neural Networks
von: Brown, Jason, et al.
Veröffentlicht: (2024)
von: Brown, Jason, et al.
Veröffentlicht: (2024)
Generating Natural-Language Surgical Feedback: From Structured Representation to Domain-Grounded Evaluation
von: Nasriddinov, Firdavs, et al.
Veröffentlicht: (2025)
von: Nasriddinov, Firdavs, et al.
Veröffentlicht: (2025)
Beyond RGB: Leveraging Vision Transformers for Thermal Weapon Segmentation
von: Kambhatla, Akhila, et al.
Veröffentlicht: (2025)
von: Kambhatla, Akhila, et al.
Veröffentlicht: (2025)
Tricks and Plug-ins for Gradient Boosting in Image Classification
von: Fang, Biyi, et al.
Veröffentlicht: (2025)
von: Fang, Biyi, et al.
Veröffentlicht: (2025)
TaylorShift: Shifting the Complexity of Self-Attention from Squared to Linear (and Back) using Taylor-Softmax
von: Nauen, Tobias Christian, et al.
Veröffentlicht: (2024)
von: Nauen, Tobias Christian, et al.
Veröffentlicht: (2024)
ShapBPT: Image Feature Attributions Using Data-Aware Binary Partition Trees
von: Rashid, Muhammad, et al.
Veröffentlicht: (2026)
von: Rashid, Muhammad, et al.
Veröffentlicht: (2026)
Contrastive Consolidation of Top-Down Modulations Achieves Sparsely Supervised Continual Learning
von: Tran, Viet Anh Khoa, et al.
Veröffentlicht: (2025)
von: Tran, Viet Anh Khoa, et al.
Veröffentlicht: (2025)
Geometric-Stochastic Multimodal Deep Learning for Predictive Modeling of SUDEP and Stroke Vulnerability
von: Girish, Preksha, et al.
Veröffentlicht: (2025)
von: Girish, Preksha, et al.
Veröffentlicht: (2025)
CADE 2.5 - ZeResFDG: Frequency-Decoupled, Rescaled and Zero-Projected Guidance for SD/SDXL Latent Diffusion Models
von: Rychkovskiy, Denis
Veröffentlicht: (2025)
von: Rychkovskiy, Denis
Veröffentlicht: (2025)
MB-DSMIL-CL-PL: Scalable Weakly Supervised Ovarian Cancer Subtype Classification and Localisation Using Contrastive and Prototype Learning with Frozen Patch Features
von: Jenkins, Marcus, et al.
Veröffentlicht: (2026)
von: Jenkins, Marcus, et al.
Veröffentlicht: (2026)
Sat-JEPA-Diff: Bridging Self-Supervised Learning and Generative Diffusion for Remote Sensing
von: Komurcu, Kursat, et al.
Veröffentlicht: (2026)
von: Komurcu, Kursat, et al.
Veröffentlicht: (2026)
A deep learning approach to track eye movements based on events
von: Seth, Chirag, et al.
Veröffentlicht: (2025)
von: Seth, Chirag, et al.
Veröffentlicht: (2025)
Multi-Modal Self-Supervised Learning for Surgical Feedback Effectiveness Assessment
von: Gupta, Arushi, et al.
Veröffentlicht: (2024)
von: Gupta, Arushi, et al.
Veröffentlicht: (2024)
Generative AI Models: Opportunities and Risks for Industry and Authorities
von: Alt, Tobias, et al.
Veröffentlicht: (2024)
von: Alt, Tobias, et al.
Veröffentlicht: (2024)
ROI-GS: Interest-based Local Quality 3D Gaussian Splatting
von: Bui, Quoc-Anh, et al.
Veröffentlicht: (2025)
von: Bui, Quoc-Anh, et al.
Veröffentlicht: (2025)
ROI-NeRFs: Hi-Fi Visualization of Objects of Interest within a Scene by NeRFs Composition
von: Bui, Quoc-Anh, et al.
Veröffentlicht: (2025)
von: Bui, Quoc-Anh, et al.
Veröffentlicht: (2025)
ParaRNN: Unlocking Parallel Training of Nonlinear RNNs for Large Language Models
von: Danieli, Federico, et al.
Veröffentlicht: (2025)
von: Danieli, Federico, et al.
Veröffentlicht: (2025)
Advancing Brain Tumor Segmentation via Attention-based 3D U-Net Architecture and Digital Image Processing
von: Gad, Eyad, et al.
Veröffentlicht: (2025)
von: Gad, Eyad, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Learning Sign Language Representation using CNN LSTM, 3DCNN, CNN RNN LSTM and CCN TD
von: Louison, Nikita, et al.
Veröffentlicht: (2024) -
Evaluating Visual Mathematics in Multimodal LLMs: A Multilingual Benchmark Based on the Kangaroo Tests
von: Sáez, Arnau Igualde, et al.
Veröffentlicht: (2025) -
Image-based Facial Rig Inversion
von: Yang, Tianxiang, et al.
Veröffentlicht: (2025) -
Combined Hyperbolic and Euclidean Soft Triple Loss Beyond the Single Space Deep Metric Learning
von: Saeki, Shozo, et al.
Veröffentlicht: (2025) -
TowerVision: Understanding and Improving Multilinguality in Vision-Language Models
von: Viveiros, André G., et al.
Veröffentlicht: (2025)