An interpretable framework using foundation models for fish sex identification
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Miao, Zheng, Hung, Tien-Chieh |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Enabling clinical use of foundation models for computational pathology
von: Henriksen, Audun L, et al.
Veröffentlicht: (2026)
von: Henriksen, Audun L, et al.
Veröffentlicht: (2026)
Generative deep learning for foundational video translation in ultrasound
von: Tomic, Nikolina, et al.
Veröffentlicht: (2025)
von: Tomic, Nikolina, et al.
Veröffentlicht: (2025)
A multimodal vision foundation model for generalizable knee pathology
von: Yu, Kang, et al.
Veröffentlicht: (2026)
von: Yu, Kang, et al.
Veröffentlicht: (2026)
Driving scenario generation and evaluation using a structured layer representation and foundational models
von: Hubert, Arthur, et al.
Veröffentlicht: (2025)
von: Hubert, Arthur, et al.
Veröffentlicht: (2025)
Deep learning framework for crater detection and identification on the Moon and Mars
von: Ma, Yihan, et al.
Veröffentlicht: (2025)
von: Ma, Yihan, et al.
Veröffentlicht: (2025)
Thinker: A vision-language foundation model for embodied intelligence
von: Pan, Baiyu, et al.
Veröffentlicht: (2026)
von: Pan, Baiyu, et al.
Veröffentlicht: (2026)
Prompting with the human-touch: evaluating model-sensitivity of foundation models for musculoskeletal CT segmentation
von: Magg, Caroline, et al.
Veröffentlicht: (2026)
von: Magg, Caroline, et al.
Veröffentlicht: (2026)
ActiveMark: on watermarking of visual foundation models via massive activations
von: Chistyakova, Anna, et al.
Veröffentlicht: (2025)
von: Chistyakova, Anna, et al.
Veröffentlicht: (2025)
An interpretable imbalanced semi-supervised deep learning framework for improving differential diagnosis of skin diseases
von: Weng, Futian, et al.
Veröffentlicht: (2022)
von: Weng, Futian, et al.
Veröffentlicht: (2022)
Attention-based multiple instance learning for predominant growth pattern prediction in lung adenocarcinoma wsi using foundation models
von: Perez-Herrera, Laura Valeria, et al.
Veröffentlicht: (2026)
von: Perez-Herrera, Laura Valeria, et al.
Veröffentlicht: (2026)
Developing a foundation model for high-resolution remote sensing data of the Netherlands
von: Vermeeren, Paul, et al.
Veröffentlicht: (2026)
von: Vermeeren, Paul, et al.
Veröffentlicht: (2026)
Near, far: Patch-ordering enhances vision foundation models' scene understanding
von: Pariza, Valentinos, et al.
Veröffentlicht: (2024)
von: Pariza, Valentinos, et al.
Veröffentlicht: (2024)
QuarterMap: Efficient Post-Training Token Pruning for Visual State Space Models
von: Chi, Tien-Yu, et al.
Veröffentlicht: (2025)
von: Chi, Tien-Yu, et al.
Veröffentlicht: (2025)
Paving the way toward foundation models for irregular and unaligned Satellite Image Time Series
von: Dumeur, Iris, et al.
Veröffentlicht: (2024)
von: Dumeur, Iris, et al.
Veröffentlicht: (2024)
Toward explainable AI approaches for breast imaging: adapting foundation models to diverse populations
von: Cavalcante, Guilherme J., et al.
Veröffentlicht: (2025)
von: Cavalcante, Guilherme J., et al.
Veröffentlicht: (2025)
Benchmarking histopathology foundation models in a multi-center dataset for skin cancer subtyping
von: Meseguer, Pablo, et al.
Veröffentlicht: (2025)
von: Meseguer, Pablo, et al.
Veröffentlicht: (2025)
Decompose the model: Mechanistic interpretability in image models with Generalized Integrated Gradients (GIG)
von: Kim, Yearim, et al.
Veröffentlicht: (2024)
von: Kim, Yearim, et al.
Veröffentlicht: (2024)
Low-Field Magnetic Resonance Image Quality Enhancement using a Conditional Flow Matching Model
von: Nguyen, Huu Tien, et al.
Veröffentlicht: (2025)
von: Nguyen, Huu Tien, et al.
Veröffentlicht: (2025)
EyeCLIP: A visual-language foundation model for multi-modal ophthalmic image analysis
von: Shi, Danli, et al.
Veröffentlicht: (2024)
von: Shi, Danli, et al.
Veröffentlicht: (2024)
Geospatial foundation models for image analysis: evaluating and enhancing NASA-IBM Prithvi's domain adaptability
von: Hsu, Chia-Yu, et al.
Veröffentlicht: (2024)
von: Hsu, Chia-Yu, et al.
Veröffentlicht: (2024)
DA-SSL: self-supervised domain adaptor to leverage foundational models in turbt histopathology slides
von: Zhang, Haoyue, et al.
Veröffentlicht: (2025)
von: Zhang, Haoyue, et al.
Veröffentlicht: (2025)
An analysis of HOI: using a training-free method with multimodal visual foundation models when only the test set is available, without the training set
von: Ai, Chaoyi
Veröffentlicht: (2024)
von: Ai, Chaoyi
Veröffentlicht: (2024)
Enhancing Interpretability of Vertebrae Fracture Grading using Human-interpretable Prototypes
von: Sinhamahapatra, Poulami, et al.
Veröffentlicht: (2024)
von: Sinhamahapatra, Poulami, et al.
Veröffentlicht: (2024)
PB-IAD: Utilizing multimodal foundation models for semantic industrial anomaly detection in dynamic manufacturing environments
von: Hofmann, Bernd, et al.
Veröffentlicht: (2025)
von: Hofmann, Bernd, et al.
Veröffentlicht: (2025)
Training a high-performance retinal foundation model with half-the-data and 400 times less compute
von: Engelmann, Justin, et al.
Veröffentlicht: (2024)
von: Engelmann, Justin, et al.
Veröffentlicht: (2024)
AttnMod: Attention-Based New Art Styles
von: Su, Shih-Chieh
Veröffentlicht: (2024)
von: Su, Shih-Chieh
Veröffentlicht: (2024)
Revisiting semi-supervised learning in the era of foundation models
von: Zhang, Ping, et al.
Veröffentlicht: (2025)
von: Zhang, Ping, et al.
Veröffentlicht: (2025)
Cognitive resilience: Unraveling the proficiency of image-captioning models to interpret masked visual content
von: Du, Zhicheng, et al.
Veröffentlicht: (2024)
von: Du, Zhicheng, et al.
Veröffentlicht: (2024)
Leveraging AI multimodal geospatial foundation models for improved near-real-time flood mapping at a global scale
von: Tulbure, Mirela G., et al.
Veröffentlicht: (2025)
von: Tulbure, Mirela G., et al.
Veröffentlicht: (2025)
Full end-to-end diagnostic workflow automation of 3D OCT via foundation model-driven AI for retinal diseases
von: Zhang, Jinze, et al.
Veröffentlicht: (2026)
von: Zhang, Jinze, et al.
Veröffentlicht: (2026)
TasselNetV4: A vision foundation model for cross-scene, cross-scale, and cross-species plant counting
von: Hu, Xiaonan, et al.
Veröffentlicht: (2025)
von: Hu, Xiaonan, et al.
Veröffentlicht: (2025)
V"Mean"ba: Visual State Space Models only need 1 hidden dimension
von: Chi, Tien-Yu, et al.
Veröffentlicht: (2024)
von: Chi, Tien-Yu, et al.
Veröffentlicht: (2024)
Training Deep Visual Networks Beyond Loss and Accuracy Through a Dynamical Systems Approach
von: La Quang, Hai, et al.
Veröffentlicht: (2026)
von: La Quang, Hai, et al.
Veröffentlicht: (2026)
TCMM: Token Constraint and Multi-Scale Memory Bank of Contrastive Learning for Unsupervised Person Re-identification
von: Zhu, Zheng-An, et al.
Veröffentlicht: (2025)
von: Zhu, Zheng-An, et al.
Veröffentlicht: (2025)
Grounding-Aware Token Pruning: Recovering from Drastic Performance Drops in Visual Grounding Caused by Pruning
von: Chien, Tzu-Chun, et al.
Veröffentlicht: (2025)
von: Chien, Tzu-Chun, et al.
Veröffentlicht: (2025)
Multi-Modal interpretable automatic video captioning
von: Hanna-Asaad, Antoine, et al.
Veröffentlicht: (2024)
von: Hanna-Asaad, Antoine, et al.
Veröffentlicht: (2024)
Explaning with trees: interpreting CNNs using hierarchies
von: Rodrigues, Caroline Mazini, et al.
Veröffentlicht: (2024)
von: Rodrigues, Caroline Mazini, et al.
Veröffentlicht: (2024)
Re-identification from histopathology images
von: Ganz, Jonathan, et al.
Veröffentlicht: (2024)
von: Ganz, Jonathan, et al.
Veröffentlicht: (2024)
ReynoldsFlow: Exquisite Flow Estimation via Reynolds Transport Theorem
von: Chen, Yu-Hsi, et al.
Veröffentlicht: (2025)
von: Chen, Yu-Hsi, et al.
Veröffentlicht: (2025)
Robustness Evaluation of OCR-based Visual Document Understanding under Multi-Modal Adversarial Attacks
von: Tien, Dong Nguyen, et al.
Veröffentlicht: (2025)
von: Tien, Dong Nguyen, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Enabling clinical use of foundation models for computational pathology
von: Henriksen, Audun L, et al.
Veröffentlicht: (2026) -
Generative deep learning for foundational video translation in ultrasound
von: Tomic, Nikolina, et al.
Veröffentlicht: (2025) -
A multimodal vision foundation model for generalizable knee pathology
von: Yu, Kang, et al.
Veröffentlicht: (2026) -
Driving scenario generation and evaluation using a structured layer representation and foundational models
von: Hubert, Arthur, et al.
Veröffentlicht: (2025) -
Deep learning framework for crater detection and identification on the Moon and Mars
von: Ma, Yihan, et al.
Veröffentlicht: (2025)