BioCAP: Exploiting Synthetic Captions Beyond Labels in Biological Foundation Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Zhang, Ziheng, Ma, Xinyue, Chowdhury, Arpita, Campolongo, Elizabeth G., Thompson, Matthew J., Zhang, Net, Stevens, Samuel, Lapp, Hilmar, Berger-Wolf, Tanya, Su, Yu, Chao, Wei-Lun, Gu, Jianyang |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
BioCLIP 2: Emergent Properties from Scaling Hierarchical Contrastive Learning
di: Gu, Jianyang, et al.
Pubblicazione: (2025)
di: Gu, Jianyang, et al.
Pubblicazione: (2025)
Finer-CAM: Spotting the Difference Reveals Finer Details for Visual Explanation
di: Zhang, Ziheng, et al.
Pubblicazione: (2025)
di: Zhang, Ziheng, et al.
Pubblicazione: (2025)
Prompt-CAM: Making Vision Transformers Interpretable for Fine-Grained Analysis
di: Chowdhury, Arpita, et al.
Pubblicazione: (2025)
di: Chowdhury, Arpita, et al.
Pubblicazione: (2025)
Static Segmentation by Tracking: A Label-Efficient Approach for Fine-Grained Specimen Image Segmentation
di: Feng, Zhenyang, et al.
Pubblicazione: (2025)
di: Feng, Zhenyang, et al.
Pubblicazione: (2025)
BioCLIP: A Vision Foundation Model for the Tree of Life
di: Stevens, Samuel, et al.
Pubblicazione: (2023)
di: Stevens, Samuel, et al.
Pubblicazione: (2023)
VLM4Bio: A Benchmark Dataset to Evaluate Pretrained Vision-Language Models for Trait Discovery from Biological Images
di: Maruf, M., et al.
Pubblicazione: (2024)
di: Maruf, M., et al.
Pubblicazione: (2024)
Interpretable and Testable Vision Features via Sparse Autoencoders
di: Stevens, Samuel, et al.
Pubblicazione: (2025)
di: Stevens, Samuel, et al.
Pubblicazione: (2025)
What Do You See in Common? Learning Hierarchical Prototypes over Tree-of-Life to Discover Evolutionary Traits
di: Manogaran, Harish Babu, et al.
Pubblicazione: (2024)
di: Manogaran, Harish Babu, et al.
Pubblicazione: (2024)
Fish-Vista: A Multi-Purpose Dataset for Understanding & Identification of Traits from Images
di: Mehrab, Kazi Sajeed, et al.
Pubblicazione: (2024)
di: Mehrab, Kazi Sajeed, et al.
Pubblicazione: (2024)
Towards Open-Ended Visual Scientific Discovery with Sparse Autoencoders
di: Stevens, Samuel, et al.
Pubblicazione: (2025)
di: Stevens, Samuel, et al.
Pubblicazione: (2025)
Image DataPalooza 2023 (Archive)
di: Lapp, Hilmar, et al.
Pubblicazione: (2024)
di: Lapp, Hilmar, et al.
Pubblicazione: (2024)
ReflectCAP: Detailed Image Captioning with Reflective Memory
di: Min, Kyungmin, et al.
Pubblicazione: (2026)
di: Min, Kyungmin, et al.
Pubblicazione: (2026)
Exploiting Pseudo Image Captions for Multimodal Summarization
di: Jiang, Chaoya, et al.
Pubblicazione: (2023)
di: Jiang, Chaoya, et al.
Pubblicazione: (2023)
Lessons and Open Questions from a Unified Study of Camera-Trap Species Recognition Over Time
di: Jeon, Sooyoung, et al.
Pubblicazione: (2026)
di: Jeon, Sooyoung, et al.
Pubblicazione: (2026)
Leveraging Latent Visual Reasoning in Silence
di: Zhu, Dongyao, et al.
Pubblicazione: (2026)
di: Zhu, Dongyao, et al.
Pubblicazione: (2026)
Fine-Tuning is Fine, if Calibrated
di: Mai, Zheda, et al.
Pubblicazione: (2024)
di: Mai, Zheda, et al.
Pubblicazione: (2024)
A continental-scale dataset of ground beetles with high-resolution images and validated morphological trait measurements
di: Rayeed, S M, et al.
Pubblicazione: (2026)
di: Rayeed, S M, et al.
Pubblicazione: (2026)
A Simple Interpretable Transformer for Fine-Grained Image Classification and Analysis
di: Paul, Dipanjyoti, et al.
Pubblicazione: (2023)
di: Paul, Dipanjyoti, et al.
Pubblicazione: (2023)
Exploiting Auxiliary Caption for Video Grounding
di: Li, Hongxiang, et al.
Pubblicazione: (2023)
di: Li, Hongxiang, et al.
Pubblicazione: (2023)
EUFCC-CIR: a Composed Image Retrieval Dataset for GLAM Collections
di: Net, Francesc, et al.
Pubblicazione: (2024)
di: Net, Francesc, et al.
Pubblicazione: (2024)
Hierarchical Conditioning of Diffusion Models Using Tree-of-Life for Studying Species Evolution
di: Khurana, Mridul, et al.
Pubblicazione: (2024)
di: Khurana, Mridul, et al.
Pubblicazione: (2024)
kabr-tools: Automated Framework for Multi-Species Behavioral Monitoring
di: Kline, Jenna, et al.
Pubblicazione: (2025)
di: Kline, Jenna, et al.
Pubblicazione: (2025)
OmniCaptioner: One Captioner to Rule Them All
di: Lu, Yiting, et al.
Pubblicazione: (2025)
di: Lu, Yiting, et al.
Pubblicazione: (2025)
CoordSpeaker: Exploiting Gesture Captioning for Coordinated Caption-Empowered Co-Speech Gesture Generation
di: Fang, Fengyi, et al.
Pubblicazione: (2025)
di: Fang, Fengyi, et al.
Pubblicazione: (2025)
Optimizing Image Capture for Computer Vision-Powered Taxonomic Identification and Trait Recognition of Biodiversity Specimens
di: East, Alyson, et al.
Pubblicazione: (2025)
di: East, Alyson, et al.
Pubblicazione: (2025)
SnapCap: Efficient Snapshot Compressive Video Captioning
di: Sun, Jianqiao, et al.
Pubblicazione: (2024)
di: Sun, Jianqiao, et al.
Pubblicazione: (2024)
Monsoon Uprising in Bangladesh: How Facebook Shaped Collective Identity
di: Abir, Md Tasin, et al.
Pubblicazione: (2025)
di: Abir, Md Tasin, et al.
Pubblicazione: (2025)
MultiModal Fine-tuning with Synthetic Captions
di: Enomoto, Shohei, et al.
Pubblicazione: (2026)
di: Enomoto, Shohei, et al.
Pubblicazione: (2026)
Improving Text Generation on Images with Synthetic Captions
di: Koh, Jun Young, et al.
Pubblicazione: (2024)
di: Koh, Jun Young, et al.
Pubblicazione: (2024)
BioBench: A Blueprint to Move Beyond ImageNet for Scientific ML Benchmarks
di: Stevens, Samuel
Pubblicazione: (2025)
di: Stevens, Samuel
Pubblicazione: (2025)
Improving Image Captioning by Mimicking Human Reformulation Feedback at Inference-time
di: Berger, Uri, et al.
Pubblicazione: (2025)
di: Berger, Uri, et al.
Pubblicazione: (2025)
The Influence of Initial Connectivity on Biologically Plausible Learning
di: Liu, Weixuan, et al.
Pubblicazione: (2024)
di: Liu, Weixuan, et al.
Pubblicazione: (2024)
SynC: Synthetic Image Caption Dataset Refinement with One-to-many Mapping for Zero-shot Image Captioning
di: Kim, Si-Woo, et al.
Pubblicazione: (2025)
di: Kim, Si-Woo, et al.
Pubblicazione: (2025)
Exploiting Label Skewness for Spiking Neural Networks in Federated Learning
di: Yu, Di, et al.
Pubblicazione: (2024)
di: Yu, Di, et al.
Pubblicazione: (2024)
Exploiting Multiple Sequence Lengths in Fast End to End Training for Image Captioning
di: Hu, Jia Cheng, et al.
Pubblicazione: (2022)
di: Hu, Jia Cheng, et al.
Pubblicazione: (2022)
CLIPS: An Enhanced CLIP Framework for Learning with Synthetic Captions
di: Liu, Yanqing, et al.
Pubblicazione: (2024)
di: Liu, Yanqing, et al.
Pubblicazione: (2024)
Hyperbolic Learning with Synthetic Captions for Open-World Detection
di: Kong, Fanjie, et al.
Pubblicazione: (2024)
di: Kong, Fanjie, et al.
Pubblicazione: (2024)
Synthetic Captions for Open-Vocabulary Zero-Shot Segmentation
di: Lebailly, Tim, et al.
Pubblicazione: (2025)
di: Lebailly, Tim, et al.
Pubblicazione: (2025)
Pixel-Level Change Detection Pseudo-Label Learning for Remote Sensing Change Captioning
di: Liu, Chenyang, et al.
Pubblicazione: (2023)
di: Liu, Chenyang, et al.
Pubblicazione: (2023)
CAP: Evaluation of Persuasive and Creative Image Generation
di: Aghazadeh, Aysan, et al.
Pubblicazione: (2024)
di: Aghazadeh, Aysan, et al.
Pubblicazione: (2024)
Documenti analoghi
-
BioCLIP 2: Emergent Properties from Scaling Hierarchical Contrastive Learning
di: Gu, Jianyang, et al.
Pubblicazione: (2025) -
Finer-CAM: Spotting the Difference Reveals Finer Details for Visual Explanation
di: Zhang, Ziheng, et al.
Pubblicazione: (2025) -
Prompt-CAM: Making Vision Transformers Interpretable for Fine-Grained Analysis
di: Chowdhury, Arpita, et al.
Pubblicazione: (2025) -
Static Segmentation by Tracking: A Label-Efficient Approach for Fine-Grained Specimen Image Segmentation
di: Feng, Zhenyang, et al.
Pubblicazione: (2025) -
BioCLIP: A Vision Foundation Model for the Tree of Life
di: Stevens, Samuel, et al.
Pubblicazione: (2023)