BioCAP: Exploiting Synthetic Captions Beyond Labels in Biological Foundation Models
Fuente:
arXiv
Guardado en:
| Autores principales: | Zhang, Ziheng, Ma, Xinyue, Chowdhury, Arpita, Campolongo, Elizabeth G., Thompson, Matthew J., Zhang, Net, Stevens, Samuel, Lapp, Hilmar, Berger-Wolf, Tanya, Su, Yu, Chao, Wei-Lun, Gu, Jianyang |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
BioCLIP 2: Emergent Properties from Scaling Hierarchical Contrastive Learning
por: Gu, Jianyang, et al.
Publicado: (2025)
por: Gu, Jianyang, et al.
Publicado: (2025)
Finer-CAM: Spotting the Difference Reveals Finer Details for Visual Explanation
por: Zhang, Ziheng, et al.
Publicado: (2025)
por: Zhang, Ziheng, et al.
Publicado: (2025)
Prompt-CAM: Making Vision Transformers Interpretable for Fine-Grained Analysis
por: Chowdhury, Arpita, et al.
Publicado: (2025)
por: Chowdhury, Arpita, et al.
Publicado: (2025)
Static Segmentation by Tracking: A Label-Efficient Approach for Fine-Grained Specimen Image Segmentation
por: Feng, Zhenyang, et al.
Publicado: (2025)
por: Feng, Zhenyang, et al.
Publicado: (2025)
BioCLIP: A Vision Foundation Model for the Tree of Life
por: Stevens, Samuel, et al.
Publicado: (2023)
por: Stevens, Samuel, et al.
Publicado: (2023)
VLM4Bio: A Benchmark Dataset to Evaluate Pretrained Vision-Language Models for Trait Discovery from Biological Images
por: Maruf, M., et al.
Publicado: (2024)
por: Maruf, M., et al.
Publicado: (2024)
Interpretable and Testable Vision Features via Sparse Autoencoders
por: Stevens, Samuel, et al.
Publicado: (2025)
por: Stevens, Samuel, et al.
Publicado: (2025)
What Do You See in Common? Learning Hierarchical Prototypes over Tree-of-Life to Discover Evolutionary Traits
por: Manogaran, Harish Babu, et al.
Publicado: (2024)
por: Manogaran, Harish Babu, et al.
Publicado: (2024)
Fish-Vista: A Multi-Purpose Dataset for Understanding & Identification of Traits from Images
por: Mehrab, Kazi Sajeed, et al.
Publicado: (2024)
por: Mehrab, Kazi Sajeed, et al.
Publicado: (2024)
Towards Open-Ended Visual Scientific Discovery with Sparse Autoencoders
por: Stevens, Samuel, et al.
Publicado: (2025)
por: Stevens, Samuel, et al.
Publicado: (2025)
Image DataPalooza 2023 (Archive)
por: Lapp, Hilmar, et al.
Publicado: (2024)
por: Lapp, Hilmar, et al.
Publicado: (2024)
ReflectCAP: Detailed Image Captioning with Reflective Memory
por: Min, Kyungmin, et al.
Publicado: (2026)
por: Min, Kyungmin, et al.
Publicado: (2026)
Exploiting Pseudo Image Captions for Multimodal Summarization
por: Jiang, Chaoya, et al.
Publicado: (2023)
por: Jiang, Chaoya, et al.
Publicado: (2023)
Lessons and Open Questions from a Unified Study of Camera-Trap Species Recognition Over Time
por: Jeon, Sooyoung, et al.
Publicado: (2026)
por: Jeon, Sooyoung, et al.
Publicado: (2026)
Leveraging Latent Visual Reasoning in Silence
por: Zhu, Dongyao, et al.
Publicado: (2026)
por: Zhu, Dongyao, et al.
Publicado: (2026)
Fine-Tuning is Fine, if Calibrated
por: Mai, Zheda, et al.
Publicado: (2024)
por: Mai, Zheda, et al.
Publicado: (2024)
A continental-scale dataset of ground beetles with high-resolution images and validated morphological trait measurements
por: Rayeed, S M, et al.
Publicado: (2026)
por: Rayeed, S M, et al.
Publicado: (2026)
A Simple Interpretable Transformer for Fine-Grained Image Classification and Analysis
por: Paul, Dipanjyoti, et al.
Publicado: (2023)
por: Paul, Dipanjyoti, et al.
Publicado: (2023)
Exploiting Auxiliary Caption for Video Grounding
por: Li, Hongxiang, et al.
Publicado: (2023)
por: Li, Hongxiang, et al.
Publicado: (2023)
EUFCC-CIR: a Composed Image Retrieval Dataset for GLAM Collections
por: Net, Francesc, et al.
Publicado: (2024)
por: Net, Francesc, et al.
Publicado: (2024)
Hierarchical Conditioning of Diffusion Models Using Tree-of-Life for Studying Species Evolution
por: Khurana, Mridul, et al.
Publicado: (2024)
por: Khurana, Mridul, et al.
Publicado: (2024)
kabr-tools: Automated Framework for Multi-Species Behavioral Monitoring
por: Kline, Jenna, et al.
Publicado: (2025)
por: Kline, Jenna, et al.
Publicado: (2025)
OmniCaptioner: One Captioner to Rule Them All
por: Lu, Yiting, et al.
Publicado: (2025)
por: Lu, Yiting, et al.
Publicado: (2025)
CoordSpeaker: Exploiting Gesture Captioning for Coordinated Caption-Empowered Co-Speech Gesture Generation
por: Fang, Fengyi, et al.
Publicado: (2025)
por: Fang, Fengyi, et al.
Publicado: (2025)
Optimizing Image Capture for Computer Vision-Powered Taxonomic Identification and Trait Recognition of Biodiversity Specimens
por: East, Alyson, et al.
Publicado: (2025)
por: East, Alyson, et al.
Publicado: (2025)
SnapCap: Efficient Snapshot Compressive Video Captioning
por: Sun, Jianqiao, et al.
Publicado: (2024)
por: Sun, Jianqiao, et al.
Publicado: (2024)
Monsoon Uprising in Bangladesh: How Facebook Shaped Collective Identity
por: Abir, Md Tasin, et al.
Publicado: (2025)
por: Abir, Md Tasin, et al.
Publicado: (2025)
MultiModal Fine-tuning with Synthetic Captions
por: Enomoto, Shohei, et al.
Publicado: (2026)
por: Enomoto, Shohei, et al.
Publicado: (2026)
Improving Text Generation on Images with Synthetic Captions
por: Koh, Jun Young, et al.
Publicado: (2024)
por: Koh, Jun Young, et al.
Publicado: (2024)
BioBench: A Blueprint to Move Beyond ImageNet for Scientific ML Benchmarks
por: Stevens, Samuel
Publicado: (2025)
por: Stevens, Samuel
Publicado: (2025)
Improving Image Captioning by Mimicking Human Reformulation Feedback at Inference-time
por: Berger, Uri, et al.
Publicado: (2025)
por: Berger, Uri, et al.
Publicado: (2025)
The Influence of Initial Connectivity on Biologically Plausible Learning
por: Liu, Weixuan, et al.
Publicado: (2024)
por: Liu, Weixuan, et al.
Publicado: (2024)
SynC: Synthetic Image Caption Dataset Refinement with One-to-many Mapping for Zero-shot Image Captioning
por: Kim, Si-Woo, et al.
Publicado: (2025)
por: Kim, Si-Woo, et al.
Publicado: (2025)
Exploiting Label Skewness for Spiking Neural Networks in Federated Learning
por: Yu, Di, et al.
Publicado: (2024)
por: Yu, Di, et al.
Publicado: (2024)
Exploiting Multiple Sequence Lengths in Fast End to End Training for Image Captioning
por: Hu, Jia Cheng, et al.
Publicado: (2022)
por: Hu, Jia Cheng, et al.
Publicado: (2022)
CLIPS: An Enhanced CLIP Framework for Learning with Synthetic Captions
por: Liu, Yanqing, et al.
Publicado: (2024)
por: Liu, Yanqing, et al.
Publicado: (2024)
Hyperbolic Learning with Synthetic Captions for Open-World Detection
por: Kong, Fanjie, et al.
Publicado: (2024)
por: Kong, Fanjie, et al.
Publicado: (2024)
Synthetic Captions for Open-Vocabulary Zero-Shot Segmentation
por: Lebailly, Tim, et al.
Publicado: (2025)
por: Lebailly, Tim, et al.
Publicado: (2025)
Pixel-Level Change Detection Pseudo-Label Learning for Remote Sensing Change Captioning
por: Liu, Chenyang, et al.
Publicado: (2023)
por: Liu, Chenyang, et al.
Publicado: (2023)
CAP: Evaluation of Persuasive and Creative Image Generation
por: Aghazadeh, Aysan, et al.
Publicado: (2024)
por: Aghazadeh, Aysan, et al.
Publicado: (2024)
Ejemplares similares
-
BioCLIP 2: Emergent Properties from Scaling Hierarchical Contrastive Learning
por: Gu, Jianyang, et al.
Publicado: (2025) -
Finer-CAM: Spotting the Difference Reveals Finer Details for Visual Explanation
por: Zhang, Ziheng, et al.
Publicado: (2025) -
Prompt-CAM: Making Vision Transformers Interpretable for Fine-Grained Analysis
por: Chowdhury, Arpita, et al.
Publicado: (2025) -
Static Segmentation by Tracking: A Label-Efficient Approach for Fine-Grained Specimen Image Segmentation
por: Feng, Zhenyang, et al.
Publicado: (2025) -
BioCLIP: A Vision Foundation Model for the Tree of Life
por: Stevens, Samuel, et al.
Publicado: (2023)