CLIBD: Bridging Vision and Genomics for Biodiversity Monitoring at Scale
Fuente:
arXiv
Saved in:
| Main Authors: | Gong, ZeMing, Wang, Austin T., Huo, Xiaoliang, Haurum, Joakim Bruslund, Lowe, Scott C., Taylor, Graham W., Chang, Angel X. |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Hyperbolic Multimodal Representation Learning for Biological Taxonomies
by: Gong, ZeMing, et al.
Published: (2025)
by: Gong, ZeMing, et al.
Published: (2025)
BIOSCAN-5M: A Multimodal Dataset for Insect Biodiversity
by: Gharaee, Zahra, et al.
Published: (2024)
by: Gharaee, Zahra, et al.
Published: (2024)
ViGiL3D: A Linguistically Diverse Dataset for 3D Visual Grounding
by: Wang, Austin T., et al.
Published: (2025)
by: Wang, Austin T., et al.
Published: (2025)
BarcodeBERT: Transformers for Biodiversity Analysis
by: Arias, Pablo Millan, et al.
Published: (2023)
by: Arias, Pablo Millan, et al.
Published: (2023)
An Empirical Study into Clustering of Unseen Datasets with Self-Supervised Encoders
by: Lowe, Scott C., et al.
Published: (2024)
by: Lowe, Scott C., et al.
Published: (2024)
Agglomerative Token Clustering
by: Haurum, Joakim Bruslund, et al.
Published: (2024)
by: Haurum, Joakim Bruslund, et al.
Published: (2024)
How to Sample High Quality 3D Fractals for Action Recognition Pre-Training?
by: Putak, Marko, et al.
Published: (2026)
by: Putak, Marko, et al.
Published: (2026)
Comparing Euclidean and Hyperbolic K-Means for Generalized Category Discovery
by: Dalal, Mohamad, et al.
Published: (2026)
by: Dalal, Mohamad, et al.
Published: (2026)
From Pixels to Privacy: Temporally Consistent Video Anonymization via Token Pruning for Privacy Preserving Action Recognition
by: Aslam, Nazia, et al.
Published: (2026)
by: Aslam, Nazia, et al.
Published: (2026)
HSM: Hierarchical Scene Motifs for Multi-Scale Indoor Scene Generation
by: Pun, Hou In Derek, et al.
Published: (2025)
by: Pun, Hou In Derek, et al.
Published: (2025)
BarcodeMamba+: Advancing State-Space Models for Fungal Biodiversity Research
by: Gao, Tiancheng, et al.
Published: (2025)
by: Gao, Tiancheng, et al.
Published: (2025)
Generating Synthetic Stereo Datasets using 3D Gaussian Splatting and Expert Knowledge Transfer
by: Slezak, Filip, et al.
Published: (2025)
by: Slezak, Filip, et al.
Published: (2025)
Self-Distillation of Hidden Layers for Self-Supervised Representation Learning
by: Lowe, Scott C., et al.
Published: (2026)
by: Lowe, Scott C., et al.
Published: (2026)
Vision Bridge Transformer at Scale
by: Tan, Zhenxiong, et al.
Published: (2025)
by: Tan, Zhenxiong, et al.
Published: (2025)
Automated Detection of Antarctic Benthic Organisms in High-Resolution In Situ Imagery to Aid Biodiversity Monitoring
by: Trotter, Cameron, et al.
Published: (2025)
by: Trotter, Cameron, et al.
Published: (2025)
Optimizing Image Capture for Computer Vision-Powered Taxonomic Identification and Trait Recognition of Biodiversity Specimens
by: East, Alyson, et al.
Published: (2025)
by: East, Alyson, et al.
Published: (2025)
Autonomous AI Bird Feeder for Backyard Biodiversity Monitoring
by: Mansouri, El Mustapha
Published: (2025)
by: Mansouri, El Mustapha
Published: (2025)
INQUIRE-Search: Interactive Discovery in Large-Scale Biodiversity Databases
by: Vendrow, Edward, et al.
Published: (2025)
by: Vendrow, Edward, et al.
Published: (2025)
Scaling Crowdsourced Election Monitoring: Construction and Evaluation of Classification Models for Multilingual and Cross-Domain Classification Settings
by: Magomere, Jabez, et al.
Published: (2025)
by: Magomere, Jabez, et al.
Published: (2025)
Shotluck Holmes: A Family of Efficient Small-Scale Large Language Vision Models For Video Captioning and Summarization
by: Luo, Richard, et al.
Published: (2024)
by: Luo, Richard, et al.
Published: (2024)
Introduction to Scientific Programming with Python
by: Sundnes, Joakim
Published: (2020)
by: Sundnes, Joakim
Published: (2020)
Point Cloud Segmentation of Agricultural Vehicles using 3D Gaussian Splatting
by: Christiansen, Alfred T., et al.
Published: (2025)
by: Christiansen, Alfred T., et al.
Published: (2025)
Computer Vision Metrics
by: Krig, Scott
Published: (2018)
by: Krig, Scott
Published: (2018)
Artificial Intelligence for Sustainable Urban Biodiversity: A Framework for Monitoring and Conservation
by: Rahmati, Yasmin
Published: (2024)
by: Rahmati, Yasmin
Published: (2024)
Isolated Channel Vision Transformers: From Single-Channel Pretraining to Multi-Channel Finetuning
by: Lian, Wenyi, et al.
Published: (2025)
by: Lian, Wenyi, et al.
Published: (2025)
Pose Matters: Evaluating Vision Transformers and CNNs for Human Action Recognition on Small COCO Subsets
by: Tang, MingZe, et al.
Published: (2025)
by: Tang, MingZe, et al.
Published: (2025)
SceneEval: Evaluating Semantic Coherence in Text-Conditioned 3D Indoor Scene Synthesis
by: Tam, Hou In Ivan, et al.
Published: (2025)
by: Tam, Hou In Ivan, et al.
Published: (2025)
Semantic Mapping in Indoor Embodied AI -- A Survey on Advances, Challenges, and Future Directions
by: Raychaudhuri, Sonia, et al.
Published: (2025)
by: Raychaudhuri, Sonia, et al.
Published: (2025)
eLasmobranc Dataset: An Image Dataset for Elasmobranch Species Recognition and Biodiversity Monitoring
by: Beviá-Ballesteros, Ismael, et al.
Published: (2026)
by: Beviá-Ballesteros, Ismael, et al.
Published: (2026)
Nested Fusion: A Method for Learning High Resolution Latent Structure of Multi-Scale Measurement Data on Mars
by: Wright, Austin P., et al.
Published: (2024)
by: Wright, Austin P., et al.
Published: (2024)
Griffon-G: Bridging Vision-Language and Vision-Centric Tasks via Large Multimodal Models
by: Zhan, Yufei, et al.
Published: (2024)
by: Zhan, Yufei, et al.
Published: (2024)
LiteFrame: Efficient Vision Encoders Unlock Frame Scaling in Video LLMs
by: Kim, Jihwan, et al.
Published: (2026)
by: Kim, Jihwan, et al.
Published: (2026)
Aged to Perfection: Machine-Learning Maps of Age in Conversational English
by: Tang, MingZe
Published: (2025)
by: Tang, MingZe
Published: (2025)
Midtraining Bridges Pretraining and Posttraining Distributions
by: Liu, Emmy, et al.
Published: (2025)
by: Liu, Emmy, et al.
Published: (2025)
BridgeTower: Building Bridges Between Encoders in Vision-Language Representation Learning
by: Xu, Xiao, et al.
Published: (2022)
by: Xu, Xiao, et al.
Published: (2022)
Measuring Chain-of-Thought Monitorability Through Faithfulness and Verbosity
by: Meek, Austin, et al.
Published: (2025)
by: Meek, Austin, et al.
Published: (2025)
Bridging the 2D-3D Gap: A Hierarchical Semantic-Geometric Map for Vision Language Navigation
by: Li, Kailing, et al.
Published: (2026)
by: Li, Kailing, et al.
Published: (2026)
Open-Insect: Benchmarking Open-Set Recognition of Novel Species in Biodiversity Monitoring
by: Chen, Yuyan, et al.
Published: (2025)
by: Chen, Yuyan, et al.
Published: (2025)
Scale Can't Overcome Pragmatics: The Impact of Reporting Bias on Vision-Language Reasoning
by: Kamath, Amita, et al.
Published: (2026)
by: Kamath, Amita, et al.
Published: (2026)
Quantification of Biodiversity from Historical Survey Text with LLM-based Best-Worst Scaling
by: Haider, Thomas, et al.
Published: (2025)
by: Haider, Thomas, et al.
Published: (2025)
Similar Items
-
Hyperbolic Multimodal Representation Learning for Biological Taxonomies
by: Gong, ZeMing, et al.
Published: (2025) -
BIOSCAN-5M: A Multimodal Dataset for Insect Biodiversity
by: Gharaee, Zahra, et al.
Published: (2024) -
ViGiL3D: A Linguistically Diverse Dataset for 3D Visual Grounding
by: Wang, Austin T., et al.
Published: (2025) -
BarcodeBERT: Transformers for Biodiversity Analysis
by: Arias, Pablo Millan, et al.
Published: (2023) -
An Empirical Study into Clustering of Unseen Datasets with Self-Supervised Encoders
by: Lowe, Scott C., et al.
Published: (2024)