Merlin L48 Spectrogram Dataset
Fuente:
arXiv
Saved in:
| Main Authors: | Sun, Aaron, Maji, Subhransu, Van Horn, Grant |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Improved Zero-Shot Classification by Adapting VLMs with Text Descriptions
by: Saha, Oindrila, et al.
Published: (2024)
by: Saha, Oindrila, et al.
Published: (2024)
Generate, Transduct, Adapt: Iterative Transduction with VLMs
by: Saha, Oindrila, et al.
Published: (2025)
by: Saha, Oindrila, et al.
Published: (2025)
Human-in-the-Loop Visual Re-ID for Population Size Estimation
by: Perez, Gustavo, et al.
Published: (2023)
by: Perez, Gustavo, et al.
Published: (2023)
Masked Autoencoders with Limited Data: Does It Work? A Fine-Grained Bioacoustics Case Study
by: Liu, Wuao, et al.
Published: (2026)
by: Liu, Wuao, et al.
Published: (2026)
Not All Birds Look The Same: Identity-Preserving Generation For Birds
by: Sun, Aaron, et al.
Published: (2025)
by: Sun, Aaron, et al.
Published: (2025)
Task2Box: Box Embeddings for Modeling Asymmetric Task Relationships
by: Daroya, Rangel, et al.
Published: (2024)
by: Daroya, Rangel, et al.
Published: (2024)
You May Speak Freely: Improving the Fine-Grained Visual Recognition Capabilities of Multimodal Large Language Models with Answer Extraction
by: Lawrence, Logan, et al.
Published: (2025)
by: Lawrence, Logan, et al.
Published: (2025)
Consensus-Driven Active Model Selection
by: Kay, Justin, et al.
Published: (2025)
by: Kay, Justin, et al.
Published: (2025)
Moment Sampling in Video LLMs for Long-Form Video QA
by: Chasmai, Mustafa, et al.
Published: (2025)
by: Chasmai, Mustafa, et al.
Published: (2025)
WildSAT: Learning Satellite Image Representations from Wildlife Observations
by: Daroya, Rangel, et al.
Published: (2024)
by: Daroya, Rangel, et al.
Published: (2024)
RealBirdID: Benchmarking Bird Species Identification in the Era of MLLMs
by: Lawrence, Logan, et al.
Published: (2026)
by: Lawrence, Logan, et al.
Published: (2026)
Active Measurement of Two-Point Correlations
by: Hamilton, Max, et al.
Published: (2026)
by: Hamilton, Max, et al.
Published: (2026)
Feedforward Few-shot Species Range Estimation
by: Lange, Christian, et al.
Published: (2025)
by: Lange, Christian, et al.
Published: (2025)
Active Measurement: Efficient Estimation at Scale
by: Hamilton, Max, et al.
Published: (2025)
by: Hamilton, Max, et al.
Published: (2025)
Merlin: A Computed Tomography Vision-Language Foundation Model and Dataset
by: Blankemeier, Louis, et al.
Published: (2024)
by: Blankemeier, Louis, et al.
Published: (2024)
SIGMA-GEN: Structure and Identity Guided Multi-subject Assembly for Image Generation
by: Saha, Oindrila, et al.
Published: (2025)
by: Saha, Oindrila, et al.
Published: (2025)
3D Space as a Scratchpad for Editable Text-to-Image Generation
by: Saha, Oindrila, et al.
Published: (2026)
by: Saha, Oindrila, et al.
Published: (2026)
Merlin:Empowering Multimodal LLMs with Foresight Minds
by: Yu, En, et al.
Published: (2023)
by: Yu, En, et al.
Published: (2023)
CATRF: Codec-Adaptive TriPlane Radiance Fields for Volumetric Content Delivery
by: Chen, Tung-I, et al.
Published: (2026)
by: Chen, Tung-I, et al.
Published: (2026)
Improving Satellite Imagery Masking using Multi-task and Transfer Learning
by: Daroya, Rangel, et al.
Published: (2024)
by: Daroya, Rangel, et al.
Published: (2024)
Counting Fish with Temporal Representations of Sonar Video
by: Van Brunt, Kai, et al.
Published: (2025)
by: Van Brunt, Kai, et al.
Published: (2025)
RiverScope: High-Resolution River Masking Dataset
by: Daroya, Rangel, et al.
Published: (2025)
by: Daroya, Rangel, et al.
Published: (2025)
CleverBirds: A Multiple-Choice Benchmark for Fine-grained Human Knowledge Tracing
by: Bossemeyer, Leonie, et al.
Published: (2025)
by: Bossemeyer, Leonie, et al.
Published: (2025)
GLaRE: A Graph-based Landmark Region Embedding Network for Emotion Recognition
by: Maji, Debasis, et al.
Published: (2025)
by: Maji, Debasis, et al.
Published: (2025)
Perceptually Optimized Color Selection for Visualization
by: Maji, Subhrajyoti, et al.
Published: (2022)
by: Maji, Subhrajyoti, et al.
Published: (2022)
Disentangling Modes and Interference in the Spectrogram of Multicomponent Signals
by: Polisano, Kévin, et al.
Published: (2025)
by: Polisano, Kévin, et al.
Published: (2025)
Believing is Seeing: Unobserved Object Detection using Generative Models
by: Bhattacharjee, Subhransu S., et al.
Published: (2024)
by: Bhattacharjee, Subhransu S., et al.
Published: (2024)
Semantic-Anchored, Class Variance-Optimized Clustering for Robust Semi-Supervised Few-Shot Learning
by: Maji, Souvik, et al.
Published: (2025)
by: Maji, Souvik, et al.
Published: (2025)
Into the Unknown: Towards using Generative Models for Sampling Priors of Environment Uncertainty for Planning in Configuration Spaces
by: Bhattacharjee, Subhransu S., et al.
Published: (2025)
by: Bhattacharjee, Subhransu S., et al.
Published: (2025)
An Intuitionistic Fuzzy Logic Driven UNet architecture: Application to Brain Image segmentation
by: Verma, Hanuman, et al.
Published: (2026)
by: Verma, Hanuman, et al.
Published: (2026)
Multi-View Spectrogram Transformer for Respiratory Sound Classification
by: He, Wentao, et al.
Published: (2023)
by: He, Wentao, et al.
Published: (2023)
Spectrogram-Based Detection of Auto-Tuned Vocals in Music Recordings
by: Gohari, Mahyar, et al.
Published: (2024)
by: Gohari, Mahyar, et al.
Published: (2024)
Improving Robustness of Spectrogram Classifiers with Neural Stochastic Differential Equations
by: Brogan, Joel, et al.
Published: (2024)
by: Brogan, Joel, et al.
Published: (2024)
FlatLands: Generative Floormap Completion From a Single Egocentric View
by: Bhattacharjee, Subhransu S., et al.
Published: (2026)
by: Bhattacharjee, Subhransu S., et al.
Published: (2026)
Knowledge-Augmented Vision Language Models for Underwater Bioacoustic Spectrogram Analysis
by: Nihal, Ragib Amin, et al.
Published: (2025)
by: Nihal, Ragib Amin, et al.
Published: (2025)
Detección y Cuantificación de Erosión Fluvial con Visión Artificial
by: Maji, Paúl, et al.
Published: (2025)
by: Maji, Paúl, et al.
Published: (2025)
Analyzing and Mitigating Bias for Vulnerable Classes: Towards Balanced Representation in Dataset
by: Katare, Dewant, et al.
Published: (2024)
by: Katare, Dewant, et al.
Published: (2024)
Eyes on the Grass: Biodiversity-Increasing Robotic Mowing Using Deep Visual Embeddings
by: Beckers, Lars, et al.
Published: (2025)
by: Beckers, Lars, et al.
Published: (2025)
The iNaturalist Sounds Dataset
by: Chasmai, Mustafa, et al.
Published: (2025)
by: Chasmai, Mustafa, et al.
Published: (2025)
SID: Stereo Image Dataset for Autonomous Driving in Adverse Conditions
by: El-Shair, Zaid A., et al.
Published: (2024)
by: El-Shair, Zaid A., et al.
Published: (2024)
Similar Items
-
Improved Zero-Shot Classification by Adapting VLMs with Text Descriptions
by: Saha, Oindrila, et al.
Published: (2024) -
Generate, Transduct, Adapt: Iterative Transduction with VLMs
by: Saha, Oindrila, et al.
Published: (2025) -
Human-in-the-Loop Visual Re-ID for Population Size Estimation
by: Perez, Gustavo, et al.
Published: (2023) -
Masked Autoencoders with Limited Data: Does It Work? A Fine-Grained Bioacoustics Case Study
by: Liu, Wuao, et al.
Published: (2026) -
Not All Birds Look The Same: Identity-Preserving Generation For Birds
by: Sun, Aaron, et al.
Published: (2025)