Model Merging Improves Zero-Shot Generalization in Bioacoustic Foundation Models
Fuente:
arXiv
Saved in:
| Main Authors: | Marincione, Davide, Crisostomi, Donato, Dessi, Roberto, Rodolà, Emanuele, Rossi, Emanuele |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LoopGen: Training-Free Loopable Music Generation
by: Marincione, Davide, et al.
Published: (2025)
by: Marincione, Davide, et al.
Published: (2025)
Model Merging: Foundations and Algorithms
by: Crisostomi, Donato
Published: (2026)
by: Crisostomi, Donato
Published: (2026)
ATM: Improving Model Merging by Alternating Tuning and Merging
by: Zhou, Luca, et al.
Published: (2024)
by: Zhou, Luca, et al.
Published: (2024)
Mergenetic: a Simple Evolutionary Model Merging Library
by: Minut, Adrian Robert, et al.
Published: (2025)
by: Minut, Adrian Robert, et al.
Published: (2025)
MERGE$^3$: Efficient Evolutionary Merging on Consumer-grade GPUs
by: Mencattini, Tommaso, et al.
Published: (2025)
by: Mencattini, Tommaso, et al.
Published: (2025)
Metric Based Few-Shot Graph Classification
by: Crisostomi, Donato, et al.
Published: (2022)
by: Crisostomi, Donato, et al.
Published: (2022)
Language Models are Injective and Hence Invertible
by: Nikolaou, Giorgos, et al.
Published: (2025)
by: Nikolaou, Giorgos, et al.
Published: (2025)
NatureLM-audio: an Audio-Language Foundation Model for Bioacoustics
by: Robinson, David, et al.
Published: (2024)
by: Robinson, David, et al.
Published: (2024)
Communicating Sound Through Natural Language
by: Rossi, Emanuele, et al.
Published: (2026)
by: Rossi, Emanuele, et al.
Published: (2026)
$C^2M^3$: Cycle-Consistent Multi-Model Merging
by: Crisostomi, Donato, et al.
Published: (2024)
by: Crisostomi, Donato, et al.
Published: (2024)
Two-Scale Latent Dynamics for Recurrent-Depth Transformers
by: Pappone, Francesco, et al.
Published: (2025)
by: Pappone, Francesco, et al.
Published: (2025)
Advancing Marine Bioacoustics with Deep Generative Models: A Hybrid Augmentation Strategy for Southern Resident Killer Whale Detection
by: Padovese, Bruno, et al.
Published: (2025)
by: Padovese, Bruno, et al.
Published: (2025)
On Task Vectors and Gradients
by: Zhou, Luca, et al.
Published: (2025)
by: Zhou, Luca, et al.
Published: (2025)
Multi-Way Representation Alignment
by: Achara, Akshit, et al.
Published: (2026)
by: Achara, Akshit, et al.
Published: (2026)
When Denoising Hinders: Revisiting Zero-Shot ASR with SAM-Audio and Whisper
by: Islam, Akif, et al.
Published: (2026)
by: Islam, Akif, et al.
Published: (2026)
Generalized Multi-Source Inference for Text Conditioned Music Diffusion Models
by: Postolache, Emilian, et al.
Published: (2024)
by: Postolache, Emilian, et al.
Published: (2024)
Implicit Inversion turns CLIP into a Decoder
by: D'Orazio, Antonio, et al.
Published: (2025)
by: D'Orazio, Antonio, et al.
Published: (2025)
Woosh: A Sound Effects Foundation Model
by: Hadjeres, Gaëtan, et al.
Published: (2026)
by: Hadjeres, Gaëtan, et al.
Published: (2026)
EuleroDec: A Complex-Valued RVQ-VAE for Efficient and Robust Audio Coding
by: Cerovaz, Luca, et al.
Published: (2026)
by: Cerovaz, Luca, et al.
Published: (2026)
STAGE: Stemmed Accompaniment Generation through Prefix-Based Conditioning
by: Strano, Giorgio, et al.
Published: (2025)
by: Strano, Giorgio, et al.
Published: (2025)
Hybrid Disagreement-Diversity Active Learning for Bioacoustic Sound Event Detection
by: Zhang, Shiqi, et al.
Published: (2025)
by: Zhang, Shiqi, et al.
Published: (2025)
Multi-objective Evolutionary Merging Enables Efficient Reasoning Models
by: Iacobelli, Mario, et al.
Published: (2026)
by: Iacobelli, Mario, et al.
Published: (2026)
Activation Patching for Interpretable Steering in Music Generation
by: Facchiano, Simone, et al.
Published: (2025)
by: Facchiano, Simone, et al.
Published: (2025)
Multi-Source Diffusion Models for Simultaneous Music Generation and Separation
by: Mariani, Giorgio, et al.
Published: (2023)
by: Mariani, Giorgio, et al.
Published: (2023)
Lightweight Hopfield Neural Networks for Bioacoustic Detection and Call Monitoring of Captive Primates
by: Lomas, Wendy, et al.
Published: (2025)
by: Lomas, Wendy, et al.
Published: (2025)
Weakly Supervised Detection and Temporal Localization of Whale Calls in Long-Duration Bioacoustic Data
by: Nihal, Ragib Amin, et al.
Published: (2025)
by: Nihal, Ragib Amin, et al.
Published: (2025)
MaskGCT: Zero-Shot Text-to-Speech with Masked Generative Codec Transformer
by: Wang, Yuancheng, et al.
Published: (2024)
by: Wang, Yuancheng, et al.
Published: (2024)
Zero-Shot Duet Singing Voices Separation with Diffusion Models
by: Yu, Chin-Yun, et al.
Published: (2023)
by: Yu, Chin-Yun, et al.
Published: (2023)
Survey on the Evaluation of Generative Models in Music
by: Lerch, Alexander, et al.
Published: (2025)
by: Lerch, Alexander, et al.
Published: (2025)
Domain Elastic Transform: Bayesian Function Registration for High-Dimensional Scientific Data
by: Hirose, Osamu, et al.
Published: (2026)
by: Hirose, Osamu, et al.
Published: (2026)
FISHER: A Foundation Model for Multi-Modal Industrial Signal Comprehensive Representation
by: Fan, Pingyi, et al.
Published: (2025)
by: Fan, Pingyi, et al.
Published: (2025)
Mapping representations in Reinforcement Learning via Semantic Alignment for Zero-Shot Stitching
by: Ricciardi, Antonio Pio, et al.
Published: (2025)
by: Ricciardi, Antonio Pio, et al.
Published: (2025)
TSPE: Task-Specific Prompt Ensemble for Improved Zero-Shot Audio Classification
by: Anand, Nishit, et al.
Published: (2024)
by: Anand, Nishit, et al.
Published: (2024)
State Space Models for Bioacoustics: A Comparative Evaluation with Transformers
by: Tang, Chengyu, et al.
Published: (2025)
by: Tang, Chengyu, et al.
Published: (2025)
Zero-Shot Quantization via Weight-Space Arithmetic
by: Solombrino, Daniele, et al.
Published: (2026)
by: Solombrino, Daniele, et al.
Published: (2026)
MoonCast: High-Quality Zero-Shot Podcast Generation
by: Ju, Zeqian, et al.
Published: (2025)
by: Ju, Zeqian, et al.
Published: (2025)
Task Singular Vectors: Reducing Task Interference in Model Merging
by: Gargiulo, Antonio Andrea, et al.
Published: (2024)
by: Gargiulo, Antonio Andrea, et al.
Published: (2024)
Foundation Models for Bioacoustics -- a Comparative Review
by: Schwinger, Raphael, et al.
Published: (2025)
by: Schwinger, Raphael, et al.
Published: (2025)
Explicit Context-Driven Neural Acoustic Modeling for High-Fidelity RIR Generation
by: Si, Chen, et al.
Published: (2025)
by: Si, Chen, et al.
Published: (2025)
NaturalSpeech 3: Zero-Shot Speech Synthesis with Factorized Codec and Diffusion Models
by: Ju, Zeqian, et al.
Published: (2024)
by: Ju, Zeqian, et al.
Published: (2024)
Similar Items
-
LoopGen: Training-Free Loopable Music Generation
by: Marincione, Davide, et al.
Published: (2025) -
Model Merging: Foundations and Algorithms
by: Crisostomi, Donato
Published: (2026) -
ATM: Improving Model Merging by Alternating Tuning and Merging
by: Zhou, Luca, et al.
Published: (2024) -
Mergenetic: a Simple Evolutionary Model Merging Library
by: Minut, Adrian Robert, et al.
Published: (2025) -
MERGE$^3$: Efficient Evolutionary Merging on Consumer-grade GPUs
by: Mencattini, Tommaso, et al.
Published: (2025)