Zwitscherkasten -- DIY Audiovisual bird monitoring
Fuente:
arXiv
Saved in:
| Main Authors: | Blum, Dominik, Häring, Elias, Jirges, Fabian, Schäffer, Martin, Schick, David, Schulenberg, Florian, Schön, Torsten |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Embedding Arithmetic: A Lightweight, Tuning-Free Framework for Post-hoc Bias Mitigation in Text-to-Image Models
by: Sambandham, Venkatesh Thirugnana, et al.
Published: (2026)
by: Sambandham, Venkatesh Thirugnana, et al.
Published: (2026)
Unlocking Past Information: Temporal Embeddings in Cooperative Bird's Eye View Prediction
by: Rößle, Dominik, et al.
Published: (2024)
by: Rößle, Dominik, et al.
Published: (2024)
DrivIng: A Large-Scale Multimodal Driving Dataset with Full Digital Twin Integration
by: Rößle, Dominik, et al.
Published: (2026)
by: Rößle, Dominik, et al.
Published: (2026)
UrbanIng-V2X: A Large-Scale Multi-Vehicle, Multi-Infrastructure Dataset Across Multiple Intersections for Cooperative Perception
by: Sekaran, Karthikeyan Chandra, et al.
Published: (2025)
by: Sekaran, Karthikeyan Chandra, et al.
Published: (2025)
A strongly annotated passive acoustic dataset for tropical bird monitoring
by: Ruiz, Daniela, et al.
Published: (2026)
by: Ruiz, Daniela, et al.
Published: (2026)
Do It Yourself (DIY): Modifying Images for Poems in a Zero-Shot Setting Using Weighted Prompt Manipulation
by: Jamil, Sofia, et al.
Published: (2025)
by: Jamil, Sofia, et al.
Published: (2025)
Operationalizing Fairness in Text-to-Image Models: A Survey of Bias, Fairness Audits and Mitigation Strategies
by: Smith, Megan, et al.
Published: (2026)
by: Smith, Megan, et al.
Published: (2026)
Audiovisual Masked Autoencoders
by: Georgescu, Mariana-Iuliana, et al.
Published: (2022)
by: Georgescu, Mariana-Iuliana, et al.
Published: (2022)
Interpretable Perception and Reasoning for Audiovisual Geolocation
by: Su, Yiyang, et al.
Published: (2026)
by: Su, Yiyang, et al.
Published: (2026)
MGNet: Monocular Geometric Scene Understanding for Autonomous Driving
by: Schön, Markus, et al.
Published: (2022)
by: Schön, Markus, et al.
Published: (2022)
MGNiceNet: Unified Monocular Geometric Scene Understanding
by: Schön, Markus, et al.
Published: (2024)
by: Schön, Markus, et al.
Published: (2024)
WSESeg: Introducing a Dataset for the Segmentation of Winter Sports Equipment with a Baseline for Interactive Segmentation
by: Schön, Robin, et al.
Published: (2024)
by: Schön, Robin, et al.
Published: (2024)
Self-supervised Audiovisual Representation Learning for Remote Sensing Data
by: Heidler, Konrad, et al.
Published: (2021)
by: Heidler, Konrad, et al.
Published: (2021)
X-Streamer: Unified Human World Modeling with Audiovisual Interaction
by: Xie, You, et al.
Published: (2025)
by: Xie, You, et al.
Published: (2025)
Robust Disentangled Counterfactual Learning for Physical Audiovisual Commonsense Reasoning
by: Qi, Mengshi, et al.
Published: (2025)
by: Qi, Mengshi, et al.
Published: (2025)
CHIRP dataset: towards long-term, individual-level, behavioral monitoring of bird populations in the wild
by: Chan, Alex Hoi Hang, et al.
Published: (2026)
by: Chan, Alex Hoi Hang, et al.
Published: (2026)
Referee: Reference-aware Audiovisual Deepfake Detection
by: Boo, Hyemin, et al.
Published: (2025)
by: Boo, Hyemin, et al.
Published: (2025)
Enhanced Multimodal Content Moderation of Children's Videos using Audiovisual Fusion
by: Ahmed, Syed Hammad, et al.
Published: (2024)
by: Ahmed, Syed Hammad, et al.
Published: (2024)
AVoCaDO: An Audiovisual Video Captioner Driven by Temporal Orchestration
by: Chen, Xinlong, et al.
Published: (2025)
by: Chen, Xinlong, et al.
Published: (2025)
Probing Multimodal Fusion in the Brain: The Dominance of Audiovisual Streams in Naturalistic Encoding
by: Abdollahi, Hamid, et al.
Published: (2025)
by: Abdollahi, Hamid, et al.
Published: (2025)
JoVALE: Detecting Human Actions in Video Using Audiovisual and Language Contexts
by: Son, Taein, et al.
Published: (2024)
by: Son, Taein, et al.
Published: (2024)
Unified Speech Recognition: A Single Model for Auditory, Visual, and Audiovisual Inputs
by: Haliassos, Alexandros, et al.
Published: (2024)
by: Haliassos, Alexandros, et al.
Published: (2024)
SkipClick: Combining Quick Responses and Low-Level Features for Interactive Segmentation in Winter Sports Contexts
by: Schön, Robin, et al.
Published: (2025)
by: Schön, Robin, et al.
Published: (2025)
Adapting the Segment Anything Model During Usage in Novel Situations
by: Schön, Robin, et al.
Published: (2024)
by: Schön, Robin, et al.
Published: (2024)
U-Mind: A Unified Framework for Real-Time Multimodal Interaction with Audiovisual Generation
by: Deng, Xiang, et al.
Published: (2026)
by: Deng, Xiang, et al.
Published: (2026)
VioPose: Violin Performance 4D Pose Estimation by Hierarchical Audiovisual Inference
by: Yoo, Seong Jong, et al.
Published: (2024)
by: Yoo, Seong Jong, et al.
Published: (2024)
QDFormer: Towards Robust Audiovisual Segmentation in Complex Environments with Quantization-based Semantic Decomposition
by: Li, Xiang, et al.
Published: (2023)
by: Li, Xiang, et al.
Published: (2023)
Mask-guided cross-image attention for zero-shot in-silico histopathologic image generation with a diffusion model
by: Winter, Dominik, et al.
Published: (2024)
by: Winter, Dominik, et al.
Published: (2024)
AUD-TGN: Advancing Action Unit Detection with Temporal Convolution and GPT-2 in Wild Audiovisual Contexts
by: Yu, Jun, et al.
Published: (2024)
by: Yu, Jun, et al.
Published: (2024)
Efficient 2D to Full 3D Human Pose Uplifting including Joint Rotations
by: Ludwig, Katja, et al.
Published: (2025)
by: Ludwig, Katja, et al.
Published: (2025)
The ADUULM-360 Dataset -- A Multi-Modal Dataset for Depth Estimation in Adverse Weather
by: Schön, Markus, et al.
Published: (2024)
by: Schön, Markus, et al.
Published: (2024)
Seamless Interaction: Dyadic Audiovisual Motion Modeling and Large-Scale Dataset
by: Agrawal, Vasu, et al.
Published: (2025)
by: Agrawal, Vasu, et al.
Published: (2025)
Benchmarking Efficient & Effective Camera Pose Estimation Strategies for Novel View Synthesis
by: Meza, Jhacson, et al.
Published: (2026)
by: Meza, Jhacson, et al.
Published: (2026)
CineBrain: A Large-Scale Multi-Modal Brain Dataset During Naturalistic Audiovisual Narrative Processing
by: Gao, Jianxiong, et al.
Published: (2025)
by: Gao, Jianxiong, et al.
Published: (2025)
AquaMonitor: A multimodal multi-view image sequence dataset for real-life aquatic invertebrate biodiversity monitoring
by: Impiö, Mikko, et al.
Published: (2025)
by: Impiö, Mikko, et al.
Published: (2025)
Schrodinger Audio-Visual Editor: Object-Level Audiovisual Removal
by: Xu, Weihan, et al.
Published: (2025)
by: Xu, Weihan, et al.
Published: (2025)
A machine learning pipeline for automated insect monitoring
by: Jain, Aditya, et al.
Published: (2024)
by: Jain, Aditya, et al.
Published: (2024)
CT respiratory motion synthesis using joint supervised and adversarial learning
by: Cao, Yi-Heng, et al.
Published: (2024)
by: Cao, Yi-Heng, et al.
Published: (2024)
OpenForest: A data catalogue for machine learning in forest monitoring
by: Ouaknine, Arthur, et al.
Published: (2023)
by: Ouaknine, Arthur, et al.
Published: (2023)
Can we make NeRF-based visual localization privacy-preserving?
by: Pietrantoni, Maxime, et al.
Published: (2025)
by: Pietrantoni, Maxime, et al.
Published: (2025)
Similar Items
-
Embedding Arithmetic: A Lightweight, Tuning-Free Framework for Post-hoc Bias Mitigation in Text-to-Image Models
by: Sambandham, Venkatesh Thirugnana, et al.
Published: (2026) -
Unlocking Past Information: Temporal Embeddings in Cooperative Bird's Eye View Prediction
by: Rößle, Dominik, et al.
Published: (2024) -
DrivIng: A Large-Scale Multimodal Driving Dataset with Full Digital Twin Integration
by: Rößle, Dominik, et al.
Published: (2026) -
UrbanIng-V2X: A Large-Scale Multi-Vehicle, Multi-Infrastructure Dataset Across Multiple Intersections for Cooperative Perception
by: Sekaran, Karthikeyan Chandra, et al.
Published: (2025) -
A strongly annotated passive acoustic dataset for tropical bird monitoring
by: Ruiz, Daniela, et al.
Published: (2026)