ChimpVLM: Ethogram-Enhanced Chimpanzee Behaviour Recognition
Fuente:
arXiv
Saved in:
| Main Authors: | Brookes, Otto, Mirmehdi, Majid, Kuhl, Hjalmar, Burghardt, Tilo |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Deep in the Jungle: Towards Automating Chimpanzee Population Estimation
by: Raynes, Tom, et al.
Published: (2026)
by: Raynes, Tom, et al.
Published: (2026)
The PanAf-FGBG Dataset: Understanding the Impact of Backgrounds in Wildlife Behaviour Recognition
by: Brookes, Otto, et al.
Published: (2025)
by: Brookes, Otto, et al.
Published: (2025)
AlphaChimp: Tracking and Behavior Recognition of Chimpanzees
by: Ma, Xiaoxuan, et al.
Published: (2024)
by: Ma, Xiaoxuan, et al.
Published: (2024)
Video-SwinUNet: Spatio-temporal Deep Learning Framework for VFSS Instance Segmentation
by: Zeng, Chengxi, et al.
Published: (2023)
by: Zeng, Chengxi, et al.
Published: (2023)
Visual-textual Dermatoglyphic Animal Biometrics: A First Case Study on Panthera tigris
by: Li, Wenshuo, et al.
Published: (2025)
by: Li, Wenshuo, et al.
Published: (2025)
The SA-FARI Dataset: Segment Anything in Footage of Animals for Recognition and Identification
by: Wasmuht, Dante Francisco, et al.
Published: (2025)
by: Wasmuht, Dante Francisco, et al.
Published: (2025)
WildLive: Near Real-time Visual Wildlife Tracking onboard UAVs
by: Dat, Nguyen Ngoc, et al.
Published: (2025)
by: Dat, Nguyen Ngoc, et al.
Published: (2025)
Universal Bovine Identification via Depth Data and Deep Metric Learning
by: Sharma, Asheesh, et al.
Published: (2024)
by: Sharma, Asheesh, et al.
Published: (2024)
FastVLM: Efficient Vision Encoding for Vision Language Models
by: Vasu, Pavan Kumar Anasosalu, et al.
Published: (2024)
by: Vasu, Pavan Kumar Anasosalu, et al.
Published: (2024)
DSCA: Dynamic Subspace Concept Alignment for Lifelong VLM Editing
by: Das, Gyanendra, et al.
Published: (2026)
by: Das, Gyanendra, et al.
Published: (2026)
CollaFuse: Collaborative Diffusion Models
by: Allmendinger, Simeon, et al.
Published: (2024)
by: Allmendinger, Simeon, et al.
Published: (2024)
Towards Application-Specific Evaluation of Vision Models: Case Studies in Ecology and Biology
by: Chan, Alex Hoi Hang, et al.
Published: (2025)
by: Chan, Alex Hoi Hang, et al.
Published: (2025)
Behavioural Cloning in VizDoom
by: Spick, Ryan, et al.
Published: (2024)
by: Spick, Ryan, et al.
Published: (2024)
A Recipe for Improving Remote Sensing VLM Zero Shot Generalization
by: Barzilai, Aviad, et al.
Published: (2025)
by: Barzilai, Aviad, et al.
Published: (2025)
GFlowVLM: Enhancing Multi-step Reasoning in Vision-Language Models with Generative Flow Networks
by: Kang, Haoqiang, et al.
Published: (2025)
by: Kang, Haoqiang, et al.
Published: (2025)
VLM Agents Generate Their Own Memories: Distilling Experience into Embodied Programs of Thought
by: Sarch, Gabriel, et al.
Published: (2024)
by: Sarch, Gabriel, et al.
Published: (2024)
LlavaGuard: An Open VLM-based Framework for Safeguarding Vision Datasets and Models
by: Helff, Lukas, et al.
Published: (2024)
by: Helff, Lukas, et al.
Published: (2024)
Anatomy-VLM: A Fine-grained Vision-Language Model for Medical Interpretation
by: Gu, Difei, et al.
Published: (2025)
by: Gu, Difei, et al.
Published: (2025)
EITNet: An IoT-Enhanced Framework for Real-Time Basketball Action Recognition
by: Liu, Jingyu, et al.
Published: (2024)
by: Liu, Jingyu, et al.
Published: (2024)
Enhancing Personality Recognition by Comparing the Predictive Power of Traits, Facets, and Nuances
by: Ansari, Amir, et al.
Published: (2026)
by: Ansari, Amir, et al.
Published: (2026)
From Masks to Pixels and Meaning: A New Taxonomy, Benchmark, and Metrics for VLM Image Tampering
by: Shang, Xinyi, et al.
Published: (2026)
by: Shang, Xinyi, et al.
Published: (2026)
TRISHUL: Towards Region Identification and Screen Hierarchy Understanding for Large VLM based GUI Agents
by: Singh, Kunal, et al.
Published: (2025)
by: Singh, Kunal, et al.
Published: (2025)
VLM-SubtleBench: How Far Are VLMs from Human-Level Subtle Comparative Reasoning?
by: Kim, Minkyu, et al.
Published: (2026)
by: Kim, Minkyu, et al.
Published: (2026)
Neural Sentinel: Unified Vision Language Model (VLM) for License Plate Recognition with Human-in-the-Loop Continual Learning
by: Sivakoti, Karthik
Published: (2026)
by: Sivakoti, Karthik
Published: (2026)
NeuroVLM-Bench: Evaluation of Vision-Enabled Large Language Models for Clinical Reasoning in Neurological Disorders
by: Dineva, Katarina Trojachanec, et al.
Published: (2026)
by: Dineva, Katarina Trojachanec, et al.
Published: (2026)
Enhancing Fine-Grained Visual Recognition in the Low-Data Regime Through Feature Magnitude Regularization
by: Chapman, Avraham, et al.
Published: (2024)
by: Chapman, Avraham, et al.
Published: (2024)
PanAf20K: A Large Video Dataset for Wild Ape Detection and Behaviour Recognition
by: Brookes, Otto, et al.
Published: (2024)
by: Brookes, Otto, et al.
Published: (2024)
PaliGemma: A versatile 3B VLM for transfer
by: Beyer, Lucas, et al.
Published: (2024)
by: Beyer, Lucas, et al.
Published: (2024)
Resolution scaling governs DINOv3 transfer performance in chest radiograph classification
by: Arasteh, Soroosh Tayebi, et al.
Published: (2025)
by: Arasteh, Soroosh Tayebi, et al.
Published: (2025)
CropVLM: Learning to Zoom for Fine-Grained Vision-Language Perception
by: Carvalho, Miguel, et al.
Published: (2025)
by: Carvalho, Miguel, et al.
Published: (2025)
Make VLM Recognize Visual Hallucination on Cartoon Character Image with Pose Information
by: Kim, Bumsoo, et al.
Published: (2024)
by: Kim, Bumsoo, et al.
Published: (2024)
Agglomerating Large Vision Encoders via Distillation for VFSS Segmentation
by: Zeng, Chengxi, et al.
Published: (2025)
by: Zeng, Chengxi, et al.
Published: (2025)
TrafficVLM: A Controllable Visual Language Model for Traffic Video Captioning
by: Dinh, Quang Minh, et al.
Published: (2024)
by: Dinh, Quang Minh, et al.
Published: (2024)
Promoting AI Equity in Science: Generalized Domain Prompt Learning for Accessible VLM Research
by: Cao, Qinglong, et al.
Published: (2024)
by: Cao, Qinglong, et al.
Published: (2024)
Your Turn: At Home Turning Angle Estimation for Parkinson's Disease Severity Assessment
by: Cheng, Qiushuo, et al.
Published: (2024)
by: Cheng, Qiushuo, et al.
Published: (2024)
ECOR: Explainable CLIP for Object Recognition
by: Rasekh, Ali, et al.
Published: (2024)
by: Rasekh, Ali, et al.
Published: (2024)
VLM in a flash: I/O-Efficient Sparsification of Vision-Language Model via Neuron Chunking
by: Yang, Kichang, et al.
Published: (2025)
by: Yang, Kichang, et al.
Published: (2025)
EgoBabyVLM: Benchmarking Cross-Modal Learning from Naturalistic Egocentric Video Data
by: Lin, Dongyan, et al.
Published: (2026)
by: Lin, Dongyan, et al.
Published: (2026)
Revolutionizing Communication with Deep Learning and XAI for Enhanced Arabic Sign Language Recognition
by: Balat, Mazen, et al.
Published: (2025)
by: Balat, Mazen, et al.
Published: (2025)
Exploring the Spectrum of Visio-Linguistic Compositionality and Recognition
by: Oh, Youngtaek, et al.
Published: (2024)
by: Oh, Youngtaek, et al.
Published: (2024)
Similar Items
-
Deep in the Jungle: Towards Automating Chimpanzee Population Estimation
by: Raynes, Tom, et al.
Published: (2026) -
The PanAf-FGBG Dataset: Understanding the Impact of Backgrounds in Wildlife Behaviour Recognition
by: Brookes, Otto, et al.
Published: (2025) -
AlphaChimp: Tracking and Behavior Recognition of Chimpanzees
by: Ma, Xiaoxuan, et al.
Published: (2024) -
Video-SwinUNet: Spatio-temporal Deep Learning Framework for VFSS Instance Segmentation
by: Zeng, Chengxi, et al.
Published: (2023) -
Visual-textual Dermatoglyphic Animal Biometrics: A First Case Study on Panthera tigris
by: Li, Wenshuo, et al.
Published: (2025)