Search papers, labs, and topics across Lattice.
This study introduces the PanAf-SBR dataset, the first annotated collection of social behaviours in wild great apes captured through camera traps, addressing a critical gap in conservation research. By extending the existing PanAf500 dataset with 100 new videos and over 81,000 annotations, the authors establish benchmarks for fine-grained social behaviour recognition using the AlphaChimp architecture. Their findings reveal that cross-dataset pre-training enhances recognition performance for specific social behaviour classes, highlighting the importance of context in automated behaviour detection.
The PanAf-SBR dataset reveals that fine-grained social behaviour recognition in wild great apes can be significantly improved through targeted cross-dataset pre-training.
Behavioural shifts in wild great ape populations, particularly the breakdown of social structures, can serve as an early indicator of population decline. Automating the detection of behaviours indicative of these shifts is therefore a critical task for conservation. Several valuable datasets have recently been introduced for the automated recognition of great ape behaviour, yet few include fine-grained social behaviour annotations, and those that do are captured either in captive settings or via aerial platforms such as UAVs. We address this gap by introducing PanAf-SBR, the first wild great ape camera trap dataset annotated with social behaviours. PanAf-SBR extends PanAf500 with 100 additional videos covering 36,063 frames. These come with 81,096 annotations including bounding boxes, segmentation masks, intra-video identities, and seven social behaviour classes defined under the action giver and receiver convention of ChimpACT. We use this data together with the AlphaChimp architecture to establish the first benchmarks for fine-grained social behaviour recognition in wild great apes from camera trap footage. We further conduct bidirectional transfer learning experiments between PanAf-SBR and the captive ChimpACT dataset, finding that cross-dataset pre-training is highly beneficial for specific classes rather than of uniform benefit. Finally, we examine the role of background context by inverting the segmentation masks to suppress non-ape pixels.