SeaTurtleID: A novel long-span dataset highlighting the importance of timestamps in wildlife re-identification

11/18/2022
by   Kostas Papafitsoros, et al.
0

This paper introduces SeaTurtleID, the first public large-scale, long-span dataset with sea turtle photographs captured in the wild. The dataset is suitable for benchmarking re-identification methods and evaluating several other computer vision tasks. The dataset consists of 7774 high-resolution photographs of 400 unique individuals collected within 12 years in 1081 encounters. Each photograph is accompanied by rich metadata, e.g., identity label, head segmentation mask, and encounter timestamp. The 12-year span of the dataset makes it the longest-spanned public wild animal dataset with timestamps. By exploiting this unique property, we show that timestamps are necessary for an unbiased evaluation of animal re-identification methods because they allow time-aware splits of the dataset into reference and query sets. We show that time-unaware splits can lead to performance overestimation of more than 100 CNN-based re-identification methods. We also argue that time-aware splits correspond to more realistic re-identification pipelines than the time-unaware ones. We recommend that animal re-identification methods should only be tested on datasets with timestamps using time-aware splits, and we encourage dataset curators to include such information in the associated metadata.

READ FULL TEXT

page 1

page 3

page 5

page 6

page 12

page 13

page 14

research
06/26/2017

VoxCeleb: a large-scale speaker identification dataset

Most existing datasets for speaker identification contain samples obtain...
research
06/13/2019

Amur Tiger Re-identification in the Wild

Monitoring the population and movements of endangered species is an impo...
research
08/17/2022

DeepSportradar-v1: Computer Vision Dataset for Sports Understanding with High Quality Annotations

With the recent development of Deep Learning applied to Computer Vision,...
research
08/13/2022

Enhanced Vehicle Re-identification for ITS: A Feature Fusion approach using Deep Learning

In recent years, the development of robust Intelligent transportation sy...
research
02/13/2019

Multi-views Embedding for Cattle Re-identification

People re-identification task has seen enormous improvements in the late...
research
03/28/2020

A Dataset of Dockerfiles

Dockerfiles are one of the most prevalent kinds of DevOps artifacts used...

Please sign up or login with your details

Forgot password? Click here to reset