sammlapp/ovenbird-annotated-1k
1000 3-second audio clips annotated for Ovenbird song presence This dataset contains 100 3 second audio clips randomly selected from a 4-year passive acoustic monitoring dataset at 126 recording location in Pennsylvania. Experts familiar with Ovenbird song reviewed the audio and spectrograms using the Dipper app. The clip_annotations.csv file contains the annotation column with 'yes' for confirmed presence (N=92), 'no' for confirmed absence (894), or 'uncertain' if presence of… See the full description on the dataset page: https://huggingface.co/datasets/sammlapp/ovenbird-annotated-1k.
1000 3-second audio clips annotated for Ovenbird song presence
This dataset contains 100 3 second audio clips randomly selected from a 4-year passive acoustic monitoring dataset at 126 recording location in Pennsylvania. Experts familiar with Ovenbird song reviewed the audio and spectrograms using the Dipper app. The clip_annotations.csv file contains the annotation column with 'yes' for confirmed presence (N=92), 'no' for confirmed absence (894), or 'uncertain' if presence of Ovenbird song could not be confirmed or rejected via review (14).
This dataset was used to evaluate acoustic species classifier performance for Ovenbird song presence. The methods and data associated with this dataset are described in detail in an associated manuscript:
Sam Lapp, R. Patrick Lyon, Scott J. Wilson, Tessa A. Rhinehart, Chapin Czarnecki, Lauren M. Chronister, Cameron J. Fiss, Jeffery L. Larkin, Erin Bayne, and Justin Kitzes, in review. "Automated identification of individual birds by song enables multi-year recapture from passive acoustic monitoring data".
A preprint of the article is also publicly available.
A codebase containing the associated scripts and analyses is available on GitHub
Contents:
clip_annotations.csv: table with the relative path of each file (file), start time of annotation in seconds relative to audio file (always 0), and the annotation (yes, uncertain, or no) corresponding to the presence of Ovenbird song
./audio/ contains 1000 3-second .wav files
Note that in the HuggingFace repository, this folder has been merged into a .tar file. After download, the .tar can be extracted to the original audio/ folder via right click and 'extract' or 'expand' on Windows/Linux or by double-click on Mac.
