0% found this document useful (0 votes)
12 views4 pages

Audio Annotation Guidelines

A detailed documentation of audio annotation for Ghana's National Science and Math Quiz.

Uploaded by

vbonskinkelsberg
Copyright
© All Rights Reserved
We take content rights seriously. If you suspect this is your content, claim it here.
Available Formats
Download as PDF, TXT or read online on Scribd
0% found this document useful (0 votes)
12 views4 pages

Audio Annotation Guidelines

A detailed documentation of audio annotation for Ghana's National Science and Math Quiz.

Uploaded by

vbonskinkelsberg
Copyright
© All Rights Reserved
We take content rights seriously. If you suspect this is your content, claim it here.
Available Formats
Download as PDF, TXT or read online on Scribd

Audio Annotation Process

Order of annotations
1. Conversation with both partners as yes
2. Speech with at least one partner as yes

Annotation Process
1. Copy the data of the couple to be annotated onto your computer (P_ID and
Z_ID)
○ Don’t do the annotation in the original folder
○ After the annotation of that subject’s data, delete the audio data from your
computer including from the trash bin
2. Find the audio in the Voice Sample folder and listen to the voices of each
partner to help with the annotation
○ If it is not clear the voice of each partner, you can also watch the video
recording of the partner’s lab conversation (if there is one)
3. For the order of audios to annotate, go in the order of the id so P001 and Z001,
and P002, Z002 etc. And for the files within each id, start with Day_1, Day_2
until Day_0. And for hours, minimum (e.g. 6) to max (e.g. 23). Check if the
corresponding audio of the corresponding partner at the day-hour-time exist and
annotate them before moving on the next signal.
4. Also, since most of the audios from each partner are collected at the same time
but with different devices (except for a few exceptions), it’s best to annotate audio
of say P_ID at a certain timestamp, and then the audio of Z_ID of the same
timestamp (if it exists, then back to P_ID) and continue in that order.
5. Open Audacity. Then click File->Open and selected the audio to be annotated.
Always listen to the audio that has “[Link]” at the end as it cleaned version
(without much background noise). There will be pop-up shown below. Always
select the option “Make a copy of the files before editing (safer)”
6. Add the name of the data to the file. Go to Edit -> Metadata and for the column
Artist name (eg P001) and Artist title, at the day, hour, and date separated by
hyphen(-) eg (Day_1-Hour_6-11-18-2019 06_44_17)
7. Annotate the data by highlighting sections of the audio waveform (click and drag)
and then press Ctrl + b (Windows)/ Cmd + b (Mac) to bring up an annotation
textbox
○ Start of the next annotation should align perfectly with the end of the
previous annotation
8. Edit the textbox with the following annotations which apply
○ m (male partner speaking)
○ f (female partner speaking)
○ p (pause no one speaking, silence)
○ c - (cross-talk both partners speaking)
○ c - u (cross-talk between any of the partners and an unknown person)
○ u - f (unknown female speaking)
○ u - m (unknown male speaking)
○ u - c (cross talk between unknown people)
○ u - tv/radio (unknown speech from radio or tv)
○ u (unknown person speaking - can’t tell gender)
○ n - noise (music, movements of the watch, vehicles, wind,
kitchen appliances, claps)

These can overlap with other annotations


○ n - noise (music, movements of the watch, vehicles, wind,
kitchen appliances, claps)
○ v - f (vocalization such as laugh, sighs, etc from female partner)
○ v - u - f (vocalization such as laugh, sighs, etc from unknown female)
○ v - m (vocalization such as laugh, sighs, etc from male partner)
○ v - u - m (vocalization such as laugh, sighs, etc from unknown male)

9. After the first annotation, save the project in the Annotated folders on your
computer. Create a new folder with the name structure below to contain the
project files and press save.
○ P001
i. Day_1_Hour_19
1. audios
a. [Link]
b. watchRecordAudio_filtered
2. annotations
a.

10. After the annotation go to File-> Export->Labels to the annotations subfolder


11. Then move the folder subject folder (eg P003/Day_1_Hour_19) to the mirror of
the folder structure on the server
○ DyMand Data-> Annotated
12. Note any key things such as no speech in the annotation sheet

Verification Process
● The audio files within the DyMand Data folder are first compared to the audio
files within the Audio subfolders found within the Annotation folder.
● This is done for all the audio files
● Next the audio files within the Audio subfolders are compared to the annotated
audio files found within each associated Day-Hour recording for each couple.
● The annotations from the audio files are then cross-checked with the associated
exported text files to see if they match. If not then appropriate corrections are
made.

Important notes
● Do not delete any data
● Save your work continuously
● It’s sometimes difficult to identify whether it’s one of the 2 partners speaking or
someone else

Common questions

Powered by AI

Audacity supports the audio annotation process by providing features like opening filtered audio files for cleaner analysis and options to make a safe copy before editing. It also facilitates annotation through textboxes, enabling precise marking of segments related to speech, pauses, noise, or other relevant audio features .

Annotators might face challenges in distinguishing between the partners' voices and unknown speakers, or differentiating various background noises and vocalizations. The process addresses these by providing specific annotation codes for different scenarios (e.g., male, female, unknown speakers, noise types), which guide annotators in categorizing audio segments accurately .

A structured folder system is important as it enables systematic organization, easy navigation, and efficient retrieval of annotated projects. By following a consistent naming and storing protocol (e.g., P001/Day_1_Hour_19 structure), the system allows collaborators to locate and verify data expediently, thereby improving data management and reducing the risk of errors .

The order of annotation is crucial because it helps maintain consistency and reduces errors by allowing annotations to follow a systematic format (e.g., P_ID audio followed by Z_ID audio at the same timestamp). This is particularly important when audios are collected simultaneously using different devices, as it ensures that all relevant data for a specific time point is considered before moving to another .

The initial steps in the audio annotation process include copying the data of the couple to be annotated onto your computer (P_ID and Z_ID) without performing the annotation in the original folder, followed by deleting the audio data after annotation . This is crucial for maintaining data security and integrity by ensuring no unauthorized alterations are made in the original data set and maintaining confidentiality by removing local copies post-annotation .

Documentation of key observations, such as noting instances of no speech, plays a crucial role in capturing exceptions and details that may impact analysis or interpretation of the data. This documentation ensures transparency and enables future reviewers to understand contextual nuances of the annotation process .

Overlapping annotations allow annotators to represent scenarios where multiple audio elements occur simultaneously, such as noise overlapping with speech or vocalizations. This enriches the data by providing a more detailed and nuanced understanding of complex audio scenes, which can be crucial for accurate analysis and contextual interpretations .

The annotation process ensures clarity and accuracy in identifying speakers by allowing annotators to listen to voice samples and, if necessary, watch video recordings to familiarize themselves with each partner’s voice. This helps in correctly attributing speech segments to the correct individual during annotation .

Metadata editing is significant as it incorporates essential identifiers such as Artist name and title, day, hour, and date into the audio file. This structured naming and identification enhance data organization, enabling easy tracking and retrieval of specific audio files for review or further analysis .

The verification process ensures reliability by comprehensively comparing audio files within the DyMand Data folder with the annotated audio files and exported text files. Any discrepancies identified in the annotations are corrected, ensuring that the annotated data accurately reflects the original recordings .

You might also like