AI training contract
Music & Sound Annotation Expert
Language & Audio · USD 30-45 an hour. · 10 openings
Music & Sound Annotation Expert is contract work for micro1, reviewing and correcting what AI models produce in language & audio. The output of the job is expert judgement written down: a model gives an answer, someone who knows the subject decides whether it is right and explains why.
What it requires. No credential is stated on this listing beyond experience in the field.
What it pays. USD 30 to USD 45 an hour, as published by micro1. Rates in this category track how rare the credential is rather than how hard the work is.
Where. The listing says this role is remote and names no country. micro1 says it hires talent from around the world, except people who reside in Afghanistan, Belarus, China, Cuba, Democratic Republic of the Congo, Hong Kong, Iran, Iraq, Libya, Macao, Myanmar, North Korea, Russian Federation, Somalia, South Sudan, Sudan, Syria, Venezuela, Ukraine, and Yemen. So this page lists every other country as eligible, on the strength of that rule and not of anything this listing says. Requirements can still differ by role, so check the application page on micro1 before you apply.
micro1 lists 10 openings on this role.
Skills named: analytical listening, music theory knowledge, audio annotation, attention to detail, written communication.
Applications are made on micro1, not here. WFA Digital is independent of micro1 and may receive a referral fee if you are hired, at no cost to you.
Full listing, as published by micro1
Role Title: Music & Sound Annotation Expert Role Type: Contractor Location: Remote micro1 is engaging Music & Sound Annotation Experts to contribute to a customer's project at the forefront of music and audio technology. Your ears are the ground truth. You'll listen to short field recordings and document exactly what's audible — instruments, tempo, key, background sounds, crowd reactions — through a structured web platform built for this project. Later phases involve grading AI model answers against what you actually heard. What you'll do Listen to 30–90 second field recordings (headphones, your own quiet workspace) and complete structured annotations: which instruments are present, tempo and key, timestamps of background events (traffic, sirens, applause), scene details Compare paired recordings of the same song captured at different locations and judge what stayed the same and what changed In later phases: review AI model descriptions of clips and mark what's correct, wrong, or invented, with a brief written explanation Attend a paid onboarding/calibration session and a short weekly sync Who we're looking for (any of these backgrounds) Working or gigging musicians, session players, producers, or mixing engineers Music educators, ear-training instructors, band/choir directors, conservatory students or graduates Location sound recordists, sound designers, foley artists, or podcast/documentary audio editors Requirements 3+ years of hands-on music or audio experience Strong analytical listening: you can pick individual instruments out of a busy, noisy, real-world mix — not just studio recordings Can identify tempo (within a few BPM) and musical key (relative pitch with a reference is fine) Comfortable with detail-oriented, repetitive annotation work in a web platform — dropdowns, timestamps, short structured notes Clear written English for brief, evidence-based notes (e.g., \"accordion enters around 0:12\") Quality headphones, quiet workspace, reliable internet Available 20–25 hrs/week for the engagement window
All AI training roles we track · Every micro1 opportunity with filters
Rate, openings and requirements checked 2026-09-25 against micro1's own public postings.