Role: AI Trainer-Audio (English) || Contractual
Job Description
We are looking for detail-oriented scholars to work on conversational speech transcription editing and timestamp alignment. The task involves reviewing ASR-generated transcripts against audio recordings, correcting errors according to client conventions, and adjusting segment start and end times to accurately match the speech. The content includes casual US English with disfluencies, false starts, slang, and regional dialects, so strong listening accuracy is essential.
Responsibilities
- Listen to naturalistic conversational audio files and review ASR-generated transcripts.
- Correct transcription errors following client-provided conventions and formatting rules.
- Adjust start and end timestamps for each segment to match the audio precisely.
- Maintain high accuracy and consistency across all assigned files.
- Follow instructions carefully and complete tasks within required timelines.
Required Skills
- Auditory C1 English or above listening proficiency is required.
- Written B2 English or above.
- Excellent attention to detail and speed.
- Ability to follow instructions exactly and work with precision.
- Comfort handling casual speech, disfluencies, slang, and regional accents.
Preferred Profile
- Experience in transcription, subtitling, annotation, or speech-to-text editing.
- Familiarity with audio-text alignment workflows.
- Strong grammar, punctuation, and proofreading skills.
- Ability to work independently and meet deadlines consistently.
This is a great opportunity for candidates who have excellent listening accuracy and enjoy working with language, detail, and precision.