Review pre-segmented audio recordings, verify word-level segment accuracy, and correct timestamp or transcription errors to support training data for speech-to-speech AI models.
Engagement Type: Freelance (Pilot).
Responsibilities
Verify segment boundaries align with actual speech onset and offset.
Review start/end timestamps when boundaries are inaccurate.
Identify and flag tasks for discard per defined criteria (unintelligible audio, non-target language, sensitive content, no voice activity, etc.).
Maintain consistency and accuracy across a detailed, evolving style guide.
Required skills
Comfortable using annotation/labeling tools.
Ability to reference and apply a detailed style guide consistently.
Ability to work independently and adapt to periodic guideline updates.
Experience and education
Prior experience with audio transcription, data annotation, or speech/NLP data labeling preferred.
Language requirements
English fluency, including strong command of spoken/colloquial forms (contractions, filler words, informal speech).
Thriveth makes AI data-training work easier to find, understand, and navigate. We replace uncertainty with clear opportunities, realistic expectations, and insights from real application journeys.