Thriveth

The AI job market, made clearer

Find work worth
your expertise.

Back to all jobs
Posted 16 Sept 2026

Simplified Chinese Audio Transcript Editor

We're looking for Audio Transcript Editors in the New York metro area to clean up, correct, and polish audio transcripts — ensuring every character, phrase, and timestamp is precisely accurate.

United States, New York metro area$10–$35 per hour
Hiring companyAlignerr
Application processN/A1 of 3 reviews needed
Work experienceN/A1 of 3 reviews needed

About the role

What if your fluency in Simplified Chinese and your sharp ear for spoken language could directly influence how AI understands Mandarin for hundreds of millions of people?

New York's large, highly educated Mandarin-speaking community is an ideal home for this kind of precision language work. You'll listen to recordings, compare them against existing transcripts, and make the careful, word-level corrections that train AI systems to genuinely understand spoken Chinese. If you notice when a single character is off or a timestamp is misaligned, this role was built for you.

This is a fully remote, flexible contract role — work from anywhere on your own schedule. No prior AI experience required.

Why Join Us

  • Work on cutting-edge AI projects alongside leading research labs.
  • Fully remote and flexible — work when and where it suits you.
  • Freelance autonomy with the structure of meaningful, task-based work.
  • Contribute to AI development that directly improves how technology understands spoken Chinese.
  • Develop and sharpen highly marketable transcription and annotation skills.
  • Potential for ongoing work and contract extension as new projects launch.

Responsibilities

  • Listen to audio recordings and compare them against existing transcripts in Simplified Chinese.
  • Correct errors in transcription — including wrong characters, missing words, misheard phrases, and punctuation issues.
  • Ensure word-level accuracy and verify exact timing and alignment of transcript segments.
  • Follow detailed style and formatting guidelines consistently across all tasks.
  • Flag audio quality issues, ambiguous speech, or segments that require further review.
  • Maintain high accuracy and consistency across large volumes of transcript data.
  • Work independently on task-based assignments at your own pace and on your own schedule.

Required skills

  • Exceptional listening skills with the ability to parse speech clearly, even in challenging audio conditions.
  • Extremely detail-oriented with a methodical, patient approach to repetitive precision work.
  • Comfortable working with word-level accuracy and exact timing or timestamps.
  • Able to follow structured, detailed style and formatting guidelines with consistency.
  • Self-motivated and reliable when working independently.

Preferred skills

  • Experience as a transcriptionist, court reporter, or stenographer.
  • Background in subtitle editing, closed-captioning, or dubbing.
  • Linguistics, phonetics, or language studies education.
  • Prior work with annotation platforms, data labeling tools, or speech-to-text systems.
  • Familiarity with audio editing software or transcript alignment tools.
  • Experience working with multiple Chinese dialects or regional accents.

Experience and education

  • No prior AI or tech experience required.

Language requirements

  • Native or near-native proficiency in Simplified Chinese — reading, writing, and listening.

Schedule details

10–40 hours per week.

Eligibility

Work arrangement: Fully remote

Eligible countries: United States

Eligible regions: New York metro area

Compensation details

$10–$35 per hour.

More about the hiring companyAlignerr
Find AI-training work. Know what to expect.

Thriveth makes AI data-training work easier to find, understand, and navigate. We replace uncertainty with clear opportunities, realistic expectations, and insights from real application journeys.