Thriveth

The AI job market, made clearer

Find work worth
your expertise.

Back to all jobs
Posted 16 Sept 2026

Simplified Chinese Audio Transcriptionist (AI Training)

We're looking for detail-oriented transcriptionists in the New York City area to produce accurate, high-quality transcriptions of spoken Chinese audio — helping AI systems learn to process natural speech the way real people actually talk.

United States, New York City area$10–$35 per hour
Hiring companyAlignerr
Application processN/ANo approved reviews yet
Work experienceN/A1 of 3 reviews needed

About the role

What if your sharp ears and command of Simplified Chinese could directly shape how AI understands real human conversation?

New York City is home to one of the largest Chinese-speaking communities in North America, and we want to tap into that incredible linguistic talent. This is a fully remote, flexible contract role — no prior AI experience needed. Just native-level Simplified Chinese, strong listening skills, and comfort following detailed style and formatting guidelines.

Why Join Us

  • Work on cutting-edge AI speech and language projects alongside leading research labs.
  • Fully remote and flexible — work when and where it suits you.
  • Freelance autonomy with the structure of meaningful, task-based work.
  • Contribute to AI development that directly improves how technology understands spoken Chinese.
  • Build skills in a growing field at the intersection of language and AI.
  • Potential for ongoing work and contract extension as new projects launch.

Responsibilities

  • Listen to audio recordings in Chinese and produce accurate, verbatim transcriptions in Simplified Chinese.
  • Identify and distinguish between multiple speakers within a single recording.
  • Accurately capture crosstalk, overlapping speech, filled pauses (e.g., 嗯, 啊, 那个), false starts, and self-corrections.
  • Handle challenging audio conditions — background noise, low-quality recordings, varying accents and speaking speeds.
  • Apply detailed style guides and formatting conventions consistently across all transcriptions.
  • Flag unclear or unintelligible segments following established protocols.
  • Maintain high accuracy and consistency across large volumes of audio content.
  • Work independently and asynchronously on your own schedule.

Required skills

  • Exceptional listening skills with a sharp ear for linguistic nuance and subtle speech patterns.
  • Naturally detail-oriented with a methodical, patient approach to precise work.
  • Comfortable distinguishing between multiple speakers and handling complex audio scenarios.
  • Able to follow detailed formatting and style guidelines accurately and consistently.
  • Self-motivated and reliable when working independently.

Preferred skills

  • Professional experience in transcription, court reporting, or stenography.
  • Background in subtitle editing, closed-captioning, or media post-production.
  • Training or education in linguistics, phonetics, or language studies.
  • Prior work with annotation platforms, data labeling tools, or speech-to-text workflows.
  • Familiarity with different Chinese regional accents and dialects.
  • Experience in translation, interpreting, or bilingual content work.

Language requirements

  • Native or near-native fluency in Simplified Chinese — both listening and writing.

Equipment requirements

Access to a quiet workspace, reliable internet, and quality headphones.

Schedule details

10–40 hours per week.

Eligibility

Work arrangement: Fully remote

Eligible countries: United States

Eligible regions: New York City area

Compensation details

$10–$35 per hour.

More about the hiring companyAlignerr
Find AI-training work. Know what to expect.

Thriveth makes AI data-training work easier to find, understand, and navigate. We replace uncertainty with clear opportunities, realistic expectations, and insights from real application journeys.