AI Tutor - Urdu

📍 World
USD 35-45 per hour
MIDDLE
✅ Remote

🕙 10-40 hours per week

Tech Stack

AI @ 3 Communication @ 6 macOS @ 3

Details

As an AI Tutor specializing in multilingual audio capabilities, you will help train and refine Grok for voice interactions, speech recognition, and auditory experiences across languages, accents, and cultural contexts. The role focuses on curating and annotating high-quality audio data to improve global accessibility, natural spoken interactions, speech processing, and multilingual audio understanding.

Responsibilities

  • Use proprietary software to provide labels, annotations, recordings, and inputs for projects involving multilingual audio clips, voice recordings, speech samples, and auditory elements.
  • Support the delivery of curated audio data with accurate linguistic and prosodic details, including intonation, rhythm, and accent, while maintaining professional audio standards.
  • Collaborate with technical staff on tasks that improve AI handling of speech modulation, accent variation, real-world recording noise, and multilingual audio processing.
  • Work with technical staff to improve annotation tools and audio workflows.

Requirements

  • Native proficiency in Urdu, including exposure to diverse accents, dialects, or regional variations.
  • English proficiency at a minimum B2 level, with clear and natural vocal delivery suitable for audio recording.
  • Strong auditory perception for identifying speech, accent, pronunciation, intonation, and audio-quality nuances.
  • Ability to handle multilingual audio content, evaluate speech accuracy and cultural vocal expressions, and interpret spoken context.
  • Demonstrated ability to transcribe audio accurately across accents and varying audio quality.
  • Comfort providing high-quality voice recordings and feedback on audio samples in multiple languages.
  • Strong comprehension, independent judgment, communication, interpersonal, analytical, organizational, and attention-to-detail skills.
  • Exceptional attention to linguistic nuance, auditory detail, and data quality.
  • Advanced transcription and annotation experience, including disfluencies, accents, intonation, stress, rhythm, and emotion.
  • Background in linguistics, phonetics, phonology, sociolinguistics, speech sciences, cognitive science, or a related field, or equivalent practical experience.
  • Experience with speech or audio datasets, annotation workflows, AI training data, voice-model training, or understanding the impact of data quality on model performance.
  • Professional voice experience, such as voice acting, voice recording, podcasting, or similar audio production.
  • Ability to make consistent and defensible annotation decisions in ambiguous audio scenarios.
  • A portfolio of voice samples, annotated transcripts, or related audio work is strongly preferred.
  • For work from a personal device, the computer must be a Chromebook, a Mac running macOS 11.0 or later, or Windows 10 or later.

Work Arrangements

  • Roles may be full-time, part-time, or contractor positions.
  • Contractor hours vary by project scope and availability. Most projects may require at least 10 hours per week, but there are no fixed commitments; contractors can set their own hours.
  • The role may be performed remotely from any location worldwide, subject to legal eligibility, time-zone compatibility, and role-specific needs.
  • US-based candidates cannot be hired in Wyoming or Illinois.
  • Visa sponsorship is not provided.

Compensation And Benefits

  • US-based candidates: $35–$45 per hour, depending on experience, skills, education, geographic location, and qualifications.
  • Compensation information for international candidates will be provided during recruitment.
  • Eligible US-based positions may include health insurance, a 401(k) plan, and paid sick leave. Benefits vary by employment type, location, and jurisdiction.

More jobs at SpaceXAI

Similar jobs