AI Tutor - Serbian

📍 World
USD 35-45 per hour
MIDDLE
✅ Remote

🕙 10-40 hours per week

Tech Stack

AI @ 3 Communication @ 6 macOS

Details

As an AI Tutor specialized in multilingual audio capabilities, you will help train and refine Grok for voice interactions, speech recognition, and auditory experiences across languages, accents, and cultural contexts. The role focuses on curating and annotating high-quality audio data to improve multilingual speech processing and natural spoken interactions.

Responsibilities

  • Use proprietary software to provide labels, annotations, recordings, and inputs for projects involving multilingual audio clips, voice recordings, speech samples, and auditory elements.
  • Translate and localize UI copy, prompts, responses, and other written content from English into Serbian, ensuring linguistic accuracy, natural phrasing, and cultural appropriateness.
  • Deliver curated audio data representing clear, natural speech, linguistic and prosodic details, and professional audio standards.
  • Collaborate with technical staff to improve AI handling of speech modulation, accent variation, real-world recording noise, and multilingual audio processing.
  • Work with technical staff to improve annotation tools and audio workflows.

Requirements

  • Native proficiency in Serbian, with exposure to diverse accents, dialects, or regional variations.
  • English proficiency at a minimum B2 level, with clear and natural vocal delivery suitable for audio recording.
  • Strong auditory perception, including the ability to identify nuances in speech, accents, pronunciation, intonation, and audio quality.
  • Experience handling multilingual audio content and evaluating speech accuracy, cultural vocal expressions, and contextual interpretation.
  • Ability to transcribe audio accurately across accents and varying audio quality.
  • Ability to translate and localize text accurately between English and Serbian while preserving meaning, tone, and cultural nuance.
  • Comfort providing high-quality voice recordings and feedback on audio samples.
  • Strong comprehension, independent judgment, communication, interpersonal, analytical, detail-oriented, and organizational skills.
  • Preferred background in linguistics, phonetics, phonology, sociolinguistics, speech sciences, cognitive science, or a related field, or equivalent practical experience.
  • Preferred experience with speech or audio datasets, annotation workflows, AI training data, voice-model training, translation or localization workflows, voice work, audio production, or speech evaluation and research.
  • Advanced transcription and annotation experience involving disfluencies, accents, intonation, stress, rhythm, and emotion is preferred.
  • A portfolio of voice samples, annotated transcripts, or audio-related work is strongly preferred for advanced candidates.

Work Arrangement

  • Roles may be full-time, part-time, or contractor positions.
  • Contractor hours vary by project scope and availability. Most projects may require at least 10 hours per week, but there are no fixed commitments; contractors can set their own schedules.
  • Roles may be performed remotely from any location worldwide, subject to legal eligibility, time-zone compatibility, and role-specific needs.
  • US-based candidates cannot be hired in Wyoming or Illinois.
  • Visa sponsorship is not available.
  • Personal devices must be a Chromebook, a Mac running macOS 11.0 or later, or Windows 10 or later.

Compensation And Benefits

  • US-based candidates: $35-$45 per hour, depending on relevant experience, skills, education, geographic location, and qualifications.
  • International compensation information will be provided during the recruitment process.
  • Eligible US-based positions may include health insurance, a 401(k) plan, and paid sick leave. Benefits vary by employment type, location, and jurisdiction.

More jobs at SpaceXAI

Similar jobs