Skip to main content
SpaceXAI

AI Tutor - Serbian

RemoteUnited States only
Published
Employment
Part-time10h/week
Company size
Startup
$35–$45/hr
Check eligibility

Open to US only. Set where you work from to check your eligibility.

No BS summary

AI tutor for Serbian audio and localization work. Needs native Serbian, English B2+, strong transcription/audio annotation skills, and ability to work with multilingual speech data. Remote worldwide subject to legal eligibility, timezone compatibility, and role needs; no visa sponsorship.

Required languages

Serbian NativeEnglish B2+

What you'll do

  • Use proprietary software to provide labels, annotations, recordings, and inputs on projects involving multilingual audio clips, voice recordings, speech samples, and auditory elements in various languages.
  • Translate and localize text strings (UI copy, prompts, responses, and other written content) from English into Serbian, ensuring linguistic accuracy, natural phrasing, and cultural appropriateness.
  • Support the delivery of high-quality curated audio data that ensures clear, natural spoken output, accurate representation of linguistic and prosodic details, and professional audio standards.
  • Collaborate with technical staff to develop tasks that improve AI's ability to handle speech modulation, accent variation, noise in real-world recordings, and multilingual audio processing.
  • Work with technical staff to improve annotation tools for efficient audio workflows.

What they require

  • Native proficiency in Serbian with exposure to diverse accents, dialects, or regional variations.
  • Proficiency in English (minimum B2 level) with clear, natural vocal delivery and pronunciation suitable for audio recording purposes.
  • Strong auditory perception to identify nuances in speech, accents, pronunciation, intonation, and audio quality across languages.
  • Demonstrated ability to handle multilingual audio content, including evaluating speech accuracy, cultural vocal expressions, and contextual interpretation in spoken form.
  • Demonstrated ability to transcribe audio with high accuracy across accents and varying audio quality.
  • Demonstrated ability to translate and localize text accurately between English and Serbian, preserving meaning, tone, and cultural nuance.
  • Comfort providing high-quality voice recordings and feedback on audio samples in multiple languages.
  • Strong comprehension skills and the ability to make independent judgments on ambiguous or varied audio material, including noisy or accented speech.
  • Strong communication, interpersonal, analytical, detail-oriented, and organizational skills, with the ability to articulate audio-related feedback effectively.
  • Commitment to developing AI that masters sophisticated multilingual audio capabilities.
  • Preferred: Demonstration of exceptional attention to linguistic nuance, auditory detail, and data quality beyond standard transcription work.
  • Preferred: Deep understanding and taste of what good/useful Audio data is.
  • Preferred: Strong command of advanced transcription and annotation practices, including handling disfluencies, accents, and prosodic features (intonation, stress, rhythm, emotion, etc) with high consistency and accuracy.
  • Preferred: Background in linguistics (e.g., phonetics, phonology, sociolinguistics), speech sciences, cognitive science, or a related field, or equivalent practical experience, with demonstrated ability to analyze accent variation, pronunciation differences, and multilingual speech patterns.
  • Preferred: Experience working with speech/audio datasets, annotation workflows, or AI training data, including knowledge/experience with training voice models, and an understanding of how data quality impacts model performance.
  • Preferred: Experience with translation or localization workflows, particularly for software UI, product copy, or conversational AI text.
  • Preferred: Professional experience in voice work, including voice acting, voice recording, podcasting with a measurable audience (e.g., X following), or similar audio production demonstrating attention to clarity and recording quality.
  • Preferred: Demonstrated ability to exercise independent judgment in ambiguous audio scenarios and make consistent, defensible annotation decisions.
  • Preferred: Portfolio: Voice samples, annotated transcripts, or audio-related work demonstrating quality, methodology, and attention to detail.
  • Candidates with professional experience in voice, linguistics, speech data, or speech evaluation and research are especially encouraged to apply.
  • Tutor roles may be offered as full-time, part-time, or contractor positions, depending on role needs and candidate fit.
  • For contractor positions, hours will vary widely based on project scope and contractor availability, with no fixed commitments required.
  • On average, most projects may require at least 10 hours per week to deliver effectively, though this is not a fixed commitment and depends on the scope of work.
  • Contractors have full flexibility to set their own hours and determine the exact amount of time needed to complete deliverables.
  • Tutor roles may be performed remotely from any location worldwide, subject to legal eligibility, time-zone compatibility, and role-specific needs.
  • For US-based candidates, we are unable to hire in Wyoming and Illinois at this time.
  • We are unable to provide visa sponsorship.
  • For those who will be working from a personal device, your computer must be a Chromebook, a Mac with macOS 11.0 or later, or Windows 10 or later.

Benefits

  • Benefits vary based on employment type, location, and jurisdiction.
  • Benefits for eligible U.S.-based positions include health insurance, 401(k) plan, and paid sick leave.

American artificial intelligence and social media division of SpaceX

AIStartupx.ai

Details

Visa sponsorshipNo
Apply routeGreenhouse
Also posted in 1 other channel
$35–$45/hr