Skip to main content
Blueprint

AI Response Labeler / Annotator – Castilian Spanish Specialty

RemoteMexico only· UTC-8…UTC-7
Published
Employment
Full-time
MXN 255–MXN 290/hr
Check eligibility

Open to MX only · UTC-8…UTC-7. Set where you work from to check your eligibility.

No BS summary

AI annotation/evaluation role for someone with native-level or professional Castilian Spanish, strong English comprehension, and strong analytical judgment. Must be in Mexico and work 9:00 a.m. to 5:00 p.m. Pacific Time during the approximately 30-day training period.

Required languages

Spanish Native-level or professional fluency; Castilian Spanish / Spain expertise required.English Strong fluency and reading comprehension.

What you'll do

  • Perform side-by-side comparisons of AI-generated responses and determine which response is stronger.
  • Evaluate responses for factual accuracy, relevance, completeness, clarity, reasoning, instruction-following, tone, and overall quality.
  • Assess content written in English, Castilian Spanish, or a combination of both, depending on the assigned scenario.
  • Evaluate a broad range of content, including general-purpose questions and answers, web-search results, file-based tasks, image-based responses, content-generation requests, and single-turn and multi-turn conversations.
  • Apply Castilian Spanish expertise when evaluating language, terminology, tone, regional conventions, idioms, and cultural context specific to Spain.
  • Evaluate the complete quality of a response rather than focusing only on grammar, translation, or language fluency.
  • Identify subtle but meaningful differences between responses, including unsupported claims, incomplete reasoning, missed instructions, unnatural phrasing, cultural inaccuracies, and differences in usefulness.
  • Apply detailed, scenario-specific annotation guidelines accurately and consistently.
  • Make independent evaluation decisions when examples or guidelines don’t provide an obvious answer.
  • Document decisions clearly and provide concise, evidence-based rationale when required.
  • Complete evaluations within established time and productivity expectations without sacrificing accuracy.
  • Maintain consistent judgment across a high volume of varied assignments.
  • Participate in training, guided practice, calibration sessions, qualification reviews, and ongoing quality-review activities.
  • Incorporate feedback and adjust evaluation decisions to remain aligned with team and client quality standards.
  • Complete similar evaluation tasks throughout the workday.
  • Complete a minimum of 25 tasks per day.
  • Meet established daily expectations while carefully applying annotation guidelines and providing accurate, well-supported evaluation decisions.

What they require

  • Native-level or professional fluency in Castilian Spanish.
  • Deep familiarity with the linguistic conventions, regional vocabulary, idioms, tone, and cultural context of Spanish as used in Spain.
  • Strong English fluency and reading comprehension, including the ability to understand complex prompts, AI-generated responses, and detailed annotation guidelines written in English.
  • Strong general analytical and critical-thinking skills that extend beyond language evaluation.
  • Ability to evaluate content across varied topics, formats, and task types.
  • Ability to assess factuality, relevance, reasoning, clarity, instruction-following, cultural appropriateness, and overall usefulness.
  • Ability to recognize subtle differences in meaning, quality, tone, and user intent.
  • Sound judgment when applying structured evaluation criteria to ambiguous or unfamiliar scenarios.
  • Strong written communication skills and the ability to explain evaluation decisions clearly and concisely.
  • Excellent attention to detail and the ability to maintain accuracy while working within established time expectations.
  • Ability to learn and consistently apply detailed evaluation frameworks.
  • Ability to work independently while remaining aligned with shared quality standards.
  • Comfort performing repetitive, detail-oriented work for extended periods while maintaining focus, accuracy and consistent judgement.
  • Ability to receive feedback, recalibrate decisions, and adapt as evaluation guidelines evolve.
  • Candidates should be comfortable maintaining focus, accuracy, and consistent judgment while reviewing a high volume of AI-generated content.
  • Most evaluation tasks are expected to take approximately 15 minutes.
  • Success in this role requires balancing productivity with quality.
  • All new hires must successfully complete a structured onboarding and qualification program before beginning production work.
  • Employees must also demonstrate the ability to evaluate broader response quality, follow detailed annotation guidelines, explain their decisions, and complete work within the expected timeframe.
  • During the approximately 30-day training and qualification period, employees must work from 9:00 a.m. to 5:00 p.m. Pacific Time.
  • After successfully completing training, employees may work standard business hours within their local time zone.
  • This role will be hired through an Employer of Record partner to support compliance with local employment, payroll, and benefits requirements.
  • Preferred: Experience performing side-by-side labeling, annotation, comparative content evaluation, or quality assessment.
  • Preferred: Experience evaluating AI-generated responses or contributing to model-quality assessment.
  • Preferred: Experience with data labeling or annotation.
  • Preferred: Experience evaluating search relevance, content quality, factual accuracy, or user-facing digital experiences.
  • Preferred: Experience working with detailed guidelines, rubrics, or structured decision-making frameworks.

Benefits

  • Medical, dental, and vision coverage
  • Flexible Spending Account (FSA)
  • 401(k) retirement plan
  • Competitive paid time off
  • Parental leave
  • Professional growth and development opportunities
  • Eligible employees will receive benefits in accordance with local requirements and the terms of their employment.
  • Employees will continue to receive feedback, quality reviews, and calibration support after entering production.

Blueprint is a technology solutions firm headquartered in Bellevue, Washington, with teams across the United States. We help organizations turn complex challenges into meaningful outcomes by connecting strategy and execution across AI, cloud, data, product development, and emerging technology.

Technology ConsultingStartup

What people say about this company

3.9/ 5

  • Employees appreciate the supportive work environment.
  • There are good opportunities for professional development.
  • Some employees have expressed concerns about management effectiveness.

Details

Apply routeGreenhouse
MXN 255–MXN 290/hr