Castilian Spanish Labeler / Annotator
Compensation
Salary undisclosedDescription
About the Role
We’re looking for a Castilian Spanish Labeler / Annotator to evaluate responses generated by AI systems. In this role, you’ll perform side-by-side evaluations of outputs from different AI models across a variety of real-world scenarios. Your analysis will help improve how AI systems understand and communicate in Castilian Spanish.
This is not a translation role. It is an evaluation and analysis role that requires strong judgment, cultural awareness, and careful attention to detail.
What You'll Do
- Perform side-by-side comparisons of AI-generated responses.
- Evaluate outputs for accuracy, relevance, clarity, instruction-following, and overall quality.
- Assess AI responses across general-purpose questions and answers, web search results, file-based and image-based responses, content-generation tasks, and single-turn and multi-turn conversations.
- Identify nuances in language, tone, meaning, and cultural context specific to Castilian Spanish.
- Apply detailed, scenario-specific annotation guidelines accurately and consistently.
- Document evaluation decisions clearly and provide rationale when required.
- Maintain consistent quality and judgment across a high volume of evaluations.
- Participate in training, calibration, and ongoing quality-review activities.
What You'll Bring
- Native-level or professional fluency in Castilian Spanish.
- Deep familiarity with the linguistic conventions, tone, idioms, and cultural context of Spanish as used in Spain.
- Strong English reading comprehension, including the ability to understand detailed guidelines written in English.
- Prior experience with side-by-side labeling, annotation, or comparative content evaluation.
- Excellent analytical thinking and attention to detail.
- Ability to identify subtle differences in quality, meaning, tone, and instruction-following.
- Ability to learn and consistently apply structured evaluation frameworks.
- Strong written communication and the ability to explain evaluation decisions clearly.
- Ability to work independently while maintaining alignment with team quality standards.
Preferred Qualifications
- Experience with AI response evaluation or model-quality assessment.
- Experience with data labeling or annotation.
- Experience evaluating search relevance, content quality, or user-facing digital experiences.
Training and Qualification
All new hires will complete a structured onboarding and qualification program that includes:
- Training sessions and guided practice exercises.
- Calibration against established quality benchmarks.
- A qualification review before moving into production work.
- Ongoing feedback and calibration to support consistency and quality.
Compensation
At Blueprint, we strive to offer competitive pay that reflects the value of our team members. Compensation for this role is influenced by a variety of factors, including skills, education, responsibilities, experience, and geographic market.
The anticipated compensation range is $38.46 to $40.87 USD per hour, with a midpoint of $39.66 USD per hour. Please note that we typically do not hire new employees at the top of the posted range. Actual starting pay will be determined based on experience, skills, and internal equity. The final compensation and job title may vary depending on the selected candidate’s qualifications.
Location and Work Arrangement
Remote within the United States.
Candidates must be authorized to work in the United States and reside within a U.S. time zone.
During the approximately 30-day training and qualification period, employees must work from 9:00 a.m. to 5:00 p.m. Pacific Time. After successfully completing training, employees may work standard business hours within their local time zone.
- Posted
- Aug 5, 2026
- Last seen
- Aug 5, 2026
- First seen
- Aug 5, 2026