AI Response Labeler – Italian Specialty

Posted 14hrs ago

Employment Information

Education
Salary
Experience
Job Type

Report this job

Job expired or something wrong with this job?

Job Description

AI response annotator comparing and evaluating AI-generated content in English and Italian. Applying Italian cultural expertise and structured quality standards for Blueprint’s technology solutions.

Responsibilities:

  • Perform side-by-side comparisons of AI-generated responses and determine which response is stronger
  • Evaluate responses for factual accuracy, relevance, completeness, clarity, reasoning, instruction-following, tone, and overall quality
  • Assess content written in English, Italian, or a combination of both
  • Evaluate general-purpose questions and answers, web-search results, file-based tasks, image-based responses, content-generation requests, and single-turn and multi-turn conversations
  • Apply Italian expertise to language, terminology, tone, regional conventions, idioms, and cultural context specific to Italy
  • Identify unsupported claims, incomplete reasoning, missed instructions, unnatural phrasing, cultural inaccuracies, and differences in usefulness
  • Apply detailed, scenario-specific annotation guidelines accurately and consistently
  • Make independent evaluation decisions for ambiguous cases
  • Document decisions and provide concise, evidence-based rationale when required
  • Complete evaluations within time and productivity expectations, generally at least 25 tasks per day
  • Participate in training, guided practice, calibration sessions, qualification reviews, and ongoing quality-review activities
  • Incorporate feedback and adjust evaluation decisions to team and client quality standards

Requirements:

  • Native-level or professional fluency in Italian
  • Deep familiarity with Italian linguistic conventions, regional vocabulary, idioms, tone, and cultural context specific to Italy
  • Strong English fluency and reading comprehension
  • Strong general analytical and critical-thinking skills
  • Ability to evaluate content across varied topics, formats, and task types
  • Ability to assess factuality, relevance, reasoning, clarity, instruction-following, cultural appropriateness, and overall usefulness
  • Ability to recognize subtle differences in meaning, quality, tone, and user intent
  • Sound judgment when applying structured evaluation criteria to ambiguous or unfamiliar scenarios
  • Strong written communication skills
  • Excellent attention to detail and ability to maintain accuracy within established time expectations
  • Ability to learn and consistently apply detailed evaluation frameworks
  • Ability to work independently while remaining aligned with shared quality standards
  • Comfort performing repetitive, detail-oriented work for extended periods while maintaining focus, accuracy, and consistent judgment
  • Ability to receive feedback, recalibrate decisions, and adapt as evaluation guidelines evolve
  • Approximately 30-day onboarding and qualification program must be successfully completed before production work
  • Ability to work 9:00 a.m. to 5:00 p.m. Pacific Time during the training and qualification period
  • Preferred: experience with side-by-side labeling, annotation, comparative content evaluation, quality assessment, AI-generated response evaluation, model-quality assessment, data labeling, search relevance, content quality, factual accuracy, user-facing digital experiences, detailed guidelines, rubrics, or structured decision-making frameworks

Benefits:

  • Medical, dental, and vision coverage
  • Flexible Spending Account (FSA)
  • 401(k) retirement plan
  • Competitive paid time off
  • Parental leave
  • Professional growth and development opportunities
  • Benefits in accordance with local requirements and the terms of employment through an Employer of Record partner