AI Evaluation Specialists

MINDTEL
Noida, Uttar Pradesh, India

AI Evaluation Specialists – Multiple Positions

Department: AI/LLM Operations

Employment: 3-Month Contract

Location: Noida Sector 62

Experience: 2+ Years, depending on the profile

We are hiring AI Evaluation Specialists across multiple domains to evaluate, benchmark, and improve next-generation AI/LLM models. Candidates will assess AI-generated outputs for accuracy, quality, reasoning, language, technical correctness, and adherence to defined evaluation guidelines.

1. Technical AI Evaluation Specialist – Code & Software Engineering

  • Experience: 2+ years Software Development or strong CS background.
  • Evaluate AI-generated code for correctness, performance, security, and best practices.
  • Review text-to-code, image-to-code, UI, and software architecture outputs.
  • Perform code reviews, SxS comparisons, debugging, and edge-case analysis.
  • Skills: Python, JavaScript/TypeScript, Java, C++, Go, React/HTML/CSS, testing, Git.
  • Strong technical writing and analytical skills required.

2. Speech & Audio AI Evaluation Specialist – International Voice

  • Experience: 2+ years international voice experience in a GCC/Captive/In-house setup.
  • Evaluate AI-generated Speech-to-Speech (S2S) and Text-to-Speech (TTS) outputs.
  • Assess pronunciation, tone, emotion, cadence, naturalness, and conversational quality.
  • Review accents, dialects, lip-sync, and audio/video synchronization.
  • Mandatory: C2 / near-native English proficiency.
  • Strong auditory attention and communication skills required.

3. Senior AI Evaluation & RLHF Specialist

  • Experience: 2+ years in AI data annotation, model evaluation, or RLHF.
  • Evaluate and rank AI responses using complex evaluation rubrics.
  • Perform SxS comparisons covering factuality, reasoning, helpfulness, and instruction following.
  • Identify model failures and provide structured corrective feedback.
  • Create golden responses and edge-case examples for model training.
  • Strong analytical ability and excellent written English required.

4. ASR & Transcription QA Specialist

  • Experience: 2–3+ years in transcription, captioning, subtitling, or related services.
  • Review and correct AI-generated speech-to-text outputs.
  • Identify transcription errors, speaker changes, overlapping speech, and audio inconsistencies.
  • Follow verbatim and clean-read transcription standards.
  • Mandatory: 65+ WPM typing speed with 99% accuracy.
  • Strong English listening and written comprehension required.

5. Spatial Reasoning Evaluation Specialist – Video RL

  • Background: STEM degree with exposure to Computer Vision, Video ML, RL, robotics, or simulation.
  • Evaluate AI-generated videos and autonomous-agent trajectories frame by frame.
  • Assess temporal sequencing, physical plausibility, action accuracy, and goal completion.
  • Review segmentation, bounding boxes, motion, and visual artifacts.
  • Identify visual hallucinations, physics errors, and spatial inconsistencies.
  • Python/scripting knowledge and strong analytical reasoning are preferred.

Common Requirements

  • Strong attention to detail and analytical reasoning.
  • Ability to follow detailed evaluation guidelines and maintain high-quality output.
  • Excellent written and/or spoken English depending on the profile.
  • Comfortable working on repetitive, high-accuracy evaluation tasks.
  • Ability to provide clear, objective, evidence-based feedback.
  • Contract Duration: 3 Months.

Score my resume against this job, free →

Get your ATS score for this role — free. Score my resume free →