Evaluation Scenario Writer - Ai Agent Testing Specialist

Mindrift · estado de méxico, estado de méxico, Mexico

Location
estado de méxico
Job Type
Full-time
Posted
June 05, 2026

Job Description

Please submit your CV in English and indicate your level of English proficiency.
Mindrift connects specialists with project-based AI opportunities for leading tech companies, focused on testing, evaluating, and improving AI systems.
Participation isproject-based, not permanent employment.
What This Opportunity Involves
You'll create challenging coding test cases that push AI coding systems to their limits:
Review and refine realistic coding tasks based on provided production codebases with realistic scope, requirements and information sources
Write comprehensive functional tests that validate actual end-to-end behavior and edge-cases, not just superficial checks
Craft fair but hard challenges where the AI has all the context it needs, but has to work for it (information scattered across files and external sources, complex reasoning required)
Analyze AI failures to understand what the model struggles with vs. what it masters
Iterate based on feedba...

Ready to Apply?

Submit your application for Evaluation Scenario Writer - Ai Agent Testing Specialist at Mindrift

Apply Now