M

Evaluation Scenario Writer - Ai Agent Testing Specialist

Mindrift

estado de méxico, estado de méxico, Mexico Full-time June 14, 2026

Found Description

Please submit your CV in English and indicate your level of English proficiency.Mindrift connects specialists with project-based AI opportunities for leading tech companies, focused on testing, evaluating, and improving AI systems.Participation isproject-based, not permanent employment.What This Opportunity InvolvesYou'll create challenging coding test cases that push AI coding systems to their limits:Review and refine realistic coding tasks based on provided production codebases with realistic scope, requirements and information sourcesWrite comprehensive functional tests that validate actual end-to-end behavior and edge-cases, not just superficial checksCraft fair but hard challenges where the AI has all the context it needs, but has to work for it (information scattered across files and external sources, complex reasoning required)Analyze AI failures to understand what the model struggles with vs. what it mastersIterate based on feedback from expert QA reviewers who score your wo...

Ready to Apply?

Submit your application for Evaluation Scenario Writer - Ai Agent Testing Specialist at Mindrift

Apply Now