04 ago
|
Mindrift
|
Madrid
ppbPlease submit your CV in English and indicate your level of English proficiency. /b /p pMindrift connects specialists with project-based AI opportunities for leading tech companies, focused on testing, evaluating, and improving AI systems. /p pbParticipation isproject-based, not permanent employment. /b /p h3What This Opportunity Involves /h3 pYou’ll create challenging coding test cases that push AI coding systems to their limits: /p ul liReview and refine realistic coding tasks based on provided production codebases with realistic scope, requirements and information sources /li liWrite comprehensive functional tests that validate vigente end-to-end behavior and edge-cases, not just superficial checks /li liCraft "fair but hard" challenges where the AI has all the context it needs, but has to work for it (information scattered across files and external sources, complex reasoning required) /li liAnalyze AI failures to understand what the model struggles with vs. what it masters /li liIterate based on feedback from expert QA reviewers who score your work on 7 quality criteria /li /ul h3What We Look For /h3 pThis opportunity is a good fit for experienced developers, software engineers, and/or test automation specialists open to part-time, non-permanent projects. Ideally, contributors will have: /p ul liDegree in Computer Science, Software Engineering or related fields /li li5+ years in software development,
primarily Python (pytest, async/await, subprocess, file operations) /li liBackground in Full-Stack development, with an equal focus on building React-based interfaces and robust Back-end systems /li liExperience writing tests (functional, integration - not just running them) /li liDocker containers (running evaluations locally in containers) /li liCI/CD understanding (GitHub Actions as a user: triggers, labels, reading results) /li liEnglish proficiency - B2 /li /ul h3How It Works /h3 pApply → Pass qualification(s) → Join a project → Complete tasks → Get paid /p h3Effort estimate /h3 pTasks for this project are estimated to take 20 hours to complete, depending on complexity. This is an estimate and not a schedule requirement; you choose when and how to work. Tasks must be submitted by the deadline and meet the listed acceptance criteria to be accepted. /p h3Payment /h3 ul liPaid contributions, with rates up to $21/hour* /li liFixed project rate or individual rates, depending on the project /li liSome projects include incentive payments /li liNote: Rates vary based on expertise, skills assessment, location, project needs, and other factors. Higher rates may be offered to highly specialized experts. Lower rates may apply during onboarding or non-core project phases. Payment details are shared per project /li /ul /p #J-18808-Ljbffr
📌 Evaluation Scenario Writer - AI Agent Testing Specialist (Madrid)
🏢 Mindrift
📍 Madrid