6 days ago
Remote, France +2 moreSenior
Responsibilities
- Build realistic simulated developer environments containing codebases, infrastructure, tickets, documentation, and conversations.
- Design tasks from intermediate environment states, including prompts and definitions of successful solutions.
- Write functional and integration tests that accept valid agent solutions and reject incorrect ones.
- Review agent solutions, analyze failures, and iterate on tasks and tests based on QA feedback.
- Create fair and robust evaluations that reveal differences between strong and weak AI coding-agent performance.
Requirements
- At least 5 years of software development experience.
- At least 3 years of professional experience in related QA automation/testing or cybersecurity roles or domains.
- Master’s degree in Computer Science, Software Engineering, Data Science/Data Analytics, Artificial Intelligence/Machine Learning, Computational Linguistics/Natural Language Processing, Information Systems, or a related field.
- A bachelor’s degree is accepted with 5 years of experience in the field.
- Experience with Python, FastAPI, JavaScript/TypeScript, React, Docker, Postgres, Kafka, and Redis.
- Experience writing functional and integration tests.
- English proficiency at B2 level or higher.
- Submit a CV in English and indicate English proficiency.
Benefits
- Part-time, remote, freelance project-based work that can fit around other professional or academic commitments.
- Work on advanced AI projects and build portfolio experience.
- Opportunity to influence how future AI models understand and communicate in the candidate’s area of expertise.
- Paid per accepted task, with rates dependent on qualification tier and task efficiency.
- Participation is project-based and not permanent employment.
