6 days ago
Remote, Argentina +2 moreSenior
Responsibilities
- Build realistic simulated developer environments containing codebases, infrastructure, tickets, documentation, conversations, and development history.
- Design challenging tasks from intermediate environment states, including prompts and objective success criteria.
- Write functional and integration tests that accept valid agent solutions and reject incorrect solutions.
- Review agent solutions, analyze failures, and iterate on tasks and tests based on QA feedback.
- Create fair and robust evaluations of AI coding agents without writing most of the implementation code yourself.
Requirements
- At least 5 years of software development experience.
- A master’s degree in Computer Science, Software Engineering, Data Science/Data Analytics, Artificial Intelligence/Machine Learning, Computational Linguistics/Natural Language Processing, Information Systems, or a related field.
- A bachelor’s degree is accepted with 5 years of experience in the field.
- At least 3 years of professional experience in related QA-automation/testing or cybersecurity roles.
- Experience writing functional and integration tests.
- Experience with Python and FastAPI, JavaScript/TypeScript and React, Docker, Postgres, Kafka, and Redis.
- English proficiency at B2 level or higher; applicants should submit a CV in English and indicate their English proficiency.
Benefits
- Part-time, remote, freelance, project-based work that can fit around other professional or academic commitments.
- Payment is per accepted task, with compensation up to the equivalent of $30 per hour depending on qualification tier and task efficiency.
- Opportunity to work on advanced AI projects, build portfolio experience, and influence how future AI models perform and communicate.
- Participation is project-based rather than permanent employment.
