1 day ago
Remote, Poland +4 moreSenior
Responsibilities
- Build realistic virtual company environments containing codebases, infrastructure, tickets, documentation, conversations, and development history.
- Create challenging coding-agent tasks from intermediate environment states, including prompts and solvability criteria.
- Write functional and integration tests that verify agent solutions while accepting all valid approaches.
- Review agent solutions, analyze failures, and iterate on tasks and tests using QA feedback.
- Define fair and robust evaluation criteria for AI coding agents.
Requirements
- At least five years of software-development experience.
- Experience with Python and FastAPI.
- Experience with JavaScript or TypeScript and React.
- Experience with Docker, Postgres, Kafka, and Redis.
- Experience writing functional and integration tests.
- English proficiency at B2 level or higher.
- A master's degree in computer science, software engineering, data science/data analytics, artificial intelligence/machine learning, computational linguistics/natural language processing, information systems, or a related field.
- A bachelor's degree is accepted with five years of experience in the field.
- At least three years of professional experience in related roles or domains, specifically for QA-automation/testing or cybersecurity roles.
- CV submitted in English and English proficiency level indicated.
Benefits
- Part-time, remote, freelance project-based work.
- Work that can fit around primary professional or academic commitments.
- Exposure to advanced AI projects and portfolio-building experience.
- Opportunity to influence how future AI models understand and communicate in the candidate's field.
- Paid per accepted task, with an effective rate up to the equivalent of $40 per hour depending on qualification tier and efficiency.
- Participation is project-based and not permanent employment.
Tech Stack
Categories
About Mindrift
Mindrift builds an expert-sourcing platform that connects domain specialists to project-based work training and evaluating generative AI models, including supervised fine-tuning, RLHF, evaluation, and red-teaming. It is built and operated by Toloka, part of Nebius Group, and run from Amsterdam, Netherlands. Work is fully remote and freelance, serving global technology companies developing and improving large AI systems.
