2 months ago
Remote, PolandStaff+
Responsibilities
- Design, build, and maintain reliability components including HTTP rate limiters, database schema migration tools, circuit breakers, and distributed Redis-based caching.
- Troubleshoot complex production issues and optimize PostgreSQL usage for high-load distributed systems.
- Lead preliminary investigations during severe production incidents by identifying likely root causes, assessing impact, and proposing mitigations.
- Create scalable, reusable tools and frameworks that help engineering teams build resilient services.
- Use AI-powered development tools and coding agents to accelerate development, analyze architectures, and automate repetitive or error-prone work.
- Influence reliability practices through knowledge sharing, design reviews, and high technical standards.
- Lead technical initiatives, drive cross-team projects, and mentor other engineers as an individual contributor.
Requirements
- Strong expertise with Java/JVM and scalable, high-performance backend systems.
- Strong understanding of distributed systems, high availability, the CAP theorem, and fault tolerance.
- Deep experience with PostgreSQL and key-value or non-relational storage such as Redis.
- Hands-on experience with Docker and Kubernetes in containerized, cloud-native environments.
- Hands-on experience with message brokers such as RabbitMQ or Kafka.
- Ability to work independently with minimal supervision and validate technical decisions through critical thinking.
- Strong written and spoken English skills for international collaboration.
- Background in infrastructure engineering or Site Reliability Engineering and infrastructure-as-code practices is a standout qualification.
- Experience leading technical initiatives, driving cross-team projects, and mentoring engineers while remaining an individual contributor is preferred.
- Familiarity with Graylog, Zabbix, Grafana, and/or BigQuery is a standout qualification.
Benefits
- Hybrid work in Prague, Czech Republic or Nicosia, Cyprus.
- Employees near certain hubs are generally expected to collaborate in person around 2–3 days per week.
- Flexible work options include remote work, hybrid environments, and co-working spaces across global hubs.
Tech Stack
About Wrike
Wrike builds a cloud-based work management and project collaboration platform for teams and enterprises, offering tasks, Gantt charts, automation, and integrations. It sells subscriptions as a SaaS product and is used across functions like marketing, PMO, and IT to plan, track, and report work. Founded in 2006 and headquartered in San Diego, Wrike was acquired by Citrix in 2021.