Lio

Site Reliability Engineer / SRE (all genders)

Lio
Apply
1 month ago
Munich, GermanySenior

Responsibilities

  • Work with the CTO on the architecture and scaling of Lio’s infrastructure.
  • Design and operate reliable cloud infrastructure for production AI workloads.
  • Build multi-region deployments and high-availability architectures.
  • Optimize databases for agentic AI workloads, including vector search, hybrid RAG, MCP, and high-performance query execution.
  • Improve backend performance, latency, throughput, and resource efficiency.
  • Build observability, SLOs, alerting, and incident response processes.
  • Automate deployments, infrastructure, and developer workflows.
  • Partner with product engineering teams to build systems that scale with company growth.

Requirements

  • Experience operating production workloads on a major cloud platform.
  • Strong Python skills and experience optimizing backend services.
  • Solid understanding of distributed systems, asynchronous processing, and scalable architectures.
  • Experience with monitoring, observability, and incident management.
  • Experience optimizing databases at scale.
  • Familiarity with CI/CD pipelines, preferably GitHub Actions.
  • Passion for automation, infrastructure, and solving complex scaling challenges.
  • MongoDB experience is a plus.

Benefits

  • Work directly with the CTO on company-defining infrastructure decisions.
  • Meaningful equity and competitive compensation.
  • Collaborate with exceptional teammates in a fast-growing enterprise AI company.
  • 100% on-site work in the Munich office.

Tech Stack

GitHub ActionsMongoDBPython

Categories

Site Reliability
Lio

About Lio

51-200 employees
Contact me