Pragmatike

ML Ops Engineer (EMEA Remote)

Pragmatike
Apply
5 months ago
Prague, Czechia +7 moreMid Level

Responsibilities

  • Build and operate production-grade model-serving infrastructure using vLLM, TGI, Triton, or equivalent frameworks.
  • Design and implement blue/green and canary deployment pipelines for ML models.
  • Develop auto-scaling systems, multi-model serving architectures, and intelligent request-routing layers.
  • Optimize GPU utilization, memory efficiency, network throughput, and model artifact storage performance.
  • Design observability systems for inference latency, throughput, GPU usage, cost metrics, and system health.
  • Manage model registries and automated, reproducible model deployment pipelines.
  • Own the full ML systems lifecycle from development through production, including operational support and on-call responsibilities.
  • Define engineering best practices and contribute to platform scalability.

Requirements

  • At least four years of experience in MLOps, platform engineering, SRE, or similar infrastructure roles focused on ML systems.
  • Hands-on experience with model-serving frameworks such as vLLM, TGI, Triton, or equivalent.
  • Strong experience with container orchestration and operating GPU-based workloads in production.
  • Experience with model registries, experiment tracking, and automated deployment pipelines.
  • Proficiency in Python and infrastructure-as-code tools such as Terraform or Helm.
  • Strong understanding of distributed systems, performance tuning, and production reliability engineering.
  • Ability to use AI coding assistants to accelerate development and debugging workflows.
  • Preferred experience with Kubeflow, MLflow, or KubeAI.
  • Preferred knowledge of GPU scheduling, CUDA or ROCm optimization, and multi-tenant inference systems.
  • Preferred experience optimizing costs across GPU types and inference workloads.
  • Preferred background in early-stage startups, greenfield infrastructure, or building production systems from scratch.
  • Fluent English is required.

Benefits

  • Fully remote work in EMEA time zones
  • ASAP start date
  • Opportunity to build foundational ML inference infrastructure from the ground up
  • Work at the intersection of distributed systems, GPU computing, and sustainable cloud architecture

Tech Stack

HelmMLflowPythonTerraform
Pragmatike

About Pragmatike

1-10 employees

Pragmatike is a Paris-based IT services and recruiting firm connecting remote-first companies with software engineers and tech specialists worldwide. Founded in 2022, it places contractors or full-time hires and staffs teams to complete projects for startups and scaleups. The private partnership offers access to a large network of 50,000+ specialists across 60+ countries.

Contact me