
ML Ops Engineer (EMEA Remote)
Pragmatike5 months ago
Prague, Czechia +7 moreMid Level
Responsibilities
- Build and operate production-grade model-serving infrastructure using vLLM, TGI, Triton, or equivalent frameworks.
- Design and implement blue/green and canary deployment pipelines for ML models.
- Develop auto-scaling systems, multi-model serving architectures, and intelligent request-routing layers.
- Optimize GPU utilization, memory efficiency, network throughput, and model artifact storage performance.
- Design observability systems for inference latency, throughput, GPU usage, cost metrics, and system health.
- Manage model registries and automated, reproducible model deployment pipelines.
- Own the full ML systems lifecycle from development through production, including operational support and on-call responsibilities.
- Define engineering best practices and contribute to platform scalability.
Requirements
- At least four years of experience in MLOps, platform engineering, SRE, or similar infrastructure roles focused on ML systems.
- Hands-on experience with model-serving frameworks such as vLLM, TGI, Triton, or equivalent.
- Strong experience with container orchestration and operating GPU-based workloads in production.
- Experience with model registries, experiment tracking, and automated deployment pipelines.
- Proficiency in Python and infrastructure-as-code tools such as Terraform or Helm.
- Strong understanding of distributed systems, performance tuning, and production reliability engineering.
- Ability to use AI coding assistants to accelerate development and debugging workflows.
- Preferred experience with Kubeflow, MLflow, or KubeAI.
- Preferred knowledge of GPU scheduling, CUDA or ROCm optimization, and multi-tenant inference systems.
- Preferred experience optimizing costs across GPU types and inference workloads.
- Preferred background in early-stage startups, greenfield infrastructure, or building production systems from scratch.
- Fluent English is required.
Benefits
- Fully remote work in EMEA time zones
- ASAP start date
- Opportunity to build foundational ML inference infrastructure from the ground up
- Work at the intersection of distributed systems, GPU computing, and sustainable cloud architecture
Categories
About Pragmatike
Pragmatike is a Paris-based IT services and recruiting firm connecting remote-first companies with software engineers and tech specialists worldwide. Founded in 2022, it places contractors or full-time hires and staffs teams to complete projects for startups and scaleups. The private partnership offers access to a large network of 50,000+ specialists across 60+ countries.