1 day ago
Remote, BrazilSenior
Responsibilities
- Build complex, highly available, and cost-optimized cloud infrastructure on Azure and Kubernetes (AKS).
- Design, operate, upgrade, autoscale, and secure production AKS clusters, including private networking, ingress, certificates, and workload identity.
- Own Terraform infrastructure as code, including reusable modules, workspaces, state, and multi-environment deployments.
- Design CI/CD pipelines for container image publishing and progressive production promotion.
- Develop monitoring and observability across Kubernetes, applications, and Azure services.
- Troubleshoot complex issues across compute, networking, identity, and data layers.
- Implement security practices including vulnerability scanning, threat modeling, and centralized identity-based secrets management.
- Design and test multi-region high availability, disaster recovery, failover, and failback for Kubernetes workloads and managed databases.
- Partner with application engineering teams to establish operational standards.
Requirements
- Bachelor’s degree in Computer Science, Software Engineering, or a related field, or equivalent work experience.
- At least 5 years of experience in a DevOps engineering role with significant Azure experience.
- Proven experience designing and implementing large-scale Azure solutions.
- Production experience running AKS, including private clusters, upgrades, autoscaling, and incident response.
- Advanced production Terraform experience covering module design, state and workspace management, and multi-environment or multi-subscription deployments.
- Ability to solve complex technical problems and architect robust solutions.
- Strong communication and collaboration skills.
- Azure certifications such as Azure Solutions Architect Expert, Azure DevOps Engineer Expert, or Certified Kubernetes Administrator are valued.
- Knowledge of Azure services and architectures, including Entra ID, RBAC, managed identities, workload identity federation, Key Vault, and landing zones is desirable.
- Experience with Azure Pipelines, GitHub Actions, Argo CD, Flux, or similar pipeline and GitOps tools is desirable.
- Strong PowerShell, Bash, or Python scripting skills and proficiency with Go, C#/.NET, Node.js, or another programming language are desirable.
- Experience with Docker, Kubernetes, Helm, container registries, configuration management tools, networking, Azure managed data services, observability platforms, disaster recovery, cloud security, penetration testing, and compliance standards is desirable.
- Exposure to Google Cloud, GKE Autopilot, cross-cloud networking and identity, or GCP migration work is a plus.
Benefits
- 100% remote work mode.
- Company-paid medical insurance, mental health programs, and five undocumented sick-leave days per year.
- Internal meetups, conferences, workshops, Udemy access, language courses, and company-paid certifications.
- 20 working days of paid vacation plus local bank holidays.
- Long-term employment with internal mobility opportunities.
- International projects, collaborative teams, and regular team-building events.
Tech Stack
AnsibleArgo CDAzureBashC#ChefDatadogDockerGitHub ActionsGoGoogle CloudGoogle Cloud PlatformGrafanaHelmKubernetes.NETNode.jsPowerShellPuppetPythonRedisTerraform
Categories
About Ciklum
Ciklum is an IT services firm that builds custom digital products and provides nearshore engineering teams, QA, DevOps, data analytics, and UX for enterprises and high-growth startups. It operates a services model spanning product engineering, e-commerce solutions, and managed delivery centers. Founded in 2002 and headquartered in London, the privately held company serves global clients including Just Eat and Zurich Insurance.
