Sr. Resource Plane Platform Engineer
Intermedia Intelligent Communications3 months ago
Remote, PortugalSenior
Responsibilities
- Own the design, implementation, and operation of shared Resource Plane services used across the organization.
- Operate and evolve Kafka, RabbitMQ, Redis, and Elasticsearch platforms in production.
- Apply reliability engineering practices to improve availability, scalability, reliability, and safe upgrades of stateful systems.
- Build and maintain Infrastructure-as-Code and GitOps automation for provisioning and lifecycle management.
- Define supported usage patterns, guardrails, and golden paths for platform consumers.
- Integrate Resource Plane services into the Internal Developer Platform for self-service consumption.
- Participate in an on-call rotation that may initially include a 24x7 weekly rotation.
- Lead incident response, root cause analysis, and reliability improvements for owned services.
- Collaborate with the Kubernetes Platform Team, Internal Developer Platform teams, and other teams.
- Contribute as a senior technical voice in architecture discussions and platform roadmap planning.
Requirements
- Bachelor’s degree in Computer Science, Software Engineering, or Information Technology.
- Senior-level experience operating production-grade, stateful distributed systems.
- Strong hands-on experience with Kafka, RabbitMQ, Redis, and Elasticsearch.
- Proven experience running infrastructure primarily in on-premises environments.
- Strong understanding of Linux systems, networking, and storage fundamentals.
- Deep experience with Infrastructure-as-Code, preferably Terraform.
- Experience with GitOps workflows and declarative infrastructure management.
- Understanding of reliability engineering concepts including SLOs, error budgets, alerting, and capacity planning.
- Experience with Kubernetes-adjacent platforms and services.
- Ability to operate independently and take ownership in a small team environment.
- Preferred experience designing shared platform services for multiple product teams.
- Preferred familiarity with Internal Developer Platform concepts and developer self-service platforms.
- Preferred experience migrating on-premises workloads toward cloud-native or hybrid models.
- Preferred exposure to security, compliance, and governance requirements for shared infrastructure.
- Prior staff- or principal-level technical leadership experience is preferred.
Benefits
- Primarily remote work with occasional visits to the Coimbra office.
- The company plans to open offices in Aveiro and Porto in the future.
- Participation in an on-call rotation, initially potentially including a 24x7 weekly schedule with a target of business-hours primary coverage.