Sparta

Senior Site Reliability Engineer (SRE)

Sparta
Apply
23 days ago
Barcelona, Spain +2 moreSenior

Responsibilities

  • Design, build, and maintain backend services for real-time and analytical data processing.
  • Optimize services and pipelines for low latency, high throughput, and scale.
  • Own features end to end, including design, implementation, and production operation.
  • Own the operational health, observability, monitoring, alerting, and resilience of services.
  • Participate in incident response, root-cause analysis, and reliability improvements.
  • Operate workloads across Lambda, ECS, EKS, and Kubernetes.
  • Help operate Kafka, Flink, Redis/Valkey, RDS, and Redshift infrastructure.
  • Extend infrastructure as code with AWS CDK and improve CI/CD pipelines.
  • Automate manual operational work and reduce toil.

Requirements

  • At least 4 years of experience as a software or reliability engineer supporting production systems.
  • Strong proficiency in at least one of Kotlin, Java, Python, or TypeScript.
  • Working knowledge of AWS compute, networking, storage, and IAM from production operations.
  • Hands-on infrastructure-as-code experience with AWS CDK, Terraform, CloudFormation, or similar tools.
  • Practical experience deploying, debugging, and tuning container workloads on Kubernetes.
  • Experience with CI/CD tooling, observability through metrics, logging, and tracing, and production incident response.
  • Experience defining or working with service-level objectives and preferably error budgets.
  • Comfort with agent-based development tools such as Claude Code or an equivalent.
  • Clear communication, pragmatic problem-solving, ownership, and comfort working autonomously.
  • Preferred experience includes Kubernetes platform operations, EKS, Datadog, Kafka, Flink, Redshift, clustered Postgres, cloud security and compliance, distributed systems, and commodities, energy, or financial markets.

Benefits

  • Hybrid working style at a remote-first company, typically involving a couple of office days per week with flexibility.
  • High ownership, autonomy, growth opportunity, and direct impact in a scaling Series B company.

Tech Stack

Categories

BackendSite Reliability
Sparta

About Sparta

201-500 employees
Contact me