20 hours ago
Base Salary
$160k - $240k/yr
Responsibilities
- Design, build, and operate highly available, scalable, and resilient cloud infrastructure for data ingestion and analytics platforms.
- Define and monitor SLIs and SLOs, use error budgets and operational metrics, and drive reliability improvements.
- Improve observability through logging, metrics, tracing, and alerting.
- Automate infrastructure provisioning and system management with Infrastructure as Code.
- Lead incident response, perform root cause analysis, and implement post-incident improvements.
- Optimize the performance, reliability, and cost efficiency of cloud-based data systems.
- Ensure the reliability of batch and streaming pipelines, storage systems, and reporting infrastructure.
- Partner with data engineers, software engineers, and stakeholders to improve system reliability and operational maturity.
- Strengthen platform security through monitoring, vulnerability management, and cloud security practices.
- Continuously improve CI/CD pipelines and deployment processes for data infrastructure.
Requirements
- At least 5 years of experience in Site Reliability Engineering, DevOps, or cloud infrastructure roles.
- Strong proficiency in at least one programming or scripting language, including Python and/or Go.
- Experience supporting production systems with a focus on reliability, scalability, and observability.
- Hands-on experience operating or designing highly available distributed systems.
- Bachelor’s degree in Computer Science, Engineering, Mathematics, or a related field, or equivalent professional experience.
- Experience with large-scale data platforms, data pipelines, or analytics infrastructure is preferred.
- Strong AWS production operations experience, including relevant AWS data services, is preferred.
- Experience with SLIs, SLOs, error budgets, monitoring, observability, incident management, and postmortems is preferred.
- Experience with Terraform or CloudFormation, CI/CD pipelines, Docker, Kubernetes, and cloud networking is preferred.
- Knowledge of Databricks or Snowflake, cloud cost optimization, compliance, and data governance is preferred.
- AWS Associate-level certification or above is preferred.
Benefits
- Benefits may include merit increases, incentive compensation for exempt roles, paid holidays, paid time off, medical, dental, vision, short- and long-term disability, 401(k) match, life insurance, and wellness programs.
- The posting states that the salary range is $160,000–$240,000 USD annually, plus benefits and bonus.
- The company does not provide benefits directly to contingent workers, contractors, or interns.
Tech Stack
Categories
Site Reliability
About Bloomberg
Bloomberg is a global leader in business and financial information, delivering trusted data, news, and insights that bring transparency and efficiency, and fairness to markets. We help connect influential communities across the global financial ecosystem via reliable technology solutions that enable our customers to make more informed decisions and foster better collaboration. We challenge the status quo through constant innovation. We collaborate broadly because we know that other perspectives matter. We put our customers first, as a guiding beacon. And we believe doing the right thing – by our people, our clients, and our communities – is the best thing for our business.
