
Staff Platform Development Engineer
Clearwater Analytics2 months ago
Base Salary
$148k - $201k/yr
Responsibilities
- Design, build, and maintain scalable, highly available infrastructure across on-premises, hybrid AWS, and co-location environments.
- Lead infrastructure automation using Harbor, Terraform, Ansible, Salt, or equivalent infrastructure-as-code tooling.
- Own Kubernetes cluster operations, workload scheduling, resource governance, and service mesh configuration.
- Architect and operate CI/CD pipelines and deployment automation with software engineering and release management teams.
- Build and maintain observability using metrics, log aggregation, and distributed tracing focused on financial platform reliability objectives.
- Support financial data connectivity including market data feeds, custodian ingestion pipelines, FIX, SWIFT, SFTP, and third-party vendor integrations.
- Provide Tier 2/3 incident escalation, lead root cause analysis, and resolve systemic reliability issues.
- Perform capacity planning, performance engineering, and cost optimization across compute, storage, and network layers.
- Partner with Security and Compliance on SOC 2 Type II, ISO 27001, and financial regulatory requirements.
- Evaluate technologies and vendors, make architecture recommendations, mentor engineers, and maintain documentation and runbooks.
Requirements
- 7+ years of experience in platform engineering, Linux systems administration, or DevOps/SRE roles, including at least 3 years in financial services or FinTech.
- Deep hands-on experience with RHEL and Debian-based Linux distributions, preferably Ubuntu, including systems-level diagnostics and command-line work.
- Production Kubernetes experience covering cluster operations, Helm, RBAC, networking, and storage; Istio or Linkerd experience is preferred.
- Infrastructure automation experience with Ansible, Salt, Terraform, or equivalent tools, including designing automation frameworks.
- Experience designing software delivery pipelines with GitLab CI, Jenkins, GitHub Actions, or Spinnaker.
- Hands-on observability experience with Prometheus, Grafana, Elastic Stack, ELK/EFK, CheckMk, or equivalent tooling.
- Strong knowledge of Java applications and JVM internals, including tuning, garbage collection analysis, and heap profiling.
- Familiarity with FIX, SWIFT, SFTP/FTPS, market data feeds, Bloomberg, Refinitiv, and custodian data integration patterns.
- Working knowledge of MySQL, PostgreSQL, and non-relational or time-series data stores.
- Strong networking fundamentals including TCP/IP, DNS, TLS/PKI, load balancing, VPN/MPLS, and firewall policy.
- Experience using AI-assisted workflows such as Claude for troubleshooting, documentation, infrastructure-as-code, and engineering productivity.
- Scripting and automation skills in Python, Bash, or equivalent, with experience managing infrastructure through APIs.
- Experience with SOC 2, PCI-DSS, or equivalent compliance frameworks and related audit, access-control, and change-management requirements.
- Preferred experience with AWS EC2, EKS, RDS, and S3; Proxmox or VMware vSphere; HashiCorp Vault, CyberArk, and Let's Encrypt or internal PKI.
- Exposure to SEC, FINRA, MiFID II, investment management, portfolio accounting, or wealth management infrastructure is preferred.
- RHCA, RHCE, CKA, CKAD, or equivalent certifications are a plus.
- Bachelor of Science in Computer Science, Engineering, or a related technical field, or equivalent practical experience.
Benefits
- Health, vision, and dental insurance.
- 401(k), paid time off, parental leave, and medical leave.
- Short-term and long-term disability insurance benefits.
- The position is based inside Revival Food Hall; geographic location may affect compensation.
Tech Stack
AnsibleAWSBashGitHub ActionsGitLab CI/CDGrafanaHelmIstioJavaJenkinsKubernetesLinuxMySQLPostgreSQLPrometheusPythonSpinnakerSwiftTerraform