Responsibilities
- Partner with engineering teams to architect scalable and reliable use of infrastructure platforms.
- Debug reliability and scalability issues across application and infrastructure layers.
- Build monitoring and alerting focused on system symptoms and maintain enterprise-grade SLAs.
- Develop infrastructure as code using Chef, Terraform, and Kubernetes.
- Create deployment pipelines and centralized tooling, services, and automation frameworks using Docker and Kubernetes.
- Support capacity management, operational efficiency, and engineering productivity through internal platform improvements.
- Participate in PagerDuty on-call rotations, respond to availability incidents, and support other engineers.
- Conduct incident retrospectives and implement automation and system improvements to prevent recurrence.
Requirements
- 3+ years of experience as a Software, DevOps, or Site Reliability Engineer.
- Strong systems-thinking skills covering interfaces, boundaries, edge cases, failure modes, and implementations.
- Ability to collaborate across global remote teams, document work, deliver quickly, and manage multiple tasks and expectations.
- Working knowledge of Linux and Unix Shell.
- Strong programming skills, with Ruby and/or Go preferred.
- Experience with Docker, Kubernetes, Terraform, or similar infrastructure-as-code technologies.
- Experience with MongoDB, Redis, Kafka, Postgres, or similar data technologies.
- Motivation to automate repetitive work and improve software engineers’ day-to-day experience.
Benefits
- Competitive compensation that may include equity.
- Retirement and Employee Stock Purchase Plans.
- Flexible paid time off.
- Medical, dental, vision, life, and disability benefit plans.
- Fertility benefits and equal paid parental leave.
- Professional development through formal career pathing, learning platforms, and a yearly learning stipend.
- Hybrid ways of working and a curated in-office employee experience.
- Volunteer Week, donation matching, and Employee Resource Groups.
- Collaborative culture recognized as a Great Place to Work®.
- The role is marked as hybrid, with benefits varying by location.
Tech Stack
Categories
About Braze
Braze is the leading customer engagement platform that empowers brands to Be Absolutely Engaging.™ Braze allows any marketer to collect and take action on any amount of data from any source, so they can creatively engage with customers in real time, across channels from one platform. From cross-channel messaging and journey orchestration to Al-powered experimentation and optimization, Braze enables companies to build and maintain absolutely engaging relationships with their customers that foster growth and loyalty. The company has been recognized as a 2024 U.S. News & World Report Best Companies to Work For, 2024 Best Small & Medium Workplaces in Europe by Great Place to Work®, 2024 Fortune Best Workplaces for Women™ by Great Place to Work® and was named a Leader by Gartner® in the 2024 Magic Quadrant™ for Multichannel Marketing Hubs and a Strong Performer in The Forrester Wave™: Email Marketing Service Providers, Q3 2024. Braze is headquartered in New York with 15 offices across AMER, LATAM, EMEA, and APAC. Learn more at braze.com.
