5 months ago
Base Salary
$120k - $250k/yr
Responsibilities
- Build and own the foundational infrastructure supporting LiveKit products.
- Modify the Go codebase to implement SRE objectives.
- Measure system performance and reliability and use data to prioritize projects.
- Participate in on-call and lead incident management for complex situations.
- Automate and manage configuration across globally distributed clusters running multiple products.
- Work with infrastructure vendors to address real-time performance and reliability issues.
Requirements
- Balance software engineering and large-scale system administration strengths.
- Experience managing complex multi-region distributed systems running on container orchestration systems such as Kubernetes.
- Ability to build maintainable, reliable infrastructure while meeting launch deadlines.
- Strong problem-solving, communication, learning, and delivery skills.
- Incident management training or Incident Commander experience is preferred.
- Linux networking, overlay networks, Kubernetes CNIs, and low-level troubleshooting of latency-sensitive workloads are preferred.
Benefits
- Health, dental, and vision benefits.
- Flexible vacation.
- Remote work environment with necessary equipment provided.
- Competitive salary and equity package.
Tech Stack
Categories
DevOpsSite Reliability
About LiveKit
LiveKit offers open source frameworks and a cloud platform for building voice, video, and physical AI agents.
