5 months ago
Base Salary
$201k - $264k/yr
Responsibilities
- Design and build scalable, fault-tolerant infrastructure systems across multiple cloud regions
- Own and evolve multi-cloud infrastructure, including Kubernetes orchestration, networking, and container management
- Lead observability, incident response, and operational excellence initiatives
- Architect and optimize distributed systems for reliability, including load balancing, quota management, and failover mechanisms
- Partner with Product Engineering and Security teams on infrastructure strategy and delivery
- Drive infrastructure-as-code practices with Terraform and Pulumi to enable reproducible, auditable deployments
- Design a model proxy architecture for millions of daily inference requests and seamless model integration
- Build distributed rate limiting and quota management systems using Redis-backed algorithms
- Architect multi-region deployments that satisfy global data residency requirements
- Develop observability infrastructure with SLA monitoring, burn rate alerts, and token attribution for cost tracking
- Lead the evolution of CI/CD pipelines while maintaining production stability
- Mentor engineers and raise the technical bar through code reviews, design reviews, and technical leadership
Requirements
- 10+ years of experience in infrastructure engineering or platform engineering in a production environment
- Extensive experience building and scaling complex, large-scale distributed systems
- Deep proficiency with cloud infrastructure platforms, with Azure preferred and GCP or AWS experience accepted
- Strong fluency with infrastructure-as-code tools such as Terraform, Pulumi, or CloudFormation
- Strong understanding of Kubernetes, container orchestration, networking, and cloud security at scale
- Experience with observability tools such as Datadog and Sentry and incident response practices using PagerDuty or Incident.io
- Strong programming skills in Python, Go, or similar languages
- Experience with AI/ML infrastructure or high-throughput inference systems is preferred
- Experience with distributed rate limiting, load balancing, or quota management systems is preferred
- Experience operating multi-tenant platforms with strict security and compliance requirements is preferred
- Track record of leading complex cross-functional projects and delivering measurable impact is preferred
Benefits
- In-person work model in New York City, New York
- Relocation assistance for new employees
Tech Stack
Categories
About Harvey
Harvey builds domain-specific generative AI for legal and other professional services, delivered as an enterprise platform and APIs to automate contract analysis, due diligence, compliance, and litigation workflows. Founded in 2022 and headquartered in San Francisco, it sells to law firms and corporate legal departments on enterprise agreements; investors include Sequoia Capital and OpenAI. Customers include multiple Am Law 100 firms and Fortune 500 companies’ legal teams.
