3 months ago
Cambridge, MA, USAStaff+
Base Salary
$300k - $500k/yr
Responsibilities
- Own mission-critical production systems end to end, including correctness, scalability, performance, reliability, and operational excellence.
- Design, build, and ship high-impact backend systems and features that improve reliability, performance, and customer value.
- Architect scalable services and cloud infrastructure using Python, REST, gRPC, Kubernetes, and Terraform.
- Identify and resolve technical bottlenecks affecting engineering quality, system performance, and organizational velocity.
- Build and operate LLM-powered systems and validation loops for correctness, consistency, durability, and production performance.
- Design and evolve relational, NoSQL, graph, and vector data architectures for enterprise applications and semantic retrieval.
- Modernize complex enterprise systems while balancing reliability, maintainability, scalability, and delivery speed.
- Set engineering quality standards through hands-on technical leadership and long-term technical ownership.
Requirements
- Direct experience using Python as a primary programming language, backend frameworks, and microservices architectures.
- Expertise in REST and gRPC, with proficiency in Node.js and JavaScript.
- Proficiency in GCP and experience with at least one additional cloud platform such as AWS or Azure.
- Advanced production experience with Kubernetes and Terraform.
- Experience operating highly available production systems, including monitoring, scalability, reliability, performance optimization, and operational tooling.
- Strong knowledge of SQL and NoSQL databases, including PostgreSQL, MySQL, MongoDB, Cassandra, or DynamoDB.
- Familiarity with graph databases such as Neo4j and vector databases or embedding infrastructure for semantic search and retrieval.
- Hands-on experience building and operating production LLM-powered systems, including evaluation, validation, regression testing, tracing, and failure analysis.
- Working knowledge of LangSmith or comparable LLM observability and evaluation tools; familiarity with OpenAI, Anthropic, or similar model providers is a plus.
- Ability to contribute across the full stack and debug, design, and ship across frontend, backend, infrastructure, and AI systems.
- Understanding of large-scale enterprise software architecture, integration, deployment, modernization, and long-term maintainability.
- Proven Staff+, Principal Engineer, or comparable experience independently driving complex technical initiatives and delivering high-impact outcomes with minimal supervision.
Benefits
- Base salary range of $300,000-$500,000 plus bonus and equity.
- The role is a fully hands-on individual contributor position without people-management responsibilities.
- The company emphasizes in-person collaboration and a fast-paced working environment.
Tech Stack
Amazon DynamoDBApache CassandraAWSAzureGoogle Cloud PlatformgRPCJavaScriptKubernetesMongoDBMySQLNeo4jNode.jsPostgreSQLPythonSQLTerraform
