
Infrastructure Software Engineer
Normal Computing CorporationBase Salary
$185k - $285k/yr
Responsibilities
- Build and maintain production infrastructure software for AI products, especially orchestration, execution, and runtime systems.
- Design internal backend services and APIs for product engineers, AI engineers, execution services, and other internal systems.
- Improve state management, failure handling, metrics, tracing, debugging tools, reliability, and operational maturity.
- Work with Kubernetes-backed execution environments, including container lifecycle, scheduling, autoscaling, resource isolation, and runtime reliability.
- Build developer-facing tools and abstractions for systems owned by the team.
- Turn prototypes into durable production systems through clear abstractions, hardened critical paths, and scalable operational patterns.
- Collaborate with product, AI, research, and platform engineers on interfaces between product features, AI workloads, and production infrastructure.
- Lead design discussions for runtime and orchestration systems, including API boundaries, state management, execution models, and operational tradeoffs.
Requirements
- 4+ years of experience in infrastructure software, backend infrastructure, production infrastructure, platform engineering, distributed systems, or related areas.
- Strong backend software engineering fundamentals, including APIs, data modeling, concurrency, debugging, and testing.
- Experience building or operating reliable, observable, and maintainable production services.
- Practical experience with Docker and Kubernetes, including containerized workload debugging, deployments, networking, resource limits, and lifecycle issues.
- Comfort with persistence systems such as Postgres, Redis/Valkey, object storage, or similar production data stores.
- Experience building orchestration systems, job schedulers, workflow engines, sandboxes, developer platforms, or distributed execution systems.
- Experience designing internal APIs and developer-facing abstractions.
- Strong systems thinking around state machines, failure modes, retries, queues, leases, scheduling, and long-running workflows.
- Clear communication, ownership, pragmatism, and sound technical judgment across product, AI, and infrastructure boundaries.
- Preferred: deep Kubernetes experience with controllers/operators, networking, storage, scheduling, autoscaling, or resource isolation.
- Preferred: experience with AI agent infrastructure, ML infrastructure, model orchestration, or LLM-based product systems.
- Preferred: production infrastructure, reliability engineering, or infrastructure software experience at meaningful scale.
- Preferred: experience in high-growth startups or teams with evolving ownership boundaries.
- Preferred: experience with chips, EDA, or device verification.
Tech Stack
About Normal Computing Corporation
At Normal, we're rewriting AI foundations to advance the frontier of reasoning and reliability in the physical world. We are tackling problems across semiconductors and industrials with a mix of interdisciplinary approaches across the full stack: from probabilistic software infrastructure and algorithms to hardware and physics, enabling AI that can reason and understand its own limits. We understand that our technology is only as powerful as the people behind it. Every employee drives significant impact within our products, often working directly with customers and embedding across our tightly-knit team. Our team members are driven by curiosity and passion for solving some of the most challenging problems in the world of atoms. Normal was founded in 2022 by engineers and scientists that pioneered industry-leading Physics + ML tools for next-gen AI at Google Brain and Google X.