
Senior Software Engineer II - Action Infrastructure - Action Runtime
DigitalOcean2 months ago
Base Salary
$184k - $231k/yr
Responsibilities
- Own and ship Action Infrastructure product features from planning and design through implementation, launch, operations, and maintenance.
- Build production-grade third-party agent tools and infrastructure, including MCP servers, tool schemas, authentication and secrets handling, rate limiting, observability, and safe execution boundaries.
- Create full-stack experiences for builders and operators to discover, configure, test, debug, and monitor agent tools.
- Define quality standards using evaluations, golden tasks, telemetry, error analysis, model and tool selection experiments, and production feedback.
- Translate customer and business requirements into clear, pragmatic technical designs and implementation plans.
- Collaborate with product, design, UX, infrastructure, security, support, and engineering partners on reliable customer-facing capabilities.
- Guide technical decisions, review code and architecture, mentor engineers, and raise standards for reliability and operational excellence.
- Operate production systems, participate in on-call practices, debug multi-system issues, lead incident diagnosis, and convert operational learnings into improvements.
- Track and apply emerging practices in MCP, evaluations, tool-use reliability, agent UX, and AI coding and agent environments.
Requirements
- Senior-level experience designing, building, shipping, and operating production software with meaningful reliability, scale, security, or customer impact.
- Production experience with Go, TypeScript, and Python, or deep expertise in one or two with the ability to work across all three.
- Hands-on experience with agent and AI coding tools such as Claude Code, Codex, Cursor, or similar systems.
- Practical familiarity with MCP, tool contracts and schemas, tool and model selection, agent evaluations, observability, and failure analysis.
- Ability to build backend services and APIs while shaping developer and operator workflows end to end.
- Strong production engineering experience with SLOs, incident response, monitoring, tracing, capacity, deployment safety, authentication, secrets, abuse prevention, and supportability.
- Ability to turn ambiguous requirements into RFCs, architecture documents, implementation plans, milestones, and tradeoff decisions.
- Experience influencing cross-functional projects and partnering with product managers, designers, managers, security, support, and engineering teams.
- Track record of mentoring or coaching engineers through technical decisions, code quality, architecture, and communication.
- Preferred experience with tool ecosystems, integration or developer platforms, workflow automation, cloud control planes, production LLM or agent evaluation, third-party API reliability patterns, open-source contributions, developer communities, or technical talks.
Benefits
- Hybrid work arrangement.
- Conference, training, and education reimbursement.
- Access to LinkedIn Learning with more than 10,000 courses.
- Employee Assistance Program, local employee meetups, and flexible time off.
- Eligible employees may receive equity grants and participate in the Employee Stock Purchase Program.
Tech Stack
Categories
About DigitalOcean
DigitalOcean is the AI-Native Cloud purpose-built for the inference and agentic era. Its five-layer integrated platform—spanning GPU and CPU infrastructure, core cloud, inference, data, and managed agent orchestration—is open throughout with no vendor lock-in, giving builders everything they need to start fast, scale production AI workloads, and improve unit economics. More than 650,000 customers and millions of developers globally trust DigitalOcean to build, ship, and scale their applications.