4 hours ago
Base Salary
$210k - $245k/yr
Responsibilities
- Design, build, and operate production agentic workflows and platform harnesses with orchestration, tool integrations, shared context, and extension points.
- Develop intent-engineering specifications, rules, constraints, and acceptance criteria that AI systems can execute reliably.
- Build observability, regression detection, human-in-the-loop controls, guardrails, safe rollout, and operational capabilities for agentic systems.
- Own agent-runtime capabilities including job isolation, scheduling, execution environments, secrets, access controls, and cloud operations.
- Design event-driven architectures using queues, streams, webhooks, and asynchronous job fan-out where appropriate.
- Improve AI-assisted software delivery across writing, testing, reviewing, debugging, and validating changes in repositories and pipelines.
- Evaluate build-versus-buy decisions and stay current with models, agent frameworks, MCP-style protocols, and AI tooling.
- Partner across Engineering to drive adoption through playbooks, examples, demonstrations, and enablement.
- Define and track adoption, quality, reliability, and satisfaction metrics and use feedback to guide investment.
- Take projects from design through implementation, rollout, and day-two operability, including potential stand-by, on-call, or off-hours support.
Requirements
- Staff-level or equivalent professional software engineering experience building and operating production distributed systems.
- Strong TypeScript, modern JavaScript, and Node.js experience shipping production services.
- Strong software engineering fundamentals, including design, testing, debugging, APIs, services, and code quality.
- Experience with automated pipelines, progressive delivery or safe rollout, monitoring, rollback, and frequent reversible production changes.
- Hands-on experience with modern LLM and agentic coding workflows in production or serious internal platforms.
- Experience building backend services, integrations, automation, and internal tools.
- Professional experience with agent orchestration, tool-calling systems, evaluation or guardrail techniques, and reliable backend integrations.
- Familiarity with agent frameworks, orchestration layers, and integrating external tools and data sources into LLM-based systems; MCP or equivalent experience is a plus.
- Production experience with event-driven systems such as queues, streams, pub/sub, webhooks, or similar asynchronous patterns.
- Fluency with metrics, logs, traces, and operating production systems.
- Clear communication, documentation, teaching, adoption-driving, independent problem-solving, and sound judgment around security, permissions, data access, and rollout safety.
- Hands-on Kubernetes experience, Terraform or similar infrastructure-as-code experience, Temporal or comparable workflow-platform experience, and deeper AWS/cloud-native operations fluency are preferred.
Benefits
- Remote-first work arrangement.
- Base salary range of $210,000—$245,000 USD.
- Health, dental, vision, short-term disability, and life insurance.
- Paid holidays and paid time off.
- Fertility treatment benefit.
- 401(k) and equity.
- Corporate Bonus Program eligibility for non-sales roles.
Categories
About Cribl
Cribl builds a vendor‑agnostic telemetry data platform used by IT, security, and observability teams to collect, route, shape, store, and search machine data in real time. Its products (including Cribl Stream, Edge, and Search) are sold as enterprise subscriptions and can run in cloud or self‑managed environments. Founded in 2018 and headquartered in San Francisco, the company is privately held and serves large global enterprises.
