2 days ago
Base Salary
$194k - $413k/yr
Responsibilities
- Own end-to-end assurance architecture spanning telemetry normalization, correlation, root cause, guarded remediation, and outcome verification.
- Define a shared model covering intent, topology, configuration, identity, policy, and state.
- Architect planning and tool-use agents, MCP tool servers, and multi-agent orchestration across network domains.
- Establish safety architecture including approval gates, policy enforcement, remediation safety tiers, blast-radius limits, and deterministic rollback.
- Determine when to apply graph reasoning, classical machine learning, or generative AI.
- Create evaluation practices using golden fault corpora, replayable incident harnesses, MTTR, and false-escalation metrics.
- Drive strategy across business units and influence executive and product stakeholders.
- Mentor engineers and develop future principal and distinguished technical leaders.
Requirements
- 15+ years of experience in engineering and architectural leadership building and operating large distributed systems in production.
- Breadth across campus, data center, WAN/SD-WAN, and security networking, with depth in at least three domains.
- Deep knowledge of EVPN-VXLAN, BGP, SAI-level behavior, overlay/underlay networking, identity and policy enforcement, and streaming telemetry using gNMI/OpenConfig.
- Knowledge of property graph modeling and graph algorithms for dependency, impact, and causality.
- Applied experience with LLM agents, tool calling, MCP, agent evaluation, classical machine learning, and time-series modeling.
- Production AIOps or observability experience involving correlation, anomaly detection, noise reduction, and incident lifecycle management.
- Cloud-native experience at scale with Kubernetes, microservices, Kafka, time-series or OLAP stores, and hybrid or public cloud environments.
- Strong hands-on Python and Go coding experience sustained through design and architecture, with proficiency in AI-first, spec-driven prototyping.
- Demonstrated Principal- or Distinguished-level impact across multiple organizations and ability to lead discovery across teams without direct reporting relationships.
- Bachelor's, Master's, or PhD in computer science, engineering, or a related discipline is preferred.
Benefits
- The role is onsite with an expectation of primarily working from an HPE office.
- Health and wellbeing benefits support employees and their families physically, financially, and emotionally.
- Personal and professional development programs support career growth and knowledge expertise.
- HPE promotes an unconditionally inclusive workplace and provides accessibility accommodations for qualified applicants.
Tech Stack
Categories
About HPE
Hewlett Packard Enterprise (HPE) builds enterprise IT infrastructure and services, including servers, storage, networking (Aruba), and hybrid edge-to-cloud platforms like HPE GreenLake for businesses and public-sector organizations. Formed in 2015 after Hewlett-Packard’s split, HPE is headquartered in Houston, Texas and trades on the NYSE. It sells hardware, software, and subscription-based managed cloud services to help customers run workloads across data centers, colocation, and public clouds.
