3 hours ago
Kansas City, MO, USA or San Jose, CA, USAStaff+
Responsibilities
- Own the architecture, governance, environment model, networking, identity, and operation of the Azure platform at large scale.
- Lead SRE initiatives, including SLOs, SLIs, on-call rotations, monitoring, observability, automated provisioning, and disaster recovery.
- Establish DevSecOps practices covering infrastructure as code, CI/CD pipelines, security testing, and deployment mechanisms.
- Automate deployment, scaling, and management of containerized microservices and event-driven systems using Docker, Kubernetes, and AKS.
- Optimize application performance, resource utilization, reliability, and cloud cost efficiency with engineering teams.
- Mentor and influence engineers on Azure architecture, reliability, security-first development, and operational best practices.
- Collaborate with Product Management and Security Operations to design scalable, cost-effective platform solutions.
- Evaluate and adopt infrastructure technologies and engineering methodologies.
Requirements
- 8+ years of progressive experience in DevOps, Site Reliability Engineering, or Platform Engineering roles.
- Deep hands-on production Azure platform expertise, including AKS, Entra ID, workload identity federation, VNet design, Private Link, Key Vault, Azure Policy, and subscription or landing-zone architecture.
- Experience standing up or migrating production workloads across cloud environments, including explaining architecture, data paths, cutovers, and failure handling.
- Experience building secure and compliant environments, such as SOC 2 or ISO 27001 environments.
- Deep understanding of microservices, containerization, Docker, Kubernetes, and event-driven systems.
- Extensive experience with infrastructure-as-code tools such as Terraform or Bicep and CI/CD practices.
- Production experience writing and shipping software in Go or Python beyond scripting and configuration.
- Experience with monitoring and observability tools such as Prometheus, Grafana, Azure Monitor, or ELK Stack.
- Familiarity with real-time data pipelines and stream-processing technologies such as Kafka, Event Hubs, Service Bus, or Pub/Sub.
- Proven experience architecting, building, and operating highly scalable, distributed, secure enterprise SaaS platforms.
- Bachelor’s or Master’s degree in Computer Science or Engineering, or relevant equivalent experience.
- Relevant Azure, Kubernetes, or security certifications are a plus.
- Experience with multiple major cloud providers, cybersecurity or MSSP environments, large-scale data warehousing or lakehouse technologies, high-growth startups or enterprise SaaS, or MLOps is preferred.
Benefits
- Monday through Thursday onsite with Friday work-from-home; Kansas City is preferred, with San Jose or Sarasota, Florida also considered.
- Candidates must live in or be willing to relocate to Kansas City, San Jose, or Sarasota.
- Early-stage startup opportunity with meaningful influence on culture, architecture, and company growth.
Tech Stack
Categories
DevOpsSite Reliability
About TENEX.AI
TENEX is the first AI-native, human-led MDR powered by AI SOC. Backed by 24/7 U.S.based expert analysts with 8+ years avg. experience. Our human-led, AI-driven platform delivers 10x faster detection, <1-minute MTTX, and 95% fewer false positives, transforming security operations economics while dramatically improving outcomes. AI handles 100% of alerts at machine speed, allowing our experts to focus on complex threats demanding human judgment. We scale through AI, not headcount, delivering premium outcomes without enterprise-level costs.
