
Principal Observability Automation Engineer
Q2 Software, Inc.11 days ago
Responsibilities
- Translate ambiguous business needs into architected, detectable, and automatable systems from prototype through production.
- Extend the proprietary cloud-hosted agentic AI platform, including locally hosted models, for observability and automated issue detection.
- Design and implement enterprise-scale, 24/7, cross-region-resilient observability and automation frameworks across AWS, Azure, GCP, VMware, and on-premises environments.
- Develop full-stack applications and APIs for automated data collection, correlation, and visualization.
- Modernize high-risk access and operations pathways into secure, controlled, self-service tooling with appropriate guardrails.
- Define and enforce best practices for system design, code quality, security, performance, availability, confidentiality, and privacy.
- Improve monitoring, alerting, and incident-response systems and identify automation opportunities with cross-functional teams.
- Own the lifecycle of observability tools and platforms, including CI/CD pipelines, configuration management, and infrastructure as code.
- Mentor senior, mid-level, and junior engineers while decomposing complex problems and maintaining a high quality bar.
- Lead complex technical projects, participate in strategic planning and roadmap development, and make technical decisions.
- Provide light event-driven on-call support for owned systems, typically a handful of times per year.
Requirements
- Bachelor’s degree in Computer Science or a related field with 15 years of experience in software engineering, systems automation, or observability tooling; or an advanced degree with 12+ years of experience; or equivalent related work experience.
- Expertise in C#, Go, and scripting languages including Python, Bash, and PowerShell.
- Deep understanding of RESTful APIs, microservices architecture, and distributed systems.
- Proficiency with SQL and NoSQL databases, including MS SQL Server and Postgres.
- Strong experience with monitoring and analytics platforms such as Grafana and Splunk.
- Hands-on experience with CI/CD tools and practices, configuration management, and infrastructure as code.
- Advanced knowledge of Linux, including Red Hat/CentOS, and Windows Server administration.
- Experience with Docker, orchestration, and Terraform-based infrastructure automation.
- Proven ability to lead complex technical projects and influence cross-functional teams.
- Excellent communication and leadership skills, including the ability to present architecture to engineers and executive leaders.
- Experience with NGINX, IIS, and public cloud datacenter operations is preferred.
- Fluent written and oral communication in English.
- Applicants must be authorized to work for any employer in the United States.
Benefits
- Hybrid work opportunities.
- Flexible time off.
- Health insurance offerings and paid parental leave for eligible new parents.
- Career development and mentoring programs.
- Community volunteering, company philanthropy programs, and employee peer recognition programs.
- Supportive, inclusive culture prioritizing career growth, collaboration, and physical, mental, and professional wellness.