
DevOps Engineer, Cloud and Toolchain
Corning Incorporated24 days ago
Shanghai, ChinaSenior
Responsibilities
- Create automated implementation, security, monitoring, alerting, and operations processes using Terraform, Ansible, YAML, Python, and environment-specific scripting languages.
- Develop and maintain integrated systems and tools supporting Agile and DevOps workflows across multiple development teams.
- Operate toolchain infrastructure using Ansible, Terraform, and Kubernetes.
- Design reusable templates and workflow integrations that reduce repetitive work and improve system resiliency.
- Test code, processes, and deployments incrementally to streamline execution and minimize errors.
- Document work and define repeatable actions that can be automated.
- Design, implement, and support monitoring infrastructure, metrics, alerting, and automated response capabilities.
- Collaborate on complex infrastructure, security, and development problems.
- Support availability incidents and urgent analytic, development, or operational needs, including after-hours escalations.
- Debug production issues across services and component levels.
- Use and continually improve the internal toolchain service offering.
Requirements
- Undergraduate degree in Computer Science or an equivalent technical area.
- Three to five years of hands-on production programming experience with agile software development in languages such as Python, .NET/C#, Go, Java, and JavaScript/Node.js.
- AWS Solution Architect certification earned within the last 12 months, with professional-level certification strongly preferred.
- Additional DevOps certifications from a major cloud infrastructure or CNCF-related technology provider within the last 24 months, preferably AWS.
- At least five years of infrastructure automation experience, including at least three years managing production environments for foundational services in AWS, Azure, and/or GCP.
- Deep understanding of the AWS Well-Architected Framework.
- Experience designing, developing, and maintaining backend systems using MLFlow, Kubernetes, and Docker.
- Ability to evaluate and implement AI-driven solutions for infrastructure provisioning, incident response, and system performance.
- At least three years of hands-on staging and production experience with Terraform and Ansible, development CI/CD tools, Docker and Kubernetes, systems administration scripting, Windows and Linux environments, and network architecture.
- Experience with CI/CD tools such as GitLab, Jenkins, or Azure DevOps.
- Knowledge of load balancing, DNS, BGP, and IPSec VPNs.
- Strong systems thinking, documentation, problem-solving, troubleshooting, communication, collaboration, and process-oriented skills.
- Ability to work in an Agile product team, balance stakeholder requests, define technology roadmaps, and promote adoption of best practices.
Tech Stack
AnsibleAWSAzureBashC#DockerGoGoogle Cloud PlatformJavaJavaScriptJenkinsKubernetesLinuxMLflow.NETNode.jsPowerShellPythonTerraformWindows