
DevOps Engineer, Cloud and Toolchain
Corning Incorporated2 hours ago
Shanghai, ChinaSenior
Responsibilities
- Create automated processes for implementation, security, monitoring, alerting, and operations using Terraform, Ansible, YAML, Python, and environment-specific scripting languages.
- Develop and maintain integrated platform systems and tools supporting Agile and DevOps workflows across multiple development teams.
- Operate toolchain infrastructure using Ansible, Terraform, and Kubernetes.
- Design reusable templates and workflow integrations that reduce repetitive work and improve system resiliency.
- Test code, processes, and deployments incrementally to streamline execution and minimize errors.
- Document work and define repeatable actions that can be automated.
- Design, implement, and support monitoring infrastructure, metrics, alerting, and automated response capabilities.
- Collaborate on complex infrastructure, security, and development problems.
- Support availability incidents, urgent analytic and operational needs, production debugging, and after-hours escalations.
- Use and continuously improve the internal toolchain service offering.
Requirements
- Undergraduate degree in Computer Science or an equivalent area of technical study.
- Three to five years of hands-on production programming experience with languages such as Python, .NET/C#, Go, Java, and JavaScript/Node.js.
- AWS Solutions Architect certification obtained within the last 12 months, with professional-level certification strongly preferred.
- Additional DevOps certifications from a major cloud infrastructure or CNCF-related technology provider obtained within the last 24 months are preferred, with AWS preferred.
- At least five years of infrastructure automation experience, including at least three years managing production environments for foundational services in AWS, Azure, and/or GCP.
- Deep understanding of the AWS Well-Architected Framework.
- Experience designing, developing, and maintaining backend systems using MLFlow, Kubernetes, and Docker.
- Ability to evaluate and implement AI-driven solutions for infrastructure provisioning, incident response, and system performance.
- At least three years of hands-on staging and production experience with Terraform and Ansible, CI/CD tools such as GitLab, Jenkins, or Azure DevOps, Docker and Kubernetes, systems administration scripting, Windows and Linux environments, and network architecture.
- Knowledge of load balancing, DNS, BGP, and IPSec VPNs.
- Strong troubleshooting, documentation, communication, collaboration, and problem-solving skills.
- Ability to work in an Agile product-team environment, manage stakeholder requests, define technology roadmaps, and promote adoption of best practices.
Tech Stack
AnsibleAWSAzureBashC#DockerGoGoogle Cloud PlatformJavaJavaScriptJenkinsKubernetesLinuxMLflow.NETNode.jsPowerShellPythonTerraformWindows