1 day ago
Bogotá, ColombiaSenior
Responsibilities
- Develop and improve software solutions, code, scripts, services, and automation that increase system stability, reliability, and performance.
- Troubleshoot complex issues across applications, integrations, infrastructure, cloud environments, APIs, gateways, and distributed systems.
- Support AWS-based solutions, Kubernetes environments, CI/CD pipelines, infrastructure configuration, secrets management, and deployment processes.
- Design and implement monitoring, alerting, logging, dashboards, and observability solutions while improving alert quality and reducing operational noise.
- Investigate integration and API reliability issues and improve API performance, availability, monitoring, and error handling.
- Own complex incidents, perform root-cause analysis, implement corrective and preventive actions, and document technical findings.
- Participate in a 24/7 on-call rotation two days per week within LAM/NAM time zones after onboarding and knowledge transfer.
- Apply security-by-design and quality practices, support vulnerability remediation, and contribute to testing, integration validation, release readiness, and post-release support.
Requirements
- Professional experience in Site Reliability Engineering, Software Engineering, DevOps, Platform Engineering, or a closely related technical field.
- Strong hands-on experience with AWS cloud environments, Kubernetes, container orchestration, infrastructure configuration, and secrets management.
- Practical experience with CI/CD processes and tools such as Jenkins or comparable technologies.
- Experience implementing or supporting monitoring, logging, alerting, observability, and application performance management solutions.
- Strong troubleshooting skills across applications, integrations, infrastructure, and cloud environments.
- Experience with APIs, gateways, middleware, or complex integration layers, with TIBCO or Kong experience beneficial.
- Ability to develop or improve code, scripts, automation, and technical solutions.
- Understanding of security, vulnerability management, system testing, release, and deployment practices.
- Strong spoken and written English and the ability to collaborate with multicultural, geographically distributed teams.
- Experience with Elastic Stack, Elasticsearch, Logstash, Kibana, Grafana, TIBCO, Kong, high-load or business-critical systems, on-call operations, or automation of recurring support activities is preferred.
- Strong accountability, ownership, independent problem-solving, prioritization, communication, and willingness to learn business-specific systems and architectures.
- A university degree in Computer Science, Software Engineering, Information Technology, or a related field, or an equivalent combination of education and professional experience.
Benefits
- Participation in a 24/7 on-call rotation two days per week within LAM/NAM time zones after the onboarding and knowledge-transfer period.
- Collaboration with globally distributed teams across multiple time zones.
- Opportunity to develop within the Site Reliability Engineering discipline.
Tech Stack
Categories
Site Reliability
About Adidas
Adidas designs and sells athletic footwear, apparel, and equipment for athletes and lifestyle consumers, distributing through wholesale partners as well as its own retail and e-commerce channels. Founded in 1949 and headquartered in Herzogenaurach, Germany, it is a public company listed in Frankfurt. The brand is the longtime official supplier of FIFA World Cup match balls and has a major presence in football, running, and training.
