1 month ago
Bangkok, ThailandMid Level
Responsibilities
- Monitor and analyze platform performance, availability, and critical metrics using observability tools.
- Diagnose and troubleshoot complex production issues and conduct root cause analyses.
- Resolve bugs and system defects promptly.
- Design and implement system optimizations and reliability enhancements for a scalable, stable, high-performing platform.
- Collaborate with engineering, DevOps, and product teams.
- Own issues end to end from detection and resolution through preventative measures.
Requirements
- Bachelor’s degree in Computer Engineering or a related field.
- Proficiency in Python is required.
- Experience with Java is advantageous.
- Strong troubleshooting skills and end-to-end ownership of issues.
- Understanding of system performance, scalability, and reliability engineering principles.
- Experience with system monitoring, logging, and observability tools.
- Experience with enterprise-grade production systems.
- Ability to learn and adopt new technologies quickly, with strong self-management, responsibility, and attention to detail.
Benefits
- Competitive salaries and global wellness benefits.
- Opportunity to join a rapidly growing, globally scaling company working on AI-powered integration tools.
