3 months ago
Taipei, TaiwanSenior
Responsibilities
- Design and build scalable, reliable, and maintainable backend services and supporting components.
- Own reliability and performance for high-traffic services through SLO/SLA definition, error budgets, and capacity planning.
- Profile, benchmark, and tune systems to resolve latency, throughput, and scalability bottlenecks.
- Diagnose production traffic issues including load spikes, hot paths, resource contention, and cascading failures.
- Lead system design and guide reliability and performance trade-offs.
- Improve logging, metrics, tracing, incident management, DevOps practices, and production operational procedures.
- Lead incident response, troubleshooting, blameless post-mortems, and follow-up engineering work.
- Build and optimize CI/CD pipelines and deployment automation.
- Lead code reviews and mentor engineers across cross-functional teams.
- Participate in the on-call rotation.
Requirements
- At least five years of backend software development experience with hands-on SRE, reliability, or infrastructure responsibilities.
- Proven experience tuning reliability and performance for production services and solving traffic and scalability problems in medium-to-large systems.
- Ability to build and operate web services on Linux.
- Proficiency in one or more of Go, Python, Java, Scala, or C++.
- Knowledge of network API design, including REST or GraphQL, and SQL/NoSQL databases such as MySQL, PostgreSQL, MongoDB, or Redis.
- Hands-on experience with observability tooling such as Prometheus, Grafana, and distributed tracing.
- Familiarity with AWS, GCP, Azure, and Git.
- Preferred qualifications include a BS/MS in Computer Science or a related field, technical leadership experience, profiling and debugging expertise, large-scale distributed-systems experience, and knowledge of distributed algorithms and data structures.
- Preferred experience includes Kubernetes, Docker, Terraform, Ansible, Jenkins, GitLab CI, GitHub Actions, ArgoCD, Nginx, HAProxy, Nagios, load balancing, caching, capacity planning, chaos or resilience engineering, and operational automation.
Benefits
- Position ideally based in Taiwan.
- Participation in an on-call rotation.
- Opportunity to mentor engineers and collaborate across cross-functional teams.
Tech Stack
AnsibleAWSAzureC++DockerGitGitHub ActionsGitLab CI/CDGoGoogle Cloud PlatformGrafanaGraphQLJavaJenkinsKubernetesLinuxMongoDBMySQLNagiosPostgreSQLPrometheusPythonRedisScalaSQLTerraform
Categories
BackendSite Reliability
About Appier
Appier builds AI-driven advertising and marketing software for enterprises, offering programmatic ad bidding, customer data, and predictive tools delivered as SaaS. Founded in 2012 and headquartered in Taipei, it is a public company listed on the Tokyo Stock Exchange (ticker: 4180). Its Ad Cloud processes millions of bid requests per second across APAC, Europe, and the U.S., and is used by in-house and agency marketers.
