Appier

Senior Backend Engineer (SRE / Reliability & Performance)

Appier
Apply
2 months ago
Taipei, TaiwanSenior

Responsibilities

  • Design and build scalable, reliable, and maintainable backend services and supporting components.
  • Own reliability and performance for high-traffic services through SLO/SLA definition, error budgets, and capacity planning.
  • Profile, benchmark, and tune systems to resolve latency, throughput, and scalability bottlenecks.
  • Diagnose production traffic issues including load spikes, hot paths, resource contention, and cascading failures.
  • Lead system design and guide reliability and performance trade-offs.
  • Improve logging, metrics, tracing, incident management, DevOps practices, and production operational procedures.
  • Lead incident response, troubleshooting, blameless post-mortems, and follow-up engineering work.
  • Build and optimize CI/CD pipelines and deployment automation.
  • Lead code reviews and mentor engineers across cross-functional teams.
  • Participate in the on-call rotation.

Requirements

  • At least five years of backend software development experience with hands-on SRE, reliability, or infrastructure responsibilities.
  • Proven experience tuning reliability and performance for production services and solving traffic and scalability problems in medium-to-large systems.
  • Ability to build and operate web services on Linux.
  • Proficiency in one or more of Go, Python, Java, Scala, or C++.
  • Knowledge of network API design, including REST or GraphQL, and SQL/NoSQL databases such as MySQL, PostgreSQL, MongoDB, or Redis.
  • Hands-on experience with observability tooling such as Prometheus, Grafana, and distributed tracing.
  • Familiarity with AWS, GCP, Azure, and Git.
  • Preferred qualifications include a BS/MS in Computer Science or a related field, technical leadership experience, profiling and debugging expertise, large-scale distributed-systems experience, and knowledge of distributed algorithms and data structures.
  • Preferred experience includes Kubernetes, Docker, Terraform, Ansible, Jenkins, GitLab CI, GitHub Actions, ArgoCD, Nginx, HAProxy, Nagios, load balancing, caching, capacity planning, chaos or resilience engineering, and operational automation.

Benefits

  • Position ideally based in Taiwan.
  • Participation in an on-call rotation.
  • Opportunity to mentor engineers and collaborate across cross-functional teams.

Tech Stack

AnsibleAWSAzureC++DockerGitGitHub ActionsGitLab CI/CDGoGoogle Cloud PlatformGrafanaGraphQLJavaJenkinsKubernetesLinuxMongoDBMySQLNagiosPostgreSQLPrometheusPythonRedisScalaSQLTerraform

Categories

BackendSite Reliability
Appier

About Appier

501-1,000 employees

Appier is an AI-native Agentic AI as a Service (AaaS) company that empowers businesses to create value with cutting-edge AdTech and MarTech solutions. Guided by the vision of “Making AI Easy by Making Software Intelligent,” our mission is to help businesses turn Agentic AI into ROI. Founded in 2012, Appier is listed on the Tokyo Stock Exchange’s Prime Market (Ticker: 4180) and operates in 17 cities worldwide, enabling over 2,000 leading companies to enhance marketing performance with the latest AI technology. As AI enablers for our customers in the AI Era, Appier delivers innovative solutions that drive measurable results.