Appier logo

Senior Backend Engineer (SRE / Reliability & Performance)

Appier

Taipei City, TaiwanFull-timeNo compensation foundPosted 1mo agoVerified open 4 days ago

Most applications go out cold — see where you stand first. No sign-up to start.

At a glance

Compensation
No compensation found
Location
Taipei City, Taiwan
Schedule
Full-time
Work Authorization
Not specified

Requirements

Credentials this posting asks for.

Bachelor's degree

Job overview

Appier is hiring a Senior Backend Engineer (SRE / Reliability & Performance). Appier seeks a Senior Backend Engineer (SRE / Reliability & Performance) to design and build scalable, reliable backend services while owning system reliability, observability, and performance for high‑traffic platforms. The role involves profiling, tuning, incident response, and guiding engineering practices, and is intended to be based in Taiwan, welcoming international candidates.

Key focus areas include Design and build scalable, reliable backend services and supporting components, Own system reliability and performance, including SLO/SLA definition and capacity planning, and Profile, benchmark, and tune components to resolve latency and throughput bottlenecks.

Successful candidates bring 5+ Years Backend Software Development Experience. Important skills include Backend Development, Reliability Engineering, Observability, Performance Optimization, System Design, and Incident Management. Preferred (not required): Technical Leadership, Profiling Tools, Debugging Tools, and High-performance Network Services.

Skills & qualifications

RequiredNice to have

Skills

Backend DevelopmentReliability EngineeringObservabilityPerformance OptimizationSystem DesignIncident ManagementCI/CD PipelinesDeployment AutomationCode ReviewsCoachingAgile CollaborationOn-Call RotationLinuxGoPythonJavaScalaC++Network API DesignRESTGraphQLSQLNoSQLMySQLPostgreSQLMongoDBRedisPrometheusGrafanaTracingAWSGCPAzureGitProactiveInterpersonal SkillsProblem-Solving SkillsTechnical LeadershipProfiling ToolsDebugging ToolsHigh-Performance Network ServicesDistributed AlgorithmsData Structures and AlgorithmsContainer OrchestrationKubernetesDockerInfrastructure as CodeConfiguration ManagementTerraformBackend Software DevelopmentInfrastructure ResponsibilitiesSystem Reliability TuningTraffic Problem SolvingScalability Problem SolvingWeb Services on LinuxSQL DatabasesNoSQL DatabasesObservability ToolingMentoring EngineersAgile ProcessesHigh-Performance Network Services on LinuxDesigning Distributed SystemsArchitecting Distributed SystemsImplementing Distributed AlgorithmsImplementing Data StructuresSystem Performance TuningSolving Traffic ProblemsSolving Scalability ProblemsDesigning Large-Scale Distributed SystemsArchitecting Large-Scale Distributed SystemsAnsibleJenkinsGitLab CIGitHub ActionsArgo CDNginxHAProxyLoad BalancingCaching StrategiesData-Intensive Application DesignMonitoring SystemsAlerting SystemsNagiosCapacity PlanningChaos EngineeringProactive Interpersonal Problem-SolvingProfiling and Debugging ToolsChaos / Resilience EngineeringOperation AutomationContinuous Integration / Continuous DeploymentCI/CDProfilingDistributed SystemsResilience EngineeringStrong InterpersonalProblem SolvingFacilitating Agile ProcessesProfiling Debugging ToolsAgile FacilitationMonitoringAlertingDebuggingPerformance EngineeringStrong Interpersonal SkillsGo / Python / Java / Scala / C++REST / GraphQLMySQL / PostgreSQL / MongoDB / RedisSLO/SLA DefinitionSQL/NoSQL DatabasesProactive Interpersonal SkillsInterpersonalMonitoring and Alerting SystemsREST API Design

Qualifications

5+ Years Backend Software DevelopmentProven Experience Tuning System Reliability and PerformanceDemonstrated Experience Solving Traffic and Scalability ProblemsBS/MS Degree in Computer Science or Related FieldTechnical Leadership ExperienceExperience Designing and Architecting Large-Scale Distributed Systems

Full job description

About Appier

Appier (TSE: 4180) is an AI-native Agentic AI as a Service (AaaS) company that empowers businesses to create value through cutting-edge AdTech and MarTech solutions. Founded in 2012 with the vision of "Making AI Easy by Making Software Intelligent," Appier helps businesses turn AI into ROI through its Ad Cloud, Personalization Cloud, and Data Cloud—each powered by Agentic AI that enables autonomous, adaptive, and real-time decision-making. Today, Appier operates 17 offices across APAC, the US, and EMEA, and is listed on the Tokyo Stock Exchange. Learn more at www.appier.com.

About the role

Engineers at Appier build a wide range of platforms and services that interconnect data and AI with our customers and users—operating at scale where every millisecond and every nine of availability matters. As a Senior Backend Engineer (SRE / Reliability & Performance), you will sit at the intersection of backend development and site reliability engineering: designing and building scalable, performant backend services while owning the reliability, observability, and performance characteristics of large-scale, high-traffic systems (QPS > 10k). You will profile and tune systems end to end, lead the response to production incidents, and drive the engineering practices that keep our services fast and resilient as traffic grows. Seniority and title are determined by job-related skills, experience, and evaluation following the interview. We welcome international candidates to our teams. This position is ideally to be based in Taiwan.

Responsibilities

  • Design and build scalable, reliable, and maintainable backend services and the components that support them
  • Own system reliability and performance for high-traffic services, including SLO/SLA definition, error budgets, and capacity planning
  • Profile, benchmark, and tune critical components to resolve latency, throughput, and scalability bottlenecks in medium-to-large systems (QPS > 10k)
  • Diagnose and solve production traffic problems—load spikes, hot paths, resource contention, and cascading failures
  • Lead system design and provide technical guidance on reliability and performance trade-offs
  • Continuously improve observability (logging, metrics, tracing), incident management, DevOps, and production operational SOPs
  • Lead incident response, troubleshooting, and blameless post-mortems, then drive the follow-up engineering work
  • Build and optimize CI/CD pipelines and deployment automation to ship safely and frequently
  • Lead code reviews to ensure high quality coding and operational standards
  • Mentor engineers and facilitate agile collaboration across cross-functional teams
  • Participate in on-call rotation to ensure product reliability and scalability

About you

[Minimum qualifications]

  • 5+ years of experience in backend software development, with hands-on SRE / reliability / infrastructure responsibilities
  • Proven experience tuning system reliability and performance for production services
  • Demonstrated experience solving traffic and scalability problems in medium-to-large systems
  • Ability to build and operate web services on Linux
  • Proficient in one or more of the following languages: Go / Python / Java / Scala / C++
  • Good knowledge of Network API design (e.g. REST or GraphQL)
  • Good understanding of SQL/NoSQL databases (MySQL / PostgreSQL / MongoDB / Redis / etc.)
  • Hands-on experience with observability tooling (e.g. Prometheus, Grafana, distributed tracing)
  • Familiar with AWS, GCP, or Azure
  • Familiar with Git
  • Proactive, with strong interpersonal and problem-solving skills

[Preferred qualifications]

  • BS/MS degree in Computer Science or related field
  • Technical leadership experience, such as mentoring engineers and facilitating agile processes
  • Strong skills with profiling and debugging tools, and building high-performance network services on Linux
  • Experience designing and architecting large-scale distributed systems and implementing distributed algorithms and data structures
  • Experience with container orchestration (Kubernetes, Docker) and infrastructure-as-code / configuration management (Terraform, Ansible)
  • Hands-on experience with CI/CD platforms (Jenkins, GitLab CI, GitHub Actions, ArgoCD)
  • Expert in some of the following CS domains:
    • Nginx / HAProxy and load balancing
    • Caching strategies and data-intensive application design
    • Monitoring and alerting systems (Prometheus / Nagios)
    • Capacity planning and chaos / resilience engineering
    • Operation automation
    • Continuous integration / continuous deployment

#LI-TC1

You've read the whole posting — now see how you match it.