Site Reliability Engineer

Taipei, Taiwan Until 8/21/2026 First posted March 26, 2025 Last posted March 26, 2025
Job description

 

Job Overview

We are seeking a skilled and passionate Site Reliability Engineer with a strong technical background and excellent communication skills. This individual will lead the development, construction, and management of reliable and distributed systems that support our business operations. 
 
In this role, you will play a vital part in supporting the following businesses: 
 
- IXT: A insurance core system solution for APAC insurance markets. 
- OneDegree HK: A user-friendly digital insurance platform for individuals and businesses in Hong Kong. 
- Cymetrics: A cybersecurity platform designed specifically for small and medium enterprises in the APAC region. 

OneDegree Tech Blog: https://medium.com/onedegree-tech-blog 

 

Responsibilities

  • Implement and enhance system reliability, availability, scalability, performance, and efficiency by leveraging monitoring, alerting, and automation tools on public cloud platforms like Azure and GCP. 
  • Participate in capacity planning, analyze software performance, and fine-tune systems to ensure optimal operation. 
  • Develop and enhance our CI/CD process and toolset to streamline software delivery and deployment. 
  • Define and monitor key metrics to assess and enhance system reliability. 
  • Collaborate closely with the engineering team to improve reliability and operational efficiency at every software development life cycle (SDLC) stage.
  • Troubleshoot, optimize infrastructure and automate repetitive tasks to increase efficiency and effectiveness

 

Requirements

  • Proficiency in programming languages such as Bash, Python, or Go. 
  • Advanced knowledge of monitoring solutions like Prometheus, Grafana, ELK (Elasticsearch, Logstash, Kibana).
  • Strong expertise and experience in cloud technologies, specifically Azure and GCP.
  • Experience in the complete software development life cycle (SDLC).
  • In-depth understanding of network concepts, particularly with a focus on security.
  • Hands-on experience implementing CI/CD processes, for example, using GitLab CI.
  • Proficiency in automation platforms like Ansible and Terraform.
  • Knowledge of orchestration tools like Kubernetes.
  • Familiarity with container technologies like Docker.
  • Experience with Git source code version control systems.
  • Strong problem-solving skills with a systematic approach, effective communication abilities, and a self-driven attitude. 

 

Interview Process

  • HR phone interview
  • 1st Interview: 1.5~2 hours, 1 hour meet with hiring team + 0.5 hours with HR
  • 2nd Interview: 0.5~1 hour, meet with CTO

 

Other Benefits 

To us, people are our greatest asset, and we are more than happy to invest in employees! We create a healthy work atmosphere and provide you with the tools and support for doing your job successfully. With a culture of flexibility and transparency, we believe there should be no barriers, and everyone’s contributions matter. 

Work Life Balance is a must  

  • 15 days annual leaves (pro-rata for partial month at first year) 
  • 5 days full-pay sick leaves, 3 days menstrual leaves 
  • Health check subsidy 
  • Ergonomic-design chair and fully-equipped devices for work 

Grow together & keep learning

  • Conferences & external subsidy 
  • Learning clubs to share technical skill (e.g: Frontend/Backend tech sharing, Product Management...etc) 

Work Hard, Play even Harder 

  • Various entertainment & sports clubs, attend basketball clubs today, and play board game tomorrow! 
  • Snacks & beverage to refill your energy anytime 

 

About this role

Summary

Develop and manage reliable systems, enhance performance, and automate processes.

Job title

Site Reliability Engineer

Experience level

3+ years

Industry

insurance

Location requirements

Located in Taipei, Taiwan; remote work not allowed.

Salary

Not specified

Management role

No

Skills & keywords

Required skills

BashPythonGoPrometheusGrafanaELKAzureGCPSDLCnetworkCI/CDGitLab CIAnsibleTerraformKubernetesDockerGit

Preferred skills

problem-solvingcommunication

Specializations

reliabilitycloudautomationmonitoringsecurity
Locations

Structured locations inferred from the posting.

Taipei, Taiwan

On-site City