Senior Systems Monitoring Engineer

BloKchain Talent

Apply to this job
Tempe, AZ, us on site Until 8/22/2026 First posted March 21, 2025 Last posted March 21, 2025
Job description

Our Client, is an award-winning workplace. They have been recognized by Comparably as #1 CEO, Company Happiness, Benefits, Compensation, Diversity, and more! Not to mention they’ve been awarded by Glassdoor as the 2nd Best US workplace & Best Large Company US CEO in 2018, Wealthfront, and Business Insider. They culture focuses on delivering happiness, our commitment to transparency, and the tangible benefits we provide our employees and our customers.

POSITION TITLE: Senior Systems Monitoring Engineer 
LOCATION: Phoenix AZ 
SALARY: Based on Experience 
SPONSORSHIP: No 

Job description:

  • Responsible for application monitoring and alarming of the production environment of Zoom global real-time online conference system, high-availability target of 99.99%, continuously find and fix problems, and ensure the stable operation of the business.
  • Ability to develop monitoring plan/system and implement maintenance in accordance with the company's product system architecture and business logic as required
  • Have excellent self-learning ability, be able to read English documents related to products and technologies, pay attention to the development of open source software, and be able to with stand high work pressure
  • Participate in the construction and continuous improvement of Zoom's global operation and maintenance system, be proficient in writing relevant operation and maintenance technical documents, and continuously improve the operation and maintenance system and process.

Job Requirements:

  • Bachelor’s degree or above, computer related major, at least three years of experience in large-scale website system operation and maintenance
  • Strong skills on some of popular monitoring systems, such as: Prometheus / Kubernetes / Grafana / Filebeat / ELK(Elasticsearch+Logstash +Kibana) / Zabbix.
  • Be good at one of the programming languages: Shell, Python,Java etc.
  • Experiences on open source software such as Nginx, Tomcat, Apache, Memcache, Redis, MySQL, experiences on system High availability, Fail-over mechanism , Load balancing.
  • Experiences on Amazon service components, such as: Awscli, S3, EC2, Route53, RDS, Cloudwatch, DymamoDB, etc.
  • Familiar with and master the use of automated operation and maintenance tools such as Ansible, Jenkins, etc., with actual large-scale (1000+) server operation experience is preferred.
  • Language requirement: English, Mandarin is plus

 

Job Requirements:

  • Bachelor’s degree or above, computer related major, at least three years of experience in large-scale website system operation and maintenance
  • Strong skills on some of popular monitoring systems, such as: Prometheus / Kubernetes / Grafana / Filebeat / ELK(Elasticsearch+Logstash +Kibana) / Zabbix.
  • Be good at one of the programming languages: Shell, Python,Java etc.
  • Experiences on open source software such as Nginx, Tomcat, Apache, Memcache, Redis, MySQL, experiences on system High availability, Fail-over mechanism , Load balancing.
  • Experiences on Amazon service components, such as: Awscli, S3, EC2, Route53, RDS, Cloudwatch, DymamoDB, etc.
  • Familiar with and master the use of automated operation and maintenance tools such as Ansible, Jenkins, etc., with actual large-scale (1000+) server operation experience is preferred.
  • Language requirement: English, Mandarin is plus

All your information will be kept confidential according to EEO guidelines.

About this role

Summary

Manage application monitoring, develop systems, ensure high availability, and improve operations for large-scale systems.

Job title

Senior Systems Monitoring Engineer

Experience level

3+ years

Industry

technology

Location requirements

Tempe, AZ, US; remote not allowed

Salary

Not specified

Management role

No

Skills & keywords

Required skills

PrometheusKubernetesGrafanaFilebeatELKZabbixShellPythonNginxTomcatApacheMemcacheRedisMySQLAwscliS3EC2Route53RDSCloudwatchDynamoDBAnsibleJenkins

Preferred skills

large-scale server operationEnglishMandarin

Specializations

monitoring systemscloud servicesopen source softwarehigh availabilityautomation tools
Locations

Structured locations inferred from the posting.

Tempe, AZ, USA

On-site City
Related searches