HOME Systems Administration & Operations Operations Engineer
  • Apple
  • Shanghai, Shanghai
  • Full-Time
  • <72 Hours
Apple VERIFIED EMPLOYER

Operations Engineer.

Systems Administration & Operations Full-Time

Operations Engineer: our view in 3 lines...

  • The Role:This role is for an operations engineer supporting technical infrastructure in a crypto services environment.
  • The Person:The person will monitor infrastructure and application services, manage incidents from initial contact to resolution, troubleshoot issues, document problems, perform system patching and upgrades, and build observability tooling.
  • Requirements:The ideal candidate has 5+ years of expertise with Linux, standard UNIX utilities, SRE principles, on-call experience, Puppet, Chef, Ansible, Bash, Python, Icinga/Nagios, and Splunk or OpenSearch.

About the role

The Operations Engineer in Crypto Services team manages key technical infrastructure. An ideal candidate will have experience in Systems Administration. The Operations Engineer will monitor infrastructure and application services and drive incident management. The Ops engineer will work closely with SRE's, PKI Engineers, systems engineers, network engineers, database administrators and information security teams to effectively ensure availability and reliability. For this position, strict application security and high availability requirements must be balanced to achieve optimal solutions.

Description

The successful candidate will participate in troubleshooting issues following established procedures, documenting problems, managing incidents and owning the issue from the initial contact to resolution. Additionally the engineer will be responsible for critical compliance tasks, system patching and upgrades, and building out observability tooling to support the team's operations.

Minimum Qualifications

5+ years of expertise with Linux (any distro, but especially RHEL), including experience with OS patching and upgrade cycles. Standard UNIX utilities and programs
Strong understanding of SRE principles and goals, coupled with prior on-call experience, including leading incident command and management efforts.
Proficiency with configuration management tools (e.g., Puppet, Chef, Ansible) and scripting languages (e.g., Bash, Python).
Experience with monitoring tools (e.g., Icinga/Nagios) and log aggregation platforms (e.g., Splunk, OpenSearch) , and building dashboards/alerts
A proven track record of practical problem-solving, coupled with excellent communication and documentation skills, particularly in conveying complex technical information to diverse audiences.

Preferred Qualifications

Experience with distributed teams or managing operations across multiple time zones.
Working knowledge of cloud platforms (AWS/GCP/AliCloud) and container technologies (Podman, Docker, Kubernetes).
A deep understanding of security and compliance best practices, especially within a cryptographic or PKI environment.
Experience supporting Java applications and/or managing critical hardware like Hardware Security Modules (HSMs).
Flexibility to travel for work

Published September 29, 2026
Location Shanghai, China
Job Type Full-Time