- Apple
- Shanghai, Shanghai
- Full-Time
- 10 days ago
Software Engineer, Data Services, IS&T Ai & Data Platforms.
Before you go
Before you leave us, sign up for our email alerts
We don't do job spam, just the best digital jobs delivered straight to your inbox.
Software Engineer, Data Services, IS&T Ai & Data Platforms: our view in 3 lines...
- The Role:This role is for a senior database systems engineer on a data services SRE team supporting transactional database platforms at scale.
- The Person:The person will build platform tooling for provisioning, backup and restore, observability, and self-service while supporting hybrid-cloud database migrations and incident management.
- Requirements:The ideal candidate has more than 7 years in an SRE or infrastructure-focused role, hands-on production experience with Oracle, PostgreSQL, MongoDB or CockroachDB, and good understanding of Python or Go.
About the role
AI & Data Platforms (AiDP) is IS&T's engine for AI-powered innovation. The team brings together data, application development, and machine learning — including generative AI — along with data services and customer success functions, to help IS&T build solutions more efficiently and streamline the adoption and embedding of generative AI across Apple.
Apple's AI & Data Platform (AiDP) Data Services organization is seeking a motivated database systems engineer to join our Data Services SRE team, focused on our transactional database fleet. Engineers on this team develop and contribute to platform tooling that manages relational and distributed-SQL/document engines powering some of Apple's most critical internet services at massive scale. In AiDP, your work will benefit hundreds of millions of users.
Description
The AiDP Data Services SRE team builds engine-agnostic platform capabilities — provisioning, backup/restore, observability, and self-service — spanning Oracle, PostgreSQL, MongoDB, and CockroachDB. This role involves supporting data migration efforts across hybrid-cloud environments (on-prem AWS Ali Cloud) with minimal downtime and guaranteed data integrity, following established runbooks and adhering to local data residency and regulatory requirements (e.g., China's Cybersecurity Law and PIPL). This role requires good communication and collaboration with Core Storage teams and colleagues across a distributed team.
Minimum Qualifications
BS or MS in Computer Science / related fields or equivalent work experience, with more than 7 years in a Site Reliability Engineering / Infrastructure focused role.
Hands-on production experience with at least one of Oracle, PostgreSQL, MongoDB, or CockroachDB, including exposure to On Call and Incident Management.
Exposure to running infrastructure with reliance on automation tooling, across Datacenter and Cloud architectures (including Alibaba Cloud/Ali Cloud and/or AWS).
Solid troubleshooting skills, a resourceful first-principles approach to problem solving, and good technical writing habits.
Good understanding in one or more of the following programming languages: Python, Go.
Preferred Qualifications
Some operational experience with stateful services on Kubernetes (operators, StatefulSets, CSI storage).
Exposure to database-as-a-service control planes, provisioning APIs, or self-service tooling.
Experience assisting with cross-cloud / hybrid-cloud data migrations for transactional systems.
Familiarity with observability concepts (SLIs/SLOs), backup, and DR practices for OLTP/distributed-SQL engines.
Proficiency in Mandarin (spoken and written), to support coordination with local Ali Cloud teams and regional vendors/partners.
Contributions to open-source database projects or internal data-platform tooling.

