Job Description
Apple's AI & Data Platform (AiDP) Data Services organization is seeking a motivated database systems engineer to join our Data Services SRE team, focused on our transactional database fleet. Engineers on this team develop and contribute to platform tooling that manages relational and distributed-SQL/document engines powering some of Apple's most critical internet services at massive scale. In AiDP, your work will benefit hundreds of millions of users.
AI & Data Platforms (AiDP) is IS&T's engine for AI-powered innovation. The team brings together data, application development, and machine learning — including generative AI — along with data services and customer success functions, to help IS&T build solutions more efficiently and streamline the adoption and embedding of generative AI across Apple.
The AiDP Data Services SRE team builds engine-agnostic platform capabilities — provisioning, backup/restore, observability, and self-service — spanning Oracle, PostgreSQL, MongoDB, and CockroachDB. This role involves supporting data migration efforts across hybrid-cloud environments with minimal downtime and guaranteed data integrity, following established runbooks and adhering to local data residency and regulatory requirements. This role requires good communication and collaboration with Core Storage teams and colleagues across a distributed team.
Preferred Qualifications
Some operational experience with stateful services on Kubernetes
Solid troubleshooting skills, a resourceful first-principles approach to problem solving, and good technical writing habits
Exposure to database-as-a-service control planes, provisioning APIs, or self-service tooling
Experience assisting with cross-cloud / hybrid-cloud data migrations for transactional systems
Familiarity with observability concepts (SLIs/SLOs), backup, and DR practices for OLTP/distributed-SQL engines
Solid troubleshooting skills, a resourceful first-principles approach to problem solving, and good technical writing habits
Minimum Qualifications
3 years of experience in a Site Reliability Engineering / Infrastructure focused role
Hands-on production experience with at least one of Oracle, PostgreSQL, MongoDB, including exposure to On Call and Incident Management
Exposure to running infrastructure with reliance on automation tooling, across Datacenter and Cloud architectures (third party cloud)
Good understanding in one or more of the following programming languages: Python or Go
BS or MS in Computer Science / related fields or equivalent work experience