Prepin
Log in
Apple

engineering opportunity

Site Reliability Engineer - Kafka

You will build and maintain next-generation Kafka infrastructure and platform services to ensure high reliability and scalability. This involves collaborating cross-functionally to develop automation, monitoring tools, and deployment architectures for large-scale distributed systems.

Seattle, Washington, United StatesonsiteFULL_TIME

Posted

About the role

What will you do at Apple?

The Apple Service Engineering – Data Streaming SRE team is looking for Site Reliability Engineers with experience developing processes, tools, and automation for managing distributed systems in production environments. Our SRE team combines software engineering, systems engineering, and Devops practices to build and run large-scale, massively distributed, fault-tolerant systems.

Our software ensures that Apple's services are reliable, scalable, and secure, and we leverage both open-source and homegrown technologies to provide managed data infrastructure services. You will help build next-generation Kafka infrastructure and platform services, collaborating cross-functionally with various ASE teams—from store and commerce to search and recommendations. You'll create platforms that can rapidly scale to serve data with very low latencies.

You should be someone who isn't afraid to question assumptions, thrives as a collaborative partner under tight deadlines, and tackles complex problems with elegant technical solutions. DESCRIPTION The Data Service SRE team develops applications and tooling that are safe, reliable, scalable, and fast. This work requires an innovative spirit and an extraordinary degree of care and difficulty in engineering.

Team members contribute to all major components of Kafka deployment infrastructure, including maintenance automation, control plane enhancements, monitoring and alerting tooling/dashboards, advanced deployment architecture, focused on safety, stability, performance, and scaling. Come join us at Apple Services Engineering and help us deliver services and applications that are fluid and responsive.

You will collaborate with engineers from across Apple to define the metrics, set targets, uncover optimization opportunities, and ship a service that will delight our customers. This role is for engineers who enjoy deep technical engineering that spans large cross-organizational projects. Your openness to learning and implementing new technologies will contribute to the continuous evolution of our organization.

Good ideas are valued and rewarded. MINIMUM QUALIFICATIONS Support of internet-facing production services and distributed systems via deployments, On Call and Incident Management.

Experience running large scale infrastructure with a heavy reliance on automation tooling Excellent troubleshooting and performance deep dive analysis Real operational experience managing services at scale on Kubernetes Proficient in one or more of the following programming languages: Java, Go (golang), Python Operational experience deploying in and running on Datacenter and Cloud architectures (networking topologies, host placement strategies, and failure modes); design of multi-datacenter systems; failure domains; and wide-area networking.

Self motivated, inquisitive with an aptitude to learn new technologies quickly and effectively. Demonstrated expertise developing and troubleshooting distributed systems and database storage engines. Experience developing critical internet services and/or platform infrastructure.

Experience with AWS, GCP and IaC such as Terraform PREFERRED QUALIFICATIONS Experience managing messaging services such as Kafka or other Data services Proficient in Java, Go (golang) & Python

Which skills does this role require?

Site Reliability EngineeringDistributed SystemsAutomationKubernetesCloud ArchitectureAWSGCPTerraformInfrastructure as CodeIncident ManagementPerformance AnalysisNetworkingSite Reliability EngineerDevOpsCloudMonitoringAlertingData StreamingScalabilityFault-tolerantDatabase Storage EnginesControl PlaneDeployment Architecture

Make your next move

Build a shortlist and prepare

Identify the requirements you can demonstrate, then choose examples from your work to discuss with the hiring team.

Review the responsibilities and requirements before adding an opening to your shortlist.

Role information can change. Confirm current details on the original application page.

Product

AI Candidate AgentCompaniesBrowse JobsDeep ProfileSkill AssessmentOpportunity Matching
Prepin.ai

© 2026 Prepin | All rights reserved.