Principal Lead Distributed Systems Engineer - Services Special Projects
AI summary of the role
Lead (Principal) distributed systems engineer on Apple's Special Services Team, building and operating high-throughput, low-latency backend services for mission-critical workloads including real-time transactions, analytics, and content delivery.
What you’ll do
- Design and build massively scalable, highly available services for Apple customers
- Operate high-throughput, low-latency backend services for real-time transactions, analytics, and content delivery
- Drive evolution of a multi-tenant platform including AI/ML-powered services
- Ship new capabilities and scale existing services from design through production
What you’ll bring
- Master's degree in Computer Science or related field
- 15+ years professional software development building scalable distributed systems, with at least 5 in Principal or Sr. Staff role
- Strong proficiency in Java and application frameworks (Spring Boot)
- Hands-on JVM performance tuning and profiling (GC, JFR, async-profiler)
Technologies
Java · Spring Boot · JVM · Netty · Project Reactor · Kafka · Cassandra · Redis · Iceberg · gRPC · Kubernetes · Docker
About Apple
Designs and sells iPhone, Mac, iPad, Watch, and Vision Pro plus a $100B+ Services layer (App Store, iCloud, Apple Pay, Music, TV+) atop a 2.5B device installed base.
Public
Source and classification
Production engineering · Evidence for this classification:
Apple's Special Services Team is seeking a Lead (Principal) Distributed Systems Engineer to design and build massively scalable, highly available services that power experiences for Apple customers both now and in the future. Description In this Lead role, you will build and operate high-throughput, low-latency backend services that ingest, process, and serve data at scale across a range of mission-critical workloads — from real-time transactions to analytics and content delivery. You'll also drive the evolution of a multi-tenant platform, including AI/ML-powered services, by shipping new capabilities, scaling what exists, and applying distributed-systems best practices from design through production. Minimum Qualifications Master's degree in Computer Science or a related field 15+ years of professional software development experience building scalable, distributed systems in
More from the job description
Apple's Special Services Team is seeking a Lead (Principal) Distributed Systems Engineer to design and build massively scalable, highly available services that power experiences for Apple customers both now and in the future. Description In this Lead role, you will build and operate high-throughput, low-latency backend services that ingest, process, and serve data at scale across a range of mission-critical workloads — from real-time transactions to analytics and content delivery. You'll also drive the evolution of a multi-tenant platform, including AI/ML-powered services, by shipping new capabilities, scaling what exists, and applying distributed-systems best practices from design through production. Minimum Qualifications Master's degree in Computer Science or a related field 15+ years of professional software development experience building scalable, distributed systems in production, with at least 5 in a Principal or Sr. Staff Level role. Experience building, authoring, and operating large-scale, multi-tiered distributed systems and customer-facing web services: including API design, authentication, authorization, scaling for high availability, concurrency, and reliability. Strong understanding of concurrency and multi-threaded programming, fundamental data structures, and efficient algorithm design Strong proficiency in Java; working knowledge of a second systems langu [... source excerpt omitted ...] nd RPC frameworks (gRPC) Experience building and maintaining CI/CD pipelines (e.g., Jenkins, GitHub Actions, GitLab CI, or similar) for automated testing, build, and deployment of production services. Experience with AWS or GCP and cloud-native tooling (Docker, Kubernetes) in the context of deploying scalable production grade services. Experience with event streaming and queueing systems (specifically Kafka) and stream processing frameworks and high-throughput, append-only write paths for durable, queryable historical records. Hands on Experience of leveraging data storage (Iceberg, Cassandra) and caching technologies (Redis) in Production services Hands on experience with Serializ [... source excerpt omitted ...] ycle development experience for a consumer product, from concept through deployment Proven history of presenting technical and business concepts to Executive Leadership Preferred Qualifications Self-motivated, with strong collaboration and communication skills, and experience in a fast-paced, agile environment Experience with machine learning systems, ML frameworks, libraries and algorithms Familiarity with deployment and optimization of Large scale Production grade AI Services that require GPUs in the path of the transaction. Hands-on experience deploying, serving, and optimizing LLMs or ML models directly in the transaction/request path Experience with security and cryptography (e.
Employer postings · Data from · Sources