Site Reliability Engineer
Back to JobsApple Get Smart Job AI Coach in the appFree on iOS and Android 




Site Reliability Engineer
Location
Bengaluru, Karnataka, 560092, India
Experience
Mid
Posted
Jul 22, 2026
Apply by
August 21, 2026
Applicants
0
Early applicantEasy applyFull-timeWork from Office
Job Description
Do you want to help build some of the largest and most consequential enterprise and customer technology systems in the world? Join Apple’s Information Systems and Technology (IS&T) organization. IS&T is the engine behind everything Apple does for customers and for the people who build for them. It’s Apple’s central nervous system. Supporting 2.5 billion active Apple devices, processing billions of secure transactions, and keeping the technology that defines modern life running flawlessly, IS&T makes the impossible feel effortless.Do you love building solutions to handle global complexity and immense scale? Imagine what you could do here.
Enterprise Technology Services (ETS) is part of IS&T and delivers global-scale platforms and services that keep Apple's operations secure and running. The team manages identity, device security, and anti-abuse platforms — covering everything from manufacturing and repairs to software updates and activations. ETS also oversees supply chain, manufacturing, and partner integration platforms, protecting data on more than 2.5 billion devices worldwide. And when Apple prepares for a global product launch, ETS owns the systems that ramp factory production — managing serial numbers, network credentials, and verified software.
The Insight team runs one of Apple's most critical Big Data ecosystems — an exabyte-
scale, highly-available infrastructure that underpins manufacturing operations for every
Apple product, globally. Every iPhone, iPad, and Mac has touched our systems.
We advance technology by relying on each other's strengths and skills to build something
bigger than ourselves. For this reason, team culture is central to our values. We value
social skills and integrity as much as technical craft.
## Description
We are looking for an extraordinary DevOps engineer with experience building large-
scale data platforms, analytics tools, and solutions that can take our environment to the
next level. You thrive in a high-demand, fast-moving setting: you prioritize well, deliver
ahead of schedule, work independently, and raise the people around you. You will operate and improve services within a very large-scale, highly available Big Data
ecosystem supporting Exabytes level of data with sustained, rapid growth. Your work
directly enables the engineering and operations teams that build every Apple product.
## Responsibilities
Own the reliability, performance, and scalability of services and infrastructure within the Insight ecosystem — with a build-to-manage mindset applied to design, deployment, and ongoing operations
Drive automation that reduces operational toil and improves ecosystem stability Instrument services for deep observability: dashboards, meaningful alerts, runbooks, and clearly defined SLOs and SLIs that give the team and stakeholders an accurate view of ecosystem health
Partner with incident management to drive effective incident response and conduct thorough post-incident reviews; translate findings into architectural and operational improvements that prevent recurrence
Build and advance AIOps capabilities across the ecosystem — developing AI- driven alerting, anomaly detection, LLM-assisted operational tooling, and automated incident triage that measurably improve observability and response
Collaborate with cross-functional engineering teams across Apple's global manufacturing services, communicating effectively across time zones and cultures
## Minimum qualifications
5+ years of experience in SRE, DevOps, or platform engineering, with proficiency
in Python or Java.
Experience with AWS or GCP, including big-data or streaming technologies such as Kafka, Elasticsearch, Redis, or Bigtable.
Experience applying AI/ML or LLM techniques to development or operations — such as anomaly detection, predictive alerting, or LLM/agent-driven tooling.
## Preferred qualifications
Experience with Distributed systems & big-data streaming: Microservices, APIs, and messaging/streaming (Kafka, Solace, Pub/Sub) at scale, plus production big-data/storage technologies (Druid, Elasticsearch, ClickHouse, or object storage).
Experience with Cloud-native delivery: Containerization and orchestration (Docker, Kubernetes)
and CI/CD or GitOps pipelines (ArgoCD, Jenkins, GitHub Actions, or similar).
Implementing and operating observability systems (Grafana, Prometheus, Kibana, or equivalent) instrumentation, dashboard design, and alert tuning.
BS/MS in Computer Science, Software Engineering, or equivalent experience; production relational databases (MySQL, PostgreSQL)multi-region operations and data-residency awareness; and a creative, flexible approach to problem-solving.
Key Responsibilities
- Own the reliability, performance, and scalability of services and infrastructure within the Insight ecosystem.
- Drive automation to reduce operational toil and improve ecosystem stability.
- Instrument services for deep observability including dashboards, alerts, runbooks, and SLOs/SLIs.
- Partner with incident management to drive effective incident response and conduct post-incident reviews.
- Build and advance AIOps capabilities including AI-driven alerting, anomaly detection, and automated incident triage.
- Collaborate with cross-functional engineering teams across global manufacturing services.
Skills Required
PythonJavaAWSGCPKafkaElasticsearchRedisBigtableAI/MLLLMCommunicationProblem solvingIndependenceCollaborationMicroservicesAPIsMessagingStreamingDruidClickHouseObject storageDockerKubernetesCI/CDGitOpsArgoCDJenkinsGitHub ActionsGrafanaPrometheusKibanaMySQLPostgreSQLCreative approach to problem-solvingFlexible approach to problem-solving
App exclusive · Free
Smart Job AI Coach
Your personal interview coach on every job — readiness tips, profile improvements, and role-specific prep. Available only in the Pulse Job app.
Interview readiness
See how prepared you are and what to improve for each role.
Personalized tips
Actionable suggestions based on your profile and the job.
After you apply
Keep coaching momentum from job detail through application success.
Similar roles for you
Matched using this role's title and skills. Open the job search anytime to see every listing.

Senior Site Reliability Engineer
Valtech
Full-timeEasy applyWork from Home
North Macedonia - RemoteSenior

DevOps Engineer
StackAdapt
Full-timeEasy applyWork from Home
CanadaSenior

Principal Data Engineer
JobGet
Full-timeEasy applyWork from Home
RemoteSenior

Senior TIBCO DevOps Engineer-10881844
ITProposal
ContractEasy applyHybrid
AmsterdamSenior
Full Stack Java Developer
Up2date Technologies LLC
$70–$90 / Hour
ContractEasy applyHybrid
Atlanta

Software Engineer
Ford Pro
Full-timeEasy applyWork from Office
ChennaiSenior