| 5 Visite |
0 Candidati |
Descrizione del lavoro:
Enterprise Technology Services (ETS) is part of IS&T and delivers global-scale platforms and services that keep Apple's operations secure and running. The team manages identity, device security, and anti-abuse platforms covering everything from manufacturing and repairs to software updates and activations. ETS also oversees supply chain, manufacturing, and partner integration platforms, protecting data on more than 2.5 billion devices worldwide. And when Apple prepares for a global product launch, ETS owns the systems that ramp factory production, managing serial numbers, network credentials, and verified software. The Insight team runs one of Apple's most critical Big Data ecosystems, a multi-petabyte, highly-available infrastructure that underpins manufacturing operations for every Apple product, globally. Every iPhone, iPad, and Mac has touched our systems. We are establishing a new EMEA operational team in Cork, not only as an extension of our US/India/China operations, but also a team with its own identity, perspective, and contribution to how Insight runs at global scale. This is a new regional team being built from the ground up, and the people who join early will help shape what it becomes. We are looking for a Site Reliability Engineer who has operated in the European high-tech landscape, who brings the rigor and perspective that comes from that environment, and who wants to shape the practices of a new team. As one of the first few SREs in Cork, you will have real influence over how this team works, what it builds, and how it contributes to Apple's manufacturing infrastructure worldwide.
You will operate and improve services within a very large-scale, highly-available Big Data ecosystem supporting Exabytes level of data with sustained, rapid growth. Your work directly enables the engineering and operations teams that build every Apple product. Own the reliability, performance, and scalability of services and infrastructure within the Insight ecosystem - with a build-to-manage mindset applied to design, deployment, and ongoing operations Build and advance AIOps capabilities across the ecosystem - developing ML-driven alerting, anomaly detection, LLM-assisted operational tooling, and automated incident triage that measurably improve observability and response Instrument services for deep observability: dashboards, meaningful alerts, runbooks, and clearly defined SLOs and SLIs that give the team and stakeholders an accurate view of ecosystem health Drive automation that reduces operational toil and improves ecosystem stability - key measures of success include ecosystem uptime, release quality, toil reduction, and sustained advancement of SRE practice Partner with incident management to drive effective incident response and conduct thorough post-incident reviews; translate findings into architectural and operational improvements that prevent recurrence Contribute to infrastructure-as-code practices and platform architecture decisions that shape the long-term direction of the Insight ecosystem Collaborate with cross-functional engineering teams across Apple's global manufacturing services, communicating effectively across time zones and cultures Help define the culture, working practices, and technical standards of the new EMEA operational team - build strong in region relationships with business partners and represent the team's perspective within the wider Insight organization
Experience in SRE, DevOps, or platform engineering - with strong communication skills, able to document clearly, lead technical discussions, and collaborate effectively across geographies Strong Linux systems proficiency - comfortable troubleshooting at the systems level including processes, networking, file systems, and performance diagnostics Programming and automation experience in Python, Go, or Bash, combined with practical exposure to AI/ML applied to IT operations - anomaly detection, intelligent alerting, or LLM-driven tooling Hands-on experience with cloud platforms such as AWS or GCP, with solid understanding of compute, storage, networking, and IAM in production environments
Kubernetes, CI/CD, and networking fundamentals - practical experience with container orchestration and deployments, pipeline tooling (ArgoCD, Jenkins, GitHub Actions, or similar), and a solid grasp of HTTP, DNS, TCP/IP, and how traffic flows through distributed systems Monitoring, observability, and infrastructure-as-code - experience operating systems instrumented with Grafana, Prometheus, Kibana, or equivalent, alongside practical IaC skills using Terraform, Pulumi, or Crossplane Exposure to big data technologies in production or substantial project contexts - Kafka, Druid, Elasticsearch, or Object Storage SRE principles applied in practice: error budgets, SLOs, SLIs, toil measurement and reduction Experience with relational databases in production environments - MySQL or PostgreSQL Configuration management, API design, and open-source practice - experience with tools such as Ansible, Chef, or Puppet; familiarity with Open API and microservice architectures; contributions to open-source projects are a meaningful signal Experience working across multiple geographies and cultures in EMEA-based technology organizations BS or MS in Computer Science, Software Engineering, or equivalent technical discipline; equivalent professional experience will be fully considered
| Provenienza: | Web dell'azienda |
| Pubblicato il: | 21 Lug 2026 |
| Tipo di impiego: | Lavoro |
| Settore: | Elettronica di consumo |
| Lingue: | Inglese |