
Job Description & Key Responsibilities
Apple is a global technology leader known for its innovative products such as the iPhone, iPad, Mac, Apple Watch, and Apple TV. The company’s Information Systems and Technology (IS&T) organization is the backbone that powers everything Apple does for customers and for the people who build for them. IS&T supports 2.5 billion active Apple devices, processes billions of secure transactions, and keeps the technology that defines modern life running flawlessly.
The Sales and Operations Engineering team within IS&T drives the technology behind Apple’s global operations, sales, and supply chain. It connects the systems that move products from factory to customer, bridging sales platforms with the operational infrastructure that keeps Apple running at scale.
As a Junior Application Support & Site Reliability Engineer, you will join the APS & SRE team that ensures mission‑critical enterprise systems and customer‑facing cloud services operate with high availability, resilience, and performance at massive scale. Your day‑to‑day responsibilities will include monitoring multi‑region environments, triaging production incidents, debugging complex application and database bottlenecks, and implementing permanent fixes. You will partner with development, QA, database, and infrastructure teams to eliminate operational toil through automation and manage platform health using modern observability and emerging GenAI technologies.
Key responsibilities include:
1. Monitoring health, performance, and capacity of mission‑critical applications across multi‑region environments.
2. Acting as first‑line responder to production alerts, outages, and anomalies, participating in on‑call rotations.
3. Investigating and debugging complex application, core system, network, and performance bottlenecks.
4. Performing database‑level troubleshooting, query analysis, and performance tuning on PostgreSQL, Oracle, and MySQL.
5. Conducting root‑cause analysis and implementing permanent fixes in collaboration with cross‑functional teams.
6. Building and refining telemetry dashboards, log aggregations, and alerting rules using Splunk, Prometheus, Grafana, and Dynatrace.
7. Tracking and upholding key reliability metrics (SLIs/SLOs) to minimize Time to Detect and Time to Mitigate.
8. Developing scripts and tools in Python and Bash to automate repetitive operational tasks.
9. Supporting production deployments, CI/CD pipeline runs, and canary/smoke testing.
10. Executing production change requests in accordance with enterprise change management standards.
Tech stack: Linux/Unix, systemd, networking, Java, Python, Maven, Gradle, PostgreSQL, Oracle, MySQL, Splunk, Grafana, Prometheus, Dynatrace, Jenkins, GitHub Actions, CI/CD, GenAI.
Growth path: Starting as a Junior SRE, you can progress to SRE, Senior SRE, Lead SRE, and eventually to SRE Manager or Platform Engineering Manager roles. Apple’s culture of continuous learning, mentorship, and cross‑team collaboration accelerates career growth.
Why join Apple? You’ll work on world‑class systems that power billions of devices, collaborate with top engineers, and have access to cutting‑edge tools and technologies. Apple’s commitment to innovation, inclusion, and employee well‑being ensures a rewarding and balanced work experience.