Padmi
Amazon logo
Amazon

e-commerce marketplace · cloud computing (AWS)

Systems Engineer MacOS, Client Engineering

SeattlePosted 2 months ago
InfrastructureUnspecified
Apply at Amazon

Opens the source posting on amazon.jobs

Source description

About the role

View original

Enterprise Engineering organization manages one of the largest macOS client fleets in the world. Our focus is to consistently raise Amazon's security bar while ensuring a smooth productivity experience for our end users. We are a team of builders managing the solutions for Amazon's global macOS client fleets. We use AWS products to evolve traditional enterprise tools and services at a large scale.

We are looking for a Systems Engineer with deep macOS troubleshooting expertise and hands-on fleet management experience. This role is heavily focused on diagnosing complex macOS issues at enterprise scale, maintaining fleet health, and driving operational excellence across our global device estate. If you thrive in high-pressure troubleshooting scenarios, enjoy root-causing elusive system issues, and have a passion for keeping large fleets running smoothly — this is the role for you.

In this role, you and your team will directly influence Amazon's macOS experience roadmap. You will manage, monitor, and maintain a growing fleet of macOS client devices using MDM platforms, native Apple tooling, and AWS-hosted infrastructure. This is a hands-on operational role where your daily activities center around macOS troubleshooting, fleet health management, incident response, performance analysis, and continuous improvement of our management stack.

Key job responsibilities Troubleshoot and resolve complex macOS issues (system crashes, kernel panics, performance degradation, network connectivity, authentication failures) in a globally scaled enterprise environment

Manage macOS fleet operations — OS patching, software deployments, configuration profile management, and MDM policy enforcement across tens of thousands of devices

Monitor fleet health — proactively identify trends, anomalies, and emerging issues through telemetry, dashboards, and alerting systems

Drive incident response — lead triage for fleet-impacting events, coordinate cross-functional resolution, and produce root cause analyses (RCAs)

Maintain and improve client management service infrastructure hosted in AWS

Develop runbooks and knowledge base documentation for recurring issues and escalation procedures

Collaborate with engineering teams on OS upgrade rollouts, compatibility testing, and risk mitigation strategies

Contribute to automation of repetitive operational tasks to improve fleet reliability and reduce manual intervention

More at Amazon

Related open roles

View all roles