Padmi
Microsoft logo
Microsoft

cloud computing (Azure) · AI and machine learning (Copilot, CoreAI)

Senior Site Reliability Engineer - CTJ - Top Secret

Seattle · Washington DC · OnsitePosted 2 months ago
InfrastructureSeniorFull TimeH-1B track record
Apply at Microsoft

Opens the source posting on apply.careers.microsoft.com

Source description

About the role

View original

Support and Automate Deployments: Execute and improve manual operations and deployments for our products, while designing automation to scale and streamline those processes across environments. Build Scalable Systems: Develop automation for monitoring, alerting, debugging, and deployment to reduce manual effort and accelerate safe, reliable delivery. Lead Post-Incident Learning: Conduct postmortems, share insights, and implement solutions that prevent recurrence—fostering a culture of learning and continuous improvement. Collaborate Across Teams: Partner with engineering and product teams to align reliability goals with customer needs and deliver seamless user experiences. Stay Ahead Technically: Continuously invest in your technical growth to improve system availability, observability, and performance at scale. Master's Degree in Computer Science, Information Technology, or related field AND 2+ years technical experience in software engineering, network engineering, or systems administration OR Bachelor's Degree in Computer Science, Information Technology, or related field AND 4+ years technical experience in software engineering, network engineering, or systems administration These requirements include, but are not limited to the following specialized security screenings: Candidates must have an active TS and be willing and eligible to upgrade to TS/SCI (with polygraph) or have an active TS/SCI and be willing and eligible to upgrade to TS/SCI (with polygraph). This role will require candidates to maintain the TS/SCI (with polygraph) clearance. Failure to maintain or obtain the appropriate clearance and/or customer screening requirements may result in employment action up to and including termination. Clearance Verification: This position requires successful verification of the stated security clearance to meet federal government customer requirements. You will be asked to provide clearance verification information prior to an offer of employment. Doctorate Degree in Computer Science, Information Technology, or related field AND 3+ years technical experience in software engineering, network engineering, or systems administration OR Master's Degree in Computer Science, Information Technology, or related field AND 6+ years technical experience in software engineering, network engineering, or systems administration OR Bachelor's Degree in Computer Science, Information Technology, or related field AND 8+ years technical experience in software engineering, network engineering, or systems administration OR equivalent experience. 3+ years technical experience working with large-scale cloud or distributed systems. Demonstrated experience applying software engineering principles to production systems, including designing, building, or improving services and platforms. Proficiency in one or more programming languages such as C#, Go, Java, or Python, with the ability to develop and maintain production-quality code. Experience with automation that results in measurable improvements (e.g., reduced toil, fewer manual steps, improved system reliability). Experience with debugging and troubleshooting complex distributed systems in production environments. Ability to independently identify problems and implement solutions that improve system reliability and operational efficiency. Hands-on experience with CI/CD pipelines, testing, deployment, and reliability tooling.

More at Microsoft

Related open roles

View all roles