Padmi
Microsoft logo
Microsoft

cloud computing (Azure) · AI and machine learning (Copilot, CoreAI)

Senior Site Reliability Engineer - Azure Storage

Seattle · OnsitePosted 10 days ago
InfrastructureSeniorFull TimeH-1B track record
Apply at Microsoft

Opens the source posting on apply.careers.microsoft.com

Source description

About the role

View original

Lead end-to-end qualification and production readiness validation for new Azure Storage hardware SKUs, platforms, SSDs, networking devices, and firmware releases Drive qualification strategy, test planning, execution, and sign-off recommendations for Azure Storage deployments Serve as a technical escalation point for qualification failures, performance regressions, reliability issues, and production-readiness concerns impacting Azure Storage deployments. Own complex investigations across software, hardware, firmware, and platform stacks to identify and resolve qualification blockers Design and develop scalable automation frameworks that improve qualification coverage, efficiency, quality, and engineering productivity. Analyze large-scale telemetry, benchmark data, and fleet health signals to identify performance trends, bottlenecks, risks, and optimization opportunities. Monitor qualification pipelines and proactively identify risks impacting production onboarding, deployment schedules, and service readiness. Perform root cause analysis for hardware, firmware, platform, and software issues discovered during qualification and drive corrective actions to closure. Validate SKU readiness against Azure Storage performance, reliability, scalability, and operational requirements Partner with Storage, Compute, Networking, Platform, and vendor engineering teams to resolve critical qualification, reliability, and performance issues. Influence qualification standards, test methodologies, performance thresholds, and operational best practices across Azure Storage NPI programs Bachelor's Degree in Computer Science, Engineering, or related field AND 4+ years technical experience in software engineering, network engineering, or systems administration OR equivalent experience Bachelor Degree in Computer Science, Engineering, or related field AND 6+ years technical experience in service engineering, performance engineering, or site reliability engineering 6+ years technical experience working with large-scale cloud or distributed systems

More at Microsoft

Related open roles

View all roles