yoinka

Senior Linux SRE (Site Reliability Engineer)

Microsoft

United States, Oregon, Hillsboro; United States, California, Mountain View; United States, North Carolina, RaleighSeniorH-1B sponsor company
Sign in to applyVerified 1h ago
Location
United States, Oregon, Hillsboro; United States, California, Mountain View; United States, North Carolina, Raleigh
Work model
On-Site
Level
Senior
H-1B history
2,066 approvals (FY2023)
Posted
2h ago

Skills

AzureLinux

About this role

Overview

Microsoft Silicon, Cloud Hardware, and Infrastructure Engineering (SCHIE) is the team behind Microsoft’s expanding Cloud Infrastructure and responsible for powering Microsoft’s “Intelligent Cloud” mission. SCHIE delivers the core infrastructure and foundational technologies for Microsoft's over 200 online businesses including Bing, MSN, Office 365, Xbox Live, Teams, OneDrive, and the Microsoft Azure platform globally with our server and data center infrastructure, security and compliance, operations, globalization, and manageability solutions. Our focus is on smart growth, high efficiency, and delivering a trusted experience to customers and partners worldwide and we are looking for engineers to help achieve that mission.    We are seeking a High-Performance Computing (HPC) professional to join Microsoft’s Silicon Development Compute Solutions (SDCS) team, embedded within the broader silicon engineering organization. As a  Senior Linux SRE (Site Reliability Engineer) , your main focus will be to support LDAP/Identity Management environment and you will play a critical role in designing, deploying, and managing scalable, Linux-based compute infrastructure that supports silicon design workloads.  This role is central to ensuring the availability, performance, and efficiency of HPC services that power Microsoft’s silicon innovation. You will work closely with CAD, Operations, Engineering, and cross-functional teams to deliver resilient and high-performing infrastructure solutions that meet the demands of a globally distributed design organization.  If you are passionate about Linux systems at scale, HPC infrastructure, and enabling cutting-edge silicon design, this is a unique opportunity to make a significant impact.  #SCHIE Responsibilities Design, implement, and maintain enterprise identity and access management solutions leveraging LDAP, Red Hat Identity Management (IdM), and Microsoft Entra ID.    Administer and support LDAP directory services, including authentication, authorization, schema management, and integration with enterprise applications.     Manage and optimize Red Hat IdM infrastructure, including Kerberos, DNS, replication, user lifecycle management, and Linux authentication services.    Configure, maintain, and troubleshoot Microsoft Entra ID services, including identity synchronization, conditional access policies, authentication mechanisms, and hybrid identity integrations.   Collaborate with infrastructure, security, and application teams to design and implement identity management solutions that meet business and compliance requirements.    Monitor, troubleshoot, and resolve identity, authentication, and directory service issues across Linux, cloud, and hybrid environments.   Administer and optimize Linux environments (Red Hat, Rocky Linux,) across on-premises and Azure cloud infrastructure, including installation, configuration, patching, and troubleshooting.   Help manage engineering services such as Exceed TurboX, VNC, and authentication platforms (Red Hat IdM, NIS), ensuring high availability, performance, and user satisfaction.   Help with HPC infrastructure security compliance planning and remediation, aligning with compliance standards and operational requirements.   Collaborate with storage teams to diagnose and resolve Linux-related issues involving Isilon, Pure Storage, and Azure NetApp Files (ANF).   Implement and maintain monitoring and alerting systems to ensure system reliability, performance, and proactive incident response.   Partner with Engineering and CAD teams to address complex technical challenges and optimize compute and storage performance for design workloads.   Communicate effectively across global teams, demonstrating strong written and verbal communication skills and a collaborative mindset.

Qualifications

Required/minimum qualifications  Doctorate in Electrical Engineering, Computer

Listing verified 1h ago. Applications go through the company's official careers site.

← Back to Yoinka

Senior Linux SRE (Site Reliability Engineer) at Microsoft, United States, Oregon, Hillsboro; United States, California, Mountain View; United States, North Carolina, Raleigh | Yoinka