Lead Infrastructure Engineer (Site Reliability Engineering) – Workplace Technology
Wells Fargo
- Location
- Bengaluru, India
- Work model
- On-Site
- Level
- Senior
- Posted
- 12h ago
Skills
About this role
About this role: Wells Fargo is seeking a Lead Infrastructure Engineer to help drive reliability, operational excellence, automation, and service resiliency across our Workplace Technology portfolio. This role will provide technical leadership for the technologies that enable our employees to work, collaborate, and communicate every day. You will partner closely with engineering, infrastructure, security, and architecture teams to improve service reliability, reduce operational effort through automation, strengthen observability, and ensure a consistent end-user experience at enterprise scale. Our Workplace Technology landscape includes: Microsoft 365 (Teams, Exchange, SharePoint, OneDrive, Copilot), Collaboration & Productivity Services Endpoint Management & Intune, Software Distribution Services Entire PC Lifecycle Management Workplace Compute (Windows Endpoints, VDI, Windows 365) Compliance, Security, and Regulatory Services This is a highly visible role that will influence the reliability strategy and operational maturity of Workplace Technology services globally. In this role, you will: Lead reliability and operational support for Workplace Technology platforms serving a global workforce. Drive the adoption of Site Reliability Engineering (SRE) practices, including service health monitoring, SLOs, SLIs, and operational metrics. Lead response and recovery efforts during major incidents, coordinating across technology teams to restore services and minimize business impact. Drive root cause analysis and continuous improvement initiatives to improve service stability and prevent recurring issues. Champion automation, self-healing capabilities, and process simplification to reduce operational toil. Build and enhance observability capabilities using monitoring, telemetry, logging, and user experience insights. Partner with engineering teams on platform modernization, cloud adoption, resiliency improvements, and operational readiness. Establish operational governance, risk controls, and compliance practices across supported services. Lead capacity, availability, and performance management activities to ensure services scale effectively. Provide technical leadership, mentoring, and guidance to engineers and support teams operating in a global environment.
Required Qualifications
5+ years of experience in Systems Operations, Infrastructure Engineering, Site Reliability Engineering, or Technology Operations. Experience supporting large-scale enterprise platforms with global operational responsibilities. Strong background in incident management, problem management, and service reliability. Proven track record of driving automation, operational improvements, and service optimization initiatives. Experience supporting critical production environments with high availability and resiliency requirements. Strong understanding of SRE principles, operational excellence, and modern support models.
Desired Qualifications
Workplace Technology Experience Hands-on experience with one or more of the following: Microsoft 365 (Exchange Online, Teams, SharePoint, OneDrive, Copilot) Windows Endpoints, Windows 365, Intune and Endpoint Management Software Distribution and Endpoint Management platforms Reliability & Observability Experience defining SLOs, SLIs, operational KPIs, and reliability measures. Experience with enterprise monitoring and observability platforms such as Splunk, ThousandEyes, Azure Monitor, Grafana, Dynatrace, Aternity, or AppDynamics. Knowledge of service resiliency, disaster recovery, capacity planning, and performance management. Automation & Cloud Automation experience with PowerShell, Python, Ansible, Terraform, CI/CD, REST APIs, SQL Queries, HTML Coding, etc., Working knowledge of Microsoft Azure and hybrid cloud environments. Experience implementing automation and self-healing solutions at scale. Job Expectations: Serve as a technical leader for Workplace Technology Operations with good knowledge on Infrastructure