yoinka

Workforce Compute Ops SRE Lead Infrastructure Engineer

Wells Fargo

ISELIN, NJSenior
Sign in to applyVerified 3h ago
Location
ISELIN, NJ
Work model
On-Site
Level
Senior
Posted
Aug 27, 2026

Skills

AgileGenAIGrafanaKubernetesPrometheusPythonSplunk

About this role

About this role: Wells Fargo is seeking a Lead Site Reliability Engineer to provide technical leadership, mentorship, and guidance to a team responsible for the stability, reliability, availability, performance, and continuous improvement of enterprise workplace technology platforms. In this role, you will establish and promote Site Reliability Engineering (SRE) standards and best practices across observability, service level indicators (SLIs), service level objectives (SLOs), error budgets, incident and problem management, automation, and operational resilience. You will leverage telemetry and data-driven insights to identify risks, reduce incident frequency and recurrence, and drive continuous reliability improvements. You will also help improve the stability of workplace technology platforms through intelligent automation, Agentic AI capabilities, and low-code/no-code solutions. This includes providing technical leadership for AI-assisted incident triage, root cause analysis, workflow automation, proactive remediation, and self-healing capabilities to improve operational efficiency and the end-user experience. In this role, you will: Lead complex initiatives to develop infrastructure solutions that support business applications. Participate in projects intended to improve, modernize, and enhance technology infrastructure. Evaluate internal and external software solutions to support target-state architecture objectives. Review and analyze high-impact outages and implement processes to reduce future operational risk. Design, build, deploy, and maintain infrastructure solutions in collaboration with technology teams and third-party vendors. Design, code, test, debug, and document solutions using Agile development practices. Influence technical designs and implementation plans while identifying project risks and resource requirements. Provide technical leadership and guidance to engineers and partners across the organization. Direct risk and control activities by ensuring adherence to policies, procedures, and operational standards. Recommend solutions that improve efficiency, manage costs, and achieve business objectives. Collaborate with peers, leaders, customers, and vendors to resolve issues and deliver technology solutions.

Required Qualifications

5+ years of Technology Infrastructure Engineering and Solutions experience, or equivalent demonstrated through one or a combination of the following: work experience, training, military experience, education 5+ years of experience developing software, automation, or data processing solutions using Python 5+ years of experience with observability and monitoring technologies such as Splunk, Grafana, Prometheus, Elastic, or similar tools Desired Qualifications: Experience providing technical leadership and mentoring engineers Strong knowledge of Site Reliability Engineering principles and practices, including observability, service level indicators (SLIs), service level objectives (SLOs), error budgets, incident management, problem management, and operational resilience Experience developing Generative AI or Agentic AI solutions Knowledge of large language models (LLMs), prompt engineering, retrieval-augmented generation (RAG), AI-assisted workflows, and API-based AI integrations Experience with workflow automation and low-code/no-code platforms Experience using operational telemetry and data analytics to identify trends, anomalies, and opportunities for proactive remediation Experience with observability and monitoring platforms such as Splunk, Grafana, Prometheus, Elastic, or similar technologies Experience with containers, Kubernetes, and Infrastructure as Code practices Strong problem-solving, communication, and collaboration skills Ability to provide technical direction and influence engineering practices across teams Job Expectations: Must be able to work onsite one of the posted locations Sponsorship is not available for this role Pay Range   Reflected is the

Listing verified 3h ago. Applications go through the company's official careers site.

← Back to Yoinka

Workforce Compute Ops SRE Lead Infrastructure Engineer at Wells Fargo, ISELIN, NJ | Yoinka