yoinka

Senior Site Reliability Engineer - Remote

UnitedHealth Group

RemoteEden Prairie, MinnesotaSenior$91.7k – $163.7k/yrClearance required
Sign in to applyVerified 3h ago
Location
Eden Prairie, Minnesota
Work model
Remote
Level
Senior
Salary
$91.7k – $163.7k/yr

Skills

AWSArgoCDAzureCybersecurityGitHelmKubernetesPrometheusPulumiSplunkTerraform

About this role

For those who want to invent the future of health care, here's your opportunity. We're going beyond basic care to health programs integrated across the entire continuum of care. Join us to start Caring. Connecting. Growing together. Our Optum Serve IT team develops cutting-edge solutions that help people live healthier lives and help make the health system work better for everyone. From advanced data analytics and AI to cybersecurity, we use innovative approaches to solve some of healthcare's most complex challenges. To support this mission, OSIT has initiated a multi-year modernization program aimed at updating and enhancing enterprise technology systems in accordance with modern design standards You'll enjoy the flexibility to work remotely * from anywhere within the U.S. as you take on some tough challenges. For all hires in the Minneapolis or Washington, D.C. area, you will be required to work in the office a minimum of four days per week.

Job Summary

The Sr. Site Reliability Engineer will architect, develop, and maintain Optum Serve's cloud environment in both the commercial and government clouds. The role will work closely with software engineers, architects, and DevOps engineers to architect and maintain a secure, resilient and high performance cloud infrastructure.

Primary Responsibilities

Build, operate and support IaaS and PaaS infrastructure in Azure and AWS commercial and government clouds Work closely with dev teams to identify and measure SLOs, SLAs and SLIs Act a solid contributor to development of platform services including architecture, provisioning, configuration, deployment, and support Perform integrations with central logging, metrics dashboards, instrumentation, incident monitoring and management Build/integrate/administer systems and tools that enable engineering teams to observe their applications in production with autonomy (Dashboards, APMs) Support software and/or cloud-infrastructure in an on-call rotation basis Assist with identification and remediation of technical problems at the root cause by continuously implementing automation, self-healing, and real-time monitoring to production systems Maintain and improve operational tooling, frameworks Build frameworks that test the performance and resiliency of our platform services/tools Automate alerts for metrics on performance, cost, vulnerabilities, risk, compliance violations Improve processes and champion automation of any manual items around support You'll be rewarded and recognized for your performance in an environment that will challenge you and give you clear direction on what it takes to succeed in your role as well as provide development for other roles you may be interested in.

Required Qualifications

6+ years of experience working in a Site Reliability Engineering, Cloud Engineering, or DevOps role Experience with infrastructure as code (IaC) tools like Terraform, Pulumi Experience with Kubernetes deployment tools like Helm, ArgoCD, Flux Experience supporting infrastructure in production cloud environments Experience working with RESTful services Some experience with monitoring tools (Azure Monitor, Splunk, Dynatrace, Graphana, Prometheus) Knowledge of Encryption, Public Key Infrastructure (PKI), understanding of OWASP Expert knowledge of at least one major cloud service provider (Azure preferred, AWS acceptable) Expert knowledge and hands on production experience in Kubernetes (bare metal or managed) cluster setup and management Understanding of identity and access management (IAM) Familiarity with IDEs and source control tools such as Visual Studio Code, GitHub, GitLab Solid awareness of networking and internet protocols Ability to participate in a 24/7 on-call rotation following documented procedures and escalation paths United States Citizenship If you are offered this position, you will be required to provide extensive personal information to obtain and maintain a suitability or determination of eligibility for a

Listing verified 3h ago. Applications go through the company's official careers site.

← Back to Yoinka

Senior Site Reliability Engineer - Remote at UnitedHealth Group, Eden Prairie, Minnesota | Yoinka