yoinka

Senior Software Engineer: Site Reliability Engineering

Jack Henry & Associates

RemoteSeniorH-1B sponsor company
Sign in to applyVerified 1h ago
Location
Remote
Work model
Hybrid
Level
Senior
H-1B history
8 approvals (FY2023)
Posted
1d ago

Skills

AnsibleGCPGitTerraform

About this role

Description & Requirements

Press space or enter keys to toggle section visibility

At Jack Henry, we’re more than a technology company, we’re a force for good in financial services. We’re redefining how community banks and credit unions connect with the people they serve. Our mission is rooted in people inspired innovation, empowering financial institutions to deliver seamless, secure, and human centered experiences. We deliver cutting-edge solutions that are paving the way for the next generation of digital banking and payments, but our true impact begins with our associates. If you're ready to help transform an industry and grow with a company that values purpose, collaboration, and excellence then we’d love to meet you. Our Enterprise Information Technology (EIT) organization is expanding, and we are seeking a Senior Site Reliability Engineer to help drive a major architectural modernization. In this role, you will move beyond traditional infrastructure maintenance to build a proactive, engineering-led ecosystem across our large-scale multi-cloud and co-location footprint. Working closely with cross-functional IT teams and business units, you will help design and implement standards for our hybrid cloud datacenter model. A primary focus of this position will be the architectural redesign, optimization, and migration of legacy on-premises workloads into Google Cloud Platform (GCP), ensuring everything we build adheres to rigorous SRE principles. Mission & Impact: Everything as Code: Drive repository-led management across our public and private cloud environments to establish consistency and eliminate manual configuration drift. Engineering over Toil: Heavily leverage Infrastructure as Code (IaC) to automate repetitive tasks, build smooth "paved roads" for product teams, and develop self-healing systems. SLO-Driven Architecture: Help shift our operational focus from traditional component monitoring to user-facing symptoms, defining meaningful SLOs and error budgets. Modernization & Migration: Lead the technical execution of re-architecting and redeploying on-premises services into GCP, ensuring scalability, performance, and long-term reliability. This position may be worked remotely within the United States, with the exception of California. This position is ineligible for immigration sponsorship and support. Please do not apply if at any time you will need immigration support now or in the future (i.e., H-1B, PERM). All positions, regardless of location, may require an onsite interview or in-person onboarding requirement to verify your identity. What you’ll be responsible for: Service Reliability and Performance:   Drive the reliability and performance of both public cloud (production, testing, and development) and internal server infrastructure environments. SRE Practice Implementation:   Design and implement robust Site Reliability Engineering practices, including defining and monitoring Service Level Objectives (SLOs) and Service Level Indicators (SLIs), focusing on proactive system health and error budgets. Automation and Toil Reduction:   Ruthlessly eliminate manual, repetitive work (toil) through automation. Develop and maintain automation scripts and tooling to streamline operations across the hybrid datacenter model (on-premises and public cloud). Implement "Everything as Code":   Treat the cloud and on-prem operational environment as a software project by using Infrastructure as Code (IaC) with tools like Terraform, Ansible, GitHub for provisioning and configuration. Configuration and State Management:   Design and maintain rigorous configuration management processes to guarantee the consistency and desired state of the hybrid datacenter infrastructure, leveraging tools like Ansible. Observability and Alerting:   Establish and manage comprehensive monitoring and alerting systems to provide deep visibility into the health and performance of services.   Build systems that are self-healing and

Senior Software Engineer: Site Reliability Engineering at Jack Henry & Associates, Remote | Yoinka