Principal SRE, Infrastructure & Platform
F5
- Location
- Singapore Homebase
- Work model
- On-Site
- Level
- Principal
- H-1B history
- 60 approvals (FY2023)
- Posted
- Sep 2, 2026
Skills
About this role
At F5, we strive to bring a better digital world to life. Our teams empower organizations across the globe to create, secure, and run applications that enhance how we experience our evolving digital world. We are passionate about cybersecurity, from protecting consumers from fraud to enabling companies to focus on innovation. Everything we do centers around people. That means we obsess over how to make the lives of our customers, and their customers, better. And it means we prioritize a diverse F5 community where each individual can thrive. F5 is bringing a better digital world to life by helping organizations create, secure, and run applications that power our lives. Within the Platform Engineering team, this role helps ensure our platform is operated safely, reliably, and with operational excellence. We’re looking for a Principal SRE who leads with kindness, operates well in a global, follow-the-sun environment, and brings strong execution, documentation, and cross-functional coordination skills. You will be responsible for designing, deploying, and operating the foundational infrastructure that underpins a large-scale, multi-datacenter platform spanning 30+ Points of Presence across the Americas, EMEA, and APAC. This is a hands-on engineering role. You will own systems from bare metal to application layer -- provisioning physical servers via out-of-band management, building and maintaining Proxmox -based hypervisor clusters, managing containerized workloads, and driving automation across a heterogeneous on-premises and cloud environment. You will operate within a PCI-DSS compliant environment and be expected to contribute to hardening, audit readiness, and security tooling. You will join an on-call rotation and be expected to respond to and lead incident resolution for production systems across a 24x7 global environment.
What You'll Do
Infrastructure Automation & Configuration Management Author, maintain , and refactor Ansible playbooks and roles across a large-scale multi-datacenter inventory, covering the full lifecycle from bare-metal provisioning to application deployment Develop and improve CI/CD pipelines (GitLab CI) for infrastructure automation, including linting, testing, and staged rollout across regions Manage secrets lifecycle using HashiCorp Vault, including AppRole authentication, secret rotation, and PKI integration Maintain CMDB/IPAM accuracy in NetBox as a source of truth for all infrastructure assets Compute & Virtualization Deploy and manage Proxmox VE hypervisor clusters on bare-metal HPE hardware, including cluster formation, OVS networking, ZFS storage, and VM replication Provision and lifecycle-manage virtual machines using cloud- init , QCOW2 images, and Proxmox API automation Manage physical server provisioning end-to-end via HPE iLO (firmware updates, SPP deployment, OS installation via virtual media) Container & Kubernetes Platforms Manage self-hosted Kubernetes clusters on-premises, including control plane operations, node provisioning, workload deployment, and upgrade management Operate Docker-based workloads on infrastructure VMs using compose-driven deployments and container health monitoring Maintain container image pipelines and registry infrastructure (Azure Container Registry or AWS ECR ) Cloud Platforms Engineer and maintain infrastructure on AWS and Azure, integrating cloud resources with on-premises systems (DNS, monitoring, identity, networking) Apply cloud cost awareness, security best practices, and IaC principles (IAM, security groups, networking, storage) across AWS and Azure environments Networking & Core Services Operate and troubleshoot core distributed services including authoritative DNS (BIND9), recursive DNS (Unbound), load balancing ( HAProxy ), and high-availability VIPs (