Sr Specialist System Engineering (ANT DevOps/Infra Engineer)
AT&T
- Location
- IND:KA:Bangalore / Epip Area, Hoodi Village, Whitefield Rd - Eqp: Plot 111/112, Epip Area, Hoodi Village, Whitefield Road
- Work model
- On-Site
- Level
- Senior
- Posted
- Aug 21, 2026
Skills
About this role
Job Responsibilities
Serve as the primary off ‑ shore DevOps/Infra support lead for Automation Tools and 4G/5G Core Network Simulators across NPRD/Labs and PROD environments. Design, build, and maintain Azure DevOps CI/CD pipelines using a 3 ‑ tier ( hyper→macro→micro ) architecture to automate packaging, publishing, and deployments. Develop and maintain Helm v3 (OCI) charts for application and infra components; package/publish artifacts to JFrog Artifactory OCI. Automate Kubernetes deployments via Python orchestration (e.g., deploy.py), including config substitution, sequencing, and post ‑ deploy health checks. Own secrets management end ‑ to ‑ end : Ansible Vault encryption, HashiCorp Vault runtime retrieval ( AppRole /KV), and token rotation with zero ‑ plaintext controls. Provision and operate infrastructure: KVM/QEMU/ libvirt VM lifecycle on Simzone hosts; implement SR ‑ IOV passthrough and host networking configurations. Author/ maintain Infrastructure ‑ as ‑ Code using OpenStack HEAT ( on ‑ prem VM provisioning) and Azure Bicep (NAKS/AON resources). Configure advanced K8s networking and security: Multus NADs ( multi ‑ NIC pods), MetalLB L2 VIPs, Calico policies, SR ‑ IOV for performance I/O. Manage storage and stateful services: Rook ‑ Ceph provisioning (RWX/RWO), local ‑ storage for single ‑ node clusters, PVC sizing, and storage class governance. Build/ maintain large Ansible automation footprint (deployment, upgrades, backup/restore, mechID mgmt , vault operations, compliance tasks). Operate AOSM lifecycle pipelines (nfPL0→nfPL4) to publish NF definitions, build deployment configs, deploy to NAKS, and onboard to HCEV. Drive environment promotion (NPRD→PROD): release branching, versioning updates, quarterly release coordination, and consistency across lab/ pre ‑ prod /prod. Maintain per ‑ instance configuration at scale (20+ K8s instances) using real ‑ parameters- values.json / input- values.json and enforce correctness across environments. Implement observability and verification: deploy kube ‑ prometheus ‑ stack , integrate DevOps metrics (DORA Build Events API), and standardize service health checks. Own certificate lifecycle management via cert ‑ manager and vault ‑ issuer patterns, including rotation workflows and Vault PKI integration. Maintain supply ‑ chain security artifacts ( bom.yaml validation, images.json version tracking, reproducible builds with pinned versions). Administer multi ‑ repo governance (30+ repos across GitHub + ADO): branch protections, PR governance, and cross ‑ repo dependency coordination. Build AI ‑ assisted DevOps tooling (agentic workflow with specialized agents) to accelerate analysis, pipeline engineering, diagnostics, and self ‑ healing operations. Provide incident response and operational troubleshooting: kubectl diagnosis, log analysis, NF status checks, backup/restore execution, and chaos testing support. Job Qualifications / Required Qualifications Hands ‑ on experience operating Kubernetes platforms and production-grade CI/CD pipelines (Azure DevOps, GitHub/ GitOps ). Strong automation skills in Python , Ansible, and Bash for deployment orchestration and operational tooling. Deep experience with Helm 3 (OCI) packaging and release management; artifact publishing ( JFrog Artifactory OCI). Expertise in secrets management ( HashiCorp Vault + Ansible Vault), credential rotation, and secure repo practices. Experience with infrastructure provisioning and IaC (OpenStack HEAT and/or Azure Bicep) across hybrid environments. Strong troubleshooting and incident management skills across K8s clusters, networking, and stateful services. Working knowledge of advanced K8s networking/storage ( Multus , SR ‑ IOV, MetalLB , Calico; Rook ‑ Ceph /local storage). Understanding of