Site Reliability Engineer - Devops, AWS, Python/GO - 4 to 8 years
Cisco
- Location
- Bangalore, India
- Work model
- On-Site
- Level
- Mid
- Posted
- Sep 12, 2026
Skills
About this role
Meet the Team As a Site Reliability Engineer on the Intersight Team you will play a key role in ensuring the reliability, scalability, and security of our cloud platforms. The broader team is composed of experienced engineers who value innovation and accountability. You will represent the Intersight SRE team, working in a dynamic environment, tackling challenges with creativity, and delivering on the team's technical roadmap. You will collaborate with cross-functional teams, including engineering, product management, customers and security, to design, influence, build, and maintain SaaS systems operating at multi-region scale. Your work will directly impact the success of our initiatives by ensuring the underlying platform infrastructure is robust, efficient, and aligned with operational excellence.
Your Impact
We are seeking an experienced Engineer to represent a high-performing team dedicated to ensuring the reliability and scalability of cloud services, with a focus on a rapidly growing next-generation project. The ideal candidate will have hands-on SRE or systems/network administration experience, with familiarity in AWS. This role involves close collaboration across product engineering, service engineering, and SRE teams in a high-trust, well-coordinated environment. If you are looking to excel in a fast-paced environment creating innovative solutions and want to make a significant long-lasting impact, this job is for you. Design, build, and optimize cloud and data infrastructure to ensure the high availability, reliability, and scalability of systems to meet customer needs, while implementing SRE principles such as monitoring, alerting, error budgets, and incident management. Collaborate closely with cross-functional teams, including customers, development, product management, and security teams, to create secure, scalable solutions and enhance operational efficiency through automation. Monitor production systems, participate in on-call rotations, troubleshoot incidents, and contribute to root cause analysis. Contribute to continuous improvement efforts through postmortem reviews and proactive performance optimization. Utilize your strong programming skills to integrate software and systems engineering, building core data platform capabilities and automation to meet enterprise customer needs and roadmap objectives.
Minimum Qualifications
Bachelor's degree in computer science or a related field with 5-8 + years of experience in Site Reliability Engineering/ DevOps, Cloud Operations, or a related role. Practical experience of working in Kubernetes (EKS or self-managed) , Docker, Networking and public cloud (preferably AWS). Strong Infrastructure as Code (IaC) skills, with experience in tools such as Terraform, Ansible or CloudFormation. Experience with observability tools using Prometheus, Grafana, CloudWatch, Splunk or Open Telemetry. Working knowledge of at least one programming or scripting language (e.g., Python, Bash, Go) for automation and operational tasks. Good understanding of Unix/Linux systems, the kernel, system libraries, file systems, and client-server protocols.
Preferred Qualifications
Experience building/managing a cloud-based data platform, automation and orchestration of their infrastructure and maintaining high availability, system reliability at scale. Experience applying AI-assisted or automation-first approaches to SRE tooling and workflows. Strong personal interest in learning, researching, and creating new technologies with high customer impact Excellent collaboration skills and the ability to bring out the best in a technically diverse team Certifications: CKA (Certified Kubernetes Administrator), CKAD (Certified Kubernetes Application Developer), AWS Certified DevOps Engineer, or equivalent certifications in cloud and security domains. Why Cisco? At Cisco, we’re revolutionizing how data and infrastructure connect and protect organizations in the AI era – and beyond. We’ve been