Site Reliability Engineer- FedRAMP
Rubrik
- Location
- Palo Alto, CA
- Work model
- On-Site
- Level
- Entry
- Salary
- $158k/yr
- H-1B history
- 23 approvals (FY2023)
- Posted
- 1h ago
Skills
About this role
About The Team
The Rubrik Engineering team is comprised of people who produce extraordinary results. Our engineers are driven to build efficient, reliable, and cost effective products. We believe in empowering our teams, giving engineers autonomy and responsibility, not just tasks. Our goal is to motivate and challenge you to do the best work of your career. As part of the Rubrik Engineering team, you will work closely with product managers, designers, and other engineers to define the next generation of products for Rubrik. At Rubrik, nothing will stop you from thinking big. We are looking for individuals who are comfortable with ambiguity and excited by the prospect of a challenge. If you have a positive attitude, high energy and limitless drive, and like to win, we want to talk to you!
About The Role
Site Reliability Engineers at Rubrik are systems/software engineers who ensure that Rubrik's infrastructure services run smoothly and have the capacity for future growth.
What you'll do
• Ensure we maintain high availability and durability of our databases
• Establish best practices for internal teams to write performant SQL queries
• Perform periodic database upgrades minimizing downtime for our customers
• Design, implement and maintain relational database systems for performance and reliability
• Manage and run backend systems like Kubernetes, MySQL and everything in between
• Drive reliability, availability and efficiency improvements to Rubrik's Polaris Cloud Platform
• Good mix of software and system engineering skills
• Participate on-call rotations across continents, using a follow-the-sun model
• Write and review code, plan and execute upgrades, develop documentation and capacity plans, and debug production issues
• Work cross-functionally with various engineering teams
• Build monitoring tools and automation to increase efficiency of all teams
• Good written and verbal communication skills
• Drive FedRamp Certification process
• Good written and verbal communication skills
Experience You'll Need
• 2+ years of experience designing and managing relational databases with a focus on performance, scalability, reliability, high-availability, and disaster recovery.
• Experience in database design and architecture supporting large enterprise customers with high SLO and SLA requirements
• Experience operating database layer of a large scale SaaS product
• Experience in one or more of the following: Golang, Python, Java, Scala, C++
• Systematic problem-solving approach, coupled with strong communication skills and a sense of ownership and drive
• Expertise in designing, analyzing and troubleshooting large-scale distributed systems
• Ability to debug and optimize code and automate routine tasks
• Strong operational experience with Unix/Linux operating systems and networking
• Experience with Google Cloud Platform or other public cloud technologies
• Minimum 1-3 years of experience as a Development, DevOps or Site Reliability Engineer Willing to provide 24/7 coverage
• Strong Documentation skills
• Experience working with multiple departments and divisions within an organization
• Strong understanding of Databases is a definite plus
• Experience leading support personnel
• Experience with FedRAMP certification is strongly desired
Additional Requirements
Due to the criteria and security levels for Rubrik’s FedRAMP program, this position will require the following: • U.S. citizenship at the time of hire. • Residence within the contiguous United States (i.e., the lower 48 states and the