Senior Support Engineer (GovOps)
Databricks
- Location
- McLean, Virginia
- Work model
- On-Site
- Level
- Senior
- Salary
- $130.2k/yr
- H-1B history
- 78 approvals (FY2023)
- Posted
- 2h ago
Skills
About this role
P-298
About Databricks
At Databricks, we are passionate about enabling data teams to solve the world's toughest problems — from making the next mode of transportation a reality to accelerating the development of medical breakthroughs. We do this by building and running the world's best data and AI infrastructure platform so our customers can use deep data insights to improve their business. Founded by engineers — and customer obsessed — we leap at every opportunity to solve technical challenges, from designing next-gen UI/UX for interfacing with data to scaling our services and infrastructure across millions of virtual machines. And we're only getting started.
The Role
s enterprise data enclaves and Generative AI applications become mission-critical to national security and public sector modernization, maintaining the reliability and operational integrity of the underlying platform is paramount. Databricks is seeking a Senior Support Engineer (GovOps) to join our Cleared Engineering team in McLean, VA. Sitting at the vital intersection of core infrastructure operations and active platform reliability, this team serves as the ultimate technical authority for our platform across GovCloud, FedRAMP High, and air-gapped environments.
In this role, you will move beyond basic ticket resolution to handle the hardest technical escalations across our platform domains—including Compute, Networking, Storage, IAM, and OS internals. You will partner directly with Core Engineering, conduct deep live troubleshooting, safely execute automation tools, and isolate complex system boundary blocks inside secure enclaves. If you are a systems specialist who thrives under pressure, enjoys writing custom tooling to unblock platform bottlenecks, and wants to safeguard the infrastructure powering next-generation federal AI workloads, this opportunity provides unmatched scope and impact.
The Impact You’ll Have
• The Impact You’ll HaveDrive High-Side Platform Reliability: Serve as the technical authority for resolving the most complex platform escalations across Compute, Networking, Storage, and OS domains within air-gapped and GovCloud enclaves.
• Partner with Core Engineering: Triage system defects, perform root-cause analysis, and collaborate directly with product engineering teams to drive permanent software bug fixes and fleet-wide supportability enhancements.
• Execute Advanced Automation & Tooling: Write and adapt custom Python, Bash, or Go scripts and diagnostic tools to automate platform triage and streamline operational recovery inside restricted boundaries.
• Manage Critical Incidents: Lead technical triage and live incident response during high-consequence platform events, maintaining composure and clear stakeholder communications under intense pressure.
• Optimize Secure Environment Pipelines: Identify systemic environment blocks, diagnose patch deployment failures, and review Infrastructure-as-Code (IaC) configurations to ensure