Site Reliability Engineering Lead – Taiwan
Obsidian Security
- Location
- Taipei, Taiwan
- Work model
- On-Site
- Level
- Senior
- Posted
- 1h ago
Skills
About this role
Obsidian Security is the leading SaaS security platform, trusted by global enterprises like Snowflake, T-Mobile, and Algolia. We protect 200+ organizations across North America, Europe, the Middle East, Southeast Asia, Australia, and New Zealand, including many of the world’s largest Fortune 1000 and Global 2000 companies.
Founded in 2017 and backed by top investors like Greylock, Obsidian was built to close a critical gap: securing SaaS apps where business happens—Microsoft 365, Salesforce, and hundreds more. The company does this by offering a complete SaaS security platform to reduce risk, detect and respond to threats, and prevent breaches at the source. Obsidian was built by leaders who redefined endpoint and identity security at CrowdStrike, Okta, Cylance, and Carbon Black. Now, they’re transforming how SaaS is secured.
With AI driving rapid SaaS growth and complexity, agentic AI tools gain privileged access to sensitive data through integrations, creating new risks most security tools miss. Obsidian uniquely detects anomalous OAuth token activity and manages integration risks. Major announcements are on the horizon. Recognizing that SaaS security needs to evolve, Obsidian enables growing organizations to start with a lightweight, prevention-focused browser extension and expand coverage over time.
With global momentum, a growing partner ecosystem including SentinelOne, Databricks, and Google Cloud, and a major fundraise ahead, Obsidian is scaling rapidly toward long-term growth and IPO readiness.
Site Reliability Engineering Lead – Taiwan
About the Role
Obsidian Security is looking for a Site Reliability Engineering Lead to establish and lead production reliability capabilities within our growing Taiwan engineering organization.
You will be responsible for the reliability, security, scalability, and operational effectiveness of cloud services that protect some of the world’s largest enterprises. You will lead work across service reliability, cloud infrastructure, observability, incident management, capacity planning, vulnerability remediation, and production readiness.
This is a hands-on leadership role for someone who can operate effectively during critical incidents while also addressing the engineering and organizational causes behind them. You will build systems and practices that enable product teams to move quickly without compromising availability, security, or customer trust.
You will work closely with the Director of Engineering – Taiwan and global engineering, infrastructure, security, and product teams. As the Taiwan site grows, you will recruit and develop a high-performing SRE team and help create a strong, shared operational culture.
What You'll Do
• Establish and lead the SRE function in Taiwan, including its technical roadmap, operating model, hiring plan, and relationship with global teams.
• Improve the availability, performance, scalability, security, and cost efficiency of Obsidian’s production services.
• Define service-level indicators, service-level objectives, error budgets, and operational health metrics for critical services.
• Build and improve observability across applications, data pipelines, APIs, infrastructure, and customer-facing workflows.
• Lead production incident response, technical coordination, customer-impact assessment, and service recovery.
• Establish effective on-call practices, escalation paths, operational runbooks, and incident command processes.
• Facilitate blameless post-incident reviews and ensure corrective actions address systemic causes.
• Develop automation that reduces manual operations, deployment risk, recovery time, and repetitive work.
• Partner with engineering teams on production readiness,