Senior Software Engineer (Cloud Infrastructure, Site Reliability Engineering)
Company: Oscar Health
Location: New York, New York, United States
Salary: $180.5k - $236.9k per year
Type: Full-time
Posted: 2026-08-07
About this role
- We’re hiring a Senior Software Engineer, Cloud Infrastructure / SRE to join our Engineering team
- Our Core Technology teams build and maintain the foundational platform upon which all Oscar engineering is built
- We are responsible for architecting a world-class, resilient ecosystem using a modern stack centered on AWS/GCP, Terraform, and Kubernetes with developer focused tooling and CI/CD
- Our mission is to provide an automated, self-service infrastructure that empowers our engineering organization to move fast without sacrificing security or stability
- You will report into a Staff/Senior Staff Engineer
- Become the expert on your team’s business and technical domains such as DevOps, site reliability, and cloud best practices
- Lead the planning, execution and release of complex technical projects across multiple teams outside of Core Technology
- Work with partners, product managers, and designers to solve challenging problems
- Lead and mentor engineers on the team to improve technology and apply best practices
- Independently responsible for large or complex technology capabilities (set of components or services) within their team’s domain or spanning multiple domains
- Facilitates, encourages, and enhances cross-team execution and collaboration; knows when cross-team projects are at risk and actively mitigates risk to deliver on time
- Prolific contributor to the objectives of their functional group, as well as organization-wide projects
- Drives prioritization of technical roadmap and influences prioritization of product roadmap and process enhancements within their team
- Actively identifies and reduces failure domains, designs and builds resilient systems, and strives to reduce adverse effects of an outage
- Builds software to minimize effort and business impact during maintenance and failures
- Guides the development of Service-Level Objectives (SLOs) for systems they are responsible for
- Own medium to large features or infrastructure projects from te...