Skip to content
← Back to job listings

Senior Software Engineer (Cloud Infrastructure / SRE)

Oscar Health · Boston, United States

Imported listingfull-time16 days ago

About The Role

Join our Engineering team as a Senior Software Engineer specializing in Cloud Infrastructure and Site Reliability Engineering (SRE). You will be responsible for architecting a resilient ecosystem using modern technologies such as AWS/GCP, Terraform, and Kubernetes. Your mission will be to provide an automated, self-service infrastructure that empowers our engineering organization to move quickly without sacrificing security or stability. You will lead complex technical projects, mentor junior engineers, and drive the prioritization of the technical roadmap.

  • Lead the planning, execution, and release of complex technical projects across multiple teams, ensuring timely delivery and risk mitigation.
  • Act as a mentor and guide for junior engineers, improving technology and applying best practices within the team.
  • Drive prioritization of the technical roadmap and influence the prioritization of the product roadmap and process enhancements.
  • 6+ years of professional software engineering experience, working with a variety of technologies, and have increasingly impactful accomplishments
  • Experience as a major contributor cross-pod or cross-company deliverables
  • Experience leading technical contributions, improving the quality of what your teams create, and are excited to build fault-tolerant, and scalable software systems
  • Experience mentoring and training more junior engineers
  • Sets and enforces the standard for writing stable, correct, and maintainable code
  • Demonstrates expertise of the practical application of CS concepts within their team
  • Programming: Understanding of at least one coding language that you are able to use to develop scripts and software
  • Education: B.S. in Computer Science, a related technical field, or equivalent high-level industry experience
  • Observability: Proficiency with monitoring using tools like Prometheus, Grafana, or similar
  • CI/CD & Automation: Experience building robust deployment pipelines via GitHub Actions
  • Security & Networking: Knowledge of cloud-native security (IAM, VPC peering) and service mesh technologies like Istio
  • SRE Discipline: Strong background in Site Reliability Engineering, including Service Level Objectives (SLOs), error budgets, and incident management
  • Orchestration & Delivery: Proven track record with Kubernetes and workflows using ArgoCD
  • Cloud Proficiency: Deep expertise in managing production environments within AWS or GCP at scale
  • Infrastructure as Code: Advanced experience with Terraform or similar IaC tools to manage complex, multi-account structures

This is an external listing. JobSpring does not represent or verify the employer. Report this listing