YOUR NEXT CHAPTER
Senior Software Engineer, DevOps
About the role
• Design, build, and maintain Google Cloud Platform infrastructure using Terraform
• Define infrastructure patterns and standards
• Own and improve GitHub Actions CI/CD pipelines
• Identify and eliminate operational toil through scalable automation and tooling
• Establish monitoring, dashboards, and alerting in Datadog
• Reduce alert noise across team systems
• Lead incident response within the on-call rotation and drive postmortems and follow-ups
• Define and track Service Level Objectives for core infrastructure and build systems
• Partner with product engineering teams on deployment pipelines, environment issues, and build troubleshooting
• Own preview and staging environments, including data synchronization, masking, and cleanup
• Improve developer experience through tooling, runbooks, and documentation
• Write tested and reviewed infrastructure code
• Lead design reviews and RFCs
• Design systems balancing reliability, performance, and security
• Conduct code reviews and mentor engineers through pairing, knowledge sharing, and documentation
• Collaborate with the Tech Lead, DevOps, Platform, Data Engineering, and domain product teams
• ~5+ years of professional experience in DevOps, SRE, infrastructure, or backend engineering in production environments
• Hands-on experience designing and operating infrastructure in at least one cloud provider, ideally Google Cloud Platform
• Track record of owning and shipping automation, pipelines, or infrastructure that made a team measurably more productive or reliable
• Enthusiasm for developer productivity
• Deep experience with infrastructure-as-code, ideally Terraform, and building CI/CD pipelines, ideally GitHub Actions
• Proficient software engineering ability and strong command of Linux
• Strong instincts for monitoring and observability, with confident debugging across logs, traces, and metrics
• Solid experience with relational databases, including MySQL and PostgreSQL, and containerized workloads
• Strong grounding in reliability, performance, and security fundamentals, with sound tradeoff judgment
• Nice to have: experience in healthcare, digital health, or regulated domains, including HIPAA, PHI, or SOC 2
• Nice to have: experience with Docker and Kubernetes
• Nice to have: exposure to incident response, on-call, and postmortem practices
• Nice to have: experience with database migrations or managing multiple environments at scale