VOTRE PROCHAIN CHAPITRE
Principal Site Reliability Engineer – Lead
À propos du poste
• Contribute to the design, development, and operation of a Golden Path platform built on Kubernetes and cloud-native technologies
• Provide technical and architectural leadership for the Next Generation Control Plane
• Lead transformation toward a highly available, resilient, distributed microservices architecture
• Provide leadership, support, and mentoring to the team
• Create and maintain SLOs and KPIs
• Collaborate with Engineering, Product, and Support teams
• Participate in on-call rotations
• Guide restoration and repair of service-impacting issues
• Work in an environment focused on innovative solutions, rapid development cycles, and open communication
• 10+ years of relevant experience
• Bachelor's degree in Computer Science or similar field, or equivalent experience
• Deep experience building and operating highly available, fault-tolerant, scalable production services using Kubernetes and other cloud-native technologies
• Solid understanding of Linux internals, especially containerization and networking
• Code-first approach to operating infrastructure, including automation
• Knowledge of Go
• Comfortable using Python or Bash for scripting
• Familiarity with infrastructure-as-code tools such as Crossplane, Pulumi, Terraform, or Ansible
• In-depth experience with modern observability tooling such as OpenTelemetry, Prometheus, Grafana, or Loki
• Ability to approach complex problems with methodical curiosity
• Benefits supporting health, well-being, finances, and life beyond work
• FlexBase flexible work arrangements: work at home, in an office, or a combination of both