O SEU PRÓXIMO CAPÍTULO
Senior Site Reliability Engineer
Sobre a vaga
• Provide technical and architectural leadership for the design, development, and scaling of the Next Generation Control Plane
• Lead transformation toward a highly available, resilient, distributed microservices architecture
• Contribute to development and operation of a globally deployed, multi-tenant platform built on Kubernetes and cloud-native technologies
• Build automation and tooling to reduce operational toil, improve deployment safety, and accelerate incident response
• Create and maintain SLOs and KPIs
• Collaborate with Engineering, Product, and Support teams
• Participate in on-call rotations and guide restoration and repair of service-impacting issues
• 5+ years of relevant experience
• Bachelor's degree in Computer Science or a similar field, or equivalent experience
• Experience building and operating highly available, fault-tolerant, scalable production services using Kubernetes and other cloud-native technologies
• Working understanding of Linux internals, especially containerization and networking
• Code-first approach to operating infrastructure
• Knowledge of Go
• Comfort leveraging Python or Bash for scripting
• Familiarity with infrastructure-as-code tools such as Crossplane, Pulumi, Terraform, or Ansible
• In-depth experience with observability tooling such as OpenTelemetry, Prometheus, Grafana, Loki, or similar
• Health, well-being, and financial support benefits
• Flexible work arrangements through Akamai's FlexBase program: work at home, in an office, or a combination of both