O SEU PRÓXIMO CAPÍTULO
Site Reliability Engineer II
Sobre a vaga
• Improve the performance, availability, and scalability of large distributed content delivery systems
• Define and establish measurable Service Level Indicators (SLIs) and Service Level Objectives (SLOs) with cross-functional teams
• Provide technical expertise and feedback on system designs and implementations
• Monitor platform availability and performance, analyze data, resolve complex issues, and prevent recurrence
• Create and implement automation solutions to improve operational efficiency
• Participate in architecture and design reviews
• Stay current on cloud computing, DevOps, and SRE practices
• 2+ years of relevant experience
• Bachelor's degree in Computer Science, Engineering, or related field
• Experience validating data integrity, analyzing anomalies, and generating reports using Oracle SQL
• Expertise in scripting languages such as Python, Bash, or JavaScript
• Experience with Prometheus, Grafana, ADBMS, and Datadog
• Expertise in Unix/Linux operating environments
• Commitment to continuous learning and operational excellence through automation and efficiency improvements
• Customer-focused approach with responsibility and accountability
• Excellent interpersonal, written, and verbal communication abilities
• Health, well-being, finances, and life-beyond-work benefits
• FlexBase workplace flexibility: work at home, in an office, or a combination of both