← Todos los empleos

Platform Operations Engineer

Sobre el puesto

• Receive and process requests via Telegram, Slack, and email
• Clarify the nature of issues and gather details from users
• Identify which component or service an incident relates to
• Own incidents end-to-end from detection through resolution
• Keep stakeholders updated throughout incident resolution
• Perform basic infrastructure troubleshooting involving logs, service status, and configurations
• Execute deterministic runbook actions such as restarting services and testing recovery
• Escalate to L2/L3 DevOps and developers with prepared context
• Create tickets and bug reports in the tracker
• Collaborate on runbooks and knowledge base entries for recurring issues

• Experience working with monitoring and logging systems
• Ability to read and analyze logs (Grafana, Kibana, Loki)
• Basic issue localization (network, DNS, service connectivity)
• Basic understanding of Kubernetes: kubectl logs, kubectl describe, kubectl get
• Understanding of application configuration: Helm values, ConfigMaps, environment variables
• Infrastructure-level troubleshooting (service availability, node status, resources)
• Background in QA/tech support
• Experience with Helm is a plus
• Familiarity with CI/CD pipelines is a plus
• Understanding of microservices architecture is a plus
• Willingness to work night shifts covering European hours
• Candidates based near the EST timezone are preferred

• 21 vacation days + public holidays + 5 sick days
• Private English lessons via Preply
• Fully remote across Europe
• Fast career progression
• Startup pace with enterprise stability
• Cutting-edge tech stack
• Real ownership and direct impact of work