About the role
At least 5 years of experience in Site Reliability Engineering, DevOps, platform engineering, or a similar role. Strong experience working with complex, distributed, production-grade systems. Very good knowledge of observability tools, especially Prometheus, Grafana, Loki, Tempo, and OpenTelemetry. Hands-on experience with Kubernetes and Docker. Practical experience with both cloud and on-premises infrastructure. AWS experience is preferred. Ability to automate tasks and workflows using Python,…
