AI SummaryVerified by Aipplify AI
The vacancy is strong in task clarity and requirements but lacks compensation details and company information.
AI quality score5.9 / 10
Check Match — Just drop your CV
See your fit for Senior DevOps Observability Engineer in seconds.
Overview
We are looking for a Senior DevOps Observability Engineer for an international product company in the iGaming sector. Strong experience in Kubernetes and observability is required.
Responsibilities
- •Ensure reliable application operation in Kubernetes in production
- •Diagnose incidents and issues in Kubernetes clusters and distributed systems
- •Develop the Observability platform for metrics, logs, and traces
- •Work with VictoriaMetrics, Grafana, Loki, Tempo, Mimir, ELK, OpenTelemetry, and other observability tools
- •Enhance monitoring, alerting, and dashboards considering service architecture
- •Support and develop the Observability strategy, standardizing approaches for engineering teams
- •Participate in the full cycle of working with production incidents - from OnCall and service recovery to RCA/Post-Mortem
- •Implement reliability improvements based on incident outcomes and reduce the likelihood of recurrence
- •Collaborate with development/backend and other engineering teams
Conditions
- •Work with a large-scale production infrastructure and modern observability stack
- •Tasks that allow influencing observability and reliability approaches, not just maintaining existing solutions
- •International product company and a strong engineering team that has been working together for a long time
- •Office format 5/2 (hybrid possible)
- •8-hour workday + 30 minutes for lunch
- •Flexible start time from 8:00 to 10:00
- •Breakfasts and lunches in the office covered by the company
- •Internal and external training
- •Compensation for sports activities
- •Programs for learning English and Greek languages
- •Regular corporate events for employees and their children
Requirements
- •At least 3 years of experience as a DevOps Engineer
- •Strong recent practical experience with Kubernetes in production
- •Experience troubleshooting and resolving incidents in Kubernetes clusters
- •Experience independently developing and maintaining Helm charts - templating, functions, custom charts of varying complexity
- •Practical experience building and enhancing monitoring/observability
- •Experience with Grafana, VictoriaMetrics/Prometheus, Loki, ELK, OpenTelemetry, or comparable stack
- •Experience with production incidents, RCA/Post-Mortem, and subsequent reliability improvements
- •Experience with CI/CD and GitOps - GitLab CI/CD, Argo CD, or comparable tools
- •Proficient with Git
- •Experience with Docker and container build tools
- •Ability to independently troubleshoot complex production system issues and find root causes
- •Experience collaborating with cross-functional engineering teams
Skills
Loading similar jobs...