Senior Observability Engineer
Egt Digital โ Bulgaria ยท Posted ~1 week ago
๐ Log in to save this job, tailor your resume & track your apply process โ 7 days free, no card needed.
Log in to add to target listDescription
EGT Digital is a next-generation tech company focused on all online gaming products.
Its portfolio includes Casino Games, Sportsbook, and the all-in-one solution โ a Gambling Platform.
EGT Digital is a part of the Euro Games Technology (EGT) Group, headquartered in Sofia, Bulgaria.
EGT Group is one of the fastest-growing enterprises in the gaming industry.
Our global network includes offices in 25 countries and our products are installed in over 85 jurisdictions in Europe, Asia, Africa, and North, Central, and South America.
Being a part of such a fast-moving industry as iGaming, the company knows no limits and is growing rapidly through its dedication to innovation and constant improvement
.
Responsibilitie
s:Build and maintain dashboards for service health, feature health, business flows, KPIs, and production operations.Create monitoring views for critical sportsbook flows such as login, bet placement, cashout, settlement, feed processing, odds updates, wallet communication, external integrations, and player journeys.Design dashboards that show health by service, business unit, cluster, provider, sport, market, event, channel, and customer flow.Define and implement meaningful alerts that detect real production issues while reducing noise and false positives.Work with Engineering, QA, NOC, Product, Trading, and Business teams to understand what visibility is missing.Analyze production behavior and transform raw technical data into clear operational insight.Help teams understand incidents by creating dashboards, reports, queries, and post-incident analysis views.Improve production readiness by ensuring important services and features have proper monitoring, alerting, and operational visibility before release.Identify recurring production issues and propose better observability, automation, or tooling around them.Create reusable monitoring templates, dashboard standards, alerting standards, naming conventions, and operational views.Help define useful SLIs, SLOs, service health indicators, business health indicators, and production KPIs.Support incident investigation by improving logs, metrics, traces, and business-flow visibility.Work with developers through pull requests and code reviews when observability-related changes are require
d.
Requiremen
ts:Strong experience with production monitoring, observability, dashboards, and alerting.Hands-on experience with tools such as Grafana, Prometheus, Kibana, ELK/OpenSearch, Loki, or similar.Good understanding of logs, metrics, traces, latency, error rates, throughput, saturation, and service health.Ability to read, understand and do small changes Java and C# backend services.Ability to make small, safe code changes related to observability, logging, metrics, tracing, and health checks.Strong SQL skills and ability to query production or analytical data sources.Good understanding of APIs, databases, queues, caches, service-to-service communication, and distributed systems.Experience building dashboards and alerts for both technical and non-technical users.Ability to understand business processes and translate them into measurable indicators.Strong analytical thinking and ability to investigate production behavior using data.Ability to work with developers through pull requests, code reviews, testing, and release coordination.Good communication skills and ability to work with Engineering, QA, NOC, Product, Trading, and Business stakeholde
rs.
Nice to h
ave:Experience in sportsbook, gaming, fintech, payments, trading, or other high-volume transactional platforms.Experience with Kafka, Redis, PostgreSQL, Kubernetes, Java/Spring Boot, .NET/C#, or microservices.Experience with SLOs, SLIs, incident management, production readiness, or SRE practices.Experience building business dashboards, operational reports, or production impact analysis views.Experience with synthetic monitoring, health checks, dependency maps, or automated incident diagnostics.Experience reducing alert noise and improving alert quality.Basic scripting or automation experie
nce.
What we o
ffer:Competitive salaryPerformance based annual bonusPerformance evaluation & salary review twice a year25 days paid annual leaveWork from home option - 2 days weeklyFlexible working scheduleAdditional health insurance โ premium packageFully paid annual transportation cardFully paid Sports cardFree company shuttle by the officeSports Teams/Sports eventsProfessional development, supportive company culture, and challenging projectsCompany-sponsored trainingsTickets for conferences and seminarsTeam building events and office partiesReferral ProgramFree snacks, soft drinks, coffee, and fruit are always availableBirthday, newborn baby, and first-grader bonusesCorporate discounts in various shops and restaurantsState-of-the-art modern officePositive working environment and chill-out zone (PS4, foosball-table, and lazy ch
airs)
All applications will be treated in strict confidentiality and only the approved candidates will be invited to an inte
rview.
We have 63,450 jobs that might be an even better fit for you
DontApply's real value goes far beyond a single job link or company name. Just upload your resume โ in under a minute we'll analyze all 63,450 jobs and tell you exactly which ones you should apply to right now.
Upload My Resume