Platform Reliability Engineer
Carecone โ Australia ยท Posted ~1 day ago
๐ Log in to save this job, tailor your resume & track your apply process โ 7 days free, no card needed.
Log in to add to target listDescription
Role name : Platform Reliability Engineer (Full Stack & Cloud Native)
Skills required (Key skill : Zabbix)
Expertise in core tool sets including (but not limited to): Go, Python 3, FastAPI, React, HTML, JSON, JavaScript, NGINX, Docker, Podman, Kafka, Grafana, Air Flow, Linux, basic networking, Databases (MySQL, ClickHouse, Victoria Metrics, Postgres, MongoDB & OpenTSDB), and experience in other open source technologies like Telemetry, Zabbix, Kubernetes container and NOC.
The current monitoring platform, NOC, is an open source network management system providing discovery, inventory, fault management, and SNMP based performance monitoring capabilities.
To simplify operations, standardise monitoring, and improve observability, SNMP monitoring functions need to be migrated to Zabbix while decommissioning or reducing reliance on NOC for monitoring functions.
Objectives
Replace NOC SNMP performance monitoring with Zabbix.Ensure continuous visibility of network devices via SNMP polling.Standardise monitoring templates and alerting models.Improve operational efficiency for NOC teams.Reduce operational complexity by consolidating monitoring tools.
We have 66,600 jobs that might be an even better fit for you
DontApply's real value goes far beyond a single job link or company name. Just upload your resume โ in under a minute we'll analyze all 66,600 jobs and tell you exactly which ones you should apply to right now.
Upload My Resume