Platform Reliability Engineer (Zabbix)

Showtime ConsultingIndia
Adzuna INPosted 8h agoOriginal Listing
it-jobs

Job Description

Company Description Showtime Consulting is a leading provider of Shielded Cloud and Digital Solutions across Australia and New Zealand. We specialise in delivering secure, enterprise-scale technology solutions across cloud, infrastructure, cybersecurity, DevSecOps, software engineering, data, and digital transformation programs. We partner with government and enterprise organisations to build high-performing technology teams that deliver secure, resilient, and scalable solutions within complex and highly regulated environments. The Role We are seeking an experienced Platform Reliability Engineer to support the uplift, consolidation, and modernisation of critical monitoring and observability platforms. This role will focus on improving platform reliability, simplifying operations, and supporting the migration of SNMP performance monitoring from an existing NOC platform to Zabbix. You will work across full-stack engineering, cloud-native platforms, monitoring, automation, and open-source technologies to improve visibility, reliability, and operational efficiency for technical teams. Key Responsibilities - Support the migration of SNMP performance monitoring from NOC to Zabbix. - Design, configure, and maintain Zabbix monitoring templates and alerting models. - Ensure continuous visibility of network devices through SNMP polling. - Improve monitoring coverage, platform reliability, and operational observability. - Support the consolidation of monitoring tools to reduce operational complexity. - Develop and maintain full-stack platform tooling using modern technologies. - Build and support backend services using Go, Python, FastAPI, and related frameworks. - Develop and maintain frontend components using React, HTML, JavaScript, and JSON. - Support containerised environments using Docker, Podman, and Kubernetes. - Work with monitoring, telemetry, messaging, and data platforms including Grafana, Kafka, and Airflow. - Support Linux systems, basic networking, NGINX, and open-source infrastructure tools. - Troubleshoot monitoring, platform, application, and infrastructure issues. - Collaborate with NOC, engineering, platform, and operations teams to improve service reliability. - Produce technical documentation, operational procedures, and support handover materials. What You'll Need - Strong experience with Zabbix monitoring, templates, alerting, and SNMP polling. - Experience working with NOC platforms or network monitoring environments is highly regarded. - Experience migrating or consolidating monitoring capabilities is advantageous. - Strong experience with Linux systems and basic networking concepts. - Experience with open-source observability, telemetry, or monitoring technologies. - Experience with Go, Python 3, FastAPI, JavaScript, React, HTML, and JSON. - Experience with NGINX, REST APIs, and full-stack platform tooling. - Hands-on experience with Docker, Podman, and Kubernetes containers. - Experience with Kafka, Grafana, Airflow, or similar platform technologies. - Database experience across MySQL, Postgres, MongoDB, ClickHouse, VictoriaMetrics, or OpenTSDB is advantageous. - Understanding of cloud-native application design and platform reliability principles. - Strong troubleshooting, analytical, and problem-solving skills. - Ability to work with engineering, operations, and NOC teams to improve monitoring outcomes. - Experience standardising monitoring templates, alerts, and operational processes is highly desirable. - Strong communication skills with the ability to document and explain technical solutions clearly. Why Join Us? - Work on a key observability and monitoring platform modernisation initiative. - Help migrate SNMP monitoring from legacy NOC tooling into Zabbix. - Improve reliability, visibility, and operational efficiency across critical platforms. - Gain exposure to cloud-native, open-source, monitoring, and full-stack technologies. - Collaborate with experienced platform, engineering, and operations teams. - Join a consulting culture focused on technical excellence, reliability, and continuous improvement.

Get AI-Matched to This Job

Upload your resume and our AI will score how well you match this and thousands of similar roles.