Systems Engineer - ELK Observability (Infra Delivery)
Date: 25 Sept 2026
Location: SG
Company: Synapxe
Position Overview
Lead and manage the Synapxe Observability Platform, overseeing its strategy, architecture, operations, and continuous enhancement across monitoring, logging, metrics, tracing, visualization, and AIOps capabilities. Drive platform reliability, performance, and operational excellence through technical leadership, governance, automation, and adoption of emerging observability technologies, while collaborating with cross-functional teams and mentoring engineers to foster technical excellence and innovation.
Role & Responsibilities
- Partner with vendors and internal stakeholders to design, develop, and enhance the Synapxe Observability Platform, including monitoring, logging, metrics, distributed tracing, visualization, and AI/ML functionalities.
- Lead the onboarding and integration of infrastructure, network, security, and application systems into the observability platform.
- Drive initiatives in anomaly detection, event correlation, root cause analysis, predictive analytics, self-healing automation, capacity optimization, noise reduction, and outage prediction.
- Design and develop dashboards, reports, and visualization capabilities that provide a unified view of infrastructure, security, application, and end-user experience metrics.
- Establish monitoring standards, alert strategies, service health indicators, and operational KPIs to improve service reliability and reduce Mean Time to Detect (MTTD) and Mean Time to Resolve (MTTR).
- Provide technical consultancy and support to stakeholders in report customization, dashboard development, and observability adoption.
- Lead incident analysis, trend analysis, and continual service improvement initiatives based on observability insights.
- Manage, mentor, and provide technical guidance to team members and junior engineers. Partner with vendors and internal stakeholders to design, develop, and enhance the Synapxe Observability Platform, including monitoring, logging, metrics, distributed tracing, visualization, and AI/ML functionalities.
- Lead the onboarding and integration of infrastructure, network, security, and application systems into the observability platform.
- Drive initiatives in anomaly detection, event correlation, root cause analysis, predictive analytics, self-healing automation, capacity optimization, noise reduction, and outage prediction.
- Design and develop dashboards, reports, and visualization capabilities that provide a unified view of infrastructure, security, application, and end-user experience metrics.
- Establish monitoring standards, alert strategies, service health indicators, and operational KPIs to improve service reliability and reduce Mean Time to Detect (MTTD) and Mean Time to Resolve (MTTR).
- Provide technical consultancy and support to stakeholders in report customization, dashboard development, and observability adoption.
- Lead incident analysis, trend analysis, and continual service improvement initiatives based on observability insights.
- Manage, mentor, and provide technical guidance to team members and junior engineers.
Requirements
- Degree in Computer Science or Computer Engineering or equivalent
- Minimum 7 years of experience in enterprise-level infrastructure, platform engineering, or IT operations environments.
- Minimum 2 years of hands-on experience in enterprise monitoring, observability, and logging platforms.
- Minimum 3 years of experience in server, virtualization, cloud, network, or security infrastructure technologies.
- Minimum 1 year of experience in programming, scripting, and automation technologies (e.g., Python, PowerShell, Bash, Ansible, Terraform).
- Strong understanding of observability concepts, including metrics, logs, distributed tracing, alerting, dashboards, and AIOps.
- Experience with enterprise observability platforms such as Splunk, Elastic, Dynatrace, Datadog, Grafana, Prometheus, OpenTelemetry, or equivalent technologies.
- Experience with RHEL
- Strong analytical, troubleshooting, problem-solving, and pattern-recognition skills.
- Strong communication, stakeholder management, and documentation skills.
- Proven ability to lead technical teams and mentor engineers.
- Possess a growth mindset with a strong interest in emerging technologies and industry best practices.
- 2 years contract
Apply Now
NOTE: It only takes a few minutes to apply for a meaningful career in HealthTech - GO FOR IT!!
#LI-SYNX08