Principal Observability Engineer
12-month contract (extension possible)
Current opening
-
Lead the architecture, design, and implementation of enterprise observability solutions and monitoring frameworks.
-
Develop and enhance dashboards, alerting, and monitoring capabilities using Splunk, Splunk Observability, OpenTelemetry, Grafana, Prometheus, and AWS CloudWatch.
-
Design and support observability across cloud and Kubernetes-based environments, driving platform performance and reliability.
-
Develop infrastructure automation using Infrastructure as Code (IaC), scripting, and configuration management tools.
-
Provide technical leadership, establish observability best practices, and mentor engineering teams.
A large global enterprise organization
-
10+ years of experience in Platform Engineering, Site Reliability Engineering (SRE), Systems Engineering, or a related infrastructure discipline.
-
Strong hands-on experience with Splunk and modern observability platforms, including Splunk Observability, OpenTelemetry, Grafana, Prometheus, and AWS CloudWatch.
-
Deep expertise with AWS, Kubernetes/OpenShift, and infrastructure automation tools such as Terraform, Ansible, Python, or Bash.
-
Proven experience designing scalable monitoring and observability solutions within large enterprise environments.
-
Strong communication and leadership skills with the ability to collaborate across technical and business teams.
-
Experience with additional cloud platforms (Azure or GCP), Red Hat technologies, virtualization, storage, backup, or identity management is considered an asset.
-
Relevant certifications such as Splunk, AWS, or Kubernetes are highly desirable.
isgSearch does not use artificial intelligence throughout the hiring process.