Skip to content
Send us a role
All live jobs

Engineering

Senior Observability Engineer

Design observability strategy for major Australian enterprises, working across Datadog, OpenTelemetry, and Amazon Bedrock from day one.

Role details

Design observability strategy for major Australian enterprises, working across Datadog, OpenTelemetry, and Amazon Bedrock from day one.

You'll build and own observability solutions across complex, cloud-native environments for some of Australia's largest organisations. This goes well beyond setting up dashboards and alerts. You'll connect platform reliability, customer experience, and business performance into a single coherent picture, then help clients act on it.

The work is genuinely varied. One engagement might involve standing up OpenTelemetry instrumentation across a distributed microservices architecture on AWS. The next might be helping a client define their SLO framework and error budgets from scratch. You'll also get early access to emerging territory, specifically observability patterns for AI and agentic workloads running on Amazon Bedrock AgentCore, which is still relatively uncharted for most engineering teams.

This is a consulting environment, so you'll be client-facing and expected to translate what you're building into language that resonates with both engineering leads and business stakeholders. The upside is real variety across industries and architectures. You won't spend years on the same platform. The team is engineering-led, meaning you're designing and building, not just advising and handing off.

What You'll Do

  • Design and implement enterprise observability strategies across applications, infrastructure, distributed systems, and cloud-native AWS workloads.
  • Define SLIs, SLOs, error budgets, and reliability KPIs, then build dashboards and operational views that make those metrics meaningful.
  • Instrument telemetry pipelines using OpenTelemetry, integrate monitoring-as-code into CI/CD workflows, and reduce alert fatigue through smarter alerting design.
  • Assess client observability maturity, identify gaps, and develop roadmaps that move teams toward genuine operational intelligence.

What You'll Need

  • Hands-on experience designing observability solutions in complex enterprise environments, using platforms such as Datadog, Grafana/Prometheus, Splunk, New Relic, or Elastic/ELK Stack.
  • Strong OpenTelemetry experience across instrumentation, collectors, metrics, traces, and logs.
  • Solid AWS background with cloud-native and distributed architectures, plus experience with Kubernetes, containers, microservices, and Terraform.
  • Experience defining SLIs, SLOs, and reliability metrics, with a working understanding of SRE principles and incident management practices.

About the Company

They're an Australian engineering consultancy that designs, builds, and scales modern technology platforms for major enterprises. Their work spans cloud, AI, security, data, and observability. They're engineering-led, which means you're in the build, not just the boardroom.

Apply

Please click the 'Apply' button. Don't worry if your CV isn't up to date - just send what you have.