Lead Observability Engineer


Details:
  • Salary:
  • Job Type: Contract
  • Job Status: Full-Time
  • Location: London
  • Date: 2 days ago
Description:

Location: London, onsite 3 days per week (Sheffield as an alternative)

Rate: £tbd/day inside IR35

Duration: 6 months+

Are you a Senior Observability Engineer / SRE Lead, with demonstrable experience of assessing and defining observability and monitoring roadmaps within enterprise scale environments? If so, apply now for this new contract opportunity.

The Lead Observability Engineer / SRE Lead will be required to assess a complex hybrid estate, understand how services, platforms, infrastructure and networks should be monitored, and work across multiple internal teams, partners and suppliers to build a consolidated view of existing telemetry, monitoring and alerting capabilities.

As well as short term tactical objectives, the role will also be focussed on longer term strategic ones.

Responsibilities of the Lead Observability Engineer / SRE Lead will be to:

Discover and document existing telemetry sources, monitoring tools, dashboards and ownership
Work with technical teams and suppliers to gain access to telemetry
Deliver meaningful dashboards and service health views
Identify gaps in telemetry, monitoring and alerting coverage and implement pragmatic improvements
Define health indicators for critical business journeys
Introduce modern observability practices where practical, including SLIs, SLOs and a roadmap towards burn-rate alerting
Develop a roadmap for OpenTelemetry adoption
Assess options for a centralised telemetry platform, including Grafana Cloud
Evaluate tooling rationalisation opportunities, operating costs and telemetry economics
Define an observability target architecture, standards and implementation roadmap

The successful Lead Observability Engineer / SRE Lead will demonstrate the following:

Proven experience leading enterprise-scale observability initiatives
Strong hands-on expertise with Grafana, Grafana Cloud, OpenTelemetry and modern telemetry pipelines
Deep understanding of metrics, logs, traces, distributed tracing, alerting, SLIs, SLOs and error-budget concepts
Experience designing observability solutions across cloud PaaS, IaaS, on-premises, legacy and third-party hosted platforms
Strong knowledge of Azure observability tooling, including Azure Monitor, Log Analytics and Application InsightsIf this sounds like you, please apply to find out more.

Lead Observability Engineer / SRE Lead / Lead Site Reliability Engineer

Report this job

By sending this message I agree to GrindJob’s Terms and Conditions and Privacy Policy.

Enter your email to get a notification when similar jobs become available.

Create a job alert for Lead Engineer in London ()

By continuing, you agree to GrindJob’s T&Cs and Privacy Policy.

When applying for a job, do not provide bank account details or any other financial information.
Never make any form of payment. GrindJob is not responsible for any external website content.

Enter your email to get a notification when similar jobs become available.

Your browser does not support Cookies or JavaScript or this option is turned off in your browser settings.

How to enable Cookies and JavaScript

Your browser is out of date!

Update your browser to view this website correctly. Update my browser now

×

Please wait...
There was an error loading the page. Would you like to reload the page?