IG Group Logo

IG Group

Senior Observability Engineer

Posted 2 Days Ago
Be an Early Applicant
In-Office
City of London, City and County of the City of London, England, GBR
Senior level
In-Office
City of London, City and County of the City of London, England, GBR
Senior level
Own and evolve IG's Honeycomb-centered observability platform; drive OpenTelemetry instrumentation across Java, Python and C++ services; define SLOs and alerting; participate in incident response and post-incident reviews; mentor engineers and produce runbooks, tooling, and standards to improve telemetry and reliability across a globally distributed microservices estate.
The summary above was generated by AI

Job Title

Senior Observability Engineer

Job Description

Senior Observability Engineer 

Location: London

Employment type: Permanent, Full Time 

Reporting into: Senior Engineering Manager – SRE and Observability 

About IG Group 

IG Group (LSE: IGG) is a leading global fintech company, established in 1974 and headquartered in London. As a constituent of the FTSE 100, IG Group provides dynamic online trading platforms and a robust educational ecosystem, empowering ambitious individuals worldwide in their pursuit of financial freedom. With operations spanning 18 countries across Europe, Africa, Asia-Pacific, the Middle East, and North America, IG Group offers clients access to approximately 19,000 financial markets, including shares, forex, indices, and commodities. 

 

IG Group is undergoing a significant transformation, led by CEO Breon Corcoran, appointed in early 2024. Our focus is on three areas: delivering quality products to better meet customer needs, embedding a high-performance culture across the organisation, and being more efficient and scalable through digitisation – all in service of growing our user base and revenue on a sustainable basis. 

About the role 

IG Group’s systems move billions of dollars every day – and our clients expect them to be fast, reliable, and transparent. As an Observability Engineer, you will own the platforms and practices that give IG’s engineering teams deep, real-time insight into how those systems behave. This is a high-impact, hands-on role at the centre of our reliability engineering agenda: building on Honeycomb and OpenTelemetry, driving instrumentation across a globally distributed microservices estate, and partnering directly with development teams to turn telemetry data into better, faster software. 

About the team 

This role sits within the Observability team, part of IG’s broader SRE and Platform Engineering function. The team is responsible for the tools, platforms, and standards that enable engineering teams across IG to understand and improve system behaviour at scale. You will report into the Senior Engineering Manager for SRE and Observability and work as an individual contributor, partnering closely with development squads, platform engineers, and incident response teams. The Bengaluru team is deeply integrated into IG’s global engineering community, with real ownership and scope to shape how observability is practised across the group. 

Key responsibilities 

Platform ownership 

  • Build, maintain, and evolve IG’s Honeycomb-centric observability platform, ensuring it is reliable, scalable, and fit for a complex, globally distributed trading environment. 

  • Define platform standards, data models, and integration patterns for telemetry collection, storage, and querying across the estate. 

Instrumentation and telemetry 

  • Drive OpenTelemetry adoption across engineering teams, providing hands-on guidance and reusable instrumentation patterns for services built in Java, Python, and C++. 

  • Partner with development teams to improve telemetry coverage, ensuring meaningful traces, metrics, and logs are in place across critical services and user journeys. 

SLOs 

  • Apply a strong understanding of SLOs, burn rates, and alert triggers to help service owners define meaningful reliability targets and translate them into actionable observability signals. 

  • Work closely with SRE teams to drive SLO adoption across the organisation, providing guidance and practical support to help engineering teams embed reliability targets into their day-to-day ways of working. 

Incident response 

  • Join the support rota and incident response, using observability tooling to accelerate diagnosis and reduce mean time to resolution. 

  • Lead post-incident reviews that produce actionable improvements to both systems and observability coverage. 

Enablement and community 

  • Mentor and upskill engineers across IG on observability principles and practices, raising the bar for how teams instrument, monitor, and debug their services. 

  • Develop training materials, runbooks, and best-practice guides that scale observability knowledge across the engineering organisation. 

Role requirements 

  • Proven hands-on experience with Honeycomb or a similar observability tool such as Grafana, including dataset design, query building, and using it as a primary tool for production debugging and reliability analysis. 

  • Strong practical experience implementing OpenTelemetry instrumentation in one or more of Java, Python, or C++, including custom collectors, exporters, and sampling strategies. 

  • Experience working in complex, distributed microservices environments with high transaction volumes or strict reliability requirements. 

  • Strong communication and collaboration skills – able to work effectively with development teams and influence engineering practice without direct authority. 

  • Practical experience with Terraform for managing observability infrastructure as code, including provisioning and maintaining platform components in a cloud environment. 

  • Ability and willingness to cover UK working hours to support collaboration with IG’s London-based engineering teams. 

  • 5–8 years of relevant experience in observability, SRE, or platform engineering roles. 

Desirable 

  • Experience with cloud platforms such as AWS or GCP, particularly in the context of observability and infrastructure monitoring. 

  • Familiarity with other observability tooling such as Grafana, Prometheus, or Splunk, and experience migrating or consolidating observability stacks. 

  • Exposure to fintech, financial services, or other regulated industry environments. 

  • Experience contributing to or maintaining open-source observability projects or OTel instrumentation libraries. 

 

The Perks

Your growth fuels our success! Thrive with tailored development programs, mentoring opportunities with leaders, and clear career progression. Expand your network through committees, sports and social clubs. Enjoy extra time off for volunteering and community work.

  • Competitive salary
  • Flexible Benefits Package on top of your salary (12%)
  • Private medical cover for you and your family
  • Life insurance
  • Contribution to gym memberships
  • 25 Days holiday, with 1 additional day off to celebrate your Birthday & 2 additional days off a year for voluntary work (28 in total
  • The option to buy or sell holiday days.
  • Unlimited access to the LinkedIn Learning Platform
  • A comprehensive global and local onboarding process
  • Employee-led LGBTQ+, Women’s, Black and Parents & Carers networks with an annual budget for organising events & projects that foster an open, diverse and inclusive culture
  • Enhanced primary (maternity), secondary (paternity), and shared parental pay and leave, as well as a range of support and benefits for parents
  • Option to participate and create ESG initiatives based on IG Brighter Future Fund

Number of openings

1
HQ

IG Group London, England Office

Cannon Bridge House, 25 Dowgate Hill, London, United Kingdom, EC4R 2YA

Similar Jobs

8 Days Ago
In-Office or Remote
United Kingdom
Senior level
Senior level
Cloud • Information Technology • Software • Infrastructure as a Service (IaaS)
Build ingestion pipelines for logs and metrics, scalable alerting engines, and observability APIs. Interface with product teams and develop microservices using Golang and Rust.
Top Skills: AnsibleGoGraphQLGrpcRustTerraformTypescript
8 Days Ago
In-Office or Remote
United Kingdom
Senior level
Senior level
Software
The Senior Infra Engineer will build and maintain ingestion pipelines, scalable alerting engines, and observability APIs, while ensuring resilience and scalability in infrastructure. They will work with tools like Golang, Rust, Terraform, and Ansible, documenting requirements and interfacing with product teams.
Top Skills: AnsibleGoGraphQLGrpcRustTerraformTypescript
2 Hours Ago
Hybrid
London, Greater London, England, GBR
Senior level
Senior level
Digital Media • Gaming • Information Technology • Software • Sports • Esports • Big Data Analytics
Develop, implement, and deploy statistical and machine learning sports models and pricing engines in Python; engineer sports data assets; build automated tests; collaborate with Trading, Product, Engineering, and QA; validate data flows and integrations; and mentor junior data scientists.
Top Skills: Ci/CdDockerKubernetesPythonVersion Control

What you need to know about the London Tech Scene

London isn't just a hub for established businesses; it's also a nursery for innovation. Boasting one of the most recognized fintech ecosystems in Europe, attracting billions in investments each year, London's success has made it a go-to destination for startups looking to make their mark. Top U.K. companies like Hoptin, Moneybox and Marshmallow have already made the city their base — yet fintech is just the beginning. From healthtech to renewable energy to cybersecurity and beyond, the city's startups are breaking new ground across a range of industries.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account