Ensono Logo

Ensono

Site Reliability Engineer

Reposted 16 Days Ago
Easy Apply
Remote or Hybrid
Hiring Remotely in United Kingdom
Mid level
Easy Apply
Remote or Hybrid
Hiring Remotely in United Kingdom
Mid level
Design, deploy, and maintain reliable cloud-native infrastructure using Terraform, Azure, and Kubernetes. Troubleshoot incidents, lead incident resolution and post‑mortems, improve CI/CD and monitoring, reduce toil, manage security risks, and lead client and supplier engagements to expand SRE practices.
The summary above was generated by AI
Who are we?

Ensono is a global technology services provider dedicated to helping organizations navigate the complexity of digital transformation. Through Ensono Product, Consulting & Technology, our dedicated consulting arm, we partner with clients to design, build, and modernize digital capabilities across application development and modernization, data platforms and AI, and identity and access management. As we continue to expand our global consulting footprint, we are committed to delivering innovative, high‑quality outcomes that enable our clients to move faster, smarter, and more securely.

About the role and what you'll be doing: 

We are seeking an experienced Site Reliability Engineer (SRE) with expertise in Infrastructure as Code tools like Terraform, core CI/CD tools such as Azure DevOps, and monitoring tools including DataDog and AWS CloudWatch. The ideal candidate will have commercial experience in technologies like Dotnet or Java, and be skilled in troubleshooting, incident resolution, and improving service and change management processes. Strong leadership in client-facing discussions and engagement with third-party suppliers is essential. An SRE Foundation certificate and a cloud provider associate-level certification are highly beneficial. 

  • Commercial experience and proficiency with industry standard: 

  • IAC tooling (Terraform preferably, or ARM/bicep and CloudFront) 

  • Core CI/CD Tooling (Azure DevOps, GitHub Actions or Gitlab) 

  • Monitoring Tooling (DataDog, Splunk, NewRelic, Azure Monitor, AWS CloudWatch) 

  • Commercial experience in at least one core technology (Dotnet, Java, AI/Data Engineering, Golang) 

  • Troubleshooting issues and identifying systemic failings indicated by incidents/failures 

  • Implementing fixes 

  • Proposing solutions for reducing toil 

  • Providing leadership in the Incident resolution process, including creating and maintaining documentation, and providing key input to Post-mortem analysis 

  • Improving Service Requests and Change Management processes, both technically and through stakeholder management). 

  • Participate in the process for, and Proactively mitigate risks in a Security management process (Vulnerabilities in Code, Infrastructure, Dependencies) 

  • Lead discussion in client-facing meetings and discussions around the SRE process, and identifying areas for increasing SRE footprint. 

  • Engaging with suppliers and 3rd parties for support, requests and opportunities 


We want all new Associates to succeed in their roles at Ensono. That's why we've outlined the job requirements below. To be considered for this role, it's important that you meet all Required Qualifications. If you do not meet all of the Preferred Qualifications, we still encourage you to apply.  


Required Qualifications  

  • 3-9 Years experience  

  • Bachelor’s degree (or equivalent) in computer science or related discipline 

  • SRE Foundation certificate (DevOps Institute) and a Cloud provider (AWS, Azure, GCP) 'associate'-level certification, or completed during the probationary period. 

  • Proficiency in Azure and Kubernetes, with hands-on experience in managing and deploying applications. 

  • Expertise in Infrastructure as Code (IaC) using Terraform for efficient and scalable infrastructure management. 


Preferred Qualifications 

  • Certified Kubernetes Administrator / Application Developer 

  • Certified Azure DevOps Engineer   

  • Experience with monitoring tools such as NewRelic or Splunk for effective system monitoring and alerting. 

  • Familiarity with Harness for continuous delivery and deployment processes. 

  • Strong programming skills in .Net, Java, or JavaScript for developing robust and scalable applications 

Ensono Spelthorne, England Office

One London Road, , United Kingdom , Spelthorne, United Kingdom, TW18 4EX

Similar Jobs

7 Days Ago
Remote or Hybrid
London, Greater London, England, GBR
Mid level
Mid level
Cloud • Software
Design, operate, and scale large distributed systems for telemetry processing. Build automation, use AI tooling to reduce toil, ensure availability and disaster recovery, participate in on-call incident response, troubleshoot production AWS/Kubernetes environments, and collaborate with application teams to meet SLOs/SLAs.
Top Skills: Ai ToolingAWSGnu/LinuxGoKubernetesPythonTerraform
6 Hours Ago
Remote
United Kingdom
Entry level
Entry level
Information Technology
Own reliability and observability for critical product journeys in a distributed consumer mobile product. Build metrics, dashboards, alerts, SLIs, and SLOs; improve monitoring, logging, tracing, and incident response; act as a first responder for P0/P1 incidents; investigate and mitigate production issues; coordinate escalations; lead postmortems; and drive infrastructure and tooling improvements using Node.js, TypeScript, and AWS.
Top Skills: AWSCloudflareCloudwatchNode.jsReact NativeSentryTypescript
Yesterday
Remote or Hybrid
United Kingdom
Entry level
Entry level
Utilities
Leads the design, development, and operation of reliable production systems. Drives improvements in observability, alerting, incident management, operational readiness, scalability, security, and maintainability. Supports teams during incidents, guides testing and architecture decisions, promotes learning from incidents, and mentors engineers across the organization. Partners with product teams to make customer-focused, data-informed, and cost-efficient technical decisions while advancing agile delivery practices.
Top Skills: AgileAIAlertingCloud-Native ArchitectureIncident ManagementObservabilitySaaSSecure Coding

What you need to know about the London Tech Scene

London isn't just a hub for established businesses; it's also a nursery for innovation. Boasting one of the most recognized fintech ecosystems in Europe, attracting billions in investments each year, London's success has made it a go-to destination for startups looking to make their mark. Top U.K. companies like Hoptin, Moneybox and Marshmallow have already made the city their base — yet fintech is just the beginning. From healthtech to renewable energy to cybersecurity and beyond, the city's startups are breaking new ground across a range of industries.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account