incident.io Logo

incident.io

Director of Engineering

Posted One Month Ago
Be an Early Applicant
Hybrid
London, Greater London, England, GBR
Senior level
Hybrid
London, Greater London, England, GBR
Senior level
Own reliability, scalability, cost, and security of incident.io infrastructure. Build and lead TechOps/infrastructure/security function, work cross-functionally, respond to outages and security incidents, and scale secure AI infrastructure for sensitive customer data.
The summary above was generated by AI
A little about us

incident.io is the AI software reliability platform trusted by engineering teams at Netflix, Etsy, and 2,000+ companies who can't afford downtime.

The stakes are real, the talent is insane, and somehow it's a genuinely lovely place to work. We all pull our weight, go the extra mile, and put in the bit of magic most companies don't have time for, all while having fun and never taking ourselves too seriously.

We've raised $96 million from Index Ventures, Insight Partners, and Point Nine, and we're growing fast. There's never been a better time to join.

The role in one line

Right now, the ceiling on how fast incident.io can scale isn't ambition or demand, it’s whether Infrastructure, Security, and IT can keep up; this role exists to close that gap.

What that looks like day to day

You'll own the reliability, scale, cost, and security of the infrastructure that runs incident.io. Customers open this infrastructure at 3am when something in their own product breaks. You're not just running infra and security, you're running the thing some of the best engineering teams in the world rely on when their product is down.

TechOps sits inside the same remit, you'll own the function and ensure every business system and integration is safe. Security and IT posture matter here because customers check them closely before they trust us with their incidents.

Day to day you'll partner with Product Development and pretty much every function in the business and if infrastructure or security breaks, it breaks for everyone.

You'll also be working on the frontier of new technology. Investigations, our AI product, runs multi-agent analysis across logs, deployments, and historical incidents to find root cause in minutes. That runs on infrastructure that needs to scale fast and stay locked down, since it handles some of our customers' most sensitive operational data. Scaling and securing that will be one of the first things you own.

This role doesn't exist yet at incident.io, so you'll be the one shaping it.

Who you are
  • You came up through infrastructure and found your calling in security. You can talk scaling, reliability, and cost, but security is where you specialise.

  • On a bad day, you're the one unblocking coordination in the middle of an outage or a security incident.

  • You've built a team or function close to scratch at a company our size, and spent time somewhere much bigger too. You know what mature looks like and how to get there without just copying it.

  • You enjoy working out how to stop an adversary before they cause damage.

  • You're comfortable owning ambiguity. This role doesn't exist yet, and you'd rather shape it than be handed a finished job spec.

  • You think in systems, not silos. Infrastructure, security, and IT are one problem to you, not three.

  • You can go deep with engineers on architecture, and explain to a non-technical exec why it matters.

Supporting you

We work hard and we think life outside work matters just as much. Our benefits pack is built to support both.

  • Private medical insurance. Seriously good cover - we want you and the people you love to be looked after.

  • Competitive annual leave. Showing up at your best requires switching off, and we make sure you have time to do that.

  • First Friday of every month off. Yes, seriously.

  • Enhanced pension. We put real money in, because future-you deserves better than an afterthought.

  • Meaningful equity. We're rapidly scaling, and everyone who helps shape the outcome should share in it.

  • Unlimited AI spend. For everyone, not just engineers. We're all-in on AI across the company, and we expect you to be too.

  • Generous parental leave. The early days with a new baby matter more than anything we're doing here, and we want you to be present for them.

  • Two budgets that have your back. £1000 to invest in your setup, £500 a year to invest in yourself.

HQ

incident.io London, England Office

London, United Kingdom

Similar Jobs

5 Days Ago
Hybrid
London, Greater London, England, GBR
Expert/Leader
Expert/Leader
Cloud • Information Technology • Security • Software • Cybersecurity
Leads and scales Cloudflare’s regional or domain-based Customer Engineering organization. Responsibilities include managing managers and multiple teams, owning revenue and commercial targets, shaping go-to-market strategy, coaching technical leaders, improving operational systems with AI, influencing product and cross-functional priorities, and serving as an executive technical sponsor for strategic customers. The role requires extensive technical go-to-market leadership experience, enterprise commercial acumen, executive communication, and fluency in networking, security, performance, or AI platforms.
Top Skills: Ai WorkflowsBgpBot ManagementCachingCloud PlatformsContent Delivery Networks (Cdn)Ddos MitigationDeveloper ToolingDnsEdge ServicesHttp(S)LlmsSaseSsl/TlsTcp/IpTraffic SteeringWafZero Trust
2 Days Ago
In-Office
Expert/Leader
Expert/Leader
Industrial • Manufacturing
Leads North American project and application engineering for engineered refrigeration, compression, and heat pump solutions. Oversees engineering strategy, technical proposals, project execution, product development, global collaboration, budgets, quality, safety, compliance, and organizational development. Partners with senior leaders, Sales, Operations, Supply Chain, Finance, and global engineering teams to drive profitable growth, innovation, standardization, and operational excellence.
Top Skills: Compression TechnologiesEngineering ControlsHeat PumpsIndustrial Refrigeration SystemsProcess SystemsQuality Management SystemsThermodynamics
2 Days Ago
Hybrid
London, Greater London, England, GBR
Entry level
Entry level
Fintech • Payments • Financial Services
Leads a sizeable engineering organization and sets strategy for technology, architecture, people, and ways of working. Builds high-performing teams through Engineering Managers, drives organizational change, improves engineering effectiveness using data and AI, and influences major technical decisions. Partners with Product, Commercial, and Technology leaders to deliver business outcomes and raise engineering standards across the company. The role is hybrid in London, requiring office attendance two to three days per week.
Top Skills: AI

What you need to know about the London Tech Scene

London isn't just a hub for established businesses; it's also a nursery for innovation. Boasting one of the most recognized fintech ecosystems in Europe, attracting billions in investments each year, London's success has made it a go-to destination for startups looking to make their mark. Top U.K. companies like Hoptin, Moneybox and Marshmallow have already made the city their base — yet fintech is just the beginning. From healthtech to renewable energy to cybersecurity and beyond, the city's startups are breaking new ground across a range of industries.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account