Fractile Logo

Fractile

Infrastructure Software Engineering – Platform & Build 1

Posted Yesterday
Be an Early Applicant
In-Office
London, Greater London, England, GBR
Senior level
In-Office
London, Greater London, England, GBR
Senior level
Own and scale CI infrastructure, multi-language Bazel monorepos, and developer platform systems for AI hardware and software engineering. Responsibilities include building reproducible CI pipelines, optimizing performance across compute clusters, maintaining observability and alerting, supporting incident response, implementing infrastructure as code, and improving reliability and developer productivity. The role requires strong experience with CI/CD, build systems, DevOps/SRE practices, infrastructure tooling, and Python, Go, or Rust development.
The summary above was generated by AI

Infrastructure Software Engineering – Platform & Build
Location: London

About Fractile

Fractile was founded in 2022 on the bet that, eventually, the world's most capable AI systems would be limited in their impact by the time taken to produce useful outputs. We bet everything on the logical conclusion: that the only way to truly unlock this latent value, to make speed viable at scale, was to radically re-invent the hardware that we run our frontier AI models on. Ever since, we have been building chips and systems that tackle this problem: how to efficiently generate output at thousands of tokens per second, while handling the complexity and capacity challenges of operating large models at very long contexts.

The workloads that push to the limits of the current frontier are already transformational; it is the technical and economic limits on inference speed that are constraining progress. The defining work of the 21st century will be marked by the engine of inference delivering immense and diffuse chains of intellectual inquiry, in drug discovery, in software engineering, in materials discovery, in any field where progress is driven by deep reasoning and intelligence to resolve complex problems.

Fractile is seeking to increase the clock speed of global progress, one chip at a time. We’ve recently raised $220M from Founders Fund and Accel and the most important work lies ahead. Come and join the mission!

The Role

As an Infrastructure Platform Engineer you will be a core part of our technical team working alongside the hardware, software and research teams to solve cutting-edge problems in AI hardware, ranging from verification to simulation, alongside the best of the best (DeepMind, Apple, Meta…).

About the Software organisation at Fractile

Infrastructure sits within the Software organisation at Fractile, which is responsible for developing a full software stack for our groundbreaking AI inference systems. That's everything from ML compilers, device drivers and systems firmware, application level runtime and ecosystem integrations, ML and compute libraries, great developer tooling and a full portfolio of simulators, through to datacenter scale workload deployment solutions. At Fractile, we know that a fantastic software stack is a critical and central part of any AI inference solution and it sits at the heart of everything we're doing.

About the team and role

The Infrastructure team is responsible for the core platform and systems at Fractile. We build, scale, and optimize the mission-critical systems that power our entire engineering org, including our multi-language Bazel monorepo and CI pipelines. From hardware design and verification to kernel development and ML compilers, we enable all Fractile engineers to ship fast and reliably.

The team brings both deep build system expertise and an SRE mindset: reliability, build performance, debugging CI failures systematically, and giving the engineering organisation clear visibility into how our infrastructure is performing.

As an Infrastructure Platform Engineer, you will:

  • Own the CI infrastructure end-to-end: create, maintain and debug reproducible multi-language CI pipelines, and optimize CI performance across large compute clusters.
  • Build and maintain infrastructure observability, alerting, runbooks, and incident response workflows for Fractile's infrastructure.
  • IaC TODO
  • Scale and maintain Fractile's Bazel monorepo as we continue growing across Python, C++, Rust, SystemVerilog, and ML workloads.
  • Shape the infrastructure and platform experience for every engineer at Fractile.

About you

You bring deep infrastructure and build systems expertise alongside an SRE mindset. You are systematic in how you approach fault isolation and root cause analysis, and you apply metrics-driven thinking to build reliability and developer productivity. You thrive on variety (ML, compilers, kernel drivers, simulators, hardware verification) and you understand that your work becomes the foundation every engineer at Fractile builds on. You take that responsibility seriously.

Key Requirements

  • 5+ years in software, platform, or infrastructure engineering
  • 5+ years hands-on experience working with CI/CD for large-scale products, including monitoring, debugging performance and resolving pipeline failures at scale
  • 3+ years working with build systems, preferably modern cross-language build systems with an emphasis on correctness, high performance and extensibility, such as Bazel, Buck, Pants, Please, etc.
  • Experience with infrastructure as code (Terraform, OpenTofu, or Pulumi)
  • Experience with monitoring and observability tooling (Prometheus, Grafana, or similar)
  • Working knowledge and practice of DevOps / SRE principles: SLOs, alerting design, incident management, and on-call practice
  • Strong proficiency in at least one systems or scripting language used for tooling (Python preferred; Go or Rust also relevant), with a track record of shipping well-tested, maintainable developer tooling that other engineers rely on: CLIs, CI integrations, build and release automation

What We Offer

  • Competitive salary: A competitive salary reflective of your experience and the specialist nature of the role
  • Equity & Ownership: meaningful equity so everyone shares in the value creation
  • Benefits: Private Medical, Dental and Vision, Contributory Pension, 25 Days holiday plus bank holidays and Life/Critical Illness Insurance
  • Diverse & fun office: we believe the hardest problems get solved by the broadest range of minds. We are committed to Equal Employment Opportunity through attracting and retaining a diverse team and building an inclusive environment

Export controls

Our work involves technologies subject to UK, US and other international export control regulations. Certain roles may require additional eligibility checks to ensure compliance with applicable law. We'll be transparent about this throughout the hiring process.

Similar Jobs

55 Minutes Ago
Hybrid
London, Greater London, England, GBR
Mid level
Mid level
Fintech • Mobile • Payments • Software • Financial Services
Support external financial reporting, disclosure governance, AI governance, month-end reporting, and Workiva automation. Lead audit-readiness initiatives, prepare accounting position papers, coordinate auditor walkthroughs, improve audit processes, and own audit issue tracking. The role requires strong reporting knowledge, ACA qualification, audit experience, documentation skills, project management, and cross-functional collaboration in a fast-moving US-listed company environment.
Top Skills: Workiva
56 Minutes Ago
Hybrid
Staines, Surrey, England, GBR
Senior level
Senior level
Artificial Intelligence • Cloud • HR Tech • Information Technology • Productivity • Software • Automation
Drive new SaaS business revenue through account and territory planning, prospect research, business development, and field sales. Build C-suite relationships, map client stakeholders, lead virtual account teams, advise customers on IT roadmaps, coordinate specialist resources, negotiate deals, achieve sales targets, and promote customer success. The role requires leveraging AI in business processes and up to 50% travel.
Top Skills: AISaaS
An Hour Ago
Hybrid
London, Greater London, England, GBR
Senior level
Senior level
Artificial Intelligence • Big Data • Enterprise Web • Fintech • Software • Financial Services
Lead pricing discovery and structure medium-to-high complexity EMEA deals, drive Rate Card adoption, coach sales on negotiation, support deal analytics, and coordinate cross-functional commercial initiatives.
Top Skills: ExcelPowerPointSalesforce

What you need to know about the London Tech Scene

London isn't just a hub for established businesses; it's also a nursery for innovation. Boasting one of the most recognized fintech ecosystems in Europe, attracting billions in investments each year, London's success has made it a go-to destination for startups looking to make their mark. Top U.K. companies like Hoptin, Moneybox and Marshmallow have already made the city their base — yet fintech is just the beginning. From healthtech to renewable energy to cybersecurity and beyond, the city's startups are breaking new ground across a range of industries.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account