OLIX Logo

OLIX

Staff / Senior Software DevOps Engineer

Posted 22 Days Ago
In-Office
London, Greater London, England, GBR
Senior level
In-Office
London, Greater London, England, GBR
Senior level
Design, build, and own CI/build/test pipelines for a compiler/runtime/simulator stack; scale and parallelize large test suites, manage heterogeneous CI runner fleets and scarce hardware, implement perf-regression baselines and observability, and define hermetic, safe-by-default build and test standards to maximize developer velocity.
The summary above was generated by AI
About OLIX

AI is growing faster than any technology in history and the explosion in demand has created a massive infrastructure gap; we can no longer build chips or power stations fast enough to keep up. The industry is still leaning on a ten-year-old hardware blueprint that has reached its limit. A new paradigm that is faster and more efficient will be the biggest economic opportunity of the next century and create the most important company of the next decade. The OLIX Decode Accelerator 1 (DX-1) is the first accelerator architected specifically for decode. Rack-scale co-design of logic, data movement, packaging, optics and interconnect enables a step change in system level performance.

The Role

We're searching for a Staff/Senior Software DevOps Engineer to own the build, test, and CI flows that the entire DX-1 software stack, including the compiler, runtime, simulator, and framework integration, depends on. Ours is a large test suite that asserts token-exact correctness against golden references, and it has to run across scarce, expensive resources that span both simulation compute and hardware-in-the-loop testing, including simulator and emulator boxes alongside DX-1 and prototype-platform boards. Your mission is to keep that system fast, trustworthy, observable, and affordable as the test suite, the team, and the resource pool all grow.

This is a build-and-test role, not product-serving SRE. You'll work where CI, the runner fleet, and the test hardware meet, partnering closely with the infrastructure, compiler, runtime, simulator, and modelling teams. At the Senior/Staff level, your impact is the velocity of every engineer who depends on this system: how fast they get a trustworthy signal, how rarely they wait on a machine or a flaky run, and how much they can self-serve without coming to you. That leverage, through the standards, platforms, and shared resource model others build on, is what we're hiring for far more than any single system you ship.

Responsibilities

Own the Build & Test Pipelines: Design, build, and own CI pipelines and test execution across PR, merge, and nightly lanes that gate the entire software stack, balancing fast feedback with coverage and cost.

Scale Test Execution: Split a large, slow suite into staged lanes, parallelize it with real test isolation, and cache aggressively using content-addressed keys so feedback stays fast and cost-effective as the suite and the team grow, rather than relying on simply adding more machines.

Manage the Fleet & Scarce Resources: Run CI across a heterogeneous fleet of cloud and self-hosted machines, and give the team fair, monitored, fail-fast shared access to scarce and expensive hardware, keeping it reliable, well utilized, and never a silent bottleneck.

Build the Performance & Readiness Signal: Stand up performance regression baselines the team trusts using pinned hardware, rolling baselines, sound metric aggregation, and deterministic testing. Turn CI and test signals into CI health and product readiness dashboards that drive real decisions.

Own Software Observability: Choose the metrics store that scales to many time series across daily runs with long-lived history, making dashboards for observable software.

Set Standards: Define the flows that keep builds and test runs hermetic and reproducible, and make the system fail closed while containing the blast radius when something is misconfigured or a job is untrusted.

Skills & Experience
  • Experience in build/test infrastructure, CI/CD, developer productivity, or large-scale systems and release engineering, with demonstrated end-to-end ownership of a large test or CI system

  • Experience scaling a large test suite through staged lanes, parallelism with real isolation, content-addressed caching, and maintaining fast, cost-effective feedback as the suite grows

  • Experience managing heterogeneous CI runner fleets across cloud and on-prem environments, including VMs, containers, and bare-metal hardware. Ability to provide teams with shared, monitored access to scarce or expensive resources such as custom accelerators, FPGA/prototype rigs, lab hardware, or contended compute, including reservations, remote access, hardware-in-the-loop testing, and artifact portability across heterogeneous hosts

  • Experience supporting performance analysis with trustworthy regression baselines, deterministic testing, noise handling, sound metric aggregation (geometric mean, arithmetic mean, and median), and fast attribution and bisection

  • Experience building observability and metrics platforms, including CI health and product readiness dashboards, metrics stores that scale to long-lived, high-cardinality time series, and self-service access for engineers

  • Strong scripting and systems programming skills (for example, Python plus a systems language), along with proficiency in containers, Linux, and cloud infrastructure such as AWS

  • A reproducible, safe-by-default mindset, including hermetic builds, least privilege, fail-closed defaults, and blast radius control

  • Excellent communication skills and the ability to align and influence cross-functional teams, including compiler, runtime, and modelling teams, without relying on formal authority

  • Bachelor's degree or higher in Computer Science, Electrical Engineering, Mathematics, or a related field

Nice to Have
  • GitHub Actions or comparable CI at scale; scaling CI runner fleets on cloud infrastructure (e.g. AWS); hardware-in-the-loop or lab automation for custom silicon or FPGA bring-up; time-series and observability stacks (Prometheus/Grafana, Datadog, columnar warehouses)

  • Adjacent depth is welcome: HPC / cluster batch scheduling, release engineering, or developer-productivity platforms

Compensation & Equity
  • Competitive Salary: Commensurate with your experience, skills, and location

  • Equity & Ownership: Meaningful stock options. You’re not just joining the mission; you’re owning a piece of it

  • Proximity Bonus: We value your time. To minimise your commute and maximise your life, we offer an annual Living-Local Bonus if your residence is within 20 minutes of the office

  • Retirement Benefits: Employer-contributed retirement plans to help you build long-term financial security

 

Due to U.S. export control regulations, candidates’ eligibility to work at OLIX depends on their most recent citizenship or permanent residency status. We are generally unable to consider applicants whose most recent citizenship or permanent residence is in certain restricted countries (currently including Iran, North Korea, Syria, Cuba, Russia, Belarus, China, Hong Kong, Macau, and Venezuela). Applicants who have subsequently obtained citizenship or permanent residency in another country not subject to these restrictions may still be eligible.

 
 

Similar Jobs

2 Hours Ago
Hybrid
Senior level
Senior level
Fintech • Professional Services • Consulting • Energy • Financial Services • Cybersecurity • Generative AI
Lead the design and evolution of enterprise CIAM capabilities: define product direction, architecture, integration patterns, and engineering standards. Deliver customer identity solutions (authentication, MFA, passkeys, onboarding, identity orchestration) and drive platform optimisation, technical governance, and stakeholder alignment across business, security, and engineering teams.
Top Skills: APIsAWSAzureForgerockGoogle Cloud PlatformMfaOauth 2.0OktaOpenid Connect (Oidc)PasskeysPing IdentitySAMLScim
2 Hours Ago
Hybrid
Senior level
Senior level
Fintech • Machine Learning • Payments • Software • Financial Services
Provide legal guidance on employment and commercial contract matters for Capital One Canada. Support business expansion, draft/review/negotiation of third-party agreements, ensure compliance with OSFI B10 guidelines, conduct legal research, advise on legislation and regulatory expectations, develop practical legal guidance for business partners, and participate in cross-functional projects to manage legal risk.
Senior level
Fintech • Machine Learning • Payments • Software • Financial Services
The French Localization Specialist translates and localizes content, ensuring grammatical accuracy and compliance with brand standards, while collaborating with various teams.
Top Skills: Adobe Acrobat ProAdobe WorkfrontGenerative Ai TechnologyGoogle WorkspaceMachine Translation ToolsMS OfficeTrados Enterprise

What you need to know about the London Tech Scene

London isn't just a hub for established businesses; it's also a nursery for innovation. Boasting one of the most recognized fintech ecosystems in Europe, attracting billions in investments each year, London's success has made it a go-to destination for startups looking to make their mark. Top U.K. companies like Hoptin, Moneybox and Marshmallow have already made the city their base — yet fintech is just the beginning. From healthtech to renewable energy to cybersecurity and beyond, the city's startups are breaking new ground across a range of industries.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account