Diffractive Labs Logo

Diffractive Labs

Platform Engineer

Posted 28 Days Ago
Be an Early Applicant
Hybrid
London, Greater London, England, GBR
Mid level
Hybrid
London, Greater London, England, GBR
Mid level
Build and own cloud infrastructure and GPU compute environments for ML model training and deployment. Implement IaC, CI/CD, orchestration, containerisation, monitoring, security, and internal tooling to enable reproducible, scalable research-to-production ML workflows.
The summary above was generated by AI

What We're Looking For

We are seeking a Platform Engineer to build and own the infrastructure that underpins our AI-driven materials discovery platform. You'll work directly with world-renowned ML researchers and software engineers to accelerate real scientific breakthroughs by making model training, experimentation, and deployment fast, reliable, and reproducible.

This is a foundational hire. You'll set the patterns others build on.

You will be joining a small, highly ambitious team of world-renowned engineers, AI researchers, and materials scientists. We move fast and value people who are energised by that.

What You'll Do

  • Design, provision, and manage cloud infrastructure (AWS/GCP) using infrastructure-as-code; Terraform, Pulumi, or equivalent.

  • Own GPU compute environments for model training and inference, including cluster configuration, job scheduling, and cost optimisation.

  • Build and maintain CI/CD pipelines that support rapid model iteration, automated testing, and safe deployments.

  • Support ML workflow orchestration; experiment tracking, training run management, and data pipeline reliability.

  • Ensure reproducibility across research and production environments through containerisation and rigorous environment management.

  • Define monitoring, alerting, and incident response processes so the team can move fast without things silently breaking.

  • Implement security best practices: secrets management, IAM, network segmentation, vulnerability scanning.

  • Build internal tooling and documentation that lets researchers self-serve infrastructure without waiting on you.

Skills & Qualifications

  • 4+ years in a DevOps, Platform Engineering, or SRE role.

  • Strong proficiency with at least one major cloud provider and its core services (compute, storage, networking, IAM).

  • Hands-on experience with infrastructure-as-code and container orchestration (Kubernetes or equivalent).

  • Solid CI/CD pipeline experience, GitHub Actions, GitLab CI, or similar.

  • Proficient in Python and Bash; comfortable reading and writing code across a polyglot stack.

  • Deep Linux systems knowledge and strong networking fundamentals.

  • A bias for building things properly the first time, even under early-stage constraints.

Nice to Have

  • Experience with GPU cluster management and ML training workloads (NVIDIA, CUDA, distributed training).

  • Familiarity with MLOps tooling:

  • Experiment tracking (MLflow, Weights & Biases).

  • Workflow orchestration (Airflow, Prefect, Argo).

  • Data versioning (DVC).

  • Background in scientific computing or HPC environments.

  • Prior experience at a deep tech or computational science company.

Why Join Us

  • Work directly on infrastructure that enables AI to make real scientific discoveries.

  • Shape how we build from day one, no legacy systems, no inherited mess.

  • Collaborate with world-class researchers across materials science and machine learning.

Diffractive is building the AI Material Scientist that autonomously learns from real-world experimentation to push the boundaries of scientific discovery. We're early, moving fast, and working on problems that genuinely matter.
You'll join a small, high-calibre team where your work has real impact from day one. We're London-based with a flexible approach to how and where you work. We offer competitive salary, generous equity and benefits. You'll have a real stake in what you build and in the company's overall success.

How to Apply

If you're excited about this role and believe you could thrive in it, we'd encourage you to apply even if you may not align with every part of the job description.

Diffractive is an equal opportunities employer. We are committed to creating an inclusive environment for all employees and welcome applications from people of all backgrounds, experiences, and identities.

If you require any adjustments or accommodations at any point during the interview process please let us know - we will be happy to help.

Hit the apply button below to submit your application. We are looking forward to hearing from you!

Similar Jobs

9 Days Ago
Hybrid
London, Greater London, England, GBR
Senior level
Senior level
AdTech • eCommerce • Information Technology • Software • Travel • Generative AI
Design, build, operate, and evolve scalable, reliable, observable data lake and catalog platform services. Lead technical design, develop well-tested software, support SLOs and incident learning, collaborate with stakeholders, create implementation plans, mentor engineers, and contribute to platform architecture and open-source initiatives.
Top Skills: Amazon EmrAmazon S3Apache IcebergAWSAws GlueHadoopHiveJavaJvmKotlinSpark
Yesterday
Hybrid
London, Greater London, England, GBR
Entry level
Entry level
Financial Services
Build and maintain DevSecOps tooling for AWS cloud environments, including threat modeling, security guardrails, policy-as-code, vulnerability analysis, compliance reporting, and runtime security visibility. Develop automation for security findings and software supply chain controls, support audits and incident response, and provide remediation guidance to engineering teams. Collaborate with platform engineers, product squads, risk, and governance teams while using authorized AI-assisted development tools securely and validating generated outputs.
Top Skills: AWSAws IamAws Key Management ServiceBashCheckovCloudtrailGoGuarddutyIso 27001KubernetesKyvernoNistPci-DssPythonSecurity HubSemgrepSentinelSnykStrideTerraformTrivyWiz
Yesterday
Hybrid
London, Greater London, England, GBR
Senior level
Senior level
Blockchain • Fintech • Payments • Consulting • Cryptocurrency • Cybersecurity • Quantum Computing
Supports the design, architecture, implementation, security, troubleshooting, and 24/7 support of a global multi-site network. Manages network projects, device upgrades, requirements, change processes, vendor coordination, risk identification, performance improvements, hardware baseline checks, and compliance activities. Collaborates across teams, mentors peers, and provides technical assistance for LAN/WAN, firewalls, load balancing, and network communications.
Top Skills: BgpCcnpCiscoDss/Pci ComplianceFirewallsLan/WanLoad BalancingMplsSecurity And Encryption TechnologiesSnmpSshTcp/IpVpn

What you need to know about the London Tech Scene

London isn't just a hub for established businesses; it's also a nursery for innovation. Boasting one of the most recognized fintech ecosystems in Europe, attracting billions in investments each year, London's success has made it a go-to destination for startups looking to make their mark. Top U.K. companies like Hoptin, Moneybox and Marshmallow have already made the city their base — yet fintech is just the beginning. From healthtech to renewable energy to cybersecurity and beyond, the city's startups are breaking new ground across a range of industries.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account