PhysicsX Logo

PhysicsX

Senior Software Engineer, Infrastructure - Python & Kubernetes

Posted Yesterday
Be an Early Applicant
In-Office
London, Greater London, England, GBR
Senior level
In-Office
London, Greater London, England, GBR
Senior level
Design, provision, and operate reliable shared infrastructure across AWS, GCP, Azure, and on-premises environments. Manage Kubernetes clusters, multi-tenancy, secrets, GPU workloads, networking, service meshes, and infrastructure as code using Terraform and Crossplane. Develop operators, establish reliability targets, respond to incidents, and partner with security and platform teams on governance and compliance.
The summary above was generated by AI
About us
Re-architecting Engineering for the Age of Intelligence

PhysicsX is the physics AI company for industrials. The company’s mission is to accelerate hardware innovation by overhauling what industrial engineering and manufacturing look like today. PhysicsX is building a new simulation software stack to deliver deep physics AI enablement across the entire engineering lifecycle. The company partners with leading organisations in aerospace & defence, automotive, semiconductors, materials, and energy & renewables, supporting them on some of their most critical and complex challenges. PhysicsX is headquartered in the United Kingdom, with offices in London, New York, and Singapore and an expanding presence in the Bay Area.

Senior Software Engineer – SRE Core Infra London (Hybrid) | Engineering | Full Time

The Role

PhysicsX is growing rapidly and so is the infrastructure that underpins our platform. We are building core infrastructure that is reliable, secure, scalable and reproducible across multiple cloud providers and on premises environments. As the platform evolves to serve increasingly complex engineering workloads, the foundational infrastructure layer becomes ever more critical.

We are looking for a Senior Software Engineer to join our Platform SRE Core Infrastructure team. This role is responsible for the design, provisioning and operation of the shared infrastructure that the entire PhysicsX platform depends on. You will work across infrastructure as code, Kubernetes cluster management, secrets management, GPU drivers, networking and multi tenancy architecture to ensure the platform is dependable, scalable and secure.

This is a role for an engineer who combines deep infrastructure expertise with a reliability engineering mindset and an appreciation for the developer experience of the teams that consume the platform.

What You Will Do

  • Own the design and delivery of core infrastructure across multi cloud providers (AWS, GCP, Azure) and on premises environments using Terraform and Crossplane.
  • Architect and operate Kubernetes clusters supporting both single tenant and multi tenant workloads, with a strong emphasis on isolation, performance and reliability.
  • Define and implement infrastructure provisioning patterns using Crossplane compositions and Terraform modules, ensuring reproducibility and auditability across environments.
  • Design and operate secrets management solutions, including dynamic secret provisioning, rotation and fine grained access control integrated with cluster identity.
  • Manage and maintain GPU driver configurations and accelerated compute node pools, ensuring compatibility and performance for AI and simulation workloads.
  • Own cluster networking design including CNI selection, Istio service mesh integration, ingress strategy and cross cluster connectivity.
  • Implement and maintain vCluster based multi tenancy to provide strong workload isolation within shared infrastructure.
  • Develop lightweight Kubernetes Operators or controllers where automation of infrastructure lifecycle tasks requires it.
  • Establish SLOs and reliability targets for core infrastructure components and lead the response to production incidents.
  • Partner with security and platform teams to enforce infrastructure governance, network policies and compliance controls.
  • Contribute to and uphold engineering standards across the platform organisation.

What You Bring to the Table

  • Kubernetes depth – 5 or more years of professional experience operating Kubernetes in production. You have a thorough understanding of cluster architecture, the scheduler, networking, storage and the API lifecycle. Kubernetes certifications such as CKAD, CKA or CKS are highly desirable.
  • Crossplane expertise – significant hands on experience designing and operating Crossplane compositions, providers and managed resources in production environments.
  • Terraform proficiency – strong experience authoring, structuring and operating Terraform at scale, including state management, module design and CI integration.
  • Multi cloud and on premises – practical experience operating infrastructure across more than one cloud provider and on premises environments, with an understanding of the differences in identity, networking and storage.
  • Multi tenancy architecture – experience designing and implementing both single tenant and multi tenant Kubernetes architectures, with strong views on isolation, resource governance and operational overhead.
  • Secrets management – experience with tools such as Vault, External Secrets Operator or cloud native secret stores, including dynamic provisioning and rotation.
  • Networking – solid knowledge of Kubernetes networking, CNI plugins, Istio service mesh and ingress patterns. Experience with cross cluster or hybrid connectivity is valuable.
  • vCluster and virtual clusters – experience using vCluster or similar tooling to provide lightweight, isolated Kubernetes environments within shared clusters.
  • GPU and accelerated compute – familiarity with GPU driver management, device plugins and the operational considerations of running accelerated workloads in Kubernetes.
  • Kubernetes Operators – at least lightweight experience writing or extending Kubernetes Operators or controllers, ideally in Golang or Python.
  • Software engineering capability – you are comfortable writing code to automate and extend infrastructure. Python and Golang are the primary languages used across the platform. Exposure to functional programming languages such as Erlang, Elixir or OCaml is appreciated.
  • Platform engineering mindset – you think about infrastructure as a product consumed by engineering teams and prioritise usability, documentation and long term maintainability.
  • Distributed systems experience – a solid grounding in distributed systems concepts, including failure modes, consistency and the operational challenges of running systems at scale.

Ideally

  • Experience with GitOps workflows using tools such as Argo CD or Flux.
  • Contributions to open source infrastructure or Kubernetes ecosystem projects.

What we offer

Build what actually matters

Help shape an AI-native engineering company at a formative stage, tackling problems that genuinely matter for industry and society. This is work with real-world impact - and something you can be proud to stand behind.

Learn alongside exceptional people

Work with a high-caliber, collaborative team of engineers, scientists, and operators who care deeply about doing great work, and about helping each other get better. We come from diverse backgrounds, but we share a commitment to operating at the highest level and addressing some of the most complex challenges out there. If you’re ambitious, thoughtful, and driven by impact, you’ll feel at home.

Influence over hierarchy

We operate with a flat structure: good ideas win - wherever they come from. Questioning assumptions and challenging the status quo isn’t just welcomed, it’s expected.

Sustainable pace, long-term ambition

Building meaningful technology is a marathon, not a sprint. We believe in balancing focused, ambitious work with a life beyond it. Our hybrid model blends time together in our Shoreditch office with work-from-home days, giving you the flexibility to work sustainably while staying connected in person.

And it doesn’t stop there …

🚀 Equity options - share meaningfully in the company you’re helping to build.

🏦 10% employer pension contribution - because investing in future matters.

🍽️ Free office lunches - to keep you energised and focused.

👶 Enhanced parental leave - 3 months full pay paternity and 6 months full pay maternity leave, to provide extra flexibility during the moments that matter most.

🍼 YellowNest nursery scheme - to help working parents manage childcare costs.

☀️ 25 days of Annual Leave (+ Public Holidays) - because taking time to rest matters.

🏥 Private medical insurance - 100% employee cover, giving you complete peace of mind.

💪 Wellhub Subscription - gain access to thousands of gyms, classes and wellness apps, supporting both physical and mental wellbeing.

👀 Eye tests - because good work depends on good health.

📈 Personal development - dedicated support for learning, development, and leveling up over time.

💛 Employee Assistance Programme (EAP) - confidential wellbeing support, available whenever you need it.

🚲 Bike2Work scheme and 🚆 Season ticket loan - to make getting to work easier and greener.

🚗 Octopus EV salary sacrifice - for a simpler, more sustainable way to drive electric.

🔎 Watch this space, we’re continuing to build this as we grow…

We value diversity and are committed to equal employment opportunity regardless of sex, race, religion, ethnicity, nationality, disability, age, sexual orientation or gender identity. We strongly encourage individuals from groups traditionally underrepresented in tech to apply.


 
We value diversity and are committed to equal employment opportunity regardless of sex, race, religion, ethnicity, nationality, disability, age, sexual orientation or gender identity. We strongly encourage individuals from groups traditionally underrepresented in tech to apply. To help make a change, we sponsor bright women from disadvantaged backgrounds through their university degrees in science and mathematics. 
 
We collect diversity and inclusion data solely for the purpose of monitoring the effectiveness of our equal opportunities policies and ensuring compliance with employment and equality legislation. This information is confidential, used only in aggregate form, and will not influence the outcome of your application. 
 
HQ

PhysicsX London, England Office

Victoria House 1 Leonard Circus, London, United Kingdom, EC2A 4DQ

Similar Jobs

47 Minutes Ago
In-Office or Remote
London, England, GBR
Senior level
Senior level
Artificial Intelligence • Fintech • Payments • Business Intelligence • Financial Services • Generative AI
Leads APAC merchant risk operations for payments, owning team performance, service levels, quality, fraud and chargeback metrics, investigations, monitoring, and escalation workflows. Partners with Legal, Risk, Compliance, Product, and Engineering to implement controls and improve risk tooling. Drives automation and AI adoption, manages senior staff across locations, influences stakeholders, and ensures portfolio risk aligns with enterprise risk appetite.
Top Skills: AIAi/MlAnalytics DashboardsCase Management SystemsData ToolsLlmsRules Engines
5 Hours Ago
Hybrid
Entry level
Entry level
AdTech • Cloud • Digital Media • Information Technology • News + Entertainment • App development
Supports end-to-end delivery of campus-wide fiber-optic and structured cabling infrastructure, including design coordination, procurement, installation, testing, commissioning, handover, scheduling, quality oversight, and vendor management. Coordinates with engineering, IT, facilities, construction, security, and operations teams. Oversees route diversity, redundancy, documentation, standards compliance, contractor performance, site readiness, and project reporting for large-scale infrastructure workstreams.
Top Skills: AutocadBluebeamFiber OpticsExcelMicrosoft VisioStructured Cabling
5 Hours Ago
Hybrid
Entry level
Entry level
AdTech • Cloud • Digital Media • Information Technology • News + Entertainment • App development
Manages the planning, design, budgeting, scheduling, construction, fabrication, and commissioning of major resort attractions and facilities. Oversees architectural and engineering consultants, construction teams, project budgets, schedules, quality, safety, documentation, change management, and project reporting. Leads project staffing and development while ensuring projects meet creative, operational, financial, and guest experience objectives.

What you need to know about the London Tech Scene

London isn't just a hub for established businesses; it's also a nursery for innovation. Boasting one of the most recognized fintech ecosystems in Europe, attracting billions in investments each year, London's success has made it a go-to destination for startups looking to make their mark. Top U.K. companies like Hoptin, Moneybox and Marshmallow have already made the city their base — yet fintech is just the beginning. From healthtech to renewable energy to cybersecurity and beyond, the city's startups are breaking new ground across a range of industries.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account