DV Trading Logo

DV Trading

Senior DevOps Engineer

Reposted 15 Days Ago
Be an Early Applicant
In-Office
London, Greater London, England, GBR
Senior level
In-Office
London, Greater London, England, GBR
Senior level
Own and operate an on-prem Kubernetes platform for trading systems: cluster lifecycle, observability (metrics/logs/traces), GitOps CI/CD, infra automation, and onboarding/support for trading desks.
The summary above was generated by AI

About Us:

Founded 20 years ago and headquartered in Chicago, the DV Group of financial services firms has grown to more than 600 people operating throughout North America, Europe and Asia. Since spinning out of a large brokerage firm in 2016, DV Trading has rapidly scaled as an independent proprietary trading firm utilizing its own capital, trading strategies, and risk management methodologies to provide liquidity to worldwide financial markets and hedging opportunities to commodity producers and users. Now, DV group affiliates include two broker dealers, a cryptocurrency market making firm, and a bourgeoning investment adviser.

Overview:

We're looking for a senior level DevOps Engineer to join a small, high-impact team that builds and operates the infrastructure powering DV's trading systems. You'll work on Kubernetes at scale, a firm-wide observability platform, CI/CD infrastructure, and workflow orchestration.

This team built the entire platform from scratch in 2025 and now owns it end-to-end. You'll be joining at a pivotal moment: the team is growing, the platform is scaling to meet increasing demand across the firm, and there's no shortage of hard problems to solve. You'll have real ownership from day one — not tickets in a queue.

Job Responsibilities:

  • Kubernetes platform operations and evolution: Cluster lifecycle management via Cluster API, fleet-wide upgrades, bare metal provisioning, CNI networking, storage, autoscaling, and disaster recovery planning.

  • Observability: Operate and scale our on-prem observability stack — Mimir (metrics), Loki (logs), Tempo (traces), Grafana (dashboards), OpenTelemetry Collector fleet, and Alertmanager. Drive adoption by onboarding teams, building dashboards, tuning alerts, and scaling ingestion for growing workloads.

  • CI/CD and GitOps: GitLab pipelines, ArgoCD for Kubernetes deployments, Artifactory for artifact management. Build reusable CI components, improve secrets integration, and support trading desks onboarding to the platform.

  • Infrastructure automation and tooling: Build and maintain automation for provisioning, configuration management, and self-service tooling that enables teams across the firm to move faster without depending on DevOps for every change.

  • Platform adoption and support: Work directly with trading desks and development teams to onboard them to the platforms we build. Write runbooks, contribute to documentation, and help build a sustainable support model as the platform scales.

Requirements:

  • 5–8 years of experience in DevOps, SRE, or Platform Engineering roles.

  • Kubernetes: Deep hands-on experience operating K8s in production. Cluster lifecycle, troubleshooting, networking, storage, RBAC. Experience with Cluster API, bare metal provisioning, or multi-cluster management is a strong plus.

  • Observability: Production experience with Prometheus, Grafana, and alerting. Familiarity with Mimir, Loki, Tempo, or Thanos for scaled metrics, logs, and traces. Understanding of OpenTelemetry (collectors, exporters, instrumentation).

  • GitOps and CI/CD: Experience with ArgoCD, Flux, or similar GitOps tools. Building and maintaining CI/CD pipelines (GitLab CI, GitHub Actions, or equivalent). Artifact management (Artifactory, Nexus, or similar).

  • Infrastructure as Code: Terraform and/or Ansible for provisioning and configuration management.

  • Linux systems: Strong fundamentals — systemd, networking, storage, performance troubleshooting.

  • Programming: Go or Python for automation, tooling, and scripting. Comfortable reading and writing YAML and Helm charts.

  • Communication: Ability to work directly with trading desks and development teams who depend on the platforms you build. You'll be explaining K8s concepts to people who aren't K8s experts.

Preferred Skills:

  • Experience working at a trading firm — understanding the urgency and reliability requirements of systems that support P&L-impacting workloads.

  • Experience with Kubeflow, Airflow, or ML pipeline orchestration.

  • Experience with advanced Kubernetes networking (CNI plugins, network policies, service mesh).

  • Experience with Kafka or event streaming platforms.

  • Experience operating on-prem infrastructure (not just cloud) — bare metal servers, IPAM, storage systems.

  • Thoughtful use of AI coding assistants and interest in AI-assisted operational tooling.

Why This Role:

  • Ownership, not tickets. This is a growing team that built and owns the entire platform. You'll have direct impact on infrastructure that trading desks depend on daily.

  • Modern stack, real scale. Kubernetes at scale, a full observability platform, GitOps everywhere — all on-prem in a high-performance trading environment.

  • Greenfield opportunities. Bare metal K8s, advanced observability, automated dependency mapping, AI infrastructure — there's no shortage of interesting projects.

  • Small team, big impact. You won't get lost in a 50-person platform org. Every engineer on this team shapes the direction of the platform.

DV is not accepting unsolicited resumes from search firms. Only search firms with valid, written agreements with DV should submit resumes in response to DV’s posted positions. All resumes submitted by search firms to DV via e-mail, the Internet, personal delivery, facsimile, or any other method without a valid written agreement shall be deemed the sole property of DV, and no fee will be paid in the event the candidate is hired by DV. DV is proud to be an equal opportunity employer and committed to creating an inclusive environment for all employees.

DV Trading London, England Office

London, United Kingdom

Similar Jobs

2 Days Ago
In-Office
Senior level
Senior level
Artificial Intelligence • Robotics • Software • Defense
Own and engineer secure, scalable AWS and Kubernetes platforms across multi-account, multi-region environments. Build infrastructure with Terraform and Helm, automate CI/CD using GitHub Actions and Argo CD, improve reliability through observability and SRE practices, and lead security hardening, compliance audits, disaster recovery, cost optimization, and incident response. Mentor developers and promote DevSecOps practices while supporting high-availability platforms for mission-critical AI systems.
Top Skills: AlbArgo CdAWSAws BackupAws Control TowerBashCloudfrontCloudwatchCompute OptimizerDastDockerDocker ComposeEc2EcrEcsEksElbGdprGithub ActionsGoGravitonGuarddutyHelmIamIam Identity CenterInspector V2Iso 27001KarpenterKubeflowKubernetesLinuxNexusNist SsdfOpenshiftOwasp SammPrometheusPythonS3SagemakerSastSavings PlansSsoTerraformTerraform CloudVpcZenml
16 Days Ago
In-Office
London, Greater London, England, GBR
Senior level
Senior level
Information Technology • Professional Services • Analytics • Consulting
Support and optimize AWS cloud infrastructure, including networking, security, automation, monitoring, and on-premises connectivity. Design and manage VPCs, subnets, load balancers, NACLs, NAT, Direct Connect, Transit Gateways, IDS/IPS, and security groups. Automate infrastructure with Chef and use Git, Jenkins, and Nexus for development workflows. Maintain Red Hat Enterprise Linux environments, promote cloud networking best practices, support AWS migrations, and help teams adopt agile ways of working.
Top Skills: AWSChefCloudcheckrDatadogDirect ConnectElastic IpElastic Load BalancingGitIds/IpsJenkinsNatNetwork AclsNexusRed Hat Enterprise LinuxRhceRhel7Security GroupsSplunkSubnetsTransit GatewayVpc
16 Days Ago
Hybrid
London, Greater London, England, GBR
Senior level
Senior level
Artificial Intelligence • Information Technology • Machine Learning • Professional Services • Software • Analytics • Consulting
Own a platform domain end to end, defining architecture, roadmap, reliability, security, and cost standards. Build scalable multi-tenant Kubernetes infrastructure, enable self-service deployments, advance agent-native engineering practices, and develop tooling for safe agent execution. Partner with engineering teams to translate infrastructure needs into reusable capabilities, mentor engineers, lead design decisions, and manage on-call and post-incident processes.
Top Skills: AWSAzureGCPGitopsHelmKubernetesMcp ServersNetwork PolicyPythonSecrets ManagementSsoTerraformWorkload Identity

What you need to know about the London Tech Scene

London isn't just a hub for established businesses; it's also a nursery for innovation. Boasting one of the most recognized fintech ecosystems in Europe, attracting billions in investments each year, London's success has made it a go-to destination for startups looking to make their mark. Top U.K. companies like Hoptin, Moneybox and Marshmallow have already made the city their base — yet fintech is just the beginning. From healthtech to renewable energy to cybersecurity and beyond, the city's startups are breaking new ground across a range of industries.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account