Kitman Labs Logo

Kitman Labs

Data Engineer

Reposted 9 Days Ago
Remote
Hiring Remotely in Greater Manchester, England, GBR
Mid level
Remote
Hiring Remotely in Greater Manchester, England, GBR
Mid level
Build and maintain BigQuery data models and Python data pipelines using Dataform and Airflow. Implement data quality checks, observability, and CI/CD; containerize workloads with Docker; support Data Science productionization and contribute to Looker reporting.
The summary above was generated by AI

Kitman Labs is the performance intelligence company, disrupting and transforming the way the sports industry uses data to unlock the potential of the world's top athletes.

Driven by a passion to innovate in the areas of sports performance, analytics and user experience, we have assembled a team of the industry's top data scientists, sports performance scientists, product specialists and engineers.

Kitman Labs' advanced Intelligence Platform (iP) is now used by over 2000 teams in 50 leagues on 6 continents, including the NFL, Premier League, National Women's Soccer League and MLS.


Data Engineer

We're looking for a Mid-Level Data Engineer to join our team and help build and evolve our data platform. You'll work across analytics engineering, data pipelines, and data quality — collaborating closely with Engineers, Data Scientists, and Product to turn raw data into reliable, scalable foundations.

What You'll Work On

    Analytics Engineering & Reporting

    • Build and maintain BigQuery data models using Dataform, following medallion architecture patterns (Bronze/Silver/Gold)

    • Contribute to Looker dashboards and LookML models, working alongside senior engineers and analysts

    • Write performant, well-structured SQL for large-scale transformations in BigQuery

    • Implement data quality checks using Dataform assertions and automated alerting

    • Support data observability across the warehouse — monitoring pipeline health, data freshness, and anomaly detection

    • Data Pipelines & Ingestion

      • Build and maintain robust Python data pipelines with testing, linting, and CI/CD integration

      • Work with orchestration tooling (Cloud Composer / Airflow) to schedule and monitor workflows

      • Develop familiarity with CDC concepts and event-driven ingestion patterns (Datastream, Pub/Sub)

      • Containerise workloads with Docker for deployment on Cloud Run or similar GCP services

      • Data Science Collaboration

        • Support Data Scientists in moving work from notebook to production pipeline

        • Contribute to feature pipelines and data preparation for ML workloads

        • Help bridge the gap between research prototypes and scalable, maintainable code

What We're Looking For

    • SQL proficiency — comfortable writing complex, performant queries against large datasets in BigQuery

    • Dataform experience — or strong dbt experience with willingness to work in Dataform; understanding of modular, version-controlled data transformation

    • Python with an engineering mindset — clean, tested, linted code; comfortable with Git and CI/CD workflows

    • GCP familiarity — hands-on experience with BigQuery is essential; broader GCP exposure (Cloud Storage, Cloud Run, Pub/Sub, Datastream) is a strong advantage

    • Orchestration experience — hands-on with Cloud Composer, Airflow, or a comparable tool

    • Data modelling fundamentals — dimensional modelling, Kimball principles, or medallion architecture patterns

    • Docker basics — able to containerise and deploy data workloads

    • Collaborative and communicative — able to translate business requirements into data models and work effectively with Analytics, Product, and Data Science stakeholders

    • Pragmatic approach to AI tooling — comfortable using AI-assisted development to improve productivity and code quality

Nice to have

    • Looker / LookML experience

    • Familiarity with CDC concepts and tools (Datastream, Debezium)

    • Exposure to ML frameworks or MLOps tooling (scikit-learn, MLflow, Vertex AI)

    • AWS experience as a complement (Redshift, Glue, RDS) — we value engineers who can draw on cross-cloud perspective

    • Curiosity about sports performance data

Why this role?

You'll work on a modern GCP-native stack — BigQuery, Dataform, Looker, Cloud Composer, Cloud Run — with room to grow into platform ownership, CDC pipelines, and MLOps as your experience develops. We care about clean code, good data, and pragmatic engineering over perfection.

Benefits
 
At Kitman Labs we pride ourselves on being the best and working with the best, so it should be no surprise that we are also dedicated to keeping the best through building a world-class work culture.
We truly believe that a successful company begins through having an outstanding and inspiring culture, so our benefits reflect this:
- Competitive salary
- Health insurance for employee & dependants
- Meaningful equity
- Pension Plan
- Life Cover
- Income protection
- Wellbeing benefits
 
Location
 
While this role allows for remote work, occasional face-to-face-gatherings are recommended.
 
Diversity
 
In addition to building a team with diverse skill-sets, Kitman Labs is committed to hiring people with diverse backgrounds. We do not discriminate based on age, civil or family status, disability, ethnicity, gender, race, religion, or sexual orientation. If you are a person with a disability and require assistance during the application process, please let us know.
 
You can find information about how we process, share and keep your personal data safe by reading our privacy policy

Similar Jobs

4 Days Ago
Remote
UK
Senior level
Senior level
HR Tech • Other • Professional Services
Build and operate a new AI-focused data platform: design ingestion, transformation, semantic/ontology layers, storage, and DAG-based orchestration. Implement pipelines, select databases and orchestration tools, ensure performance, observability, and data quality, and collaborate with product and AI teams while remaining hands-on in coding and production operations.
Top Skills: AirflowAnalytical WarehousesBi ToolsClickhouseDagsterDbtDuckdbObject StoragePythonRelational DatabasesSalesforceSAP
3 Hours Ago
Remote
UK
Mid level
Mid level
Cloud • Information Technology
Design, build and maintain cloud-native ETL/ELT pipelines and data lake solutions using dbt, Snowflake/Databricks and AWS services. Optimize SQL and data models for performance and cost, create dashboards, follow CI/CD and software engineering best practices, and collaborate with architects, engineers and clients to deliver production-ready enterprise data solutions.
Top Skills: Amazon QuicksightAmazon S3AWSAws GlueAws LambdaAws Step FunctionsCi/CdDatabricksDatabricks DashboardsDbtGitMwaa (Apache Airflow)PythonSnowflakeSnowsightSQL
11 Days Ago
Remote or Hybrid
England, GBR
Mid level
Mid level
Financial Services
Design, build, maintain and evolve cloud data platform and ELT pipelines. Collaborate with stakeholders to provision, model and govern data, automate CI/CD, create data quality tests, and recommend new technologies.
Top Skills: Aws CdkBigQueryBitbucketCopilotDbtDockerFivetranHvrJenkinsKimball MethodologyMongoDBMs SqlPythonRedshiftSnowflakeSnowflake CortexVisual Studio Code

What you need to know about the London Tech Scene

London isn't just a hub for established businesses; it's also a nursery for innovation. Boasting one of the most recognized fintech ecosystems in Europe, attracting billions in investments each year, London's success has made it a go-to destination for startups looking to make their mark. Top U.K. companies like Hoptin, Moneybox and Marshmallow have already made the city their base — yet fintech is just the beginning. From healthtech to renewable energy to cybersecurity and beyond, the city's startups are breaking new ground across a range of industries.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account