Build and scale a predictive modeling platform supporting multiple problem types and tenants. Develop embedding-based, temporal survival, and gradient-boosted models for social-domain risk prediction. Create training datasets, evaluation pipelines, and production ML infrastructure, while researching and safely deploying emerging ML and LLM techniques. Implement containerized APIs and Kubernetes services, and incorporate responsible AI, ethical, and regulatory considerations throughout development.
In this role you will work in the Platform team – a function for the deployment and evolution of the backend platform that underpins the core of the Xantura business.
Key Responsibilities
- Own and advance a predictive modelling platform that scales across problem types and tenants, using it to design, implement, and iterate models (embedding-based sequence encoders, temporal survival models, gradient-boosted decision trees) that predict key vulnerabilities in housing, health, and other social domains.
- Track developments in ML and frontier models, running structured experiments to bring promising techniques into production safely.
- Build robust evaluation pipelines, training datasets, and model infrastructure to support continuous improvement of natural language & predictive analytics.
- Ensure responsible AI deployment, embedding ethical and regulatory considerations into every stage of development.
Skills, Knowledge & Expertise
- Bachelor’s or Master’s degree in Computer Science, Machine Learning, or a related technical field – or equivalent practical experience.
- 3+ years of professional experience as an ML Engineer, or related role.
- Strong programming skills and production experience in Python.
- Experience building and maintaining data or ML pipelines with an orchestration tool such as Dagster (or Airflow, Prefect, etc.).
- Hands-on experience with common ML libraries and frameworks, e.g. PyTorch, scikit-learn, and gradient-boosting libraries such as XGBoost or LightGBM.
- Clear evidence of practical experience defining and deploying containerised systems, i.e.:
- Implementing APIs for internal services, e.g. via FastAPI;
- Deploying containerised systems to production, in particular via Kubernetes.
In addition, the following would be an advantage:
- PhD in Computer Science, Machine Learning, or a related field with a strong publication record in text analytics, representation learning, or applied predictive modelling.
- Practical experience productionising LLMs, i.e.:
- Working with vector databases and developing retrieval-augmented generation (RAG) pipelines – experience setting up/configuring vector DBs, as well as using, would be advantageous;
- Finding and productionising recent AI models (e.g. via Huggingface (transformers), OpenAI APIs);
- Building agentic systems (e.g. via LangChain, AutoGen, PydanticAI).
- Evidence of participating in Open-Source Software (OSS) development, public hackathons, or other sharable coding samples.
- Deep expertise in embedding-based architectures, including bi-encoders, cross-encoders, etc. for long-horizon text or temporal prediction tasks.
- Practical experience building and serving production-ready, asynchronous APIs for embedding and/or other compute-intensive services.
- Proficiency in Python for building high-performance data and model pipelines, with strong software engineering discipline (testing, versioning, CI/CD).
- Good familiarity with the Azure ecosystem (Azure Kubernetes Service, Azure Batch, Azure AI Foundry, Azure Machine Learning, Azure Blob Storage, Azure Key Vault) .
This is a Hybrid opportunity with the expectations of being in the office 1 - 2 days a week.
Job Benefits
- Competitive salary reviewed annually
- Work for a passionate, mission-driven company solving society’s big problems
- Work flexible hours around life commitments with a focus on delivering company value rather than hours worked
- Training and development opportunities
- 25 days annual leave (plus bank holidays)
- Company pension
- Private medical insurance
- Generous enhanced parental leave policies
- Cycle to work scheme
- Flu Vaccinations,
- Eye Test and contribution towards Glasses for VDU use
- Employee Assistance Programme
- Mental health and wellbeing support
- Remote GP access
- Counselling/therapy
- Physiotherapy
- Medical second opinions
About
At Xantura, we’re on a mission to reduce societal inequality by helping local authorities use data more effectively. Our AI-driven platform empowers frontline workers with the insights they need to prevent complex issues like homelessness or children being taken into care — before they happen. We make this possible by connecting siloed datasets, applying advanced machine learning to enrich the data, and using predictive analytics to identify those most at risk. Our platform then distills this into clear, actionable insights that help frontline staff intervene early and make a real difference. It’s an exciting time to join Xantura. We’re scaling quickly, bringing on new clients, strengthening our platform, and expanding into new areas. While we’re a technology company at heart, our true focus is on improving lives — and we’re looking for people who share that vision.
Xantura Limited London, England Office
London, United Kingdom
Similar Jobs
Fintech • Mobile • Payments • Software • Financial Services
Lead the architecture and production deployment of deep learning and machine learning systems for real-time financial crime detection. Design sequence-based, graph-based, and attention-based models; establish reusable experimentation-to-production pipelines; evaluate foundation models and embeddings; make architecture decisions under latency and throughput constraints; partner with data scientists on evaluation and measurement; and mentor engineers and data scientists.
Top Skills:
Attention MechanismsBatchingDeep LearningDistributed TrainingEmbeddingsFoundation ModelsGraph Message-PassingGraph Neural NetworksLlm EvaluationMachine LearningMl Pipeline OrchestrationPythonPyTorchQuantizationSequence Modeling
Financial Services
Leads the design, productionization, and operation of LLM-powered agentic commerce applications. Builds retrieval systems, agent memory, organizational context, multi-agent workflows, evaluation frameworks, guardrails, and ML pipelines. Deploys resilient, observable solutions on AWS or Azure using MLOps practices, while partnering with product, business, data science, and engineering stakeholders to deliver secure, auditable production agents.
Top Skills:
A2AAWSAws EksAzureBitbucketConfluenceDatabricksGraph RagJavaScriptKubernetesMcpMlflowNext.JsOpensearchPostgresPythonRedisSagemakerSnowflakeSplunkSQLSvelteTypescriptVector Stores
Artificial Intelligence • Cloud • Machine Learning • Mobile • Software • Virtual Reality • App development
Develop machine learning models and computer vision algorithms for photon-efficient, event-driven imaging in AR glasses. Responsibilities include tracking, depth estimation, SLAM, detection, probabilistic modeling, and novel imaging algorithms for ultra-low-light, high-temporal-resolution systems. The role requires collaboration with global hardware and software teams and emphasizes fast, reliable implementation in C++ and Python.
Top Skills:
3D GeometryAugmented RealityC++Computer VisionEvent-Driven ImagingMachine LearningPhoton-Efficient ImagingPythonPyTorchSlamStatistical Signal ProcessingTensorFlowVisual-Inertial Odometry
What you need to know about the London Tech Scene
London isn't just a hub for established businesses; it's also a nursery for innovation. Boasting one of the most recognized fintech ecosystems in Europe, attracting billions in investments each year, London's success has made it a go-to destination for startups looking to make their mark. Top U.K. companies like Hoptin, Moneybox and Marshmallow have already made the city their base — yet fintech is just the beginning. From healthtech to renewable energy to cybersecurity and beyond, the city's startups are breaking new ground across a range of industries.



