MUFG Investor Services Logo

MUFG Investor Services

Senior MLOps Engineer

Posted 4 Days Ago
Be an Early Applicant
Hybrid
London, England, GBR
Senior level
Hybrid
London, England, GBR
Senior level
Build and operate secure, scalable infrastructure for AI agents and machine learning workloads. Responsibilities include deploying AI agents, managing AWS cloud resources with Terraform and Boto3, implementing OpenTelemetry and Datadog observability, optimizing performance and costs, supporting CI/CD automation, conducting load testing and vulnerability assessments, troubleshooting platform issues, and collaborating with AI, data science, and engineering teams.
The summary above was generated by AI
Company Description

MUFG Investor Services is a trusted partner to many of the world’s largest public and private funds, providing asset servicing and operational solutions built for alternatives. With over $1 trillion in client assets under administration, we offer fund administration, banking, payments, fund financing, foreign exchange overlay, corporate and regulatory services, custody, business consulting, and more. Operating from 17 locations worldwide, we help clients mitigate risk, enhance efficiency, and navigate the operational complexities of today’s investment management landscape. As a division of Mitsubishi UFJ Financial Group (MUFG), one of the world’s largest financial institutions with approximately $3 trillion in assets, we combine deep expertise with the strength and stability of a leading financial institution. To learn more, visit us at www.mufg-investorservices.com.

#LI-Hybrid

Job Description

We are seeking a highly skilled MLOps / Platform Engineer with a strong background in DevOps workflows and platform engineering best practices to join our AI initiative. This is a high-visibility project focused on deploying and managing AI agents across our infrastructure. You will work closely with the Research & Data Science team, backend and frontend engineers, and other technical teams to build a secure, scalable, and cost-optimized platform for AI workloads.

This position supports AI Engineering and Data Science initiatives by focusing on infrastructure, operations, and platform reliability. The Platform Engineer will work closely with AI Engineers and Data Scientists to ensure they have robust, scalable infrastructure to deploy their work.

You Will:

  • Design, deploy, and maintain AI agents on Agent Core MCP servers and MCP gateways.
  • Implement and manage observability using OpenTelemetry for logs and traces, integrating with Datadog.
  • Ensure security, high availability, and cost optimization across all AI platform components.
  • Provide infrastructure and deployment support to AI researchers and engineering teams, enabling integration of cutting-edge technologies into production.
  • Perform load testing, token cost measurement, and optimize resource utilization.
  • Facilitate external vulnerability assessments and ensure compliance with security best practices.
  • Troubleshoot and resolve platform issues promptly to maintain operational stability.
  • Contribute to DevOps workflows, CI/CD pipelines, and automation for AI deployments.
  • Support evaluation of third-party products related to hosting AI agents or enhancing project capabilities.
  • Assist in external audits and maintain documentation for platform architecture and processes.
  • Develop and execute automation scripts using the AWS Boto3 SDK to deploy, test, and validate AI platform components across multiple environments.
  • Implement Infrastructure as Code (IaC) using Terraform to provision and manage cloud resources for AI workloads, ensuring consistency and scalability.

#LI-Hybrid

Qualifications

You Have:

  • 5+ Years of experience in Platform Engineering / DevOps practice
  • Deep understanding of DevOps principles, workflows, and best practices.
  • Proven experience in platform engineering and full-stack development.
  • Proficiency in API design and integration.
  • Hands-on experience with AWS services
  • Familiarity with OpenTelemetry, Datadog, and observability tooling.
  • Solid coding skills in languages commonly used for backend and automation (e.g., Python, Node.js, Go).
  • Knowledge of microservices, container orchestration (Kubernetes/EKS), and cloud-native architectures.
  • Extensive knowledge of security practices, cost optimization, and performance testing.
  • Interest and familiarity with latest trends in MCPs (Model Context Protocol) and AI agent frameworks.

Preferred Experience

  • Working with AI/ML platforms or deploying AI agents in production environments.
  • Exposure to high-scale distributed systems and cloud infrastructure.
  • Experience in observability and monitoring for complex systems.
  • AWS certifications

Project Details

  • High visibility within the organization.
  • Opportunity to work with cutting-edge AI technologies and collaborate with leading experts.

 

Additional Information

What’s in it for you to join MUFG Investor Services? 

Take a look at our careers site and you’ll find everything you’d expect working with one of the fastest-growing businesses at one of the world’s largest financial groups. Now take another look. Because it’s how we defy expectations that really defines us. You’ll feel that difference in all kinds of ways.  Our vibrant CULTURE. Connected team. Love of innovation, laser client focus. We also talk the talk when it comes to HYBRID WORKING.

So, why settle for the ordinary?  Apply now for your next Brilliantly Different opportunity. 

We thank all candidates for applying; however, only those proceeding to the interview stage will be contacted.

MUFG is an equal opportunity employer.

Similar Jobs

5 Days Ago
In-Office
London, Greater London, England, GBR
Senior level
Senior level
Insurance • Software
Develop and evolve Ki’s end-to-end MLOps platform, including feature stores, model registries, monitoring, governance, and lifecycle management. Enable safe production deployment of machine learning, actuarial, and rules-based models. Own roadmap development, cost and vendor decisions, regulatory alignment, stakeholder adoption, and knowledge sharing. Coach early-career engineers and drive operational improvements across the digital underwriting capability.
Top Skills: Feature StoresInference GraphsLarge Language ModelsMachine LearningMlopsModel MonitoringModel RegistriesModel WorkflowsTerraform
One Month Ago
In-Office or Remote
London, Greater London, England, GBR
Senior level
Senior level
Hardware • Information Technology • Software • Sports • Wearables
Design, build, and maintain scalable edge ML infrastructure and model compilation pipelines for embedded devices. Deploy and validate optimized inference engines (TensorRT, quantization), run CI/CD and automated testing on fleets, monitor telemetry and performance, solve edge constraints (bandwidth, storage, reliability), and mentor the team on Python tooling and IaC.
Top Skills: Aws Iot GreengrassBalenaCi/CdDeepstream SdkDockerFfmpegGstreamerInfrastructure-As-CodeJetson OrinLinuxPythonTensorrt
29 Days Ago
In-Office or Remote
United Kingdom
Senior level
Senior level
Hardware • Healthtech • Machine Learning • Software
Design and implement ML infrastructure for scalable deployment, collaborate across teams, optimize ML systems, and enhance tooling.
Top Skills: AWSCloudwatchDynamoDBEcsFlinkKafkaKinesisKubeflowLambdaMlflowPythonPyTorchSagemakerTensorFlowVertex Ai

What you need to know about the London Tech Scene

London isn't just a hub for established businesses; it's also a nursery for innovation. Boasting one of the most recognized fintech ecosystems in Europe, attracting billions in investments each year, London's success has made it a go-to destination for startups looking to make their mark. Top U.K. companies like Hoptin, Moneybox and Marshmallow have already made the city their base — yet fintech is just the beginning. From healthtech to renewable energy to cybersecurity and beyond, the city's startups are breaking new ground across a range of industries.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account