Duku AI Logo

Duku AI

Applied Research Engineer

Posted 8 Hours Ago
Be an Early Applicant
In-Office
London, Greater London, England, GBR
Senior level
In-Office
London, Greater London, England, GBR
Senior level
Develop and deploy reinforcement learning systems that autonomously test modern web applications. Build agents capable of perceiving applications through vision and structure, exploring edge cases, learning from sparse rewards, scaling across environments, and transferring knowledge between applications. The role requires shipping production ML systems, advancing RL methods, and optimizing agents for large-scale GPU infrastructure.
The summary above was generated by AI

Change Software Forever


Most research roles publish papers. This one ships the intelligence that makes AI-generated software trustworthy in production. AI is making engineers dramatically more productive. The problem is nobody fully trusts what AI ships.


Why this matters now


AI is rewriting how software gets built. But shipping still breaks on testing. Engineering leaders are desperate to release faster without dropping quality, and they can't get there with manual QA or brittle test automation.


Here's the bigger shift: as AI writes more of the code, the bottleneck moves from generating software to trusting it. Nobody owns that trust layer yet. Duku is building it, autonomous agents that simulate real users, catch critical failures before production, and self-heal as the product evolves.


Backed by experienced operators and investors, we're building what we believe will become a core layer of the modern software stack. We're looking for an exceptional Applied Research Engineer to help invent it. Not someone to apply existing methods. Someone to discover new ones.

Why This Role is Different

Most “AI engineer” jobs are just applying models someone else built. This isn’t that. This is about pushing RL to its edge:


  • Agents that think: networks that see and understand apps through vision, structure, and behavior.
  • Agents that explore: curiosity-driven RL that uncovers edge cases no human would think of.
  • Agents that learn: smarter with every bug, sharper with every correction.
  • Agents that scale: millions of states, thousands of sessions, decisions in sub-seconds.


If you’ve ever wanted to take RL out of papers and into the wild, this is it.


What You’ll Achieve

In your first three months, you’ll see your reinforcement learning prototypes running live inside real applications, surfacing bugs no human ever noticed.

By six months, those agents will have evolved , scaling across multiple environments, learning and adapting in ways that prove this isn’t theory but reality.

And within a year, the intelligence you’ve built will sit at the heart of every release for our first customers, powering their ability to ship AI-generated code with confidence.


What You Bring ( Non-Negotiables)

  • 5+ years shipping ML to production (real systems, not papers).
  • Deep RL expertise , you think in Q-values and policy gradients.
  • Experience building autonomous agents that actually work at scale.
  • Python/PyTorch mastery.


The Stuff That Matters

  • You’re obsessed with solving “impossible” problems.
  • You’d rather ship and learn than debate in theory.
  • You can explain RL to a CEO and optimize it for a GPU cluster.
  • You thrive in chaos and see it as opportunity.


Why Join Now

  • Impact: You won’t be “joining a team.” You’ll be the team that defines how software is built in the age of AI. Your code won’t sit in a corner , it will become the backbone of a new category.
  • Market: Software testing hasn’t changed in 30 years. AI-generated code has rewritten the rules overnight. Whoever solves this bottleneck doesn’t just win a market , they reshape the entire industry.
  • Team: Small, elite, no passengers. You’ll be working side by side with a CTO who built this at Meta and a founding team that’s scaled some of the fastest-growing tech companies on the planet.
  • Timing: Rarely do technology shifts and career timing line up. This is one of those moments. Five years from now, autonomous QA will be a given. Right now, it’s unsolved , and you could be the one who solves it.


The Challenge

Big tech tried to brute-force this problem and hit a wall. Most startups never got past brittle scripts. The reason is simple: building true autonomy takes more than patching frameworks , it takes intelligence. That’s the path we’re on. Your system will need to:


  • Navigate the chaos of modern web apps.
  • Learn from sparse, delayed rewards.
  • Balance exploration with validation.
  • Transfer knowledge across completely different applications.


It won’t be easy. That’s the point.


What You Get

  • Equity that actually moves the needle , not token options, but a real ownership stake in what could be the category-defining AI company of the decade.
  • Unlimited firepower , the hardware, compute, and resources you need to push RL further than anyone has before.
  • A seat at the table, not a cog in the machine, you’ll be in the room where every decision is made, shaping both the product and the company.
  • Speed over politics , a London base where execution beats process, every time.
  • A shot at legacy , work that will outlive your CV, the kind of achievement you’ll still be talking about 20 years from now.


To win the space, we’re looking for the best people in London, with 10/10 ambition and work ethic to join us and build a product people love.

Similar Jobs

Yesterday
In-Office
London, Greater London, England, GBR
Internship
Internship
Artificial Intelligence • Healthtech • Software • Automation
Build and evaluate machine learning models using healthcare administrative data, curate training data, and develop tools for data generation, training, evaluation, and deployment. Collaborate with research and engineering teams to develop agent-based solutions and deploy work into live NHS environments. The role requires hands-on research ownership, rigorous evaluation, clean code, and a strong focus on safety and real-world patient impact.
Top Skills: PythonPyTorchScikit-Learn
7 Days Ago
In-Office or Remote
London, Greater London, England, GBR
Senior level
Senior level
Artificial Intelligence
Lead applied research and engineering to build production-grade video foundation models. Work includes developing latent video diffusion models, designing conditioning for controllability, scaling distributed multi-GPU training, improving training stability, building evaluation frameworks, and optimizing inference for low-latency, high-resolution deployment.
Top Skills: AWSCi/CdCudaDdpDeepspeedDiffusion ModelsDockerFsdpGansGitLatent Video DiffusionMulti-Gpu Multi-Node TrainingPythonPyTorchSequence ParallelismSlurmVaes
2 Days Ago
Remote or Hybrid
London, Greater London, England, GBR
Senior level
Senior level
Artificial Intelligence • Software
Work across modeling, data, systems, and evaluation to make video foundation models expressive, controllable, and personalized. Build adaptation and personalization pipelines, define end-user quality metrics and evaluations, and collaborate with product and design to deploy production-ready model variants for creative partners.
Top Skills: Control AdaptersDiffusion ModelsKubernetesModel DistillationPythonPyTorchRayReinforcement LearningSftSlurmTransformers

What you need to know about the London Tech Scene

London isn't just a hub for established businesses; it's also a nursery for innovation. Boasting one of the most recognized fintech ecosystems in Europe, attracting billions in investments each year, London's success has made it a go-to destination for startups looking to make their mark. Top U.K. companies like Hoptin, Moneybox and Marshmallow have already made the city their base — yet fintech is just the beginning. From healthtech to renewable energy to cybersecurity and beyond, the city's startups are breaking new ground across a range of industries.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account