Research Scientist Intern – PhD (AI Agentic Systems)

Research Scientist Intern – PhD (AI Agentic Systems) at Leeroo — London, England, GB

  • Company: Leeroo
  • Location: London, England, GB
  • Employment type: INTERN
  • Salary: USD 6500–10000 / year
  • Posted: 2026-07-17

About this role

Location: Remote (EU, UK, US) / In‑person in London

Duration: 3 months (option to extend to 6)

At Leeroo—a Y Combinator‑backed company founded by former Meta AI and LinkedIn AI researchers and engineers—we’re building continuous‑learning agents driven by our Learning Engine, a modular framework that lets agents learn from human knowledge bases, feedback, and through self‑discovery, then apply that learning—evolving into expert teammates that get better every day rather than demo‑stage agents. We’re seeking Research Scientist Interns to join us in advancing the Learning Engine to build superhuman agents across domains.

About the Role

You’ll partner directly with the founders to design, implement, and evaluate new algorithms that let agents learn continuously. Successful prototypes move quickly into production, so your research will have an immediate impact on real users.

Responsibilities

Design and run state‑of‑the‑art machine‑learning experiments to extend our agentic Learning Engine.

Publish results through papers, datasets, and open‑source code.

Collaborate with product and engineering to translate your research findings into the Learning Engine, delivering value to customers across diverse domains.

Qualifications

PhD in AI, Computer Science, or a related field — or an equivalent record of research excellence

Publications at top AI conferences (Neurips, ICML, ACL, …) or an open‑source project with significant traction

Strong Python skills and experience with ML frameworks such as PyTorch

Exceptional problem‑solving ability; comfortable working independently and in teams

Familiarity with graph databases, vector stores, or Elasticsearch for structured‑data retrieval

Knowledge of advanced training algorithms (e.g., DPO, RLHF, GRPO)

Clear written and verbal communication

What We Offer

Remote‑friendly setup — work from anywhere in compatible time zones, with our office based in London for in‑person collaboration

Competitive, top‑of‑market compensation package

Freedom to pursue ambitious research aligned with our mission

Close collaboration with the founders and senior researchers

Limitless cloud compute — GPU clusters and a war chest of cloud credits ready for your experiments

Conference travel, publication, and open‑source support

Option to extend to 6 months or transition to a full‑time role

If you’re excited about creating AI that learns continuously and autonomously, we’d love to hear from you.

Apply for this position