PhD Studentship: Causal Reinforcement Learning

23 hours, 18 minutes ago
Internship
Entry Level
Artificial Intelligence and Machine Learning
Phaidra

Phaidra

Phaidra is an industrial AI company that creates self-learning, intelligent control systems for industrial facilities. By leveraging AI technology, Phaidra helps operators reduce risk, improve energy efficiency, and meet sustainability goals by maximiz...

Internet Software & Services
51-250
Founded 2019
$30M raised

Description

  • Develop theoretical foundations for policy learning from biased, small datasets through a causal lens.
  • Characterize how confounding and mediators affect offline reinforcement learning.
  • Build RL algorithms that use known or learned causal structure to improve out-of-distribution generalization and provide policy guarantees.
  • Design and evaluate methods in controlled simulated environments with known causal structure.
  • Benchmark proposed approaches against standard and offline RL baselines.
  • Contribute to foundational RL research while staying grounded in real-world industrial challenges.
  • Work under the supervision of academic and industrial co-supervisors at Cambridge and Phaidra.

Requirements

  • A first-class or upper second-class honours degree (or equivalent) in Computer Science, Mathematics, Engineering, Statistics, or a related technical field.
  • Strong background in at least one of: reinforcement learning, machine learning, probabilistic modelling, or control theory.
  • Proficiency in Python and standard ML libraries including PyTorch, NumPy, SciPy, and scikit-learn.
  • Clear scientific writing skills and the ability to communicate research to academic and applied audiences.
  • Eligibility to study at the University of Cambridge; international students are welcome and English language requirements apply.
  • Familiarity with causal inference, causal graphical models, or structural equation models (preferred).
  • Prior research experience such as an undergraduate thesis, MSc dissertation, research internship, or publications (preferred).
  • Experience with offline RL, batch RL, or safe RL (preferred).
  • Exposure to applying ML to real-world physical or industrial systems (preferred).

Benefits

  • Fully funded 4-year PhD studentship.
  • Expected start date in January 2027.
  • 100% remote company with no physical office.
  • Competitive compensation with meaningful equity.
  • Medical, dental, and vision insurance (varies by region).
  • Unlimited paid time off with a required minimum of 20 days per year.
  • Paid parental leave (varies by region).
  • Flexible stipends for workspace, well-being, and professional development.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

Research Contractor

FUTURESEARCH Internet Software & Services

FutureSearch is hiring a research contractor to help develop tasks for training and benchmarking frontier AI models in a highly research-driven environment.

1 day, 23 hours ago

RL Environment Software Engineer

Talentpluto, Inc. 1-10 Recruiting

A remote RL Environment Software Engineer at an applied AI research lab building reinforcement-learning environments and agents for leading AI labs, with the goal of creating high-quality systems the frontier will need next.

Docker Kubernetes Machine Learning Node.js Python React Reinforcement Learning TypeScript
1 day, 23 hours ago

Research Scientist

FUTURESEARCH Internet Software & Services

FutureSearch is hiring Research Scientists to evaluate and improve forecasting and agentic LLM web research as part of its frontier AI research team.

LLM Machine Learning
1 day, 23 hours ago

General Application

FUTURESEARCH Internet Software & Services

FutureSearch is hiring a remote AI-focused team member to help build, evaluate, and improve systems that predict the future and support decision-making as transformative AI advances.

LLM
1 day, 23 hours ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers