Principal Research Scientist, Trust & Safety Alignment

Engineer

Principal Research Scientist, Trust & Safety Alignment

Apply Now

- £0.00

  • Date posted
    October 6, 2026
  • Expiration date
    January 6, 2027
  • Application ends
    January 6, 2027

We are seeking a Principal Research Scientist to lead advanced research in AI alignment, trust and safety, and foundation model reliability.

The role focuses on improving the safety, robustness, controllability, and interpretability of large language models, vision-language models, and agentic AI systems. Research areas include continuous and post-training, reinforcement learning, preference optimisation, model editing, machine unlearning, interpretability, controlled learning, and agentic AI safety.

This position is suited to an experienced research leader who combines deep technical expertise with a strong track record of advancing the state of the art and translating research into scalable, production-ready systems.

Key Responsibilities

  • Lead research and development of advanced alignment and safety methods for large language models, vision-language models, and agentic AI systems.
  • Develop techniques for continuous training, post-training, reinforcement learning, preference optimisation, and controlled model adaptation.
  • Research methods for aligning model behaviour with safety requirements, human preferences, and intended system objectives.
  • Investigate new neural network architectures, learning algorithms, and optimisation techniques for foundation models.
  • Develop approaches for neural network editing, model behaviour modification, machine unlearning, and controlled knowledge updates.
  • Advance research in interpretability, explainable AI, mechanistic analysis, and reverse engineering of neural networks.
  • Develop techniques to understand, diagnose, and modify internal model behaviour.
  • Define long-term research directions and contribute to strategic roadmaps in AI safety, alignment, robustness, and trustworthy AI.
  • Identify emerging research opportunities and establish new technical programmes.
  • Design and run hands-on experiments to validate novel research ideas.
  • Translate successful research outcomes into robust, efficient, and scalable systems.
  • Lead collaborations with universities, research institutions, engineering teams, and external research partners.
  • Provide technical leadership and mentorship to researchers and engineers.
  • Communicate complex research findings clearly to both technical and non-technical stakeholders.

Essential Requirements

  • PhD in Computer Science, Artificial Intelligence, Machine Learning, Deep Learning, Mathematics, or a related technical field.
  • Eight or more years of relevant research experience in artificial intelligence, machine learning, AI safety, security, or a closely related area.
  • Strong research track record with evidence of significant technical contributions.
  • Deep expertise in large language model or vision-language model continuous training, post-training, and alignment.
  • Strong knowledge of reinforcement learning, preference optimisation, RLHF, adversarial training, or related techniques.
  • Proven experience with neural network architecture, algorithm design, model optimisation, and foundation model development.
  • Experience with large-scale model training, adaptation, evaluation, or deployment.
  • Strong understanding of neural network editing, model interpretability, explainable AI, or model reverse engineering.
  • Ability to independently define and lead complex research programmes.
  • Strong experimental design, analytical, and problem-solving skills.
  • Ability to translate research concepts into practical and scalable implementations.
  • Strong written and verbal communication skills.
  • Are you interested in this position?

     

    Apply by clicking on the “Apply Now” button below!

     

    #AlbionarcJobs#FintechJobs

    #AsiaJobs#MiddleEastCareers

    #TechTalent#FintechRecruitment

    #FinanceOpportunities#

     

     

     

     

Apply Now

- £0.00

Select your currency