Computer Vision Research Scientist

London, England
Exciting role within global market leading company
Job ID: 491310

A global technology company at the forefront of artificial intelligence and advanced computing. With a strong commitment to research and innovation, the organisation operates internationally, investing in cutting-edge AI technologies and collaborating with leading academic and industry partners to develop next-generation intelligent systems.

Our London-based AI research team is expanding and is seeking a Research Scientist – Computer Vision to contribute to pioneering research in multimodal artificial intelligence. Working alongside a highly experienced international team, you will help develop state-of-the-art models spanning computer vision, multimodal learning, and foundation models, with opportunities to translate research into impactful real-world applications.

This is a permanent, full-time position based in Central London.

Key responsibilities

Frontier AI Research

  • Design and develop Vision Transformer (ViT) and multimodal model architectures with enhanced reasoning, efficiency, and scalability
  • Advance research in multimodal representation learning, alignment techniques, and long-context modelling
  • Investigate scalable training approaches for large multimodal foundation models
  • Improve model performance, robustness, and generalisation across diverse tasks

Data & Model Development

  • Process and curate large-scale multimodal datasets comprising images, video, audio, and text
  • Build robust pipelines for data cleaning, filtering, annotation, and quality assurance
  • Maintain reproducible datasets through effective versioning and documentation
  • Optimise data sampling strategies and improve dataset quality through iterative evaluation and feedback

Systems & Infrastructure

  • Develop distributed training systems for large-scale multimodal models
  • Optimise GPU utilisation, resource scheduling, and training efficiency
  • Contribute to training and inference frameworks that support scalable model development
  • Improve the reliability, performance, and scalability of AI infrastructure

Research Translation

  • Apply advanced multimodal AI capabilities to intelligent products and user-facing applications
  • Work closely with engineering and product teams to bring research innovations into production
  • Contribute to the continuous improvement and deployment of cutting-edge AI technologies

Person specification

Essential

  • Degree in Computer Science, Mathematics, Statistics, Artificial Intelligence, or a related technical discipline
  • Strong Python programming skills with hands-on experience using PyTorch and modern deep learning frameworks
  • Excellent algorithmic thinking, mathematical reasoning, and problem-solving ability
  • Strong communication and collaboration skills, with the ability to work effectively across multidisciplinary teams
  • A proactive, self-motivated approach and enthusiasm for tackling challenging research problems

Desirable

  • Publications at leading AI or computer vision conferences (e.g. CVPR, ICCV, ECCV, NeurIPS, ICML, or ICLR)
  • Experience training or fine-tuning large-scale vision, language, or multimodal models
  • Contributions to open-source AI projects or research experience within industry or academic laboratories

Please contact Charles Duran for more information.