A global technology company at the forefront of artificial intelligence and advanced computing. With a strong commitment to research and innovation, the organisation operates internationally, investing in cutting-edge AI technologies and collaborating with leading academic and industry partners to develop next-generation intelligent systems.
Our London-based AI research team is expanding and is seeking a Research Scientist – Computer Vision to contribute to pioneering research in multimodal artificial intelligence. Working alongside a highly experienced international team, you will help develop state-of-the-art models spanning computer vision, multimodal learning, and foundation models, with opportunities to translate research into impactful real-world applications.
This is a permanent, full-time position based in Central London.
Key responsibilities
Frontier AI Research
- Design and develop Vision Transformer (ViT) and multimodal model architectures with enhanced reasoning, efficiency, and scalability
- Advance research in multimodal representation learning, alignment techniques, and long-context modelling
- Investigate scalable training approaches for large multimodal foundation models
- Improve model performance, robustness, and generalisation across diverse tasks
Data & Model Development
- Process and curate large-scale multimodal datasets comprising images, video, audio, and text
- Build robust pipelines for data cleaning, filtering, annotation, and quality assurance
- Maintain reproducible datasets through effective versioning and documentation
- Optimise data sampling strategies and improve dataset quality through iterative evaluation and feedback
Systems & Infrastructure
- Develop distributed training systems for large-scale multimodal models
- Optimise GPU utilisation, resource scheduling, and training efficiency
- Contribute to training and inference frameworks that support scalable model development
- Improve the reliability, performance, and scalability of AI infrastructure
Research Translation
- Apply advanced multimodal AI capabilities to intelligent products and user-facing applications
- Work closely with engineering and product teams to bring research innovations into production
- Contribute to the continuous improvement and deployment of cutting-edge AI technologies
Person specification
Essential
- Degree in Computer Science, Mathematics, Statistics, Artificial Intelligence, or a related technical discipline
- Strong Python programming skills with hands-on experience using PyTorch and modern deep learning frameworks
- Excellent algorithmic thinking, mathematical reasoning, and problem-solving ability
- Strong communication and collaboration skills, with the ability to work effectively across multidisciplinary teams
- A proactive, self-motivated approach and enthusiasm for tackling challenging research problems
Desirable
- Publications at leading AI or computer vision conferences (e.g. CVPR, ICCV, ECCV, NeurIPS, ICML, or ICLR)
- Experience training or fine-tuning large-scale vision, language, or multimodal models
- Contributions to open-source AI projects or research experience within industry or academic laboratories
Please contact Charles Duran for more information.