About the Role
The London Office of UKRD is a core global R&D institution at Huawei, dedicated to advancing artificial intelligence. This role focuses on breaking frontiers in AI research, particularly in multimodal understanding and generation, vision-language large models, and embodied intelligence, backed by an international research team and abundant computing resources.
Responsibilities
- Develop ViT and multimodal large model architectures with improved reasoning and efficiency
- Advance multimodal alignment, representation learning, and long-context modeling
- Explore scalable training methods for large multimodal models
- Optimize model architectures for generalization and performance
- Process large-scale multimodal data across images, videos, audio, and text
- Build pipelines for data cleaning, filtering, annotation, and quality control
- Construct and maintain datasets with versioning and reproducibility
- Optimize data mixtures and sampling strategies for model training
- Improve data quality through feedback-driven curation loops
- Build distributed training systems for large-scale multimodal models
- Optimize GPU utilization, cluster efficiency, and resource scheduling
- Develop open-source training frameworks for scalable model development
- Engineer training, inference, and serving infrastructure
- Improve scalability, stability, and performance of model systems
- Integrate multimodal capabilities into assistant and content generation scenarios
- Translate research into production and user-facing applications
- Collaborate with product and engineering teams to deploy and iterate models
Requirements
- Proficient in Python programming with strong hands-on experience in PyTorch and deep learning frameworks
- Strong algorithm development and implementation skills
- Solid mathematical and logical reasoning ability
- Excellent cross-functional communication and collaboration skills
- Self-driven and highly motivated toward advancing artificial intelligence (AI)
- Strong resilience and the ability to tackle challenging technical problems
- Strong track record of publications in top-tier AI or computer vision conferences (CVPR, ICCV, ECCV, NeurIPS, ICML, ICLR)
- Hands-on experience in large-scale model pre-training or fine-tuning
- High-impact open-source projects or internship experience in leading technology companies within CV, NLP, or multimodal domains
Skills
- Artificial Intelligence (AI)
- Multimodal Understanding
- Multimodal Generation
- Vision-Language Large Models
- Embodied Intelligence
- ViT
- Large Model Architectures
- Multimodal Alignment
- Representation Learning
- Long-Context Modeling
- Scalable Training Methods
- Data Processing
- Data Cleaning
- Data Filtering
- Data Annotation
- Quality Control
- Dataset Construction
- Versioning
- Reproducibility
- Data Mixtures
- Sampling Strategies
- Feedback-driven Curation
- Distributed Training Systems
- GPU Utilization Optimization
- Cluster Efficiency
- Resource Scheduling
- Open-source Training Frameworks
- Model Systems Engineering
- Inference Engineering
- Serving Infrastructure Engineering
- Scalability Improvement
- Stability Improvement
- Performance Improvement
- Multimodal Capabilities Integration
- Assistant Scenarios
- Content Generation Scenarios
- Research Translation
- Production Deployment
- User-facing Applications
- Collaboration
- Python Programming
- PyTorch
- Deep Learning Frameworks
- Algorithm Development
- Mathematical Reasoning
- Logical Reasoning
- Cross-functional Communication
- Computer Vision (CV)
- Natural Language Processing (NLP)
Location
- King's Cross, London
Work Type
- Permanent
- Full-time
Experience Level
- Bachelor’s degree or above
Education Level
- Bachelor’s degree or above in Computer Science, Mathematics, Statistics, or related technical disciplines
Benefits
- 33 days annual leave entitlement per year (including UK public holidays)
- Group Personal Pension
- Life insurance
- Private medical insurance
- Medical expense claim scheme
- Employee Assistance Program
- Cycle to work scheme
- Company sports club and social events
- Additional time off for learning and development
About the Company
- Founded in 1987, Huawei is a leading global provider of information and communications technology (ICT) infrastructure and smart devices.
- We have 207,000 employees and operate in over 170 countries and regions, serving more than three billion people around the world.
- Our vision and mission is to bring digital to the world.
- We drive ubiquitous connectivity and promote equal access to networks.
- We bring cloud and artificial intelligence to provide superior computing power.
- We build digital platforms to help all industries and organizations become more agile, efficient, and dynamic.
- We redefine user experience with AI, making it more personalized.
- Huawei works in close partnership with leading academic institutions in the UK to develop and refine the latest technologies.
- Huawei has the largest Research and Development organization in the world with 96,000+ employees in research centers around the globe.
- In the UK, we already have design centers in Cambridge, London, Edinburgh and Ipswich.
- We continue to explore and define new research directions and new services.
- We have expanded our collaborations with academic researchers; researched new network architectures, integration of communications and key enabling technologies; and developed the fundamental theories of these technologies.
- Huawei's vision is a fully connected, intelligent world.
- We work to inspire passion for basic research around the world.
- Our combined passion drives development across the global innovation value chain.
