About the Role
We are looking for talented and experienced data scientists with solid knowledge and experience of AI and ML to join our programme. The role involves working with the existing delivery team to deliver models, documentation, and associated productionised services.
Responsibilities
- Design and develop AI / ML based solutions
- Work with other data scientists to build and deploy production-level solutions
- Troubleshoot and debug code
- Work with other teams to understand and solve business problems
Requirements
- Python (pandas, NumPy, scikit-learn) for data wrangling, modelling, and feature engineering
- SQL for querying structured data sources
- Experience with classification, unsupervised learning (e.g. outlier detection), and ranking models
- Familiarity with containerised deployment (e.g. Podman, SageMaker, DSW pipelines)
- Version Control (Git) to maintain reproducible and collaborative workflows
- Time-Series Analysis to assess risk trends over financial years
- Exploratory Data Analysis (EDA) to spot early signals or risk clusters
- Understanding methods like Robust Rank Fusion (RRF)
- Familiarity with model explainability tools (e.g., SHAP, LIME) to support interpretability
- Experience with Model Monitoring & Drift Detection
- Experience in RegTech / FinCrime / Data-led Supervision Projects is a plus
- Experience developing solutions for record linkage and/or network analytics tasks
- Experience with graph query languages (e.g., Gremlin, Cypher), graph database platforms (e.g., Neptune, Neo4j), and/or graph visualisation platforms
Skills
- AI
- ML
- Python
- pandas
- NumPy
- scikit-learn
- SQL
- Model Development
- Model Validation
- Classification
- Unsupervised Learning
- Outlier Detection
- Ranking Models
- Machine Learning Deployment
- Containerised Deployment
- Podman
- SageMaker
- DSW pipelines
- Version Control
- Git
- Time-Series Analysis
- Exploratory Data Analysis (EDA)
- Rank Aggregation
- Ensemble Techniques
- Robust Rank Fusion (RRF)
- Model Explainability
- SHAP
- LIME
- Model Monitoring
- Drift Detection
- RegTech
- FinCrime
- Data-led Supervision
- Record Linkage
- Network Analytics
- Graph Query Languages
- Gremlin
- Cypher
- Graph Database Platforms
- Neptune
- Neo4j
- Graph Visualisation Platforms
Experience Level
- 3-5 years
