About the Role
The Data Scientist will design, develop, evaluate, and optimize machine learning models to solve complex business problems for enterprise clients. This role involves close collaboration with business partners, engineers, and client stakeholders to deliver scalable AI solutions and ensure effective production model performance.
Responsibilities
- Collaborate with business and technical teams to define machine learning objectives, success metrics, and measurable outcomes.
- Identify, ingest, transform, and enrich structured and unstructured data for model development.
- Engineer features from structured and text-based data, including entity extraction, normalization, embeddings, and feature generation.
- Apply statistical and machine learning techniques including classification, regression, clustering, and deep learning models.
- Design and execute experiments including A/B testing, hypothesis testing, causal analysis, and model benchmarking.
- Build production-ready machine learning models using Python and Spark.
- Design automated evaluation methodologies including benchmark datasets, acceptance tests, and promotion gates.
- Monitor production model performance, identify model drift and degradation, and continuously improve model accuracy through retraining and optimization.
- Develop scalable machine learning solutions that integrate with enterprise AI platforms.
- Present methodologies, findings, and recommendations to both technical teams and executive stakeholders.
Requirements
- Passionate about solving business problems through data science and artificial intelligence.
- Experienced working with large-scale enterprise data environments.
- Comfortable owning machine learning models throughout their lifecycle from development through production.
- Strong analytical thinker with excellent problem-solving abilities.
- Fast learner with attention to detail.
- Outstanding verbal and written communication skills.
- Able to collaborate effectively with cross-functional teams and client stakeholders.
Skills
- Advanced SQL
- Python (NumPy, Pandas, Scikit-learn, TensorFlow, PyTorch)
- Apache Spark / PySpark
- AWS SageMaker (Databricks, Azure ML, or Vertex AI experience is a plus)
- Natural Language Processing (NLP), embeddings, entity extraction, and feature engineering
- Statistical modeling, regression, experimental design, hypothesis testing, and drift analysis
- Automated testing and data validation
- Experience deploying and supporting production machine learning models
- Git and modern software development practices
Experience Level
- 3+ years of experience in Data Science, Machine Learning, or Advanced Analytics
Education Level
- Bachelor's or Master's degree in Data Science, Computer Science, Statistics, Mathematics, Engineering, or another quantitative field
Salary/Compensations
- A competitive salary and an annual bonus
Benefits
- Full health and commuter benefits
- Standard time off, sick leave, and time off on all national holidays
- Visa sponsorship
About the Company
- At BDIPlus, our mission is to help enterprises utilize their resources more efficiently, implement effective information management and empower them by enabling richer insights and intelligence.
- We are driven by a single purpose: empower the technology transformation.
- We are passionate about creating foundational technology platforms for enterprise data and information management.
- Our employees are at the heart of the work we do at BDIPlus.
- We are committed to encouraging and celebrating innovation, creativity, and hard work among our team members.
- A diverse, fun to work with, highly intelligent and innovative team.
- An environment where creative thinking is encouraged and innovation is a driving force of everything we do.
