About the Role
As a Software Engineer within the Canadian Banking Engineering team, you will design, develop, and modernize enterprise-grade data integration platforms. You will focus on building scalable batch and distributed data pipelines, enabling high-performance processing with Spark, and transforming legacy ETL workloads into modern, cloud-ready architectures.
Responsibilities
- Design, develop, and support scalable ETL and data pipelines using Talend and Apache Spark.
- Lead and contribute to the migration of legacy ETL workloads to modern frameworks like Spark.
- Build and optimize large-scale batch and distributed data processing pipelines.
- Analyze ETL performance bottlenecks and implement tuning strategies.
- Develop reusable ETL frameworks, components, and orchestration patterns.
- Implement data ingestion, transformation, and quality checks for structured and semi-structured data.
- Develop and maintain complex SQL transformations, stored procedures, and data models.
- Integrate ETL pipelines with enterprise systems using messaging, APIs, and batch orchestration.
- Ensure data integrity, lineage, reconciliation, and auditability.
- Collaborate with architecture and engineering teams to design target-state data platforms.
- Participate in end-to-end system integration and migration testing.
- Contribute to technical design discussions and provide stakeholder input.
- Collaborate with cross-functional teams including data engineering, application support, and infrastructure.
- Mentor junior developers and promote best practices in ETL design and performance tuning.
- Ensure adherence to coding standards, version control, and CI/CD practices.
Requirements
- Strong hands-on experience with Talend ETL development.
- Strong experience with Apache Spark (PySpark or Scala) for large-scale data processing.
- Experience with distributed data processing and big data frameworks.
- Strong Unix/Linux scripting experience for orchestration and automation.
- Strong experience with relational databases such as DB2 or Oracle.
- Advanced SQL proficiency including complex transformations, performance tuning, and stored procedures.
- Experience working with large datasets and data warehousing concepts.
- Proven experience in enterprise ETL and data pipeline development.
- Experience with data migration and modernization initiatives.
- Experience with batch processing frameworks, data ingestion, messaging systems (Kafka, MQ), and API-based integrations.
- Experience with source control systems like Git, Bitbucket, or GitHub.
- Hands-on experience supporting ETL modernization or platform migration programs.
- Understanding of data pipeline re-engineering and performance optimization strategies.
Skills
- Talend
- Apache Spark
- PySpark
- Scala
- Unix/Linux Scripting
- DB2
- Oracle
- SQL
- Kafka
- MQ
- Git
- Bitbucket
- GitHub
- Azure
- AWS
- GCP
- Airflow
- Spark Streaming
- Kafka Streams
Location
- Toronto, Ontario, Canada
About the Company
- Scotiabank is a leading bank in the Americas.
- Guided by the purpose: for every future.
- Provides personal and commercial banking, wealth management, private banking, corporate and investment banking, and capital markets services.
Equal Opportunity
- Scotiabank is committed to creating and maintaining an inclusive and accessible environment for everyone.
- The company provides accommodations during the recruitment and selection process upon request.
