About the Role
This is a foundational infrastructure role where the data layer serves as the nervous system for a payments platform processing real-time agent transactions, policy decisions, and risk signals. You will define the standards, architecture, and culture of data at Sapiom, focusing on data quality, governance, and scalable pipeline design.
Responsibilities
- Build, scale, and optimize production-quality ETL pipelines from ingestion to availability.
- Design data schemas and architecture capable of handling 10x growth.
- Establish and maintain data quality, governance, and security standards.
- Develop standardized, self-serve data models to enable AI-powered analytics.
- Instrument pipeline observability and surface health metrics to cross-functional teams.
- Partner with Data Science, Analytics, and DevOps teams to act as a force multiplier.
Requirements
- 5+ years of experience transforming raw data into governed, production-ready datasets.
- Deep experience building and deploying production data pipelines using SQL, Python, Spark, AWS Glue, EMR, DBT, and Airflow.
- 3+ years of hands-on production experience with MPP databases like Snowflake, AWS Redshift, or Teradata.
- Proven ability to partner effectively with Engineering, Analytics, Data Science, and DevOps teams.
- Strong architectural instincts for designing scalable systems.
- Comfortable participating in an on-call rotation for incident response.
- Strong communication skills for translating complex infrastructure decisions to diverse stakeholders.
Skills
- SQL
- Python
- Spark
- AWS Glue
- EMR
- DBT
- Airflow
- Snowflake
- AWS Redshift
- Teradata
- Data Governance
- ETL Pipeline Design
- Data Modeling
- System Architecture
Experience Level
- 5+ years of experience
About the Company
- Sapiom builds financial payments infrastructure for the machine economy, enabling AI agents to transact safely.
- Backed by $15.75M in investment from Accel, Menlo, and Anthropic.
