About the Role
We are building an AI-native data platform that powers fraud detection and response. This role owns the data platform and data lake, working hands-on with a small group of senior engineers, data scientists, and product partners. The data foundation is critical for detection accuracy, reducing customer losses, and protecting brand trust.
Responsibilities
- Own the design, build, and operation of the data lake and ingestion platform end-to-end.
- Build low-latency batch and streaming pipelines to ingest, normalize, enrich, and serve data.
- Streamline the addition of new data sources to continuously expand risk visibility.
- Establish data quality, freshness, completeness, lineage, and observability for platform trustworthiness.
- Build data pipelines for generative AI, including text processing, embedding generation, and vector storage/retrieval.
- Own deployment, CI/CD, and operational reliability of the platform on Kubernetes.
- Partner with data science, product, and architecture to establish the platform as a shared foundation.
Requirements
- Extensive experience building and operating large-scale data platforms and data lakes with high data volumes.
- Deep, hands-on expertise with Apache Spark, Apache Flink, and modern big-data systems.
- Proven command of best practices for building and maintaining batch and streaming data pipelines.
- Strong production engineering skills across the full delivery lifecycle, including Kubernetes and CI/CD tooling.
- Ability to ship end-to-end solutions.
- Track record of owning data infrastructure end-to-end with limited supervision.
- Experience with generative AI and embedding models, including embedding pipelines, vector databases, and retrieval.
- Cybersecurity or threat intelligence background, with exposure to threat types like phishing, mobile threats, and malware.
- Familiarity with transaction data and transaction fraud signals.
Skills
- Apache Spark
- Apache Flink
- Kubernetes
- CI/CD
- Generative AI
- Embedding models
- Vector databases
- Data pipelines
- Batch processing
- Streaming processing
- Data lakes
- Data platforms
Location
- New York City
Work Type
- On-site
- Hybrid
Experience Level
- Staff
- Principal
Salary/Compensations
- $180,000 – $270,000
- 15% bonus/commission
About the Company
- AppGate is an Equal Opportunity/Affirmative Action Employer.
Equal Opportunity
- All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability or veteran status, age or any other federally protected class.
- AppGate has developed a written affirmative action program available for review upon request.
