Data Pipeline Engineer
Your Impact
Join a small team building the next generation of cybersecurity products from the ground up. Led by industry veterans with a proven track record of success - you will get to architect, build, and deliver hugely impactful products with this world-class team. You will have the opportunity to grow your career and skills along with the company from the very start.
Role Overview
Design, build, and maintain a scalable, open-source data lakehouse architecture supporting petabyte-scale analytics workloads. Responsible for architecting end-to-end data pipelines from ingestion through transformation to consumption, ensuring high performance, reliability, and data quality.
Required Experience
A proven track record of success architecting, building, and running large-scale data systems (PB scale)
Experience with both batch and real-time processing architectures
Experience with open source Data Lakehouse components, including Apache Iceberg, PostgreSQL, Neo4j, Apache Parquet, etc.
Experience with tools for stream processing and data analytics - Apache Kafka, Spark, Flink, etc.