Data Engineer (AWS, Spark)
At Peregrine Advisors, you will build the pipelines that move a federal agency's data from source to platform. But we hire people, not seats. As the work evolves, you will learn new systems and tools, take on greater responsibility, and help develop the firm's capabilities, tools, and lines of business. We move our best to where the hardest problems are.
This is a full-time W-2 position with a salary of $103,000 to $140,000 per year. It is a hybrid work arrangement based in the Washington, DC metropolitan area, and the commuting cadence varies by assignment. The initial engagement requires United States citizenship and the ability to obtain a Public Trust determination.
We are a data and technology innovation hub and a Benefit Corporation working at the center of the federal government's mission to deliver for client stakeholders and the US public.
Your first project
Your first project will likely have you building ingest, processing, and storage at scale. Depending on the assignment, you may:
- Build Spark-based extract, transform, and load (ETL) pipelines with Glue, Amazon EMR, Lambda, and Step Functions.
- Write the processing in Python and PySpark.
- Design the S3 layer, including Parquet, partitioning, and lifecycle, feeding Apache Iceberg tables.
- Connect PostgreSQL on Amazon Aurora and DynamoDB, with Trino for federated Structured Query Language (SQL) across them.
You deliver pipelines