Jobs Companies OpenDataJobs Data Engineer (AWS, Spark)

Sobre este puesto de Data Engineer (AWS, Spark) en OpenDataJobs

OpenDataJobs · Híbrido · Washington, District of Columbia, United States

Peregrine Advisors is a firm founded on a simple conviction: the best solutions come from the people closest to the problem, given real ownership and the tools to deliver. We are a data and technology innovation hub and a Benefit Corporation working at the center of the federal government's mission to deliver for client stakeholders and the US public, looking for highly motivated contributors who thrive when trusted to own a hard problem and equipped to deliver the solution.

Your first assignment

Your first assignment is to build the pipelines that move a federal agency's data from source to platform on Amazon Web Services (AWS): ingest, process, store, and keep it clean and trustworthy at scale. The work is real, hard, and it matters. It is also where you start, not the shape of your career here: we hire people, not seats, and we move our best to where the hardest problems are.

What you'll build

  • Ingest-process-store pipelines on AWS: Spark-based extract, transform, and load (ETL) with Glue, Amazon EMR, Lambda, and Step Functions, in Python and PySpark.
  • A data lake that feeds the platform: S3 design (Parquet, partitioning, lifecycle) into Apache Iceberg tables, PostgreSQL on Amazon Aurora, and DynamoDB, with Trino for federated SQL across them, plus event orchestration, secrets and monitoring, and data quality, validation, and lineage built in.
  • In time, the firm itself: new capabilities, tools, and lines of business you help spin up.

Who you are

You are a data engineer who is serious about the craft and eager to go deeper. You care as much about whether the data is trustworthy as whether it arrives, and you do your best work alongside people who push you. You experiment, fail, learn, and repeat quickly. You would rather own an outcome than be handed a task.

What you bring

  • Hands-on data engineering on AWS: Spark ETL (Glue, EMR), Python and PySpark, and S3 data-lake design feeding the platform stores. — and if a second lake bullet sits under "What you bring", it becomes: The reliability craft around it: event orchestration, data quality and lineage, monitoring, and infrastructure-as-code.
  • The reliability craft around it: event orchestration, data quality and lineage, monitoring, and infrastructure-as-code.
  • The judgment to build in a regulated environment where accuracy and auditability are not optional.

What you'll need

Sole United States citizenship and the ability to obtain a Public Trust determination are required for this initial engagement. This is a hybrid role based in the Washington, DC metropolitan area, and it requires commuting into DC regularly. Everything else, the years, the certifications, the specific tools, we ask in the application and get into during the interview

What we offer

A high-performing team of developers, engineers, data scientists, architects, and strategists solving complex, real-world problems, with work that runs from strategy formulation to hands-on implementation. We develop people across roles and clients, with extensive onboarding and sponsored training and professional development. And Peregrine has been a Benefit Corporation from day one: public value is built into the work itself, not bolted on afterward. Work worth your best years.

What we commit to

As a Benefit Corporation, our commitment runs three ways: real, measurable value for our clients; government that works better for the public; and a team that makes everyone in it better.

We hire people who want to help build the firm, not just work at it. If that is you, apply.

Peregrine Advisors is an equal opportunity employer.

Peregrine exclusively works with OPEN Data Jobs to recruit our team. Register with OPEN Data Jobs to be considered for this opening and future roles. 

Requirements

  • 4+ years of data engineering experience
  • Bachelor's degree
  • Spark ETL on AWS (Glue, Amazon EMR) in Python and PySpark
  • S3 data-lake design (Parquet, partitioning, lifecycle) feeding Apache Iceberg tables, Amazon Aurora PostgreSQL, and DynamoDB
  • Event orchestration (Lambda, Step Functions, SQS/SNS) with secrets management and monitoring
  • Data quality, validation, and lineage
  • Infrastructure-as-code (CloudFormation or Terraform)
  • Basic proficiency in writing, PowerPoint, and Excel

Preferred

  • Master's degree in a relevant field
  • Trino or comparable federated SQL across the lake and relational stores
  • Apache Ranger-governed access
  • Legacy ETL migration (for example DataStage)
  • Federal information technology or high-volume data experience
  • Familiarity with AI-assisted developer tooling

Benefits

Medical, dental, and vision with the employee premium fully paid and half of dependent premiums; employer-paid life, accidental death, and short-term and long-term disability insurance; a 401(k) matched 100% up to 4% of salary, vesting immediately; unlimited paid time off; and sponsored professional certifications and continuing education.

¿Listo para postularte en OpenDataJobs?
Postúlate en OpenDataJobs

Sobre OpenDataJobs

Connecting Data and Tech Professionals with a world of opportunities

Ver todos los empleos en OpenDataJobs →

Empleos similares

OpenDataJobs
Data Engineer Role
OpenDataJobs
⚡ Postúlate pronto Washington, District of Columb... Presencial
● Nuevo 👁 Visto ✓ Postulado hace 2sem
Node.Digital
Senior Data Engineer
Node.Digital
⚡ Postúlate pronto Washington, District of Columb... · restringido por ubicación
● Nuevo 👁 Visto ✓ Postulado hace 4sem
Delaware Nation Industries
Data Platform Engineer
Delaware Nation Industries
⚡ Postúlate pronto Washington, District of Columb... Presencial
● Nuevo 👁 Visto ✓ Postulado hace 1 mes
9W
Expert ETL Developer
9th Way Insignia
⚡ Postúlate pronto Washington, District of Columb... Presencial $118,737–$150,000
● Nuevo 👁 Visto ✓ Postulado hace 1 mes
9W
Expert ETL Developer
9th Way Insignia
⚡ Postúlate pronto Washington, District of Columb... Presencial $120,518–$150,000
● Nuevo 👁 Visto ✓ Postulado hace 1 mes
9W
ETL Developer
9th Way Insignia
⚡ Postúlate pronto Washington, District of Columb... Presencial $77,135–$89,000
● Nuevo 👁 Visto ✓ Postulado hace 1 mes
H2 Performance Consulting
Data Systems Engineer
H2 Performance Consulting
⚡ Postúlate pronto Washington, District of Columb... Presencial
● Nuevo 👁 Visto ✓ Postulado hace 1 mes
Rhombus Power, Inc.
Data Engineer (U.S. Citizen, Secret/Top Secret), Washington D.C.
Rhombus Power, Inc.
⚡ Postúlate pronto Washington, District of Columb... Presencial
● Nuevo 👁 Visto ✓ Postulado hace 2 meses
Datacom
Senior Data & Analytics Engineer
Datacom
⚡ Postúlate pronto Taguig, Philippines Presencial
● Nuevo 👁 Visto ✓ Postulado hace 1h

Regístrate para recibir sugerencias adaptadas a los empleos que abres y las búsquedas que guardas.

Más empleos en OpenDataJobs

Ver todos los empleos en OpenDataJobs →

Postúlate ahora
🤖

Un momento — para

JobsRadar se creó para personas reales que están pasando un mal momento en su búsqueda de empleo — no para solicitudes automatizadas. Estás haciendo clic demasiado rápido y ahora estás bloqueado temporalmente.

Vuelve más tarde. Si de verdad estás buscando empleo, cuentas con nosotros — solo compórtate como una persona.

Catch your next role the second it’s posted.

Create a free account and we’ll watch the boards for you — the instant a job matches your search, it lands in your inbox or Telegram. No digging, no refreshing.

Create free account

Free forever · takes 30 seconds · already have one?

Toma ventaja en tu búsqueda de empleo.

Únete a nuestro canal de Telegram para lo que te ayuda a conseguir el puesto — referencias salariales, el pulso semanal del mercado y avisos de nuevas funciones. Sin spam, solo señal.

Únete al canal — es gratis