Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Data Engineer (AWS, Spark)

$103k - $140k

OpenDataJobs

Job Description

Job Description

At Peregrine Advisors, you will build the pipelines that move a federal agency's data from source to platform. But we hire people, not seats. As the work evolves, you will learn new systems and tools, take on greater responsibility, and help develop the firm's capabilities, tools, and lines of business. We move our best to where the hardest problems are.

This is a hybrid work arrangement based in the Washington, DC metropolitan area, and the commuting cadence varies by assignment. The initial engagement requires United States citizenship and the ability to obtain a Public Trust determination.

We are a data and technology innovation hub and a Benefit Corporation working at the center of the federal government's mission to deliver for client stakeholders and the US public.

Your first project

Your first project will likely have you building ingest, processing, and storage at scale. Depending on the assignment, you may:

  • Build Spark-based extract, transform, and load (ETL) pipelines with Glue, Amazon EMR, Lambda, and Step Functions.
  • Write the processing in Python and PySpark.
  • Design the S3 layer, including Parquet, partitioning, and lifecycle, feeding Apache Iceberg tables.
  • Connect PostgreSQL on Amazon Aurora and DynamoDB, with Trino for federated Structured Query Language (SQL) across them.

You deliver pipelines that keep data clean and trustworthy at the volumes a federal platform runs at. You automate ingestion that used to be handled case by case, and you build the storage design everything downstream depends on.

This is where you start, not the shape of your career here.

Requirements

  • 4+ years of data engineering experience
  • Bachelor's degree
  • Spark ETL on AWS (Glue, Amazon EMR) in Python and PySpark
  • S3 data-lake design (Parquet, partitioning, lifecycle) feeding Apache Iceberg tables, Amazon Aurora PostgreSQL, and DynamoDB
  • Event orchestration (Lambda, Step Functions, SQS/SNS) with secrets management and monitoring
  • Data quality, validation, and lineage
  • Infrastructure-as-code (CloudFormation or Terraform)
  • Basic proficiency in writing, PowerPoint, and Excel

Preferred

  • Master's degree in a relevant field
  • Trino or comparable federated SQL across the lake and relational stores
  • Apache Ranger-governed access
  • Legacy ETL migration (for example DataStage)
  • Federal information technology or high-volume data experience
  • Familiarity with AI-assisted developer tooling

Who you are

You are a data engineer who wants to get better at it, and you know which parts you have not mastered yet. You care as much about whether the data is trustworthy as whether it arrives, and you do your best work alongside people who push you. You experiment, fail, learn, and repeat quickly. You would rather own an outcome than be handed a task.

What you bring

  • Hands-on data engineering on AWS: Spark ETL (Glue, EMR), Python and PySpark, and S3 data-lake design feeding the platform stores.
  • The reliability craft around it: event orchestration, data quality and lineage, monitoring, and infrastructure-as-code.
  • The judgment to build in a regulated environment where accuracy and auditability are not optional.

You adapt your development workflow as AI tools evolve, using them to help implement, test, and improve the components you own. You give the tools clear context, review and test their output, and remain accountable for the code you deliver.

When we talk

Be prepared to discuss a difficult problem you worked through, the decisions you made, what happened, and what you learned.

Benefits

This is a full-time W-2 position with a salary of $103,000 to $140,000 per year.

Benefits include medical, dental, and vision insurance with the employee premium fully paid and half of dependent premiums; employer-paid life, accidental death, and short-term and long-term disability insurance; a 401(k) matched 100% up to 4% of salary, vesting immediately; unlimited paid time off; and sponsored professional certifications and continuing education.

What we offer

You will work alongside developers, engineers, data scientists, architects, and strategists, on work ranging from strategy to implementation. We support your development across assignments and clients through extensive onboarding, sponsored professional certifications such as the Data Management Capability Assessment Model (DCAM) and AWS technical certifications, and rotation across functions to expand your skills and perspective. The mission is real, the problems are hard, and you own what you ship.

What we commit to

As a Benefit Corporation, our commitment runs three ways: real, measurable value for our clients; government that works better for the public; and a team that makes everyone in it better.

We hire people who want to help build the firm, not just work at it. If that is you, apply.

Peregrine Advisors is an equal opportunity employer.

Peregrine exclusively works with OPEN Data Jobs to recruit our team. Register with OPEN Data Jobs to be considered for this and future openings.

Vacancy posted 3 days ago
Similar jobs that could be interesting for youBased on the Data Engineer (AWS, Spark) in Washington DC vacancy
  • Programmers.io seeks a Big Data Engineer for a full-time hybrid role in Arlington, VA. You will...  ...end-to-end pipelines with Python/Scala, Spark and Hadoop, ensuring data quality and scalable...  ..., with cloud familiarity (Databricks or AWS) as a bonus. Onsite 3 days per week. #J-... 
    Amazon Web Service
    Full time
    3 days per week

    Programmers.io

    Arlington, VA
    2 days ago
  •  ...Opportunity SDA is seeking a Senior Big Data Engineer / Java Data Integration Engineer to support...  ...using technologies such as Kafka, Spark, Elasticsearch, and cloud-based platforms...  ...Integration | Data Mapping | REST APIs | MongoDB | AWS | Linux What You’ll Do Design, develop,... 
    Amazon Web Service
    Full time
    Remote work
    Flexible hours

    Systems Development and Analysis

    Arlington, VA
    2 days ago
  • Eliassen Group seeks a Data Engineer to design and support a Python-based, AWS-native data ingestion platform that moves financial data from source systems into...  ...work with CDK-provisioned infrastructure, Glue and Spark processing, and Lambda/ECS/Flink pipelines, ensuring... 
    Amazon Web Service

    Eliassen Group

    Mc Lean, VA
    18 hours ago
  • $86.7k - $170.9k

     ...Join Deloitte’s Core AI & Data practice and help organizations...  ...capabilities. As a Databricks Data Engineer, you will support the design,...  ...using Databricks, Apache Spark, Python, and Structured Query...  ...platform: Amazon Web Services (AWS), Microsoft Azure, or Google Cloud... 
    Amazon Web Service
    Local area
    Visa sponsorship

    Deloitte

    Rosslyn, VA
    13 hours ago
  •  ...Description Job Description Role Name : Data Engineer, Healthcare and Life Science Domain :...  ...Architect end-to-end solutions on AWS and Databricks Define architectural standards...  ...Databricks, DBT Core or Cloud, Python, Spark, SQL, distributed compute paradigms including... 
    Amazon Web Service
    Full time

    YO AI Labs

    Washington DC
    a month ago
  • $120k - $140k

     ...Section Details Job Title Data Engineer Job Type Full Time...  ...Work with cloud data platforms such as AWS, Azure, or GCP and services including data...  ...), or Azure. ~ Proficiency in Apache Spark, Kafka, NoSQL databases, and orchestration... 
    Amazon Web Service
    Full time

    SFE

    Washington DC
    2 days ago
  •  ...datasets. •Analyze complex data structures and source-to-target...  ...into effective data engineering solutions. •Monitor pipeline...  ...-based data platforms such as AWS Redshift, Google BigQuery, Azure...  ...big data technologies (Hadoop, Spark, or distributed processing frameworks... 
    Amazon Web Service

    Bow Wave LLC

    Arlington, VA
    16 days ago
  •  ...easily move between business, data management, and technical teams...  ...in using Python or Scala, Spark, Hadoop platforms & tools (Hive...  ...like data ingestion, feature engineering, modeling, tuning, evaluating,...  ...Cloud knowledge (Databricks or AWS ecosystem) is a plus but not required... 
    Amazon Web Service
    3 days per week

    programmers.io

    Arlington, VA
    2 days ago
  •  ...Description Description: Anivas Tech is seeking a Data Engineer to support our government agency IT...  ...integration tools and platforms (e.g., Apache Spark, Kafka, Airflow, or equivalent) ~ Familiarity with cloud platforms (AWS, Azure, or GCP) and cloud-native data... 
    Amazon Web Service

    Anivas Tech LLC

    Upper Marlboro, MD
    a month ago
  •  ...Makpar has an exciting opportunity for a Senior Data Engineer to join our growing team.his role requires...  ...Expert-level proficiency in Databricks, Apache Spark, and SQL . ~ Experience with Kafka, Airflow, or AWS Glue for ETL. ~ Proficiency in CI/CD pipelines... 
    Amazon Web Service
    Contract work
    For subcontractor
    Work from home
    Flexible hours

    IQUASAR LLC

    Washington DC
    a month ago
  •  ...under the guidance of more experienced engineers. •Support analysis of data structures, mappings, and data...  ...storage or compute platforms (e.g., AWS S3/Redshift, Google BigQuery, Azure...  ...Exposure to big data technologies (Hadoop, Spark, or distributed processing... 
    Amazon Web Service
    Internship

    Bow Wave LLC

    Arlington, VA
    16 days ago
  •  ...The Mission: Square Peg is seeking a Data Engineer to be responsible for designing, implementing,...  ...based database services such as Databricks and AWS Preferred Qualifications Experience with using Apache Spark and Hadoop to conduct large-scale data processing... 
    Amazon Web Service

    Square Peg Technologies

    Washington DC
    a month ago
  •  ...functional and technical expertise in cloud engineering, data management, cybersecurity and emerging...  ...Experience with Databricks, Apache Spark, and SQL. •      Proficiency in Python...  ...with ETL tools such as Airflow, Kafka, or AWS Glue. •      Familiarity with GitHub/... 
    Amazon Web Service
    Start working today
    Local area
    Flexible hours

    Makpar

    Washington DC
    9 days ago
  •  ...Description Job Overview We are seeking a Data Engineer to support our Federal Government...  ...full-stack applications, microservices, and AWS infrastructure, setting architectural direction...  ...data processing pipelines using Spark and PySpark, defining partitioning, caching... 
    Amazon Web Service
    1 day per week

    Ryde Technologies, LLC

    Arlington, VA
    a month ago
  • $101.3k - $160k

    Business Data Technologies (BDT) makes it easier for teams across...  ...managed solutions combine standard AWS tooling, open-source products,...  ...BDT customers move beyond the engineering and operational burden...  ...pipelines using SQL, Python and Spark. · Build and deliver high quality... 
    Amazon Web Service
    Flexible hours

    Amazon

    Washington DC
    2 days ago
  •  ...Expression is seeking an experienced Data Engineer to support the design, development, and operational...  ...modeling approaches using Python, Spark, and cloud-native ML frameworks such as SageMaker...  ...cloud-based environments, including AWS, Azure, or Palantir Foundry. ~ Strong... 
    Amazon Web Service
    Full time
    For contractors
    Work at office
    Immediate start

    Expression

    Arlington, VA
    a month ago
  •  ...Expression is seeking a Data Engineer to support the Drug Enforcement Administration (DEA) Investigative...  ...to cloud-based data platforms such as AWS Redshift, Google BigQuery, Azure SQL, or...  ...of big-data technologies such as Hadoop, Spark, or other distributed-processing... 
    Amazon Web Service
    Full time
    For contractors
    Work at office
    Immediate start

    Expression

    Washington DC
    a month ago
  •  ...help organizations move from data complexity to decision advantage...  ...excited to find our next Data Engineer to join the company. As a...  ...processing frameworks (e.g., Spark). ~ Strong experience with relational...  ...data modeling. ~ Advanced AWS experience (S3, RDS, EMR,... 
    Amazon Web Service
    Remote work
    Flexible hours
    Shift work

    Virtualitics, Inc

    Washington DC
    9 days ago
  •  ...Analytica is seeking a Databricks Data Engineer to support a high-profile program for a financial...  ...and platform services within the AWS and Databricks ecosystem. This role works...  ...concepts. Experience with Databricks, Apache Spark, PySpark, or similar big data... 
    Amazon Web Service
    Full time
    For contractors
    Local area

    Analytica

    Washington DC
    a month ago
  •  ...Description Job Description 540 is seeking a Data Engineer to support a mission-critical federal...  ..., production-ready data pipelines using Spark, Python, and SQL in a Databricks...  ...NICE TO HAVE SKILLS & EXPERIENCE AWS cloud experience Experience working with... 
    Amazon Web Service
    Temporary work
    Work at office
    Local area
    Remote work
    Flexible hours

    540

    Arlington, VA
    29 days ago
  •  ...Analytica is seeking a  Senior Data Engineer  to support large-scale Health care data modernization...  ....  ~ Strong proficiency with Apache Spark, PySpark, SQL, and Python.  ~...  ...experience with cloud data platforms in AWS or Azure cloud environments.   ~ Experience... 
    Amazon Web Service
    Full time
    For contractors
    Local area

    Analytica

    Washington DC
    more than 2 months ago
  • $130k - $165k

     ...transformation, human-centered design, data analytics and visualization, and...  ...CTG is seeking a Senior Data Engineer to design, build, and maintain scalable...  ...models using Python, Apache Spark (PySpark), dbt, SQL (PostgreSQL), and AWS Glue . Develop and optimize... 
    Amazon Web Service
    Full time
    Temporary work
    Remote work

    Capital Technology Group

    Silver Spring, MD
    6 days ago
  • $103k - $153k

     ...As a Data Engineer II , you will support Homes.com by helping shape sitewide tracking architecture and ensuring data flows...  ...processes using Databricks , Snowflake , and other AWS services Implement and optimize Spark jobs , data transformations, and data processing workflows... 
    Amazon Web Service

    CoStar Group

    Arlington, VA
    1 day ago
  •  ...Job Description 540 is seeking a Senior Data Engineer to support a mission-critical federal health...  ...data pipelines and data products using Spark (Python/SQL) in a Databricks environment...  ...NICE TO HAVE SKILLS & EXPERIENCE AWS cloud experience Experience working with... 
    Amazon Web Service
    Temporary work
    Work at office
    Local area
    Remote work
    Flexible hours

    540

    Arlington, VA
    29 days ago
  • Programmers.io in Arlington, VA is seeking a data engineer who can bridge business, data, and...  ...on data quality, using Python or Scala, Spark and Hadoop, plus SQL for analytics. Experience...  ...and cloud platforms like Databricks or AWS is a plus as you deploy production-grade... 
    Amazon Web Service

    programmers.io

    Arlington, VA
    3 days ago
  • $150k - $200k

     ...Capital Technology Group seeks a Lead Data Engineer to build scalable AWS-native data platforms for mission-critical analytics in federal environments...  ...ETL/ELT workflows, and data models using Python , Apache Spark (PySpark) , SQL (PostgreSQL) , and AWS Glue Develop and optimize... 
    Amazon Web Service
    Temporary work
    Remote work

    DataJobs.com

    Washington DC
    3 days ago
  •  ...Leverage a variety of modern big data tools / approaches to solve complex...  ...technologies (e.g. - Hadoop, Spark, Kafka, Kubernetes, Terraform, Airflow, AWS, Azure, GCP, etc.) Conduct peer code...  ...stakeholders) Develop data engineering designs that positively impacts business... 
    Amazon Web Service
    Local area

    Confidential

    Washington DC
    1 day ago
  • Senior Data Engineer with 3-5 years of experience // Secret with SCI or TOP secret with SCI / Mandate The Opportunity: Ever...  ...with distributed data/computing tools including Spark, Databricks, Hadoop, Hive, AWS EMR, or Kafka Experience with data visualization or geospatial... 
    Amazon Web Service

    AllianceIT Inc

    Washington DC
    4 days ago
  • Experienced data engineers to join a large-scale Databricks platform build supporting federal government...  ...experience Advanced Python, SQL, and Spark/PySpark with some exposure to Databricks...  ...At least one Associate-level or higher AWS, Azure, or GCP certification Able to... 
    Amazon Web Service
    Full time
    Remote work
    Relocation

    Bizoforce - The Innovation Platform Accelerating Digital Sol...

    Washington DC
    2 days ago
  •  ...About this Role Imagineeer is seeking a Data Engineer to support the design, development, and maintenance...  ...with IL4/IL5 cloud environments (AWS GovCloud, Azure Government, or equivalent...  ...Experience building pipelines using Spark or distributed processing frameworks Experience... 
    Amazon Web Service
    Local area
    Work from home
    Flexible hours

    IMAGINEEER LLC

    Arlington, VA
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Data Engineer (AWS, Spark). Be the first to apply!