Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Data Engineer (AWS, Spark)

OpenDataJobs

Job Description

Job Description

Peregrine Advisors is a firm founded on a simple conviction: the best solutions come from the people closest to the problem, given real ownership and the tools to deliver. We are a data and technology innovation hub and a Benefit Corporation working at the center of the federal government's mission to deliver for client stakeholders and the US public, looking for highly motivated contributors who thrive when trusted to own a hard problem and equipped to deliver the solution.

Your first assignment

Your first assignment is to build the pipelines that move a federal agency's data from source to platform on Amazon Web Services (AWS): ingest, process, store, and keep it clean and trustworthy at scale. The work is real, hard, and it matters. It is also where you start, not the shape of your career here: we hire people, not seats, and we move our best to where the hardest problems are.

What you'll build
  • Ingest-process-store pipelines on AWS: Spark-based extract, transform, and load (ETL) with Glue, Amazon EMR, Lambda, and Step Functions, in Python and PySpark.
  • A data lake that feeds the platform: S3 design (Parquet, partitioning, lifecycle) into Apache Iceberg tables, PostgreSQL on Amazon Aurora, and DynamoDB, with Trino for federated SQL across them, plus event orchestration, secrets and monitoring, and data quality, validation, and lineage built in.
  • In time, the firm itself: new capabilities, tools, and lines of business you help spin up.
Who you are

You are a data engineer who is serious about the craft and eager to go deeper. You care as much about whether the data is trustworthy as whether it arrives, and you do your best work alongside people who push you. You experiment, fail, learn, and repeat quickly. You would rather own an outcome than be handed a task.

What you bring
  • Hands-on data engineering on AWS: Spark ETL (Glue, EMR), Python and PySpark, and S3 data-lake design feeding the platform stores. — and if a second lake bullet sits under "What you bring", it becomes: The reliability craft around it: event orchestration, data quality and lineage, monitoring, and infrastructure-as-code.
  • The reliability craft around it: event orchestration, data quality and lineage, monitoring, and infrastructure-as-code.
  • The judgment to build in a regulated environment where accuracy and auditability are not optional.
What you'll need

Sole United States citizenship and the ability to obtain a Public Trust determination are required for this initial engagement. This is a hybrid role based in the Washington, DC metropolitan area, and it requires commuting into DC regularly. Everything else, the years, the certifications, the specific tools, we ask in the application and get into during the interview

What we offer

A high-performing team of developers, engineers, data scientists, architects, and strategists solving complex, real-world problems, with work that runs from strategy formulation to hands-on implementation. We develop people across roles and clients, with extensive onboarding and sponsored training and professional development. And Peregrine has been a Benefit Corporation from day one: public value is built into the work itself, not bolted on afterward. Work worth your best years.

What we commit to

As a Benefit Corporation, our commitment runs three ways: real, measurable value for our clients; government that works better for the public; and a team that makes everyone in it better.

We hire people who want to help build the firm, not just work at it. If that is you, apply.

Peregrine Advisors is an equal opportunity employer.

Peregrine exclusively works with OPEN Data Jobs to recruit our team. Register with OPEN Data Jobs to be considered for this opening and future roles. 

Requirements

  • 4+ years of data engineering experience
  • Bachelor's degree
  • Spark ETL on AWS (Glue, Amazon EMR) in Python and PySpark
  • S3 data-lake design (Parquet, partitioning, lifecycle) feeding Apache Iceberg tables, Amazon Aurora PostgreSQL, and DynamoDB
  • Event orchestration (Lambda, Step Functions, SQS/SNS) with secrets management and monitoring
  • Data quality, validation, and lineage
  • Infrastructure-as-code (CloudFormation or Terraform)
  • Basic proficiency in writing, PowerPoint, and Excel

Preferred

  • Master's degree in a relevant field
  • Trino or comparable federated SQL across the lake and relational stores
  • Apache Ranger-governed access
  • Legacy ETL migration (for example DataStage)
  • Federal information technology or high-volume data experience
  • Familiarity with AI-assisted developer tooling

Benefits

Medical, dental, and vision with the employee premium fully paid and half of dependent premiums; employer-paid life, accidental death, and short-term and long-term disability insurance; a 401(k) matched 100% up to 4% of salary, vesting immediately; unlimited paid time off; and sponsored professional certifications and continuing education.

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Data Engineer (AWS, Spark) in Washington DC vacancy
  •  ...easily move between business, data management, and technical teams...  ...in using Python or Scala, Spark, Hadoop platforms & tools (Hive...  ...like data ingestion, feature engineering, modeling, tuning, evaluating,...  ...Cloud knowledge (Databricks or AWS ecosystem) is a plus but not required
    Amazon Web Service

    Programmers.io

    Arlington, VA
    2 days ago
  •  ...Opportunity SDA is seeking a Senior Big Data Engineer / Java Data Integration Engineer to support...  ...using technologies such as Kafka, Spark, Elasticsearch, and cloud-based platforms...  ...Integration | Data Mapping | REST APIs | MongoDB | AWS | Linux What You’ll Do Design, develop,... 
    Amazon Web Service
    Full time
    Remote work
    Flexible hours

    SYSTEMS DEVELOPMENT AND ANALYSIS, INC.

    Arlington, VA
    5 days ago
  • Yoh Services LLC in Maryland is seeking a Senior Associate L2 - Full Stack Data Platform Support Engineer to maintain enterprise data platforms and cloud-native applications on AWS. You will ensure availability, stability, and performance of batch and real-time processing... 
    Amazon Web Service

    Yoh

    Bethesda, MD
    3 days ago
  • Edgesource Corp. in McLean, VA is seeking a Data Engineer to build and optimize complex data pipelines for a high-scale environment. You will work with Python, Spark, and AWS services, focusing on data security, governance, and compliant data handling. The ideal candidate... 
    Amazon Web Service
    Full time

    Edgesource

    Mc Lean, VA
    1 day ago
  • $93.4k - $176.2k

     ...Services seeks a skilled Pipeline Architect for their Washington D.C. office to design, build, and maintain data pipelines using technologies like Databricks and AWS. Candidates must have experience with cloud-based ETL services, data warehousing, and programming in... 
    Amazon Web Service
    Work at office

    Accenture Federal Services

    Washington DC
    2 days ago
  • $158.6k - $181k

    Senior Data Engineer (Python, AWS, Spark) Do you love building and pioneering in the technology space? Do you enjoy solving complex business problems in a fast-paced, collaborative, inclusive, and iterative delivery environment? At Capital One, you'll be part of a big group... 
    Amazon Web Service
    Full time
    Part time
    Internship
    H1b
    Local area

    Capital One National Association

    Mc Lean, VA
    4 days ago
  • $229.9k - $262.4k

    Senior Lead Data Engineer (Python, AWS, Spark, Kakfa, SQL, Snowflake, DynamoDB, Databricks, GenAI) Do you love building and pioneering in the technology space? Do you enjoy solving complex business problems in a fast-paced, collaborative,inclusive, and iterative delivery... 
    Amazon Web Service
    Full time
    Part time
    Internship
    Local area

    Capital One

    Mc Lean, VA
    2 days ago
  • NIRA is seeking a highly experienced Senior Data Engineer to support the U.S. Department of Justice...  ...federal datasets.Experience with AWS GovCloud, Azure Government, FedRAMP-authorized...  ...Experience with Python-based data frameworks, Spark, Databricks, Kafka, Airflow, cloud-native... 
    Amazon Web Service

    NIRA

    Washington DC
    4 days ago
  • Analytica is seeking a Databricks Data Engineer to support a high-profile program for a financial...  ...and platform services within the AWS and Databricks ecosystem. This role works...  ...concepts.Experience with Databricks, Apache Spark, PySpark, or similar big data technologies... 
    Amazon Web Service

    ANALYTICA

    Washington DC
    4 hours ago
  • Analytica is seeking a Senior Data Engineer to support large-scale Health care data modernization...  ...Platform. Strong proficiency with Apache Spark, PySpark, SQL, and Python. Experience...  ...experience with cloud data platforms in AWS or Azure cloud environments. Experience implementing... 
    Amazon Web Service

    ANALYTICA

    Washington DC
    4 hours ago
  • $115k - $150k

     ...stronger results for our clients.As a Senior Data Engineer, you will architect, build, and optimize...  ...data workflows using tools like Apache Spark, Kafka, Airflow, and dbt.Monitor,...  ...automation.Deep understanding of cloud platforms (AWS, GCP, Azure) and cloud-native data... 
    Amazon Web Service
    Full time
    Visa sponsorship

    Insomniac Design

    Washington DC
    4 hours ago
  • $152k - $205.6k

     ...combines the talents of science, engineering, and UX to build and deliver...  ...Amazonians worldwide.As a Data Engineer on PXTCS, you'll work...  ...scalable data pipelines using native AWS services (Glue, EMR, Lambda);...  ...such as: Hadoop, Hive, Spark, EMRAmazon is an equal opportunity... 
    Amazon Web Service
    Local area
    Worldwide
    Flexible hours

    Amazon

    Arlington, VA
    36 minutes ago
  • $62k - $141k

    Data EngineerThe Opportunity: Ever-expanding technology like IoT,...  ...today than ever before. As a data engineer, you know that organizing data...  ...a service (SaaS), including AWS EMR, Redshift, SageMaker,...  ...frameworks, including Apache Spark or NVIDIA CUDAExperience with... 
    Amazon Web Service
    Full time
    Contract work
    Part time
    Work at office
    Local area
    Remote work

    Booz Allen Hamilton

    Arlington, VA
    3 days ago
  • $116.2k - $229.1k

     ...Summary Join Deloitte’s AI & Engineering practice and help...  ...technology platforms, modernize data environments, and unlock value...  ...with tools such as Azure DevOps, AWS Code Pipeline, Jenkins, TFS, or...  ...Lakehouse architecture, Apache Spark, Delta Lake, cloud-native databases... 
    Amazon Web Service
    Local area
    Visa sponsorship

    Deloitte

    Rosslyn, VA
    1 day ago
  • $119.7k - $199.3k

     ...Communications, Customer Care, Engineering & Product, Finance, Human...  ...skilled and experienced Senior Data Engineer for our Data and AI team...  ...engineering efforts using Python, Spark, Flink and other data...  ...and PySpark. Experience with AWS Data Lakehouse technologies including... 
    Amazon Web Service
    Full time
    Part time

    Washington Post

    Washington DC
    4 hours ago
  •  ...Job Title8+ Years experience working in Data Analysis.Hands on Experience in Hadoop Stack of Technologies (Hadoop, Spark, HBase, Hive, Pig, Sqoop, Scala, Flume, HDFS, Map Reduce...  ..., Data Processing & Data Analytics on AWS is good to have.Experience with data modeling... 
    Amazon Web Service
    Work experience placement

    Saxon Global

    Washington DC
    2 days ago
  •  ...Data Engineer - EtlWe are seeking a data engineer to design, build, and scale high quality data...  ...with cloud based data platforms (azure, aws, or gcp) and modern data architectures (data...  ...data processing frameworks (e.g., spark or equivalent).Experience building and operating... 
    Amazon Web Service

    Saxon Global

    Washington DC
    3 days ago
  • $110k - $115k

     ...Overview Join to apply for the Data Engineer role at Allnessjobs . This range is provided by Allnessjobs...  ...ingestion and cleaning using Python and Spark. Support Requirements gathering...  ...Previous experience working with Cloud (e.g. AWS) is a plus. Experience with Databricks is... 
    Amazon Web Service
    Full time
    Zero hours contract
    Work at office

    Allnessjobs

    Washington DC
    1 day ago
  •  ...Data EngineerLocation: On-Site, Washington, DCInterview Type: Phone...  ...: 12+ monthsThe Data Engineer designs, builds, and operates...  ...) using Databricks and Apache Spark. This role is hands-on in Python...  ...(SCD) pattern implementation.AWS data services experience including... 
    Amazon Web Service
    For contractors

    InterSources

    Washington DC
    2 days ago
  •  ...Data EngineerImagineeer is seeking a Data Engineer to support the design, development, and maintenance of secure, compliant...  ...with IL4/IL5 cloud environments (AWS GovCloud, Azure Government, or...  ...PythonExperience building pipelines using Spark or distributed processing... 
    Amazon Web Service
    Local area
    Work from home
    Flexible hours

    Imagineeer

    Arlington, VA
    2 days ago
  •  ...Data EngineerWe are seeking a Data Engineer to support our Federal Government Customer with delivering secure, cloud...  ...stack applications, microservices, and AWS infrastructure, setting...  ...distributed data processing pipelines using Spark and PySpark, defining partitioning,... 
    Amazon Web Service

    Ryde Technologies

    Arlington, VA
    2 days ago
  •  ...DMV area, and a bachelor degree. The Data Engineer The Data Engineer is responsible for designing...  .... Experience utilizing Python, SQL, Spark/Scala, Databricks, and distributed data...  ...technologies. Experience supporting AWS cloud environments utilizing S3, Kinesis,... 
    Amazon Web Service
    Local area
    Flexible hours

    Amivero

    Suitland, MD
    3 days ago
  • $112k - $179k

     ...Peraton is seeking a Data Engineer to support our classified customer in Washington, DC. Peraton...  ...availability. Experience with Python, SQL, and/or Spark. Experience working with both structured,...  ...experience Have knowledge of Splunk AWS experience Neo4J knowledge or any graph... 
    Amazon Web Service

    Peraton

    Washington DC
    4 days ago
  •  ...Data Engineer Position Summary Join our team as a Data Engineer and become a key contributor to...  ...with Azure Synapse Analytics (Notebooks/SQL/Spark) and ADLS Gen2. Production experience...  ...Qualifications Experience integrating with AWS data services (S3, Glue/Glue Data Catalog... 
    Amazon Web Service

    AEM

    Washington DC
    2 days ago
  • $93.4k - $176.2k

     ...Data EngineerAt Accenture Federal Services, nothing matters more...  ...data pipelines using Databricks, Spark, and related technologies....  ...ecosystem.Cloud Commander: Leverage AWS services (e.g., S3, Glue,...  ...with COTS and open-source data engineering tools such as ElasticSearch and... 
    Amazon Web Service
    Local area

    Accenture Federal Services

    Washington DC
    2 days ago
  • $186k - $221.91k

     ..., D.C. K.ATS Foundry is Keystone’s engineering center of excellence, embedding data, platform, and forensic expertise into...  ...across cloud environments (AWS, GCP, Azure, Snowflake). Develop reproducible...  ...orchestration tools (Airflow, dbt, Spark, or equivalent). Design data models... 
    Amazon Web Service

    Keystone

    Washington DC
    4 days ago
  •  ...Data EngineerAs part of the Data Management area, the Data Engineer will provide leadership in the conceptualization and design of innovative...  ...connect to SaaS solutions), Hadoop, Spark, Kafka, Linux scripting, etc. Experience with AWS/Azure cloud services and platforms,... 
    Amazon Web Service

    Software Technology Inc

    Washington DC
    2 days ago
  •  ...Data EngineerExpression is seeking an experienced Data Engineer to support the design, development, and operational deployment...  ...modeling approaches using Python, Spark, and cloud-native ML frameworks...  ...cloud-based environments, including AWS, Azure, or Palantir Foundry.... 
    Amazon Web Service
    For contractors
    Work at office
    Immediate start

    Expression Networks

    Arlington, VA
    1 day ago
  •  ...We apply our domain expertise, data acumen, and technology know-...  ...us at We are seeking a Data Engineer to join our team in Washington...  ...using Apache Kafka and Apache Spark to support STU reporting cycles...  ...data platform infrastructure on AWS, GCP, and/or Azure, including... 
    Amazon Web Service
    Work experience placement
    Work at office
    Flexible hours

    UNISSANT

    Washington DC
    5 days ago
  • $117.9k - $169.6k

     ...Data EngineerSuitland, MD At Accenture Federal Services, nothing...  ...a skilled and passionate Data Engineer to join our team. You will play...  ...preference for experience in AWS. Experience in GCP or Azure is...  ...Leverage frameworks like Apache Spark, Databricks, or Apache Kafka... 
    Amazon Web Service
    Live in
    Work at office
    Local area

    Accenture Federal Services

    Suitland, MD
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Data Engineer (AWS, Spark). Be the first to apply!