Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Data Engineer (AWS, Spark)

OpenDataJobs

Job Description

Job Description

Peregrine Advisors is a firm founded on a simple conviction: the best solutions come from the people closest to the problem, given real ownership and the tools to deliver. We are a data and technology innovation hub and a Benefit Corporation working at the center of the federal government's mission to deliver for client stakeholders and the US public, looking for highly motivated contributors who thrive when trusted to own a hard problem and equipped to deliver the solution.

Your first assignment

Your first assignment is to build the pipelines that move a federal agency's data from source to platform on Amazon Web Services (AWS): ingest, process, store, and keep it clean and trustworthy at scale. The work is real, hard, and it matters. It is also where you start, not the shape of your career here: we hire people, not seats, and we move our best to where the hardest problems are.

What you'll build
  • Ingest-process-store pipelines on AWS: Spark-based extract, transform, and load (ETL) with Glue, Amazon EMR, Lambda, and Step Functions, in Python and PySpark.
  • A data lake that feeds the platform: S3 design (Parquet, partitioning, lifecycle) into Apache Iceberg tables, PostgreSQL on Amazon Aurora, and DynamoDB, with Trino for federated SQL across them, plus event orchestration, secrets and monitoring, and data quality, validation, and lineage built in.
  • In time, the firm itself: new capabilities, tools, and lines of business you help spin up.
Who you are

You are a data engineer who is serious about the craft and eager to go deeper. You care as much about whether the data is trustworthy as whether it arrives, and you do your best work alongside people who push you. You experiment, fail, learn, and repeat quickly. You would rather own an outcome than be handed a task.

What you bring
  • Hands-on data engineering on AWS: Spark ETL (Glue, EMR), Python and PySpark, and S3 data-lake design feeding the platform stores. — and if a second lake bullet sits under "What you bring", it becomes: The reliability craft around it: event orchestration, data quality and lineage, monitoring, and infrastructure-as-code.
  • The reliability craft around it: event orchestration, data quality and lineage, monitoring, and infrastructure-as-code.
  • The judgment to build in a regulated environment where accuracy and auditability are not optional.
What you'll need

Sole United States citizenship and the ability to obtain a Public Trust determination are required for this initial engagement. This is a hybrid role based in the Washington, DC metropolitan area, and it requires commuting into DC regularly. Everything else, the years, the certifications, the specific tools, we ask in the application and get into during the interview

What we offer

A high-performing team of developers, engineers, data scientists, architects, and strategists solving complex, real-world problems, with work that runs from strategy formulation to hands-on implementation. We develop people across roles and clients, with extensive onboarding and sponsored training and professional development. And Peregrine has been a Benefit Corporation from day one: public value is built into the work itself, not bolted on afterward. Work worth your best years.

What we commit to

As a Benefit Corporation, our commitment runs three ways: real, measurable value for our clients; government that works better for the public; and a team that makes everyone in it better.

We hire people who want to help build the firm, not just work at it. If that is you, apply.

Peregrine Advisors is an equal opportunity employer.

Peregrine exclusively works with OPEN Data Jobs to recruit our team. Register with OPEN Data Jobs to be considered for this opening and future roles. 

Requirements

  • 4+ years of data engineering experience
  • Bachelor's degree
  • Spark ETL on AWS (Glue, Amazon EMR) in Python and PySpark
  • S3 data-lake design (Parquet, partitioning, lifecycle) feeding Apache Iceberg tables, Amazon Aurora PostgreSQL, and DynamoDB
  • Event orchestration (Lambda, Step Functions, SQS/SNS) with secrets management and monitoring
  • Data quality, validation, and lineage
  • Infrastructure-as-code (CloudFormation or Terraform)
  • Basic proficiency in writing, PowerPoint, and Excel

Preferred

  • Master's degree in a relevant field
  • Trino or comparable federated SQL across the lake and relational stores
  • Apache Ranger-governed access
  • Legacy ETL migration (for example DataStage)
  • Federal information technology or high-volume data experience
  • Familiarity with AI-assisted developer tooling

Benefits

Medical, dental, and vision with the employee premium fully paid and half of dependent premiums; employer-paid life, accidental death, and short-term and long-term disability insurance; a 401(k) matched 100% up to 4% of salary, vesting immediately; unlimited paid time off; and sponsored professional certifications and continuing education.

Vacancy posted 9 days ago
Similar jobs that could be interesting for youBased on the Data Engineer (AWS, Spark) in Washington DC vacancy
  •  ...service consulting company in Washington, DC, is seeking a Data Engineer to design, develop, and maintain scalable data...  ...engineering, strong Python/SQL skills, and experience with Spark, Databricks, and CI/CD. AWS/Azure/GCP certifications are required. #J-18808-Ljbffr... 
    Amazon Web Service

    Galent

    Washington DC
    3 days ago
  •  ...the tools to deliver. We are a data and technology innovation hub...  ...platform on Amazon Web Services (AWS): ingest, process, store, and...  ...-store pipelines on AWS: Spark-based extract, transform, and...  ...Who you are You are a data engineer who is serious about the craft... 
    Amazon Web Service
    Temporary work
    Work at office
    Immediate start

    OpenDataJobs

    Washington DC
    6 days ago
  •  ...Opportunity SDA is seeking a Senior Big Data Engineer / Java Data Integration Engineer to support...  ...using technologies such as Kafka, Spark, Elasticsearch, and cloud-based platforms...  ...Integration | Data Mapping | REST APIs | MongoDB | AWS | Linux What You’ll Do Design, develop,... 
    Amazon Web Service
    Full time
    Remote work
    Flexible hours

    SYSTEMS DEVELOPMENT AND ANALYSIS, INC.

    Arlington, VA
    4 days ago
  • Edgesource Corp. in McLean, VA is seeking a Data Engineer to build and optimize complex data pipelines for a high-scale environment. You will work with Python, Spark, and AWS services, focusing on data security, governance, and compliant data handling. The ideal candidate... 
    Amazon Web Service
    Full time

    Edgesource

    Mc Lean, VA
    11 hours ago
  • $158.6k - $181k

    Senior Data Engineer (Python, AWS, Spark) Do you love building and pioneering in the technology space? Do you enjoy solving complex business problems in a fast-paced, collaborative, inclusive, and iterative delivery environment? At Capital One, you'll be part of a big group... 
    Amazon Web Service
    Full time
    Part time
    Internship
    H1b
    Local area

    Capital One National Association

    Mc Lean, VA
    3 days ago
  • $229.9k - $262.4k

    Senior Lead Data Engineer (Python, AWS, Spark, Kakfa, SQL, Snowflake, DynamoDB, Databricks, GenAI) Do you love building and pioneering in the technology space? Do you enjoy solving complex business problems in a fast-paced, collaborative,inclusive, and iterative delivery... 
    Amazon Web Service
    Full time
    Part time
    Internship
    Local area

    Capital One

    Mc Lean, VA
    1 day ago
  • NIRA is seeking a highly experienced Senior Data Engineer to support the U.S. Department of Justice...  ...federal datasets.Experience with AWS GovCloud, Azure Government, FedRAMP-authorized...  ...Experience with Python-based data frameworks, Spark, Databricks, Kafka, Airflow, cloud-native... 
    Amazon Web Service

    NIRA

    Washington DC
    3 days ago
  • Analytica is seeking a Databricks Data Engineer to support a high-profile program for a financial...  ...and platform services within the AWS and Databricks ecosystem. This role works...  ...concepts.Experience with Databricks, Apache Spark, PySpark, or similar big data technologies... 
    Amazon Web Service

    ANALYTICA

    Washington DC
    4 days ago
  • Analytica is seeking a Senior Data Engineer to support large-scale Health care data modernization...  ...Platform. Strong proficiency with Apache Spark, PySpark, SQL, and Python. Experience...  ...experience with cloud data platforms in AWS or Azure cloud environments. Experience implementing... 
    Amazon Web Service

    ANALYTICA

    Washington DC
    4 days ago
  • $152k - $205.6k

     ...combines the talents of science, engineering, and UX to build and deliver...  ...Amazonians worldwide.As a Data Engineer on PXTCS, you'll work...  ...scalable data pipelines using native AWS services (Glue, EMR, Lambda);...  ...such as: Hadoop, Hive, Spark, EMRAmazon is an equal opportunity... 
    Amazon Web Service
    Local area
    Worldwide
    Flexible hours

    Amazon

    Arlington, VA
    4 days ago
  • $62k - $141k

    Data EngineerThe Opportunity: Ever-expanding technology like IoT,...  ...today than ever before. As a data engineer, you know that organizing data...  ...a service (SaaS), including AWS EMR, Redshift, SageMaker,...  ...frameworks, including Apache Spark or NVIDIA CUDAExperience with... 
    Amazon Web Service
    Full time
    Contract work
    Part time
    Work at office
    Local area
    Remote work

    Booz Allen Hamilton

    Arlington, VA
    2 days ago
  • $115k - $150k

     ...stronger results for our clients.As a Senior Data Engineer, you will architect, build, and optimize...  ...data workflows using tools like Apache Spark, Kafka, Airflow, and dbt.Monitor,...  ...automation.Deep understanding of cloud platforms (AWS, GCP, Azure) and cloud-native data... 
    Amazon Web Service
    Full time
    Visa sponsorship

    Insomniac Design

    Washington DC
    4 days ago
  • $116.2k - $229.1k

     ...Summary Join Deloitte’s AI & Engineering practice and help...  ...technology platforms, modernize data environments, and unlock value...  ...with tools such as Azure DevOps, AWS Code Pipeline, Jenkins, TFS, or...  ...Lakehouse architecture, Apache Spark, Delta Lake, cloud-native databases... 
    Amazon Web Service
    Local area
    Visa sponsorship

    Deloitte

    Rosslyn, VA
    5 days ago
  •  ...Capital Technology Group is seeking a Junior Data Engineer to design, build, and maintain scalable data pipelines and systems...  ...clients. At CTG, you will work with Databricks, dbt, Spark, SQL, Python, Java, and AWS to build robust data platforms, supporting mission-critical... 
    Amazon Web Service

    Capital Technology Group, Inc.

    Silver Spring, MD
    1 day ago
  •  ...Agile Defense is seeking a Data Engineer to design, develop, and deploy AI-enabled data solutions within DoD programs...  ...strong Python/SQL skills, experience with Spark/Databricks, and familiarity with cloud platforms (AWS/Azure). A Bachelor's degree (Masters preferred) and... 
    Amazon Web Service

    Agile Defense

    Washington DC
    1 day ago
  •  ...Big Data EngineerLocation: Arlington, VA (3 Days onsite/week)Ability...  ...in using Python or Scala, Spark, Hadoop platforms & tools (Hive...  ...like data ingestion, feature engineering, modeling, tuning, evaluating,...  ...presentingCloud knowledge (Databricks or AWS ecosystem) is a plus but not... 
    Amazon Web Service
    3 days per week

    Futran Tech Solutions Pvt. Ltd.

    Arlington, VA
    1 day ago
  •  ...Data Pipeline EngineerWe are seeking a Data Pipeline Engineer to design, build, and maintain scalable data pipelines in support...  ...solutions using Python, PySpark, Spark, and/or PolarsDebug and troubleshoot...  ...with cloud platforms (AWS or similar)Experience with data visualization... 
    Amazon Web Service

    ClearanceJobs

    Washington DC
    4 days ago
  •  ...Data Engineer -- 100% RemoteRate: $54/Hr on C2CEST or CST time zoneInterview Process: 1-2 roundsCandidate...  ..., just someone with a decent amount of AWS experience.• Eastern or Central time zone...  ....• Biggest tools: Glue, Python, PI Spark.• PI Spark is adjacent to Python, but... 
    Amazon Web Service
    For contractors
    Remote work

    Anveta

    Washington DC
    1 day ago
  •  ...Senior Data Engineer with 3-5 years of experience // Secret with SCI or TOP secret with SCI / Mandate The Opportunity: Ever...  ...with distributed data/computing tools including Spark, Databricks, Hadoop, Hive, AWS EMR, or Kafka Experience with data visualization or geospatial... 
    Amazon Web Service

    Alliance IT

    Washington DC
    1 day ago
  •  ...We have an immediate opening for a Data Engineer with a leading IT Service Consulting company in...  ...experience Strong Python and SQL Hands-on Apache Spark / PySpark with exposure to Databricks....  ...At least one Associate-level or higher AWS, Azure, or GCP certification Nice-to-... 
    Amazon Web Service
    Full time
    Immediate start

    Galent

    Washington DC
    1 day ago
  •  ...Senior Data EngineerApex Systems is currently seeking a Senior Data Engineer who enjoys data and building data storage platforms from...  ...and execute ETLs using Apache Spark on Hadoop among other Data technologiesDetermine...  ...using tools like in the AWS ecosystem.Create and automate... 
    Amazon Web Service
    Work experience placement

    Software Technology Inc

    Washington DC
    1 day ago
  •  ...About this Role Imagineeer is seeking a Data Engineer to support the design, development, and maintenance...  ...with IL4/IL5 cloud environments (AWS GovCloud, Azure Government, or equivalent...  ...Experience building pipelines using Spark or distributed processing frameworks Experience... 
    Amazon Web Service
    Local area
    Work from home
    Flexible hours

    IMAGINEEER LLC

    Arlington, VA
    3 days ago
  •  ...Job Summary The Data Engineer / ETL Developer is responsible for designing, developing, and maintaining...  ...platforms. Experience with Python, Spark, or other data processing technologies....  ...Skills & Certifications Experience with AWS, Azure, or DoD cloud environments. Experience... 
    Amazon Web Service
    Shift work

    vTech Solution

    Arlington, VA
    1 day ago
  •  ...Data Engineer OpportunityMakpar is a comprehensive professional and technical solutions provider...  ...engineering.Experience with Databricks, Apache Spark, and SQL.Proficiency in Python and SQL....  ...ETL tools such as Airflow, Kafka, or AWS Glue.Familiarity with GitHub/GitLab and CI... 
    Amazon Web Service
    Start working today
    Local area
    Flexible hours

    Makpar Corporation

    Washington DC
    2 days ago
  • $145k - $165k

     ...Data Engineer - CDAOAt Agile Defense we know that action defines the outcome and new challenges...  ...platforms such as Databricks, Palantir, or AWS-native data services is highly preferred....  ..., and distributed data frameworks (e.g., Spark, Databricks, PySpark).Experience... 
    Amazon Web Service
    Temporary work
    Work experience placement

    Agile Defense

    Washington DC
    5 days ago
  •  ...AWS Data Engineer The AWS Data Engineer designs and develops scalable data solutions using data integration tools and technologies. Being...  ...engineering and integration tools such as Python, Informatica IICS, Spark, Dynamo DB etc. in the AWS environment. The data engineer... 
    Amazon Web Service
    Contract work

    Software Technology Inc

    Washington DC
    5 days ago
  •  ...Senior Data Engineer BizTech Fusion is hiring for this position for one of its clients. Location...  ...strong hands‑on experience in Python, Apache Spark/PySpark, SQL, Databricks, cloud data...  ...Hands‑on Databricks experience on Azure, AWS, or GCP. Experience with cloud data services... 
    Amazon Web Service
    Remote work
    Relocation
    2 days per week

    BizTech Fusion LLC

    Washington DC
    4 days ago
  •  ...DMV area Bachelor Degree The Gist... The Data Engineer is responsible for designing, developing,...  ...architectures. Experience utilizing Python, SQL, Spark/Scala, Databricks, and distributed data...  ...technologies. Experience supporting AWS cloud environments utilizing S3, Kinesis,... 
    Amazon Web Service
    Local area
    Flexible hours

    Amivero

    District Heights, MD
    3 days ago
  • $119.7k - $199.3k

     ...seeking a skilled and experienced Senior Data Engineer for our Data and AI team to contribute to...  ...Lead data engineering efforts using Python, Spark, Flink, and other data processing...  ...ETL development using Python and PySpark; AWS Data Lakehouse technologies (Redshift, Athena... 
    Amazon Web Service

    Nashville Public Radio

    Washington DC
    3 days ago
  •  ...Data EngineerHatch I.T. is partnering with Expression to find a Data Engineer. See details below:About The Role:Expression is seeking an...  ...modeling approaches using Python, Spark, and cloud-native ML...  ...based environments, including AWS, Azure, or Palantir Foundry.Strong... 
    Amazon Web Service
    For contractors
    Work at office
    Immediate start

    Hatchit Co

    Washington DC
    5 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Data Engineer (AWS, Spark). Be the first to apply!