Data Engineer (AWS, Spark)
OpenDataJobs
Job Description
Job Description
Peregrine Advisors is a firm founded on a simple conviction: the best solutions come from the people closest to the problem, given real ownership and the tools to deliver. We are a data and technology innovation hub and a Benefit Corporation working at the center of the federal government's mission to deliver for client stakeholders and the US public, looking for highly motivated contributors who thrive when trusted to own a hard problem and equipped to deliver the solution.
Your first assignmentYour first assignment is to build the pipelines that move a federal agency's data from source to platform on Amazon Web Services (AWS): ingest, process, store, and keep it clean and trustworthy at scale. The work is real, hard, and it matters. It is also where you start, not the shape of your career here: we hire people, not seats, and we move our best to where the hardest problems are.
What you'll build- Ingest-process-store pipelines on AWS: Spark-based extract, transform, and load (ETL) with Glue, Amazon EMR, Lambda, and Step Functions, in Python and PySpark.
- A data lake that feeds the platform: S3 design (Parquet, partitioning, lifecycle) into Apache Iceberg tables, PostgreSQL on Amazon Aurora, and DynamoDB, with Trino for federated SQL across them, plus event orchestration, secrets and monitoring, and data quality, validation, and lineage built in.
- In time, the firm itself: new capabilities, tools, and lines of business you help spin up.
You are a data engineer who is serious about the craft and eager to go deeper. You care as much about whether the data is trustworthy as whether it arrives, and you do your best work alongside people who push you. You experiment, fail, learn, and repeat quickly. You would rather own an outcome than be handed a task.
What you bring- Hands-on data engineering on AWS: Spark ETL (Glue, EMR), Python and PySpark, and S3 data-lake design feeding the platform stores. — and if a second lake bullet sits under "What you bring", it becomes: The reliability craft around it: event orchestration, data quality and lineage, monitoring, and infrastructure-as-code.
- The reliability craft around it: event orchestration, data quality and lineage, monitoring, and infrastructure-as-code.
- The judgment to build in a regulated environment where accuracy and auditability are not optional.
Sole United States citizenship and the ability to obtain a Public Trust determination are required for this initial engagement. This is a hybrid role based in the Washington, DC metropolitan area, and it requires commuting into DC regularly. Everything else, the years, the certifications, the specific tools, we ask in the application and get into during the interview
What we offerA high-performing team of developers, engineers, data scientists, architects, and strategists solving complex, real-world problems, with work that runs from strategy formulation to hands-on implementation. We develop people across roles and clients, with extensive onboarding and sponsored training and professional development. And Peregrine has been a Benefit Corporation from day one: public value is built into the work itself, not bolted on afterward. Work worth your best years.
What we commit toAs a Benefit Corporation, our commitment runs three ways: real, measurable value for our clients; government that works better for the public; and a team that makes everyone in it better.
We hire people who want to help build the firm, not just work at it. If that is you, apply.
Peregrine Advisors is an equal opportunity employer.
Peregrine exclusively works with OPEN Data Jobs to recruit our team. Register with OPEN Data Jobs to be considered for this opening and future roles.
Requirements
- 4+ years of data engineering experience
- Bachelor's degree
- Spark ETL on AWS (Glue, Amazon EMR) in Python and PySpark
- S3 data-lake design (Parquet, partitioning, lifecycle) feeding Apache Iceberg tables, Amazon Aurora PostgreSQL, and DynamoDB
- Event orchestration (Lambda, Step Functions, SQS/SNS) with secrets management and monitoring
- Data quality, validation, and lineage
- Infrastructure-as-code (CloudFormation or Terraform)
- Basic proficiency in writing, PowerPoint, and Excel
Preferred
- Master's degree in a relevant field
- Trino or comparable federated SQL across the lake and relational stores
- Apache Ranger-governed access
- Legacy ETL migration (for example DataStage)
- Federal information technology or high-volume data experience
- Familiarity with AI-assisted developer tooling
Benefits
Medical, dental, and vision with the employee premium fully paid and half of dependent premiums; employer-paid life, accidental death, and short-term and long-term disability insurance; a 401(k) matched 100% up to 4% of salary, vesting immediately; unlimited paid time off; and sponsored professional certifications and continuing education.
- ...easily move between business, data management, and technical teams... ...in using Python or Scala, Spark, Hadoop platforms & tools (Hive... ...like data ingestion, feature engineering, modeling, tuning, evaluating,... ...Cloud knowledge (Databricks or AWS ecosystem) is a plus but not requiredAmazon Web Service
- ...Opportunity SDA is seeking a Senior Big Data Engineer / Java Data Integration Engineer to support... ...using technologies such as Kafka, Spark, Elasticsearch, and cloud-based platforms... ...Integration | Data Mapping | REST APIs | MongoDB | AWS | Linux What You’ll Do Design, develop,...Amazon Web ServiceFull timeRemote workFlexible hours
- Yoh Services LLC in Maryland is seeking a Senior Associate L2 - Full Stack Data Platform Support Engineer to maintain enterprise data platforms and cloud-native applications on AWS. You will ensure availability, stability, and performance of batch and real-time processing...Amazon Web Service
- Edgesource Corp. in McLean, VA is seeking a Data Engineer to build and optimize complex data pipelines for a high-scale environment. You will work with Python, Spark, and AWS services, focusing on data security, governance, and compliant data handling. The ideal candidate...Amazon Web ServiceFull time
$93.4k - $176.2k
...Services seeks a skilled Pipeline Architect for their Washington D.C. office to design, build, and maintain data pipelines using technologies like Databricks and AWS. Candidates must have experience with cloud-based ETL services, data warehousing, and programming in...Amazon Web ServiceWork at office$158.6k - $181k
Senior Data Engineer (Python, AWS, Spark) Do you love building and pioneering in the technology space? Do you enjoy solving complex business problems in a fast-paced, collaborative, inclusive, and iterative delivery environment? At Capital One, you'll be part of a big group...Amazon Web ServiceFull timePart timeInternshipH1bLocal area$229.9k - $262.4k
Senior Lead Data Engineer (Python, AWS, Spark, Kakfa, SQL, Snowflake, DynamoDB, Databricks, GenAI) Do you love building and pioneering in the technology space? Do you enjoy solving complex business problems in a fast-paced, collaborative,inclusive, and iterative delivery...Amazon Web ServiceFull timePart timeInternshipLocal area- NIRA is seeking a highly experienced Senior Data Engineer to support the U.S. Department of Justice... ...federal datasets.Experience with AWS GovCloud, Azure Government, FedRAMP-authorized... ...Experience with Python-based data frameworks, Spark, Databricks, Kafka, Airflow, cloud-native...Amazon Web Service
- Analytica is seeking a Databricks Data Engineer to support a high-profile program for a financial... ...and platform services within the AWS and Databricks ecosystem. This role works... ...concepts.Experience with Databricks, Apache Spark, PySpark, or similar big data technologies...Amazon Web Service
- Analytica is seeking a Senior Data Engineer to support large-scale Health care data modernization... ...Platform. Strong proficiency with Apache Spark, PySpark, SQL, and Python. Experience... ...experience with cloud data platforms in AWS or Azure cloud environments. Experience implementing...Amazon Web Service
$115k - $150k
...stronger results for our clients.As a Senior Data Engineer, you will architect, build, and optimize... ...data workflows using tools like Apache Spark, Kafka, Airflow, and dbt.Monitor,... ...automation.Deep understanding of cloud platforms (AWS, GCP, Azure) and cloud-native data...Amazon Web ServiceFull timeVisa sponsorship$152k - $205.6k
...combines the talents of science, engineering, and UX to build and deliver... ...Amazonians worldwide.As a Data Engineer on PXTCS, you'll work... ...scalable data pipelines using native AWS services (Glue, EMR, Lambda);... ...such as: Hadoop, Hive, Spark, EMRAmazon is an equal opportunity...Amazon Web ServiceLocal areaWorldwideFlexible hours$62k - $141k
Data EngineerThe Opportunity: Ever-expanding technology like IoT,... ...today than ever before. As a data engineer, you know that organizing data... ...a service (SaaS), including AWS EMR, Redshift, SageMaker,... ...frameworks, including Apache Spark or NVIDIA CUDAExperience with...Amazon Web ServiceFull timeContract workPart timeWork at officeLocal areaRemote work$116.2k - $229.1k
...Summary Join Deloitte’s AI & Engineering practice and help... ...technology platforms, modernize data environments, and unlock value... ...with tools such as Azure DevOps, AWS Code Pipeline, Jenkins, TFS, or... ...Lakehouse architecture, Apache Spark, Delta Lake, cloud-native databases...Amazon Web ServiceLocal areaVisa sponsorship$119.7k - $199.3k
...Communications, Customer Care, Engineering & Product, Finance, Human... ...skilled and experienced Senior Data Engineer for our Data and AI team... ...engineering efforts using Python, Spark, Flink and other data... ...and PySpark. Experience with AWS Data Lakehouse technologies including...Amazon Web ServiceFull timePart time- ...Job Title8+ Years experience working in Data Analysis.Hands on Experience in Hadoop Stack of Technologies (Hadoop, Spark, HBase, Hive, Pig, Sqoop, Scala, Flume, HDFS, Map Reduce... ..., Data Processing & Data Analytics on AWS is good to have.Experience with data modeling...Amazon Web ServiceWork experience placement
- ...Data Engineer - EtlWe are seeking a data engineer to design, build, and scale high quality data... ...with cloud based data platforms (azure, aws, or gcp) and modern data architectures (data... ...data processing frameworks (e.g., spark or equivalent).Experience building and operating...Amazon Web Service
$110k - $115k
...Overview Join to apply for the Data Engineer role at Allnessjobs . This range is provided by Allnessjobs... ...ingestion and cleaning using Python and Spark. Support Requirements gathering... ...Previous experience working with Cloud (e.g. AWS) is a plus. Experience with Databricks is...Amazon Web ServiceFull timeZero hours contractWork at office- ...Data EngineerLocation: On-Site, Washington, DCInterview Type: Phone... ...: 12+ monthsThe Data Engineer designs, builds, and operates... ...) using Databricks and Apache Spark. This role is hands-on in Python... ...(SCD) pattern implementation.AWS data services experience including...Amazon Web ServiceFor contractors
- ...Data EngineerImagineeer is seeking a Data Engineer to support the design, development, and maintenance of secure, compliant... ...with IL4/IL5 cloud environments (AWS GovCloud, Azure Government, or... ...PythonExperience building pipelines using Spark or distributed processing...Amazon Web ServiceLocal areaWork from homeFlexible hours
- ...Data EngineerWe are seeking a Data Engineer to support our Federal Government Customer with delivering secure, cloud... ...stack applications, microservices, and AWS infrastructure, setting... ...distributed data processing pipelines using Spark and PySpark, defining partitioning,...Amazon Web Service
- ...DMV area, and a bachelor degree. The Data Engineer The Data Engineer is responsible for designing... .... Experience utilizing Python, SQL, Spark/Scala, Databricks, and distributed data... ...technologies. Experience supporting AWS cloud environments utilizing S3, Kinesis,...Amazon Web ServiceLocal areaFlexible hours
$112k - $179k
...Peraton is seeking a Data Engineer to support our classified customer in Washington, DC. Peraton... ...availability. Experience with Python, SQL, and/or Spark. Experience working with both structured,... ...experience Have knowledge of Splunk AWS experience Neo4J knowledge or any graph...Amazon Web Service- ...Data Engineer Position Summary Join our team as a Data Engineer and become a key contributor to... ...with Azure Synapse Analytics (Notebooks/SQL/Spark) and ADLS Gen2. Production experience... ...Qualifications Experience integrating with AWS data services (S3, Glue/Glue Data Catalog...Amazon Web Service
$93.4k - $176.2k
...Data EngineerAt Accenture Federal Services, nothing matters more... ...data pipelines using Databricks, Spark, and related technologies.... ...ecosystem.Cloud Commander: Leverage AWS services (e.g., S3, Glue,... ...with COTS and open-source data engineering tools such as ElasticSearch and...Amazon Web ServiceLocal area$186k - $221.91k
..., D.C. K.ATS Foundry is Keystone’s engineering center of excellence, embedding data, platform, and forensic expertise into... ...across cloud environments (AWS, GCP, Azure, Snowflake). Develop reproducible... ...orchestration tools (Airflow, dbt, Spark, or equivalent). Design data models...Amazon Web Service- ...Data EngineerAs part of the Data Management area, the Data Engineer will provide leadership in the conceptualization and design of innovative... ...connect to SaaS solutions), Hadoop, Spark, Kafka, Linux scripting, etc. Experience with AWS/Azure cloud services and platforms,...Amazon Web Service
- ...Data EngineerExpression is seeking an experienced Data Engineer to support the design, development, and operational deployment... ...modeling approaches using Python, Spark, and cloud-native ML frameworks... ...cloud-based environments, including AWS, Azure, or Palantir Foundry....Amazon Web ServiceFor contractorsWork at officeImmediate start
- ...We apply our domain expertise, data acumen, and technology know-... ...us at We are seeking a Data Engineer to join our team in Washington... ...using Apache Kafka and Apache Spark to support STU reporting cycles... ...data platform infrastructure on AWS, GCP, and/or Azure, including...Amazon Web ServiceWork experience placementWork at officeFlexible hours
$117.9k - $169.6k
...Data EngineerSuitland, MD At Accenture Federal Services, nothing... ...a skilled and passionate Data Engineer to join our team. You will play... ...preference for experience in AWS. Experience in GCP or Azure is... ...Leverage frameworks like Apache Spark, Databricks, or Apache Kafka...Amazon Web ServiceLive inWork at officeLocal area
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Data Engineer (AWS, Spark). Be the first to apply!
- data center engineer Washington DC
- senior cloud data engineer Washington DC
- data infrastructure engineer Washington DC
- sr information security engineer Washington DC
- hadoop big data developer Washington DC
- data visualization developer Washington DC
- remote data engineer Washington DC
- etl data engineer Washington DC
- data engineer analytics Washington DC
- sr data engineer Washington DC


