Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

AWS Lakehouse Data Engineer

Jobleads-US

Position: AWS Lakehouse Data Engineer

Clearance: Ability to Obtain Public Trust

Location: 100% Remote (prefer DMV)

We are seeking an AWS Lakehouse Data Engineer to design, implement, and operate the cloud-native data platform that powers AI/ML, analytics, reporting, and data visualization. You will build a modern lakehouse on Amazon S3 using AWS-native services and open table formats, providing Databricks-like capabilities while maintaining portability, strong governance, cost efficiency, and operational control. You will also develop scalable batch and streaming ingestion, Python and PySpark ETL/ELT pipelines, metadata and governance services, and automated cloud provisioning and CI/CD across environments.

This role is ideal for an engineer who enjoys platform building, automation, performance optimization, and enabling advanced analytics through trusted, secure, and well-governed data.

What You Will Do

  • Build and Operate Data Pipelines (Batch and Streaming)Design and implement batch and streaming ingestion from APIs, relational databases, file drops, event streams, and external partners.
  • Implement, test, and optimize ETL/ELT pipelines using Python and PySpark to produce curated, analytics-ready datasets for reporting, visualization, and machine learning.
  • Implement incremental processing, change data capture (CDC), data contracts, schema validation, and reusable transformation frameworks.
  • Improve pipeline reliability through automated testing, orchestration, monitoring, retry handling, and operational runbooks.
  • Deliver an AWS-Native Lakehouse Data Platform
  • Design and implement a Delta Lakehouse-style data platform using AWS-native services to provide Databricks-like capabilities for data engineering, analysis, and data visualization.
  • Build and manage a scalable lakehouse on Amazon S3 using Apache Iceberg and open columnar formats such as Apache Parquet.
  • Implement SQL-like table reliability for data stored in Amazon S3, including ACID transactions, schema evolution, partition evolution, snapshot isolation, time travel, and rollback capabilities using Apache Iceberg.
  • Enable fast, interactive querying of lakehouse data using AWS-native query and compute services such as Amazon Athena, Amazon EMR, AWS Glue, and Amazon Redshift where appropriate.
  • Optimize performance and cost through partitioning, compaction, file sizing, statistics, caching, lifecycle policies, and efficient separation of compute and storage.
  • Establish standardized development, test, and production environments with consistent configuration and controlled promotion across stages.
  • Metadata, Governance, Access Control, Lineage, and Quality
  • Implement data governance and fine-grained access control using AWS-native services, including AWS Lake Formation, AWS Glue Data Catalog, AWS Identity and Access Management (IAM), AWS Key Management Service (KMS), and related security services.
  • Implement a managed metadata repository for dataset cataloging, ownership, business definitions, tagging, classification, and discoverability.
  • Enable end-to-end lineage from source through transformation and consumption to support auditability, impact analysis, and regulatory requirements.
  • Apply policy-based access, least-privilege permissions, row-, column-, and cell-level controls where required, data classification, retention, encryption, and secure data handling.
  • Build operational data quality checks for freshness, completeness, uniqueness, validity, consistency, and anomaly detection, and publish measurable SLAs/SLOs.
  • AWS Automation, CI/CD, and Operations
  • Implement automated AWS provisioning using Infrastructure as Code (IaC) to create consistent environments and secure-by-default baselines.
  • Build and enhance CI/CD for data pipelines and lakehouse components, including automated tests, security checks, validation gates, packaging, deployment, promotion, and rollback strategies.
  • Implement observability with centralized metrics, logs, traces, alerts, dashboards, runbooks, and incident-response procedures.
  • Continuously evaluate platform performance, scalability, reliability, security, and cost, and implement measurable improvements.
  • Cross-Team Collaboration and Documentation
  • Work closely with data, application, analytics, AI/ML, security, networking, and cloud platform teams to support mission needs and delivery timelines.
  • Maintain high-quality engineering documentation, including architecture diagrams, data models, SOPs, interface specifications, operational runbooks, and secure configuration baselines.
  • Present technical findings, trade-offs, risks, and recommendations clearly to technical and non-technical stakeholders.

What You Will Need

  • Bachelor's degree in Engineering, Information Technology, Computer Science, Data Engineering, or a related field, or FOUR (4) years equivalent practical experience in leu of degree.
  • SIX (6) years of relevant experience.
  • Hands-on experience implementing AWS-native data lake or lakehouse architectures using Amazon S3 and services such as AWS Glue, Amazon Athena, Amazon EMR, AWS Lake Formation, and Amazon Redshift.
  • Strong experience developing production ETL/ELT pipelines using Python and PySpark, including data modeling, transformation, testing, performance tuning, and error handling.
  • Hands-on experience with Apache Iceberg, including ACID transactions, snapshots, schema and partition evolution, time travel, table maintenance, and query optimization.
  • Advanced SQL skills and experience supporting analytical queries, semantic layers, reporting tools, and data visualization workloads.
  • Experience implementing metadata management and governance capabilities, including cataloging, lineage, ownership, classification, policy enforcement, and fine-grained access controls.
  • Experience with AWS security fundamentals, including IAM and least privilege, KMS encryption, secrets management, network security, logging, and secure SDLC practices.
  • Experience provisioning AWS resources using IaC and operating data platforms across multiple environments.
  • Experience building or operating CI/CD pipelines for data workflows, including testing, packaging, deployment automation, environment promotion, and rollback.
  • Experience troubleshooting distributed data-processing workloads and optimize performance, reliability, and cost.

What Would Be Nice to Have

  • Hands-on experience with Databricks, Delta Lake, or migrating Databricks workloads to AWS-native services and Apache Iceberg.
  • Experience with AWS Step Functions, Amazon Managed Workflows for Apache Airflow (MWAA), Amazon Kinesis, AWS Database Migration Service (DMS), AWS Lambda, Amazon MSK, or similar ingestion and orchestration services.
  • Experience with modern DevOps practices and tools such as Git, Terraform, AWS CloudFormation or AWS CDK, Jenkins, AWS CodePipeline, GitHub Actions, and Docker.
  • Experience integrating lakehouse data with business intelligence and visualization tools such as Amazon QuickSight, Tableau, or Power BI.
  • Experience using AI-assisted coding tools, such as GitHub Copilot, ChatGPT, Cursor, or Kiro, to accelerate implementation while maintaining code quality, testing, review, privacy, and security controls.
  • Knowledge graph and Graph RAG experience, including graph modeling, ontology and taxonomy alignment, entity resolution, relationship extraction, and hybrid retrieval that combines graph traversal with semantic or vector search.
#J-18808-Ljbffr Jobleads-US
Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the AWS Lakehouse Data Engineer in Atlanta, GA vacancy
  •  ...engineeringjobs.net, Inc. is hiring an AWS Lakehouse Data Engineer to design, build, and operate a cloud-native data platform powering AI/ML, analytics, and reporting. You will create a scalable lakehouse on Amazon S3 with Iceberg, implement CDC, schema evolution, and... 
    Amazon Web Service
    Remote job

    Jobleads-US

    Atlanta, GA
    2 days ago
  •  ...team of employee-owners. We are seeking a Data Engineer III to partner with stakeholders and...  ...across multiple cloud providers such as: AWS, Microsoft/Azure, and GCP. This includes...  ...broader data stack: object storage and lakehouse layers, processing and data governance.... 
    Amazon Web Service
    Full time
    Flexible hours

    HNTB Companies

    Atlanta, GA
    3 days ago
  •  ...employee-owners. We are seeking a Senior Data Engineer to lead the partnership with...  ...pipelines across multiple cloud providers (AWS, Microsoft/Azure, GCP), with deep expertise...  ...broader data stack, including object storage, lakehouse architectures, processing, and data... 
    Amazon Web Service
    Full time
    Flexible hours

    HNTB Companies

    Atlanta, GA
    3 days ago
  • $86.7k - $170.9k

     ...Join Deloitte’s Core AI & Data practice and help organizations...  ...capabilities. As a Databricks Data Engineer, you will support the design,...  ...: Amazon Web Services (AWS), Microsoft Azure, or Google Cloud...  ...Experience with data warehousing, lakehouse architecture, and performance... 
    Amazon Web Service
    Local area
    Visa sponsorship

    Deloitte

    Atlanta, GA
    3 days ago
  •  ...Description Job Description Position: Data Engineer IV Location: Atlanta, GA Duration:...  ...cloud-based data solutions using Azure, AWS, or Google Cloud Platform. · Integrate...  ...Experience with Databricks, data lakes, lakehouse platforms, or cloud data warehouses. ·... 
    Amazon Web Service

    4P Consulting Inc.

    Atlanta, GA
    a month ago
  • $115k - $184k

     ...greatest potential. Title and Summary Senior Data Engineer Overview: Marketing Services...  ...data warehouse, data lake, and lakehouse solutions for large-scale processing and...  ...discipline; advanced degree a plus. • AWS data engineering experience. • Azure Data... 
    Amazon Web Service
    Full time
    Part time
    Worldwide
    Flexible hours

    Mastercard

    Atlanta, GA
    7 days ago
  •  ...Description: We are seeking an experienced Senior Data Engineer to design, develop, and optimize...  ...practices, cloud-native architectures, Lakehouse platforms, and distributed data...  ...Spark, Delta Lake, Snowflake, Kafka, and AWS, while also contributing to the adoption... 
    Amazon Web Service

    Openkyber

    Atlanta, GA
    1 day ago
  • Managing Director, Data Engineering & AIWho You'll Work WithAs a Managing Director in Slalom's...  ...clients on cloud-native architectures across AWS, Azure, and Google Cloud.* Ensure...  ...platforms.* Guide enterprise adoption of lakehouse architectures, data products, data... 
    Amazon Web Service
    Temporary work

    Slalom

    Atlanta, GA
    1 day ago
  •  ...Position: Senior Data Engineer Experience: 10+ Years Employment Type: Contract W2...  ...optimize data pipelines using tools such as AWS Glue, Informatica, Azure Data Factory,...  ...implement data warehouses, data lakes, and lakehouse solutions. Work with cloud... 
    Amazon Web Service
    Contract work

    Openkyber

    Atlanta, GA
    1 day ago
  • Lead enterprise data engineering for a Databricks-based lakehouse and streaming analytics platform at Norfolk Southern Corp . In this hybrid role in Atlanta,...  ...for high-volume event processing; and 3+ years using AWS analytics services including S3 , IAM , Glue/Lambda/MSK... 
    Amazon Web Service
    Work at office
    Remote work
    Shift work
    Weekend work

    Norfolk Southern Corp

    Atlanta, GA
    4 days ago
  • AWS Data EngineerLooking for someone with hands-on development experience in the following AWS technologies. The role will be based in Atlanta and the client needs someone who can start immediately.· Redshift · Python · Spark · Kinesis
    Amazon Web Service
    Immediate start

    ClifyX

    Atlanta, GA
    3 days ago
  •  ...and leading the implementation of scalable data architectures for cutting-edge AI and...  ...spearhead the design of complex feature engineering workstreams, ensuring our data assets are...  ...code (Terraform) across cloud providers (AWS, Azure, GCP)Exceptional time management and... 
    Amazon Web Service
    Apprenticeship
    Work at office
    Local area
    Easy work

    McKinsey & Company

    Atlanta, GA
    4 days ago
  • $113.5k - $181.5k

     ...Center for the Advancement of Data and Research in Economics (CADRE...  .... We seek a Data/Software Engineer to design, build, and maintain...  .... ~ Experience using core AWS services to build and support...  ...stores like S3, Spark, Airflow, Lakehouse architectures, real-time databases... 
    Amazon Web Service
    Permanent employment
    Full time
    Temporary work
    Part time
    Remote work
    Visa sponsorship
    Relocation package
    Shift work

    Federal Reserve System

    Atlanta, GA
    3 days ago
  •  ...Role: Data Engineer Duration: FULL TIME (Accepting H1B Transfer ) Location:...  ...streaming and batch workflows in Databricks Lakehouse using Pyspark and Spark SQL ~...  ...generic Cloud Platform certifications (AWS, Azure, GCP, etc.,) Any one of the Data... 
    Amazon Web Service
    Full time
    Work experience placement
    H1b

    ACI Infotech

    Atlanta, GA
    3 days ago
  • $176k - $179.5k

     ...QuantumBlack, AI by McKinsey, as a senior software engineer, building AI that transforms Pharma and...  .... You will design and manage scalable data pipelines and secure analytics platforms,...  ...deploying across major cloud platforms (AWS, Azure, GCP) A strong foundation in system... 
    Amazon Web Service
    Apprenticeship
    Easy work
    Shift work

    McKinsey & Company

    Atlanta, GA
    4 days ago
  •  ...: United StatesZip/Postal Code: 30319Job DescriptionMust Have: AWS Data service experience (S3, GLUE, Lambda, Redshift, KMS etc.), Strong...  ...and Aws CDK knowledge. 8+ years of experience in Data Engineer with AWS services like Simple Storage Service(S3), Lambda, Glue... 
    Amazon Web Service
    Remote work
    Shift work

    LVT Labs Consulting

    Atlanta, GA
    3 days ago
  •  ...TechnologyJob Number: 194669Apply: JobGeorgia-Pacific is seeking a Sr Data Engineer to join the IT Data & Analytics Packaging & Cellulose Delivery...  ...Core business.You will guide technical decisions across AWS, Amazon Redshift, Power BI, Microsoft BI technologies, source integrations... 
    Amazon Web Service
    Work at office
    Worldwide
    Flexible hours

    Georgia-Pacific

    Atlanta, GA
    1 day ago
  •  ...95316Apply: JobGeorgia-Pacific is seeking a Senior Databricks Data Engineer to join our Data Platform team in Atlanta, Georgia. In this role...  ...advanced analytics, Genie Spaces, or AI agentsExperience with AWS, Microsoft Azure, or Google Cloud and native storage services such... 
    Amazon Web Service
    Work at office
    Flexible hours
    3 days per week

    Georgia-Pacific

    Atlanta, GA
    16 hours ago
  •  ...Job Description Job Description Data Engineer Location: Atlanta, Georgia (Hybrid Remote- 2 days onsite/ 3 days remote) Job Type: Full...  ...(e.g., Python, C#). Familiarity with cloud services like AWS or Azure. Experience with data modeling, ETL processes, and... 
    Amazon Web Service
    Full time
    Remote work

    TIER4 GROUP

    Atlanta, GA
    a month ago
  •  ...Sr. Data Engineer Location: Plano, TX OR Atlanta, GA (Onsite from Day 1 & F2F Required) Type: Long Term Job Description: ~5+...  ...requiring complex queries, and optimization. ~ Experience with AWS-based data services technologies (e.g., Glue, RDS, Athena, etc.... 
    Amazon Web Service

    Yochana

    Atlanta, GA
    3 days ago
  • $205.32k

     ...GroupEngineering / Product DevelopmentJob ProfileSr Lead Data EngineerManagement LevelSr Manager - Non...  ...EngineerJob Description: Sr. Lead Data Engineer positions offered by Cox Automotive...  ...data cataloging platforms including AWS Glue or Collibra to support data discoverability... 
    Amazon Web Service
    Full time
    Work at office
    Local area
    Remote work
    Flexible hours

    Cox Enterprises

    Atlanta, GA
    1 day ago
  •  ...only be submitted via our official job platform. Job Title: Data Engineer (hybrid) Location: US-GA-Atlanta (Sandy Springs) FLSA : Exempt...  ...data pipeline tools (e.g., Apache Airflow, DBT, Snowflake, AWS Glue). ~ Familiarity with cloud platforms (AWS, Azure, GCP).... 
    Amazon Web Service
    Contract work
    Temporary work
    Local area
    Flexible hours

    Safe-Guard Products International LLC

    Atlanta, GA
    a month ago
  •  ...Description: Senior Data Engineer - Data Platforms Role Overview We are seeking a Senior Data Engineer to design, build, and maintain scalable...  ...has deep experience building cloud native data solutions on AWS, strong fundamentals in distributed systems, and hands on experience... 
    Amazon Web Service

    Openkyber

    Atlanta, GA
    1 day ago
  • $100k - $176.3k

    Healthcare Data Engineer Position Description Healthcare Data Scientist/Engineer CGI is a global IT and business consulting services...  ...analytics solutions using cloud and big data platforms such as AWS, Azure, GCP, Databricks, or Snowflake • Communicate... 
    Amazon Web Service
    Work at office
    Local area
    Atlanta, GA
    14 days ago
  •  ...Job Description Job Description Senior Software Data Engineer (Python) We are looking for a Senior Software Data Engineer – Python--...  ...infrastructures. You will leverage your expertise in Python, SQL, AWS, and Infrastructure as Code to deliver scalable and secure data... 
    Amazon Web Service

    Pointwest Technologies Corp

    Atlanta, GA
    a month ago
  •  ...Data Engineer (Hybrid) We are seeking a skilled Data Engineer to design, develop, and optimize scalable data pipelines and data platforms...  ...Architecture Data Quality and Data Governance Cloud Technologies AWS (Glue, EMR, Lambda, S3, Redshift, Athena) or Azure (ADF, Data... 
    Amazon Web Service

    Axelon

    Atlanta, GA
    3 days ago
  •  ...being for you and your family.What you'll doAs a Knowledge Graph Data Engineer, you will build graph-centric data pipelines and a semantic...  ...create and modify GitHub Actions CI/CD pipelines, and operate AWS- and Docker-based infrastructure to deliver reliable, scalable... 
    Amazon Web Service
    Apprenticeship
    Work experience placement
    Easy work

    McKinsey & Company

    Atlanta, GA
    4 days ago
  •  ...Amazon Data Services, Inc. in Atlanta, GA seeks a Sr. Mechanical Engineer for Data Center Colocation. You will lead site assessments, conduct risk analyses, and influence design standards to meet AWS business needs. Expect collaboration with vendors, cross-functional... 
    Amazon Web Service

    Jobleads-US

    Atlanta, GA
    2 days ago
  • $116k - $145k

     ...roster has an opening with your name on it We are looking for a Data Engineer to join our growing data engineering team and help build the...  .... ~ Experience working with cloud-based data platforms (AWS, GCP, or Azure). Preferred Qualifications Experience... 
    Amazon Web Service
    Temporary work
    Local area
    Worldwide

    FanDuel

    Atlanta, GA
    17 hours ago
  •  ...Role We are looking for a technically skilled and motivated Data Engineer with 7-12 years of experience to join our growing Data &...  ..., reliable data pipelines using tools like Azure Data Factory, AWS Glue, or Palantir Foundry Develop ETL/ELT processes to transform... 
    Amazon Web Service
    Local area
    Relocation

    Thought Logic Consulting

    Atlanta, GA
    19 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to AWS Lakehouse Data Engineer. Be the first to apply!