Data Platform Engineer
Take2 Consulting LLC
Take2 has proven experience bridging the intersection of technology and people solutions. As a proven, trusted provider for our Federal and commercial clients, we provide the right solutions, at the right time through trusted partnerships, customized to solve our client’s unique business challenges. Take2 invests time, discipline, and rigor into our technology and people solutions, as well as utilizes our proprietary People Cloud. Whether we are bridging the gap between IT talent and our customers’ business challenges, Take2 will work as a partner to best resolve client needs.
Take2 is hiring an AWS Lakehouse Data Engineer who is eligible to be sponsored for a Public Trust Clearance . This position is Remote , but it will require you to work East Coast Hours while being located in the United States.
Job Description:
Take2 is seeking an AWS Lakehouse Data Engineer to design, implement, and operate the cloud-native data platform that powers AI/ML, analytics, reporting, and data visualization. You will build a modern lakehouse on Amazon S3 using AWS-native services and open table formats, providing Databricks-like capabilities while maintaining portability, strong governance, cost efficiency, and operational control. You will also develop scalable batch and streaming ingestion, Python and PySpark ETL/ELT pipelines, metadata and governance services, and automated cloud provisioning and CI/CD across environments.
This role is ideal for an engineer who enjoys platform building, automation, performance optimization, and enabling advanced analytics through trusted, secure, and well-governed data.
What You Will Do
Build and Operate Data Pipelines (Batch and Streaming)
- Design and implement batch and streaming ingestion from APIs, relational databases, file drops, event streams, and external partners.
- Implement, test, and optimize ETL/ELT pipelines using Python and PySpark to produce curated, analytics-ready datasets for reporting, visualization, and machine learning.
- Implement incremental processing, change data capture (CDC), data contracts, schema validation, and reusable transformation frameworks.
- Improve pipeline reliability through automated testing, orchestration, monitoring, retry handling, and operational runbooks.
Deliver an AWS-Native Lakehouse Data Platform
- Design and implement a Delta Lakehouse-style data platform using AWS-native services to provide Databricks-like capabilities for data engineering, analysis, and data visualization.
- Build and manage a scalable lakehouse on Amazon S3 using Apache Iceberg and open columnar formats such as Apache Parquet.
- Implement SQL-like table reliability for data stored in Amazon S3, including ACID transactions, schema evolution, partition evolution, snapshot isolation, time travel, and rollback capabilities using Apache Iceberg.
- Enable fast, interactive querying of lakehouse data using AWS-native query and compute services such as Amazon Athena, Amazon EMR, AWS Glue, and Amazon Redshift where appropriate.
- Optimize performance and cost through partitioning, compaction, file sizing, statistics, caching, lifecycle policies, and efficient separation of compute and storage.
- Establish standardized development, test, and production environments with consistent configuration and controlled promotion across stages.
Metadata, Governance, Access Control, Lineage, and Quality
- Implement data governance and fine-grained access control using AWS-native services, including AWS Lake Formation, AWS Glue Data Catalog, AWS Identity and Access Management (IAM), AWS Key Management Service (KMS), and related security services.
- Implement a managed metadata repository for dataset cataloging, ownership, business definitions, tagging, classification, and discoverability.
- Enable end-to-end lineage from source through transformation and consumption to support auditability, impact analysis, and regulatory requirements.
- Apply policy-based access, least-privilege permissions, row-, column-, and cell-level controls where required, data classification, retention, encryption, and secure data handling.
- Build operational data quality checks for freshness, completeness, uniqueness, validity, consistency, and anomaly detection, and publish measurable SLAs/SLOs.
AWS Automation, CI/CD, and Operations
- Implement automated AWS provisioning using Infrastructure as Code (IaC) to create consistent environments and secure-by-default baselines.
- Build and enhance CI/CD for data pipelines and lakehouse components, including automated tests, security checks, validation gates, packaging, deployment, promotion, and rollback strategies.
- Implement observability with centralized metrics, logs, traces, alerts, dashboards, runbooks, and incident-response procedures.
- Continuously evaluate platform performance, scalability, reliability, security, and cost, and implement measurable improvements.
Cross-Team Collaboration and Documentation
- Work closely with data, application, analytics, AI/ML, security, networking, and cloud platform teams to support mission needs and delivery timelines.
- Maintain high-quality engineering documentation, including architecture diagrams, data models, SOPs, interface specifications, operational runbooks, and secure configuration baselines.
- Present technical findings, trade-offs, risks, and recommendations clearly to technical and non-technical stakeholders.
What You Will Need
- Bachelor's degree in Engineering, Information Technology, Computer Science, Data Engineering, or a related field, or FOUR (4) years equivalent practical experience in leu of degree.
- SIX (6) years of relevant experience.
- Hands-on experience implementing AWS-native data lake or lakehouse architectures using Amazon S3 and services such as AWS Glue, Amazon Athena, Amazon EMR, AWS Lake Formation, and Amazon Redshift.
- Strong experience developing production ETL/ELT pipelines using Python and PySpark, including data modeling, transformation, testing, performance tuning, and error handling.
- Hands-on experience with Apache Iceberg, including ACID transactions, snapshots, schema and partition evolution, time travel, table maintenance, and query optimization.
- Advanced SQL skills and experience supporting analytical queries, semantic layers, reporting tools, and data visualization workloads.
- Experience implementing metadata management and governance capabilities, including cataloging, lineage, ownership, classification, policy enforcement, and fine-grained access controls.
- Experience with AWS security fundamentals, including IAM and least privilege, KMS encryption, secrets management, network security, logging, and secure SDLC practices.
- Experience provisioning AWS resources using IaC and operating data platforms across multiple environments.
- Experience building or operating CI/CD pipelines for data workflows, including testing, packaging, deployment automation, environment promotion, and rollback.
- Ability to troubleshoot distributed data-processing workloads and optimize performance, reliability, and cost.
What Would Be Nice to Have
- Hands-on experience with Databricks, Delta Lake, or migrating Databricks workloads to AWS-native services and Apache Iceberg.
- Experience with AWS Step Functions, Amazon Managed Workflows for Apache Airflow (MWAA), Amazon Kinesis, AWS Database Migration Service (DMS), AWS Lambda, Amazon MSK, or similar ingestion and orchestration services.
- Experience with modern DevOps practices and tools such as Git, Terraform, AWS CloudFormation or AWS CDK, Jenkins, AWS CodePipeline, GitHub Actions, and Docker.
- Experience integrating lakehouse data with business intelligence and visualization tools such as Amazon QuickSight, Tableau, or Power BI.
- Experience using AI-assisted coding tools, such as GitHub Copilot, ChatGPT, Cursor, or Kiro, to accelerate implementation while maintaining code quality, testing, review, privacy, and security controls.
- Knowledge graph and Graph RAG experience, including graph modeling, ontology and taxonomy alignment, entity resolution, relationship extraction, and hybrid retrieval that combines graph traversal with semantic or vector search.
$150k - $180k
...Lead Control Engineer // Data Center Developer Location: Austin, Texas (On-site)Relocation Needed Compensation: $150,000–$180,000... ...telemetry, and reliability engineering ~ Experience with platforms such as Delta Controls, ALC, or similar building automation...SuggestedRelocationFlexible hours- ...Senior Data Engineer With Databricks Remote, Anywhere in the Continental US Due to the nature of the role, this position requires U.S. Citizenship. Data Surge is disrupting the services industry with cutting-edge technology that brings together the very best...SuggestedFull timeImmediate startRemote work
- ...Data Engineer – Remote Remote | Full-time | Data Engineering We're looking for an experienced Data Engineer to join a growing data... ...datasets across bookings, customer interactions, pricing, digital platforms and third-party APIs, helping create a reliable and scalable...SuggestedFull timeRemote work
- ...Location: Remote We are seeking an experienced Data Engineer to support challenging data engineering, data management, and analytics initiatives. In this role, you will develop and maintain data pipelines and data structures that enable advanced analytics and data...SuggestedTemporary workLocal areaRemote work
$136k - $221k
...Data Acquisition Engineer Fully Remote (U.S. only) | Full-time | $136,000 – $221,000 + performance bonus Job description: Data Acquisition... ...Architect and develop custom end-to-end software platforms powering ticket-marketplace operations, pricing algorithms...SuggestedFull timeRemote workFlexible hours- ...Aeroflow Health – Data Engineer Aeroflow Health is a national leader in home medical equipment and clinical services. We combine technology... ...engineering experience with another relational database platform such as PostgreSQL, MySQL, or Oracle, with the ability and...Work at office3 days per week
$110k - $145k
...Job Description Summary The Senior Data Engineer designs and builds the AWS-native data foundation behind our enterprise AI applications... ...and Identity Integration Partner with security and platform teams to integrate data access with enterprise identity and access...Permanent employmentContract workRemote workVisa sponsorshipWork visaRelocation package- ...Data Solutions Engineer Remote (US) Klickly is a fast-growing, award-winning startup that is revolutionizing eCommerce. Based in the heart... ...customer-facing datasets, recurring data feeds, analytics platforms, or other complex data environments ~ Ability to translate...Remote work
- ...Platform Engineer NeuBird AI is scaling rapidly and we need a platform engineer who can build the internal tools and infrastructure that make our engineering teams more productive. You'll create self-service platforms, streamline developer workflows, and eliminate friction...Flexible hours
- ...Description Remote Our client seeks a Senior Information Security Engineer to operate and maintain an LLM-based code vulnerability... ...Preferred: experience in 24/7 operations, AWS Bedrock or hosted LLM platforms, ECS or other container orchestration, Splunk, Linux...Remote workShift workNight shiftRotating shift
- ...Great Place to Work® certification year after year. Principal Data Scientist Job requirements ~ Experience Range: 15 - 18... ...PySpark, and R, ensuring efficient data processing and feature engineering Develop, validate, and maintain probabilistic graph models and...
- ...and B2B technology. Every search we run produces data: each interaction with candidates, with our clients,... ...as unstructured data in Rolodex, our proprietary platform. We’re evolving Rolodex into a learning engine that turns that raw material into models our consultants...Remote work
- ...borrowers overlooked by traditional methods. Our platform enables financial institutions of all... ...of the U.S. financial system. Data Scientists at Zest AI use the power of machine... ...constraints Collaborate with ML engineers, product leaders, and domain experts to bring...Work at office
- ...Data Scientist / ManagerAre you looking for a challenge? Looking for an innovative organization and the opportunity to learn and grow professionally? We can help! We are seeking a Data Scientist / Manager to support the United States Department of Agriculture (USDA) Forest...Full timePart time
- ...client's team as a Senior Machine Learning Engineer , you will play a pivotal role in... ...and MLOps best practices to build robust data pipelines and deploy models at scale. You... ...experts to integrate AI capabilities into our platform (including Care Access products). Ensure...Remote work
- ...to design, build, and modernize mission-critical platforms using cloud-native architectures and modern engineering practices. We combine deep domain expertise with... ...-end solution architecture across application, data, integration, infrastructure, and security layers...
- ...focuses on streamlining workflows, enhancing data integrity, and supporting regulatory... ...methodologies to continuously enhance the platform and services offered to customers. Role... ...This full-time remote DevOps Engineer role is responsible for building, maintaining...Full timeRemote work
- CASE Consultants International in Asheville, NC is looking for AI/ML and environmental data professionals to support NOAA Fisheries’ Ecosystem Science Division. The role involves developing workflows for imagery analysis, ecological monitoring, and data processing. Candidates...Remote jobFlexible hours
- ...across the U.S. and APAC. As a Reference Platform NVIDIA Cloud Partner (NCP) and a validated... ...are seeking a skilled Site Reliability Engineer to join the GMI Global Infrastructure team... ...high performance AI/ML clusters in our data center. The ideal candidate will bring expertise...
- ...businesses by unlocking the hidden value in data. We deliver measurable outcomes 10X... ...seeking a highly motivated and skilled AI Engineer. You will have strong fundamentals in applied... ...agents directly into the RapidCanvas platform. Implement LLM & RAG Pipelines Develop...
- ...AI Engineer Location: Remote, Nationwide Our client is building the next generation of intelligent software, where AI is expected to do more than generate a compelling response. The goal is to create dependable AI experiences that understand context, take action...Remote work
$240k - $300k
...AI Engineer $240k–$300k base + equity Remote, United States High-growth applied AI... ...infrastructure that connects fragmented enterprise data, creates business context, and enables AI... ...handling. Build reusable agent and platform infrastructure. Work directly with...Remote work- ...Position Overview A large grocery retailer is seeking a Lead AI Engineer / Agentic Commerce Lead to drive the architecture, development... ...LLM-powered applications, AI agents, or conversational AI platforms. Experience with orchestration frameworks such as Temporal...Local area
- ...About the Role We’re looking for a hands-on Applied AI Engineer to build the AI tools, agents, and automations that... ...work spans dozens of interconnected systems: workflow platforms (n8n, Clay, Zapier, Make), data and BI tools (BQ, CloudSQL, Domo, Google Sheets), project...
- ..., iterating until they work well enough to ship or hand off to engineering Run experiments, read the results, and use what you learn to... ...escalates when something isn’t working Enough range to work across data pipelines, modeling, and serving infrastructure Builds with...
- ...About the role Stellar Blue.ai is looking for a AI Solutions Engineer who can combine strong client communication with practical AI workflow... ...fluency to ask good questions about tools, integrations, data flow, and workflow behavior. • The judgment to recognize when...Remote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Data Platform Engineer. Be the first to apply!
- ai data Asheville, NC
- data loss prevention engineer Asheville, NC
- provider data management Asheville, NC
- data cabling Asheville, NC
- health data Asheville, NC
- oncology data Asheville, NC
- test data management Asheville, NC
- data cabling installation Asheville, NC
- data internship Asheville, NC
- data modeling Asheville, NC

