Data Platform Engineer
Take2 Consulting LLC
Take2 has proven experience bridging the intersection of technology and people solutions. As a proven, trusted provider for our Federal and commercial clients, we provide the right solutions, at the right time through trusted partnerships, customized to solve our client’s unique business challenges. Take2 invests time, discipline, and rigor into our technology and people solutions, as well as utilizes our proprietary People Cloud. Whether we are bridging the gap between IT talent and our customers’ business challenges, Take2 will work as a partner to best resolve client needs.
Take2 is hiring an AWS Lakehouse Data Engineer who is eligible to be sponsored for a Public Trust Clearance . This position is Remote , but it will require you to work East Coast Hours while being located in the United States.
Job Description:
Take2 is seeking an AWS Lakehouse Data Engineer to design, implement, and operate the cloud-native data platform that powers AI/ML, analytics, reporting, and data visualization. You will build a modern lakehouse on Amazon S3 using AWS-native services and open table formats, providing Databricks-like capabilities while maintaining portability, strong governance, cost efficiency, and operational control. You will also develop scalable batch and streaming ingestion, Python and PySpark ETL/ELT pipelines, metadata and governance services, and automated cloud provisioning and CI/CD across environments.
This role is ideal for an engineer who enjoys platform building, automation, performance optimization, and enabling advanced analytics through trusted, secure, and well-governed data.
What You Will Do
Build and Operate Data Pipelines (Batch and Streaming)
- Design and implement batch and streaming ingestion from APIs, relational databases, file drops, event streams, and external partners.
- Implement, test, and optimize ETL/ELT pipelines using Python and PySpark to produce curated, analytics-ready datasets for reporting, visualization, and machine learning.
- Implement incremental processing, change data capture (CDC), data contracts, schema validation, and reusable transformation frameworks.
- Improve pipeline reliability through automated testing, orchestration, monitoring, retry handling, and operational runbooks.
Deliver an AWS-Native Lakehouse Data Platform
- Design and implement a Delta Lakehouse-style data platform using AWS-native services to provide Databricks-like capabilities for data engineering, analysis, and data visualization.
- Build and manage a scalable lakehouse on Amazon S3 using Apache Iceberg and open columnar formats such as Apache Parquet.
- Implement SQL-like table reliability for data stored in Amazon S3, including ACID transactions, schema evolution, partition evolution, snapshot isolation, time travel, and rollback capabilities using Apache Iceberg.
- Enable fast, interactive querying of lakehouse data using AWS-native query and compute services such as Amazon Athena, Amazon EMR, AWS Glue, and Amazon Redshift where appropriate.
- Optimize performance and cost through partitioning, compaction, file sizing, statistics, caching, lifecycle policies, and efficient separation of compute and storage.
- Establish standardized development, test, and production environments with consistent configuration and controlled promotion across stages.
Metadata, Governance, Access Control, Lineage, and Quality
- Implement data governance and fine-grained access control using AWS-native services, including AWS Lake Formation, AWS Glue Data Catalog, AWS Identity and Access Management (IAM), AWS Key Management Service (KMS), and related security services.
- Implement a managed metadata repository for dataset cataloging, ownership, business definitions, tagging, classification, and discoverability.
- Enable end-to-end lineage from source through transformation and consumption to support auditability, impact analysis, and regulatory requirements.
- Apply policy-based access, least-privilege permissions, row-, column-, and cell-level controls where required, data classification, retention, encryption, and secure data handling.
- Build operational data quality checks for freshness, completeness, uniqueness, validity, consistency, and anomaly detection, and publish measurable SLAs/SLOs.
AWS Automation, CI/CD, and Operations
- Implement automated AWS provisioning using Infrastructure as Code (IaC) to create consistent environments and secure-by-default baselines.
- Build and enhance CI/CD for data pipelines and lakehouse components, including automated tests, security checks, validation gates, packaging, deployment, promotion, and rollback strategies.
- Implement observability with centralized metrics, logs, traces, alerts, dashboards, runbooks, and incident-response procedures.
- Continuously evaluate platform performance, scalability, reliability, security, and cost, and implement measurable improvements.
Cross-Team Collaboration and Documentation
- Work closely with data, application, analytics, AI/ML, security, networking, and cloud platform teams to support mission needs and delivery timelines.
- Maintain high-quality engineering documentation, including architecture diagrams, data models, SOPs, interface specifications, operational runbooks, and secure configuration baselines.
- Present technical findings, trade-offs, risks, and recommendations clearly to technical and non-technical stakeholders.
What You Will Need
- Bachelor's degree in Engineering, Information Technology, Computer Science, Data Engineering, or a related field, or FOUR (4) years equivalent practical experience in leu of degree.
- SIX (6) years of relevant experience.
- Hands-on experience implementing AWS-native data lake or lakehouse architectures using Amazon S3 and services such as AWS Glue, Amazon Athena, Amazon EMR, AWS Lake Formation, and Amazon Redshift.
- Strong experience developing production ETL/ELT pipelines using Python and PySpark, including data modeling, transformation, testing, performance tuning, and error handling.
- Hands-on experience with Apache Iceberg, including ACID transactions, snapshots, schema and partition evolution, time travel, table maintenance, and query optimization.
- Advanced SQL skills and experience supporting analytical queries, semantic layers, reporting tools, and data visualization workloads.
- Experience implementing metadata management and governance capabilities, including cataloging, lineage, ownership, classification, policy enforcement, and fine-grained access controls.
- Experience with AWS security fundamentals, including IAM and least privilege, KMS encryption, secrets management, network security, logging, and secure SDLC practices.
- Experience provisioning AWS resources using IaC and operating data platforms across multiple environments.
- Experience building or operating CI/CD pipelines for data workflows, including testing, packaging, deployment automation, environment promotion, and rollback.
- Ability to troubleshoot distributed data-processing workloads and optimize performance, reliability, and cost.
What Would Be Nice to Have
- Hands-on experience with Databricks, Delta Lake, or migrating Databricks workloads to AWS-native services and Apache Iceberg.
- Experience with AWS Step Functions, Amazon Managed Workflows for Apache Airflow (MWAA), Amazon Kinesis, AWS Database Migration Service (DMS), AWS Lambda, Amazon MSK, or similar ingestion and orchestration services.
- Experience with modern DevOps practices and tools such as Git, Terraform, AWS CloudFormation or AWS CDK, Jenkins, AWS CodePipeline, GitHub Actions, and Docker.
- Experience integrating lakehouse data with business intelligence and visualization tools such as Amazon QuickSight, Tableau, or Power BI.
- Experience using AI-assisted coding tools, such as GitHub Copilot, ChatGPT, Cursor, or Kiro, to accelerate implementation while maintaining code quality, testing, review, privacy, and security controls.
- Knowledge graph and Graph RAG experience, including graph modeling, ontology and taxonomy alignment, entity resolution, relationship extraction, and hybrid retrieval that combines graph traversal with semantic or vector search.
$150k - $180k
...Lead Control Engineer // Data Center Developer Location: Austin, Texas (On-site)Relocation Needed Compensation: $150,000–$180,000... ...telemetry, and reliability engineering ~ Experience with platforms such as Delta Controls, ALC, or similar building automation...SuggestedRelocationFlexible hours- ...best-in-class, proprietary, AI-driven data analytics. We've spent years building our... ...the United States. Our machine learning platform helps insurers evaluate policies at a... ...and growing quickly. As our second Data Engineer, you will own the data layer that turns customer...Suggested
- ...communication and collaboration across the supply chain. DESCRIPTION: Transflo is seeking a Senior Data Engineer to architect and own our enterprise data platform — from raw ingestion through curated, analytics-ready data products. You will be the foundational...SuggestedRemote work
- ...Quantiphi is an award-winning, AI-First digital engineering and consulting company focused on... ...expertise, disciplined cloud and data engineering practices, and cutting-edge... ...Premier partner to leading cloud and AI platforms such as NVIDIA, Google Cloud, AWS, and Snowflake...SuggestedFull timeRemote work
- ...Senior Data Engineer With Databricks Remote, Anywhere in the Continental US Due to the nature of the role, this position requires U.S. Citizenship. Data Surge is disrupting the services industry with cutting-edge technology that brings together the very best...SuggestedFull timeImmediate startRemote work
- ...operate production-grade batch and streaming data pipelines across SQL Server, cloud... ...clustering, compute utilization, and overall platform cost. Implement security controls... ...Work closely with Product, Application Engineering, Quality Assurance, DevOps/Site Reliability...
- We are looking for an experienced Senior Compute Platform Engineer, this Long-term Contract position. This role is 4 days/week onsite with our... ...for assigned platforms.Perform and coordinate hands-on data-center infrastructure work for assigned platforms, including...Long term contractRemote work
$120k - $150k
...English (Required)Work Shift:1st shift (United States of America)Please review the following job description:The Senior ServiceNow Platform Engineer is responsible for the engineering, operations, administration, automation, reliability, and continuous improvement of the...Permanent employmentFull timePart timeWork experience placementH1bWork at officeLocal areaImmediate startWork visaShift workDay shift- ...in Financial Services & Insurance Principal Data Scientist Job Responsibilities Lead the... ...probabilistic modeling. Partner with AI Engineering teams to productionize models and integrate them into enterprise AI platforms and operational systems. Design feature...Remote work
- ...DIRECTLY ON OUR W2 (EAD, OPT, USC, GC, H4) NO THIRD PARTY! Data Scientist We are seeking a Data Scientist to join a growing... ...degree in Data Science, Statistics, Mathematics, Computer Science, Engineering, or a related field. - FROM THE USA ~3+ years of experience...
$100k
...for recent grads in Mathematics, Statistics, Computer Science or Engineering or candidates with gaps in their career or people wanting to... ...Docker, Jenkins, Github, Kubernates and REST API's experienceFor data Science/Data Engineer, Data Analyst/AI/Machine learning...H1b- ...Current job opportunities are posted here as they become available. Market America is seeking a Senior Data Scientist supporting our product, engineering, leadership and marketing teams with insights gained from analyzing company data. In this role you will be...Temporary work
- ...enjoy your career with us! As a Principal Data Scientist at Lumentum Operations LLC,... .... Partner with manufacturing and engineering leaders to frame problems, define analytical... .... Work with data engineering and platform teams to productionize datasets, features...
- ...borrowers overlooked by traditional methods. Our platform enables financial institutions of all... ...of the U.S. financial system. Data Scientists at Zest AI use the power of machine... ...constraints Collaborate with ML engineers, product leaders, and domain experts to bring...Work at office
- ...Data Scientist Remote, Anywhere in the Continental US Due to the nature of the... ...audiences. Collaborate with policy, engineering, and operations teams to ensure analytics... ...verification systems. Experience with cloud platforms like AWS, Azure, or GCP in secure...Full timeImmediate startRemote work
- ...Great Place to Work® certification year after year. Principal Data Scientist Job requirements ~ Experience Range: 15 - 18... ...PySpark, and R, ensuring efficient data processing and feature engineering Develop, validate, and maintain probabilistic graph models and...
$111.1k - $132.7k
...into clear technical requirements; collaborates closely with data engineering to design or acquire the right datasets; and leads the model... ...learning mindset to stay current with evolving AI platforms, tools and methods. At the Volvo Group, we strive for a...Temporary workLive in$18 - $50 per hour
...university graduation and completion of the program Position Overview: Siemens Digital Industries Software is seeking an AI & Engineering Data Intern to support emerging Artificial Intelligence and Digital Thread initiatives within the Capital PreSales organization....Remote jobHourly payFull timeInternshipLocal area- ...Position: Senior Data Scientist Work Location: 1302 Pleasant Ridge Rd., Greensboro, JOB DESCRIPTION: Perform statistical analysis... .... Collaborate with other teams including search and data engineering, web and marketing analysts, data analyst, project managers, system...For contractors
$155k - $410k
...SummaryThe OpportunityAs an AI & GenAI Data Scientist-Director, you will leverage advanced... ...growth. Within our Data and Analytics Engineering practice, you will apply data,... ...engineering, machine learning, and cloud platforms, including AWS, Google Cloud, Microsoft...Full timeH1b$100k
...we are looking for entry-level software programmers, Java full-stack developers, Python/Java developers, data analysts/data scientists, and machine learning engineers for full-time positions with clients.Who should apply? Recent computer science/engineering/mathematics/...Full timeH1b- ...We are looking for entry-level software programmers, Java Full stack developers, Python/Java developers, Data analysts/ Data Scientists, Machine Learning engineers for full time positions with clients. Who Should Apply Recent Computer science/Engineering /...Full timeH1bRemote work
$124k - $280k
...SummaryThe OpportunityAs a Pricing and Revenue Data Science Senior Manager - Consumer... ...following fields of study: Accounting, Engineering, Data Processing/Analytics/Science, Computer... ..., visualization tools (Power BI), cloud platforms (AWS, Azure, GCP), or predictive...Full timeH1b$55k - $151.47k
...SummaryThe OpportunityAs a Site Reliability Engineer - Senior Associate, you will play a... ...the performance of networks, servers, and data centers, minimizing downtime and confirming... ...performance- Utilizing cloud infrastructure platforms such as AWS, Google Cloud Platform, and...Full timeH1b- ...approach to where we work. About the Role: As a Machine Learning Engineer, you will have the opportunity to collaborate closely with... ...problems executing on a high throughput system with dynamic data. About the Job: Design, develop, and deploy machine learning...Permanent employmentFull timeWork at officeRemote workWork from homeFlexible hours
$107.66k - $161.7k
...Quora : a global knowledge sharing platform with over 400M monthly unique visitors,... ...About the Team and Role: Our small engineering team works on challenging problems every... ...Machine Learning systems -- from prototyping, data pipelines and training, to realtime LLM...Remote jobFull timeWork experience placementInternship$77k - $202k
...SummaryThe OpportunityAs a GenAI Python Systems Engineer - Senior Associate, you will play a pivotal role in transforming raw data into actionable insights, enabling informed... ...engineering, machine learning, and cloud platforms, including AWS, Google Cloud, Microsoft Azure...Full timeH1b$63k - $140k
...SummaryThe OpportunityAs a GenAI Python Systems Engineer - Experienced Associate, you will... ...techniques to design and develop robust data solutions for clients, transforming raw data... ...engineering, machine learning, and cloud platforms, including AWS, Google Cloud, Microsoft...Full timeH1b$73.5k - $212.28k
...LevelManagerJob Description & SummaryAt PwC, our people in data and analytics engineering focus on leveraging advanced technologies and techniques... ...end-to-end AI applications integrated into various platforms- Managing CI/CD pipelines for AI systems using GitHub Actions...Full timeH1b$91k - $321.5k
...Technology (IT)Management LevelSenior ManagerJob Description & SummaryThe OpportunityAs a CTIO - AI Engineer- Senior Manager, you will play a pivotal role in transforming raw data into actionable insights, enabling informed decision-making and driving business growth. Within...Full timeH1b
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Data Platform Engineer. Be the first to apply!
- junior data engineer remote Greensboro, NC
- data engineering intern summer Greensboro, NC
- ai data Greensboro, NC
- data loss prevention engineer Greensboro, NC
- data officer Greensboro, NC
- provider data management Greensboro, NC
- data cabling Greensboro, NC
- health data Greensboro, NC
- oncology data Greensboro, NC
- test data management Greensboro, NC




