Data Platform Engineer
Take2 Consulting LLC
Take2 has proven experience bridging the intersection of technology and people solutions. As a proven, trusted provider for our Federal and commercial clients, we provide the right solutions, at the right time through trusted partnerships, customized to solve our client’s unique business challenges. Take2 invests time, discipline, and rigor into our technology and people solutions, as well as utilizes our proprietary People Cloud. Whether we are bridging the gap between IT talent and our customers’ business challenges, Take2 will work as a partner to best resolve client needs.
Take2 is hiring an AWS Lakehouse Data Engineer who is eligible to be sponsored for a Public Trust Clearance . This position is Remote , but it will require you to work East Coast Hours while being located in the United States.
Job Description:
Take2 is seeking an AWS Lakehouse Data Engineer to design, implement, and operate the cloud-native data platform that powers AI/ML, analytics, reporting, and data visualization. You will build a modern lakehouse on Amazon S3 using AWS-native services and open table formats, providing Databricks-like capabilities while maintaining portability, strong governance, cost efficiency, and operational control. You will also develop scalable batch and streaming ingestion, Python and PySpark ETL/ELT pipelines, metadata and governance services, and automated cloud provisioning and CI/CD across environments.
This role is ideal for an engineer who enjoys platform building, automation, performance optimization, and enabling advanced analytics through trusted, secure, and well-governed data.
What You Will Do
Build and Operate Data Pipelines (Batch and Streaming)
- Design and implement batch and streaming ingestion from APIs, relational databases, file drops, event streams, and external partners.
- Implement, test, and optimize ETL/ELT pipelines using Python and PySpark to produce curated, analytics-ready datasets for reporting, visualization, and machine learning.
- Implement incremental processing, change data capture (CDC), data contracts, schema validation, and reusable transformation frameworks.
- Improve pipeline reliability through automated testing, orchestration, monitoring, retry handling, and operational runbooks.
Deliver an AWS-Native Lakehouse Data Platform
- Design and implement a Delta Lakehouse-style data platform using AWS-native services to provide Databricks-like capabilities for data engineering, analysis, and data visualization.
- Build and manage a scalable lakehouse on Amazon S3 using Apache Iceberg and open columnar formats such as Apache Parquet.
- Implement SQL-like table reliability for data stored in Amazon S3, including ACID transactions, schema evolution, partition evolution, snapshot isolation, time travel, and rollback capabilities using Apache Iceberg.
- Enable fast, interactive querying of lakehouse data using AWS-native query and compute services such as Amazon Athena, Amazon EMR, AWS Glue, and Amazon Redshift where appropriate.
- Optimize performance and cost through partitioning, compaction, file sizing, statistics, caching, lifecycle policies, and efficient separation of compute and storage.
- Establish standardized development, test, and production environments with consistent configuration and controlled promotion across stages.
Metadata, Governance, Access Control, Lineage, and Quality
- Implement data governance and fine-grained access control using AWS-native services, including AWS Lake Formation, AWS Glue Data Catalog, AWS Identity and Access Management (IAM), AWS Key Management Service (KMS), and related security services.
- Implement a managed metadata repository for dataset cataloging, ownership, business definitions, tagging, classification, and discoverability.
- Enable end-to-end lineage from source through transformation and consumption to support auditability, impact analysis, and regulatory requirements.
- Apply policy-based access, least-privilege permissions, row-, column-, and cell-level controls where required, data classification, retention, encryption, and secure data handling.
- Build operational data quality checks for freshness, completeness, uniqueness, validity, consistency, and anomaly detection, and publish measurable SLAs/SLOs.
AWS Automation, CI/CD, and Operations
- Implement automated AWS provisioning using Infrastructure as Code (IaC) to create consistent environments and secure-by-default baselines.
- Build and enhance CI/CD for data pipelines and lakehouse components, including automated tests, security checks, validation gates, packaging, deployment, promotion, and rollback strategies.
- Implement observability with centralized metrics, logs, traces, alerts, dashboards, runbooks, and incident-response procedures.
- Continuously evaluate platform performance, scalability, reliability, security, and cost, and implement measurable improvements.
Cross-Team Collaboration and Documentation
- Work closely with data, application, analytics, AI/ML, security, networking, and cloud platform teams to support mission needs and delivery timelines.
- Maintain high-quality engineering documentation, including architecture diagrams, data models, SOPs, interface specifications, operational runbooks, and secure configuration baselines.
- Present technical findings, trade-offs, risks, and recommendations clearly to technical and non-technical stakeholders.
What You Will Need
- Bachelor's degree in Engineering, Information Technology, Computer Science, Data Engineering, or a related field, or FOUR (4) years equivalent practical experience in leu of degree.
- SIX (6) years of relevant experience.
- Hands-on experience implementing AWS-native data lake or lakehouse architectures using Amazon S3 and services such as AWS Glue, Amazon Athena, Amazon EMR, AWS Lake Formation, and Amazon Redshift.
- Strong experience developing production ETL/ELT pipelines using Python and PySpark, including data modeling, transformation, testing, performance tuning, and error handling.
- Hands-on experience with Apache Iceberg, including ACID transactions, snapshots, schema and partition evolution, time travel, table maintenance, and query optimization.
- Advanced SQL skills and experience supporting analytical queries, semantic layers, reporting tools, and data visualization workloads.
- Experience implementing metadata management and governance capabilities, including cataloging, lineage, ownership, classification, policy enforcement, and fine-grained access controls.
- Experience with AWS security fundamentals, including IAM and least privilege, KMS encryption, secrets management, network security, logging, and secure SDLC practices.
- Experience provisioning AWS resources using IaC and operating data platforms across multiple environments.
- Experience building or operating CI/CD pipelines for data workflows, including testing, packaging, deployment automation, environment promotion, and rollback.
- Ability to troubleshoot distributed data-processing workloads and optimize performance, reliability, and cost.
What Would Be Nice to Have
- Hands-on experience with Databricks, Delta Lake, or migrating Databricks workloads to AWS-native services and Apache Iceberg.
- Experience with AWS Step Functions, Amazon Managed Workflows for Apache Airflow (MWAA), Amazon Kinesis, AWS Database Migration Service (DMS), AWS Lambda, Amazon MSK, or similar ingestion and orchestration services.
- Experience with modern DevOps practices and tools such as Git, Terraform, AWS CloudFormation or AWS CDK, Jenkins, AWS CodePipeline, GitHub Actions, and Docker.
- Experience integrating lakehouse data with business intelligence and visualization tools such as Amazon QuickSight, Tableau, or Power BI.
- Experience using AI-assisted coding tools, such as GitHub Copilot, ChatGPT, Cursor, or Kiro, to accelerate implementation while maintaining code quality, testing, review, privacy, and security controls.
- Knowledge graph and Graph RAG experience, including graph modeling, ontology and taxonomy alignment, entity resolution, relationship extraction, and hybrid retrieval that combines graph traversal with semantic or vector search.
- ...institutions, you’ve come to the right place. As a Principal Software Engineer at JPMorganChase within the Corporate Technology Organization,... .../traceability of changes, and secure handling of sensitive data.Strong understanding of responsible AI use and control expectations...SuggestedWork at officeLocal area
- ...adventure where you can push the limits of what's possible.As a Lead Software Engineer at JPMorgan Chase, within the Commercial & Investment Banking's Payments Technology team to build a Data & AI Platform powering analytics and automation at scale. You’ll design and implement...Suggested
$61k - $101k
...Formal training or certification in software engineering concepts, with at least 5 years of... ...traceability, and secure handling of sensitive data Understanding of responsible AI... ...Lead Software Engineer for our AI/ML Data Platforms team at JPMorganChase, where you will...SuggestedFull time- ...JPMorganChase is seeking a Lead Software Engineer in its Corporate Technology organization. You will be a core technical contributor delivering trusted market-leading technology products in a secure, scalable manner. You will lead and participate in cross-functional...Suggested
- ...JPMorganChase, a leading global financial institution, seeks a Principal Software Engineer within the Corporate Technology Organization to deliver scalable, secure, and high-quality technology products across portfolios. You will collaborate with cross-team stakeholders...Suggested
$61k - $101k
...101,000 per year Requirements: We require formal software engineering training or certification with 7+ years of applied experience for... ...in engineering workflows, including security, resiliency, data sensitivity, and risk-based governance. We require experience...Full timeWork at office- ...JPMorgan Chase & Co. in Jersey City, NJ is seeking a Lead Software Engineer to drive the Data and Payments Business Observability Platform. You will lead the design of a scalable React UI and Java/Spring Boot backend, build event-driven microservices with Kafka, and...
- ...JPMorgan Chase & Co. in Jersey City seeks a Director of Software Engineering – Payment Data Platform Infrastructure to guide the BRIE data platform and NEO agent runtime. You’ll own reliability, security, and cost trade-offs across multi-region AWS and on‑prem environments...
- ...client sites four days per week and work remotely one day. A member of our recruitment team will provide more details.Principal Platform Engineer - AI & DataKey ResponsibilitiesDesign and develop scalable backend services using Java (11/17), Spring Boot, and Spring Cloud....Full timeWork at officeLocal areaRemote work1 day per week
- ...role SimplifyVMS is building an AI-first platform for managing the contingent workforce. We... ...for an exceptional and passionate engineer to own the design and delivery of our next... ...software development. You will turn complex VMS data into trusted, fast, intuitive analytics...Full timeFlexible hours
- ...know more about us. Role : Senior Python Site Reliability Engineer Location : Pennington, NJ/ Jersey City, NJ Mode : Hybrid... ...REST APIs, MySQL, Linux, automation, observability, and cloud platforms and a proven ability to improve platform reliability,...Full time3 days per week
$82.42k - $126.6k
...driving digital transformation for financial institutions. We specialize in leveraging advanced technologies such as AI, cloud, and data-led innovation to help our clients accelerate growth and unlock business value. Our AI-driven solutions empower financial institutions...Full timeTemporary workRelocation- ...Information Technology Enabled Services.Job DescriptionOur Big Data capability team needs hands-on developers who can produce beautiful... ...) is a strong advantage.Operating knowledge of cloud computing platforms (AWS, especially EMR, EC2, S3, SWF services and the AWS CLI)...Permanent employmentFull timeH1b
$105.4k - $207.8k
Position Summary Our Deloitte AI & Engineering team works to transform technology platforms, drive innovation, and help make a significant impact on our clients... ...fuel growth through innovation.Work you'll do As a Data Engineer III on the AI & Data team, you will be...Local area$142.32k - $213.48k
...successful to designing our digital architecture and ensuring our platforms provide a first-class customer experience. We reimagine... ...come join us. We’ll enable growth and progress together.The Sr Data Engineer, AVP is an intermediate level position responsible for...Full time- Job ID: 25727970Reference Number: 25-00659Title: Data EngineerLocation: Jersey City, NJ, 08830Posted Date: 2025-06-17Company: HAN Staffing Role: Data EngineerLocation: NJ5 Day's OnsiteExperience: 8+ Years AWS - 25%, Databricks- 25%, Pyspark -25%, Splunk- 25% Skill : AWS...
- Be part of a dynamic team where your distinctive skills will contribute to a winning culture and team. As a Data Engineer III at JPMorganChase within the Corporate Technology , you serve as a seasoned member of an agile team to design and deliver trusted data collection...
- TikTok USDS Joint Venture LLC is seeking a Software Engineer Intern (Data Foundation) for Summer 2027 in Seattle, WA. You will help build scalable data platforms and contribute to high-concurrency systems while collaborating with cross-functional teams. The internship offers...Summer workInternship
- ...focused technologist to design and evolve a production analytics platform used across its macro business. You will shape APIs and... ...will collaborate with researchers, traders, risk teams, and C++ engineers, contribute to releases and production support, and drive performance...
$120k - $140k
...Acunor is Hiring: Lead Data Engineer Location: NY / NJ / Pittsburgh – Onsite Employment: Full-Time Compensation: $120K–$140K... ...✅ Must-Have Skills ~8+ years of Data Engineering / Data Platform experience. ~5+ years hands-on PySpark and ETL/ELT development...Full time- Job ID: 25823899Reference Number: 25-00728Title: Senior Data Engineer (Python)Location: Jersey City, NJPosted Date: 2025-07-02Contact: Mahi MahiContact Email: ****@*****.*** Phone: (***) ***-****Company: HAN Staffing Advanced Python programming skills...
- Job ID: 25956017Reference Number: 25-00834Title: Python PySpark/ Data EngineerLocation: Columbus, OH and Jersey City, NJPosted Date: 2025-07-23Contact: Mahi MahiContact Email: ****@*****.*** Phone: (***) ***-****Company: HAN Staffing Job Description:6...Work experience placement
- ...proficiency in Oracle databases including advanced SQL PLSQL and performance tuningStrong in performance tuning of queriesExperience in data warehouse developmentStrong experience in data analysis and business requirementsBanking knowledge preferably futures derivatives...
- ...a brighter future and make a meaningful difference.As a Lead Data Engineer at JPMorganChase within the Corporate Sector, you are an integral... ...data collection, storage, access, and analytics data platform solutions in a secure, stable, and scalable wayBuild and optimize...
$116k - $152.25k
...opening with your name on itAs an Analytics Engineer for our Sports Modeling & Innovation team... ...building, maintaining, and improving the data models, datasets and data visualizations... ...FanDuel Sportsbook; its leading iGaming platform, FanDuel Casino; the industry’s...Temporary workLocal areaWorldwide- ...WiproContact: Meghana GorusuCompany: SRI Tech SolutionsRole: Data Engineer - Python/PySparkLocation: Jersey City, NJ (3 Days onsite/week... ...domains- Hands-on experience with Spark, Hadoop, Kafka, and cloud platforms- Solid knowledge of SQL and NoSQL databasesKey...Full time3 days per week
- ...Job Title5+ years of experience in data engineeringStrong proficiency with Databricks, Spark, PySpark/ScalaAdvanced skills in Power BI, DAX, and data visualization best practicesExperience with Azure Data Lake, Data Factory, and SQLStrong problem-solving skills and experience...
$88.59 - $96.59 per hour
...Data EngineerGenesis10 is currently seeking a Data Engineer for a contract position with a Global Financial Institution located in Jersey City, NJ. This is a 1... ...Spark ecosystem as they migrate a large-scale data platform from Vertica to Databricks. The team processes enormous...Contract work- We Are Looking For A Resource With Data Engineering ExperienceWe are looking for a resource with Data Engineering experience, preferably a Senior Data Lead skilled in investment and wealth management • Snowflake • dbt • Python • Data Warehousing / ETL
$80 - $85 per hour
...Senior Data EngineerSeeking a highly skilled Senior Data Engineer with 8+ years of hands-on experience in enterprise data engineering, including deep expertise... ...modeling and implementation, and cloud-native container platforms (Kubernetes / OpenShift). location: Jersey City,...Hourly payContract work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Data Platform Engineer. Be the first to apply!
- data engineer machine learning Jersey City, NJ
- aws data engineer Jersey City, NJ
- big data developer Jersey City, NJ
- data engineer hedge fund Jersey City, NJ
- sr data engineer Jersey City, NJ
- big data cloud engineer Jersey City, NJ
- finance data engineer Jersey City, NJ
- junior big data engineer Jersey City, NJ
- sr information security engineer Jersey City, NJ
- data center engineer Jersey City, NJ




