Data Engineer
Compunnel
div:has([data-free-thinking-preview-answer=true])+:is(.text-message,.relative:has(>.text-message))]:-mt-2 grow"> Job Summary We are seeking an experienced Lead Data Engineer with approximately 70% data engineering and 30% analytical responsibilities. The role requires strong proficiency in modern cloud-native data platforms, large-scale data processing, and analytical and data science workflows. The ideal candidate will have deep Databricks and PySpark expertise, finance or payroll domain experience, and the analytical fluency to partner closely with data scientists, economists, and business stakeholders on macroeconomic and financial data research. This role will span production pipeline development, platform optimization, exploratory data analysis, and scalable data platform development for financial analytics and research. Key Responsibilities
- Design, develop, and maintain scalable data pipelines for the ingestion, transformation, and distribution of payroll, macroeconomic, and financial datasets.
- Build and support Databricks-based platforms that enable financial research and analytical workloads.
- Implement ETL/ELT frameworks using PySpark and Delta Lake for structured and unstructured data from internal and external sources.
- Develop data models and data marts optimized for analytical, reporting, and machine learning use cases.
- Ensure data quality, consistency, lineage, governance, and observability across data assets.
- Optimize performance for large-scale datasets, including billions of records, multi-terabyte environments, and time-series data.
- Translate business requirements into scalable and maintainable data solutions.
- Perform exploratory data analysis, including profiling datasets and identifying distributions, outliers, missing patterns, and data drift.
- Translate data scientist logic into efficient and scalable PySpark implementations, including cross-sectional metrics and time-windowed aggregations.
- Build validation dashboards and exploratory notebooks to verify pipeline outputs and data quality.
- Support feature engineering by implementing complex aggregation and transformation logic at scale.
- Independently validate and sanity-check analytical outputs and identify results requiring further investigation.
- Perform ad hoc analytical work using pandas and NumPy alongside PySpark to support research initiatives.
- Contribute to architecture design discussions and evaluate tradeoffs related to catalog design, lakehouse architecture, medallion architecture, and data mesh concepts.
- Implement CI/CD pipelines using Databricks Asset Bundles, Bitbucket Pipelines, and Jenkins.
- Manage Unity Catalog governance, access patterns, and schema design.
- Maintain security, compliance, and data governance standards across data platforms.
- Leverage AI-assisted coding tools such as GitHub Copilot, Amazon Q, Kiro, or equivalent platforms to accelerate development.
- Review AI-generated code for correctness, performance, scalability, and maintainability.
- Integrate AI-assisted development workflows into engineering and analytical activities.
- Bachelor's or Master's degree in Computer Science, Data Engineering, Information Systems, Statistics, Economics, Finance, or a related field.
- 5+ years of experience in Data Engineering or Data Platform development.
- Finance or payroll domain experience, including familiarity with payroll data structures, pay-period logic, compensation and deduction relationships, or financial-services data environments.
- Experience handling large-scale datasets, including billions of records, multi-terabyte environments, and time-series data.
- Strong proficiency with Databricks, including Unity Catalog governance, access patterns, and catalog/schema design.
- Strong understanding of Delta Lake internals, including optimization, clustering, Change Data Feed, and versioning.
- Experience with Databricks Workflows, including orchestration, dependencies, and monitoring.
- Experience with Databricks Asset Bundles or equivalent deployment frameworks.
- Experience participating in architecture-level design decisions and evaluating technical tradeoffs.
- Strong proficiency in Python, SQL, PySpark, data modeling, and ETL/ELT development.
- Analytical fluency, including exploratory data analysis, basic statistical concepts, distributions, correlations, time-series patterns, and feature engineering.
- Proficiency with pandas and NumPy for ad hoc analytical work alongside production PySpark.
- Experience implementing CI/CD using tools such as Bitbucket Pipelines, Jenkins, and automated deployment frameworks.
- Experience with AI-assisted development tools such as GitHub Copilot, Amazon Q, Kiro, or equivalent platforms.
- Experience implementing data quality and validation frameworks.
- Strong analytical and problem-solving skills with statistical literacy.
- Ability to build reliable, scalable, and maintainable data pipelines with consideration for long-term ownership and support.
- Strong understanding of data governance, data quality, and metadata management.
- Excellent communication and documentation skills.
- Ability to work effectively in a fast-paced, data-driven environment.
- Ability to work across both engineering and analytical responsibilities, including production ETL development, data profiling, and dataset investigation.
- Ability to independently conduct exploratory analysis from an initial dataset through findings and recommendations.
- Ability to identify questionable data, investigate anomalies, and propose data-driven improvements.
- Experience with macroeconomic, capital markets, or financial-services data.
- Exposure to data architecture design, including lakehouse patterns, data mesh concepts, and medallion architecture.
- Experience with streaming and event-driven pipelines using Kafka or Structured Streaming.
- Census or geographic data processing experience, including TIGER and FIPS codes.
- Infrastructure-as-code experience with Terraform, CDK, or similar technologies.
- Experience migrating legacy data platforms to modern technology stacks, including Glue, EMR, or HDInsight to Databricks.
- Experience supporting machine learning and AI-driven analytics solutions.
- Data visualization experience using Power BI, Tableau, Databricks Dashboards, or similar platforms.
- Experience with Scala.
- Experience with SQL Server, PostgreSQL, Delta Tables, or NoSQL databases.
- Experience with Python visualization libraries, including Matplotlib and Plotly.
- Experience with Git and related DevOps and automation practices.
- Experience with payroll data structures and processing.
- Experience with macroeconomic analysis and forecasting.
- Experience with financial markets and alternative data.
- Experience with large-scale time-series data engineering.
- Experience with census and geographic data processing.
- Ability to use AI tools effectively to accelerate engineering and analytical workflows.
- Strong ownership of complex data initiatives and ability to collaborate with technical and business stakeholders.
$95k - $120.65k
...Position Overview Support and maintain the mission-critical Cost Roll process by ensuring accurate aggregation of healthcare claims data and precise matching of claims to provider demographic data. This role is responsible for preserving data pipeline integrity that supports...SuggestedFull timeWork at officeLocal areaVisa sponsorshipFlexible hours- ...Job Description We are seeking a skilled Snowflake Data Engineer to join our team in developing a centralized data management hub that supports digital marketing analytics and campaign measurement. This role is critical in transforming diverse digital datasets into actionable...Suggested
- ...it should be: simple, transparent, and enjoyable. Welcome to the future of car buying. Welcome to AiAuto. Job Summary The Senior Data Engineer leads the design, development, and maintenance of scalable and secure data architecture, pipelines, and systems to support the...SuggestedLive inRemote workWeekend work
- ...visa sponsorship is there for this role. job Description: Design develop and optimize scalable ETLELT pipelines Build and maintain data ingestion frameworks for structured and unstructured data Develop and manage data lakes data warehouses and data marts Ensure data quality...SuggestedVisa sponsorship
- Join Experis ' Data & Analytics team in Hanover, NJ (onsite) and work on innovative data projects using modern, cutting-edge technologies... ...use cases. Develop ETL/ELT solutions using modern data engineering tools and frameworks. Collaborate with business stakeholders, architects...SuggestedRelocation
- Data Platform Engineer Location: Parsippany, NJ Duration: Fulltime Job Description Must Have Technical/Functional Skills: 5+ years of total IT experience, including 3+ years of experience with AWS platform Good understanding of the AWS cloud platform capabilities is mandatory...Full time
- Role : Data Platform Engineer Location : Parsippany, NJ FTE ONLY Job Description Must Have Technical/Functional Skills 7+ years of Total IT experience, including 5+ years of experience in data platform engineering Strong hands-on experience with AWS services such as...
$75.6k - $106.78k
...Data Engineer II Req #: 0000247737 Category: Information Technology and Systems / Clinical Informatics Status: Full-Time Shift: Day Facility: RWJBarnabas Health Corporate Services Department: ITS Research and Data Science...Full timeTemporary workWork experience placementLocal areaFlexible hoursShift work- ...Job Description Looking for data engineers with 9+ years of experience. 1. Experience developing and deploying application code using SQL, Python, spark. 2. 3 to 5 years’ experience developing and deploying data pipeline in cloud. 3. 3 to 5 years experience...Hourly payRelocationWork visa
$125k - $135k
About The Role The Data Engineer Lead is a hands-on engineering role within AIG's Data Engineering organization. You will design, build, and operate production-grade data pipelines and platforms - and bring sound engineering judgment to every solution you deliver. AIG...Work at office- ...across cards, payments, lending, and core banking. We are an engineering-first organization that values ownership, bias for action, and... ...us on LinkedIn, Instagram, YouTube, and X. About the Role Lead Data Engineer with expertise in building scalable data lakes and data...
- ...Location: Hybrid — 3 days/week onsite in Newark, NJ About the Role Join a major investment firm as a Senior AWS Data Engineer supporting enterprise data platform initiatives. This role focuses on building scalable data pipelines, CI/CD automation, and end-to-end data solutions...3 days per week
$182k - $242k
...Founded in 2017, CoreWeave became a publicly traded company (Nasdaq: CRWV) in March 2025. Learn more at What You’ll Do The Data Engineering Team builds and operates the foundational data infrastructure powering analytics, AI, and operational decision-making across...Permanent employmentFull timeTemporary workCasual workWork at officeFlexible hours$130.7k - $196.1k
...experiences. Position Summary This strategic role will own the design, build, and run of AdvanSix’s digital data platform, spanning IT and OT, and the engineering lifecycle for advanced analytics and AI applications. This leader manages a 3-5 person team and coordinates...- Role : Data Engineer Location Madison , NJ (Hybrid) Job Summary Seeking an Associate Principal experienced in Python, Java, and Azure Data Factory to lead advanced data solutions. Job Description Design and implement data pipelines with Azure Data Factory Build scalable...
$94k - $145k
Data Engineering-Cloud Data Engineering Fabric Analytics Data Engineer Hybrid; Newark, NJ $94k - $145k plus bonus For more information on benefits and what we offer please visit us at The posted range is the hiring range for this role — a subset of the broader range available...$85 per hour
...Hybrid - 3 days onsite) Engagement: Contract Pay Rate: Up to $85/hr (C2C) or $80/hr (W2) We are seeking an experienced AWS Lead Data Engineer to join a high-performing Data Engineering team. This role requires a strong technical leader who can architect, build, and scale...Contract work- Position: Data Engineer Location: Newark, NJ - Hybrid Duration: 6 months Responsibilities Prior experience working with Market Data feeds and asset management industry or financial services experience. Prior experience designing and implementing a medallion architecture...
- Job Title: Data Engineer (Financial background) Location: Newark, NJ Hybrid Locals only Duration: 12 months Hands-on experience in Python, AWS (Lambda, Glue, Redshift), data streaming (Kinesis, SQS), REST APIs, full-stack (React, Spring Boot, Node.js), microservices,...Local area
- Data Engineer Client: Charles Schwab Location: Lone Tree, CO Onsite: 4 days a week Interview process: 2 video interviews to hire Contract: 6 months extendable Must-have Skills / Experience: Strong SQL and data warehouse / database development (≈50% of role focus) Hands‐...Contract workTemporary workFor contractors
- ...we will also conduct one round of internal evaluation with our panelists. Job summary: We are seeking an experienced AWS Lead Data Engineer to join our Data Engineering team. As a technical leader, you will be responsible for architecting, implementing, and managing scalable...For subcontractorWork at officeLocal areaShift work
- AWS Data Engineer We are seeking a highly skilled AWS Data engineer with 8 to 12 years of experience to join our team. The ideal candidate will have expertise in AWS, S3, IAM, Glue, Lambda, Cloud Formation, Python & SQL, Athena, AWS CloudWatch, and AWS. This hybrid role...Day shift
- ...Data Center Engineer, Newark, NJ We are seeking a talented and self-motivated Data Center Engineer to join a growing engineering team. The successful Data Center Engineer for this position will be working closely with engineers and traders to optimize our data center...Temporary workWork experience placementLocal area
- ...Datacenter Engineer The engineer in this position should have a wealth of experience engineering/installing datacenter equipment. The... ...cohesive package. This would include the complete engineering of data centers including, but not limited to, infrastructure, overhead...For contractorsWork at officeLocal area
- ...Title: AWS Data Engineer Location: Newark, NJ (Hybrid) Duration: 12 Months Job Description: Responsibilities: Designing, building, and maintaining efficient, reusable, and reliable code Ensure the best possible performance and quality of high scale data applications and...
- ...AbbVie, please visit us at . Follow @abbvie on LinkedIn, Facebook , Instagram , X and YouTube. Job Description The Data Engineer will design and build scalable data-pipelines across initiatives such as agentic trail design, digital twins and other modeling...Full timeTemporary workLocal area
- Position : Sr. Data Engineer Location: Newark, NJ (Hybrid) Duration: 12 Months Contract Job Description Bachelor's degree in Computer Science, Software Engineering, MIS, or equivalent combination of education and experience. 5 years of experience implementing and supporting...Contract work
- Max rate : ***/hr Local to NJ (Hybrid mode preferred but we are open for remote too) We are seeking a hands-on Senior Data Engineer with strong Denodo experience to design, build, and maintain scalable virtualized data products and semantic-layer solutions. The ideal candidate...Local areaRemote work
- Senior Data Engineer Location: Columbus, Ohio- need local Onsite: 4 days a week Contract: 6 months extendable (could go perm) Interview process: 1 video interview then onsite panel interview Duties and Responsibilities: Architecting, creating and maintaining data pipelines...Permanent employmentContract workLocal area
$60 - $70 per hour
JSR Tech Consulting is seeking a contract position for a data engineering role hybrid model, 3 days per week in Newark, NJ. The ideal candidate will have a strong background in AWS services and data engineering, responsible for designing and implementing data ingestion...Hourly payContract work3 days per week
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Data Engineer. Be the first to apply!
- ai data Parsippany, NJ
- data visualization Parsippany, NJ
- data collection researcher Parsippany, NJ
- data network cabling Parsippany, NJ
- data collection Parsippany, NJ
- data Parsippany, NJ
- clinical data coordinator remote Parsippany, NJ
- sap master data Parsippany, NJ
- data recovery Parsippany, NJ
- data modeling Parsippany, NJ

