Sr Data Engineer
$135k - $160kMcGraw-Hill Education
Overview
Build the Future
At McGraw Hill, we are dedicated to delivering digital learning experiences that transform education for learners and educators. Our focus is on creating seamless, impactful products that truly benefit our users while supporting growth and collaboration across teams. We foster a culture that values innovation, teamwork, and a balance between career growth and personal well-being.
How can you make an impact?
The Senior Data Engineer in Data and Analytics is responsible for advancing McGraw-Hill Education's (MHE) business intelligence and data platform capabilities, delivering scalable, reliable, and actionable insights across financial, product, customer, user, and third-party data domains. This role is deeply hands-on — designing, building, and optimizing end-to-end data pipelines and architectures on AWS (including services such as S3, Glue, Redshift, Lambda, EMR, and Step Functions) and Databricks (leveraging Delta Lake, Unity Catalog, and MLflow where applicable).
The Senior Data Engineer will architect and implement dynamic reporting, analytics, and data modeling solutions that drive measurable outcomes in the education domain, while ensuring the performance, efficiency, and reliability of the broader Data Platform. The ideal candidate brings a strong data engineering foundation with deep, hands-on expertise in AWS cloud infrastructure and Databricks, including experience with Delta Lake architecture, medallion (Bronze/Silver/Gold) data design patterns, and Databricks Workflows for pipeline orchestration. Advanced proficiency in SQL and experience with Python or Scala for large-scale data transformation are essential. Familiarity with infrastructure-as-code (e.g., Terraform) and CI/CD practices for data pipelines is a strong plus.
This role requires close collaboration with business stakeholders, data analysts, and product teams to translate complex data requirements into robust, production-grade engineering solutions — ensuring timely, high-quality delivery across all data initiatives.
This is a remote position open to applicants authorized to work for any employer within the United States.
What You'll Do
- Senior Data Engineer must have prior hands-on experience designing and delivering data solutions on Databricks, including building and maintaining lakehouses using Delta Lake with a medallion (Bronze/Silver/Gold) architecture.
- Strong knowledge working with data from financial and operational systems, with proven experience implementing Slowly Changing Dimensions (SCD Types 1, 2, and 3) using Delta Lake MERGE operations and Databricks SQL within a unified lakehouse model.
- Experience running and optimizing cloud data platforms on Databricks, including cluster configuration, autoscaling policies, job scheduling via Databricks Workflows, and adherence to daily runbook SLAs through proactive monitoring and alerting.
- Strong experience with Git-based version control integrated into Databricks (Databricks Repos / Git folders) and project management tools such as Jira, operating within Agile/Kanban delivery frameworks.
- Strong experience with modern data architecture principles, including Unity Catalog for data governance, Delta Sharing, and cloud-native lakehouse design patterns on AWS with Databricks.
- Ability to translate business requirements into technical designs and deliver production-grade data solutions within Databricks, from initial scoping through deployment.
- Design and develop parallel and distributed ETL/ELT pipelines using Apache Spark (PySpark/Scala) on Databricks, applying partitioning, caching, and broadcast join strategies for optimal resource efficiency and throughput.
- Understand data mapping and transformation requirements and implement them using Databricks-native constructs including Spark transformations (aggregations, joins, unions, window functions, lookups, and pivot/unpivot operations) and Delta Live Tables (DLT) for declarative pipeline development.
- Develop and maintain Databricks Workflows and job orchestration logic (including dependency management, retry policies, and alerting), replacing traditional shell-based wrapper patterns with cloud-native, maintainable pipeline automation.
- Proven experience designing and building integrations that support standard data modeling constructs — fact tables, dimension tables, star and snowflake schemas, and aggregations — implemented as Delta tables within Unity Catalog.
- Ability to provide end-to-end technical guidance across the full software development life cycle, from requirements gathering and architecture design through implementation, testing, and production deployment on Databricks.
- Ability to produce high-quality solution design documentation, including data flow diagrams, pipeline architecture specs, and Unity Catalog data asset definitions, ensuring clarity for both technical and business stakeholders.
What You Bring
- Deep expertise in modern data lakehouse architecture, including Delta Lake, medallion design patterns, Unity Catalog governance, and the transition from traditional data warehousing to cloud-native lakehouse solutions on Databricks.
- 5+ years of experience in Data Engineering, with a focus on the following tools and technologies:
- Databricks — Delta Live Tables (DLT), Databricks Workflows, Unity Catalog, Delta Lake (MERGE, OPTIMIZE, VACUUM, Z-ordering), Databricks SQL, and MLflow
- AWS services — S3, Redshift, Glue, Lambda, EMR, Athena (with Iceberg), Step Functions, and IAM — integrated with Databricks as the primary compute and transformation layer
- Scripting and programming languages — Python (PySpark), Scala (Spark), or SQL as primary languages for pipeline development and data transformation within Databricks
- 3+ years of experience working with cloud platforms — primarily AWS — architecting and operating Databricks environments including workspace configuration, cluster policies, instance profiles, and cost optimization strategies.
- 1+ years of experience with workflow automation and pipeline orchestration using Databricks Workflows, Apache Airflow (with the Databricks provider), or equivalent cloud-native schedulers, replacing traditional Unix shell scripting with scalable, observable pipeline management.
Preferred Experience & Skills:
- Experience with Publication and Education domain.
- Prior experience or familiarity with Tableau/Alteryx.
Why work for us?
The work you do at McGraw Hill will be work that matters. We are collectively building experiences that will help shape the future of education. Play your part and experience a sense of fulfilment that will inspire you to even greater heights.
The pay range for this position is between $135,000 - $160,000 annually, however, base pay offered may vary depending on job-related knowledge, skills, experience, and location. An annual bonus plan may be provided as part of the compensation package, in addition to a full range of medical and/or other benefits, depending on the position offered. Click here to learn more about our benefit offerings.
McGraw Hill recruiters always use a “@mheducation.com” email address and/or from our Applicant Tracking System, iCIMS. Any variation of this email domain should be considered suspicious. Additionally, McGraw Hill recruiters and authorized representatives will never request sensitive information in email.
51188- ...areas of inspiration and expand your capabilities, then consider a career in Advisory. KPMG is currently seeking a Sr. Associate, Data Science Engineer to join our Consulting practice. Responsibilities: ~ Translate advanced business analytics problems into...SeniorFull timeH1bLocal area
- ...Data Engineer Contract length: 6+ month contract to hire - conversion contingent on performance Location: 100% remote - will require travel to Charlotte, NC 1-2 times a quarter Top Requirements: 5+ years experience ~ Hadoop Ecosystem – Spark/PySpark – queries...SeniorContract workRemote work
- ...Data Engineer Location: West Conshohocken PA Duration: 6 + months Job Description: Candidate should have previous experience building a data lake using Oracle database technologies. Candidate should be expert in improving extraction time and DB performance tuning...Senior
- Sr. Data Engineer Location: Charlotte, NC/Marlborough, MA (Initial Remote) Duration: 12+ Months Contract TJMAXx Job Description: Minimum years of experience: 7-10+ Hands on experience in Databricks Hands on experience in SnowflakeSeniorContract workRemote work
- ...Sr. Data Engineer Location: Austin TX (Remote) Duration: 6+ Months Job Description: ~ Key Skills to evaluate – Python (advanced level), Pyspark, data flow pipeline in AWS, distributed system, Snowflake, Redshift, ETL testing, QE knowledge JD:...SeniorRemote work
- ...Job Title Sr Data Engineer Location (100% Remote) Duration: 12 Months + Extension Job Type: Contract Top 4 Skills Required Spark Scala – Must 5 years with Spark Scala Python – Must AWS – Must (S3, Lambdas, EMR, EC2) Shell Scripting – Must...SeniorContract workRemote work
- ...Lead Data Engineer Location: Chicago, IL (Remote ok) Duration: 12+ Month Must have skill: Snowflake SQL, Data warehouse, Python, AWS-Lamda, ETL/ELT Required Skills: ~6-8 years+ experience on Snowflake SQL – advanced SQL expertise ~6-8 years+ experience...SeniorRemote work
- ...Sr. Data Engineer Data Engineers who can build scalable, reliable data pipelines and infrastructure, requiring mastery in SQL, Python/Scala, cloud platforms (AWS), and big data technologies (Spark, Kafka, Hadoop). Key skills include data modeling, ETL/ELT pipeline...Senior
- ...Sr Data Engineer / Arch with Databricks & Python / PySpark Experience Requirements: Hands on experience in analyzing complex requirements, designing, and developing data engineering solution on Databricks. Excellent analytical skills Python/PySpark are...Senior
- ...Senior Data Engineer We are seeking a senior data engineer with 4-8 years of experience, specializing in Spark and CoreJava. The ideal candidate will leverage technologies to deliver innovative mapping solutions, contributing significantly to our project outcomes and...Senior
- ...Sr. Data Engineer Insight Global is hiring for a Sr. Data Engineer to join a leading Health Insurance Agency. This team is looking to build a centralized data platform from the ground up, and building out a team to support this initiative. This is the cornerstone technical...SeniorRemote work
- ...POsition : Data Engineer Location : Minnetonka, MN AND Raleigh, NC Duration : 6 Months to hire Experience and Required qualifications: To be considered for this position, applicants need to meet the qualifications listed in this posting....Senior
- ...Sr. Data Engineer Our leading Fortune 100 Financial Services client is looking to bring on a Sr, Data Engineer to assist with several business critical pushes. In this role you will be contributing a specific line of business of our clients and developing new best practices...SeniorHourly payRemote work
- ...Sr. Data Engineer Location: Bentonville preferred, Dallas could be option Length: 1+ years contract Tech Stack Spark, Kafka, Python Data design, building data services, some data warehousing SQL - creating structures, tables,...SeniorContract work
- ...problems. -Applies specialized knowledge in a single discipline such as assembly/integration, cross-discipline functions, knowledge engineering, industry expertise, or legacy evolution. -Interacts with the customer to gain an understanding of the business environment and...Senior
- ...Job Details: • Databricks, Azure Data Factory, Databricks workflows, PySpark, Python, Databricks SQL • Data Modeling • Databricks, Azure Data Factory, Databricks workflows, PySpark, Python, Databricks SQL • Config driven data pipelines • Databricks Genie...SeniorRemote work
- ...Big Data Engineer Location: Day 1 onsite in McLean, VA on Tuesdays, and Wednesdays. Duration: 06+ months Contract (Possibility of Extension) Job Description Must Haves: Python, PySpark, SQL. Hadoop Platform. Responsibilities include: Cleanse, manipulate...SeniorContract work
- ...Purple Drive Sr Data Engineer Share Contractual Columbus, OH Role: Senior Engineer Total Exp: 6 - 8 Years Skills: Python, Databricks Must have databricks, Python, PySpark, CICD experience 3-5 years' experience with data management, best practices and...Senior
- ...Data Engineer Job Description : We are looking for a Senior Data Engineer who will help make our next decade just as revolutionary... ...better world, you're the right fit. Come grow with us. The Sr. Data Engineer for Enterprise Data will be an active...Senior
- ...Senior Data Engineer The Senior Data Engineer is responsible for collecting, designing, and converting complex data into information that can be interpreted by Data Scientists and Business Analysts. Data accessibility is the ultimate goal, which enables our organization...Senior
- ...Sr Data Engineer Experience on AWS and its web service offering S3, Redshift, EC2, EMR, Lambda, CloudWatch, RDS, Step functions, Spark streaming etc. Good knowledge of Configuring and working on Multi node clusters and distributed data processing framework Spark...Senior
- ...Data Engineer 100% Remote Chicago, IL 18 months to 20 months Contract Very Urgent. If the candidate is right fit, I will not put them into L1 …. Rather will directly submit to client. Need your support. Please look for candidates with minimum 9 to 10+ years… Project...SeniorContract workRemote work
- ...Azure Services, Power BI Proven experience with Symphony Engine development (architect-level skills in development and design).... ...Experience in Power BI for creating reports, dashboards, and data visualizations to aid in decision-making. Strong understanding...SeniorRemote work
- ...Sr. Data Engineer Project Optimus is the next step in our multi-year, Oracle Cloud journey that will implement multiple modules in Oracle’s Cloud ERP and EPM SaaS environments. One of the main objectives of Project Optimus is to redesign the Chart of Accounts (COA)...Senior
- ...team that sets the standard rather than follows it, Oteemo is where you belong. Job Description We're looking for a Senior Data Engineer to design and implement AI features end to end - from data pipelines through GenAI and agentic workflows to production...SeniorRemote work
- ...Sr Data Engineer This role is for a Change Data Capture (CDC) Senior Data Engineer that will join a team responsible for streaming data for a cloud-based data ecosystem consisting of a metadata driven data lake and databases that support real time analytics, extracts...SeniorRemote work
- ...Sr. Data Engineer Type: Remote Duration: Long Term Requirement: Candidate must have: Proficiency in Python, SQL, GCP tools, clean rooms (Google, Databricks), Kafka or similar streaming tools, and strong understanding of retail media and campaign KPIs....SeniorLocal areaRemote work
- ...Senior Data Engineer Cupertino, CA or Austin, TX Onsite role Must have strong Python and SQL and take a code pad test we will give them before submittal. Also, must have Airflow Scala, Kafka and or Spark. We are seeking an experienced, detail...SeniorRemote workWorldwide
- ...Sr Data Engineer Location: Jersey City, NJ Duration: Long Term Must Have Skills ~10+ years of professional experience in data engineering, designing and supporting enterprise-scale data platforms in production environments ~ The role requires hands-on...Senior
- ...Sr. Data Engineer Work hours: 8am-5pm CT This candidate will be working on data integration solutions for Optum Consume office. The work will involve data integration/engineering work involving Azure Databricks, Py Spark, spark/scala and other technologies. Also...SeniorWork at office
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Sr Data Engineer. Be the first to apply!
- aws data engineer United States
- director data engineering United States
- data platform engineer United States
- data engineer machine learning United States
- data science developer United States
- senior data engineer United States
- finance data engineer United States
- bi data engineer United States
- principal data engineer United States
- IT data engineer United States


