Data Engineer
United IT
Data Engineer
We are seeking a highly skilled and motivated Data Engineer to play a pivotal role in designing, building, and optimizing our next-generation scalable data pipelines. This position requires expertise in processing massive datasets using cutting-edge technologies like Apache Spark, PySpark, and Hive within Cloudera Platform. Your primary objective will be to ensure the utmost data reliability, speed, and efficiency, providing a robust foundation for downstream business intelligence and advanced analytics initiatives.
Roles & Responsibilities:
- Data Pipeline Development & Maintenance: Design, build, and maintain highly scalable and efficient ETL/ELT data pipelines utilizing PySpark and Spark SQL, Hive for complex data transformations.
- Data Warehousing & Storage Optimization: Strategically manage data layout, partitioning, and indexing within Apache Hive and various cloud data lake solutions to optimize performance and accessibility.
- Performance Tuning & Optimization: Proactively identify and resolve performance bottlenecks in Spark jobs, leveraging Spark UI for in-depth analysis, effectively managing data skewness, and optimizing memory utilization.
- Diverse Data Integration: Develop robust solutions for ingesting high-volume and diverse datasets from both structured relational databases and unstructured flat files into our data ecosystem.
- Automated Workflow Orchestration: Implement and manage automated data workflows using industry-standard scheduling tools like Apache Airflow or platform-native schedulers, ensuring timely and reliable data delivery.
- Strategic Collaboration: Partner closely with data scientists, business analysts, and cross-functional enterprise teams to translate complex business requirements into technically sound and efficient data solutions.
Qualifications:
- Big Data Frameworks Expertise: Demonstrated high proficiency in Apache Spark architecture, including a deep understanding of drivers, executors, and Directed Acyclic Graphs (DAGs).
- Advanced Programming: Exceptional coding skills in Python and extensive experience with the PySpark API for developing intricate data transformations and processing logic.
- Querying & Schema Management: Strong command of HiveQL and ANSI SQL, coupled with expertise in data partitioning techniques and effective schema definition.
- Optimized Storage Formats: In-depth understanding and practical experience with optimized big data storage file formats such as Parquet, ORC, and Avro.
- Data Warehousing Fundamentals: Solid foundation in Dimensional Data Modeling, including Star and Snowflake schemas, and practical experience with Data Lakes concepts and implementation.
Preferred Qualifications:
- CI/CD & DevOps Automation: Experience with Continuous Integration/Continuous Deployment (CI/CD) practices and automation tools like Git, Jenkins, or Ansible.
- Cloud Ecosystem Development: Experience in development experience utilizing cloud-native big data utilities (e.g., AWS EMR, AWS Databricks) within major cloud platforms.
- NoSQL Database Integration: Exposure to and experience with NoSQL databases such as HBase, Cassandra, or MongoDB.
- ...leader, take a look at the exciting employment opportunities that are currently available and apply online.Job SummaryAs a Lead Data Engineer, you will be responsible for leading the design and implementation of complex data solutions and managing the organization's data...SuggestedFull timeLocal area
- Job Title: Data EngineerWork Location:Irving, TXRequired Skills:Expertise in Designing and developing scalable Apache spark ETL based Data processing pipelines.Strong command line knowledge in Unix Linux with Shell scripting using Bash Korn shell or Perl and File processing...Suggested
- Overview We are looking for a Senior Data Engineer with strong 7-10 year hands-on experience in Azure and Microsoft Fabric to help modernize and migrate data products from Azure Synapse to Microsoft Fabric. The ideal candidate will have experience with Azure Data Factory...SuggestedFull time
- ...every day. If you're ready to grow, lead and make a difference, come join our team and help shape the future of convenience.The Data Engineer - Marketing Technologies builds and maintains the data pipelines, models, and real-time streams that serve as the foundation for...SuggestedHourly payWork experience placement
$75.17k - $130.5k
Req ID:376258NTT DATA strives to hire exceptional, innovative and passionate individuals who want to grow with us. If you want to be... ...thinking organization, apply now.We are currently seeking a Data Engineer - Hybrid to join our team in Irving, Texas (US-TX), United States...SuggestedTemporary workWork at officeRemote workFlexible hours- ...Job Title: Lead Data Engineer Location: Irving, TX - Hybrid Must have: ~ Expert in Databricks, PySpark and Azure data services like ADF, Synapse etc. ~ Must have lead level experience ~8+ years of professional data engineering experience, with proven...
- ...Lead Data Engineer Lennar is one of the nation's leading homebuilders, dedicated to making an impact and creating an extraordinary experience for their Homeowners, Communities, and Associates by building quality homes and providing exceptional customer service, giving...Live in
$111.8k - $186.4k
..., development, and optimization of enterprise-scale healthcare data platforms. This role will play a critical part in building scalable... ...pipelines, modern analytics solutions, and cloud-based data engineering capabilities that power reporting, insights, and data-driven...Full timeH1bWork at officeRemote workWork from home2 days per week- Job Title: Sr. DATA Engineer Work LocationIrving TXJob Description:Seeking a Senior professional with 3 to 5 years of experience in Python Scala and Spark SQL to design and optimize scalable big data processing solutions within the Databricks Spark SQL Scala Python Java...
- Desirable Technical Skills • Familiarity with and invoking web-APIs• Exposure to machine learning engineering• Exposure to NLP and text processing• Experience with pipelines, job scheduling and workflow managementPersonal Skills Experienced in managing work with distributed...
$107.12k - $160.68k
...together.This is an intermediate-level Applications Development role responsible for building, implementing, and supporting enterprise data and ETL systems in close coordination with the Technology organization. The position blends hands-on systems analysis and...Full time- • Analyze and understand data sources & APIs• Design and Develop methods to connect & collect data from different data sources• Design... ...with and invoking web-APIs• Exposure to machine learning engineering• Exposure to NLP and text processing• Experience with pipelines...
- Job Title: Sr. DATA Engineer Work LocationTampa FLJob Description:Mandatory Karat InterviewRole Overview We are seeking a highly skilled and motivated AWS Certified Engineer to design build and optimize scalable data solutions within the Amazon Web Services AWS ecosystem...
- ...Job Title Required Skills: ~10+ years of experience in solutioning data pipeline for large enterprise data warehouse applications using AWS, Data Lake, Data Ingestion, Data Transformation, Computation, Orchestration, Reporting & Data Analytics ~ Good experience...
- ...Job Description Insight Global is seeking a Data Engineer to support a large-scale modernization program within the Analytics & Behavioral Change organization at a major enterprise healthcare company. This is a builder role, where engineers will execute against a well...
- ...Data Engineer Location: Dallas, TX (3 days onsite/week) Duration: Contract Interview Mode: In person Interview. Must Have: ~5+ years of experience in data engineering ~ Python ~ SQL ~ Data Warehouse ~ Experience with development, operation support and data...Contract work3 days per week
- ...Data Solutions Engineer As a part of our Big Data Product team, the Data Solutions Engineer will be responsible for developing and validating Big data products and applications which runs on the large Hadoop cluster. The qualified engineer will be developing and testing...Work experience placement
- ..., Kudu, Impala, HBase spark scala batch pipelines GCP, AWS - data pipelines Oozie, Airflow, Composer - scheduling tools CICD... ...highly preferred. Job Description ~ Experience in data engineer, Data Analytics ~ Big Data Technologies Expert in Python ~(...
- ...Data Engineer/Developer With Python And SQL Irving, TX (Onsite) Full time Key Responsibilities Design, develop, and implement new ETL (Extract, Transform, Load) jobs using Python and/or Spark to support various data initiatives. Migrate existing ETL processes...Full time
- ...Data Engineer (x3) Profile data ingestion requirements from SORs, Data Lakes or Data Warehouses Configure data pipelines using Azure Synapse Analytics or Data Bricks Configure Azure SQL or PostgreSQL tables, views, indexes, etc. Required Qualifications:...
- ...Job description : Data Engineer Location: Irving, TX 12+ Months Only W2 Python/Pyspark/sql candidate. Cloud nice to have. Need experience building data pipelines Data engineer with experience in building, tuning scalable high volume complex...
- ...Data Engineer – Implementation Consultant We are currently onboarding North America based OPCO’ s (Operating companies) to their new digital platform (Spark) which is built on Kubernetes running in AZURE, the customer and associate experience (CX/AX) on their new SPARK...
- ...Data Engineer – Onsite Locations: Raleigh, NC | Irving, TX | Chandler, AZ | New Jersey, NJ | Charlotte, NC | San Leandro, CA Duration: Long term Required Skills: Strong Python and/or Java development experience. Hands-on experience with Spark, Flink, streaming...
- ...Data Engineer (Contract)- 5 new contract openings Location : Onsite 2 days a week Irving, TX or Hartford, CT - Only Citizenship : Any (GC/USC will be given preference) Duration : 6+ months Data Engineer with 8+ years' experience including...Contract work2 days per week
$77.4k - $135.4k
...In this role, you will design, develop, and support scalable data and database solutions that enable analytics and innovation across... ...solutions. Partner with application developers, data engineers, BI teams, and analytics partners to deliver integrated solutions...- ...Position: Big Data Engineer Location: Remote Role Summary Excellent Scala, PySpark, Java coding experience for REST API and SOAP webservices Excellent AWS experience (Batch and API microservices) using EMR/EC2/ECS/S3/Airflow/Step Functions, AWS...Remote work
$145k - $165k
...healthcare fintech company, Wellfit is investing heavily in AI, data, and modern technology to better understand our business, support... .... About the Role: We are seeking a Senior Data & AI/ML Engineer to help Wellfit unlock the value of its data through applied AI,...Full timeImmediate start$120k - $130k
...including query development, performance tuning, and optimization required. • Hands-on experience using Python for data processing, analytics, or data engineering workflows required. • Experience with Databricks, Apache Spark, or similar distributed data processing...- ...Sr Data Engineer Location Canada Toronto Requirements ~6-8 years of experience with Ab Initio in ETL projects ~ Expertise as an Ab Initio Technical Expert ~ Strong problem-solving skills: Able to analyze issues, identify root causes, and implement timely...
- ...Senior Data Engineer We are seeking a Senior Data Engineer to design, build, and maintain scalable data infrastructure that powers analytics, reporting, and strategic decision-making. This role focuses on developing robust data pipelines, API integrations, and governance...
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Data Engineer. Be the first to apply!
- big data cloud engineer Irving, TX
- big data developer Irving, TX
- finance data engineer Irving, TX
- software data engineer Irving, TX
- hadoop big data developer Irving, TX
- data engineer machine learning Irving, TX
- senior data quality engineer Irving, TX
- data visualization developer Irving, TX
- aws data engineer Irving, TX
- data center engineer Irving, TX


