Hadoop / ETL Developer and PySpark
SRI Tech
Job ID: SRJ_7721Posted: 2025-10-10Location: Dallas, TX; Atlanta, GA; Cleveland, OH; Pittsburgh, PASalary: USD 98-100 / Yearly (Full Time)Employment Type: Full TimeIndustry: Computer SoftwareClient: WiproContact: Meghana GorusuCompany: SRI Tech SolutionsHadoop / ETL Developer and PySparkLocation – Dallas, TX / Atlanta, GA / Cleveland, OH / Pittsburgh, PA (Day 1 Onsite)Note – Pls submit only visa independent candidatesEmployment Type – FTEJob Description:We are seeking a highly experienced Senior Big Data & DevOps Engineer with 8+ years of professional experience in HDFS, Hive, Impala, PySpark, Python, and DevOps automation tools such as uDeploy and Jenkins. This role is responsible for managing end-to-end data operations, including HDFS table management, ETL pipeline development, multi-environment codebase governance, platform upgrades, and production support. The ideal candidate will have strong expertise in Linux system operations, Big Data ecosystem tools, and experience with incident/change management using ServiceNow. This role plays a key part in ensuring the stability, scalability, and efficiency of enterprise data platforms while enabling seamless development-to-production workflows.Key Responsibilities: Big Data Platform Operations
- Design, manage, and optimize HDFS directories, tables, and partitioning strategies.
- Implement and enforce data retention and lifecycle policies across large datasets.
- Administer Hive and Impala environments, ensuring high availability, performance tuning, and security compliance.
- Develop scalable ETL pipelines using PySpark, Hive, and Python.
- Build reusable frameworks for data ingestion, transformation, and aggregation.
- Optimize job performance through query tuning, resource management, and parallelization.DevOps & Environment Management
- Maintain and promote code across DEV, QA, UAT, and PROD environments.
- Develop and support CI/CD pipelines using Jenkins and uDeploy for automated deployments.
- Perform environment upgrades, patching, and dependency management aligned with release schedules.Linux & Infrastructure Operations
- Execute Linux administration tasks including performance tuning, disk management, and scripting (Bash/Python).
- Troubleshoot cluster-level issues including node failures, job errors, and distributed system anomalies.Change & Incident Management
- Drive incident resolution and change execution using ServiceNow workflows.
- Conduct root cause analysis (RCA) for critical issues and implement preventive solutions.
- Ensure compliance with ITIL processes for change, incident, and problem management.
- Partner with data engineers, developers, DevOps teams, and business analysts to ensure operational excellence.
- Mentor junior engineers and contribute to technical leadership across the Big Data ecosystem.
- Document operational procedures, troubleshooting guides, and architectural decisions for internal knowledge sharing.Required Qualifications:
- Bachelor’s degree in Computer Science, Information Technology, or related field.
- 8+ years of experience in Big Data engineering and DevOps practices.
- Advanced proficiency in HDFS, Hive, Impala, PySpark, Python, and Linux.
- Proven experience with CI/CD tools such as Jenkins and uDeploy.
- Strong understanding of ETL development, orchestration, and performance optimization.
- Experience with ServiceNow for incident/change/problem management.
- Excellent analytical, troubleshooting, and communication skills.Nice to Have:
- Exposure to cloud-based Big Data platforms (AWS EMR).
- Familiarity with containerization (Docker, Kubernetes) and infrastructure automation tools (Ansible, Terraform).SRI Tech Solutions is an equal opportunity employer and does not discriminate on the basis of race, color, gender, religion, age, sexual orientation, national origin or citizenship status or ethnic origin, disability, marital status, veteran status, or any other occupationally irrelevant criteria.
- ...hands-on contributor who can develop, implement, and optimize robust... ...time data solutions within our Hadoop ecosystem and streaming... ...hands-on knowledge of Python, PySpark, Unix, and SQL. Familiarity with... ...developing and implementing complex ETL/ELT processes for both batch...Suggested
$103.33k - $128.66k
...solutions that drive business insights. Key Responsibilities Develop, optimize, and maintain complex SQL and PL/SQL queries, stored procedures... ..., and physical data models to support database design Support ETL processes, transformation logic, and performance tuning Build...SuggestedFull timeLocal areaFlexible hours- ...a bright future for NTT DATA Services and for the people who work here.NTT DATA Services currently seeks a Java Full Stack Developer with Hadoop to join our team in Dallas, Texas (US-TX), United States (US).Ntt data is hiring a backend Java development lead to join our...SuggestedFor contractors
- ...Description:-At least 2 years of experience in Big Data space.Strong Hadoop - MAP REDUCE/Hive/Pig/SQOOP/OOZIE - MUSTCandidate should have... ...with Apache Parquet Data formatPast experience and exposure to ETL and data warehouse projectsExperience with Spark, FlumeCloudera/...SuggestedPermanent employmentFull timeH1b
$107.12k - $160.68k
...We are seeking a Senior Data Developer to join one of our small, co-located... ...data solutions within our Hadoop ecosystem and streaming... ...hands-on knowledge of Python, PySpark, Unix, and SQL.Familiarity with... ...developing and implementing complex ETL/ELT processes for both batch...SuggestedFull time- Key Responsibilities1. Backend & Microservices DevelopmentDesign, develop, and implement scalable, resilient microservices using Java and... ...application performance for low latency and high throughput.2. ETL & Data Pipeline EngineeringArchitect and maintain complex ETL/ELT...
- Key Responsibilities1. Backend & Microservices DevelopmentDesign, develop, and implement scalable, resilient microservices using Java and... ...application performance for low latency and high throughput.2. ETL & Data Pipeline EngineeringArchitect and maintain complex ETL/ELT...
- ...contribute to our shared success. This includes attracting and developing exceptional talent, recognizing and rewarding performance, and... ...Summary:Seeking a software engineer with experience in experience in ETL/Data Engineering, Mainframe, and real-time data replication...Full timeWork at officeFlexible hoursShift workDay shift
$103.33k - $128.66k
...Data / Database Engineer - SQL, Oracle, PL/SQL, Informatica, ETL, Powercenter Job Location - Westlake TX (Day One Onsite - Hybrid) Key Responsibilities Strong in Oracle SQL and PLSQL development Experienced with Informatica PowerCenter ETL development Comfortable supporting...Flexible hours- ...WiproContact: Meghana GorusuCompany: SRI Tech SolutionsRole: Senior PySpark Developer / EngineerLocation: Dallas, TX (3 Days onsite/ week)Job... ...for PySpark integrationETL Knowledge:Hands-on experience with ETL processes using any programming language or frameworkAbility to...Full time3 days per week
- ...Duties: Research, design, and develop computer and network software... ...each of the following: Build ETL pipelines to extract data from... ...data technologies such as Apache Hadoop, Apache Spark, Apache Hive,... ...using Spark, Scala, Java, Python, Pyspark, SQL, databricks, shell...Full timeLocal areaRemote workFlexible hours
$125.76k - $188.64k
...We are seeking a Senior Data Developer to join one of our small, co-located... ...data solutions within our Hadoop ecosystem and streaming... ...architectural knowledge of Python, PySpark, Unix, and SQL.Familiarity with... ...designing and implementing complex ETL/ELT processes for both batch...Full time$156.5k - $230k
...our commitment to being an inclusive workplace, attracting and developing exceptional talent, supporting our teammates’ physical,... ...Platform, which today includes on prem data warehouses, Informatica ETL, Hadoop ecosystems, mission critical mainframe processing, and legacy...Full timeWork at officeDay shift- ...Experience with processing large data sets using Hadoop, HDFS, Spark, Kafka, Flume or similar... ...Lily Enterprise.Creating and maintaining ETL processesKnowledgeable of best practices... ...along with team counterparts to develop solutions in an end-to-end framework on a...Work at officeFlexible hours2 days per week
- ...WiproContact: Surendra PenumalaCompany: SRI Tech SolutionsJob Title: Sr ETL DeveloperLocation - Dallas, TX ( DAY 1 onsite ) Prefer local... ...are seeking a highly skilled and experienced Senior ETL Tools Developer with deep expertise in Microsoft Azure Data Factory (ADF) and...Full timeLocal area
$116k - $170k
...Locust, JMeter)AWS AthenaAWS S3 (bucket policies), IAM roles and policiesSplunkAWS Secrets Manager and Parameter StoreAWS EMR (PySpark-based ETL)DockerKubernetesTerraformOpenSearch / Search Engine indexing, retrieval (RAG systems)Cache technologiesNice to Have (Willing...Full timeApprenticeshipImmediate startWork from homeWorldwide- ...SolutionsRole: Data Engineer - Lakehouse (PySpark / Databricks / Airflow / AWS)Location:... ...)Role OverviewResponsible for designing, developing, and maintaining data pipelines into a Lakehouse... ..., Build and optimize ETL/ELT pipelines using PySpark on Databricks...
$77.4k - $135.4k
...future. Summary: In this role, you will design, develop, and support scalable data and database... ...Azure SQL, Databricks, Apache Spark, Python, and PySpark. Develop and maintain scalable Extract, Transform, Load (ETL) and Extract, Load, Transform (ELT) processes to...Full time$125.76k - $188.64k
...Large Database (VLDB) Management: Managing, developing, and optimizing solutions for very large... ...for VLDB environments.Data Loading & ETL Leadership: Overseeing and guiding robust... ...verbal communication.Nice to Have:Python/PySpark skills for data processing and analytics....Full time- Job DescriptionSenior Hadoop Developer/Analyst6+ MonthsDallas, TX.Need GC and USCNeed LocalsThe Senior Hadoop Developer/Analyst is responsible... ...HDFS, Hive, Pig, Flume, HBase, Spark, Impala and Hadoop ETL development via tools such as Informatica,• Translate, load and...Work experience placementLocal area
- ...Execution: Directly execute the migration of legacy ETL and microservices to AWS. This includes refactoring monolithic... ...Connect and AWS Cloud Map.Data Pipeline Engineering: Develop end-to-end data flows using AWS Glue (PySpark), Amazon EMR, and Snowflake. Implement "Lakehouse"...
$156.16k - $234.24k
...manage performance, capacity, and hirinHands-on Apache Spark (PySpark/Scala) development; Spark on Kubernetes (EKS); data lake/warehouse... ..., Glue, Snowflake, Databricks; MWAA orchestrationLead on-prem (Hadoop, Teradata, RDBMS) to cloud migrations; secure hybrid...Full time- ...Integration : Design and implement ETL/ELT processes to ingest, process, and transform... ...to understand business requirements and develop data solutions. Manage data... ...Technical Skills : Proficiency in Spark, PySpark, Scala, or SQL for large-scale data processing...
- ...seeking an experienced and highly motivated Java Spark Lead Developer to join our dynamic data engineering team. The ideal... ...Java.Design and develop scalable data pipelines and ETL workflows using Spark and Hadoop ecosystem tools.Develop RESTful APIs using Spring Boot...Full time
$112.2k - $202.6k
...individual with the ability to design and develop data integrations to deliver Enterprise... ...as RDBMS, MPP platforms (e.g., Teradata), Hadoop data lakes, Azure Cloud, and Azure Databricks... ...like HDFS, Hive, HQL, Spark, Scala, Pyspark/python, Sqoop.5 + years of experience working...Full timeFlexible hours$112.7k - $193.2k
...retrieval stores, embeddings pipelines, and ETL/ELT jobs on Spark/Databricks; design analytics models and rules enginesDefine and develop APIs for integrations across the... ...scaling)4+ years with big data & streaming (Hadoop, MapReduce/HDFS, Spark, Kafka); Docker/Kubernetes...Minimum wageFull timeContract workWork experience placementWork at officeLocal areaRemote work- ...IT - Technology Lead | Big Data - Hadoop | Hadoop Job Description: IT - Technology Lead | Big Data - Hadoop | Hadoop Work Location: Richardson, TX 75081 Preferred Rate: 70 Contract duration (in months): 6 Target Start Date: 13 Feb 2017 Mandatory Job...Contract work
- ...Hadoop Developer W2 Candidates - with minimum validity of 12 months Addison, TX- Look for nearby candidates Onsite/hybrid: 5 Days Onsite Years of Experience required : 7+ Yrs Must Have skills: Hadoop ,Spark,(Scala or Python) Job...Work experience placement
- ...WiproContact: Meghana GorusuCompany: SRI Tech SolutionsSenior Informatica ETL DeveloperLocation - Dallas, TX / Cleveland, OH / Pittsburgh, PA... ...:We are seeking a highly skilled Senior Informatica ETL Developer to join our data engineering team. In this role, you will...Full time
- ...Apache Iceberg, Delta Lake, and OLAP SQL Engines), you will build ETL pipelines that integrate data from shop floor systems spanning... ...production metrics, electrical test, etc.Build ETL/ELT pipelines using PySpark and SQL to load data into Iceberg and Delta Lake tables with...Local area
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Hadoop / ETL Developer and PySpark. Be the first to apply!
- interactive developer Dallas, TX
- developer relations Dallas, TX
- developer contractor Dallas, TX
- nosql developer Dallas, TX
- progress developer Dallas, TX
- entry level mulesoft developer Dallas, TX
- sas programmer Dallas, TX
- jira workflow developer Dallas, TX
- remedy developer Dallas, TX
- junior mulesoft developer Dallas, TX

