Hadoop / ETL Developer and PySpark
SRI Tech
Job ID: SRJ_7721Posted: 2025-10-10Location: Dallas, TX; Atlanta, GA; Cleveland, OH; Pittsburgh, PASalary: USD 98-100 / Yearly (Full Time)Employment Type: Full TimeIndustry: Computer SoftwareClient: WiproContact: Meghana GorusuCompany: SRI Tech SolutionsHadoop / ETL Developer and PySparkLocation – Dallas, TX / Atlanta, GA / Cleveland, OH / Pittsburgh, PA (Day 1 Onsite)Note – Pls submit only visa independent candidatesEmployment Type – FTEJob Description:We are seeking a highly experienced Senior Big Data & DevOps Engineer with 8+ years of professional experience in HDFS, Hive, Impala, PySpark, Python, and DevOps automation tools such as uDeploy and Jenkins. This role is responsible for managing end-to-end data operations, including HDFS table management, ETL pipeline development, multi-environment codebase governance, platform upgrades, and production support. The ideal candidate will have strong expertise in Linux system operations, Big Data ecosystem tools, and experience with incident/change management using ServiceNow. This role plays a key part in ensuring the stability, scalability, and efficiency of enterprise data platforms while enabling seamless development-to-production workflows.Key Responsibilities: Big Data Platform Operations
- Design, manage, and optimize HDFS directories, tables, and partitioning strategies.
- Implement and enforce data retention and lifecycle policies across large datasets.
- Administer Hive and Impala environments, ensuring high availability, performance tuning, and security compliance.
- Develop scalable ETL pipelines using PySpark, Hive, and Python.
- Build reusable frameworks for data ingestion, transformation, and aggregation.
- Optimize job performance through query tuning, resource management, and parallelization.DevOps & Environment Management
- Maintain and promote code across DEV, QA, UAT, and PROD environments.
- Develop and support CI/CD pipelines using Jenkins and uDeploy for automated deployments.
- Perform environment upgrades, patching, and dependency management aligned with release schedules.Linux & Infrastructure Operations
- Execute Linux administration tasks including performance tuning, disk management, and scripting (Bash/Python).
- Troubleshoot cluster-level issues including node failures, job errors, and distributed system anomalies.Change & Incident Management
- Drive incident resolution and change execution using ServiceNow workflows.
- Conduct root cause analysis (RCA) for critical issues and implement preventive solutions.
- Ensure compliance with ITIL processes for change, incident, and problem management.
- Partner with data engineers, developers, DevOps teams, and business analysts to ensure operational excellence.
- Mentor junior engineers and contribute to technical leadership across the Big Data ecosystem.
- Document operational procedures, troubleshooting guides, and architectural decisions for internal knowledge sharing.Required Qualifications:
- Bachelor’s degree in Computer Science, Information Technology, or related field.
- 8+ years of experience in Big Data engineering and DevOps practices.
- Advanced proficiency in HDFS, Hive, Impala, PySpark, Python, and Linux.
- Proven experience with CI/CD tools such as Jenkins and uDeploy.
- Strong understanding of ETL development, orchestration, and performance optimization.
- Experience with ServiceNow for incident/change/problem management.
- Excellent analytical, troubleshooting, and communication skills.Nice to Have:
- Exposure to cloud-based Big Data platforms (AWS EMR).
- Familiarity with containerization (Docker, Kubernetes) and infrastructure automation tools (Ansible, Terraform).SRI Tech Solutions is an equal opportunity employer and does not discriminate on the basis of race, color, gender, religion, age, sexual orientation, national origin or citizenship status or ethnic origin, disability, marital status, veteran status, or any other occupationally irrelevant criteria.
- ...hands-on contributor who can develop, implement, and optimize robust... ...time data solutions within our Hadoop ecosystem and streaming... ...hands-on knowledge of Python, PySpark, Unix, and SQL. Familiarity with... ...developing and implementing complex ETL/ELT processes for both batch...Suggested
$86.25k - $158.13k
...contribute to the company’s success. As a Software Engineer Lead (ETL) within PNC's Retail Tech Core Debit Product organization, you... ...warehousing solutions using platforms such as Teradata, Oracle, Hadoop, and mainframe scheduling tools (CA7). Leads end-to-end solution...SuggestedFull timeTemporary workPart timeWork experience placementWork at office- ...a bright future for NTT DATA Services and for the people who work here.NTT DATA Services currently seeks a Java Full Stack Developer with Hadoop to join our team in Dallas, Texas (US-TX), United States (US).Ntt data is hiring a backend Java development lead to join our...SuggestedFor contractors
- ...Description:-At least 2 years of experience in Big Data space.Strong Hadoop - MAP REDUCE/Hive/Pig/SQOOP/OOZIE - MUSTCandidate should have... ...with Apache Parquet Data formatPast experience and exposure to ETL and data warehouse projectsExperience with Spark, FlumeCloudera/...SuggestedPermanent employmentFull timeH1b
$115k - $175k
...ETL Developer – Remote Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data... ...or Google BigQuery. Knowledge of Apache Spark, Databricks, PySpark, or distributed data processing frameworks. Experience...SuggestedFull timeH1bLocal areaImmediate startRemote workVisa sponsorship- ...ETL Developer Location: Strongsville OH/Pittsburgh, PA/Dallas, TX Duration: Full Time Job Description: ~ Informatica powercenter... ..., Migration project, Unix, Data Quality Rules, Oracle, PLSQL, Hadoop. Roles & Responsibilities Looking for a resource...Full time
- Key Responsibilities1. Backend & Microservices DevelopmentDesign, develop, and implement scalable, resilient microservices using Java and... ...application performance for low latency and high throughput.2. ETL & Data Pipeline EngineeringArchitect and maintain complex ETL/ELT...
- Key Responsibilities1. Backend & Microservices DevelopmentDesign, develop, and implement scalable, resilient microservices using Java and... ...application performance for low latency and high throughput.2. ETL & Data Pipeline EngineeringArchitect and maintain complex ETL/ELT...
- ...contribute to our shared success. This includes attracting and developing exceptional talent, recognizing and rewarding performance, and... ...Summary:Seeking a software engineer with experience in experience in ETL/Data Engineering, Mainframe, and real-time data replication...Full timeWork at officeFlexible hoursShift workDay shift
$103.33k - $128.66k
...Data / Database Engineer - SQL, Oracle, PL/SQL, Informatica, ETL, Powercenter Job Location - Westlake TX (Day One Onsite - Hybrid... ...solutions that drive business insights. Key Responsibilities Develop, optimize, and maintain complex SQL and PL/SQL queries, stored...Permanent employmentFull timeContract workLocal areaFlexible hours- ...to a diverse set of customers across various industries in the United States.Job DescriptionDescription:Design and development on Hadoop software ecosystem and development on MapReduce, HBase, Hive, Pig,Programming in Spark/StormProgramming in distributed messaging system...
- ...WiproContact: Meghana GorusuCompany: SRI Tech SolutionsRole: Senior PySpark Developer / EngineerLocation: Dallas, TX (3 Days onsite/ week)Job... ...for PySpark integrationETL Knowledge:Hands-on experience with ETL processes using any programming language or frameworkAbility to...Full time3 days per week
$112.7k - $193.2k
...Big Data and cloud solutions Develop scalable microservices and... ...stores, embeddings pipelines, and ETL/ELT jobs on Spark/Databricks;... ...) with Scala, Python, PySpark ~4+ years hands on with Databricks... ...years with big data & streaming (Hadoop, MapReduce/HDFS, Spark, Kafka)...Minimum wageFull timeContract workWork experience placementWork at officeLocal areaRemote work$125.76k - $188.64k
...We are seeking a Senior Data Developer to join one of our small, co-located... ...data solutions within our Hadoop ecosystem and streaming... ...architectural knowledge of Python, PySpark, Unix, and SQL.Familiarity with... ...designing and implementing complex ETL/ELT processes for both batch...Full time$156.5k - $230k
...our commitment to being an inclusive workplace, attracting and developing exceptional talent, supporting our teammates’ physical,... ...Platform, which today includes on prem data warehouses, Informatica ETL, Hadoop ecosystems, mission critical mainframe processing, and legacy...Full timeWork at officeDay shift- ...Experience with processing large data sets using Hadoop, HDFS, Spark, Kafka, Flume or similar... ...Lily Enterprise.Creating and maintaining ETL processesKnowledgeable of best practices... ...along with team counterparts to develop solutions in an end-to-end framework on a...Work at officeFlexible hours2 days per week
- ...Duties: Research, design, and develop computer and network software... ...each of the following: Build ETL pipelines to extract data from... ...data technologies such as Apache Hadoop, Apache Spark, Apache Hive,... ...using Spark, Scala, Java, Python, Pyspark, SQL, databricks, shell...Full timeLocal areaRemote workFlexible hours
- ...specifications, and aid in creation of process documentation. Develop, refine, and document code for new software applications, following... ....Technical/Domain Skill 1Technology|Big Data - Data Processing|PySpark Technical/Domain Skill 2Technology|Cloud Platform|AWS Data...Full timeTemporary workRelocation
- ...WiproContact: Surendra PenumalaCompany: SRI Tech SolutionsJob Title: Sr ETL DeveloperLocation - Dallas, TX ( DAY 1 onsite ) Prefer local... ...are seeking a highly skilled and experienced Senior ETL Tools Developer with deep expertise in Microsoft Azure Data Factory (ADF) and...Full timeLocal area
- ...coordinate bug fixes to uphold the software quality standards Develop user training programs, documentation, and support frameworks to... ....Technical/Domain Skill 1Technology|Big Data - Data Processing|PySpark Technical/Domain Skill 2Technology|Cloud Platform|AWS Database...Full timeTemporary workRelocation
$116k - $170k
...Locust, JMeter)AWS AthenaAWS S3 (bucket policies), IAM roles and policiesSplunkAWS Secrets Manager and Parameter StoreAWS EMR (PySpark-based ETL)DockerKubernetesTerraformOpenSearch / Search Engine indexing, retrieval (RAG systems)Cache technologiesNice to Have (Willing...Full timeApprenticeshipImmediate startWork from homeWorldwide- ...SolutionsRole: Data Engineer - Lakehouse (PySpark / Databricks / Airflow / AWS)Location:... ...)Role OverviewResponsible for designing, developing, and maintaining data pipelines into a Lakehouse... ..., Build and optimize ETL/ELT pipelines using PySpark on Databricks...
- Job DescriptionSenior Hadoop Developer/Analyst6+ MonthsDallas, TX.Need GC and USCNeed LocalsThe Senior Hadoop Developer/Analyst is responsible... ...HDFS, Hive, Pig, Flume, HBase, Spark, Impala and Hadoop ETL development via tools such as Informatica,• Translate, load and...Work experience placementLocal area
- ...Execution: Directly execute the migration of legacy ETL and microservices to AWS. This includes refactoring monolithic... ...Connect and AWS Cloud Map.Data Pipeline Engineering: Develop end-to-end data flows using AWS Glue (PySpark), Amazon EMR, and Snowflake. Implement "Lakehouse"...
$156.16k - $234.24k
...manage performance, capacity, and hirinHands-on Apache Spark (PySpark/Scala) development; Spark on Kubernetes (EKS); data lake/warehouse... ..., Glue, Snowflake, Databricks; MWAA orchestrationLead on-prem (Hadoop, Teradata, RDBMS) to cloud migrations; secure hybrid...Full time- ...seeking an experienced and highly motivated Java Spark Lead Developer to join our dynamic data engineering team. The ideal... ...Java.Design and develop scalable data pipelines and ETL workflows using Spark and Hadoop ecosystem tools.Develop RESTful APIs using Spring Boot...Full time
- ...Integration : Design and implement ETL/ELT processes to ingest, process, and transform... ...to understand business requirements and develop data solutions. Manage data... ...Technical Skills : Proficiency in Spark, PySpark, Scala, or SQL for large-scale data processing...
$112.2k - $202.6k
...individual with the ability to design and develop data integrations to deliver Enterprise... ...as RDBMS, MPP platforms (e.g., Teradata), Hadoop data lakes, Azure Cloud, and Azure Databricks... ...like HDFS, Hive, HQL, Spark, Scala, Pyspark/python, Sqoop.5 + years of experience working...Full timeFlexible hours- ...platform with Design and automate scalable ETL/ELT data pipelines using Microsoft Fabric... ...matter experts, and decision-makers to develop scope of work, success criteria and create... ...and "Medallion" architectures using PySpark, SQL, and Fabric Notebooks to transform raw...
- ...WiproContact: Meghana GorusuCompany: SRI Tech SolutionsSenior Informatica ETL DeveloperLocation - Dallas, TX / Cleveland, OH / Pittsburgh, PA... ...:We are seeking a highly skilled Senior Informatica ETL Developer to join our data engineering team. In this role, you will...Full time
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Hadoop / ETL Developer and PySpark. Be the first to apply!
- interactive developer Dallas, TX
- developer relations Dallas, TX
- developer contractor Dallas, TX
- nosql developer Dallas, TX
- progress developer Dallas, TX
- entry level mulesoft developer Dallas, TX
- sas programmer Dallas, TX
- jira workflow developer Dallas, TX
- remedy developer Dallas, TX
- junior mulesoft developer Dallas, TX

