Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Hadoop / ETL Developer and PySpark

SRI Tech

Job ID: SRJ_7721Posted: 2025-10-10Location: Dallas, TX; Atlanta, GA; Cleveland, OH; Pittsburgh, PASalary: USD 98-100 / Yearly (Full Time)Employment Type: Full TimeIndustry: Computer SoftwareClient: WiproContact: Meghana GorusuCompany: SRI Tech SolutionsHadoop / ETL Developer and PySparkLocation – Dallas, TX / Atlanta, GA / Cleveland, OH / Pittsburgh, PA (Day 1 Onsite)Note – Pls submit only visa independent candidatesEmployment Type – FTEJob Description:We are seeking a highly experienced Senior Big Data & DevOps Engineer with 8+ years of professional experience in HDFS, Hive, Impala, PySpark, Python, and DevOps automation tools such as uDeploy and Jenkins. This role is responsible for managing end-to-end data operations, including HDFS table management, ETL pipeline development, multi-environment codebase governance, platform upgrades, and production support. The ideal candidate will have strong expertise in Linux system operations, Big Data ecosystem tools, and experience with incident/change management using ServiceNow. This role plays a key part in ensuring the stability, scalability, and efficiency of enterprise data platforms while enabling seamless development-to-production workflows.Key Responsibilities: Big Data Platform Operations

  • Design, manage, and optimize HDFS directories, tables, and partitioning strategies.
  • Implement and enforce data retention and lifecycle policies across large datasets.
  • Administer Hive and Impala environments, ensuring high availability, performance tuning, and security compliance.
ETL Development & Data Engineering
  • Develop scalable ETL pipelines using PySpark, Hive, and Python.
  • Build reusable frameworks for data ingestion, transformation, and aggregation.
  • Optimize job performance through query tuning, resource management, and parallelization.DevOps & Environment Management
  • Maintain and promote code across DEV, QA, UAT, and PROD environments.
  • Develop and support CI/CD pipelines using Jenkins and uDeploy for automated deployments.
  • Perform environment upgrades, patching, and dependency management aligned with release schedules.Linux & Infrastructure Operations
  • Execute Linux administration tasks including performance tuning, disk management, and scripting (Bash/Python).
  • Troubleshoot cluster-level issues including node failures, job errors, and distributed system anomalies.Change & Incident Management
  • Drive incident resolution and change execution using ServiceNow workflows.
  • Conduct root cause analysis (RCA) for critical issues and implement preventive solutions.
  • Ensure compliance with ITIL processes for change, incident, and problem management.
Collaboration & Technical Leadership
  • Partner with data engineers, developers, DevOps teams, and business analysts to ensure operational excellence.
  • Mentor junior engineers and contribute to technical leadership across the Big Data ecosystem.
  • Document operational procedures, troubleshooting guides, and architectural decisions for internal knowledge sharing.Required Qualifications:
  • Bachelor’s degree in Computer Science, Information Technology, or related field.
  • 8+ years of experience in Big Data engineering and DevOps practices.
  • Advanced proficiency in HDFS, Hive, Impala, PySpark, Python, and Linux.
  • Proven experience with CI/CD tools such as Jenkins and uDeploy.
  • Strong understanding of ETL development, orchestration, and performance optimization.
  • Experience with ServiceNow for incident/change/problem management.
  • Excellent analytical, troubleshooting, and communication skills.Nice to Have:
  • Exposure to cloud-based Big Data platforms (AWS EMR).
  • Familiarity with containerization (Docker, Kubernetes) and infrastructure automation tools (Ansible, Terraform).SRI Tech Solutions is an equal opportunity employer and does not discriminate on the basis of race, color, gender, religion, age, sexual orientation, national origin or citizenship status or ethnic origin, disability, marital status, veteran status, or any other occupationally irrelevant criteria.

Vacancy posted 4 days ago
Similar jobs that could be interesting for youBased on the Hadoop / ETL Developer and PySpark in Dallas, TX vacancy
  •  ...hands-on contributor who can develop, implement, and optimize robust...  ...time data solutions within our Hadoop ecosystem and streaming...  ...hands-on knowledge of Python, PySpark, Unix, and SQL. Familiarity with...  ...developing and implementing complex ETL/ELT processes for both batch... 
    Suggested

    Citigroup Inc

    Irving, TX
    3 days ago
  • $86.25k - $158.13k

     ...contribute to the company’s success. As a Software Engineer Lead (ETL) within PNC's Retail Tech Core Debit Product organization, you...  ...warehousing solutions using platforms such as Teradata, Oracle, Hadoop, and mainframe scheduling tools (CA7). Leads end-to-end solution... 
    Suggested
    Full time
    Temporary work
    Part time
    Work experience placement
    Work at office

    PNC

    Dallas, TX
    3 days ago
  •  ...a bright future for NTT DATA Services and for the people who work here.NTT DATA Services currently seeks a Java Full Stack Developer with Hadoop to join our team in Dallas, Texas (US-TX), United States (US).Ntt data is hiring a backend Java development lead to join our... 
    Suggested
    For contractors

    NTT DATA

    Dallas, TX
    2 days ago
  •  ...Description:-At least 2 years of experience in Big Data space.Strong Hadoop - MAP REDUCE/Hive/Pig/SQOOP/OOZIE - MUSTCandidate should have...  ...with Apache Parquet Data formatPast experience and exposure to ETL and data warehouse projectsExperience with Spark, FlumeCloudera/... 
    Suggested
    Permanent employment
    Full time
    H1b

    Sonsoft

    Irving, TX
    4 days ago
  • $115k - $175k

     ...ETL Developer – Remote Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data...  ...or Google BigQuery. Knowledge of Apache Spark, Databricks, PySpark, or distributed data processing frameworks. Experience... 
    Suggested
    Full time
    H1b
    Local area
    Immediate start
    Remote work
    Visa sponsorship

    Bright Vision Technologies

    Richardson, TX
    1 day ago
  •  ...ETL Developer Location: Strongsville OH/Pittsburgh, PA/Dallas, TX Duration: Full Time Job Description: ~ Informatica powercenter...  ..., Migration project, Unix, Data Quality Rules, Oracle, PLSQL, Hadoop. Roles & Responsibilities Looking for a resource... 
    Full time

    JConnect Infotech

    Dallas, TX
    3 days ago
  • Key Responsibilities1. Backend & Microservices DevelopmentDesign, develop, and implement scalable, resilient microservices using Java and...  ...application performance for low latency and high throughput.2. ETL & Data Pipeline EngineeringArchitect and maintain complex ETL/ELT... 

    Goldman Sachs

    Dallas, TX
    3 days ago
  • Key Responsibilities1. Backend & Microservices DevelopmentDesign, develop, and implement scalable, resilient microservices using Java and...  ...application performance for low latency and high throughput.2. ETL & Data Pipeline EngineeringArchitect and maintain complex ETL/ELT... 

    Goldman Sachs

    Dallas, TX
    2 days ago
  •  ...contribute to our shared success. This includes attracting and developing exceptional talent, recognizing and rewarding performance, and...  ...Summary:Seeking a software engineer with experience in experience in ETL/Data Engineering, Mainframe, and real-time data replication... 
    Full time
    Work at office
    Flexible hours
    Shift work
    Day shift

    Bank of America

    Addison, TX
    1 day ago
  • $103.33k - $128.66k

     ...Data / Database Engineer - SQL, Oracle, PL/SQL, Informatica, ETL, Powercenter Job Location - Westlake TX (Day One Onsite - Hybrid...  ...solutions that drive business insights. Key Responsibilities Develop, optimize, and maintain complex SQL and PL/SQL queries, stored... 
    Permanent employment
    Full time
    Contract work
    Local area
    Flexible hours

    Capgemini

    Irving, TX
    3 days ago
  •  ...to a diverse set of customers across various industries in the United States.Job DescriptionDescription:Design and development on Hadoop software ecosystem and development on MapReduce, HBase, Hive, Pig,Programming in Spark/StormProgramming in distributed messaging system... 

    LanceSoft

    Irving, TX
    4 days ago
  •  ...WiproContact: Meghana GorusuCompany: SRI Tech SolutionsRole: Senior PySpark Developer / EngineerLocation: Dallas, TX (3 Days onsite/ week)Job...  ...for PySpark integrationETL Knowledge:Hands-on experience with ETL processes using any programming language or frameworkAbility to... 
    Full time
    3 days per week

    SRI Tech

    Dallas, TX
    2 days ago
  • $112.7k - $193.2k

     ...Big Data and cloud solutions Develop scalable microservices and...  ...stores, embeddings pipelines, and ETL/ELT jobs on Spark/Databricks;...  ...) with Scala, Python, PySpark ~4+ years hands on with Databricks...  ...years with big data & streaming (Hadoop, MapReduce/HDFS, Spark, Kafka)... 
    Minimum wage
    Full time
    Contract work
    Work experience placement
    Work at office
    Local area
    Remote work

    Optum

    Richardson, TX
    2 days ago
  • $125.76k - $188.64k

     ...We are seeking a Senior Data Developer to join one of our small, co-located...  ...data solutions within our Hadoop ecosystem and streaming...  ...architectural knowledge of Python, PySpark, Unix, and SQL.Familiarity with...  ...designing and implementing complex ETL/ELT processes for both batch... 
    Full time

    Citigroup

    Irving, TX
    2 days ago
  • $156.5k - $230k

     ...our commitment to being an inclusive workplace, attracting and developing exceptional talent, supporting our teammates’ physical,...  ...Platform, which today includes on prem data warehouses, Informatica ETL, Hadoop ecosystems, mission critical mainframe processing, and legacy... 
    Full time
    Work at office
    Day shift

    Bank of America

    Addison, TX
    3 days ago
  •  ...Experience with processing large data sets using Hadoop, HDFS, Spark, Kafka, Flume or similar...  ...Lily Enterprise.Creating and maintaining ETL processesKnowledgeable of best practices...  ...along with team counterparts to develop solutions in an end-to-end framework on a... 
    Work at office
    Flexible hours
    2 days per week

    GM Financial

    Irving, TX
    4 days ago
  •  ...Duties: Research, design, and develop computer and network software...  ...each of the following: Build ETL pipelines to extract data from...  ...data technologies such as Apache Hadoop, Apache Spark, Apache Hive,...  ...using Spark, Scala, Java, Python, Pyspark, SQL, databricks, shell... 
    Full time
    Local area
    Remote work
    Flexible hours

    Publicis Groupe Holdings B.V

    Irving, TX
    4 days ago
  •  ...specifications, and aid in creation of process documentation. Develop, refine, and document code for new software applications, following...  ....Technical/Domain Skill 1Technology|Big Data - Data Processing|PySpark Technical/Domain Skill 2Technology|Cloud Platform|AWS Data... 
    Full time
    Temporary work
    Relocation

    Infosys Technologies

    Richardson, TX
    1 day ago
  •  ...WiproContact: Surendra PenumalaCompany: SRI Tech SolutionsJob Title: Sr ETL DeveloperLocation - Dallas, TX ( DAY 1 onsite ) Prefer local...  ...are seeking a highly skilled and experienced Senior ETL Tools Developer with deep expertise in Microsoft Azure Data Factory (ADF) and... 
    Full time
    Local area

    SRI Tech

    Dallas, TX
    4 days ago
  •  ...coordinate bug fixes to uphold the software quality standards Develop user training programs, documentation, and support frameworks to...  ....Technical/Domain Skill 1Technology|Big Data - Data Processing|PySpark Technical/Domain Skill 2Technology|Cloud Platform|AWS Database... 
    Full time
    Temporary work
    Relocation

    Infosys Technologies

    Richardson, TX
    1 day ago
  • $116k - $170k

     ...Locust, JMeter)AWS AthenaAWS S3 (bucket policies), IAM roles and policiesSplunkAWS Secrets Manager and Parameter StoreAWS EMR (PySpark-based ETL)DockerKubernetesTerraformOpenSearch / Search Engine indexing, retrieval (RAG systems)Cache technologiesNice to Have (Willing... 
    Full time
    Apprenticeship
    Immediate start
    Work from home
    Worldwide

    Gartner

    Irving, TX
    15 hours ago
  •  ...SolutionsRole: Data Engineer - Lakehouse (PySpark / Databricks / Airflow / AWS)Location:...  ...)Role OverviewResponsible for designing, developing, and maintaining data pipelines into a Lakehouse...  ..., Build and optimize ETL/ELT pipelines using PySpark on Databricks... 

    SRI Tech

    Dallas, TX
    4 days ago
  • Job DescriptionSenior Hadoop Developer/Analyst6+ MonthsDallas, TX.Need GC and USCNeed LocalsThe Senior Hadoop Developer/Analyst is responsible...  ...HDFS, Hive, Pig, Flume, HBase, Spark, Impala and Hadoop ETL development via tools such as Informatica,• Translate, load and... 
    Work experience placement
    Local area

    USM Systems

    Dallas, TX
    4 days ago
  •  ...Execution: Directly execute the migration of legacy ETL and microservices to AWS. This includes refactoring monolithic...  ...Connect and AWS Cloud Map.Data Pipeline Engineering: Develop end-to-end data flows using AWS Glue (PySpark), Amazon EMR, and Snowflake. Implement "Lakehouse"... 

    Goldman Sachs

    Dallas, TX
    3 days ago
  • $156.16k - $234.24k

     ...manage performance, capacity, and hirinHands-on Apache Spark (PySpark/Scala) development; Spark on Kubernetes (EKS); data lake/warehouse...  ..., Glue, Snowflake, Databricks; MWAA orchestrationLead on-prem (Hadoop, Teradata, RDBMS) to cloud migrations; secure hybrid... 
    Full time

    Citigroup

    Irving, TX
    1 day ago
  •  ...seeking an experienced and highly motivated Java Spark Lead Developer to join our dynamic data engineering team. The ideal...  ...Java.Design and develop scalable data pipelines and ETL workflows using Spark and Hadoop ecosystem tools.Develop RESTful APIs using Spring Boot... 
    Full time

    SRI Tech

    Irving, TX
    2 days ago
  •  ...Integration : Design and implement ETL/ELT processes to ingest, process, and transform...  ...to understand business requirements and develop data solutions. Manage data...  ...Technical Skills : Proficiency in Spark, PySpark, Scala, or SQL for large-scale data processing... 

    Cedent Consulting

    Dallas, TX
    4 days ago
  • $112.2k - $202.6k

     ...individual with the ability to design and develop data integrations to deliver Enterprise...  ...as RDBMS, MPP platforms (e.g., Teradata), Hadoop data lakes, Azure Cloud, and Azure Databricks...  ...like HDFS, Hive, HQL, Spark, Scala, Pyspark/python, Sqoop.5 + years of experience working... 
    Full time
    Flexible hours

    Health Care Service

    Richardson, TX
    2 days ago
  •  ...platform with Design and automate scalable ETL/ELT data pipelines using Microsoft Fabric...  ...matter experts, and decision-makers to develop scope of work, success criteria and create...  ...and "Medallion" architectures using PySpark, SQL, and Fabric Notebooks to transform raw... 

    Omni Hotels

    Dallas, TX
    19 hours ago
  •  ...WiproContact: Meghana GorusuCompany: SRI Tech SolutionsSenior Informatica ETL DeveloperLocation - Dallas, TX / Cleveland, OH / Pittsburgh, PA...  ...:We are seeking a highly skilled Senior Informatica ETL Developer to join our data engineering team. In this role, you will... 
    Full time

    SRI Tech

    Dallas, TX
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Hadoop / ETL Developer and PySpark. Be the first to apply!