Data Engineer (PySpark)
Capgemini
Choosing Capgemini means choosing a company where you will be empowered to shape your career in the way you'd like, where you'll be supported and inspired by a collaborative community of colleagues around the world, and where you'll be able to reimagine what's possible. Join us and help the world's leading organizations unlock the value of technology and build a more sustainable, more inclusive world.
Job Description
Responsibilities
Monitor and control all phases of development process and analysis design construction testing and implementation as well as provide user and operational support on applications to business users
Utilize indepth specialty knowledge of applications development to analyze complex problemsissues provide evaluation of business process system process and industry standards and make evaluative judgement
Recommend and develop security measures in post implementation analysis of business usage to ensure successful system design and functionality
Consult with usersclients and other technology groups on issues recommend advanced programming solutions and install and assist customer exposure systems
Ensure essential procedures are followed and help define operating standards and processes
Serve as advisor or coach to new or lower level analysts
Has the ability to operate with a limited level of direct supervision
Can exercise independence of judgement and autonomy
Acts as SME to senior stakeholders and or other team members
Experience in systems analysis and programming of software applications
Experience in managing and implementing successful projects
Working knowledge of consultingproject management techniquesmethods
Ability to work under pressure and manage deadlines or unexpected changes in expectations or requirements Education Bachelors degreeUniversity degree or equivalent experience This job description provides a highlevel review of the types of work performed Other jobrelated duties may be assigned as required Indepth understanding of HDFS architecture data storage and fault tolerance mechanisms Experience with HDFS commands and administration Solid understanding of YARN resource management and job scheduling Fundamental understanding of MapReduce programming paradigm even if primary development is in SparkFlinkKnowledge of Zookeeper for distributed coordination services StrongproficiencyinSpark Core Spark SQL Spark Streaming and SparkGraphXbeneficialExpertlevel programming skills inPython and Pyspark specifically for developing Spark applications
Experience with Spark performance optimization techniques eg caching partitioning shuffle optimizations memory management
Experience withPySparkfor data processingFamiliarity with data manipulation libraries Pandas NumPyScripting for automation and data orchestration
Complex query writing subqueries window functions and performance tuningHBase for realtime access to large datasets within HadoopCassandra MongoDB or similarFamiliarity with RDBMS concepts and SQL for data integration
Understanding of dimensional modeling fact and dimension tables starsnowflake schemas' The base compensation range for this role in the posted location is: 80,000 - 90,000 Capgemini provides compensation range information in accordance with applicable national, state, provincial, and local pay transparency laws. The base compensation range listed for this position reflects the minimum and maximum target compensation Capgemini, in good faith, believes it may pay for the role at the time of this posting. This range may be subject to change as permitted by law. The actual compensation offered to any candidate may fall outside of the posted range and will be determined based on multiple factors legally permitted in the applicable jurisdiction. These may include, but are not limited to: Geographic location, Education and qualifications, Certifications and licenses, Relevant experience and skills, Seniority and performance, Market and business consideration, Internal pay equity. It is not typical for candidates to be hired at or near the top of the posted compensation range. In addition to base salary, this role may be eligible for additional compensation such as variable incentives, bonuses, or commissions, depending on the position and applicable laws.
Capgemini offers a comprehensive, non-negotiable benefits package to all regular, full-time employees. In the U.S. and Canada, available benefits are determined by local policy and eligibility and may include:
- Paid time off based on employee grade (A-F), defined by policy: Vacation: 12-25 days, depending on grade, Company paid holidays, Personal Days, Sick Leave
- Medical, dental, and vision coverage (or provincial healthcare coordination in Canada)
- Retirement savings plans (e.g., 401(k) in the U.S., RRSP in Canada)
- Life and disability insurance
- Employee assistance programs
- Other benefits as provided by local policy and eligibility
Important Notice: Compensation (including bonuses, commissions, or other forms of incentive pay) is not considered earned, vested, or payable until it becomes due under the terms of applicable plans or agreements and is subject to Capgemini's discretion, consistent with applicable laws. The Company reserves the right to amend or withdraw compensation programs at any time, within the limits of applicable legislation.
Disclaimers Capgemini is an Equal Opportunity Employer encouraging inclusion in the workplace. Capgemini also participates in the Partnership Accreditation in Indigenous Relations (PAIR) program which supports meaningful engagement with Indigenous communities across Canada by promoting fairness, accessibility, inclusion and respect. We value the rich cultural heritage and contributions of Indigenous Peoples and actively work to create a welcoming and respectful environment. All qualified applicants will receive consideration for employment without regard to race, national origin, gender identity/expression, age, religion, disability, sexual orientation, genetics, veteran status, marital status or any other characteristic protected by law. This is a general description of the Duties, Responsibilities and Qualifications required for this position. Physical, mental, sensory or environmental demands may be referenced in an attempt to communicate the manner in which this position traditionally is performed. Whenever necessary to provide individuals with disabilities an equal employment opportunity, Capgemini will consider reasonable accommodations that might involve varying job requirements and/or changing the way this job is performed, provided that such accommodation does not pose an undue hardship. Capgemini is committed to providing reasonable accommodation during our recruitment process. If you need assistance or accommodation, please reach out to your recruiting contact. Please be aware that Capgemini may capture your image (video or screenshot) during the interview process and that image may be used for verification, including during the hiring and onboarding process. Click the following link for more information on your rights as an Applicant in the United States.
Capgemini is a global business and technology transformation partner, helping organizations to accelerate their dual transition to a digital and sustainable world, while creating tangible impact for enterprises and society. It is a responsible and diverse group of 340,000 team members in more than 50 countries. With its strong over 55-year heritage, Capgemini is trusted by its clients to unlock the value of technology to address the entire breadth of their business needs. It delivers end-to-end services and solutions leveraging strengths from strategy and design to engineering, all fueled by its market leading capabilities in AI, generative AI, cloud and data, combined with its deep industry expertise and partner ecosystem.
$60k - $135k
Job Description Role: Data Engineer - PySpark Location: Irving, TX (3 Days onsite/week) Key Responsibilities Strong programming skills in PySpark in a BigData environment with exceptional hands-on capabilities Familiarity with big data processing tools and techniques. Experience...SuggestedMinimum wageLocal area3 days per week- ...requiredIndepth understanding of HDFS architecture data storage and fault tolerance mechanisms... ...programming skills inPython and Pyspark specifically for developing Spark... ...leveraging strengths from strategy and design to engineering, all fueled by its market leading...SuggestedFull timeLocal area
- ...A technology services company is seeking a Senior Data Engineer to create scalable ETL pipelines and work with vast datasets. The role requires... ...Python and SQL skills, along with experience in tools like PySpark and AWS. This full-time position is remote, offering...SuggestedFull timeRemote work
- LTIMindtree is seeking a Sr. DATA Engineer to design and optimize scalable data solutions on AWS. You will implement PySpark-based ETL pipelines, manage data lakes and support analytics with Iceberg and Hive. The role requires strong AWS data services experience, scripting...Suggested
- LTIMindtree in Tampa, FL is seeking a PySpark Developer to design, build and optimize large-scale data pipelines using PySpark and Spark SQL. You will work with... ...with DataOps practices, and collaborate with data engineers, analysts and DevOps in an Agile environment....Suggested
- Wipro is seeking a Data Engineer specialized in PySpark to join our data engineering team in Irving, TX. You will design and implement robust data pipelines in a Big Data environment, leveraging Hadoop ecosystem tools, Hive, HDFS, Spark, and Scala. The role emphasizes...
- ...development, analysis, and implementation while providing user support to business users. The role requires deep knowledge of large-scale data processing with Spark, HDFS, YARN, and SQL-based data integration. The candidate should have strong programming, systems analysis,...
$76.73k - $127.88k
...wants to work in a collaborative environment? As an experienced Data Engineer, you will have the ability to share new ideas and collaborate... ...on experience with SQL and Python, including Snowflake and/or PySpark for scalable data processing and ELT.4+ years of experience...Local areaVisa sponsorship- Job Title: DATA Engineer Work LocationTampa FLJob Description:Mandatory Karat Interview Role Overview We are seeking a highly skilled and... ...candidate will have strong expertise in big data processing using PySpark and a deep understanding of data warehousing concepts...
$113.84k - $170.76k
...DesignDesign and implement scalable, fault-tolerant batch and real-time data processing pipelines.Develop robust data models and schema... ...a modern Data Lakehouse environment.ETL/ELT Transformation: Re-engineer existing stored procedures and complex legacy ETL jobs into...Full timeFlexible hours- ...Data Engineer Discover your future at Citi. Working at Citi is far more than just a job. A career with us means joining a team of more... ...Oracle, PostgreSQL Expertise in Unix shell scripting, Python, PySpark ETL expertise on AbInitio tool (EME, GDE, Co-op) bringing...
- ...Data Engineer (Python + React, Healthcare Domain)We are seeking a Data Engineer with strong experience in Python and React to build and maintain... ...with data processing frameworks and libraries (Pandas, PySpark, Airflow, SQLAlchemy, etc.)Basic to strong experience with React...
$118.9k - $162.3k
...Data Engineer At Accenture Federal Services, nothing matters more than helping the US federal government make the nation stronger and... ...in Python and familiarity with common data libraries (Pandas, PySpark, requests, etc.). Preferred Qualifications Strong experience...Live inWork at officeLocal area- ...Silverthorne Advisory Group is seeking a skilled and highly motivated Data Engineer to join an exciting and growing opportunity within the Defense... ...processes and technologies such as Databricks, Python, Spark, PySpark, Scala, JavaScript/JSON, SQL, and Jupyter Notebooks. Create...Work from homeFlexible hours
$118.9k - $162.3k
...Data EngineerAt Accenture Federal Services, nothing matters more than helping the US federal... ...forward!We are looking for a Data Engineer with strong hands-on experience designing... ...familiarity with common data libraries (Pandas, PySpark, requests, etc.).Preferred...Live inWork at officeLocal area- ...Overview: Job Title: Big Data Engineer Location: Atlanta, GA / Tampa, FL/ Dallas, TX Job Summary We are seeking an experienced... ..., develop, and maintain big data pipelines using Spark (PySpark/Scala), Hadoop, Kafka, and distributed computing frameworks....
$159.42k - $215.69k
...the lives of patients while transforming your career.Principal Data Engineer What you will doLet’s do this. Let’s change the world. In this... ...optimization strategies. Proficient on experience in Python, PySpark, SQL. Handon experience with bid data ETL performance tuning....Full timeFlexible hours- Job Title: Pyspark DeveloperWork Location : Tampa, FloridaDesign develop and maintain data pipelines using PySpark for largescale data processingPerform data transformations... ...root cause analysisWork closely with data engineers analysts QA and DevOps teams in Agile delivery...
$118.9k - $162.3k
...moves missions and the government forward! We are looking for a Data Engineer with strong hands‑on experience designing, developing, and... ...in Python and familiarity with common data libraries (Pandas, PySpark, requests, etc.). Preferred Qualifications Strong experience...Live inWork at officeLocal area$118.9k - $162.3k
...moves missions and the government forward! We are looking for a Data Engineer with strong hands‑on experience designing, developing, and... ...in Python and familiarity with common data libraries (Pandas, PySpark, requests, etc.). Preferred Qualifications Strong experience...Work at officeLocal area- ...with free counseling, legal, and financial services. NowHiring:Data Engineer We’relookingfora Data Engineer to lead the design, development... ...Azure Data Lake. Hands-on experience with Azure Databricks (PySpark, Python, or Scala); proficiency in Python, Spark, or Scala for...Full timeTemporary workPart timeRemote work
- Hone Health is seeking a Data Engineering Intern to join our growing data team. You will report to the Senior Director of Data, Analytics &... ...internship features a modern data stack (Microsoft Fabric, dbt, PySpark, SQL) and opportunities to build analytics‑ready datasets...Remote jobInternship
- ...description/tech stack: We are looking for a proficient Azure Data Engineer to design, build, and maintain scalable data pipelines and... ...and technologies including Python, SQL, Postgres, MongoDB, PySpark, Databricks, and Snowflake. You will collaborate with data...
- Job Summary Pyspark Developer - Tampa, Florida Responsibilities Design, develop and maintain data pipelines using PySpark for large‑scale data processing Perform data transformations... ...‑cause analysis Work closely with data engineers, analysts, QA and DevOps teams in Agile...Temporary workLocal area
- ...Lead Data Engineer Snowflake shop ETL processes Data migration modern data engineering practices, etc. Opentable formats like datalake and iceberg. Exposed to databricks pipeline Data mesh exposure. Must see the data frameworks they have worked with in the resume. Exposure...
- ...Role: Lead Streaming Data Engineer / Technical Lead Experience: 10+ years of overall experience with 5+ years in real-time streaming and distributed data processing solutions. Location: Tampa, FL (Preferable) Pay rate: $50/hr on W2 60/hr on C...
- Job Title: Sr. DATA EngineerWork Location : Tampa FLJob Description:Seeking a Senior professional with 3 to 5 years of experience in... ...in Spark jobs and data pipelines Collaborate with data engineers’ data scientists and stakeholders to translate business requirements...
$77.5k - $176k
Data Engineer, SeniorThe Opportunity: Ever-expanding technology like IoT, machine learning, and artificial intelligence means that there’s more structured and unstructured data available today than ever before. As a data engineer, you know that organizing data can yield...Full timeContract workPart timeWork at officeLocal areaRemote work- ...Data Engineer Employment Type: Full-Time, Mid-level CGS is seeking a passionate and driven Data Engineer to support a rapidly growing Data Analytics and Business Intelligence platform focused on providing solutions that empower our federal customers with the tools...Full timeFlexible hours
- ...Date Engineer Hillsborough County Sherrif Office 6-12 month contract to hire Onsite Tampa, FL... ...It's better to think of this job as more Date Engineer and some Data Analyst. Major focus is designing and implementing data pipelines...Contract workWork at office
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Data Engineer (PySpark). Be the first to apply!
- aws data engineer Tampa, FL
- data science developer Tampa, FL
- sr information security engineer Tampa, FL
- senior data engineer Tampa, FL
- data engineer machine learning Tampa, FL
- junior data engineer remote Tampa, FL
- data center engineer Tampa, FL
- data developer Tampa, FL
- senior data center engineer Tampa, FL
- big data developer Tampa, FL

