Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Apache Spark Developer

$125k - $185k
Full-time

Bright Vision Technologies

Role Description

We are seeking an experienced Apache Spark Developer to design, develop, and optimize large-scale distributed data processing applications supporting enterprise analytics, machine learning, real-time reporting, and cloud-based data platforms. This role focuses on building high-performance Spark applications capable of processing billions of records across structured and semi-structured data sources while delivering scalable, reliable, and cost-efficient data pipelines.

You will work closely with data architects, data engineers, cloud platform teams, machine learning engineers, and business intelligence developers to build modern data processing solutions leveraging Apache Spark, cloud-native technologies, and distributed computing frameworks. The ideal candidate possesses deep expertise in Spark architecture, distributed systems, performance optimization, and cloud-based big data ecosystems.

Key Responsibilities

  • Design, develop, and maintain high-performance distributed data processing applications using Apache Spark.
  • Build scalable batch and real-time ETL/ELT pipelines processing large volumes of enterprise data.
  • Develop Spark applications using PySpark, Scala, or Spark SQL for data transformation, aggregation, and analytics.
  • Optimize Spark jobs for memory utilization, partitioning strategies, shuffle performance, and execution efficiency.
  • Process structured, semi-structured, and streaming data from enterprise databases, APIs, Kafka, cloud storage, and data lakes.
  • Develop reusable Spark libraries, data processing frameworks, and metadata-driven ingestion pipelines.
  • Collaborate with cloud engineering teams to deploy Spark workloads on Databricks, EMR, Azure Synapse, or Kubernetes.
  • Implement data quality validation, reconciliation, monitoring, and automated error handling across distributed pipelines.
  • Integrate Spark applications with enterprise data warehouses, lakehouses, and reporting platforms.
  • Participate in architecture reviews, code reviews, technical design discussions, and Agile development activities.
  • Troubleshoot production issues involving distributed processing, cluster performance, resource utilization, and data quality.
  • Support cloud migration initiatives by modernizing legacy ETL workloads into Spark-based architectures.

Qualifications

  • Six or more years of professional software or data engineering experience.
  • Four or more years of hands-on Apache Spark development experience in enterprise production environments.
  • Strong proficiency in PySpark, Scala, or Spark SQL for distributed data processing.
  • Deep understanding of Apache Spark architecture including RDDs, DataFrames, Datasets, Catalyst Optimizer, DAG execution, and Tungsten engine.
  • Strong experience with distributed computing concepts including partitioning, shuffling, caching, broadcast joins, and fault tolerance.
  • Advanced SQL skills with databases such as SQL Server, Oracle, PostgreSQL, Snowflake, or Teradata.
  • Experience working with Hadoop ecosystem technologies including Hive, HDFS, YARN, and Parquet.
  • Experience processing streaming data using Spark Structured Streaming, Apache Kafka, or Event Hubs.
  • Hands-on experience with cloud platforms including Azure Databricks, AWS EMR, AWS Glue, Azure Synapse Analytics, or Google Dataproc.
  • Experience integrating Spark applications with Delta Lake, Apache Iceberg, or Apache Hudi.
  • Strong understanding of data warehousing concepts, dimensional modeling, and data lake architecture.
  • Experience using Git, CI/CD pipelines, Azure DevOps, GitHub Actions, or Jenkins.
  • Strong debugging, troubleshooting, and Spark performance tuning skills.
  • Experience working in Agile Scrum development environments.

Preferred Qualifications

  • Experience building enterprise Lakehouse architectures using Databricks or Delta Lake.
  • Familiarity with Apache Airflow, Azure Data Factory, AWS Step Functions, or Control-M for workflow orchestration.
  • Experience with machine learning workflows using Spark MLlib, MLflow, or feature engineering pipelines.
  • Knowledge of Kubernetes, Docker, and containerized Spark deployments.
  • Experience implementing Data Quality frameworks using Great Expectations or Deequ.
  • Familiarity with Apache NiFi, Apache Flink, Trino, or Presto.
  • Experience working with cloud object storage including Amazon S3, Azure Data Lake Storage (ADLS Gen2), or Google Cloud Storage.
  • Knowledge of Infrastructure as Code using Terraform or ARM templates.
  • Experience with enterprise monitoring tools including Prometheus, Grafana, Datadog, or OpenTelemetry.
  • Cloud certifications in Azure, AWS, Databricks, or Apache Spark-related technologies are highly desirable.

Project Environment

You will be joining a modern data engineering team responsible for building cloud-native big data platforms supporting enterprise analytics, AI, and business intelligence initiatives. Current projects include:

  • Enterprise data lakehouse implementation using Databricks and Delta Lake
  • Real-time streaming analytics processing billions of daily events
  • Large-scale customer analytics and behavioral data platforms
  • Financial risk modeling and fraud detection pipelines
  • Healthcare clinical and operational analytics solutions
  • Cloud migration of legacy Hadoop and ETL workloads
  • Machine learning feature engineering and model training pipelines
  • Enterprise reporting platforms supporting executive dashboards and self-service analytics
  • Distributed data processing infrastructure deployed on Azure and AWS

How to Apply

Would you like to know more about this opportunity? For immediate consideration, please send your resume to [email protected] or contact us at View phone number on remotive.com. Learn more about Bright Vision Technologies at .

Vacancy posted 15 hours ago
Similar jobs that could be interesting for youBased on the Apache Spark Developer in Remote vacancy
  • $130k - $270k

     ...timely manner. Our team of engineers take pride in what they develop and constantly innovate to provide the best solution. Captivation...  ...with Distributed Big Data processing engines including Apache Spark Experience using Jupyter Notebook Experience with data wrangling... 
    Suggested
    Hourly pay
    Full time
    Temporary work

    Captivation Software

    Remote
    15 hours ago
  • $130k - $270k

     ...team of engineers take pride in what they develop and constantly innovate to provide the...  ...workflows and automation pipelines using Apache Airflow. This role focuses on building reliable...  ...Data processing engines including Apache Spark Experience with containerization... 
    Suggested
    Hourly pay
    Full time
    Temporary work

    Captivation Software

    Remote
    15 hours ago
  • $184k - $230k

     ...Engineer with deep expertise in distributed systems to join the Apache Spark Team. You will be at the forefront of innovation, building our...  ...in the open-source community. ~Build with Modern Stacks: Develop high-performance features using Scala, Java, and Python on... 
    Suggested
    Full time
    Work from home
    Flexible hours

    Cloudera

    Remote
    2 days ago
  •  ...Job Description Job Description This is a remote position. Job Title: Databricks Developer (Java & Apache Spark) Location: Remote Duration: Full-Time NEED IRS MBI Clearance. We are looking for a Databricks Developer who can design, build, and... 
    Suggested
    Full time
    Remote work

    3M Consultancy

    Washington DC
    28 days ago
  • $130k - $270k

     ...timely manner. Our team of engineers take pride in what they develop and constantly innovate to provide the best solution. Captivation...  ...expertise in dataflow design, data transport mechanisms, and Apache Spark based distributed processing. In this role, the Software... 
    Suggested
    Hourly pay
    Full time
    Temporary work

    Captivation Software

    Remote
    15 hours ago
  • About the TeamThe Spark Platform team owns and operates DoorDash's Apache Spark ecosystem — the execution runtime, remote shuffle service, cluster scheduler, and reliability tooling that powers the company's data, analytics, and ML workloads. We run Spark across the company... 
    Hourly pay
    Work at office
    Local area
    Remote work
    Relocation
    Flexible hours

    Doordash

    San Francisco, CA
    2 days ago
  • About the TeamThe Spark Platform team owns and operates DoorDash's Apache Spark ecosystem — the execution runtime, remote shuffle service, cluster scheduler, and reliability tooling that powers the company's data, analytics, and ML workloads. We run Spark across the company... 
    Hourly pay
    Work at office
    Local area
    Remote work
    Relocation
    Flexible hours

    Doordash

    Seattle, WA
    2 days ago
  • $120k

     ...your time and have a great week ahead!!! Role: Senior Java/Spark Developer - Remote (Should be inside US to apply for this role) Need...  ...and data processing solutions using Java, Kotlin, Scala, and Apache Spark. ~Design and implement data loading and transformation... 
    Full time
    Remote work

    Global Alliant Inc

    Remote
    5 days ago
  • Role Description Como Senior Backend Spark Developer, serás una pieza clave en el diseño, desarrollo y optimización de soluciones de procesamiento...  ...pipelines de Big Data robustos y escalables utilizando Apache Spark. ~Optimizar el rendimiento de las soluciones de... 
    Full time

    Zemsania

    Remote
    15 hours ago
  • $135k - $150k

     ...AWS. This role centers on hardening and operating Amazon EMR (Spark) and OpenSearch workloads, building secure CI/CD pipelines, managing...  ...Git Data & Development Foundation ~ Working knowledge of Apache Spark core concepts: RDDs, DataFrames, Spark SQL ~... 
    Remote work

    One Dynamic

    United States
    4 days ago
  •  ...Job Description The Senior Front-End Developer will be part of a team supporting established projects and creating products from the...  ...process, such as Jira. - Familiarity with web servers such as Apache, Nginx, etc. Bonus/Nice To Have - Interest in design and... 
    Full time

    Adept Solutions Inc

    Remote
    15 hours ago
  •  ...prior to applying Description: Looking for a full stack developer supporting a backend development team focused on middleware and...  ...in a big data environment using tools such as Hadoop, Pyspark, Spark, and Hbase (prefer at least two of the list) 3. Minimum of 5 years... 
    Full time

    Leading Path Consulting

    Remote
    15 hours ago
  •  ...Job Description Hi,   Role : Sr. Apache Druid Administrator Location: Irving, TX Duration: 12 Months Contract...  ...Configure and maintain Druid metadata stores using MySQL. Develop operational automation using Ansible and Infrastructure-as-Code... 
    Full time
    Contract work

    Cystems Logic Inc

    Remote
    15 hours ago
  • $100k - $150k

     ...~5+ years of professional experience designing and operating big-data pipelines on Hadoop. ~ Strong hands-on expertise with Apache Spark (Scala, Python, or Java) in production environments. ~ Solid experience with Hive, HDFS, Sqoop, HBase, and the broader Hadoop... 
    Full time
    H1b
    Local area
    Immediate start
    Remote work
    Visa sponsorship

    Bright Vision Technologies

    United States
    2 days ago
  •  ...Migration, Continuous Delivery, Big Data, Apache, Web Logic, Jenkins) in Reston, VA...  ...NoSQL,  HADOOP, NOSQL, Apache Kafka, Apache Spark, WebLogic, Oracle, Linux, Unix, Windows...  ...will take ownership of conceptualizing, developing, standardizing and driving the adoption of... 
    Permanent employment
    Full time
    Remote work

    DBA Web Technologies

    Canada
    more than 2 months ago
  •  ...Description Job Title: Technical Lead – Data Engineering (Palantir, Spark, PySpark, Python) Please dont apply if you dont have PALANTIR...  ...~Strong hands-on experience with:  ~ Python & PySpark ~ Apache Spark ~Advanced SQL ~Experience with Palantir Foundry or... 
    Work from home
    Flexible hours

    CLOUDSCOUTS SOFTWARE SOLUTIONS LLC

    Frisco, TX
    14 days ago
  • $130k - $270k

     ...a timely manner. Our team of engineers take pride in what they develop and constantly innovate to provide the best solution. Captivation...  ...substituted for a bachelor’s degree. Required Skills: Spark MapReduce SQL/NoSQL Pandas/Numpy/SciPy   This position... 
    Hourly pay
    Full time
    Temporary work

    Captivation Software

    Remote
    15 hours ago
  • $97.3k - $176.9k

     ...Data Platform, where you will design and develop the platform and infrastructure used by our...  ...edge Big Data tools like Airflow, Spark, Jenkins, Docker, and KubernetesPartner with...  ...and/or JavaExperience with Kafka, Spark, Apache Airflow, and KubernetesExperience with distributed... 
    Work experience placement
    Local area
    Remote work

    UnitedHealth Group

    San Francisco, CA
    3 days ago
  •  ...Fastest Growing Company by Inc.com 2015- SPARK FastTrack Award from Ann Arbor SPARK 20...  ...Job Description Role: UI Developer  Location: Charlotte, NC Duration:...  ...• Understanding of Web/App Servers like Apache and Apache Tomcat would be a plus  • Liferay... 
    Full time

    Stemxpert Llc

    Remote
    15 hours ago
  •  ...We are hiring a Sr. Software Engineer (Apache, Python, Trino, Kubernetes) to work in...  ...Description: The Software Engineer develops, maintains, and enhances complex and diverse...  ...degree. Apache AirFlow Python Apache Spark Trino Kubernetes Themis Insight... 
    Full time
    Local area
    Work from home
    Flexible hours

    Themis Insight

    Annapolis Junction, MD
    8 days ago
  • $111k - $114k

     ...Innovations is seeking Mid-Level Software Developers to provide remote support for a federal...  ...cloud platforms.Hands-on experience with Apache Kafka for event-driven architectures and...  ...Experience with distributed data systems such as Spark, HBase, or Solr.Quality &... 
    Work experience placement
    Local area
    Remote work

    Trilogy Innovations

    Bridgeport, WV
    4 days ago
  • $180.5k - $225.6k

     ...production scale. This greenfield provisioning layer will power all non-Spark compute workloads on Serverless (Notebooks, AI Agents, Remote...  ...globe and was founded by the original creators of Lakehouse, Apache Spark, Delta Lake and MLflow. To learn more, follow Databricks... 
    Local area
    Remote work
    Worldwide

    DataBricks

    Bellevue, WA
    2 days ago
  • $150.1k - $225.1k

     ...spaceRequirementsBachelor's degree in Computer Science3+ years of direct experience developing scalable, distributed systemsProficient in at least one...  ...with distributed data processing frameworks like Apache Spark or Apache FlinkUnderstanding of CI/CD pipelines, monitoring,... 
    Remote work

    Sony Interactive Entertainment America

    Los Angeles, CA
    4 days ago
  • $159.6k - $239.4k

     ...risks, the tools to grow, the skills to develop and the support of a company invested in...  ...authority on modern open-table formats (Apache Iceberg or Delta Lake). Define the standards...  ...the AWS data stack, specifically AWS EMR (Spark/Presto/Trino tuning), AWS Athena, S3... 
    Full time
    Work at office
    Remote work
    Home office
    Flexible hours

    Workday

    Boulder, CO
    15 hours ago
  • $158k - $197k

     ...datastores like DyanmoDB, Postgres, MySQL, etc.Distributed systems design & distributed processing through technologies such as Apache Spark or Databricks / Snowflake is a plusKnowledge on messaging technologies like Kafka / AWS Kinesis or similar is a plusOur Values Act... 
    H1b
    Remote work
    Worldwide
    Visa sponsorship
    Work visa

    Addepar

    New York, NY
    4 days ago
  • $134.63k - $224.38k

     ...frameworksGenie / AI-assisted development: accelerate developer productivity and data accessibilityEnable...  ...batch and streaming data pipelines using Spark and Delta LakeDevelop data products and...  ...enterprise environmentsDeep expertise in:Apache Spark (Scala/Python)Delta Lake and... 
    Temporary work
    Work at office
    Remote work
    Relocation
    Flexible hours

    NTT DATA

    Plano, TX
    3 days ago
  • $94k - $120k

     ...frameworks, open-source libraries, and APIs to develop basic application solutions.Learn and...  ...technologies (e.g. MS Cosmos DB, Apache Cassandra, Amazon DynamoDB)Understanding...  ...distributed computing (MS HPC, Sagemaker, Spark)Two years of experience with integration... 
    Full time
    Contract work
    Work at office
    Local area
    Remote work
    Worldwide
    Work visa
    Relocation package
    Flexible hours
    3 days per week

    Transamerica

    Denver, CO
    2 days ago
  • $85 - $90 per hour

     ...Experience working on Microservices.• Good experience on AWS Cloud, MSK, Kinesis, Lamda• Good hands-on experience on AWS Glue using Apache Spark, Amazon EMR• Good experience on container orchestration system like Kubernetes, ECS• Good working knowledge in SPRING Framework... 
    Hourly pay
    Full time
    Remote work

    SRI Tech

    San Diego, CA
    5 days ago
  •  ....Research industry best practices, evaluate new technologies, develop standards and engineering best practices and recommend innovative...  ...Azure/API GatewaysExperience with data processing technology (Apache Spark etc.)Experience with data virtualization technology (Tibco DV,... 
    Work at office
    Remote work
    Flexible hours

    NTT DATA

    Concord, CA
    3 days ago
  • $140k - $215k

     ...in your work will be Java microservices, Spark/Scala data processing (also some Flink),...  ...Write highly fault-tolerant Java code within Apache Spark, Flink and custom OLAP, Time Series...  ...trends and related analyticsDesign, develop, and maintain ultra-high-scale data platforms... 
    Full time
    Work experience placement
    Work at office
    Local area
    Remote work
    3 days per week
    1 day per week

    CrowdStrike

    New York, NY
    11 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Apache Spark Developer. Be the first to apply!