Apache Spark Developer
$125k - $185kBright Vision Technologies
Location: 100% Remote (U.S.)
Position Type: Full-time, Direct W2
Salary Range: $125,000–$185,000 Annually
Experience Required: 6+ years Sponsorship: U.S. Citizens, Green Card Holders, EAD Holders, and H-1B transfer candidates are encouraged to apply. We are unable to sponsor new H-1B visa petitions for this position. Job Summary
We are seeking an experienced Apache Spark Developer to design, develop, and optimize large-scale distributed data processing applications supporting enterprise analytics, machine learning, real-time reporting, and cloud-based data platforms. This role focuses on building high-performance Spark applications capable of processing billions of records across structured and semi-structured data sources while delivering scalable, reliable, and cost-efficient data pipelines.
You will work closely with data architects, data engineers, cloud platform teams, machine learning engineers, and business intelligence developers to build modern data processing solutions leveraging Apache Spark, cloud-native technologies, and distributed computing frameworks. The ideal candidate possesses deep expertise in Spark architecture, distributed systems, performance optimization, and cloud-based big data ecosystems. Key Responsibilities
- Design, develop, and maintain high-performance distributed data processing applications using Apache Spark.
- Build scalable batch and real-time ETL/ELT pipelines processing large volumes of enterprise data.
- Develop Spark applications using PySpark, Scala, or Spark SQL for data transformation, aggregation, and analytics.
- Optimize Spark jobs for memory utilization, partitioning strategies, shuffle performance, and execution efficiency.
- Process structured, semi-structured, and streaming data from enterprise databases, APIs, Kafka, cloud storage, and data lakes.
- Develop reusable Spark libraries, data processing frameworks, and metadata-driven ingestion pipelines.
- Collaborate with cloud engineering teams to deploy Spark workloads on Databricks, EMR, Azure Synapse, or Kubernetes.
- Implement data quality validation, reconciliation, monitoring, and automated error handling across distributed pipelines.
- Integrate Spark applications with enterprise data warehouses, lakehouses, and reporting platforms.
- Participate in architecture reviews, code reviews, technical design discussions, and Agile development activities.
- Troubleshoot production issues involving distributed processing, cluster performance, resource utilization, and data quality.
- Support cloud migration initiatives by modernizing legacy ETL workloads into Spark-based architectures.
- Six or more years of professional software or data engineering experience.
- Four or more years of hands-on Apache Spark development experience in enterprise production environments.
- Strong proficiency in PySpark , Scala , or Spark SQL for distributed data processing.
- Deep understanding of Apache Spark architecture including RDDs, DataFrames, Datasets, Catalyst Optimizer, DAG execution, and Tungsten engine.
- Strong experience with distributed computing concepts including partitioning, shuffling, caching, broadcast joins, and fault tolerance.
- Advanced SQL skills with databases such as SQL Server, Oracle, PostgreSQL, Snowflake, or Teradata.
- Experience working with Hadoop ecosystem technologies including Hive, HDFS, YARN, and Parquet.
- Experience processing streaming data using Spark Structured Streaming, Apache Kafka, or Event Hubs.
- Hands-on experience with cloud platforms including Azure Databricks, AWS EMR, AWS Glue, Azure Synapse Analytics, or Google Dataproc.
- Experience integrating Spark applications with Delta Lake, Apache Iceberg, or Apache Hudi.
- Strong understanding of data warehousing concepts, dimensional modeling, and data lake architecture.
- Experience using Git, CI/CD pipelines, Azure DevOps, GitHub Actions, or Jenkins.
- Strong debugging, troubleshooting, and Spark performance tuning skills.
- Experience working in Agile Scrum development environments.
- Experience building enterprise Lakehouse architectures using Databricks or Delta Lake.
- Familiarity with Apache Airflow, Azure Data Factory, AWS Step Functions, or Control-M for workflow orchestration.
- Experience with machine learning workflows using Spark MLlib, MLflow, or feature engineering pipelines.
- Knowledge of Kubernetes, Docker, and containerized Spark deployments.
- Experience implementing Data Quality frameworks using Great Expectations or Deequ.
- Familiarity with Apache NiFi, Apache Flink, Trino, or Presto.
- Experience working with cloud object storage including Amazon S3, Azure Data Lake Storage (ADLS Gen2), or Google Cloud Storage.
- Knowledge of Infrastructure as Code using Terraform or ARM templates.
- Experience with enterprise monitoring tools including Prometheus, Grafana, Datadog, or OpenTelemetry.
- Cloud certifications in Azure, AWS, Databricks, or Apache Spark-related technologies are highly desirable.
You will be joining a modern data engineering team responsible for building cloud-native big data platforms supporting enterprise analytics, AI, and business intelligence initiatives. Current projects include:
- Enterprise data lakehouse implementation using Databricks and Delta Lake
- Real-time streaming analytics processing billions of daily events
- Large-scale customer analytics and behavioral data platforms
- Financial risk modeling and fraud detection pipelines
- Healthcare clinical and operational analytics solutions
- Cloud migration of legacy Hadoop and ETL workloads
- Machine learning feature engineering and model training pipelines
- Enterprise reporting platforms supporting executive dashboards and self-service analytics
- Distributed data processing infrastructure deployed on Azure and AWS
Bright Vision Technologies is an Equal Opportunity Employer.
Equal Employment Opportunity (EEO) Statement
Bright Vision Technologies (BV Teck) is committed to equal employment opportunity (EEO) for all employees and applicants without regard to race, color, religion, sex, sexual orientation, gender identity or expression, national origin, age, genetic information, disability, veteran status, or any other protected status as defined by applicable federal, state, or local laws. This commitment extends to all aspects of employment, including recruitment, hiring, training, compensation, promotion, transfer, leaves of absence, termination, layoffs, and recall.
BV Teck expressly prohibits any form of workplace harassment or discrimination. Any improper interference with employees' ability to perform their job duties may result in disciplinary action up to and including termination of employment.
$130k - $270k
...timely manner. Our team of engineers take pride in what they develop and constantly innovate to provide the best solution. Captivation... ...with Distributed Big Data processing engines including Apache Spark Experience using Jupyter Notebook Experience with data wrangling...SuggestedHourly payFull timeTemporary work$130k - $270k
...team of engineers take pride in what they develop and constantly innovate to provide the... ...workflows and automation pipelines using Apache Airflow. This role focuses on building reliable... ...Data processing engines including Apache Spark Experience with containerization...SuggestedHourly payFull timeTemporary work$184k - $230k
...Engineer with deep expertise in distributed systems to join the Apache Spark Team. You will be at the forefront of innovation, building our... ...in the open-source community. Build with Modern Stacks: Develop high-performance features using Scala, Java, and Python on...SuggestedRemote workWork from homeFlexible hours$125k - $185k
Role Description We are seeking an experienced Apache Spark Developer to design, develop, and optimize large-scale distributed data processing applications supporting enterprise analytics, machine learning, real-time reporting, and cloud-based data platforms. This role...SuggestedFull timeImmediate start$130k - $270k
...timely manner. Our team of engineers take pride in what they develop and constantly innovate to provide the best solution. Captivation... ...expertise in dataflow design, data transport mechanisms, and Apache Spark based distributed processing. In this role, the Software...SuggestedHourly payFull timeTemporary work$130.6k - $192k
...About the Team The Spark Platform team owns and operates DoorDash's Apache Spark ecosystem — the execution runtime, remote shuffle service, cluster scheduler, and reliability tooling that powers the company's data, analytics, and ML workloads. We run Spark across...Hourly payWork at officeLocal areaRemote workRelocationFlexible hours- About the TeamThe Spark Platform team owns and operates DoorDash's Apache Spark ecosystem — the execution runtime, remote shuffle service, cluster scheduler, and reliability tooling that powers the company's data, analytics, and ML workloads. We run Spark across the company...Hourly payWork at officeLocal areaRemote workRelocationFlexible hours
- ...build and maintain reusable frameworks, libraries and internal developer tooling to standardise and accelerate data pipeline... ...Demonstrated experience with enterprise data pipeline tooling (Apache Spark, dbt, Airflow), providing guidance on best practices and helping...Permanent employmentFull timePart timeWork at officeFlexible hours
$120k
...your time and have a great week ahead!!! Role: Senior Java/Spark Developer - Remote (Should be inside US to apply for this role) Need... ...and data processing solutions using Java, Kotlin, Scala, and Apache Spark. ~Design and implement data loading and transformation...Full timeRemote work$80k - $110k
...optimize application performance and scalability. ~ Design and build distributed data processing pipelines using Apache Spark and Java/Scala. ~ Develop high-performance, multi-threaded server-side applications. ~ Integrate big data ecosystems with relational or...Full timeRelocation package- Role Description Como Senior Backend Spark Developer, serás una pieza clave en el diseño, desarrollo y optimización de soluciones de procesamiento... ...pipelines de Big Data robustos y escalables utilizando Apache Spark. ~Optimizar el rendimiento de las soluciones de...Full time
- ...solutions using technologies such as Kafka, Spark, Elasticsearch, and cloud-based platforms... ...| AWS | Linux What You’ll Do Design, develop, and maintain enterprise data ingestion,... ...solutions. Experience building and supporting Apache Kafka producers, consumers, and streaming...Full timeRemote workFlexible hours
- ...Job Description The Senior Front-End Developer will be part of a team supporting established projects and creating products from the... ...process, such as Jira. - Familiarity with web servers such as Apache, Nginx, etc. Bonus/Nice To Have - Interest in design and...Full time
- ...prior to applying Description: Looking for a full stack developer supporting a backend development team focused on middleware and... ...in a big data environment using tools such as Hadoop, Pyspark, Spark, and Hbase (prefer at least two of the list) 3. Minimum of 5 years...Full time
$135k - $150k
...AWS. This role centers on hardening and operating Amazon EMR (Spark) and OpenSearch workloads, building secure CI/CD pipelines, managing... ...Git Data & Development Foundation ~ Working knowledge of Apache Spark core concepts: RDDs, DataFrames, Spark SQL ~...Full time- ...Job Description Hi, Role : Sr. Apache Druid Administrator Location: Irving, TX Duration: 12 Months Contract... ...Configure and maintain Druid metadata stores using MySQL. Develop operational automation using Ansible and Infrastructure-as-Code...Full timeContract work
- About the job Tech Lead (Spark and Python expertise) Position: Tech Lead (Spark and Python expertise) Location... ...working in heavy data background needed. Develop, program, and maintain applications using the Apache Spark open-source framework. Work with different...Remote work
$100k - $150k
...architect, deploy, and operate large-scale Apache Kafka and Confluent platform environments... ...~Automation ~Observability ~Developer enablement The ideal candidate will... ...with stream processing frameworks (Flink, Spark Streaming) (preferred). ~Experience with...Full timeLocal areaImmediate start- ...ETL Engineer / Java Spark/ Ab Initio Wilmington DE (3 days WFO, 2 days WFH) Look for candidate who can be onsite from day 1 Abinitio - is needed (if not abinitio then any other ETL tools such as Informatica, Talend etc. ) Spark AWS...Work from homeFlexible hours
$130k - $270k
...a timely manner. Our team of engineers take pride in what they develop and constantly innovate to provide the best solution. Captivation... ...substituted for a bachelor’s degree. Required Skills: Spark MapReduce SQL/NoSQL Pandas/Numpy/SciPy This position...Hourly payFull timeTemporary work- ...design and scale the Audience Builder platform, enabling billions of identity records to activate to ad platforms. You will work on Spark pipelines, data models, and integration with the in-house AdRise stack. The role involves optimizing performance, collaborating across...Remote job
- ...Fastest Growing Company by Inc.com 2015- SPARK FastTrack Award from Ann Arbor SPARK 20... ...Job Description Role: UI Developer Location: Charlotte, NC Duration:... ...• Understanding of Web/App Servers like Apache and Apache Tomcat would be a plus • Liferay...Full time
- ...ad platform teams to drive advertisers' value. The role requires 6+ years building production data or backend systems, strong Scala or JVM background, and experience with Spark, Delta Lake, Parquet, or Iceberg on Databricks. Remote options available. #J-18808-Ljbffr FOXRemote job
- ...learn, share knowledge and teach within your team and within the developer community at Electric Mind via educational sessions, study... ...systems at scale handling large data sets leveraging Apache Spark, Kafka, Kinesis, and Hadoop toolsets Experience with Infrastructure...Contract work
- ...accessibility. The Software Engineer role will be responsible for developing end to end systems for one of Nava's major government... ...AWS cloud native services and big data technologies such as Apache Spark and Parquet. Tenacity to dive into problems and iterate in...Remote jobFull timeContract workTemporary workFor contractorsWork at officeLocal areaHome officeVisa sponsorshipFlexible hours
$110k - $270k
...architecture development of the data platform for Opendata Develop core platform components including data ingestion, storage and... ...developing scalable data systems ~ Experience working with Apache Spark, Airflow (or similar), Data Lakes and open table formats such...Remote jobFull timeWork at officeLocal areaWork from homeFlexible hours- ...and operate core components of our lakehouse platform, including Apache Iceberg table management (data compaction, data layout... ...formats across internal teams, owning the integration of Trino, Spark and other query engines (DuckDB, Puppygraph…) with our Iceberg-based...Full timeWork at office
$107k - $120.6k
...accessibility.The Software Engineer role will be responsible for developing end to end systems for one of Nava's major government... ...technical and otherwise Desired skills Experience using Apache Spark and Parquet for data engineering Experience with Node.js...Remote jobFull timeTemporary workFor contractors$136k - $220k
...knowledge of high throughput processing technologies such as Hadoop, Spark, Flink and/or Kafka. Proficiency in Java and strong... ...as Snowflake, Google BigQuery, Databricks Lakehouse, AWS Athena, Apache Trino, or Presto You’ve experience with open source data storage...Full timeWorldwideFlexible hours$140k - $200k
...interview process involves several technical interviews and we aim to complete them within 1 week. What Yo u’ ll Do Design, develop, and maintain robust APIs including public TTS API, internal APIs like Payment, Subscription, Auth and Consumption Tracking,...Full timeWork at office
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Apache Spark Developer. Be the first to apply!







