Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Apache Spark Developer

$125k - $185k
Full-time

Bright Vision Technologies

Apache Spark Developer – Remote

Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States.

This is a fantastic opportunity to join an established and well-respected organization offering tremendous career growth potential.

Job Title: Apache Spark Developer
Location: 100% Remote (U.S.)
Position Type: Full-time, Direct W2
Salary Range: $125,000–$185,000 Annually
Experience Required: 6+ years

Sponsorship: U.S. Citizens, Green Card Holders, EAD Holders, and H-1B transfer candidates are encouraged to apply. We are unable to sponsor new H-1B visa petitions for this position.

Job Summary
We are seeking an experienced Apache Spark Developer to design, develop, and optimize large-scale distributed data processing applications supporting enterprise analytics, machine learning, real-time reporting, and cloud-based data platforms. This role focuses on building high-performance Spark applications capable of processing billions of records across structured and semi-structured data sources while delivering scalable, reliable, and cost-efficient data pipelines.
You will work closely with data architects, data engineers, cloud platform teams, machine learning engineers, and business intelligence developers to build modern data processing solutions leveraging Apache Spark, cloud-native technologies, and distributed computing frameworks. The ideal candidate possesses deep expertise in Spark architecture, distributed systems, performance optimization, and cloud-based big data ecosystems.

Key Responsibilities
  • Design, develop, and maintain high-performance distributed data processing applications using Apache Spark.
  • Build scalable batch and real-time ETL/ELT pipelines processing large volumes of enterprise data.
  • Develop Spark applications using PySpark, Scala, or Spark SQL for data transformation, aggregation, and analytics.
  • Optimize Spark jobs for memory utilization, partitioning strategies, shuffle performance, and execution efficiency.
  • Process structured, semi-structured, and streaming data from enterprise databases, APIs, Kafka, cloud storage, and data lakes.
  • Develop reusable Spark libraries, data processing frameworks, and metadata-driven ingestion pipelines.
  • Collaborate with cloud engineering teams to deploy Spark workloads on Databricks, EMR, Azure Synapse, or Kubernetes.
  • Implement data quality validation, reconciliation, monitoring, and automated error handling across distributed pipelines.
  • Integrate Spark applications with enterprise data warehouses, lakehouses, and reporting platforms.
  • Participate in architecture reviews, code reviews, technical design discussions, and Agile development activities.
  • Troubleshoot production issues involving distributed processing, cluster performance, resource utilization, and data quality.
  • Support cloud migration initiatives by modernizing legacy ETL workloads into Spark-based architectures.
Required Skills
  • Six or more years of professional software or data engineering experience.
  • Four or more years of hands-on Apache Spark development experience in enterprise production environments.
  • Strong proficiency in PySpark , Scala , or Spark SQL for distributed data processing.
  • Deep understanding of Apache Spark architecture including RDDs, DataFrames, Datasets, Catalyst Optimizer, DAG execution, and Tungsten engine.
  • Strong experience with distributed computing concepts including partitioning, shuffling, caching, broadcast joins, and fault tolerance.
  • Advanced SQL skills with databases such as SQL Server, Oracle, PostgreSQL, Snowflake, or Teradata.
  • Experience working with Hadoop ecosystem technologies including Hive, HDFS, YARN, and Parquet.
  • Experience processing streaming data using Spark Structured Streaming, Apache Kafka, or Event Hubs.
  • Hands-on experience with cloud platforms including Azure Databricks, AWS EMR, AWS Glue, Azure Synapse Analytics, or Google Dataproc.
  • Experience integrating Spark applications with Delta Lake, Apache Iceberg, or Apache Hudi.
  • Strong understanding of data warehousing concepts, dimensional modeling, and data lake architecture.
  • Experience using Git, CI/CD pipelines, Azure DevOps, GitHub Actions, or Jenkins.
  • Strong debugging, troubleshooting, and Spark performance tuning skills.
  • Experience working in Agile Scrum development environments.
Preferred Qualifications
  • Experience building enterprise Lakehouse architectures using Databricks or Delta Lake.
  • Familiarity with Apache Airflow, Azure Data Factory, AWS Step Functions, or Control-M for workflow orchestration.
  • Experience with machine learning workflows using Spark MLlib, MLflow, or feature engineering pipelines.
  • Knowledge of Kubernetes, Docker, and containerized Spark deployments.
  • Experience implementing Data Quality frameworks using Great Expectations or Deequ.
  • Familiarity with Apache NiFi, Apache Flink, Trino, or Presto.
  • Experience working with cloud object storage including Amazon S3, Azure Data Lake Storage (ADLS Gen2), or Google Cloud Storage.
  • Knowledge of Infrastructure as Code using Terraform or ARM templates.
  • Experience with enterprise monitoring tools including Prometheus, Grafana, Datadog, or OpenTelemetry.
  • Cloud certifications in Azure, AWS, Databricks, or Apache Spark-related technologies are highly desirable.
Project Environment
You will be joining a modern data engineering team responsible for building cloud-native big data platforms supporting enterprise analytics, AI, and business intelligence initiatives. Current projects include:
  • Enterprise data lakehouse implementation using Databricks and Delta Lake
  • Real-time streaming analytics processing billions of daily events
  • Large-scale customer analytics and behavioral data platforms
  • Financial risk modeling and fraud detection pipelines
  • Healthcare clinical and operational analytics solutions
  • Cloud migration of legacy Hadoop and ETL workloads
  • Machine learning feature engineering and model training pipelines
  • Enterprise reporting platforms supporting executive dashboards and self-service analytics
  • Distributed data processing infrastructure deployed on Azure and AWS
This is a hands-on engineering role where you will contribute to distributed system architecture, Spark application development, cloud migration, performance optimization, production support, and continuous improvement of enterprise-scale data processing platforms.

How to Apply

Would you like to know more about this opportunity? For immediate consideration, please send your resume to View email address on brightvisiontechnologies.applytojob.com or contact us at Show phone number . Learn more about Bright Vision Technologies at .

 

Bright Vision Technologies is an Equal Opportunity Employer.

 

Equal Employment Opportunity (EEO) Statement

Bright Vision Technologies (BV Teck) is committed to equal employment opportunity (EEO) for all employees and applicants without regard to race, color, religion, sex, sexual orientation, gender identity or expression, national origin, age, genetic information, disability, veteran status, or any other protected status as defined by applicable federal, state, or local laws. This commitment extends to all aspects of employment, including recruitment, hiring, training, compensation, promotion, transfer, leaves of absence, termination, layoffs, and recall.

BV Teck expressly prohibits any form of workplace harassment or discrimination. Any improper interference with employees' ability to perform their job duties may result in disciplinary action up to and including termination of employment.

Vacancy posted 4 days ago
Similar jobs that could be interesting for youBased on the Apache Spark Developer in Richardson, TX vacancy
  •  ...the technologies at the heart of 5G. Their next vision was to develop an IoT sticker, a computing element that can power itself by harvesting...  ...(AWS, Azure, or GCP). ~ Demonstrated experience with Apache Spark. ~ Familiarity with RESTful APIs, microservices, and event-... 
    Suggested
    Full time
    Work at office

    Wiliot

    Plano, TX
    1 day ago
  • $120k - $155k

     ...Solutions Engineer, where you will play a key role in designing and developing enterprise data solutions. This role works across the data...  ...Lake Storage • Azure Synapse Analytics, Azure Databricks, Apache Spark, Apache Airflow • Azure Logic Apps, Azure DevOps, Power... 
    Suggested
    Full time
    Temporary work
    Flexible hours

    BakerHostetler

    Dallas, TX
    3 days ago
  •  ...full-stack Java, Web/UI designers, Big Data or Cloud or Mobility developers/architects, we have them all.  Job Description Minimum...  ...like NetBeans, Eclipse. Experience with server management (Apache/GlassFish/WAS) to provide basic administration and code builds... 
    Suggested
    Full time

    Jobsbridge

    Plano, TX
    1 day ago
  •  ...Software Engineer Job Description Duties: Design, develop and implement software solutions. Solve business...  ...maintaining data pipelines using technologies including Spark, data integration tools including Apache NiFi, and messaging systems including Kafka and JMS;... 
    Suggested
    Full time

    JPMorgan Chase

    Plano, TX
    5 days ago
  • $179.4k - $204.7k

     ...Enable Application Teams: Work directly with internal application developers to refactor data consumers, ensuring a smooth cutover from...  ...at a multi-billion record scale. 3+ years of experience with Apache Kafka (both self-managed and AWS MSK) and AWS Lambda. 3+ years... 
    Suggested
    Full time
    Part time
    Internship
    H1b
    Local area

    Capital One

    Plano, TX
    3 days ago
  • $269.1k - $307.2k

     ..., JSON, XML, Ruby, Perl, NoSQL databases, relational databases, Spark, Artifactory, Maven, iOS, Android, and AWS/Cloud Infrastructure...  ...practices. Work within and across Agile teams to select, design, develop, test, implement, and support technical solutions across full-... 
    Full time
    Part time
    Local area
    Shift work

    COMFORT SYSTEMS

    Plano, TX
    4 days ago
  •  ...capabilities. This role sits within the engineering effort to develop a modern Lakehouse and AI data platform that enables reliable,...  ...working with distributed data processing frameworks such as Apache Spark . Working knowledge of common data formats such as JSON... 
    Full time
    Work at office

    Goldman Sachs Group, Inc.

    Dallas, TX
    6 hours ago
  •  ...The Java Full Stack Developer will be responsible for designing, developing, and deploying scalable microservices and responsive user interfaces...  ...(v10+), TypeScript, HTML5, and CSS3. Strong expertise in Apache Kafka and message-driven architecture. Experience integrating... 

    Compunnel

    Plano, TX
    2 days ago
  • $136.2k - $204.3k

     ...CarMax's customer-first mission. Role Responsibilities Develop and implement self-service capabilities and automation tools to...  ...the modern customer and drive progress. Your work fuels change—sparking ideas, overcoming challenges, and shaping what's next. Join us... 
    Hourly pay
    Full time
    Home office
    2 days per week

    CarMax

    Plano, TX
    15 hours ago
  • $179.4k - $204.7k

     ...What You’ll Do Collaborate with and across Agile teams to design, develop, test, implement, and support technical solutions in full‑stack...  ...data/computing tools (MapReduce, Hadoop, Hive, EMR, Kafka, Spark, Gurobi, or MySQL) 4+ years of experience working on real‑time data... 
    Internship
    Local area

    Capital One

    Plano, TX
    2 days ago
  • $88.2k - $121.2k

    Full Stack Developer Category: Software Development/ Engineering Main location: United...  ...datasets Develop services that leverage Apache Iceberg tables and modern data lakehouse...  ...as Code Familiarity with Spark and PySpark Experience with distributed... 
    Full time
    Local area

    CGI Technologies and Solutions, Inc.

    Dallas, TX
    1 day ago
  • $134.63k - $224.38k

     ...Genie / AI-assisted development: accelerate developer productivity and data accessibility...  ...batch and streaming data pipelines using Spark and Delta Lake Develop data products and...  ...enterprise environments Deep expertise in: Apache Spark (Scala/Python) Delta Lake and... 
    Temporary work
    Work at office
    Remote work
    Relocation
    Flexible hours

    NTT DATA, Inc.

    Plano, TX
    7 days ago
  • $134.63k - $224.38k

     ...Genie / AI-assisted development: accelerate developer productivity and data accessibility...  ...batch and streaming data pipelines using Spark and Delta Lake Develop data products and...  ...enterprise environments Deep expertise in: Apache Spark (Scala/Python) Delta Lake and... 
    Temporary work
    Work at office
    Remote work
    Relocation
    Flexible hours

    NTT DATA, Inc.

    Plano, TX
    a month ago
  • $179.4k - $204.7k

     ...delivering capabilities that exemplify those practices. What You’ll Do: Lead a portfolio of diverse technology projects and a team of developers with deep experience in distributed microservices, and full stack systems to create solutions that help meet regulatory needs for... 
    Full time
    Part time
    Internship
    H1b
    Local area

    Capital One

    Plano, TX
    2 days ago
  • $112.2k - $202.6k

     ...Job Summary An individual with the ability to design and develop data integrations to deliver Enterprise BI Data Marts and BI Semantic...  ...experience working with Hadoop tech stack like HDFS, Hive, HQL, Spark, Scala, PySpark/python, Sqoop. 5+ years of experience working with... 
    Flexible hours

    HCSC

    Richardson, TX
    4 days ago
  •  ...gather requirements, build and optimize Spark-based batch/real-time workflows and data...  ...framework(s), i.e., Python, Java, Scala, Apache Spark, Databricks, Grafana, Prometheus, Elasticsearch...  ...on smart, driven people like you to develop applications and provide tech support for... 
    Work at office

    JPMorgan Chase & Co.

    Plano, TX
    6 days ago
  •  ...What You’ll Do Collaborate with and across Agile teams to design, develop, test, implement, and support technical solutions in full‑stack...  ...data/computing tools (MapReduce, Hadoop, Hive, EMR, Kafka, Spark, Gurobi, or MySQL) 4+ years experience working on real‑time data... 
    Internship
    H1b
    Local area

    Capital One

    Plano, TX
    3 days ago
  • Company Description SonSoft is an IT Staffing and consulting firm and duly organized under the laws of the Commonwealth of Georgia. We are growing at a steady pace specializing in the fields of Software Development, Software Consultancy and Information Technology ...
    Permanent employment
    Full time
    H1b

    SonSoft Inc.

    Plano, TX
    1 day ago
  •  ...Job Responsibilities: Lead architecture and delivery of high-throughput, low-latency data pipelines using Databricks and Apache Spark (Core, SQL, Structured Streaming). Establish lakehouse patterns with Delta Lake (ACID transactions, schema evolution, time travel... 

    JPMorgan Chase & Co.

    Plano, TX
    11 days ago
  • $67.7k - $90.27k

     ...defined by trust, collaboration, and accountability, a Software Developer II has the opportunity to challenge convention, solve complex...  ...operational reliability and SLA compliance. Leverage Apache Spark and Databricks to support large-scale data processing and modernization... 
    Full time
    Temporary work
    Remote work
    Work from home

    Lumen

    Dallas, TX
    3 days ago
  •  ...You’ll Do: Collaborate with and across Agile teams to design, develop, test, implement, and support technical solutions in full-stack...  ...experience does not apply). Tech Stack: Full Stack (Java) - AWS/Python/Spark. Nice to Have: Data Engineering. Preferred Qualifications: 5... 
    Internship

    Ampcus, Inc

    Plano, TX
    5 days ago
  •  ...stores such as OpenSearch, Postgres (or PSQL), CockroachDB and Spark. 3+ years of professional experience with Kafka. 5+ years of experience...  ...Experience with Data Pipelines such as ElasticSearch, NATS or Apache Airflow Experience with data mining or machine learning... 
    Immediate start
    Flexible hours

    DomainTools

    Dallas, TX
    3 days ago
  • $179.4k - $204.7k

     ...driving a major transformation within Capital One. What You’ll Do: Lead a portfolio of diverse technology projects and a team of developers with deep experience in distributed microservices, and full stack systems to create solutions that help meet regulatory needs for... 
    Full time
    Part time
    Internship
    H1b
    Local area

    Capital One

    Plano, TX
    2 days ago
  • $209k - $238.5k

     ...driving a major transformation within Capital One. What You’ll Do: Lead a portfolio of diverse technology projects and a team of developers with deep experience in distributed microservices, and full stack systems to create solutions that help meet regulatory needs for... 
    Full time
    Part time
    Internship
    Local area

    Capital One

    Plano, TX
    2 days ago
  •  ...Job Description DESCRIPTION: Duties: Design, develop and implement software solutions. Solve business problems...  ...data pipelines using technologies including Spark, data integration tools including Apache NiFi, and messaging systems including Kafka and JMS;... 
    Full time

    JPMorgan Chase & Co.

    Plano, TX
    13 days ago
  •  ...building scalable, high performing and robust applications; design, develop, test and operate code across multiple productions systems •...  ...• Experience with big data methodologies involving Hadoop and Spark. • Experience with Continuous Integration and related tools (... 

    Redolent, Inc

    Dallas, TX
    10 days ago
  •  ...volunteer, disaster relief, grant management, payroll deductions and more. Job Description YourCause is looking for a Front-End Developer to help us build engaging experiences for both desktop and mobile. Responsibilities will include: Be a key member in re-... 
    Full time
    Work experience placement
    Flexible hours

    Yourcause, Llc

    Plano, TX
    1 day ago
  •  ...and financial stakeholders to understand their priorities and guide them toward technology solutions that deliver impact. You will develop a deep understanding of Tyler’s Justice solutions and competitive strengths, telling a compelling story that resonates with diverse... 
    Currently hiring
    Local area
    Remote work

    Tyler Technologies

    Plano, TX
    1 day ago
  • $138k - $190.5k

     ...tooling for software development (e.g., code generation, test creation, refactoring assistance). Platform & Cloud Experience developing and deploying applications on Microsoft Azure, including: Azure Functions Azure Storage (Blobs) Azure Service Bus API... 
    Local area

    Loandepot

    Plano, TX
    1 day ago
  • $140k - $200k

     ...interview process involves several technical interviews and we aim to complete them within 1 week.  What Yo u’ ll Do Design, develop, and maintain robust APIs including public TTS API, internal APIs like Payment, Subscription, Auth and Consumption Tracking,... 
    Remote job
    Full time
    Work at office

    Speechify

    Plano, TX
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Apache Spark Developer. Be the first to apply!