Data Engineer
Soteris
ABOUT SOTERIS
Soteris is a YC-backed AI company building the future of pricing and product management for the
insurance industry. Our mission is to infuse the $5 trillion P&C insurance industry with best-in-class,
proprietary, AI-driven data analytics. We’ve spent years building our own proprietary AI models on
personal auto claims and exposure data to help insurers improve their loss ratios, with over 100 million
submissions and $180 billion in premium scored to date.
Each year, roughly $750 billion in insurance policies are written in the United States. Our machine
learning platform helps insurers evaluate policies at a granular level, moving beyond broad segmentation approaches that often lead to risks being over- or underpriced. Our modeling approach incorporates multiple model families and calibration methods to rank policy risk within a book of business.
We are a team of 10 and growing quickly. As our second Data Engineer, you will own the data layer that turns customer policy, claims, quote, and financial data into trusted inputs for actuarial analysis, model development, and production scoring. This is a builder role: you will work directly with customer data teams, create repeatable ingestion and transformation pipelines, reconcile outputs to source-of-truth control totals, and make the platform easier to operate as we add customers and products.
WHAT YOU'LL BE DOING
Customer Data Onboarding and Integration
- Leading the technical data workstream for new customer implementations by understanding policy, claims, rating, quote, and financial systems and establishing secure access to the data.
- Building reusable extraction and synchronization workflows for databases, backups, secure file transfer, APIs, and other delivery methods while preserving source lineage and supporting backfills.
- Mapping customer data into Soteris’s internal ontology and working with customer technical teams and internal project leads to resolve definitions, transformations, data-quality issues, and onboarding blockers.
Lakehouse and Pipeline Engineering
- Owning Databricks and AWS pipelines that move customer data from raw and Bronze ingestion through standardized Silver tables and curated Gold or model-ready datasets.
- Designing idempotent, incremental, observable workflows that handle schema evolution, late-arriving data, backfills, orchestration, performance, and cost.
- Developing shared components and configuration-driven patterns, then publishing well-defined datasets for actuarial analysis, backtesting, model training, production scoring, reporting, and monitoring.
Data Quality and Modeling Readiness
- Building automated quality gates for completeness, uniqueness, referential integrity, valid ranges, freshness, balance, schema drift, and other customer-specific controls.
- Reconciling written and earned premium, exposure, incurred and ultimate loss, claim counts, fee income, and other economic drivers to customer control statistics at the required state, program, year, and coverage levels.
- Partnering with data science to produce leakage-resistant, point-in-time-correct datasets and productionize approved actuarial reference data
Production Platform and Operational Ownership
- Owning data flows for quote requests, model inputs, and bound policy outcomes, including pre-live comparisons that confirm production request fields match corresponding policy data and post-launch drift checks.
- Operating pipelines with monitoring, alerting, recovery behavior, runbooks, and clear incident diagnostics across customer synchronization, Databricks processing, and downstream model- serving dependencies.
- Testing, reviewing, and deploying data code through GitHub and GitHub Actions, with automated tests and deployment controls for production data assets.
- Managing Databricks permissions, Unity Catalog controls, sensitive data, and least-privilege access in support of Soteris’s security and SOC 2 requirements.
OUR CURRENT STACK
- Python, SQL, PySpark, pandas, and related data-engineering libraries
- Databricks, Delta Lake, Databricks Workflows, and Unity Catalog
- AWS, including S3, Lambda, EC2, ECS, SageMaker, and secure customer file transfer
- Customer databases, backups, SFTP, APIs, and file-based ingestion
- Terraform, GitHub, GitHub Actions, automated testing, and infrastructure as code
- MLflow, SageMaker, and production scoring APIs at the model handoff boundary
- Modern generative AI development tools, including Claude, ChatGPT, Codex, or similar models
ABOUT YOU
You must have the following:
- Strong Python and SQL skills and the ability to write production-quality transformations, tests, utilities, and operational tooling.
- Hands-on experience building and operating production pipelines using Spark, Databricks, or a comparable distributed data platform.
- Strong understanding of data modeling, lakehouse or warehouse design, incremental processing, schema evolution, idempotency, backfills, and lineage.
- Experience designing data-quality controls and reconciling complex datasets to source systems or independent control totals.
- Experience with AWS, Git-based development, automated testing, CI/CD, and practical tradeoffs involving reliability, security, performance, and cost.
- The ability to work directly with customer technical teams, understand unfamiliar schemas, ask precise questions, and document decisions clearly.
- Comfort operating with significant ownership and ambiguity where customer implementation, platform development, security, and production operations overlap.
- The judgment to use AI development tools effectively while verifying generated code, tests, and transformations against source evidence.
You’d be a great fit if you also have:
- Experience with P&C insurance data, including policy transactions, coverages, claims, premium, exposure, rating, or underwriting data.
- Experience integrating with policy administration, claims management, rating, or other operational source systems.
- Deep experience with Databricks, Delta Lake, Unity Catalog, Databricks Workflows, or configuration-driven data pipelines.
- Experience with Terraform, secure file transfer, database replication, or customer-specific ingestion infrastructure.
- Experience supporting or building machine-learning feature pipelines, point-in-time datasets, model monitoring, MLflow, SageMaker, or production scoring systems.
$130k - $140k
...Lead Data Engineer Target Salary: $130,000 - $140,000 per year Join a dynamic organization dedicated to innovation and growth in the plasma donation industry. As a Lead Data Engineer, you will play a vital role in designing and optimizing scalable data pipelines that...SuggestedFull time$134.5k - $265.1k
Position Summary Deloitte is seeking a Lead Data Engineer- Databricks to support the design, build, and delivery of modern data and analytics solutions for clients across industries. In this role, you will help teams translate business needs into scalable data platform...SuggestedLocal areaVisa sponsorship$62k - $80k
...maintaining, and operating integrations, reporting pipelines, and data transformation systems. Qualifications: ● Passion for... ...operational & performance issues. ● Work with architecture/engineering leads and other teams to ensure quality solutions are...SuggestedFull timeVisa sponsorshipWork visa- ...Aventis Solutions is supporting a leading financial services organisation as it continues to scale a major data transformation. We’re looking for a Data Engineering Lead with 3+ years Databricks expertise to join a large scale data platform programme in Charlotte,...SuggestedPermanent employmentContract workRemote work2 days per week3 days per week
$90k - $95k
...technologies and optimization strategies span end-to-end Artificial Intelligence, Consulting, Digital, Cloud & DevOps, Data, and Software Engineering, servicing an array of noteworthy financial services and technology firms. Through research and development initiatives...SuggestedFull timeTemporary workFlexible hours$90k - $100k
...technologies and optimization strategies span end-to-end Artificial Intelligence, Consulting, Digital, Cloud & DevOps, Data, and Software Engineering, servicing an array of noteworthy financial services and technology firms. Through research and development initiatives...Full timeTemporary workFlexible hours- ...Responsibilities: Architecture & Engineering: Design, build, and optimize data structures, heavily utilizing advanced Snowflake capabilities to execute complex data transformations and business logic. System Integration: Manage NetSuite Data Model changes and...Full time
- ...Data Engineer Location: Hybrid - Charlotte, NC (onsite 3 days/week) Employment Type: 12+ Month Contract-to-Hire Pay: 55-68 hourly depending on experience Position Overview We are seeking a Data Engineer to join a Finance Data Engineering team supporting enterprise...Hourly payContract workLocal area3 days per week
- ...Data Engineer Mid-Level | Full-Time | On-Site | Reports to Director of FP&A Position Summary The Data Engineer is a mid-level individual contributor responsible for the systems and structures that make the company's data reliable, accessible, and useful. This...Full time
$82.42k - $126.6k
...driving digital transformation for financial institutions. We specialize in leveraging advanced technologies such as AI, cloud, and data-led innovation to help our clients accelerate growth and unlock business value. Our AI-driven solutions empower financial institutions...Full timeTemporary workRelocation$140k - $170k
OverviewThe Data Engineer II independently designs, builds, tests, deploys, and supports scalable data pipelines, integrations, models, and reusable platform components. This role translates defined business and analytical requirements into secure, reliable solutions and...$135k - $157k
...management solutions that advisors use in helping clients achieve wealth, independence, and purpose.The Job/What You'll Do:The Data Engineer (Analytics Engineering) will be a key technical leader responsible for the architecture, development, and optimization of mission...Full timeWork at officeFlexible hours$81.5k - $138.55k
...bring expertise in cloud, cybersecurity, enterprise architecture, data modernization, and digital transformation to support mission-... ...enterprise adopt modern data and AI solutions at scale. The Data Engineer builds and maintains data pipelines, transformations, semantic views...Full timeContract workWork experience placementLive inWork at officeImmediate startRemote work- Snowflake Data EngineerLocation RemoteWe are looking for a highly skilled Data Engineer with handson experience in Snowflake Python DBT and modern data architecture The ideal candidate will be responsible for designing building and maintaining scalable data pipelines and...
$105.4k - $207.8k
Position Summary Our Deloitte AI & Engineering team works to transform technology platforms, drive innovation, and help make a significant... ..., and fuel growth through innovation.Work you'll do As a Data Engineer III on the AI & Data team, you will be responsible for...Local area- Skanska is searching for a dynamic Data Engineer. This is a great opportunity to start a career with a company that builds things that matter and values its team. We are proud to share our culture of diversity and inclusion.Our work makes a clear contribution to society...Second jobLocal areaVisa sponsorship
- ...all.Who You’ll Work WithAt Slalom, we co-create custom software, data, AI, and cloud solutions with clients who are ready to... ...products, experiences, and organizations.Our Data Architecture & Engineering team sees data as the foundation for decision-making, artificial...Temporary workWork at officeLocal area
- ..., and government agencies, a list that spans across the country.Job DescriptionCapTech Machine Learning Engineers are responsible for designing and implementing data-driven solutions for our clients, with a specific focus on building and deploying scalable machine learning...Work at officeRemote workVisa sponsorshipWork visaFlexible hours
- ...Lead Data EngineerWithin COO Technology, Wells Fargo is seeking a Lead Data Engineer to help shape and scale our cloud-native data ecosystem. In this role, you will focus on Google Cloud Platform (GCP) services and frameworks, leading the design, build, and operation of...Work experience placementWork at officeRemote workFlexible hours
- ...Lead Data Engineer Location: Charlotte NC Duration: Long term Mandatory Skills: Advanced Oracle RDMS expertise, including object creation (tables, views, indexes, partitioned objects). Proficient in SQL and PL/SQL for data manipulation and analysis....Local area
$90 - $100 per hour
...Job Description Job Title: Lead Data Engineer Location: Charlotte, NC Duration: 12 months Pay Rate: $90 - $100/HR (W2 Only) Job/Role Description: Lead the design, architecture, and implementation of enterprise-scale data engineering solutions across...- ...Job Title: Lead Data Engineer Location: Charlotte NC Job Summary: We are seeking a highly skilled Lead Data Engineer to join our dynamic team. The ideal candidate will possess strong hands-on experience in advanced Oracle RDMS, SQL, PL...
- ***MUST HAVE CURRENT HEALTHCARE EXPERIENCE*** As a Lead Data Engineer, you will play a hands-on technical leadership role in evolving the data capabilities that power our client's solutions. You will lead complex initiatives across two closely connected areas: onboarding...
$189k - $238k
...The Talent Mine is partnering with a major US steel manufacturer to hire a Senior Data Engineer for an immediate, on-site FTE role in either Charlotte, NC or Irvine, TX . This client is the nation's #1 recycler of any product, a company with unmatched job security...Full timeH1bImmediate startVisa sponsorshipWork visa- FCG is seeking a Senior Data Engineer to work on our Central Data Platform team, utilizing Snowflake, dbt, and Fivetran, and responsible for end-to-end ELT development from ingestion of raw source system data through our integrated and data consumption layers in Snowflake...
$49.84 - $54 per hour
DescriptionWe are seeking a midlevel Data Engineer with hands‑on expertise in Hadoop, PySpark, Python, and real‑time streaming technologies such as Kafka. This role supports a large‑scale migration of on‑prem data and processing workloads into Google Cloud Platform (GCP...Contract workTemporary work$170k - $190k
...management solutions that advisors use in helping clients achieve wealth, independence, and purpose.The Job/What You'll Do:The Senior Data Engineer / Technical Lead is a pivotal, hands-on leadership role responsible for the end-to-end design, governance, and operational...Full timeWork at officeFlexible hours- THE GLOBAL LEADER IN DATA & ANALYTICS RECRUITMENTHarnham Search and Selection Company Number: 05723485Harnham Search and Selection is a registered company in England and Wales. Reg no. 05723485Harnham Europe Limited Company Number: 09956940Harnham GmbH HRB: 196954Harnham...Work at office
$128k - $249k
...search for talented visionaries and your search for important and impactful work lead to the same place.We are seeking a Senior Data Engineer to join the firm. The Senior Data Engineer contributes to the Enterprise Applications, Data, and AI Platforms group, designing scalable...Full timeTemporary workLocal areaRemote workFlexible hoursAfternoon shift$80k - $133k
Job Family:Data Science & AnalysisTravel Required:Up to 10%Clearance Required:Ability to Obtain Public TrustWhat You Will Do:Guidehouse seeks a Senior Data Engineer to design, develop, and optimize modern data platforms, pipelines, and cloud-based analytics solutions. The...Full timeContract workFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Data Engineer. Be the first to apply!
- data engineer machine learning Charlotte, NC
- aws data engineer Charlotte, NC
- big data developer Charlotte, NC
- sr data engineer Charlotte, NC
- big data cloud engineer Charlotte, NC
- finance data engineer Charlotte, NC
- junior big data engineer Charlotte, NC
- entry level data engineer Charlotte, NC
- sr information security engineer Charlotte, NC
- data center engineer Charlotte, NC



