Data Engineer
Soteris
ABOUT SOTERIS
Soteris is a YC-backed AI company building the future of pricing and product management for the
insurance industry. Our mission is to infuse the $5 trillion P&C insurance industry with best-in-class,
proprietary, AI-driven data analytics. We’ve spent years building our own proprietary AI models on
personal auto claims and exposure data to help insurers improve their loss ratios, with over 100 million
submissions and $180 billion in premium scored to date.
Each year, roughly $750 billion in insurance policies are written in the United States. Our machine
learning platform helps insurers evaluate policies at a granular level, moving beyond broad segmentation approaches that often lead to risks being over- or underpriced. Our modeling approach incorporates multiple model families and calibration methods to rank policy risk within a book of business.
We are a team of 10 and growing quickly. As our second Data Engineer, you will own the data layer that turns customer policy, claims, quote, and financial data into trusted inputs for actuarial analysis, model development, and production scoring. This is a builder role: you will work directly with customer data teams, create repeatable ingestion and transformation pipelines, reconcile outputs to source-of-truth control totals, and make the platform easier to operate as we add customers and products.
WHAT YOU'LL BE DOING
Customer Data Onboarding and Integration
- Leading the technical data workstream for new customer implementations by understanding policy, claims, rating, quote, and financial systems and establishing secure access to the data.
- Building reusable extraction and synchronization workflows for databases, backups, secure file transfer, APIs, and other delivery methods while preserving source lineage and supporting backfills.
- Mapping customer data into Soteris’s internal ontology and working with customer technical teams and internal project leads to resolve definitions, transformations, data-quality issues, and onboarding blockers.
Lakehouse and Pipeline Engineering
- Owning Databricks and AWS pipelines that move customer data from raw and Bronze ingestion through standardized Silver tables and curated Gold or model-ready datasets.
- Designing idempotent, incremental, observable workflows that handle schema evolution, late-arriving data, backfills, orchestration, performance, and cost.
- Developing shared components and configuration-driven patterns, then publishing well-defined datasets for actuarial analysis, backtesting, model training, production scoring, reporting, and monitoring.
Data Quality and Modeling Readiness
- Building automated quality gates for completeness, uniqueness, referential integrity, valid ranges, freshness, balance, schema drift, and other customer-specific controls.
- Reconciling written and earned premium, exposure, incurred and ultimate loss, claim counts, fee income, and other economic drivers to customer control statistics at the required state, program, year, and coverage levels.
- Partnering with data science to produce leakage-resistant, point-in-time-correct datasets and productionize approved actuarial reference data
Production Platform and Operational Ownership
- Owning data flows for quote requests, model inputs, and bound policy outcomes, including pre-live comparisons that confirm production request fields match corresponding policy data and post-launch drift checks.
- Operating pipelines with monitoring, alerting, recovery behavior, runbooks, and clear incident diagnostics across customer synchronization, Databricks processing, and downstream model- serving dependencies.
- Testing, reviewing, and deploying data code through GitHub and GitHub Actions, with automated tests and deployment controls for production data assets.
- Managing Databricks permissions, Unity Catalog controls, sensitive data, and least-privilege access in support of Soteris’s security and SOC 2 requirements.
OUR CURRENT STACK
- Python, SQL, PySpark, pandas, and related data-engineering libraries
- Databricks, Delta Lake, Databricks Workflows, and Unity Catalog
- AWS, including S3, Lambda, EC2, ECS, SageMaker, and secure customer file transfer
- Customer databases, backups, SFTP, APIs, and file-based ingestion
- Terraform, GitHub, GitHub Actions, automated testing, and infrastructure as code
- MLflow, SageMaker, and production scoring APIs at the model handoff boundary
- Modern generative AI development tools, including Claude, ChatGPT, Codex, or similar models
ABOUT YOU
You must have the following:
- Strong Python and SQL skills and the ability to write production-quality transformations, tests, utilities, and operational tooling.
- Hands-on experience building and operating production pipelines using Spark, Databricks, or a comparable distributed data platform.
- Strong understanding of data modeling, lakehouse or warehouse design, incremental processing, schema evolution, idempotency, backfills, and lineage.
- Experience designing data-quality controls and reconciling complex datasets to source systems or independent control totals.
- Experience with AWS, Git-based development, automated testing, CI/CD, and practical tradeoffs involving reliability, security, performance, and cost.
- The ability to work directly with customer technical teams, understand unfamiliar schemas, ask precise questions, and document decisions clearly.
- Comfort operating with significant ownership and ambiguity where customer implementation, platform development, security, and production operations overlap.
- The judgment to use AI development tools effectively while verifying generated code, tests, and transformations against source evidence.
You’d be a great fit if you also have:
- Experience with P&C insurance data, including policy transactions, coverages, claims, premium, exposure, rating, or underwriting data.
- Experience integrating with policy administration, claims management, rating, or other operational source systems.
- Deep experience with Databricks, Delta Lake, Unity Catalog, Databricks Workflows, or configuration-driven data pipelines.
- Experience with Terraform, secure file transfer, database replication, or customer-specific ingestion infrastructure.
- Experience supporting or building machine-learning feature pipelines, point-in-time datasets, model monitoring, MLflow, SageMaker, or production scoring systems.
$120k - $155k
The Senior Consultant I - Data Engineer provides technical expertise on project tasks under the supervision of a project manager to deliver quality services on schedule, within budget, and in alignment with customer requirements. In this role, the incumbent designs, builds...SuggestedWork experience placementRemote workFlexible hours$200k
...Lead Data Engineer - Remote 100% Remote (U.S. Based) Up to $200,000 Base Salary U.S. Citizens & Green Card Holders Only I'm partnering with a rapidly growing healthcare technology company that is looking to hire a Lead Data Engineer to help drive the evolution...SuggestedRemote work$100k - $120k
...Data Engineering - DevOps Fractal Analytics is a strategic AI partner to Fortune 500 companies with a vision to power every human decision in the enterprise. Fractal is building a world where individual choices, freedom, and diversity are the greatest assets. An ecosystem...SuggestedHourly payFull timeLocal areaRemote work- ...Quantiphi: Quantiphi is an award-winning, AI-First digital engineering and consulting company focused on delivering high-impact Services... ...by combining deep industry expertise, disciplined cloud and data engineering practices, and cutting-edge applied AI research. Our...SuggestedFull timeRemote work
- ...Senior Data EngineerAs a Senior Data Engineer, you will play a critical role in designing, building, and optimizing scalable data pipelines and architectures that support advanced analytics, machine learning, and business intelligence initiatives. You will own the full...SuggestedSummer workWork at officeRemote work
- ...Senior Data EngineerCGI is seeking a Senior Data Engineer with deep expertise in Palantir Foundry to join a fast paced data engineering team supporting large scale, cloud based financial data platforms. In this role, you'll build and maintain data pipelines, develop ontology...
- ...solutions that enable seamless communication and collaboration across the supply chain. DESCRIPTION: Transflo is seeking a Senior Data Engineer to architect and own our enterprise data platform — from raw ingestion through curated, analytics-ready data products. You will...Remote work
- ...Responsibilities: Build and operate production-grade batch and streaming data pipelines across SQL Server, cloud applications, device/event... ...on-call activities. Work closely with Product, Application Engineering, Quality Assurance, DevOps/Site Reliability Engineering,...
- ...Senior Data EngineerLocation: Knoxville, TN Type: Permanent Full Time Work Model: Hybrid – onsite and remoteResponsibilities• Build... ...financial services space.Requirements• 5+ years of hands-on data engineering experience, with a strong foundation in building and...Permanent employmentFull time
- ...Job Title: AWS Data Engineer Location: Knoxville, TN, Columbia SC or Lafayette, LA or Birmingham, AL Type: Direct Hire Responsibilities • Support the development, enhancement, and maintenance of Enterprise Risk Analytics applications within a financial services...Full timeLocal area
$150k - $180k
...Lead Control Engineer // Data Center Developer Location: Austin, Texas (On-site)Relocation Needed Compensation: $150,000–$180,000 base About the Company A leading developer of hyperscale data centers, delivering scalable, efficient, and sustainable solutions...RelocationFlexible hours- ...Operations division of the Information Technology Directorate at Oak Ridge National Laboratory (ORNL) to recruit a hands‑on Data Center Engineer who will contribute to the modernization of data center operations, especially around DCIM enhancements and PSC (Power, Space...Permanent employmentFull timeCasual work
- ...Senior AWS Data Engineer Location: Knoxville, TN, Columbia SC or Lafayette, LA or Birmingham, AL Senior AWS Data Engineer to support the development, enhancement, and maintenance of enterprise risk analytics applications within a large-scale financial services environment...
- ...Company Description Who We Are AMS Corporation is a nuclear engineering services company based in Knoxville, Tennessee, with a mission... ...aging diagnostics, calibration verification, and advanced data analysis. Our engineers work across nuclear power plants, research...Full timeWorldwide
- ...customers' business challenges, Take2 will work as a partner to best resolve client needs. Take2 is hiring an AWS Lakehouse Data Engineer who is eligible to be sponsored for a Public Trust Clearance. This position is Remote, but it will require you to work East Coast...Remote work
- ...Job Title: AWS AI/Data Engineer Location: Knoxville, Tennessee Type: Direct Hire Work Model: Hybrid – onsite and remote Overview System One is seeking an AWS AI/Data Engineer to support the development of AI-driven and analytical capabilities for...Full timeLocal areaRemote work
- Job Title:Measurement-Based Grid Monitoring Analytics and Control Engineer - Staff Level IIILocation:Charlotte, NC, Dallas, TX, Knoxville,... ...tools such as PSCAD.Analyze field-recorded measurement data and support event, disturbance, and post-contingency investigations...Full timeRemote workWork from homeFlexible hours
- ...Description \n \n Who We Are \n \n AMS is a nuclear engineering services company based in Knoxville, Tennessee, with a mission... ...online monitoring for condition-based maintenance, and custom data acquisition equipment development. Our engineers work across all...Worldwide
- Job DescriptionSenior Business Intelligence & Data EngineerIn-Office/Non-Remote in Maryville, TNWe are looking for a Modern BI / Data Warehouse Developer who can bridge traditional BI engineering (T‑SQL, ETL, SSIS—nice to have) with modern cloud analytics patterns, including...Work at officeRemote work
- ...Job Title: Senior AWS IAM Engineer Location: Lafayette, LA, Knoxville, TN, Columbia, SC, Birmingham, AL Job Type: Permanent Full... ...Engineer to support enterprise cloud identity services and AWS based data and analytics capabilities within a large scale financial...Permanent employmentFull timeLocal area
- ...Job Title: Senior AWS IAM Engineer Location: Lafayette, LA, Knoxville, TN, Columbia, SC, Birmingham, AL Job Type: Permanent... ...to support enterprise cloud identity services and AWS based data and analytics capabilities within a large scale financial services...Permanent employmentFull time
$145k - $165k
Job Title:Critical Power and Electromagnetic Threats - Simulation and Data AnalystLocation:Knoxville, TNJob Summary and Description:The position is for a senior level engineer with expertise in power system simulation, data analysis, and design. The candidate will provide...Full timeFlexible hours$18 - $50 per hour
...university graduation and completion of the program Position Overview: Siemens Digital Industries Software is seeking an AI & Engineering Data Intern to support emerging Artificial Intelligence and Digital Thread initiatives within the Capital PreSales organization....Remote jobHourly payFull timeInternshipLocal area- ...DIRECTLY ON OUR W2 (EAD, OPT, USC, GC, H4) NO THIRD PARTY! Data Scientist We are seeking a Data Scientist to join a growing... ...degree in Data Science, Statistics, Mathematics, Computer Science, Engineering, or a related field. - FROM THE USA ~3+ years of experience...
- ...Fortune Best Workplaces in Financial Services & Insurance Principal Data Scientist Job Responsibilities Lead the design and... ...anomaly detection, and probabilistic modeling. Partner with AI Engineering teams to productionize models and integrate them into enterprise...Remote work
- ...Data Scientist Remote, Anywhere in the Continental US Due to the nature of the role, this position requires that you are a U... ...technical and non-technical audiences. Collaborate with policy, engineering, and operations teams to ensure analytics solutions align with...Full timeImmediate startRemote work
- ...Great Place to Work® certification year after year. Principal Data Scientist Job requirements ~ Experience Range: 15 - 18... ...PySpark, and R, ensuring efficient data processing and feature engineering Develop, validate, and maintain probabilistic graph models and...
- ...strengthening the resilience of the U.S. financial system. Data Scientists at Zest AI use the power of machine learning (ML)... ...explainability, and regulatory constraints Collaborate with ML engineers, product leaders, and domain experts to bring models from...Work at office
- ...We are seeking a Data Scientist to join our Data Science team, with a specific focus on causal inference and marketing measurement... ...leaders. Cross-Functional Collaboration: Partner with Product and Engineering teams to help translate data science solutions into scalable,...
- ...: Product Development Manager Cooperates with: Project Engineering, Project Management, Field Service, & Operations What You... ...within a cross-functional team Develop test plans, analyze data gathered, and develop mathematical models from data Characterize...Work at officeFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Data Engineer. Be the first to apply!
- data engineer machine learning Knoxville, TN
- aws data engineer Knoxville, TN
- sr data engineer Knoxville, TN
- finance data engineer Knoxville, TN
- sr information security engineer Knoxville, TN
- data center engineer Knoxville, TN
- senior data integration developer Knoxville, TN
- senior cloud data engineer Knoxville, TN
- data engineer Knoxville, TN
- data engineer analytics Knoxville, TN






