Data Engineer
Soteris
ABOUT SOTERIS
Soteris is a YC-backed AI company building the future of pricing and product management for the
insurance industry. Our mission is to infuse the $5 trillion P&C insurance industry with best-in-class,
proprietary, AI-driven data analytics. We’ve spent years building our own proprietary AI models on
personal auto claims and exposure data to help insurers improve their loss ratios, with over 100 million
submissions and $180 billion in premium scored to date.
Each year, roughly $750 billion in insurance policies are written in the United States. Our machine
learning platform helps insurers evaluate policies at a granular level, moving beyond broad segmentation approaches that often lead to risks being over- or underpriced. Our modeling approach incorporates multiple model families and calibration methods to rank policy risk within a book of business.
We are a team of 10 and growing quickly. As our second Data Engineer, you will own the data layer that turns customer policy, claims, quote, and financial data into trusted inputs for actuarial analysis, model development, and production scoring. This is a builder role: you will work directly with customer data teams, create repeatable ingestion and transformation pipelines, reconcile outputs to source-of-truth control totals, and make the platform easier to operate as we add customers and products.
WHAT YOU'LL BE DOING
Customer Data Onboarding and Integration
- Leading the technical data workstream for new customer implementations by understanding policy, claims, rating, quote, and financial systems and establishing secure access to the data.
- Building reusable extraction and synchronization workflows for databases, backups, secure file transfer, APIs, and other delivery methods while preserving source lineage and supporting backfills.
- Mapping customer data into Soteris’s internal ontology and working with customer technical teams and internal project leads to resolve definitions, transformations, data-quality issues, and onboarding blockers.
Lakehouse and Pipeline Engineering
- Owning Databricks and AWS pipelines that move customer data from raw and Bronze ingestion through standardized Silver tables and curated Gold or model-ready datasets.
- Designing idempotent, incremental, observable workflows that handle schema evolution, late-arriving data, backfills, orchestration, performance, and cost.
- Developing shared components and configuration-driven patterns, then publishing well-defined datasets for actuarial analysis, backtesting, model training, production scoring, reporting, and monitoring.
Data Quality and Modeling Readiness
- Building automated quality gates for completeness, uniqueness, referential integrity, valid ranges, freshness, balance, schema drift, and other customer-specific controls.
- Reconciling written and earned premium, exposure, incurred and ultimate loss, claim counts, fee income, and other economic drivers to customer control statistics at the required state, program, year, and coverage levels.
- Partnering with data science to produce leakage-resistant, point-in-time-correct datasets and productionize approved actuarial reference data
Production Platform and Operational Ownership
- Owning data flows for quote requests, model inputs, and bound policy outcomes, including pre-live comparisons that confirm production request fields match corresponding policy data and post-launch drift checks.
- Operating pipelines with monitoring, alerting, recovery behavior, runbooks, and clear incident diagnostics across customer synchronization, Databricks processing, and downstream model- serving dependencies.
- Testing, reviewing, and deploying data code through GitHub and GitHub Actions, with automated tests and deployment controls for production data assets.
- Managing Databricks permissions, Unity Catalog controls, sensitive data, and least-privilege access in support of Soteris’s security and SOC 2 requirements.
OUR CURRENT STACK
- Python, SQL, PySpark, pandas, and related data-engineering libraries
- Databricks, Delta Lake, Databricks Workflows, and Unity Catalog
- AWS, including S3, Lambda, EC2, ECS, SageMaker, and secure customer file transfer
- Customer databases, backups, SFTP, APIs, and file-based ingestion
- Terraform, GitHub, GitHub Actions, automated testing, and infrastructure as code
- MLflow, SageMaker, and production scoring APIs at the model handoff boundary
- Modern generative AI development tools, including Claude, ChatGPT, Codex, or similar models
ABOUT YOU
You must have the following:
- Strong Python and SQL skills and the ability to write production-quality transformations, tests, utilities, and operational tooling.
- Hands-on experience building and operating production pipelines using Spark, Databricks, or a comparable distributed data platform.
- Strong understanding of data modeling, lakehouse or warehouse design, incremental processing, schema evolution, idempotency, backfills, and lineage.
- Experience designing data-quality controls and reconciling complex datasets to source systems or independent control totals.
- Experience with AWS, Git-based development, automated testing, CI/CD, and practical tradeoffs involving reliability, security, performance, and cost.
- The ability to work directly with customer technical teams, understand unfamiliar schemas, ask precise questions, and document decisions clearly.
- Comfort operating with significant ownership and ambiguity where customer implementation, platform development, security, and production operations overlap.
- The judgment to use AI development tools effectively while verifying generated code, tests, and transformations against source evidence.
You’d be a great fit if you also have:
- Experience with P&C insurance data, including policy transactions, coverages, claims, premium, exposure, rating, or underwriting data.
- Experience integrating with policy administration, claims management, rating, or other operational source systems.
- Deep experience with Databricks, Delta Lake, Unity Catalog, Databricks Workflows, or configuration-driven data pipelines.
- Experience with Terraform, secure file transfer, database replication, or customer-specific ingestion infrastructure.
- Experience supporting or building machine-learning feature pipelines, point-in-time datasets, model monitoring, MLflow, SageMaker, or production scoring systems.
$90k - $154k
...SynergisticIT commonly supports hiring pipelines for roles such as junior software programmer, Java full stack engineer, Python/Java developer, DevOps/cloud engineer, plus data-track roles like data analyst, BI analyst, data engineer, data scientist, and ML/AI engineer. The...SuggestedFull timeH1bShift work- ...Location: Remote We are seeking an experienced Data Engineer to support challenging data engineering, data management, and analytics initiatives. In this role, you will develop and maintain data pipelines and data structures that enable advanced analytics and data...SuggestedTemporary workLocal areaRemote work
- ...Responsibilities: Build and operate production-grade batch and streaming data pipelines across SQL Server, cloud applications, device/event... ...on-call activities. Work closely with Product, Application Engineering, Quality Assurance, DevOps/Site Reliability Engineering,...Suggested
- ...solutions that enable seamless communication and collaboration across the supply chain. DESCRIPTION: Transflo is seeking a Senior Data Engineer to architect and own our enterprise data platform — from raw ingestion through curated, analytics-ready data products. You will...SuggestedRemote work
$150k - $180k
...Lead Control Engineer // Data Center Developer Location: Austin, Texas (On-site)Relocation Needed Compensation: $150,000–$180,000 base About the Company A leading developer of hyperscale data centers, delivering scalable, efficient, and sustainable solutions...SuggestedRelocationFlexible hours- ...Data EngineerLocation: Hybrid (3x onsite) 2 location options: Houston TX (Downtown) Evansville IN 47708Project Details: Looking for a Data Engineer that will work in the IT and Data Analytics group at. They will help with developing, constructing, testing, and maintaining...Work experience placement
- ...optimization, regulatory compliance, and organizational success.This role is ideal for someone who enjoys solving complex problems, analyzing data, improving processes, and using technology to drive measurable business outcomes.What You'll Do Analyze charge capture, billing, and...Full time
$18 - $50 per hour
...university graduation and completion of the program Position Overview: Siemens Digital Industries Software is seeking an AI & Engineering Data Intern to support emerging Artificial Intelligence and Digital Thread initiatives within the Capital PreSales organization....Remote jobHourly payFull timeInternshipLocal area$90 per hour
...Database & Data Platform EngineerWe're looking for a skilled Database & Data Platform Engineer to join our Data Management team. In this role, you'll be responsible for supporting and optimizing SQL Server environments while helping advance our modern data platform strategy...Work at office- ...Great Place to Work® certification year after year. Principal Data Scientist Job requirements ~ Experience Range: 15 - 18... ...PySpark, and R, ensuring efficient data processing and feature engineering Develop, validate, and maintain probabilistic graph models and...
- ...DIRECTLY ON OUR W2 (EAD, OPT, USC, GC, H4) NO THIRD PARTY! Data Scientist We are seeking a Data Scientist to join a growing... ...degree in Data Science, Statistics, Mathematics, Computer Science, Engineering, or a related field. - FROM THE USA ~3+ years of experience...
- ...We are seeking a Data Scientist to join our Data Science team, with a specific focus on causal inference and marketing measurement... ...leaders. Cross-Functional Collaboration: Partner with Product and Engineering teams to help translate data science solutions into scalable,...
- ...Data Scientist Remote, Anywhere in the Continental US Due to the nature of the role, this position requires that you are a U... ...technical and non-technical audiences. Collaborate with policy, engineering, and operations teams to ensure analytics solutions align with...Full timeImmediate startRemote work
- ...ensure that Beaker and related laboratory applications are configured, maintained, and improved to support efficient workflows, accurate data, and exceptional patient care. Your expertise will play a critical role in connecting technology and laboratory operations to...Full timeRemote work
- ...skillsAbility to manage multiple priorities in a fast-paced healthcare environmentExperience working with databases, reporting tools, data analysis, and system integrationsAbility to work independently while collaborating effectively with cross-functional teamsWhat We...Full time
- Help Enhance Inpatient Care Through Healthcare TechnologyWe are seeking a skilled and detail-oriented Clinical Application Analyst III - ClinDoc to join our Digital Technology Services (DTS) team. In this role, you will support and optimize Epic ClinDoc workflows that help...Full timeWork at officeRemote work
$27 - $28 per hour
Mentor Community Services , a part of the Sevita family, provides community-based services for individuals with intellectual and developmental disabilities. Here we believe every person has the right to live well, and everyone deserves to have a fulfilling career. You...Hourly payFull timePart time- ...Evansville, IN. The Application Support Analyst works closely with senior and executive management to evaluate business process, data integrity, and compliance issues as they relate to the loan origination systems. The successful candidate will maintain cost-effective...Temporary workWork experience placementWeekend workAfternoon shift
$240k - $300k
...AI Engineer $240k–$300k base + equity Remote, United States High-growth applied AI company building production agents to automate... ...is building infrastructure that connects fragmented enterprise data, creates business context, and enables AI agents to execute high...Remote work$91k - $121k
...of competitive analysis, market positioning, salary structures, data maintenance, as well as developing compensation reports to... ...century, we've been at the forefront of designing, building, and engineering premium, award-winning products. Today, Marvin is also proud to...Temporary workRelocationRelocation package- ...Work-From-Home Data Entry Research PanelistJoin Our Team as a Work-From-Home Data Entry Research Panelist! Are you ready to earn money from the comfort of your own home? This exciting opportunity is perfect for anyone with a variety of skills and backgrounds – whether...Full timePart timeImmediate startRemote workWork from homeFlexible hours
$20 - $60 per hour
...What You’ll Do Collaborate with a multidisciplinary team of engineers and scientists to design, develop, and implement technology solutions... ...Notice This position may involve access to technology or data that is subject to U.S. export control laws, including the Export...Permanent employmentFull timeSummer workInternshipWork at office- Part Time Behavior Analyst Mentor Community Services, a part of the Sevita family, provides community-based services for individuals with intellectual and developmental disabilities. Here we believe every person has the right to live well, and everyone deserves to have...Part timeInternship
- ...Operations & Technology team, you’ll help share the client's marketing data strategy. This includes driving data governance and consistency... ...: Collaborate with cross-functional teams such as IT and engineering to integrate data from various sources into the customer data...Full time
$17.2 per hour
Job Description Lab Support Technician - Evansville, IN, Monday to Friday, 6:00 AM to 2:30 PM, with rotational weekends Pay range: $17.20+ per hour Salary offers are based on a wide range of factors including relevant skills, training, experience, education...Hourly payFull timePart timeWork experience placementMonday to FridayFlexible hoursWeekend work- ...Earn at Home by Taking Polls - Data Entry Clerk - Customer Service Rep - Work at Home & Part Time We are looking for people nationwide to participate in polls - Apply ASAP! We offer you the opportunity to earn extra income from home (teleworking) and also to decide...Extra incomePart timeImmediate startWork from home
- ...solutions.Our Corporate IT team has an opening for a Business Intelligence Analyst! The Business Intelligence Analyst will transform data into clear, actionable insights by developing reliable reporting solutions and partnering with stakeholders to align business needs...Full time
$32 - $38 per hour
Insight Global is Seeking a Junior C Developer with 1–3 years of hands‑on experience in Unix/Linux (Red Hat) environments, shell scripting, Oracle SQL, and Team Foundation Server (TFS). You’ll build and maintain C applications, automate workflows, and support production...- ...group/initiatives/projects: The team is partnering with J the Al data group, right to enable client with M365 agent adoption and then... ...onshore team. Other roles include product owners, data engineer’s various levels, infrastructure architects. Role/Responsibilities...Contract workRemote workFlexible hours
- ...approach to where we work. About the Role: As a Machine Learning Engineer, you will have the opportunity to collaborate closely with... ...problems executing on a high throughput system with dynamic data. About the Job: Design, develop, and deploy machine learning...Permanent employmentFull timeWork at officeRemote workWork from homeFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Data Engineer. Be the first to apply!
- data engineering intern summer Evansville, IN
- ai data Evansville, IN
- data loss prevention engineer Evansville, IN
- provider data management Evansville, IN
- health data Evansville, IN
- test data management Evansville, IN
- data cabling installation Evansville, IN
- data internship Evansville, IN
- data modeling Evansville, IN
- data recovery Evansville, IN





