Data Engineer
Soteris
ABOUT SOTERIS
Soteris is a YC-backed AI company building the future of pricing and product management for the
insurance industry. Our mission is to infuse the $5 trillion P&C insurance industry with best-in-class,
proprietary, AI-driven data analytics. We’ve spent years building our own proprietary AI models on
personal auto claims and exposure data to help insurers improve their loss ratios, with over 100 million
submissions and $180 billion in premium scored to date.
Each year, roughly $750 billion in insurance policies are written in the United States. Our machine
learning platform helps insurers evaluate policies at a granular level, moving beyond broad segmentation approaches that often lead to risks being over- or underpriced. Our modeling approach incorporates multiple model families and calibration methods to rank policy risk within a book of business.
We are a team of 10 and growing quickly. As our second Data Engineer, you will own the data layer that turns customer policy, claims, quote, and financial data into trusted inputs for actuarial analysis, model development, and production scoring. This is a builder role: you will work directly with customer data teams, create repeatable ingestion and transformation pipelines, reconcile outputs to source-of-truth control totals, and make the platform easier to operate as we add customers and products.
WHAT YOU'LL BE DOING
Customer Data Onboarding and Integration
- Leading the technical data workstream for new customer implementations by understanding policy, claims, rating, quote, and financial systems and establishing secure access to the data.
- Building reusable extraction and synchronization workflows for databases, backups, secure file transfer, APIs, and other delivery methods while preserving source lineage and supporting backfills.
- Mapping customer data into Soteris’s internal ontology and working with customer technical teams and internal project leads to resolve definitions, transformations, data-quality issues, and onboarding blockers.
Lakehouse and Pipeline Engineering
- Owning Databricks and AWS pipelines that move customer data from raw and Bronze ingestion through standardized Silver tables and curated Gold or model-ready datasets.
- Designing idempotent, incremental, observable workflows that handle schema evolution, late-arriving data, backfills, orchestration, performance, and cost.
- Developing shared components and configuration-driven patterns, then publishing well-defined datasets for actuarial analysis, backtesting, model training, production scoring, reporting, and monitoring.
Data Quality and Modeling Readiness
- Building automated quality gates for completeness, uniqueness, referential integrity, valid ranges, freshness, balance, schema drift, and other customer-specific controls.
- Reconciling written and earned premium, exposure, incurred and ultimate loss, claim counts, fee income, and other economic drivers to customer control statistics at the required state, program, year, and coverage levels.
- Partnering with data science to produce leakage-resistant, point-in-time-correct datasets and productionize approved actuarial reference data
Production Platform and Operational Ownership
- Owning data flows for quote requests, model inputs, and bound policy outcomes, including pre-live comparisons that confirm production request fields match corresponding policy data and post-launch drift checks.
- Operating pipelines with monitoring, alerting, recovery behavior, runbooks, and clear incident diagnostics across customer synchronization, Databricks processing, and downstream model- serving dependencies.
- Testing, reviewing, and deploying data code through GitHub and GitHub Actions, with automated tests and deployment controls for production data assets.
- Managing Databricks permissions, Unity Catalog controls, sensitive data, and least-privilege access in support of Soteris’s security and SOC 2 requirements.
OUR CURRENT STACK
- Python, SQL, PySpark, pandas, and related data-engineering libraries
- Databricks, Delta Lake, Databricks Workflows, and Unity Catalog
- AWS, including S3, Lambda, EC2, ECS, SageMaker, and secure customer file transfer
- Customer databases, backups, SFTP, APIs, and file-based ingestion
- Terraform, GitHub, GitHub Actions, automated testing, and infrastructure as code
- MLflow, SageMaker, and production scoring APIs at the model handoff boundary
- Modern generative AI development tools, including Claude, ChatGPT, Codex, or similar models
ABOUT YOU
You must have the following:
- Strong Python and SQL skills and the ability to write production-quality transformations, tests, utilities, and operational tooling.
- Hands-on experience building and operating production pipelines using Spark, Databricks, or a comparable distributed data platform.
- Strong understanding of data modeling, lakehouse or warehouse design, incremental processing, schema evolution, idempotency, backfills, and lineage.
- Experience designing data-quality controls and reconciling complex datasets to source systems or independent control totals.
- Experience with AWS, Git-based development, automated testing, CI/CD, and practical tradeoffs involving reliability, security, performance, and cost.
- The ability to work directly with customer technical teams, understand unfamiliar schemas, ask precise questions, and document decisions clearly.
- Comfort operating with significant ownership and ambiguity where customer implementation, platform development, security, and production operations overlap.
- The judgment to use AI development tools effectively while verifying generated code, tests, and transformations against source evidence.
You’d be a great fit if you also have:
- Experience with P&C insurance data, including policy transactions, coverages, claims, premium, exposure, rating, or underwriting data.
- Experience integrating with policy administration, claims management, rating, or other operational source systems.
- Deep experience with Databricks, Delta Lake, Unity Catalog, Databricks Workflows, or configuration-driven data pipelines.
- Experience with Terraform, secure file transfer, database replication, or customer-specific ingestion infrastructure.
- Experience supporting or building machine-learning feature pipelines, point-in-time datasets, model monitoring, MLflow, SageMaker, or production scoring systems.
$120k - $140k
...Acunor is Hiring: Lead Data Engineer Location: NY / NJ / Pittsburgh – Onsite Employment: Full-Time Compensation: $120K–$140K + Benefits Experience: 8+ Years We are looking for a hands-on Lead Data Engineer to design and deliver enterprise-scale data...SuggestedFull time- ...and talents will be valued and celebrated. Together we will create a brighter future and make a meaningful difference.As a Lead Data Engineer at JPMorganChase within the Corporate Sector, you are an integral part of an agile team that works to enhance, build, and deliver...Suggested
- ..., and agility needed to move forward with confidence. Visit to know more about us. Role : Senior Python Site Reliability Engineer Location : Pennington, NJ/ Jersey City, NJ Mode : Hybrid ( (Hybrid, minimum 3 days onsite per week) Job Description We...SuggestedFull time3 days per week
$82.42k - $126.6k
...driving digital transformation for financial institutions. We specialize in leveraging advanced technologies such as AI, cloud, and data-led innovation to help our clients accelerate growth and unlock business value. Our AI-driven solutions empower financial institutions...SuggestedFull timeTemporary workRelocation$105.4k - $207.8k
Position Summary Our Deloitte AI & Engineering team works to transform technology platforms, drive innovation, and help make a significant... ..., and fuel growth through innovation.Work you'll do As a Data Engineer III on the AI & Data team, you will be responsible for...SuggestedLocal area- Job ID: 25727970Reference Number: 25-00659Title: Data EngineerLocation: Jersey City, NJ, 08830Posted Date: 2025-06-17Company: HAN Staffing Role: Data EngineerLocation: NJ5 Day's OnsiteExperience: 8+ Years AWS - 25%, Databricks- 25%, Pyspark -25%, Splunk- 25% Skill : AWS...
$142.32k - $213.48k
...pride. If you are a problem solver who seeks passion in your work, come join us. We’ll enable growth and progress together.The Sr Data Engineer, AVP is an intermediate level position responsible for participation in the establishment and implementation of new or revised...Full time- Be part of a dynamic team where your distinctive skills will contribute to a winning culture and team. As a Data Engineer III at JPMorganChase within the Corporate Technology , you serve as a seasoned member of an agile team to design and deliver trusted data collection...
- Job ID: 25659978Reference Number: 25-00597Title: Data Analyst EngineerPosted Date: 2025-06-06Company: HAN Staffing Position : Data Analyst Engineer Location : Jersey City, Wilmington, Chicago, Plano, Seattle, Palo Alto. Contract : w2 Job Description: Experience : 7+ Experience...Contract work
$200k
...Lead Data Engineer - Remote 100% Remote (U.S. Based) Up to $200,000 Base Salary U.S. Citizens & Green Card Holders Only I'm partnering with a rapidly growing healthcare technology company that is looking to hire a Lead Data Engineer to help drive the evolution...Remote work- ...specializing in the fields of Software Development, Software Consultancy, and Information Technology Enabled Services.Job DescriptionOur Big Data capability team needs hands-on developers who can produce beautiful & functional code to solve complex analytics problems. If you...Permanent employmentFull timeH1b
- Job TitleOkay with relocation.5 Openings total. USC/GC/H4 Only!Client: Fidelity location: Jersey City, NJ (Hybrid) Duration: 12 Month+ Pay: $68/Hr W2Need LinkedInMust Have Skills: Skills wise we need a Senior Level Oracle Database Developer, that is the number one skillset...Relocation
- ...Job Title5+ years of experience in data engineeringStrong proficiency with Databricks, Spark, PySpark/ScalaAdvanced skills in Power BI, DAX, and data visualization best practicesExperience with Azure Data Lake, Data Factory, and SQLStrong problem-solving skills and experience...
$88.59 - $96.59 per hour
...Data EngineerGenesis10 is currently seeking a Data Engineer for a contract position with a Global Financial Institution located in Jersey City, NJ. This is a 12+ month contract opportunity.This role is heavily focused on building and expanding the organization's Apache...Contract work- ...Data EngineerWe are seeking an experienced professional with strong expertise in Snowflake, Python, Oracle PL/SQL, and reporting/BI... ...Generative AI concepts and tools where relevant to enhance data engineering, analytics, automation, and productivity use cases.Continuously...
$104k - $157.7k
...Job Overview The Data Management, Database Specialist will drive initiatives for data lineage, platform migration, and regulatory... ...enabling data‑informed decision‑making. Mentor Data Engineers, monitor key performance indicators, and enforce internal controls...Shift workDay shift- ...Data Engineer - LOCAL ONLY - NYC OR NJLong Term2-3 years of experience, Snowflake, PYTHON, db and Azure products like ADF IS A MUSTSQL is a must- will be sql assessmentComputer Science degree- from reputable school (BIG PLUS)Bachelor's Degree in Mathematics, Operations...Local area
- ...Data EngineerDuration: Long Term ContractLocation: Jersey City, NJClient: Mphasis/ BankingJob ResponsibilitiesExecutes creative software... ...coding hygiene and system architecture.Contributes to software engineering communities of practice and events that explore new and...
$100k - $120k
...Data Engineering - DevOps Fractal Analytics is a strategic AI partner to Fortune 500 companies with a vision to power every human decision in the enterprise. Fractal is building a world where individual choices, freedom, and diversity are the greatest assets. An ecosystem...Hourly payFull timeLocal areaRemote work- ...Senior Data EngineerDuration: Long Term contractLocation: Durham, NC or Jersey City, NJ or Westlake, TX or Boston, MA or Salt Lake... ...month)Job Description:Bachelor’s degree in Computer Science or Engineering.9+ years’ experience developing data solutions and data movement...
$80 - $85 per hour
...Senior Data EngineerSeeking a highly skilled Senior Data Engineer with 8+ years of hands-on experience in enterprise data engineering, including deep expertise in Apache Airflow DAG development, dbt Core modeling and implementation, and cloud-native container platforms...Hourly payContract work$70.23 per hour
...Data EngineerGenesis10 is currently seeking a Data Engineer position with a Global Financial Institution located in Jersey City, NJ. This is a hybrid 6 month contract opportunity.This role is responsible for building scalable ingestion, transformation, and storage pipelines...Contract work- ...Snowflake Data Platform Operations EngineerAs a Snowflake Data Platform Operations Engineer, you will play a critical role in maintaining the reliability and resilience of DTCC's cloud-based data infrastructure. Your contributions will directly support the flawless clearing...Remote work
- ...Data Engineering Manager Overview Seeking a technical leader to oversee enterprise data solutions, lead engineering initiatives, and deliver scalable platforms that support analytics and business operations. Key Responsibilities Lead data engineering projects from design...For contractorsWork at officeLocal area
- ...Senior Data Engineer With Databricks Remote, Anywhere in the Continental US Due to the nature of the role, this position requires U.S. Citizenship. Data Surge is disrupting the services industry with cutting-edge technology that brings together the very best...Full timeImmediate startRemote work
- ...Location: Remote We are seeking an experienced Data Engineer to support challenging data engineering, data management, and analytics initiatives. In this role, you will develop and maintain data pipelines and data structures that enable advanced analytics and data...Temporary workLocal areaRemote work
- ...Key Requirements• 0-2 years of overall experience in data engineering, building and maintaining large-scale data pipelines and platforms.• Proficiency in programming languages such as Python and PySpark.• Familiar with AWS services for data processing and storage (e.g...
- ...Data EngineerWe are seeking a skilled Data Engineer to design, develop, and support modern data platforms and pipelines that power analytics, reporting, and AI-driven solutions. The ideal candidate will have strong expertise in Snowflake, Python, SQL, AWS, Oracle, and...
- ...Purple DriveWe are seeking a highly skilled Palantir Data Engineer to design, develop, and maintain scalable data solutions using Palantir Foundry and associated technologies. The ideal candidate will have hands-on experience in data modeling, pipeline building, and analytics...
$142.32k - $213.48k
...Job Overview: The Sr. Data Engineer, AVP is an intermediate level position responsible for participation in the establishment and implementation of new or revised application systems and programs in coordination with the Technology team. The overall objective of this role...
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Data Engineer. Be the first to apply!
- data engineer machine learning Jersey City, NJ
- aws data engineer Jersey City, NJ
- big data developer Jersey City, NJ
- data engineer hedge fund Jersey City, NJ
- sr data engineer Jersey City, NJ
- big data cloud engineer Jersey City, NJ
- finance data engineer Jersey City, NJ
- junior big data engineer Jersey City, NJ
- sr information security engineer Jersey City, NJ
- data center engineer Jersey City, NJ



