Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Founding Engineer, Commerce Data Supply Chain

$400k

Product.ai

Own every code and merchant fact from source to shelf, across hundreds of thousands of stores: what enters, how it is normalized, and how fast it goes stale. Product.ai is the verified truth layer for shopping. When a person or an AI agent needs to know what is actually true about a purchase, we answer with proof. SimplyCodes is the first proof at scale, the code verification service that shows shoppers codes that actually work. It earns about $22 million a year at roughly 60% margins. We are 100% founder-owned, profitable, and bootstrapped since 2009. No outside investors, no board. Fewer than twenty operators, outbuilding companies 10x our size. Why This Role Exists Every claim we publish stands on a supply chain of facts. Codes and merchant facts arrive by the hundreds of thousands from networks, feeds, and merchant surfaces we do not control, each in its own shape and wrong in its own way. They resolve into one canonical shelf: this code, this merchant, these terms. And they start dying the moment they land. Codes expire and terms change without notice. Freshness is the fuel of the engine, because a stale shelf is a wall of dead codes, and dead codes are what we exist to kill. That supply chain has never had a single owner. The pipelines run and the revenue flows, but no one seat has ever decided what enters, what record it becomes, and how fast each fact is rechecked or retired. This is a founding seat, the whole intake-to-shelf path in one pair of hands, working directly with the founder. One layer is mid-transition. Row-level validation, the industry's last manual stronghold, is becoming an agent fleet here. LLM pipelines read and check at machine scale; humans own the verdicts. You inherit an agent-fleet design problem, not a headcount. One boundary, drawn on purpose. Downstream of you, a robot fleet proves codes at real checkouts, and a sibling seat owns that proof. You own everything upstream, including the clock that decides when a fact must be proved again. They prove what you shelve; you decide what is worth proving. Two seats, one loop. The System You'll Need to Model Ingestion from sources you do not control. The problem class every data platform faces: dozens of third-party sources, each with its own schema, lag, and error habits. The craft is contracts at the edge, so you build schema-drift detection, quarantine lanes for suspect rows, and pipelines that fail loudly instead of silently. Entity resolution, the wall the shelf hangs on. The same merchant arrives under five names; the same code from three sources with two expiry dates. Record linkage and dedupe decide whether the shelf ends up with one truth or three almost-truths, and the same discipline runs behind every serious catalog, listings, and knowledge-graph system. A precision error here becomes a public claim with our name on it. Freshness as an economic frontier. Commerce facts decay on their own clocks, and a code can die an hour after it arrives. Every class of fact has a staleness budget and a recheck cost. Verify too often and the pipeline eats its own margin; too rarely and the shelf rots. You are setting data-quality SLOs where the SLO is the product. Agent fleets as the validation workforce. Does this code look real, does this merchant match, do these terms parse? Those row-by-row judgment calls run as LLM pipelines here, and a human owns every verdict. The craft is evaluation design, which means golden sets, adversarial cases, precision floors that prove the fleet grades rows right, and a cost line that proves the fleet pays for itself. Cortex, the brain you build inside. You work inside Cortex, the shared AI brain that runs the company and the product family we sell; it answers its own questions from more than 8,600 documents. The company moves at that speed, and no spec stays current for a quarter. Nobody hands you a brief, so you model where the system is going and meet it there. If reading that energizes you, keep going. If it feels overwhelming or underspecified, this isn't the right fit. What You Will Own The commerce data supply chain, end to end. Every code and merchant fact from source to shelf: what enters, the one canonical record it becomes, when it is retired. The system class behind every catalog, price-intelligence, and listings platform, pointed here at the data under the revenue engine that pays for everything we build. Agents write much of the code; you own the design, the failure modes, and the verdict on what ships. The agent validation fleet. The fleet of LLM validators that does row-level judgment at machine scale. You design it, write the evals that prove its precision, and price it: fleet architecture, eval harnesses, human-verdict escalation paths, a cost line you can defend. No direct reports. The fleet is the team. Freshness as your number. Verified-freshness coverage across hundreds of thousands of stores, meaning how much of the shelf is machine-verified and inside its staleness budget. It is the fuel gauge of the business, and yours to move. Developers and AI agents now buy this data through paid keys, so a stale row is a broken promise to a paying machine. The seat, chartered. Inside your first quarter you co-sign a seat charter with one machine-checkable number that proves the seat works, plus a written split of what you decide alone and what you bring to the founder first. Walking in, you must already own large-scale ingestion, entity resolution and dedupe, freshness SLAs, and pipeline unit economics. You grow into agent fleets as a production workforce, LLM evaluation design, and a data platform run like a P&L. Who You Are You reason about data systems in invariants and lifecycles. Handed a shelf of facts you have never seen, you first ask where each row came from, what would make it wrong, and when it dies. You see the pattern behind the pile, like the five sources behind one duplicate or the decay class behind one stale code. You notice when your model is wrong and update fast, and you write clearly, because on a team this small the written spec is the meeting. You move between architecture and shipped code without ceremony, and a resolution design in the morning can be processing real rows by night. Agents are your production workforce; you direct them and verify what comes back. You can do this job by hand and prove it, and that mastery is what lets you trust, or reject, what an agent hands you. The expensive thing here is a redo cycle, never the compute. You have built and run production data pipelines at real scale: ingestion from third parties you did not control, entity resolution or dedupe where a bad merge cost something, data-quality guarantees someone else depended on. That record can come from catalog and listings platforms, price intelligence, ad or affiliate data, search indexing, or knowledge graphs. All the same discipline. What counts is that your guarantees held. Python and SQL are daily tools; Airflow, Dagster, or Temporal are familiar ground; your batch-versus-streaming opinions come with incident stories attached. If you have run LLM pipelines against golden sets, better still. If not, you will learn that here fast. We care about the artifact and the reasoning far more than where you did it, and there is no degree to check. Who this isn't for. This seat is wrong if you guard one lane and call the rest someone else's department; the supply chain runs from raw source to public shelf, and you own all of it. It is wrong if you need a finished spec and a groomed queue before you can move, or a platform team underneath you to feel senior. It is wrong if you pick tools for the resume line rather than for what the pipeline needs tonight. And it is wrong if you would ship what an agent handed you without being able to say why it is right, or let the fleet grade its own homework. You will be happiest here if your idea of craft is a shelf that is never silently wrong, and a supply chain you can defend row by row. How We Evaluate We don't run traditional engineering interviews. We evaluate demonstrated performance on work-relevant tasks, in four steps. Async video screen . Brief and on your own time — about fifteen minutes. We want to see how you think, not how you present. Calls with company stakeholders . Short conversations with the people you would build beside. Conversation with the founder . How you reason about freshness, cost, and truth at supply-chain scale — and where you push back. Paid work trial . Four days, paid, on real work in our real environment — a live piece of the supply chain, taken from grounding to a change you can prove. We watch how you get grounded, whether you write the spec before the build, how you verify what your agents produce, and whether your self-assessment is honest. We both learn more in four days than in forty hours of interviews. We hire on the work and the reasoning, not the pedigree. Compensation & Ownership Total first-year comp: $400,000 – $500,000 — base, plus performance-based ownership and profit-share programs. Base: $280,000 – $330,000 , top of market for senior data engineering. Eligibility for the company's ownership and profit-share programs — grants are performance-based, with terms discussed at the offer stage; 100% family premium coverage; an AI tooling budget steered by return, never capped. The model is built to mint partners. Based in Santa Monica, Los Angeles — in person, five days a week. The rooms are real rooms. Relocation support available for the right builder. #J-18808-Ljbffr Product.ai

Vacancy posted 5 days ago
Similar jobs that could be interesting for youBased on the Founding Engineer, Commerce Data Supply Chain in Los Angeles, CA vacancy
  • $128k - $160k

    Grailed is looking for a Senior Data Engineer to help us build and scale the data infrastructure...  ...in retail, resale, or marketplace commerce, genuine curiosity about the domain a must...  ...million members across 170 countries. Founded in 2013, Grailed is the leading... 
    Suggested
    Local area

    Grailed

    Los Angeles, CA
    6 days ago
  • Senior Data Engineer Beverly Hills About Us Fashion Nova is a leading, digital and social commerce platform, and is one of the most visible fashion brands today. Fashion Nova was founded in 2006 when our first store opened its doors. Fast forward to 2013, Fashionnova.com... 
    Suggested
    Summer work

    Fashion Nova

    Beverly Hills, CA
    4 days ago
  • $105.4k - $207.8k

     ...Consulting's Innovation & Delivery Transformation (I&DT) practice. I&DT brings an engineering- and innovation-led mindset to how Deloitte builds, delivers, and scales technology-enabled solutions. Data Studio is Converge for Healthcare's foundational platform — combining... 
    Suggested
    Local area

    Deloitte

    Los Angeles, CA
    7 hours ago
  • $15k

     ...ways of thinking, we build better experiences for our members and our team. The Team Data influences all of the decisions we make at Tinder. As a Senior Software Engineer, Data on the Analytics Engineering team, you will have the opportunity to build critical... 
    Suggested
    Full time
    Work experience placement
    Work at office
    Relocation

    Match Group

    Los Angeles, CA
    1 day ago
  •  ...access the full potential of digital money. We are looking for engineers who enjoy building and shipping quickly, learning how other...  ...entire company operates. As a Senior Software Engineer on the Data & AI team, you'll expand Lighthouse, our internal data and AI platform... 
    Suggested
    Full time
    Worldwide
    Relocation

    Lightspark

    Culver City, CA
    1 day ago
  •  ...Staff Data Engineer The Staff Data Engineer is an expert data handler and developer who has demonstrated their capacity for solving problems...  ...flagship business, now known as Rocket Mortgage®, which was founded in 1985. Today, we're a publicly traded company involved in... 

    Rocket

    Los Angeles, CA
    3 days ago
  • $140k - $200k

     ...- Speechify has no office. These include frontend and backend engineers, AI research scientists, and others from Amazon, Microsoft, and...  ...their own companies. Overview We're looking to hire for our Data side of our AI team at Speechify. This role is responsible for... 
    Full time
    Work at office
    Shift work

    Speechify

    Los Angeles, CA
    1 day ago
  • $98.4k - $147.6k

     ...we co-create moments that matter – both for our audiences and our employees – and aim to leave a positive mark on culture. Data Engineer – Data Pipeline & ETL 46034 Overview and Responsibilities Job Summary The Data Engineering team is hiring a Data... 
    Burbank, CA
    more than 2 months ago
  • SpaceX is seeking a Data Engineer to join the Starlink Enterprise team in Hawthorne, CA, focusing on improving data quality and building AI-enabled integrations across sales, channel, and enablement workflows. You will partner with Enterprise sales, channel and reseller... 

    SPACE EXPLORATION TECHNOLOGIES CORP

    Hawthorne, CA
    2 days ago
  • $180k - $210k

    Job Title: Sr. Data Engineer Team: Engineering Location: Onsite-Santa Monica, CA Employment Type: Full-time, Salaried, Exempt Reports To:...  ...gives dealers a dedicated channel for sourcing used EV inventory. Founded in 2023 and led by former Tesla executives, Plug has... 
    Full time
    Work at office
    Relocation
    Relocation package

    Plugmotors

    Santa Monica, CA
    1 day ago
  • The Walt Disney Company seeks a Senior Data Engineer in Santa Monica, CA to join the Data Measurement team. You’ll help define, measure, and certify data pipelines that drive business insights for Disney's streaming, sports, and advertising platforms. Core responsibilities... 

    The Walt Disney Company

    Glendale, CA
    3 days ago
  • $115k - $150k

     ...Attachments Salary Range: $115,000.00 To $150,000.00 Annually Data Engineer — Everytable Salary range : $115,000 - $150,000 Department:...  ...Technology OUR STORY + WHAT MAKES US SPECIAL Everytable was founded on the belief that healthy food is a human right and shouldn’t... 
    Full time
    Local area

    Everytable

    Los Angeles, CA
    5 days ago
  • $165k - $190k

    About Thrive Market Thrive Market was founded in 2014 with a mission to make healthy and sustainable living easy and affordable for...  ...of Americans in the years to come. THE ROLE Thrive Market’s Data Engineering team is seeking a Senior Data Engineer! We are looking for a... 
    Summer work
    Work at office
    Flexible hours

    Thrive Market

    Los Angeles, CA
    6 days ago
  • $141.4k - $204.4k

     ...ideas matter. A team where everyone makes play happen. Job Title: Gameplay Data Science Engineer Location: Hybrid - Vancouver or Los Angeles Reports to: Technical Director Job Summary: Founded in 2010, Respawn was created with the philosophy that when talented people... 
    Full time
    Local area

    Electronic Arts

    Los Angeles, CA
    4 days ago
  • A leading AI research organization is seeking a Signal Integrity System Design Engineer to lead system design for advanced AI workloads. Responsibilities include collaborating with engineering partners to develop high-speed technologies and evaluating new methodologies... 
    Relocation package

    OpenAI

    Los Angeles, CA
    5 days ago
  • Build the Data Backbone of a Healthcare AI Startup Join a stealth-mode healthcare venture as their first Data Engineer and lay the groundwork for real-time, high-impact data systems that...  ....S. only) About the Role Youll be the founding Data Engineer at a healthcare AI... 
    Immediate start
    Remote work
    Visa sponsorship
    Work visa
    Flexible hours

    AllyNd Partners

    Los Angeles, CA
    3 days ago
  • $155.4k - $208.4k

    Sr Data Engineer - Req ID 10153284 Technology is at the heart of Disney’s past, present, and future. Disney Entertainment and ESPN Product...  ...on the level and position offered. Job Posting Segment: Commerce, Data & Identity Primary Business: PE - Sports, News & Entertainment... 
    Full time
    Work experience placement

    5014 Disney Entertainment & Sports LLC

    Santa Monica, CA
    3 days ago
  • WME Enterprise IT is seeking a Senior Data Engineer to own and modernize the SQL data estate that underpins analytics and reporting across our global footprint. You’ll stabilize legacy environments, build ELT pipelines from multiple sources, and partner with the enterprise... 

    IMG LIVE

    Beverly Hills, CA
    6 days ago
  • $125k - $150k

    SpaceX was founded under the belief that a future where humanity is out exploring the stars is fundamentally more exciting than one...  ...possible, with the ultimate goal of enabling human life on Mars. DATA ENGINEER (STARLINK GO-TO-MARKET) Starlink is the world largest... 
    Permanent employment
    Temporary work
    Weekend work

    InvestedintheMission

    Hawthorne, CA
    3 days ago
  • $125k - $150k

    SpaceX was founded under the belief that a future where humanity is exploring the stars is fundamentally more exciting than one where...  ...terminals, and the software that ties it all together. As a Data Engineer on the Starlink Enterprise team, you will improve the quality... 
    Permanent employment
    Temporary work
    Internship
    Worldwide
    Weekend work

    SPACE EXPLORATION TECHNOLOGIES CORP

    Hawthorne, CA
    2 days ago
  • $170k - $200k

    Tatari is on a mission to revolutionize TV advertising. Founded in 2016 to help transform the antiquated world of TV advertising through...  ...businesses of any size to advertise on TV. As a Senior Data Engineer in our Reporting and Measure pillar, you will design, build and... 
    Work at office
    2 days per week

    Tatari

    Los Angeles, CA
    6 days ago
  • True Anomaly in Los Angeles area seeks an entry-level Data Engineer I to build real-world data systems using Python and SQL in a fast-paced startup. You will learn quickly, own small data pipelines, and collaborate with experienced engineers to deliver reliable data products... 

    Trueanomalyinc

    Los Angeles, CA
    6 days ago
  • $144.77k - $209.11k

     ...Contact Government Services) is seeking a passionate and driven Data Engineer to support a rapidly growing Data Analytics and Business...  ...Email: ****@*****.*** Tools and more information can be found on our website at or the job board at #J-18808-Ljbffr CGS Federal... 
    Full time
    Flexible hours

    CGS Federal (Contact Government Services)

    Los Angeles, CA
    6 days ago
  • SpaceX was founded under the belief that a future where humanity is out exploring the stars is fundamentally more exciting...  ...goal of enabling human life on Mars.PROPULSION ENGINEER, PROPULSION SIMULATION & DATA ANALYSIS (RAPTOR) The Raptor Systems Modeling and Control... 
    Permanent employment
    Temporary work

    SpaceX

    Hawthorne, CA
    2 days ago
  •  ...Lead and manage a team of data engineers delivering enterprise-grade data products, data marts, and AI-ready data assets Lead the strategic...  ...of scalable data pipelines and architectures supporting commerce analytics, machine learning, and AI workloads Partner with... 

    Jobtailor

    Santa Monica, CA
    5 days ago
  • $121.5k - $200k

     ...we LEAD: We are seeking an experienced and driven Senior Data Engineering Manager - Enterprise Data Products within the Global Data &...  ...delivery of scalable data pipelines and architectures that support commerce analytics, machine learning, and AI workloads. Partner... 
    Summer work
    Immediate start
    Flexible hours

    Universal Music Group

    Santa Monica, CA
    4 days ago
  • FIGS is seeking a Senior Analytics Engineer to join our Data Engineering team. You will translate business needs into reliable data assets and own curated datasets, semantic models, and analytics frameworks used across Product, Marketing, Operations, Finance, and Ecommerce... 

    FIGS

    Santa Monica, CA
    5 days ago
  • Activision Publishing Inc. is seeking an Analytics Engineer to support mobile marketing analytics operations and data systems powering marketing measurement, attribution, and reporting from Santa Monica. You will balance production analytics operations with analytics engineering... 

    Activision Publishing Inc.

    Santa Monica, CA
    5 days ago
  • $73.8k - $218.8k

     ...platform, and we're scaling fast. This is a founding team, these first hires will shape the...  ...conversation. You might come from engineering, consulting, product, or pre-sales — what...  ...more information on how we process your data during the Recruiting and Hiring process... 
    Work experience placement
    Live in
    Work at office
    Local area

    Accenture

    Culver City, CA
    4 days ago
  • $95k - $105k

    A non-profit organization in Los Angeles seeks an Analytics Engineer responsible for data consolidation and maintenance across its data infrastructure. The role involves collaborating with data analysts to design data models, automate data pipelines, and support reporting... 

    Brilliant Corners

    Los Angeles, CA
    5 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Founding Engineer, Commerce Data Supply Chain. Be the first to apply!