Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Staff ML Data Engineer

$171.9k - $221k

Jobleads-US

ABOUT LVT

LVT is redefining how businesses operate in the physical world, moving beyond traditional security solutions to deliver AI-driven, actionable intelligence that makes sites smarter, safer, and more secure. Since pioneering our first mobile, solar-powered units, our commitment to scrappy, hands-on innovation has made us an established leader and one of the fastest-growing companies in intelligent site technology. We are building the next generation of solutions—from our physical units in the field to a powerful Agentic AI platform—that allows our customers to gain unprecedented visibility and control over safety, compliance, and operations. This is your chance to join a cutting-edge team that isn't just watching the world change, but actively building the technology that is changing it.

We’re a team that’s focused on growth and innovation, and we’re proud that our crew, products, and leadership are being recognized for it.

  • A Top-Tier Growth Company: Named one of Financial Times’ Fastest Growing Companies 2025 and #10 on the Inc. 5000 Rocky Mountain Regional list for 2025.
  • Innovative Leadership: Our CEO, Ryan Porter, was named an EY Entrepreneur of the Year 2025, and our CTO, Steve Lindsey, was inducted into the Silicon Slopes CTO Hall of Fame in 2024.
  • Product & Software Excellence: We were named one of The Software Report’s Top 100 Software Companies of 2023 and are a winner of the Security Today Govies Award for 2025.

ABOUT THIS ROLE

LVT's AI systems are only as good as the data behind them. As we move toward Physical AI, the binding constraint shifts from model architecture to the data flywheel.

We are seeking a Staff Data Engineer to own that flywheel end to end including logs, sensor telemetry, labels and annotations, evaluation and benchmark sets. Every AI team trains and evaluates from a single stack that transforms data from the raw source through standardized, versioned, governed datasets.

This is a senior individual‑contributor and technical‑leadership role; formal people management is not required. You will partner closely with AI/ML research, the ML platform / MLOps function. You own the data side of the contract that defines what a model consumes and emits and annotation, edge, and infrastructure teams. You should be equally comfortable discussing dataset schema design, storage and partitioning trade‑offs for multimodal data, versioning and migration strategy, and the governance controls that keep sensitive video and sensor data safe.

ROLE RESPONSIBILITIES

  • Data Flywheel Ownership: Own the end-to-end loop that converts raw edge telemetry and video into labeled training data, frozen evaluation sets and feeds model outputs back into the next round.
  • Layered Dataset Pipelines: Build and own the pipelines that register raw source data, standardize it into a single well-defined schema, and join and aggregate it into curated datasets so every team trains, validates, and benchmarks from one consistent store through one reader, rather than copying and reformatting data per use case.
  • Labels & Annotation Data Lifecycle: Own how labels and semantic annotations are appended to datasets without rewriting source data, then versioned, quality-checked, and served, partnering with annotation and data‑operations teams on label production and verification while you own the dataset, storage, and serving side.
  • Evaluation & Benchmark Sets: Own the frozen, versioned validation and benchmark datasets that make model comparisons valid over time stable enough that an accuracy delta reflects the model, not a shifting dataset including the review and scrubbing discipline required before any set is shared externally.
  • Dataset Versioning: Own schema and content versioning so producers can evolve datasets without breaking consumers opt‑in versions, append-without-rewrite for new fields, and the reader/writer indirection that lets data migrate underneath clients on a controlled rollout instead of forced lockstep migrations.
  • Framework Integration & Self‑Serve Access: Own the read/write libraries and integrations researchers depend on PyTorch/Lightning dataloaders, a simple record‑level CRUDL API, and Spark/analytics access and self‑service so AI teams stay focused on model development.
  • Governance Enforced: Make governance machine‑enforced in the flywheel rather than documented after the fact classification of clips, frames, labels, and embeddings; scrubbing and anonymization in load jobs; and lineage and provenance for every dataset version, annotation campaign, and training input.
  • Technical Mentorship: Set the data‑engineering standards for the flywheel schema conventions, dataset contracts, quality gates and mentor IC work toward them, growing the function as the team forms.

OUR IDEAL CANDIDATE

  • Data Engineering Depth: 8+ years building and operating large-scale data pipelines and data‑lake or lakehouse systems in production ingestion, ETL/ELT, partitioning and storage-format decisions, and the reader/writer libraries consumers rely on.
  • ML Data Specialty: Has built data pipelines for model training and evaluation, labeled data, and evaluation/benchmark sets with a working understanding of how data quality and versioning move model results.
  • Lakehouse Architecture: Strong experience with medallion-style layered data architectures and modern table/lake formats (e.g. Iceberg, Delta, Parquet, or comparable), including schema evolution and dataset versioning.
  • Multimodal Data at Scale: Experience with large multimodal data video, image, sensor/telemetry and the storage and access patterns that make it queryable at scale (denesting, repartitioning, binary-inline vs. reference storage).
  • Framework Integration: Hands‑on with the data side of ML frameworks PyTorch/Lightning dataloaders and Spark and strong Python knowledge.
  • Governance & Provenance: Practical experience enforcing data governance in pipelines classification, access control, lineage and provenance, retention, particularly for privacy sensitive data.
  • Technical Leadership: A track record of setting data‑engineering direction and leveling up engineers (technical leadership; formal management not required).
  • Education: Bachelor's or Master's in Computer Science, Engineering, or a related field, or equivalent practical experience.

PREFERRED QUALIFICATIONS

  • Streaming or near‑real‑time ingestion from edge/IoT sources into a data lake (e.g. Kafka, Lambda, EMR, or similar).
  • Append‑without‑rewrite and hash-indexed dataset techniques on open table formats, and dataset/feature‑versioning systems.
  • Generative‑AI data work: fine‑tuning and evaluation dataset curation for LLMs/VLMs.
  • Exposing datasets to AI agents through MCP‑style query interfaces, with semantic schema and plain‑language documentation for retrieval.
  • Computer‑vision / video annotation tooling and workflows (e.g. Encord, Labelbox, or similar).

COMPENSATION

The beginning annual salary range for this role is $171,900 - $221,000 USD and is determined by location, job‑related experience, and education/training. Your total earning potential is amplified by a bonus structure tied to meeting goals, and you will become an owner from day one through our employee equity program.

BENEFITS

We believe you do your best work when your whole life is supported. We invest in our crew’s health, families, and financial futures with a benefits package designed to support you inside and outside the office. Full‑time benefits include, but not limited to: Comprehensive health, dental and vision coverage, retirement benefits (401k match up to 4%), and flexible PTO.

LVT IS PROUD TO BE AN EQUAL OPPORTUNITY EMPLOYER. All applicants will be considered for employment without attention to race, color, religion, sex, sexual orientation, gender identity, national origin, veteran or disability status. All candidates must pass a drug screening and background check upon employment. Some roles may also require passing a federal background check and fingerprinting. Must be authorized to work in the U.S. If reasonable accommodation is needed to participate in the job application or interview process, and/or to perform essential job functions, please reach out to your recruiter.

#J-18808-Ljbffr Jobleads-US
Vacancy posted 4 days ago
Similar jobs that could be interesting for youBased on the Staff ML Data Engineer in Seattle, WA vacancy
  • Our team plays a crucial role in the data ecosystem of the TikTok Recommendation System, focusing on creating offline and real-time...  ...Qualifications:- A bachelor's degree or above in computer science, software engineering, or related fields, with experience in building scalable... 
    Suggested
    Flexible hours

    TikTok

    Seattle, WA
    4 days ago
  •  ...LiveView Technologies in Seattle is seeking a Staff Data Engineer to own the data flywheel from raw edge telemetry and video to labeled training...  ...and evaluation sets. This senior IC role partners with AI/ML research, the ML platform, and data operations to define dataset... 
    Suggested

    Jobleads-US

    Seattle, WA
    4 days ago
  • $143.7k - $194.4k

    We are looking for an **AI/ML Engineer** to build, deploy, and operate the ML/AI systems that power the agentic decision intelligence workflow...  ...Generation) systems that ground LLM outputs in operational data, historical playbooks, and domain knowledge- Build guardrails,... 
    Suggested
    Internship
    Flexible hours

    Amazon

    Seattle, WA
    3 days ago
  •  ...Irvine is the place. We go beyond typical data-driven approaches or pure transformer-only architectures, combining rigorous engineering with learning systems proven in globally deployed...  ...in the field. About This Role Modern ML systems improve through their data flywheel... 
    Suggested
    Local area

    FieldAI

    Seattle, WA
    24 days ago
  • $95 - $100 per hour

     ...Pay rate range - $95/hr. to $100/hr. 100% Onsite Must Have 1. Python/SQL 2. Problem solving/solution designing 3. Data engineering best practices Nice to Have 1. Requirements capture 2. Data modeling/architecture 3. Communication Years of experience... 
    Suggested
    Flexible hours

    Intelliswift - An LTTS Company

    Seattle, WA
    2 days ago
  •  ...We are seeking a Senior Data Engineer with strong expertise in Databricks, Python, ETL, SQL, and Data/Financial Modeling to design, build...  .../AI data pipelines and collaboration with Data Scientists or ML Engineers. Experience optimizing large-scale data pipelines... 
    Full time

    BrickRed Systems

    Bellevue, WA
    14 hours ago
  •  ...Role :: AWS Data Engineer Location :: Seattle, WA - Onsite Fulltime_Permanent_role All Visa workable Min exp 8+ year Job Description Must-have: Hands-on data engineer with strong Python skills and experience building large scale Apache Spark pipelines... 
    Permanent employment
    Full time

    K&K Global Talent Solutions INC.

    Seattle, WA
    8 hours ago
  •  ...We are seeking a Data Engineer with strong expertise in Microsoft Fabric, Power BI, Semantic Modeling, SQL, dimensional data modeling, and cloud data platforms . The role will focus on designing, building, and governing an enterprise semantic layer using Microsoft... 
    Full time

    BrickRed Systems

    Bellevue, WA
    4 days ago
  •  ...Title: Data Engineer Location: Seattle, WA/Hybrid Key Responsibilities: Data engineer with python, Apache spark, aws, data lake stack, Athena, Emr, datashift, practical experience in designing, testing Aws lake models. Required Skills & Qualifications... 

    K&K Global Talent Solutions INC.

    Seattle, WA
    8 hours ago
  • $150k - $160k

     ...About Ascendion Ascendion is an AI-native software engineering disruptor helping businesses innovate faster, smarter, and with greater impact...  ..., the UK, Europe, and APAC to solve complex challenges in data, experience design, software product engineering, and workforce... 
    Hourly pay
    Full time
    Temporary work

    Ascendion

    Seattle, WA
    2 days ago
  •  ...Must-Have Skills & Experience: Strong hands-on experience as a Data Engineer . Strong proficiency in Python . Experience building and maintaining large-scale Apache Spark pipelines on AWS. Strong knowledge of the modern AWS Data Lake ecosystem , including... 

    Tenarai

    Seattle, WA
    2 days ago
  •  ...We are seeking an experienced Sr. Data Engineer with strong expertise in Databricks, Python, Salesforce Automation, ETL, SQL, and Azure...  ...understand requirements and deliver data solutions. Support ML Engineers and Data Scientists with relevant datasets and data engineering... 
    Full time

    BrickRed Systems

    Bellevue, WA
    2 days ago
  • $132.1k - $178.8k

     ...direct revenue to the business.We are looking for an excellent Data Engineer who is passionate about data and the insights that large amounts...  ...will work in lock-step with BI Engineers, Data scientists, ML scientists, Business analysts, Product Managers and other stakeholders... 
    Worldwide
    Flexible hours

    AmazonWebServices

    Bellevue, WA
    3 days ago
  • $132.1k - $178.8k

    Are you passionate about standardizing data platforms and automating data engineering to drive analytics and reporting? Do you excel in dynamic, fast-paced environments and find joy in converting data into actionable insights? Are you adept at implementing data governance... 
    Worldwide
    Flexible hours

    Amazon

    Bellevue, WA
    3 days ago
  • $145.3k - $196.6k

     ...dashboards, no pattern to follow. If you're energized by shaping data infrastructure from zero to one inside a fast-moving org, keep...  ...products- Self-service reporting that scales spanning Engineering, Science, PM-T, and Design across multiple AI native advertising... 
    Flexible hours

    Amazon

    Seattle, WA
    3 days ago
  • $132.1k - $178.8k

     ...Amazon's Customer Service (CS) organization, responsible for the data infrastructure that powers measurement, analytics, and...  ...interaction — chat, voice, bots, and digital self-service. Our Data Engineers own the foundational data layer that BIEs, scientists, and product... 
    Flexible hours
    Day shift

    Amazon

    Seattle, WA
    2 days ago
  • $132.1k - $178.8k

    Are you passionate about building data platforms that empower thousands of users to discover, trust, and act on data? Do you thrive in...  ...opportunity for you!We are seeking a customer-centric Data Engineer to help build and scale our unified data management and discovery... 
    Worldwide
    Flexible hours

    Amazon

    Bellevue, WA
    3 days ago
  • $132.1k - $178.8k

     ...Product organization. We bring together Business Intelligence Engineers, Data Engineers, and Data Scientists to deliver comprehensive data tools...  ..., reporting, experimentation, full-funnel attribution, and ML use cases.* Monitor pipeline health, resolve job and cluster... 
    Flexible hours

    Amazon

    Seattle, WA
    4 days ago
  • $132.1k - $178.8k

    UTR Planning Tech builds the data infrastructure that powers labor planning across 12 Amazon...  ...systems that determine how Amazon staffs its delivery network, serving hundreds of...  ...We are building reusable frameworks where engineers and business users define what they need... 
    Flexible hours
    Shift work

    Amazon

    Bellevue, WA
    3 days ago
  •  ...Bellevue, WAEmployment Type: C2CIndustry: Computer SoftwareClient: WiproContact: Meghana GorusuCompany: SRI Tech SolutionsRole: Senior Data Engineer Location: Bellevue, WA Job Description: We are seeking a skilled and motivated Senior Data Engineer/Data Engineer to join our... 

    SRI Tech

    Bellevue, WA
    2 days ago
  • $86.7k - $170.9k

    Position Summary Join Deloitte’s Core AI & Data practice and help organizations modernize data platforms, strengthen enterprise...  ...and artificial intelligence capabilities. As a Databricks Data Engineer, you will support the design, build, and optimization of cloud-based... 
    Local area
    Visa sponsorship

    Deloitte

    Seattle, WA
    3 days ago
  • As a data engineer in the Data Platform E-Commerce team, you will have the opportunity to build, optimize and grow one of the largest data platforms in the world. You'll have the opportunity to gain hands-on experience on all kinds of systems in the data platform ecosystem... 

    TikTok

    Seattle, WA
    4 days ago
  • $132.1k - $178.8k

    WorldWide Amazon Stores FinTech (WWASFT) team is looking for an Data Engineer who is data-driven, uncompromisingly detail oriented, smart, efficient, and driven to help our business succeed. You have passion for technology. You are keen to leverage existing skills while... 
    Worldwide
    Flexible hours

    Amazon

    Seattle, WA
    1 day ago
  • $152k - $205.6k

     ...global business teams.We are looking for a Data Engineer to help us extend our geospatial...  ...data pipelines that enable statistical and ML-based modeling with geospatial data- Produce...  ...cooperatively with other employees, supervisors, and staff; adhere to standards of excellence... 
    Local area
    Worldwide
    Flexible hours

    Amazon

    Seattle, WA
    2 days ago
  • Data Engineer3-5 years | SQL Server, SSIS, T-SQL, VisualCron, C#/.NETRole overview Own and enhance core enterprise data pipelines while...  ...event-driven workflows. This role combines SQL Server and SSIS engineering with advanced VisualCron orchestration, automation, monitoring,... 

    Fuel Talent

    Seattle, WA
    2 days ago
  • $177k - $239.4k

     ...we’re the people who keep the cloud running. We support all AWS data centers and all of the servers, storage, networking, power, and...  ....You’ll join a diverse team of software, hardware, and network engineers, supply chain specialists, security experts, operations... 
    Remote work
    Flexible hours

    Amazon

    Seattle, WA
    4 days ago
  • $116.5k

     ...careers and have the flexibility, benefits, and support to do their best work. Join us and build for travelers everywhere.Global Data Sharing Engineer The Global Data Sharing Team (GDST) supports Expedia by delivering evolving global legal and regulatory requirements for... 
    Full time

    Expedia

    Seattle, WA
    2 days ago
  • $137.4k - $161.7k

    Are you ready to make an impact?West Monroe is seeking a Senior Data Engineer to join our Data & AI practice. In this role, you will design, build, and optimize modern cloud data platforms that enable advanced analytics, AI, and business intelligence. Working with technologies... 
    Local area
    Immediate start
    Flexible hours

    West Monroe Partners

    Seattle, WA
    2 days ago
  • $157.4k - $236k

     ...professionally.The TeamAs part of the IT Enterprise Data Solutions (EDS) department, the Data Engineering team’s mission is to lead on data technology and utilization...  ...plans.Mentor data engineers and contingent staff through pairing, technical guidance, review, and knowledge... 
    Full time
    Temporary work
    H1b

    Bill & Melinda Gates Foundation

    Seattle, WA
    2 days ago
  • $116.2k - $229.1k

    Position Summary Join our AI & Engineering team in transforming technology platforms, driving innovation, and helping make a significant...  ...operate integrated/verticalized sector solutions in software, data, AI, network, and hybrid cloud infrastructure. These solutions... 
    Local area
    Visa sponsorship

    Deloitte

    Seattle, WA
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Staff ML Data Engineer. Be the first to apply!