Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Model Bringup Engineer / ML Compiler Engineer

General Compute Inc.

Take a new model and get it running — correctly — on our ASIC in record time. When a frontier model drops, the only question that matters is how fast we can land it on our silicon and start serving it. You own that loop: from reference weights, through the compiler, to first correct tokens. The low-level runtime is co-owned with our hardware partner today; your job is everything it takes to get a brand-new architecture compiled, verified, and fast on top of it. The bet of this role is that bringup should be an agentic loop, not a hand-port. You'll build the harness of agents that compiles, runs, diffs against reference, and localizes failures — so the marginal model comes up faster than the last one did. Correctness first, optimization second: get it right, prove it's right, then make it cheap. This is a senior IC role on a small team. You'll own the bringup pipeline, not tickets. Responsibilities Own model bringup end-to-end. Take a new architecture — a frontier LLM, an MoE, a multimodal model — from reference weights to first correct tokens running on our ASIC, in days, not quarters. Build the agentic bringup loop. The differentiator isn't hand-porting one model — it's the harness of agents that compiles, runs, diffs against reference, localizes the failing op, and iterates without you in the inner loop. Each model you land should make the loop better at landing the next one. Live in the compiler. Graph capture, IR lowering, op coverage, kernel selection — when a model won't compile or produces wrong numbers, the fix is yours, whether it's a missing lowering, a fused-kernel bug, or a numerics mismatch. Own correctness before speed. Build the verification harness — layer-by-layer activation diffs, logit parity, end-to-end evals — that proves a freshly brought-up model matches reference before anyone trusts a token of it. Then optimize. Once it's correct, make it fast: operator fusion, quantization, memory layout, batching and KV-cache behavior on our hardware. Bringup gets it running; this is where it earns its cost-per-token. Work shoulder-to-shoulder with our hardware partner's compiler and runtime team. You're the person who turns 'the chip can technically run this' into 'this model is live and correct in production.' What we're looking for 5+ years in systems or ML systems, with real depth in at least one of: ML compilers, model porting/bringup, or high-performance kernels. You've taken a model architecture you didn't design and made it run — and run correctly — on a target it wasn't written for. Numerics debugging doesn't scare you. Strong on the internals of modern LLM inference: transformers, attention, KV cache, MoE routing, quantization, batching. You can read a new model's reference implementation and know what will be hard to lower. Comfortable inside a compiler stack — MLIR/LLVM, XLA, or a vendor graph compiler — at the level of IR, lowering, and op coverage, not just calling into one. Fluent with agentic tooling. You'd rather build the agent that runs the tedious bringup loop than run it by hand — and you have the taste to know where the loop still needs a human. Self-directed. We don't assign tickets — you'll see the next model coming and have it half brought-up before anyone asks. Nice to have Have worked on a non-NVIDIA accelerator — TPU, Trainium/Inferentia, Tenstorrent, Groq, Cerebras, or similar — at the compiler or model-bringup layer. Kernel-level experience in CUDA, Triton, or a vendor kernel language. You know why a fused attention kernel beats three unfused ops. Have built eval and numerics-verification harnesses (logit parity, activation diffing) for models in production. Contributed to a graph compiler or serving runtime — XLA, TVM, MLIR, vLLM, TGI, TensorRT-LLM, or SGLang. Have built agent loops or LLM-driven tooling that did real engineering work, not demos. #J-18808-Ljbffr General Compute Inc.

Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the Model Bringup Engineer / ML Compiler Engineer in New York, NY vacancy
  • $175k - $280k

     ...New York is seeking an expert in optimizing machine learning models to turbocharge their serving layer, integrating LLM, speech, and...  ...significant experience in systems programming and performance engineering, aiming to improve high-throughput, low-latency serving. Join... 
    Suggested

    SESAME

    New York, NY
    2 days ago
  • $175k - $215k

     ...reinforcement learning, learning from demonstration, generative modeling, Bayesian inference, hierarchical learning, and robust evaluation...  ...in an industrial or research setting developing recipes for ML models We prefer: - Track record of publications in top-tier... 
    Suggested
    Full time
    Remote work

    Waymo

    New York, NY
    15 hours ago
  • $228.7k - $343.1k

     ...financial crime at enormous scale, and one bad model can mean millions in credit losses,...  ...agents that run in parallel. Reason about ML systems end to end — how features,...  ...aggregate accuracy. Solid software and data engineering: production-quality Python, SQL on large... 
    Suggested
    Remote job
    Full time
    Local area
    Shift work

    Block

    New York, NY
    15 hours ago
  •  ...as a protected veteran.Job Description:ML Engineer3M Health Care is now SolventumAt Solventum...  ...this role, you won’t just be building models; you will be ensuring those models work...  ...and familiarity with SQL. Knowledge of a compiled language (like Go or Java) is a plus.Cloud... 
    Suggested
    H1b

    Solventum-

    New York, NY
    3 days ago
  •  ...through agent architecture changes, prompt engineering, fine-tuning, or rule-based post-...  ...weeks. Requirements ~2+ years building ML/AI systems in production ~ Built and deployed...  ..., infrastructure glue, not just model training scripts ~ Practical LLM experience... 
    Suggested
    Full time

    Triomics

    New York, NY
    15 hours ago
  • $100k - $250k

     ...science. This innovative field blends AI, engineering, and materials science, revolutionizing...  ...sustainability. The opportunity As an ML Engineer at Radical AI, you will be...  ...lifecycles, but also data ingestion, enrichment, model exporting and serving, and much more.... 
    Full time

    Radical Ai

    New York, NY
    15 hours ago
  •  ...the gap between bleeding-edge generative models and production-grade execution in a fast...  ...(Must-Have)6+ years of software engineering and machine learning experience (or 4+ years...  ...premier technology firms.Proficiency in a compiled programming language (such as C++, Go, or... 

    Objective Paradigm

    New York, NY
    3 days ago
  •  ...recently ran Platform at Hyperscience AI, building a workflow orchestration layer for computer vision models. Job Description We are seeking a Staff ML Engineer with a passion for building application-layer AI for high consequence, complex user interactions.... 
    Full time
    Work at office
    Work from home
    Flexible hours
    3 days per week

    Spara

    New York, NY
    15 hours ago
  • $251k - $310k

     ...learning from demonstration, generative modeling, Bayesian inference, hierarchical learning...  ...and Radar.. Partner effectively with engineering and research teams across Waymo to deploy...  ...profiling (e.g., xprof), and debugging of ML models. You have: PhD or Masters... 
    Full time
    Temporary work
    Remote work

    Waymo

    New York, NY
    15 hours ago
  • $77k - $202k

     ...of impactful solutions in a consulting settingWhat You Must Have- Bachelor's Degree- At least 4 years of experience in software engineering or data engineeringWhat Sets You Apart- Master's Degree in Computer Science, Data Engineering, Software Engineering preferred- Experience... 
    Full time
    H1b

    PwC

    New York, NY
    15 hours ago
  • $184.9k - $250.2k

    We are seeking a Machine Learning Engineer to work directly alongside our research scientists to train, evaluate, and deploy the models that make our robots move, perceive, and act in the real world. This is a hands-on ML role: you will train policies, debug convergence... 
    Internship
    Flexible hours

    Amazon

    New York, NY
    1 day ago
  • $175k - $250k

     ...and GraphQL. About the role Swayable is seeking a Senior Engineer blending Python software development expertise with scientific...  ...systems. You keep up with the constantly evolving toolset for ML and AI Ops. You are knowledgeable about software architecture... 
    Full time

    Swayable

    New York, NY
    15 hours ago
  •  ...Senior ML Eng Location: NYC (onsite only – not remote) Alliance is the leading accelerator...  ...15B+. We're hiring a  Machine Learning Engineer to join our in-house engineering team. You...  ...problem into a dataset, an experiment, a model, and a production system without relying... 
    Full time
    Relocation

    Alliance

    New York, NY
    15 days ago
  • $300k - $375k

     ...and self-directed Senior Machine Learning Engineer to join their in-house engineering team...  .... Key Responsibilities Own Applied ML End-to-End: Translate ambiguous, high-...  ...business problems into datasets, experiments, models, and production systems independently.... 
    Full time
    Work at office
    Relocation

    MLabs

    New York, NY
    12 days ago
  • $101.49k - $147k

     ...We have an exciting opportunity to join our team as a Sr. Engineer I - AI. As a Senior ML Engineer, you will support the implementation of diverse...  ...scientists to build and / or fine-tune high performing AI models and implement modern techniques from relevant published... 
    Work at office

    NYU Langone Health

    New York, NY
    3 days ago
  •  ...and connected: hardened batch pipelines, model governance and lineage, and the external...  ...into completing jobs took real data-skew engineering (pre-aggregation, windowing, connection-pool...  ...~5+ years of experience in data / ML pipeline engineering ~ Hands-on MLOps:... 
    Permanent employment
    Full time
    Contract work
    Work at office

    TWG Global AI

    New York, NY
    12 days ago
  • $160k - $300k

     ...ML Engineer Location: New York City (Flatiron) Company Stage: Series B — AI Infrastructure / Financial Operations Technology Office...  ...financial workflows. Their platform leverages large language models and advanced AI infrastructure to streamline bookkeeping, audit... 
    Work at office

    Recruiting from Scratch

    New York, NY
    1 day ago
  • $160k - $210k

     ...Senior Machine Learning Engineer / ML Scientist Remote, United States $160,000 to $210,000 base salary plus bonus and equity...  ...operate with a high level of ownership, designing and deploying models that solve complex problems in a production environment. The role... 
    Remote work
    Flexible hours

    Harnham

    New York, NY
    15 hours ago
  • $200k - $240k

     ...ML Engineer New York City or Remote. $200,000 to $240,000 base plus equity and benefits. This is a rare opportunity to join a very...  ...behavioural issues into structured training data and deploy effective model improvements Create scalable evaluation and testing systems... 
    Immediate start
    Remote work

    Harnham

    New York, NY
    4 days ago
  • $170k - $220k

     ...Founding ML Engineer Title of Role: Founding ML Engineer Location: New York, hybrid Company Stage of Funding: Venture-Backed —...  ...What You Will Do Design and implement machine learning models to analyze and interpret complex data sets. Develop backend... 
    Work at office

    Recruiting from Scratch

    New York, NY
    1 day ago
  •  ...Machine Learning Engineer, Applied AI As a machine learning engineer you'll join our product innovations team and work across the full applied ML stack - deploying models, building the evaluation systems that tell us whether they actually work, and making the data... 
    Work at office
    Work from home
    Home office

    CreatorIQ

    New York, NY
    2 days ago
  • $190k - $225k

     ...years of experience as a Machine Learning Engineer or Data Scientist, ideally in highly...  ...We want strong hands-on experience with ML and data science tooling such as sklearn,...  ...frameworks, along with solid data analysis and modeling skills. You should be confident... 
    Full time
    Visa sponsorship

    EvolutionIQ

    New York, NY
    2 days ago
  •  ...com We're looking for candidates with experience in: # Model Development & Deployment : Design, build, and deploy scalable...  ...Collaboration : Work closely with data scientists, software engineers, and founders to integrate machine learning solutions that solve... 
    Work at office
    Relocation

    WindMill

    New York, NY
    4 days ago
  •  ...company. We build cutting-edge foundation AI models and end-to-end products that are designed...  .... Cohere is a team of researchers, engineers, designers, and more, who are all passionate...  ...enjoy working across the full stack of ML systems, this role gives you the opportunity... 
    Full time
    Work at office
    Local area
    Remote work
    Home office

    Cohere

    New York, NY
    2 days ago
  • AI/ML Ops EngineerLocation: Remote / Hybrid (Client-Facing Consulting Engagement)Employment...  ...RoleWe are seeking an experienced AI/ML Engineer to design, deploy, and operate production...  ...transforming validated machine learning models into scalable, governed, and monitored... 
    Full time
    Contract work
    Local area
    Remote work
    Flexible hours

    Slalom

    New York, NY
    2 days ago
  • $157.95k - $259.48k

     ...role or a pipeline maintenance job. It is the highest-leverage engineering position in Phase 1 of a platform that will define what performance...  ...~ Experience building probabilistic evaluation frameworks or model calibration infrastructure, you understand the difference... 

    Catapult Sports

    New York, NY
    17 days ago
  • $200k - $225k

     ...power of neural networks and ML algorithms onto our custom hardware...  ..., and deploy machine learning engines on custom hardware, achieving...  ...engineers to translate ML models into efficient implementations...  ...Experience with ML-relevant compiler intermediate representations and... 
    Permanent employment
    Full time

    Imc

    New York, NY
    15 hours ago
  • $98k - $140k

     ...quality bar for Notion AI products. You’ll work with product and engineering teams to build systems to define what “good” looks like,...  ...shape our quality strategy. As part of that you'll shape Notion's model strategy and work directly with frontier AI labs (OpenAI, Anthropic... 
    Live in
    Local area

    Notion Labs

    New York, NY
    2 days ago
  • General Information Job Title ML Staff Engineer - LLM & Production Systems Job ID 107242 Work Areas Technology & Engineering...  ...of handling evolving datasets while continuously improving model performanceDeploy & Operate Production ML PlatformsLead production... 
    Permanent employment
    Full time
    Work at office
    Local area
    1 day per week

    Bain & Company

    New York, NY
    3 days ago
  • $50 per hour

    United States Digital Space LLC is seeking an experienced ML Engineer to design and develop AI-powered agents for the CRO/Pharma industry...  ...in AI/ML development, particularly with Large Language Models and strong programming skills in Python and SAS. This remote position... 
    Remote job
    Hourly pay

    United States Digital Space LLC

    New York, NY
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Model Bringup Engineer / ML Compiler Engineer. Be the first to apply!