Model Bringup Engineer / ML Compiler Engineer
General Compute Inc.
Take a new model and get it running — correctly — on our ASIC in record time. When a frontier model drops, the only question that matters is how fast we can land it on our silicon and start serving it. You own that loop: from reference weights, through the compiler, to first correct tokens. The low-level runtime is co-owned with our hardware partner today; your job is everything it takes to get a brand-new architecture compiled, verified, and fast on top of it. The bet of this role is that bringup should be an agentic loop, not a hand-port. You'll build the harness of agents that compiles, runs, diffs against reference, and localizes failures — so the marginal model comes up faster than the last one did. Correctness first, optimization second: get it right, prove it's right, then make it cheap. This is a senior IC role on a small team. You'll own the bringup pipeline, not tickets. Responsibilities Own model bringup end-to-end. Take a new architecture — a frontier LLM, an MoE, a multimodal model — from reference weights to first correct tokens running on our ASIC, in days, not quarters. Build the agentic bringup loop. The differentiator isn't hand-porting one model — it's the harness of agents that compiles, runs, diffs against reference, localizes the failing op, and iterates without you in the inner loop. Each model you land should make the loop better at landing the next one. Live in the compiler. Graph capture, IR lowering, op coverage, kernel selection — when a model won't compile or produces wrong numbers, the fix is yours, whether it's a missing lowering, a fused-kernel bug, or a numerics mismatch. Own correctness before speed. Build the verification harness — layer-by-layer activation diffs, logit parity, end-to-end evals — that proves a freshly brought-up model matches reference before anyone trusts a token of it. Then optimize. Once it's correct, make it fast: operator fusion, quantization, memory layout, batching and KV-cache behavior on our hardware. Bringup gets it running; this is where it earns its cost-per-token. Work shoulder-to-shoulder with our hardware partner's compiler and runtime team. You're the person who turns 'the chip can technically run this' into 'this model is live and correct in production.' What we're looking for 5+ years in systems or ML systems, with real depth in at least one of: ML compilers, model porting/bringup, or high-performance kernels. You've taken a model architecture you didn't design and made it run — and run correctly — on a target it wasn't written for. Numerics debugging doesn't scare you. Strong on the internals of modern LLM inference: transformers, attention, KV cache, MoE routing, quantization, batching. You can read a new model's reference implementation and know what will be hard to lower. Comfortable inside a compiler stack — MLIR/LLVM, XLA, or a vendor graph compiler — at the level of IR, lowering, and op coverage, not just calling into one. Fluent with agentic tooling. You'd rather build the agent that runs the tedious bringup loop than run it by hand — and you have the taste to know where the loop still needs a human. Self-directed. We don't assign tickets — you'll see the next model coming and have it half brought-up before anyone asks. Nice to have Have worked on a non-NVIDIA accelerator — TPU, Trainium/Inferentia, Tenstorrent, Groq, Cerebras, or similar — at the compiler or model-bringup layer. Kernel-level experience in CUDA, Triton, or a vendor kernel language. You know why a fused attention kernel beats three unfused ops. Have built eval and numerics-verification harnesses (logit parity, activation diffing) for models in production. Contributed to a graph compiler or serving runtime — XLA, TVM, MLIR, vLLM, TGI, TensorRT-LLM, or SGLang. Have built agent loops or LLM-driven tooling that did real engineering work, not demos. #J-18808-Ljbffr General Compute Inc.
$175k - $280k
...New York is seeking an expert in optimizing machine learning models to turbocharge their serving layer, integrating LLM, speech, and... ...significant experience in systems programming and performance engineering, aiming to improve high-throughput, low-latency serving. Join...Suggested$175k - $215k
...reinforcement learning, learning from demonstration, generative modeling, Bayesian inference, hierarchical learning, and robust evaluation... ...in an industrial or research setting developing recipes for ML models We prefer: - Track record of publications in top-tier...SuggestedFull timeRemote work$228.7k - $343.1k
...financial crime at enormous scale, and one bad model can mean millions in credit losses,... ...agents that run in parallel. Reason about ML systems end to end — how features,... ...aggregate accuracy. Solid software and data engineering: production-quality Python, SQL on large...SuggestedRemote jobFull timeLocal areaShift work- ...as a protected veteran.Job Description:ML Engineer3M Health Care is now SolventumAt Solventum... ...this role, you won’t just be building models; you will be ensuring those models work... ...and familiarity with SQL. Knowledge of a compiled language (like Go or Java) is a plus.Cloud...SuggestedH1b
- ...through agent architecture changes, prompt engineering, fine-tuning, or rule-based post-... ...weeks. Requirements ~2+ years building ML/AI systems in production ~ Built and deployed... ..., infrastructure glue, not just model training scripts ~ Practical LLM experience...SuggestedFull time
$100k - $250k
...science. This innovative field blends AI, engineering, and materials science, revolutionizing... ...sustainability. The opportunity As an ML Engineer at Radical AI, you will be... ...lifecycles, but also data ingestion, enrichment, model exporting and serving, and much more....Full time- ...the gap between bleeding-edge generative models and production-grade execution in a fast... ...(Must-Have)6+ years of software engineering and machine learning experience (or 4+ years... ...premier technology firms.Proficiency in a compiled programming language (such as C++, Go, or...
- ...recently ran Platform at Hyperscience AI, building a workflow orchestration layer for computer vision models. Job Description We are seeking a Staff ML Engineer with a passion for building application-layer AI for high consequence, complex user interactions....Full timeWork at officeWork from homeFlexible hours3 days per week
$251k - $310k
...learning from demonstration, generative modeling, Bayesian inference, hierarchical learning... ...and Radar.. Partner effectively with engineering and research teams across Waymo to deploy... ...profiling (e.g., xprof), and debugging of ML models. You have: PhD or Masters...Full timeTemporary workRemote work$77k - $202k
...of impactful solutions in a consulting settingWhat You Must Have- Bachelor's Degree- At least 4 years of experience in software engineering or data engineeringWhat Sets You Apart- Master's Degree in Computer Science, Data Engineering, Software Engineering preferred- Experience...Full timeH1b$184.9k - $250.2k
We are seeking a Machine Learning Engineer to work directly alongside our research scientists to train, evaluate, and deploy the models that make our robots move, perceive, and act in the real world. This is a hands-on ML role: you will train policies, debug convergence...InternshipFlexible hours$175k - $250k
...and GraphQL. About the role Swayable is seeking a Senior Engineer blending Python software development expertise with scientific... ...systems. You keep up with the constantly evolving toolset for ML and AI Ops. You are knowledgeable about software architecture...Full time- ...Senior ML Eng Location: NYC (onsite only – not remote) Alliance is the leading accelerator... ...15B+. We're hiring a Machine Learning Engineer to join our in-house engineering team. You... ...problem into a dataset, an experiment, a model, and a production system without relying...Full timeRelocation
$300k - $375k
...and self-directed Senior Machine Learning Engineer to join their in-house engineering team... .... Key Responsibilities Own Applied ML End-to-End: Translate ambiguous, high-... ...business problems into datasets, experiments, models, and production systems independently....Full timeWork at officeRelocation$101.49k - $147k
...We have an exciting opportunity to join our team as a Sr. Engineer I - AI. As a Senior ML Engineer, you will support the implementation of diverse... ...scientists to build and / or fine-tune high performing AI models and implement modern techniques from relevant published...Work at office- ...and connected: hardened batch pipelines, model governance and lineage, and the external... ...into completing jobs took real data-skew engineering (pre-aggregation, windowing, connection-pool... ...~5+ years of experience in data / ML pipeline engineering ~ Hands-on MLOps:...Permanent employmentFull timeContract workWork at office
$160k - $300k
...ML Engineer Location: New York City (Flatiron) Company Stage: Series B — AI Infrastructure / Financial Operations Technology Office... ...financial workflows. Their platform leverages large language models and advanced AI infrastructure to streamline bookkeeping, audit...Work at office$160k - $210k
...Senior Machine Learning Engineer / ML Scientist Remote, United States $160,000 to $210,000 base salary plus bonus and equity... ...operate with a high level of ownership, designing and deploying models that solve complex problems in a production environment. The role...Remote workFlexible hours$200k - $240k
...ML Engineer New York City or Remote. $200,000 to $240,000 base plus equity and benefits. This is a rare opportunity to join a very... ...behavioural issues into structured training data and deploy effective model improvements Create scalable evaluation and testing systems...Immediate startRemote work$170k - $220k
...Founding ML Engineer Title of Role: Founding ML Engineer Location: New York, hybrid Company Stage of Funding: Venture-Backed —... ...What You Will Do Design and implement machine learning models to analyze and interpret complex data sets. Develop backend...Work at office- ...Machine Learning Engineer, Applied AI As a machine learning engineer you'll join our product innovations team and work across the full applied ML stack - deploying models, building the evaluation systems that tell us whether they actually work, and making the data...Work at officeWork from homeHome office
$190k - $225k
...years of experience as a Machine Learning Engineer or Data Scientist, ideally in highly... ...We want strong hands-on experience with ML and data science tooling such as sklearn,... ...frameworks, along with solid data analysis and modeling skills. You should be confident...Full timeVisa sponsorship- ...com We're looking for candidates with experience in: # Model Development & Deployment : Design, build, and deploy scalable... ...Collaboration : Work closely with data scientists, software engineers, and founders to integrate machine learning solutions that solve...Work at officeRelocation
- ...company. We build cutting-edge foundation AI models and end-to-end products that are designed... .... Cohere is a team of researchers, engineers, designers, and more, who are all passionate... ...enjoy working across the full stack of ML systems, this role gives you the opportunity...Full timeWork at officeLocal areaRemote workHome office
- AI/ML Ops EngineerLocation: Remote / Hybrid (Client-Facing Consulting Engagement)Employment... ...RoleWe are seeking an experienced AI/ML Engineer to design, deploy, and operate production... ...transforming validated machine learning models into scalable, governed, and monitored...Full timeContract workLocal areaRemote workFlexible hours
$157.95k - $259.48k
...role or a pipeline maintenance job. It is the highest-leverage engineering position in Phase 1 of a platform that will define what performance... ...~ Experience building probabilistic evaluation frameworks or model calibration infrastructure, you understand the difference...$200k - $225k
...power of neural networks and ML algorithms onto our custom hardware... ..., and deploy machine learning engines on custom hardware, achieving... ...engineers to translate ML models into efficient implementations... ...Experience with ML-relevant compiler intermediate representations and...Permanent employmentFull time$98k - $140k
...quality bar for Notion AI products. You’ll work with product and engineering teams to build systems to define what “good” looks like,... ...shape our quality strategy. As part of that you'll shape Notion's model strategy and work directly with frontier AI labs (OpenAI, Anthropic...Live inLocal area- General Information Job Title ML Staff Engineer - LLM & Production Systems Job ID 107242 Work Areas Technology & Engineering... ...of handling evolving datasets while continuously improving model performanceDeploy & Operate Production ML PlatformsLead production...Permanent employmentFull timeWork at officeLocal area1 day per week
$50 per hour
United States Digital Space LLC is seeking an experienced ML Engineer to design and develop AI-powered agents for the CRO/Pharma industry... ...in AI/ML development, particularly with Large Language Models and strong programming skills in Python and SAS. This remote position...Remote jobHourly pay
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Model Bringup Engineer / ML Compiler Engineer. Be the first to apply!
- graduate machine learning engineer New York, NY
- senior ml engineer New York, NY
- data scientist machine learning engineer New York, NY
- machine learning engineer New York, NY
- machine learning ai engineer New York, NY
- entry level machine learning engineer New York, NY
- machine learning software engineer New York, NY
- ai ml engineer New York, NY
- computer vision machine learning engineer New York, NY
- junior machine learning research engineer New York, NY



