Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

ML Performance Engineer

$100k - $150k
Full-time

Bright Vision Technologies

Bothell, WA
ML Performance Engineer - Remote 

 
Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States. 
This is a fantastic opportunity to join an established and well-respected organization offering tremendous career growth potential. 

 
Job Title: ML Performance Engineer
Location: 100% Remote (U.S.) 
Position Type:  Full-time, Direct W2 
Salary Range: $100,000–$150,000 Annually 
Experience Required: 6+ years 

 
Sponsorship:  U.S. Citizens, Green Card Holders, EAD Holders, and H-1B transfer candidates are encouraged to apply. We are unable to sponsor new H-1B visa petitions for this position. 

 
Job Summary 
We are seeking an AI Performance Optimization Engineer to focus on extracting maximum throughput, minimizing latency, and reducing cost across training and inference workloads for large neural network systems. The role spans the full stack from low-level kernel optimization to distributed system tuning, requiring deep understanding of GPU architecture, model parallelism, memory management, and compiler-level optimization. The ideal candidate has demonstrated impact on production AI workloads, with strong instrumentation and measurement discipline that enables rigorous, data-driven optimization decisions. In this role you will work closely with cross-functional partners — product, design, engineering, operations, and business stakeholders — to translate ambiguous requirements into well-engineered solutions, and will be expected to raise the bar through code review, design review, and mentorship of more junior engineers. The successful candidate brings strong engineering discipline, a clear communication style, and a track record of shipping meaningful work that holds up well in production. 

Key Responsibilities  
  • Profile and optimize end-to-end AI training and inference pipelines for throughput, latency, and cost. 
  • Identify and eliminate bottlenecks across data loading, model compute, communication, and memory. 
  • Implement and tune quantization, sparsity, and pruning strategies to reduce model footprint and accelerate inference. 
  • Optimize distributed training using tensor parallelism, pipeline parallelism, FSDP, and ZeRO-style sharding. 
  • Tune attention implementations using FlashAttention, paged attention, and related techniques. 
  • Implement KV cache optimization, continuous batching, and speculative decoding for LLM serving. 
  • Drive compiler-level optimizations using Triton, XLA, TorchInductor, or TVM, working with the broader ML framework community to land improvements that translate into measurable end-to-end performance gains. 
  • Optimize data pipelines, sharding strategies, and storage access patterns for high-throughput training. 
  • Build and maintain rigorous benchmark suites and regression frameworks across workloads. 
  • Collaborate with ML and platform engineering teams to embed best practices in standard pipelines. 
  • Drive cost-efficiency improvements through model architecture, hardware selection, and scheduling strategies. 
  • Evaluate new hardware and software offerings, and advise on adoption. 
  • Document performance tuning playbooks and share findings broadly across engineering teams. 
  • Stay current with AI systems research and translate advances into production improvements. 
Required Qualifications 
  • Bachelor’s or Master’s degree in Computer Science, Computer Engineering, or a related field. 
  • Six or more years of experience in performance engineering, ML systems, or HPC. 
  • Strong proficiency in Python and C++. 
  • Hands-on experience optimizing deep learning workloads on modern GPUs. 
  • Deep understanding of distributed training and inference techniques. 
  • Experience with profiling tools across CPU, GPU, and distributed systems. 
  • Familiarity with model compression techniques and their accuracy implications. 
  • Strong grasp of memory hierarchies, communication primitives, and parallelism strategies. 
  • Excellent measurement, debugging, and analytical reasoning skills. 
  • Strong communication and collaboration skills. 
Preferred Qualifications 
  • Experience optimizing LLM inference at production scale. 
  • Contributions to vLLM, TensorRT-LLM, DeepSpeed, or similar projects. 
  • Familiarity with custom kernel authoring in Triton or CUTLASS. 
  • Experience with FinOps for AI workloads. 
  • Publications or talks on AI systems performance. 
How to Apply 
Would you like to know more about this opportunity? For immediate consideration, please send your resume to View email address on brightvisiontechnologies.applytojob.com or contact us at Show phone number. Learn more about Bright Vision Technologies at  .
Bright Vision Technologies is an Equal Opportunity Employer.

 

Equal Employment Opportunity (EEO) Statement

Bright Vision Technologies (BV Teck) is committed to equal employment opportunity (EEO) for all employees and applicants without regard to race, color, religion, sex, sexual orientation, gender identity or expression, national origin, age, genetic information, disability, veteran status, or any other protected status as defined by applicable federal, state, or local laws. This commitment extends to all aspects of employment, including recruitment, hiring, training, compensation, promotion, transfer, leaves of absence, termination, layoffs, and recall.

BV Teck expressly prohibits any form of workplace harassment or discrimination. Any improper interference with employees' ability to perform their job duties may result in disciplinary action up to and including termination of employment.

Vacancy posted 6 days ago
Similar jobs that could be interesting for youBased on the ML Performance Engineer in Bothell, WA vacancy
  • $175k - $215k

     ...from a diverse set of sensors, enabling engineers like you to (1) develop methods for efficiently...  ...require a diverse skill set of ML and geometric algorithms, onboard sensing...  ...and metrics for measuring and improving performance of pre-trained and fine-tuned models... 
    Performance
    Full time
    Remote work

    Waymo

    Kirkland, WA
    1 day ago
  • $148.5k - $223.9k

     ...future of Salesforce.This role is for a Senior Machine Learning Engineer within the Trust Intelligence Platform team who will architect...  ..., promotion, benefits, training, assessment of job performance, discipline, termination, and everything in between. Recruiting... 
    Performance
    Full time

    Salesforce

    Bellevue, WA
    1 day ago
  • $182k - $242k

     ...CoreWeave combines superior infrastructure performance with deep technical expertise to...  ...qualifications, we’re looking for strong engineers with great taste. The most important qualification...  ...and hands on experience with modern ML frameworks such as PyTorch or JAX.... 
    Performance
    Permanent employment
    Temporary work
    Casual work
    Work at office
    Flexible hours

    CoreWeave

    Bellevue, WA
    4 days ago
  • $175k - $280k

    Sesame, located in Bellevue, Washington, is seeking a talented engineer to join our team focused on revolutionizing the way computers...  ...models. Candidates should have significant systems programming and performance engineering experience. You will enjoy benefits like unlimited... 
    Performance

    SESAME

    Bellevue, WA
    1 day ago
  • $175k - $280k

     ...variety of LLM, speech, and vision models. Partner with ML infrastructure and training engineers to build a fast, cost‑effective, accurate, and...  ...SGLang to take advantage of the latest techniques in high‑performance model serving. Work with the training team to identify... 
    Performance
    Contract work
    Flexible hours

    SESAME

    Bellevue, WA
    2 days ago
  • $148.5k - $225k

     ...is looking for a Senior Machine Learning Engineer to help launch various innovative ads-offerings...  ..., campaign optimization, ads channel performance, ads performance maximization and wide...  ...applications.Publish research papers in leading ML/AI/Advertising conferences solving... 
    Performance
    Local area
    Flexible hours

    Chewy

    Bellevue, WA
    1 day ago
  • $168.1k - $227.4k

     ...cheaper is as much a systems problem as a modeling problem - agent performance depends on the model, the harness, the evaluation...  ...designed together.We are looking for a Senior Machine Learning Engineer to build and own core systems in this agentic platform. You will... 
    Performance
    Internship
    Flexible hours
    Day shift

    Amazon

    Bellevue, WA
    2 days ago
  • $130k - $260k

     ...development of scalable, production-grade ML systems powering customer-facing...  ...and continuous improvement.Design high-performance batch and real-time ML architectures capable...  ...maintainable production systems.Establish engineering patterns and best practices for model deployment... 
    Performance
    Full time
    Temporary work
    Part time
    Work experience placement

    Walmart

    Bellevue, WA
    3 days ago
  • $232.3k - $406.6k

     ...organizational lead for this site, you will oversee diverse engineering teams, guiding them as they deliver results for their respective...  ...excels at leading through influence, inspiring high-performing and diverse technical teams at all levels of the organization... 
    Performance
    Work experience placement
    Work at office
    Local area
    Immediate start

    Visa

    Bellevue, WA
    5 days ago
  • $135.31k - $251.29k

     ...looking for a passionate Machine Learning engineer to support, build and scale the...  ...observability, scalability, cost-optimization, and performance tuning and other improvements to...  ...scale recommender systems, or large-scale ML ranking/retrieval/search systems and familiarity... 
    Performance
    Full time
    Temporary work
    Local area
    Worldwide

    Warner Bros. Discovery

    Bellevue, WA
    5 days ago
  • $127.4k - $191.2k

     ...monthly users on the world's leading game engine. Recommendation and ranking systems are...  ...Experience working with large-scale data and ML systems, whether through research or...  ...written exchanges in this language since the performance of the duties related to this position... 
    Performance
    Full time
    Internship
    Work at office
    Worldwide
    Shift work

    Unity Technologies

    Bellevue, WA
    4 days ago
  • Plutus is seeking a founding machine learning engineer to scale from v1 to a billion under management by year-end. You will design and deploy...  ...in ambiguity at an early-stage startup. You should have strong ML fundamentals, experience taking models to production, and a... 

    Plutus

    Kirkland, WA
    4 days ago
  • $200.4k - $260.5k

    Senior Machine Learning Engineer, Data InfrastructureUnity Vector builds an Data platform that...  ...workflows that power production ML systems.To support this growth, we need strong...  ...ensure the reliability, scalability, and performance of our data platform.What You’ll Do Develop... 
    Performance
    Full time
    Work at office
    Worldwide
    Relocation package

    Unity Technologies

    Bellevue, WA
    1 day ago
  • Electronic Arts seeks a Senior Machine Learning Engineer to design and operate production-grade data and ML infrastructure for fraud, anti-cheat, and account security across EA games. The role reports to the Senior Manager, EA Player Security Data Labs and follows a hybrid... 

    Electronic Arts

    Kirkland, WA
    4 days ago
  • $255k - $300k

     ...best work of their careers. We're a high-performing, fast-moving team with ethics at the...  ...functional partner to growth, product, and data engineering—translating complex financial data into...  ...product, data engineering, and fellow ML engineers to take ambitious ideas from... 
    Performance
    Work at office
    Flexible hours
    Shift work
    3 days per week

    Robinhood

    Bellevue, WA
    5 days ago
  • $100k - $150k

     ...MLOps Engineer - Remote    Bright Vision Technologies is a technology consulting and...  ...Engineer to design, build, and operate high-performance, highly reliable inference platforms for...  ..., throughput, cost, and quality in ML serving.  Key Responsibilities  Design... 
    Performance
    Full time
    H1b
    Local area
    Immediate start
    Remote work
    Visa sponsorship

    Bright Vision Technologies

    Bothell, WA
    6 days ago
  •  ...Description Team Red Dog is hiring a Robotics Machine Learning Engineer (Embodied AI) for our client, a Fortune 50 technology leader...  ...and integration issues. Develop new research features, perform evaluations, and create technical documentation supporting research... 
    Performance
    Contract work
    Local area
    Immediate start
    Flexible hours

    Team Red Dog

    Redmond, WA
    11 days ago
  •  ...with different functional teams to implement models and monitor outcomes.Develop processes and tools to monitor and analyze model performance and data accuracyRequirementsKnowledge of advanced statistical techniques and concepts (regression, properties of distributions,... 
    Performance

    InterSources

    Bothell, WA
    1 day ago
  • $117.3k - $195.5k

     ...management Gather, prepares, and maintains datasets required to perform analytics Oversee data systems holding in close...  ...Science Relevant certification &/or work experience in data engineering Experience (7+ years) in data sciences, notably in the healthcare... 
    Performance
    Work experience placement
    Remote work
    Relocation package
    Flexible hours

    Pfizer

    Bothell, WA
    4 days ago
  • $141k - $202k

    A leading tech company in Kirkland is seeking a software engineer to develop next-generation technologies that transform user interactions...  ...will have expertise in Python or C++, along with experience in ML infrastructure and Generative AI concepts. The role offers an annual... 

    Google

    Kirkland, WA
    5 days ago
  •  ...ultra-wealthy. We launched in January, AUM is already in the tens of millions, and we're now looking for a founding machine learning engineer to help take us from v1 to a billion under management by end of year. We have an enormous amount of customer and portfolio data... 
    Live in
    Weekend work

    Plutus

    Kirkland, WA
    4 days ago
  • Snap Inc. seeks a Machine Learning Engineer to join the Generative ML Platform team. You will develop innovative ML technologies and build AR experiences using generative models, with a focus on on-device inference and scalable systems across mobile, web, and wearables... 

    Snap

    Bellevue, WA
    2 days ago
  • $122.3k - $158.5k

     ...strengthens system integrity, supports fair play, and enables teams to build and operate securely at scale. The Senior Machine Learning Engineer will report to the Senior Manager, EA Player Security Data Labs. You will follow a hybrid work model with a mix of remote work and... 
    Full time
    Work at office
    Local area
    Remote work

    Electronic Arts

    Kirkland, WA
    4 days ago
  • Chewy is seeking a Senior Machine Learning Engineer for our Bellevue, WA team to advance sponsored ads offerings across onsite and offsite channels. You will shape model development, ranking, bidding dynamics, and campaign optimization, collaborating with product and engineering... 

    Chewy

    Bellevue, WA
    1 day ago
  • NVIDIA AI is seeking an engineer to implement quantized and sparse recipes in inference engines and to manage model export pipelines for...  ...The role requires strong Python and C++ skills, experience with ML accelerators, and familiarity with PyTorch internals, with 4+ years... 

    NVIDIA AI

    Redmond, WA
    3 days ago
  • $104k - $208k

     ...prefer prior experience as a software developer, coder, software engineer, or programmer. We only consider applicants located in the...  ...rates, with opportunities for higher earnings based on performance. We operate remotely for qualified candidates in the US, Canada... 
    Performance
    Hourly pay
    Full time
    Remote work
    Flexible hours

    DataAnnotation

    Redmond, WA
    7 days ago
  • Google is seeking a Software Engineering Manager II to lead a team focused on AI/ML for Google Cloud Compute. You will influence stakeholders and drive the decision-making process while owning project goals and contributing to product strategy. Ideal candidates should... 
    Full time

    Google

    Kirkland, WA
    5 days ago
  • Robinhood is seeking a Staff/ Senior Staff Machine Learning Engineer to own design and delivery of personalization and recommendation systems...  ..., data engineering, and platform teams to drive end-to-end ML lifecycle from feature engineering to deployment and monitoring.... 

    Robinhood

    Bellevue, WA
    5 days ago
  • $117.3k - $195.5k

     ...analysis) Gather, prepares, and maintains datasets required to perform analytics Oversee data systems holding in close...  ...implementation and deployment of emerging tools and analytic data engineering process in order to improve team productivity BASIC QUALIFICATIONS... 
    Performance
    Work experience placement
    Remote work
    Relocation package
    Flexible hours

    Pfizer

    Bothell, WA
    4 days ago
  •  ...WiproContact: Meghana GorusuCompany: SRI Tech SolutionsJob RoleAgentic AI/ML EngineerLocationRedmond, WAPrimary SkillsMachine Learning,...  ...pipelines, and operational excellence Quality & Integration Engineering - focusing on agent behavior testing, UI/UX integration, and... 
    Full time

    SRI Tech

    Redmond, WA
    5 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to ML Performance Engineer. Be the first to apply!