Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

ML Performance Engineer

$100k - $150k
Full-time

Bright Vision Technologies

Renton, WA
ML Performance Engineer - Remote 

 
Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States. 
This is a fantastic opportunity to join an established and well-respected organization offering tremendous career growth potential. 

 
Job Title: ML Performance Engineer
Location: 100% Remote (U.S.) 
Position Type:  Full-time, Direct W2 
Salary Range: $100,000–$150,000 Annually 
Experience Required: 6+ years 

 
Sponsorship:  U.S. Citizens, Green Card Holders, EAD Holders, and H-1B transfer candidates are encouraged to apply. We are unable to sponsor new H-1B visa petitions for this position. 

 
Job Summary 
We are seeking an AI Performance Optimization Engineer to focus on extracting maximum throughput, minimizing latency, and reducing cost across training and inference workloads for large neural network systems. The role spans the full stack from low-level kernel optimization to distributed system tuning, requiring deep understanding of GPU architecture, model parallelism, memory management, and compiler-level optimization. The ideal candidate has demonstrated impact on production AI workloads, with strong instrumentation and measurement discipline that enables rigorous, data-driven optimization decisions. In this role you will work closely with cross-functional partners — product, design, engineering, operations, and business stakeholders — to translate ambiguous requirements into well-engineered solutions, and will be expected to raise the bar through code review, design review, and mentorship of more junior engineers. The successful candidate brings strong engineering discipline, a clear communication style, and a track record of shipping meaningful work that holds up well in production. 

Key Responsibilities  
  • Profile and optimize end-to-end AI training and inference pipelines for throughput, latency, and cost. 
  • Identify and eliminate bottlenecks across data loading, model compute, communication, and memory. 
  • Implement and tune quantization, sparsity, and pruning strategies to reduce model footprint and accelerate inference. 
  • Optimize distributed training using tensor parallelism, pipeline parallelism, FSDP, and ZeRO-style sharding. 
  • Tune attention implementations using FlashAttention, paged attention, and related techniques. 
  • Implement KV cache optimization, continuous batching, and speculative decoding for LLM serving. 
  • Drive compiler-level optimizations using Triton, XLA, TorchInductor, or TVM, working with the broader ML framework community to land improvements that translate into measurable end-to-end performance gains. 
  • Optimize data pipelines, sharding strategies, and storage access patterns for high-throughput training. 
  • Build and maintain rigorous benchmark suites and regression frameworks across workloads. 
  • Collaborate with ML and platform engineering teams to embed best practices in standard pipelines. 
  • Drive cost-efficiency improvements through model architecture, hardware selection, and scheduling strategies. 
  • Evaluate new hardware and software offerings, and advise on adoption. 
  • Document performance tuning playbooks and share findings broadly across engineering teams. 
  • Stay current with AI systems research and translate advances into production improvements. 
Required Qualifications 
  • Bachelor’s or Master’s degree in Computer Science, Computer Engineering, or a related field. 
  • Six or more years of experience in performance engineering, ML systems, or HPC. 
  • Strong proficiency in Python and C++. 
  • Hands-on experience optimizing deep learning workloads on modern GPUs. 
  • Deep understanding of distributed training and inference techniques. 
  • Experience with profiling tools across CPU, GPU, and distributed systems. 
  • Familiarity with model compression techniques and their accuracy implications. 
  • Strong grasp of memory hierarchies, communication primitives, and parallelism strategies. 
  • Excellent measurement, debugging, and analytical reasoning skills. 
  • Strong communication and collaboration skills. 
Preferred Qualifications 
  • Experience optimizing LLM inference at production scale. 
  • Contributions to vLLM, TensorRT-LLM, DeepSpeed, or similar projects. 
  • Familiarity with custom kernel authoring in Triton or CUTLASS. 
  • Experience with FinOps for AI workloads. 
  • Publications or talks on AI systems performance. 
How to Apply 
Would you like to know more about this opportunity? For immediate consideration, please send your resume to View email address on brightvisiontechnologies.applytojob.com or contact us at Show phone number. Learn more about Bright Vision Technologies at  .
Bright Vision Technologies is an Equal Opportunity Employer.

 

Equal Employment Opportunity (EEO) Statement

Bright Vision Technologies (BV Teck) is committed to equal employment opportunity (EEO) for all employees and applicants without regard to race, color, religion, sex, sexual orientation, gender identity or expression, national origin, age, genetic information, disability, veteran status, or any other protected status as defined by applicable federal, state, or local laws. This commitment extends to all aspects of employment, including recruitment, hiring, training, compensation, promotion, transfer, leaves of absence, termination, layoffs, and recall.

BV Teck expressly prohibits any form of workplace harassment or discrimination. Any improper interference with employees' ability to perform their job duties may result in disciplinary action up to and including termination of employment.

Vacancy posted 8 days ago
Similar jobs that could be interesting for youBased on the ML Performance Engineer in Renton, WA vacancy
  • $148.5k - $223.9k

     ...future of Salesforce.This role is for a Senior Machine Learning Engineer within the Trust Intelligence Platform team who will architect...  ..., promotion, benefits, training, assessment of job performance, discipline, termination, and everything in between. Recruiting... 
    Performance
    Full time

    Salesforce

    Bellevue, WA
    8 hours ago
  • $175k - $200k

    Senior Machine Learning Engineer Truveta is the world’s first health provider led data platform...  ...large-scale engineering. You’ll blend ML craftsmanship with platform engineering...  ...scaling laws influence quality, cost, and performance. Work fluently with embeddings and vector... 
    Performance
    For contractors
    Visa sponsorship
    Work visa
    Flexible hours

    Truveta

    Seattle, WA
    1 day ago
  • $106.9k - $160.4k

     ...to scale AI across the enterprise, we are seeking a skilled ML Engineer to design, build, and operationalize machine learning solutions...  ...-native services and containerized architectures, ensuring performance, reliability, and cost efficiency.ML System Design & IntegrationDesign... 
    Performance
    Full time
    Temporary work

    Weyerhaeuser

    Seattle, WA
    3 days ago
  • $190k - $230k

     ...Machine Learning Engineer Location: Hybrid - Seattle, WA Comp: $190-230k base + startup equity Our client,...  ...inference platform. If you enjoy working at the intersection of ML, systems, and performance engineering - and want to shape core infrastructure from... 
    Performance
    Contract work
    Local area

    Prime Team Partners

    Seattle, WA
    4 days ago
  • $148.5k - $223.9k

     ...efforts. Job Category Software Engineering Job Details About Salesforce Salesforce...  ...to be a Trailblazer, too - driving your performance and career growth, charting new paths,...  ...managing automated, production-grade ML pipelines. Software Engineering... 
    Performance

    Salesforce.Com Inc

    Bellevue, WA
    4 days ago
  • $71.39 - $107.64 per hour

     ...to be a Trailblazer, too — driving your performance and career growth, charting new paths, and...  ...learning pipelines across the security engineering organization. We are looking for a...  ...for managing automated, production-grade ML pipelines.Software Engineering Excellence... 
    Performance
    Full time

    Salesforce

    Bellevue, WA
    5 days ago
  • $146.83k - $192.72k

     ...Description & Requirements Who we arelululemon is an innovative performance apparel company for yoga, running, training, and other...  ...drives enterprise efficiency. Core responsibilities As an AI/ML Engineer, you will contribute to the design and implementation of AI/ML... 
    Performance
    Permanent employment
    Full time
    Part time
    Work visa

    Lululemon Athletica

    Seattle, WA
    1 day ago
  • $176.76k - $232k

     ...Description & Requirements Who we arelululemon is an innovative performance apparel company for yoga, running, training, and other...  ...drives enterprise efficiency.Core responsibilities As a Senior AI/ML Engineer, you will lead the delivery of scalable AI/ML solutions to... 
    Performance
    Permanent employment
    Full time
    Contract work
    Part time
    Work visa

    Lululemon Athletica

    Seattle, WA
    2 days ago
  • $114.1k - $160k

     ...you like to use network and Unix systems engineering to deliver simple, sustainable, and...  ...developing, building, deploying, operating, performance optimization and scaling the Amazon networks...  ...performance of the Machine Learning (ML) network infrastructure across all of... 
    Performance
    Local area
    Flexible hours

    Amazon

    Seattle, WA
    4 days ago
  • $143.7k - $194.4k

    We are looking for an **AI/ML Engineer** to build, deploy, and operate the ML/AI systems that power the agentic decision intelligence...  ...to production. Implement model monitoring: drift detection, performance degradation alerts, automated retraining triggers- Build A/B... 
    Performance
    Internship
    Flexible hours

    Amazon

    Seattle, WA
    8 hours ago
  • $148.7k - $201.2k

     ...leading work delivering continuous price performance improvements in the cloud for AI model...  ...high performance and scalability in AI/ML and HPC workloads.You are intrigued by the...  ...for builders like you. The AWS Hardware Engineering team creates server designs for Amazon’s... 
    Performance
    Internship
    Local area
    Flexible hours

    Amazon

    Seattle, WA
    8 hours ago
  • $184.5k

     ...and AI tooling used in practice every day.As a Senior AI / ML Data Engineer, you will build the data foundations that power Layla’s AI-native...  ...of the range, based on ongoing, demonstrated, and sustained performance in the role. The total cash range for this position in... 
    Performance
    Full time

    Expedia

    Seattle, WA
    2 days ago
  • $120k - $180k

     ...team of best-in-class machine learning engineers. We are looking for developers who are excited...  ...Everything involved in applying a ML model to a production use case, including...  ...into production, with measurably improved performance over baseline, either in industry or as... 
    Performance
    Full time

    Hive

    Seattle, WA
    1 day ago
  • $179.8k - $236k

     ...Description & Requirements who we arelululemon is an innovative performance apparel company for yoga, running, training, and other...  ...enterprise efficiency.core responsibilitiesAs a Senior Manager, AI/ML Engineering, you lead a team of AI/ML Engineers, setting technical... 
    Performance
    Permanent employment
    Full time
    Part time
    Work visa

    Lululemon Athletica

    Seattle, WA
    1 day ago
  •  ...opportunity for you to take your software engineering career to the next level. As a Software...  ...experience, with emphasis on ML systems.Hands-on experience using enterprise...  ...refine AI-generated outputs for correctness, performance, and security.Responsible for AI use in... 
    Performance

    JP Morgan Chase

    Seattle, WA
    1 day ago
  •  ...-only architectures, combining rigorous engineering with learning systems proven in globally...  ...in the field. We are seeking a Staff ML Systems Engineer to architect and build...  ...distributed pipeline development. Optimize performance across distributed CPU and GPU... 
    Performance
    Local area

    FieldAI

    Seattle, WA
    21 days ago
  • $195.7k - $338.4k

     ...On-Device ML Engineering Manager (Tools & Services) The On-Device Machine Learning team at Apple is responsible for enabling the Research...  .... Our group is seeking an Engineering Manager to lead the Performance Tools and Services team, with a focus on the tools, services... 
    Performance
    Relocation

    Apple

    Seattle, WA
    3 days ago
  •  ...smooth operations and customer satisfaction. As a Machine Learning Engineer focused on demand forecasting, you will contribute to the...  ...training, and prediction processes.6. Evaluate and optimize the performance of existing time series forecasting models, and propose... 
    Performance

    TikTok

    Seattle, WA
    1 day ago
  • $157.44k - $236.2k

     ...highly skilled and motivated Machine Learning Engineer to join our team and play a key role in...  ....Optimize knowledge graph algorithms for performance, scalability, and reliability.Conduct...  ...and implement scalable, production-ready ML services;3 years of experience optimizing... 
    Performance
    Full time
    Work at office
    Remote work

    Zoom

    Seattle, WA
    3 days ago
  • $209k - $313k

     ...Saturn, and other digital services.Snap Engineering teams build fun and technically sophisticated...  ...millions of SnapchattersApply modern ML techniques to solve large-scale, real-...  ...generated output for architectural integrity, performance bottlenecks, and security... 
    Performance
    Full time
    Live in
    Work at office
    Local area

    Snap

    Seattle, WA
    2 days ago
  • $184.5k

     ...satisfaction.This Senior Machine Learning Engineer role is part of the Distribution & Supply...  ...that directly improve the quality and performance of our distribution platform for both travelers...  ...and customer problems into clear ML‑driven solutions, selecting appropriate modeling... 
    Performance
    Full time

    Expedia

    Seattle, WA
    2 days ago
  • $151.8k - $265.35k

     ...team is seeking Senior Machine Learning Engineers for our GenAI Services area. In this high...  ...engineers in building scalable, high-performance generative AI systems—powering features...  ...set technical direction, and mentor other ML engineers.Job ResponsibilitiesDesign and... 
    Performance
    Full time
    Temporary work
    Local area
    Worldwide

    Adobe Systems

    Seattle, WA
    5 days ago
  • $173k - $259k

     ...Saturn, and other digital services.Snap Engineering teams build fun and technically sophisticated...  ...millions of SnapchattersApply modern ML techniques to solve large-scale, real-...  ...generated output for architectural integrity, performance bottlenecks, and security risks.... 
    Performance
    Full time
    Live in
    Work at office
    Local area

    Snap

    Seattle, WA
    2 days ago
  • $183.6k - $275.4k

     ...functional teams to integrate research findings into scalable engineering solutions that align with business objectives.Participate in...  ...existing systems and proposing innovative solutions to enhance performance, scalability, and reliability.Stay up to date with the newest... 
    Performance
    Full time
    Work at office
    Remote work
    1 day per week

    Zoom

    Seattle, WA
    3 days ago
  • $168.1k - $227.4k

     ...RoleWe're seeking a talented Senior Machine Learning Engineer with expertise in agentic system, production ML systems, and scalable deploymentarchitectures. You...  ...AI and ML methods, ensuring reliability and performance• Establish scalable, efficient, automated processes... 
    Performance
    Internship
    Worldwide
    Flexible hours

    Amazon

    Seattle, WA
    4 days ago
  • $101.9k - $224k

     ...team that wants you to grow and succeed. What you'll doAs a ML Engineer in Bellevue, you will join a new team dedicated to building and...  ...actual payout amount is dependent on company and personal performance. Please reference this link for a summary of SAP benefits and... 
    Performance
    Permanent employment
    Full time
    Worldwide
    Flexible hours

    SAP

    Bellevue, WA
    1 day ago
  • $168.1k - $227.4k

     ...cheaper is as much a systems problem as a modeling problem - agent performance depends on the model, the harness, the evaluation...  ...designed together.We are looking for a Senior Machine Learning Engineer to build and own core systems in this agentic platform. You will... 
    Performance
    Internship
    Flexible hours
    Day shift

    Amazon

    Bellevue, WA
    1 day ago
  • $185.6k - $255k

     ...into a live segment without waiting on an engineering queue. The bet underneath all of it is...  ...to speak with you.The RoleAt Amperity, ML Engineers work in small, collaborative,...  ...automated retraining pipelines triggered by performance degradation, and architect monitoring... 
    Performance
    Work at office
    Local area
    Remote work

    Amperity

    Seattle, WA
    8 hours ago
  • $141.9k - $190.3k

     ...Technology is a global organization of engineers, product developers, designers, technologists...  ...and products - driving advertising performance, innovation, and value in Disney’s sports...  ...maintainable, and testable software and apply ML solutions where needed.Daily, you should... 
    Performance
    Work experience placement
    Work at office

    Disney Interactive

    Seattle, WA
    6 days ago
  • $119.5k - $190k

     ...Summary:We are seeking a talented and passionate Machine Learning Engineer II to join our growing team. In this role, you will play a...  ...various techniques and metrics. Fine-tune models to achieve optimal performance and efficiency. Deploy machine learning models to production... 
    Performance
    Local area
    Flexible hours

    Chewy

    Bellevue, WA
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to ML Performance Engineer. Be the first to apply!