Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

ML Performance Engineer

$100k - $150k
Full-time

Bright Vision Technologies

ML Performance Engineer - Remote 

 
Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States. 
This is a fantastic opportunity to join an established and well-respected organization offering tremendous career growth potential. 

 
Job Title: ML Performance Engineer
Location: 100% Remote (U.S.) 
Position Type:  Full-time, Direct W2 
Salary Range: $100,000–$150,000 Annually 
Experience Required: 6+ years 

 
Sponsorship:  U.S. Citizens, Green Card Holders, EAD Holders, and H-1B transfer candidates are encouraged to apply. We are unable to sponsor new H-1B visa petitions for this position. 

 
Job Summary 
We are seeking an AI Performance Optimization Engineer to focus on extracting maximum throughput, minimizing latency, and reducing cost across training and inference workloads for large neural network systems. The role spans the full stack from low-level kernel optimization to distributed system tuning, requiring deep understanding of GPU architecture, model parallelism, memory management, and compiler-level optimization. The ideal candidate has demonstrated impact on production AI workloads, with strong instrumentation and measurement discipline that enables rigorous, data-driven optimization decisions. In this role you will work closely with cross-functional partners — product, design, engineering, operations, and business stakeholders — to translate ambiguous requirements into well-engineered solutions, and will be expected to raise the bar through code review, design review, and mentorship of more junior engineers. The successful candidate brings strong engineering discipline, a clear communication style, and a track record of shipping meaningful work that holds up well in production. 

Key Responsibilities  
  • Profile and optimize end-to-end AI training and inference pipelines for throughput, latency, and cost. 
  • Identify and eliminate bottlenecks across data loading, model compute, communication, and memory. 
  • Implement and tune quantization, sparsity, and pruning strategies to reduce model footprint and accelerate inference. 
  • Optimize distributed training using tensor parallelism, pipeline parallelism, FSDP, and ZeRO-style sharding. 
  • Tune attention implementations using FlashAttention, paged attention, and related techniques. 
  • Implement KV cache optimization, continuous batching, and speculative decoding for LLM serving. 
  • Drive compiler-level optimizations using Triton, XLA, TorchInductor, or TVM, working with the broader ML framework community to land improvements that translate into measurable end-to-end performance gains. 
  • Optimize data pipelines, sharding strategies, and storage access patterns for high-throughput training. 
  • Build and maintain rigorous benchmark suites and regression frameworks across workloads. 
  • Collaborate with ML and platform engineering teams to embed best practices in standard pipelines. 
  • Drive cost-efficiency improvements through model architecture, hardware selection, and scheduling strategies. 
  • Evaluate new hardware and software offerings, and advise on adoption. 
  • Document performance tuning playbooks and share findings broadly across engineering teams. 
  • Stay current with AI systems research and translate advances into production improvements. 
Required Qualifications 
  • Bachelor’s or Master’s degree in Computer Science, Computer Engineering, or a related field. 
  • Six or more years of experience in performance engineering, ML systems, or HPC. 
  • Strong proficiency in Python and C++. 
  • Hands-on experience optimizing deep learning workloads on modern GPUs. 
  • Deep understanding of distributed training and inference techniques. 
  • Experience with profiling tools across CPU, GPU, and distributed systems. 
  • Familiarity with model compression techniques and their accuracy implications. 
  • Strong grasp of memory hierarchies, communication primitives, and parallelism strategies. 
  • Excellent measurement, debugging, and analytical reasoning skills. 
  • Strong communication and collaboration skills. 
Preferred Qualifications 
  • Experience optimizing LLM inference at production scale. 
  • Contributions to vLLM, TensorRT-LLM, DeepSpeed, or similar projects. 
  • Familiarity with custom kernel authoring in Triton or CUTLASS. 
  • Experience with FinOps for AI workloads. 
  • Publications or talks on AI systems performance. 
How to Apply 
Would you like to know more about this opportunity? For immediate consideration, please send your resume to View email address on brightvisiontechnologies.applytojob.com or contact us at Show phone number. Learn more about Bright Vision Technologies at  .
Bright Vision Technologies is an Equal Opportunity Employer.

 

Equal Employment Opportunity (EEO) Statement

Bright Vision Technologies (BV Teck) is committed to equal employment opportunity (EEO) for all employees and applicants without regard to race, color, religion, sex, sexual orientation, gender identity or expression, national origin, age, genetic information, disability, veteran status, or any other protected status as defined by applicable federal, state, or local laws. This commitment extends to all aspects of employment, including recruitment, hiring, training, compensation, promotion, transfer, leaves of absence, termination, layoffs, and recall.

BV Teck expressly prohibits any form of workplace harassment or discrimination. Any improper interference with employees' ability to perform their job duties may result in disciplinary action up to and including termination of employment.

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the ML Performance Engineer in Washington State vacancy
  • $148.5k - $223.9k

     ...future of Salesforce.This role is for a Senior Machine Learning Engineer within the Trust Intelligence Platform team who will architect...  ..., promotion, benefits, training, assessment of job performance, discipline, termination, and everything in between. Recruiting... 
    Performance
    Full time

    Salesforce

    Bellevue, WA
    1 day ago
  • $106.9k - $160.4k

     ...to scale AI across the enterprise, we are seeking a skilled ML Engineer to design, build, and operationalize machine learning solutions...  ...-native services and containerized architectures, ensuring performance, reliability, and cost efficiency.ML System Design & IntegrationDesign... 
    Performance
    Full time
    Temporary work

    Weyerhaeuser

    Seattle, WA
    4 days ago
  • $175k - $200k

    Senior Machine Learning Engineer Truveta is the world’s first health provider led data platform...  ...large-scale engineering. You’ll blend ML craftsmanship with platform engineering...  ...scaling laws influence quality, cost, and performance. Work fluently with embeddings and vector... 
    Performance
    For contractors
    Visa sponsorship
    Work visa
    Flexible hours

    Truveta

    Seattle, WA
    2 days ago
  • $148.5k - $223.9k

     ...to be a Trailblazer, too — driving your performance and career growth, charting new paths, and...  ...learning pipelines across the security engineering organization. We are looking for a...  ...for managing automated, production-grade ML pipelines.Software Engineering Excellence... 
    Performance
    Full time

    Salesforce

    Bellevue, WA
    4 days ago
  • $182k - $242k

     ...CoreWeave combines superior infrastructure performance with deep technical expertise to...  ...qualifications, we’re looking for strong engineers with great taste. The most important qualification...  ...and hands on experience with modern ML frameworks such as PyTorch or JAX ~... 
    Performance
    Full time
    Temporary work
    Casual work
    Work at office
    Flexible hours

    CoreWeave

    Bellevue, WA
    5 days ago
  • $175k - $308.5k

     ...reconstruction Experience with training data curation and pipeline engineering Experience with model optimization for production deployment Proficiency in C++ or experience integrating ML models into performance‑critical systems Strong communication skills and ability to... 
    Performance
    Relocation

    Apple

    Seattle, WA
    2 days ago
  • $251k - $310k

     ...-art Multimodal LLMs and World models to perform 3D Perception using sensor information from...  ...and Radar.. Partner effectively with engineering and research teams across Waymo to deploy...  ...(e.g., xprof), and debugging of ML models. You have: PhD or Masters... 
    Performance
    Full time
    Temporary work
    Remote work

    Waymo

    Kirkland, WA
    2 days ago
  •  ...-only architectures, combining rigorous engineering with learning systems proven in globally...  ...in the field. About This Role Modern ML systems improve through their data...  ...record of field deployments and strong performance in DARPA challenge segments. Be Part... 
    Performance
    Local area

    FieldAI

    Seattle, WA
    7 days ago
  • $114.1k - $160k

     ...you like to use network and Unix systems engineering to deliver simple, sustainable, and...  ...developing, building, deploying, operating, performance optimization and scaling the Amazon networks...  ...performance of the Machine Learning (ML) network infrastructure across all of... 
    Performance
    Local area
    Flexible hours

    Amazon

    Seattle, WA
    1 day ago
  • $146.83k - $192.72k

     ...Description & Requirements Who we arelululemon is an innovative performance apparel company for yoga, running, training, and other...  ...drives enterprise efficiency. Core responsibilities As an AI/ML Engineer, you will contribute to the design and implementation of AI/ML... 
    Performance
    Permanent employment
    Full time
    Part time
    Work visa

    Lululemon Athletica

    Seattle, WA
    2 days ago
  • $176.76k - $232k

     ...Description & Requirements Who we arelululemon is an innovative performance apparel company for yoga, running, training, and other...  ...drives enterprise efficiency.Core responsibilities As a Senior AI/ML Engineer, you will lead the delivery of scalable AI/ML solutions to... 
    Performance
    Permanent employment
    Full time
    Contract work
    Part time
    Work visa

    Lululemon Athletica

    Seattle, WA
    3 days ago
  • $148.7k - $201.2k

     ...leading work delivering continuous price performance improvements in the cloud for AI model...  ...high performance and scalability in AI/ML and HPC workloads.You are intrigued by the...  ...for builders like you. The AWS Hardware Engineering team creates server designs for Amazon’s... 
    Performance
    Internship
    Local area
    Flexible hours

    Amazon

    Seattle, WA
    1 day ago
  • $143.7k - $194.4k

    We are looking for an **AI/ML Engineer** to build, deploy, and operate the ML/AI systems that power the agentic decision intelligence...  ...to production. Implement model monitoring: drift detection, performance degradation alerts, automated retraining triggers- Build A/B... 
    Performance
    Internship
    Flexible hours

    Amazon

    Seattle, WA
    1 day ago
  • $144.7k - $261.3k

     ...Foundations, solves critical evaluation challenges for autonomous vehicle development. We engineer high-performance tools that identify top-performing models and partner with data-intensive ML teams to drive rapid innovation. Job Description About the Team The Evaluation... 
    Performance
    Local area
    Remote work
    Work from home
    Flexible hours

    General Motors

    Seattle, WA
    1 day ago
  •  ...opportunity for you to take your software engineering career to the next level. As a Software...  ...experience, with emphasis on ML systems.Hands-on experience using enterprise...  ...refine AI-generated outputs for correctness, performance, and security.Responsible for AI use in... 
    Performance

    JP Morgan Chase

    Seattle, WA
    3 days ago
  • $213.76k - $280.55k

     ...Description & Requirements who we arelululemon is an innovative performance apparel company for yoga, running, training, and other...  ...enterprise efficiency.core responsibilitiesAs a Senior Manager, AI/ML Engineering, you lead a team of AI/ML Engineers, setting technical... 
    Performance
    Permanent employment
    Full time
    Part time
    Work visa

    Lululemon Athletica

    Seattle, WA
    2 days ago
  • $144.7k - $261.3k

     ...org provides developer environments, cloud infrastructure, and ML/AI GPU platforms for AV research and development teams to...  ...test, and run faster in GM. The Role GM is looking for a Senior Performance Engineer to join the AV Capacity and Performance Engineering team in... 
    Performance
    Work at office
    Local area
    Remote work
    Work from home
    Flexible hours
    3 days per week

    General Motors Ventures

    Seattle, WA
    4 days ago
  •  ...smooth operations and customer satisfaction. As a Machine Learning Engineer focused on demand forecasting, you will contribute to the...  ...training, and prediction processes.6. Evaluate and optimize the performance of existing time series forecasting models, and propose... 
    Performance

    TikTok

    Seattle, WA
    2 days ago
  • $209k - $313k

     ...Saturn, and other digital services.Snap Engineering teams build fun and technically sophisticated...  ...millions of SnapchattersApply modern ML techniques to solve large-scale, real-...  ...generated output for architectural integrity, performance bottlenecks, and security... 
    Performance
    Full time
    Live in
    Work at office
    Local area

    Snap

    Seattle, WA
    3 days ago
  • $168.1k - $227.4k

     ...RoleWe're seeking a talented Senior Machine Learning Engineer with expertise in agentic system, production ML systems, and scalable deploymentarchitectures. You...  ...AI and ML methods, ensuring reliability and performance• Establish scalable, efficient, automated processes... 
    Performance
    Internship
    Worldwide
    Flexible hours

    Amazon

    Seattle, WA
    11 hours ago
  • $173k - $259k

     ...Saturn, and other digital services.Snap Engineering teams build fun and technically sophisticated...  ...millions of SnapchattersApply modern ML techniques to solve large-scale, real-...  ...generated output for architectural integrity, performance bottlenecks, and security risks.... 
    Performance
    Full time
    Live in
    Work at office
    Local area

    Snap

    Seattle, WA
    3 days ago
  • $151.8k - $265.35k

     ...team is seeking SeniorMachine Learning Engineers for our GenAI Services area. In this high...  ...talented engineers in building scalable, high-performance generative AI systems—powering features...  ...technical direction, and mentor other ML engineers.Job ResponsibilitiesDesign and... 
    Performance
    Full time
    Temporary work
    Local area
    Worldwide

    Adobe Systems

    Seattle, WA
    11 hours ago
  • $203k - $258.6k

     ...together AI researchers, machine learning engineers, data engineers, and networking domain...  ...you will build and improve the data and ML systems that power our LLMs and AI models...  ...and how to measure its impact on model performance.What You’ll DoDesign, build, and own end... 
    Performance
    Full time
    Temporary work
    Work at office
    Local area
    Flexible hours
    Shift work

    CISCO Systems

    Seattle, WA
    2 days ago
  •  ...the Role:As a member of the Product and Engineering team at PitchBook, you will be part of a...  ...Machine Learning Engineer (MLE) on the AI & ML (Insights) team, you will play a...  ...our systems meet the highest standards of performance, reliability, and security.Primary Job Responsibilities... 
    Performance
    Work at office
    Remote work
    Visa sponsorship

    PitchBook Data

    Seattle, WA
    11 hours ago
  •  ...outcomes.About the RoleAs a Machine Learning Engineer on the Drive team, you'll own machine...  ...model iteration to continuously improve performance.Partner closely with software engineers,...  ...excited to work across a diverse set of ML techniques—from neural networks and optimization... 
    Performance
    Hourly pay
    Work at office
    Local area
    Remote work
    Relocation
    Flexible hours

    Doordash

    Seattle, WA
    2 days ago
  • $98.8k - $148.2k

     ...enterprise, we are seeking a Machine Learning Engineer to help operationalize machine learning...  ...candidate has practical experience with ML deployment pipelines, cloud-native...  ...containerized architectures, with attention to performance, reliability, and cost efficiency.... 
    Performance
    Full time
    Temporary work

    Weyerhaeuser

    Seattle, WA
    11 hours ago
  • $157.44k - $236.2k

     ...highly skilled and motivated Machine Learning Engineer to join our team and play a key role in...  ....Optimize knowledge graph algorithms for performance, scalability, and reliability.Conduct...  ...and implement scalable, production-ready ML services;3 years of experience optimizing... 
    Performance
    Full time
    Work at office
    Remote work

    Zoom

    Seattle, WA
    4 days ago
  • $184.5k

     ...satisfaction.This Senior Machine Learning Engineer role is part of the Distribution & Supply...  ...that directly improve the quality and performance of our distribution platform for both travelers...  ...and customer problems into clear ML‑driven solutions, selecting appropriate modeling... 
    Performance
    Full time

    Expedia

    Seattle, WA
    3 days ago
  • $183.6k - $275.4k

     ...functional teams to integrate research findings into scalable engineering solutions that align with business objectives.Participate in...  ...existing systems and proposing innovative solutions to enhance performance, scalability, and reliability.Stay up to date with the newest... 
    Performance
    Full time
    Work at office
    Remote work
    1 day per week

    Zoom

    Seattle, WA
    4 days ago
  • $150.75k - $241.2k

     ...Your Impact As an Senior Machine Learning Engineer II focused on Sensor Fusion & Tracking,...  ...tracking.Conduct simulation studies and performance benchmarking across varying operational conditions...  ...equivalent hands-on experience building ML systems in production.8+ years of... 
    Performance
    Work experience placement
    Work at office
    Remote work
    Worldwide

    Axon

    Seattle, WA
    11 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to ML Performance Engineer. Be the first to apply!