Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

ML Performance Engineer

$100k - $150k
Full-time

Bright Vision Technologies

Redmond, WA
ML Performance Engineer - Remote 

 
Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States. 
This is a fantastic opportunity to join an established and well-respected organization offering tremendous career growth potential. 

 
Job Title: ML Performance Engineer
Location: 100% Remote (U.S.) 
Position Type:  Full-time, Direct W2 
Salary Range: $100,000–$150,000 Annually 
Experience Required: 6+ years 

 
Sponsorship:  U.S. Citizens, Green Card Holders, EAD Holders, and H-1B transfer candidates are encouraged to apply. We are unable to sponsor new H-1B visa petitions for this position. 

 
Job Summary 
We are seeking an AI Performance Optimization Engineer to focus on extracting maximum throughput, minimizing latency, and reducing cost across training and inference workloads for large neural network systems. The role spans the full stack from low-level kernel optimization to distributed system tuning, requiring deep understanding of GPU architecture, model parallelism, memory management, and compiler-level optimization. The ideal candidate has demonstrated impact on production AI workloads, with strong instrumentation and measurement discipline that enables rigorous, data-driven optimization decisions. In this role you will work closely with cross-functional partners — product, design, engineering, operations, and business stakeholders — to translate ambiguous requirements into well-engineered solutions, and will be expected to raise the bar through code review, design review, and mentorship of more junior engineers. The successful candidate brings strong engineering discipline, a clear communication style, and a track record of shipping meaningful work that holds up well in production. 

Key Responsibilities  
  • Profile and optimize end-to-end AI training and inference pipelines for throughput, latency, and cost. 
  • Identify and eliminate bottlenecks across data loading, model compute, communication, and memory. 
  • Implement and tune quantization, sparsity, and pruning strategies to reduce model footprint and accelerate inference. 
  • Optimize distributed training using tensor parallelism, pipeline parallelism, FSDP, and ZeRO-style sharding. 
  • Tune attention implementations using FlashAttention, paged attention, and related techniques. 
  • Implement KV cache optimization, continuous batching, and speculative decoding for LLM serving. 
  • Drive compiler-level optimizations using Triton, XLA, TorchInductor, or TVM, working with the broader ML framework community to land improvements that translate into measurable end-to-end performance gains. 
  • Optimize data pipelines, sharding strategies, and storage access patterns for high-throughput training. 
  • Build and maintain rigorous benchmark suites and regression frameworks across workloads. 
  • Collaborate with ML and platform engineering teams to embed best practices in standard pipelines. 
  • Drive cost-efficiency improvements through model architecture, hardware selection, and scheduling strategies. 
  • Evaluate new hardware and software offerings, and advise on adoption. 
  • Document performance tuning playbooks and share findings broadly across engineering teams. 
  • Stay current with AI systems research and translate advances into production improvements. 
Required Qualifications 
  • Bachelor’s or Master’s degree in Computer Science, Computer Engineering, or a related field. 
  • Six or more years of experience in performance engineering, ML systems, or HPC. 
  • Strong proficiency in Python and C++. 
  • Hands-on experience optimizing deep learning workloads on modern GPUs. 
  • Deep understanding of distributed training and inference techniques. 
  • Experience with profiling tools across CPU, GPU, and distributed systems. 
  • Familiarity with model compression techniques and their accuracy implications. 
  • Strong grasp of memory hierarchies, communication primitives, and parallelism strategies. 
  • Excellent measurement, debugging, and analytical reasoning skills. 
  • Strong communication and collaboration skills. 
Preferred Qualifications 
  • Experience optimizing LLM inference at production scale. 
  • Contributions to vLLM, TensorRT-LLM, DeepSpeed, or similar projects. 
  • Familiarity with custom kernel authoring in Triton or CUTLASS. 
  • Experience with FinOps for AI workloads. 
  • Publications or talks on AI systems performance. 
How to Apply 
Would you like to know more about this opportunity? For immediate consideration, please send your resume to View email address on brightvisiontechnologies.applytojob.com or contact us at Show phone number. Learn more about Bright Vision Technologies at  .
Bright Vision Technologies is an Equal Opportunity Employer.

 

Equal Employment Opportunity (EEO) Statement

Bright Vision Technologies (BV Teck) is committed to equal employment opportunity (EEO) for all employees and applicants without regard to race, color, religion, sex, sexual orientation, gender identity or expression, national origin, age, genetic information, disability, veteran status, or any other protected status as defined by applicable federal, state, or local laws. This commitment extends to all aspects of employment, including recruitment, hiring, training, compensation, promotion, transfer, leaves of absence, termination, layoffs, and recall.

BV Teck expressly prohibits any form of workplace harassment or discrimination. Any improper interference with employees' ability to perform their job duties may result in disciplinary action up to and including termination of employment.

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the ML Performance Engineer in Redmond, WA vacancy
  • $175k - $215k

     ...from a diverse set of sensors, enabling engineers like you to (1) develop methods for efficiently...  ...require a diverse skill set of ML and geometric algorithms, onboard sensing...  ...and metrics for measuring and improving performance of pre-trained and fine-tuned models... 
    Performance
    Full time
    Remote work

    Waymo

    Kirkland, WA
    16 hours ago
  • $148.5k - $223.9k

     ...future of Salesforce.This role is for a Senior Machine Learning Engineer within the Trust Intelligence Platform team who will architect...  ..., promotion, benefits, training, assessment of job performance, discipline, termination, and everything in between. Recruiting... 
    Performance
    Full time

    Salesforce

    Bellevue, WA
    2 days ago
  •  ...Snowflake is seeking an Senior Software Engineer for Cortex Training to advance the ML platform that lets customers run demanding ML/AI workloads inside...  ...into reliable, enterprise-grade components, with a focus on performance, reliability, and cost #J-18808-Ljbffr... 
    Performance

    Neura Market

    Bellevue, WA
    21 hours ago
  • $100k - $150k

     ...ML Infrastructure Engineer - Remote    Bright Vision Technologies is a technology consulting and software development company delivering...  ..., distributed training frameworks, scheduling, storage performance, and developer experience for ML engineers and researchers... 
    Performance
    Full time
    H1b
    Local area
    Immediate start
    Remote work
    Visa sponsorship

    Bright Vision Technologies

    Bellevue, WA
    a month ago
  • $175k - $280k

    Sesame, located in Bellevue, Washington, is seeking a talented engineer to join our team focused on revolutionizing the way computers...  ...models. Candidates should have significant systems programming and performance engineering experience. You will enjoy benefits like unlimited... 
    Performance

    SESAME

    Bellevue, WA
    3 days ago
  • $175k - $280k

     ...variety of LLM, speech, and vision models. Partner with ML infrastructure and training engineers to build a fast, cost‑effective, accurate, and...  ...SGLang to take advantage of the latest techniques in high‑performance model serving. Work with the training team to identify... 
    Performance
    Contract work
    Flexible hours

    SESAME

    Bellevue, WA
    4 days ago
  • $148.5k - $225k

     ...is looking for a Senior Machine Learning Engineer to help launch various innovative ads-offerings...  ..., campaign optimization, ads channel performance, ads performance maximization and wide...  ...applications.Publish research papers in leading ML/AI/Advertising conferences solving... 
    Performance
    Local area
    Flexible hours

    Chewy

    Bellevue, WA
    3 days ago
  • $117k - $152k

     ...opportunityUnity Vector builds an offline ML platform that powers insight,...  ...systems.We’re looking for a Machine Learning Engineer to join our Offline Infrastructure team....  ...experimentation and model iterationHelp optimize performance and efficiency across data processing and... 
    Performance
    Full time
    Work at office
    Remote work
    Worldwide

    Unity Technologies

    Bellevue, WA
    16 hours ago
  • $104k - $208k

     ...prefer prior experience as a software developer, coder, software engineer, or programmer. We only consider applicants located in the...  ...rates, with opportunities for higher earnings based on performance. We operate remotely for qualified candidates in the US, Canada... 
    Performance
    Hourly pay
    Full time
    Remote work
    Flexible hours

    DataAnnotation

    Redmond, WA
    20 days ago
  •  ...Plutus is seeking a founding machine learning engineer to scale from v1 to a billion under management by year-end. You will design and deploy...  ...in ambiguity at an early-stage startup. You should have strong ML fundamentals, experience taking models to production, and a... 

    Plutus

    Kirkland, WA
    21 hours ago
  • $200.4k - $260.5k

    Senior Machine Learning Engineer, Data InfrastructureUnity Vector builds an Data platform that...  ...workflows that power production ML systems.To support this growth, we need strong...  ...ensure the reliability, scalability, and performance of our data platform.What You’ll Do Develop... 
    Performance
    Full time
    Work at office
    Worldwide
    Relocation package

    Unity Technologies

    Bellevue, WA
    2 days ago
  •  ...Electronic Arts seeks a Senior Machine Learning Engineer to design and operate production-grade data and ML infrastructure for fraud, anti-cheat, and account security across EA games. The role reports to the Senior Manager, EA Player Security Data Labs and follows a hybrid... 

    Electronic Arts

    Kirkland, WA
    1 day ago
  •  ...The role emphasizes C++ and Python development, automated test creation, and collaboration across Windows, Linux, and embedded targets. Applicants should have 3+ years in software development and BS in Engineering/CS, with a strong QA mindset. #J-18808-Ljbffr Nintendo

    Nintendo

    Redmond, WA
    3 days ago
  •  ...WiproContact: Meghana GorusuCompany: SRI Tech SolutionsJob RoleAgentic AI/ML EngineerLocationRedmond, WAPrimary SkillsMachine Learning,...  ...pipelines, and operational excellence Quality & Integration Engineering - focusing on agent behavior testing, UI/UX integration, and... 
    Full time

    SRI Tech

    Redmond, WA
    2 days ago
  •  ...isn't just encouraged it's expected. The Role As one of our AI ML Engineer’s, you'll be a key technical leader and thought leader, shaping our ML strategy and building intelligent, high-performance multi-agent systems that perceive, learn, and act in real time.... 
    Performance
    Full time
    Shift work

    C-serv

    Bellevue, WA
    16 hours ago
  •  ..., and Google Cloud. We are a collaborative team of architects, engineers, financial analysts, and delivery experts who partner with world...  .... If you are hired by Accenture and require accommodation to perform the essential functions of your role, you will be asked to participate... 
    Full time
    Work experience placement
    Live in
    Work at office
    Local area

    Accenture

    Redmond, WA
    4 days ago
  • $119.8k - $234.7k

     ...toward interpretability and causal validity over purely predictive performance.Balance methodological rigor with pragmatism, selecting...  ...results) OR equivalent experience. 5+ years’ experience building ML models. 5+ years’ experience writing SQL to analyze data. 5+ years... 
    Performance
    Ongoing contract
    Local area
    3 days per week

    Microsoft

    Redmond, WA
    1 day ago
  • $142.8k - $274.8k

     ...as a horizontal analytics team supporting multiple product, engineering, marketplace, monetization, and business teams across the ecosystem...  ..., content consumption, marketplace dynamics, and revenue performance to identify growth opportunities and business risks.Develop... 
    Performance
    Ongoing contract
    Work at office
    Local area

    Microsoft

    Redmond, WA
    1 day ago
  • $142.8k - $274.8k

     ...and the need to deliver trustworthy, high-performing models, Microsoft's Commercial Business...  ...to partner directly with the engineering and product management groups responsible...  ...Intelligence (AI) and Machine Learning (ML) landscape. Drive meaningful impact for... 
    Performance
    Ongoing contract
    Local area
    Worldwide
    3 days per week

    Microsoft

    Redmond, WA
    2 days ago
  • $102.1k - $202.2k

     ...EngineeringDiscipline: Data EngineeringCompany: MicrosoftOverviewCommercial Engineering & AI (CEAI) partners closely with stakeholders to accelerate...  ....Monitor platform availability, capacity utilization, performance, and service health.Drive incident management,... 
    Performance
    Ongoing contract
    Local area
    3 days per week

    Microsoft

    Redmond, WA
    16 hours ago
  • $100k - $150k

     ...MLOps Engineer - Remote    Bright Vision Technologies is a technology consulting and...  ...Engineer to design, build, and operate high-performance, highly reliable inference platforms for...  ..., throughput, cost, and quality in ML serving.  Key Responsibilities  Design... 
    Performance
    Full time
    H1b
    Local area
    Immediate start
    Remote work
    Visa sponsorship

    Bright Vision Technologies

    Redmond, WA
    2 days ago
  • itD is seeking a Data Engineer (AI/ML) III to drive the development of machine learning solutions and scalable data processing pipelines that enhance AR display systems. The role requires expertise in ML, computer vision, image and signal processing, and large-scale data... 
    Remote work

    itD Tech

    Redmond, WA
    4 days ago
  • $141k - $202k

     ...A leading tech company in Kirkland is seeking a software engineer to develop next-generation technologies that transform user interactions...  ...will have expertise in Python or C++, along with experience in ML infrastructure and Generative AI concepts. The role offers an annual... 

    Google

    Kirkland, WA
    1 day ago
  • $122.3k - $158.5k

     ...strengthens system integrity, supports fair play, and enables teams to build and operate securely at scale. The Senior Machine Learning Engineer will report to the Senior Manager, EA Player Security Data Labs. You will follow a hybrid work model with a mix of remote work and... 
    Full time
    Work at office
    Local area
    Remote work

    Electronic Arts

    Kirkland, WA
    1 day ago
  •  ...Chewy is seeking a Senior Machine Learning Engineer for our Bellevue, WA team to advance sponsored ads offerings across onsite and offsite channels. You will shape model development, ranking, bidding dynamics, and campaign optimization, collaborating with product and... 

    Chewy

    Bellevue, WA
    21 hours ago
  •  ...ultra-wealthy. We launched in January, AUM is already in the tens of millions, and we're now looking for a founding machine learning engineer to help take us from v1 to a billion under management by end of year. We have an enormous amount of customer and portfolio data... 
    Live in
    Weekend work

    Plutus

    Kirkland, WA
    21 hours ago
  • $264.1k - $350k

    AWS Open Data Analytics Engines is a suite of fully managed, high-performance analytics services built on popular open-source frameworks, enabling customers to process, analyze, and derive insights from massive datasets with speed, flexibility, and cost efficiency. Open... 
    Performance
    Work experience placement
    Flexible hours

    Amazon

    Redmond, WA
    1 day ago
  •  ...Google is seeking a Software Engineering Manager II to lead a team focused on AI/ML for Google Cloud Compute. You will influence stakeholders and drive the decision-making process while owning project goals and contributing to product strategy. Ideal candidates should... 
    Full time

    Google

    Kirkland, WA
    21 hours ago
  • Overview This role involves robotics research in robot learning, specifically training AI models that enable robots to perform various physical tasks by collecting and curating data through teleoperation. The team conducts research in this area and the role includes working... 
    Performance
    Full time
    Flexible hours

    TALENT Software Services

    Redmond, WA
    2 days ago
  • $129.2k - $174.8k

     ...communities around the world. As a systems development engineer, you will work on a variety of projects and...  ...patterns and issues for RF links in space- Apply AI/ML techniques for detecting RF and PHY performance anomalies from satellites on orbit - Work with communication... 
    Performance
    Permanent employment
    Remote work
    Flexible hours

    Amazon

    Redmond, WA
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to ML Performance Engineer. Be the first to apply!