Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

AI Optimization Engineer

$100k
Full-time

Jobgether

This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for an AI Optimization Engineer based in United States.

This is a fully remote opportunity focused on improving the performance, scalability, and economics of large-scale AI systems.
You will optimize training and inference workloads across the stack, from low-level GPU kernels to distributed infrastructure.
The role combines systems engineering, performance analysis, machine learning infrastructure, and compiler-level optimization.
You’ll work with modern GPUs and large neural networks, using rigorous measurement and profiling to identify and resolve performance bottlenecks.
The position offers the opportunity to influence production AI workloads where improvements in throughput, latency, and cost have meaningful business impact.
You’ll collaborate closely with engineering, product, operations, and business teams while contributing to technical direction and engineering standards.
As a senior technical contributor, you’ll also mentor engineers and help drive a culture of measurable, production-ready optimization.

\n Accountabilities:
  • Optimize training and inference workloads to maximize throughput, minimize latency, and improve cost efficiency across large-scale neural network systems.
  • Analyze and improve performance across the full technology stack, including GPU kernels, memory management, communication, distributed systems, and model execution.
  • Profile CPU, GPU, and distributed workloads to identify bottlenecks and use quantitative analysis to guide optimization decisions.
  • Design and implement performance improvements using Python, C++, and relevant AI systems technologies.
  • Optimize distributed training and inference architectures, including model parallelism, communication strategies, and resource utilization.
  • Evaluate and implement model compression techniques while carefully considering their impact on model accuracy and production performance.
  • Investigate complex performance and reliability issues through systematic debugging, instrumentation, benchmarking, and root-cause analysis.
  • Contribute to production-scale optimization of large language model inference and other demanding AI workloads.
  • Develop and improve low-level optimization techniques, including custom GPU kernels where appropriate.
  • Collaborate with product, design, engineering, operations, and business stakeholders to translate ambiguous requirements into scalable, well-engineered technical solutions.
  • Participate in architecture and code reviews, establish engineering best practices, and contribute to long-term technical strategy.
  • Mentor junior and mid-level engineers, helping raise technical quality and strengthen performance engineering capabilities.
  • Identify opportunities to improve the cost structure of AI workloads through infrastructure optimization and FinOps-oriented analysis.

Requirements:

  • Bachelor’s or Master’s degree in Computer Science, Computer Engineering, or a related technical discipline.
  • 6+ years of professional experience in performance engineering, machine learning systems, high-performance computing, or a closely related field.
  • Strong programming proficiency in Python and C++ , with the ability to develop production-quality, maintainable code.
  • Hands-on experience optimizing deep learning workloads on modern GPU architectures.
  • Deep understanding of distributed training and inference techniques, including parallelism strategies and communication primitives.
  • Strong knowledge of memory hierarchies, GPU/CPU performance characteristics, and systems-level optimization.
  • Experience using profiling and instrumentation tools across CPU, GPU, and distributed environments.
  • Familiarity with model compression methods and their implications for accuracy, performance, and production deployment.
  • Excellent measurement, debugging, analytical reasoning, and problem-solving abilities.
  • Strong communication and collaboration skills, with the ability to explain complex technical concepts to cross-functional stakeholders.
  • Demonstrated ability to work independently, make data-driven technical decisions, and deliver meaningful improvements in production environments.
  • Experience with production-scale LLM inference is strongly preferred.
  • Contributions to projects such as vLLM, TensorRT-LLM, DeepSpeed, or comparable AI systems projects are a plus.
  • Experience with custom kernel development using technologies such as Triton or CUTLASS is preferred.
  • Familiarity with FinOps and cost optimization for AI workloads is advantageous.
  • Publications, conference presentations, or technical talks focused on AI systems or performance engineering are a plus.
  • Must be currently based in the United States and authorized to work in the U.S.; U.S. citizens, permanent residents, EAD holders, and candidates eligible for H-1B transfer are encouraged to apply. New H-1B sponsorship is not available.

Benefits:

  • $100,000 annual salary for this full-time direct W2 position.
  • 100% remote work within the United States.
  • Opportunity to work on challenging AI optimization and high-performance computing problems.
  • Exposure to large-scale neural networks, modern GPU architectures, distributed systems, and production AI infrastructure.
  • Significant opportunities for technical ownership, mentorship, and career growth.
  • Collaborative environment spanning engineering, product, operations, design, and business teams.
  • Opportunity to contribute to impactful production AI systems and advance performance, scalability, and cost efficiency.
  • Equal employment opportunity and an inclusive workplace committed to fair treatment of employees and applicants.
\n

How Jobgether works:

We use an AI-powered matching process to ensure your application is reviewed quickly, objectively, and fairly against the role's core requirements. Our system identifies the top-fitting candidates, and this shortlist is then shared directly with the hiring company. The final decision and next steps (interviews, assessments) are managed by their internal team.

We appreciate your interest and wish you the best!

Why Apply Through Jobgether?

Data Privacy Notice: By submitting your application, you acknowledge that Jobgether will process your personal data to evaluate your candidacy and share relevant information with the hiring employer. This processing is based on legitimate interest and pre-contractual measures under applicable data protection laws (including GDPR). You may exercise your rights (access, rectification, erasure, objection) at any time.

#LI-CL1

Vacancy posted 6 days ago
Similar jobs that could be interesting for youBased on the AI Optimization Engineer in Remote vacancy
  • $100k

    Role Description We are seeking an AI Optimization Engineer to focus on extracting maximum throughput, minimizing latency, and reducing cost across training and inference workloads for large neural network systems. The role spans the full stack from low-level kernel optimization... 
    Suggested
    Full time
    Local area
    Immediate start

    Bright Vision Technologies

    Remote
    8 hours ago
  • $180k - $250k

     ...the first and best place to understand and optimize digital customer experiences. Our...  ...technology, business, and operations teams AI-powered insights into any issues impacting...  ...not a research role, and it is not prompt engineering; the team builds production systems that... 
    Suggested
    Full time

    Conviva

    Remote
    1 day ago
  • $100k

     ...position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for an AI Optimization Engineer based in United States. This is a fully remote opportunity focused on improving the performance, scalability, and... 
    Suggested
    Permanent employment
    Full time
    H1b
    Remote work

    jobgether

    United States
    7 days ago
  • $100k

     ...position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for an AI Optimization Engineer based in United States. This is a fully remote opportunity focused on improving the performance, scalability, and... 
    Suggested
    Remote job
    Permanent employment
    Full time
    H1b

    jobgether

    United States
    6 days ago
  • $115k - $200k

     ...Forward Deployed AI Engineer (ML / Full Stack) Compensation: $115K–$200K base (varies by level and interview performance) + equity...  ...with customer data backends and integrations. Troubleshoot, optimize, and scale AI-driven systems in production. Contribute to scripting... 
    Suggested
    Full time
    Work at office
    Visa sponsorship

    Jenn Nguyen And Friends

    Remote
    1 day ago
  •  ...We are rebuilding biotech for the AI era. When a breakthrough is delayed, the world...  .... Benchling is building Intelligence Engineering & Enablement, a small autonomous team within...  ...~ Demonstrated understanding of how to optimize workloads across deterministic and non-... 
    Remote job
    Full time
    Local area

    Benchling

    United States
    1 day ago
  •  ...Partner delivering cloud and Generative AI solutions for SMB, startup, and enterprise...  ...Overview We’re looking for a GenAI Engineer to join our Professional Services team,...  ...observability, logging, and performance optimization for AI workloads · Collaborate with clients... 
    Remote job
    Full time

    Innovative Solutions

    Remote
    1 day ago
  • $135k - $200k

     ...missing children, and more. The Role Forward Deployed AI Engineers work directly with customers owning Gen AI strategy and...  ...Paying attention to the needs of our community enables us to optimize our opportunities to grow and helps ensure many pathways to success... 
    Full time
    Work experience placement
    Work at office
    Remote work
    Work from home
    Relocation package
    Flexible hours

    Palantir Technologies

    New York, NY
    1 day ago
  • $114k - $148k

     ...AI Engineer   Location :  Remote, United States  Employment   Type : Full-Time  Benefits   Offered : Vision, Medical, Life,...  ...Implement robust AI systems and functionalities, ensuring optimal performance and scalability. Troubleshoot and debug issues... 
    Remote job
    Full time
    Temporary work
    Work experience placement
    Work at office

    Onestream

    Remote
    1 day ago
  •  ...Forward Deployed AI Engineer AI Foundry | NewRocket Location: Remote with travel (~25%) Reports to: Global AI Center of Excellence...  ...AI solutions are secure, scalable, and production-ready. Optimize deployed systems for performance and reliability. Cross-Team... 
    Remote job
    Full time
    Immediate start

    Newrocket

    United States
    1 day ago
  •  ...Patlytics:   Patlytics is the fastest-growing AI-native patent intelligence platform,...  .... Our advanced LLMs and generative AI engine—custom-built for intellectual property—power...  ...vector search, BM25 hybrids, re-ranking) optimized for patent corpus characteristics, long... 
    Full time
    Immediate start
    Remote work

    Patlytics

    New York, NY
    1 day ago
  •  ...solutions. The Role We are hiring our first dedicated AI Engineer to do the same thing internally: bring our marketing, brand,...  ...it in code Headless CMS, edge rendering, and performance optimization at scale Comfortable inheriting an existing codebase and making... 
    Permanent employment
    Full time
    Worldwide

    Energyx

    Remote
    1 day ago
  • $110k - $140k

     ...developing, and deploying production-grade AI solutions including autonomous agents,...  ...using techniques like LoRA and QLoRA to optimize performance for domain-specific use cases...  ...in code reviews and maintain high-quality engineering standards. Keep updated with advances... 
    Full time
    Work from home

    Confie

    Addison, TX
    1 day ago
  •  ...behavior biometrics, machine learning, and AI to stop fraud before it happens. Today,...  ...payment fraud, account takeovers, and social engineering scams. We have raised $145M from world-...  ...evaluation frameworks to continuously optimize AI agent performance, ensuring high accuracy... 
    Full time
    Remote work
    Worldwide
    Home office
    Flexible hours

    Sardine

    United States
    1 day ago
  • $180k - $280k

     ...AI Engineer Title of Role: AI Engineer Location: New York, hybrid Company Stage of Funding: Venture-Backed — Healthcare, Fintech...  ...datasets to derive insights and improve model performance. Optimize existing algorithms for better accuracy and efficiency in... 
    Full time
    Work at office

    Recruiting From Scratch

    Remote
    1 day ago
  •  ...Figure is an AI robotics company developing autonomous general-purpose humanoid robots...  ...autonomy. We are looking for a Helix AI Engineer, Pretraining to build large-scale...  ...dynamics for frontier models Build and optimize large-scale distributed training pipelines... 
    Full time
    Work at office

    Figure

    Remote
    1 day ago
  • $171k - $240k

     ...and control spend effortlessly. Brex’s AI-native automation and world-class service...  ...to grow your career. AI at Brex AI Engineering at Brex is redefining how businesses run...  ...just surface insights—they take action, optimizing spend, managing workflows, and making... 
    Full time
    Work at office
    Remote work
    Work from home

    Brex Inc.

    San Francisco, CA
    1 day ago
  • $145k - $240k

     ...Founding AI Engineer Title of Role: Founding AI Engineer Location: San Francisco, hybrid Company Stage of Funding: Venture-Backed...  ...user experience. Implement machine learning algorithms to optimize insurance processes and decision-making. Collaborate with a... 
    Full time
    Work at office

    Recruiting From Scratch

    Remote
    1 day ago
  •  ...better decisions. Now, we are building an AI-native platform to redefine how teams...  ...performance infrastructure. As our first AI Engineer, you will play a pivotal role in...  ...task routing across multiple LLMs/Agents to optimize scalability, cost, latency, and throughput... 
    Full time
    Flexible hours

    Quantifi

    Remote
    1 day ago
  •  ...The AI Engineer plays a critical role in making modeling and simulation accessible to non-data scientists. You will be designing, developing...  ..., and government use Epistemix to reduce uncertainty, optimize decisions, and accelerate time to value. Whether estimating total... 
    Remote job
    Full time
    Flexible hours

    Epistemix

    United States
    1 day ago
  • $200k - $350k

     ...Figure is an AI Robotics company developing a general purpose humanoid. Our humanoid...  ...Our Helix team is looking for Perception Engineers to empower Figure humanoid robots to perform...  ...cross-functionally to evolve and optimize our autonomy stack. Requirements:... 
    Full time
    Work at office

    Figure

    Remote
    1 day ago
  •  ...strong background in machine learning, artificial intelligence, and data science, the AI engineer will design, develop, and deploy AI solutions to address business challenges and optimize operations. Work closely with cross-functional teams to integrate AI models into... 
    Full time
    Work experience placement
    Work at office

    Luck Companies

    Remote
    1 day ago
  •  ...how Socure builds, deploys, and scales AI-driven identity solutions while enabling...  ...Reporting to the Head of New Product Engineering, you'll join a new Internal AI Engineering...  ...assessment, failure-mode analysis, and optimization of speed and accuracy of specific agentic... 
    Full time

    Socure

    Remote
    1 day ago
  • $195k - $230k

     ...About Elicit Elicit is an AI research assistant that uses language models to help researchers...  ...on our mission. What is an "AI Engineer"? AI engineering is a new category of...  ...+ equity, depending on your level. We're optimizing for a hire who can contribute at a L4/... 
    Full time
    Temporary work
    Work at office
    Remote work
    Flexible hours

    Elicit

    Oakland, CA
    1 day ago
  •  ...and increasing revenue through cutting-edge tools like AI chatbots, customized funnel optimization, targeted paid advertising, and powerful email/SMS...  ...seeking a talented and experienced Machine Learning/AI Engineer with a specialized focus on OpenAI technologies. As... 
    Full time

    Sales Hub Careers

    Chino Hills, CA
    1 day ago
  • $187k - $253k

     ...Eigen Labs is building the coordination engine for a world run by humans and agents alike. AI is becoming abundant. Trust is not. We are building the...  ...Make agents cheap (cost-aware execution, performance optimization) Make agents useful in production (not demos -... 
    Remote job
    Full time
    Temporary work
    Flexible hours
    Shift work

    Eigen Labs

    Remote
    1 day ago
  •  ...Company Description We're an Industrial AI start-up founded by a Stanford professor...  ...relies on proprietary Explainable AI engines for SME-explainable insights from...  ...Bayesian estimation Mathematical optimization Decision and Control Experience... 
    Remote job
    Full time
    Contract work
    Part time
    For contractors
    Work at office

    Brain Gain Recruiting

    Palo Alto, CA
    1 day ago
  • $150k - $200k

     ...what mental healthcare looks like in the AI era. Last year, 1 in 10 teens...  ...strong opinions on what to build, raise the engineering bar, and help define what a thoughtful,...  ...accessibility, or frontend performance optimization Why Marble - Direct impact on a critical... 
    Full time
    Work at office
    3 days per week

    Marble Health

    Remote
    1 day ago
  •  ...chat, extend RAG based response repositories, and leverage AI to optimize workflows, processes, and drive system improvements. Continuously...  ...-context documents, and structured data. Champion prompt engineering best practices across internal teams including Operations,... 
    Full time
    Contract work
    Second job
    Work at office
    Local area

    Lifescience Logistics

    Remote
    1 day ago
  • Figure is an AI robotics company developing autonomous general-purpose humanoid robots...  .... We are looking for a Helix AI Engineer, Reinforcement Learning to develop learning...  ...understanding of RL fundamentals (e.g., policy optimization, value methods, model-based RL)... 
    Full time
    Work at office

    Figure

    Remote
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to AI Optimization Engineer. Be the first to apply!