AI Optimization Engineer
$100kJobgether
This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for an AI Optimization Engineer based in United States.
This is a fully remote opportunity focused on improving the performance, scalability, and economics of large-scale AI systems.
You will optimize training and inference workloads across the stack, from low-level GPU kernels to distributed infrastructure.
The role combines systems engineering, performance analysis, machine learning infrastructure, and compiler-level optimization.
You’ll work with modern GPUs and large neural networks, using rigorous measurement and profiling to identify and resolve performance bottlenecks.
The position offers the opportunity to influence production AI workloads where improvements in throughput, latency, and cost have meaningful business impact.
You’ll collaborate closely with engineering, product, operations, and business teams while contributing to technical direction and engineering standards.
As a senior technical contributor, you’ll also mentor engineers and help drive a culture of measurable, production-ready optimization.
- Optimize training and inference workloads to maximize throughput, minimize latency, and improve cost efficiency across large-scale neural network systems.
- Analyze and improve performance across the full technology stack, including GPU kernels, memory management, communication, distributed systems, and model execution.
- Profile CPU, GPU, and distributed workloads to identify bottlenecks and use quantitative analysis to guide optimization decisions.
- Design and implement performance improvements using Python, C++, and relevant AI systems technologies.
- Optimize distributed training and inference architectures, including model parallelism, communication strategies, and resource utilization.
- Evaluate and implement model compression techniques while carefully considering their impact on model accuracy and production performance.
- Investigate complex performance and reliability issues through systematic debugging, instrumentation, benchmarking, and root-cause analysis.
- Contribute to production-scale optimization of large language model inference and other demanding AI workloads.
- Develop and improve low-level optimization techniques, including custom GPU kernels where appropriate.
- Collaborate with product, design, engineering, operations, and business stakeholders to translate ambiguous requirements into scalable, well-engineered technical solutions.
- Participate in architecture and code reviews, establish engineering best practices, and contribute to long-term technical strategy.
- Mentor junior and mid-level engineers, helping raise technical quality and strengthen performance engineering capabilities.
- Identify opportunities to improve the cost structure of AI workloads through infrastructure optimization and FinOps-oriented analysis.
Requirements:
- Bachelor’s or Master’s degree in Computer Science, Computer Engineering, or a related technical discipline.
- 6+ years of professional experience in performance engineering, machine learning systems, high-performance computing, or a closely related field.
- Strong programming proficiency in Python and C++ , with the ability to develop production-quality, maintainable code.
- Hands-on experience optimizing deep learning workloads on modern GPU architectures.
- Deep understanding of distributed training and inference techniques, including parallelism strategies and communication primitives.
- Strong knowledge of memory hierarchies, GPU/CPU performance characteristics, and systems-level optimization.
- Experience using profiling and instrumentation tools across CPU, GPU, and distributed environments.
- Familiarity with model compression methods and their implications for accuracy, performance, and production deployment.
- Excellent measurement, debugging, analytical reasoning, and problem-solving abilities.
- Strong communication and collaboration skills, with the ability to explain complex technical concepts to cross-functional stakeholders.
- Demonstrated ability to work independently, make data-driven technical decisions, and deliver meaningful improvements in production environments.
- Experience with production-scale LLM inference is strongly preferred.
- Contributions to projects such as vLLM, TensorRT-LLM, DeepSpeed, or comparable AI systems projects are a plus.
- Experience with custom kernel development using technologies such as Triton or CUTLASS is preferred.
- Familiarity with FinOps and cost optimization for AI workloads is advantageous.
- Publications, conference presentations, or technical talks focused on AI systems or performance engineering are a plus.
- Must be currently based in the United States and authorized to work in the U.S.; U.S. citizens, permanent residents, EAD holders, and candidates eligible for H-1B transfer are encouraged to apply. New H-1B sponsorship is not available.
Benefits:
- $100,000 annual salary for this full-time direct W2 position.
- 100% remote work within the United States.
- Opportunity to work on challenging AI optimization and high-performance computing problems.
- Exposure to large-scale neural networks, modern GPU architectures, distributed systems, and production AI infrastructure.
- Significant opportunities for technical ownership, mentorship, and career growth.
- Collaborative environment spanning engineering, product, operations, design, and business teams.
- Opportunity to contribute to impactful production AI systems and advance performance, scalability, and cost efficiency.
- Equal employment opportunity and an inclusive workplace committed to fair treatment of employees and applicants.
How Jobgether works:
We use an AI-powered matching process to ensure your application is reviewed quickly, objectively, and fairly against the role's core requirements. Our system identifies the top-fitting candidates, and this shortlist is then shared directly with the hiring company. The final decision and next steps (interviews, assessments) are managed by their internal team.
We appreciate your interest and wish you the best!
Why Apply Through Jobgether?
Data Privacy Notice: By submitting your application, you acknowledge that Jobgether will process your personal data to evaluate your candidacy and share relevant information with the hiring employer. This processing is based on legitimate interest and pre-contractual measures under applicable data protection laws (including GDPR). You may exercise your rights (access, rectification, erasure, objection) at any time.
#LI-CL1
$100k
Role Description We are seeking an AI Optimization Engineer to focus on extracting maximum throughput, minimizing latency, and reducing cost across training and inference workloads for large neural network systems. The role spans the full stack from low-level kernel optimization...SuggestedFull timeLocal areaImmediate start$180k - $250k
...the first and best place to understand and optimize digital customer experiences. Our... ...technology, business, and operations teams AI-powered insights into any issues impacting... ...not a research role, and it is not prompt engineering; the team builds production systems that...SuggestedFull time$100k
...position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for an AI Optimization Engineer based in United States. This is a fully remote opportunity focused on improving the performance, scalability, and...SuggestedPermanent employmentFull timeH1bRemote work$100k
...position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for an AI Optimization Engineer based in United States. This is a fully remote opportunity focused on improving the performance, scalability, and...SuggestedRemote jobPermanent employmentFull timeH1b$115k - $200k
...Forward Deployed AI Engineer (ML / Full Stack) Compensation: $115K–$200K base (varies by level and interview performance) + equity... ...with customer data backends and integrations. Troubleshoot, optimize, and scale AI-driven systems in production. Contribute to scripting...SuggestedFull timeWork at officeVisa sponsorship- ...We are rebuilding biotech for the AI era. When a breakthrough is delayed, the world... .... Benchling is building Intelligence Engineering & Enablement, a small autonomous team within... ...~ Demonstrated understanding of how to optimize workloads across deterministic and non-...Remote jobFull timeLocal area
- ...Partner delivering cloud and Generative AI solutions for SMB, startup, and enterprise... ...Overview We’re looking for a GenAI Engineer to join our Professional Services team,... ...observability, logging, and performance optimization for AI workloads · Collaborate with clients...Remote jobFull time
$135k - $200k
...missing children, and more. The Role Forward Deployed AI Engineers work directly with customers owning Gen AI strategy and... ...Paying attention to the needs of our community enables us to optimize our opportunities to grow and helps ensure many pathways to success...Full timeWork experience placementWork at officeRemote workWork from homeRelocation packageFlexible hours$114k - $148k
...AI Engineer Location : Remote, United States Employment Type : Full-Time Benefits Offered : Vision, Medical, Life,... ...Implement robust AI systems and functionalities, ensuring optimal performance and scalability. Troubleshoot and debug issues...Remote jobFull timeTemporary workWork experience placementWork at office- ...Forward Deployed AI Engineer AI Foundry | NewRocket Location: Remote with travel (~25%) Reports to: Global AI Center of Excellence... ...AI solutions are secure, scalable, and production-ready. Optimize deployed systems for performance and reliability. Cross-Team...Remote jobFull timeImmediate start
- ...Patlytics: Patlytics is the fastest-growing AI-native patent intelligence platform,... .... Our advanced LLMs and generative AI engine—custom-built for intellectual property—power... ...vector search, BM25 hybrids, re-ranking) optimized for patent corpus characteristics, long...Full timeImmediate startRemote work
- ...solutions. The Role We are hiring our first dedicated AI Engineer to do the same thing internally: bring our marketing, brand,... ...it in code Headless CMS, edge rendering, and performance optimization at scale Comfortable inheriting an existing codebase and making...Permanent employmentFull timeWorldwide
$110k - $140k
...developing, and deploying production-grade AI solutions including autonomous agents,... ...using techniques like LoRA and QLoRA to optimize performance for domain-specific use cases... ...in code reviews and maintain high-quality engineering standards. Keep updated with advances...Full timeWork from home- ...behavior biometrics, machine learning, and AI to stop fraud before it happens. Today,... ...payment fraud, account takeovers, and social engineering scams. We have raised $145M from world-... ...evaluation frameworks to continuously optimize AI agent performance, ensuring high accuracy...Full timeRemote workWorldwideHome officeFlexible hours
$180k - $280k
...AI Engineer Title of Role: AI Engineer Location: New York, hybrid Company Stage of Funding: Venture-Backed — Healthcare, Fintech... ...datasets to derive insights and improve model performance. Optimize existing algorithms for better accuracy and efficiency in...Full timeWork at office- ...Figure is an AI robotics company developing autonomous general-purpose humanoid robots... ...autonomy. We are looking for a Helix AI Engineer, Pretraining to build large-scale... ...dynamics for frontier models Build and optimize large-scale distributed training pipelines...Full timeWork at office
$171k - $240k
...and control spend effortlessly. Brex’s AI-native automation and world-class service... ...to grow your career. AI at Brex AI Engineering at Brex is redefining how businesses run... ...just surface insights—they take action, optimizing spend, managing workflows, and making...Full timeWork at officeRemote workWork from home$145k - $240k
...Founding AI Engineer Title of Role: Founding AI Engineer Location: San Francisco, hybrid Company Stage of Funding: Venture-Backed... ...user experience. Implement machine learning algorithms to optimize insurance processes and decision-making. Collaborate with a...Full timeWork at office- ...better decisions. Now, we are building an AI-native platform to redefine how teams... ...performance infrastructure. As our first AI Engineer, you will play a pivotal role in... ...task routing across multiple LLMs/Agents to optimize scalability, cost, latency, and throughput...Full timeFlexible hours
- ...The AI Engineer plays a critical role in making modeling and simulation accessible to non-data scientists. You will be designing, developing... ..., and government use Epistemix to reduce uncertainty, optimize decisions, and accelerate time to value. Whether estimating total...Remote jobFull timeFlexible hours
$200k - $350k
...Figure is an AI Robotics company developing a general purpose humanoid. Our humanoid... ...Our Helix team is looking for Perception Engineers to empower Figure humanoid robots to perform... ...cross-functionally to evolve and optimize our autonomy stack. Requirements:...Full timeWork at office- ...strong background in machine learning, artificial intelligence, and data science, the AI engineer will design, develop, and deploy AI solutions to address business challenges and optimize operations. Work closely with cross-functional teams to integrate AI models into...Full timeWork experience placementWork at office
- ...how Socure builds, deploys, and scales AI-driven identity solutions while enabling... ...Reporting to the Head of New Product Engineering, you'll join a new Internal AI Engineering... ...assessment, failure-mode analysis, and optimization of speed and accuracy of specific agentic...Full time
$195k - $230k
...About Elicit Elicit is an AI research assistant that uses language models to help researchers... ...on our mission. What is an "AI Engineer"? AI engineering is a new category of... ...+ equity, depending on your level. We're optimizing for a hire who can contribute at a L4/...Full timeTemporary workWork at officeRemote workFlexible hours- ...and increasing revenue through cutting-edge tools like AI chatbots, customized funnel optimization, targeted paid advertising, and powerful email/SMS... ...seeking a talented and experienced Machine Learning/AI Engineer with a specialized focus on OpenAI technologies. As...Full time
$187k - $253k
...Eigen Labs is building the coordination engine for a world run by humans and agents alike. AI is becoming abundant. Trust is not. We are building the... ...Make agents cheap (cost-aware execution, performance optimization) Make agents useful in production (not demos -...Remote jobFull timeTemporary workFlexible hoursShift work- ...Company Description We're an Industrial AI start-up founded by a Stanford professor... ...relies on proprietary Explainable AI engines for SME-explainable insights from... ...Bayesian estimation Mathematical optimization Decision and Control Experience...Remote jobFull timeContract workPart timeFor contractorsWork at office
$150k - $200k
...what mental healthcare looks like in the AI era. Last year, 1 in 10 teens... ...strong opinions on what to build, raise the engineering bar, and help define what a thoughtful,... ...accessibility, or frontend performance optimization Why Marble - Direct impact on a critical...Full timeWork at office3 days per week- ...chat, extend RAG based response repositories, and leverage AI to optimize workflows, processes, and drive system improvements. Continuously... ...-context documents, and structured data. Champion prompt engineering best practices across internal teams including Operations,...Full timeContract workSecond jobWork at officeLocal area
- Figure is an AI robotics company developing autonomous general-purpose humanoid robots... .... We are looking for a Helix AI Engineer, Reinforcement Learning to develop learning... ...understanding of RL fundamentals (e.g., policy optimization, value methods, model-based RL)...Full timeWork at office
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to AI Optimization Engineer. Be the first to apply!




