Model Optimization Engineer
$150k - $175kBright Vision Technologies
Location: 100% Remote (U.S.)
Position Type: Full-time, Direct W2
Salary Range: $150,000–$175,000 Annually
Experience Required: 6+ years Sponsorship: U.S. Citizens, Green Card Holders, EAD Holders, and H-1B transfer candidates are encouraged to apply. We are unable to sponsor new H-1B visa petitions for this position. Job Summary:
We are seeking an Model Optimization Engineer to focus on extracting maximum throughput, minimizing latency, and reducing cost across training and inference workloads for large neural network systems. The role spans the full stack from low-level kernel optimization to distributed system tuning, requiring deep understanding of GPU architecture, model parallelism, memory management, and compiler-level optimization. The ideal candidate has demonstrated impact on production AI workloads, with strong instrumentation and measurement discipline that enables rigorous, data-driven optimization decisions. In this role you will work closely with cross-functional partners — product, design, engineering, operations, and business stakeholders — to translate ambiguous requirements into well-engineered solutions, and will be expected to raise the bar through code review, design review, and mentorship of more junior engineers. The successful candidate brings strong engineering discipline, a clear communication style, and a track record of shipping meaningful work that holds up well in production. Required Qualifications
- Bachelor’s or Master’s degree in Computer Science, Computer Engineering, or a related field.
- Six or more years of experience in performance engineering, ML systems, or HPC.
- Strong proficiency in Python and C++.
- Hands-on experience optimizing deep learning workloads on modern GPUs.
- Deep understanding of distributed training and inference techniques.
- Experience with profiling tools across CPU, GPU, and distributed systems.
- Familiarity with model compression techniques and their accuracy implications.
- Strong grasp of memory hierarchies, communication primitives, and parallelism strategies.
- Excellent measurement, debugging, and analytical reasoning skills.
- Strong communication and collaboration skills.
- Experience optimizing LLM inference at production scale.
- Contributions to vLLM, TensorRT-LLM, DeepSpeed, or similar projects.
- Familiarity with custom kernel authoring in Triton or CUTLASS.
- Experience with FinOps for AI workloads.
- Publications or talks on AI systems performance.
Equal Employment Opportunity (EEO) Statement
Bright Vision Technologies (BV Teck) is committed to equal employment opportunity (EEO) for all employees and applicants without regard to race, color, religion, sex, sexual orientation, gender identity or expression, national origin, age, genetic information, disability, veteran status, or any other protected status as defined by applicable federal, state, or local laws. This commitment extends to all aspects of employment, including recruitment, hiring, training, compensation, promotion, transfer, leaves of absence, termination, layoffs, and recall.
BV Teck expressly prohibits any form of workplace harassment or discrimination. Any improper interference with employees' ability to perform their job duties may result in disciplinary action up to and including termination of employment.
$158.4k - $237.6k
Company Qualcomm Technologies, Inc. Job Area Engineering Group, Engineering Group > Machine Learning Engineering... ...efficient inference of large-scale foundation models. We are seeking a Staff Engineer - AI Model Optimization Architect to lead end-to-end model...SuggestedWork experience placement- Qualcomm Technologies, Inc. seeks a Staff Engineer to lead end-to-end AI model optimization for LLMs, VLMs, and diffusion on Qualcomm accelerators. You will transform PyTorch models, manage KVcache behavior, and drive deployment with PyTorch, ONNX, and torch.compile across...Suggested
$200k - $230k
...tremendous career growth potential. Job Title: Foundation Model Engineer Location: 100% Remote (U.S.) Position Type: Full-time,... .... ~ Experience with RLHF, DPO, or other preference optimization techniques. ~ Strong understanding of evaluation methodology...SuggestedFull timeH1bLocal areaImmediate startRemote workVisa sponsorship- ...7 Quarter Students: June 21, 2027 - December 10, 2027 WHAT YOU WILL BE DOING: We are seeking highly motivated AI Model Optimization & Software Engineer Interns/Co-op to join our teams. We are recruiting for multiple opportunities across AI model optimization, framework...SuggestedFull timeSummer workInternshipSummer internshipWorldwide
- AMD in Austin, TX is seeking a Fellow to define and drive end-to-end software optimization strategy for top-tier customers. You will shape GPU kernel performance, contribute to ROCm core software, and collaborate across ML workloads and AI frameworks to deliver high-performance...Suggested
- RWE Americas, LLC is seeking a Sr BESS Optimization Engineer to lead technical solutions for utility-scale energy storage projects with a project development mindset. The role focuses on state-of-the-art technology and design features to fulfill dispatch and evolving market...
$140k - $150k
...full time, permanent Functional area: Engineering; Innovation / R&D; Project Development... ...strong candidate for our BESS Technology Optimization team as the Sr BESS Optimization Engineer... ...strategy, and thermal performance modeling Mentor junior engineers and establish...Permanent employmentFull timeTemporary workWork experience placementWork at officeLocal areaImmediate startWork visaFlexible hours- ...diverse team performing fast-paced investigations to empower our engineers to develop at the speed of light. Participate in the full life... ...with other team members and chip engineers to understand and optimize how workflows use the compute and storage environments. Build...
- ...How does GridBeyond do this? We do this by living our Mission, leveraging AI to innovate and collaborate with our customers creating optimal value from renewable energy generation, demand and storage to deliver a zero-carbon future. What’s the business and culture like?...Temporary workLive inWork at officeNight shiftAfternoon shift
$184k - $287.5k
Sr. Inference Engineer, GPU Kernel Optimization We're now looking for a Sr. Inference Engineer for GPU Kernel Optimization! What does it take to push... ...silicon-measured kernel benchmarking infrastructure, model-level performance projection tooling, and agentic optimization...- ...highly skilled Fellow to define and lead end-to-end software optimization for premier GPU workloads. You will shape kernel performance,... ..., and framework teams, with a strong focus on transformer-era models and cutting-edge ML techniques. #J-18808-Ljbffr Advanced Micro...
$195.2k - $292.8k
Qualcomm is seeking a GPU Engineer to work in Austin with a focus on optimizing GPU cores and supporting driver and compiler development. Candidates should possess a Bachelor’s degree in Computer Engineering or related fields with at least 6 years of experience. Preferred...Remote job- General Motors, through Embodied AI, seeks a Senior Engineer to measure and visualize AV model performance. You will design and implement evaluation workflows, collaborate across Data, Infra and validation teams, and connect offline evaluation with real-world behavior....Remote job
- ...US NOC, helping monetize flexibility and generation across energy markets and ancillary services. You will work with the market optimization team to analyze dynamics and execute trading strategies. Based in Austin, this Mon-Fri role covers evening peaks and offers exposure...Night shiftAfternoon shift
- USF in Austin is looking for a skilled Process Engineer to optimize manufacturing processes and ensure high-quality production. This hands-on role involves troubleshooting complex issues and training team members while adhering to safety and efficiency standards. The ideal...
- US Farathane is seeking a skilled Process Engineer to join our team in Austin, Texas. In this hands-on role, you will be responsible for optimizing and controlling manufacturing processes to produce high-quality parts efficiently, balancing performance, cycle time, cost...
- ABM Industries seeks a Building Engineer II to independently operate, monitor, maintain, troubleshoot, and optimize commercial building systems, including HVAC, electrical, plumbing, and BAS, ensuring tenant comfort and energy efficiency. The role collaborates with vendors...
- Product Manager - AI Inference & Model Serving Full-time Job Summary Mirantis is looking... ..., distributed systems, and performance engineering. You will define how NeoClouds and... ..., and full-stack performance optimization. This person will define how customers...Full time
- Grow with us AI Model Architect - Silicon-Software Co-Design Austin, Texas The Voice... ...Machine. The Mission Most AI architects optimize models for hardware that already exists.... ...'t model fine-tuning. This isn't prompt engineering. This is deep, architecture-level...Contract workTemporary work
$87.4k - $253k
...at accenture.com. We Are: Accenture's Enterprise Operating Model practice. We partner with Boards, CEOs, and other C-Suite... ...an end-to-end partner from strategy to design to activation to optimization, guiding clients to deliver the most impactful programs of their...Full timeWork experience placementLive inWork at officeLocal area$122.7k - $317.2k
...at accenture.com. We Are: Accenture's Enterprise Operating Model practice. We partner with Boards, CEOs, and other C-Suite... ...an end-to-end partner from strategy to design to activation to optimization, guiding clients to deliver the most impactful programs of their...Full timeWork experience placementLive inWork at officeLocal area- Mirantis in Austin, Texas, is hiring a Product Manager to drive AI inference and model serving for k0rdent AI. The role requires a blend of technical expertise and commercial acumen to optimize performance in cloud-native infrastructure. Ideal candidates will have 7+...
$130.5k - $193.6k
...job applies security best practices to optimize systems, partners with teams to drive security... ...effectiveness, and collaborates with engineers for continuous improvements. Job... ...employees, PayPal's balanced hybrid work model offers 3 days in the office for effective...Full timeWork at officeLocal areaImmediate startFlexible hours- Flex is hiring an SMT Process Engineer in Austin, TX to join our manufacturing team. You will focus on improving cost, quality, and delivery on SMT lines and will work closely with production and maintenance to ensure stable line performance. The role reports to the Process...Flexible hours
- Tokyo Electron Austin is seeking a highly skilled process development engineer to advance TEL equipment processes at customer sites. You will develop and characterize ALD, CVD, PECVD, PE-ALD and related thin-film deposition processes for advanced logic and memory applications...
$140k - $145k
H R PUNDITS INC is seeking a Design Engineer III in Austin, Texas, to conduct thorough power analysis from RTL to GDSII, improve automation flows, and collaborate with the RTL design team. The ideal candidate should have a Bachelor's degree in Electrical Engineering or...- ...safety, and performance. Your role involves maintaining production efficiency, minimizing defects, and collaborating with engineering teams to optimize process reliability.ResponsibilitiesOperate, maintain, and troubleshoot high-precision plastic injection molding...Hourly payFull timeFlexible hours
- ...Title: Databricks Engineer Location: Hybrid (Austin, TX) (Location of job: Hybrid, 3... ...Databricks Engineer design, develop, and optimize scalable data solutions on Databricks, leveraging... ...and reporting. Develop robust data models and implement data quality, validation,...Remote work
- ...Job Description Insight Global is seeking an OT Engineer to join one of their top supply chain customers. This individual will be joining... ...in each environment, the differences between them and how to optimize the existing landscape. This person will also support network...Remote work
$195k - $255k
...Your Job The Senior LLM Engineer will fine-tune and deploy LLMs on Azure OpenAI... ...You Will Do Fine-tune foundation models (LoRA/QLoRA, RLHF/DPO, instruction tuning... ...tuned Azure OpenAI models. Deploy and optimize inference (quantization, KV-cache, vLLM/...Flexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Model Optimization Engineer. Be the first to apply!



