Senior Software Engineer - Model Performance
$220k - $320kSOLANA FOUNDATION
Help us make inference blazingly fast. If you love squeezing every last drop of performance out of GPUs, diving deep into CUDA kernels, and turning optimization techniques into production systems, we'd love to meet you. About Inference.net Inference.net trains and hosts specialized language models for companies that need frontier-quality AI at a fraction of the cost. The models we train match GPT-5 accuracy but are smaller, faster, and up to 90% cheaper. Our platform handles everything end-to-end: distillation, training, evaluation, and planet-scale hosting. We are a well‑funded ten‑person team of engineers who work in‑person in downtown San Francisco on difficult, high‑impact engineering problems. Everyone on the team has been writing code for over 10 years, and has founded and run their own software companies. We are high‑agency, adaptable, and collaborative. We value creativity alongside technical prowess and humility. We work hard, and deeply enjoy the work that we do. Most of us are in the office 4 days a week in SF; hybrid works for Bay Area candidates. About the Role You will be responsible for making our inference stack as fast and efficient as possible. Your work spans from implementing known optimization techniques to experimenting with novel approaches, always with the goal of serving models faster and cheaper at scale. Your north star is inference performance: latency, throughput, cost efficiency, and how quickly we can bring new model architectures into production. You'll work across the full inference stack—from CUDA kernels to serving frameworks—to find and eliminate bottlenecks. This role reports directly to the founding team. You’ll have the autonomy, a large compute budget, and technical support to push the limits of what’s possible in model serving. Key Responsibilities Implement and productionize optimization techniques including quantization, speculative decoding, KV cache optimization, continuous batching, and LoRA serving Deep dive into inference frameworks (vLLM, SGLang, TensorRT‑LLM) and underlying libraries to debug and improve performance Profile and optimize CUDA kernels and GPU utilization across our serving infrastructure Add support for new model architectures, ensuring they meet our performance standards before going to production Experiment with novel inference techniques and bring successful approaches into production Build tooling and benchmarks to measure and track inference performance across our fleet Collaborate with applied ML engineers to ensure trained models can be served efficiently Requirements 2+ years of experience in ML systems, inference optimization, or GPU programming Strong proficiency in Python and familiarity with C++ Hands‑on experience with LLM inference frameworks (vLLM, SGLang, TensorRT‑LLM, or similar) Deep understanding of GPU architecture and experience profiling GPU workloads Familiarity with LLM optimization techniques (quantization, speculative decoding, continuous batching, KV cache management) Experience with PyTorch and understanding of how models execute on hardware Track record of measurably improving system performance Nice‑to‑Have Experience with CUDA programming Familiarity with serving non-LLM models (TTS, vision, embeddings) Experience with distributed inference and multi‑GPU serving Contributions to open‑source inference frameworks Experience with Docker and Kubernetes You don't need to tick every box. Curiosity and the ability to learn quickly matter more. Compensation We offer competitive compensation, equity in a high‑growth startup, and comprehensive benefits. The base salary range for this role is $220,000 - $320,000, plus equity and benefits, depending on experience. Equal Opportunity Inference.net is an equal opportunity employer. We welcome applicants from all backgrounds and don't discriminate based on race, color, religion, gender, sexual orientation, national origin, genetics, disability, age, or veteran status. If you're excited about making AI inference faster for everyone, we'd love to hear from you. Please send your resume and GitHub to View email address on click.appcast.io and/or apply here on Ashby. #J-18808-Ljbffr
$166k - $225k
...to improve their business. Databricks’ Model Serving product provides enterprises with... ...strong SLAs and cost efficiency.As a Senior Engineer, you’ll play a critical role in shaping... ...architectural decisions and trade-offs to optimize performance, throughput, autoscaling, and...SeniorPerformanceLocal areaWorldwide$172.43k - $230.95k
...customers and partners advance their AI strategies, and be part of a high-performing team that believes in each other, come build with us at Crusoe.About This Role:The Senior Software Engineer for the AI Model Lifecycle team will play a crucial role in building a comprehensive...SeniorPerformanceTemporary work- ...leading data and AI company in San Francisco is seeking a Senior Engineer to enhance their Model Serving platform. This role requires expertise in... ...distributed systems and collaboration across teams to optimize performance and reliability. Ideal candidates will have a strong...SeniorPerformance
- ...AI company in San Francisco is seeking a Staff Engineer to design and implement systems for their AI/ML Model Serving platform. You will collaborate with product... ..., and research teams to ensure high-performance system delivery. The ideal candidate has over 10...SeniorPerformance
- ...access our start-of-the-art AI models, allowing them to do things... ...able to before. We focus on performant and efficient model inference... ...Role We are looking for an engineer who wants to take the world's... ...least 3 years of professional software engineering experience....PerformanceFull time
- ...About the Team We’re hiring software engineers to make OpenAI’s Model Performance teams more productive. These teams work on the systems, tooling, and infrastructure that help improve model performance across OpenAI’s training and inference workloads at frontier scale...PerformanceFull time
- ...at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently... .... Join us and help build the platform engineers turn to to ship AI products. THE ROLE: Baseten’s Model Performance (MP) team is responsible for ensuring the...PerformanceFull timeFlexible hours
- ...data, and run AI agents and models directly in their workflows.... ...therapeutics. As a full-stack engineer on the team, you’ll focus on... ...published, improving model performance and scalability, and enabling... ...QUALIFICATIONS ~3+ years of software engineering or equivalent research...PerformanceFull timeWork at officeLocal areaMonday to FridayShift work
$124k - $280k
...OpportunityAs a Strategy& Strategy Consulting - Business Model Reinvention - Senior Manager you will provide strategic guidance and insights... ...organizations, analyzing market trends and assessing business performance to develop recommendations that help clients achieve...SeniorPerformanceFull timeH1b$77k - $202k
...OpportunityAs a Strategy& - Strategy Consulting Business Model Reinvention - Senior Associate, you will provide strategic guidance and insights... ..., analyzing market trends and assessing business performance to develop recommendations that help clients achieve their...SeniorPerformanceFull timeH1b$298k - $368k
...diverse set of sensors, enabling engineers like you to (1) develop... ...real-world data, to (2) develop models and model training at scale,... ...architectures. Optimize model performance for on-device use cases (... ...Engage directly with research, software engineering, hardware...SeniorPerformanceFull timeRemote work$325k
...company in San Francisco seeks an engineer to optimize their powerful AI models for high-volume production... ...ideal candidate has over 5 years of software engineering experience, strong familiarity... ...with researchers and focus on performance optimization. Compensation ranges...SeniorPerformance- ...risk inherent to generative models. CTGT is the deterministic governance... ...on HaluEval, the CTGT Policy Engine (paired with GPT-120B OSS)... ...reliable, controllable, and performant in practice. Our mission... ...is the fundamentals of software engineering, applied at the level...SeniorPerformanceFull time
$160k - $250k
...next step is to speak to Jack . Senior Software Engineer Salary: $160K – $250K + Equity... ...role involves optimizing multimodal models for low latency, multilingual support... ...production codebase. Optimize system performance by centralizing inter-process...SeniorPerformanceFull time$180k - $300k
...Software Engineer, Consensus - Anza Who We Are Anza is a Solana R&D lab pushing the boundaries of blockchain performance and scalability. Anza was founded by experienced executives and core engineers solving the toughest problems in Web3. Crypto ecosystems rely on...SeniorPerformanceRemote jobFull timeWorldwideFlexible hours- ...About the role This isn't just another engineering role. This is a unique opportunity to... ...function and shape the future of performance and scalability at Persona. As our products... ...What you'll bring to Persona A strong software engineering background, demonstrated by...SeniorPerformanceFull timeTemporary workFor contractorsInternship
- ...Forest Labs is a leader in foundational AI models for image and video. Based in Freiburg... ...presence in San Francisco, we seek engineers who can bridge research breakthroughs and... ...rapidly. This role spans backend systems, GPU performance, and production ML serving, offering a...SeniorPerformance
$204k - $259k
...and collaborative group of software engineers, machine learning (ML) engineers... ...measure and enhance the performance of the Waymo Driver. We... ...achieve those goals by jointly modeling the real world, including... ...role, you will report to a Senior Staff Engineering Manager....SeniorPerformanceFull timeRemote work$165k - $210k
...ll play a critical role in shaping our engineering + broader company culture and help make... ...You have 8+ years of experience in a Software Engineering position. Experience with... ..., experience/qualifications, interview performance, market data, etc. Total compensation for...SeniorPerformanceFull time$119.77k - $140.9k
...Specifically, this position supports the Model Risk Management (“MRM”) program at the... ...framework, process/governance, ongoing performance, etc. The Analyst will generate reporting... ...stakeholders across multiple levels (e.g., senior leadership, teammates, etc.), with...SeniorPerformanceFull timeLocal area3 days per week- PLEASE CLICK HERE TO SEE *ALL* OF OUR JOB OPENINGS!Senior Software Engineer — Backend PerformanceAs a Senior Software Engineer on Backend Performance, you own the hottest paths in data products — the code that has to be fast because everything downstream depends on it....SeniorPerformance
- ...a roadmap brimming with opportunity. We're looking for a senior software engineer to not only amplify our development horsepower but also catalyze... ...deployment, and drive them to completion with a focus on performance, reliability, and security. Contribute to the evolution...SeniorPerformanceFull timeLocal areaRemote work
$195k - $225k
...MoM, and expanding the core engineering team in SF. The surface area... ...enterprise The Role As the Senior Software Engineer – Backend (Systems... ...deep ownership over APIs, performance, and data integrity — while... ...backend feature (e.g., model queue, caching layer, or realtime...SeniorPerformanceFull time- ...next generation of frontier models. By co-designing chips, systems... ...runtime within the inference engine that executes complex,... ...layers of the cluster serving software stack, translating demanding... ...silicon can deliver meaningful performance in production.In this role, you...Performance
$150k - $300k
...Senior Software Engineer Location: San Francisco, CA (required 5 days/week in-office) Type: Full-Time Compensation: $150,000–... ...workflows - Improve internal tooling, developer experience, and performance - Mentor junior engineers and help shape engineering...SeniorPerformanceFull timeH1bWork at office- ...agents. We’re on a mission to enable one billion humans to become software creators, and we are starting with a focus on data science and... ...from OpenAI and Google DeepMind. Building intuitive and performant user interfaces is critical to making Julius easy to use for...SeniorPerformanceFull time
$220k - $275k
...and after playing games. We are looking for a Senior Software Engineer specializing in Machine Learning to join our... ...a critical role in developing foundational ML models that enhance ad relevance, optimize performance, and drive revenue. This is a unique opportunity...SeniorPerformanceFull timeShift work$213k - $270k
...+ U.S. states. Waymo's Systems Engineering team works together to blend software and hardware systems in groundbreaking new ways. We set the high performance standards that ensure our vehicles... ...building or evaluating machine learning models in production ~ Experience...SeniorPerformanceFull timeRemote work$160k - $220k
...The next step is to speak to Jack . Job Title: Senior Software Engineer Salary: $160K - $220K + Equity Company Description... ...team where your code directly impacts autonomous aircraft performance in the most demanding environments on Earth. What you...SeniorPerformanceFull time$175k - $205k
...About Us At 3Y Health, we are building AI-driven software to empower healthcare providers and solve the overwhelming... ...VC. About the Role We are seeking a Frontend Engineer to help us craft intuitive, high-performance user interfaces that our providers love. The ideal...SeniorPerformanceFull timeWork experience placementPrivate practice
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Software Engineer - Model Performance. Be the first to apply!
- cybersecurity software engineer San Francisco, CA
- graduate software engineer San Francisco, CA
- software developer fintech San Francisco, CA
- new graduate software engineer San Francisco, CA
- senior robotics software engineer San Francisco, CA
- software engineer visa sponsorship San Francisco, CA
- software qa engineer San Francisco, CA
- network software engineer San Francisco, CA
- software engineer remote San Francisco, CA
- part time software developer San Francisco, CA


