Member of Technical Staff ML Data Infra
$250k - $350kNuance Labs, Inc.
Member of Technical Staff — Model Optimization and Inference (Experienced) Seattle, Washington About Nuance Labs Nuance Labs is building photorealistic, real-time AI avatars with emotional intelligence: a full-duplex audiovisual system that can listen, speak, react, interrupt, and respond like a real person. We're a research company, with PhDs from MIT, UW, Oxford, CMU, and Johns Hopkins, and industry experience from Apple, Meta, Amazon AGI, and more. Backed by Accel, Lightspeed, South Park Commons, and NVIDIA, we combine frontier research with ruthless engineering needed for consumer-grade, real-time systems. The team is small, the work is real, and the problems are unsolved. How Nuance Differentiates Most conversational AI avatars today are hacks — a face slapped on a speech-to-speech pipeline, stuck in the uncanny valley: emotionless, mechanical, one-turn-at-a-time. Current systems take 2–5 seconds to respond; natural conversation requires sub-500ms. That's a 10x improvement, and it demands rethinking the entire stack. That rethinking starts with full-duplex: an AI that listens and speaks simultaneously, perceives emotion in real time, and responds with a face that actually reflects it. It's an extremely hard problem, and we're developing foundation models designed for it from the ground up. About the Role We can train a great model. The next problem is making it fast enough to actually use in a real-time conversation — and that gap is enormous. A model that responds in 3 seconds is a demo. A model that responds in under 500ms is a product. We’re looking for someone who specializes in taking trained models and squeezing every last millisecond out of them. You understand the full stack from model weights to serving infrastructure — quantization, KV cache optimization, kernel-level acceleration, batching strategies — and you know which lever to pull for which problem. You’ve worked with vLLM, SGLang, or similar frameworks at scale and have strong opinions about where they fall short. This posting is aimed at experienced engineers and researchers who’ve operated at a senior to senior-staff level at big tech, a leading AI lab, or a high-traffic inference team. Everyone at Nuance is MTS — we don’t run title ladders — but we’re hiring people who have already done this work at scale. Our stack is more complex than a standard LLM deployment: we’re serving a full-duplex multimodal system that must satisfy strict real-time latency constraints. There’s a lot of unsolved optimization work here, and we need someone who finds that genuinely exciting. What You’ll Do Own end-to-end inference optimization across our model stack — LLMs, audio models, and diffusion-based components Implement and tune KV cache strategies for long-context conversations, including eviction policies, compression, and memory-efficient attention Evaluate, deploy, and extend inference serving frameworks (vLLM, SGLang, TensorRT-LLM, etc.) for our specific workloads Profile and benchmark end-to-end latency and throughput; identify and systematically eliminate bottlenecks Build internal tooling that makes optimization work faster and more rigorous — profiling viewers, end-to-end inference test harnesses, and other infrastructure that helps the team move quickly Accelerate diffusion model inference — consistency models, step distillation, caching strategies, and custom kernel optimizations Apply and develop quantization techniques (INT8, INT4, GPTQ, AWQ, and beyond) to reduce memory footprint and increase throughput without meaningfully degrading quality Work closely with research and infrastructure to ensure new models ship with optimized serving from day one What We’re Looking For Significant hands-on experience with LLM inference optimization — you’ve shipped work on KV caching, memory layout, attention kernels, or batching strategies in a production or high-traffic research context Proven proficiency with inference serving frameworks — vLLM, SGLang, TensorRT-LLM, or similar — including going well beyond default configurations and adapting them to non-standard workloads Experience optimizing diffusion model inference (latency reduction, step distillation, caching, or kernel-level work) Strong Python and PyTorch skills; comfort reading and writing CUDA or Triton kernels is a significant plus A systematic approach to profiling and optimization — you measure first, then optimize Familiarity with speculative decoding or other inference-time acceleration techniques Bonus Points Hands-on experience with post-training quantization (GPTQ, AWQ, or similar) and a clear sense of quality/performance tradeoffs Familiarity with multimodal or streaming inference architectures Experience deploying real-time AI systems with hard latency SLAs Prior work at an AI lab, inference startup, or on a high-traffic model serving platform Contributions to open-source inference frameworks Compensation $250,000 – $350,000 base salary, plus meaningful equity. We think long-term ownership matters and structure equity accordingly. Logistics Location: In-person in Seattle, five days a week — we believe in the compounding value of working shoulder-to-shoulder. Visa sponsorship: We sponsor visas (O-1, H-1B, green card, etc.) from day one. AI-native tooling: Do your best work with the best tools, including unlimited tokens. Health: We offer a variety of plans that meet your needs, including an HDHP with ~$2,000 in annual HSA contributions by the company (roughly 2x what most big tech companies put in). Wellness: $2,000/yr to spend on gym memberships, fitness equipment, training/coaching, classes, and more. Work setup: One-time $500 to spend on any tech or accessories that you might need to optimize your workflow. Time off: 15 days of PTO, 10 public holidays, and we close the office for a full week at year-end. Food: Lunch, drinks, and snacks on us every workday. We observe boba tea Tuesdays and Thursdays. Commuter benefits: Utilize pre-tax money (up to $340/month) for parking and transportation. 401(k): 4% match (100% of 1st 3% + 50% of next 2% contributions). Nuance Labs is an equal opportunity employer. We believe diverse teams build better AI. #J-18808-Ljbffr Nuance Labs, Inc.
$159.75k - $255.6k
Sr. Full Stack Member of Technical Staff Seattle, Washington, United States Join Axon and be a Force for... ...operate across the full stack, from data, models, and infrastructure to system integration... ...Python/C++ and experience with modern ML frameworks (e.g., PyTorch, TensorFlow)....DataWork at office$250k - $350k
Member of Technical Staff — RL Research (New PhD Grad) Seattle, Washington About Nuance Labs Nuance Labs... ...modeling, policy optimization, evaluation, data feedback loops, serving, observability,... ..., or in its final stretch — in ML, RL, or a related field, with research...DataInternshipH1bWork at officeVisa sponsorshipShift work- ...our products and our core platform, the data models that both work well for RL-at-... ...interactions with Lightspeed's products. As a Member of Technical Staff, you will work directly with our... ...them harmonize! Graduate degree in CS/ML or industry-experience equivalent: You'...DataShift work
- Building data-driven AI applications and agents is too complex, even for advanced developers... ...source projects. Experience with Apache infra, CNCF-stack, or cloud-native development.... ...Spice.ai OSS project 30-60 days - take technical and engineering ownership of an entire...Data
$180k
...modalities, while also incorporating audio where it enhances visual content (e.g., synchronized audio for video). Responsibilities span data curation, modeling, training, inference serving, and product integration, covering both pre‑training and post‑training phases. You...DataTemporary work$200k - $300k
Member of Technical Staff — Model Optimization and Inference (New Grad) Seattle, Washington About Nuance Labs Nuance Labs is building photorealistic... ...from day one What We’re Looking For BS, MS, or PhD in CS, ML, or a related field — completed or in the final stretch Strong...InternshipH1bWork at officeVisa sponsorship$162.7k - $220.2k
The AWS Worldwide Data & AI Strategy Team is focused on working... ...strategies into actionable technical plans using AWS Database, Analytics, AI/ML, and Generative AI services.As a member of this team, you will... ...employees, supervisors, and staff; adhere to standards of excellence...DataLocal areaWorldwideFlexible hours$40k - $80k
...development. Our GPU cloud bolsters technical capabilities and directly... ..., where every team member takes pride in their work and... ...alignment with IT security and data protection standards. Key Performance... ...Exposure to GPU infrastructure or AI/ML environments. Understanding...DataRemote workFlexible hours$220k - $405k
...Role The Connector Platform team builds the data layer that lets Perplexity's agents reach... ...works inside agent loops. Set the technical bar for connector reliability and operability... ...years for mid-level, more for senior and staff). Strong system design skills, with a...Data$89.52k - $134.28k
...Overview OlmoEarth is growing and seeks a Technical Partner Operations Specialist to... ...institutions. Conceptual familiarity with AI/ML (models, training data, fine‑tuning). Comfort scanning... ...work under deadlines. Benefits Team members and families are covered by medical,...DataWork experience placementWork at officeRemote workFlexible hours$220k - $405k
...Develop empathy for the nuances of our technical and business teams' work across disciplines... ...improvements. Collaborate closely with PM, Design, Data Science, Sales, and Enterprise customers... ...of the work of non-technical staff across enterprises, along with a curiosity...Data$117.2k - $176.7k
...incident response and operational investigations by collecting data, analyzing system behavior, and supporting remediation Partner... ...to collaborate and learn in a team environment ~ A related technical degree required Even Better If... Exposure to deploying...DataFull time- ...We apply modern capabilities, including AI/ML, cloud, cybersecurity, and IT... ..., and execution over bureaucracy. Title: Technical Support Analyst (Tier II) Location: Onsite... ...systems regularly to ensure reliability and data integrity Maintain accurate documentation...DataFull timeWork experience placementWork at officeFlexible hours
$180k
...dental coverage, access to a 401(k) retirement plan, short & long-term disability insurance, life insurance, and various other discounts and perks. SpaceXAI is an equal opportunity employer. For details on data processing, view our Recruitment Privacy Notice....DataTemporary work$168.1k - $227.4k
...of large-scale distributed systems, data pipelines, and applied ML/AI, building platforms that power metadata... ...Leadership & Architecture- Lead the technical direction for major features and... ...different tiers of service: Prime members get access to all the music in shuffle...DataInternshipFlexible hours$154.6k - $209.1k
...speed.We are seeking a Senior Data Engineer to design, build and... ...proof data ecosystem.As a key member of our data team, you'll collaborate... ...demand. You'll help shape our technical roadmap for data systems that... ...and feature pipelines for ML/GenAI training, fine-tuning, and...DataFlexible hours- ...enterprises prepare and optimize data at the most fundamental layer... ...FOR 7+ years in backend or infra engineering roles, with 2+ years... ...environments (e.g. databases, ML infra, lakehouse systems) Deep... ...-tier investors and built for technical founders and builders...DataFull timeFlexible hours
$202.16k - $368.22k
...Acceleration Engineer - DPU & AI Infra ByteDance Seattle, WA, US... ...About the Team The ByteDance DPU (Data Processing Unit) team is... ...virtualization and scheduling for AI/ML workloads We work at the... ...Contribute to architecture design, technical proposals, and long-term...DataTemporary workLocal area$186.1k - $300.55k
...Docusign unleashes business-critical data that is trapped inside of... ...business requirements into technical specificationsImplement and maintain... ...analytics techniques using ML and AI to assist data... ...responsibility to ensure every team member has an equal opportunity to succeed...DataPermanent employmentFull timeContract workWork at officeLocal areaRemote work2 days per week$214.51k
...Smith is seeking a Senior Data and AI Architect to... ..., data science, and AI/ML to give our business and... ...champion implementation. As a member of the Data & AI... ...thinker with a strong technical background, has exceptional... ...and subsidiaries staff will not be considered...DataWork experience placementH1b- ...a defining moment in the platform journey. OpenAI has the data, product surface area, and pace of innovation to learn faster... ...We are looking for a Machine Learning Engineer to lead the technical direction for ML-powered experimentation and insights capabilities. You will...DataFull timeWork at office
$19.55 - $24.8 per hour
...Sense Innovation Accountability Position Data Office Support Technician-J23-U00-H Regular... ...as knowledgeable resource to Yakima County staff, job applicants and the general public;... ...records requests or refers to appropriate staff member for response. Assists the general public...DataHourly payPart timeFor contractorsWork at office$54k - $80k
Job Overview The Technical Support Specialist provides technical support... ...with senior level team members as needed to troubleshoot and... ...maintaining historical ticket data Monitor assigned open tickets... ...Technical Support Specialist staff Complete all daily tasks and...DataContract workRemote workMonday to FridayFlexible hoursShift workWeekend workDay shiftAfternoon shiftEarly shift- ...experience by working at a premier 7/24 San Diego data center. Although beneficial, prior IT or... ...and data center operations. Serving as a technical liaison to both internal and external... .... Communicating and working with team members to coordinate efforts to support clients....Data
- ...Seattle team. This role provides hands‑on technical support for hardware, software, and... ...Escalate more complex problems to senior team members as needed. Log issues and resolutions in... ...Frequent use of a keyboard for data entry and troubleshooting. May be required...DataContract workTemporary workSeasonal workWork at officeWorldwideFlexible hoursShift work
- ...high-performance inference for AI foundation models. You will work on building large-scale distributed training systems, optimizing data/model parallelism and memory efficiency, and collaborate with researchers and engineers to translate model requirements into scalable...DataInternship
- The Service Desk Administrator provides technical oversight and guidance to other Service Desk team members and applies specialized knowledge and skills to resolve escalated... ..., and advanced infrastructure Cloud-based data centers such as Azure and AWS Server...DataWork at officeLocal areaRemote work
$32.23 - $47.59 per hour
...point for the Service Desk's most complex technical issues, applying considerable judgment... ...complex technical issues, ensuring team members can resolve recurring problems independently... ..., procurement, and regular audits, using data to identify and recommend improvements while...DataFull time- ...worldwide. With more than 1,700+ team members, 1,500+ AI & data experts, and 100+ prime contracts, we... ...measurable performance indicators. AI/ML Strategy, Delivery, and Responsible Use... ...security, privacy, records management, technical maturity, interoperability,...DataContract workTemporary workFor contractorsWorldwide
- ...years of experience to provide exceptional support for our members. You'll troubleshoot technical issues, collaborate with cross-functional teams, and... ...library of how Range worksMaintain success metrics and data as directedBe an active part of a collaborative team and...DataWork at officeRelocationMonday to Friday
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Member of Technical Staff ML Data Infra. Be the first to apply!
- helpdesk support technician Seattle, WA
- technical associate Seattle, WA
- customer support technician Seattle, WA
- work from home technical support specialist Seattle, WA
- customer support analyst Seattle, WA
- technical support associate Seattle, WA
- operations support technician Seattle, WA
- user support analyst Seattle, WA
- help desk technical support Seattle, WA
- support technician Seattle, WA


