Member of Technical Staff (TPM, Inference)
Apply
Perplexity is looking for a technical program manager to be the connective tissue between our model providers, engineering, and product teams, driving our core inference platform forward. Perplexity runs one of the highest-throughput inference stacks in the industry, serving Ask, Computer, and API traffic across a large and constantly shifting portfolio of first-party and third-party models. This role sits at the intersection of product, engineering, and finance: you'll orchestrate across model providers and internal teams to keep new models and capacity moving smoothly into production, while executing the roadmap for the inference platform itself. The ideal candidate has strong technical judgment, thrives coordinating across teams and external partners with competing timelines, and is energized by building the operating model for a function that doesn't have much precedent yet. Our Mission Perplexity's mission is to power curiosity. Curious people are the people who drive change in the world. Driving change is a continuous cycle of learning, building, and integrating. Learn: curious people constantly learn new things by asking more. They question the status quo in their own expertise and they constantly learn outside of it. Research is essential to them and never ending. Build: curious people make and create things, to show the world their new answers to problems no one else ever questioned. They take action on what they've learned. Makers need tools to create their products, their companies, their reality. Integrate: they must interact with the world as it is to drive change and adoption. True leaders do not simply build something and hope. They must have armies of agents and workers who can constantly work in millions of small ways. Repeat. For curious people this is a cycle that never ends. What You'll Do Execute the roadmap for the inference platform - request handling, rate limits and quotas, usage controls, and the reliability and observability surface engineering and product teams depend on Be the connective tissue between model providers and Perplexity's engineering and product teams - coordinating onboarding, launch readiness, and rollout for new models and capacity Drive latency, throughput, uptime, and cost-efficiency as core execution metrics, surfacing tradeoffs between them rather than letting them become side effects Run the operating model for model-release and optimization programs , including day-zero launches, across performance engineering, infrastructure, and product teams Lead cross-functional delivery for inference-stack changes , from planning through launch and post-launch validation Build the mechanisms that make releases predictable - rituals, dashboards, launch checklists - so inference releases stay low-risk at Perplexity's scale Partner with GPU capacity and compute teams to reconcile execution decisions against cost, capacity, and vendor constraints Qualifications Strong experience with technical program management or product management in infrastructure, distributed systems, or ML/model-serving products Direct experience with production LLM or ML inference - understanding what makes serving fast, reliable, and cheap rather than just what a roadmap slide says about it Comfort orchestrating across external partners and internal engineering teams with competing priorities and timelines Experience with data and metrics, and the judgment to surface difficult tradeoffs between latency, throughput, uptime, and cost Thrives in a small, agile team; has initiative and desire for ownership without much precedent to lean on 6+ years of combined technical program management or product management experience #J-18808-Ljbffr Apply
- Member of Technical Staff - ML Systems & Inference Bay Area, CA | Onsite Join a well-funded AI infrastructure startup building the orchestration layer for next-generation AI workloads This role sits at the intersection of ML systems, inference, distributed systems, and...Suggested
$150k - $350k
Member of Technical Staff | Distributed Systems San Francisco - Onsite $150k-$350k base + equity I'm working with a small, deeply technical... ...million Series A AI infrastructure company , building an inference cloud for agentic workloads that can partition and orchestrate...Suggested- ...evaluations, efficient training, reliable inference, and products that make those... ...About the role You will join as a full member of the technical team—not as an observer. Depending on... ...stronger fit for our Member of Technical Staff, Research — Early Career role. What you...SuggestedFull timeInternshipWork at officeVisa sponsorshipFlexible hours3 days per week
$200k
...employment type. We train action policies and world models on the largest corpus of action-labeled gaming data in the world. Member of Technical Staff is the title everyone in our technical team holds. Each MTS is not fit to a single role - we help you find the most...Suggested- ...About the Role Anthropic is looking for a Member of Technical Staff to join our Applied AI team. In this role, you will work directly with our largest enterprise customers and research partners to deploy Claude into production environments, ensuring high reliability, safety...SuggestedRemote workFlexible hours
$227.5k - $401k
...your career. We are motivated individuals who tackle unique technical challenges at scale and solve them as a team, delivering... ...AI research within the financial technology sector. As a Member of Technical Staff , you will operate with a high degree of autonomy and responsibility...Work at officeImmediate startRelocationFlexible hours- ...your career. We are motivated individuals who tackle unique technical challenges at scale and solve them as a team, delivering... ...AI research within the financial technology sector. As a Member of Technical Staff , you will operate with a high degree of autonomy and responsibility...Flexible hours
- # Research Member of Technical Staff, Materials and ManufacturingTurn physical limits in compute into materials, process, reliability, and manufacturing... ...still breaks. We follow repeated limits through software, inference, kernels, memory, hardware, materials, and manufacturing...Full timeRemote work
$180k - $280k
...drowning in repetitive tasks, paper-based bottlenecks, and fragmented data - and we're building the software to fix it. As a Member of the Technical Staff at Finch, you'll own critical features end-to-end, ship fast, and work directly alongside product, ops, and design in a...Work at officeRemote workFlexible hours1 day per week- Sail builds the world's most efficient software for inference (processing LLM tokens) and agent hosting (cloud VMs). Together, our technologies... ...and share as much detail about Sail as you want to hear. A technical interview with one of our Sailbox engineers. This will also be...Work at office
- ...orchestration systems (e.g. Kubernetes, Slurm) for topology-aware placement, preemption, quotas, and multi-tenancy across training and inference workloads Build software that abstracts cluster management and presents a unified, self-serve interface to researchers and...
- Member of Technical Staff - Simulation About Us Veeda AI is building the next generation of multimodal foundation world models for Physical AI. We're a small, fast-moving team of engineers and researchers from leading AI labs, tackling some of the most challenging problems...
- Member of Technical Staff: DataMonk is an AI-native accounts receivable platform for B2B companies. We automate billing, collections, and payments so finance teams can turn outstanding invoices into cash, faster. We believe the AR stack is broken and the finance function...Immediate start
$140k - $185k
...publishable evaluations and work with model labs and enterprises to shape how foundation models are measured. Job Description Role Member of Technical Staff - Research responsible for designing and implementing novel, high-impact benchmarks that assess challenging real-world...Full timeRelocation package- ...evaluations, efficient training, reliable inference, and products that make those... ...influential research, open-source contributions, technically ambitious independent projects, or production... ...hybrid role. We currently expect all staff to work from one of our offices at least...Work at officeVisa sponsorshipFlexible hours3 days per week
- ...evaluations, efficient training, reliable inference, and products that make those... ...findings in clear internal documents and technical reviews; contribute to papers, technical... ...based hybrid role. We currently expect all staff to work from one of our offices at least...Work at officeVisa sponsorshipFlexible hours3 days per week
$200k - $400k
Inferact's mission is to grow vLLM as the world's AI inference engine and accelerate AI progress by making inference cheaper and faster. Founded by the creators and core maintainers of vLLM, we sit at the intersection of models and hardware—a position that took years to...Remote workVisa sponsorship- ...the future of AI. Our mission: make intelligence open and accessible to all. Role Overview Reflection.AI is looking for a Member of Technical Staff - Infrastructure Security to secure our geographically diverse multi‑cloud Kubernetes and cloud environments. In this role...Work at officeVisa sponsorship
$140k - $185k
...full‑stack features (Python/Django backend and React/TypeScript frontend) on an in‑person team based in San Francisco. Role Member of Technical Staff - Platform responsible for owning and developing the platform that runs large-scale LLM benchmarks. You will build and...Full timeRelocationRelocation package- ...will build task classification, capability profiles, evaluation harnesses, and the cost governor that enforces spend ceilings before inference happens rather than after the invoice arrives. Routing decisions must be explainable and reproducible, because they end up in the...
$160k - $200k
Member of Technical Staff, Agent Delivery Full time · Remote · San Francisco or New York Important: if an employer asks you to log into their system via iCloud or Google, send a code, an SMS or Telegram password, run some code, or install software — refuse. These are signs...Full timeWork at officeLocal areaRemote work- THE ROLE You\'ll build the secure, reliable systems beneath Valkai\'s agents and products. The scope spans compute, storage, deployment, data pipelines, indexing, and retrieval. You will make these systems observable and dependable across cloud and self-hosted environments...Immediate start
- Our Mission Reflection is a research lab making intelligence open and accessible for everyone to use, customize, and build on. We build open models that let anyone control their intelligence and help shape the future of AI. Our mission: make intelligence open and accessible...Work at officeVisa sponsorship
- ...across aerospace, defense, and advanced manufacturing from day one. Competitive salary and meaningful equity participation. A small, technical team with backgrounds from NASA, Formula 1, and top engineering institutions. Three meals at the office, plus snacks, tea, and...Full timeWork at office
- ...optimization insights through compelling visualizations Own features from conception through deployment and iteration Contribute to technical architecture decisions and best practices Qualifications Strong full-stack engineering skills (React/TypeScript frontend, Python/...
- Our Mission Reflection is a research lab making intelligence open and accessible for everyone to use, customize, and build on. We build open models that let anyone control their intelligence and help shape the future of AI. Our mission: make intelligence open and accessible...Work at officeVisa sponsorship
- Footprint is the agentic platform that learns your compliance program and runs it end to end.Compliance teams at banks and fintechs are drowning in financial crimes investigations. Transaction volumes are surging, regulators keep raising the bar, and the work still runs...Full time
- THE ROLE Nirva is building the first AI fashion accessory for your mind: a necklace or bracelet that privately listens only to the wearer to enable proactive journaling, emotion tracking, and life coaching. Think Oura Ring for your mind. Backed by Lightspeed and South Park...
- ...accelerator developers, and infrastructure engineers whose designs have to hold up to scrutiny, not just pass a testbench. We are a small, technical team where every engineer ships across the full stack. There is no distinction between "research" and "production" here. You will...Full timeRemote work
- ...you will Build & ship end-to-end features (frontend UIs to backend services/APIs) Own projects full lifecycle Design discussions & technical scoping Implementation & testing Post-launch iteration Translate ideas to intuitive, performant, scalable UX Contribute to...Work at officeRemote workVisa sponsorshipMonday to Friday
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Member of Technical Staff (TPM, Inference). Be the first to apply!
- mri tech aide Brooklyn, NY
- salesforce technical analyst Brooklyn, NY
- service desk assistant Brooklyn, NY
- end user support technician Brooklyn, NY
- operations support technician Brooklyn, NY
- help desk technical support Brooklyn, NY
- technical assistant Brooklyn, NY
- support analyst Brooklyn, NY
- technical associate Brooklyn, NY
- life support technician Brooklyn, NY

