Member of Technical Staff (Software Engineer, Inference & Training Platform) Perplexity AI San [...]
Neura Market
Perplexity serves hundreds of millions of queries a month, and every one of them fans out into multiple AI inference requests running in real time. Behind that sits a large GPU fleet spread across several cloud providers. Today, our inference engineers and researchers build models while also managing networking, securing capacity, and operating the underlying GPU clusters, responsibilities we want a dedicated platform team to own. Your job is to take ownership of that infrastructure and hide its complexity behind a unified, self-serve platform for running training and inference workloads. Responsibilities Build a self-serve compute platform. Design and own the systems that let inference engineers and researchers launch training jobs and operate inference services without managing GPU provisioning, cluster configuration, or provider-specific infrastructure. Operate the GPU fleet . Own provisioning, lifecycle management, reliability, and capacity integration across providers, giving teams a consistent way to use compute regardless of where it runs. Solve for GPU scarcity. Build the scheduling and placement logic that finds available capacity across providers, packs it efficiently, and gets the right workload onto the right hardware under real constraints. Support two very different workloads. Keep long-running distributed training jobs healthy while simultaneously guaranteeing the availability and latency of production inference services on the same fleet. Own the Kubernetes for GPU orchestration. Write the operators and CRDs, and manage many clusters across providers so the platform behaves the same everywhere we run. Make failure boring. Build the fault tolerance, autoscaling, and observability that keep the fleet utilized and let workloads survive node loss, provider hiccups, and capacity shifts without human intervention. Set technical direction across teams. Partner with inference and cloud infrastructure engineers to turn operational constraints into a coherent platform architecture and roadmap. Qualifications We expect you to have real depth in most of these: Deep Kubernetes experience—custom operators, CRDs, and multi-cluster federation, not just running kubectl apply. You’ve managed GPU clusters at scale: NVIDIA hardware, CUDA, and the networking that makes them fast (InfiniBand or RoCE). You’ve orchestrated compute across multiple clouds (CoreWeave, AWS, GCP, or similar) and understand how different each one really is. Strong distributed systems fundamentals: scheduling, resource allocation, and fault tolerance under load. You write infrastructure and systems-level code in Go, Rust or C++. You’ve supported both long-running training jobs and high-availability inference services, and you know why they pull infrastructure in opposite directions. You own problems end-to-end and do well when the path forward isn’t laid out for you. Additional experience we value Inference serving stacks: vLLM, SGLang, or TensorRT-LLM. Slurm or other HPC schedulers. GPU kernel work in CUDA or Triton—not required, but notable. High-speed interconnects: InfiniBand, RoCE, or RDMA in production. Observability for ML workloads: Prometheus, Grafana, or Weights & Biases. If you’re excited about this role, we encourage you to apply even if your experience doesn’t match every qualification listed above. #J-18808-Ljbffr Neura Market
$150k - $300k
...anyone to create, train, and deploy... ...serving, LLM inference optimization... ...stack. Core Technical... ...LLM serving platform that operates... ...LLM Inference engine development and... ...arrangement (remote or San Francisco office... ...AI and RL at Prime... ...encourage team members to contribute...PlatformTrainingWork at officeRemote workVisa sponsorshipRelocation packageFlexible hoursShift work- Member of Technical Staff - Agents at Prime Intellect - San Francisco Building the Future... ...Decentralized AI At Prime... ...or capital to train powerful, open... ...research, and other engineering teams to... ...Requirements Agent & Platform Skills Python... ...training or inference on GPUs. Advanced...PlatformTrainingRemote workFlexible hours
- ...humanity. We’re training and deploying... ...are building AI systems to power... ...researchers, engineers, designers,... ...generation of AI platforms powering... ...are looking for Members of Technical Staff to join the Model... ...throughput of inference. ~ Strong understanding... ..., New York, San Francisco,...PlatformTrainingFull timeWork experience placementWork at officeRemote workFlexible hours
- ...Decentralized AI Development At... ...at scale. Our platform combines powerful distributed training infrastructure... ...researchers and engineers to train state‑... ...systems Core Technical Responsibilities... ...arrangement (remote or San Francisco... ...encourage team members to contribute to...PlatformTrainingWork at officeRemote workVisa sponsorshipRelocation packageFlexible hours
- About Us: AI needs a new infrastructure... ...serve low-latency inference, fine-tune models,... ..., and experienced engineering and product leaders... ...We're building a platform that covers the whole... ...life of an LLM -- train it, deploy it,... ...person, in our NYC or San Francisco office....PlatformTrainingWork at office
$190.9k - $232.8k
...About This RoleAs a staff software engineer for GenAI inference, you will lead the architecture... ...collaboration: with platform engineers, cloud... ...certifications and training, and specific work... ...is the data and AI company. More than 1... ...is headquartered in San Francisco, with offices...PlatformTrainingLocal areaWorldwide$190k - $265k
...enabling data and AI teams to solve... ...infrastructure platform so our... ...business. Founded by engineers — and customer-... ...opportunity to solve technical challenges,... ...agents, model training, model serving,... ...Foundation Model Inference team is the... ...headquartered in San Francisco, with...PlatformTrainingLocal areaWorldwide- ...Pixeltable Inc. Member of Technical Staff San Francisco, CA·Full... ...founding member of the engineering team, you will... ...revolutionizing the AI development... ...our data-centric platform designed to simplify... ..., transformation, training/fine-tuning, and inference? You will also: Find...PlatformTrainingFull timePart timeWork at officeWork from homeFlexible hours2 days per week
- ...breakthrough AI application,... ...today's data platforms (like Databricks... ...open-source engine, Daft, is... ...Databricks and Perplexity, we're looking... ...Role As a Member of Technical Staff, you will be... ...building and training state-of-the-... ...Familiarity with inference optimization...PlatformTrainingWork at officeImmediate startFlexible hoursNight shift
- ...era of agentic AI. Millions of people now use Perplexity to transform knowledge... .... As a growth engineer at Perplexity,... ...projects from training and... ...of professional software engineering experience... ...experimentation platforms (Eppo, Statsig,... ..., effective technical solutions. Self...PlatformTraining
$220k - $260k
...interaction — a unified platform that combines... ...enterprise AI agents need in... ...experienced Staff Software Engineer to help... ...also shaping the technical culture behind... ...based in our San Francisco office... ...education or training to determine individual... ...with team members around the...PlatformTrainingWork at office- Careers / Member of Technical Staff (AI research) Member of Technical... ...parts of Fearn’s platform from deploying our... ...) Location San Francisco, California... ...a crucial role in training models and designing inference pipelines, pushing... ...Compute Engine, GKE, Vertex AI, or...PlatformTrainingFull timeWork at office
- Perplexity is seeking energetic engineers to join our highly driven Agents engineering team.... ...backend, full-stack, and AI/ML engineers who collaborate... ...Perplexity Computer (our platform for generalized frontier... ...of work for our users; Training action and decision...PlatformTrainingFlexible hours
- Member of Technical Staff, Applied AI The opportunity We are looking... ..., protein engineers and... ...protein screening platforms. At Latent Labs... ...our London and San Francisco... ...architectures, training dynamics and inference behaviour. You... ...enterprise software. You have experience...PlatformTrainingFlexible hours
$150k - $300k
...anyone create, train, and deploy... ...Scientist, Together AI), Dylan Patel... ...training platform - the product... ...jobs. Core Technical... ...training and inference orchestration... ...looking for engineers who are fluent... ...arrangement (remote or San Francisco office... ...team members to contribute...PlatformTrainingWork at officeLocal areaRemote workVisa sponsorshipRelocation packageFlexible hours- ...'s Frontier AI & Robotics team... .... As a Member of Technical Staff, you'll be at... ...collaborating with platform teams to... ...model inference, video tokenization... ...robotics engineers to integrate... ...datasets to train and deploy state... ...software development... ...Pursuant to the San Francisco Fair...PlatformTrainingLocal area
$250k
Eragon — Member of Technical Staff Type: Full-time | On-site | San Francisco, CA Compensation... ...-grade AI operating system. It post-trains open-source models... ...Systems engineering: Design scalable... ...for training, inference, and data processing... ...also shows a platform "Pending...PlatformTrainingFull timeH1bWork at officeLocal areaVisa sponsorship- About Perplexity AI Perplexity is an AI-powered answer engine built to serve the world... ...that can use software like a human... ...Role The Data Platform team owns the... ...In this senior/staff role, you will... ...the long-term technical direction of Perplexity... ...features, AI training and evaluation...PlatformTraining
- ...looking for a Member of Technical Staff with 2+ years... ...team building AI that autonomously... ...working on a platform already... ...Applying context engineering, agent harnesses... ..., small model training, and RL techniques... ...in full-stack software engineering,... ...in-office in San Francisco Full...PlatformTrainingFull timeWork at officeImmediate startVisa sponsorship
- Member of Technical Staff, Lead Researcher San Francisco, CA; Sunnyvale, CA About... ...is building an AI Research org from... ...mentor researchers, engineers, and fellows... ...DoorDash with ML platform, product, and operations... ...budgets for training and inference, sized to support...PlatformTrainingLocal area
- ...capabilities both in training and production,... ...optimize for evolving technical and product... ...integration of ML inference, monitoring systems... ...Kubernetes, serverless platforms), and database and... ...week in downtown San Francisco.... ...notify ****@*****.***.ai for support. #J-18...PlatformTrainingLocal areaRelocation
$220k - $405k
Location San Francisco; New York City... ...Department Product Engineering Compensation $2... ...listed above. Perplexity is AI for people who expect... ...the billing platform this role owns.... ...definition through technical design, implementation... ...of professional software engineering...PlatformFull timeLocal area$150k - $300k
...anyone to create, train, and deploy... ...the developer platform that makes all... ...(Eureka AI, Tesla, OpenAI... ...a generalist software engineering role focused on... ...improvements Technical Requirements... ...arrangement (remote or San Francisco... ...encourage team members to contribute...PlatformWork at officeRemote workVisa sponsorshipRelocation packageFlexible hours- ...fast, efficient inference. As AI workloads become... ...together. Gimlet's platform intelligently partitions... ...headcount. The engineers we hire today... ...will be built on software capable of... ...combination of education, training, and professional... .... As an early member of the team, you...PlatformTraining
$200k
Member of Technical Staff, Supercomputing Platform & Infrastructure Magic’s mission... ...-scale pre-training, domain-... ...context, and inference-time compute to... ...the role As an engineer on the Supercomputing... ...and manage AI workloads Develop... ...for Strong software engineering skills...PlatformTrainingRelocationVisa sponsorship- Introducing Moonlake, AI for creating... .... Our platform enables the creation... ...used to train the next generation... ...looking for a Member of Technical Staff - Robotics to... ...and engineers developing next... ...Debug hardware, software, sensing, and... ...currently based in San Francisco. #J...PlatformTraining
$140k - $225k
Member of Technical Staff — SketchPro.ai Location: San Francisco, CA (Onsite, 5 days/week — office next... ...broader AEC stack. The platform automates architecture... ...What You'll Own Agent engineering across context design,... ...researchers focused on model training only Pure computer...PlatformTrainingFull timeH1bWork at officeVisa sponsorship- ...Perplexity is seeking creative, AI-pilled engineers to join our Acceleration team. Our... ...team will lead both software and process engineering... ...possess sound technical judgment forged in... ..., infrastructure/platform engineering, developer... ...work of technical staff across disciplines...PlatformFull timeShift work
- ...of machine learners, software engineers, biologists and bioinformaticians... ...protein screening platforms. At Latent Labs you... ...minds in generative AI and biology. Our... ...our London and San Francisco sites. We’... ...have experience running training and inference on cloud hardware, distributing...PlatformTrainingFlexible hours
$200k - $400k
...the world. With AI, anyone can create... ...100 enterprises, trained a first-of-its-kind... ...researchers, engineers, designers, and operators... ...means running inference over populations... ...the Role As a Member of Technical Staff in Research Infrastructure... ...will build the platform our researchers...PlatformTrainingLive inFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Member of Technical Staff (Software Engineer, Inference & Training Platform) Perplexity AI San [...]. Be the first to apply!
- work from home technical support specialist San Francisco, CA
- product support technician San Francisco, CA
- helpdesk support technician San Francisco, CA
- help desk assistant San Francisco, CA
- senior technical associate San Francisco, CA
- IT help desk technician San Francisco, CA
- technical solutions specialist San Francisco, CA
- desktop support analyst San Francisco, CA
- trade support analyst San Francisco, CA
- senior IT support technician San Francisco, CA



