Member of Technical Staff (Software Engineer, Inference & Training Platform) Perplexity AI San [...]
Neura Market
Perplexity serves hundreds of millions of queries a month, and every one of them fans out into multiple AI inference requests running in real time. Behind that sits a large GPU fleet spread across several cloud providers. Today, our inference engineers and researchers build models while also managing networking, securing capacity, and operating the underlying GPU clusters, responsibilities we want a dedicated platform team to own. Your job is to take ownership of that infrastructure and hide its complexity behind a unified, self-serve platform for running training and inference workloads. Responsibilities Build a self-serve compute platform. Design and own the systems that let inference engineers and researchers launch training jobs and operate inference services without managing GPU provisioning, cluster configuration, or provider-specific infrastructure. Operate the GPU fleet . Own provisioning, lifecycle management, reliability, and capacity integration across providers, giving teams a consistent way to use compute regardless of where it runs. Solve for GPU scarcity. Build the scheduling and placement logic that finds available capacity across providers, packs it efficiently, and gets the right workload onto the right hardware under real constraints. Support two very different workloads. Keep long-running distributed training jobs healthy while simultaneously guaranteeing the availability and latency of production inference services on the same fleet. Own the Kubernetes for GPU orchestration. Write the operators and CRDs, and manage many clusters across providers so the platform behaves the same everywhere we run. Make failure boring. Build the fault tolerance, autoscaling, and observability that keep the fleet utilized and let workloads survive node loss, provider hiccups, and capacity shifts without human intervention. Set technical direction across teams. Partner with inference and cloud infrastructure engineers to turn operational constraints into a coherent platform architecture and roadmap. Qualifications We expect you to have real depth in most of these: Deep Kubernetes experience—custom operators, CRDs, and multi-cluster federation, not just running kubectl apply. You’ve managed GPU clusters at scale: NVIDIA hardware, CUDA, and the networking that makes them fast (InfiniBand or RoCE). You’ve orchestrated compute across multiple clouds (CoreWeave, AWS, GCP, or similar) and understand how different each one really is. Strong distributed systems fundamentals: scheduling, resource allocation, and fault tolerance under load. You write infrastructure and systems-level code in Go, Rust or C++. You’ve supported both long-running training jobs and high-availability inference services, and you know why they pull infrastructure in opposite directions. You own problems end-to-end and do well when the path forward isn’t laid out for you. Additional experience we value Inference serving stacks: vLLM, SGLang, or TensorRT-LLM. Slurm or other HPC schedulers. GPU kernel work in CUDA or Triton—not required, but notable. High-speed interconnects: InfiniBand, RoCE, or RDMA in production. Observability for ML workloads: Prometheus, Grafana, or Weights & Biases. If you’re excited about this role, we encourage you to apply even if your experience doesn’t match every qualification listed above. #J-18808-Ljbffr Neura Market
$150k - $300k
...anyone to create, train, and deploy... ...serving, LLM inference optimization... ...stack. Core Technical... ...LLM serving platform that operates... ...LLM Inference engine development and... ...arrangement (remote or San Francisco office... ...AI and RL at Prime... ...encourage team members to contribute...PlatformTrainingWork at officeRemote workVisa sponsorshipRelocation packageFlexible hoursShift work- ..., Mexico, San Francisco,... ...Semiconductor and AI industries... ...Models, software, and... ...our deep technical research and... ...Overview Member of Technical Staff will play... ...developing training & inference benchmarks... ...Science, Engineering or other relevant... ...hardware platforms and large-...PlatformTrainingFull timeWork at officeRemote workWorldwide
- Member of Technical Staff - Agents at Prime Intellect - San Francisco Building the Future... ...Decentralized AI At Prime... ...or capital to train powerful, open... ...research, and other engineering teams to... ...Requirements Agent & Platform Skills Python... ...training or inference on GPUs. Advanced...PlatformTrainingRemote workFlexible hours
- ...Decentralized AI Development At... ...at scale. Our platform combines powerful distributed training infrastructure... ...researchers and engineers to train state‑... ...systems Core Technical Responsibilities... ...arrangement (remote or San Francisco... ...encourage team members to contribute to...PlatformTrainingWork at officeRemote workVisa sponsorshipRelocation packageFlexible hours
- About Us: AI needs a new infrastructure... ...serve low-latency inference, fine-tune models,... ..., and experienced engineering and product leaders... ...We're building a platform that covers the whole... ...life of an LLM -- train it, deploy it,... ...person, in our NYC or San Francisco office....PlatformTrainingWork at office
$190.9k - $232.8k
...About This RoleAs a staff software engineer for GenAI inference, you will lead the architecture... ...collaboration: with platform engineers, cloud... ...certifications and training, and specific work... ...is the data and AI company. More than 1... ...is headquartered in San Francisco, with offices...PlatformTrainingLocal areaWorldwide- ...Pixeltable Inc. Member of Technical Staff San Francisco, CA·Full... ...founding member of the engineering team, you will... ...revolutionizing the AI development... ...our data-centric platform designed to simplify... ..., transformation, training/fine-tuning, and inference? You will also: Find...PlatformTrainingFull timePart timeWork at officeWork from homeFlexible hours2 days per week
$190k - $265k
...enabling data and AI teams to solve... ...infrastructure platform so our... ...business. Founded by engineers — and customer-... ...opportunity to solve technical challenges,... ...agents, model training, model serving,... ...time and batch inference, powering model... ...headquartered in San Francisco, with...PlatformTrainingLocal areaWorldwide$220k - $260k
...interaction — a unified platform that combines... ...enterprise AI agents need in... ...experienced Staff Software Engineer to help... ...also shaping the technical culture behind... ...based in our San Francisco office... ...education or training to determine individual... ...with team members around the...PlatformTrainingFull timeWork at office- ...era of agentic AI. Millions of people now use Perplexity to transform knowledge... .... As a growth engineer at Perplexity,... ...projects from training and... ...of professional software engineering experience... ...experimentation platforms (Eppo, Statsig,... ..., effective technical solutions. Self...PlatformTraining
- Member of Technical Staff (Infra) We're looking... ...experienced Backend Engineer to join... ...Location San Francisco,... ...AI infra engineer... ...infrastructure and inference systems... ...power our platform, while mentoring... ...+ years of software engineering... ...training systems. Familiarity...PlatformTrainingFull timeWork at office
- Careers / Member of Technical Staff (AI research) Member of Technical... ...parts of Fearn’s platform from deploying our... ...) Location San Francisco, California... ...a crucial role in training models and designing inference pipelines, pushing... ...Compute Engine, GKE, Vertex AI, or...PlatformTrainingFull timeWork at office
- Member of Technical Staff, Applied AI The opportunity We are looking... ..., protein engineers and... ...protein screening platforms. At Latent Labs... ...our London and San Francisco... ...architectures, training dynamics and inference behaviour. You... ...enterprise software. You have experience...PlatformTrainingFlexible hours
$150k - $300k
...anyone create, train, and deploy... ...Scientist, Together AI), Dylan Patel... ...training platform - the product... ...jobs. Core Technical... ...training and inference orchestration... ...looking for engineers who are fluent... ...arrangement (remote or San Francisco office... ...team members to contribute...PlatformTrainingWork at officeLocal areaRemote workVisa sponsorshipRelocation packageFlexible hours- Perplexity is seeking energetic engineers to join our highly driven Agents engineering team.... ...backend, full-stack, and AI/ML engineers who collaborate... ...Perplexity Computer (our platform for generalized frontier... ...of work for our users; Training action and decision...PlatformTrainingFlexible hours
- ...breakthrough AI application,... ...today’s data platforms (like Databricks... ...open-source engine, Daft, is... ...Databricks and Perplexity, we’re looking... ...Your Role: As a Member of Technical Staff, you will be... ...building and training state-of-the-... ...with inference optimization...PlatformTrainingWork at officeImmediate startFlexible hoursNight shift
- Job Description - Member of Technical Staff (Inference) Location: San Francisco (on-site at our offices)... ...is the leading independent AI benchmarking company. We support labs, engineers and enterprises to... ...our inference benchmarking platform together with our engineers...PlatformShift work
- About Perplexity AI Perplexity is an AI-powered answer engine built to serve the world... ...that can use software like a human... ...Role The Data Platform team owns the... ...In this senior/staff role, you will... ...the long-term technical direction of Perplexity... ...features, AI training and evaluation...PlatformTraining
$2,000 per month
...About Build AI Build AI is the data hyperscaler... ..., and model training to scale the... ...Summary We’re hiring Members of Technical Staff — the general research/engineering seat. If you have... ...of waiting for a platform team Work with Head... ...for those moving to San Francisco (Financial...PlatformTrainingWork at officeRelocation package- Member of Technical Staff, Lead Researcher San Francisco, CA; Sunnyvale, CA About... ...is building an AI Research org from... ...mentor researchers, engineers, and fellows... ...DoorDash with ML platform, product, and operations... ...budgets for training and inference, sized to support...PlatformTrainingLocal area
$176.5k
...new team members? At Scribd... ...platform for all content... ...and our inference-driven features... ...Senior, Staff, and Principal engineers building... ...owning software end-to-end... ...agentic AI tools part... ...end, from technical design... ...location. San Francisco... ...education or training; and...PlatformTrainingFull timeFor contractorsLocal areaHome officeFlexible hours$250k - $385k
Location San Francisco, New York... ...Department AI Compensation... ...listed above. Perplexity is seeking... ...creative, AI-pilled engineers to join our... ...lead both software and process... ...possess sound technical judgment forged... .../platform engineering,... ...of technical staff across disciplines...PlatformFull timeLocal areaShift work$200k - $250k
...venture-backed AI startup building... ...benchmarks used to train computer-using... ...scaling the founding engineering team. Founded... ...As a founding Member of Technical Staff on platform engineering, you... ...training and inference infrastructure that... ...Location: San Francisco, CA...PlatformTrainingFull timeH1bRelocationVisa sponsorshipRelocation package$150k - $300k
...anyone to create, train, and deploy... ...the developer platform that makes all... ...(Eureka AI, Tesla, OpenAI... ...a generalist software engineering role focused on... ...improvements Technical Requirements... ...arrangement (remote or San Francisco... ...encourage team members to contribute...PlatformWork at officeRemote workVisa sponsorshipRelocation packageFlexible hours$300k
...lab developing AI capable of... ...infrastructure engineers who can build... ...rollouts, training orchestration, inference, evals, data... ...the durable platform that enables... ...Requirements Strong software engineering... ...by other technical users. Strong... ...based in our San Francisco office...PlatformTrainingWork at officeLocal area- ...fast, efficient inference. As AI workloads become... ...together. Gimlet's platform intelligently partitions... ...headcount. The engineers we hire today... ...will be built on software capable of... ...combination of education, training, and professional... .... As an early member of the team, you...PlatformTraining
$200k
Member of Technical Staff, Supercomputing Platform & Infrastructure Magic’s mission... ...-scale pre-training, domain-... ...context, and inference-time compute to... ...the role As an engineer on the Supercomputing... ...and manage AI workloads Develop... ...for Strong software engineering skills...PlatformTrainingRelocationVisa sponsorship- Introducing Moonlake, AI for creating... .... Our platform enables the creation... ...used to train the next generation... ...looking for a Member of Technical Staff - Robotics to... ...and engineers developing next... ...Debug hardware, software, sensing, and... ...currently based in San Francisco. #J...PlatformTraining
- Location San Francisco; London; New York Employment... ...On-site Department Engineering Our Mission... ...states. Our team of AI researchers and company... ...high-throughput model inference and mid-training workloads. Develop systems... ...inference platforms capable of serving and...PlatformTrainingFull timeRelocation package
- ...of machine learners, software engineers, biologists and bioinformaticians... ...protein screening platforms. At Latent Labs you... ...minds in generative AI and biology. Our... ...our London and San Francisco sites. We’... ...have experience running training and inference on cloud hardware, distributing...PlatformTrainingFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Member of Technical Staff (Software Engineer, Inference & Training Platform) Perplexity AI San [...]. Be the first to apply!
- mri tech aide San Francisco, CA
- salesforce technical analyst San Francisco, CA
- service desk assistant San Francisco, CA
- end user support technician San Francisco, CA
- operations support technician San Francisco, CA
- help desk technical support San Francisco, CA
- technical assistant San Francisco, CA
- support analyst San Francisco, CA
- technical associate San Francisco, CA
- life support technician San Francisco, CA



