Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Member of Technical Staff (Software Engineer, Inference & Training Platform) Perplexity AI San [...]

Neura Market

Perplexity serves hundreds of millions of queries a month, and every one of them fans out into multiple AI inference requests running in real time. Behind that sits a large GPU fleet spread across several cloud providers. Today, our inference engineers and researchers build models while also managing networking, securing capacity, and operating the underlying GPU clusters, responsibilities we want a dedicated platform team to own. Your job is to take ownership of that infrastructure and hide its complexity behind a unified, self-serve platform for running training and inference workloads. Responsibilities Build a self-serve compute platform. Design and own the systems that let inference engineers and researchers launch training jobs and operate inference services without managing GPU provisioning, cluster configuration, or provider-specific infrastructure. Operate the GPU fleet . Own provisioning, lifecycle management, reliability, and capacity integration across providers, giving teams a consistent way to use compute regardless of where it runs. Solve for GPU scarcity. Build the scheduling and placement logic that finds available capacity across providers, packs it efficiently, and gets the right workload onto the right hardware under real constraints. Support two very different workloads. Keep long-running distributed training jobs healthy while simultaneously guaranteeing the availability and latency of production inference services on the same fleet. Own the Kubernetes for GPU orchestration. Write the operators and CRDs, and manage many clusters across providers so the platform behaves the same everywhere we run. Make failure boring. Build the fault tolerance, autoscaling, and observability that keep the fleet utilized and let workloads survive node loss, provider hiccups, and capacity shifts without human intervention. Set technical direction across teams. Partner with inference and cloud infrastructure engineers to turn operational constraints into a coherent platform architecture and roadmap. Qualifications We expect you to have real depth in most of these: Deep Kubernetes experience—custom operators, CRDs, and multi-cluster federation, not just running kubectl apply. You’ve managed GPU clusters at scale: NVIDIA hardware, CUDA, and the networking that makes them fast (InfiniBand or RoCE). You’ve orchestrated compute across multiple clouds (CoreWeave, AWS, GCP, or similar) and understand how different each one really is. Strong distributed systems fundamentals: scheduling, resource allocation, and fault tolerance under load. You write infrastructure and systems-level code in Go, Rust or C++. You’ve supported both long-running training jobs and high-availability inference services, and you know why they pull infrastructure in opposite directions. You own problems end-to-end and do well when the path forward isn’t laid out for you. Additional experience we value Inference serving stacks: vLLM, SGLang, or TensorRT-LLM. Slurm or other HPC schedulers. GPU kernel work in CUDA or Triton—not required, but notable. High-speed interconnects: InfiniBand, RoCE, or RDMA in production. Observability for ML workloads: Prometheus, Grafana, or Weights & Biases. If you’re excited about this role, we encourage you to apply even if your experience doesn’t match every qualification listed above. #J-18808-Ljbffr Neura Market

Vacancy posted 3 days ago
Similar jobs that could be interesting for youBased on the Member of Technical Staff (Software Engineer, Inference & Training Platform) Perplexity AI San [...] in San Francisco, CA vacancy
  • $150k - $300k

     ...anyone to create, train, and deploy...  ...serving, LLM inference optimization...  ...stack. Core Technical...  ...LLM serving platform that operates...  ...LLM Inference engine development and...  ...arrangement (remote or San Francisco office...  ...AI and RL at Prime...  ...encourage team members to contribute... 
    Platform
    Training
    Work at office
    Remote work
    Visa sponsorship
    Relocation package
    Flexible hours
    Shift work

    Prime Intellect

    San Francisco, CA
    4 days ago
  • Member of Technical Staff - Agents at Prime Intellect - San Francisco Building the Future...  ...Decentralized AI At Prime...  ...or capital to train powerful, open...  ...research, and other engineering teams to...  ...Requirements Agent & Platform Skills Python...  ...training or inference on GPUs. Advanced... 
    Platform
    Training
    Remote work
    Flexible hours

    Victrays

    San Francisco, CA
    1 day ago
  •  ...humanity. We’re training and deploying...  ...are building AI systems to power...  ...researchers, engineers, designers,...  ...generation of AI platforms powering...  ...are looking for Members of Technical Staff to join the Model...  ...throughput of inference. ~ Strong understanding...  ..., New York, San Francisco,... 
    Platform
    Training
    Full time
    Work experience placement
    Work at office
    Remote work
    Flexible hours

    Cohere

    San Francisco, CA
    12 hours ago
  •  ...Decentralized AI Development At...  ...at scale. Our platform combines powerful distributed training infrastructure...  ...researchers and engineers to train state‑...  ...systems Core Technical Responsibilities...  ...arrangement (remote or San Francisco...  ...encourage team members to contribute to... 
    Platform
    Training
    Work at office
    Remote work
    Visa sponsorship
    Relocation package
    Flexible hours

    Kubelt

    San Francisco, CA
    12 hours ago
  • About Us: AI needs a new infrastructure...  ...serve low-latency inference, fine-tune models,...  ..., and experienced engineering and product leaders...  ...We're building a platform that covers the whole...  ...life of an LLM -- train it, deploy it,...  ...person, in our NYC or San Francisco office.... 
    Platform
    Training
    Work at office

    Mixpeek

    San Francisco, CA
    2 days ago
  • $190.9k - $232.8k

     ...About This RoleAs a staff software engineer for GenAI inference, you will lead the architecture...  ...collaboration: with platform engineers, cloud...  ...certifications and training, and specific work...  ...is the data and AI company. More than 1...  ...is headquartered in San Francisco, with offices... 
    Platform
    Training
    Local area
    Worldwide

    DataBricks

    San Francisco, CA
    2 days ago
  • $190k - $265k

     ...enabling data and AI teams to solve...  ...infrastructure platform so our...  ...business. Founded by engineers — and customer-...  ...opportunity to solve technical challenges,...  ...agents, model training, model serving,...  ...Foundation Model Inference team is the...  ...headquartered in San Francisco, with... 
    Platform
    Training
    Local area
    Worldwide

    DataBricks

    San Francisco, CA
    2 days ago
  •  ...Pixeltable Inc. Member of Technical Staff San Francisco, CA·Full...  ...founding member of the engineering team, you will...  ...revolutionizing the AI development...  ...our data-centric platform designed to simplify...  ..., transformation, training/fine-tuning, and inference? You will also: Find... 
    Platform
    Training
    Full time
    Part time
    Work at office
    Work from home
    Flexible hours
    2 days per week

    Pixeltable, Inc.

    San Francisco, CA
    4 days ago
  •  ...breakthrough AI application,...  ...today's data platforms (like Databricks...  ...open-source engine, Daft, is...  ...Databricks and Perplexity, we're looking...  ...Role As a Member of Technical Staff, you will be...  ...building and training state-of-the-...  ...Familiarity with inference optimization... 
    Platform
    Training
    Work at office
    Immediate start
    Flexible hours
    Night shift

    Eventual

    San Francisco, CA
    4 hours ago
  •  ...era of agentic AI. Millions of people now use Perplexity to transform knowledge...  .... As a growth engineer at Perplexity,...  ...projects from training and...  ...of professional software engineering experience...  ...experimentation platforms (Eppo, Statsig,...  ..., effective technical solutions. Self... 
    Platform
    Training

    aijoblist

    San Francisco, CA
    4 days ago
  • $220k - $260k

     ...interaction — a unified platform that combines...  ...enterprise AI agents need in...  ...experienced Staff Software Engineer to help...  ...also shaping the technical culture behind...  ...based in our San Francisco office...  ...education or training to determine individual...  ...with team members around the... 
    Platform
    Training
    Work at office

    Redpanda Data

    San Francisco, CA
    1 day ago
  • Careers / Member of Technical Staff (AI research) Member of Technical...  ...parts of Fearn’s platform from deploying our...  ...) Location San Francisco, California...  ...a crucial role in training models and designing inference pipelines, pushing...  ...Compute Engine, GKE, Vertex AI, or... 
    Platform
    Training
    Full time
    Work at office

    Kindredventures

    San Francisco, CA
    1 day ago
  • Perplexity is seeking energetic engineers to join our highly driven Agents engineering team....  ...backend, full-stack, and AI/ML engineers who collaborate...  ...Perplexity Computer (our platform for generalized frontier...  ...of work for our users; Training action and decision... 
    Platform
    Training
    Flexible hours

    The Consensus

    San Francisco, CA
    3 days ago
  • Member of Technical Staff, Applied AI The opportunity We are looking...  ..., protein engineers and...  ...protein screening platforms. At Latent Labs...  ...our London and San Francisco...  ...architectures, training dynamics and inference behaviour. You...  ...enterprise software. You have experience... 
    Platform
    Training
    Flexible hours

    Latent Labs

    San Francisco, CA
    12 hours ago
  • $150k - $300k

     ...anyone create, train, and deploy...  ...Scientist, Together AI), Dylan Patel...  ...training platform - the product...  ...jobs. Core Technical...  ...training and inference orchestration...  ...looking for engineers who are fluent...  ...arrangement (remote or San Francisco office...  ...team members to contribute... 
    Platform
    Training
    Work at office
    Local area
    Remote work
    Visa sponsorship
    Relocation package
    Flexible hours

    Kubelt

    San Francisco, CA
    12 hours ago
  •  ...'s Frontier AI & Robotics team...  .... As a Member of Technical Staff, you'll be at...  ...collaborating with platform teams to...  ...model inference, video tokenization...  ...robotics engineers to integrate...  ...datasets to train and deploy state...  ...software development...  ...Pursuant to the San Francisco Fair... 
    Platform
    Training
    Local area

    Amazon Science

    San Francisco, CA
    2 days ago
  • $250k

    Eragon — Member of Technical Staff Type: Full-time | On-site | San Francisco, CA Compensation...  ...-grade AI operating system. It post-trains open-source models...  ...Systems engineering: Design scalable...  ...for training, inference, and data processing...  ...also shows a platform "Pending... 
    Platform
    Training
    Full time
    H1b
    Work at office
    Local area
    Visa sponsorship

    davidjoseph-co

    San Francisco, CA
    4 days ago
  • About Perplexity AI Perplexity is an AI-powered answer engine built to serve the world...  ...that can use software like a human...  ...Role The Data Platform team owns the...  ...In this senior/staff role, you will...  ...the long-term technical direction of Perplexity...  ...features, AI training and evaluation... 
    Platform
    Training

    Perplexity

    San Francisco, CA
    4 days ago
  •  ...looking for a Member of Technical Staff with 2+ years...  ...team building AI that autonomously...  ...working on a platform already...  ...Applying context engineering, agent harnesses...  ..., small model training, and RL techniques...  ...in full-stack software engineering,...  ...in-office in San Francisco Full... 
    Platform
    Training
    Full time
    Work at office
    Immediate start
    Visa sponsorship

    Spot Hunting USA

    San Francisco, CA
    1 day ago
  • Member of Technical Staff, Lead Researcher San Francisco, CA; Sunnyvale, CA About...  ...is building an AI Research org from...  ...mentor researchers, engineers, and fellows...  ...DoorDash with ML platform, product, and operations...  ...budgets for training and inference, sized to support... 
    Platform
    Training
    Local area

    DoorDash USA

    San Francisco, CA
    4 days ago
  •  ...capabilities both in training and production,...  ...optimize for evolving technical and product...  ...integration of ML inference, monitoring systems...  ...Kubernetes, serverless platforms), and database and...  ...week in downtown San Francisco....  ...notify ****@*****.***.ai for support. #J-18... 
    Platform
    Training
    Local area
    Relocation

    additiveai

    San Francisco, CA
    2 days ago
  • $220k - $405k

    Location San Francisco; New York City...  ...Department Product Engineering Compensation $2...  ...listed above. Perplexity is AI for people who expect...  ...the billing platform this role owns....  ...definition through technical design, implementation...  ...of professional software engineering... 
    Platform
    Full time
    Local area

    B Capital

    San Francisco, CA
    3 days ago
  • $150k - $300k

     ...anyone to create, train, and deploy...  ...the developer platform that makes all...  ...(Eureka AI, Tesla, OpenAI...  ...a generalist software engineering role focused on...  ...improvements Technical Requirements...  ...arrangement (remote or San Francisco...  ...encourage team members to contribute... 
    Platform
    Work at office
    Remote work
    Visa sponsorship
    Relocation package
    Flexible hours

    Prime Intellect

    San Francisco, CA
    4 days ago
  •  ...fast, efficient inference. As AI workloads become...  ...together. Gimlet's platform intelligently partitions...  ...headcount. The engineers we hire today...  ...will be built on software capable of...  ...combination of education, training, and professional...  .... As an early member of the team, you... 
    Platform
    Training

    The Consensus

    San Francisco, CA
    2 days ago
  • $200k

    Member of Technical Staff, Supercomputing Platform & Infrastructure Magic’s mission...  ...-scale pre-training, domain-...  ...context, and inference-time compute to...  ...the role As an engineer on the Supercomputing...  ...and manage AI workloads Develop...  ...for Strong software engineering skills... 
    Platform
    Training
    Relocation
    Visa sponsorship

    Magic AI, Inc

    San Francisco, CA
    2 days ago
  • Introducing Moonlake, AI for creating...  .... Our platform enables the creation...  ...used to train the next generation...  ...looking for a Member of Technical Staff - Robotics to...  ...and engineers developing next...  ...Debug hardware, software, sensing, and...  ...currently based in San Francisco. #J... 
    Platform
    Training

    Embedding VC

    San Francisco, CA
    1 day ago
  • $140k - $225k

    Member of Technical Staff — SketchPro.ai Location: San Francisco, CA (Onsite, 5 days/week — office next...  ...broader AEC stack. The platform automates architecture...  ...What You'll Own Agent engineering across context design,...  ...researchers focused on model training only Pure computer... 
    Platform
    Training
    Full time
    H1b
    Work at office
    Visa sponsorship

    davidjoseph-co

    San Francisco, CA
    1 day ago
  •  ...Perplexity is seeking creative, AI-pilled engineers to join our Acceleration team. Our...  ...team will lead both software and process engineering...  ...possess sound technical judgment forged in...  ..., infrastructure/platform engineering, developer...  ...work of technical staff across disciplines... 
    Platform
    Full time
    Shift work

    Perplexity®️

    San Francisco, CA
    12 hours ago
  •  ...of machine learners, software engineers, biologists and bioinformaticians...  ...protein screening platforms. At Latent Labs you...  ...minds in generative AI and biology. Our...  ...our London and San Francisco sites. We’...  ...have experience running training and inference on cloud hardware, distributing... 
    Platform
    Training
    Flexible hours

    Latent Labs Ltd.

    San Francisco, CA
    12 hours ago
  • $200k - $400k

     ...the world. With AI, anyone can create...  ...100 enterprises, trained a first-of-its-kind...  ...researchers, engineers, designers, and operators...  ...means running inference over populations...  ...the Role As a Member of Technical Staff in Research Infrastructure...  ...will build the platform our researchers... 
    Platform
    Training
    Live in
    Flexible hours

    Simile

    San Francisco, CA
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Member of Technical Staff (Software Engineer, Inference & Training Platform) Perplexity AI San [...]. Be the first to apply!