Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Member of Technical Staff (Software Engineer, Inference & Training Platform) Perplexity AI San [...]

Neura Market

Perplexity serves hundreds of millions of queries a month, and every one of them fans out into multiple AI inference requests running in real time. Behind that sits a large GPU fleet spread across several cloud providers. Today, our inference engineers and researchers build models while also managing networking, securing capacity, and operating the underlying GPU clusters, responsibilities we want a dedicated platform team to own. Your job is to take ownership of that infrastructure and hide its complexity behind a unified, self-serve platform for running training and inference workloads. Responsibilities Build a self-serve compute platform. Design and own the systems that let inference engineers and researchers launch training jobs and operate inference services without managing GPU provisioning, cluster configuration, or provider-specific infrastructure. Operate the GPU fleet . Own provisioning, lifecycle management, reliability, and capacity integration across providers, giving teams a consistent way to use compute regardless of where it runs. Solve for GPU scarcity. Build the scheduling and placement logic that finds available capacity across providers, packs it efficiently, and gets the right workload onto the right hardware under real constraints. Support two very different workloads. Keep long-running distributed training jobs healthy while simultaneously guaranteeing the availability and latency of production inference services on the same fleet. Own the Kubernetes for GPU orchestration. Write the operators and CRDs, and manage many clusters across providers so the platform behaves the same everywhere we run. Make failure boring. Build the fault tolerance, autoscaling, and observability that keep the fleet utilized and let workloads survive node loss, provider hiccups, and capacity shifts without human intervention. Set technical direction across teams. Partner with inference and cloud infrastructure engineers to turn operational constraints into a coherent platform architecture and roadmap. Qualifications We expect you to have real depth in most of these: Deep Kubernetes experience—custom operators, CRDs, and multi-cluster federation, not just running kubectl apply. You’ve managed GPU clusters at scale: NVIDIA hardware, CUDA, and the networking that makes them fast (InfiniBand or RoCE). You’ve orchestrated compute across multiple clouds (CoreWeave, AWS, GCP, or similar) and understand how different each one really is. Strong distributed systems fundamentals: scheduling, resource allocation, and fault tolerance under load. You write infrastructure and systems-level code in Go, Rust or C++. You’ve supported both long-running training jobs and high-availability inference services, and you know why they pull infrastructure in opposite directions. You own problems end-to-end and do well when the path forward isn’t laid out for you. Additional experience we value Inference serving stacks: vLLM, SGLang, or TensorRT-LLM. Slurm or other HPC schedulers. GPU kernel work in CUDA or Triton—not required, but notable. High-speed interconnects: InfiniBand, RoCE, or RDMA in production. Observability for ML workloads: Prometheus, Grafana, or Weights & Biases. If you’re excited about this role, we encourage you to apply even if your experience doesn’t match every qualification listed above. #J-18808-Ljbffr Neura Market

Vacancy posted 4 days ago
Similar jobs that could be interesting for youBased on the Member of Technical Staff (Software Engineer, Inference & Training Platform) Perplexity AI San [...] in San Francisco, CA vacancy
  • $150k - $300k

     ...anyone to create, train, and deploy...  ...serving, LLM inference optimization...  ...stack. Core Technical...  ...LLM serving platform that operates...  ...LLM Inference engine development and...  ...arrangement (remote or San Francisco office...  ...AI and RL at Prime...  ...encourage team members to contribute... 
    Platform
    Training
    Work at office
    Remote work
    Visa sponsorship
    Relocation package
    Flexible hours
    Shift work

    Prime Intellect

    San Francisco, CA
    20 hours ago
  •  ..., Mexico, San Francisco,...  ...Semiconductor and AI industries...  ...Models, software, and...  ...our deep technical research and...  ...Overview Member of Technical Staff will play...  ...developing training & inference benchmarks...  ...Science, Engineering or other relevant...  ...hardware platforms and large-... 
    Platform
    Training
    Full time
    Work at office
    Remote work
    Worldwide

    S27a

    San Francisco, CA
    2 days ago
  • Member of Technical Staff - Agents at Prime Intellect - San Francisco Building the Future...  ...Decentralized AI At Prime...  ...or capital to train powerful, open...  ...research, and other engineering teams to...  ...Requirements Agent & Platform Skills Python...  ...training or inference on GPUs. Advanced... 
    Platform
    Training
    Remote work
    Flexible hours

    Victrays

    San Francisco, CA
    2 days ago
  •  ...Decentralized AI Development At...  ...at scale. Our platform combines powerful distributed training infrastructure...  ...researchers and engineers to train state‑...  ...systems Core Technical Responsibilities...  ...arrangement (remote or San Francisco...  ...encourage team members to contribute to... 
    Platform
    Training
    Work at office
    Remote work
    Visa sponsorship
    Relocation package
    Flexible hours

    Kubelt

    San Francisco, CA
    1 day ago
  • About Us: AI needs a new infrastructure...  ...serve low-latency inference, fine-tune models,...  ..., and experienced engineering and product leaders...  ...We're building a platform that covers the whole...  ...life of an LLM -- train it, deploy it,...  ...person, in our NYC or San Francisco office.... 
    Platform
    Training
    Work at office

    Mixpeek

    San Francisco, CA
    3 days ago
  • $190.9k - $232.8k

     ...About This RoleAs a staff software engineer for GenAI inference, you will lead the architecture...  ...collaboration: with platform engineers, cloud...  ...certifications and training, and specific work...  ...is the data and AI company. More than 1...  ...is headquartered in San Francisco, with offices... 
    Platform
    Training
    Local area
    Worldwide

    DataBricks

    San Francisco, CA
    3 days ago
  •  ...Pixeltable Inc. Member of Technical Staff San Francisco, CA·Full...  ...founding member of the engineering team, you will...  ...revolutionizing the AI development...  ...our data-centric platform designed to simplify...  ..., transformation, training/fine-tuning, and inference? You will also: Find... 
    Platform
    Training
    Full time
    Part time
    Work at office
    Work from home
    Flexible hours
    2 days per week

    Pixeltable, Inc.

    San Francisco, CA
    20 hours ago
  • $190k - $265k

     ...enabling data and AI teams to solve...  ...infrastructure platform so our...  ...business. Founded by engineers — and customer-...  ...opportunity to solve technical challenges,...  ...agents, model training, model serving,...  ...time and batch inference, powering model...  ...headquartered in San Francisco, with... 
    Platform
    Training
    Local area
    Worldwide

    DataBricks

    San Francisco, CA
    6 days ago
  • $220k - $260k

     ...interaction — a unified platform that combines...  ...enterprise AI agents need in...  ...experienced Staff Software Engineer to help...  ...also shaping the technical culture behind...  ...based in our San Francisco office...  ...education or training to determine individual...  ...with team members around the... 
    Platform
    Training
    Full time
    Work at office

    Redpanda Data

    San Francisco, CA
    20 hours ago
  •  ...era of agentic AI. Millions of people now use Perplexity to transform knowledge...  .... As a growth engineer at Perplexity,...  ...projects from training and...  ...of professional software engineering experience...  ...experimentation platforms (Eppo, Statsig,...  ..., effective technical solutions. Self... 
    Platform
    Training

    aijoblist

    San Francisco, CA
    20 hours ago
  • Member of Technical Staff (Infra) We're looking...  ...experienced Backend Engineer to join...  ...Location San Francisco,...  ...AI infra engineer...  ...infrastructure and inference systems...  ...power our platform, while mentoring...  ...+ years of software engineering...  ...training systems. Familiarity... 
    Platform
    Training
    Full time
    Work at office

    Kindredventures

    San Francisco, CA
    2 days ago
  • Careers / Member of Technical Staff (AI research) Member of Technical...  ...parts of Fearn’s platform from deploying our...  ...) Location San Francisco, California...  ...a crucial role in training models and designing inference pipelines, pushing...  ...Compute Engine, GKE, Vertex AI, or... 
    Platform
    Training
    Full time
    Work at office

    Kindredventures

    San Francisco, CA
    2 days ago
  • Member of Technical Staff, Applied AI The opportunity We are looking...  ..., protein engineers and...  ...protein screening platforms. At Latent Labs...  ...our London and San Francisco...  ...architectures, training dynamics and inference behaviour. You...  ...enterprise software. You have experience... 
    Platform
    Training
    Flexible hours

    Latent Labs

    San Francisco, CA
    1 day ago
  • $150k - $300k

     ...anyone create, train, and deploy...  ...Scientist, Together AI), Dylan Patel...  ...training platform - the product...  ...jobs. Core Technical...  ...training and inference orchestration...  ...looking for engineers who are fluent...  ...arrangement (remote or San Francisco office...  ...team members to contribute... 
    Platform
    Training
    Work at office
    Local area
    Remote work
    Visa sponsorship
    Relocation package
    Flexible hours

    Kubelt

    San Francisco, CA
    1 day ago
  • Perplexity is seeking energetic engineers to join our highly driven Agents engineering team....  ...backend, full-stack, and AI/ML engineers who collaborate...  ...Perplexity Computer (our platform for generalized frontier...  ...of work for our users; Training action and decision... 
    Platform
    Training
    Flexible hours

    Apply

    San Francisco, CA
    4 days ago
  •  ...breakthrough AI application,...  ...today’s data platforms (like Databricks...  ...open-source engine, Daft, is...  ...Databricks and Perplexity, we’re looking...  ...Your Role: As a Member of Technical Staff, you will be...  ...building and training state-of-the-...  ...with inference optimization... 
    Platform
    Training
    Work at office
    Immediate start
    Flexible hours
    Night shift

    Mixpeek

    San Francisco, CA
    20 hours ago
  • Job Description - Member of Technical Staff (Inference) Location: San Francisco (on-site at our offices)...  ...is the leading independent AI benchmarking company. We support labs, engineers and enterprises to...  ...our inference benchmarking platform together with our engineers... 
    Platform
    Shift work

    Artificial Analysis, Inc.

    San Francisco, CA
    4 days ago
  • About Perplexity AI Perplexity is an AI-powered answer engine built to serve the world...  ...that can use software like a human...  ...Role The Data Platform team owns the...  ...In this senior/staff role, you will...  ...the long-term technical direction of Perplexity...  ...features, AI training and evaluation... 
    Platform
    Training

    Perplexity

    San Francisco, CA
    20 hours ago
  • $2,000 per month

     ...About Build AI Build AI is the data hyperscaler...  ..., and model training to scale the...  ...Summary We’re hiring Members of Technical Staff — the general research/engineering seat. If you have...  ...of waiting for a platform team Work with Head...  ...for those moving to San Francisco (Financial... 
    Platform
    Training
    Work at office
    Relocation package

    Build AI

    San Francisco, CA
    3 days ago
  • Member of Technical Staff, Lead Researcher San Francisco, CA; Sunnyvale, CA About...  ...is building an AI Research org from...  ...mentor researchers, engineers, and fellows...  ...DoorDash with ML platform, product, and operations...  ...budgets for training and inference, sized to support... 
    Platform
    Training
    Local area

    DoorDash USA

    San Francisco, CA
    20 hours ago
  • $176.5k

     ...new team members? At Scribd...  ...platform for all content...  ...and our inference-driven features...  ...Senior, Staff, and Principal engineers building...  ...owning software end-to-end...  ...agentic AI tools part...  ...end, from technical design...  ...location. San Francisco...  ...education or training; and... 
    Platform
    Training
    Full time
    For contractors
    Local area
    Home office
    Flexible hours

    Scribd, Inc.

    San Francisco, CA
    20 hours ago
  • $250k - $385k

    Location San Francisco, New York...  ...Department AI Compensation...  ...listed above. Perplexity is seeking...  ...creative, AI-pilled engineers to join our...  ...lead both software and process...  ...possess sound technical judgment forged...  .../platform engineering,...  ...of technical staff across disciplines... 
    Platform
    Full time
    Local area
    Shift work

    Pantera Capital

    San Francisco, CA
    2 days ago
  • $200k - $250k

     ...venture-backed AI startup building...  ...benchmarks used to train computer-using...  ...scaling the founding engineering team. Founded...  ...As a founding Member of Technical Staff on platform engineering, you...  ...training and inference infrastructure that...  ...Location: San Francisco, CA... 
    Platform
    Training
    Full time
    H1b
    Relocation
    Visa sponsorship
    Relocation package

    David Joseph & Company

    San Francisco, CA
    9 days ago
  • $150k - $300k

     ...anyone to create, train, and deploy...  ...the developer platform that makes all...  ...(Eureka AI, Tesla, OpenAI...  ...a generalist software engineering role focused on...  ...improvements Technical Requirements...  ...arrangement (remote or San Francisco...  ...encourage team members to contribute... 
    Platform
    Work at office
    Remote work
    Visa sponsorship
    Relocation package
    Flexible hours

    Prime Intellect

    San Francisco, CA
    20 hours ago
  • $300k

     ...lab developing AI capable of...  ...infrastructure engineers who can build...  ...rollouts, training orchestration, inference, evals, data...  ...the durable platform that enables...  ...Requirements Strong software engineering...  ...by other technical users. Strong...  ...based in our San Francisco office... 
    Platform
    Training
    Work at office
    Local area

    Vmax

    San Francisco, CA
    10 hours ago
  •  ...fast, efficient inference. As AI workloads become...  ...together. Gimlet's platform intelligently partitions...  ...headcount. The engineers we hire today...  ...will be built on software capable of...  ...combination of education, training, and professional...  .... As an early member of the team, you... 
    Platform
    Training

    The Consensus

    San Francisco, CA
    3 days ago
  • $200k

    Member of Technical Staff, Supercomputing Platform & Infrastructure Magic’s mission...  ...-scale pre-training, domain-...  ...context, and inference-time compute to...  ...the role As an engineer on the Supercomputing...  ...and manage AI workloads Develop...  ...for Strong software engineering skills... 
    Platform
    Training
    Relocation
    Visa sponsorship

    Magic AI, Inc

    San Francisco, CA
    3 days ago
  • Introducing Moonlake, AI for creating...  .... Our platform enables the creation...  ...used to train the next generation...  ...looking for a Member of Technical Staff - Robotics to...  ...and engineers developing next...  ...Debug hardware, software, sensing, and...  ...currently based in San Francisco. #J... 
    Platform
    Training

    Moonlake AI

    San Francisco, CA
    2 days ago
  • Location San Francisco; London; New York Employment...  ...On-site Department Engineering Our Mission...  ...states. Our team of AI researchers and company...  ...high-throughput model inference and mid-training workloads. Develop systems...  ...inference platforms capable of serving and... 
    Platform
    Training
    Full time
    Relocation package

    B Capital

    San Francisco, CA
    10 hours ago
  •  ...of machine learners, software engineers, biologists and bioinformaticians...  ...protein screening platforms. At Latent Labs you...  ...minds in generative AI and biology. Our...  ...our London and San Francisco sites. We’...  ...have experience running training and inference on cloud hardware, distributing... 
    Platform
    Training
    Flexible hours

    Latent Labs Ltd.

    San Francisco, CA
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Member of Technical Staff (Software Engineer, Inference & Training Platform) Perplexity AI San [...]. Be the first to apply!