Member of Technical Staff - Training Platform
$150k - $300kKubelt
Building Open Superintelligence Infrastructure Prime Intellect is building the open superintelligence stack - from frontier agentic models to the infrastructure that lets anyone create, train, and deploy them. We aggregate and orchestrate global compute into a single control plane and pair it with the full RL post-training stack: environments, secure sandboxes, verifiable evals, and our async RL trainer. We enable researchers, startups, and enterprises to run end-to-end reinforcement learning at frontier scale, adapting models to real tools, workflows, and deployment contexts. We recently raised $15M in funding (taking total funding to $20M), led by Founders Fund with participation from Menlo Ventures and prominent angels including Andrej Karpathy (Eureka Labs, Tesla, OpenAI), Tri Dao (Chief Scientist, Together AI), Dylan Patel (SemiAnalysis), Clem Delangue (Hugging Face), Emad Mostaque (Stability AI), and many others. Role Impact You'll help build our hosted training platform - the product that lets users launch LoRA and full fine-tuning runs on managed GPU clusters with a single API call or a few clicks. The role spans the developer-facing platform and the underlying Kubernetes-based training infrastructure that runs the jobs. Core Technical Responsibilities Hosted Training Infrastructure Design and operate Kubernetes-based training and inference orchestration across multi-cluster, multi-cloud GPU fleets Build and maintain Helm charts that compose trainers, inference servers, environment servers, and supporting services into reproducible "Training stacks" Develop the Python control-plane agents that watch pods, report run state to the platform, and keep clusters in sync Implement scheduling and autoscaling for heterogeneous hardware (H100/H200/B200) using KEDA, LeaderWorkerSet, taints/tolerations, and gang scheduling Run a tight GitOps workflow - every change ships through PRs, Helm values, and CI Build node-local model caches, checkpoint pipelines, and shared storage for fast cold starts Operate the observability stack (Prometheus, Grafana, Loki, DCGM) and make GPU cluster debugging fast Platform Development Build the developer-facing surfaces for hosted training: job submission, live run monitoring, logs, metrics, model/adapter management, comparisons Develop FastAPI backend services and REST APIs that bridge the platform to running clusters Build real-time monitoring and debugging tools (streaming logs, step-level metrics, failure analysis) Ship product UI in Next.js / React / TypeScript with shadcn, Tailwind, tRPC, and TanStack Query Research Bridge Interface with the RL trainer, inference servers, and environment servers running inside our clusters Productize new training capabilities (new model architectures, RL algorithms, modes) Technical Requirements We're looking for engineers who are fluent across three areas - you don't need to be the world's best at any one, but you should have real depth in all three and a clear point of view on how they connect. AI & GPU Landscape Strong working knowledge of the modern AI stack - open model families, finetuning techniques (LoRA, QLoRA, full FT, RLHF/RLAIF), inference engines (vLLM, SGLang, TensorRT-LLM) Familiarity with GPU hardware tradeoffs (H100 / H200 / B200, NVLink, interconnects, memory hierarchy) and what they mean for training and inference workloads Understanding of distributed training fundamentals (data/tensor/pipeline/expert parallelism, NCCL, multi-node scheduling) Aware of what's happening at the frontier - new models, training methods, infra patterns - and the ability to translate that into product decisions Kubernetes & Infrastructure Strong Kubernetes operations experience - Helm, CRDs, operators, KEDA, gang scheduling, GPU operator Comfortable debugging real production clusters (kubectl, pod lifecycle, node issues, networking) Cloud platform experience (GCP preferred - GCS, GKE, Cloud Run, Cloud Tasks) Infrastructure automation (Helm, Terraform, Ansible) and a GitOps mindset Observability: Prometheus, Grafana, Loki, OpenTelemetry, DCGM Linux fundamentals: networking, namespaces, performance tuning Programming & Platform Strong Python backend development (FastAPI, async, SQLAlchemy) Comfortable building Python control-plane agents that talk to Kubernetes APIs Modern frontend development (TypeScript, React/Next.js, Tailwind, shadcn) - enough to ship product surfaces end-to-end REST and tRPC API design Experience building developer tools, dashboards, and live-monitoring UIs What We Offer Cash compensation $150K–$300K with significant equity Flexible work arrangement (remote or San Francisco office) Full visa sponsorship and relocation support Professional development budget for courses and conferences Regular team off-sites and conference attendance Opportunity to shape the future of decentralized AI development Growth Opportunity You'll join a team of experienced engineers and researchers working on cutting‑edge problems in AI infrastructure. We believe in open development and encourage team members to contribute to the broader AI community through research and open‑source work. We value potential over perfection - if you're passionate about democratizing AI development and have experience in either platform or infrastructure development (ideally both), we want to talk to you. Ready to help shape the future of AI? Apply now and join us in our mission to make powerful AI models accessible to everyone. #J-18808-Ljbffr Kubelt
- ...medalists, and experienced engineering and product leaders with decades of experience. The Role: We're building a platform that covers the whole life of an LLM: training it, deploying it, and observing it in production. We already run multi-node training, elastic inference,...PlatformTechnical trainingWork at office
- ...GPU infrastructure for high-throughput model inference and mid-training workloads. Develop systems that power synthetic data... ...learning pipelines at scale. Build high-performance inference platforms capable of serving and evaluating models across thousands of GPUs...PlatformTechnical trainingWork at officeVisa sponsorship
$200k - $275k
...and data company building computer-use and tool-use RL gyms to train AI agents on real financial services work — Excel modeling, pitch... ...share with the broader research community. Contribute to core platform engineering as needed — AI tooling, data collection pipelines,...PlatformTechnical trainingH1bWork at officeRelocationVisa sponsorship- ...decoupling AI workloads from the underlying hardware. Our platform intelligently partitions workloads into components and orchestrates... ...gigawatt-class AI datacenters. Gimlet Labs is seeking a Member of Technical Staff focused on ML systems and inference. In this role, you...Platform
$150k - $350k
...Member of Technical Staff | Distributed Systems San Francisco - Onsite $150k-$350k base + equity I'm working with a small, deeply technical $... ...help build the foundational infrastructure underneath that platform. You'll work on infrastructure spanning distributed scheduling...Platform- ...that will reshape how people discover and buy online. Role As a Member of Technical Staff, you will ship core systems, set engineering culture, and move the mission from prototype to platform. You will work across the stack and own problems end to end. You...PlatformWork at office
- ...Member of Technical Staff – Machine Learning I’m partnering with a rapidly scaling healthtech startup... ...engineering team. Their AI-powered platform is already helping clinicians by... ...lifecycle: from data pipelines, to model training, to deployment Adapt and fine-tune...PlatformWork at office
- ...Job Description – Member Of Technical Staff (Language Model Evaluations) Location: San Francisco (preferred), Sydney, Melbourne, Brisbane About... ...Analysis Intelligence Index and other areas of our platform, defining how the industry measures frontier capability Publish...Platform
- ...Job Description - Member of Technical Staff Location: San Francisco (preferred), Sydney, Melbourne, Brisbane About Artificial Analysis... ...in the world, and help drive the product direction of our platform. The bar for success is becoming a world expert in modern...Platform
$150k - $350k
...Member of Technical Staff - Distributed Systems San Francisco, CA- 5 days per week onsite $150,000–$350,000 + equity The Opportunity Join a rapidly... ...AI infrastructure company building a multi-silicon cloud platform for fast, efficient inference. The future of AI inference...Platform$200k - $300.09k
...agent relationship while continuously evolving the OpenClaw platform; Maintaining and hardening the codebase to ensure security... .... The Role The OpenClaw Foundation is seeking exceptional Members of Technical Staff (MTS) to serve as full‑time maintainers, builders, and...PlatformFull time- ...We are looking for a Member of Technical Staff with strong Python skills and a passion for building scalable platforms for AI and ML workloads. As MTS, you'll influence strategic decisions, partner closely with the founding team, and play a critical role in shaping Activeloop...Platform
- ..., making us the financial technology platform of choice. At Adyen, everything we do... ...motivated individuals who tackle unique technical challenges at scale and solve them as... ...financial technology sector. As a Member of Technical Staff , you will operate with a high degree...PlatformFlexible hours
- ...Perplexity is seeking an intrepid, polymathic Member of Technical Staff to take on one of the AI industry’s most unique engineering roles. You... ...regulatory issues pertaining to Perplexity. Build automated platforms for harvesting, prioritizing, and patenting Perplexity’s...Platform
- ...frontier models. Millions of domain experts on the platform are paid over $4 million per day to train frontier AI models. Mercor's APEX benchmark... ...Location Preference SF only. About the Role As a Member of Technical Staff on Talent Experience, you'll build the backend systems...PlatformWork at officeRelocation package
- ...Ryan Hoover (Founder, Product Hunt), Charlie Songhurst (Board Member, Meta), and Michael Jones (Former Chair, Huntington Bank Ventures... ...CxOs at leading NBFIs and FIs to deploy and integrate Krew's platform Support with change management and PMO objectives Develop software...PlatformFull timeWork experience placementInternshipWorldwide
- ...are responsible for designing, building, and scaling core infrastructure that powers a high-volume data platform for AI applications. We are looking for team members who love building enabling systems that empower our engineers and power our rapidly growing product. We’...PlatformWork at office
$240k - $300k
...Member of Technical Staff Global Placement Firm is conducting a confidential search for a high-performing Member of Technical Staff to join a... ...Experience building systems that integrate with external platforms, APIs, or complex customer environments. Willingness to work...PlatformFull timeH1bWork at officeLocal areaRelocationVisa sponsorship$200k
...Join to apply for the Member of Technical Staff role at Listen Labs . TL;DR: We are seeing strong market demand and an aggressive 6‑month... .... Background: Listen Labs is an AI‑powered research platform that helps teams uncover insights from customer interviews...PlatformFlexible hours- ...We’ve redesigned the foundation model training stack to turn the world’s raw... ...responsibility to defend. About the role As a Member of Technical Staff, ML Product Engineer, this role owns... ...developer experience feels like. The platform you build is how Omnii reaches the...PlatformLocal area
- ...Member of Technical Staff – Full Stack / AI Systems Company : AdsGency AI Relocation : San Francisco City Required Authorization : Applicants... ...that orchestrate multi‑agent ad automation across global platforms. You’ll work shoulder‑to‑shoulder with founders and senior...PlatformFull timeWork experience placementRelocationVisa sponsorship
- ...human attention, and the right software platform can lift that ceiling by an order of... ..., and Ramp. About the Role Members of Technical Staff (MTS) are the senior engineers who build... ...around a model someone else trained, and have an informed opinion on where...Platform
$150k
...possible in robotic intelligence. As a Member of Technical Staff, you'll be at the forefront of... ...practical implementation, collaborating with platform teams to ensure your models and algorithms... ...and rich real‑world datasets to train and deploy state‑of‑the‑art foundation...PlatformLocal area- ...fastest-growing and most trusted legal AI platform for in-house legal teams. We're building... ...— contribute to engineering culture and technical direction in a company where those... ...GC AI is a distributed company with team members across North America, and soon Europe. We...PlatformPermanent employmentCurrently hiringWork at officeImmediate startShift work
$275k - $315k
...and distillation down to smaller and edge-deployable models. We're hiring a Member of Technical Staff to own the modeling side of that end to end. What You’ll Do You'll own the post-training pipeline end-to-end: data curation, SFT, preference optimization, RL, evals,...Technical trainingFull timeWork at officeRelocation package- Member of Technical Staff - Post‑Training Join to apply for the Member of Technical Staff - Post‑Training role at Reflection AI . Our Mission Reflection’s mission is to build open superintelligence and make it accessible to all. We’re developing open weight models for...Technical trainingFull timeRelocation package
- ...mission is to scale intelligence to serve humanity. We’re training and deploying frontier models for developers and enterprises... ...New York but also embrace being remote-friendly! As a Member of Technical Staff, you will: Design and write high-performant and scalable...Technical trainingFull timeWork at officeRemote workFlexible hours
$200k - $350k
Member of Technical Staff — LLM Research & Training About the Role We are looking for an exceptional Member of Technical Staff specializing in Machine Learning and Large Language Models to join an early-stage AI company building and training state‑of‑the‑art foundation...Technical trainingH1bVisa sponsorship- Pixeltable, Inc. is seeking a Member of Technical Staff based in San Francisco, CA. As a founding member of our engineering team, you will directly... ...the design and development of a revolutionary AI data platform. With over 5 years of experience in systems engineering...PlatformFlexible hours
$150k - $350k
...decoupling AI workloads from the underlying hardware. Our platform intelligently partitions workloads into components and orchestrates... ...‑class AI datacenters. Mission Gimlet Labs is seeking a Member of Technical Staff focused on kernels and GPU performance. In this role, you...Platform
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Member of Technical Staff - Training Platform. Be the first to apply!
- senior technical analyst San Francisco, CA
- systems support technician San Francisco, CA
- senior help desk analyst San Francisco, CA
- help desk technical support San Francisco, CA
- trade support analyst San Francisco, CA
- support analyst San Francisco, CA
- technical analyst San Francisco, CA
- technical support assistant San Francisco, CA
- help desk assistant San Francisco, CA
- IT assistant San Francisco, CA

