Member of Technical Staff, Reinforcement Learning Infra
$200k - $350kInception
Inception creates the world’s fastest, most efficient AI models. Our Mercury model is the world’s fastest reasoning LLM and first commercially available diffusion LLM, delivering 5x greater speed and efficiency than today’s LLMs, with best-in-class quality. We are the AI researchers and engineers behind such breakthrough AI technologies as diffusion models, flash attention, and DPO. The Role We're looking for engineers and scientists to design, optimize, and maintain the core systems that enable scalable, efficient reinforcement learning for large models. This role sits at the intersection of research and large-scale systems engineering: you'll wear many hats, from optimizing rollout and reward pipelines to enhancing reliability, observability, and orchestration, collaborating closely with researchers to make RL stable, fast, and production-ready. Key Responsibilities Design, build, and optimize the infrastructure that powers large-scale reinforcement learning and post-training workloads. Improve the reliability and scalability of RL training pipelines, distributed RL workloads, and training throughput. Develop shared monitoring and observability tools to ensure high uptime, debuggability, and reproducibility for RL systems. Qualifications BS/MS/PhD in Computer Science, Engineering, or a related field (or equivalent experience). Understanding of ML frameworks (PyTorch, TensorFlow, Ray, Megatron) from a systems perspective. Experience working with reinforcement learning workloads (PPO, DPO, RLHF, or reward modeling). Experience with containerization (Docker), orchestration (Kubernetes), and CI/CD pipelines. Preferred Skills Experience building and maintaining large-scale language models with tens of billions of parameters or more. Experience with ML workflow orchestration tools (Kubeflow, Airflow). Background in performance optimization and profiling of ML systems. Compensation The annual base salary range for this role is $200,000 – $350,000 USD. Final compensation is determined based on experience, skills, and qualifications. Equity and benefits are included in the total package. Why Join Inception Work with World-Class Talent : Collaborate with the inventors of diffusion models and leading AI researchers Shape Foundational Technology: Your decisions will influence how the next generation of AI products are built and used Immediate Impact: Join at the ground floor where your contributions directly shape product direction and company trajectory Competitive salary and equity in a rapidly growing startup Flexible vacation and paid time off (PTO) Health, dental, and vision insurance Catered meals (breakfast, lunch, & dinner) A collaborative and inclusive culture About Us Inception creates the world’s fastest, most efficient AI models. Today’s autoregressive LLMs generate tokens sequentially, which makes them painfully slow and expensive. Inception’s diffusion-based LLMs (dLLMs) generate answers in parallel. They are 5x faster and more efficient, while delivering best-in-class quality. Inception was co-founded by Stanford professor Stefano Ermon, who co-invented such breakthrough AI technologies as diffusion models, flash attention, and DPO, UCLA professor Aditya Grover, who co-invented node2vec, decision transformers, and d1 reasoning, and Cornell professor and Afresh co-founder Volodymyr Kuleshov, who co-invented MDLM and Block Diffusion. We pioneered the application of diffusion to language, with world’s first (and only) commercially available dLLM, Mercury. We are currently deploying our large-scale diffusion LLMs at Fortune 500 companies. Diffusion is the technology behind today’s image and video AI, and we’re making it the standard for LLMs as well. Our team includes engineers from AWS, Google DeepMind, Meta AI, Microsoft, HashiCorp, and OpenAI. Based in Palo Alto, CA, we are backed by top-tier venture capitalists, including Menlo Ventures, Mayfield, M12 (Microsoft’s venture fund), Snowflake Ventures, Databricks, and Innovation Endeavors, and by tech luminaries such as Andrew Ng, Andrej Karpathy, and Eric Schmidt. If you are talented, innovative, and ambitious, come help us invent the future of AI. We are an equal opportunity employer and encourage candidates of all backgrounds to apply. #J-18808-Ljbffr Inception
- We’re looking for candidates with experience building reinforcement learning-based LLM training pipelines. As Part Of Our Founding Team You May Train reinforcement learning-based LLMs to solve tasks in the domain of materials science, chemical engineering, and engineering...Suggested
- Position Summary You'll advance reinforcement learning for manipulation alongside a world-class team—using RL to push our robot policies on contact... ...team of exceptional professionals who combine world-class technical skills with creative vision, grounded in humility and...SuggestedWork from homeFlexible hours
- Member of Technical Staff (Infra) We're looking for an experienced Backend Engineer to join Fearn as we build up a scalable infrastructure. Role... .... Introductory chat: We will tell you a bit about Fearn, learn about you and your prior experiences, and answer any questions...SuggestedFull timeWork at office
$150k - $300k
...frontier agentic models to the infra that enables anyone to... ...enterprises to run end‑to‑end reinforcement learning at frontier scale, adapting... ...RL training stack. Core Technical Responsibilities LLM Serving... ...and encourage team members to contribute to the broader...SuggestedWork at officeRemote workVisa sponsorshipRelocation packageFlexible hoursShift work- ...re looking for engineers and scientists to design, optimize, and maintain the core systems that enable scalable, efficient reinforcement learning for large models. This role sits at the intersection of research and large-scale systems engineering: you\'ll wear many hats...Suggested
- ...push AI closer to achieving its transformative potential. About the Role We’re hiring new graduate Machine Learning Engineers to design and build reinforcement learning environments to safely advance model capabilities specifically on machine learning research and engineering...Visa sponsorshipRelocation package
- Member of Technical Staff - Applied AI Research San Francisco, CA; Sunnyvale, CA DoorDash’s mission is to empower local... ..., measuring performance Strong knowledge of reinforcement learning techniques Experience building infra for large‑scale reinforcement learning and multi...Hourly payWork at officeLocal areaImmediate startRemote workFlexible hours
$150k - $300k
...enterprises to run end-to-end reinforcement learning at frontier scale, adapting... ...that runs the jobs. Core Technical Responsibilities Hosted... ...models, training methods, infra patterns - and the ability... ...development and encourage team members to contribute to the broader...Work at officeLocal areaRemote workVisa sponsorshipRelocation packageFlexible hours- ...for high-throughput model inference and mid-training workloads. Develop systems that power synthetic data generation and reinforcement learning pipelines at scale. Build high-performance inference platforms capable of serving and evaluating models across thousands of...Work at officeVisa sponsorship
- Member of Technical Staff - Post‑Training Join to apply for the Member of Technical Staff - Post... ...generation pipelines, reward models, reinforcement learning algorithms, and inference‑time... ...to work fluidly across research and infra boundaries. Strong communication capabilities...Full timeRelocation package
$150k - $300k
...frontier agentic models to the infra that enables anyone to... ...enterprises to run end-to-end reinforcement learning at frontier scale, adapting... ...reliable at scale. Core Technical Responsibilities Infrastructure... ...and encourage team members to contribute to the broader...Work at officeRemote workVisa sponsorshipRelocation packageFlexible hours$200k - $300k
...Be humble - Prioritize growth, learning, and doing whatever is needed to further... ...The Role Nimble is looking for a Member of Technical Staff to help us advance our robotics... ...general-purpose robotic AI models, and reinforcement learning algorithms controlling multi...Local areaImmediate startFlexible hoursWeekend work- ...person can. We're looking for founding members of technical staff: engineers who will co-own Hivemind's... ...answer. Build and scale systems that learn continually from many users' feedback... ...relationships. Work across backend, AI infra, data, and devops as the work demand,...Full time
- ...DoorDash. The Token Company trains machine learning models to compress raw LLM inputs... ...a research and product focus. As a Member of Technical Staff on our infrastructure team, you'll own... ...latency, high-throughput GPU ML inference infra that sits in the critical path of...Visa sponsorship
$150k - $300k
...- from frontier agentic models to the infra that enables anyone to create, train,... ...startups and enterprises to run end-to-end reinforcement learning at frontier scale, adapting models to... ...for GPU Infrastructure, you'll be the technical expert who transforms customer...- ...into new problems. About the Role As a Member of Technical Staff at Phonic, you'll work across the full... ...your strengths and interests point — infra one week, a research prototype or a customer... ...spec. Execution speed: you ship, learn, and iterate quickly rather than...Work at officeShift work
- Traverse is a research data lab building reinforcement learning environments for frontier AI labs. We focus on the non-deterministic,... ...thought partner. Backed by Y Combinator. About the Role As a Member of Technical Staff, you will build the core infrastructure and...
- ...next AlphaFold breakthrough. As a Member of Technical Staff, you will build this autonomous AI research... ...in ML. A graduate degree in Machine Learning, CS, Math, Stats, Physics, or... ...Evolutionary algorithms, meta-learning, reinforcement learning, or related methods for...
$200k
...many of the company’s most important decisions. As a Member of Technical Staff on Evals, you will build both the platform and the evaluations... ...validate eval tasks for pre-training, post-training, reinforcement learning, inference, and product systems Develop infrastructure...Visa sponsorshipRelocation package$220k
...foundational model capabilities, post-training techniques, building RL infra and infrastructure that benefits the entire organization.... ...Post-train SOTA LLMs using the latest supervised and reinforcement learning techniques (SFT/DPO/GRPO) Leverage our rich query/answer...Full time- ...conferences like CVPR , and moves with the speed and clarity of a startup obsessed with impact. About the Role We’re seeking a Machine Learning Engineer who thrives at the frontier of foundation‑model research and production engineering . You’ll help define how machines...Immediate start
$200k - $350k
Member of Technical Staff — LLM Research & Training About the Role We are looking for an exceptional... ...Staff specializing in Machine Learning and Large Language Models to join an... ...research Pre‑training Post‑training Reinforcement learning / preference optimization Training...H1bVisa sponsorship- ...early, own problems end-to-end, and watch your work ship directly into the models defining the frontier. About the Role As a Machine Learning Engineer at Sieve, you\'ll own the entire ML lifecycle — from understanding customer problems, to designing datasets, improving...
- Pixeltable, Inc. is seeking a Member of Technical Staff based in San Francisco, CA. As a founding member of our engineering team, you will directly influence the design and development of a revolutionary AI data platform. With over 5 years of experience in systems engineering...Flexible hours
- ...architecture reviews for other members of the team Help establish... ...meets their needs Requirements Technical Strong engineering fundamentals... ...a React frontend. All of the infra is on AWS using CDK for IaC. What We’re Looking For Learning velocity: The role encompasses...Work experience placementRelocationRelocation packageShift work
- ...ideas to work What we’re looking for We value a relentless approach to problem-solving, rapid execution, and the ability to quickly learn in unfamiliar domains. Strong grasp of state-of-the‑art techniques for optimizing training and inference workloads Demonstrated...
$200k
Member of Technical Staff, Supercomputing Platform & Infrastructure Magic’s mission is to build safe AGI... ...with experience building production infra systems Deep, hands-on experience with... ...can do their best work. We value quick learning and grit just as much as skill and...RelocationVisa sponsorship- Careers / Member of Technical Staff (AI research) Member of Technical Staff (AI research) You will... ...production. Staying close enough to infra to get your own work into production.... ...: We will tell you a bit about Fearn, learn about you and your prior experiences,...Full timeWork at office
- ...to deployment is what wins. Role: Member of Technical Staff, Research Why this role matters We're... ...sim-first workflow and self-improving learning on real cells. Required Education.... .... Build Imitation Learning and Reinforcement Learning models. Speaks to architecture...Full timeRelocation packageShift work
- ...WORKING LINKEDIN PROFILE WILL NOT BE CONSIDERED! PLEASE, ADD IT TO YOUR CV OR APPLICATION. THANK YOU. We're looking for a Machine Learning Engineer to own the ML lifecycle end to end - from understanding customer problems through designing datasets, improving models,...Full timeImmediate start
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Member of Technical Staff, Reinforcement Learning Infra. Be the first to apply!
- operations support technician San Francisco, CA
- product support technician San Francisco, CA
- senior technical analyst San Francisco, CA
- systems support technician San Francisco, CA
- user support analyst San Francisco, CA
- technical support specialist San Francisco, CA
- remote support technician San Francisco, CA
- help desk assistant San Francisco, CA
- personal computer support technician San Francisco, CA
- mri tech aide San Francisco, CA

