Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Software Engineer, GPT Infrastructure

$293k - $385k

OpenAI

About the TeamThe GPT Infrastructure team builds systems that turn advances in model inference and optimization into reliable production capabilities. We enable OpenAI workloads to be qualified and optimized across new accelerator platforms without requiring a one-off port and tuning effort for every hardware target.Our work spans distributed systems, model execution, compilers and runtimes, performance engineering, secure partner integrations, evaluation systems, and developer tooling. We build the infrastructure that makes optimization workflows automated, reproducible, and trustworthy.About the RoleWe are seeking a software engineer to help build the platform that qualifies and optimizes inference workloads across heterogeneous compute environments.You will develop both OpenAI-hosted services and secure partner-side software for running long-lived optimization workflows. These workflows generate candidate kernels, runtime configurations, and serving-stack changes; compile and execute them on target hardware; verify their correctness; measure their performance; and use the results to guide further optimization.You will work across model architecture, distributed execution, compilers, runtimes, networking, and accelerator systems. A central part of the role is turning research prototypes and one-off hardware bring-up efforts into reliable, reusable infrastructure with clear contracts, reproducible results, strong observability, and well-defined security boundaries.Key ResponsibilitiesDesign, build, and operate APIs and control-plane services for long-running workload qualification and optimization campaigns, including scheduling, retries, checkpointing, resource budgets, and observability.Build secure partner-side execution and evaluation software that can compile, run, verify, profile, and benchmark candidate artifacts on accelerator hardware.Integrate model workloads, hardware profiles, compiler toolchains, runtimes, serving engines, and distributed-execution backends into a repeatable platform.Develop correctness and performance evaluation systems spanning output fidelity, latency, throughput, memory footprint, accelerator utilization, communication efficiency, scaling behavior, and cost efficiency.Automate the generation, evaluation, and improvement of kernels, runtime configurations, parallelization strategies, and serving-stack changes.Diagnose performance and correctness issues across model code, kernels, compilers, runtimes, memory systems, networking, collective communication, and hardware.Build artifact-management, provenance, regression-testing, and qualification workflows for kernels, binaries, configurations, evaluation results, and deployment reports.Turn experimental research workflows into reliable product surfaces with clear interfaces, actionable failure modes, and strong developer ergonomics.Collaborate with Research, Inference Engineering, Runtime and Compiler teams, Infrastructure, Security, Product, and Strategic Partnerships to onboard and optimize new compute platforms.Drive technical architecture and execution across ambiguous initiatives spanning OpenAI systems and partner environments.QualificationsStrong software engineering experience building distributed systems, infrastructure platforms, production services, developer platforms, or orchestration systems.Proficiency in one or more systems-oriented languages such as Python, C++, Go, or Rust.Experience designing and operating APIs, job orchestration systems, durable workflows, or large-scale backend services.Strong understanding of Linux, networking, storage, containers, distributed execution, and modern infrastructure architectures.Ability to reason about model execution and diagnose problems across software and hardware boundaries.Experience using profiling, tracing, benchmarking, and measurement to guide engineering decisions.Strong ownership and the ability to work effectively across research, engineering, security, product, and external-partner teams.Preferred SkillsExperience with AI infrastructure, model inference, distributed ML systems, or inference-serving platforms.Familiarity with GPU or accelerator architecture, memory hierarchies, interconnects, collective communication, and distributed model execution.Experience with compilers, runtimes, kernels, or performance engineering using technologies such as CUDA, ROCm, Triton, LLVM, or MLIR.Familiarity with inference engines or serving systems such as vLLM, SGLang, Triton Inference Server, or comparable internal systems.Experience with model partitioning, sharding, tensor or expert parallelism, and compute–communication tradeoffs.Experience building remote-execution systems, secure partner-facing infrastructure, evaluation harnesses, or artifact pipelines.Experience with automated optimization, search systems, coding agents, or evaluator-driven systems that iteratively improve kernels or runtime configurations.About OpenAIOpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity. We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic. For additional information, please see OpenAI’s Affirmative Action and Equal Employment Opportunity Policy Statement.Background checks for applicants will be administered in accordance with applicable law, and qualified applicants with arrest or conviction records will be considered for employment consistent with those laws, including the San Francisco Fair Chance Ordinance, the Los Angeles County Fair Chance Ordinance for Employers, and the California Fair Chance Act, for US-based candidates. For unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. In addition, job duties require access to secure and protected information technology systems and related data security obligations.To notify OpenAI that you believe this job posting is non-compliant, please submit a report through this form. No response will be provided to inquiries unrelated to job posting compliance.We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made via this link.OpenAI Global Applicant Privacy PolicyAt OpenAI, we believe artificial intelligence has the potential to help people solve immense global challenges, and we want the upside of AI to be widely shared. Join us in shaping the future of technology.Compensation Range: $293K - $385KLocationSan Francisco; SeattleEmployment TypeFull timeLocation TypeHybridDepartmentScalingCompensation$293K – $385K • Offers EquityThe base pay offered may vary depending on multiple individualized factors, including market location, job-related knowledge, skills, and experience. If the role is non-exempt, overtime pay will be provided consistent with applicable laws. In addition to the salary range listed above, total compensation also includes generous equity, performance-related bonus(es) for eligible employees, and the following benefits.Medical, dental, and vision insurance for you and your family, with employer contributions to Health Savings AccountsPre-tax accounts for Health FSA, Dependent Care FSA, and commuter expenses (parking and transit)401(k) retirement plan with employer matchPaid parental leave (up to 24 weeks for birth parents and 20 weeks for non-birthing parents), plus paid medical and caregiver leave (up to 8 weeks)Paid time off: flexible PTO for exempt employees and up to 15 days annually for non-exempt employees13+ paid company holidays, and multiple paid coordinated company office closures throughout the year for focus and recharge, plus paid sick or safe time (1 hour per 30 hours worked, or more, as required by applicable state or local law)Mental health and wellness supportEmployer-paid basic life and disability coverageAnnual learning and development stipend to fuel your professional growthDaily meals in our offices, and meal delivery credits as eligibleRelocation support for eligible employeesAdditional taxable fringe benefits, such as charitable donation matching and wellness stipends, may also be provided.More details about our benefits are available to candidates during the hiring process.This role is at-will and OpenAI reserves the right to modify base pay and other compensation components at any time based on individual performance, team or company results, or market conditions.

Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the Software Engineer, GPT Infrastructure in San Francisco, CA vacancy
  •  ...About the Role This role broadly owns infrastructure across the stack. If it’s running in the cloud, you probably care...  ...Robotics ), launched and scaled ChatGPT and GPT-4 to hundreds of millions of users, engineered the foundations of autonomous driving, built next-generation... 
    Suggested
    Full time

    The Generalist

    San Francisco, CA
    1 day ago
  •  ..., and Google Lens. Before that, Clay led the product and design teams for Google Workspace.  What you’ll do As a Software Engineer, Infrastructure at Sierra, you will be responsible for designing, building, and maintaining the core systems that make our AI platform... 
    Suggested
    Full time
    Flexible hours

    Sierra Limited

    San Francisco, CA
    1 day ago
  • $100k - $300k

     ...Abnormal AI, Zscaler Preeminent research labs like Deepmind and SAIL About the Role We're hiring a Senior Storage Infrastructure Engineer on our Core Platform team to own how we store, protect, and operate data at scale. You'll build the backup, observability,... 
    Suggested
    Full time

    Cogent Security

    San Francisco, CA
    1 day ago
  •  ...an experienced, creative, and versatile engineer who is eager to tackle the challenge of...  ..., building, and scaling the core infrastructure and systems powering Lightfield's AI-driven...  ...Who you are ~3+ years of experience in software development, with a strong background... 
    Suggested
    Full time
    Work from home

    Lightfield

    San Francisco, CA
    1 day ago
  • $190k - $260k

     ...Notion. Centralize was founded by Rachit Kataria, a founding engineer on Facebook Shops who helped scale it to 250M MAUs, and Will...  ...what Centralize becomes. The Role We are hiring an infrastructure engineer to own scalability across Centralize. Our system processes... 
    Suggested
    Remote job
    Full time
    Relocation
    Visa sponsorship

    Centralize

    San Francisco, CA
    1 day ago
  • $137.1k - $201.6k

     ...Learning Platform and builds the shared infrastructure that helps DoorDash, Wolt, and...  ...direction across model serving and inference engines, fine-tuning and training pipelines, GPU...  ...~6+ years of industry experience in software engineering ~ Deep backend engineering... 
    Hourly pay
    Work at office
    Local area
    Remote work
    Flexible hours

    Doordash

    San Francisco, CA
    19 hours ago
  • $160k - $180k

     ...out before, during, and after playing games. The Database Infrastructure team develops and operates all of Discord’s databases and data...  ...matters most to our users. Work with a talented team of engineers who have built one of the largest communication platforms in... 
    Full time

    Discord

    San Francisco, CA
    1 day ago
  •  ...including stablecoins. You’ll help design, deploy and operate the infrastructure that makes this possible. This is a hands-on devops role...  ...role sits at the intersection of infrastructure, platform engineering and security. You’ll define how Modern Treasury scales and... 
    Full time
    Local area
    Immediate start
    Remote work

    Modern Treasury

    San Francisco, CA
    1 day ago
  •  ...About the Team We’re hiring software engineers to join our broader Infrastructure organization, which supports multiple high-impact teams. Depending on your interests and experience, you could work on one of several focus areas—including Core Distributed Systems, Databases... 
    Full time

    OpenAI

    San Francisco, CA
    1 day ago
  • $130k - $175k

     ...goal is to enable and empower Kiddom’s engineering by building a scalable and sustainable...  ...new and existing services. Practicing Infrastructure as Code (IaC) wherever possible, giving...  ...related field ~5+ years professional software engineering experience ~ Experience... 
    Permanent employment
    Full time
    Work at office
    Local area
    Remote work

    Kiddom

    San Francisco, CA
    1 day ago
  • $145.4k - $188.1k

     ...together. About the Machine Learning Infrastructure Team At Thumbtack, our challenges...  ...these experiences. We empower product engineering teams by providing scalable, high-performance...  ...blog . The challenge As a Software Engineer on the ML Infrastructure team,... 
    Full time
    Seasonal work
    Local area

    Thumbtack

    San Francisco, CA
    1 day ago
  •  ...Orb: Orb is redefining what billing software can be, turning one of the most complex...  ...the role: As a senior member of the infrastructure team, you’ll play a key role in maintaining...  ...customer growth Partner with other engineering teams to ensure we build reliable... 
    Full time
    Work at office
    Remote work
    3 days per week

    Orb

    San Francisco, CA
    1 day ago
  • $140k - $225k

     .... We’re assembling a diverse, world-class team—engineers, designers, researchers, and product minds—focused...  ...their best work. About The Role As the Senior Software Engineer, Tooling and Development Infrastructure, you will play a critical role in shaping the... 
    Full time
    Temporary work
    Local area
    Flexible hours

    HP IQ

    San Francisco, CA
    1 day ago
  • $190k - $280k

     ...About Sentry Bad software is everywhere, and we’re tired of it. Sentry is on a mission to help developers write better software...  ...and an informative deployment pipeline. As an engineer on the Infrastructure Engineering team, you’ll help deliver on our mission: We... 
    Hourly pay
    Full time
    Work at office

    Sentry

    San Francisco, CA
    1 day ago
  • $170k - $216k

     ...across 15+ U.S. states. The Simulation Infrastructure team creates reliable, scalable, and...  ...products that evaluate the Waymo Driver's software stack at a massive scale. We solve...  ...for a broad range of customers Software Engineers, Product, Data Science, System Engineering... 
    Full time
    Remote work

    Waymo

    San Francisco, CA
    1 day ago
  •  ...teams for Google Workspace.  What you'll do The Payments Infrastructure team builds the trust boundary between a live conversation...  ...plaintext cardholder data. Make payments something other engineers can use without becoming compliance experts: drive the platform... 
    Full time
    Flexible hours

    Sierra Limited

    San Francisco, CA
    1 day ago
  •  ...Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable...  ...Conviction. Join us and help build the platform engineers turn to to ship AI products. THE ROLE As a Software Engineer at on the Training Infrastructure... 
    Full time
    Flexible hours

    Baseten

    San Francisco, CA
    1 day ago
  • $160k - $220k

     ...Horowitz to Blackrock and Fidelity, and employs a team of 450 engineers and entrepreneurs. Astranis designs, builds, and...  .... ft. headquarters in Northern California, USA. Senior Software Engineer - Infrastructure As a Senior Software Engineer on the Infrastructure... 
    Permanent employment
    Full time
    Remote work
    Flexible hours

    Astranis

    San Francisco, CA
    19 hours ago
  •  ...reliability, performance, and security for multi‑tenant compute. What You’ll Do Design and operate secure, multi‑tenant container infrastructure with fast startup and smart autoscaling. Ship cloud deployments (Helm/Terraform) with SSO, network controls, and audit... 
    Full time
    Remote work

    Julius Ai

    San Francisco, CA
    1 day ago
  •  ...coding agents that replace traditional software development by generating, testing, and...  ...You'll Be Responsible For Platform & Infrastructure Maintain stability of our platform...  ...~4+ years of software/platform engineering experience with production systems ~... 
    Full time
    Flexible hours

    Emergent Labs

    San Francisco, CA
    1 day ago
  •  ...About Flow Flow Engineering is an AI-native requirements platform for modern engineering organizations, enabling hardware...  ...and speed.​ About the role Flow is hiring a Software Engineer with an infrastructure focus to build and scale the core platform behind Flow... 
    Full time
    Flexible hours

    Flow Engineering

    San Francisco, CA
    1 day ago
  • $180k - $250k

     ...powering the next generation of AI products. We build the infrastructure, tools, and model access that teams need to move from...  ...that ambitious teams build on. You are a hands-on engineer who builds the software and processes that keep a large fleet of GPU servers healthy... 
    Full time
    Local area
    Relocation package

    Falò

    San Francisco, CA
    1 day ago
  •  ...goals. Whoever deploys frontier compute infrastructure fastest will decide whether AI expands...  ...them - with teams spanning hardware and software. Speed and scale are our key...  ...matters to the world. The Production Engineering Team Examples of key exciting problems... 
    Full time
    Local area

    Fluidstack

    San Francisco, CA
    1 day ago
  •  ...About the Role We are seeking a Cloud Infrastructure Engineer to help design and evolve the platforms that power OpenAI’s products. In this role, you will be a hands-on technical leader, driving the architecture, scalability, reliability, and security of critical... 
    Full time

    OpenAI

    San Francisco, CA
    1 day ago
  • $250k

     ...product-market fit with a substantial customer pipeline already in place.   Role Overview We’re looking for an ML infrastructure engineer to design and build the core systems that enable scalable, efficient training of large models for deployment and research.... 
    Full time

    Epsilon Labs, Inc.

    San Francisco, CA
    1 day ago
  •  ...government agencies address real-world challenges. The Infrastructure Engineering team is crucial to the overall success of Hayden products:...  ...Uphold Engineering Standards: Establish and enforce elite software engineering and DevOps standards, including rigorous... 
    Full time
    Shift work

    Hayden Ai

    San Francisco, CA
    1 day ago
  •  ...into real, working systems and build any software needed for running large-scale frontier...  ...About the Role We are looking for engineers to operate the next generation of compute...  ...systems engineering with hands-on infrastructure work on our largest datacenters. You will... 
    Full time

    OpenAI

    San Francisco, CA
    1 day ago
  • $209k - $240k

     ...office workdays. About the Product Infrastructure Team: The Product Infrastructure...  ...classes of problems up-front for product engineers. Solve hard technical challenges such...  ...values, and enthusiastic about making software toolmaking ubiquitous, we want to hear... 
    Full time
    Work at office
    Local area

    Notion

    San Francisco, CA
    1 day ago
  •  ...of the great threats AI presents: mass-manufactured social engineering. Countless scams, deepfakes, and other social engineering attacks...  ...for an experienced backend engineer to build out the infrastructure required to rapidly scale up our engineering organization. A... 
    Full time
    Work at office
    Flexible hours

    Doppel

    San Francisco, CA
    1 day ago
  •  ...business with billions in revenue The Role Handshake is building the infrastructure layer that powers the next generation of AI agents across our platform. As a Senior Software Engineer on our Agentic Infrastructure team, you'll be at forefront of AI at... 
    Full time
    Freelance
    Internship
    Work at office
    Remote work
    Flexible hours

    Handshake

    San Francisco, CA
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Software Engineer, GPT Infrastructure. Be the first to apply!