Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Principal AI Product Engineer

$290k
Full-time

Nscale

About Nscale

Nscale is taking on the hyperscalers by building a vertically integrated GenAI cloud platform. We own the data centers, software, and applications that power today's AI stack using sustainable technology solutions. We thrive on a culture of relentless innovation, ownership, and accountability, where every team member takes pride in their work and drives it with excellence and urgency. As a Nscaler, you'll build trust through openness and transparency, where everyone is inspired to do their best work. Collaboration is key, and we work together swiftly and respectfully, embracing adaptability and resilience in all we do.

About the Role

Nscale is looking for a Principal AI Engineer (Specialised) to lead the inference and post-training pillar of our AI systems engineering organization. You’ll define the multi-year technical roadmap for how models are served, evaluated, and post-trained on Nscale’s GPU cloud, across dedicated and serverless inference and bring-your-own-model deployments. You’ll lead the most consequential architectural programmes in that space and set the engineering standards that 20–50+ engineers build to.

As a Principal engineer, you are one of the deepest technical authorities in the company on AI systems. Your decisions set the cost, latency, and reliability at which Nscale serves tokens and runs post-training workloads, and those numbers have to compete with the world’s leading AI infrastructure providers. The problems span the full stack: kernel efficiency on state-of-the-art GPU systems, fleet-level KV cache and serving architecture, the evals that prove model quality, and RL loops where inference and training share hardware. You frame the solutions the organization executes against, including the API contracts customers see.

How We Work

  • Dog years. We move quickly and compress a lot of learning into a short time.
  • Don’t let perfect be the enemy of good. Ship, measure, iterate.
  • Be relentless. Own the problem end to end and see it through.
  • One team, one mission. Outcomes over process, and no “not my job”.

Responsibilities

  • Define and own the multi-year technical roadmap for Nscale’s inference, evals, and post-training platform, and translate it into architecture that multiple teams can execute against
  • Lead company-scale architectural initiatives in the pillar, such as next-generation serving (disaggregated prefill/decode, KV cache orchestration across GPU, host, and storage tiers, speculative decoding, multi-tenant scheduling), GPU kernel and model efficiency work (custom kernels, FP8/NVFP4/INT8/4 quantization, sparsity, distillation, MoE serving), evals and benchmarking frameworks, and post-training and RL infrastructure
  • Establish engineering standards adopted across all AI teams: API design and compatibility guarantees, benchmarking and evals methodology, training stability norms, and performance testing practices
  • Own the framework by which cost, latency, throughput, and model quality trade-offs are made and measured across the pillar
  • Identify long-horizon systemic risks early (serving engine and framework bets, accelerator support, capability gaps) and resolve them before they block the organization
  • Align AI engineering, research, product, and infrastructure leadership on multi-team technical strategy; frame technical trade-offs in product and commercial terms
  • Mentor and develop Staff and Senior AI Engineers, and grow the next generation of inference technical leaders at Nscale
  • Represent Nscale’s technical approach externally: open-source leadership in the frameworks we depend on, publications, conference talks, and partnerships with GPU vendors and AI labs

Requirements

  • 10–15 years of engineering experience, with a clear track record of pillar-level impact on production AI systems
  • 4+ years of hands-on work with LLMs in inference, GPU performance, evals, or post-training and RL, in production or research
  • Demonstrated ability to define multi-year technical strategy for complex, multi-team AI systems organizations
  • World-class depth in production LLM inference, GPU performance, evals, and/or post-training and RL infrastructure, with strong working knowledge across the rest
  • Demonstrated ownership of the architecture of a large-scale production inference or training platform
  • Proven ability to create architectural frameworks and engineering standards adopted across large engineering organizations
  • Deep understanding of the hardware/software boundary for AI accelerators: CUDA or ROCm, memory bandwidth and interconnect constraints, and distributed compute paradigms
  • Strong history of growing technical leaders (Staff and above) and multiplying technical capability across teams
  • External recognition in the AI systems community through research, open source, or industry contribution

Preferred

  • Prior experience at a top-tier AI lab or major hyperscaler AI infrastructure team
  • Maintainer or core contributor to a foundational inference, kernel, or RL framework (vLLM, SGLang, TensorRT-LLM, LMCache, FlashInfer, Triton, verl, OpenRLHF, TRL, DeepSpeed, Megatron-LM, etc.)
  • Hands-on depth in RL for LLMs (DPO/GRPO-style methods, reward modelling, multi-turn and tool-use RL) and the interaction between inference and training infrastructure
  • Experience defining developer API platforms adopted at scale by external developers
  • Deep experience with control plane / data plane architecture and cell-based deployment patterns in large-scale inference infrastructure
  • Published work in AI systems: MLSys, NeurIPS Systems Track, OSDI, EuroSys, SC, or equivalent
  • Experience with hardware-software co-design: custom accelerator kernels (CUDA, Triton), compiler-level optimization, AI hardware roadmap engagement
  • Experience defining pricing, SLO, and capacity models for a commercial inference product

The range below reflects the base salary for the position. Actual compensation may vary based on job-related factors such as skill set, experience, education, and location. In addition to base salary, this role may be eligible for bonus, equity, and/or commission programs. Nscale may offer a competitive benefits package including medical, dental, vision, flexible paid time off, parental leave, and retirement plan participation.

Salary Range

$290,000—$443,333 USD

For information on how Nscale handles candidate personal data, please see our Employee & Candidate Privacy Notice: Here.

Nscale does not accept unsolicited candidate submissions from recruitment agencies.

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Principal AI Product Engineer in Houston, TX vacancy
  • $174.1k - $306.3k

    Be the one building AI-powered experiences where they matter most. At Genesys,...  ...countries - moving AI from possibility to production in real-world enterprise environments...  ...efficiency. We are looking for a Principal Software Engineer with deep, hands-on Genesys Cloud experience... 
    Principal
    Work from home
    Worldwide
    Flexible hours

    Genesys Cloud Services, Inc.

    Houston, TX
    1 day ago
  •  ...Transfer is looking for a hands-on AI Developer with strong...  ...developer will work alongside AI engineers/developers and business...  ...build, ship, and iterate on real production AI solutions.This role is ideal...  ...job related experience.• The Principal Specialist level requires a... 
    Principal
    Work experience placement
    Work at office
    Night shift

    Energy Transfer Partners

    Houston, TX
    8 hours ago
  • $139k - $258k

     ...delivering quality services of unmatched value and technical competence. This is the Lead position on assigned projects or performs engineering assignments of any complexity. • Develop and review estimates and schedules, progress reports, including workforce forecasts •... 
    Principal
    For contractors
    Local area
    Remote work

    Fluor

    Houston, TX
    1 day ago
  •  ...Principal Design Engineer Do you enjoy working with complex electro-mechanical systems? Do you have experience defining system architecture...  ...Design Engineer with a minimum of 15 years of hands‑on product development experience with electro‑mechanical systems. The... 
    Principal
    Permanent employment
    Contract work
    Worldwide

    Baker Hughes Holdings LLC

    Houston, TX
    2 days ago
  • $175k - $275k

     ...further.OverviewEast West Bank is seeking an experienced Senior AI Engineering to design, build, and operationalize enterprise-grade AI and...  ...translate emerging AI capabilities into secure, scalable, production-ready banking applications that improve operational efficiency... 
    Suggested
    Full time

    East West Bank

    Houston, TX
    8 hours ago
  •  ...That work changes the operating model an engineering organization runs on, the ways of working...  ...sets the target. What we design from it is AI-native by construction. We redesign...  ...We treat the lifecycle as a system and a product itself - built to operate at a speed required... 
    Full time
    Work experience placement
    Live in
    Work at office
    Local area

    Accenture

    Houston, TX
    3 days ago
  • Position Summary Join our AI & Engineering team in transforming technology platforms, driving innovation, and helping make a significant...  ...spanning solution design, prototyping, testing, and production rolloutLead small teams through business requirements gathering... 
    Visa sponsorship

    Deloitte

    Houston, TX
    1 day ago
  • OverviewInfosys Topaz is an AI-first suite of services, solutions, and platforms designed...  ...ensure smooth deployment into scalable, production-ready solutions. • Conduct testing and...  ..., memory, multi-agent), and prompt engineering.• Working with Azure OpenAI, AWS Bedrock... 
    Full time
    Temporary work
    Relocation

    Infosys Technologies

    Houston, TX
    3 days ago
  • $150k - $200k

    About Permidia Permidia builds software that checks construction plans against the local building code before they reach the city. Builders use it to catch problems before they submit, and building departments use it to give their examiners a clear, cited first pass...
    Full time
    Work at office
    Local area
    Remote work

    Permidia Inc

    Houston, TX
    2 days ago
  •  ...digital core and unleashing the power of AI to create value at speed across the enterprise...  ...you can imagine.You are:An AI Native Engineer with a minimum of 3 years of experience...  ...monitoring, and agent observability.Deploying to production — CI/CD, infrastructure as code (... 
    Full time
    Work experience placement
    Live in
    Work at office
    Local area

    Accenture

    Houston, TX
    1 day ago
  •  ...Advanced Technology Centers (ATCs) are the engine for reinvention in our clients’...  ...deepest industry knowledge, the latest in Gen AI solutions, and tech expertise from around...  ...and shape cross-functional influence across Product, R&D, and GTM. You lead AI PS initiatives,... 
    Full time
    Work experience placement
    Live in
    Work at office
    Local area

    Accenture

    Houston, TX
    13 hours ago
  •  ...Oracle is seeking a Senior Product Engineer – Server Platforms to drive supplier qualification, component validation, New Product Introduction (NPI), manufacturing readiness, and supplier quality initiatives for advanced hardware products. This role serves as a key... 

    Ll Oefentherapie

    Houston, TX
    2 days ago
  • $75 - $120 per hour

     ...enterprise organization is seeking a hands-on Forward Deployed AI Engineer to join its growing Enterprise AI team. This role will work...  ...ambiguous business problems into functional prototypes and production-ready solutions. Rapidly build and iterate using Python,... 
    Temporary work
    Shift work

    Addison Group

    Houston, TX
    a month ago
  •  ...Job Title: Principal Process Engineer Duration: Initially 1-year Schedule: M-F 5/40 – hybrid Location: Houston, TX Oversees the successful execution of contractor and supplier Process engineering work starting in the early project development phases through the project... 
    Principal
    For contractors
    Work at office
    Local area
    Immediate start

    Airswift

    Houston, TX
    1 day ago
  •  ...AI Engineer Houston, Texas, United States ON.energy is building the backbone of energy and AI infrastructure powering grid-safe...  ...travel should be expected. Key Responsibilities Building production AI systems end to end, from prototype through deploy and... 
    Full time
    Internship
    Work at office
    Local area
    Remote work
    Worldwide
    Flexible hours
    3 days per week

    ON.energy

    Houston, TX
    3 days ago
  • $124k - $132k

     ...world. We are proud to offer a market-leading range of premium products across categories, styles and price points, remaining...  ...President Enterprise Data Platform within the IT department, the AI Engineer III is responsible for collaborating with data engineering, business... 
    Temporary work
    Visa sponsorship
    Monday to Friday
    Flexible hours
    Weekend work
    Afternoon shift
    Early shift

    Visual Comfort

    Houston, TX
    2 days ago
  •  ...AI Engineer III Job opportunity in the Retail / E-commerce industry! This position would be on-site full-time in Houston, TX. Join...  ...as well as work independently to design, develop, and deploy production-ready solutions. The AI Engineer III would report to the VP of... 
    Full time
    Visa sponsorship

    Perceptive Recruiting, LLC

    Houston, TX
    2 days ago
  • $120k - $125k

     ...AI Engineer Location: Houston, Texas Type: Direct Hire Salary: $120,000 - $125,000 Job Description The AI Engineer will...  ...and deploying machine learning or LLM-based applications in production, including experience architecting systems rather than solely... 
    Work experience placement
    Work at office

    Clearpoint

    Houston, TX
    2 days ago
  •  ...carbon management to advance lower-carbon technologies and products. Headquartered in Houston, Oxy primarily operates in the United...  ...experienced and innovative leader to fill the position of AI Optimization Engineer within our Applied AI Center of Excellence group based... 
    Local area

    Oxy

    Houston, TX
    3 days ago
  •  ...Description Responsibilities Design, develop, and deploy advanced AI and machine learning models to solve complex business...  ...with cross-functional teams, including data scientists, product managers, and software engineers, to integrate AI solutions into production systems.... 

    Pivotal Solutions Inc

    Houston, TX
    4 days ago
  • $73.5k - $212.28k

     ...At PwC, our people in data and analytics engineering focus on leveraging advanced...  ...will lead the development of innovative AI solutions that drive remarkable client outcomes...  ...initiatives within teams Understanding production-grade quality gates and monitoring... 
    H1b

    PwC

    Houston, TX
    2 days ago
  • $100.8k - $245.5k

    Data & AI Engineer Position Description CGI is seeking a hands on Data & AI Engineer with strong experience in modern cloud data...  ...technical requirements, design practical solutions, and develop production ready data and AI applications. The role requires someone... 
    Work at office
    Local area
    Houston, TX
    a month ago
  •  ...become part of a global team of over 50,000 planners, designers, engineers, scientists, digital innovators, program and construction...  ...better world. Join us.Job DescriptionAECOM is seeking a Senior Principal Engineer / Lead Engineer to serve as a senior technical leader... 
    Principal
    Work at office
    Local area
    Worldwide
    Relocation
    Flexible hours

    AECOM

    Houston, TX
    3 days ago
  •  ...Stream-Flo Wellhead Engineer Position Stream-Flo is one of the largest privately held companies providing wellheads, gate valves, check...  ...of wellhead industry to design, develop, and maintain wellhead product lines. Create design files, engineering documentation, and... 
    Worldwide

    Stream-Flo Industries

    Houston, TX
    2 days ago
  • $59.35 - $77.03 per hour

    Senior Agentic AI Engineer Location: Houston, TX Onsite Flexibility: Hybrid - 3 days onsite per week Contract Details Position Type...  ...deploy modern Agentic AI solutions. This role focuses on building production-grade Generative AI applications, multi-agent workflows, and... 
    Contract work
    Work visa
    3 days per week

    Global Technical Talent

    Houston, TX
    1 day ago
  • Job Overview Job ID: J51816 Job Title: AI / Python Data Engineering Location: Houston, TX Duration: 10 Months + Extension Hourly Rate: Depending on Experience (DOE) Work Authorization: US Citizen, Green Card, OPT-EAD, CPT, H-1B, H4-EAD, L2-EAD, GC-EAD Client: To Be Discussed... 
    Hourly pay
    Permanent employment
    Contract work
    H1b
    Local area

    MACHINE LEARNING TECHNOLOGIES LLC

    Houston, TX
    13 hours ago
  •  ...Title: Principal Process Engineer KBR ISA (Integrated Solutions Americas) delivers integrated, end-to-end engineering, procurement, and...  ...including flare load estimation and flare header sizing Production of process design basis documents and various narratives... 
    Principal
    Local area
    Flexible hours

    KBR

    Houston, TX
    4 days ago
  • $115k - $145k

     ...business of advanced ceramics manufacturing. Job Title AI Deployment Engineer The AI Deployment Engineer will support the AI and...  ...CoorsTek by designing, developing, deploying, and supporting production software solutions that incorporate AI and automation capabilities... 
    Permanent employment
    Work experience placement
    Remote work

    CoorsTek

    Houston, TX
    2 days ago
  • $101.9k - $163k

     ...how we work every day. To learn more, please see . The AI Automation Engineer – Sales & Marketing is a hands-on technical role focused...  ...proposal cycle time reduction, pipeline velocity gains, and rep productivity lift Partner with sales and marketing stakeholders to... 
    Full time
    Contract work
    Live in
    Local area
    Worldwide

    Cengage Group

    Houston, TX
    2 days ago
  • $150k - $200k

     ...how we work every day. To learn more, please see . The AI Automation Engineer - Finance & Accounting applies AI to finance operations at...  ...role demands a rare combination: the technical chops to build production AI automation, the financial fluency to understand GL... 
    Full time
    Live in
    Local area
    Worldwide

    Cengage Group

    Houston, TX
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Principal AI Product Engineer. Be the first to apply!