Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Software Engineer - AI/ML GPU Serving

$61k - $101k
Full-time

J.P. Morgan

Salary: $61,000 - 101,000 per year Requirements:

  • We need formal training or certification in software engineering concepts, plus 3+ years of applied experience.
  • We need professional software development experience, with a strong focus on ML systems.
  • We need hands-on experience using enterprise-authorized AI-assisted development tools in the work environment for coding, testing, troubleshooting, or documentation, along with the ability to review, validate, and refine AI-generated output for correctness, performance, and security.
  • We need responsible AI usage practices in engineering workflows, including sensitivity awareness, secure handling of inputs and outputs, and adherence to resiliency and security expectations, with the ability to guide peers on safe usage.
  • We need experience contributing to the design and build of scalable systems and services such as web services and APIs.
  • We need proficiency in Python, including experience with ML frameworks such as TensorFlow, PyTorch, or similar.
  • We need experience with cloud-native technologies such as Docker and Kubernetes.
  • We need experience with one or more public cloud platforms, such as AWS or GCP.
  • We need a solid understanding of system design fundamentals, including reliability, scalability, and performance trade-offs.
  • We need experience working with distributed systems concepts and microservices architectures.
  • We need experience with Cassandra or a comparable NoSQL database.
  • Preferred: working knowledge of additional languages such as Java and/or C++.
  • Preferred: exposure to deploying or operating GPU workloads in Kubernetes environments.
  • Preferred: exposure to ML model serving frameworks such as TorchServe, TensorFlow Serving, or Triton Inference Server.
  • Preferred: familiarity with distributed training and inference systems.
  • Preferred: awareness of model compression techniques and edge deployment concepts.
  • Preferred: familiarity with API management concepts, Domain-Driven Design (DDD), and Agile development environments.
Responsibilities:
  • We execute software solutions, including design, development, and technical troubleshooting, while taking a creative approach to solving complex technical problems.
  • We create secure, high-quality production code and maintain algorithms that run synchronously with the relevant systems.
  • We use enterprise-authorized AI coding assistance tools to improve code quality, delivery speed, and productivity across complex work, while validating results through peer review, automated testing, and secure coding standards and sharing reusable learnings with the team.
  • We apply knowledge of Software Development Life Cycle tools, including enterprise-authorized AI-assisted development and automation capabilities, to increase the value delivered by automation.
  • We produce architecture and design artifacts for complex applications and ensure software code meets design constraints.
  • We gather, analyze, synthesize, and visualize large, diverse data sets to support continuous improvement of software applications and systems.
  • We proactively identify hidden issues and patterns in data and use those insights to improve coding hygiene and system architecture.
Technologies:
  • AI
  • API
  • AWS
  • Cloud
  • Cassandra
  • DDD
  • Docker
  • GCP
  • Support
  • Java
  • Kubernetes
  • Model Serving
  • NoSQL
  • PyTorch
  • Python
  • Security
  • TensorFlow
  • Web
  • microservices

More:

We are collaborating with hackajob to connect with exceptional professionals for this opportunity. We offer an exciting role on our Software Engineer III team within JPMorgan Chases Commercial and Investment Banking business, where we design and deliver secure, stable, and scalable technology products in an agile environment. Our company is one of the oldest financial institutions, serving consumers, small businesses, and major corporate, institutional, and government clients around the world under the J.P. Morgan and Chase brands. We provide a competitive total rewards package, with base salary based on role, experience, skill set, and location, and eligible roles may include commission or discretionary incentive compensation. Our benefits can include comprehensive health care, wellness centers, retirement savings, backup childcare, tuition reimbursement, mental health support, financial coaching, and more. We value diversity and inclusion, offer reasonable accommodations where needed, and operate as an equal opportunity employer. Our Commercial & Investment Bank is a global leader across banking, markets, securities services, and payments, serving organizations in more than 100 countries.

last updated 37 week of 2026

Vacancy posted 4 days ago
Similar jobs that could be interesting for youBased on the Software Engineer - AI/ML GPU Serving in Palo Alto, CA vacancy
  •  ...opportunity for you to take your software engineering career to the next...  ...enterprise-authorized AI coding assist tools...  ...experience, with emphasis on ML systems.Hands-on...  ...deploying or operating GPU workloads in Kubernetes...  ...environmentsExposure to ML model serving frameworks (e.g.,... 
    Suggested

    JP Morgan Chase

    Palo Alto, CA
    7 days ago
  • $61k - $101k

     ...training or certification in software engineering concepts, along with 5+ years...  ..., with a particular focus on ML systems. We need hands-on...  ...using enterprise-authorized AI-assisted development tools in...  ...API DDD Java Model Serving More: We are working with... 
    Suggested
    Full time

    J.P. Morgan

    Palo Alto, CA
    6 days ago
  • $145k - $170k

     ...profound opportunity for AI to drive positive...  ...the Role Our software team is growing,...  ...looking for talented engineers who are graduating...  ..., Onboard Systems, ML Infrastructure, Simulation...  ...The Data Platform serves as a comprehensive...  ...modalities (CPU, GPU, FPGA) etc. You... 
    Suggested
    Immediate start
    Flexible hours

    Nuro

    Mountain View, CA
    10 days ago
  •  ...profound opportunity for AI to drive positive...  ...About the Role As a software engineering intern, you will work closely...  ..., Onboard Systems, ML Infrastructure, Simulation...  ...: The Data Platform serves as a comprehensive management...  ...modalities (CPU, GPU, FPGA) etc. You have... 
    Suggested
    Internship
    Immediate start
    Flexible hours

    Nuro

    Mountain View, CA
    4 days ago
  • $224k - $356.5k

     ...unlimited potential of AI to define the next...  ...era in which our GPU acts as the brains...  ...team is building the software stack that makes...  ...bring-up questions; serve as the technical point...  ...Science, Computer Engineering, Electrical...  ...in GPU computing, ML systems, or high-performance... 
    Suggested
    Full time
    Local area

    Nvidia

    Santa Clara, CA
    5 days ago
  • $145k - $170k

     ...profound opportunity for AI to drive positive...  ...the Role Our software team is growing,...  ...looking for talented engineers who are graduating...  ..., Onboard Systems, ML Infrastructure, Simulation...  ...The Data Platform serves as a comprehensive...  ...modalities (CPU, GPU, FPGA) etc. You... 
    Full time
    Immediate start
    Flexible hours

    Nuro

    Mountain View, CA
    4 days ago
  • $140k - $150k

     ...Software EngineerDeepInfra is building the infrastructure...  ...next generation of AI. We believe open-...  ...inference platform serves trillions of tokens...  ...everything from GPU infrastructure to...  ...-career Software Engineers (0-2 years of...  ...re excited about AI/ML, have taken related... 
    Full time
    Internship

    DeepInfra

    Palo Alto, CA
    1 day ago
  • $227.5k - $280k

     ...evolving beyond using AI for simple productivity...  ...are seeking a Principal Software Architect to lead our transition...  ...fully embedded in our engineering development cycles....  ...and maintaining the AI/ML stack while optimizing...  ...initiatives, and serve as a catalyst for organizational... 
    Temporary work
    Local area
    Remote work
    Flexible hours

    Ring Inc

    Menlo Park, CA
    3 days ago
  •  ...Networking Systems & Software Architecture group is solving some of AI’s hardest...  ...research and production engineering! What you will be...  ...‑optimization with GPU, DPU, NIC, and switch...  ...capabilities into AI serving stacks such as vLLM...  ...~ Understanding of ML systems concepts—... 

    NVIDIA Gruppe

    Santa Clara, CA
    4 days ago
  • $345k

     ...digital bank in Latin America, serving 135 million customers across...  ...About the Role At Nu, the AI Private Banker is our...  ...action seamlessly. As a Staff Software Engineer on this team, you'll define the...  ...and markets. Integrate AI/ML model outputs into consumer-facing... 
    Work at office
    Flexible hours

    Nubank

    Palo Alto, CA
    5 days ago
  • $182k - $242k

     ...Essential Cloud for AI™. Built for pioneers...  ...for high-performance GPU infrastructure across AI/ML, visual effects,...  ...inference. Our stack is engineered for speed, scale,...  ...of large-scale model serving—squeezing maximum throughput...  ..., GPU/accelerator software, or performance-... 
    Permanent employment
    Full time
    Temporary work
    Casual work
    Work at office
    Flexible hours

    CoreWeave

    Sunnyvale, CA
    a month ago
  •  ...generation computing experiences—from AI and data centers, to PCs,...  ...'re looking for a senior software engineer who combines deep systems performance...  ...who can shape software from GPU kernels through distributed...  ...CUTLASS, Thrust, CUB, NCCL), and ML framework cores such as... 
    Shift work

    AMD

    Santa Clara, CA
    4 days ago
  • $119.8k - $234.7k

     ...Less than 25%Profession: Software EngineeringDiscipline:...  ...builds the end-to-end AI stack and is core to Azure...  ...as a Service, Azure ML, Cognitive Services, and...  ...customers are better served. Within CoreAI, the Foundry...  ...a Principal Software Engineer - Responsible AI who is... 
    Ongoing contract
    Work at office
    Local area
    3 days per week

    Microsoft

    Mountain View, CA
    6 days ago
  • $92k - $135k

     ...CoreWeave is The Essential Cloud for AI™. Built for pioneers by...  ..., and cost for model serving on our GPU platform. As an IC1, you'll implement...  ...mentorship from experienced engineers. About the role:...  ...that deployed a microservice or ML inference demo. Coursework... 
    Permanent employment
    Full time
    Temporary work
    Casual work
    Internship
    Work at office
    Flexible hours

    CoreWeave

    Sunnyvale, CA
    24 days ago
  •  ...Principal Software EngineerIf you are looking for a game-changing career...  ...place.As a Principal Software Engineer at JPMorganChase within the Corporate Sector – AI/ML & Data Platforms for LLM Suite,...  ...complex ML projects and initiatives, serving as the primary decision-maker... 

    Chase

    Palo Alto, CA
    1 day ago
  • $55 - $60 per hour

     ...inference optimization (quantization, batching, GPU utilization). Exposure to...  ...track record of working with roughly 10+ AI/ML projects deployed to production. Deep algorithmic...  ...approach, agent/orchestration design, serving strategy, API design, and infrastructure footprint... 

    Cynet Systems

    Santa Clara, CA
    5 days ago
  •  ...AI Infrastructure EngineerWe're looking for an AI Infrastructure Engineer to build and operate the serving infrastructure. You will work hands-on with vLLM and...  ...parallelism strategies across GPU systems.Build reproducible...  ...in LLM inference, ML systems, GPU systems, or performance... 
    Work at office
    Immediate start
    Relocation package

    Netpreme

    Santa Clara, CA
    3 days ago
  • $180k - $240k

     ...Senior AI Infrastructure EngineerSanta Clara...  ..., Gatik Carrier™ serves as an all-...  ...integrates advanced software and hardware powering...  ...AI Infrastructure Engineer to design, build, and...  ...DoDistributed Training & ML Systems...  ...and optimize multi-GPU setups, ensuring efficient... 
    Odd job
    Work at office

    Gatik AI

    Santa Clara, CA
    18 hours ago
  •  ...potential of generative AI to power the...  ...are at the forefront of software and hardware innovation...  ...Principal System Software Engineer, AI Inference Execution...  ...closely with other software (ML and compilers) and hardware...  ...servers/model serving frameworks (such as TensorRT... 
    Work experience placement
    3 days per week

    Entrada Ventures

    Santa Clara, CA
    2 days ago
  • $150k - $250k

     ...will build state-of-the-art AI capabilities for Cylake's next...  ...with security researchers and software engineers to develop, productize, and...  ...agents, model inference, and serving optimization. Required Experiences...  ...systems, and production ML pipelines. ~ Strong... 
    Full time

    Cylake, Inc

    Sunnyvale, CA
    1 day ago
  •  ...is the world’s leading AI-powered Fraud and Risk Platform...  ..., powerful decision engine, and investigation tools...  .... Our award-winning software platform is powered by a...  ...hiring a AI Engineer to serve as a technical architect...  ...role. Classic ML modeling experience is a... 

    DataVisor

    Mountain View, CA
    4 days ago
  • $193.93k - $352.29k

     ...profound opportunity for AI to drive positive...  ...autonomously inside Nuro's own engineering organization, under the...  ...You ~3+ years of software engineering experience...  .... ~ Experience with ML training or research infrastructure...  ...running inference serving, cost, or capacity at... 
    Immediate start
    Flexible hours

    Nuro

    Mountain View, CA
    3 days ago
  • $100k - $200k

     ...and experienced backend engineer to join our growing team....  ...performance RESTful APIs that serve as the backbone for our generative AI-powered Android...  ...closely with Android and ML teams on API contracts and...  ...of professional backend software development experience, with... 
    Full time

    OPPO US Research Center

    Palo Alto, CA
    23 days ago
  •  ...Nexxa is building the best AI systems for heavy industries...  ...We're looking for Backend AI Engineers to design, build, and own...  ...solutions possible at scale: model-serving pipelines, inference and...  ...role is a blend of backend software engineering, ML infrastructure, and systems... 
    Full time

    Nexxa.ai

    Sunnyvale, CA
    1 day ago
  • $120k - $160k

     ...information powered by advanced AI, recommendation systems...  ...Recruiter, AI & Engineering to help build the teams...  ...pipeline across AI/ML, backend, infrastructure...  ...Manager Experience Serve as a trusted advisor to...  ...full-cycle recruiting for software engineers, AI/ML... 
    Full time
    Work at office
    Local area
    Work from home
    Monday to Friday

    NewsBreak

    Mountain View, CA
    10 days ago
  • $184k - $287.5k

     ...seeking highly skilled and motivated software engineers to join us and build AI inference systems that serve large-scale models with extreme...  ...inference stacks, optimize GPU kernels and compilers, drive...  ...pareto frontier for the field of ML Systems; survey recent... 
    Full time

    Nvidia

    Santa Clara, CA
    5 days ago
  • $184k - $287.5k

     ...looking for outstanding AI systems engineers to develop...  ...the inference systems software stack! We build innovative...  ...code generators, and GPU kernel technologies for...  ...abstractions for LLM serving enginesBuilding efficient...  ...industry) experience with ML/DL systems development... 
    Full time

    Nvidia

    Santa Clara, CA
    4 days ago
  •  ...builds Industrial AI for the world's leading...  ...state-of-the-art ML to large volumes of...  ..., fine-tuning, and serving them in large-scale...  ...a Senior/Staff AI Engineer, you will turn ML research...  ...AI Scientists, Software Engineers, and...  ...parallel training (multi-GPU/multi-node).... 
    Full time
    Shift work

    Gauss Labs

    Palo Alto, CA
    a month ago
  • $1,000 - $2,030 per month

     ...kitchens so they can better serve restaurateurs, customers and...  ...on the Machine Learning and AI Infrastructure team, which is...  ...for data scientists and other engineers to use ML and AI technologies effectively. Build and maintain software and tools for classical ML stack... 
    Full time
    Temporary work
    Work at office
    Flexible hours

    Cloudkitchens

    Mountain View, CA
    1 day ago
  •  ...backend or distributed systems (pre-AI experience required)...  ...Has shipped AI/LLM features serving real users at scale - not just...  ...detection) Hands-On Engineer Not just an architect - writes...  ..., prompt compression ML observability tools: Langfuse,... 

    ClifyX

    Sunnyvale, CA
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Software Engineer - AI/ML GPU Serving. Be the first to apply!