Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Software Engineer

$167.7k - $245.2k

Jobleads-US

The application window is expected to close on: 10/29/2026

We run the platform that serves foundational models to Cisco IT. Our Foundational Model Service, gives engineering teams across the company access to small, large, and embedding models. youWe serve those models on Kubernetes clusters, with Nim, Vllm and other runtimes. Beyond serving, We benchmark, evaluate, monitor, and release new models as improvements and demand warrant. We use published model artifacts where possible, refining or rebuilding them for compatibility, performance, or quality. Our customers depend on the platform under a 99.9% uptime SLA, and we build and operate accordingly.

As a Senior Software Engineer, you'll guide model serving, runtime tuning, accelerator optimization, and release evaluation. You'll lead incident response and improve reliability, performance, and operability. You'll also use agentic workflows to identify problems earlier and automate remediation.

You'll follow Cisco Design Thinking Principles, simplify user experience, apply secure coding practices, and protect privacy. You'll work across design, product, and engineering to improve customer solutions, documentation, development practices, and production reliability.

This role calls for production experience in AI infrastructure, model serving, evaluation, and related services. We're looking for a self starter who works independently, turns ambiguous problems into plans, mentors engineers, and raises technical standards. The platform evolves with new runtimes, accelerators, and models. On call is shared across the team.

Responsibilities

  • Contribute to model serving direction and roadmap, including runtime selection and tuning across vLLM, NVIDIA NIM, and other runtimes.
  • Guide quantization and accelerator optimization across GPU vendors, validating performance, quality, capacity, and cost with data
  • Develop and enhance platform services, APIs, gateways, and operational tooling around model inference
  • Evolve evaluation, benchmarking, and load test frameworks that gate model releases against service level objectives
  • Define model promotion criteria across quality, safety, latency, throughput, and resource use
  • Evaluate how fine tuning, distillation, and quantization affect production behavior
  • Shape routing and capacity behavior, including prefix caching, KV aware routing, and prefill and decode separation
  • Improve model registry, packaging, evaluation, release, and development workflows using Infrastructure as Code, GitHub Actions, and agentic workflows
  • Refine model artifacts when runtime compatibility or performance requires changes
  • Develop observability that shows platform health, model performance, capacity, customer adoption, and usage
  • Monitor production, serve as an escalation point for on call issues, lead postmortems and root cause analyses, and drive durable improvements
  • Coordinate across customer, product, design, and engineering teams to gather input, forecast capacity, track milestones, and guide platform direction
  • Apply AI to platform operations through anomaly detection, automated remediation, and predictive operations
  • Lead features and projects from technical design through completion, working with minimal guidance and driving results through delegation and review
  • Write clean code and unit tests independently, and review code for quality, threat models, scale, reliability, and release velocity
  • Act as a technical resource, mentor engineers, run design reviews, and share knowledge across teams
  • Create technical designs, runbooks, user documentation, project updates, and remediation plans

Requirements

  • 7 or more years of related systems, platform, or software engineering experience, or equivalent practical experience, with solid knowledge across related technologies
  • A production background developing and operating AI infrastructure
  • A solid understanding of LLM, SLM, embedding, and reranker model internals, including context length, batching, token throughput, and memory use
  • Hands on work serving models on inference runtimes including vLLM, NVIDIA NIM, or Triton
  • Proven results building services around models, including the APIs, gateways, and operational tooling that make them consumable
  • Depth in evaluating and benchmarking models, and using the results to make release decisions
  • Familiarity moving model artifacts through evaluation, optimization, packaging, and production serving
  • Deep understanding of supervised fine tuning, parameter efficient fine tuning, distillation, and quantization
  • Command of distributed GPU training concepts, mixed precision, and parallelism strategies
  • Production Kubernetes work running GPU workloads at scale
  • Linux administration and troubleshooting
  • Programming in Python or Go
  • Fluency with CI/CD pipelines and Infrastructure as Code, for example Terraform or Ansible
  • mastery of monitoring and observability tooling, including Prometheus, Grafana, or Splunk
  • A track record of mentoring engineers and setting technical standards
  • A demonstrated pattern of learning new systems and the initiative to lead unfamiliar work
  • Ability to take part in an on call rotation for a service the company depends on
  • Clear written communication and the habit of documenting what you develop

Preferred Knowledge and Experience

  • Work with AMD accelerators and ROCm alongside NVIDIA Cuda
  • Exposure to distributed inference, disaggregated serving, or KV cache aware routing
  • Familiarity with evaluation frameworks including lm-eval or DeepEval, and with safety and capability suites
  • A background evaluating RAG or agent systems
  • Time spent with API gateways, ingress, or load balancing, for example Envoy, APISIX, or NGINX
  • Fluency with GitOps and Helm, particularly ArgoCD
  • Capacity planning, traffic pattern understanding, and cost optimization for GPU fleets
  • Disaster recovery for stateful platform services
  • Contributions to open source inference or evaluation projects
  • Certified Kubernetes Administrator (CKA) or an equivalent cloud certification
  • Training or fine tuning transformer models with PyTorch and Hugging Face
  • Practical use of LoRA, QLoRA, and PEFT
  • Distributed training with PyTorch FSDP, DeepSpeed, or comparable frameworks
  • Model registries and experiment tracking systems, for example MLflow or Weights and Biases
  • Multi node GPU training and collective communication libraries

Education

  • Bachelor's degree in Computer Science, Information Systems, or a related field, or equivalent practical experience

Why Cisco?

At Cisco, we’re revolutionizing how data and infrastructure connect and protect organizations in the AI era – and beyond. We’ve been innovating fearlessly for 40 years to create solutions that power how humans and technology work together across the physical and digital worlds. These solutions provide customers with unparalleled security, visibility, and insights across the entire digital footprint.

Fueled by the depth and breadth of our technology, we experiment and create meaningful solutions. Add to that our worldwide network of doers and experts, and you’ll see that the opportunities to grow and build are limitless. We work as a team, collaborating with empathy to make really big things happen on a global scale. Because our solutions are everywhere, our impact is everywhere.

We are Cisco, and our power starts with you.

Message to applicants applying to work in the U.S. and/or Canada:

The starting salary range posted for this position is $167,700.00 to $245,200.00 and reflects the projected salary range for new hires in this position in U.S. and/or Canada locations, not including incentive compensation*, equity, or benefits.

Individual pay is determined by the candidate's hiring location, market conditions, job-related skillset, experience, qualifications, education, certifications, and/or training. The full salary range for certain locations is listed below. For locations not listed below, the recruiter can share more details about compensation for the role in your location during the hiring process.

U.S. employees are offered benefits, subject to Cisco’s plan eligibility rules, which include medical, dental and vision insurance, a 401(k) plan with a Cisco matching contribution, paid parental leave, short and long-term disability coverage, and basic life insurance. Please see the Cisco careers site to discover more benefits and perks. Employees may be eligible to receive grants of Cisco restricted stock units, which vest following continued employment with Cisco for defined periods of time.

U.S. employees are eligible for paid time away as described below, subject to Cisco’s policies:

  • 10 paid holidays per full calendar year, plus 1 floating holiday for non-exempt employees
  • 1 paid day off for employee’s birthday, paid year-end holiday shutdown, and 4 paid days off for personal wellness determined by Cisco
  • Non-exempt employees** receive 16 days of paid vacation time per full calendar year, accrued at rate of 4.92 hours per pay period for full-time employees
  • Exempt employees participate in Cisco’s flexible vacation time off program, which has no defined limit on how much vacation time eligible employees may use (subject to availability and some business limitations)
  • 80 hours of sick time off provided on hire date and each January 1st thereafter, and up to 80 hours of unused sick time carried forward from one calendar year to the next
  • Additional paid time away may be requested to deal with critical or emergency issues for family members
  • Optional 10 paid days per full calendar year to volunteer

For non-sales roles, employees are also eligible to earn annual bonuses subject to Cisco’s policies.

Employees on sales plans earn performance-based incentive pay on top of their base salary, which is split between quota and non-quota components, subject to the applicable Cisco plan. For quota-based incentive pay, Cisco typically pays as follows:

  • .75% of incentive target for each 1% of revenue attainment up to 50% of quota;
  • 1.5% of incentive target for each 1% of attainment between 50% and 75%;
  • 1% of incentive target for each 1% of attainment between 75% and 100%; and
  • Once performance exceeds 100% attainment, incentive rates are at or above 1% for each 1% of attainment with no cap on incentive compensation.

For non-quota-based sales performance elements such as strategic sales objectives, Cisco may pay 0% up to 125% of target. Cisco sales plans do not have a minimum threshold of performance for sales incentive compensation to be paid.

The applicable full salary ranges for this position, by specific state, are listed below:

New York City Metro Area:

$167,700.00 - $282,000.00

Non-Metro New York state & Washington state:

$149,100.00 - $250,900.00

* For quota-based sales roles on Cisco’s sales plan, the ranges provided in this posting include base pay and sales target incentive compensation combined.

** Employees in Illinois, whether exempt or non-exempt, will participate in a unique time off program to meet local requirements.

#J-18808-Ljbffr Jobleads-US
Vacancy posted 3 days ago
Similar jobs that could be interesting for youBased on the Software Engineer in San Jose, CA vacancy
  • $182k - $226k

     ...our world! Our outstanding people deliver creative decarbonization energy solutions to our customers.We are seeking a Senior Software Engineer to join Electric Hydrogen's Digital Team, where you will develop mission-critical software that powers the world's most powerful... 
    Suggested
    Temporary work

    Electric Hydrogen

    San Jose, CA
    3 days ago
  • $184k - $287.5k

    We are looking for a Senior Software Engineer to become part of our storage management plane team. The management plane is a web-based application crafted to provide our storage customers the capabilities to handle and supervise our distributed storage infrastructure. Our... 
    Suggested
    Full time

    Nvidia

    Santa Clara, CA
    4 days ago
  • $388k

     ...be a part of what’s next.Our Content & Business Products (CBP) Engineering teams build the products and services to optimize and automate...  ...that other engineering teams build on, across the full software stack — from GraphQL/RESTful APIs and embeddable UI components... 
    Suggested
    Hourly pay
    Full time
    Immediate start
    Remote work
    Flexible hours

    Netflix

    Los Gatos, CA
    2 days ago
  • $388k

     ...entertainment is imagined, created, and delivered to a global audience. Engineering teams within Netflix work hard every day to scale and innovate...  ...production and member experiences in an ever-growing complex software landscape. Our Content & Business Products (CBP) Engineering... 
    Suggested
    Hourly pay
    Full time
    Immediate start
    Flexible hours

    Netflix

    Los Gatos, CA
    5 days ago
  •  ...Full-Stack Software Engineer Precision Neuroscience is building a next-generation brain–computer interface (BCI) to heal and empower millions of people living with neurological conditions. Our team brings together experts in neurosurgery, AI and machine learning, microfabrication... 
    Suggested
    Work at office
    Remote work

    Precision Neuroscience

    Santa Clara, CA
    4 days ago
  • $136.3k - $231.7k

     ...invest 15% of sales back into R&D. Our expert teams of physicists, engineers, data scientists and problem-solvers work together with the...  ...is looking for the best and the brightest research scientist, software engineers, application development engineers, and senior... 
    Minimum wage
    Full time
    Flexible hours

    KLA-Tencor

    Milpitas, CA
    4 days ago
  • $184k - $287.5k

    The Autonomous Vehicles Platform team is looking for a hands-on Cybersecurity Software Engineer. As part of our team, you will work on our Autonomous Driving Platform software, implementing performant and scalable solutions for data collection and autonomous vehicle fleets... 
    Full time

    Nvidia

    Santa Clara, CA
    1 day ago
  • $250k - $413k

     ...and what do they need? This is the responsibility of Identity Engineering, managing some of the most foundational infrastructure at Netflix...  ...might be a great fit if you:Have 3-6+ years of professional software engineering experience building production systems, ideally... 
    Hourly pay
    Full time
    Immediate start
    Flexible hours

    Netflix

    Los Gatos, CA
    3 days ago
  • $250k - $413k

     ...localized experience helps millions of members connect with content in their own language and culture. The Product Localization Engineering team is reimagining how localization works at Netflix. We build the platforms, products, and AI-powered systems that enable translation... 
    Hourly pay
    Full time
    Immediate start
    Flexible hours

    Netflix

    Los Gatos, CA
    5 days ago
  • Senior Software Engineer - Data Protection Software Engineering (C, C++)Infrastructure Solutions Group (ISG) builds the products that power infrastructure, solutions, and data management our customers need most. Our teams design and develop the hardware and software that... 

    Dell Technologies

    Santa Clara, CA
    2 days ago
  • $120.75k - $161k

     ...Platform team to design, develop, and maintain the large-scale software platforms that serve millions of users globally. In this...  ...tolerance, horizontal scalability, and load balancing.Performance Engineering: Drive the scaling, optimization, and innovation of the Data Platform... 
    Permanent employment
    Work at office
    3 days per week

    Eightfold

    Santa Clara, CA
    1 day ago
  • $160k - $320k

     ...tenant security infrastructure.You will combine strong backend engineering skills with deep Kubernetes operational expertise to solve...  ...operational needs.What You Will Bring9-16 years of professional software engineering experience, with a strong focus on backend systems... 
    Work at office
    Local area
    Remote work
    Relocation package
    3 days per week

    Nutanix

    San Jose, CA
    2 days ago
  • Job DescriptionTitle: Senior Software Architect EngineerType: permanentLocation: San Jose, CAJob description:A Medical Device Company Located...  ...in San Jose, California is looking for a Senior Software Engineer Architect to drive the software architecture development effort... 
    Permanent employment

    Maganti IT Resources

    San Jose, CA
    4 days ago
  • $184k - $287.5k

     ...& SOC technology? NVIDIA is looking for a Senior GPU Platforms Engineer to be part of our innovative and ambitious group. This is an outstanding...  ...to work on groundbreaking GPU & SOC systems and platform software, driving the next era of accelerated computing!What you'll be... 
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago
  • $280k - $380k

     ...fast at internet scale. We focus on reliability and automation, engineering systems that perform under stress and continuously improve. We...  ...experienced DevOps/SRE (Site Reliability Engineering) Senior Software Engineer to join our dynamic team. The ideal candidate will... 
    Work at office
    Local area
    Remote work
    Monday to Thursday
    Flexible hours

    Roku

    San Jose, CA
    4 days ago
  • $165.6k - $227.7k

     ...at the Intelligent Edge. ADI combines analog, digital, AI, and software technologies into solutions that combat climate change, reliably...  ...applications across ADI. We are seeking a platform-minded engineer with strong expertise in TypeScript, React, APIs, and AWS who is... 
    Permanent employment
    Full time
    Work at office
    Day shift

    Analog Devices

    San Jose, CA
    2 days ago
  • $224k - $356.5k

    The Autonomous Vehicles Platform team is looking for a Senior System Software Engineer. As part of our team, you will work on our Autonomous Driving Platform software, implementing high-performance and scalable solutions for data collection and autonomous vehicle fleets... 
    Full time

    Nvidia

    Santa Clara, CA
    5 days ago
  • $168k - $270.25k

    As a Senior Software Engineer on NVIDIA’s Global Network Visibility (GNV) team within NVIDIA's Global Network Infrastructure (GNI) organization, you will lead the platform that turns network topology, configuration, telemetry, and direct measurement into trusted, self-service... 
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • $152k - $241.5k

     ...revolution across many applications and industries. Within our software stack, CUTLASS stands out as a popular open-source ecosystem...  ...need to see:Masters or PhD degree in Computer Science, Computer Engineering, or related field (or equivalent experience).3+ years of... 
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • $186k - $388k

     ...adopt new technologies, design shared architectural layers (queuing, event systems, shared memory), and collaborate across Product, Engineering, QA, and Ops to deliver resilient services spanning streaming, APIs, notifications, and batch workloads. As a hands‑on technical... 
    Work at office
    Local area
    Remote work
    Monday to Thursday
    Flexible hours

    Roku

    San Jose, CA
    4 days ago
  • $388k

     ...client applications, partner device implementations, mobile games, and more. We view ourselves as a force multiplier for Netflix engineering, providing composable capabilities and pluggable abstractions that allow teams to manage, orchestrate, and analyze their automated... 
    Hourly pay
    Full time
    Immediate start
    Flexible hours

    Netflix

    Los Gatos, CA
    5 days ago
  • $193.3k - $261.5k

     ...-system C++ and SystemC models of these custom SoCs — that let software teams start development months before silicon arrives. For Trainium...  ...within 12 hours of first silicon. We're looking for a software engineer to build and own the models and infrastructure that make this... 
    Internship
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    4 days ago
  • $293.9k - $406.8k

     ...expected to close on: 10/29/2026This is a hybrid role based out of Cisco's San Jose office.Meet the TeamCisco’s AI Software & Platform (AISWP) team is the engine driving our transformation into an AI-native company. We are a central organization of researchers, ML... 
    Full time
    Temporary work
    Work at office
    Local area
    Worldwide
    Flexible hours

    CISCO Systems

    Milpitas, CA
    5 days ago
  • $184k - $287.5k

     ...best work. Come join the team and see how you can make a lasting impact on the world.We are looking for an experienced Senior Software Engineer for the Embedded Platform team. This is an outstanding opportunity to accelerate the pace of Jetson, IGX and DGX Spark Product... 
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago
  • $224k - $356.5k

     ...lasting impact on the world.NVIDIA's Local AI team is building the software stack that makes large language models and generative AI...  ...concernsWhat we need to see:BS, MS, or PhD in Computer Science, Computer Engineering, Electrical Engineering, or equivalent experience.12+ years of... 
    Full time
    Local area

    Nvidia

    Santa Clara, CA
    4 days ago
  • $167.7k - $245.2k

     ...network performance for the AI era. We design and deliver the software intelligence behind Cisco’s next-generation high-performance routing...  ...One NPUs, and distributed system software. We are looking for engineers who are passionate about solving the challenges of exponential... 
    Full time
    Temporary work
    Work at office
    Local area
    Flexible hours
    3 days per week

    CISCO Systems

    Milpitas, CA
    3 days ago
  • $174k - $252k

     ...product or system development code. Review code developed by other engineers and provide feedback to ensure best practices (e.g., style...  ...or C)3 years of experience testing, maintaining, or launching software products, and 1 year of experience with software design and... 

    Google

    San Jose, CA
    5 days ago
  • $174.72k - $295.68k

     ...machine learning, and smart connectivity. You will be a senior engineer on the team building our internal AI engineering platform — the...  ..., and automation that power how our engineers build and ship software. A core mission is connecting agents to real engineering... 
    Full time

    XPENG Motors

    Santa Clara, CA
    2 days ago
  •  ...price point while also creating a compelling path for advertisers to reach deeply engaged audiences.Our TeamThe Ads Serving Platform engineering team sits within the Ad Serving & Decisioning org at Netflix Ads. We own the high-throughput, low-latency distributed systems... 
    Hourly pay
    Full time
    Immediate start
    Flexible hours

    Netflix

    Los Gatos, CA
    4 days ago
  • $186k - $388k

     ...adopt new technologies, design shared architectural layers (queuing, event systems, shared memory), and collaborate across Product, Engineering, QA, and Ops to deliver resilient services spanning streaming, APIs, notifications, and batch workloads. As a hands‑on technical... 
    Work at office
    Local area
    Remote work
    Monday to Thursday
    Flexible hours

    Roku

    San Jose, CA
    5 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Software Engineer. Be the first to apply!