Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Principal AI Performance Engineer

$262.7k - $355.4k

Arm, Inc.

Job Summary;Are you passionate about optimizing AI workloads and delivering real-world performance improvements on edge devices?We’re looking for an experienced engineer to help customers achieve best-in-class inference performance for production AI models running on Arm technology.This role is based in San Jose, with significant time spent working directly with customers across the Bay Area.Job Description:In this role, you will work closely with customers to optimize AI workloads targeting Arm technology, focusing on achieving best-in-class performance and power efficiency.Using your experience with Arm’s AI optimization tools and your understanding of hardware architectures, you will develop kernel level implementations across a range of DNN models, optimizing for power and performance.In collaboration with multiple teams across Arm’s engineering organization, you will diagnose and resolve performance challenges, and use these insights to influence Arms IP and tooling roadmaps.This role requires strong coding and communication skills; you’ll translate complex technical challenges into clear insights, presenting progress and recommendations to audiences ranging from engineers to senior leadership.Responsibilities:Develop highly optimized solutions for AI workloads, from kernel level to system level, to meet the needs of the customer application.Create production quality reference implementations, documentation, and performance focused technical contentAct as a technical bridge between customers and internal teams, driving resolution of complex performance issuesInfluence Arm’s IP and software roadmap through insights gained from real-world customer use casesRequired Skills and Experience :Experience optimizing DNNs in Triton, CUDA or other kernel level programming languageDeep understanding of parallel computing, memory hierarchies and performance optimization techniques for DNNsStrong programming skills in Python and C++, a solid experience with modern AI frameworks and execution models, experience with profiling and analysis toolsStrong communication and interpersonal skills“Nice To Have”:Experience in a customer facing or field engineering environmentExperience within the Arm ecosystemBackground in AI performance optimization for edge devicesWe will ensure that individuals with disabilities are provided reasonable accommodation to participate in the job application or interview process, to perform essential job functions, and to receive other benefits and privileges of employment. Please contact us to request accommodation.In Return:Joining Arm means stepping into a career‑defining opportunity. You’ll occupy a central role in the company’s most critical initiatives. These initiatives build how Arm innovates, scales, and partners globally.Additional InformationPlease note this role does not meet the eligibility requirements for sponsorship, and therefore the successful candidate must have the right to work in the US without relying on sponsorship by Arm.10x Mindset at ArmAt Arm, we believe progress happens when people are empowered to think bigger and push beyond what seems possible. Our 10x mindset is about curiosity, ambition and creating impact that will be used by millions. We learn fast, prioritise collaboration and turn bold ideas into real technology. We look for people who are inspired by this way of working and want to grow in an environment where bold ideas are welcomed. Read more about how we bring the 10x mindset to life on the Arm blog.Didn't find what you were looking for?Join our talent community. Click here to get started.Accommodations at ArmAt Arm, we want to build extraordinary teams. If you need an adjustment or an accommodation during the recruitment process, please email View email address on click.appcast.io. To note, by sending us the requested information, you consent to its use by Arm to arrange for appropriate accommodations. All accommodation or adjustment requests will be treated with confidentiality, and information concerning these requests will only be disclosed as necessary to provide the accommodation. Although this is not an exhaustive list, examples of support include breaks between interviews, having documents read aloud, or office accessibility. Please email us about anything we can do to accommodate you during the recruitment process.Equal Opportunities at ArmArm is an equal opportunity employer, committed to providing an environment of mutual respect where equal opportunities are available to all applicants and colleagues. We are a diverse organization of dedicated and innovative individuals, and don’t discriminate on the basis of race, color, religion, sex, sexual orientation, gender identity, national origin, disability, or status as a protected veteran.Hybrid Working at ArmArm’s hybrid approach to working is centred around flexibility, where we split our time between the office and other locations to get our work done. Within that framework, we empower groups and teams to determine their own particular hybrid working pattern, depending on the work and the team’s needs. Details of what this means for each role will be shared upon application. In some cases, the flexibility we can offer is limited by local legal, regulatory, tax, or other considerations, and where this is the case, we will collaborate with you to find the best solution. Please talk to us to find out more about what this could look like for you.Salary Range:$262,700-$355,400 per yearWe value people as individuals and our dedication is to reward people competitively and equitably for the work they do and the skills and experience they bring to Arm. Salary is only one component of Arm's offering. The total reward package will be shared with candidates during the recruitment and selection process.Accommodations at ArmAt Arm, we want to build extraordinary teams. If you need an adjustment or an accommodation during the recruitment process, please email View email address on click.appcast.io. To note, by sending us the requested information, you consent to its use by Arm to arrange for appropriate accommodations. All accommodation or adjustment requests will be treated with confidentiality, and information concerning these requests will only be disclosed as necessary to provide the accommodation. Although this is not an exhaustive list, examples of support include breaks between interviews, having documents read aloud, or office accessibility. Please email us about anything we can do to accommodate you during the recruitment process.Hybrid Working at ArmArm’s approach to hybrid working is designed to create a working environment that supports both high performance and personal wellbeing. We believe in bringing people together face to face to enable us to work at pace, whilst recognizing the value of flexibility. Within that framework, we empower groups/teams to determine their own hybrid working patterns, depending on the work and the team’s needs. Details of what this means for each role will be shared upon application. In some cases, the flexibility we can offer is limited by local legal, regulatory, tax, or other considerations, and where this is the case, we will collaborate with you to find the best solution. Please talk to us to find out more about what this could look like for you.Equal Opportunities at ArmArm is an equal opportunity employer, committed to providing an environment of mutual respect where equal opportunities are available to all applicants and colleagues. We are a diverse organization of dedicated and innovative individuals, and don’t discriminate on the basis of race, color, religion, sex, sexual orientation, gender identity, national origin, disability, or status as a protected veteran.

Vacancy posted 4 days ago
Similar jobs that could be interesting for youBased on the Principal AI Performance Engineer in San Jose, CA vacancy
  •  ...California or Austin, TexasNXP is searching for a hands-on AI Compiler Engineer who thrives at the convergence of cutting-edge AI, compiler...  ...aligning silicon and code for maximum impact.Diagnose and crush performance bottlenecks with AI-enabled profiling and diagnostics,... 
    Principal
    Performance
    Full time
    Work at office
    Local area

    NXP Semiconductors

    San Jose, CA
    7 hours ago
  •  ...next-generation computing experiences—from AI and data centers, to PCs, gaming and...  ...runtime, libraries, models, frameworks, and performance optimization layers. The role also...  ...tier AI customers. Workload Performance Engineering: Lead the profiling, analysis, and tuning... 
    Principal
    Performance

    AMD

    San Jose, CA
    7 hours ago
  •  ...that accelerate next-generation computing experiences—from AI and data centers, to PCs, gaming and embedded systems. Grounded...  ..., we advance your career. THE ROLE:AMD is looking for a performance-obsessed engineer to drive AI inference performance to the absolute limit on... 
    Principal
    Performance

    AMD

    San Jose, CA
    2 days ago
  • $270k - $340k

     ...of our team members.What You’ll Do:As a Principal AI and ML fundamentalist who is an expert...  ...deployment. Collaborate with other researcher engineers to prototype and validate complex...  ...grow our business. We drive a pay-for-performance culture and reward performance that supports... 
    Principal
    Performance
    Local area

    Archer Aviation

    San Jose, CA
    4 days ago
  • $250.44k - $375.67k

     ...for every OK-er.About the OpportunityWe are looking for a Principal AI Engineer to lead the architecture and deployment of large-scale, LLM...  ...simulations, and continuous learning pipelines for Chatbot performance optimization.Design multi-level intent routing, classifier... 
    Principal
    Performance

    OKX

    San Jose, CA
    2 days ago
  •  ...About the Opportunity We are seeking a Principal Engineer with a deep expertise in autonomous AI agent architecture and deployment, to spearhead the design,...  ...simulations, and continuous learning pipelines for agent performance optimization. Stay ahead of the curve on... 
    Principal
    Performance

    United States Digital Space LLC

    San Jose, CA
    3 days ago
  • $249k

     ...their best work. Join us and build for travelers everywhere.Principal Data & AI Engineer, Reporting and InsightsIntroduction to the Team: Our...  ...how Expedia Group transforms operational, portfolio, and performance data into trusted executive insights. You will bring together... 
    Principal
    Performance
    Full time
    Work at office

    Expedia

    San Jose, CA
    3 days ago
  • $313.06k

     ...every OK-er. About The Opportunity We are seeking a Principal Engineer with a deep expertise in autonomous AI agent architecture and deployment to spearhead the...  ..., and continuous learning pipelines for agent performance optimization. Stay ahead of the curve on developments... 
    Principal
    Performance

    Okx

    San Jose, CA
    1 day ago
  • $123.24k - $200k

     ...Overview Of Role As a Sr./Principal AI Engineer within TSMC's Artificial Intelligence for Business Intelligence Innovation (AI4BII) Center,...  ...engineering, machine learning engineering, or related fields in high‑performance environments. 7+ years of hands‑on experience in... 
    Principal
    Performance
    Work at office

    TSMC

    San Jose, CA
    4 days ago
  • $190.2k - $360.5k

    The Opportunity We are looking for a Principal AI Systems Engineer with deep C++ expertise to help build the next generation of AI-enabled product...  ...Engineering Build high-quality C++ components for performance-sensitive, cross-platform environments. Own critical client... 
    Principal
    Performance
    Full time
    Temporary work
    Local area
    Remote work
    Worldwide

    Adobe Systems

    San Jose, CA
    3 days ago
  • $147k - $237.5k

     ..., Integrity, and Inclusion. We weave AI into the fabric of everything we do and...  ...and resolve incidents at scale. As a Principal Software Engineer, you will own the technical vision...  ...or similar techniques to optimize LLM performance and reduce inference costSolid skills... 
    Principal
    Performance
    Full time
    Work at office

    Palo Alto Networks

    Santa Clara, CA
    2 days ago
  • $160k - $220k

     ...everywhere.Fortinet is seeking an experienced and innovative Principal AI Security Engineer to join our Corporate Information Security team. As an AI...  ...closely with development teams, conduct code reviews, perform AI Red Teaming assessments, to identify vulnerabilities... 
    Principal
    Performance
    Full time
    Work experience placement
    Worldwide

    Fortinet

    Sunnyvale, CA
    2 days ago
  • $296.3k - $423.9k

     ...a global scale.  We are looking for a Principal Technical Lead Manager (TLM) to lead the...  ...Generation team within the Embodied AI organization. This role combines deep...  ...world scenarios. You will lead a high-performing team of engineers building ML-driven trajectory generation... 
    Principal
    Performance
    Full time
    Local area
    Remote work
    Work from home
    Relocation
    Relocation package
    Flexible hours

    General Motors

    Sunnyvale, CA
    2 days ago
  • $142.8k - $274.8k

     ...Hardware, and Infrastructure Engineering (SCHIE) is the team behind Microsoft...  ...organization is developing AI-native silicon and hyperscale...  ...custom silicon, high-performance networking, advanced compiler...  ...Engineering (PSE) team is seeking a Principal AI Accelerator Tools... 
    Principal
    Performance
    Ongoing contract
    Work at office
    Local area
    Worldwide
    3 days per week

    Microsoft

    Mountain View, CA
    4 days ago
  • $275.8k - $340.5k

     ...to meet the unique demands of AI and ML innovation, supporting...  ...the productivity of ML engineers, and drive the adoption of cutting...  ...Inference: Ensures robust model performance by running large-scale...  ...Position Overview: The Principal AI/ML Engineer will lead a growing... 
    Principal
    Performance
    Local area
    Remote work
    Work from home
    Relocation
    Relocation package
    Flexible hours

    General Motors

    Sunnyvale, CA
    3 days ago
  • Principal AI EngineerLocation: Hybrid, Santa Clara Industry: Medical Device REQUIRED: proficiency in Chinese to speak to offshore teams...  ...companyWhat We're Looking For7+ years of experience in AI/ML engineering or related fieldsMust be very hands on and technical day to day... 
    Principal

    Real Staffing Group

    Santa Clara, CA
    1 day ago
  • $133.2k - $192.8k

    Job Details: Job Description: AI Engineer - Agentic AI Systems Why This Role Matters AI is shifting from models to autonomous systems...  .../ GPU / FPGA Explore efficient inference and system-level performance tradeoffs Build Platforms, Not Just Features Develop... 
    Performance
    Full time
    Internship
    Local area
    Shift work

    Altera

    San Jose, CA
    1 day ago
  • $152k - $241.5k

    We are seeking a Senior AI/ML Performance and Efficiency Engineer, GPU Clusters at NVIDIA to join our AI Efficiency efforts. As an Engineer, you will have a pivotal role in enhancing efficiency for our researchers by implementing progressions throughout the entire stack... 
    Performance
    Full time
    Remote work

    Nvidia

    Santa Clara, CA
    2 days ago
  • $184k - $287.5k

    We're looking for outstanding AI systems engineers to develop groundbreaking technologies in the inference systems software stack! We build...  ...TVM, MLIR)Strong experience in GPU kernel development and performance optimizations (especially using CUDA C/C++, cuTile, Triton,... 
    Performance
    Full time

    Nvidia

    Santa Clara, CA
    1 day ago
  • $152k - $241.5k

    NVIDIA's GPUs are at the core of modern AI infrastructure, from training large-scale...  ...as much as hardware, and compiler engineering is a big part of what makes it work.We are...  ...optimization passes, and target-specific performance signals.Apply RL techniques to optimize... 
    Performance
    Full time
    Remote work

    Nvidia

    Santa Clara, CA
    1 day ago
  • $152k - $241.5k

     ...breakthroughs in gaming, computer graphics, high-performance computing, and artificial intelligence....  ...powers everything from generative AI to autonomous systems, and we continue...  ..., and tools that enable researchers and engineers to develop the next generation of AI/ML... 
    Performance
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • $152k - $241.5k

    We are now looking for a Senior AI Frameworks Engineer (C++/Python)! NVIDIA's high-performance computing platforms are powering the AI revolution across many applications and industries. Within our software stack, CUTLASS stands out as a popular open-source ecosystem dedicated... 
    Performance
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • $149.52k - $175.9k

     ...ResponsibilitiesLead the design and implementation of enterprise AI and Generative AI solutions, driving adoption of AI...  ...organization in anticipation of future use cases.Partner with engineering, product, architecture, and business stakeholders to identify,... 
    Principal
    Full time
    Work experience placement
    Local area
    3 days per week

    US Bank

    Cupertino, CA
    8 hours ago
  • $152k - $241.5k

    We’re currently seeking a Senior AI Developer Technology Engineer, Financial Sector!Would you like to help shape the future of financial AI and data...  ...system bottlenecks to achieve the best possible performance of computer hardware? Could you be thrilled about an opportunity... 
    Performance
    Full time
    Work experience placement
    Remote work

    Nvidia

    Santa Clara, CA
    2 days ago
  • $150k - $250k

     ...a significant plus.Need to work closely with system and test engineers to develop high speed interface, package/board, and system clocks...  ...into layout optimizations for high speed or high precision performance directly.• Interface verifications between analog and digital.... 
    Principal
    Performance

    Omnivision Technologies

    Santa Clara, CA
    2 days ago
  • $149.75k - $275.58k

     ...innovations to market. Our diverse team of engineers and researchers have pioneered sparse,...  ...will power the coming era of physical AI systems-beyond the reach of GPUs and mainstream...  ...that enable customers to build high-performance physical AI applications powered by... 
    Performance
    Full time
    Internship
    Work at office
    Local area
    Immediate start
    Shift work

    Intel

    Santa Clara, CA
    7 hours ago
  • $152k - $241.5k

     ...tapping into the unlimited potential of AI to define the next era of computing. An era...  ...for an AI & Deep Learning Compiler Engineer. NVIDIA is hiring software engineers for...  ...compiler must deliver leading inference performance, fast build time, reduced memory footprints... 
    Performance
    Full time
    Remote work

    Nvidia

    Santa Clara, CA
    2 days ago
  • $195.2k - $275.58k

    Job Details:Job Description: The Software and AI (SAI) organization is seeking a highly skilled Software Development Engineer to contribute to the development and...  ...oneDNN, a complex, cross‑platform, open‑source performance library for deep learning applications ().oneDNN... 
    Performance
    Full time
    Local area
    Immediate start
    Remote work
    Worldwide
    Flexible hours
    Shift work

    Intel

    Santa Clara, CA
    3 days ago
  •  ...systems for production environmentsImprove performance across GPU and CPU pathwaysWork on KV...  ...systems that support RAG and retrieval-heavy AI workloadsContribute to infrastructure...  ...materially affect AI performanceSolve engineering problems at the intersection of AI, high... 
    Performance

    DataDirect Networks

    Santa Clara, CA
    4 days ago
  •  ...next-generation computing experiences—from AI and data centers, to PCs, gaming and...  ...looking for a Forward Deployed Research Engineer to build, evaluate, and deploy cutting-edge...  ...modeling.Diagnose and optimize AI system performance across models, data, infrastructure, orchestration... 
    Performance

    AMD

    Santa Clara, CA
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Principal AI Performance Engineer. Be the first to apply!