Platform Reliability, Availability, Serviceability (RAS), and Manageability Software Architect. Prin
$211.8k - $317.8kQualcomm
Company:
Qualcomm Technologies, Inc.
Job Area:
Engineering Group, Engineering Group> Software Engineering
General Summary:
Hiring in Santa Clara, San Diego, Portland, Austin
Qualcomm Data Center team is developing High performance, Energy efficient server solution for data center applications. We are looking for highly talented, innovative, teamwork-oriented individuals for our cutting-edge technology work!
Our Mission
We are dedicated to transforming the industry by reimagining silicon and developing next-generation computing platforms. By joining our team, you'll collaborate with world-class engineers to create innovative solutions that push the limits of performance, energy efficiency, and scalability. Our focus is on developing reference platforms based on Qualcomm's Snapdragon SoC, delivering a comprehensive solution that includes hardware, software, reference designs, user guides, SDKs, and more.
Job Summary
Qualcomm is seeking an experienced Platform RAS and Manageability Software Architect to design and implement robust RAS and manageability solutions for our cutting-edge compute platforms based on Qualcomm Oryon CPUs. The ideal candidate will have a deep understanding of ARM architecture, RAS principles, and system manageability considerations.
As a Software Architect, you will take on a technical leadership role within the product software team, working closely with cross-functional engineering teams, including SoC, CPU, software, firmware, and software product management. You will be responsible for architecting a high-reliability software stack, minimizing downtime and ensuring efficient maintenance cycles.
Minimum Qualifications:
• Bachelor's degree in Engineering, Information Systems, Computer Science, or related field and 8+ years of Software Engineering or related work experience.
OR
Master's degree in Engineering, Information Systems, Computer Science, or related field and 7+ years of Software Engineering or related work experience.
OR
PhD in Engineering, Information Systems, Computer Science, or related field and 6+ years of Software Engineering or related work experience.
• 4+ years of work experience with Programming Language such as C, C++, Java, Python, etc.
Preferred Qualifications:
10+ years of experience in designing software and firmware for various compute environments.
Strong expertise in modern operating systems, ARM64 architectures, hypervisors, software reliability and manageability, and software development methodologies.
Deep proficiency in Linux kernels, RAS, System Manageability, DDR, PCIe, and communication protocols such as I2C, SPI, and MDIO.
Practical experience with in-lab debugging tools.
Strong technical documentation skills and excellent written and verbal
Good to Have:
Master's degree in Computer Science/Engineering, Electrical Engineering, or a related field.
15+ years of experience in software development&design for commercially deployed compute platforms.
In-depth knowledge of ARM architectures for various compute environments and relevant BSA, BBR, and Manageability specifications.
Deep understanding of ARM RAS specification, ARM CPU RAS extensions, and Software components (SDEI, APEI, UEFI CPER) specifications.
Proven success in architecting and delivering solutions for a commercially deployed compute environment.
Principal Duties and Responsibilities:
Architect and help design RAS features for compute platforms and develop manageability solutions to monitor and maintain system health.
Actively engage with the ARM and OCP community to stay updated on the latest developments and ensure alignment of the architected solution with the community direction.
Collaborate with cross-functional teams, including SoC, CPU HW, HLOS, and BIOS software teams, to ensure seamless adoption and integration of the architected solution.
Provide technical leadership and oversight to various HW and SW teams involved, ensuring compatibility with ARM specifications.
Collaborate with customers to guide and support the development of custom software solutions leveraging Qualcomm CPUs.
Prepare and present clear and comprehensive technical documentation and reports tailored to the needs of stakeholders, including engineering teams, senior management, customers, and suppliers.
Partner with internal teams, marketing, end-customers, OEMs, and suppliers to create software roadmaps and detailed requirement documentation.
Qualcomm is an equal opportunity employer. If you are an individual with a disability and need an accommodation during the application/hiring process, rest assured that Qualcomm is committed to providing an accessible process. You may e-mail View email address on techcareers.com or call Qualcomm's toll-free number found here . Upon request, Qualcomm will provide reasonable accommodations to support individuals with disabilities to be able participate in the hiring process. Qualcomm is also committed to making our workplace accessible for individuals with disabilities. (Keep in mind that this email address is used to provide reasonable accommodations for individuals with disabilities. We will not respond here to requests for updates on applications or resume inquiries).
To all Staffing and Recruiting Agencies Our Careers Site is only for individuals seeking a job at Qualcomm. Staffing and recruiting agencies and individuals being represented by an agency are not authorized to use this site or to submit profiles, applications or resumes, and any such submissions will be considered unsolicited. Qualcomm does not accept unsolicited resumes or applications from agencies. Please do not forward resumes to our jobs alias, Qualcomm employees or any other company location. Qualcomm is not responsible for any fees related to unsolicited resumes/applications.
EEO Employer: Qualcomm is an equal opportunity employer; all qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, Veteran status, or any other protected classification.
Qualcomm expects its employees to abide by all applicable policies and procedures, including but not limited to security and other requirements regarding protection of Company confidential information and other confidential and/or proprietary information, to the extent those requirements are permissible under applicable law.
Pay range and Other Compensation&Benefits
$211,800.00 - $317,800.00
The above pay scale reflects the broad, minimum to maximum, pay scale for this job code for the location for which it has been posted. Even more importantly, please note that salary is only one component of total compensation at Qualcomm. We also offer a competitive annual discretionary bonus program and opportunity for annual RSU grants (employees on sales-incentive plans are not eligible for our annual bonus). In addition, our highly competitive benefits package is designed to support your success at work, at home, and at play. Your recruiter will be happy to discuss all that Qualcomm has to offer– and you can review more details about our US benefits at this link
If you would like more information about this role, please contact Qualcomm Careers
- ...for system deployment, management and maintenance.What... ...the validation of platform quality.Design, build... ...scalable control plane services, operators, and custom... ..., and platform reliability.You6+ years of experience... ...values are publicly available: We offer generous cash...PlatformWork at officeLocal areaWork from homeFlexible hours
$262k - $364k
Lead a team of Software/Systems Engineers on... ...uptime.Own end-to-end availability and performance of key services and build... ...technical execution.Manage on-call rotations... ...related field.Site Reliability Engineering (SRE)... ...flagship collaboration platform.Behind everything...Platform$207k - $300k
Manage a team of Software/Systems Engineers on projects for users and... ...uptime.Own the end-to-end availability and performance of key services, build automation to... ...expertise in Site Reliability Engineering practices,... ...-generation of Google platforms, we make Google's product...Platform- ...develop a team of Site Reliability/Production Engineers,... ...Own the reliability, availability, and operational... ...health of critical Cortex services and infrastructure.... ..., alerting, incident management, and observability to... ...technical knowledge of cloud platforms, preferably Google...PlatformFull timeWork at officeVisa sponsorshipWork visa
$216.14k - $282.98k
...Senior Staff Software Architect – Network Monitoring... ...leading quantum platform and merchant... ...including Amazon Web Services, and AstraZeneca... ...services are available through all major... ...platform operates reliably at global scale... ..., fleet management, and real-time alerting...PlatformPermanent employmentContract workWork at office$245k - $325k
...generative AI platform, from chip to... ...hardware and software platform. Our... ...hiring a Software Architect for the... ...enterprises and service providers host... ...seamless deployment, management, and scaling... ...scalability, reliability, and... ...being benefits available to you and your...PlatformFull timeTemporary workLocal areaFlexible hours$212k - $340k
...Staff Tech Lead Manager to join the Aurora Services Engineering... ...is to build the software that drives the... ...commercial side of our platform, such as... ...operation of high-availability solutions for... ...mission-critical reliability, while... ...Contribution: Architect, design, and operate...PlatformWork at officeLocal areaRemote work3 days per week$168k - $258.75k
...a Senior Technical Program Manager (TPM) to lead complex, cross... ...NVIDIA’s next-generation AI software platforms. In this role, you will drive... ...initiatives across platform services, cloud infrastructure, and system... ...is on enabling scalable, reliable, and supportable software...PlatformFull time$255.7k - $300k
...team of engineers to maintain service uptime while managing global on-call rotations and evaluating... ....Design and deliver software that optimizes availability, scalability, and efficiency while... ...operational practices to drive reliability, maintainability, and stakeholder...Full timeWork at office$151.8k - $265.35k
...Firefly’s Generative AI Services team is seeking... ...performance, scalability, and reliability.Foster a culture of... ..., and GPU resource management in large-scale... ...distributed systems, and MLOps platforms.#FireflyGenAIAbout... ...needs and position availability.Massachusetts: Massachusetts...PlatformFull timeTemporary workLocal areaWorldwide- ...generation AI server platforms powered by AMD... ...Team Leadership: Manage a team of... ...programs that ensure reliability at scale.Customer... ...architecture, ROCm software stack, and performance... ...recruitment services. AMD and its subsidiaries... ...AI Policy” is available here.This posting...PlatformFor contractors
- ...PositionAs a Web Access Management Engineer... ...on-prem and SaaS platforms. You will play a... ...ReliabilityDevelop and maintain reliable and scalable... ...Science, Software Engineering, or a... ...JIT) and directory services (LDAP, Active... ...benefits are not available for this job posting...PlatformFull timeImmediate startRelocation package
$207k - $300k
Manage and mentor a fast-expanding team of Software Engineers, coordinating closely with... ...networking, and reliability engineering.... ...internationally.The Network Service Level Objective (... ...end-to-end availability and performance,... ...providing the essential platforms that enable...PlatformWorldwide$222k - $300.5k
...financial technology platform that powers... ...Infrastructure and Site Reliability organization owns the... ...movement and fintech services — where availability, data integrity, and... ...'re hiring a Senior Manager, Site Reliability Engineering... ...closely with software engineering, product...PlatformWorldwideShift work$207k - $300k
...highly scalable software and... ...and mitigate reliability issues across... ...for fault management. Guide the... ...reliability.Architect pipelines to... ...platform, SQL pipelines... ...the Server RAS team, you will... ..., maximize availability, and lower... ...Availability, and Serviceability (RAS)...PlatformWorldwide$101k - $161k
...artificial intelligence, and software-defined networking to... ...’re looking for Site Reliability Engineers to join our... ...’s CloudVision-as-a-Service (CVaaS) global SRE... ...an enterprise network management and streaming telemetry... ...with GCP (Google Cloud Platform) and GKE (Google Kubernetes...Platform$212.7k - $287.7k
...control-plane services that keep HyperPlane... ..., and always available at massive... ...looking for a Software Development Manager (SDM III) to grow... ...Virtualization (NFV) platform. Networking... ...systems fast and reliable as demand keeps... ...designing or architecting (design patterns...PlatformLocal areaFlexible hoursShift work$207k - $300k
Lead a team of Software/Systems Engineers on... ...uptime.Own end-to-end availability and performance of key services and build... ...technical execution.Manage on-call rotations... ...experience. Site Reliability Engineering (SRE)... ...generation of Google platforms, we make Google's...Platform$185.9k - $300.68k
...SummaryAs the Senior Manager of Product Management for Operations & Reliability, you will build and lead... ...our AI Firewall services meet a 99.99% SLA. This... ...cloud-delivered security platform.Partner with Engineering... ...to improve availability, resilience, and efficiency...PlatformFull timeWork at office- Locations available: San Diego and San Jose,... ...track record in architecting and delivering cutting... ...with product management, R&D, and customer... ...quality, reliability, and security standards... ...knowledge of hardware/software co-design, cloud-... ...tools and platforms (e.g., Verilog, VHDL...PlatformFull timeWork at officeLocal area
$152k - $241.5k
As a Software Solution Architect, NVIS at NVIDIA, you will lead the transformation... ..., an agentic software platform with tools, services, and AI agents that... ...augmented generation, context management, agent memory, function... ...and guardrails to build reliable AI systems.Integrate...PlatformFull time$207k - $300k
...response, maintain high reliability standards, and... ...years of experience with software development in one or... ...3 years of experience managing and growing engineering... ...ensures that Google's services—both our internally critical... ...elements, developer platforms, product components,...Platform$350k
...#1 TV streaming platform in the U.S., Canada... ...create simple, reliable, and delightful experiences... ...hardware and software to create a... ...Systems Software Architect who leads complex... ...HALs, and low-level services that expose the... ...every benefit is available in all locations...PlatformWork at officeLocal areaRemote workMonday to ThursdayFlexible hours$171.5k - $245k
...Zero Trust Exchange platform. This innovation... ...Principal Product Manager, Cloud Reliability and Cloud Operations... ...Implement the ZIA/ZPA RAS roadmap by defining... ...with cloud architects and software engineers to ensure... ...with cloud-native services at hyper-scalers or...PlatformFull timeWork at officeLocal area- ...of experience in Site Reliability Engineering, DevOps,... ...Practical vulnerability-management experience;... ...thousands of systems. Available to be online from 9am... ...continued growth of a platform serving a rapidly growing... ...tools to improve software development, automation...PlatformFull timeWorldwide
$169k - $338k
...Global Tech's Site Reliability Engineering... .... You will architect and implement... ...machine learning platforms and autonomous... ...systems and software engineers who... ...related to uptime, availability and fast rate... ...capacity management and predictive... ...technology, financial services, and all...PlatformFull timeTemporary workPart time$188k - $274k
...learning system power management and optimization... ...across silicon, platform hardware, rack power, system software, and... ...performance, cost, reliability, and availability.Partner with power... ...collaboration with system architects, hardware... ...all of Google's services. In this role,...PlatformWorldwide$83.8k - $150.2k
...design, intelligent software, and next-... ...global scaleThe AV Service Operations Program Manager coordinates the programs... ...Management, Fleet Reliability, technicians, suppliers... ...process, tool, and platform training; identify... ...repair documentation available when needed....PlatformFull timeFor contractorsLocal areaWork from homeFlexible hours$193.3k - $261.5k
AWS Infrastructure Services owns the design, planning... ...hardest problems that fuse software, hardware, and the... ...the next-generation platform. You will have a direct... ...security experts, operations managers, and other vital roles... ...(design patterns, reliability and scaling) of new...PlatformInternshipLocal areaFlexible hours$171.5k - $245k
...Principal Product Manager, Platform Infrastructure and Premium Connectivity Services San Jose, California... ...for driving reliability and scalability and... ...Implement the ZIA/ZPA RAS roadmap by defining... ...Collaborate with cloud architects and software engineers to ensure...PlatformWork at officeLocal area
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Platform Reliability, Availability, Serviceability (RAS), and Manageability Software Architect. Prin. Be the first to apply!
- software architect Santa Clara, CA
- application architect Santa Clara, CA
- senior software architect Santa Clara, CA
- .net software architects (remote) Santa Clara, CA
- platform manager Santa Clara, CA
- power platform Santa Clara, CA
- platform product manager Santa Clara, CA
- software sales representative Santa Clara, CA
- embedded software Santa Clara, CA
- software applications developer Santa Clara, CA


