ML and AI Knowledge Systems Engineer
Socket.dev
ADVANCE YOUR CAREER. ADVANCE THE WORLD.
At AMD, we believetechnology has the power to solve the world’s most important challenges. From advancing healthcare and scientific discovery to powering AI and the technologies people rely on every day, innovation at AMDis shapingthefuture. Whetheryou’redesigning next-gen processors, enabling AI breakthroughs, orbringing leading edge products to market, every role at AMD contributes to something bigger— technologythat moves the world forward.Join us and, together, we’ll advance your career. The ROLE: We are building GoldenEye, an internal AI knowledge platform that lets engineers ask natural-language questions across GPU design knowledge, including RTL, microarchitecture specifications, verification collateral, Confluence, and JIRA, and receive trustworthy answers with citations. A prototype is already used by hardware and software teams. We are now scaling it to thousands of users, with the potential to serve more than 40,000 engineers worldwide. We are seeking a hands-on PMTS-level ML engineer to lead GoldenEye’s retrieval, ranking, and answer-quality architecture. You will turn a promising prototype into an authoritative production platform while setting technical direction, mentoring engineers, and partnering with hardware, software, infrastructure, and security teams. GoldenEye will make critical engineering knowledge easier to find, verify, and use, accelerating GPU design, verification, debugging, and bring-up. You will serve as the technical anchor for its ML core and shape its adoption across the engineering organization.KEY RESPONSIBILITIES:
Own the architecture for embeddings, chunking, hybrid retrieval, reranking, multi-query planning, and reciprocal-rank fusion. Build retrieval workflows across specifications, RTL, verification artifacts, wikis, and issue trackers. Improve search for domain-specific content such as ISA mnemonics, registers, signal names, acronyms, and code identifiers. Reduce hallucinations through grounded generation, citation enforcement, source validation, and conflict resolution. Define source-authority and freshness policies for conflicting or outdated information. Build an evaluation framework for retrieval relevance, answer correctness, citation faithfulness, latency, cost, and regressions. Use expert feedback, production data, hard negatives, and targeted failure cases to create representative evaluation datasets. Track advances in retrieval, RAG, agentic systems, evaluation, and efficient inference, and evaluate promising techniques against production requirements. Optimize embedding, reranking, and inference workloads on AMD Instinct GPUs, including multi-GPU and multi-node serving. Partner with platform teams on scalable services, indexing, data reconciliation, observability, and cluster orchestration. Advance Model Context Protocol (MCP) tools used by coding agents. Ensure retrieval, caching, and answer generation respect per-user access controls. Mentor engineers, review designs, and communicate technical decisions across organizations.PREFERRED EXPERIENCE:
Deep relevant experience building production ML, search, ranking, or information-retrieval systems. Hands-on experience with RAG, embeddings, vector and keyword search, reranking, chunking, and retrieval evaluation. Practical experience with LLM prompting, structured generation, tool calling, evaluation, and hallucination reduction. Experience designing ML evaluation datasets, metrics, experiments, and regression tests. Demonstrated ability to translate current ML or information-retrieval research into rigorous experiments and production improvements. Strong software and systems engineering skills, including APIs, asynchronous processing, observability, and production operations. Experience leading architecture, mentoring engineers, and influencing senior technical stakeholders. Experience with distributed inference systems such as vLLM and GPU performance optimization. Experience with AMD Instinct GPUs or comparable accelerators. Experience with vector databases, hybrid search, distributed indexing, or authorization-aware retrieval. Familiarity with MCP, coding agents, or tool-based agent architectures. Experience with hardware design, EDA, RTL, microarchitecture, firmware, compilers, or chip verification. Technical thought leadership demonstrated through publications, patents, open-source contributions, conference participation, or influential production systems. Publications are valued but not required.PREFERRED ACADEMIC CREDENTIALS:
Master’s or Doctoral degree in machine learning, information retrieval, NLP, computer science, electrical or computer engineering, or a related field. Location: San Jose, CA#LI-G11
#LI-HYBRID
This role is not eligible for visa sponsorship. Benefits offered are described: AMD benefits at a glance. AMD does not accept unsolicited resumes from headhunters, recruitment agencies, or fee-based recruitment services. AMD and its subsidiaries are equal opportunity, inclusive employers and will consider all applicants without regard to age, ancestry, color, marital status, medical condition, mental or physical disability, national origin, race, religion, political and/or third-party affiliation, sex, pregnancy, sexual orientation, gender identity, military or veteran status, or any other characteristic protected by law. We encourage applications from all qualified candidates and will accommodate applicants’ needs under the respective laws throughout all stages of the recruitment and selection process. AMD may use Artificial Intelligence to help screen, assess or select applicants for this position. AMD’s “Responsible AI Policy” is available here. This posting is for an existing vacancy. #J-18808-Ljbffr Socket.devVacancy posted 5 days ago
Similar jobs that could be interesting for youBased on the ML and AI Knowledge Systems Engineer in San Jose, CA vacancy
- ...scientific discovery to powering AI and the technologies people... ...GoldenEye, an internal AI knowledge platform that lets engineers ask natural-language... ...seeking a hands‑on PMTS-level ML engineer to lead GoldenEye’... ...in retrieval, RAG, agentic systems, evaluation, and efficient...SuggestedWorldwide
- Advanced Micro Devices in San Jose, CA, seeks a hands-on PMTS-level ML engineer to lead GoldenEye’s retrieval, ranking, and answer-quality architecture. You will transform a prototype into a production platform, setting technical direction and mentoring engineers across...Suggested
- AMD in San Jose, CA is hiring a hands-on ML engineer to lead GoldenEye’s retrieval, ranking, and answer-quality architecture. You will scale a prototype into a production platform, setting technical direction and mentoring engineers across hardware, software, and security...Suggested
- AMD is building GoldenEye, an internal AI knowledge platform helping engineers query GPU design knowledge with trusted citations. The role leads ML architecture for embeddings, retrieval, and answer quality, partnering with hardware, software, and security teams to scale...Suggested
$165.2k - $223.6k
...learning and mastering challenging engineering domains, and shaping the future of AI-driven cloud services, then we... ...domains (like security, AI/ML, distributed systems, emerging AWS services, etc.),... ...engineers and foster a culture of knowledge-sharing, while maintaining...SuggestedInternshipLocal areaFlexible hours$186.2k - $316.5k
...without us. KLA invents systems and solutions for the... ...expert teams of physicists, engineers, data scientists and... ...a builder-first AI Engineering Architect to... ...digital thread, MBSE, and knowledge-grounded engineering workflows... ...Computer Science, AI/ML, Data Science,...Minimum wageFull timeWork experience placementFlexible hours$132.3k - $241.5k
...team members to build and deliver AI-powered software initiatives... ...management Bachelor's degree in software engineering, computer science, or a related field. Knowledge of software engineering standard... ...to detail Experience shipping ML-powered products Thorough understanding...Relocation$125k - $175k
...Blue Acorn iCi is seeking an AI Engineer to design and ship AI-... ...AI solutions with enterprise systems including Adobe Experience Platform... .... ~2+ years shipping AI or ML-powered solutions in... ..., or equivalent. ~ Working knowledge of prompt engineering, tool use...Full timeTemporary workH1b3 days per week- ..., high-performance computing, cloud, and AI. Whether you’re designing next-gen processors... ...forward.THE ROLE:We are seeking an AI Systems Engineer to join our AMD IT compute platforms... ...ensuring optimal performance Administer ML/AI platforms - Distributed ML services, LLMs...
$152k - $208.5k
...global leader in materials engineering solutions used to... ...connect our world - like AI and IoT. If you want to... ...expertise in intricate systems, deciphering code, and... ...experience integrating AI/ML models into developer... ...optimization.Practical knowledge of RAG and vector search...Full time$160k - $198k
...team members.What You’ll DoAs a Senior AI Systems Engineer, you will architect, deploy, and manage... ...experience with a dedicated focus on AI/ML systems, high-performance computing (... ...determined by factors such as job-related knowledge, skills, and experience.Archer is proud...Local area$190k - $237k
...team members.What You’ll DoAs a Staff AI Systems Engineer, you will architect, deploy, and manage... ...experience with a dedicated focus on AI/ML systems, high-performance computing (HPC... ...by factors such as job-related knowledge, skills, and experience.Archer is proud...Local area$189.3k - $320.7k
...breakthrough hardware and battery systems to intuitive design,... ...driving? Join the Embodied AI team at General Motors. Our... ...world scenarios.As a Staff AI/ML Future Sensing Engineer in the Embodied AI... ...and code reviews, fostering knowledge sharing and engineering excellence...Full timeLocal areaRemote workWork from homeRelocationRelocation packageFlexible hours- ...computing, cloud, and AI. Whether you’re designing... ...of large-scale AI/ML clustered infrastructure... ...team of multi-disciplined engineers that operates across industry... ...supportWorking knowledge of web application security... ...: Linux operating systems, networking, filesystems...Flexible hours
$125.5k - $261.6k
...The opportunity We are seeking AI Systems Engineers to build and operate the foundational substrate... ...have Experience supporting AI/ML and GPU-intensive workloads (GPU... ...not limited to education, experience, knowledge, skills and geography. In addition, our...Full timeContract workSummer holidayFlexible hours- ...AI SYSTEMS ENGINEERAt AMD, we believe technology can change lives for... ...ROLEAMD is seeking an AI Systems Engineer to help develop and optimize... ..., designing high-performance ML operator kernels, optimizing... ...and performance optimization.Knowledge of machine learning inference...Worldwide
- AMD is seeking an AI Systems Engineer to design, deploy, and manage HPC/AI infrastructure, GPU clusters, and AI workload schedulers. You will... ...security. You will optimize distributed AI workloads, implement ML platforms, and advance automation for provisioning and...
- AMD seeks an AI Systems Engineer to advance ML workloads on AMD AI accelerators, bridging hardware and software from kernel design to production inference across NPU and GPU platforms. You will collaborate with compiler, runtime, silicon, and architecture teams, delivering...
$184k - $287.5k
...into the unlimited potential of AI to define the next era of... ...looking for talented Software Engineers to work on developing and deploying... ...building AI-driven software systems, ideally applied to complex engineering... .... Proficiency with modern ML frameworks (PyTorch,...- AMD is seeking an AI Systems Engineer in San Jose to develop and optimize ML workloads on next‑gen AMD accelerators. You will design high‑performance operator kernels, optimize dataflow, and enable AI inference across NPU and GPU platforms. You will work across compiler...
- ...designed, high-performance AI applications.... ...single-node and distributed systems. Requirements Bachelor... ...science, electrical engineering, or a related field... ...with at least one major ML framework: PyTorch, TensorFlow... ..., cuDNN, or cuBLAS. Knowledge of memory hierarchy...Full timeTemporary workFlexible hours
- NVIDIA Corporation in Santa Clara, CA seeks a Software Engineer to develop and deploy AI agents that optimize VLSI design flows. You will work with ML models, Python production systems, and frameworks like PyTorch/TensorFlow to enhance efficiency and quality across design...
- The era of pervasive AI has arrived. In this era, organizations... ...a talented and driven ML performance engineer to optimize and scale state-... ...gap between deep learning and systems performance, collaborating across... ...libraries is a plus. Knowledge of memory hierarchy optimization...Full timeTemporary workLocal areaFlexible hours
$144k - $180k
...team members. What You’ll Do As a Senior AI Systems Engineer, you will architect, deploy, and manage... ...with a dedicated focus on AI/ML systems, high-performance computing (HPC... ...determined by factors such as job-related knowledge, skills, and experience. Archer is proud...Local area- ...Matrix, headquartered in Santa Clara, CA, seeks a Principal System Software Engineer for AI Inference Execution. You will join the software team to... ...stack, developing deployment software and collaborating with ML, compiler, and hardware experts. Required: strong...
$144k - $180k
Archer Aviation in San Jose seeks a Senior AI Systems Engineer to architect, deploy, and manage the critical infrastructure for large-scale AI model training and inference, ensuring robust, low-latency performance. You will work with researchers and software engineers to...- NVIDIA’s Cosmos team seeks engineers to build agentic AI-native software and tooling that accelerates model... .../PyTorch codebases, design end-to-end ML pipelines, and create self-improving... ...deployment. We expect deep expertise in ML systems, robust software engineering, and...
- TikTok is building a Knowledge Graph team focused on the neural network backend for e-commerce, including product categorization, influencer and merchant data, and tag/categorical attributes. The team collaborates with product, data science, and operations to craft advanced...
- ...generation computing experiences—from AI and data centers, to PCs, gaming and embedded systems. Grounded in a culture of... ...is seeking an experienced AI/ML Engineer to join our Applied Research and... ...pipelines in production environments.Knowledge of computer vision, image...
$136.5k - $276.5k
AI/ML Engineer - AgenticThis role has been designed as ‘Hybrid’ with an expectation that you... ...services, and high‑performance backend systems that power agent execution. This position... ....Typically, 4-7 years’ experience.Knowledge and Skills:Core Agentic/Orchestration:Production...Full timeWork experience placementWork at officeLocal areaImmediate start2 days per week
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to ML and AI Knowledge Systems Engineer. Be the first to apply!
Related searches
- machine learning ai engineer San Jose, CA
- ai developer San Jose, CA
- senior ai engineer San Jose, CA
- ai engineer San Jose, CA
- ai ml engineer San Jose, CA
- ai prompt engineer San Jose, CA
- ai engineer remote San Jose, CA
- system engineer remote San Jose, CA
- senior windows systems engineer San Jose, CA
- systems engineer intern San Jose, CA


