Job Title
ML and AI Knowledge Systems Engineer
Role Summary
Lead engineering of GoldenEye, an internal AI knowledge platform that provides natural-language access to GPU design knowledge (RTL, specifications, verification collateral, Confluence, JIRA) and returns answers with citations. Turn a prototype into a production platform used across hardware and software teams.
Based in San Jose, CA with a hybrid work arrangement; collaborate with hardware, software, infrastructure, and security teams to scale retrieval, ranking, and answer-quality systems.
Experience Level
Senior (PMTS-level). The posting indicates a senior technical lead role; specific years of experience are not stated.
Responsibilities
Primary responsibilities focus on designing, evaluating, and operating retrieval and answer-generation systems that serve engineering knowledge at scale.
- Own architecture for embeddings, chunking, hybrid retrieval, reranking, multi-query planning, and reciprocal-rank fusion.
- Build retrieval workflows across specifications, RTL, verification artifacts, wikis, and issue trackers.
- Improve search for domain-specific tokens (ISA mnemonics, registers, signal names, acronyms, code identifiers).
- Reduce hallucinations through grounded generation, citation enforcement, source validation, and conflict resolution.
- Define source-authority and freshness policies for conflicting or outdated information.
- Build an evaluation framework for retrieval relevance, answer correctness, citation faithfulness, latency, cost, and regressions.
- Use expert feedback, production data, hard negatives, and targeted failure cases to create representative evaluation datasets.
- Track advances in retrieval, RAG, agentic systems, evaluation, and efficient inference; evaluate techniques against production requirements.
- Optimize embedding, reranking, and inference workloads on GPUs, including multi-GPU and multi-node serving.
- Partner with platform teams on scalable services, indexing, data reconciliation, observability, and cluster orchestration.
- Advance Model Context Protocol (MCP) tools used by coding agents and ensure access-control compliance for retrieval and caching.
- Mentor engineers, review designs, and communicate technical decisions across organizations.
Requirements
Must-have technical skills and responsibilities for the role.
- Hands-on experience building production ML, search, ranking, or information-retrieval systems.
- Experience with RAG, embeddings, vector and keyword search, reranking, chunking, and retrieval evaluation.
- Practical experience with LLM prompting, structured generation, tool calling, and hallucination reduction.
- Experience designing ML evaluation datasets, metrics, experiments, and regression tests.
- Strong software and systems engineering skills (APIs, asynchronous processing, observability, production operations).
- Experience leading architecture, mentoring engineers, and influencing senior technical stakeholders.
- Experience with distributed inference systems and GPU performance optimization.
- Experience with vector databases, hybrid search, distributed indexing, or authorization-aware retrieval.
- Role is not eligible for visa sponsorship.
Nice-to-have:
- Experience with AMD Instinct GPUs or comparable accelerators.
- Familiarity with MCP, coding agents, or tool-based agent architectures.
- Experience with hardware design, EDA, RTL, microarchitecture, firmware, compilers, or chip verification.
- Technical thought leadership via publications, patents, or open-source contributions.
Education Requirements
Preferred: Master’s or Doctoral degree in machine learning, information retrieval, NLP, computer science, electrical or computer engineering, or a related field. These academic credentials are listed as preferred; the posting does not state a strict degree requirement or include explicit equivalent-experience language.
About the Company
Company: Advanced Micro Devices
Headquarters: Sunnyvale, California, USA
Advanced Micro Devices, or AMD, is a global semiconductor company that designs and manufactures microprocessors, graphics processors, and related technologies for a variety of computing devices. Known for pushing the boundaries of innovation, AMD's mission is to deliver high-performance computing solutions for AI, data centers, gaming, and embedded applications. They foster a collaborative, inclusive culture focused on creativity and problem-solving, aiming to drive progress and excellence in technology.

Date Posted: 2026-08-20