Lead system-level architecture for connecting XPUs to memory and interconnect fabrics across node, rack, and cluster scales. Define topology and communication strategies for distributed LLM training and inference, resolve integration issues across silicon, networking, memory, and software teams, and mentor system-interconnection architects. The role requires substantial NPU architecture, ASIC tapeout, silicon bring-up, validation, and production deployment experience.
Credo is looking for a Principal AI System Architect to join our team in San Jose, CA, reporting to AVP, XPU system AI interface. This role needs to solve the system-level interconnection challenges that connect XPUs (NPU/GPU/custom silicon) into working AI infrastructure. This role needs to follow and define how compute elements are wired, networked, and orchestrated together at node, rack, and cluster scale so that large-scale LLM training and inference run efficiently across the platform. We strongly prefer candidates with a hands-on NPU hardware background and real experience taking NPU silicon into production deployment at scale.
Base salary range is $220,000 - $280,000 a year. The base salary offer will depend on factors such as education, experience, training, skills, qualifications, and location. This position is also eligible for a discretionary bonus, equity and a full range of medical and other benefits.
Why Credo
Qualifications
Basic Qualifications
Responsibilities
Benefits
Credo’s mission is to transform connectivity at scale through fast, reliable, and energy-efficient system solutions. Our high-speed copper and optical interconnect products deliver industry-leading power and performance at up to 1.6T to meet the ever-expanding data infrastructure demands of AI.
Our product portfolio includes ZeroFlap (ZF) Active Electrical Cables (AECs) and ZF optical transceivers, OmniConnect memory solutions, and a suite of retimers and DSPs for optical and copper Ethernet and PCIe, all leveraging the PILOT diagnostic and analytics software platform. Credo innovations enable our customers to connect the systems that connect the world.
Credo is committed to creating an inclusive environment for all employees and welcome applicants from diverse backgrounds without regard to race, color, religion, gender, sex, gender identity, sexual orientation, pregnancy, marital status, national origin, ethnicity, genetic information, age, disability, veteran status, or any other legally protected basis. If you have a disability or special need that requires accommodation to navigate our website or complete the application process, email [email protected].
Base salary range is $220,000 - $280,000 a year. The base salary offer will depend on factors such as education, experience, training, skills, qualifications, and location. This position is also eligible for a discretionary bonus, equity and a full range of medical and other benefits.
Why Credo
- Purpose: We invest in what matters. From meaningful-future shaping projects to competitive compensation, we empower you to grow your career while making a lasting impact.
- People: Connection starts within. We collaborate, celebrate wins, and create an environment where everyone can do their best work.
- Possibilities: Our belief shapes what’s next. Our technology powers the most reliable and energy-efficient connections around the world – and our team powers new products and markets that come next.
Qualifications
Basic Qualifications
- Bachelor's degree in Computer Engineering, Electrical Engineering, Computer Science, or related field required
- Ten years in XPU or AI interconnect system design, architecture and micro architecture working experience.
- Five cycles of complete ASIC tapeouts.
- Strong foundational understanding of computer architecture, including GPU, NPU, and/or CPU design principles; deep, hands-on NPU hardware architecture expertise.
- Familiarity with large language model architectures and distributed training/inference concepts (e.g., parallelism strategies, model serving) through coursework, research, or personal projects.
- Demonstrated analytical ability, for example, through published research, thesis work, or quantitative project work- with a track record of independently studying and synthesizing technical material.
- Strong cross-functional leadership and communication skills, daily experience working with compiler, software, SOC & backend team.
- Master's degree or PhD in a relevant technical field.
- Direct NPU tapeout and production ramp experience.
- Hands-on NPU hardware background, with direct experience carrying an NPU design from architecture definition through silicon bring-up, validation, and volume production deployment.
- Working knowledge of high-speed interconnect or networking concepts (e.g., PCIe, Ethernet, RDMA).
Responsibilities
- Define the memory and interconnection architecture linking XPUs to memory and across node, rack, and cluster boundaries -not the internal design of the NPUs/GPUs/CPUs themselves.
- Solve communication and topology bottlenecks that arise from all kinds of parallelism strategies during distributed LLM training and inference.
- Define how XPUs utilizes the memory and interconnect fabric, ensuring workloads map cleanly onto the network topology.
- Work with silicon, memory, networking, and software teams to specify interconnect fabric requirements and resolve integration issues between compute nodes.
- Bring hands-on NPU hardware architecture judgment to interconnect and system design decisions, informed by direct experience taking NPU silicon from definition through bring-up, validation, and production deployment.
- Track leading open LLM architectures to anticipate how model structure will stress interconnect and system topology.
- Lead and mentor a team of architects focused on system interconnection; represent this strategy to leadership and partners.
Benefits
Credo’s mission is to transform connectivity at scale through fast, reliable, and energy-efficient system solutions. Our high-speed copper and optical interconnect products deliver industry-leading power and performance at up to 1.6T to meet the ever-expanding data infrastructure demands of AI.
Our product portfolio includes ZeroFlap (ZF) Active Electrical Cables (AECs) and ZF optical transceivers, OmniConnect memory solutions, and a suite of retimers and DSPs for optical and copper Ethernet and PCIe, all leveraging the PILOT diagnostic and analytics software platform. Credo innovations enable our customers to connect the systems that connect the world.
Credo is committed to creating an inclusive environment for all employees and welcome applicants from diverse backgrounds without regard to race, color, religion, gender, sex, gender identity, sexual orientation, pregnancy, marital status, national origin, ethnicity, genetic information, age, disability, veteran status, or any other legally protected basis. If you have a disability or special need that requires accommodation to navigate our website or complete the application process, email [email protected].
Similar Jobs
Software • Quantum Computing • Metaverse • Infrastructure as a Service (IaaS)
Owns AI rack- and cluster-scale hardware system architecture for Microsoft Azure infrastructure. Defines requirements and specifications, evaluates architecture tradeoffs across accelerators, networking, memory, storage, power, cooling, and datacenter constraints, and develops proof-of-concepts and performance models. The role drives cross-functional engineering alignment, influences industry standards and vendor roadmaps, and helps deliver scalable AI accelerator systems.
Top Skills:
Ai AcceleratorsCloud InfrastructureHardware ModelingInfinibandPerformance SimulationsRoce
Hardware • Information Technology • Semiconductor • Manufacturing
Lead system-level architecture for AI inference and advanced memory solutions, optimize compute/memory bottlenecks, map ML workloads to hardware, drive hardware/software co-design, and provide technical leadership across chiplet, SoC, and memory subsystem designs.
Top Skills:
Ai AcceleratorsCC++ChipletCxlDdrDramEthernetGddrGpuIn-Memory ComputingLpddrNandNear-Memory ComputingNeural Processing EngineNocPciePythonRdmaRtl DesignSocSystem-Level Modeling
Artificial Intelligence • Cloud • HR Tech • Information Technology • Productivity • Software • Automation
Leads strategic customer adoption engagements for ServiceNow’s AI Platform and Now Platform capabilities. Manages complex cross-functional programs, project governance, implementation planning, risk mitigation, stakeholder communication, customer value propositions, and successful rollouts. Partners with product teams, customers, field teams, and external partners to drive Generative AI adoption, resolve delivery gaps, mentor team members, and ensure knowledge transfer and engagement completion.
Top Skills:
Generative AiNow PlatformPmpSaaSScrumServicenow
What you need to know about the San Francisco Tech Scene
San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.
Key Facts About San Francisco Tech
- Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
- Major Tech Employers: Google, Apple, Salesforce, Meta
- Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
- Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
- Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
- Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine



