Tenstorrent Inc. Logo

Tenstorrent Inc.

Performance Architect, CPU Cluster

Posted 4 Days Ago
Be an Early Applicant
In-Office
Santa Clara, CA, USA
100K-500K Annually
Entry level
In-Office
Santa Clara, CA, USA
100K-500K Annually
Entry level
Analyze and optimize multicore CPU cluster performance across caches, coherence, interconnects, memory bandwidth, latency, and scalability. Build performance models and simulations using Gem5 or equivalent, study real workloads, identify bottlenecks, lead architectural tradeoff studies, and translate performance insights into design decisions through collaboration with CPU, RTL, compiler, software, and system architecture teams.
The summary above was generated by AI

Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. With AI redefining the computing paradigm, solutions must evolve to unify innovations in software models, compilers, platforms, networking, and semiconductors. Our diverse team of technologists have developed a high performance RISC-V CPU from scratch, and share a passion for AI and a deep desire to build the best AI platform possible. We value collaboration, curiosity, and a commitment to solving hard problems. We are growing our team and looking for contributors of all seniorities.

Tenstorrent is looking for a CPU Cluster Performance Architect to help shape the performance and scalability of our next-generation CPUs. You’ll focus on how multiple CPU cores work together as a cluster — from cache hierarchies and coherent interconnects, to memory bandwidth and overall system behavior. This is a hands-on architecture role for someone who enjoys using performance models and simulation to understand bottlenecks, explore tradeoffs, and influence decisions early in the design process. You’ll work closely with CPU architects, RTL designers, software and compiler teams, and system architects to understand workload behavior and turn performance insights into architectural improvements. If you enjoy looking at the entire cluster and how performance scales across cores, memory systems, and chiplets, this is an opportunity to have a significant impact.

This role is remote, based out of North America.

We welcome candidates at various experience levels for this role. During the interview process, candidates will be assessed for the appropriate level, and offers will align with that level, which may differ from the one in this posting.


Who You Are

  • You have a strong foundation in CPU microarchitecture and performance, with an understanding of how cores, caches, interconnects, and memory systems interact.
  • You enjoy using modeling, simulation, and workload analysis to understand cluster-level performance bottlenecks and scaling challenges.
  • You’re comfortable moving between detailed microarchitecture and broader system-level questions around latency, bandwidth, quality of service, utilization, and scalability.
  • You’re analytical and hands-on, with the ability to turn large amounts of performance data into clear architectural recommendations.
  • You’re a strong communicator who enjoys working across architecture, RTL, compiler, software, and system teams.

What We Need

  • Analyze and optimize CPU cluster performance, including cache hierarchies, interconnects, coherence, and memory access behavior.
  • Build and use performance models and simulation environments such as Gem5 or equivalent to evaluate architectural concepts and identify performance opportunities.
  • Study real workloads to understand bottlenecks related to cache misses, coherence traffic, memory bandwidth, latency, contention, and core-to-core communication.
  • Lead architectural tradeoff studies across performance, scalability, bandwidth, latency, power, and implementation complexity.
  • Collaborate with CPU, cache, interconnect, memory, software, and system architects to translate performance analysis into concrete design decisions.

What You Will Learn

  • How to architect, model, and optimize a multi-core CPU cluster from early performance exploration through implementation and silicon.
  • How cache coherence, interconnects, memory hierarchy, and core architecture interact to determine overall system performance.
  • How to model and reason about scaling across cores, including contention, bandwidth limits, latency, and workload behavior.
  • How architectural decisions at the cluster level translate into real-world application performance.
  • How Tenstorrent approaches scalable CPU and chiplet-based system architecture across compute, interconnect, and memory.

Compensation for all engineers at Tenstorrent ranges from $100k - $500k including base and variable compensation targets. Experience, skills, education, background and location all impact the actual offer made.

Tenstorrent offers a highly competitive compensation package and benefits, and we are an equal opportunity employer.

This offer of employment is contingent upon the applicant being eligible to access U.S. export-controlled technology.  Due to U.S. export laws, including those codified in the U.S. Export Administration Regulations (EAR), the Company is required to ensure compliance with these laws when transferring technology to nationals of certain countries (such as EAR Country Groups D:1, E1, and E2).   These requirements apply to persons located in the U.S. and all countries outside the U.S.  As the position offered will have direct and/or indirect access to information, systems, or technologies subject to these laws, the offer may be contingent upon your citizenship/permanent residency status or ability to obtain prior license approval from the U.S. Commerce Department or applicable federal agency.  If employment is not possible due to U.S. export laws, any offer of employment will be rescinded.

Tenstorrent Inc. Santa Clara, California, USA Office

Santa Clara, California, United States

Similar Jobs

Yesterday
Hybrid
2 Locations
230K-286K Annually
Senior level
230K-286K Annually
Senior level
Fintech • Machine Learning • Payments • Software • Financial Services
Leads multiple technology projects and engineering teams while designing and implementing large-scale, distributed, cloud-native systems. Provides technical leadership across teams, shapes architecture and platform roadmaps, mentors engineers, drives reliability and performance, and partners with product leaders. Uses modern programming languages, cloud platforms, microservices, databases, containers, CI/CD, observability, and AI-based development tools to deliver secure solutions supporting regulatory and customer needs.
Top Skills: Ai Coding ToolsAutomated Testing FrameworksAWSAzureBitbucketC#Ci/CdDockerGCPGitGitGoIde CopilotsInfrastructure As CodeJavaJavaScriptKubernetesNode.jsNoSQLObservabilityOpen Source FrameworksPythonRdbmsRustScalaTypescript
Yesterday
In-Office
San Jose, CA, USA
110K-234K Annually
Senior level
110K-234K Annually
Senior level
Artificial Intelligence • Hardware • Information Technology • Machine Learning
Design, develop, deploy, and scale enterprise AI products, agents, automations, and intelligent workflows for the People Organization. Build multi-agent systems integrating Workday, ServiceNow, Microsoft Copilot, and other enterprise platforms through APIs and event-driven architectures. Apply HR domain expertise to employee lifecycle processes, ensure secure and responsible AI governance, and implement monitoring, testing, observability, and evaluation frameworks. Partner with business and technology leaders to prototype AI solutions and drive adoption and measurable workforce transformation.
Top Skills: Agent OrchestrationAi AgentsAi Reasoning SystemsAPIsAzureCi/CdDevOpsDockerEvent-Driven ArchitecturesKnowledge RetrievalLarge Language Models (Llms)Microsoft 365Microsoft CopilotMulti-Agent ArchitecturesPythonRetrieval-Augmented Generation (Rag)Semantic SearchServicenowTool CallingVector DatabasesWorkdayWorkflow Automation
Yesterday
Remote or Hybrid
7 Locations
185K-327K Annually
Senior level
185K-327K Annually
Senior level
eCommerce • Fintech • Hardware • Payments • Software • Financial Services
Design, build, scale, and operate high-volume payment APIs and distributed systems serving global traffic. Lead complex engineering projects from requirements through production, improve reliability, performance, observability, fault tolerance, and security, and participate in on-call and incident response. Collaborate across product and engineering teams, use AI-assisted development tools responsibly, and mentor engineers through technical reviews and documentation.
Top Skills: AWSGoKafkaKotlinTypescript

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account