Tensordyne Logo

Tensordyne

Director, AI Systems Solutions Engineering

Posted 10 Days Ago
Be an Early Applicant
In-Office
Sunnyvale, CA, USA
Expert/Leader
In-Office
Sunnyvale, CA, USA
Expert/Leader
Leads strategic technical engagements for an AI inference platform from architecture evaluation through benchmarking, model enablement, deployment, and production rollout. Builds and manages a high-performing solutions engineering team, guides customer evaluations and deployments, translates customer requirements into product priorities, and develops scalable technical go-to-market assets. Requires deep expertise in AI inference, accelerator architecture, distributed systems, serving frameworks, performance optimization, and customer-facing technical leadership.
The summary above was generated by AI

About Tensordyne

Tensordyne is building a new class of AI inference system designed for high-performance, power-efficient deployment of the world’s most demanding generative AI workloads.

Our platform combines purpose-built silicon, new AI math, optimized scale-up networking, and memory architecture into a tightly integrated system purpose built for large-scale AI inference. We work with hyperscalers, neoclouds, frontier model developers, enterprises, and infrastructure partners operating at the leading edge of AI.

As Tensordyne moves from system development into silicon bring-up, customer validation, beta deployments, and production rollout, we are building the technical customer organization that will sit directly between our engineering teams and the companies deploying the platform.

We are looking for an exceptional technical leader to help build and lead that function.

The Role

Tensordyne is hiring a Director of AI Systems Solutions Engineering to own and grow our most important technical customer engagements.

This is a senior, highly technical role for someone who understands modern AI infrastructure from model architecture through accelerator hardware, distributed inference, serving software, and datacenter deployment — and who can credibly engage with the engineers and architects building the next generation of AI platforms.

You will work directly with frontier model builders, hyperscalers, neoclouds, developers, infrastructure partners, and strategic customers as they evaluate and deploy Tensordyne systems.

You will also build and lead a small team of exceptional Sales and Solutions Engineers responsible for customer benchmarking, technical evaluation, NPI, model enablement, AI DC architecture, and production deployment.

This is not a traditional pre-sales engineering role. The team will operate at the frontier of a rapidly changing technology landscape, working with constantly evolving new models and requirements. The right person will be equally comfortable in a customer architecture review, helping prioritize product capabilities and roadmaps, and leading a technical evaluation with hyperscalers and frontier AI companies.

What You Will Own

  • Strategic technical customer engagements: Own the technical relationship with key customers and partners from initial architecture discussions through benchmarking, evaluation, integration, deployment, and expansion.
  • Technical evaluation strategy: Define how Tensordyne demonstrates system performance across KPI's like throughput, tokens/sec/user, ttft, memory utilization, power efficiency, system density, model accuracy/quality, and other relevant inference metrics.
  • AI workload and model architecture engagement: Work with customers and model developers to understand current and emerging HW and model architectures, serving requirements, context lengths, parallelism strategies, model topology, quantization approaches, and inference optimization requirements.
  • Benchmarking and competitive analysis: Maintain a technically rigorous understanding of Tensordyne performance relative to leading GPU and AI accelerator platforms. Ensure customer-facing comparisons are credible, reproducible, current, and aligned with real deployment requirements.
  • Model enablement and optimization: Partner with compiler, runtime, kernel, systems, and SDK teams to bring important customer models and workloads onto the Tensordyne platform and identify opportunities for performance improvement.
  • Forward deployment: Lead technical PoCs, remote evaluations, on-premises beta deployments, integration programs, and production readiness efforts with strategic customers.
  • Customer-to-product feedback loop: Translate recurring customer requirements into clear priorities for the SDK, compiler, runtime, inference server, model support, orchestration, networking, observability, and system architecture teams.
  • Technical market intelligence: Stay deeply current on model architectures, inference techniques, accelerator roadmaps, serving frameworks, competitive systems, benchmarking methodologies, and changes in the AI infrastructure market.
  • Team leadership: Help recruit, develop, and lead a small team of elite Sales/Solutions Engineers capable of independently managing sophisticated technical engagements with the world’s most demanding AI infrastructure customers.
  • Scalable technical GTM: Turn early customer engagements into repeatable benchmarks, evaluation frameworks, reference architectures, deployment playbooks, documentation, demos, and technical collateral that can support a rapidly growing customer base.

What We Are Looking For

We are looking for a proven AI systems leader with substantial technical depth and strong judgment.

You should bring:

  • Deep understanding of modern AI inference systems, including LLM and multimodal architectures.
  • Strong knowledge of AI accelerator and system architecture, including compute, memory hierarchy, interconnect, parallelism, and distributed inference.
  • Experience reasoning about inference performance across latency, throughput, memory bandwidth, utilization, batching, context length, prefill, decode, and system scaling.
  • Hands-on familiarity with modern AI frameworks and serving environments such as PyTorch, vLLM, SGLang, Triton, or comparable systems.
  • Experience working across the boundary between AI software and accelerator hardware, ideally including GPUs, custom silicon, or emerging AI accelerators.
  • Experience benchmarking and optimizing workloads on large-scale AI infrastructure.
  • Strong understanding of production inference techniques including quantization, tensor/model/expert parallelism, disaggregated serving, KV-cache management, distributed execution, and related optimization strategies.
  • Demonstrated ability to engage technically sophisticated external organizations including Sr technical leaders at hyperscalers, cloud providers, model developers, AI infrastructure companies, or large enterprise engineering teams.
  • Experience leading high-performing Solutions Engineering, Field Engineering, Forward Deployed Engineering, or comparable technical customer teams.
  • Ability to operate effectively in a fast-moving environment where the product, software stack, competitive landscape, and customer requirements are evolving simultaneously.

Particularly Relevant Experience

Candidates may come from organizations building or deploying:

  • GPU or custom AI accelerator platforms
  • Large-scale AI inference infrastructure
  • Frontier or foundation models
  • Hyperscale cloud infrastructure
  • AI serving and orchestration platforms
  • Compiler, runtime, or distributed AI systems
  • High-performance computing or distributed systems

Experience bringing a new accelerator architecture or AI infrastructure platform from early access through customer validation and production deployment is especially valuable.

Why This Role Matters

Tensordyne is entering the stage where our technology moves from internal development into the hands of customers.

The technical customer organization will play a central role in that transition.

This team will help determine which workloads we prioritize, how customers evaluate our platform, how quickly new models become production-ready, how effectively customer feedback reaches engineering, and ultimately how Tensordyne systems are adopted at scale.

We are looking for someone who wants to build that capability from the beginning — and establish the technical standard for how Tensordyne engages with the companies defining the future of AI infrastructure.

Tensordyne's culture was built on the following values

  • Put people first. We only succeed when our people succeed.
  • Ethics and integrity always; Being open, honest, and respectful of everyone.
  • Think Big. Be ambitious and have audacious goals.
  • Aim for excellence. Quality and excellence count in everything we do.
  • Own it and get it done. Results matter!
  • Make each person better together, than they would be as an individual.
  • Embrace each others’ differences, and embrace that there will be differences.

Tensordyne is an equal opportunity employer. We believe that a diverse team is better at tackling complex problems and coming up with innovative solutions. All qualified applicants will receive consideration for employment without regard to age, color, gender identity or expression, marital status, national origin, disability, protected veteran status, race, religion, pregnancy, sexual orientation, or any other characteristic protected by applicable laws, regulations and ordinances.

 

A note to Recruitment Agencies: Please don’t reach out to Tensordyne employees or leaders about our roles -- we’ve got it covered. We don’t accept unsolicited agency resumes and we are not responsible for any fees related to unsolicited resumes. Thank you for your understanding.

HQ

Tensordyne San Jose, California, USA Office

San Jose, CA, United States

Similar Jobs

37 Seconds Ago
In-Office
79K-110K Annually
Entry level
79K-110K Annually
Entry level
Aerospace • Artificial Intelligence • Machine Learning • Robotics • Software • Defense Technology
Develop Python-based automated testing, simulation analysis, log-processing, scenario-generation, and domain-randomization tools for the Hivemind Platform. Support test orchestration, distributed simulation, containerized environments, CI/CD, build troubleshooting, and technical documentation. Investigate failures using logs, telemetry, rosbags, network captures, and test results while collaborating with autonomy, systems, and platform engineers.
Top Skills: BashBuildkiteC++CsvDockerGitGithub ActionsGitlab CiGoogletestHelmJenkinsJSONKubernetesLinuxMatplotlibNumpyPandasPlotlyPodmanPytestPythonRayRest ApisRobot FrameworkRosRos 2RosbagScipyTcpdumpTerraformUnittestWiresharkYaml
37 Seconds Ago
In-Office
92K-140K Annually
Junior
92K-140K Annually
Junior
Aerospace • Artificial Intelligence • Machine Learning • Robotics • Software • Defense Technology
Supports integration, verification, and validation testing for autonomous unmanned aircraft systems. Assists with test planning, procedures, simulation and hardware-in-the-loop environments, hardware and software configuration, data collection, troubleshooting, defect investigation, reporting, and technical documentation. The role collaborates with systems, software, autonomy, hardware, and program teams and periodically supports flight-test sites, customer activities, and integration events.
Top Skills: BashC++Command-Line ToolsContinuous IntegrationEthernetGitHardware-In-The-Loop (Hil)Ip NetworkingLinuxPythonSimulationSoftware-In-The-Loop (Sil)TcpTelemetry AnalysisTest AutomationUdpWindows
39 Seconds Ago
In-Office
San Mateo, CA, USA
130K-240K Annually
Senior level
130K-240K Annually
Senior level
Aerospace • Artificial Intelligence • Machine Learning • Robotics • Software • Defense Technology
Build and support full-stack applications for developing, training, simulating, and evaluating autonomous systems. Develop React and TypeScript interfaces, backend services, APIs, data models, and Python workflows. Deploy and troubleshoot containerized services across local, cloud, on-premises, and air-gapped Kubernetes environments. Collaborate with users, product, design, and engineering teams to translate complex workflows into reliable, extensible applications and reusable platform capabilities.
Top Skills: Argo WorkflowsCi/CdDagsterDockerGoGrpcHelmJson SchemaKubernetesOpenapiProtocol BuffersPythonRayReactTypescriptWebsockets

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account