TetraMem Logo

TetraMem

Compiler Engineer

Reposted 26 Days Ago
Be an Early Applicant
In-Office
San Jose, CA, USA
160K-300K Annually
Senior level
In-Office
San Jose, CA, USA
160K-300K Annually
Senior level
Develop and optimize a compiler toolchain for translating deep learning models, collaborating with ML and hardware teams. Requires significant industry experience in compiler development.
The summary above was generated by AI

Responsibilities:

  • Design, develop, and maintain compiler toolchains that translate machine learning models from industry-standard frameworks into optimized workloads for TetraMem’s analog in-memory computing hardware.

  • Develop runtime systems, software libraries, and SDK components that enable efficient deployment, execution, and management of AI applications on TetraMem accelerators.

  • Implement compiler optimizations, including graph transformations, operator fusion, memory optimization, scheduling, and code generation to maximize performance and energy efficiency.

  • Research and develop innovative techniques to improve machine learning inference speed, latency, throughput, and power consumption across a wide range of AI workloads.

  • Collaborate closely with machine learning engineers to support model conversion, validation, optimization, benchmarking, and deployment.

  • Partner with hardware architects and silicon engineering teams to co-design software and hardware features that improve system performance, programmability, and usability.

  • Develop performance analysis, profiling, debugging, and benchmarking tools to evaluate and optimize AI workloads on current and future TetraMem platforms.

  • Integrate and support industry-standard machine learning frameworks and model formats, including PyTorch, TensorFlow, ONNX, and other emerging AI ecosystems.

  • Lead technical design reviews, contribute to software architecture decisions, and establish best practices for scalable, maintainable, and high-quality software development.

  • Mentor junior engineers, contribute to technical documentation, and help define the long-term roadmap for TetraMem’s compiler, runtime, and SDK technologies.

Requirements:

  • MS or PhD in Computer Engineering/CS/EE
  • 5+ years industry experience as a compiler engineer or developer
  • Experience developing compilers for GPU, dataflow compilers, or ML compilers
  • Startup mindset/experience

Experience in one or more of the following areas considered a strong plus:

  • Experience in RISC-V CPU/VPU kernel development and optimization
  • Experience providing technical leadership and/or guidance to other engineers
  • Knowledge of popular CPU/GPU compilers such as GCC, Clang
  • Knowledge of ML compilers such as MLIR
  • Experience with LLVM and other open-source compiler libraries and tools
  • Publications on compilation of ML or dataflow programs for HW acceleration

Salary Range: $160,000 - $300,000 / year

TetraMem celebrates diversity and is committed to creating an inclusive environment for all employees. We are proud to be an Equal Opportunity Employer and welcome applicants from all backgrounds. Qualified candidates will receive consideration for employment without regard to race, color, religion, creed, sex, gender identity or expression, sexual orientation, national origin, ancestry, age, marital status, medical condition, disability, genetic information, military or veteran status, or any other characteristic protected by applicable federal, state, or local law.

TetraMem is committed to providing reasonable accommodations to qualified applicants with disabilities throughout the recruitment process. Applicants requiring accommodation may contact Human Resources for assistance.

To ensure a fair, consistent, and efficient hiring process, all candidates must apply through TetraMem’s official ClearCompany Applicant Tracking System (ATS). Applications submitted through the ATS allow our hiring team to evaluate candidates using a standardized process and ensure timely communication throughout the recruitment process. To promote equal consideration for all applicants, applications submitted outside of the ClearCompany ATS, including direct emails, LinkedIn messages, or unsolicited submissions to employees, may not be reviewed or considered.

We encourage all interested candidates to apply through the official TetraMem Careers page.

HQ

TetraMem Newark, California, USA Office

Newark, CA, United States

Similar Jobs

Yesterday
In-Office or Remote
4 Locations
152K-242K Annually
Senior level
152K-242K Annually
Senior level
Artificial Intelligence • Computer Vision • Hardware • Robotics • Metaverse
Design and build intelligent compiler systems, compiler-oriented agents, code transformation workflows, and numerical verification infrastructure. Develop systems for low-level IR reasoning, code generation, optimization, differential testing, regression detection, and correctness validation across GPU-centric workloads. Collaborate with compiler, CUDA, runtime, library, and hardware teams to integrate capabilities into NVIDIA products while optimizing performance, scalability, and numerical correctness.
Top Skills: CC++CudaGpu ProgrammingLlvmMlirPtxPythonRustTriton
2 Days Ago
In-Office or Remote
2 Locations
108K-196K Annually
Entry level
108K-196K Annually
Entry level
Artificial Intelligence • Computer Vision • Hardware • Robotics • Metaverse
Develop and optimize machine learning compilers, AI systems, and compiler abstractions for NVIDIA hardware. Build agent-driven automation for machine learning system development, co-design compiler and agent solutions, optimize AI workloads, and collaborate with software and hardware engineering teams. The role focuses on ML compiler technologies, LLM inference and training, GPU acceleration, and productizing machine learning systems.
Top Skills: Apache TvmCC++FlashattentionFlashinferGpu ProgrammingMlirPython
3 Days Ago
In-Office or Remote
4 Locations
152K-242K Annually
Senior level
152K-242K Annually
Senior level
Artificial Intelligence • Computer Vision • Hardware • Robotics • Metaverse
Develop compiler optimization algorithms for deep learning workloads, optimizing JAX and OpenXLA performance on NVIDIA GPUs. Design graph partitioning, tensor sharding, code generation, and performance-tuning solutions for distributed training and inference. Collaborate with deep learning framework and GPU hardware teams, implement user-facing JAX features, and contribute to production-grade software using MLIR, LLVM, and OpenAI Triton.
Top Skills: CC++CudaJaxLlvmMlirNvidia GpusOpenai TritonOpenclOpenxlaPyTorchTensorFlowTvmXla

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account