Harmonic (harmonic.fun) Logo

Harmonic (harmonic.fun)

Research Engineer, Training & Inference

Reposted 11 Days Ago
Be an Early Applicant
In-Office
Palo Alto, CA, USA
200K-450K Annually
Mid level
In-Office
Palo Alto, CA, USA
200K-450K Annually
Mid level
This role involves optimizing reinforcement learning systems for high performance by managing the entire stack, improving training throughput, and enhancing inference engine efficiency.
The summary above was generated by AI
About Harmonic

At Harmonic, we are building a mathematical reasoning engine that operates with absolute precision. While most AI makes maximum-likelihood guesses, Harmonic's Aristotle uses Lean 4 and reinforcement learning to verify its reasoning and results.
Following our Gold Medal-level performance on the 2025 International Math Olympiad (IMO) and the successful resolution of long-standing open problems, we are proving that AI can master the most rigorous domains of human thought. Backed by some of the world’s most prominent investors, we are intentionally scaling an elite technical team.
Visit our company blog to learn more about what we are working on!

About the Role

We are developing reinforcement learning systems at a scale where standard abstractions frequently fail. Unlike labs that operate primarily through high-level wrappers, we own the entirety of our RL stack. This ownership spans from low-level environment simulators and custom communication primitives to our distributed training loops and inference engines.

We are seeking engineers who view existing libraries as a baseline and the hardware speed itself as the true target. You will be responsible for the architecture powering our agents, with a relentless focus on maximizing the throughput of our reinforcement learning and production workflows.

Key Responsibilities
  • Total Stack Ownership: Maintain and optimize our proprietary RL training and serving infrastructure. You have the authority to refactor any layer—from the Python API down to the CUDA kernels—to achieve peak performance for foundation model workloads.

  • Optimized Training: maximize the throughput of our reinforcement learning system from data generation to model training with sharded multi-node training and inference algorithms.

  • High-Performance Serving: optimize our inference stack for high-throughput reinforcement learning and low-latency LLM production traffic. Tune the inference engine, router, and scheduler, down to custom kernels if need be.

  • Compute Optimization: Identify and resolve performance bottlenecks within our distributed clusters, ensuring optimal throughput and memory efficiency for multi-billion parameter models, balancing memory constraints with compute-heavy training cycles.

Minimum Qualifications
  • BS in Computer Science or a related technical field, or equivalent industry experience

  • 2+ years of relevant, hands-on industry experience

  • Proficiency in Python

  • Experience building or maintaining components within ML frameworks (e.g., PyTorch, JAX, or TensorFlow).

  • Proficiency in either:

    • Understanding of distributed training concepts and collective communication primitives (e.g., NCCL).

      OR

    • Practical experience deploying and profiling models on GPU-accelerated cloud infrastructure.

Preferred Qualifications
  • MS or PhD in Computer Science, Mathematics, or a related field.

  • 5+ years of relevant, hands-on industry experience

  • Proficiency in C++

  • Experience writing or improving kernels (Triton, CuTeDSL, TileLang, CUDA, CUTLASS, ThunderKittens) to resolve low-level bottlenecks.

  • Proven success deploying performant inference at scale using open-source or custom inference engines, routers, etc.

  • Direct experience scaling models via FSDP, Tensor Parallelism, or related sharding techniques on multi-node GPU clusters.

  • Experience designing reinforcement learning systems for high-throughput training and asynchronous data sampling.

What We Offer
  • Unlimited PTO

  • 401(k) matching

  • 100% employer-paid health, vision, and dental benefits for employees and 50% coverage for dependents. Harmonic offers varied health coverage options to select what is best for you and your family.

  • Health Savings Account (HSA) available for qualifying health plans

Equal Opportunity Statement

Harmonic is committed to diversity and inclusivity in the workplace. We are an equal opportunity employer and do not discriminate on the basis of race, religion, national origin, gender, sexual orientation, age, veteran status, disability or any other legally protected status.

HQ

Harmonic (harmonic.fun) San Jose, California, USA Office

San Jose, CA, United States

Similar Jobs

6 Days Ago
In-Office
San Francisco, CA, USA
200K-290K Annually
Junior
200K-290K Annually
Junior
Artificial Intelligence • Information Technology
Develop and maintain a platform for customizing open-source models, integrating Model Shaping with Inference, adding inference engine features and RL optimizations, ensuring production stability and 24/7 availability, and collaborating with product, research, and engineering teams to support fine-tuning, RL, and evaluation workflows.
Top Skills: CudaCuteFp4Fp8GoKubernetesLoraPythonReinforcement LearningSglangTensorrt-LlmTritonVllm
10 Minutes Ago
Hybrid
Redwood City, CA, USA
90K-150K Annually
Mid level
90K-150K Annually
Mid level
Artificial Intelligence • Big Data • Healthtech • Machine Learning • Analytics • Biotech • Generative AI
Partner with pharma clients to design and execute computational translational research using large clinical and molecular datasets. Drive account strategies, co-architect client solutions on the Tempus platform, communicate technical results to diverse audiences, author whitepapers, and coordinate with Product, Engineering, and Research teams. Travel ~25% for client engagements.
Top Skills: AWSCSS3D3DaskDockerFlaskGgplotGitHTML5JavaScriptJupyter NotebooksMatplotlibNumpyPandasPlot.LyPythonRRstudioScikit-LearnScipySeabornSQLTidyverse
An Hour Ago
In-Office
2 Locations
182K-242K Annually
Mid level
182K-242K Annually
Mid level
Cloud • Information Technology • Machine Learning
Embed with customer AI teams to design, prototype, and productionize AI agents using W&B Weave. Advise on architecture, build reference implementations and demos, run technical workshops, gather field feedback for product roadmap, and troubleshoot customer environments.
Top Skills: Hugging FaceLangchainLlamaindexLlmsPythonPyTorchTensorFlowVector DatabasesW&B Weave

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account