Fireworks AI Logo

Fireworks AI

MTS, Research Engineer

Posted 19 Days Ago
Be an Early Applicant
In-Office
San Mateo, CA, USA
250K-400K Annually
Senior level
In-Office
San Mateo, CA, USA
250K-400K Annually
Senior level
Conduct open-ended ML research and reproduce/extend state-of-the-art models while building scalable distributed training infrastructure. Implement and optimize training loops, data pipelines, and communication across GPU clusters, translating research into robust, efficient code and collaborating with research scientists.
The summary above was generated by AI
About Us:

At Fireworks, we’re building the future of generative AI infrastructure. Our platform delivers the highest-quality models with the fastest and most scalable inference in the industry. We’ve been independently benchmarked as the leader in LLM inference speed and are driving cutting-edge innovation through projects like our own function calling and multimodal models. Fireworks is a Series C company valued at $4 billion and backed by top investors including Benchmark, Sequoia, Lightspeed, Index, and Evantic. We’re an ambitious, collaborative team of builders, founded by veterans of Meta PyTorch and Google Vertex AI.

About the Role

We are looking for a Research Engineer to join our team, operating at the critical intersection of model research and training infrastructure.

In this role, your time will be split between tackling open-ended research problems—such as designing novel architectures and improving algorithmic efficiency — and building the distributed training systems required to make those research breakthroughs a reality. You won't just be handed a paper to implement; you will be expected to reproduce state-of-the-art results from the literature, identify their limitations, and build the infrastructure needed to push beyond them.

The most significant advances in deep learning require massive scale. We need engineers who are as comfortable reasoning about gradient descent and loss landscapes as they are about distributed systems, GPU cluster utilization, and data pipelines.

 

What You'll Do

  • Conduct Open-Ended Research: Explore new model architectures, training objectives, and optimization techniques. Formulate hypotheses, design experiments, and iterate quickly based on empirical results.
  • Reproduce and Extend State-of-the-Art: Implement and reproduce results from recent machine learning papers. Identify bottlenecks, propose improvements, and scale these methods to larger datasets and models.
  • Build and Scale Training Infrastructure: Design, implement, and maintain high-performance, distributed machine learning systems. Optimize training loops, data loaders, and communication overhead across large GPU clusters.
  • Bridge Science and Engineering: Translate abstract mathematical concepts and research ideas into robust, bug-free, and efficient code.
  • Collaborate Cross-Functionally: Work closely with Research Scientists to unblock their experiments by providing tooling, optimizing code, and co-designing experiments that are hardware-aware.

We Expect You To Have:

  • Strong programming skills (Python, C++, or Rust) and a commitment to writing clean, maintainable code.
  • Deep practical knowledge of machine learning frameworks (PyTorch, JAX, or TensorFlow).
  • Experience working with large distributed systems and parallel computing (e.g., CUDA, NCCL, MPI).
  • A strong foundation in linear algebra, calculus, probability, and statistics.
  • A proven track record of implementing complex deep learning algorithms from scratch.

Nice to Have:

  • A Master’s or PhD in Computer Science, Machine Learning, Physics, Mathematics, or a related field (or equivalent industry experience).
  • Experience with low-level GPU programming (CUDA/Triton) or hardware co-design.
  • Familiarity with the challenges of training Large Language Models (LLMs)
  • Familiarity with the challenges of inference, and OSS inference engines such as SGLang and vLLM

Total compensation for this role also includes meaningful equity in a fast-growing startup, along with a competitive salary and comprehensive benefits package. Base salary is determined by a range of factors including individual qualifications, experience, skills, interview performance, market data, and work location. The listed salary range is intended as a guideline and may be adjusted.

Base Pay Range (Plus Equity)
$250,000$400,000 USD
Why Fireworks AI?
  • Solve Hard Problems: Tackle challenges at the forefront of AI infrastructure, from low-latency inference to scalable model serving.
  • Build What’s Next: Work with bleeding-edge technology that impacts how businesses and developers harness AI globally.
  • Ownership & Impact: Join a fast-growing, passionate team where your work directly shapes the future of AI—no bureaucracy, just results.
  • Learn from the Best: Collaborate with world-class engineers and AI researchers who thrive on curiosity and innovation.

Fireworks AI is an equal-opportunity employer. We celebrate diversity and are committed to creating an inclusive environment for all innovators.

HQ

Fireworks AI Redwood, California, USA Office

Redwood, CA, United States, 94063

Similar Jobs

2 Hours Ago
Hybrid
Senior level
Senior level
Financial Services
Lead CIAM architecture for customer authentication across cloud platforms: define target state, engage stakeholders, set CIAM standards and security controls, evaluate identity protocols and vendors, design and review secure code and integrations, leverage enterprise-authorized AI responsibly, automate remediation, and ensure operational stability and compliance for large-scale customer authentication systems.
Top Skills: AIAuth0Ci/CdCiamCloud-NativeForgerockIdentity ProofingiOSMachine LearningOauth 2.0OktaOpenid ConnectRisk-Based AuthenticationSAMLToken/Session Management
4 Hours Ago
In-Office
160K-230K Annually
Mid level
160K-230K Annually
Mid level
Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Real Estate
The Accounting Lead will manage core accounting operations, automate processes, enhance workflow efficiency, and support financial integrity as the company scales.
Top Skills: Accounting ToolsErp Systems
4 Hours Ago
In-Office
225K-325K Annually
Mid level
225K-325K Annually
Mid level
Artificial Intelligence • Legal Tech • Software
The Legal Engineer will act as a thought partner, building relationships with clients, facilitating product adoption, providing training, and guiding legal teams through technological transformation.
Top Skills: AIGenerative AiLegal Technology

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account