Vizcom Logo

Vizcom

Research Engineer, Post-Training

Posted 8 Days Ago
Be an Early Applicant
Hybrid
San Francisco, CA, USA
250K-450K Annually
Entry level
Hybrid
San Francisco, CA, USA
250K-450K Annually
Entry level
Research Engineer responsible for developing and productionizing post-training methods for diffusion and flow models, including reward modeling, preference optimization, supervised fine-tuning, distillation, and reinforcement learning. The role owns research planning, evaluation standards, training pipelines, and the feedback loop between model behavior and product data. Models will be deployed to professional designers, requiring rigorous engineering, reproducible evaluations, and close collaboration with product and founders.
The summary above was generated by AI
Research Engineer, Post-Training

San Francisco, CA · In Person · Full-Time

Applying to this role will also allow us to consider you for other research opportunities at Vizcom. We believe the best roles are shaped around exceptional people, not just job descriptions.

About Vizcom

Vizcom is where design teams at companies like Nike, GM, New Balance, and Hasbro bring ideas from sketch to product. Designers use Vizcom to sketch, render, explore color and materials, work in 3D, and prepare concepts for production.

The render itself was never the point. The point is the physical thing that comes after it. We call this pencil to product.

Vizcom is a Series B company with more than $52M raised from investors including Radical Ventures, Index Ventures, and Nat Friedman.

Five years of professional designers working this way has created something difficult to reproduce in a traditional research environment: millions of moments where a trained designer, in the middle of real work, decided what should survive into a product that ultimately has to become real.

Those decisions create a uniquely interesting research problem. A designer's preference among several candidates can reflect the generator's style, where they are in the design process, what they are trying to make, and the professional judgment they bring to the decision. Existing approaches don't cleanly separate those signals.

Understanding that judgment — and learning how to model it — is the challenge this role will help solve.

The Role

As a Research Engineer, Post-Training, you'll build models that help us understand and learn from the judgment that carries a design from pencil to product.

You'll work closely with the engineer who built our current post-training stack and contribute to a growing body of research documenting the approaches we've tested, what we've learned, and where we've found meaningful signal.

This role sits directly between research and product. The models you train will ship to working designers, and what we learn through research will influence what the product captures next.

If your primary goal is research that ends with publication, this may not be the right environment. If you're excited by the idea of a reward model influencing what professional designers see in the product, it probably is.

We also believe the strongest results won't come from clever objectives alone. They'll come from excellent engineering: correct training code, rigorous evaluations, reliable pipelines, and experiments we can trust.

What You'll Own
  • Help define and execute the post-training roadmap for our models, from research plan through production.

  • Build reward and preference models using years of professional design decisions.

  • Explore and apply methods including supervised fine-tuning, distillation, preference optimization, and reinforcement learning.

  • Build rigorous evaluations and research practices that help us determine when a result is real and worth scaling.

  • Partner closely with Product to understand where today's signals fall short and shape what we capture next.

  • Evaluate emerging post-training techniques and determine which approaches are worth bringing into our stack.

  • Build reliable training and experimentation infrastructure that allows research results to translate into production systems.

This is a charter, not a week-one checklist. We don't expect one person to tackle everything at once. Part of the role is helping determine what matters most and in what sequence.

What Your First 90 Days Could Look Like

Days 1–30: Learn and map

Understand our data, post-training stack, existing research, evaluation methods, and the approaches we've already tested.

Days 30–60: Validate

Produce an initial result using historical data that holds up against our evaluation and reproducibility standards.

Days 60–90: Set direction

Help define the roadmap for moving from historical signals toward a closed feedback loop between our models, our product, and the designers using Vizcom.

What We're Looking For
  • Strong programming and software engineering skills, particularly for machine learning systems.

  • Hands-on experience training or post-training generative models, including diffusion or flow models.

  • Experience with one or more post-training methods such as supervised fine-tuning, preference optimization, reward modeling, distillation, or reinforcement learning.

  • Experience designing evaluations and experimentation pipelines you can trust.

  • An interest in product-coupled research, where research questions are informed by real users and models make their way into production.

  • Comfort working on ambiguous research problems where the right methodology may not yet exist.

Nice to Have
  • Experience building high-performance training or inference systems.

  • Experience optimizing ML workloads or working with large-scale training infrastructure.

  • Experience working with preference data or human-feedback systems.

  • An interest in industrial design, physical products, or the people who make them.

Above all, we're looking for someone who is more interested in understanding how professionals decide than optimizing for what the internet likes.

What You'll Get
  • A unique dataset: Five years of professional design decisions, with new signals generated every day.

  • Research that reaches users: The professionals whose judgment you're modeling are also the people using Vizcom. Successful research can reach their workflows quickly.

  • High ownership: You'll have meaningful influence over both our post-training research direction and the systems used to evaluate it.

  • Close product feedback loops: Research insights can directly shape what Vizcom builds and what data we capture next.

  • Direct access to the founders: You'll work closely with Vizcom's founders and technical leadership as we build out the research function.

Benefits at Vizcom
  • 100% employer-sponsored medical coverage for employees, plus 25% coverage toward dependents

  • Dental and vision coverage, plus mental health benefits

  • Meaningful equity ownership

  • Flexible PTO

  • 401(k) with employer match

  • Generous annual Learning & Development allowance

  • Paid parental leave

  • Weekly catered lunch at our San Francisco headquarters

  • Monthly gym membership stipend

Compensation

Base salary: $220,000–$300,000 USD + equity

We regularly benchmark compensation against relevant peer companies using current market data from industry-standard sources, including Carta and Pave. This range reflects our Tier 1 compensation market, which includes San Francisco.

The actual offer and overall compensation package will be determined based on multiple factors, including relevant experience, skills, qualifications, and business considerations. The compensation and benefits described in this posting apply to U.S.-based W-2 employees and may vary based on applicable employment laws and requirements.

How We Work
  • We document what we learned, not just what we worked on.

  • Negative results are valuable when they help us close off the wrong paths.

  • Results should be reproducible before they earn additional compute.

  • We share meaningful research through technical write-ups, demonstrations, and showcases where appropriate.

  • Our interview process emphasizes real-world problem solving and practical technical work rather than LeetCode-style interviews.

Location

This is an in-person role based in San Francisco, CA.

As part of Vizcom's SOC 2 Type II compliance program, employment is contingent upon successful completion of a background check, as permitted by applicable law.

Join Us

At Vizcom, we move quickly, give people meaningful ownership, and offer the opportunity to shape both our product and our company as we grow. We believe deeply in the craft of industrial design and in building tools that help designers bring better ideas into the physical world.

Join us in shaping a world designed by you.

HQ

Vizcom San Francisco, California, USA Office

San Francisco, California, United States, 94103

Similar Jobs

25 Days Ago
In-Office
San Francisco, CA, USA
200K-290K Annually
Junior
200K-290K Annually
Junior
Artificial Intelligence • Information Technology
Develop and maintain a platform for customizing open-source models, integrating Model Shaping with Inference, adding inference engine features and RL optimizations, ensuring production stability and 24/7 availability, and collaborating with product, research, and engineering teams to support fine-tuning, RL, and evaluation workflows.
Top Skills: CudaCuteFp4Fp8GoKubernetesLoraPythonReinforcement LearningSglangTensorrt-LlmTritonVllm
6 Days Ago
Hybrid
San Francisco, CA, USA
231K-340K Annually
Mid level
231K-340K Annually
Mid level
Artificial Intelligence • Legal Tech • Professional Services • Software
Drive post-training experiments to improve agent performance for legal tasks: optimize harnesses, design grading/reward systems, study agent behavior, and collaborate with internal and external researchers to convert findings into training data, evals, and model improvements.
Top Skills: DistillationDistributed TrainingGpu WorkloadsLlmsPreference OptimizationPythonReward ModelingRlaifRlhfSft
14 Days Ago
In-Office
Redwood City, CA, USA
275K-400K Annually
Expert/Leader
275K-400K Annually
Expert/Leader
Artificial Intelligence • Software • Conversational AI • Generative AI
Lead technical vision and execution for post-training systems that adapt OSS LLMs into production conversational products. Drive research in alignment, RL and fine-tuning, architect scalable training/inference infrastructure, build data pipelines and evaluation frameworks, and mentor teams to improve model behavior, safety, and user engagement at scale.
Top Skills: A/B Testing FrameworksCloud-Native Ml InfrastructureData PipelinesDistributed TrainingDockerGpu-Based SystemsKubernetesLarge Language Models (Llms)MistralModel ObservabilityModel ServingOrchestration PlatformsPreference OptimizationQwenReinforcement LearningSupervised Fine-TuningTransformers

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account