Research Engineer, Post-Training

Posted 3 Days Ago
Be an Early Applicant
San Francisco, CA, USA
Hybrid
250K-450K Annually
Entry level
Artificial Intelligence • Information Technology • Software
The Role
Research Engineer responsible for developing and productionizing post-training methods for diffusion and flow models, including reward modeling, preference optimization, supervised fine-tuning, distillation, and reinforcement learning. The role owns research planning, evaluation standards, training pipelines, and the feedback loop between model behavior and product data. Models will be deployed to professional designers, requiring rigorous engineering, reproducible evaluations, and close collaboration with product and founders.
Summary Generated by Built In
Research Engineer, Post-Training

San Francisco, CA · In Person · Full-Time

Applying to this role will also allow us to consider you for other research opportunities at Vizcom. We believe the best roles are shaped around exceptional people, not just job descriptions.

About Vizcom

Vizcom is where design teams at companies like Nike, GM, New Balance, and Hasbro bring ideas from sketch to product. Designers use Vizcom to sketch, render, explore color and materials, work in 3D, and prepare concepts for production.

The render itself was never the point. The point is the physical thing that comes after it. We call this pencil to product.

Vizcom is a Series B company with more than $52M raised from investors including Radical Ventures, Index Ventures, and Nat Friedman.

Five years of professional designers working this way has created something difficult to reproduce in a traditional research environment: millions of moments where a trained designer, in the middle of real work, decided what should survive into a product that ultimately has to become real.

Those decisions create a uniquely interesting research problem. A designer's preference among several candidates can reflect the generator's style, where they are in the design process, what they are trying to make, and the professional judgment they bring to the decision. Existing approaches don't cleanly separate those signals.

Understanding that judgment — and learning how to model it — is the challenge this role will help solve.

The Role

As a Research Engineer, Post-Training, you'll build models that help us understand and learn from the judgment that carries a design from pencil to product.

You'll work closely with the engineer who built our current post-training stack and contribute to a growing body of research documenting the approaches we've tested, what we've learned, and where we've found meaningful signal.

This role sits directly between research and product. The models you train will ship to working designers, and what we learn through research will influence what the product captures next.

If your primary goal is research that ends with publication, this may not be the right environment. If you're excited by the idea of a reward model influencing what professional designers see in the product, it probably is.

We also believe the strongest results won't come from clever objectives alone. They'll come from excellent engineering: correct training code, rigorous evaluations, reliable pipelines, and experiments we can trust.

What You'll Own
  • Help define and execute the post-training roadmap for our models, from research plan through production.

  • Build reward and preference models using years of professional design decisions.

  • Explore and apply methods including supervised fine-tuning, distillation, preference optimization, and reinforcement learning.

  • Build rigorous evaluations and research practices that help us determine when a result is real and worth scaling.

  • Partner closely with Product to understand where today's signals fall short and shape what we capture next.

  • Evaluate emerging post-training techniques and determine which approaches are worth bringing into our stack.

  • Build reliable training and experimentation infrastructure that allows research results to translate into production systems.

This is a charter, not a week-one checklist. We don't expect one person to tackle everything at once. Part of the role is helping determine what matters most and in what sequence.

What Your First 90 Days Could Look Like

Days 1–30: Learn and map

Understand our data, post-training stack, existing research, evaluation methods, and the approaches we've already tested.

Days 30–60: Validate

Produce an initial result using historical data that holds up against our evaluation and reproducibility standards.

Days 60–90: Set direction

Help define the roadmap for moving from historical signals toward a closed feedback loop between our models, our product, and the designers using Vizcom.

What We're Looking For
  • Strong programming and software engineering skills, particularly for machine learning systems.

  • Hands-on experience training or post-training generative models, including diffusion or flow models.

  • Experience with one or more post-training methods such as supervised fine-tuning, preference optimization, reward modeling, distillation, or reinforcement learning.

  • Experience designing evaluations and experimentation pipelines you can trust.

  • An interest in product-coupled research, where research questions are informed by real users and models make their way into production.

  • Comfort working on ambiguous research problems where the right methodology may not yet exist.

Nice to Have
  • Experience building high-performance training or inference systems.

  • Experience optimizing ML workloads or working with large-scale training infrastructure.

  • Experience working with preference data or human-feedback systems.

  • An interest in industrial design, physical products, or the people who make them.

Above all, we're looking for someone who is more interested in understanding how professionals decide than optimizing for what the internet likes.

What You'll Get
  • A unique dataset: Five years of professional design decisions, with new signals generated every day.

  • Research that reaches users: The professionals whose judgment you're modeling are also the people using Vizcom. Successful research can reach their workflows quickly.

  • High ownership: You'll have meaningful influence over both our post-training research direction and the systems used to evaluate it.

  • Close product feedback loops: Research insights can directly shape what Vizcom builds and what data we capture next.

  • Direct access to the founders: You'll work closely with Vizcom's founders and technical leadership as we build out the research function.

Benefits at Vizcom
  • 100% employer-sponsored medical coverage for employees, plus 25% coverage toward dependents

  • Dental and vision coverage, plus mental health benefits

  • Meaningful equity ownership

  • Flexible PTO

  • 401(k) with employer match

  • Generous annual Learning & Development allowance

  • Paid parental leave

  • Weekly catered lunch at our San Francisco headquarters

  • Monthly gym membership stipend

Compensation

Base salary: $220,000–$300,000 USD + equity

We regularly benchmark compensation against relevant peer companies using current market data from industry-standard sources, including Carta and Pave. This range reflects our Tier 1 compensation market, which includes San Francisco.

The actual offer and overall compensation package will be determined based on multiple factors, including relevant experience, skills, qualifications, and business considerations. The compensation and benefits described in this posting apply to U.S.-based W-2 employees and may vary based on applicable employment laws and requirements.

How We Work
  • We document what we learned, not just what we worked on.

  • Negative results are valuable when they help us close off the wrong paths.

  • Results should be reproducible before they earn additional compute.

  • We share meaningful research through technical write-ups, demonstrations, and showcases where appropriate.

  • Our interview process emphasizes real-world problem solving and practical technical work rather than LeetCode-style interviews.

Location

This is an in-person role based in San Francisco, CA.

As part of Vizcom's SOC 2 Type II compliance program, employment is contingent upon successful completion of a background check, as permitted by applicable law.

Join Us

At Vizcom, we move quickly, give people meaningful ownership, and offer the opportunity to shape both our product and our company as we grow. We believe deeply in the craft of industrial design and in building tools that help designers bring better ideas into the physical world.

Join us in shaping a world designed by you.

Skills Required

  • Strong programming skills and the ability to write reliable machine learning code
  • Experience post-training diffusion or flow models using supervised fine-tuning, preference optimization, or reinforcement learning
  • High-performance implementations of training or inference code
  • Experience making physical products or deep interest in professionals who make physical products
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: San Francisco, California
56 Employees
Year Founded: 2021

What We Do

Building tools that shorten the distance between having ideas and bringing them to life. https://linktr.ee/vizcom_

Similar Jobs

Together AI Logo Together AI

Research Engineer, Post-Training Inference

Artificial Intelligence • Information Technology
In-Office
San Francisco, CA, USA
84 Employees
200K-290K Annually

Harvey Logo Harvey

Research Engineer, Post-Training

Artificial Intelligence • Legal Tech • Professional Services • Software
Hybrid
San Francisco, CA, USA
373 Employees
231K-340K Annually

Character.AI Logo Character.AI

Principal Research Engineer, Post-Training

Artificial Intelligence • Software • Conversational AI • Generative AI
In-Office
Redwood City, CA, USA
30 Employees
275K-400K Annually
Hybrid
San Francisco, CA, USA
350 Employees
200K-275K Annually

Similar Companies Hiring

Golden Pet Brands Thumbnail
Digital Media • eCommerce • Information Technology • Marketing Tech • Pet • Retail • Social Media
El Segundo, California
178 Employees
Kepler  Thumbnail
Fintech • Software
New York, New York
6 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account