Weave is our new lightweight toolkit for tracking and evaluating GenAI applications. Our users are relying on Weave to understand and improve their AI-powered applications. They’re on the cutting edge, building Agents, Retrieval Augmented Generation (RAG) systems, and more.
We’re looking for a software engineer to help build Weave. You’ll play a foundational role on this small but growing team, developing a new generation of tools for the most consequential emerging market today. You’ll ship core functionality, gather feedback directly from early users, and use your own intuition to continually improve the product.
If you have a strong interest in GenAI, love building and shipping quickly, and enjoy talking to users, this is the perfect role for you.
What you’ll achieve (Responsibilities)Implement robust and ergonomic open-source client APIs.
Design and implement integrations with popular LLM tools such as LlamaIndex, LangSmith, and Instructor.
Scale and improve our backend performance to support low latency and high volume production use cases.
Build polished and delightful features such as multi-modal evaluations, LLM playgrounds, and dataset editing/drill-down capabilities.
Actively meet with and iterate based on customer feedback, understanding their workflows and addressing key bugs.
Bachelor’s degree in Computer Science, or related field.
4 years or more of prior experience in Software Engineering.
Deep knowledge of Python and/or TypeScript.
Experience (personal or professional) building GenAI-powered applications.
Ability to operate effectively amid ambiguity and adapt rapidly to a changing business. environment. Previous startup experience a plus.
Bonus points: Experience managing large-scale online production services or OLAP databases
This role will be based in our San Francisco office a minimum of 2-3 days per week.
Skills Required
- Bachelor's degree in Computer Science or a related field
- At least 4 years of software engineering experience
- Deep knowledge of Python and/or TypeScript
- Experience building GenAI-powered applications
- Ability to work effectively amid ambiguity and adapt rapidly to a changing business environment
- Experience managing large-scale online production services or OLAP databases
- Ability to work from the San Francisco office at least 2–3 days per week
What We Do
Weights & Biases helps machine learning teams build better models faster. With a few lines of code, practitioners can instantly debug, compare and reproduce their models — architecture, hyperparameters, git commits, model weights, GPU usage, and even datasets and predictions — and collaborate with their teammates.








