Wafer runs dedicated and serverless inference for open-weights models. We use AI to optimize AI infrastructure: we tune the whole serving stack so the same model runs faster on cheaper hardware, and our customers keep the difference. We are seven people in San Francisco and plan to double that as fast as we can hire.
About the role
We are hiring our first dedicated go-to-market hire to own the top of the funnel.
Your job is to create qualified first meetings with AI-native companies running production inference.
For the right person this is a fast way into technical sales. The cycle is short and you will be in conversations with CTOs in your first month. We would rather promote you than hire over you.
This role is in person in San Francisco, five days a week.
In this role, you will:
Own outbound volume. Roughly 80+ touches a day across email, LinkedIn, cold calls, and whatever else gets a reply.
Book qualified first meetings that convert into benchmarks. A meeting that was never going to benchmark is worse than no meeting.
Build and test multi-step sequences, and report reply and meeting rates every week with real numbers.
Work with the closer and forward deployed engineer before every first meeting so nobody walks in cold.
Keep the CRM up to date.
You might thrive in this role if you:
Have at 12+ months of B2B outbound ai a high-growth AI-native company, active now or within the last six months, and can recite your own numbers from memory: touches per day, reply rate, meetings per week, and what you changed to move them.
Have done this at an early-stage startup before, seed or Series A, where the playbook did not exist yet and nobody handed you a list.
Are unreasonably persistent, and treat a 2 to 5 percent reply rate as a problem to solve rather than a reason to stop.
Write like a person. Short, specific, no marketing voice.
Are technically curious enough to learn the inference stack and hold your own with a CTO.
Track your own conversion rates before anyone asks for them.
Want the job after this one. This should turn into full cycle sales.
How we hire
We score every candidate on seven values: Infinitely Resourceful, Exceptionalism, Unreasonable Standards, Company Over Self, High EQ, Learns Quickly, and First Principles Thinker.
The process is short. An intro call, a working session where you write real outbound against real accounts, and an onsite. We aim to go from first conversation to offer in under two weeks.
About Wafer
Our mission is to maximize intelligence per watt by using AI to optimize AI infrastructure, achieving orders of magnitude better energy and cost efficiency per token. We believe cheap intelligence is the most essential piece of technology for a future of abundance, and we care about building a world where intelligence is "too cheap to meter." We commercialize that work by serving optimized LLMs as an inference cloud, and we serve the highest performance inference to fast-growing AI startups.
Compensation and benefits
$85K base and $140K OTE, plus a equity. Variable is uncapped.
Medical, dental, and vision insurance for employees and dependents, with 100% of premiums covered.
Unlimited paid time off
Daily lunch and dinner.
$1K/month housing stipend if you live within walking distance (0.5 miles) of the office.
Visa sponsorship available.
Equal opportunity
Wafer is committed to maintaining a workplace free from discrimination and harassment.
We make employment decisions based on business needs, job requirements, and individual qualifications, without regard to race, color, religion, belief, national origin, social or ethnic origin, age, physical, mental, or sensory disability, sexual orientation, gender identity or expression, marital status, civil union or domestic partnership status, past or present military service, HIV status, family medical history or genetic information, family or parental status including pregnancy, or any other status protected by law.
We welcome the opportunity to consider qualified applicants with prior arrest or conviction records. Our commitment to diversity includes hiring talented individuals regardless of their criminal history, in accordance with local, state, and federal laws, including San Francisco's Fair Chance Ordinance and California's ban-the-box laws.
Skills Required
- 12+ months of B2B outbound experience at a high-growth AI-native company (active now or within last six months)
- Able to recite personal metrics from memory: touches per day, reply rate, meetings per week, and what changed to move them
- Experience booking qualified meetings that convert into technical benchmarks or demos
- Willingness to work in person in San Francisco five days a week
- Track and report reply and meeting rates weekly with real numbers
- Experience at an early-stage startup (seed or Series A) where outbound playbooks were built from scratch
- Strong written communication: short, specific, non-marketing voice
- Technical curiosity to learn the inference stack and engage with CTOs
- High persistence and ability to optimize low reply rates (2-5%)
What We Do
Wafer is an AI infrastructure company that makes model inference faster and more cost-efficient. Its autonomous performance-engineering agents optimize GPU kernels and the broader serving stack—including batching, scheduling, and memory layout—for open-source large language models. Wafer offers serverless and dedicated inference endpoints, routing workloads across NVIDIA, AMD, TPUs, and other silicon to provide low-latency, high-throughput APIs for enterprise and AI-native applications.









