Member of Technical Staff
AI Summary
Member of Technical Staff building AI infrastructure optimization for serverless and dedicated inference, shipping day-zero model support and tuning serving stack performance and reliability.
About this role
Member of Technical Staff
Our mission at Wafer is to maximize intelligence per watt by using AI to optimize AI infrastructure, achieving orders of magnitude better energy and cost efficiency per token.
We believe cheap intelligence is the most essential piece of technology for a future of abundance. We care about building a future where intelligence is "too cheap to meter."
Wafer commercializes these efforts by serving serverless and dedicated inference for open source LLMs at the best performance per dollar. Our core bet is doing this through autonomous optimization of heterogeneous hardware.
What you'll do
Ship day-zero support for new open-source models, tuned for latency and throughput
Optimize the serving stack: batching, KV cache, speculative decoding, quantization
Write and tune kernels in CUDA, HIP, and Triton for NVIDIA, AMD, TPU, Trainium, D-Matrix, and more.
Design, deploy, and operate heterogeneous clusters across vendors
Run production inference across a mixed fleet: reliability, observability, and cost per token at scale
How we evaluate
We score every candidate on seven values:
Infinitely Resourceful
Exceptionalism
Unreasonable Standards
Company Over Self
High EQ
Learns Quickly
First Principles Thinker
Compensation and benefits
$200K base salary + 1–2% equity.
Fully covered medical, dental, and vision insurance.
Daily lunch and dinner, unlimited PTO, and parental leave.
$1K/month housing stipend (post-tax) if you live within walking distance (0.5 miles) from the office.
Covered Uber/Waymo from/to office.
Visa sponsorship available.
How we work
On-site in San Francisco, five days a week. Small team with massive surface area and ownership. You operate with complete autonomy of how to solve problems, and work with the team to set the direction of your work. We don't see engineers as code writers, but as problem solvers. You will do everything from talking to customers to writing custom GPU kernels in esoteric hardware.
Skills
Explore related jobs
More jobs at Wafer
Similar AMD jobs
Jobs in San Francisco
Proposal ManagerPGH Wong Engineering, Inc. · San Francisco, California
Software Engineer Intern, Mobile (Winter 2027)Notion · San Francisco, California- Residential Security Agent (San Francisco, CA)Concentric · San Francisco, California
Partner Marketing Manger, GSI & SIAnthropic · San Francisco, California | New York City- Marketing ManagerAsset Living · Denver, CO
Program Manager, ConnectMeter · San Francisco
Browse these categories
Market data for this role
All reports →- SeriesRole reportsOne role family at a time: how many openings, what changed this week, who is hiring, what it pays.
- SeriesSalary reportsWhat employers publish in job postings, by level and workplace. Not self-reported pay.
- Market overviewState of tech hiring, September 2026: up 4.8%Tech hiring rose 4.8% month over month in September 2026, with 411,122 new listings. Customer support and account executive roles led the growth.
