
Wafer
Wafer is a YC Summer 2025 (S25) company. Based in San Francisco, CA, USA. Team of 4.
About Wafer
Summary
Wafer builds a continual inference platform that automatically profiles AI workloads and optimizes the GPU serving stack in real time, replacing the manual tuning and static configurations that teams otherwise maintain themselves. It targets AI companies and cloud providers running large-scale LLM inference, delivering lower latency and higher throughput as traffic patterns shift. The platform replaces both bespoke kernel engineering and prefetching infrastructure that engineers build to keep inference fast under load.
Go-to-market
AI teams and cloud providers start by routing inference traffic through Wafer's platform, which then continuously re-tunes kernels and the serving stack without manual intervention.
From the Y Combinator listing
AI that makes AI fast
Profile, optimize, and ship GPU kernels at speed of light
Facts
- B2B / B2B -> Engineering, Product and Design
- San Francisco, CA, USA
- 4
- Active
- Early
Founders (2)
Score every batch against your own thesis
GPAgent's sourcing agent reads every company in a batch, scores each one against your fund's thesis in a private workspace, and lines up the founders worth meeting before Demo Day. You decide who to contact. Your scores, notes, and pipeline never appear on these pages.
How Accelerator intelligence works