Skip to content
Wafer logo

Wafer

Hiring

Wafer is a YC Summer 2025 (S25) company. Based in San Francisco, CA, USA. Team of 4.

  • AI

About Wafer

GPAgent's own fund-neutral profile, written from public information. Updated Oct 5, 2026.

Summary

Wafer builds a continual inference platform that automatically profiles AI workloads and optimizes the GPU serving stack in real time, replacing the manual tuning and static configurations that teams otherwise maintain themselves. It targets AI companies and cloud providers running large-scale LLM inference, delivering lower latency and higher throughput as traffic patterns shift. The platform replaces both bespoke kernel engineering and prefetching infrastructure that engineers build to keep inference fast under load.

Go-to-market

AI teams and cloud providers start by routing inference traffic through Wafer's platform, which then continuously re-tunes kernels and the serving stack without manual intervention.

From the Y Combinator listing

How Wafer describes itself in Y Combinator's public directory. Original listing (opens in a new tab)

AI that makes AI fast

Profile, optimize, and ship GPU kernels at speed of light

Facts

Industry
B2B / B2B -> Engineering, Product and Design
Location
San Francisco, CA, USA
Team size
4
Status
Active
Stage
Early

Founders (2)

Score every batch against your own thesis

GPAgent's sourcing agent reads every company in a batch, scores each one against your fund's thesis in a private workspace, and lines up the founders worth meeting before Demo Day. You decide who to contact. Your scores, notes, and pipeline never appear on these pages.

How Accelerator intelligence works