B2B
StartupLamb Labs
Lamb Labs is building MPUs (Model Processing Units). We hardcode the entire model, including its weights, into silicon. The model is the chip. This solves the memory-bandwidth bottleneck GPUs face, targeting 20,000+ tokens per second and 63× higher intelligence per watt.
Milestones
Milestone
YC Summer 202660 upvotes
Lamb Labs: Custom Chips for AI Inference
Co-designed models and silicon: we hardcode the model architecture into chips targeting 20,000+ tokens per second and 63x higher intelligence per watt.