At our latest YC Paper Club, researchers and builders presented on multi-GPU kernels, intelligence per watt, heterogeneous inference, and more.
Thank you to our presenters:
0:00 –
@FrancoisChauba1: The case for chip and kernel specialization
7:16 –
@stuart_sul: Parallel Kittens - Systematic and Practical Simplification of Multi-GPU Al Kernels (
21:29 –
@JonSaadFalcon: Intelligence per Watt - Measuring the Intelligence Efficiency of Local and Cloud AI (
31:05 –
@MarkSaroufim: When Al Starts Writing Systems Code
47:04 – Misha Smelyanskiy: Why AI Inference Needs Heterogeneous Hardware
1:04:33 –
@shacklettbp: Building a High-Throughput Game Engine that Runs ENTIRELY on the GPU (