This video from the Y Combinator Paper Club features researchers and industry builders presenting advancements in AI systems, hardware specialization, and GPU-accelerated computing. The presentations cover:
The Networking Bottleneck
Key Multi-GPU Design Trade-offs
The Parallel Kittens Framework
Local Inference Viability
Measuring Efficiency (Intelligence per Watt & Joule)
The Rise of Automated Kernel Engineering
The Challenge of "Reward Hacks"
Building Robust Evaluators
The Distinct Phases of LLM Inference
SRAM Machines vs. GPUs
Heterogeneous Speculative Decoding Systems
Avoiding the CPU-GPU Bottleneck
Entity Component System (ECS) Architecture
Massive Performance Speedups