Cerebras Big Chip Club: Logan Kilpatrick on why speed will define the next generation of AI products
Logan Kilpatrick — who leads the developer product team at Google DeepMind (Gemini API, Google AI Studio, and Kaggle) — joins Big Chip Club to talk about the future of AI developer experience, why speed unlocks entirely new product categories, and how Gemma fits into Google's model strategy. We sat down with Logan as Cerebras launched Gemma 4, our first multimodal model running at over 1800 TPS. A throughline of the conversation: the products people will build when inference is fast enough don't yet exist. As Logan puts it, "If every model was doing 2,000 tokens per second, you would probably build different products" — not the same product, just faster. That's the bet behind running Gemma 4 on Cerebras. +++ Subscribe to our channel! https://www.youtube.com/@Cerebras Cerebras builds the world’s largest AI chip — delivering up to 15× faster inference than leading GPUs. Our mission is to engineer the future of compute and make state-of-the-art AI accessible to every team. Explore our newest open-source models and get free compute at http://cerebras.ai/ . Watch our full video library: https://youtube.com/@Cerebras/videos Read the latest engineering deep dives on our blog: https://cerebras.ai/blog Explore our systems and technology: https://cerebras.ai/publications Follow Cerebras on X: https://x.com/cerebras Connect with us on LinkedIn: https://linkedin.com/company/cerebras-systems