The Tokenomics Foundation → https://goo.gle/4w4qnHd The FinOps Foundation → https://goo.gle/3U3IeR0 The Five-Layer Cake of Tokenomics: How Pinterest Thinks About AI Efficiency → https://goo.gle/44YEFy2 Are your AI budgets exploding? In this episode of The Agent Factory, host Luke Schlangen sits down with J.R. Storment, the newly appointed head of the Tokenomics Foundation under the Linux Foundation. Together, they discuss the industry's rapid shift from the era of tokenmaxxing to the great token panic, exploring how enterprises can scale AI agents without blowing through their revenue. J.R. introduces the emerging discipline of Tokenomics—defined simply as energy to intelligence to value—and maps it across three critical pillars: 1. Production: Data centers, edge devices (such as local Mac Minis or phones), and procured tokens. 2. Consumption: Programmatic model routing, caching, and prompt architecture. 3. Value: Monetization, hybrid credit systems, and the shift from seat-based to usage-based pricing models. They also cover Adobe's Big T framework for token efficiency, Pinterest's five-layer cake of consumption, and how hardware memory bottlenecks are ending the era of subsidized AI. Plus, don't miss J.R.'s predictions in our Rapid Fire Round on the future of algorithmic resource optimization! Chapters: 0:00 - Intro 1:15 - Meet J.R. Storment & The Linux Foundation 2:18 - The parable of the blind men and the elephant 5:34 - From token maxing to the great token panic 10:29 - The nonlinear growth of global tokens 12:36 - The explosion of context windows & agentic loops 14:02 - Global token projections (Goldman Sachs data) 15:38 - What is a token? The 4 core functions 19:59 - The Broad AI spend landscape & KV cache 28:09 - Defining tokenomics: Energy to intelligence to value 29:58 - FinOps vs. tokenomics 31:19 - Production: Data centers, edge, & local tokens 35:38 - Consumption: Pinterest's five-layer cake & Adobe's levers (Big T) 40:40 - Value: The Shift from Seats to Usage-Based Pricing 45:19 - The birth of the Tokenomics Foundation & Tokenomicon 52:46 - Rapid fire round: Predicting the future of token efficiency 56:43 - Outro & advice for engineers More resources: * FOCUS Open Source Spec for Cloud Billing → https://goo.gle/4vXdrmo * Token Economics: The Atomic Unit of AI Value → https://goo.gle/4xbsjOS * FinOps Foundation Framework → https://goo.gle/3TOdovE Watch more The Agent Factory → https://www.youtube.com/playlist?list=PLIivdWyY5sqLXR1eSkiM5bE6pFlXC-OSs 🔔 Subscribe to Google Cloud Tech → https://www.youtube.com/@googlecloudtech #AIAgents #Tokenomics #CloudFinOps #GenerativeAI #GoogleCloudTech #TheAgentFactory Speakers: J.R. Storment, Luke Schlangen Products Mentioned: Google Kubernetes Engine, Gemini, Antigravity

The Tokenomics Foundation → https://goo.gle/4w4qnHd The FinOps Foundation → https://goo.gle/3U3IeR0 The Five-Layer Cake of Tokenomics: How Pinterest Thinks About AI Efficiency → https://goo.gle/44YEFy2 Are your AI budgets exploding? In this episode of The Agent Factory, host Luke Schlangen sits down with J.R. Storment, the newly appointed head of the Tokenomics Foundation under the Linux Foundation. Together, they discuss the industry's rapid shift from the era of tokenmaxxing to the great token panic, exploring how enterprises can scale AI agents without blowing through their revenue. J.R. introduces the emerging discipline of Tokenomics—defined simply as energy to intelligence to value—and maps it across three critical pillars: 1. Production: Data centers, edge devices (such as local Mac Minis or phones), and procured tokens. 2. Consumption: Programmatic model routing, caching, and prompt architecture. 3. Value: Monetization, hybrid credit systems, and the shift from seat-based to usage-based pricing models. They also cover Adobe's Big T framework for token efficiency, Pinterest's five-layer cake of consumption, and how hardware memory bottlenecks are ending the era of subsidized AI. Plus, don't miss J.R.'s predictions in our Rapid Fire Round on the future of algorithmic resource optimization! Chapters: 0:00 - Intro 1:15 - Meet J.R. Storment & The Linux Foundation 2:18 - The parable of the blind men and the elephant 5:34 - From token maxing to the great token panic 10:29 - The nonlinear growth of global tokens 12:36 - The explosion of context windows & agentic loops 14:02 - Global token projections (Goldman Sachs data) 15:38 - What is a token? The 4 core functions 19:59 - The Broad AI spend landscape & KV cache 28:09 - Defining tokenomics: Energy to intelligence to value 29:58 - FinOps vs. tokenomics 31:19 - Production: Data centers, edge, & local tokens 35:38 - Consumption: Pinterest's five-layer cake & Adobe's levers (Big T) 40:40 - Value: The Shift from Seats to Usage-Based Pricing 45:19 - The birth of the Tokenomics Foundation & Tokenomicon 52:46 - Rapid fire round: Predicting the future of token efficiency 56:43 - Outro & advice for engineers More resources: * FOCUS Open Source Spec for Cloud Billing → https://goo.gle/4vXdrmo * Token Economics: The Atomic Unit of AI Value → https://goo.gle/4xbsjOS * FinOps Foundation Framework → https://goo.gle/3TOdovE Watch more The Agent Factory → https://www.youtube.com/playlist?list=PLIivdWyY5sqLXR1eSkiM5bE6pFlXC-OSs 🔔 Subscribe to Google Cloud Tech → https://www.youtube.com/@googlecloudtech #AIAgents #Tokenomics #CloudFinOps #GenerativeAI #GoogleCloudTech #TheAgentFactory Speakers: J.R. Storment, Luke Schlangen Products Mentioned: Google Kubernetes Engine, Gemini, Antigravity