The Tokenomics Foundation → https://goo.gle/4w4qnHd The FinOps Foundation → https://goo.gle/3U3IeR0 The Five-Layer Cake of Tokenomics: How Pinterest Thinks About AI Efficiency → https://goo.gle/44YEFy2 Are your AI budgets exploding? In this episode of The Agent Factory, host Luke Schlangen sits down with J.R. Storment, the newly appointed head of the Tokenomics Foundation under the Linux Foundation. Together, they discuss the industry's rapid shift from the era of tokenmaxxing to the great token panic, exploring how enterprises can scale AI agents without blowing through their revenue. J.R. introduces the emerging discipline of Tokenomics—defined simply as energy to intelligence to value—and maps it across three critical pillars: 1. Production: Data centers, edge devices (such as local Mac Minis or phones), and procured tokens. 2. Consumption: Programmatic model routing, caching, and prompt architecture. 3. Value: Monetization, hybrid credit systems, and the shift from seat-based to usage-based pricing models. They also cover Adobe's Big T framework for token efficiency, Pinterest's five-layer cake of consumption, and how hardware memory bottlenecks are ending the era of subsidized AI. Plus, don't miss J.R.'s predictions in our Rapid Fire Round on the future of algorithmic resource optimization! Chapters: 0:00 - Intro 1:15 - Meet J.R. Storment & The Linux Foundation 2:18 - The parable of the blind men and the elephant 5:34 - From token maxing to the great token panic 10:29 - The nonlinear growth of global tokens 12:36 - The explosion of context windows & agentic loops 14:02 - Global token projections (Goldman Sachs data) 15:38 - What is a token? The 4 core functions 19:59 - The Broad AI spend landscape & KV cache 28:09 - Defining tokenomics: Energy to intelligence to value 29:58 - FinOps vs. tokenomics 31:19 - Production: Data centers, edge, & local tokens 35:38 - Consumption: Pinterest's five-layer cake & Adobe's levers (Big T) 40:40 - Value: The Shift from Seats to Usage-Based Pricing 45:19 - The birth of the Tokenomics Foundation & Tokenomicon 52:46 - Rapid fire round: Predicting the future of token efficiency 56:43 - Outro & advice for engineers More resources: * FOCUS Open Source Spec for Cloud Billing → https://goo.gle/4vXdrmo * Token Economics: The Atomic Unit of AI Value → https://goo.gle/4xbsjOS * FinOps Foundation Framework → https://goo.gle/3TOdovE Watch more The Agent Factory → https://www.youtube.com/playlist?list=PLIivdWyY5sqLXR1eSkiM5bE6pFlXC-OSs 🔔 Subscribe to Google Cloud Tech → https://www.youtube.com/@googlecloudtech #AIAgents #Tokenomics #CloudFinOps #GenerativeAI #GoogleCloudTech #TheAgentFactory Speakers: J.R. Storment, Luke Schlangen Products Mentioned: Google Kubernetes Engine, Gemini, Antigravity
The Tokenomics Foundation → https://goo.gle/4w4qnHd
The FinOps Foundation → https://goo.gle/3U3IeR0
The Five-Layer Cake of Tokenomics: How Pinterest Thinks About AI Efficiency → https://goo.gle/44YEFy2
Are your AI budgets exploding? In this episode of The Agent Factory, host Luke Schlangen sits down with J.R. Storment, the newly appointed head of the Tokenomics Foundation under the Linux Foundation. Together, they discuss the industry's rapid shift from the era of tokenmaxxing to the great token panic, exploring how enterprises can scale AI agents without blowing through their revenue.
J.R. introduces the emerging discipline of Tokenomics—defined simply as energy to intelligence to value—and maps it across three critical pillars:
1. Production: Data centers, edge devices (such as local Mac Minis or phones), and procured tokens.
2. Consumption: Programmatic model routing, caching, and prompt architecture.
3. Value: Monetization, hybrid credit systems, and the shift from seat-based to usage-based pricing models.
They also cover Adobe's Big T framework for token efficiency, Pinterest's five-layer cake of consumption, and how hardware memory bottlenecks are ending the era of subsidized AI. Plus, don't miss J.R.'s predictions in our Rapid Fire Round on the future of algorithmic resource optimization!
Chapters:
0:00 - Intro
1:15 - Meet J.R. Storment & The Linux Foundation
2:18 - The parable of the blind men and the elephant
5:34 - From token maxing to the great token panic
10:29 - The nonlinear growth of global tokens
12:36 - The explosion of context windows & agentic loops
14:02 - Global token projections (Goldman Sachs data)
15:38 - What is a token? The 4 core functions
19:59 - The Broad AI spend landscape & KV cache
28:09 - Defining tokenomics: Energy to intelligence to value
29:58 - FinOps vs. tokenomics
31:19 - Production: Data centers, edge, & local tokens
35:38 - Consumption: Pinterest's five-layer cake & Adobe's levers (Big T)
40:40 - Value: The Shift from Seats to Usage-Based Pricing
45:19 - The birth of the Tokenomics Foundation & Tokenomicon
52:46 - Rapid fire round: Predicting the future of token efficiency
56:43 - Outro & advice for engineers
More resources:
* FOCUS Open Source Spec for Cloud Billing → https://goo.gle/4vXdrmo
* Token Economics: The Atomic Unit of AI Value → https://goo.gle/4xbsjOS
* FinOps Foundation Framework → https://goo.gle/3TOdovE
Watch more The Agent Factory → https://www.youtube.com/playlist?list=PLIivdWyY5sqLXR1eSkiM5bE6pFlXC-OSs
🔔 Subscribe to Google Cloud Tech → https://www.youtube.com/@googlecloudtech
#AIAgents #Tokenomics #CloudFinOps #GenerativeAI #GoogleCloudTech #TheAgentFactory
Speakers: J.R. Storment, Luke Schlangen
Products Mentioned: Google Kubernetes Engine, Gemini, Antigravity