
From tokenmaxxing to tokenomics for your AI agents
The Tokenomics Foundation → https://goo.gle/4w4qnHd
The FinOps Foundation → https://goo.gle/3U3IeR0
The Five-Layer Cake of Tokenomics: How Pinterest Thinks About AI Efficiency → https://goo.gle/44YEFy2
Are your AI budgets exploding? In this episode of The Agent Factory, host Luke Schlangen sits down with J.R. Storment, the newly appointed head of the Tokenomics Foundation under the Linux Foundation. Together, they discuss the industry's rapid shift from the era of tokenmaxxing to the great token panic, exploring how enterprises can scale AI agents without blowing through their revenue.
J.R. introduces the emerging discipline of Tokenomics—defined simply as energy to intelligence to value—and maps it across three critical pillars:
1. Production: Data centers, edge devices (such as local Mac Minis or phones), and procured tokens.
2. Consumption: Programmatic model routing, caching, and prompt architecture.
3. Value: Monetization, hybrid credit systems, and the shift from seat-based to usage-based pricing models.
They also cover Adobe's Big T framework for token efficiency, Pinterest's five-layer cake of consumption, and how hardware memory bottlenecks are ending the era of subsidized AI. Plus, don't miss J.R.'s predictions in our Rapid Fire Round on the future of algorithmic resource optimization!
Chapters:
0:00 - Intro
1:15 - Meet J.R. Storment & The Linux Foundation
2:18 - The parable of the blind men and the elephant
5:34 - From token maxing to the great token panic
10:29 - The nonlinear growth of global tokens
12:36 - The explosion of context windows & agentic loops
14:02 - Global token projections (Goldman Sachs data)
15:38 - What is a token? The 4 core functions
19:59 - The Broad AI spend landscape & KV cache
28:09 - Defining tokenomics: Energy to intelligence to value
29:58 - FinOps vs. tokenomics
31:19 - Production: Data centers, edge, & local tokens
35:38 - Consumption: Pinterest's five-layer cake & Adobe's levers (Big T)
40:40 - Value: The Shift from Seats to Usage-Based Pricing
45:19 - The birth of the Tokenomics Foundation & Tokenomicon
52:46 - Rapid fire round: Predicting the future of token efficiency
56:43 - Outro & advice for engineers
More resources:
* FOCUS Open Source Spec for Cloud Billing → https://goo.gle/4vXdrmo
* Token Economics: The Atomic Unit of AI Value → https://goo.gle/4xbsjOS
* FinOps Foundation Framework → https://goo.gle/3TOdovE
Watch more The Agent Factory → https://www.youtube.com/playlist?list=PLIivdWyY5sqLXR1eSkiM5bE6pFlXC-OSs
? Subscribe to Google Cloud Tech → https://www.youtube.com/@googlecloudtech
#AIAgents #Tokenomics #CloudFinOps #GenerativeAI #GoogleCloudTech #TheAgentFactory
Speakers: J.R. Storment, Luke Schlangen
Products Mentioned: Google Kubernetes Engine, Gemini, Antigravity
The FinOps Foundation → https://goo.gle/3U3IeR0
The Five-Layer Cake of Tokenomics: How Pinterest Thinks About AI Efficiency → https://goo.gle/44YEFy2
Are your AI budgets exploding? In this episode of The Agent Factory, host Luke Schlangen sits down with J.R. Storment, the newly appointed head of the Tokenomics Foundation under the Linux Foundation. Together, they discuss the industry's rapid shift from the era of tokenmaxxing to the great token panic, exploring how enterprises can scale AI agents without blowing through their revenue.
J.R. introduces the emerging discipline of Tokenomics—defined simply as energy to intelligence to value—and maps it across three critical pillars:
1. Production: Data centers, edge devices (such as local Mac Minis or phones), and procured tokens.
2. Consumption: Programmatic model routing, caching, and prompt architecture.
3. Value: Monetization, hybrid credit systems, and the shift from seat-based to usage-based pricing models.
They also cover Adobe's Big T framework for token efficiency, Pinterest's five-layer cake of consumption, and how hardware memory bottlenecks are ending the era of subsidized AI. Plus, don't miss J.R.'s predictions in our Rapid Fire Round on the future of algorithmic resource optimization!
Chapters:
0:00 - Intro
1:15 - Meet J.R. Storment & The Linux Foundation
2:18 - The parable of the blind men and the elephant
5:34 - From token maxing to the great token panic
10:29 - The nonlinear growth of global tokens
12:36 - The explosion of context windows & agentic loops
14:02 - Global token projections (Goldman Sachs data)
15:38 - What is a token? The 4 core functions
19:59 - The Broad AI spend landscape & KV cache
28:09 - Defining tokenomics: Energy to intelligence to value
29:58 - FinOps vs. tokenomics
31:19 - Production: Data centers, edge, & local tokens
35:38 - Consumption: Pinterest's five-layer cake & Adobe's levers (Big T)
40:40 - Value: The Shift from Seats to Usage-Based Pricing
45:19 - The birth of the Tokenomics Foundation & Tokenomicon
52:46 - Rapid fire round: Predicting the future of token efficiency
56:43 - Outro & advice for engineers
More resources:
* FOCUS Open Source Spec for Cloud Billing → https://goo.gle/4vXdrmo
* Token Economics: The Atomic Unit of AI Value → https://goo.gle/4xbsjOS
* FinOps Foundation Framework → https://goo.gle/3TOdovE
Watch more The Agent Factory → https://www.youtube.com/playlist?list=PLIivdWyY5sqLXR1eSkiM5bE6pFlXC-OSs
? Subscribe to Google Cloud Tech → https://www.youtube.com/@googlecloudtech
#AIAgents #Tokenomics #CloudFinOps #GenerativeAI #GoogleCloudTech #TheAgentFactory
Speakers: J.R. Storment, Luke Schlangen
Products Mentioned: Google Kubernetes Engine, Gemini, Antigravity
Google Cloud Tech
Helping you build what's next with secure infrastructure, developer tools, APIs, data analytics and machine learning....