Get your free personalized podcast brief

We scan new podcasts and send you the top 5 insights daily.

As compute costs rise, driven by AI's ability to perform high-value tasks like automating scientific research, many current AI applications will be priced out. AI labs will be willing to outbid consumer use cases to allocate scarce compute for their own R&D, shifting the landscape of viable AI services.

Related Insights

The demand for AI tokens is growing faster than the supply of GPU infrastructure. This profound imbalance creates a market where not just top-tier AI labs, but also second and third-tier players will likely sell out their capacity. Superior models will command better margins, but the overall resource constraint means even lesser models will find customers.

The tech industry wrongly compares AI to software, which has near-zero marginal costs for new users. In reality, providing access to frontier AI models is a zero-sum game during compute crunches because of immense computational requirements. Servicing another user is expensive, leading to rationed access.

The 'Andy Warhol Coke' era, where everyone could access the best AI for a low price, is over. As inference costs for more powerful models rise, companies are introducing expensive tiered access. This will create significant inequality in who can use frontier AI, with implications for transparency and regulation.

Despite massive infrastructure investments, Greg Brockman believes demand for AI will consistently outstrip supply, leading to a long-term state of "compute scarcity." As AI tackles bigger problems like curing diseases, the appetite for computation will prove effectively infinite, making it a chronically scarce resource.

The focus in AI has evolved from rapid software capability gains to the physical constraints of its adoption. The demand for compute power is expected to significantly outstrip supply, making infrastructure—not algorithms—the defining bottleneck for future growth.

Escalating compute requirements for frontier models are creating a new market dynamic where access to the best AI becomes restricted and expensive. This shifts power to the labs that control these models, creating a "seller's market" where they act as "kingmakers," granting massive competitive advantages to the highest corporate bidders.

The current compute crunch isn't just a supply issue. It's because new AI models are so much more capable that they unlock a total addressable market (TAM) of valuable tasks that grows exponentially, far outpacing the linear or geometric growth of compute supply.

As AI models achieve human-level capabilities in valuable roles like software engineering, they can generate significantly more revenue from the same hardware. This increased monetization potential will cause the rental price of GPUs to skyrocket, potentially by over 15x, to match the economic value they produce.

The future of compute demand is a tale of two opposing forces. Enterprises will use AI to compress redundant data and streamline operations, reducing compute costs. Consumers, however, will demand generative AI for entertainment and personalization (e.g., 'Star Wars with my face'), creating massive new compute needs.

As demand for AI far outpaces compute supply, costs will rise. Only labs with the most lucrative algorithms, like OpenAI and Anthropic, can afford it. They reinvest massive revenues into the next training run, creating a self-reinforcing loop that raises the barrier to entry for any potential competitor, solidifying their duopoly.