[BidClub_]
Empire · · 66 分钟

算力是一个在群聊里交易的万亿美元市场 | Brett Harrison & Andrawes Bahou

Brett HarrisonAndrawes Bahou

其他资产半导体金融投资
YouTube ↗
TL;DR
  • 算力已经是一个规模约1万亿美元的实体市场,但其中相当大一部分仍通过电话、短信、Slack和WhatsApp交易,市场也没有标准远期曲线。 单笔合约规模可从1亿美元到10亿美元,有时甚至达到100亿美元;买方对几乎相同的产能,可能收到天差地别的报价。正如 Brett Harrison 所说,“算力实际上的大部分买卖,都是通过短信完成的”。

  • 从约3家超大规模云厂商扩张到约500家 GPU 服务商的爆发,正在让算力商品化,同时催生一个庞大的经纪层。 进入这一市场可能需要资本、GPU、电力、场地和需求,但买方必须区分真实供给与转卖另一家经纪商承诺的经纪商。Andrawes Bahou 将这种已经预先承诺、实际却不存在的产能称为“幽灵电力”(phantom power)。

  • 高速增长的 AI 公司,正面临立即高价采购与承担多年产能风险之间的残酷取舍。 按主持人举的例子,如果公司每天增长5–8%、甚至10%,为避免收入损失,可能愿意多付10–20%,而已观察到的案例中溢价最高达到2倍;更低的价格则可能要求签下3年或5年的承诺。期货和期权有望将算力合约与风险拆开,解决“眼下的特定需求”与提前数年出售的产能之间的错配。

  • 参考价格的机会取决于交易数据,而不是公有云价目表。 Brett 表示,自2025年末以来,算力价格整体上涨,因为需求增长超过了可部署供给,尽管超大规模云厂商的 GPU 毛利率在压缩。因此,指数业务向主要 Neocloud 厂商付费,以获取实际计费系统中的价格,并在 Bloomberg 发布 H100、H200、B200 和 B300 基准价。公有云按需价格“总体上毫无用处”,无法用于给主导市场的大额私下议价合约编制指数。

  • Architect 押注的是有到期日的期货和期权,而不只是加密市场式永续合约,因为前者更适合算力由合约驱动的现金流。 一座数据中心目前每GPU小时收入5美元,可能需要对冲续约价格跌至2美元的风险,例如设定4美元或3.50美元的底价;一家无法确定最终需要4个还是8个 H200 节点的 AI 买方,则可以购买期权。Brett 表示,“永续合约很棒”,但它们无法对冲与某个明确未来日期挂钩的现金流。

  • 资本可能是最具系统性的瓶颈,因为贷款方无法有把握地给贷款底层资产定价或对冲。 除非借款人或承购方具备投资级信用,融资成本可能达到20–25%,甚至根本拿不到融资;与此同时,随着 AI 资本开支上升,超大规模云厂商的自由现金流已接近零,甚至转负。市场正在从围绕 Google 的广告现金流做授信,转向思考如何给算力本身做授信;Brett 的比喻是,一辆只有“油门和手刹”的车,不可能安全地高速行驶。

  • Andrawes 预计,随着开放模型和更高效的模型不断进步,前沿实验室的定价权会下降;但 Brett 表示,数十亿 AI 用户目前普遍还没有意识到这一变化。 Brett 随后将约1万亿美元的实体算力交易与2.8–3万亿美元的实体原油交易作比较,后者的纸货市场约为实体市场的40倍。如果算力本身再增长10倍,他认为名义金额可能增长2个、甚至3个数量级:“人们根本没有意识到这个市场有多大,也没有意识到我们实现这一规模的速度有多快。”

摘要 · 为研究而整理的核心内容

1. 算力已商品化,但尚未成为可互换商品

  • Brett 的出发点是:AI 经济最大的支出项目,就是租用或购买 GPU;大额合约规模从1亿美元到10亿美元不等,有时甚至达到100亿美元。因此,买卖双方都对未来算力价格承担巨大敞口,但目前无法干净地对冲“潜在风险上的差异”。

  • Andrawes 将市场演进追溯到 Google、AWS 和 Azure,再到规模化时代逐渐形成的认识:更大的计算机可以持续改进模型。高资本需求随后吸引贷款方进入,而 FluidStack、Lambda 等专业 GPU 云厂商,则为机器学习团队提供原始 GPU 访问能力,而不要求它们使用完整的超大规模云厂商技术栈。

  • 结构性变化在于供给者激增:过去可能只能从3家超大规模云厂商采购的产能,如今可以来自约500家纳入跟踪的 GPU 服务商。当大量供应商都能凭借资本、设备、电力和场地进入市场,而需求仍然极其旺盛时,Andrawes 认为这“说明某种东西正在商品化”。

  • Brett 用2017年的 Bitcoin 作类比,强调了市场的不透明和地域套利。Andrawes 则指出关键差异:单个 Bitcoin 可以互换,也不受所在地影响;算力却是一项服务,受芯片显存、网络互联、散热、编排软件、利用率、供应商和地理位置共同影响。它是异质化商品,而不是可以随身搬运的标准单位。

2. 采购是通过关系网络完成的询价,市场充斥“幽灵电力”

  • 买方可能向10位 AI 创始人朋友打听“你那边的人”,然后得知产能要16周后才有,而自己的需求是今天。经纪商横跨超大规模云厂商、直接供应商和中间商,后者可能承诺2周交付40,000块 GPU;买方必须判断每个报价背后是真实机器,还是“一个经纪商的经纪商”。

  • Andrawes 所说的“幽灵电力”,指的是集群尚未存在就已经被拿出来销售。开发商首先需要客户签下3年协议;这份承诺才能撬动贷款方资本,用于向 Dell、Supermicro 和 Nvidia 等供应商采购设备,随后再安装到数据中心。一座中型集群的交付可能需要4–16周,甚至更久。

  • 对于每天增长5–8%、甚至10%的公司,速度可能足以支撑其多付10–20%;在已观察到的高增长推理业务交易中,溢价最高可达2倍,因为错失收入的代价高于单位经济性恶化。另一种选择是签下更便宜的3年或5年合约,但一旦增长速度或 AI 需求发生变化,承诺的容量可能严重错配。

3. 金融化将产能决策与价格风险拆开

  • Brett 对商品市场的类比是 Hershey 采购可可豆:公司确实需要豆子,但不必承担未来可可价格带来的全部风险。银行和专业交易公司可以对风险建模、承接、分散或对冲,让生产者和消费者专注于自己的主营业务。

  • 主持人反问,创业公司或许应该完全暴露在风险中——如果 AI 助手失败,公司反正也会失败。Brett 给出的更窄答案是,对冲不必消除公司的创业判断:如果采购产能需要2个月,而买方担心期间价格可能翻倍,就可以在锁定实体合约之前先“锁定价格”。

  • 定价证据仍然高度碎片化。超大规模云厂商的 GPU 算力毛利率已经压缩,支持算力商品化的判断;但数据中心各类 GPU 的平均价格自2025年末以来上涨,因为需求扩张快于供给。即便产能几乎等价,报价仍可能大幅分化,因为市场没有一个被普遍关注的基准价来锚定谈判。

4. 计费系统数据是可交易指数的基础

  • AWS 式公有云按需价格和抓取的 API 价格,无法覆盖市场定价重心:数千块 GPU 的私下议价、长期合约。Brett 表示,决定性价格存在于“买方、卖方以及记录交易的计费系统”之中。

  • 指数业务向主要 Neocloud 厂商付费,以获取实际计费记录,汇总已完成交易的价格,并在 Bloomberg 发布 H100、H200、B200 和 B300 算力的参考价格。其目标类似于原油价格报告机构:把分散的双边交易汇总成 WTI 或 Brent 一类的基准。

  • 交易所通过传统的最低费用加收入分成机制向指数提供方付费,使数据提供方能够参与交易所的成交量和未平仓量。Brett 认为,指数设计是关键:如果基准价基于按需租赁,就无法对冲一个通过3年期合约经营、一次采购5,000或10,000块 GPU 的现货业务。

  • Brett 还提到,Nvidia 宣布对这些合约提供25%的折扣,试图提供一条远期曲线并建立底价;但他认为,这本应是市场的功能,而不是 Nvidia 的职责。

5. 有到期日的合约比永续合约更匹配算力的期限

  • Architect 已取得指定合约市场牌照,并向 CFTC 提交方案;CFTC 正在评估指数可靠性、操纵控制,以及现金交割还是实物交割。Brett 承认 CME 是直接竞争对手,但认为 Architect 的直接接入、云端交付、API 和平等接入机制,更适合不熟悉传统期货基础设施的商业算力用户。

  • 永续合约适合没有自然到期日的敞口——有人可以一直持有 Apple 多头,直到自己改变主意——但算力买方通常面对到期日明确的合约。与贷款和利率风险类似,这些义务需要用期货或期权对冲特定期限,而不是用永续合约对冲价格敞口。

  • 一家数据中心目前每GPU小时收入5美元,18个月后续约时可能只能收到2美元,进而威胁其偿债能力;期货或期权则可以设定4美元或3.50美元的底价。这种产品转移了续约风险,却不要求数据中心修改底层客户合约。

  • Brett 举了一个合约设计案例:一家 AI 公司想在1年内使用4个 H200 节点,每个节点包含8块 GPU,但60天后可能需要增加到8个节点,而且所有节点必须位于同一个数据中心。它可以签订4节点的年度合约,同时购买以4个节点为标的的看跌期权和以8个节点为标的的看涨期权,从而保留灵活性;期权费低于提前承诺未必用得上的产能。

  • 对于操纵风险,Brett 认为,公开数据和工具可以让风险得到标准化和平滑处理,而重大危机往往出现在不透明的场外市场。CFTC 的意见征询期正在讨论如何确保算力指数可靠并抵抗操纵。

6. 信贷市场被迫开始给算力本身做授信

  • Brett 表示,没有投资级信用支持的借款人,可能面临20–25%的利率,或根本拿不到融资,因为贷款方缺乏运营历史,无法确定残值,也无法确信今天的芯片或算法未来仍然有用。次贷式 GPU 贷款可以为较小运营商提供资金,但他明确将其危险性与20年前的次贷过度扩张相提并论。

  • 在市场高端,Apollo、KKR、Blackstone 和 Brookfield 等公司与 JPMorgan、Goldman Sachs 一起,为拥有 Microsoft 或 Meta 等高信用承购方的大型运营商提供融资,起始规模通常在10亿–30亿美元。中型市场机构可能承做数亿美元至10亿美元的融资,资产租赁公司则服务于长尾市场。

  • Andrawes 认为,今天突然出现的金融化,其催化剂是超大规模云厂商的自由现金流在 AI 资本开支压力下跌向零甚至转负。资本市场不能再假定 Google 的广告业务能够吸收所有风险,而必须思考如何给算力这一资产类别估值、投保和对冲,因为“经济体量如今已经大到任何单一参与者都无法独自承受”。

7. 约束是地方性的,但资本决定供给调整速度

  • 让 GPU 上线需要资本、电力和设备,具体瓶颈环节取决于项目。高显存系统和 InfiniBand 训练网络可能面临更长交付周期;GW 级园区要接入电网,可能需要多年,于是项目转向表后天然气发电;规模更小的部署则通过经纪商搜寻分散的 MW,后者从中赚取“离谱的钱”。

  • Brett 认为,资本是最大的系统性约束:贷款方无法给算力定价或对冲,因此要求长期承诺。他用一辆只有油门和手刹的汽车作比喻——加装正常刹车后,车反而可以开得更快。Brett 更广泛的反驳是,需求会绕开任何瓶颈重新配置;中国面临的约束,就推动了在旧硬件上也具备竞争力的高效算法。

  • 电力、芯片、内存和许可审批的优先级,都无法脱离地理位置做全球统一排序:美国公司正在接受菲律宾或冰岛的产能,过去只考虑100 MW以上项目的超大规模云厂商,如今30 MW的项目也会接电话。Andrawes 提到,距离11月选举只剩5周,预计选举相关的谨慎情绪会在选后消退;他还开玩笑说,反对项目落地的 NIMBY 人群或许更愿意接受“别在我家后院,但可以在我的经纪账户里”。

  • Andrawes 预计,随着开源模型和更高效的模型不断进步,前沿实验室的定价权和主导地位会下降。Brett 只是对这一变化被多少人理解提出异议:AI 圈内的少数人会讨论开放模型,但数十亿 AI 用户普遍还没有意识到这一点。他随后回到规模问题:算力并不只是“新石油”;约1万亿美元的算力已经在实体市场交易,而原油实体市场规模为2.8–3万亿美元,其衍生品市场约为实体市场的40倍。算力再扩张10倍,名义金额就可能增长2个或3个数量级。

核验说明

  • 逐字稿对指数/数据业务的归属说法并不一致:主持人称 Andrawes 是该指数业务的搭建者,而 Brett 称“我的公司”发布这些指数,并描述其向 Neocloud 厂商购买计费数据。因此,本文统一使用“指数业务”这一说法,不指明其归属。
完整逐字稿
Speaker 1

Welcome to Token 2049. Token 2049 returns on October 7th and 8th, bringing together 25,000 attendees, 300 speakers, and 500 exhibitors at the world's largest cryptocurrency event. Token 2049 is taking place in parallel and in partnership with our own Asia Digital Asset Summit. So you can attend both conferences in Singapore within one week. There will be over 1,000 side events during Token 2049 week, culminating in the post-2049 and Formula 1 weekend, and the speaker list is very packed. Shane Kopan of Poly Market, Jeff Yang of Hyperlid, Arthur Hayes, NASDAQ CEO Adena Freidman, and many others. Join us in Singapore on October 7th and 8th for Token 2049 and Digital Asset Summit Asia. Nothing said on Empire constitutes a recommendation to buy or sell any investment or product.

1. Why Does Compute Need Markets?

Okay, everyone. I'm really happy about this. I've wanted to do this episode for a long time. I feel like I don't understand the computing markets well enough, so I'd like to invite two of the smartest people in those markets that I know: Brett Harrison, founder of Architect, and Andrawes Bahou, founder and CEO of Compute Desk.

Andrawes Bahou

I'm glad to be here. Thank you very much.

Brett Harrison

Nice to see you. Glad to be here.

Speaker 1

So, guys, I think Brett, I could ask you to start by saying that you two are very knowledgeable about the computing markets. We have derivatives, indexes, and markets, and I think that for those who haven't been following us, a lot of computing markets will be launching in the relatively near future. What market problem do computing technology markets solve? Why do we build markets around computing technologies?

Brett Harrison

The main unit of expenditure in the AI economy is companies renting or buying GPUs to perform various types of training or inference tasks. As I'm sure people know, the amount of money we're talking about here is extremely large. People are spending $100 million to $1 billion, or even $10 billion, at a time on large-scale computing contracts.

The companies that create them are sometimes called neoclouds or hyperscalers. Many different companies in this field are involved in selling computing technology to other companies. This raises the question of what will happen if the prices of computing technologies decrease over time. What if they grow significantly over time?

That difference in potential risk, and the price at which they could offer these computing technologies, is a commodity that cannot be hedged today. We're trying to collectively track the prices of computing technology and create derivatives based on computing technology so that people can hedge or speculate and have full transparency about the future value of computing technology.

Speaker 1

Good. Who are these people? Can we identify them? We have labs, hyperscalers, neoclouds, and cloud companies. There are many of them, and they are synonymous with one another. Can you, perhaps, Andrawes, tell us who the players are here?

Andrawes Bahou

Once upon a time, there were hyperscalers. That's Google, AWS, whatever. They bought GPUs, accumulated hardware, and then made it available to you on AWS, GCP, or Azure.

Then people started needing a lot more GPUs. The buyers of hyperscale capacity were, of course, small research labs, then larger research labs, AI companies, and developers who needed to train new models, perform inference, and do everything related to machine learning workloads.

You probably know the hard lesson, which is that you really should increase the amount of compute and the amount of data. If you increase the amount of compute, your models will continue to increase their capabilities. Along came OpenAI, and this started the race to create something new, until the consensus was, “Let’s just create better algorithms.” But then, at that point, they said, “Let’s just build bigger computers.” So they started making them much, much bigger.

In the process, because of how capital-intensive this new infrastructure was, a new type of participant started to emerge: lenders. Lenders are large private credit firms—not necessarily just private credit firms, but credit firms of all kinds.

Eventually, the bond markets started coming in because hyperscalers needed a lot more capital to create these clusters. They came in and started guaranteeing free cash flow for hyperscalers. The idea is that Google can generate a lot of money from advertising, but it has a big, expensive data center project. The lender is happy to guarantee this because they think, “Google will always pay its dues. We’re going to lend it the capital to build these clusters.”

This led to the emergence of three types of participants in this market: lenders, infrastructure providers, and then buyers. These are businesses and AI labs. Now things started to get a little more fractalized.

Speaker 1

Right. So, out of these three types of players, just to be clear, the players today are Google, Google Cloud, Amazon, OpenAI, Anthropic, JPMorgan, and Goldman Sachs, for example.

Andrawes Bahou

Yes, for example. Now, the world dominated by these big hyperscalers is starting to come to an end. Previously, if you needed compute, you went to one of them.

In the 2010s, new, smaller companies that were just providing GPUs—not a full cloud experience, just GPUs as a service—came along. Companies like FluidStack and Lambda saw a need for machine learning engineers and AI companies that needed to run machine learning workloads or train machine learning models. They said, “We’ll just have a cloud of GPUs for you. You don’t need to bother with all this AWS stuff. We just have a cloud of GPUs for you.”

Now fast-forward to today, and what has happened is the hyperproliferation of these types of GPU-as-a-service providers. Some call them neoclouds. I don’t like the term. Some people call them data centers with GPUs. Some are neo-hyperscalers, such as CoreWeave.

The thing is that there are many more new participants in this market, and this will be an important moment. When there are a very large number of participants in a market, it is usually a sign that something is being commoditized, and that is exactly what is happening now.

Many people see a low barrier to entry into the computing services market. This is a product. There’s not much differentiation here. There’s crazy demand for computing equipment. Everyone wants to serve it, and a lot of people can. That’s the important part: all you need is some capital to buy the equipment, a place to put the equipment, and demand for it. That’s all you need to play this market.

So suddenly, with this hyperproliferation of new GPU-as-a-service providers, we’re tracking something like 500 of them now. Previously, GPUs could be obtained from 3 companies. There are 500 of them now.

Speaker 1

So it just so happens that you can get GPUs.

Andrawes Bahou

Yes. Or you can get compute on a GPU, and then rent the compute for anywhere from 3 to 500.

Speaker 1

Yes.

2. Compute Still Trades In Group Chats

Andrawes Bahou

What emerged from that is now a very interesting category. This is a phenomenon of the last 2 years: a whole bunch of brokers for computation. There are a lot of people who need computation, so the cardinality of the number of people who need computation and the cardinality of the people who supply it both increased. Brokers came in to essentially match supply and demand, and a very inefficient market was formed.

Speaker 1

Well said. Yes. One of my friends was the chief of staff at one of the big AI companies, and I asked, “What do you do all day?” He said, “I’m on the phone with brokers trying to get compute.” And he said, “This is the worst job. They all quote different prices at different times.”

That was a few months ago—probably over a year ago—so I think things have probably changed now. But the market is evolving every day in all directions.

And so, first, to Andrawes’s earlier point: if you have access to any source of electricity, to the power grid, or to the ability to buy GPUs, people are creating these little neoclouds all over the place, in every country, through various kinds of renewable energy sources, in abandoned gas stations, old car depots, and refurbished Bitcoin mining rigs. It’s everywhere.

Now this spread is happening. The question is, first of all, what is the actual fair price for any of these things? You have this giant, opaque market. Most of the actual buying and selling of compute happens through text messages, phone calls, and opaque channels.

There is a complete price gap between what hyperscalers offer for compute and what neoclouds offer for compute. It is a very heterogeneous market in terms of what compute actually is, and it can vary depending on whom you get it from. Regardless, there’s a critical need right now to bring this pricing out into the open and standardize it, and to have a liquid tool that allows you to track and hedge what’s actually happening in the computing market.

And that extends from computing outward in both directions from a supply-chain perspective, because there are a lot of things that go into creating compute. There is electricity and energy. There are foreign currencies so that you can purchase GPUs. There are metals that can be used in memory production, on the one hand. On the other hand, there are things such as inference.

Andrawes Bahou

So, there are many different aspects that affect the cost of computing and what computing is used for. That’s the industry we’re in together: how we hedge the prices of everything related to computing and everything that’s part of the computing supply chain.

Brett Harrison

It’s shockingly reminiscent to me of Bitcoin in 2017. There were a few years of peer-to-peer trading, and there was sometimes a 20% price dislocation between Asia and the U.S., but it didn’t matter because nobody really cared about the markets until 2017. Bitcoin was skyrocketing, going from 1,000 to 20,000, and then, oh my God, we have to create really efficient markets around these things. This is just 100 times larger in scale, maybe 1,000 times larger in scale, than that.

Is this it? What are the similarities, and maybe what are the differences?

Andrawes Bahou

Yes, it’s probably more than 1,000 times. I don’t know—if I do a quick count, will it be more than 1,000 times? That’s more than 1,000 times, and it has some similarities and differences.

As far as similarities go, I’ve heard people talk about what the OTC markets looked like in digital assets about 10 years ago: very broker-to-broker, peer-to-peer, opaque pricing, and large dislocations and arbitrages between different geographic locations and markets.

One of the big differences here is that it is absolutely a commodity, but it’s not a perfectly interchangeable good. If you have 1 bitcoin, 1 bitcoin is 1 bitcoin. That’s true. It’s in the address; it’s in the wallet. It is not tied to a specific location or provider. It exists on its own.

Computation does not exist on its own. It’s not something you pull out of one place and carry to another. It’s a service provided by a data center that determines everything from the amount of memory on each chip to the type of network, the cooling—liquid or air—whether there’s good software for managing GPUs, and the utilization rate.

There are so many variables that make computing a heterogeneous commodity, which is precisely what makes it difficult to immediately value and bring to market in the form of an exchange.

Speaker 1

Good. Can you tell me, maybe going back to Andrawes for a second? By the way, is your name Andy or Andre?

3. Why Finding GPUs Takes So Long

Andrawes Bahou

Both. My name is a bit difficult to pronounce, but my own mother calls me Andy.

Speaker 1

I think I understand. I think I’m close. Okay, let’s go back to Andy for a second.

I was listening to a podcast with the founder of Instinct on Invest Like the Best. Have you been at Meta before?

Andrawes Bahou

I was at Meta.

Brett Harrison

You were at Meta. There’s a lot of buzz around Instinct. The founder of Instinct said that 40% of his time is now spent trying to get computing resources.

I’m wondering what that really means. Who is he calling? Who is he writing to? What company is he calling? Does he contact a broker? Do you just call someone? Are you calling the data center? Are you calling a hyperscaler? Who is this? What does it look like?

Andrawes Bahou

We said earlier that there are now hundreds of computing-power providers, and there are probably just as many brokers of that power, if not more. What he is probably doing is calling his top 10 friends who run AI companies and asking, “Who’s your guy who provides you with computing resources?”

They’ll send him a list of 10 people, and he calls them. They say, “He’s a good guy. You should talk to him.” So he picks up the phone and says, “Hey, I need compute.” They reply, “Great. It’ll be available in 16 weeks.” He says, “Okay, I need it today. We’re growing very quickly. I just raised $1 billion. I need a lot more, and I need it a lot faster.”

He’s going to bypass all these brokers and all these end-to-end compute providers, from the hyperscalers all the way down to the random broker who promises they’ll have 40,000 GPUs in 2 weeks. Because of the opaque market, you don’t know how trustworthy someone is behind the scenes. They could be a broker for another broker for someone who actually has the compute.

He’ll probably spend a lot of time trying to verify whether that compute exists. There’s a big problem now. I call it phantom power, so it’s probably worth spending 30 seconds on.

Many of these GPU clusters are presented in such a way that the developer of the project—which could be a data center, a hyperscaler, or a neocloud—wants to raise debt capital to buy the equipment. To do this, the lender tells them, “Of course, we’ll do it, provided that you find someone to sign an agreement with you.”

This is called a buyout agreement: someone who will sign a long-term lease on this equipment for the next 3 years, even though the equipment doesn’t exist yet. They have to look for an imaginary cluster that only comes into existence when a customer signs up and says, “I’ll buy it for the next 3 years.”

Then they take that agreement to a lender and say, “I have a buyout agreement. Give me the capital.” They unlock the debt capital, spend it on Dell, Supermicro, and Nvidia, buy the hardware, store it in a data center somewhere, and then make it available.

Why? It takes 4 to 16 weeks to complete an order for a medium-sized cluster, if not a little longer these days. So that’s what he spends his time on. This is what he spends 40% of his time on.

It works primarily through email, Slack, and WhatsApp—something like a request for quotation, or RFQ—where he has to separate the real capacity from the fake capacity and then get a competitive offer.

Brett Harrison

Do you think that, at his scale, with Instinct being very popular and having just raised a lot of money, he’ll turn to brokers? Does he need to go to brokers, or can he just go directly to computing-service providers?

Andrawes Bahou

I’m sure it combines all of the above. He probably knows, and we know this from our own experience, because even in the process of building an exchange for computing derivatives, we naturally have computing sellers coming to us looking for buyers and computing buyers coming to us looking for sellers in the spot-computing market.

We have relationships with several non-Aklar companies that have capacity or are building capacity, so we know we can act as a broker in this market. He’ll approach brokers and maybe look for suppliers he has already gotten compute from in the past.

As Andy said, he’ll turn to his friends who run other AI companies and have their own verified sources. He’ll probably put it out on Twitter, saying, “Hey, I’m looking for compute. I just raised $1 billion. Contact me if you have these GPUs next week.” Whatever means are necessary, it will come to him.

Again, it’s such a strange market because there’s such a large volume of capital expenditure and such a large volume of debt being issued to finance these things. Yet there is no standardized market for either spot computing or the actual buying and selling of computing, nor any instrument that could hedge or lock in that price, including things like insurance.

The market for cluster provisioning, providing residual value for chips, or establishing a lower bound on the cost of computing is extremely volatile.

Speaker 1

4. How Much Compute Should You Buy?

Good. Staying with this intuitive story a little longer, in the podcast, the founder of Manus said that they’re growing at 10% per day—let’s say 5% to 8% per day, but maybe it’s actually 10% per day. Those numbers add up very quickly if he stays at that pace.

He said, essentially, “How much leverage do I need to increase? Should I buy computing resources? Should I buy enough computing resources to serve 100 million users, or 1 billion users?”

I don’t know how many users they have today, but if he buys enough compute resources just to serve his current users, he’ll run out of compute in a week. But if he buys enough to make it to, I don’t know, the third quarter or something, he’ll actually be done in 3 weeks.

Since their growth rate is so extreme, how would you think about this problem if you were in his shoes? How far out should you buy? I’ll turn it over to Andy to answer that question, and I’ll answer the question of what should be done in this particular market.

Andrawes Bahou

Yes. He was stuck between a rock and a hard place. The reason is that it’s not an easy decision. He’ll have to compromise between several options.

Either he moves fast and compromises on the price of the compute he gets, or he compromises on the price by 10% to 20%, or by 2 or 3 times—probably 2 times. You can compromise up to 2 times. We’ve seen this in transactions from some very large inference providers that are growing very rapidly.

For them, the opportunity cost of lost revenue outweighs some degradation in unit economics. He has to get the compute very quickly, because otherwise it’s a revenue wipeout.

He knows what his demand will be in the short and medium term, maybe in the near future, but not in the long term. So if he wants a good price and buys a lot of compute from the other side, whoever is selling it to him will say, “Hey, take a 5-year commitment or a 3-year commitment.”

That’s the problem. There are several dimensions you need to consider when buying compute: the length of the commitment and the volume of compute you’re buying.

Brett Harrison

For such a term of commitment, the inflexibility here is something that can be untangled or broken if you have a working futures market.

Speaker 1

Good. And enter Brett, stage left.

Brett Harrison

Yeah, maybe another step back is that anyone who buys compute resources in these absolutely astronomical quantities—or sells compute resources because they're building data centers and trying to get contracts with buyers—has to combine 2 different things. One is buying or selling a product, and the other is managing all the risks. In all traditional commodity markets, these 2 things are separate.

Generally speaking, when Hershey buys cocoa beans to make chocolate bars, they don't have to retain all the risk of what happens with the price of cocoa beans going up and down or with supply. In general, this is something they can pass on to the other party. Obviously, there are financial companies whose entire job is to model risk, store that risk, hedge it, and create a diversified portfolio of risks. That's their whole reason for existing, while producers and consumers of goods do not have to do so.

The problem with compute is that there's no way for a company like Instinct, or any other company that's grown so quickly in the last week or 2 weeks and is now thinking about the next 5 years of compute it's buying, to think, “What if, in 1 year's time, nobody cares about AI anymore? What happens if, a year later, I discover that I've overbought compute resources by a very large amount? What happens if, a year later, I discover that I've under-purchased compute resources by a very large amount?”

No matter which of these cases happens, I can't say, “No matter what, I can sell the last part of this contract if I don't need all of it,” or, “No matter what, I'm guaranteed a floor price or a ceiling price.” That's why we have financialization. It's not just a vehicle for speculation, gambling, or betting on the prices of things. It's a way of transferring risk and addressing things like duration mismatches.

I have a specific need right now, and it might be related to what my compute resources are going to be over the next year. Someone really wants to sell compute resources in 5 years. There is a complete mismatch in duration, and who will take the risk from the second to the fifth year? There are probably many firms that would be happy to hold that risk on their own behalf if there were a liquid derivatives market or even an over-the-counter derivatives market, which is obviously what we're trying to do.

Speaker 1

I have a stupid question about hedging. Do Fortune 500 companies do this on their own? Does Hershey, for example, hedge all of this in-house with the help of a risk department and trading department? Do they send it to the trading department at Goldman Sachs? Do they send it to a specialized firm?

Brett Harrison

Both. This is exactly the combination. The canonical example is when airlines used to hedge their own fuel prices through the crude oil markets, and then they stopped doing that, which was a bad move given what's happened to oil over the last year. There were also a lot of internal trading divisions in big energy companies and commodity companies.

They either do it themselves, outsource the risk-management function to a large bank or FCM that will handle it for them, or do both. There are internal analysts who figure out how to properly manage treasury, from interest rates to commodities, and then they work with the trading department to figure out how to make or close those deals in the most efficient way.

Speaker 1

Do you think a startup should hedge risks? You both build startups. Shouldn't you just put everything on the line if you're a startup founder? For example, if AI assistants don't take off, then there's no incentive.

Brett Harrison

I think it's not about being as protected as possible or as vulnerable as possible. Here's a very simple example. Let's say you're going to spend the next 2 months looking for this compute, and you're just worried that by the time you finally find the person who will sell you the compute, the price will have doubled. So fix the price. You can lock in a price right now using a futures contract.

It's not about being as hedged as possible or avoiding risks. You simply know that there is a price movement in the market, and you can hedge against it and fix the price today before you can find a supply contract for yourself. This is just one example of how this can help even startups.

5. Why Are Compute Prices Rising?

Andrawes Bahou

Good. Let's look at what's happening in the compute markets themselves with pricing. I have a few questions, but maybe the first thing is: what's happening with the pricing of compute over the last couple of years? Obviously, we know that demand has increased. Has it stabilized? Has it decreased at all?

Brett Harrison

Let's back up a bit. One interesting thing is that, before I get to the price of compute, there's something interesting that's happened to compute margins. If you look at the reports of hyperscalers, the margins they get on their GPU compute have shrunk. This is an encouraging signal that this is indeed a product. In commodity goods, you don't have the luxury of nice, big profits, so those profits are shrinking.

There's a narrative—and it's an absolutely true narrative, by the way—that this year, or since the very end of 2025, the price of compute technology has gone up. We see this in the prices we track. On average, the price of compute for every type of chip and every type of GPU for data centers has gone up. The reason is that there's high demand and not enough supply to meet that demand in a timely manner, so demand is growing a little faster than supply. That drives prices up.

You don't think about compute prices the way you think about, “What's the price of crude oil right now?” You can check what WTI is showing or what Brent is today, and prices around the world will be trading with a very tight spread. Today, in compute, you can get 2 very different quotes for practically the same thing, but the spread is very large.

The reason it's so large is simply because the market is very opaque and poorly informed about compute pricing. It's essentially an illiquid market. There is no single global standard price that everyone looks at and says, “Here's the price. Let me see how much the purchase price of my compute differs from this price, or how differently I sell it.”

We're trying to change that. My company publishes price indices for H100, H200, B200, and B300 compute on Bloomberg, oddly enough, so people can get an actual reference price for what it costs to buy compute today. But that average price is very different from the spreads that are being realized.

We expect this market to become much smarter and more efficient. Those spreads will tighten, and you'll have market-maker input and more efficient price discovery on exchanges. All of this will contribute to the compression of these spreads, or rather, the pricing inefficiencies that exist now.

Andrawes Bahou

So, probably a stupid question: you can't determine one price for compute, right? There is no single compute price. There is a compute price for B200 or H100. You need to set a price for each of them. Is that correct?

Brett Harrison

Of course. Crude oil brokers will privately negotiate a price. Now there are so-called price-reporting agencies. They take all the crude oil transactions—these are technically midstream companies—and check the prices at which those transactions were completed that day. They combine them into what you see as the price of crude oil, whether that's WTI or Brent. That's the price that actually exists.

Somewhere, there is a private negotiation between supply and demand, and 2 people agree on a price. There is a price-reporting agency that collects this price, and it becomes an index and a reference price. That's what is happening in compute right now. People can negotiate privately, and we try to report how those private negotiations ended and publish a reference price.

Andrawes Bahou

So how do you get this data? We're trying to think about where compute operations are being performed. There is compute that you can get on demand through AWS. There are prices listed online. You could scrape websites and take data from people's APIs.

Brett Harrison

Overall, that's useless. The reason it's useless is that most compute is not purchased on a website. Most compute is privately agreed upon through large contract agreements.

So where does this price exist? We thought it was in the minds of the person who bought it, the person who sold it, and the billing system that recorded it. We have exclusive agreements with the biggest neoclouds to have exclusive access to their billing systems, where we get the real, agreed-upon transaction price. We take all these prices, compile them, and then publish an index.

Andrawes Bahou

And why would they give it to you?

Brett Harrison

We pay them a lot of money.

Andrawes Bahou

You pay them money. Good. And then you monetize on the back end by charging an exchange?

Brett Harrison

The exchange pays us.

Andrawes Bahou

Us?

Brett Harrison

Yes. It's very typical for derivatives exchanges. When you build a futures contract, option, or perpetual based on another company's index data, you want to split the economics between the index provider and the exchange.

6. Can Architect Take On CME?

There is a combination of a minimum commission plus a revenue share. Andrawes will benefit from the growth in volume and open interest on our exchange, and that's how this partnership works. Again, this is very typical if you buy index data from S&P, MarketVector Indexes, or some other index company. This is a very standard relationship.

Speaker 1

For anyone listening, since we skipped all the introductions at the beginning, Brett, you build an exchange, and Andrawes, you're building an index business. What does the computing market look like today? We have examined the players in the computing markets. What is the computing space like with hyperscalers, neoclouds, lenders, and all that stuff? What does the market look like for the markets you guys are building?

When I think about futures trading, CME, ICE, and people like that come to mind. So I'm curious how you see them being your direct competitors.

Brett Harrison

Architect acquired a designated contract market license a few months ago—a DCM, which is the CFTC's designation for the ability to organize an exchange of futures and options on commodities and related products. We are in a small group of potential players who can offer futures and options to U.S. clients.

The market for derivatives on computing technologies does not yet exist. This is what we are trying to create, and this is what we have various proposals to the CFTC to launch. There is a public comment period going on right now that the CFTC has announced to get comments on how to create a reliable index, how to prevent manipulation, and how to actually track the price of computing technology.

Which is better: cash-settled futures or physically settled futures? All the issues related to the creation of a completely new derivatives market are being raised now, and hopefully the process will be completed soon. After that, we will be able to launch these products for the first time.

Speaker 1

So, is CME your main competitor here?

Brett Harrison

Yes, absolutely.

Speaker 1

And why would anyone trade on Architect and not CME?

Brett Harrison

There are several reasons, everything from the contract design to the user experience—how you actually trade and sign up. We could talk about a few examples.

Speaker 1

This is a better platform. This is the best technology platform. One hundred percent. Good. Let's talk about a few of them.

Brett Harrison

First of all, one of the reasons we work with Andrawes as our index provider is because it's really important what the actual index tracks in terms of creating a good future. If you have a future that tracks the price but doesn't actually provide a good hedging product—that is, it doesn't hedge enough of the variance in the prices of your actual computing systems that you see in the spot market every day—then it's not going to be a useful future, and nobody's going to want to trade it.

For example, there are several futures offerings that track the on-demand rate for computing systems. This is, as Andrawes said, not how people trade computing systems in the real market. They don't go out and buy 7 GPUs for 2 days. The market is not on demand: they buy 3-year contracts for 10,000 or 5,000 GPUs. So this index that Andrawes and his team created is much more similar to actual contract rates for computing. That's the first thing.

Secondly, you can't go to CMEGroup.com, register a trading account, enter your information, and then start trading through their API. It's not that they don't provide it—they do not distribute their own products. They work through intermediaries, brokers, and futures commission merchants. You have to pay a lot of fees for the data itself, and it is often not cloud-based. The same thing applies to the API.

We built our exchange much more in the image of a modern exchange, similar to how crypto exchanges were built, where you can register an account directly and start trading almost immediately after signing up. You can use a GUI. You can use the API. There's no advantage for a big player over a small one. We provide open access to everyone in the cloud, and these are the things that we think are necessary to create completely new products in the modern era.

The last thing I'll say is that, especially for commercial-computing hedgers, who we hope to make among the first customers of these products, these are not people who have traded futures before in their lives. A lot of them don't even know what futures are, but they intuitively understand the idea of futures because they live and breathe these sales contracts. They understand what it means to fix prices.

7. How Do You Hedge Compute?

These are not customers with existing loyalty to a particular exchange. They will be looking for people who live and breathe computing and artificial intelligence every day, can speak their language, and can engage them with a product that specializes in computing. No Silicon Valley AI firm wants to look through soybeans and crude oil to get to computing on the front end. They want something tailored specifically to their needs.

Speaker 1

Speaking of adapting to those needs, what does someone actually come up to you and say? Do they say, “Hey, I just got this big contract. Can you show me what it actually looks like?”

Brett Harrison

Yes, of course. Here's an example. Someone bought, say, a 3-year contract for the availability of computing resources and sold someone an 18-month contract for compute. A data center—a neocloud—has locked in someone who bought 18 months of compute resources, and let's say they bought 100% of the capacity in their data center.

Now let's say they have the option to renew this contract in 18 months at whatever the current rate is. You don't know what that rate will be in 18 months. Let's say today it's $5 an hour for a GPU. What if in 18 months it's $2 per hour per GPU, and you don't have the ability to hedge that price and lock it in now?

Here's what they come to us with and say: “How can I protect myself against the risk of prices going down? When I sell this contract again in 18 months, I won't even be able to pay my debt anymore because it's just not that valuable. Can I set a minimum price for computing power at $4? Can I set a minimum price for computing power at $3.50?”

It will be some combination of futures or options to lock in that price.

Speaker 1

I remember when you launched Architect. If I remember correctly, the thesis was this: people come to trade, essentially.

Brett Harrison

That's true.

Speaker 1

Now you're saying that futures and options are more like a standard future. Or is it something like a perp, which I think most people listening to this are familiar with, as opposed to a standard future that expires? So why not a perp?

Brett Harrison

With a perpetual contract, it is impossible to fix a specific cash flow tied to a specific term in the future. Perpetual contracts are great when you need perpetual exposure to the price of something. For example, single-stock perpetuals were a great product because most people don't say, “I really want to hedge my March exposure to Apple.” People just say, “I want to be long Apple, and I want to be long Apple until I don't want to be long Apple anymore.”

That doesn't apply to computing, where people say, “I have a very specific contract that ends on a certain date,” because it's completely based on a contract. The same goes for interest rates. I think perpetual interest-rate contracts will be relevant, and we have some proposals for that as well. But mostly, people have loans, and loans have a specific term, so people need to hedge the duration of that particular loan.

This is why derivatives in the form of standard futures or standard options are so important. That's why our exchange chooses the right tool for the job. Is this a perpetual contract? Is this a future or an option? Or some combination of all 3?

Speaker 1

What other deals are there that people are making? I'm trying to imagine what some of the other ones look like.

Brett Harrison

There is an example. It's really very simple. There was an AI company that came in and said, “Hey, we want 4 nodes.” A node is a group of 8 GPUs. They said, “We want 4 H200 nodes per year, but in 60 days it may turn out that we actually need 8 instead of 4, and they all have to be in the same data center.”

How do you give them that flexibility? People will only want to sell them a 4-node annual contract or an 8-node annual contract. We solved this problem for them by structuring a 1-year contract for 4 nodes, a put option on 4 nodes expiring in 60 days, and a call option on 8 nodes at the same time.

This gave them the flexibility they needed not to commit in advance to something where there was a certain level of uncertainty. They paid a premium, and that premium was worth it to them so they wouldn't spend the extra money on 8 extra nodes if it turned out that they didn't need them. It was a way to manage a risk that made sense.

Speaker 1

That makes sense. They had a risk, and we were able to move it away from them and give it to someone else who was willing to take that risk.

Brett Harrison

Yes.

Speaker 1

8. The Risks Behind GPU Lending

Okay, now I understand. It became clear to me. Can you guys do one thing we haven't done? At the beginning, you talked about players and lenders, and we didn't talk too much about how they're borrowing and lending GPUs now.

What happened in this market?

Brett Harrison

This is what we hear from all the different clients in this area. The lenders themselves will also become derivatives customers, because one of the ways they hedge the fixed-income contracts they issue is through derivatives.

Obviously, people want to borrow as much money as possible to build as much capacity as possible and try to meet the growing demand for computing technology. The problem, of course, is whether you're willing to guarantee anyone who comes up with the idea, “I want to buy some computing technology.” And, of course, it is very difficult.

First, most people don't have years of history to prove their creditworthiness. Second, people don't know what GPUs are going to cost in a few years, whether the capacity is going to mean the same thing in a couple of years, or whether the algorithms are going to be different. So unless you're an investment-grade company, good luck getting a loan from a lender at less than 20% or 25%, and so on.

We hear that sometimes people can get these loans, but they are at exorbitant interest rates, or they just can't get loans at all. People don't want to guarantee them. So we start to see a few different things.

First, we see companies emerging that do things like subprime lending, which in itself is perhaps important for getting these small companies off the ground, but it's also dangerous, and it reminds us very much of the subprime mortgage problems of about 20 years ago. At the same time, we're starting to see lenders trying to figure out how they can secure their contracts, either through actual insurance or eventually through a derivative product, to reduce some of the risk and be able to lend more.

Finally, we're seeing a lot of other companies entering this market, especially if you think about it as a parallel to the private-lending boom. Large-cap trading firms that understand computing and maybe can implement it themselves, along with many existing large private-credit issuers, are entering this market and combining venture capital and debt to finance these transactions. This is what we see.

Speaker 1

Okay, Andy, what about you? And one more question: maybe you can take this if you want, Andy, but who are the biggest players in the computing-lending markets today? Are these banks or private-credit funds?

Andrawes Bahou

There are several layers to the world of debt. There are big, very big banks and financial institutions like Apollo, KKR, top-tier Blackstone, Brookfield, and so on. There are maybe about 10 such firms, and then you can add JPMorgan and Goldman Sachs to them.

These firms provide loans to large operators whose clients are very large and very creditworthy, so-called offtakers. So it's an operator like CoreWeave, for example, that sells a contract to Microsoft or Meta, all of whom have incredible balance sheets and creditworthiness. This is one segment of the lending world.

In this world, nobody considers loans smaller than $1 billion, $2 billion, or $3 billion. This is the starting amount. Now there's a sort of middle market of computing, and then there's a very long tail of computing, where there are very different types of credit firms.

That's where some of these underwriting shops come in, like Blue Owl, for example. They underwrite these loans for GPU data centers—GPU deals, mostly—ranging from, I don't know, hundreds of millions to $1 billion.

And then there's a very long tail of asset-leasing companies making loans to everyone else. Therefore, the risk profile in each of these segments is very different. The type of risk is very different, the type of operator is very different, and what is fundamentally underwritten is different.

There is one thing that is happening right now. I think a lot of people are wondering why we didn't hear about this financialization of computing 9 months ago, and suddenly it's so trendy now. There is one thing that seemed to catalyze this moment.

I don't know if your viewers saw this, but there's a chart that was circulating that showed the free cash flows of hyperscalers. It showed free cash flow dropping to zero or even becoming negative. The idea is that hyperscalers' ability to generate free cash flow is being affected by how much they spend on AI development, and that capital expenditure on AI development has caused it to drop to zero or become negative.

So the lenders who guaranteed Google's creditworthiness guaranteed free cash flow. Now there is no free cash flow. Suddenly, there's a big shift happening where, instead of capital markets underwriting Google's ability to make money from advertising and then pay the GPU bills, capital markets are thinking: How do we underwrite computing technology itself as an asset class?

This raises several questions: How do we evaluate this? How do we hedge this? How can we insure this? Previously, it was believed that all these risks would be completely absorbed by Google or Amazon. But not anymore, because the economy is now much larger than any single player can absorb on their own.

Hence financialization, if you think about the role of capital markets in transferring risks. That's why this is happening now.

Speaker 1

These hyperscaler free cash flows going to zero or becoming negative are a very real issue right now. It just happened.

Yes. Good. Brett, are there any concerns? It seems like you spent some time in Washington. Are there any concerns that if you create markets around something, it will mean that the markets will become more manipulated?

Brett Harrison

Of course, there are concerns. I think historically, when there was publicly available data and tools to track something, it helped normalize and smooth out risk over time. In all cases where there was a real financial crisis, it happened because everything that was happening was completely opaque and over-the-counter, because there was no market.

That's right. So, do you know what happened? What if there were publicly disclosed swap markets that represented something like credit default swaps and all these crazy financial instruments that were created during the financial crisis, but not during the implosion?

I think this is precisely one of the main questions being addressed in the current CFTC comment period: How do we ensure that indices are not subject to manipulation, that they reflect the fair price of computing, and that they are properly administered and monitored?

But we believe that this will only help the market. I think another turning point for me was Nvidia's announcement of its 25% discount on these contracts. What Nvidia essentially meant was that since there's no forward compute curve that tells everyone how much these things cost, we'll build one for everyone and set a floor price on it in the meantime.

9. What Really Constrains The AI Market?

This really shouldn't be Nvidia's job. This should be the job of the market. So you're absolutely right that this is a problem, but we think this is the wrong way to look at the market.

Speaker 1

10. What Is Everyone Missing About Compute?

Yes. Good. Guys, maybe I want to take a little time while we think about wrapping this up with what you see that other people don't see yet. Andy, you have data that no one else in the world has. Maybe let's talk about where exactly the main constraint in the market is.

I think if you ask people on Twitter, or anyone who's not really involved, maybe a third would say we're computationally constrained, a third would say we're memory constrained, and a third would say the whole network would grind to a halt. What do you see, and what is the real limitation over the next few years?

Andrawes Bahou

Bringing the power of the GPUs online depends on 3 or 4 things. The first is access to capital. The second is access to power inside the data center. The third is access to equipment.

Depending on the scale you're working at, one of these three paths becomes the critical path, or the bottleneck. In a sense, it's true that all three things are bottlenecks. But by the definition of a bottleneck, only one of them can be true at a time.

There's a hardware bottleneck, and there's a memory bottleneck. If you talk to Supermicro or Dell, or any of the GPU hardware vendors right now, you're going to have a very real conversation of this sort: “Hey, we have 2 configurations of this equipment, with a large amount of memory and with a smaller amount of memory.” The lower-memory configuration will be delivered in a much shorter period of time.

There's a second type of conversation where it's something like: If you need this specialized network fabric called InfiniBand, if you want to build training clusters, as opposed to using some other general-purpose network fabric, then it's going to take a lot more time and much longer lead times to acquire the specialized fabric.

Now, if you don't have these restrictions, your access to electricity could become a limitation. Access to electricity is not the same depending on scale. If you're trying to build gigawatt-scale campuses, interconnection to the grid is a multiyear nightmare.

People are resorting to building gas turbines onsite and building their own, essentially, natural-gas power plants behind the meter.

Brett Harrison

This is another type of restriction. If you’re operating on a small scale, the amount of data center capacity that’s ready to go and running in the United States, and then in other parts of the world—but it’s particularly acute in the United States—exists for small deployments but not for large ones. So we see even the biggest players—hyperscale companies, OpenAI, AI-powered companies, and so on—moving downmarket. You couldn’t call them before unless you had 100 megawatts. Now they’ll pick up the phone if you have 30 megawatts, and it’s getting smaller and smaller because of all this extra capacity, which, by the way, is related to the fact that a lot of this demand is based on inference. But that’s another one of these phenomena that we’re seeing.

Your ability to find data center capacity in the United States is interesting. I have a funny story. There’s a class of people—you can think of them as Miami-style real estate brokers—who happen to know a bunch of data center vendors with a megawatt here and 5 megawatts there for a small deployment. These guys take a commission for introducing you to their buddy who runs a data center that has some space. They’re making crazy money right now by essentially hooking up someone who wants to host some equipment to a data center in some random location.

This is the second limitation. The third constraint, probably the most systemically important, is capital. There’s a kind of paradox: There is currently huge demand for computing technologies, but at the same time, lenders are hesitant to underwrite computing technologies themselves. Their hesitation means that these projects take longer to underwrite, and therefore to finance, and therefore to deploy, and so on.

These 2 things at the same time don’t make sense to me, because if lenders knew about the state of demand, they wouldn’t have to be so hesitant. They would know that there’s a market for this compute, and this stems from their inability to guarantee, price, and hedge compute. So I have a little analogy, which is this: For these lenders, a buyout agreement, which kind of forces the person they’re lending to to enter into a very long-term contract with the client, is a kind of hedge. It’s something like, “I need the whole project committed for the next 3 years.”

It’s kind of like if you have a car. It’s like a handbrake. What’s funny is that if you have a car and you only have a gas pedal and a handbrake, you’re not going to go very fast. If you have a small brake, you can go a lot faster because you can stop pretty quickly. That’s what hedging provides. Hedging instead of a full buyout. If you provide hedging, they will feel much more confident.

By that I mean that lenders and capital markets in general will feel much more comfortable making loans because they know they have a place to hedge. This is the third and largest systematic constraint, which is a systemic constraint: access to capital. That’s it.

From our position, I want to give a very different answer to what we see as limitations. Regardless of the limitations, this market is so powerful, so important, and so in demand that it will find a way around them. This is perhaps a naive answer, but I actually hardly believe in any of these restrictions.

Look at what happened in China. Do you think people just find their way around a constraint? 100%. Look at what happened in China. They were told that they couldn’t have access to any of the latest and greatest chips. What they did instead was create extremely cost-effective algorithms that can run on older-generation hardware and are now almost completely competitive with cutting-edge models. They understood it.

Whatever the constraint is—whether it’s not enough capacity, politics forcing people to put a moratorium on building new data centers, or not enough credit or capital—it will somehow find a way. Because if companies are just greedy for tokens and they need them, logically, they will solve it somehow. Maybe if all 3 of these things are limited, people will just create better algorithms that can run more cheaply on the available electricity.

So, if I had to rank the limitations, let’s look at power, chips, memory, and local permitting. How would you rank them?

Andrawes Bahou

It’s impossible to rank them because this set of constraints is different in each geographical location, and the market is becoming increasingly global. We worked with neoclouds that said, “Oh, we have a new data center in the Philippines. We have one in Iceland. We have American companies that are now very happy to do these weekend contracts from locations outside the U.S., even though they could never do that before.”

So suddenly, there’s no single answer to the global ordering of all these constraints.

Brett Harrison

How do you think the U.S. midterm elections will impact data centers, compute, and these markets? If data centers become this huge political topic that leads to more regulation, is that something that will perhaps put— you talked about a pause and a break—or is that a moment of pause?

Andrawes Bahou

The simple answer to that is that it’s the end of September, and the elections are in November. There’s not much time left. We have 5 weeks. When this is over, I think the floodgates will open again.

I think we’ve temporarily stalled because nobody knows how the general public feels about these big data center contracts when they go to the polls. Once this is over, we’ll get back to business.

Brett Harrison

Do you have an opinion, Andy? This is a bit of a spicy question that I think might be a good closing point.

Andrawes Bahou

I think the average NIMBY-type data center opponent will finally be very happy to trade compute futures. You’re bothering me by building a data center in my backyard, but I know it’s profitable, and I want you to know that I’m not stupid. I want to have some involvement in this activity—not in my yard, but maybe in my brokerage account.

Speaker 1

I don’t disagree with you. As we think about wrapping this up, maybe we could ask one final question that seems very obvious to you. It’s a very broad, general question and may refer to the markets for compute, compute technology, or artificial intelligence in general. From your perspective, what seems very obvious to you that you think the general public is currently overlooking?

Andrawes Bahou

I think what seems obvious to me is that the pricing power and dominance of cutting-edge labs like Anthropic and OpenAI are bound to decrease. The availability, power, and efficiency of the remaining models, like open-source models and so on, are just going to continue to increase.

Better tools will come along, and we’re launching something very soon that will help with that. But there will be greater accessibility for people to draw conclusions and apply them to all sorts of models that don’t necessarily limit them to just a few of the biggest companies.

Brett Harrison

I don’t think so. I think in niche AI circles, on X, and in places like that, people are certainly talking about open-source models and how good they are. But I don’t think the general public—the several billion users of artificial intelligence—has a real understanding of this.

There’s one thing that I think everyone has heard but not understood. There’s this banal saying that compute is the new oil. But in reality, it’s not that trivial. I want to put this number in perspective.

Physical trading of crude oil is a market of about $2.8 trillion, possibly $3 trillion. Physical trading of compute is approximately $1 trillion. We’re actually halfway there.

Now, if you consider that this is the order of magnitude of physical trading of compute, and then look at other markets that have both a physical market and a derivatives market, the derivatives market for compute on a notional basis will be several orders of magnitude larger—perhaps 1.5 orders of magnitude larger.

Crude oil, for example, has a paper market, or a crude oil derivatives market, that’s 40 times larger than the physical trading of crude oil. I expect the same thing to happen with compute, and that’s at its current scale.

What if compute grew another 10 times? We have the potential for 2, maybe 3, orders of magnitude of growth, given the notional transaction volume in compute. People just don’t realize how big it is or how quickly we’ve achieved this. How fast.

Okay, guys, thank you for that. Thank you. Congratulations to everyone.

Andrawes Bahou

Thank you very much. Let’s do it again in a year and see where the compute markets are.

Brett Harrison

Thank you, and see you later.