
The author is extremely bullish on Cerebras, citing reported Ultrafast inference revenue of about $200 million per megawatt annually and throughput of 750 tokens per second; they say customers including Jane Street pay a premium and that they are buying more capacity. The images add that OpenAI’s GPT-6.1 Sol Ultrafast is running on NVIDIA GPUs at low batch size, while Cerebras capacity is reportedly sold out.

By FPuklowski
ceo @ https://t.co/2EB7cIGTCv - the world's fastest ai neocloud - with no GPUs