Chronicles. Aug. 24 - Aug. 30 2026
It’s Monday, and so it’s time for the next issue of Chronicles. Last week was very significant because it represented a tectonic shift from using expensive, efficient models to cheap and still very efficient models. In other words, AI is getting democratized.
Let’s start with the hardware. This week we’ve seen interesting news from OpenAI, which published the first benchmarks for their Jalapeño chip, a custom inference chip designed with Broadcom. It’s pretty impressive. They published over 700 tokens per second per user on DeepSeek R1 at concurrency of one, and about 1400 on Kimi K2.5 and GPT-OSS. The variety here matters because the chip has been benchmarked not only on OpenAI’s own models but on open-weight models from other vendors as well. The chip is critical in the ongoing price war with the Chinese providers. Remember that last week we had data on cutting prices on their Sol model by 20% on input and 33% on output, which is a promotional window that expires on 21 November, and a week before that they cut the price of Luna by a whopping 80%.