[00:02] dropped Kimi K3, a massive open-source monster that instantly parameter mogged every other open model in existence. And not just that, it has OpenAI and Anthropic terrified because its trust-me-bro benchmark performance is on [00:15] par with, and in some cases beating, Claude Fable and GPT-5.6 Soul. And ago, the government was telling you these models were too dangerous for the Chinese a few weeks to match their performance and then give away all the [00:28] weights for free. But big AI is not too happy about this, and there are already calls to ban Chinese models completely in the United States. In today's video, breakthrough that is Kimi K3 and find out what it means for the future of [00:41] 22nd, [music] 2026, and you're watching The Code Report. Like Claude Fable and GPT-5.6 Soul, Kimi K3 is a native multi-model million token context window and a staggering 2.8 trillion parameters, and [00:57] is optimized for jobs like long-horizon reasoning and, of course, coding. One interesting characteristic of K3 as a mixture of experts model is that it has mixture of experts model is that it has 896 total experts, of which exactly 16 [01:09] activate per token, which means it works just like a big corporation where 16 good programmers do all the work while 880 other managers sit there and do well for Kimi because it makes scaling about 2.5 times more efficient than K2. [01:23] Despite these gains in efficiency, Kimi was so popular upon release that their GPUs ran out of juice and they had to start turning away paying customers. And plans are currently sold out. That's unfortunate, but in theory, you could [01:36] are open. The weights are expected to be released on July 27th, but there's no chance in hell you'll be able to run it on your little gaming GPU. To run a monster like this, you'll need a massive array of data center caliber GPUs. But [01:49] would have unlimited access to a Fable Soul caliber model, and that would be amazing because the benchmark situation with K3 is pretty wild. K3 is ranked number one on front end code Arena at a 1,679 [02:02] Elo, which puts it ahead of Fable 5 and GPT-5.6 Soul. In addition, it lands in intelligence index. And if we look at every other coding benchmark, it's at frontier models. But you should never trust the trust me bro benchmarks [02:17] because many of the K3 numbers were produced with Moon Shot's own Kimiko different harnesses. That could make Kimiko look slightly better at coding, but to their credit, Moon Shot admits that K3 still trails Fable and GPT-5.6 [02:31] Soul overall, especially on benchmarks like Humanity's Last Exam where it's down by about 10 points. On top of that, artificial analysis measured a 51% not a good thing, especially when it comes to coding. In addition, it also [02:44] tends to spit out way more tokens than it needs to, which could ultimately end model itself being cheaper. When it comes to things like UI design and data visualization, it's extremely impressive for an open model, but in my opinion, [02:57] it's still one step behind Fable and GPT Soul. But one of the most interesting things about this release is the geopolitics surrounding it. Recently at Communist Party became the loudest advocate for free and open artificial [03:10] intelligence. Meanwhile, in the land of the free, Silicon Valley wants to regulate and gate keep it by pushing the fear narrative that it's about to take all of our jobs. In Washington, they're reportedly considering entity listing [03:22] Chinese AI labs and OpenAI's Dean Ball argued that open weights are inherently decelerationist. >> What? Bro, what are you talking about, man? >> And coincidentally, that's very similar [03:34] make in the '90s about Linux when he said Linux is communism. And also coincidentally, both of these guys have balls in their names. Frontier labs don't like open models simply because they divert the flow of money from them [03:47] to someone else. As of today, the odds the US government bans Chinese models is only sitting at 29% on Poly Market, but that could change quickly if they responsible for some kind of cyber attack. But, the best thing about K3 is [03:59] that it pushes the arms race forward. Alibaba also just released Qwen 3.8, which itself has 2.4 trillion parameters and open weights. And I think this model UI right for Horse Tender. But, before you let AI slop out your UI, you need to [04:13] check out mobbin.com, the sponsor of today's video. I've been using Mobbin for over 5 years now because it provides highly detailed breakdowns of every screen in thousands of popular web and mobile apps. And they just launched an [04:25] MCP server, which connects your AI agent to over 600,000 screens and user flows from apps in every category. This gives your agent real-world references, so it can design high-quality UIs for your specific use case, instead of just [04:39] spitting out generic purple gradient vibes law. It also lets you do deep UI prototype and your agent will use Mobbin's library to provide a ranked list of other apps that do it better. You can also ask it how your competitors [04:52] paywalls, and it'll show you every screen in their full user flow. And so, if you're tired of your UIs looking like the same as everyone else, I'd highly below. This has been the Code Report. Thanks for watching, and I will see you [05:06] Thanks for watching, and I will see you in the next one.