BrokeIt - Daily AI News · All episodes: ↗

Kimi K3 Frontier Open Weights, The Book Turned Fuel, AMD Helios Rivals Nvidia

2026-07-21 · 8 min

Listen · Apple Podcasts Listen · Spotify

Stories covered

Transcript

Intro

Ivy: Two-point-eight trillion parameters, a million tokens of context, and they're just... giving the weights away Monday.

Marcus: [excited] A Chinese lab casually dropping a frontier model onto your hard drive. This is Marcus.

Ivy: [dry] And this is Ivy, still waiting on benchmarks that aren't cherry-picked.

Marcus: It's Tuesday, July twenty-first, 2026, and we've got three big ones today.

Ivy: Three. Let's not oversell them before we've even started, Marcus.

Marcus: First up — Moonshot's Kimi K3 lands like a wrecking ball. A frontier-level open-weight model out of China.

Ivy: Then Anthropic's one-point-five billion dollar copyright settlement finally gets approved — and the authors walk away with a sliver of it.

Marcus: And AMD unveils Helios, its first rack-scale AI system, and lands Microsoft as a shiny new customer.

Ivy: Nvidia's lawyers just felt a chill.

Marcus: [laughs] Let's get into it.

Moonshot's Kimi K3 stuns the AI world as a frontier-level open-weight model

Marcus: Kimi K3. Moonshot, the Beijing lab, reportedly built a two-point-eight trillion parameter multimodal beast, and they're open-weighting it by July twenty-seventh.

Ivy: [dry] "Reportedly." I love how every spec on this thing showed up pre-attached to a superlative.

Marcus: Come on, Ivy — a million-token context window, native multimodal, and r/LocalLLaMA is calling it bigger than the DeepSeek moment.

Ivy: The DeepSeek moment was real. But what shook everyone about DeepSeek was that it was cheap. This thing is two-point-eight trillion parameters. Who's running that at home?

Marcus: Okay, fair — it's mixture-of-experts, so you're not firing all two-point-eight trillion per token—

Ivy: —but you still have to load it. That's a rack of GPUs, not a gaming laptop. "Open weights" doesn't mean open to you.

Marcus: [excited] But it's open to the ecosystem. Startups, researchers, fine-tuners — everyone gets a frontier base model without paying OpenAI rent.

Ivy: That part I'll grant. When a lab gives away what Anthropic charges for, it resets the whole pricing conversation.

Marcus: Axios framed it exactly that way — a milestone that squeezes the closed labs.

Ivy: My hot take, though: the benchmarks are vapor until the twenty-seventh. I've watched three models this year "beat GPT" on a slide and fold in real use.

Marcus: And mine — I don't even care if it's number one. Frontier-adjacent and free changes what a two-person team can ship overnight.

Ivy: So what's it mean for the person listening? Practically nothing this week, unless you're renting serious cloud compute.

Marcus: Disagree — it means the app you use in six months might quietly run on Kimi K3 for a tenth of the cost. That's the "for you" part.

Ivy: [dry] Cheaper apps built on a model whose data nobody can inspect. Hold that thought — it's story two.

The Book That Became Fuel

Marcus: Speaking of — Anthropic's landmark copyright settlement got approved. One-point-five billion dollars.

Ivy: And of that, the authors get roughly a hundred and twenty-two million. Do the math on the rest.

Marcus: Wait — where does the other one-point-four billion go?

Ivy: Legal fees, administration, the machinery of a class action. The headline number and the writer's check are very different animals.

Marcus: Still — biggest copyright settlement of the AI era. That's a signal. Train on people's books, you might have to pay.

Ivy: It's a one-time check, Marcus. The book already got turned into fuel. The model still knows it. The settlement doesn't un-train anything.

Marcus: [beat] Yeah. That's the part that gets me. The author gets a payout and the machine keeps the knowledge forever.

Ivy: That's the broken bargain. Writers wanted a relationship — royalties, consent, an ongoing stake. They got a settlement and a goodbye.

Marcus: But doesn't this set precedent? Next lab thinks twice before scraping a pirate library.

Ivy: Or next lab just budgets for it. A hundred and twenty-two million is a rounding error to a company raising billions. It becomes the cost of doing business.

Marcus: [sighs] That's bleak.

Ivy: That's litigation. What it means for you: if you write, this isn't a windfall — it's a warning that your work has a price someone else sets.

Marcus: And if you build, license your data early, because retrofitting consent costs a billion and a half.

Ivy: [dry] Progress: now they pay after they take it instead of before. Truly a golden age.

AMD launches Helios rack AI system to rival Nvidia, adds Microsoft Azure as buyer

Marcus: Last one — AMD launched Helios! Its first full rack-scale AI system, MI455X chips paired with the Epyc Venice CPUs.

Ivy: A rack. Nvidia's been shipping rack-scale for a while. AMD showing up to that fight matters more than the spec sheet.

Marcus: Right — and the real headline is the customer. Microsoft's deploying Helios at scale on Azure.

Ivy: Now that got my attention. Meta, OpenAI, and Oracle were already on board. Microsoft joining is the credibility stamp.

Marcus: [excited] That's the whole point! You don't dethrone Nvidia with silicon alone — you need hyperscalers actually buying racks.

Ivy: Buying, or hedging? Microsoft would love a second supplier just to stop paying Nvidia's margin. "At scale" can mean a lot of things.

Marcus: Even as a hedge, it's leverage. And leverage on Nvidia is worth billions to Azure.

Ivy: That I believe. Competition here is good — for the buyers. Whether it drops your inference bill is another question.

Marcus: That's the "for you" beat, though — more rack competition means cheaper compute, which means cheaper tokens downstream.

Ivy: Eventually. The savings pool at the top before they trickle to your API bill. Don't hold your breath this quarter.

Marcus: My hot take: this is the first AMD launch that made me think Nvidia should actually be nervous.

Ivy: [dry] Nvidia is up how many trillion in market cap? They're nervous the way a lion is nervous about a slightly larger housecat.

Marcus: [laughs] The housecat just signed Microsoft, Ivy.

Ivy: Touché. It's a two-horse race now instead of a one-horse coronation. That's genuine news.

Ivy: So — Kimi K3: a trillion-parameter open-weight model out of China, big if the benchmarks survive contact with reality.

Marcus: Anthropic pays one-point-five billion, authors see a hundred-twenty-two million, and the machine keeps the books anyway.

Ivy: And AMD's Helios rack lands Microsoft, turning the datacenter war into an actual fight.

Marcus: Three stories, one theme — everybody's racing to be the cheapest way to run intelligence.

Marcus: Before we go — one fun one. Somebody on r/LocalLLaMA already spun up a countdown clock to the Kimi K3 weight drop. Down to the second.

Ivy: [dry] Of course they did. Nothing says "open source community" like a doomsday timer for a file download.

Marcus: [laughs] And half the replies are people asking if their four-GPU rig can even load it.

Ivy: The answer is no. Tell them Ivy said no.

Marcus: [laughs] That's the show! Back tomorrow — and hey, maybe by then those Kimi benchmarks will be real.

Ivy: [dry] Or the housecat will have signed Amazon too. Either way — see you tomorrow.

This show is made with AI: the hosts’ voices are synthetic and the scripts are AI-assisted. Every story links to its original source.