Ivy: Two-point-eight trillion parameters, a million tokens of context, and they're just... giving the weights away Monday.
Marcus: [excited] A Chinese lab casually dropping a frontier model onto your hard drive. This is Marcus.
Ivy: [dry] And this is Ivy, still waiting on benchmarks that aren't cherry-picked.
Marcus: It's Tuesday, July twenty-first, 2026, and we've got three big ones today.
Ivy: Three. Let's not oversell them before we've even started, Marcus.
Marcus: First up — Moonshot's Kimi K3 lands like a wrecking ball. A frontier-level open-weight model out of China.
Ivy: Then Anthropic's one-point-five billion dollar copyright settlement finally gets approved — and the authors walk away with a sliver of it.
Marcus: And AMD unveils Helios, its first rack-scale AI system, and lands Microsoft as a shiny new customer.
Ivy: Nvidia's lawyers just felt a chill.
Marcus: [laughs] Let's get into it.
Marcus: Kimi K3. Moonshot, the Beijing lab, reportedly built a two-point-eight trillion parameter multimodal beast, and they're open-weighting it by July twenty-seventh.
Ivy: [dry] "Reportedly." I love how every spec on this thing showed up pre-attached to a superlative.
Marcus: Come on, Ivy — a million-token context window, native multimodal, and r/LocalLLaMA is calling it bigger than the DeepSeek moment.
Ivy: The DeepSeek moment was real. But what shook everyone about DeepSeek was that it was cheap. This thing is two-point-eight trillion parameters. Who's running that at home?
Marcus: Okay, fair — it's mixture-of-experts, so you're not firing all two-point-eight trillion per token—
Ivy: —but you still have to load it. That's a rack of GPUs, not a gaming laptop. "Open weights" doesn't mean open to you.
Marcus: [excited] But it's open to the ecosystem. Startups, researchers, fine-tuners — everyone gets a frontier base model without paying OpenAI rent.
Ivy: That part I'll grant. When a lab gives away what Anthropic charges for, it resets the whole pricing conversation.
Marcus: Axios framed it exactly that way — a milestone that squeezes the closed labs.
Ivy: My hot take, though: the benchmarks are vapor until the twenty-seventh. I've watched three models this year "beat GPT" on a slide and fold in real use.
Marcus: And mine — I don't even care if it's number one. Frontier-adjacent and free changes what a two-person team can ship overnight.
Ivy: So what's it mean for the person listening? Practically nothing this week, unless you're renting serious cloud compute.
Marcus: Disagree — it means the app you use in six months might quietly run on Kimi K3 for a tenth of the cost. That's the "for you" part.
Ivy: [dry] Cheaper apps built on a model whose data nobody can inspect. Hold that thought — it's story two.
Marcus: Speaking of — Anthropic's landmark copyright settlement got approved. One-point-five billion dollars.
Ivy: And of that, the authors get roughly a hundred and twenty-two million. Do the math on the rest.
Marcus: Wait — where does the other one-point-four billion go?
Ivy: Legal fees, administration, the machinery of a class action. The headline number and the writer's check are very different animals.
Marcus: Still — biggest copyright settlement of the AI era. That's a signal. Train on people's books, you might have to pay.
Ivy: It's a one-time check, Marcus. The book already got turned into fuel. The model still knows it. The settlement doesn't un-train anything.
Marcus: [beat] Yeah. That's the part that gets me. The author gets a payout and the machine keeps the knowledge forever.
Ivy: That's the broken bargain. Writers wanted a relationship — royalties, consent, an ongoing stake. They got a settlement and a goodbye.
Marcus: But doesn't this set precedent? Next lab thinks twice before scraping a pirate library.
Ivy: Or next lab just budgets for it. A hundred and twenty-two million is a rounding error to a company raising billions. It becomes the cost of doing business.
Marcus: [sighs] That's bleak.
Ivy: That's litigation. What it means for you: if you write, this isn't a windfall — it's a warning that your work has a price someone else sets.
Marcus: And if you build, license your data early, because retrofitting consent costs a billion and a half.
Ivy: [dry] Progress: now they pay after they take it instead of before. Truly a golden age.
Marcus: Last one — AMD launched Helios! Its first full rack-scale AI system, MI455X chips paired with the Epyc Venice CPUs.
Ivy: A rack. Nvidia's been shipping rack-scale for a while. AMD showing up to that fight matters more than the spec sheet.
Marcus: Right — and the real headline is the customer. Microsoft's deploying Helios at scale on Azure.
Ivy: Now that got my attention. Meta, OpenAI, and Oracle were already on board. Microsoft joining is the credibility stamp.
Marcus: [excited] That's the whole point! You don't dethrone Nvidia with silicon alone — you need hyperscalers actually buying racks.
Ivy: Buying, or hedging? Microsoft would love a second supplier just to stop paying Nvidia's margin. "At scale" can mean a lot of things.
Marcus: Even as a hedge, it's leverage. And leverage on Nvidia is worth billions to Azure.
Ivy: That I believe. Competition here is good — for the buyers. Whether it drops your inference bill is another question.
Marcus: That's the "for you" beat, though — more rack competition means cheaper compute, which means cheaper tokens downstream.
Ivy: Eventually. The savings pool at the top before they trickle to your API bill. Don't hold your breath this quarter.
Marcus: My hot take: this is the first AMD launch that made me think Nvidia should actually be nervous.
Ivy: [dry] Nvidia is up how many trillion in market cap? They're nervous the way a lion is nervous about a slightly larger housecat.
Marcus: [laughs] The housecat just signed Microsoft, Ivy.
Ivy: Touché. It's a two-horse race now instead of a one-horse coronation. That's genuine news.
Ivy: So — Kimi K3: a trillion-parameter open-weight model out of China, big if the benchmarks survive contact with reality.
Marcus: Anthropic pays one-point-five billion, authors see a hundred-twenty-two million, and the machine keeps the books anyway.
Ivy: And AMD's Helios rack lands Microsoft, turning the datacenter war into an actual fight.
Marcus: Three stories, one theme — everybody's racing to be the cheapest way to run intelligence.
Marcus: Before we go — one fun one. Somebody on r/LocalLLaMA already spun up a countdown clock to the Kimi K3 weight drop. Down to the second.
Ivy: [dry] Of course they did. Nothing says "open source community" like a doomsday timer for a file download.
Marcus: [laughs] And half the replies are people asking if their four-GPU rig can even load it.
Ivy: The answer is no. Tell them Ivy said no.
Marcus: [laughs] That's the show! Back tomorrow — and hey, maybe by then those Kimi benchmarks will be real.
Ivy: [dry] Or the housecat will have signed Amazon too. Either way — see you tomorrow.
This show is made with AI: the hosts’ voices are synthetic and the scripts are AI-assisted. Every story links to its original source.