Product Tag

Product

Posts related to product

59 posts

← Back to all posts

Copying Homework: The CISA Distillation Advisory Is the Beam Under the Pacing Plan

Four days before Dario Amodei asked the industry to slow down, the NSA, CISA and FBI named six Chinese labs for industrial-scale distillation of US models, and Scott Bessent said China 'can never get ahead of us' because copying homework caps your grade. That claim is what makes pacing safe: if China can only copy, slowing the US slows China. Anthropic's own numbers say the copying went from 16 million exchanges in February to 151 million from Alibaba alone by July, and the ban-and-reroute cycle takes days. The beam is real. It is also under load.

Read more →

The Rumour Was the Prompt: OpenAI's Navier-Stokes Proof, 10,000 Agents, and the Team It Scooped

OpenAI heard a rumour on September 1 that Anthropic's models had cracked a Millennium Prize problem. It launched 10,000 agents. Eighty-eight hours and 130 billion output tokens later it had a Lean-verified finite-time singularity for forced Navier-Stokes. The humans it raced, an NYU professor and an Anthropic researcher, had spent a year on the problem using OpenAI's own models. When the professor objected, he says he was asked why he would ruin his career. The result is real. The economics are the story.

Read more →

Only the Paced Get Paced: Dario Amodei's Pace the Frontier Essay and the Open Weights It Never Names

Four lab CEOs endorsed slowing the frontier inside one news cycle. Critics called it an attack on open source. The essay never mentions open weights once. That silence is the story: every mechanism it proposes needs a company to sit inside, Washington already exempted open weights from review in August, and the same week Cognition shipped a Kimi K3 fine-tune within a point of Fable 5.1. Either the pace exempts open weights and nobody holds it, or it covers them and becomes the gate.

Read more →

GPT-6 Astra Is On Every Plan: What It Costs, What It's Good At, and Which Effort Level to Use

OpenAI's GPT-6 Astra reached every paid ChatGPT plan, the API, Copilot and OpenRouter 27 hours after launch, at Fable 5.1's exact price. It sits two points behind Fable on the independent index at 42% of the cost per task, refuses exploit-writing by default, and hides its reasoning. I ran the same code review at low, high and max effort. Low found five real bugs in 60 seconds. Max found seven in seven minutes and got the ranking right.

Read more →

The Raise That Is a Cut: Claude Code Weekly Limits, Up 25% and Down 17%

On August 29 Anthropic announced it will permanently raise Claude Code weekly limits by 25% from September 14. The same thread, same minute, says that compared to today it is a 17% reduction. Both are true. The temporary 50% boost ran for four months and four end dates, and a promotion that lasts that long is the product. The interesting mistake is not Anthropic's. It is everybody who thought the baseline was 150.

Read more →

Two Meters: Claude Fable 5.1 Is Cheaper on the API and Hungrier on Max, and What to Turn Down

Anthropic says Fable 5.1 costs up to 45% less. Reddit says it empties a five-hour window in fifteen minutes. Both are true. The cache-read cut applies 'wherever usage is billed by token', and a subscription is not. Ten days of my own Claude Code transcripts show where the money goes, why the discount lands on one meter and the appetite on the other, and which dial to turn.

Read more →

The Cage Is in the Table Now: Claude Fable 5.1, Mythos 5.1, and the Five Points Safety Costs

Claude Fable 5.1 and Mythos 5.1 share identical weights. For the first time Anthropic put both in one benchmark table and footnoted the gap: 5.1 points on Terminal-Bench, lost to cyber safeguards on tasks with no cyber in them. The system card goes further. In its own attack evals, every successful break ran on the fallback model the safeguards hand you.

Read more →

Too Dangerous to Release: The OpenAI Astra Playbook

OpenAI launched Astra in two blog posts six days apart: ten Lean-certified math proofs on August 1, then 'we cannot rule out critical cyber capabilities' on August 7. The capability claim ships with machine-checkable proof. The danger claim ships with none, and none is possible from outside. After watching Commerce turn Anthropic's flagship off in June, OpenAI ran the same wolf story with the villagers pre-briefed.

Read more →

Paragraph Nine Was the Payload: Nvidia's Open-Weights Letter and Anthropic Alone

Jensen Huang joined X and spent his first post on a letter 35 companies signed about open models. The openness argument is the wrapper. Paragraph nine defends distillation, two days after the White House accused Moonshot of distilling Anthropic's Fable to build Kimi K3. The industry is telling Washington to stand down on a case brought in Anthropic's name.

Read more →

Grok 4.5 Trained on the Answer Key

xAI's launch page for Grok 4.5 is a wall of green bars led by a token-efficiency chart. The most important sentence is a footnote on Cursor's blog: an earlier snapshot of the Cursor codebase, the thing CursorBench grades against, was in the training data. The exam graded itself, and the answer key came stapled to it.

Read more →

The Coding Moat Was Never the Code

Anthropic studied 400,000 Claude Code sessions and found the best users weren't the best programmers. Managers, lawyers, and salespeople land within a few points of software engineers, and management scored highest of all. The skill that transfers isn't syntax. It's knowing what the right thing to build is, which is the one thing a bootcamp never taught.

Read more →

The Permission Tier: Claude Fable 5 Comes Back Changed

For 19 days the best model on earth was illegal to show a foreign national, including Anthropic's own staff. Then Fable 5 came back with a new classifier, a silent reroute to Opus 4.8, and no proof the weights were the same. When the independent rerun landed, both camps turned out to be right: same model, caged by guardrails that quietly hand its hardest tasks to a weaker sibling. Access used to be gated by price. Now it's gated by permission.

Read more →

The Expensive Middle: Claude Opus 4.8 vs Sonnet 5

Sonnet 5 lands within a few points of Opus 4.8 on most work and looks 2.5x cheaper, but that discount inverts on real tasks: at high effort Sonnet is so token-hungry it often bills more per task than Opus. The usage squeeze, meanwhile, is self-inflicted: agentic work now fans out dozens of subagents across parallel workstreams. Opus 4.8 became the expensive middle, though its real problem was never the price. It's the position.

Read more →

The Archetype Under the Title

Boris Cherny, who built Claude Code, says engineering, product, design and data science are melting into one role, and what's left is five archetypes: Prototyper, Builder, Sweeper, Grower, Maintainer. I read the list and realised I'm all five, because building solo with agents leaves no one to hand a phase to. The framework is thirty years old. What's new is that it just became the primary axis instead of the secondary one.

Read more →

Claude Doesn't Know It Isn't DeepSeek

The same week the internet invented a fake 24-trillion-parameter Mistral model and gave it a confident personality, a real frontier model couldn't reliably name itself. Ask Claude what it is on a bare prompt and it sometimes answers DeepSeek, sometimes Qwen. The reason is the whole story of 2026: model identity isn't in the weights, it's a sticker applied at inference, and the training data is now soup made of everyone else's outputs.

Read more →

It Wasn't in Your Head

Every Claude power user has felt it: the limits ratcheting down week after week while Anthropic insisted nothing had changed. On June 14 that feeling got a docket number. Kahn v. Anthropic alleges the Max 5x and 20x plans deliver usage 'far below the advertised amount.' The lawsuit may or may not win. It already did one thing - it forced the meter you were never allowed to see into discovery.

Read more →

AI Is Licensed Now

The Fable 5 ban was supposed to lift in weeks. Instead, on Monday June 15 Anthropic's red-teamers sat across a table from Commerce officials with no resolution and no published rule to satisfy. The export control didn't get walked back. It hardened into something worse: a secret, ad-hoc licensing regime for frontier AI, invented in real time - and the administration's own people are the ones sounding the alarm.

Read more →

One Went Dark, Two Went Open

In the same 72 hours the US export-controlled Fable 5 off the planet, China's open-weight labs shipped two major coding models into the commons: Kimi K2.7 on June 12, GLM-5.2 on June 13. One model went dark behind a national-security letter; two more went open under MIT. The diffusion layer didn't pause for America's panic. It shipped through it.

Read more →

The Call Came From Inside the Cap Table

The report that got Anthropic's Fable 5 export-controlled off the planet came from Amazon - Anthropic's single biggest investor. Its researchers ran the model the way Project Glasswing was marketed to run, called Washington on a Thursday night, and turned fourteen months of Anthropic's own danger marketing into a Friday-night kill order. The wolf was always fake. This week we learned who was holding the trigger.

Read more →

The Trophy and the Territory

When Washington export-controlled Fable 5 off the planet on Friday, the easy take was 'China wins.' That's the small version. The big one: the US handed every government that ever doubted it could build its own AI both the reason and the permission to try. Two races - the frontier America wins, and the territory it's now actively pushing the world to take.

Read more →

Too Dangerous to Keep

For fourteen months Anthropic told Washington its frontier models were national-security-grade dangerous. It was marketing - the moat behind the safety brand. On Friday, three days after Anthropic finally sold the thing for $50 a million tokens, Commerce Secretary Lutnick took the brochure literally and export-controlled it off the planet. The wolf was always fake. A villager finally believed it.

Read more →

It Was Always an IPO

Anthropic filed a confidential S-1 on June 1 at a $965B valuation, eclipsing OpenAI. Read backwards from the filing, the last two years stop looking like a safety lab's awkward compromises and start looking like a pre-IPO playbook executed on schedule.

Read more →

Cheap Is a Hardware Strategy

Google led I/O 2026 with a cheap, fast Gemini Flash instead of a frontier behemoth, and everyone read it as conceding the top of the market. Wrong read. Cheap isn't a model strategy, it's a silicon strategy. Google owns every layer from the TPU to the search box, which is why it can give intelligence away while its rivals rent the compute to compete with it, some of them for $40 billion.

Read more →

The Last Slow Thing

Everything in software got a fast mode this year except understanding what to build. The proof is in the labs' own org charts: the companies selling the models that supposedly end software engineering are paying $600k for engineers to go sit in customers' offices. The bottleneck moved all the way up to the conversation.

Read more →