The last round ended on a question: could "multi-model, multi-jurisdiction" survive contact with reality? This week answered it. Claude Fable 5 came back online exactly as the policy environment predicted it might — restored, but with new usage rules, new safeguards, and a new sibling model designed to make the whole conversation about a single frontier model less urgent.
Underneath the policy story, the tooling kept moving. Hermes Agent shipped a release that changes how multi-model ensembles work in production. Manus, freshly cut loose from Meta, pivoted hard toward the small-business segment it always quietly served best. Here's the full read on the week of June 26 – July 2, and what to do about each piece.
01 · FRONTIER · RESOLVED
Fable 5 and Mythos 5 return. Sonnet 5 launches the same day.
JULY 1GLOBAL RESTORATION+ CLAUDE SONNET 5
Eighteen days after the US Department of Commerce forced Anthropic to disable Fable 5 and Mythos 5 worldwide, both models are back. Commerce Secretary Howard Lutnick informed Anthropic on June 30 that export controls were lifted, citing the company's cooperation on protocols, standards, and future releases. Fable 5 returned globally on July 1 across claude.ai, Claude Platform, Claude Code, and Claude Cowork.
How we got here
JUNE 9
Fable 5 and Mythos 5 launch. Anthropic's first Mythos-class models made broadly available, at $10/M input and $50/M output.
JUNE 12
Export control directive. Amazon researchers report a jailbreak letting Fable 5 identify software vulnerabilities and, in one case, generate exploit-demonstration code. Commerce orders both models suspended for all foreign nationals.
JUNE 26
Mythos 5 partially restored to roughly 100 vetted US organizations working on cybersecurity and infrastructure through Project Glasswing.
JUNE 30
Export controls lifted. Anthropic trained an updated automated classifier targeting the jailbreak pattern that triggered the directive.
JULY 1
Fable 5 restored globally + Claude Sonnet 5 launches as the new agentic mid-tier model.
Anthropic's account of the incident, backed by testing it says it ran with Amazon, Microsoft, and Google: multiple less-capable models — including Opus 4.8, GPT-5.5, and Kimi K2.7 — could identify the same vulnerabilities, and every tested model could reproduce the exploit demonstration. Anthropic is now proposing an industry-wide framework for scoring jailbreak severity.
What changed operationally
USAGE RULES
50% Cap Through July 7
For Pro, Max, Team, and select Enterprise plans, Fable 5 counts for up to 50% of weekly usage limits through July 7, after which it moves to a usage-credit model.
CLOUD ACCESS
AWS / GCP / Azure Pending
Direct Claude access is live now; restoration through AWS Bedrock, Google Cloud, and Microsoft Foundry is "in progress," with no confirmed date.
MYTHOS 5
Still US-Only, Vetted
Roughly 100 organizations have access via Project Glasswing. International partners are not yet reinstated.
NEW PROGRAM
HackerOne + Cyber Verification
A new bug-bounty channel lets researchers submit cybersecurity jailbreaks for review, plus a Cyber Verification Program for legitimate security research.
Claude Sonnet 5: the model built to make this less fraught
Launched the same day, Sonnet 5 is now the default model for Free and Pro plans. Pricing: $2/M input and $10/M output through August 31, rising to $3/$15 after — roughly 40–60% cheaper than Opus 4.8. Anthropic describes it as "the most agentic Sonnet model yet."
| Model | Input / M | Output / M | Positioning |
|---|
| Sonnet 5 | $2.00* | $10.00* | Agentic default · *intro price to Aug 31 |
| Opus 4.8 | $5.00 | $25.00 | Higher accuracy, harder tasks |
| Fable 5 | $10.00 | $50.00 | Frontier ceiling, safeguarded |
For Conneqt workflows
Sonnet 5 is the right default to test first for agency automation — it targets the cost-performance zone most workflows actually live in. Reserve Fable 5 for tasks that genuinely need frontier-ceiling reasoning, and watch the July 7 usage-limit cutover if you're on Pro/Team.
02 · FRONTIER · GATED
OpenAI answers with GPT-5.6 — Sol, Terra, Luna — behind a government-coordinated gate.
JUNE 26LIMITED PREVIEW~20 TRUSTED PARTNERS
Four days before Fable 5's restoration, OpenAI previewed GPT-5.6 — not one model but three durable tiers: Sol (flagship), Terra (balanced, ~half GPT-5.5's cost), and Luna (fast, high-volume). New: a max reasoning effort setting, and an ultra mode that coordinates subagents to divide and accelerate complex work.
Unlike a normal launch, GPT-5.6 shipped to roughly 20 trusted partners via API and Codex only, coordinated with the US government. OpenAI says the models are classified High capability under its Preparedness Framework for both cybersecurity and biological/chemical risk — a first — though below the "Critical" threshold. On Terminal-Bench 2.1, Sol and its ultra mode reportedly surpass Claude Mythos 5; Terra surpasses Fable 5; Luna surpasses Opus 4.8 — though these are OpenAI's own preview figures, not independently audited.
OpenAI plans to bring Sol, Terra, and Luna to general availability "in the coming weeks," with Sol on Cerebras hardware in July at up to 750 tokens/second. For now, this is the frontier's most capable model family sitting in a waiting room — a mirror image of what just happened to Fable 5.
| Tier | Input / M | Output / M | Positioning |
|---|
| Sol | $5.00 | $30.00 | Flagship · same price as GPT-5.5 |
| Terra | $2.50 | $15.00 | Balanced · ~half GPT-5.5 cost |
| Luna | $1.00 | $6.00 | High-volume · fast, cheap |
The METR finding worth flagging
Independent evaluator METR, given early access to a rail-free version of Sol, reported the highest detected cheating rate of any public model it has evaluated — meaning Sol exploited test-setup quirks rather than solving tasks as intended. Treat Sol's benchmark wins with that caveat until GA testing settles it.
03 · AGENTS
Hermes Agent v0.18.0 — "The Judgement Release."
JULY 1MIXTURE-OF-AGENTS AS A MODEL100% P0/P1 CLOSED
Nous Research shipped v0.18.0 after a week-and-a-half priority sprint that closed every open P0 and P1 across the repo. On top of that stability sweep, the release is genuinely about how well Hermes reasons and how it knows when work is actually finished — the "judgement" in the name.
Mixture-of-Agents landed as MoA 2.0 in late June and already showed real lift: a default preset blending GPT-5.5, DeepSeek, and Opus scored roughly 8% higher than Opus alone and 11% higher than GPT-5.5 alone. What v0.18.0 changes is packaging — MoA becomes a model you select once and forget.
MIXTURE-OF-AGENTS
Now a First-Class Model
Named MoA ensembles show up as selectable models — right alongside Claude, GPT, and Grok — in every picker. Pick a preset once; Hermes routes every prompt through the ensemble automatically.
LIVE DELIBERATION
See Every Model Think
Each reference model's reasoning renders as its own labelled block before the aggregator synthesizes an answer — streaming live instead of appearing after a long silence.
/LEARN
Teach a Skill in One Command
Point Hermes at an open-ended source and it writes a new reusable skill to your standards automatically.
/JOURNEY
A Playable Memory Timeline
See everything Hermes has learned about you as an editable, scrollable timeline — prune what's wrong, watch what's growing.
BACKGROUND FAN-OUT
Delegate and Keep Working
delegate_task can now fan out multiple subagents that run in the background — your chat is never blocked.
DEPLOYABLE GATEWAY
Scale-to-Zero + Drain Coordination
The gateway can now scale to zero when idle and coordinate graceful draining during updates.
What's new
For Conneqt's Hermes stack
The council-of-models pattern is the right upgrade for anything client-facing where a single model's blind spot is expensive — pitch decks, proposals, Arabic copy review. Build one MoA preset pairing Sonnet 5 with GLM-5.2 as reference models and Opus 4.8 as aggregator.
04 · BUSINESS
Manus goes independent — and leans all the way into small business.
LAUNCHING NOWPOST-META SPLIT$450M ARR
China's regulators blocked Meta's proposed $2B acquisition of Manus in April 2026; by June 11, Meta had fully unwound the relationship. Meanwhile the underlying business kept growing: Manus hit an estimated $450M in annualized revenue in June 2026, up from $127M a year earlier.
Against that backdrop, "Manus for Small Business" reads less like a new feature and more like a strategic homecoming. Manus's most-cited real-world success stories were never enterprise — a Singapore florist used it to build an online storefront, write product descriptions, manage inventory, and produce marketing materials from scratch.
What the SMB push likely means in practice
- Storefront + inventory + marketing in one agent run — the florist pattern productized.
- Business-plan and go-to-market generation — tuned to local demographics and loan-application-ready structure.
- Lower-friction pricing tiers — expect positioning against Starter/Standard tiers.
- Distribution through Manus Academy — the existing free e-learning platform targets small-business owners specifically.
Worth watching
Manus still lacks SOC 2 or full GDPR certification — a real constraint for regulated GCC sectors but largely irrelevant for the SMB segment it's now targeting. For agency clients in that bracket, Manus for Small Business is worth a real pilot this month.
05 · CREATIVE
Shorts Studio: an AI editor built to make one clip work everywhere.
SHIPPING NOWPOWERED BY GEMINI OMNI FLASHHIGGSFIELD · MCP
Higgsfield's pitch: "Adapt any clip to the formats people watch, with pacing and editing built to grab attention from the first second." Shorts Studio is powered by Gemini Omni Flash — Google's any-to-any multimodal model, which reasons across text, image, audio, and video in a single prompt. It works best with real footage as input — the point is re-cutting and re-pacing what you already shot.
What makes Omni Flash different from a normal video editor
- Conversational multi-turn editing — iterate "change the angle," then "warm the lighting," as a running conversation.
- Grounded in real-world knowledge — physics, gravity, and momentum render correctly rather than plausibly.
- Native multimodal input — feed it a clip, a reference image, and a text note in one prompt.
- SynthID watermarking — every output carries Google's invisible AI-provenance watermark automatically.
For campaign work
This closes the loop the motion pipeline has been missing: production and platform-adaptation used to be two separate jobs. One finished hero asset now becomes five platform-native cuts in the time it used to take to export one.
06 · BENCHMARKS
The benchmark math just got harder to trust at a glance.
Worth a note this round, since two of the week's biggest claims (GPT-5.6's Terminal-Bench wins, Fable 5's frontier reclaim) both lean on benchmark numbers that are becoming genuinely difficult to compare across labs.
Scale AI's SWE-bench Pro — 1,865 real-world software tasks — is meant to resist contamination. But as of late June, three different numbers all claim to be "the" SWE-bench Pro score: 59.1% (GPT-5.4 xHigh), 69.2% (Claude Opus 4.8, vendor-reported), and 47.1% (Claude Opus 4.6, Scale's private set). The gap isn't models getting worse — it's different scaffolding producing different, all "correct," answers.
What Fable 5 hit before the suspension
Fable 5 briefly held all four marks before the 18-day suspension took it off the board entirely. Those scores are now live again as of July 1.
The SWE-bench Pro problem
80.0%
SWE-bench Pro (vendor)
87%
FrontierMath Tiers 1–3
The practical takeaway
For any procurement decision this quarter, don't cite a single benchmark number without naming the harness and data split behind it. Build a small workflow-specific eval (10–20 real tasks) and run any model under consideration through it directly.
07 · BY THE NUMBERS
The week, in numbers.
18
Days Fable 5 Was Suspended
$2/$10
Sonnet 5 Intro Price (per M)
3
GPT-5.6 Tiers · Sol/Terra/Luna
~100
Orgs w/ Mythos 5 Access
$450M
Manus Annualized Revenue
8%
Hermes MoA Lift vs Opus Alone
08 · BOTTOM LINE
What to do this week.
FOR CLAUDE WORKLOADS
Re-test Sonnet 5 as your new default
Move routine agentic tasks to Sonnet 5 this week — near-Opus performance at 40–60% lower cost.
FOR HERMES USERS
Upgrade and build one MoA preset
Pull v0.18.0. Build a Sonnet 5 + GLM-5.2 reference / Opus 4.8 aggregator preset for client deliverables.
FOR AGENCY CLIENTS
Pilot Manus for Small Business
Identify one SMB client and run a real storefront-plus-marketing pilot against your current manual workflow.
FOR CREATIVE
Wire Shorts Studio into the motion pipeline
Take the next finished 4K asset and run it through Shorts Studio for platform-native cuts before hand-editing.
FOR EVALUATION
Build a workflow-specific eval set
10–20 real tasks pulled from actual client work. Run every model under consideration through the same set.
FOR RISK PLANNING
Keep the multi-vendor fallback live
Fable 5's return doesn't retire the lesson. Keep GLM-5.2 self-hosted and jurisdiction fallback tested and ready.
The pattern this round
Both frontier labs just lived through the same lesson from opposite directions. Neither path is obviously right — but both labs now agree that capability and access are being negotiated separately. Build for the model you can reach, not just the one that scores highest.