Если вы в 2026 году поднимаете agent stack на LLM — Claude Code, Cursor Agent, internal copilot — эта неделя бьёт сразу по двум осям: Claude Opus 5 (релиз 24.07.2026) даёт ~Fable-5 quality при ~50% token cost ($5/$25 за миллион); параллельно Kimi K3 попал в distillation gate — White House обвиняет Moonshot в «industrial-scale distillation», а Ryan Greenblatt показывает, что K3 называет себя Claude и сплёвывает internal deploy strings вроде claude-opus-4-5-20250929. Ниже — Opus-vs-Fable матрица, K3 timeline, разбор identity leak на уровне training metadata и 6-шаговый playbook для production.
01 Три bottleneck: OPEX frontier, provenance open weights, retention policy
Opus 5 и K3 gate — разные сюжеты, один вектор: сколько вы платите за intelligence и можно ли доказать её происхождение. На практике инженеры упираются в три слоя:
- Token OPEX на frontier: Fable 5 = 2× Opus pricing ($10/$50 per 1M). Agent с высоким output rate за месяц превращает разницу в заметную строку P&L при marginal CursorBench gain <0,5%.
- Provenance до open weights: K3 на старте — 2,8T MoE, GPQA-Diamond 93,5%, но без публичных весов (до 27.07) нельзя воспроизвести architecture audit. Self-reported benchmarks ≠ reproducible science.
- Data retention и routing: Fable 5 / Mythos 5 требуют 30-day retention; Opus 5 — default без mandatory retention. Для multi-model fallback смотрите OpenRouter integration guide — но routing не заменяет policy review.
Тезис: Opus 5 оптимизирует $/quality; K3 gate впервые ставит вопрос «дешёвый SOTA — engineering или borrowed weights?»
02 Claude Opus 5 vs Fable 5: specs и decision matrix
24.07.2026 Anthropic выкатил Claude Opus 5 как default в Claude Max + top tier для Claude Pro. Pricing как у Opus 4.8: $5 input / $25 output per 1M tokens. Context window 1M (single profile), max output 128K, Thinking on by default. Model ID: claude-opus-5; endpoints: Claude API, AWS Bedrock, Google Vertex AI, Microsoft Foundry.
| Dimension | Claude Opus 5 | Claude Fable 5 |
|---|---|---|
| Input / output ($/1M tokens) | $5 / $25 | ~$10 / $50 |
| CursorBench 3.2 (max effort) | <0,5% ниже Fable 5 peak | Coding frontier reference |
| Frontier-Bench v0.1 | >2× Opus 4.8, лидер | — |
| OSWorld 2.0 cost/perf | <⅓ cost Fable 5 при лучшем score | GUI agent benchmark |
| Data retention | Default без mandatory retention | 30-day retention required |
| Dual-use / cyber | Classifier ~85% мягче Fable 5; exploit gen blocked | Stricter filters |
| Typical pick | Daily agent / coding: best $/perf | Peak tasks или absolute max quality |
Published highlights: ARC-AGI 3 ~3× second place; Zapier AutomationBench ~1,5× pass rate; life-science internal +10,2 PP (spectrum→molecule), +7,7 PP (protein variants) vs Opus 4.8. Alignment audit: lowest deception rate в Opus line; dual-use ceiling остаётся у restricted Mythos 5. Для Claude Code deployment context — Claude Code на bare-metal Mac mini.
03 Kimi K3 distillation gate: politics, timeline, identity leak
Kimi K3 (Moonshot, 16.07.2026): 2,8T MoE (896 experts, 16 active ≈ 50B effective), 1M context, KDA architecture. GPQA-Diamond 93,5%, BrowseComp 91,2% — top open-source scores at launch. Full weights promised 27.07.2026. Deep dive: обзор Kimi K3.
| Date | Event |
|---|---|
| Feb 2026 | Anthropic accuses Moonshot / DeepSeek / MiniMax of industrial distillation — 3,4M+ anomalous API interactions |
| 01.07.2026 | Claude Fable 5 public GA |
| 16.07.2026 | Kimi K3 API / product launch |
| 22–23.07.2026 | White House (OSTP Michael Kratsios) public distillation accusation; GB300 mention |
| 23.07.2026 | TechCrunch: independent researchers skeptical of 2-week distillation timeline |
| ~24.07.2026 | Ryan Greenblatt (Redwood Research) publishes «Which Claude is K3?» stats |
| 27.07.2026 (planned) | Open weights — external verification |
Timeline critique почти unanimous: Fable 5 GA 01.07 → K3 launch 16.07 = 15 days. Braden Hancock (Snorkel) и Nathan Lambert (AI2) считают full industrial distillation + train + ship за это окно unrealistic. Lambert добавляет: gap сужается через RL infra, не только через cheap imitation.
Technically interesting — Greenblatt identity probe: на «кто ты?» K3 анomalously часто отвечает «I'm Claude» и выдаёт internal deploy IDs (claude-opus-4-5-20250929, claude-sonnet-4-5-20250929) — точнее, чем настоящие Claude models. Это smell of training data с API deployment metadata (logs, labeled synthetic), не style cloning. K2 указывал на более ранние Claude gens; K3 locked на Claude 4.5 gen (late 2025). Greenblatt: not legal proof, но strongest technical signal so far.
04 6 шагов: model pick, provenance gate, bare-metal compute
- Classify workloads: coding agent, doc RAG, ops automation — tag sensitive data и retention requirements (Opus 5 vs Fable 5 delta measurable).
- Run Opus-vs-Fable matrix: CursorBench, Frontier-Bench, domain tests. Gap <1% + cost ÷2 → Opus 5 as default.
- Treat K3 as unverified until 27.07.: benchmarks = vendor self-report; no compliance-critical path без weight audit plan.
- Track distillation signals: Greenblatt repo, Moonshot statements; vendor audits — provenance clause в contract.
- Multi-model routing: Opus 5 primary, Fable 5 или GPT-5.6 peak fallback — см. OpenRouter guide.
- Production compute layer: Claude Code / Cursor Agent нужны stable Metal pipeline и 7×24 uptime. Virtualized Mac = hypervisor tax + broken iOS CI chains; dedicated bare-metal Mac mini убирает этот слой из failure domain.
05 Hard data, FAQ, weekly takeaway
- Claude Opus 5 pricing: $5/$25 per 1M tokens; Fast mode ~2,5× speed, 2× price; ID
claude-opus-5. - Kimi K3 scale: 2,8T total, ~50B active; Terminal-Bench 2.1: 88,3%; SWE Marathon 42,0%.
- Distillation backdrop: Anthropic Feb 2026: 3,4M+ suspicious calls; White House Jul 2026 без public exhibits; Scott Bessent mentions «US model watermark» без spec.
FAQ
Насколько Claude Opus 5 дешевле Fable 5?
Opus 5: $5/M input, $25/M output — примерно вдвое дешевле Fable 5 (~$10/$50). CursorBench 3.2 max effort отстаёт от пика Fable 5 менее чем на 0,5%.
Claude Opus 5 — дефолтная модель Claude Max?
Да. С 24 июля 2026 Opus 5 — default в Claude Max и самая мощная публичная версия для Claude Pro.
Kimi K3 дистиллировал Claude?
По состоянию на июль 2026 обвинение не доказано и спорно. White House не опубликовал доказательств; timeline в две недели неправдоподобен. Статистика Ryan Greenblatt — K3 называет себя Claude и выдаёт internal deploy ID — сильнейший технический индикатор.
Когда выйдут полные веса Kimi K3?
Moonshot обещает open weights 27 июля 2026 — после этого можно независимо проверить архитектуру и бенчмарки.
Primary sources; при policy updates — official pages authoritative.
Anthropic: Claude Opus 5 announcement
TechCrunch: researchers skeptical of K3 distillation timeline
Ryan Greenblatt GitHub: Which Claude is K3?
2026 competition axis = $/frontier token + provenance auditability. Opus 5 cuts daily burn; K3 pushes open-source ceiling; distillation gate politicizes model lineage. Для команд с Claude Code, Cursor Agent или custom agents 24/7 API routing не решает Metal compile chain и iOS CI stability: virtualized Mac hosts добавляют hypervisor overhead и long-term fragility. Для production с native Apple Silicon, stable iOS CI/CD и agent automation 7×24 bare-metal Mac mini cloud nodes ZUKCLOUD — типично optimal base: dedicated physical machine, zero hypervisor loss, 7×24 online, flexible rental. Architecture deep dive: bare-metal architecture manifesto.