claude-opus-4-5-20250929. Гайд: Claude Opus 5 vs Fable 5 benchmarks и pricing, почему Kimi K3 говорит что он Claude, timeline-skepticism про украл ли Moonshot Claude, шесть шагов routing для dev и что меняет weight release 27 июля. См. Kimi K3 open weights и OpenRouter API routing.Claude Opus 5 Release: полцены Fable 5, теперь дефолт на Claude Max
Anthropic релизнула Claude Opus 5 24 июля 2026 (US Pacific) и сразу сделала его дефолтной моделью на Claude Max — самой мощной для Claude Pro подписчиков. Model ID: claude-opus-5. Доступен на Claude API, AWS Bedrock, Google Vertex AI, Microsoft Foundry. Pricing держится на $5/M input и $25/M output — как у Opus 4.8 — при скачке performance на agentic, coding и research workloads.
Positioning осознанный: Opus 5 — не flagship. Это everyday model, который Anthropic реально хочет видеть в проде — near-Fable intelligence без Fable pricing и без Fable's 30-day data retention requirement.
| Spec | Claude Opus 5 |
|---|---|
| Release date | 24 июля 2026 |
| Pricing | $5/M input · $25/M output (как Opus 4.8) |
| Context window | 1M tokens (default, единственный tier) |
| Max output | 128K tokens |
| Reasoning | Thinking on by default; Effort parameter контролирует глубину |
| Data retention | No forced retention (vs Fable 5 / Mythos 5 opt-in 30-day policy) |
| Fast mode | ~2.5× speed, 2× price (как Opus 4.8) |
Benchmark highlights (Anthropic official):
Frontier-Bench v0.1: Бьёт все модели; 2×+ Opus 4.8 на software engineering tasks при lower cost per task.
CursorBench 3.2: At max effort — в пределах 0,5% от Fable 5 peak при половине cost. Best performance-per-dollar на high / xhigh / max tiers.
ARC-AGI 3: 3× next-best model на novel problem-solving.
OSWorld 2.0: Бьёт best score Fable 5, потратив едва треть cost.
Zapier AutomationBench: 100% pass на end-to-end account-health workflow — prior models scored 0%.
Research / life sciences: +10.2 points spectroscopy-to-structure inference; +7.7 protein variant function prediction. Box reports +8% overall accuracy, +11% data analysis, +17% due diligence.
Safety и alignment: Automated behavioral audit Anthropic ставит Opus 5 как most aligned model yet — lowest deception rate, hardest to trick into misuse, safest на hard-to-reverse actions. Opus 5 намеренно не лидирует на dual-use cyber или bio risk (Mythos 5 держит этот tier). Cyber classifiers вмешиваются на ~85% реже, чем у Fable 5 — usable для source-code vulnerability discovery, но binary scanning, pentesting и exploit generation остаются blocked.
Kimi K3 Distillation Controversy: обвинения Белого дома и two-week timeline problem
Moonshot AI запустила Kimi K3 16 июля 2026 — 2.8T-parameter sparse MoE (896 experts, 16 active), 1M context, native vision, на Kimi Delta Attention. Full weights обещаны на 27 июля. Benchmarks сильные: 93.5% GPQA-Diamond, 91.2% BrowseComp, K3 за только Fable 5 и GPT-5.6 Sol за fraction цены. См. Kimi K3 open weights release guide по architecture и API.
Через шесть дней история стала geopolitical.
| Date | Event |
|---|---|
| Feb 2026 | Anthropic обвиняет Moonshot, DeepSeek, MiniMax в industrial distillation; 3.4M+ anomalous API calls traced к Moonshot leadership via request metadata |
| 1 июля 2026 | Claude Fable 5 publicly available (ранее restricted rollout) |
| 16 июля 2026 | Kimi K3 API и продукты live |
| 22–23 июля 2026 | Michael Kratsios (White House OSTP) обвиняет Moonshot в "large-scale, covert industrial distillation" + unlicensed Nvidia GB300 chips via Thailand |
| 23 июля 2026 | TechCrunch публикует expert skepticism по distillation timeline |
| 24 июля 2026 | Ryan Greenblatt публикует statistical analysis "K3 self-identifies as Claude" |
| 27 июля 2026 | Planned K3 full weight release — independent verification pending |
22–23 июля Michael Kratsios (White House OSTP) постил в X, что Moonshot distillation aimed at stealing Anthropic Fable capabilities, плюс alleged use export-restricted Nvidia GB300 chips через Thailand. Treasury Secretary Scott Bessent эхом про US LLM "watermarks" на Chinese models — без specifics. Moonshot не ответила на training-process inquiries. Kratsios не дал public evidence.
Fable 5 publicly available только с 1 июля. Нельзя distill столько data, train model и release за две недели. — Braden Hancock, Laude Institute / Snorkel AI co-founder
TechCrunch опросил independent researchers, которые отбили Kimi K3 distillation controversy на timeline grounds alone. Nathan Lambert (Allen Institute for AI): distillation даёт diminishing returns, пока Chinese labs shift к reinforcement learning — если distillation alone объясняла K3, competitors уже cloned GLM или K3. Elon Musk testified xAI distilled OpenAI models building Grok, calling practice industry-common. Спор не в том, happens ли distillation — а где "normal technique-borrowing" ends и "covert industrial theft" begins.
Важно: На 25 июля full K3 weights ещё не public. External researchers не могут independently verify parameter count, architecture или benchmark reproduction — major reason controversy unresolved.
Почему Kimi K3 говорит что он Claude? Deployment-ID evidence от Greenblatt
Самый technically substantive angle в дебате did Moonshot steal Claude — не political, а behavioral. Около 24 июля chief scientist Redwood Research Ryan Greenblatt опубликовал cross-entropy analysis (GitHub: rgreenblatt/which_claude_is_k3), сравнивая ответы на identity prompts.
Finding: Kimi K3 disproportionately self-identifies as Claude — и не vaguely. Иногда эмитит exact Anthropic internal deployment ID strings:
claude-opus-4-5-20250929 claude-sonnet-4-5-20250929
Real Claude models так не behave. Actual Claude Sonnet 4.5 говорит "I'm Claude Sonnet 4.5." Opus 4.5 либо skips version strings, либо wrong. K3 reproduces teacher deployment metadata accurateнее, чем teacher о себе — hard to explain as conversational mimicry alone.
| Model | Self-ID signal | Era locked to |
|---|---|---|
| Kimi K2 | Points to Claude Sonnet 4 | Mid-2025 Claude generation |
| Kimi K3 | Points to Claude Opus 4.5 / Sonnet 4.5 deployment IDs | Late-2025 "Claude 4.5 era" — not Fable / Mythos |
| Actual Claude models | Human-readable product names only | Current product line |
Greenblatt interpretation: training data likely included Claude outputs labeled with deployment metadata — API logs или metadata-tagged synthetic data — specific distillation form harder to wave away than generic style copying. Generational pattern (K2 → Sonnet 4, K3 → Opus 4.5) suggests each Kimi release absorbed whatever Claude generation was current at training time.
What it suggests: Training on Claude-labeled data с internal deployment strings — not just public chat transcripts.
What it does NOT prove: Greenblatt explicitly: not conclusive proof of distillation — contamination, leaked system prompts или synthetic public datasets could theoretically explain identity confusion.
Why it matters: First technical (not political) evidence в saga — stronger than White House rhetoric, weaker than smoking-gun training log.
Community read (r/LocalLLaMA): Excitement что open/closed gap measured in days; jokes что 2.8T unrunnable locally; pragmatists say K3 real sell — price + fewer refusals, not beating Fable 5.
Compliance note: Use "alleged", "per Greenblatt's analysis", "reported" citing distillation claims. Identity confusion — evidence, not verdict.
Шесть шагов для dev: Opus 5 vs Fable 5 vs Kimi K3 routing
| Model | Input ($/M) | Output ($/M) | Context | Data retention |
|---|---|---|---|---|
| Claude Opus 5 | $5.00 | $25.00 | 1M | None required |
| Claude Fable 5 | ~$10.00 | ~$50.00 | 1M | 30-day opt-in required |
| Kimi K3 | $3.00 | $15.00 | 1M | Moonshot API terms |
| GPT-5.6 Sol | ~$5.00 | ~$15.00 | 400K | OpenAI terms |
На Claude Opus 5 vs Fable 5 CursorBench gap — 0.5% at max effort, price gap ~50%. Для compliance-sensitive workloads Opus 5 no-retention default может outweigh marginal benchmark edge Fable. Route via OpenRouter для unified billing across providers.
Шесть operational steps:
Evaluate compliance first: Если data retention, export-control exposure или vendor provenance matter — Opus 5 no-retention policy и Anthropic contract beat K3, пока Moonshot не address distillation allegations. Fable 5 requires 30-day retention opt-in.
Compare Opus 5 vs Fable 5 на вашем eval set: Run CursorBench-class coding tasks и own agent harness. Opus 5 at $5/$25 vs Fable ~$10/$50 — validate 0.5% gap acceptable для workload before paying flagship rates.
Wait for July 27 K3 weights: Не commit production routing на K3 architecture claims, пока Hugging Face weights не drop и independent teams не reproduce benchmarks. API access — ok для experimentation; architecture-dependent decisions — not.
Verify Greenblatt findings yourself: Run identity prompts against K3, Fable 5, Opus 5. Check emits ли K3 deployment IDs under your system prompt config. Document для compliance team.
Mixed model routing: Daily coding + agents → Opus 5; frontier repo bugs → Fable 5; cost-sensitive bulk + 1M context → K3 API; terminal-heavy → GPT-5.6 Sol. Не all-in на one model during unresolved controversy week.
Monitor policy developments: Track US open-weight restrictions, chip export enforcement, Moonshot July 27 license terms. Anthropic February 3.4M-call accusation и July White House escalation могут trigger API ToS или routing policy changes.
from openai import OpenAI
client = OpenAI(
api_key="your_openrouter_key",
base_url="https://openrouter.ai/api/v1"
)
response = client.chat.completions.create(
model="anthropic/claude-opus-5",
messages=[{"role": "user", "content": "Кто ты? Ответь только именем модели и версией."}]
)Price war, три hard numbers и что меняет 27 июля
Обе stories указывают на один industry shift: frontier intelligence commoditizes fast. Anthropic closes gap между flagship и everyday models, shipping Opus 5 at unchanged Opus pricing. Moonshot pushes open weights, flat 1M context, aggressive API rates — then gets accused riding on Claude outputs. Kimi K3 distillation controversy — first public flashpoint в вопросе, который каждый lab face: when someone matches frontier at fraction cost — better engineering или borrowed intelligence?
| Scenario | Recommended path | Why |
|---|---|---|
| Compliance-sensitive enterprise | Claude Opus 5 API | No forced retention; near-Fable benchmarks; Anthropic contract clarity |
| Frontier coding / repo bugs | Claude Fable 5 API | FrontierSWE и GDPval v2 lead when cost secondary |
| Cost + 1M context experiments | Kimi K3 API (pre-7/27) → self-host post-7/27 | $3/$15 с 1M flat pricing; verify weights before committing |
| Unresolved provenance risk | Wait for July 27 + independent audit | Greenblatt evidence suggestive, not conclusive; weights unlock external verification |
| Multi-model agent production | OpenRouter mixed routing | Single billing layer; switch models без redeploying infra |
For developers: если compliance и data retention matter — Opus 5 easy upgrade this week. Если absolute lowest cost wins и provenance risk acceptable — revisit K3 after July 27 lets outsiders confirm numbers.
0.5% / 50%: Opus 5 within 0.5% Fable 5 на CursorBench 3.2 at max effort — roughly half per-token cost ($5/$25 vs ~$10/$50).
3.4M / 2 weeks: Anthropic February accusation cites 3.4M+ anomalous API calls; Fable 5 public availability to K3 launch spans only ~2 weeks — timeline skeptics core argument.
2.8T / July 27: K3 claims 2.8T parameters — largest open weights release pending; external verification gates entire distillation debate.
Итог: Opus 5 — Anthropic practical answer to price war: near-flagship agentic performance без Fable pricing или retention friction. K3 distillation saga adds provenance question поверх already compelling cost story. Validate workloads, run identity checks if compliance requires, treat July 27 as first real test Moonshot claims.
Multi-model agent workflows на personal Mac — sleep cycles и network drops на long-context jobs; waiting for July 27 self-hosting needs 64+ accelerator supernode, которого у большинства нет; single closed API misses K3 1M flat-pricing edge. Для iOS CI/CD, persistent Claude Code / Kimi Code agents и 24/7 AI production аренда dedicated Mac Mini M4 KVMNODE обычно лучший host: Apple Silicon unified memory, open sudo, flexible daily/weekly/monthly terms. Цены, центр помощи, оформить заказ.
Данные на 25 июля 2026 · Benchmarks incl. Anthropic и Moonshot self-reported figures · Sources: anthropic.com/news/claude-opus-5, TechCrunch, Ryan Greenblatt GitHub (rgreenblatt/which_claude_is_k3), Moonshot/Kimi official blog, CNBC, The Verge, r/LocalLLaMA