Project Glasswing: Why Anthropic Gated Mythos for Critical Software Defenders
Glasswing is Anthropic’s trusted path for Mythos-class models on critical software defense — partners, credits, and why Mythos stays gated while Fable stays public.
81 results for “Claude”
Glasswing is Anthropic’s trusted path for Mythos-class models on critical software defense — partners, credits, and why Mythos stays gated while Fable stays public.
June 12, 2026: U.S. export controls forced Anthropic to suspend Fable 5 and Mythos 5; access restored July 1 after controls lifted June 30. Confirmed from Anthropic’s redeploy post — not a footnote.
NVIDIA's AVO agent architecture achieved a perfect 100% score on the ARC-AGI-3 public benchmark, completing all 183 levels across 25 environments. The result shows system-level design — not model scale alone — can unlock long-horizon autonomous reasoning that stumped every frontier model.
NVIDIA released NeMo Switchyard, an open-source library that routes AI agent steps across models to balance cost, latency, and accuracy. LangChain benchmarks show 74% cost savings with a 6-point accuracy tradeoff when routing between a 30B model and Claude Opus 4.8.
Anthropic announced Enterprise Frontier Safeguards, allowing regulated enterprises to store 30-day safety monitoring logs in their own AWS, Azure, or GCP accounts under customer-managed encryption keys with zero vendor human review.
Princeton researchers shadow-evaluated an AI scientist on two NeurIPS directions; original authors scored the outputs 2/6 and 1/6. Engineering loops worked, research judgment did not — while Anthropic's own data shows steep gains in code execution, not goal selection.
Related angle: OpenAI and Anthropic’s summer 2026 release tempo — the competitive pressure behind DeepMind’s shipping-speed remake.
Series end: two axes — who gets Mythos vs Fable (access), and Opus vs Fable pricing (tier). Confirmed architecture vs analysis, and what is still unknown.
Related angle: why coding/agentic execution became the urgency gap around Kavukcuoglu’s SVP mandate — confirmed vs report-based “Code Strike” claims.
A research-stage arXiv paper tests whether the factors LLMs cite for their decisions match measured behavioural influence. Across eight Claude, GPT, and Gemini models, cited top-three factors often missed the strongest drivers.
The Pentagon added OpenAI's ChatGPT Mil and xAI's Grok for Government to its GenAI.mil portal on August 31, giving up to 3 million personnel secure access to multiple frontier models. The move expands model choice for unclassified work while Anthropic's Claude remains absent amid a supply-chain risk dispute.
Z.ai released GLM-5.2, a 753B open-weight model with MIT licensing that matches Anthropic's Opus 4.8 on agentic benchmarks at roughly one-fifth the token cost. The 1M-context model is already running in Cursor and Claude Code, shifting enterprise leverage toward self-hosted inference.