日本語
2026-08-13 Evening edition
Evening edition — Research Report

AI News Daily 2026-08-13

Date
2026-08-13
Edition
Evening edition
Audience
Executives, decision makers and business leads
Format
Detailed research report
Executive Summary
  1. Google DeepMind has handed day-to-day leadership of Gemini development to CTO-turned-SVP Koray Kavukcuoglu, while co-founder Demis Hassabis steps back into a Chairman and Alphabet Chief Scientist role — a reorganization reportedly tied to slower model progress and talent attrition.
  2. NVIDIA has released a free, open-weight 30-billion-parameter (3-billion-active) MoE model, Nemotron 3.5 Lightning, alongside an open-source agent-routing library, NeMo Switchyard, claiming roughly 4x faster output than comparison models.
  3. Anthropic's new flagship Claude Opus 5 brings a 1-million-token context window and switchable low/medium/high reasoning effort, priced at $5 per million input tokens and $25 per million output tokens.
  4. OpenAI has lifted the text-chat cap for free ChatGPT users and made GPT-5.6 Luna the default model on Free and Go plans, as the service is said to have surpassed 1 billion weekly users.

01Google DeepMind: Koray Kavukcuoglu takes over day-to-day operations as SVP; Demis Hassabis moves to Chairman and Alphabet Chief Scientist

Published: 2026-08-12

Facts

Google DeepMind's Chief Technology Officer, Koray Kavukcuoglu, has taken on the title of Senior Vice President and assumed responsibility for day-to-day operations, including Gemini model development and productization. Co-founder Demis Hassabis has moved into the role of Chairman of Google DeepMind and Chief Scientist at Alphabet.

Background

The reshuffle is reported to be part of a broader reorganization at Google DeepMind, amid concerns about the pace of frontier model development and departures of key personnel.

Implications

With operational leadership of frontier AI development now in new hands, the change offers an early signal of how Gemini's development cadence and commercialization strategy may evolve, and is worth watching as a marker in the competitive dynamics against rival labs.

Sources: CNBC — report on Kavukcuoglu's appointment as SVP / Bloomberg — report on Hassabis's move to the Chairman role

02NVIDIA releases open-weight model "Nemotron 3.5 Lightning" and routing tool "NeMo Switchyard"

Published: 2026-08-11

Facts

NVIDIA has released Nemotron 3.5 Lightning, a 30-billion-parameter mixture-of-experts model with 3 billion active parameters, together with NeMo Switchyard, an open-source library that routes agent workloads to the model best suited to each task. Both are free for commercial use and available through Hugging Face and other channels; NVIDIA claims output speeds roughly four times faster than comparison models.

Background

The release continues NVIDIA's pattern of publishing lightweight, inference-optimized open models, positioning the company not only as a hardware supplier but as a source of software tooling that helps enterprises deploy agentic AI more cheaply.

Implications

A free, lightweight, high-speed inference model lowers the barrier for enterprises looking to deploy agentic AI systems in-house, and intensifies competition around reducing the cost of running AI workloads.

Sources: CNBC — report on the Nemotron 3.5 Lightning release / VentureBeat — explainer on Switchyard's cost-reduction effect

03Anthropic unveils flagship model "Claude Opus 5"

Published: 2026-07-24 (retrospective item)

Facts

Anthropic has announced its new flagship model, Claude Opus 5. It ships with a 1-million-token context window and a switchable reasoning-effort setting — low, medium, or high — that lets users trade off cost against depth of reasoning. Pricing is set at $5 per million input tokens and $25 per million output tokens, making it the fourth model released in the Claude 5 line.

Background

This item falls outside the usual 48-hour news window but is included as a retrospective entry because of its significance within the current reporting period (since April 2026).

Implications

The addition of a finely adjustable cost-versus-performance option gives enterprises more room to match model choice to specific workloads, underscoring the growing importance of deliberate model-selection strategy in production use cases.

Source: Anthropic official — "Introducing Claude Opus 5"

04OpenAI removes text-chat limits for free ChatGPT users and sets GPT-5.6 Luna as the default model

Published: 2026-08-06 (retrospective item)

Facts

OpenAI has removed the text-chat usage cap for free ChatGPT users and begun rolling out its new model, GPT-5.6 Luna, as the default model for Free and Go plan users. GPT-5.6 Sol is being rolled out progressively to paid tiers such as Pro. ChatGPT is reported to have surpassed 1 billion weekly users.

Background

Like the Claude Opus 5 entry above, this item is included as a retrospective addition given its significance within the current reporting period, even though it falls outside the standard 48-hour window.

Implications

Loosening restrictions on free-tier usage is likely to broaden general-public adoption of AI further, and could influence enterprise decisions about free-tier deployment in areas such as customer support and education.

Sources: OpenAI Help Center — "GPT-5.6 in ChatGPT" / TechCrunch — report on unlimited free-tier text chat

05Editor's Note

Today's items, taken together, sketch out an industry that is shifting from a pure model-capability race toward organizational and cost discipline. At Google DeepMind, operational leadership has moved from co-founder Demis Hassabis to CTO-turned-SVP Koray Kavukcuoglu, a change reportedly driven by concerns about development pace and talent retention — a reminder that even frontier labs are not immune to execution pressure. NVIDIA's free release of a lightweight, fast Nemotron model and its Switchyard routing library points in a complementary direction: as more organizations look to deploy agentic AI, the competitive battleground is increasingly about running models cheaply and efficiently, not just building the largest ones. Anthropic's Claude Opus 5 and OpenAI's GPT-5.6 rollout, both retrospective entries from earlier in this reporting period, reflect the same undercurrent — both companies are segmenting their offerings by cost and capability tier (adjustable reasoning effort at Anthropic, a free unlimited tier alongside paid tiers at OpenAI) rather than simply pushing a single frontier model forward. For decision makers, the throughline is that model selection and deployment-cost strategy are becoming as consequential as raw capability.