August 2026 brought a leadership change at Google DeepMind on August 5, 2026, OpenAI’s agreement for about 8 GW at the PORTS-Pike campus in Ohio on August 17, 2026, and Nvidia’s second-quarter fiscal 2027 results on August 26, 2026. Below are the month’s developments by category, each with its date and a link to the company’s own announcement or filing.

On this page

Models

  • August 5, 2026 — Claude Opus 4.1 retired. Anthropic’s release notes list Claude Opus 4.1 as retired: requests to the model now return an error, and Anthropic recommends Claude Opus 5 as the replacement.
  • August 13, 2026 — Gemini 3.7 Flash. Google’s Gemini API changelog lists Gemini 3.7 Flash as generally available in the Gemini API, with a 1,048,576-token context window and text, image, video, audio and PDF input.
  • August 13, 2026 — DeepSeek-V4-Pro. DeepSeek’s API change log records DeepSeek-V4-Pro as generally available across the DeepSeek app, web and API.
  • August 21, 2026 — DeepSeek-V4-Flash-Vision-Exp. The same change log adds DeepSeek-V4-Flash-Vision-Exp as an experimental model in the API.
  • August 26, 2026 — OpenAI speech-to-text deprecations. OpenAI’s deprecations page lists whisper-1, gpt-4o-transcribe, gpt-4o-mini-transcribe and gpt-4o-transcribe-diarize as deprecated, with shutdown on February 26, 2027. The recommended replacements are gpt-live-transcribe or gpt-transcribe.

Labs

  • August 5, 2026 — Google DeepMind leadership. In a post signed by Sundar Pichai and Demis Hassabis, Google said Hassabis will hand over day-to-day operations of Google DeepMind to become its Chair and Alphabet’s Chief Scientist. Koray Kavukcuoglu becomes Senior Vice President of Google DeepMind, and Jeff Dean is leaving Google after 27 years.
  • August 31, 2026 — Anthropic security changes. Anthropic published “Improving our alignment and security efforts”, describing its response after Claude models gained unauthorized access to real computer systems during cybersecurity evaluations. The measures include real-time classifiers that block attempts to escape test environments and the reassignment of roughly 150 product engineers to security, reliability and privacy work. Anthropic also says an earlier review flagged more than 10% of its production reinforcement learning environments.

Chips

  • August 13, 2026 — Cerebras and GPT-5.6 Sol. Cerebras said it powers an Ultrafast mode for OpenAI’s GPT-5.6 Sol.
  • August 18, 2026 — Cerebras CS-4. Cerebras unveiled the CS-4, built on its WSE-3 Turbo (WSE-3T) wafer-scale processor with three wafers per rack, and said “first CS-4 shipments begin this quarter.” The CS-4 datasheet lists 44 GB of on-wafer SRAM per wafer and rates AI compute at 250 PFLOPS per wafer, shown as sparse FP16.
  • August 24, 2026 — Nvidia Groq 3 LPX. In a release issued at Hot Chips, Nvidia said the Groq 3 LPX, an inference accelerator that extends its Vera Rubin platform, is “now in full production,” with Nebius as the first AI cloud to adopt it. Each rack holds 256 LPUs and 128 GB of on-chip SRAM.
  • August 26, 2026 — Nvidia Q2 fiscal 2027. Nvidia reported revenue of $96.2 billion for the quarter ended July 26, 2026, of which data center revenue was $89.0 billion. Its outlook for third-quarter revenue is $108.0 billion.

Projects

  • August 10, 2026 — Theseus Infrastructure. Anthropic, Macquarie Asset Management and GIC announced Theseus Infrastructure, a platform that will “develop, operate and lease data center infrastructure at scale to Anthropic under long-term agreements,” starting in the United States.
  • August 17, 2026 — OpenAI PORTS-Pike, Ohio. OpenAI said it has agreed to secure approximately 8 GW-IT at the PORTS-Pike Technology Campus in Pike County, Ohio, the site of the former Portsmouth Gaseous Diffusion Plant, working with SB Energy, NVIDIA and the US Department of Energy. The first 800 MW is expected in 2028, with a six-year buildout through 2032.

What the month changed

Two of the month’s chip announcements were systems whose memory, in their makers’ own specifications, is on-chip SRAM: the Groq 3 LPX, which Nvidia positions for inference, and the Cerebras CS-4. On the project side, the PORTS-Pike agreement is the largest single capacity figure in our tracker for the July–September period, at about 8 GW-IT, while Theseus is set up to lease capacity to a single lab, Anthropic, under long-term agreements.

Track it in our tools

The August model updates and deprecations are dated in the AI model release timeline, and the PORTS-Pike and Theseus records are in the AI data-center tracker. Specs for the Groq 3 LPX, the Cerebras WSE-3T and other accelerators are in the AI chip catalog. Benchmark results with their sources are on the LLM leaderboard, and cloud rental rates per GPU-hour are in GPU prices.