Ground Truth.
AI, checked against the source.
← 2026-08-072026-08-08later →

Hassabis hands DeepMind to a non-CEO and takes an Alphabet job

2026-08-08

Demis Hassabis is giving up day-to-day control of Google DeepMind to become chair and Alphabet's chief scientist, with Koray Kavukcuoglu running the lab as a senior vice-president rather than as CEO.

deepmind · google · industry · leadership · alphabet

Claude Code stops asking permission on August 14

2026-08-08

Anthropic is making auto mode the default for new Claude Code sessions on Pro, Max, and Team plans from 14 August 2026, replacing per-action approval prompts with a separate classifier that blocks actions driven by hostile content the agent read.

cybersecurity · prompt-injection · ai-security · agents · supply-chain · anthropic

WeatherNext called Melissa's Category 5 landfall five days out

2026-08-08

Google DeepMind's WeatherNext predicted Hurricane Melissa's Category 5 landfall in Jamaica five days ahead at 80% confidence, and its five-day track forecasts averaged about 140 kilometres closer than the European ensemble -- roughly a day and a half of extra warning.

deepmind · weather · science · forecasting · google · ensembles

Sixteen AI-designed viruses worked, and one borrowed a part from a cousin

2026-08-08

Arc Institute researchers used a genome language model to design bacteriophages from scratch, synthesized the DNA, and got 16 working viruses out of 285 tested -- one of which swapped in a structural protein from a distantly related phage.

cybersecurity · ai-security · biosecurity · open-weights · science · arc-institute

A diffusion model picks its answer a fifth of the way through

2026-08-08

Researchers logged every token commitment in a masked diffusion language model and found it locks in the final answer 15 to 24 percent of the way through generation, while half the reasoning is still blank -- so the visible reasoning is written around a frozen conclusion.

diffusion-models · reasoning · interpretability · research · chain-of-thought

A video model counted events correctly two-tenths of one percent of the time

2026-08-08

Asked to count simple events in short synthetic clips, Google's Gemini 3.6 Flash got the final count right 0.2 percent of the time in the hardest setting and recovered only 18 percent of the events that actually occurred.

video · benchmarks · vision-language-models · research · evaluation

Stack Overflow took 1,490 questions in July

2026-08-08

Stack Overflow received 1,490 new questions in July 2026, down from 6,414 in July 2025 and 176,610 in July 2014 -- a 118-fold collapse in the public programming corpus that trained today's coding models.

stack-overflow · training-data · developers · industry · coding

China's biggest memory maker is booked through 2027

2026-08-08

ChangXin Memory Technologies has reportedly sold out its DRAM output through the end of 2027 as PC brands rushed to secure supply, and consumer memory prices have stayed near their highs since.

hardware · memory · supply-chain · china · local-ai

A 4B search agent matches 30B by grading its own failed attempts

2026-08-08

ABSeeker trains a 4-billion-parameter web-search agent on 8,500 examples by working backwards from the known answer to score each individual search step, letting useful steps inside failed runs earn credit -- and matches agents roughly seven times its size.

agents · reinforcement-learning · research · search · credit-assignment

Walmart put a token allowance on its in-house coding AI

2026-08-08

Walmart replaced unlimited access to its in-house AI coding tool with a fixed per-employee token allotment, and its CTO says the reason is duplicated requests rather than the bill.

enterprise · industry · coding · cost · adoption

The data firms behind frontier AI sell judgment, not labels

2026-08-08

Mercor, Surge AI and AfterQuery have all converged on the same product line -- reinforcement-learning environments, scoring rubrics, expert demonstrations and human evaluations -- turning graded professional judgment into a commodity input for frontier models.

data · industry · reinforcement-learning · labeling · evaluation

← 2026-08-072026-08-08later →