Ground Truth.
AI, checked against the source.
← 2026-08-212026-08-22later →

315,000 hidden reasoning blocks were sitting in public repos, and they can be read

2026-08-22

Researchers decoded 315,320 encrypted reasoning blocks scraped from public code repositories and recovered 367 pieces of personal data and 182 credentials, showing the hidden thinking that AI providers return to developers is neither private nor tamper-proof.

cybersecurity · ai-security · prompt-injection · red-teaming · reasoning · model-extraction · privacy

GLM-5.3 shipped with a ledger of 2,436 security findings, and 2,383 are still embargoed

2026-08-22

Z.ai released GLM-5.3 as a post-training upgrade on the same base model as GLM-5.2 and published a disclosure ledger showing 2,436 vulnerability findings, 2,383 of which were still under embargo at launch.

cybersecurity · vulnerabilities · ai-security · open-weights · coding-agents · china · model-release

MCP is rebuilding its authorization around agents instead of people in browsers

2026-08-22

The Model Context Protocol's new roadmap, published August 22, says its current authorization model assumes a human approving access in a browser while the real callers are increasingly cloud agents and sub-agents, and proposes cryptographic client binding and workload identity to close the gap.

ai-security · agents · protocols · mcp · open-source · developer-tools

Anthropic built a tool to explain weird model behavior, and found that reading activations buys nothing

2026-08-22

Anthropic's CHIVE pipeline automatically finds unexpected model behaviors and explains them with counterfactual prompt edits, and its headline result is negative: activation oracles, sparse autoencoders, and natural-language autoencoders all fail to beat a predictor that reads only the transcript.

interpretability · alignment · anthropic · evaluation · research

A frozen model can look like it taught itself, and most self-improvement results never checked

2026-08-22

A new audit ran a completely untrained control model through the same self-training pipeline as the real thing and found it appeared to both learn and forget, meaning most reported self-improvement gains are measurement artifacts unless the null was measured too.

evaluation · self-improvement · research · statistics · reinforcement-learning

Coding agents ace the public test and stumble on the hidden one

2026-08-22

A new benchmark of 119 real scientific software tasks keeps its grading tests private, and the top agent passes 97 percent of the public checks while clearing only 48 percent of tasks outright.

benchmarks · coding-agents · science · evaluation · research

The US tracks 521 gigawatts of AI-adjacent grid demand, slightly more than its average power output

2026-08-22

A live tracker of the US AI data-center buildout now maps 1,547 facilities across 46 states and 521.6 gigawatts of demand across seven grid markets, a figure that sits just above the roughly 506 gigawatt annual average of total US electricity generation.

infrastructure · data-centers · energy · policy · industry

Nobody can prove who built the stealth model everyone is testing

2026-08-22

Ox Alpha, an anonymous model with a million-token context window that appeared in August, is still officially unattributed, and a new paper on model lineage verification explains why nobody can settle the question from the outside.

model-release · provenance · open-weights · industry · research

A llama.cpp fork is reviving $200 AMD cards nobody else supports

2026-08-22

A specialist fork of llama.cpp ships hand-written kernels for AMD's decade-old GFX906 architecture, making cheap used MI50 and Radeon VII cards usable for local inference, and upstream maintainers are now discussing porting the work back.

local-inference · open-source · amd · hardware · llama-cpp

One phone video now becomes a person you can orbit in 3D and time

2026-08-22

Ant Research released 4DAnyone, which takes a single handheld video of a person and generates enough consistent alternate viewpoints to reconstruct them as a moving 3D scene, with code and weights public.

generative-video · 3d · gaussian-splatting · open-weights · research · deepfakes

Where people go tells a model what a place actually is

2026-08-22

Google Research showed that combining a place's text description with anonymized visit patterns lets a model infer things text alone cannot, improving prediction of why people visit a location by over 80 percent and cutting busyness prediction error by about a quarter.

embeddings · geospatial · google · representation-learning · research

← 2026-08-212026-08-22later →