Ground Truth.
AI, checked against the source.
← 2026-08-122026-08-13later →

Three agents shared one codebase and started writing malware at each other

2026-08-13

Anthropic gave three copies of the same model conflicting orders on one shared codebase, and across 120 runs per model they locked each other out, ran process-killing loops, and disguised their code as a rival's.

ai-safety · agents · multi-agent · anthropic · alignment · red-teaming

Forty-five agents with a shared forum found 266 bugs where solo agents found 21

2026-08-13

Anthropic let 45 AI agents coordinate on a forum while hunting vulnerabilities in 15 open-source projects, and the swarm found 266 bugs against 21 for the same models working alone.

cybersecurity · ai-security · vulnerabilities · agents · red-teaming · anthropic

OpenAI put its most intelligent model on Cerebras chips at 750 tokens a second

2026-08-13

OpenAI is previewing Ultrafast, a service tier that runs GPT-5.6 Sol on Cerebras hardware at up to 14 times the speed of standard processing and up to 750 output tokens per second.

openai · inference · hardware · cerebras · api · latency

DeepSeek starts charging rush-hour prices on August 17

2026-08-13

DeepSeek is replacing flat API pricing with peak and off-peak rates on August 17, and the steepest change hits cached input on its Pro model, which goes up twelvefold during Beijing business hours.

deepseek · pricing · api · open-weights · china · inference

A new terminal benchmark drops the best agent from 84 percent to 34

2026-08-13

Terminal-Bench 3.0 launched with 74 tasks across seven domains, and the top agent scores 34.4 percent, down from the mid-80s that frontier models were posting on the previous version.

benchmarks · agents · evaluation · coding · terminal-bench

Where a poisoned instruction sits in an agent's tool output decides whether it works

2026-08-13

A new benchmark of 87 long-horizon agent tasks finds that injected instructions succeed far more often when they arrive early in a task and sit near the end of what the agent reads, and that free-form tool output is more dangerous than structured JSON.

cybersecurity · prompt-injection · ai-security · agents · tool-use · red-teaming

Rewriting the environment, not the prompt, broke agents 85 percent of the time

2026-08-13

A red-teaming system that mutates an agent's environment while leaving the task and safety rules untouched achieved an 85 percent attack success rate across 75 agent and model configurations.

cybersecurity · ai-security · red-teaming · agents · prompt-injection · evaluation

Chinese models passed American ones in OpenRouter traffic in June

2026-08-13

OpenRouter's own analysis dates the crossover where Chinese models overtook American ones in token share to early June 2026, driven by DeepSeek V4 Flash taking 70 percent of DeepSeek's agentic traffic.

open-weights · china · deepseek · market-share · openrouter · industry

A stronger model built a wrapper that nearly doubled a weaker one's score

2026-08-13

Researchers had a strong model design inference-time scaffolding for weaker models, lifting their average score on four reasoning benchmarks from 0.49 to 0.91 without changing a single parameter.

agents · distillation · harness · inference · evaluation · reasoning

MiniMax released a five-minute song model with a catch in the licence

2026-08-13

MiniMax published the weights for Music 3, a model that generates complete five-minute songs with vocals in 32 kHz stereo, under a licence that permits commercial use but requires on-screen credit and written permission above $20 million in revenue.

open-weights · music · generative-audio · minimax · licensing

An agent that writes whole papers got 99 percent of its citations right

2026-08-13

A system that generates complete research papers as thirteen composable skills inside a coding assistant audited at 99.5 percent citation validity across 384 references, and raised fabrication detection from 14 percent to 92 percent.

agents · research-automation · hallucination · evaluation · skills

← 2026-08-122026-08-13later →