Ground Truth.
AI, checked against the source.

← All topics

amd

Everything on Ground Truth tagged “amd” — 8 items.

A llama.cpp fork is reviving $200 AMD cards nobody else supports News

A specialist fork of llama.cpp ships hand-written kernels for AMD's decade-old GFX906 architecture, making cheap used MI50 and Radeon VII cards usable for local inference, and upstream maintainers are now discussing porting the work back.

Three separate tricks dropped the hardware floor for local AI in one day News

A 26-billion-parameter Gemma model ran on an iPhone by streaming expert weights from storage, AMD shipped a sparse mixture-of-experts model activating 2.8 billion parameters per token, and a llama.cpp fork began saving conversation caches to disk - three unrelated attacks on three different bottlenecks.

Anthropic plans up to two gigawatts of AMD chips, with AMD committing up to $5 billion back News

AMD said Anthropic plans to deploy up to 2 gigawatts of MI450-series capacity starting in the first half of 2027, and that AMD has committed to a future equity investment of up to $5 billion in Anthropic.

AMD and Cerebras split AI inference across two different chips News

AMD and Cerebras announced a joint inference offering on July 23 in which AMD's Helios racks process the prompt and Cerebras's wafer-scale engine generates the tokens, claiming up to five times the tokens per watt of a Cerebras-only setup.

AMD Absorbs FastFlowLM Team to Build GPU-Free NPU Inference News

AMD announced on July 17, 2026 that the FastFlowLM team has joined its Artificial Intelligence Group to build out an NPU-first, GPU-free local inference stack for Ryzen AI laptops.

llama.cpp-gfx906 Tool

A llama.cpp fork with hand-written kernels for AMD's GFX906 architecture, making used Instinct MI50, MI60, and Radeon VII cards usable for local inference. Ships custom flash-attention, RoPE, and matrix-multiply paths plus overclocking and power-scaling scripts.

Unsloth (AMD support) Tool

The fine-tuning and RL toolkit now documents AMD support across training, RL, chat, and deployment on Windows, WSL, and Linux, plus a cross-platform Studio beta.

FastFlowLM Tool

An NPU-first, GPU-free inference runtime built exclusively for AMD Ryzen AI (XDNA) NPUs, targeting long-context local LLMs at low power on laptop-class hardware; the team just joined AMD, with open install guides for Ubuntu, Arch, and more.