tool-calling
Everything on Ground Truth tagged “tool-calling” — 2 items.
tool-eval-bench Tool
A public benchmark harness for testing tool calling against OpenAI-compatible local and hosted serving endpoints.
Mercury 2.5 Tool
Inception's diffusion language model, which refines a whole draft in parallel rather than writing left to right, reporting 1,107 tokens per second on standard NVIDIA GPUs with a 260K context window. Closed weights, available through Inception's API, Baseten and OpenRouter with 100 million free tokens.