hosted
fal post-trained MiniMax H3 and kept the weights News
Inference company fal released H3 Max, a post-trained version of the open-weight MiniMax H3 video model that renders a five-second 768p clip in under three seconds -- available only as a hosted API, with no weights published.
fal H3 Max Tool
A post-trained MiniMax H3 that renders a five-second 768p clip in under three seconds through fal's API, at 480p or 768p and up to 15 seconds. Hosted only -- there are no downloadable weights -- and priced at $3.60 per minute of generated video. There is a browser sandbox for trying it without writing code.
Tinker Tool
Thinking Machines Lab's hosted fine-tuning service, now serving Inkling alongside its other models. It is the managed path to customizing Inkling if you do not want to provision the GPUs yourself -- with the caveat that the API caps context at 256K tokens, versus 1M for the open weights you run yourself.
OpenRouter Tool
A production gateway to hundreds of models behind one API, with public rankings built from real usage and the ability to sort by price, throughput, latency and popularity.
Mistral OCR 4 Tool
A hosted document-reading model that converts scanned pages, PDFs, and complex layouts into clean structured text ready for a language model. Send a document, get back tidy text with the structure preserved.
Hailuo AI Video Tool
MiniMax's hosted front end for H3, for trying the model in a browser before downloading tens of gigabytes of weights. Supports text-to-video, image-to-video, first-and-last-frame and reference-to-video workflows.
AuK Demo Space Tool
A browser demo of AuK on Hugging Face Spaces for judging the output quality directly without downloading 6.8 GB of weights first, which for a generative audio model is the only assessment that counts.