Run your inputs across every model and config, and get back the cheapest one that clears your quality bar, with a shareable proof. This endpoint exposes that loop to Claude and other agents over the Model Context Protocol.
prove_taskOne call. A plain-language task plus a few examples, and RedCrown runs every model and returns the cheapest that clears your bar, with a shareable proof link. Leave the expected output blank to rank against the model you use now.
try_sampleA zero-input demo on a public dataset. Returns a shareable proof link with no keys and no setup.
Fifteen more advanced tools drive the full loop (import results, scaffold and run experiments, live proxy capture, and the reviewer decision report).
claude mcp add --transport http redcrown https://mcp.redcrown.ai
{
"mcpServers": {
"redcrown": { "type": "http", "url": "https://mcp.redcrown.ai" }
}
}Once connected, try: "Use RedCrown to prove the cheapest model for
classifying these support tickets." The agent calls prove_task and returns a proof link.