Skip to content
MCP server

Cheapest LLM API

By tehtelabsAll Cheapestllmapi servers

Compare LLM API prices, search models and providers, and access reviewed benchmark results.

First seen 4 Oct 2026. One server, whatever directories list it: each directory listing keeps its own page and history.

1
Directories
Collected by InvokeRank
7
Tools
From an anonymous probe
-
ToolBench grade
Not graded by Arcade
-
GitHub stars
No repository data

Tools

ToolDescriptionBehaviour
benchmark_listDiscover exact benchmark IDs, editions, score units, configurations, provenance, and reviewed result counts before calling benchmark_results. Empty benchmarks are retained with resultCount 0.Read-only
benchmark_resultsGet reviewed scores for an exact benchmark_id from benchmark_list. Join result.modelId to catalog_search model.id. Compare scores only within the same benchmark edition and configuration. Includes source provenance and reported coding-agent usage when published. Follow pagination.hasMore.Read-only
catalog_searchFind canonical model slugs, context limits, and provider prices. Optional provider_slug filters availability, not the provider in bestOffer. bestOffer is selected for 10M input + 2M output tokens; use offer_compare for your workload. Follow pagination.hasMore.Read-only
executeRun a JavaScript async function in an isolated sandbox, composing catalog_search, model_detail, offer_compare, benchmark_results, provider_list, and benchmark_list as catalog.* methods. Return a small JSON value. Example: async () => { const page = await catalog.catalog_search({query: 'claude', page_size: 5}); return page.data.models.map(m => ({slug: m.slug, prices: m.bestOffer?.prices})); }. No network, secrets, or database binding is exposed. Shared worldwide execution budgets: 100 per UTC minute and 5000 per UTC day. Limits: 30 executions per client per minute (IPv6 /64), 20 catalog calls and read units (offer_compare uses 5), at most two concurrent calls, 5 seconds wall time, 50ms sandbox CPU, 32KiB code, 64KiB JSON output. JavaScript only; no imports or TypeScript syntax. Use /mcp?mode=code for only search + execute.Read-only
model_detailGet capabilities, context/output limits, and current sourced provider offers by model_slug from catalog_search. Prices are base-context rates; long-context tiers, cache writes, batch discounts, and free-tier quotas can change actual cost.Read-only
offer_compareRank current provider offers for up to five canonical model slugs by the USD cost of your input/output token workload. Input tokens are uncached. Estimates use base-context rates and exclude long-context tiers, cache writes, batch discounts, and free-tier quotas. provider_slugs restricts the ranked offers.Read-only
provider_listDiscover accepted provider slugs, roles, and current comparable model/offer counts. Provider minimum input and output prices are independent minima and may belong to different models; use offer_compare for an actual quote.Read-only

Directory listings

DirectoryListingTierFirst seen
Official MCP RegistryCheapest LLM API-4 Oct 2026