Home Catalog LLM Hosting Pricing

LLM Hosting Pricing

Developer tools Listed by MCPBay

Compare LLM API prices across providers and GPU rental prices, with price history

By: llmhosting.ai

The MCPBay team added this server to the catalog. Its author is not affiliated with MCPBay and did not submit it here. The server runs on the author's side, so MCPBay has no call statistics for it.
Connect

The connection URL lives in your cabinet. Use "Add" above, then open "My MCPs" — the address for your AI client appears there.

About

LLM Hosting Pricing tracks what it costs to run language models: per-token API prices across inference providers and hourly prices for renting GPUs.

Tools:
- search_models — find models and the providers that serve them.
- get_model_prices — input/output token prices for a model across providers.
- list_gpus — GPU types and where they can be rented.
- cheapest_gpu — the lowest current hourly price for a given GPU.
- gpu_price_history — how a GPU's rental price has changed over time.

Typical use: "What is the cheapest way to serve Llama 3.1 70B: an API provider or renting an H100? Estimate monthly cost for 50M tokens."

Who it is for: developers and teams budgeting LLM inference or choosing a hosting provider.

Notes: free, no account needed. Prices change often; always confirm on the provider's own pricing page before committing.

Tools (5)

  • cheapest_gpu

    Cheapest GPU rentals right now, optionally filtered by minimum VRAM (GB).

  • get_model_prices

    All provider prices for one LLM model: input/output/cache price per 1M tokens and context window per provider, cheapest first.

  • gpu_price_history

    Daily minimum rental price history for a GPU (recorded daily since 2026-07-06).

  • list_gpus

    Every GPU with live rental pricing: cheapest $/hr, provider, tier, and VRAM.

  • search_models

    Search LLM models by name. Returns matching models with provider count and the cheapest input/output price per 1M tokens.