LLM Hosting Pricing
Developer tools Listed by MCPBayCompare LLM API prices across providers and GPU rental prices, with price history
By: llmhosting.ai
The connection URL lives in your cabinet. Use "Add" above, then open "My MCPs" — the address for your AI client appears there.
About
LLM Hosting Pricing tracks what it costs to run language models: per-token API prices across inference providers and hourly prices for renting GPUs.
Tools:
- search_models — find models and the providers that serve them.
- get_model_prices — input/output token prices for a model across providers.
- list_gpus — GPU types and where they can be rented.
- cheapest_gpu — the lowest current hourly price for a given GPU.
- gpu_price_history — how a GPU's rental price has changed over time.
Typical use: "What is the cheapest way to serve Llama 3.1 70B: an API provider or renting an H100? Estimate monthly cost for 50M tokens."
Who it is for: developers and teams budgeting LLM inference or choosing a hosting provider.
Notes: free, no account needed. Prices change often; always confirm on the provider's own pricing page before committing.
Tools (5)
cheapest_gpuCheapest GPU rentals right now, optionally filtered by minimum VRAM (GB).
get_model_pricesAll provider prices for one LLM model: input/output/cache price per 1M tokens and context window per provider, cheapest first.
gpu_price_historyDaily minimum rental price history for a GPU (recorded daily since 2026-07-06).
list_gpusEvery GPU with live rental pricing: cheapest $/hr, provider, tier, and VRAM.
search_modelsSearch LLM models by name. Returns matching models with provider count and the cheapest input/output price per 1M tokens.