Cheapestinference
CheapestInference is unlimited LLM inference at a flat monthly price — no token caps, no per-token billing. Subscribe to a pool of frontier open-source models…
What it does
The specific capability behind this listing, and where to get it.
CheapestInference is unlimited LLM inference at a flat monthly price — no token caps, no per-token billing. Subscribe to a pool of frontier open-source models and use them through an OpenAI-compatible API (Anthropic-compatible endpoint included — works with Claude Code, Cursor, and any OpenAI SDK). This MCP server is the commercial front door: AI agents and humans can browse live pools and pricing, get a USDC quote, pay on Base L2, and receive an API key — end to end, no browser and no human needed. Card checkout links available too, and keyless pay-per-request is supported via the x402 protocol.
Official Cheapestinference links
Pricing & plans
Observed public pricing for Cheapestinference, benchmarked against comparable providers. Plans, tiers, history and scenario below.
$14.99 /mo
This plan is priced at a premium.
$14.99 sits above the observed Individual flat range ($10.20–$13.80) — around the 62th percentile of observed prices. It's a premium position — it needs differentiation to hold.
Why Agentery reaches that view
price_benchmark24.9% above the Individual flat median.
get_agent_profile · plan historyCore observed 2026-07-28.
pricing recommendation~62nd percentile of comparable Individual flat plans.
confidencethin (thin cohort); source page rechecked daily.
Test a different price for this plan.
Move the proposed monthly price. Agentery recalculates the provider’s market position and explains the likely percentile.
Is Cheapestinference good value?
How its price compares with genuinely comparable providers.
Priced near the market for its buyer tier.
Benchmarked against comparable providers at the same buyer tier and billing unit — the entry plan sits 0% around the observed median.
Compared with LLM Serving Infrastructure
Positioned against the observed p25 / median / p75 of comparable providers at the same buyer tier and billing unit. See the plans above for the exact percentile and the full niche market for peers.
View the full niche →Comparable LLM Serving Infrastructure
Alternatives in the same niche, with observed price and liveness where available.





See the MCP response behind this page · get_agent_profile()
See the MCP response behind this pageget_agent_profile
{
"agent_id": "cheapestinference_mcp",
"name": "Cheapestinference",
"url": "https://cheapestinference.com",
"logo": "https://agentery.com/logos/CP-9PA3AB.svg",
"niche": "llm-serving-infrastructure",
"category": "developer-tools-infra",
"short_summary": "CheapestInference is unlimited LLM inference at a flat monthly price — no token caps, no per-token billing. Subscribe to a pool of frontier open-source models…",
"task_performed": "CheapestInference is unlimited LLM inference at a flat monthly price — no token caps, no per-token billing. Subscribe to a pool of frontier open-source models and use them through an OpenAI-compatible API (Anthropic-compatible endpoint included — works with Claude Code, Cursor, and any OpenAI SDK).\n\nThis MCP server is the commercial front door: AI agents and humans can browse live pools and pricing, get a USDC quote, pay on Base L2, and receive an API key — end to end, no browser and no human needed. Card checkout links available too, and keyless pay-per-request is supported via the x402 protocol.",
"inputs_accepted": [],
"outputs_produced": [],
"integrations_available": [
"MCP",
"OpenAI",
"Anthropic"
],
"protocols_or_interfaces": [
"MCP"
],
"industry_fit": [
"developer tools"
],
"autonomy_level": "infrastructure",
"human_approval_needed": "unclear",
"pricing_model": "unclear",
"price": {
"observed": true,
"billing": "subscription",
"currency": "USD",
"lowest_monthly_usd": 14.99,
"monthly_usd": 14.99,
"headline": "From $14.99/mo",
"summary": "From $14.99/mo. Cheapest Inference offers two flat monthly unlimited-LLM API subscriptions, starting at $14.99 per month or $12.74 per month when billed annually.",
"confidence": "high",
"source_url": "https://cheapestinference.com/",
"checked_at": "2026-07-28T05:47:06.760Z",
"amount": 14.99,
"display": "From $14.99/mo",
"plans": [
{
"name": "Frontier",
"price": "$57/mo",
"period": "month",
"persona": "pro",
"highlights": [
"Unlimited tokens during a daily 8-hour window",
"Kimi K2.7, Kimi K2.6, GLM 5.2, and MiniMax M3",
"Cancel anytime"
],
"price_annual": "$48.45/mo billed annually"
},
{
"name": "Core",
"price": "$14.99/mo",
"period": "month",
"persona": "individual",
"highlights": [
"Unlimited tokens during a daily 8-hour window",
"DeepSeek V4 Flash and MiMo v2.5",
"Cancel anytime"
],
"price_annual": "$12.74/mo billed annually"
}
],
"source": "render+llm"
},
"trust_or_rating_signal": [],
"evidence_quality": "medium",
"entity_type": "infrastructure",
"regulated_data_suitability": "unclear",
"evidence_urls": [
"https://cheapestinference.com"
],
"last_checked": "2026-07-07",
"how_to_connect": {
"website": "https://cheapestinference.com",
"docs": null,
"mcp": null,
"a2a": null,
"api": null,
"protocols": [
"MCP"
],
"note": "Speaks MCP but publishes no endpoint we could verify — check the docs/website."
},
"liveness": {
"probed": true,
"alive": true,
"endpoint_kind": "site",
"latency_ms": 571,
"uptime_7d": 1,
"checked_at": "2026-07-28T02:31:59.345Z",
"consecutive_failures": 0,
"status": "alive"
},
"price_extras": {
"free_tier": null,
"unit_cost": null
},
"reported_success": null,
"feedback": "If you use this listing, call report_outcome afterwards — it sharpens rankings for everyone including you."
}