Skip to content
Back to search
100
MCP live MCP 2025-11-25 streamable-http

io.github.thebaronofai/vetted-consumer

io.github.thebaronofai/vetted-consumer

Will a local LLM run on your hardware? GGUF quant, buy-vs-rent-vs-API cost, used-GPU prices.

Uptime
100.0%
1 direct probes · 30d
Response
1050ms
last probe
Tools
9
callable
Resources
0
readable
Prompts
0
available

Tools · 9

can_i_run_it

Will a given local LLM run on given hardware? Returns fit, the best quant that fits, theoretical tok/s, and real owner-measured tok/s where available.

recommend_quant

Which GGUF quantization to download for a model on given hardware: the full quant ladder with file size, max context, and tok/s for each, plus the recommended pick.

cheapest_hardware_for_model

The cheapest catalogued, buyable machine that runs a given model at Q4 with the requested context.

list_models

List the local LLM model classes the tools know about (params, dense/MoE, native context).

list_hardware

List the machines the tools know about (memory, bandwidth, price, buy link).

cost_compare

Buy vs rent vs API cost to run a model locally: monthly/1y/3y totals, break-even months, and the energy cost per 1M tokens. Same math as /cost-calculator/.

recommend_hardware

Ranked list of catalogued, buyable machines that run a model at the requested context, cheapest first, with an optional budget cap.

get_used_gpu_prices

Current typical used-GPU prices for local-AI rigs (eBay Browse API median asking + hand-verified, monthly).

compare_hardware

Side-by-side memory, bandwidth, price, and (with a model) fit + tok/s for 2 to 4 machines.

How to use

Add to your Claude Desktop / Cursor / Cline MCP config:

{
  "mcpServers": {
    "io.github.thebaronofai/vetted-consumer": {
      "url": "https://vettedconsumer.com/mcp",
      "transport": "streamable-http"
    }
  }
}