
//beforeyouship — LLM cost modeling from your editor
vlatest
MCP Serversmithery🔥3k
Model the realistic monthly cost of an LLM app **before you build it**. Not a token calculator: retries, prompt caching, batch discounts, infra overhead, and 3×/10× growth are modeled in, across GPT-5.x, Claude, Gemini, DeepSeek, and more.
**Works without a key.** Connect and ask — demo mode covers the six free-tier models. A Pro API key ([beforeyouship.dev](https://beforeyouship.dev)) unlocks the full 18-model catalog.
## Tools
| Tool | What it does |
|---|---|
- **`estimate_cost`** Full cost model for an architecture at a given usage level. Returns Naive / Realistic / Worst Case $/mo per model, growth scenarios, and an opinionated recommendation. |
- **`get_model_prices`** Current per-1M-token pricing (input, output, cached, batch) with context windows and staleness metadata. |
- **`list_archetypes`** Seven preset architecture patterns (chatbot, RAG pipeline, multi-step agent, …) used as starting points for estimates. |
## Try it
Paste into Claude Code or Cursor after connecting:
> Estimate the monthl
streamable-http