compute.wick.pics

Will it fit?

Weights plus KV cache plus stated overhead, checked against the live book — the working shown, every approximation named. The fit is deterministic arithmetic; the speed is yours to measure. Optional: give your measured tokens/s and each machine prices itself per million tokens — we do the division, you own the throughput claim.

Spec-sheet arithmetic, not a benchmark: VRAM need is public math (the assumptions above are the whole list), the book is the live tape, and we never invent a tokens/s number — that one is yours. Agents: GET /api/fit or the will_it_fit MCP tool, docs.