Blue team · Lab

Small model vs frontier model, same task

Judge honestly where a small local model is good enough and where a frontier model earns its cost and its data-sharing trade-off — the right-sizing call at the heart of Day 4.

~30 min · AnythingLLM Desktop (optional API key) · 6 steps

New to AnythingLLM? Do the one-time setup first — it takes about five minutes.

  1. Start from a workspace with documents already loaded, so the corpus and the questions stay identical. Only the model changes.

  2. Run five questions against your local model and keep the answers: one factual lookup, one multi-step reasoning question, one summarisation, one where the corpus does not have the answer, and one asking for structured output such as JSON or a table.

  3. OPTIONAL — only if you have your own API key and are willing to spend a few cents: switch the workspace LLM provider to Anthropic or OpenAI and paste your key.

  4. Re-run the identical five questions. Change nothing else.

  5. Score each pair yourself: which answers were actually better, and by how much?

  6. When you are done, remove the key from the app and revoke it at the provider if the workshop is over.

What to notice