Skip to content

DoesItLocal

Score the task, not the model

Does it run local?

Look up a task and get one honest verdict: can a local LLM safely do it, which model clears it, and what to use instead when none can.

Early prototype. This is the task catalog — 6 of 23 tasks have a verdict from a real local-model eval run; the rest are honestly marked "not yet evaluated." Nothing here is fabricated. How a task becomes a verdict →
23 tasks in the catalog · 6 evaluated

The verdict scale (safe for local / local with a check / needs a bigger model / needs more data) is described in Reading a verdict. Start with the introduction for why this exists.

Built by Sam Carlton