Skip to main content

This Lab calls a model. Point it at your own endpoint (OpenCode Go / kimi-k3 works) — the key stays in your browser.

prompting studio · pro

Prompt Batch Runner

Diff two prompts across a small suite and score structured wins/losses.

Does the new prompt beat the old one on 30 cases?

Engine 1.0.0 · studio needs model key

Checking for a configured model key…

Method

  • Runs two system prompts over the same baked case suite, one case at a time, and optionally judges each pair with a strict JSON judge.
  • Same cases, same model, only the prompt changes — which is the only honest way to A/B a prompt.