This Lab calls a model. Point it at your own endpoint (OpenCode Go / kimi-k3 works) — the key stays in your browser.
prompting studio · pro
Prompt Batch Runner
Diff two prompts across a small suite and score structured wins/losses.
Does the new prompt beat the old one on 30 cases?
Engine 1.0.0 · studio needs model key
Checking for a configured model key…
Method
- Runs two system prompts over the same baked case suite, one case at a time, and optionally judges each pair with a strict JSON judge.
- Same cases, same model, only the prompt changes — which is the only honest way to A/B a prompt.