Prompt Editor
Edit the prompts sent to Claude for article analysis. Changes take effect within 5 minutes (cache TTL).
Tokens — substituted at runtime from Remote Config weights
{{W_FA}}
{{W_OMR}}
{{W_MT}}
{{W_SVT}}
{{W_IC}}
{{W_DFO}}
{{W_LI}}
{{W_AI}}
{{TOT}}
{{FORMULA}}
{{MAX}}
Version history
Every save is kept. Restore loads it into the editor — review, then Save.
Loading…
Tokens — substituted at runtime per request
{{TODAY}}
{{EXCERPT}}
Version history
Every save is kept. Restore loads it into the editor — review, then Save.
Loading…
Scores each article twice — once under each candidate analysis prompt — and reports the
difference. Nothing is persisted: no cache write, no facts-KB write, no outlet
stats, so a run cannot contaminate production data.
Research runs once per article and both variants receive the identical context,
and prior-facts KB context is omitted, so the only thing that differs between A and B is the prompt.
Single model, no ensemble. Only the deterministic grounding layers apply (L1 score recompute,
L2 research-contradiction penalty).
Single-model scoring carries roughly ±14 points of run-to-run noise — use several articles and 2–3 repeats before trusting a delta.
Single-model scoring carries roughly ±14 points of run-to-run noise — use several articles and 2–3 repeats before trusting a delta.
A
0 characters
B
0 characters
—
| Article | A score | B score | Δ score | A intent | B intent | Δ intent | Cost |
|---|
| Category (mean across articles) | A | B | Δ |
|---|