Evaluate
AdvancedLocation: /studio/[project]/evaluate
The Evaluate tab scores your scenario 0-100 against a '4 hormones (dopamine, serotonin, oxytocin, endorphin) + narrative grammar' rubric via a BYOK judge. It shows the overall score, summary, per-axis scores with rationale, suggestions, and top fixes; you can save evaluation versions and compare them on a score-trend sparkline, and refine mode loops evaluate-then-rewrite 1-5 times. It also shows a deterministic (LLM-free) Structure Quality card (slideshow-risk, scene variation, delivery-promise check) and a slop-gate banner for shot review status.
- 1Open the Evaluate tab; the deterministic Structure Quality card and slop gate show first (no key needed).
- 2Check the provider (BYOK key or CLI agent) and press Evaluate to get a score.
- 3Review per-axis scores, rationale, suggestions, and top fixes.
- 4Use Save version to persist a result and compare across versions on the trend.
- 5Use Refine with an iteration count to auto-loop evaluate-then-rewrite.
LLM evaluation goes through the /api/owntent/evaluate proxy: CLI agent (Codex/Claude Code) preferred, or BYOK key. BYOK default models are OpenAI gpt-4o, Anthropic claude-3-7-sonnet-latest, Google gemini-2.0-flash. Structure Quality and the slop gate are LLM-free.
Even without a key the Structure Quality card and slop gate appear instantly; LLM scoring separately needs a BYOK key or CLI agent. In short-drama format the rubric weights immediate hook and retention more heavily.