home/roleplay/prompt-ab-analysis

Prompt ab Analysis

GPTClaudeGemini··419 copies·updated 2026-07-14
prompt-ab-analysis.prompt
# Prompt Robustness Analysis

Prompt variants are evaluated on stable cases rather than anecdotal examples.

| prompt | mean_score | failures | dominant_failure |
| --- | --- | --- | --- |
| v1 | 0.5 | 6 | missing_evidence |
| v2 | 1.0 | 0 | none |
| v3 | 0.88 | 1 | too_verbose |

## Interpretation

`v2` wins in this miniature suite because it consistently asks for evidence.
`v3` is more cautious but sometimes verbose. This mirrors real prompt tradeoffs:
quality, concision, grounding, and calibration often move independently.

when to use it

Community prompt sourced from the open-source GitHub repo YutoTerashima/prompt-robustness-suite (MIT). A "Prompt ab Analysis" style prompt — adapt the placeholders and specifics to your task. Imported as-is and not independently retested here, so check the output before relying on it.

tags

roleplaycommunitygeneral

source

YutoTerashima/prompt-robustness-suite · MIT