Compare Two ChatGPT Prompts
A side-by-side way to decide between two ChatGPT prompt drafts — scored on clarity, specificity, output control, and risk instead of gut feeling.
'Research X and tell me what's best' against a scoped prompt with criteria and source rules — compared on clarity, where research prompts live or die.
Research prompts fail by being unfalsifiable: 'find the best option' gives the model nothing to be wrong about, so it returns a confident overview of everything. The fix is scope, criteria, and source rules — and a comparison makes the gap measurable. The loaded pair asks the same question two ways; the clarity and specificity scores show why one returns a decision input and the other returns a Wikipedia summary.
Compare with Clarity focus
The loaded pair shares a goal. A names a category; B names candidates, criteria, a team size, and source rules.
Spot A's contradiction
'Comprehensive but keep it short' is flagged in the risk section — the most common research-prompt conflict.
Note what makes B falsifiable
Named criteria mean the answer can be checked. That's the line between research and content.
Scope your own ask
Rewrite your open-ended research prompt with candidates, criteria, and an output shape, then compare against the original.
It scores the pair on clarity and specificity and flags risks; it does not run the research or verify an answer. Comparing the open ask against the scoped version (Linear, Jira, Asana with named criteria and source rules) shows why one returns a decision input and the other a generic overview. Which to adopt, and running it, stays with you.
It surfaces that contradiction in the risk section, the most common research-prompt conflict, where "comprehensive but short" silently averages into neither. Against Prompt B's named evaluation criteria (onboarding effort, sprint/reporting features, per-seat cost), the Clarity focus makes the gap measurable: named criteria mean the answer can be checked, the line between research and a generic overview.
A side-by-side way to decide between two ChatGPT prompt drafts — scored on clarity, specificity, output control, and risk instead of gut feeling.
Seven questions that decide between two prompts — audience, format, length control, constraints, criteria, ambiguity, and contradictions.
Two blog prompt variations for the same topic, compared: which one actually controls angle, audience, structure, and length?
A set of before-and-after examples showing exactly what prompt cleanup removes — and what it deliberately leaves alone.
Formats fuzzy agent instructions into a structured prompt with objective, available tools, constraints, success criteria, and failure handling.
Convert scattered bug notes, Slack messages, or user complaints into structured engineering tasks with reproduction steps, severity, and root cause hypothesis.
Compare two prompts side by side — quality scores, strengths, risks, and a clear recommendation.