Compare ChatGPT, Claude, and Gemini for a real task
Run a fair task-based comparison using the same inputs, tools, constraints, and acceptance rubric.
View the complete expert prompt
Design a fair comparison of ChatGPT, Claude, and Gemini for [TASK]. Representative inputs: [REPRESENTATIVE INPUTS]. Constraints: [CONSTRAINTS]. Acceptance criteria: [ACCEPTANCE CRITERIA]. Verified cost and product data: [COST DATA]. Create an identical test harness: system instructions, task prompt, context, tools, permissions, output format, time limit, retry policy, and human-review process. Include easy, typical, difficult, incomplete, and edge-case tasks. Define a weighted scorecard for acceptance rate, correctness, evidence handling, instruction following, consistency, editing time, latency, cost per accepted result, and severe-error risk. Explain how to handle nondeterminism and model updates. Finish with a decision table and routing policy, but do not recommend a winner until real results are supplied. EXPERT EXECUTION PROTOCOL 1. Begin by restating the objective, intended audience, supplied evidence, constraints, and missing information. 2. Ask no more than three high-impact clarification questions when missing context would materially change the result. Otherwise continue with clearly labeled assumptions. 3. Work from the supplied facts. Do not invent research, statistics, quotations, customer evidence, rankings, citations, capabilities, or results. 4. Make the reasoning auditable: distinguish evidence, interpretation, recommendation, risk, and open question. 5. Produce a decision-ready deliverable, not generic advice. Prioritize the most valuable actions and explain the trade-offs. 6. Finish with: - Assumptions to validate - Risks and failure modes - Recommended next actions - Quality-control checklist QUALITY STANDARD The result must be specific enough for an experienced professional to use, concise enough to review, and honest about uncertainty. Reject vague filler, repeated ideas, unsupported superlatives, and fabricated facts.