The same prompt, answered by ChatGPT and Claude side by side.
Pick one of 3 prompts, watch both models answer at once, and keep the one you would actually send. Nothing here calls a model: the answers are fixed transcripts, so the page is free to run and identical every time.
How this works
Each prompt is scored out of 100 on the one thing that prompt is really testing, and the winner is not always the same model. That is the point: ChatGPT structures work better, Claude decides what matters more, and which one you want depends on the task in front of you.
These are fixed example answers written to show how the two models differ in kind rather than in quality. They are not live model output, and they are not a benchmark.
Summarize this 40 page contract
Judged on accuracy and focus
ChatGPT (OpenAI), 92/100
Here's a breakdown of the key terms:
Parties: Meridian Corp (client) and Halcyon Ltd (vendor)
Term: 24 months, auto renews unless cancelled 60 days out
Fees: $12,400/mo with a 4% annual increase
Liability: capped at 12 months of fees
Termination: either party, 90 days written notice
Let me know if you'd like a clause by clause table.
Claude Sonnet (Anthropic), 96/100
This is a 24 month services agreement between Meridian Corp and Halcyon Ltd at $12,400 a month, rising 4% each year.
Three things deserve your attention before signing: the contract renews automatically unless you cancel 60 days before term, liability is capped at one year of fees, and the 90 day termination notice is longer than standard.
Everything else is boilerplate.
Verdict: Claude wins, 96 to 92. Both got the facts right, but Claude told you what to worry about instead of listing everything equally.
Decline a meeting, keep the bridge
Judged on tone and readiness
ChatGPT (OpenAI), 88/100
Subject: Re: Partnership discussion
Hi [Name],
Thank you so much for reaching out! Unfortunately, I won't be able to make this work with my current schedule. I truly appreciate you thinking of me, and I'd love to keep the door open for future opportunities.
Wishing you all the best with the project!
Best regards, [Your name]
Claude Sonnet (Anthropic), 95/100
Hi Dana,
Thanks for the invite. I'm going to pass on Thursday, my week is already overcommitted and I'd rather skip than show up half prepared.
If it's useful, I can review the deck async and send notes by Friday. Otherwise, catch me in the next round.
Noura
Verdict: Claude wins, 95 to 88. No placeholders, a real reason, and an alternative offered. You could send it exactly as written.
Five things stand between these notes and Friday: lock scope on Monday so QA has a stable build, run the payments regression Tuesday, get the press kit out of Design by Wednesday, schedule the Product Hunt assets Thursday night, and keep Friday clear for the launch itself.
The notes also mention localized pricing. I'd cut it from this launch, it isn't load bearing.
Verdict: ChatGPT wins, 94 to 90. Owners, deadlines, and a numbered list you can paste straight into a tracker.
Questions
Does this call the real models?
No. The answers are fixed transcripts, which is why the page loads instantly and costs nothing to run. To send your own prompt to both models at once, open Whizi.
Which model is better, ChatGPT or Claude?
Neither, across the board. On these three prompts Claude wins the contract summary and the email because it decides what matters, and ChatGPT wins the checklist because it structures work. Picking one for everything is the mistake this page is about.
Can I share a specific comparison?
Yes. Copy the shareable link and the page opens on the same prompt for whoever you send it to.