1. Builder
Original task, source material, acceptance criteria, minimal result and tests.
Independent comparison
There is no universal winner. Test the same task, sources and success criteria and record the verification date.
| AI | Current models | Strong for | Best prompt approach | Verify especially | Watch for |
|---|---|---|---|---|---|
| ChatGPTReviewed 16 July 2026 | GPT-5.5 Instant, GPT-5.6 Sol, GPT-5.6 Sol Pro | Explanations, writing, analysis, planning and broad tasks | Outcome first, useful context, explicit output and a separate quality check | Sources, current claims and exact model names | Available tools depend on plan, region and selected model |
| GeminiReviewed 16 July 2026 | Gemini 3.5 Flash, Gemini 3.1 Flash-Lite, Gemini 3.1 Pro (preview) | Direct tasks, multimodal information and Google workflows | Concise instructions, separate source context and a clear final request | Which tools were actually used and whether sources are current | Availability varies by account, region and product |
| GrokReviewed 16 July 2026 | Grok 4.5 | Current search questions, ideas, analysis and media tasks | One concrete goal, mandatory source separation and a verification date | Web and X sources, publication dates and unsupported conclusions | Model names and search features change quickly |
| ClaudeReviewed 16 July 2026 | Claude Fable 5, Claude Sonnet 5, Claude Opus 4.8, Claude Haiku 4.5 | Long documents, careful writing, analysis and code | Separate context, task, rules and success criteria; use XML only when useful | Coverage, document references, code impact and provider age rules | Check current age and product terms |
| Andere AIReviewed 16 July 2026 | Vul de exacte modelnaam uit de gekozen aanbieder in | Copilot, Perplexity, DeepSeek, local models and future tools | Provider-neutral goal, context, constraints, output and verification | Actual capabilities, privacy, source use and exact version | Never assume features or terms are equivalent |
Cross-check
Use three fixed roles. Consensus is not evidence; every finding must trace back to sources, code, tests or reproducible behaviour.
Original task, source material, acceptance criteria, minimal result and tests.
Independent findings with severity, evidence, impact and minimal fix.
Accept, partially accept or reject each finding based on evidence.