Which AI Writes Best? I Tested Claude, Grok, and ChatGPT

Which AI Writes Best? I Tested Claude, Grok, and ChatGPT

More

Summary

Craig Hewitt, founder of podcast platform Castos, runs a structured head-to-head comparison of three AI writing setups: Claude (via the Claude desktop app and Claude Code), Grok 4.5 running inside Cursor, and GPT-5.6 running inside Codex. He tests each across three real content types — a long-form SEO article on podcast advertising CPM rates, an email newsletter, and a social post — using an actual Castos brief as input.

The video captures concrete workflow differences between platforms, not just output quality. Cursor with Grok raises clarifying questions before writing but encounters a web-access block from an egress proxy, meaning the final article draws on fabricated data without flagging it. Codex with GPT-5.6 writes without asking questions but proactively cites and links external sources in the article body — a deliberate SEO strategy. Claude’s approach is evaluated on voice fidelity and long-form structure. Hewitt also flags Anthropic’s August 2026 announcement of invisible watermarking in all Claude model outputs, raising questions about whether search engines or AI-powered discovery tools like Perplexity might eventually deprioritize watermarked content.

For teams doing content marketing at scale, the video offers a practical tour of which platform fits which workflow, along with an honest account of where each tool’s guardrails and defaults can introduce unexpected failure modes.


📺 Source: Craig Hewitt · Published August 13, 2026
🏷️ Format: Comparison

1 Item

Channels

1 Item

Companies

1 Item

People