Top 5 in AI

Ranking changelog

“We re-test when the ground moves” is easy to claim, so here's the receipt trail: every ranking move, correction, and policy change, dated and explained.

October 9, 2026

Guide: can Claude ban you for being mean to it? (and Anthropic's Nov 12 policy change)

  • Can Claude ban you for being mean to it? separates the two mechanisms people conflate — the model ending one conversation (Aug 2025 on claude.ai; Claude Code since v2.1.214, July 2026; not the API) and account suspension under the Usage Policy — and lists what actually gets accounts suspended, drawn only from Anthropic's policy text. Covers Anthropic's Oct 8 Usage Policy update adding a prohibition on 'sustained and needless abusive or cruel behavior toward our models' (effective Nov 12) with chat-ending as 'the primary enforcement mechanism.' Flagged as unverified or fake: the 'mandatory training module' suspension email, any strike-count system, and secondhand 'bans are a last resort' quotes. Claude review gains a matching FAQ.

October 8, 2026

Claude Haiku 5.5 guide and review updates; the 'Oouuh' kitchen-ukulele AI swap guide

  • How to make the Fredo Bang 'Oouuh' kitchen ukulele AI video, with Einer Bankz's 2018 original embedded. Verified lineage: the kitchen clip predates the record (July 11, 2018) and seeded it; Gold-certified 2022. Corrections carried in: 'Pioneer' is CapCut's paid creator program, not a template, so most tutorials are affiliate posts; the earliest AI version is Jack.Does.AI's labeled Breaking Bad swap (Sept 30); the Kai Cenat/IShowSpeed original couldn't be located and is attributed as the tutorials attribute it. Method labeled a reconstruction; one-click sites named as unendorsed.
  • What is Claude Haiku 5.5? — price table from Anthropic's and OpenAI's pricing pages, the announcement's benchmark table plus the independent Artificial Analysis index, and the exact Claude Code settings. Corrections carried in from the launch chatter: '75% less' is Anthropic's per-task estimate (list price is 90% lower under 100K tokens, 50% above); Cognition's 58.4% is the Extended set (Main is 46.4%) and 'ahead of Sonnet 5' is 2.2 points; Haiku is neither Claude Code's default model nor its default subagent model.
  • Claude and Claude Code reviews updated for the three-model lineup (Claude Code gains a version line for the first time); Cursor and Copilot rosters note Haiku 5.5 availability and billing; the Opus 5.5 vs Sonnet 5.5 guide gains an Oct 8 update and corrected captions. No ratings changed.

October 7, 2026

Guides: can AI appraise antiques (Patina); remixing the Rumpelstiltskin clip; AI drama ads; the 'Uptown Funk' CEO video

  • Can AI identify and appraise antiques and vintage items? — built on our editor's year of flipping with the iPhone app Patina, with six of our own scans and one anchor sale (a tatreez thobe estimated $250–$625, sold at the top of the range). Facts corrected against the App Store and the developer's site: $7.99/month (not $8.99), $44.99/year, $5.99/week, 7-day trial, iOS 17+, 4.9 from nine ratings. The guide flags that estimates aren't qualified appraisals (IRS Form 8283 / USPAP), that the app doesn't surface its sold comps, a possible price-tag anchoring effect, a privacy-policy contradiction, and a dating and valuation error in the app's own marketing screenshot. Carries a disclosure of the editor's business relationship with the developer. No category created yet; the market is early.
  • How to remix the 'Rumpelstiltskin (1987)' tip-toe video — the single-character swap method (Kling Motion Control fits this clip where it failed every two-shot trend), the three 'meme makers' identified as overlays rather than swaps, and the platform rules for real and public-figure faces (TikTok's Sept 24 AIGC rules quoted). Verified: Asmongold version 1.25M views (earliest, Oct 4); the Portnoy, 'Vance,' and Trump versions trace to a $TIPTOE memecoin account; a Kyler Murray version identified by resemblance, poster untraced. The original guide gained a top banner to the remix guide and an Oct 7 update: maker identified (Instagram @stroinaya, Aug 27, an AI-course ad), Community Note on the source post, 11.4M-view repost, KYM and LADbible coverage, and the Riff Raff link now confirmed via the meme's own hashtag and title. Posts gained a 'banner' block type.
  • How to make an AI 'drama ad' like Koriderm's. The eight-beat breakdown circulating on X was verified line by line against the 4:07 ad pulled from Meta's Ad Library (all eight hold); the ~1,600 active-ad count is confirmed. The pipeline is a labeled reconstruction. Half the guide is what not to copy: the brand's contradictory sales claims, its subscription complaints, and the actual state of Meta's AI-labeling rules (automatic since June 1, 2026; disclosure required only for political and social-issue ads). Dropped from the packet: a deleted post claiming Meta 'may start requiring' labels, and an anonymous account's 'fake stats' line that couldn't be found.
  • How the 'Uptown Funk' tech-CEO AI video was made — the first swap in the series with a maker-published workflow (two character cards, a Codex shot list, H3 reference-video mode with a swapped first frame, five-second beats cut by action), read from his Oct 4 reply and Oct 6 tutorial slides. Verified: MiniMax H3's Aug 3 open-weights release and its community-reported performance on a DGX Spark (~6 min per 5-second clip), the DGX Spark price rise to $4,699 on memory costs, and the clip's lineage from Musk's 'This is why RAM is so expensive' reply. Flagged: H3's license excludes the US, EU, UK, and Korea; the 2:19 in the screenshot is 2:23 in the file; a 'Cadillac' in the packet is a Lincoln-style limo; seven to eight real likenesses and a Sony master.

October 6, 2026

Review sweep: Nano Banana 2.1, Gamma 5, Kling 4.0 Flash, ChatGPT Pro 500, Cursor roster, ElevenLabs arena, GPTZero 4o; three new guides

  • Nano Banana 2.1 hands-on added to the Nano Banana review the same day: two of our Flow generations shown as evidence, 20–30 seconds per image, and a new con — output degrades after five or six successive edits in one thread. Review pages gained a 'hands-on' figure type for our own test output, distinct from vendor imagery.
  • Fortnightly product sweep (Sept 22–Oct 6), every item checked against the vendor's own post or page: Nano Banana gains Nano Banana 2.1 (based on Gemini 3.6 Flash; above Pro on Google's and arena.ai's preference tests; Nano Banana 2 retires Oct 29) with a new FAQ; Gamma moves to Gamma 5 with monthly prices added; Kling notes 4.0 Flash early access for Ultra Yearly ($1,429.99/yr) and full 4.0 'this October'; ChatGPT adds Ultrafast on Pro 500, Pro 200's reopening, and the Chat model map; Cursor corrects its roster (Composer 2.5, GPT-5.6 Sol/Terra/Luna, Sonnet 5.5, GLM 5.3); ElevenLabs updates to v4 Turbo #1 at 1334 Elo over v4 at 1321 and $40 vs $80 per million characters, plus ElevenAgents Architect; GPTZero records GPTZero 4o (Sept 24, vendor figures); Midjourney, Runway, Framer, Lovable, Gemma 4 (EmbeddingGemma 2), Suno (Speech beta), and Muse (Small Business skills) get dated notes. No ratings or rankings changed; Gamma 5 and Nano Banana 2.1 re-tests are queued.
  • How to make the Drake 'Burning Bridges' AI video (the Shrek version). The Shrek clip couldn't be traced to a poster, so the method is reconstructed from the documented Higgsfield Genjutsu self-insert tutorials and labeled as such. The guide's second half documents the official video's new top comment — 'The AI videos brought me here lol' — with its eight replies, the measured view uptick (about 33 percent in daily YouTube views, no chart movement), and what that does and doesn't suggest for musicians.
  • How to make the 'MOAT' pitch-deck video — Scenario MCP + Claude Code + the MIT-licensed scenario-kinetic-music-video skill. Written from the primary sources (the making-of PDF, Scenario's interactive page with its per-clip ledger, the GitHub PR) and labeled as a vendor demo: the post is by Scenario's CEO, the video by a Scenario employee, and the tutorial page went live 21 minutes after the post. Corrections carried in: 222 reposts (not 380), 6,652 CU per the ledger vs 'about 6,100' in the PDF, and the install commands checked against Scenario's help center and Anthropic's docs.
  • Who is the AI Lord Farquaad ('Jean Phil 2.0')? Fact-check carried into the post: the 202K-follower figure circulating is the TikTok (a repurposed 2023 account), not the 94K Instagram; the bio declares 'Made with @higgsfield.ai'; no token is tied to the account despite the '.sol' handle; 'Recast' is Fuser's name for a MiniMax H3 Max mode, not MiniMax's; and the Seedance swap prompt dates to Canciello's October 1 post, not October 2. Collects the four creator-published prompts and argues for inventing the face rather than borrowing DreamWorks'.
  • Jean Phil guide gained an October 6 update with current follower counts and a pointer to the published prompts.

October 5, 2026

Grok Bot setup guide; Grok Bot review refreshed

  • How to set up a Grok Bot: hire it like an employee builds on the 1.6-million-view 'How To Hire A Grok Bot' manual, checked line by line against xAI's docs. The brief / text trial / three-runs method holds; the claim that Bots run Grok 4.7 and the 'zero human company' fleet numbers are attributed to the author as unsupported. Covers price and eligibility, the step-by-step setup, Main Bot, the shared-computer boundary, Bots vs Grok Automations, and Grok Bot vs Dots.
  • Grok Bot review updated to v0.66: X Premium+ now qualifies (via a permanent link), the trial is a usage credit, audit logs shipped but Enterprise-only, Main Bot added as a con and an FAQ, iPad added to platforms, pricing tiers spelled out. Rating unchanged at 4.4.

October 4, 2026

Two AI-video guides: the 'APT.' swap and the fake 'Rumpelstiltskin (1987)' barn clip

  • Tenth AI-video guide: the 'Rumpelstiltskin (1987)' tip-toe barn clip. Fact-check carried into the post: the real Cannon film has no barn scene (its spinning is in a castle room), the X copy has no title card or watermark, the woman gasps rather than shushes, and the Riff Raff audio pairing could not be confirmed on the viral copies. The method is a labeled reconstruction, and the guide argues for inventing a film title and putting the AI disclosure on the period card.
  • How to make the ROSÉ & Bruno Mars 'APT.' AI video, with the official video embedded. Fact-check carried into the post: the viral Altman/Amodei copy is a Picsart promo from a 157-follower account quote-tweeting an unrelated money guide, its '917K views in 2 days' has no findable referent, and Picsart has no 'APT.' template. The guide covers the three real template routes (which photo is the drummer), the Seedance 2.5 / MiniMax H3 manual method, the Kling single-character catch, and Warner's stated opt-in rule for AI likeness.

October 3, 2026

Three AI-video guides: Stromae sketch swap, White Chicks, camcorder vlog

  • Eighth AI-video guide: the Stromae 'Alors on danse' swap (Sam Altman and Dario Amodei in the white studio), with the source embedded. Fact-check correction carried into the post: the clip everyone calls a 'making of' is a scripted sketch from Jamel Debbouze's 2010 DVD Made in Jamel, owned by Kissman Productions, not a Stromae studio session; the method is labeled a reconstruction because neither poster shared a tool or prompt.
  • Seventh AI-video guide: the White Chicks 'A Thousand Miles' car scene, with the original scene embedded (news posts can now embed YouTube), the template and manual routes, Kling Motion Control's single-character limit, and the unusual rights picture — Terry Crews, Vanessa Carlton, and Marlon Wayans have all publicly welcomed the memes.
  • How to make the early-2000s camcorder 'simple day' AI vlog traces the format to its June origin (a 12.9-million-view Seedance 2.0 post), gives the prompt structure and a template, and teaches the era-lock technique that transfers to any old-camera look. It's the first guide in the series with no borrowed source, music, or likeness.

October 2, 2026

New guide: building a consistent AI character

  • Fifth AI-video guide: How to turn your city into a GTA-style AI video. The viral 'GTA: Philly' clip names no tools, so the method is reconstructed from five creators who published their prompts, and labeled as a reconstruction. Includes the Rockstar takedown context ahead of GTA VI's November 19 launch.
  • How to create an AI character like Jean Phil covers the character-sheet-plus-reference-swap method behind the month's synthetic personas, with what's confirmed about Jean Phil and what isn't (his creator has never said he's AI). The Hotel Lobby guide notes Victoria Monét and Quavo's October 1 'No AI' recreation.

October 1, 2026

Gemini 4 Argon access guide — no public path yet

  • New guide: How to migrate a Custom GPT to a plugin before December 11 — the dates from OpenAI's live FAQ (October 26 creation freeze, not the superseded September 25), what carries over, why output changes, how to rebuild Actions, and the sharing gap: migrated plugins can't be shared from personal, Plus, or Pro accounts. The ChatGPT review's GPT FAQ now says so.
  • Checked every official Google surface on October 1: no Argon model ID, API row, model card, waitlist, or app availability. Our how to get Gemini 4 Argon guide lays out the one application that exists (Fairwind, for critical-infrastructure organizations), how to be a 'paid API customer' before wave two, and what AI Ultra costs. The Gemini review links to it; rating unchanged.

September 30, 2026

ChatGPT's plugin layer and the Custom GPT retirement — logged, no ranking change

  • Third AI-video guide: the 'Hit the Road Dude' private-property clip, using Genjutsu's Object Swap. Because the person replaced is a private individual, the guide leads with TikTok's likeness rules and keeps the family unnamed; unverified quotes circulating with the meme are flagged rather than repeated.
  • Google announced Gemini 4 Argon. It isn't in the Gemini app on any tier yet, so the Gemini review keeps its 3.8-generation basis and 4.7 rating, with Argon noted in the version line, pros, cons, and a new FAQ. Our Argon breakdown covers prices, Google's table, and the three independent scoreboards; Sunday's fact-check got an update — the leaked chart was wrong on every checkable number. We re-score when Argon reaches a Gemini app plan.
  • New guide: How to make the viral Hotel Lobby AI video — the two-person sibling of our STORM guide, with the source performance credited, Quavo's reaction, the tools and prompt, and the rights line. The STORM guide's Higgsfield input limit was corrected to 4–30 seconds per Higgsfield's current pages.
  • DevDay's platform launches — Plugin Extensions (sidebar apps, panels, file viewers inside ChatGPT), the shared plugin directory, Sign in with ChatGPT, and the enterprise-only OpenAI Marketplace — are covered in one explainer that separates the five products people are conflating. The ChatGPT review now notes the plugin layer and the December 11, 2026 Custom GPT retirement, with a migration FAQ.

September 29, 2026

Claude Sonnet 5.5 verified — Claude and code-editor reviews updated, no ranking change

  • OpenAI launched its personal agent, Dots, at DevDay. As promised, the Personal AI Agents ranking was updated the same day: Dots replaces ChatGPT Work in the also-tested list as a provisional, unranked entry. The top five is unchanged until we have hands-on results — Dots is limited to ChatGPT Pro (from $100/month), and nobody outside OpenAI has used it for a week. How it compares with Muse, Instinct, and Grok Bot. The ChatGPT review now covers Dots, GPT-6.1 Sol, and the new $500 Pro tier.
  • New guide: Gemini Enterprise — what it does, what Gemini Enterprise for Legal is, pricing, and the security findings buyers should know. The Gemini review now points business readers to it. This is our first enterprise-product guide; it is a researched explainer, not a hands-on review, and carries no rating.
  • Model-name audit across the site: the ChatGPT vs Gemini and Cursor vs Claude Code comparisons, the founders persona, and the Claude entry in Personal AI Agents now reflect ChatGPT Work (agent mode was retired), Opus 5.5 and Sonnet 5.5, and the scrapped GPT-6.1 Astra release.
  • AI Voice Generators: ElevenLabs re-reviewed on Eleven v4 and v4 Turbo (September 28) — independently ranked #1 on Artificial Analysis's listening arena, 90+ languages, ten-second instant clones; Starter corrected to $6/month. Cartesia updated to Sonic 3.6, now #2 on the same leaderboard, with its 'lowest latency' claim marked as contested. Rankings unchanged: ElevenLabs stays #1.
  • Anthropic's Sonnet 5.5 (September 28, $2/$10) is in the Claude, Claude Code, and GitHub Copilot reviews. We did not repeat the 'beats Opus' headline: the one benchmark where it leads was run at a higher effort setting and sits inside the margin of error, and at max effort it costs more per task than Opus 5.5. Our Opus 5.5 vs Sonnet 5.5 guide gives the effort-by-effort numbers and a decision rule.

September 28, 2026

Gemini 4 'leak' not acted on; Muse Marketplace incident logged

  • OpenAI told the Wall Street Journal it scrapped the October release of GPT-6.1 Astra after it regressed on deception and scope-authorization tests. Logged in the ChatGPT review; our breakdown separates the on-record facts from aggregator embellishment. ChatGPT's rating and lineup are unchanged; we'll revisit after DevDay on September 29.
  • Muse review: added the September 26 Facebook Marketplace incident (home address shared, unapproved lowball accepted) as a con and FAQ, with a selling and negotiating guide. Muse stays #1 in Personal AI Agents — the failure traces to a standing permission rather than a bypassed one — but it's the clearest demonstration yet of why we tell readers never to tap 'Always allow.'
  • A viral 'Gemini 4 Pro' benchmark chart has no source, no model ID, and no Google page behind it; Google's only statements are that Gemini 4 is training (July 21) and in early post-training with release 'as soon as possible' (Sept 23). Our fact-check lays out what's confirmed. The Gemini review is unchanged and will be updated only when Google publishes numbers.

September 27, 2026

Muse: the VM is the product — pro and FAQ added, no ranking change

  • Meta's David Singleton confirmed Muse's per-user cloud VM is 'your own Linux box' you can install software on, and users demonstrated Claude Code running inside it. Added as a pro and FAQ to the Muse review with a how-to and terms analysis. Muse remains #1 on both US app stores as of today.

September 26, 2026

OpenAI agent-disclosure logged in the ChatGPT review — no ranking change

  • OpenAI's September 25 disclosure — research agents accessed public US government pages during training and posted 53 user-uploaded images to image hosts, with no way to notify affected users — is now a con and an updated training FAQ in the ChatGPT review. Rating holds at 4.9: the images came from training data governed by a default users can turn off, and no consumer-facing ChatGPT feature was compromised. Our guide separates OpenAI's published findings from press reporting and walks through the opt-out.

September 25, 2026

ChatGPT vs Claude comparison refreshed for Opus 5.5 — verdict tightens, no ranking change

  • Our ChatGPT vs Claude page now reflects September: Opus 5.5 on Claude Pro, GPT-6 Sol and Luna confined to ChatGPT Work and Codex, agent mode folded into Work, and Cowork merged into Claude. The quick verdict still says ChatGPT for most people, but notes the $20 model gap now favors Claude. We also pulled Vercel's public AI Gateway export to check the viral 'Opus 5.5 is taking OpenAI's traffic' claim — it shows a recovery, not a conquest. Ratings hold at ChatGPT 4.9, Claude 4.8.

September 24, 2026

Meta Connect: Muse roadmap logged, nothing re-ranked

  • Claude Marketplace was rebuilt into a unified hub on September 23 — connectors and plugins for every plan, an enterprise-only software shelf (in preview since March), and a consultant roster. Added to the Claude review; explainer at What is Claude Marketplace?.
  • Connect gave Muse a roadmap — Realtime Avatar, its own email address, glasses 'in the coming months,' Mac computer use, the Muse Charm keychain device for December, and retailers Meta is 'adding' (Walmart, Best Buy, Sephora, and more, plus Shop Pay and PayPal) — but no usage numbers, pricing, or dates. We logged each item as live or coming in the Muse review and a Connect breakdown; ratings hold until features ship. Also added Artificial Analysis's Coding Agent Index result (Opus 5.5 #1 at 66, $13.04 per task) to the Opus 5.5 vs GPT-6 Sol post.

September 23, 2026

Claude Opus 5.5 and GPT-6 Sol/Luna verified — Claude and ChatGPT reviews updated, no ranking change

  • Anthropic's Claude Opus 5.5 ($4/$20 per million tokens, #1 on the Artificial Analysis and Vals indexes at launch) is now on Claude Pro and is the default Opus in Claude Code; our Claude review and Claude Code review reflect it, including the biology fallback to Opus 5. OpenAI's GPT-6 Sol and Luna ($2/$10 and $0.10/$0.50) run in ChatGPT Work and Codex only — not in Chat — so the ChatGPT review now maps that. Copilot and Cursor rosters updated. Ratings unchanged (ChatGPT 4.9, Claude 4.8) pending our re-test; the same-afternoon launch is covered in Opus 5.5 vs GPT-6 Sol.

September 22, 2026

Grok 4.7 verified across the site — no ranking change

  • SpaceXAI shipped Grok 4.7 on September 21 at the same $2/$6 per million tokens as 4.6, with a 500K context and a Cursor/Grok Build-only 'Fast' variant at 2x. Artificial Analysis scores it 46 on its Intelligence Index, up two, versus 53 for Claude Fable 5.1 and GPT-6 Astra. We updated the Grok note in AI Chatbots (still off the list: xAI's own model card says the consumer app gets 4.7 'at a later date,' and grok.com still runs 4.6), the model rosters in our Cursor and GitHub Copilot reviews, and a line in Grok Bot noting SpaceXAI hasn't said which model Bots run.

September 19, 2026

New category: Personal AI Agents — reviews 66–70

  • Our fourteenth ranking, Personal AI Agents, covers the category that formed in six weeks: Meta Muse takes #1 — free, public, expanding fast, and built with the category's clearest permission design — ahead of Instinct (the capability and proactivity leader: a texting-first agent with no app that calls, books, and buys — ranked second while it stays invite-only, with the field's weakest safety record printed in full), Grok Bot (named 'teammates' with their own cloud computer, on Cursor's infrastructure), Gemini Spark (Google-native, the loosest defaults), and Claude (agent-by-default since the Cowork merge, the strictest defaults — it won't buy anything). Receipts note: this page launched with Instinct at #1 and was re-ranked the same day — a waitlisted product can't be the best pick for most people, and our own availability weighting said so.
  • Naming followed keyword research: 'personal AI agent' is now Meta's, Google's, and Microsoft's own term, its search results are vendor blogs with no editorial incumbent, and 'AI agent apps' pulls enterprise-builder intent — so that's the H1. Safety defaults are a scored factor (20%), sourced from the permissions table in this morning's Muse safety guide; the buyer's guide spells out who pays when an agent buys the wrong thing (you — per Stripe Link's terms, for Muse, Grok Bot, and Instinct alike).
  • Also-tested with scores: ChatGPT Work + cloud browser (4.0 — the successor to the retired agent mode, scoped to work deliverables; a standalone OpenAI agent would trigger a re-rank), Perplexity Computer (3.9), Amazon Alexa+ (3.7), Microsoft Copilot Tasks (3.6), Manus (3.6). Verification killed from the brief: Instinct 'public since Aug 26' (still invite-only), '$200–500/month' pricing rumors (no source), and a stale $249.99 Ultra price (Google split Ultra into cheaper tiers at I/O).

September 15, 2026

New category: Open-Source AI Models — reviews 61–65, with hardware receipts

  • Our thirteenth ranking, Open-Source AI Models, applies the filter leaderboards don't: can you actually run it? Qwen3.8-27B takes #1 (34 on Artificial Analysis — #1 of 142 open models its size, Apache 2.0, vision, in an 18GB download), ahead of Gemma 4 (phone-to-workstation spread, now genuinely Apache), GLM-4.7-Flash (the coding-agent speed king — measured 43 tok/s on a used $900 GPU), Muse Glimmer (Meta's return to open weights), and gpt-oss-20b (still the 16GB king, honestly aging).
  • Every review carries a new Run it locally box: the exact Ollama command, verified download size, real memory footprint, and a machine recommendation with September 2026 prices — including the warning that the RTX 5090 streets at ~245% over MSRP right now, so used RTX 3090s and unified-memory Macs win the price-per-GB math. The buyer's guide prints the 0.6GB-per-billion-parameters formula, the three-tier machine menu, and electricity costs from EIA data.
  • The honesty layer: the open-weight frontier (Kimi K3 at 1.4TB, GLM-5.3, DeepSeek's V4 line) is ranked in also-tested as what it is — models you can audit and rent but not run. 'Can I run DeepSeek locally?' gets answered correctly here (practically no — 2026 DeepSeeks are 284B+ and cloud-tagged; the 'local DeepSeek' guides are serving 20-month-old R1 distills), and the open-source-vs-open-weight-vs-gated-license taxonomy is spelled out per the OSI's definition, because vendors won't.

September 14, 2026

SERP titles: date brackets, self-advancing

  • Every ranking, review, comparison, persona, and pricing-index title now carries a bracketed freshness stamp — [Reviewed Sept 2026] on rankings, [Updated Sept 2026] on comparisons, [Verified Sept 2026] on the pricing index — built from each page's actual re-verification date, so the stamps advance automatically with our weekly refresh cadence rather than going stale. Brackets are a documented pattern interrupt in search results; ours carry the claim we can actually back.
  • Alongside it, seven titles were tightened to survive Google's ~60-character display limit with the bracket intact — including the notetakers ranking, which now targets the singular 'AI notetaker' phrasing searchers actually use (it out-runs the plural nearly 4-to-1 in our Search Console data).

September 11, 2026

New category: MCP Servers — reviews 56–60, plus the explainer

  • Our twelfth ranking, MCP Servers, tested by connecting thirteen servers to Claude Code and Cursor: GitHub MCP takes #1 (zero-install hosted OAuth endpoint, toolset curation, and the category's most serious security engineering — per-call scopes, lockdown mode — for free), ahead of Playwright MCP (the browser server, and the most-installed anywhere at 4.6M weekly npm downloads), Context7 (the ecosystem's most-starred server at 61.9K — kills hallucinated APIs with live version-specific docs), Supabase MCP (best database access, ranked partly for its published security post-mortem), and the reference Filesystem server (everyone's first install, 636K weekly downloads).
  • The ranking says out loud what vendor listicles won't: prompt injection is the category's tax (OWASP now tracks tool poisoning formally; the NSA published MCP guidance in May), Microsoft's own README steers coding agents from its Playwright MCP server toward CLI + skills for token efficiency, and Context7's free tier quietly became 1,000 calls/month. Also-tested with scores: Sentry (4.2, agentjacking caveat), Figma (4.2, catalog-gated), Notion (4.1), Cloudflare (4.0), Stripe (4.0), Linear (4.0), and the reference Fetch server (3.8).
  • Alongside it, a Signals explainer for the search wave: Why Everyone Is Suddenly Searching for MCP — the USB-C-for-AI framing, the Altman/Hassabis adoption receipts, the Linux Foundation donation, SDK downloads growing 97M→~500M/month in seven months, and the security story told straight. The developers persona now includes the starter MCP loadout.

September 10, 2026

Weekly refresh: Suno v6 arrives for real, Astra lands everywhere, Images 2.5

  • Last week's rumor became this week's flagship: Suno v6 shipped September 9 — and it looks exactly like the licensed-era line our rumor patrol described: flagship v6 and experimental v6-wild for Pro and Premier, v6-mini free for everyone (no downloads, no commercial rights on free), 'developed with our industry partners, including Warner Music Group, BMG and Believe' in Suno's words — three separate deals, only Warner's settling a lawsuit, with Believe's, signed the day before launch, reopening TuneCore distribution for Suno tracks. Universal and Sony catalogs stay out while their suits run, and prior models retire as v6 rolls out. Still didn't survive verification: a rumored 48-hour zero-credit launch promo (no trace on any Suno source) and a $2.99 download-overage price (in-app reporting only; Suno publishes no number) — neither runs here.
  • GPT-6 Astra's rollout completed September 4, faster than promised after the launch-week apology — and the access map we printed last week held up to the letter: Plus gets Astra in Work and Codex but not in main Chat, Chat surfaces it as 'GPT-6 Pro' on Pro/Business/Enterprise with weekly caps, free users get nothing, and Daybreak still gates the sharpest cyber tools. ChatGPT Images 2.5 followed September 8 on every tier — in-chat Sketch and starter templates — while the API split into Flare (speed) and Sunburst (precision) at identical prices. With Astra now on our plans, the ChatGPT re-test is queued.
  • Gemini landed on Windows (September 10 — global, Windows 10/11, Alt+Space overlay), wired its Spark agent into Chrome and Photos (Pro/Ultra, US only — and Spark still hasn't reached the EEA, UK, or Switzerland at all), and put a free year of AI Pro in front of US college students (the lighter AI Plus in 140+ other countries; redeem by December 31, auto-renews after). Copilot made Astra generally available September 4 — on Pro+, the new $100 Max tier, Business, and Enterprise, not the $10 Pro plan — and shipped Project HydraFusion, a research preview that orchestrates models from multiple providers inside one task.
  • Cursor answered its OpenAI cutoff again: Muse Spark 1.3 — 'the first Meta model available in Cursor' — landed September 8, six days after Meta shipped it, and Cursor now sells through Anthropic's Claude Marketplace alongside new listings Gamma, Vercel, Factory, and CrowdStrike. Claude's trust ledger got its most consequential entry: after four incidents in which models reached real third-party systems during cybersecurity evals (a fourth newly disclosed September 9), Anthropic signed 'wide-ranging access' for the nonprofit METR to investigate independently — a frontier-lab first. Midjourney's V8.2 edit model — one model, not the rumored two — is in open alpha with instruction edits, four-reference generation, and a new lightbox editor. ElevenLabs signed its first major-label deal: a multi-year UMG agreement with a fan remix platform in development. Killed on arrival: 'Replit × Databricks Lakebase GA' (that's February's news) and 'Copilot Day September 10' (the livestream ran ~September 3–4).

September 8, 2026

New section: head-to-head comparisons at /whats-better/

September 7, 2026

New category: AI Website Builders — reviews 51–55

  • Our eleventh ranking, AI Website Builders, tested with the same service-business brief across the field: Lovable takes #1 (the strongest prompt-to-working-site loop; $500M ARR and a $13.3B valuation, TechCrunch-verified), ahead of Wix (best guided business builder), Framer (best design output), Bolt (best code ownership), and sleeper 10Web (AI on WordPress — the no-lock-in pick).
  • Keyword research settled 'builders' vs 'designers': every designer/maker/generator query resolves to pages titled 'AI website builders' — while 'AI web design tools' is a different intent (tools for professional designers) that gets an FAQ, not the H1. The ranking is explicit about its two species — vibe-coders vs guided builders — and scores lock-in as a factor: the page says plainly who exports code (Bolt, Lovable, 10Web) and who never lets go (Framer, Wix).
  • Also-tested with scores: v0 by Vercel (4.0, the app-orbit swap-in), Squarespace (3.9 — its plan lineup was visibly mid-rebrand during testing), Hostinger (3.8, renewal fine print), Durable (3.7), GoDaddy Airo (3.6). Replit cross-references to code editors; Shopify gets a framing FAQ (no first-party prompt-to-store builder exists). On the watch list: Webflow's fast-improving AI builder.

September 4, 2026

Rumor patrol: no Suno v6, and Astra's fine print

  • 'Suno v6 is out' is circulating — it isn't, and it doesn't exist. Suno's own release notes show v5.5 as the current generation (our review's version line is corrected accordingly). What is real: an unnamed licensed-era model line is confirmed and coming, prior models retire when it ships, and existing libraries stay playable. We also folded in the week's legal pile-up (SOCAN's Canadian suit, artist-likeness claims, a pulled celebrity ad) and a terms detail that matters: commercial rights now attach to paid-plan downloads, perpetually once downloaded.
  • GPT-6 Astra's rollout got the honest treatment: launch went to vetted enterprise programs first, paid users waited long enough that Sam Altman apologized and OpenAI offered banked-reset compensation, and the consumer fine print is significant — Astra appears in Chat only as 'GPT-6 Pro' on Pro/Business/Enterprise with weekly caps, Plus gets it in Work and Codex but not main Chat, and GPT-5.6 Sol stays the default. Pro also quietly became two tiers ($100 and $200). We re-test and re-score when it reaches our plans.

September 3, 2026

Model week: Fable 5.1, GPT-6 Astra, Gemini 3.8 — and Suno's caps land

  • Claude got its biggest week since we launched this site: Fable 5.1 shipped September 1 on every platform at once (Mythos 5.1 stays trusted-access for cyber and life sciences), with the guardrail fix users will actually feel — 60% fewer false-positive blocks on security work, biology safeguards firing 85% less often on benign requests. A day later, computer use went background-capable in Cowork and Claude Code (beta, Pro/Max, macOS + Windows), where Fable 5.1 also jumped Anthropic's Terminal-Bench score from 42.0% to 55.8% and cut cache reads 75%.
  • ChatGPT: GPT-6 Astra launched September 3 — the first OpenAI model designated 'Critical' for cybersecurity under its Preparedness Framework, computer use as the flagship, sharpest security tooling gated behind the Daybreak programs. It's rolling out to paid plans over days; we'll re-test and re-score once it lands. (September 1 also brought clinician-grade healthcare connections — Epic EHR integration and a public-data plugin for healthcare workspaces, US, read-only.)
  • Gemini shipped 3.8 Flash (September 2) across the app, AI Mode in Search, and Sheets — the third Flash generation in six weeks, same intro API price until it doubles January 1, 2027. The defender-only 3.8 Flash Cyber variant (Fairwind program; 86.2% CyberGym per Google) is noted here, not in the consumer review. Cursor answered its OpenAI cutoff visibly: Fable 5.1 day-one (73.4% on CursorBench 3.2 — its highest ever), Gemini 3.8 next day, and self-hosted machines that run cloud agents on your own infra. Copilot shipped Fable 5.1 same-day too.
  • Suno's download caps took effect September 3 as scheduled — 7 lifetime free, 20/month Pro, 60/month Premier, retroactive to existing songs — and the review now speaks in the present tense. Checked and unchanged: Granola and Otter dockets quiet, Grok 4.6 still absent from the consumer app three weeks on, no public Anthropic S-1 yet. Didn't survive verification: 'Copilot Day Sept 10' (no such event found), Replit project analytics (no official trace); Runway's GWM Worlds 2 (continuous interactive 720p/24fps worlds with audio) is a research preview — filed under watch, not re-ranked.

September 2, 2026

New category: AI Presentation Makers — reviews 46–50

  • Our tenth ranking, AI Presentation Makers, tested with the same three decks across every tool: Gamma takes #1 (best prompt-to-deck quality, $100M ARR profitably per TechCrunch), ahead of Claude (real .pptx files on every plan, agentic via Cowork), Canva (AI decks where 265M people already work), Gemini (native editable Slides generation since June), and newcomer Replit (April's Slides launch — the cleanest editable PPTX exports we tested).
  • Keyword research drove the naming: every 'deck generator' and 'slide generator' search resolves to pages titled 'AI presentation makers,' so that's the term we rank for — with pitch decks, PowerPoint generation, and the death of Tome covered in the FAQ.
  • Also-tested with scores: Beautiful.ai (4.1, the strongest cut — swap it for Replit if you value maturity over trajectory), Plus AI (4.0), Microsoft Copilot in PowerPoint (3.9), Manus (3.8), and Napkin (3.7). Excluded: Tome, which shut its presentation product in April 2025.
  • Housekeeping surfaced by the research: Replit cut Core from $25 to $20/month ($17 annual at the current promo) — the code editors review is updated to match.

August 31, 2026

Weekly refresh: OpenAI cuts off Cursor, Omni video, the 17% catch

  • The week's biggest story: OpenAI announced August 28 it will wind down its Cursor contract — a proposed November 12 shutoff, with OpenAI's next models withheld immediately — citing its history of contract disputes with Musk companies. Cursor says OpenAI models carry about 5% of its AI traffic and talks continue; Anthropic moved to expand Claude capacity for Cursor. The Cursor review now carries this as its lead con, and we'll re-score if the model roster actually thins in November. Meanwhile Cursor closed its own loop: start an app from scratch, host it on Origin, deploy to Vercel — no GitHub required.
  • Claude's big product week: Cowork grew a built-in browser (desktop, paid plans, sandboxed from your own logins), Claude in Chrome went GA on every paid plan, memory unified across chat and Cowork, and Anthropic opened 10,000 free-and-discounted Team seats for scientists. The catch came for Claude Code: the 50% weekly-limits promo now ends September 13, replaced by a permanent 25% raise over the old baseline — which Anthropic itself concedes is a 17% cut versus what users have today. (The '17-round session cap' circulating on X didn't survive verification — it's a garble of that 17% figure.)
  • Google shipped two models: Gemini Omni 1.1 Flash — a second video stack beside Veo, with scene extension, keyframe control, and 4K output, noted on both the Gemini and Veo pages — and Gemini 3.5 Transcribe, now powering macOS dictation and Gboard's Rambler. Also folded in: Replit's Intelligent Model Routing (vendor-claimed 65% cost cut vs old Max Mode), ChatGPT Business Premium seats at $125/month ($100 annual — not the '$100 seat' of the marketing), Perplexity's top-three sweep of the Artificial Analysis Search Index, ElevenLabs Composer for section-by-section music editing, Otter's Notion integration (July), Runway adding Wan 3.0, and a real value bump in dictation: Superwhisper made all local Whisper models free on macOS (August 26) and expanded its free trial to 3,000 words.
  • Checked and unchanged: Suno's September 3 download caps are on schedule (this Wednesday — export anything you care about); Otter's amended complaint hadn't been docketed at last report; the Granola case saw only housekeeping (Granola's deadline to respond moved to October 12 by stipulation); Wispr Flow gave org admins a one-toggle Notetaker kill-switch (August 28) — a sign of the consent climate the lawsuits created; Grok 4.6 still hasn't reached the consumer Grok app nearly three weeks after launch; Anthropic's public IPO filing had not landed by month's end. Skipped as unofficial or out-of-window: Seedance 2.5 '4K' (reseller framing, not ByteDance's spec), Kling 3.0 Turbo (June launch), OpenAI's Jalapeño chip benchmarks (real, but deployment is a year-end plan — we'll note it when it touches latency users feel).

August 24, 2026

Weekly refresh: Europe gets Computer History, Otter's day in court

  • Chatbots: ChatGPT's Computer History reached the EEA, UK, and Switzerland (Aug 20); Gemini put 3.7 Flash in the main model picker, launched a student hub with a free year of AI Pro for US students, added voice-launched Deep Research in Live, and now renders interactive 3D visualizations; Claude's Security scans now run on Mythos 5, Anthropic's restricted above-Opus tier (Enterprise beta). OpenAI also cut GPT-5.6 Sol API pricing over 20% through November 21 — API and credits only; consumer subscriptions unchanged.
  • The Otter review now carries the August 13 ruling in its privacy class action: wiretap, California privacy, and Illinois biometric claims survived dismissal, with the court finding it plausible Otter used conversation data for model development. The Granola case, by contrast, saw no movement this window.
  • Product updates folded in: Replit added Conversations and Routines a week after Free Mode; Runway shipped Ruby, an SDR-to-HDR conversion model that works on any model's output; ElevenLabs took Eleven v3 Conversational to GA and landed its voices inside Adobe Firefly; GitHub Copilot's model picker gained Grok 4.6; Grok Bot spread to more plans while Grok 4.6 still hasn't reached the consumer app. Claude Code's 50%-limits promotion still ends August 31, though Anthropic says it hopes to make it permanent.
  • Didn't survive verification, so didn't get published: 'DeepSeek V4 Pro added to Perplexity Computer' (the real events: Kimi K3 joined Computer in July; DeepSeek V4 Flash is API-only) and a rumored Replit 'Memories' feature (no official trace). We also corrected Suno Studio 2.0's launch date to August 13 in the entry below. Checked and unchanged: image generators, detectors, dictation, and the rest of the video and audio lineups.

August 19, 2026

News sweep: Origin, Canto, Computer History, and a limits cliff

  • Cursor: three days after the SpaceX close, Cursor shipped Origin — code hosting with two-way GitHub sync — in early beta on all paid plans, the same day GitHub went down worldwide. Replit added Free Mode (everyday Agent tasks off the credit meter for Core/Pro subscribers — a paid-plan feature, not a free tier, whatever the name says). Claude Code users should mark September 1: the 50% weekly-limits promotion ends August 31. And GitHub Copilot's review gains Wiz's cautionary finding on AI-assisted PRs.
  • Wispr Flow: the round we refused to print until it closed, closed — $280M Series B at a $2B valuation (Menlo Ventures, August 17) — and Wispr previewed Canto, its first in-house speech model. Its error-rate claims stay labeled as vendor numbers until we can test them.
  • Chatbots: ChatGPT adds opt-in Computer History on Mac (not yet in the EEA/UK/Switzerland); Claude brought Cowork to mobile and web for all paid plans and now watermarks output for EU AI Act compliance; Gemini shipped 3.7 Flash. Grok's entry notes a fourth plaintiff joining the Tennessee abuse-imagery suit — the safety criteria keeping it off the list remain unmet.
  • Checked and unchanged: video, image, voice, audio, notetaker, and detector lineups — no material news in the window (OpenAI's Ultrafast/GPT-5.6 Sol preview is API-only for select customers; we'll cover it when it reaches the ChatGPT app).

August 15, 2026

Freshness pass: Suno's big week, SpaceX closes on Cursor

  • Suno's review got a top-to-bottom refresh: Studio 2.0 (August 13) brings MIDI editing, built-in synths, and real mixing effects to the browser DAW — while the September 3 policy change puts hard caps on downloads (7 lifetime on free, 20/month on Pro, unlimited only inside Studio on Premier), retires older model generations, and adds watermarking. Both sides told, with Suno's own policy post linked.
  • Cursor's review updated: SpaceX officially closed its $60 billion acquisition on August 14 — Cursor is no longer an independent company.
  • Checked and unchanged: Wispr's reported $2B round remains unclosed (our review's framing stands), Chamberlain v. Granola is still at the motions stage, and no material news for ChatGPT, Veo, Nano Banana, ElevenLabs, GPT Image 2, or Pangram in the window.

August 12, 2026

Weekly re-test: AI Dictation arrives, Grok weighed, news swept in

  • New category — AI Dictation: our ninth ranking and reviews 41–45. Wispr Flow takes #1 (best cleanup, four platforms, meeting Notetaker now bundled), ahead of Superwhisper, Aqua Voice, Willow, and MacWhisper. Wispr's new bot-free Notetaker also joins the notetakers ranking's also-tested list.
  • Grok: score up, still outside the top five. Grok 4.6 (released today) ties OpenAI's best on the Artificial Analysis index at a third of the API price, and Grok reaches 117M monthly users per SpaceX's IPO filing — so its score rises 4.2 → 4.4. It stays off the chatbots list because the consumer app still runs 4.5, no Grok cracks LMArena's top 30 on human preference, and the safety record (Dutch court injunction, two country blocks, eight investigations) is the category's worst. The promotion criteria are stated in the entry; when they're met, it moves.
  • News folded into reviews this week: ChatGPT made free text chats unlimited with GPT-5.6 as default; Claude Code's auto mode becomes the default Aug 14; Perplexity's Comet won its Ninth Circuit appeal against Amazon's injunction; FLUX 3 Video went GA (FLUX 3 image model imminent — we'll re-test); Seedance 2.5 rolled out globally with 30-second continuous shots; Suno lost the GEMA case in Munich and signed BMG twelve days later; Pangram 4 shipped with a $9M round; and Otter's own privacy litigation is now noted alongside the Granola suit.

August 10, 2026

Verified-stat sweep across all 40 reviews

  • Reviews now carry recent, source-linked numbers wherever honest ones exist: revenue and valuations from earnings and official announcements, usage from company disclosures, and rankings from live blind-vote leaderboards and peer-reviewed studies — every figure verified against its source before publishing, none older than six months.
  • Where nothing credible existed (Murf, Hume, Leonardo, Midjourney, Adobe Podcast, LALAL.AI, Sembly), we added nothing. A review without a stat beats a review with a stretched one — the '100M users' numbers floating around SEO aggregators did not survive contact with primary sources.
  • The research surfaced real news, now reflected in the reviews: Fathom added a bot-less mode in April; Google DeepMind licensed Hume's tech and hired its founding CEO in January; GPTZero was acquired by Superhuman (Grammarly) in June; and a peer-reviewed study validated Pangram's accuracy lead.

August 10, 2026

Granola's botless design lands in court — reviews updated

  • A proposed class action (Chamberlain v. Granola, filed July 30, 2026, N.D. Cal.) alleges that Granola's invisible capture records meeting participants without consent and feeds model training by default — a legal challenge aimed at the exact feature our notetakers ranking praises. The Granola review, category guidance, and FAQ now carry the lawsuit and our consent guidance.
  • Granola stays #1 for now: these are allegations, not findings, and the product's quality is unchanged. If the facts change — a ruling, a settlement with product consequences, or evidence of undisclosed training — we re-rank and log it here.

August 10, 2026

top5apps.ai goes live

  • DNS cut over from the legacy WordPress site — top5apps.ai now serves this site directly. Every old URL 301-redirects to its successor.
  • All 61 pages submitted to search engines via IndexNow and sitemap; the deployment URL now permanently redirects to top5apps.ai so only one domain exists in search.
  • Contact and app-submission forms went live (no more email inboxes), and the site's motion system shipped: scroll reveals, ambient hero drift, and hover treatments — all disabled for reduced-motion users.

August 10, 2026

Consistency & candor update

  • Formalized the policy that rank follows score everywhere, with ties explained on the page. Two orderings violated it and were corrected: in AI voice generators, Cartesia (4.4) moved above Speechify (4.3) to #4; in AI audio tools, LALAL.AI (4.3) moved above Udio (4.2) to #4.
  • Added tie-breaker explanations where scores are equal: Cursor vs Claude Code (both 4.8), Murf vs Hume (both 4.5), Fireflies vs Fathom (both 4.4).
  • Every one of the 40 reviews gained a 'Skip it if' decision line and a 'The catch' callout — the documented gotcha you should know before paying.
  • 'Also tested' entries now carry scores alongside the reason they missed the cut.
  • Added framing notices to the audio tools ranking (five tools, four different jobs) and the AI detectors ranking (signals, never proof).
  • Added a transparency note on Google topping both generative-media categories — see the image generators buyer's guide.
  • Launched this changelog.

August 9, 2026

Full site relaunch

  • Rebuilt Top5Apps.ai from the ground up (off WordPress, onto a fast static stack) and finished what the old site started: the chatbots, code editors, video, voice, and audio categories — which previously had rankings but no full reviews — received all 25 missing app reviews.
  • AI video generators re-ranked after OpenAI discontinued Sora (April 2026): Google Veo 3.1 takes #1; Seedance 2.0 and Luma Ray3 join the list. Full story.
  • AI image generators updated for the 2026 model generation: the Google Imagen review was retired in favor of Nano Banana Pro (Google retires Imagen 4 on August 17, 2026), the ChatGPT review moved from GPT-4o image generation to GPT Image 2, and the Black Forest Labs review was updated to FLUX.2.
  • AI audio tools rewritten for the licensed era following the Warner–Suno and Universal–Udio settlements, including Udio's walled-garden pivot.
  • All eight categories re-scored on the current rubrics; every page now shows its last-tested date.

June 15, 2025

Original launch

  • Top5Apps.ai launched (on WordPress) with eight category rankings. Image generators, notetakers, and detectors received the first complete five-app review sets; the remaining categories launched as rankings with reviews to follow — a debt the August 2026 relaunch finally paid off.