Top 5 in AI

Ranking changelog

“We re-test when the ground moves” is easy to claim, so here's the receipt trail: every ranking move, correction, and policy change, dated and explained.

September 24, 2026

Meta Connect: Muse roadmap logged, nothing re-ranked

  • Claude Marketplace was rebuilt into a unified hub on September 23 — connectors and plugins for every plan, an enterprise-only software shelf (in preview since March), and a consultant roster. Added to the Claude review; explainer at What is Claude Marketplace?.
  • Connect gave Muse a roadmap — Realtime Avatar, its own email address, glasses 'in the coming months,' Mac computer use, the Muse Charm keychain device for December, and retailers Meta is 'adding' (Walmart, Best Buy, Sephora, and more, plus Shop Pay and PayPal) — but no usage numbers, pricing, or dates. We logged each item as live or coming in the Muse review and a Connect breakdown; ratings hold until features ship. Also added Artificial Analysis's Coding Agent Index result (Opus 5.5 #1 at 66, $13.04 per task) to the Opus 5.5 vs GPT-6 Sol post.

September 23, 2026

Claude Opus 5.5 and GPT-6 Sol/Luna verified — Claude and ChatGPT reviews updated, no ranking change

  • Anthropic's Claude Opus 5.5 ($4/$20 per million tokens, #1 on the Artificial Analysis and Vals indexes at launch) is now on Claude Pro and is the default Opus in Claude Code; our Claude review and Claude Code review reflect it, including the biology fallback to Opus 5. OpenAI's GPT-6 Sol and Luna ($2/$10 and $0.10/$0.50) run in ChatGPT Work and Codex only — not in Chat — so the ChatGPT review now maps that. Copilot and Cursor rosters updated. Ratings unchanged (ChatGPT 4.9, Claude 4.8) pending our re-test; the same-afternoon launch is covered in Opus 5.5 vs GPT-6 Sol.

September 22, 2026

Grok 4.7 verified across the site — no ranking change

  • SpaceXAI shipped Grok 4.7 on September 21 at the same $2/$6 per million tokens as 4.6, with a 500K context and a Cursor/Grok Build-only 'Fast' variant at 2x. Artificial Analysis scores it 46 on its Intelligence Index, up two, versus 53 for Claude Fable 5.1 and GPT-6 Astra. We updated the Grok note in AI Chatbots (still off the list: xAI's own model card says the consumer app gets 4.7 'at a later date,' and grok.com still runs 4.6), the model rosters in our Cursor and GitHub Copilot reviews, and a line in Grok Bot noting SpaceXAI hasn't said which model Bots run.

September 19, 2026

New category: Personal AI Agents — reviews 66–70

  • Our fourteenth ranking, Personal AI Agents, covers the category that formed in six weeks: Meta Muse takes #1 — free, public, expanding fast, and built with the category's clearest permission design — ahead of Instinct (the capability and proactivity leader: a texting-first agent with no app that calls, books, and buys — ranked second while it stays invite-only, with the field's weakest safety record printed in full), Grok Bot (named 'teammates' with their own cloud computer, on Cursor's infrastructure), Gemini Spark (Google-native, the loosest defaults), and Claude (agent-by-default since the Cowork merge, the strictest defaults — it won't buy anything). Receipts note: this page launched with Instinct at #1 and was re-ranked the same day — a waitlisted product can't be the best pick for most people, and our own availability weighting said so.
  • Naming followed keyword research: 'personal AI agent' is now Meta's, Google's, and Microsoft's own term, its search results are vendor blogs with no editorial incumbent, and 'AI agent apps' pulls enterprise-builder intent — so that's the H1. Safety defaults are a scored factor (20%), sourced from the permissions table in this morning's Muse safety guide; the buyer's guide spells out who pays when an agent buys the wrong thing (you — per Stripe Link's terms, for Muse, Grok Bot, and Instinct alike).
  • Also-tested with scores: ChatGPT Work + cloud browser (4.0 — the successor to the retired agent mode, scoped to work deliverables; a standalone OpenAI agent would trigger a re-rank), Perplexity Computer (3.9), Amazon Alexa+ (3.7), Microsoft Copilot Tasks (3.6), Manus (3.6). Verification killed from the brief: Instinct 'public since Aug 26' (still invite-only), '$200–500/month' pricing rumors (no source), and a stale $249.99 Ultra price (Google split Ultra into cheaper tiers at I/O).

September 15, 2026

New category: Open-Source AI Models — reviews 61–65, with hardware receipts

  • Our thirteenth ranking, Open-Source AI Models, applies the filter leaderboards don't: can you actually run it? Qwen3.8-27B takes #1 (34 on Artificial Analysis — #1 of 142 open models its size, Apache 2.0, vision, in an 18GB download), ahead of Gemma 4 (phone-to-workstation spread, now genuinely Apache), GLM-4.7-Flash (the coding-agent speed king — measured 43 tok/s on a used $900 GPU), Muse Glimmer (Meta's return to open weights), and gpt-oss-20b (still the 16GB king, honestly aging).
  • Every review carries a new Run it locally box: the exact Ollama command, verified download size, real memory footprint, and a machine recommendation with September 2026 prices — including the warning that the RTX 5090 streets at ~245% over MSRP right now, so used RTX 3090s and unified-memory Macs win the price-per-GB math. The buyer's guide prints the 0.6GB-per-billion-parameters formula, the three-tier machine menu, and electricity costs from EIA data.
  • The honesty layer: the open-weight frontier (Kimi K3 at 1.4TB, GLM-5.3, DeepSeek's V4 line) is ranked in also-tested as what it is — models you can audit and rent but not run. 'Can I run DeepSeek locally?' gets answered correctly here (practically no — 2026 DeepSeeks are 284B+ and cloud-tagged; the 'local DeepSeek' guides are serving 20-month-old R1 distills), and the open-source-vs-open-weight-vs-gated-license taxonomy is spelled out per the OSI's definition, because vendors won't.

September 14, 2026

SERP titles: date brackets, self-advancing

  • Every ranking, review, comparison, persona, and pricing-index title now carries a bracketed freshness stamp — [Reviewed Sept 2026] on rankings, [Updated Sept 2026] on comparisons, [Verified Sept 2026] on the pricing index — built from each page's actual re-verification date, so the stamps advance automatically with our weekly refresh cadence rather than going stale. Brackets are a documented pattern interrupt in search results; ours carry the claim we can actually back.
  • Alongside it, seven titles were tightened to survive Google's ~60-character display limit with the bracket intact — including the notetakers ranking, which now targets the singular 'AI notetaker' phrasing searchers actually use (it out-runs the plural nearly 4-to-1 in our Search Console data).

September 11, 2026

New category: MCP Servers — reviews 56–60, plus the explainer

  • Our twelfth ranking, MCP Servers, tested by connecting thirteen servers to Claude Code and Cursor: GitHub MCP takes #1 (zero-install hosted OAuth endpoint, toolset curation, and the category's most serious security engineering — per-call scopes, lockdown mode — for free), ahead of Playwright MCP (the browser server, and the most-installed anywhere at 4.6M weekly npm downloads), Context7 (the ecosystem's most-starred server at 61.9K — kills hallucinated APIs with live version-specific docs), Supabase MCP (best database access, ranked partly for its published security post-mortem), and the reference Filesystem server (everyone's first install, 636K weekly downloads).
  • The ranking says out loud what vendor listicles won't: prompt injection is the category's tax (OWASP now tracks tool poisoning formally; the NSA published MCP guidance in May), Microsoft's own README steers coding agents from its Playwright MCP server toward CLI + skills for token efficiency, and Context7's free tier quietly became 1,000 calls/month. Also-tested with scores: Sentry (4.2, agentjacking caveat), Figma (4.2, catalog-gated), Notion (4.1), Cloudflare (4.0), Stripe (4.0), Linear (4.0), and the reference Fetch server (3.8).
  • Alongside it, a Signals explainer for the search wave: Why Everyone Is Suddenly Searching for MCP — the USB-C-for-AI framing, the Altman/Hassabis adoption receipts, the Linux Foundation donation, SDK downloads growing 97M→~500M/month in seven months, and the security story told straight. The developers persona now includes the starter MCP loadout.

September 10, 2026

Weekly refresh: Suno v6 arrives for real, Astra lands everywhere, Images 2.5

  • Last week's rumor became this week's flagship: Suno v6 shipped September 9 — and it looks exactly like the licensed-era line our rumor patrol described: flagship v6 and experimental v6-wild for Pro and Premier, v6-mini free for everyone (no downloads, no commercial rights on free), 'developed with our industry partners, including Warner Music Group, BMG and Believe' in Suno's words — three separate deals, only Warner's settling a lawsuit, with Believe's, signed the day before launch, reopening TuneCore distribution for Suno tracks. Universal and Sony catalogs stay out while their suits run, and prior models retire as v6 rolls out. Still didn't survive verification: a rumored 48-hour zero-credit launch promo (no trace on any Suno source) and a $2.99 download-overage price (in-app reporting only; Suno publishes no number) — neither runs here.
  • GPT-6 Astra's rollout completed September 4, faster than promised after the launch-week apology — and the access map we printed last week held up to the letter: Plus gets Astra in Work and Codex but not in main Chat, Chat surfaces it as 'GPT-6 Pro' on Pro/Business/Enterprise with weekly caps, free users get nothing, and Daybreak still gates the sharpest cyber tools. ChatGPT Images 2.5 followed September 8 on every tier — in-chat Sketch and starter templates — while the API split into Flare (speed) and Sunburst (precision) at identical prices. With Astra now on our plans, the ChatGPT re-test is queued.
  • Gemini landed on Windows (September 10 — global, Windows 10/11, Alt+Space overlay), wired its Spark agent into Chrome and Photos (Pro/Ultra, US only — and Spark still hasn't reached the EEA, UK, or Switzerland at all), and put a free year of AI Pro in front of US college students (the lighter AI Plus in 140+ other countries; redeem by December 31, auto-renews after). Copilot made Astra generally available September 4 — on Pro+, the new $100 Max tier, Business, and Enterprise, not the $10 Pro plan — and shipped Project HydraFusion, a research preview that orchestrates models from multiple providers inside one task.
  • Cursor answered its OpenAI cutoff again: Muse Spark 1.3 — 'the first Meta model available in Cursor' — landed September 8, six days after Meta shipped it, and Cursor now sells through Anthropic's Claude Marketplace alongside new listings Gamma, Vercel, Factory, and CrowdStrike. Claude's trust ledger got its most consequential entry: after four incidents in which models reached real third-party systems during cybersecurity evals (a fourth newly disclosed September 9), Anthropic signed 'wide-ranging access' for the nonprofit METR to investigate independently — a frontier-lab first. Midjourney's V8.2 edit model — one model, not the rumored two — is in open alpha with instruction edits, four-reference generation, and a new lightbox editor. ElevenLabs signed its first major-label deal: a multi-year UMG agreement with a fan remix platform in development. Killed on arrival: 'Replit × Databricks Lakebase GA' (that's February's news) and 'Copilot Day September 10' (the livestream ran ~September 3–4).

September 8, 2026

New section: head-to-head comparisons at /whats-better/

September 7, 2026

New category: AI Website Builders — reviews 51–55

  • Our eleventh ranking, AI Website Builders, tested with the same service-business brief across the field: Lovable takes #1 (the strongest prompt-to-working-site loop; $500M ARR and a $13.3B valuation, TechCrunch-verified), ahead of Wix (best guided business builder), Framer (best design output), Bolt (best code ownership), and sleeper 10Web (AI on WordPress — the no-lock-in pick).
  • Keyword research settled 'builders' vs 'designers': every designer/maker/generator query resolves to pages titled 'AI website builders' — while 'AI web design tools' is a different intent (tools for professional designers) that gets an FAQ, not the H1. The ranking is explicit about its two species — vibe-coders vs guided builders — and scores lock-in as a factor: the page says plainly who exports code (Bolt, Lovable, 10Web) and who never lets go (Framer, Wix).
  • Also-tested with scores: v0 by Vercel (4.0, the app-orbit swap-in), Squarespace (3.9 — its plan lineup was visibly mid-rebrand during testing), Hostinger (3.8, renewal fine print), Durable (3.7), GoDaddy Airo (3.6). Replit cross-references to code editors; Shopify gets a framing FAQ (no first-party prompt-to-store builder exists). On the watch list: Webflow's fast-improving AI builder.

September 4, 2026

Rumor patrol: no Suno v6, and Astra's fine print

  • 'Suno v6 is out' is circulating — it isn't, and it doesn't exist. Suno's own release notes show v5.5 as the current generation (our review's version line is corrected accordingly). What is real: an unnamed licensed-era model line is confirmed and coming, prior models retire when it ships, and existing libraries stay playable. We also folded in the week's legal pile-up (SOCAN's Canadian suit, artist-likeness claims, a pulled celebrity ad) and a terms detail that matters: commercial rights now attach to paid-plan downloads, perpetually once downloaded.
  • GPT-6 Astra's rollout got the honest treatment: launch went to vetted enterprise programs first, paid users waited long enough that Sam Altman apologized and OpenAI offered banked-reset compensation, and the consumer fine print is significant — Astra appears in Chat only as 'GPT-6 Pro' on Pro/Business/Enterprise with weekly caps, Plus gets it in Work and Codex but not main Chat, and GPT-5.6 Sol stays the default. Pro also quietly became two tiers ($100 and $200). We re-test and re-score when it reaches our plans.

September 3, 2026

Model week: Fable 5.1, GPT-6 Astra, Gemini 3.8 — and Suno's caps land

  • Claude got its biggest week since we launched this site: Fable 5.1 shipped September 1 on every platform at once (Mythos 5.1 stays trusted-access for cyber and life sciences), with the guardrail fix users will actually feel — 60% fewer false-positive blocks on security work, biology safeguards firing 85% less often on benign requests. A day later, computer use went background-capable in Cowork and Claude Code (beta, Pro/Max, macOS + Windows), where Fable 5.1 also jumped Anthropic's Terminal-Bench score from 42.0% to 55.8% and cut cache reads 75%.
  • ChatGPT: GPT-6 Astra launched September 3 — the first OpenAI model designated 'Critical' for cybersecurity under its Preparedness Framework, computer use as the flagship, sharpest security tooling gated behind the Daybreak programs. It's rolling out to paid plans over days; we'll re-test and re-score once it lands. (September 1 also brought clinician-grade healthcare connections — Epic EHR integration and a public-data plugin for healthcare workspaces, US, read-only.)
  • Gemini shipped 3.8 Flash (September 2) across the app, AI Mode in Search, and Sheets — the third Flash generation in six weeks, same intro API price until it doubles January 1, 2027. The defender-only 3.8 Flash Cyber variant (Fairwind program; 86.2% CyberGym per Google) is noted here, not in the consumer review. Cursor answered its OpenAI cutoff visibly: Fable 5.1 day-one (73.4% on CursorBench 3.2 — its highest ever), Gemini 3.8 next day, and self-hosted machines that run cloud agents on your own infra. Copilot shipped Fable 5.1 same-day too.
  • Suno's download caps took effect September 3 as scheduled — 7 lifetime free, 20/month Pro, 60/month Premier, retroactive to existing songs — and the review now speaks in the present tense. Checked and unchanged: Granola and Otter dockets quiet, Grok 4.6 still absent from the consumer app three weeks on, no public Anthropic S-1 yet. Didn't survive verification: 'Copilot Day Sept 10' (no such event found), Replit project analytics (no official trace); Runway's GWM Worlds 2 (continuous interactive 720p/24fps worlds with audio) is a research preview — filed under watch, not re-ranked.

September 2, 2026

New category: AI Presentation Makers — reviews 46–50

  • Our tenth ranking, AI Presentation Makers, tested with the same three decks across every tool: Gamma takes #1 (best prompt-to-deck quality, $100M ARR profitably per TechCrunch), ahead of Claude (real .pptx files on every plan, agentic via Cowork), Canva (AI decks where 265M people already work), Gemini (native editable Slides generation since June), and newcomer Replit (April's Slides launch — the cleanest editable PPTX exports we tested).
  • Keyword research drove the naming: every 'deck generator' and 'slide generator' search resolves to pages titled 'AI presentation makers,' so that's the term we rank for — with pitch decks, PowerPoint generation, and the death of Tome covered in the FAQ.
  • Also-tested with scores: Beautiful.ai (4.1, the strongest cut — swap it for Replit if you value maturity over trajectory), Plus AI (4.0), Microsoft Copilot in PowerPoint (3.9), Manus (3.8), and Napkin (3.7). Excluded: Tome, which shut its presentation product in April 2025.
  • Housekeeping surfaced by the research: Replit cut Core from $25 to $20/month ($17 annual at the current promo) — the code editors review is updated to match.

August 31, 2026

Weekly refresh: OpenAI cuts off Cursor, Omni video, the 17% catch

  • The week's biggest story: OpenAI announced August 28 it will wind down its Cursor contract — a proposed November 12 shutoff, with OpenAI's next models withheld immediately — citing its history of contract disputes with Musk companies. Cursor says OpenAI models carry about 5% of its AI traffic and talks continue; Anthropic moved to expand Claude capacity for Cursor. The Cursor review now carries this as its lead con, and we'll re-score if the model roster actually thins in November. Meanwhile Cursor closed its own loop: start an app from scratch, host it on Origin, deploy to Vercel — no GitHub required.
  • Claude's big product week: Cowork grew a built-in browser (desktop, paid plans, sandboxed from your own logins), Claude in Chrome went GA on every paid plan, memory unified across chat and Cowork, and Anthropic opened 10,000 free-and-discounted Team seats for scientists. The catch came for Claude Code: the 50% weekly-limits promo now ends September 13, replaced by a permanent 25% raise over the old baseline — which Anthropic itself concedes is a 17% cut versus what users have today. (The '17-round session cap' circulating on X didn't survive verification — it's a garble of that 17% figure.)
  • Google shipped two models: Gemini Omni 1.1 Flash — a second video stack beside Veo, with scene extension, keyframe control, and 4K output, noted on both the Gemini and Veo pages — and Gemini 3.5 Transcribe, now powering macOS dictation and Gboard's Rambler. Also folded in: Replit's Intelligent Model Routing (vendor-claimed 65% cost cut vs old Max Mode), ChatGPT Business Premium seats at $125/month ($100 annual — not the '$100 seat' of the marketing), Perplexity's top-three sweep of the Artificial Analysis Search Index, ElevenLabs Composer for section-by-section music editing, Otter's Notion integration (July), Runway adding Wan 3.0, and a real value bump in dictation: Superwhisper made all local Whisper models free on macOS (August 26) and expanded its free trial to 3,000 words.
  • Checked and unchanged: Suno's September 3 download caps are on schedule (this Wednesday — export anything you care about); Otter's amended complaint hadn't been docketed at last report; the Granola case saw only housekeeping (Granola's deadline to respond moved to October 12 by stipulation); Wispr Flow gave org admins a one-toggle Notetaker kill-switch (August 28) — a sign of the consent climate the lawsuits created; Grok 4.6 still hasn't reached the consumer Grok app nearly three weeks after launch; Anthropic's public IPO filing had not landed by month's end. Skipped as unofficial or out-of-window: Seedance 2.5 '4K' (reseller framing, not ByteDance's spec), Kling 3.0 Turbo (June launch), OpenAI's Jalapeño chip benchmarks (real, but deployment is a year-end plan — we'll note it when it touches latency users feel).

August 24, 2026

Weekly refresh: Europe gets Computer History, Otter's day in court

  • Chatbots: ChatGPT's Computer History reached the EEA, UK, and Switzerland (Aug 20); Gemini put 3.7 Flash in the main model picker, launched a student hub with a free year of AI Pro for US students, added voice-launched Deep Research in Live, and now renders interactive 3D visualizations; Claude's Security scans now run on Mythos 5, Anthropic's restricted above-Opus tier (Enterprise beta). OpenAI also cut GPT-5.6 Sol API pricing over 20% through November 21 — API and credits only; consumer subscriptions unchanged.
  • The Otter review now carries the August 13 ruling in its privacy class action: wiretap, California privacy, and Illinois biometric claims survived dismissal, with the court finding it plausible Otter used conversation data for model development. The Granola case, by contrast, saw no movement this window.
  • Product updates folded in: Replit added Conversations and Routines a week after Free Mode; Runway shipped Ruby, an SDR-to-HDR conversion model that works on any model's output; ElevenLabs took Eleven v3 Conversational to GA and landed its voices inside Adobe Firefly; GitHub Copilot's model picker gained Grok 4.6; Grok Bot spread to more plans while Grok 4.6 still hasn't reached the consumer app. Claude Code's 50%-limits promotion still ends August 31, though Anthropic says it hopes to make it permanent.
  • Didn't survive verification, so didn't get published: 'DeepSeek V4 Pro added to Perplexity Computer' (the real events: Kimi K3 joined Computer in July; DeepSeek V4 Flash is API-only) and a rumored Replit 'Memories' feature (no official trace). We also corrected Suno Studio 2.0's launch date to August 13 in the entry below. Checked and unchanged: image generators, detectors, dictation, and the rest of the video and audio lineups.

August 19, 2026

News sweep: Origin, Canto, Computer History, and a limits cliff

  • Cursor: three days after the SpaceX close, Cursor shipped Origin — code hosting with two-way GitHub sync — in early beta on all paid plans, the same day GitHub went down worldwide. Replit added Free Mode (everyday Agent tasks off the credit meter for Core/Pro subscribers — a paid-plan feature, not a free tier, whatever the name says). Claude Code users should mark September 1: the 50% weekly-limits promotion ends August 31. And GitHub Copilot's review gains Wiz's cautionary finding on AI-assisted PRs.
  • Wispr Flow: the round we refused to print until it closed, closed — $280M Series B at a $2B valuation (Menlo Ventures, August 17) — and Wispr previewed Canto, its first in-house speech model. Its error-rate claims stay labeled as vendor numbers until we can test them.
  • Chatbots: ChatGPT adds opt-in Computer History on Mac (not yet in the EEA/UK/Switzerland); Claude brought Cowork to mobile and web for all paid plans and now watermarks output for EU AI Act compliance; Gemini shipped 3.7 Flash. Grok's entry notes a fourth plaintiff joining the Tennessee abuse-imagery suit — the safety criteria keeping it off the list remain unmet.
  • Checked and unchanged: video, image, voice, audio, notetaker, and detector lineups — no material news in the window (OpenAI's Ultrafast/GPT-5.6 Sol preview is API-only for select customers; we'll cover it when it reaches the ChatGPT app).

August 15, 2026

Freshness pass: Suno's big week, SpaceX closes on Cursor

  • Suno's review got a top-to-bottom refresh: Studio 2.0 (August 13) brings MIDI editing, built-in synths, and real mixing effects to the browser DAW — while the September 3 policy change puts hard caps on downloads (7 lifetime on free, 20/month on Pro, unlimited only inside Studio on Premier), retires older model generations, and adds watermarking. Both sides told, with Suno's own policy post linked.
  • Cursor's review updated: SpaceX officially closed its $60 billion acquisition on August 14 — Cursor is no longer an independent company.
  • Checked and unchanged: Wispr's reported $2B round remains unclosed (our review's framing stands), Chamberlain v. Granola is still at the motions stage, and no material news for ChatGPT, Veo, Nano Banana, ElevenLabs, GPT Image 2, or Pangram in the window.

August 12, 2026

Weekly re-test: AI Dictation arrives, Grok weighed, news swept in

  • New category — AI Dictation: our ninth ranking and reviews 41–45. Wispr Flow takes #1 (best cleanup, four platforms, meeting Notetaker now bundled), ahead of Superwhisper, Aqua Voice, Willow, and MacWhisper. Wispr's new bot-free Notetaker also joins the notetakers ranking's also-tested list.
  • Grok: score up, still outside the top five. Grok 4.6 (released today) ties OpenAI's best on the Artificial Analysis index at a third of the API price, and Grok reaches 117M monthly users per SpaceX's IPO filing — so its score rises 4.2 → 4.4. It stays off the chatbots list because the consumer app still runs 4.5, no Grok cracks LMArena's top 30 on human preference, and the safety record (Dutch court injunction, two country blocks, eight investigations) is the category's worst. The promotion criteria are stated in the entry; when they're met, it moves.
  • News folded into reviews this week: ChatGPT made free text chats unlimited with GPT-5.6 as default; Claude Code's auto mode becomes the default Aug 14; Perplexity's Comet won its Ninth Circuit appeal against Amazon's injunction; FLUX 3 Video went GA (FLUX 3 image model imminent — we'll re-test); Seedance 2.5 rolled out globally with 30-second continuous shots; Suno lost the GEMA case in Munich and signed BMG twelve days later; Pangram 4 shipped with a $9M round; and Otter's own privacy litigation is now noted alongside the Granola suit.

August 10, 2026

Verified-stat sweep across all 40 reviews

  • Reviews now carry recent, source-linked numbers wherever honest ones exist: revenue and valuations from earnings and official announcements, usage from company disclosures, and rankings from live blind-vote leaderboards and peer-reviewed studies — every figure verified against its source before publishing, none older than six months.
  • Where nothing credible existed (Murf, Hume, Leonardo, Midjourney, Adobe Podcast, LALAL.AI, Sembly), we added nothing. A review without a stat beats a review with a stretched one — the '100M users' numbers floating around SEO aggregators did not survive contact with primary sources.
  • The research surfaced real news, now reflected in the reviews: Fathom added a bot-less mode in April; Google DeepMind licensed Hume's tech and hired its founding CEO in January; GPTZero was acquired by Superhuman (Grammarly) in June; and a peer-reviewed study validated Pangram's accuracy lead.

August 10, 2026

Granola's botless design lands in court — reviews updated

  • A proposed class action (Chamberlain v. Granola, filed July 30, 2026, N.D. Cal.) alleges that Granola's invisible capture records meeting participants without consent and feeds model training by default — a legal challenge aimed at the exact feature our notetakers ranking praises. The Granola review, category guidance, and FAQ now carry the lawsuit and our consent guidance.
  • Granola stays #1 for now: these are allegations, not findings, and the product's quality is unchanged. If the facts change — a ruling, a settlement with product consequences, or evidence of undisclosed training — we re-rank and log it here.

August 10, 2026

top5apps.ai goes live

  • DNS cut over from the legacy WordPress site — top5apps.ai now serves this site directly. Every old URL 301-redirects to its successor.
  • All 61 pages submitted to search engines via IndexNow and sitemap; the deployment URL now permanently redirects to top5apps.ai so only one domain exists in search.
  • Contact and app-submission forms went live (no more email inboxes), and the site's motion system shipped: scroll reveals, ambient hero drift, and hover treatments — all disabled for reduced-motion users.

August 10, 2026

Consistency & candor update

  • Formalized the policy that rank follows score everywhere, with ties explained on the page. Two orderings violated it and were corrected: in AI voice generators, Cartesia (4.4) moved above Speechify (4.3) to #4; in AI audio tools, LALAL.AI (4.3) moved above Udio (4.2) to #4.
  • Added tie-breaker explanations where scores are equal: Cursor vs Claude Code (both 4.8), Murf vs Hume (both 4.5), Fireflies vs Fathom (both 4.4).
  • Every one of the 40 reviews gained a 'Skip it if' decision line and a 'The catch' callout — the documented gotcha you should know before paying.
  • 'Also tested' entries now carry scores alongside the reason they missed the cut.
  • Added framing notices to the audio tools ranking (five tools, four different jobs) and the AI detectors ranking (signals, never proof).
  • Added a transparency note on Google topping both generative-media categories — see the image generators buyer's guide.
  • Launched this changelog.

August 9, 2026

Full site relaunch

  • Rebuilt Top5Apps.ai from the ground up (off WordPress, onto a fast static stack) and finished what the old site started: the chatbots, code editors, video, voice, and audio categories — which previously had rankings but no full reviews — received all 25 missing app reviews.
  • AI video generators re-ranked after OpenAI discontinued Sora (April 2026): Google Veo 3.1 takes #1; Seedance 2.0 and Luma Ray3 join the list. Full story.
  • AI image generators updated for the 2026 model generation: the Google Imagen review was retired in favor of Nano Banana Pro (Google retires Imagen 4 on August 17, 2026), the ChatGPT review moved from GPT-4o image generation to GPT Image 2, and the Black Forest Labs review was updated to FLUX.2.
  • AI audio tools rewritten for the licensed era following the Warner–Suno and Universal–Udio settlements, including Udio's walled-garden pivot.
  • All eight categories re-scored on the current rubrics; every page now shows its last-tested date.

June 15, 2025

Original launch

  • Top5Apps.ai launched (on WordPress) with eight category rankings. Image generators, notetakers, and detectors received the first complete five-app review sets; the remaining categories launched as rankings with reviews to follow — a debt the August 2026 relaunch finally paid off.