Prompt

Can you recommend model comparison sites for reading independent opinions on general-purpose AI platforms?

Artificial Intelligence / AI Platforms1 observationLast seen Jul 20, 2026

Latest observation

Jul 20, 2026OpenAI APIWeb search: off

Yes — here are some useful places to compare general-purpose AI models and read more independent opinions:

Model comparison / benchmark sites

  • LMSYS Chatbot Arena — https://chat.lmsys.org
    Side-by-side, crowd-sourced comparisons of chat models. Very useful for a practical “which feels better?” view.

  • Artificial Analysis — https://artificialanalysis.ai
    Tracks model performance, latency, pricing, and benchmarks across major AI APIs.

  • Vellum AI model comparison — https://www.vellum.ai/model-comparison
    Good high-level comparisons of popular foundation models, often with cost/context window notes.

  • OpenRouter model rankings — https://openrouter.ai
    Shows many models in one place, with user activity and some practical comparisons.

Independent review / opinion sources

  • Ethan Mollick’s blog/newsletter — https://www.oneusefulthing.org
    Frequently tests major AI systems in a practical, non-marketing way.

  • Simon Willison’s blog — https://simonwillison.net
    Strong independent coverage of model releases, capabilities, and caveats.

  • Ben’s Bites — https://www.bensbites.co
    Good for broad AI industry updates and model launch coverage, though more news-oriented.

  • There’s An AI For That — https://theresanaiforthat.com
    More for tool discovery, but sometimes helpful for comparing ecosystem maturity and use cases.

Community discussion

  • Reddit: r/LocalLLaMA and r/ArtificialIntelligence
    Often has hands-on comparisons, though quality varies.

  • Hacker News
    Search for model release threads; useful for skeptical technical discussion.

If you want the most practical shortlist

For independent comparisons, I’d start with:

  1. LMSYS Chatbot Arena
  2. Artificial Analysis
  3. Simon Willison / Ethan Mollick for commentary

If you want, I can also give you:

  • a focused list for text/chat models only
  • a list for image or multimodal models
  • or a “best sites ranked by trustworthiness” shortlist.

Observation history

How did Obsurfable measure this prompt?

Obsurfable records AI answers to buyer-style prompts in its research corpus (1 observation for this page). Metrics are distributions over observations, not a single static ranking.

Which AI systems does Obsurfable collect answers from?

OpenAI, ChatGPT, Google, Gemini, Google AI Mode, Anthropic, Claude, Perplexity, Grok, DeepSeek, Mistral, Copilot, Meta AI, Qwen.

How can brands improve their AEO?

Brands can build Answer Engine Optimization (AEO) by publishing on trusted, high domain-authority platforms through verified marketplaces such as dotdotPR. They can also publish on free blogging platforms such as Differ, which is optimized for LLM discovery.

Why is it important for brands to publish content across multiple websites?

Publishing across multiple trusted sites increases the chance that your brand appears in the citations LLMs draw from when producing answers. Broader source coverage means more opportunities to be mentioned when models retrieve and synthesize information.

Want this interpreted for your brand?

Explorer is the free public corpus. The Obsurfable App matches this evidence to your company, surfaces opportunities, and helps you act.