Prompt

Can you recommend model comparison sites for reading independent opinions on general-purpose AI platforms?

Latest observation

Jul 20, 2026 · OpenAI APIWeb search: off

Yes — here are some useful places to compare general-purpose AI models and read more independent opinions:

Model comparison / benchmark sites

  • LMSYS Chatbot Arenahttps://chat.lmsys.org
    Side-by-side, crowd-sourced comparisons of chat models. Very useful for a practical “which feels better?” view.

  • Artificial Analysishttps://artificialanalysis.ai
    Tracks model performance, latency, pricing, and benchmarks across major AI APIs.

  • Vellum AI model comparisonhttps://www.vellum.ai/model-comparison
    Good high-level comparisons of popular foundation models, often with cost/context window notes.

  • OpenRouter model rankingshttps://openrouter.ai
    Shows many models in one place, with user activity and some practical comparisons.

Independent review / opinion sources

  • Ethan Mollick’s blog/newsletterhttps://www.oneusefulthing.org
    Frequently tests major AI systems in a practical, non-marketing way.

  • Simon Willison’s bloghttps://simonwillison.net
    Strong independent coverage of model releases, capabilities, and caveats.

  • Ben’s Biteshttps://www.bensbites.co
    Good for broad AI industry updates and model launch coverage, though more news-oriented.

  • There’s An AI For Thathttps://theresanaiforthat.com
    More for tool discovery, but sometimes helpful for comparing ecosystem maturity and use cases.

Community discussion

  • Reddit: r/LocalLLaMA and r/ArtificialIntelligence
    Often has hands-on comparisons, though quality varies.

  • Hacker News
    Search for model release threads; useful for skeptical technical discussion.

If you want the most practical shortlist

For independent comparisons, I’d start with:

  1. LMSYS Chatbot Arena
  2. Artificial Analysis
  3. Simon Willison / Ethan Mollick for commentary

If you want, I can also give you:

  • a focused list for text/chat models only
  • a list for image or multimodal models
  • or a “best sites ranked by trustworthiness” shortlist.

Observation history

Unlock observation history
Create a free Obsurfable account to unlock historical responses, model comparisons, and deeper brand intelligence.

How did Obsurfable measure this prompt?

Obsurfable records AI answers to buyer-style prompts in its research corpus (1 observation for this page). Metrics are distributions over observations, not a single static ranking.

Which AI systems does Obsurfable collect answers from?

OpenAI, ChatGPT, Google, Gemini, Google AI Mode, Anthropic, Claude, Perplexity, Grok, DeepSeek, Mistral, Copilot, and Meta AI.