Page cover
For the complete documentation index, see llms.txt. This page is also available as Markdown.

Large Language Models (LLMs)

Navigation:

Large Language Models (LLMs) Ranking

Independent analysis of AI language models and API. Provides quality, speed, and price comparisons.

LMArena.ai is a comprehensive AI model leaderboard, ranking over 240 large language models across various domains—including text, vision, web development, search, and coding—based on more than 3.5 million user votes.

Unlike traditional benchmarks, LMArena employs a crowdsourced, blind-voting system where users compare anonymous model responses to the same prompt and vote for the better one. This approach provides a dynamic, real-world evaluation of model performance.

Comparison

AI Model
Core Strength (2026)
Context Window
Agent Capability
Starting Price

ChatGPT (GPT-5.2)

Generalist & Logic (o3)

256K Tokens

Operator (Browser Agent)

$20/mo

Claude 4 (Opus/Sonnet)

Human-like Writing & Code

500K+ Tokens

Computer Use (Desktop)

$20/mo

Gemini 3 (Ultra)

Google Ecosystem & Video

2M - 10M Tokens

Google Workspace Agent

$19.99/mo

Grok 4

Real-time X Data & Wit

128K+ Tokens

Integrated Social Agent

$8/mo (X Premium)


Large Language Models

Close Source Apps:

Open Source Models:


Learning Resources:

How I use LLMs by Andrej Karpathy

LLM101n: Let's build a Storyteller

Generative AI Handbook: A Roadmap for Learning Resources

Deep Dive into LLMs like ChatGPT by Andrej Karpathy

Google - Prompt Engineering by Lee Boonstra

Last updated