New: connect Claude & other AIs to GenAIList over MCP โ research the catalog and contribute to the shared knowledge base. Learn how โ
Large language models (LLMs)
Browse every major large language model in one place. This LLM list tracks frontier and open-source foundation models โ GPT, Claude, Gemini, Llama, Mistral, Qwen and 400+ more โ with parameter counts, context windows, benchmark scores and provider pricing.
WizardLM-2 70B
๐บ๐ธ Microsoft
WizardLM-2 7B
๐บ๐ธ Microsoft
WizardLM-2 8x22B
๐บ๐ธ Microsoft
DDIM
๐บ๐ธ University Paris-Saclay
DDPM
๐บ๐ธ University Paris-Saclay
Bencao Zhiku
๐จ๐ณ Chengdu University of Traditional Chinese Medicine
HGRN2 1B
๐จ๐ณ Shanghai AI Lab
HGRN2 3B
๐จ๐ณ Shanghai AI Lab
tsuzumi 7B upgrade 2024
๐ฏ๐ต NTT Communication Science Laboratories
AF2RAVE
๐บ๐ธ University of Maryland
Zephyr 141B-A39B
๐บ๐ธ Hugging Face
DiffBindFR
๐จ๐ณ Peking University
OpenThaiGPT v1.0.0 (13B)
๐น๐ญ Mahidol University
OpenThaiGPT v1.0.0 (7B)
๐น๐ญ Mahidol University
SambaLingo-Thai-Chat (7B)
๐บ๐ธ "SambaNova Systems
SambaLingo-Thai-Chat-70B
๐บ๐ธ "SambaNova Systems
Stable LM 2 12B
๐บ๐ธ Stability AI
YaART
๐ท๐บ Yandex
ESM-AA
๐จ๐ณ Peking University
Command R+
๐จ๐ฆ Cohere
Sailor-7B-Chat
๐ธ๐ฌ Sea AI Lab
Viking
๐ซ๐ฎ Silo AI
eFold
๐บ๐ธ Harvard Medical School
AutoDiff
๐บ๐ธ Galixir Technologies
Mixture-of-Depths
๐บ๐ธ Google DeepMind
XVERSE-MoE-A4.2B
๐จ๐ณ XVERSE Technology
Youyuanjian (้ฎ่ฟ่ง)
๐จ๐ณ "China Post Consumer Finance Co.
Le_Triomphant-ECE-TW3
๐ซ๐ท French Engineering School ECE
Minerva 1B
๐ฎ๐น Sapienza NLP
Minerva 3B
๐ฎ๐น Sapienza NLP
TW3-JRGL-v2
๐ซ๐ท French Engineering School ECE
TeleChat-12B
๐จ๐ณ China Telecom
TeleChat-3B
๐จ๐ณ China Telecom
TeleChat-7B
๐จ๐ณ China Telecom
Volare
๐ฎ๐น Moxoff
ReALM
๐บ๐ธ Apple
Grok-1.5
๐บ๐ธ xAI
Jamba
๐ฎ๐ฑ AI21 Labs
YandexGPT 3
๐ท๐บ Yandex
DBRX
๐บ๐ธ Databricks
MultiVerse 70B
๐ท๐บ MTS
BindDM
๐บ๐ธ Peng Cheng Laboratory
CrossBind
๐จ๐ณ Shanghai AI Lab
ProstT5
๐บ๐ธ Technical University of Munich
Xuanji Yuheng (็็็่กก)
๐จ๐ณ Zhuoshi Technology
JetFire (GPT2-LARGE)
๐จ๐ณ Tsinghua University
MiniGPT4 + LRV-Instruction
๐บ๐ธ University of Maryland
ERNIE-RNA
๐บ๐ธ Microsoft Research
PocketVec
๐ช๐ธ Barcelona Institute of Science and Technology
Quiet-STaR
๐บ๐ธ Stanford University
Recraft V2 (Recraft 20B)
๐ฌ๐ง Recraft
Command R
๐จ๐ฆ Cohere
Dream Home LLM (่ดๅฃณๆขฆๆณๅฎถๅคงๆจกๅ)
๐จ๐ณ KE Holdings Inc. (โBeikeโ)
Inflection-2.5
๐บ๐ธ Inflection AI
BaseFold
๐ฌ๐ง Basecamp Research
Aramco Metabrain AI
๐ธ๐ฆ Saudi Aramco
Claude 3 Haiku
๐บ๐ธ Anthropic
Azzurro
๐ฎ๐น Moxoff
Maestrale Chat v0.4
๐ฎ๐น mii-llm
Mistral ITA 7B
๐ฎ๐น DeepMount00
About Large language models (LLMs)
Large language models (LLMs) are the foundation of modern generative AI โ general-purpose text models trained on vast corpora that can write, reason, summarise, translate and code. Choosing the best LLM is rarely about a single winner: the best AI model for one task may lag on another. Reasoning-heavy work rewards models that score well on benchmarks like MMLU, GPQA and AIME, while agentic and tool-use workloads care more about instruction following, function calling and long-context recall. When you compare LLMs, weigh raw capability against the practical constraints that decide cost and feasibility โ context window, throughput, latency, licensing and price per million tokens. Open-source LLMs such as Llama, Qwen, Mistral and DeepSeek let you self-host and fine-tune, while proprietary frontier models from OpenAI, Anthropic and Google often lead on raw quality. Use our benchmarks to see where each model ranks, and put two candidates side by side with compare before you commit to a provider.
Frequently asked questions
What is the best LLM right now?
There is no single best LLM โ it depends on the task. Frontier proprietary models from OpenAI, Anthropic and Google tend to lead on reasoning benchmarks, while open-source LLMs like Llama, Qwen and DeepSeek are best when you need to self-host or fine-tune. Compare candidates on the benchmarks page for your specific workload.
What is the best open source LLM?
The strongest open-source and open-weights LLMs at any given time typically come from the Llama, Qwen, Mistral and DeepSeek families. They can be downloaded, self-hosted and fine-tuned, and the top ones rival proprietary models on many benchmarks. Filter the list above by availability to see current open-weights options.
How do I compare two LLMs?
Use the compare tool to put two models side by side on parameters, context window, availability and benchmark scores, then check the benchmarks page for task-specific rankings such as MMLU, GPQA and coding scores.