New: connect Claude & other AIs to GenAIList over MCP โ research the catalog and contribute to the shared knowledge base. Learn how โ
Large language models (LLMs)
Browse every major large language model in one place. This LLM list tracks frontier and open-source foundation models โ GPT, Claude, Gemini, Llama, Mistral, Qwen and 400+ more โ with parameter counts, context windows, benchmark scores and provider pricing.
llama-3-airoboros-70b-3.3
Unknown
Fugaku-LLM
๐ฏ๐ต Tohoku University
MatterSim (Grpaphomer)
๐บ๐ธ Microsoft Research AI for Science
MatterSim (M3GNet - MatterSim-v1.0.0-5M)
๐บ๐ธ Microsoft Research AI for Science
Falcon 2 11B
๐ฆ๐ช Technology Innovation Institute
AlphaFold 3
๐บ๐ธ Google DeepMind
BiosimDock
๐บ๐ธ DeepOrigin
Amazon Titan Text Premier
๐บ๐ธ Amazon
DeepSeek-V2 (MoE-236B)
๐จ๐ณ DeepSeek
xLSTM 1.4B
๐ฆ๐น Johannes Kepler University Linz
Med-Gemini-2D
๐บ๐ธ Google DeepMind
Med-Gemini-3D
๐บ๐ธ Google DeepMind
Microsoft MAI-1
๐บ๐ธ Microsoft
MetaMath 70B
๐จ๐ณ University of Cambridge
MetaMath 7B (LLaMa finetune)
๐จ๐ณ University of Cambridge
MetaMath 7B (Mistral finetune)
๐จ๐ณ University of Cambridge
OpenELM-1.1B
๐บ๐ธ Apple
OpenELM-270M
๐บ๐ธ Apple
OpenELM-3B
๐บ๐ธ Apple
OpenELM-450M
๐บ๐ธ Apple
DanteLLM 7B
๐ฎ๐น RSTLess Research
GenCast
๐บ๐ธ Google DeepMind
LLaMAntino-3 ANITA 8B
๐ฎ๐น University of Bari
Llama-3 8B ITA
๐ฎ๐น DeepMount00
Med-Gemini-M 1.5
๐บ๐ธ Google DeepMind
Amazon Q Developer
๐บ๐ธ Amazon
DiffPepBuilder
๐จ๐ณ Peking University
Multi-Token Prediction 13B
๐บ๐ธ Facebook AI Research
Multi-Token Prediction 7B
๐บ๐ธ Facebook AI Research
Llama 3-TAIDE-LX-8B-Chat-Alpha1
๐น๐ผ National Science and Technology Council
TAIDE LX-13B
๐น๐ผ National Science and Technology Council
TAIDE-LX-7B
๐น๐ผ National Science and Technology Council
Swallow
๐ฏ๐ต Tokyo Institute of Technology
Qwen1.5-110B
๐จ๐ณ Qwen
Arctic
๐บ๐ธ Snowflake
Beyond ESM2: Graph-Enhanced Protein Sequence Modeling with Efficient
๐จ๐ณ Huazhong University of Science and Technology
NEC cotomi
๐ฏ๐ต NEC Corporation
Yuanjing LLM (่้ๅ ๆฏๅคงๆจกๅ)
๐จ๐ณ China Unicom
Firefly Image 3
๐บ๐ธ Adobe
Phi-3.5-MoE
๐บ๐ธ Microsoft
SI-PLM
๐บ๐ธ University of Pittsburgh
SenseChat 5.0
๐จ๐ณ SenseTime
phi-3-medium 14B
๐บ๐ธ Microsoft
phi-3-mini 3.8B
๐บ๐ธ Microsoft
phi-3-small 7.4B
๐บ๐ธ Microsoft
phi-3.5-mini
๐บ๐ธ Microsoft
VISTA-2D
๐บ๐ธ NVIDIA
InstructPLM
๐จ๐ณ Zhejiang Lab
SaProt
๐บ๐ธ Zhejiang University (ZJU)
FRED-T5-XL
๐ท๐บ Sber
LLaMA-3-Instruct-8B
๐บ๐ธ Meta AI
Llama 3-70B
๐บ๐ธ Meta AI
Llama 3-8B
๐บ๐ธ Meta AI
GRITLM 7B
๐บ๐ธ Contextual AI
GRITLM 8x7B
๐บ๐ธ Contextual AI
LINGO-2
๐ฌ๐ง Wayve
METL-Global
๐บ๐ธ University of Wisconsin Madison
Mixtral 8x22B
๐ซ๐ท Mistral AI
Tiangong 3.0 (MoE)
๐จ๐ณ Kunlun Inc.
abab6.5
๐จ๐ณ MiniMax
About Large language models (LLMs)
Large language models (LLMs) are the foundation of modern generative AI โ general-purpose text models trained on vast corpora that can write, reason, summarise, translate and code. Choosing the best LLM is rarely about a single winner: the best AI model for one task may lag on another. Reasoning-heavy work rewards models that score well on benchmarks like MMLU, GPQA and AIME, while agentic and tool-use workloads care more about instruction following, function calling and long-context recall. When you compare LLMs, weigh raw capability against the practical constraints that decide cost and feasibility โ context window, throughput, latency, licensing and price per million tokens. Open-source LLMs such as Llama, Qwen, Mistral and DeepSeek let you self-host and fine-tune, while proprietary frontier models from OpenAI, Anthropic and Google often lead on raw quality. Use our benchmarks to see where each model ranks, and put two candidates side by side with compare before you commit to a provider.
Frequently asked questions
What is the best LLM right now?
There is no single best LLM โ it depends on the task. Frontier proprietary models from OpenAI, Anthropic and Google tend to lead on reasoning benchmarks, while open-source LLMs like Llama, Qwen and DeepSeek are best when you need to self-host or fine-tune. Compare candidates on the benchmarks page for your specific workload.
What is the best open source LLM?
The strongest open-source and open-weights LLMs at any given time typically come from the Llama, Qwen, Mistral and DeepSeek families. They can be downloaded, self-hosted and fine-tuned, and the top ones rival proprietary models on many benchmarks. Filter the list above by availability to see current open-weights options.
How do I compare two LLMs?
Use the compare tool to put two models side by side on parameters, context window, availability and benchmark scores, then check the benchmarks page for task-specific rankings such as MMLU, GPQA and coding scores.