Category: Model families

This category contains 13 pages.

  • Claude (model family) · Anthropic's assistant model line, from Claude 1 (2023) through the Claude 4.x series to the Fable/Mythos 5 generation (2026).
  • DeepSeek (model family) · DeepSeek's open-weight model line, known for efficiency innovations (MLA, DeepSeekMoE) and the R1 reasoning model.
  • Gemini (model family) · Google DeepMind's flagship multimodal model line, successor to LaMDA and PaLM; natively multimodal since Gemini 1.0 (2023).
  • GLM (model family) · Zhipu AI's model line, from the GLM-130B open release of 2022 through the agentic GLM-4.5/5 generation.
  • GPT (model family) · OpenAI's Generative Pre-trained Transformer series, from GPT-1 (2018) through GPT-4 and successors.
  • Grok (model family) · xAI's model line, distributed through the X platform; known for rapid scaling on the Colossus cluster and a lighter-refusal posture.
  • Kimi (model family) · Moonshot AI's model line: the long-context Kimi assistant, k1.5 reasoning, and the trillion-parameter open-weight K2 series.
  • Llama (model family) · Meta's open-weight model line; the LLaMA leak of 2023 seeded the modern open-source LLM ecosystem.
  • Mistral (model family) · Mistral AI's model line: efficient open-weight releases (Mistral 7B, Mixtral) alongside commercial Large and specialist models.
  • Moondream (model family) · The tiny-VLM line from M87 Labs: Moondream 1 (2024) through the MoE Moondream 3 Preview, built for edge and high-volume vision work.
  • Phi (model family) · Microsoft's small-model line built on the 'textbooks are all you need' thesis: curated and synthetic data over parameter count.
  • Qwen (model family) · Alibaba's open-weight model line; by 2025 the most-downloaded and most-fine-tuned base family in the open ecosystem.
  • Trinity (model family) · Arcee AI's open-weight sparse mixture-of-experts line: edge-scale Nano, agent-focused Mini, the 400B Trinity Large, and a Thinking reasoning branch.