Category: 2024 model releases

This category contains 51 pages.

  • Amazon Nova · Amazon's in-house foundation-model family (December 2024): the Micro/Lite/Pro ladder plus creative models, distributed exclusively through Bedrock.
  • Apple foundation models · Apple's family of proprietary language models (an approximately 3B-parameter on-device model plus a larger server model) powering Apple Intelligence, announced June 2024.
  • Aya · Cohere For AI's open multilingual model family (2024-2025), built with thousands of volunteer contributors to extend instruction-tuned LLMs beyond English.
  • Claude 3 · Anthropic's March 2024 generation (Haiku, Sonnet, Opus); Opus was the first model to consistently challenge GPT-4-class results.
  • Claude 3.5 Haiku · Anthropic's fast, low-cost model of the Claude 3.5 generation, announced in October 2024 as the successor to Claude 3 Haiku.
  • Claude 3.5 Sonnet · Anthropic's mid-2024 workhorse; beat Claude 3 Opus at twice the speed, and its October refresh introduced computer use.
  • Codestral · Mistral AI's May 2024 code-generation model, a 22-billion-parameter weight-available system trained on more than 80 programming languages.
  • Command (models) · Cohere's enterprise model line; Command R (2024) made retrieval-augmented generation and tool use the product, and Command A (2025) optimized for two-GPU serving.
  • DBRX · Databricks' March 2024 open-weight mixture-of-experts language model, with 132 billion total and 36 billion active parameters.
  • DeepSeek-Coder-V2 · DeepSeek's June 2024 open-weight mixture-of-experts code model, developer-reported as competitive with closed frontier models on coding benchmarks.
  • DeepSeek-V2 · DeepSeek's May 2024 open MoE that introduced multi-head latent attention and DeepSeekMoE; started China's LLM price war.
  • DeepSeek-V3 · DeepSeek's December 2024 open-weight flagship: 671B-parameter MoE reporting frontier quality from a ~$5.6M disclosed final training run.
  • Doubao · ByteDance's assistant model line; China's largest chatbot user base, priced to force market-wide cuts.
  • FLUX · Black Forest Labs' family of rectified flow transformer text-to-image models, launched in August 2024 with open-weight and proprietary variants.
  • Gemini 1.5 · Google DeepMind's February 2024 generation; made million-token context windows a product reality and introduced the Flash tier.
  • Gemini 2.0 · Google DeepMind's December 2024 'agentic era' generation: Flash-first release, native tool use, and the Flash Thinking reasoning experiments.
  • Gemma · Google DeepMind's open-weight line drawing on Gemini research: Gemma 1 (2024) through the multimodal Gemma 3 (2025).
  • Genie · Google DeepMind's series of generative interactive world models (Genie, Genie 2, Genie 3) that produce playable environments from images or text.
  • GLM-4 · Zhipu AI's January 2024 flagship language model generation, spanning a proprietary API flagship and later open-weight GLM-4-9B variants.
  • GPT-4o · OpenAI's May 2024 'omni' model: natively multimodal across text, vision, and audio, with real-time voice; became ChatGPT's long-running default.
  • GPT-4o mini · OpenAI's July 2024 small multimodal model, a low-cost successor to GPT-3.5 Turbo in ChatGPT and the API.
  • Grok-1.5 · xAI's March 2024 successor to Grok-1, adding a 128,000-token context window and developer-reported gains in math and coding.
  • Grok-2 · xAI's August 2024 frontier chatbot model, released in beta on the X platform alongside a smaller Grok-2 mini variant.
  • Hailuo · MiniMax's consumer AI brand, covering the Hailuo AI assistant app and the Hailuo video-generation model line (video-01, Hailuo 02).
  • Hunyuan · Tencent's model line: the open Hunyuan-Large MoE (2024), the leading open video model HunyuanVideo, and the Apache-2.0 Hunyuan 3.0 (2026).
  • HunyuanVideo · Tencent's December 2024 open-weight text-to-video generation model, a roughly 13-billion-parameter diffusion transformer released with code and weights.
  • IBM Granite · IBM's family of open-weight enterprise language models, spanning dense, mixture-of-experts, code, and hybrid Mamba/transformer variants.
  • Jamba · AI21 Labs' 2024 hybrid: the first production-scale model interleaving Transformer attention with Mamba state-space layers.
  • Kling · Kuaishou's text-to-video generation model, launched in June 2024 as one of the first widely accessible answers to OpenAI's Sora.
  • Llama 3 · Meta's 2024 generation: 8B/70B in April, the 405B frontier-scale release in July; the herd paper documented open training at frontier scale.
  • Llama 3.1 · Meta AI's July 2024 open-weight model family (8B, 70B, 405B), whose 405B variant was widely described as the first open-weight model competitive with frontier closed systems.
  • Llama 3.2 · Meta AI's September 2024 Llama release adding small on-device text models (1B, 3B) and the family's first vision-capable models (11B, 90B).
  • Llama 3.3 · Meta AI's December 2024 open-weight 70B instruct model, developer-reported to approach Llama 3.1 405B quality at a fraction of the serving cost.
  • Mistral Large · Mistral AI's flagship proprietary language model line, launched February 2024 and updated in July 2024 as the 123-billion-parameter Mistral Large 2.
  • Mistral Small · Mistral AI's mid-size model line, launched February 2024 and later released open-weight, positioned for low-latency deployment below Mistral Large.
  • Mixtral 8x22B · Mistral AI's April 2024 sparse mixture-of-experts model with about 141B total and 39B active parameters, released under Apache 2.0.
  • Molmo · Ai2's September 2024 family of open-weight vision-language models trained on the human-annotated PixMo dataset rather than synthetic captions distilled from proprietary systems.
  • Movie Gen · Meta AI's October 2024 suite of media foundation models for text-to-video and synchronized audio generation, shown as a research preview without a public release.
  • Nemotron · NVIDIA's family of open-weight language models, best known for the June 2024 Nemotron-4 340B release aimed at synthetic data generation.
  • o1 · OpenAI's first reasoning model (September 2024): trained with RL to think in long chains before answering, opening the test-time compute era.
  • o3 · OpenAI's frontier reasoning model, previewed in December 2024 and released in April 2025, noted for developer-reported gains on ARC-AGI, math, and coding benchmarks.
  • OLMo · AI2's fully open model line: weights, data, code, and checkpoints all public; the scientific control group of the LLM era.
  • PaliGemma · Google's open vision-language model of May 2024, pairing a SigLIP vision encoder with a Gemma language decoder and designed for fine-tuning on downstream tasks.
  • Pixtral · Mistral AI's first natively multimodal model line, opened in September 2024 with the Apache-2.0 Pixtral 12B vision-language model.
  • Qwen2 · Alibaba's June 2024 open-weight model series spanning 0.5B to 72B parameters, including one mixture-of-experts variant.
  • Qwen2.5 · Alibaba's September 2024 open-weight model series (0.5B to 72B), trained on a developer-reported 18 trillion tokens and widely used as a base for fine-tunes and reasoning distillations.
  • SmolLM · Hugging Face's family of small open-weight language models (135M to 3B parameters), launched in July 2024 and built on openly documented training corpora.
  • Sora · OpenAI's text-to-video generation model, previewed in February 2024 and built on a diffusion transformer over spacetime patches.
  • Step-2 · StepFun's 2024 trillion-parameter-scale mixture-of-experts language model, one of the first Chinese models announced at that scale.
  • Tulu 3 · Ai2's November 2024 family of openly post-trained Llama 3.1 derivatives, released with its full recipe and known for introducing reinforcement learning with verifiable rewards (RLVR).
  • Veo · Google DeepMind's flagship text-to-video generation model series, first announced in May 2024 and extended with native audio in Veo 3.