Category: 2024 model releases
This category contains 51 pages.
- Amazon Nova · Amazon's in-house foundation-model family (December 2024): the Micro/Lite/Pro ladder plus creative models, distributed exclusively through Bedrock.
- Apple foundation models · Apple's family of proprietary language models (an approximately 3B-parameter on-device model plus a larger server model) powering Apple Intelligence, announced June 2024.
- Aya · Cohere For AI's open multilingual model family (2024-2025), built with thousands of volunteer contributors to extend instruction-tuned LLMs beyond English.
- Claude 3 · Anthropic's March 2024 generation (Haiku, Sonnet, Opus); Opus was the first model to consistently challenge GPT-4-class results.
- Claude 3.5 Haiku · Anthropic's fast, low-cost model of the Claude 3.5 generation, announced in October 2024 as the successor to Claude 3 Haiku.
- Claude 3.5 Sonnet · Anthropic's mid-2024 workhorse; beat Claude 3 Opus at twice the speed, and its October refresh introduced computer use.
- Codestral · Mistral AI's May 2024 code-generation model, a 22-billion-parameter weight-available system trained on more than 80 programming languages.
- Command (models) · Cohere's enterprise model line; Command R (2024) made retrieval-augmented generation and tool use the product, and Command A (2025) optimized for two-GPU serving.
- DBRX · Databricks' March 2024 open-weight mixture-of-experts language model, with 132 billion total and 36 billion active parameters.
- DeepSeek-Coder-V2 · DeepSeek's June 2024 open-weight mixture-of-experts code model, developer-reported as competitive with closed frontier models on coding benchmarks.
- DeepSeek-V2 · DeepSeek's May 2024 open MoE that introduced multi-head latent attention and DeepSeekMoE; started China's LLM price war.
- DeepSeek-V3 · DeepSeek's December 2024 open-weight flagship: 671B-parameter MoE reporting frontier quality from a ~$5.6M disclosed final training run.
- Doubao · ByteDance's assistant model line; China's largest chatbot user base, priced to force market-wide cuts.
- FLUX · Black Forest Labs' family of rectified flow transformer text-to-image models, launched in August 2024 with open-weight and proprietary variants.
- Gemini 1.5 · Google DeepMind's February 2024 generation; made million-token context windows a product reality and introduced the Flash tier.
- Gemini 2.0 · Google DeepMind's December 2024 'agentic era' generation: Flash-first release, native tool use, and the Flash Thinking reasoning experiments.
- Gemma · Google DeepMind's open-weight line drawing on Gemini research: Gemma 1 (2024) through the multimodal Gemma 3 (2025).
- Genie · Google DeepMind's series of generative interactive world models (Genie, Genie 2, Genie 3) that produce playable environments from images or text.
- GLM-4 · Zhipu AI's January 2024 flagship language model generation, spanning a proprietary API flagship and later open-weight GLM-4-9B variants.
- GPT-4o · OpenAI's May 2024 'omni' model: natively multimodal across text, vision, and audio, with real-time voice; became ChatGPT's long-running default.
- GPT-4o mini · OpenAI's July 2024 small multimodal model, a low-cost successor to GPT-3.5 Turbo in ChatGPT and the API.
- Grok-1.5 · xAI's March 2024 successor to Grok-1, adding a 128,000-token context window and developer-reported gains in math and coding.
- Grok-2 · xAI's August 2024 frontier chatbot model, released in beta on the X platform alongside a smaller Grok-2 mini variant.
- Hailuo · MiniMax's consumer AI brand, covering the Hailuo AI assistant app and the Hailuo video-generation model line (video-01, Hailuo 02).
- Hunyuan · Tencent's model line: the open Hunyuan-Large MoE (2024), the leading open video model HunyuanVideo, and the Apache-2.0 Hunyuan 3.0 (2026).
- HunyuanVideo · Tencent's December 2024 open-weight text-to-video generation model, a roughly 13-billion-parameter diffusion transformer released with code and weights.
- IBM Granite · IBM's family of open-weight enterprise language models, spanning dense, mixture-of-experts, code, and hybrid Mamba/transformer variants.
- Jamba · AI21 Labs' 2024 hybrid: the first production-scale model interleaving Transformer attention with Mamba state-space layers.
- Kling · Kuaishou's text-to-video generation model, launched in June 2024 as one of the first widely accessible answers to OpenAI's Sora.
- Llama 3 · Meta's 2024 generation: 8B/70B in April, the 405B frontier-scale release in July; the herd paper documented open training at frontier scale.
- Llama 3.1 · Meta AI's July 2024 open-weight model family (8B, 70B, 405B), whose 405B variant was widely described as the first open-weight model competitive with frontier closed systems.
- Llama 3.2 · Meta AI's September 2024 Llama release adding small on-device text models (1B, 3B) and the family's first vision-capable models (11B, 90B).
- Llama 3.3 · Meta AI's December 2024 open-weight 70B instruct model, developer-reported to approach Llama 3.1 405B quality at a fraction of the serving cost.
- Mistral Large · Mistral AI's flagship proprietary language model line, launched February 2024 and updated in July 2024 as the 123-billion-parameter Mistral Large 2.
- Mistral Small · Mistral AI's mid-size model line, launched February 2024 and later released open-weight, positioned for low-latency deployment below Mistral Large.
- Mixtral 8x22B · Mistral AI's April 2024 sparse mixture-of-experts model with about 141B total and 39B active parameters, released under Apache 2.0.
- Molmo · Ai2's September 2024 family of open-weight vision-language models trained on the human-annotated PixMo dataset rather than synthetic captions distilled from proprietary systems.
- Movie Gen · Meta AI's October 2024 suite of media foundation models for text-to-video and synchronized audio generation, shown as a research preview without a public release.
- Nemotron · NVIDIA's family of open-weight language models, best known for the June 2024 Nemotron-4 340B release aimed at synthetic data generation.
- o1 · OpenAI's first reasoning model (September 2024): trained with RL to think in long chains before answering, opening the test-time compute era.
- o3 · OpenAI's frontier reasoning model, previewed in December 2024 and released in April 2025, noted for developer-reported gains on ARC-AGI, math, and coding benchmarks.
- OLMo · AI2's fully open model line: weights, data, code, and checkpoints all public; the scientific control group of the LLM era.
- PaliGemma · Google's open vision-language model of May 2024, pairing a SigLIP vision encoder with a Gemma language decoder and designed for fine-tuning on downstream tasks.
- Pixtral · Mistral AI's first natively multimodal model line, opened in September 2024 with the Apache-2.0 Pixtral 12B vision-language model.
- Qwen2 · Alibaba's June 2024 open-weight model series spanning 0.5B to 72B parameters, including one mixture-of-experts variant.
- Qwen2.5 · Alibaba's September 2024 open-weight model series (0.5B to 72B), trained on a developer-reported 18 trillion tokens and widely used as a base for fine-tunes and reasoning distillations.
- SmolLM · Hugging Face's family of small open-weight language models (135M to 3B parameters), launched in July 2024 and built on openly documented training corpora.
- Sora · OpenAI's text-to-video generation model, previewed in February 2024 and built on a diffusion transformer over spacetime patches.
- Step-2 · StepFun's 2024 trillion-parameter-scale mixture-of-experts language model, one of the first Chinese models announced at that scale.
- Tulu 3 · Ai2's November 2024 family of openly post-trained Llama 3.1 derivatives, released with its full recipe and known for introducing reinforcement learning with verifiable rewards (RLVR).
- Veo · Google DeepMind's flagship text-to-video generation model series, first announced in May 2024 and extended with native audio in Veo 3.