Category: 2022 model releases
This category contains 16 pages.
- BLIP · Salesforce Research's vision-language pre-training family (2022-2023), whose BLIP-2 Q-Former popularized bridging frozen image encoders to frozen LLMs.
- BLOOM · BigScience's 176-billion-parameter open-access multilingual language model, released in July 2022 as a large-scale collaborative science project coordinated by Hugging Face.
- Chinchilla · DeepMind's 2022 compute-optimal model: 70B parameters trained on 1.4T tokens, proving the GPT-3 generation was undertrained and resetting industry training budgets.
- DALL-E 2 · OpenAI's April 2022 text-to-image diffusion model, which generated 1024x1024 images from CLIP embeddings and popularized AI image generation.
- Flamingo · DeepMind's April 2022 visual language model that pioneered few-shot multimodal prompting by bridging frozen vision and language models.
- Flan-T5 · Google's October 2022 family of instruction-finetuned T5 models, released openly in five sizes and a standard baseline for instruction-tuning research.
- GLM-130B · A 130-billion-parameter open bilingual (English-Chinese) language model released by Tsinghua University and Zhipu AI in August 2022.
- GPT-3.5 · OpenAI's 2022 model series bridging GPT-3 and GPT-4; the original engine of ChatGPT.
- GPT-NeoX-20B · EleutherAI's February 2022 open-source 20-billion-parameter autoregressive language model, at release the largest publicly available dense model with fully open weights.
- Imagen · Google's text-to-image diffusion model line, introduced in May 2022 and continued through Imagen 2, 3, and 4.
- InstructGPT · OpenAI's 2022 RLHF-aligned GPT-3 variants; demonstrated that alignment training beats scale on human preference and set the template for ChatGPT.
- OPT · Meta AI's May 2022 suite of open decoder-only language models, from 125M to 175B parameters, released with weights, code, and a candid training logbook.
- PaLM · Google's 540-billion-parameter Pathways Language Model of April 2022, a landmark in dense scaling and few-shot reasoning.
- Stable Diffusion · Open-weight latent diffusion text-to-image model family released in August 2022 by Stability AI with CompVis and Runway, spanning v1 through Stable Diffusion 3.5.
- Whisper · OpenAI's open-source speech recognition model of September 2022, trained on 680,000 hours of weakly supervised multilingual audio.
- YaLM-100B · Yandex's June 2022 open 100-billion-parameter GPT-like language model, at release the largest dense model with freely downloadable Apache 2.0 weights.