Closed-weight language models in 2017
13
Models
11 organisations
0%
Open weights
0 open, 13 closed
0
Frontier
By Epoch's flag
9B
Largest traceable
MoE-Multi
2017
Span
8 with a traceable size
Releases
by monthOrganisation
Records
13| Model | Organisation | Released | Parameters | Weights |
|---|---|---|---|---|
| DL scaling LM | Baidu | 2017-12 | 177M * | Closed |
| AWD-LSTM-MoS + dynamic evaluation (WT2, 2017) | Carnegie Mellon University (CMU) | 2017-11 | 35M * | Closed |
| DCN+ | Salesforce Research | 2017-10 | — | Closed |
| Fraternal dropout + AWD-LSTM 3-layer (WT2) | Jagiellonian University | 2017-10 | 34M * | Closed |
| AWD-LSTM+WT+Cache+IOG (WT2) | NTT Communication Science Laboratories | 2017-09 | 53M | Closed |
| ISS | Duke University | 2017-09 | 11M | Closed |
| LSTM + dynamic eval | University of Edinburgh | 2017-09 | 50M * | Closed |
| AWD-LSTM - 3-layer LSTM (tied) + continuous cache pointer (WT2) | Salesforce Research | 2017-08 | 33M | Closed |
| EI-REHN-1000D | Korea Advanced Institute of Science and Technology (KAIST) | 2017-08 | 19M | Closed |
| GL-LWGC-AWD-MoS-LSTM + dynamic evaluation (WT2) | Ben-Gurion University of the Negev | 2017-08 | 38M | Closed |
| AWD-LSTM | DeepMind | 2017-07 | 24M | Closed |
| Transformer | Google Research | 2017-06 | 213M | Closed |
| MoE-Multi | Jagiellonian University | 2017-01 | 9B | Closed |
An asterisk marks a parameter figure the source does not consider traceable to the people who built the model. That judgement is the source’s and it has false negatives: Kimi K3 is marked speculative at 2.8T while Moonshot’s own model card states the figure. Why that distinction carries the whole exhibit.
Counts are of models Epoch AI records as notable, not of every model released. The most recent month is always incomplete, because a model enters the dataset when it is reviewed rather than when it ships.
Epoch AI, 'Data on AI Models'. Published online at epoch.ai. Retrieved from 'https://epoch.ai/data/ai-models-documentation' Used under CC BY 4.0.