Introducing MPT-7B: A New Standard for Open-Source, Commercially Usable LLMs

1270 shaares
211 private links

1270 shaares · 211 private links

Filters

Links per page

20 50 100

Introducing MPT-7B: A New Standard for Open-Source, Commercially Usable LLMs

Introducing MPT-7B, the latest entry in our MosaicML Foundation Series. MPT-7B is a transformer trained from scratch on 1T tokens of text and code. It is open source, available for commercial use, and matches the quality of LLaMA-7B. MPT-7B was trained on the MosaicML platform in 9.5 days with zero human intervention at a cost of ~$200k. Starting today, you can train, finetune, and deploy your own private MPT models, either starting from one of our checkpoints or training from scratch. For inspiration, we are also releasing three finetuned models in addition to the base MPT-7B: MPT-7B-Instruct, MPT-7B-Chat, and MPT-7B-StoryWriter-65k+, the last of which uses a context length of 65k tokens!

ml · llm · ai

May 6, 2023 at 18:36:40 EDT * · permalink

https://www.mosaicml.com/blog/mpt-7b

Filters

Links per page

20 50 100