search
Get Started
search
MPT-30B - Model
zoom_in Click to enlarge

MPT-30B

description MPT-30B Overview

MPT-30B is a 30-billion-parameter open-source large language model released by MosaicML in June 2023. It was trained on a massive dataset consisting of 1 trillion tokens and utilized the FlashAttention technique to optimize training and inference efficiency. The model features an 8,000-token context window and employs ALiBi (Attention with Linear Biases) for sequence length extrapolation. Distributed under the commercially permissive Apache 2.0 license, it is designed to be freely used, modified, and deployed by enterprise developers.

help MPT-30B FAQ

What is the MPT-30B language model?

MPT-30B is an open-source large language model developed by MosaicML. It features 30 billion parameters and was designed to offer commercial-grade performance to developers.

What makes MPT-30B's context window special?

The model supports an 8K-token context window, allowing it to process large documents and long text inputs in a single prompt. This was a highly sought-after feature for developers building coding or document-analysis tools.

When was MPT-30B released?

MosaicML released the MPT-30B model in June 2023. The release showcased the company's efficient training methods for large-scale transformer models before the company was acquired by Databricks.

Can MPT-30B be used for commercial purposes?

Yes, MPT-30B was released under a commercially permissive Apache 2.0 license. This allows businesses to freely integrate the model into proprietary, for-profit products without paying licensing fees.

Reviews & Comments

Write a Review

rate_review

Be the first to review

Share your thoughts with the community and help others make better decisions.

Save to your list

Save your favorites and follow how their scores change over time.

Save favorites
Track changes
Compare scores

Already have an account? Sign in

Compare Items

See how they stack up against each other

Comparing
VS
Select 1 more item to compare