DBRX

Databricks' 132B mixture-of-experts LLM with 36B active parameters, released with open weights in 2024.

Open SourceSelf HostedOffline CapableGPU Required
0.0 (0)

About

DBRX was Databricks' statement release in March 2024: a 132B-parameter mixture-of-experts LLM with 36B active parameters on any input, pretrained on 12T tokens of text and code with a 32K maximum context, using a fine-grained MoE architecture built on the MegaBlocks research project. Base and Instruct variants are downloadable from Hugging Face under the Databricks Open Model License and an accompanying acceptable use policy, which allow commercial use subject to their terms, and the model is also distributed through the Databricks Marketplace. DBRX Instruct specializes in few-turn interactions and was aimed squarely at enterprise workloads, with Databricks publishing benchmarks against contemporary open models at launch. Running it is a serious infrastructure exercise given the parameter count, and the original GitHub repository has since been taken down, leaving the Hugging Face model cards as the primary artifact. The release matters historically as one of the first demonstrations that an enterprise data platform vendor could train a frontier-scale open-weight MoE model entirely in-house.

Reviews (0)

Leave a Review

No reviews yet. Be the first to review!

Details

Price
Free
Platform
Local/Desktop
Difficulty
Advanced (4/5)
License
Databricks Open Model License
Added
Jul 29, 2026

Related Tools

Featured

Lightweight open-weight LLM by Google available in 1B to 27B sizes.

Open SourceSelf HostedOfflineGPU 4GB+
Easy
0.0 (0)

Open-source code LLM family by IBM for enterprise code generation.

Open SourceSelf HostedOfflineGPU 8GB+
Intermediate
0.0 (0)

Open-weight LLM by Meta in 8B and 70B sizes with strong general capabilities.

Open SourceSelf HostedOfflineGPU 8GB+
Intermediate
0.0 (0)
Featured

High-performance open-weight MoE LLM with 671B total parameters.

Open SourceSelf HostedOfflineGPU 24GB+
Advanced
0.0 (0)

Hybrid SSM-Transformer model by AI21 Labs combining Mamba with attention layers.

Open SourceSelf HostedOfflineGPU 12GB+
Intermediate
0.0 (0)

Open-weight code LLM trained on 2 trillion tokens of code and natural language.

Open SourceSelf HostedOfflineGPU 8GB+
Intermediate
0.0 (0)
Browse all Large Language Models (LLMs) tools