DBRX: The new best open LLM and Databricks' ML strategy

Topics covered
Popular Clips
Episode Highlights
Efficiency
Nathan Lambert discusses the efficiency trends observed with Databricks' DBRX model. He highlights the significant performance gains achieved through meticulous training and engineering efforts. Lambert notes, "Our MoEs are two times more efficient than modern non-MoEs. Our data is two times more token efficient than for MPT. Inference is up to two times faster versus Llama 270B, up to 150 tokens per second" 1. These advancements suggest that GPT-4 level models could become effectively free within a decade.
  Â
Evaluation
Lambert also delves into the evaluation methods applied to DBRX, both qualitative and quantitative. He shares his experience testing the model's knowledge and limits, confirming that DBRX instruct is a solid sub-GPT-4 model. Lambert's attempts to jailbreak the model revealed consistent refusals, indicating robust safety filters 2 3. He concludes that while DBRX has some limitations, it remains a top contender in the open LLM space.
Related Episodes

The end of the "best open LLM"
Answers 383 questions

A recipe for frontier model post-training
Answers 383 questions
Open Language Models (OLMos) and the LLM landscape
Answers 383 questions
Llama 3.1 405b, Meta's AI strategy, and the new open frontier model ecosystem
Answers 383 questions
Big Tech's LLM evals are just marketing
Answers 383 questions
We aren't running out of training data, we are running out of open training data
Answers 383 questions
Google ships it: Gemma open LLMs and Gemini backlash
Answers 383 questions
Interconnects year in review: 2023
Answers 383 questions
Model merging lessons in The Waifu Research Department
Answers 383 questions

Futures of the data foundry business model
Answers 383 questions
Llama 3: Scaling open LLMs to AGI
Answers 383 questions
SB 1047, AI regulation, and unlikely allies for open models
Answers 383 questions
Local LLMs, some facts some fiction
Answers 383 questions
Alignment-as-a-Service: Scale AI vs. the new guys
Answers 383 questions
