Databricks has unveiled DBRX, a groundbreaking open-source large language model (LLM) that it asserts establishes a new benchmark for such models. Surpassing established options like GPT-3.5 on industry benchmarks, DBRX, with its 132 billion parameters, claims superiority over popular open-source LLMs such as LLaMA 2 70B, Mixtral, and Grok-1 across various tasks including language understanding, programming, and mathematics. Notably, it even outshines Anthropic’s closed-source model Claude on specific benchmarks. In coding tasks, DBRX exhibits state-of-the-art performance among open models, surpassing specialized models like CodeLLaMA despite its general-purpose nature. Furthermore, it either matches or exceeds GPT-3.5 across nearly all evaluated benchmarks. The remarkable capabilities of DBRX are attributed to its more efficient mixture-of-experts architecture, enabling it to achieve up to twice the speed of inference compared to LLaMA 2 70B, despite having fewer active parameters. ...
Comments
Post a Comment