Ai2 releases Olmo-core 3 to make developing large mixture-of-experts LLMs more efficient

SiliconANGLEen

Ai2 releases Olmo-core 3 to make developing large mixture-of-experts LLMs more efficient

Seattle-based artificial intelligence research firm Allen Institute for AI announced a development framework for large language models Thursday that significantly improves how mixture-of-experts large language models are trained. The new framework, Olmo-core 3, allows MoE training to reach the trillion-parameter scale while keeping costs low by preserving computational efficiency. Mixture-of-experts models operate differently from dense […] The post Ai2 releases Olmo-core 3 to make developing large mixture-of-experts LLMs more efficient appeared first on SiliconANGLE .

This is a short summary published by AI Global Wire. The full article is owned and hosted by SiliconANGLE — open it there to read it in full.

Read the full story at SiliconANGLE
  • Verktyg
  • Forskning

Related AI news