{
  "id": 11479893,
  "title": "Ai2 releases Olmo-core 3 to make developing large mixture-of-experts LLMs more efficient",
  "url": "https://urgent.news/2026/10/02/ai2-releases-olmo-core-3-to-make-developing-large-mixture-of-experts",
  "topic": "ai",
  "section": "AI",
  "published": "2026-10-02T16:15:44.000Z",
  "source": {
    "name": "SiliconANGLE",
    "slug": "siliconangle",
    "url": "https://siliconangle.com/2026/10/02/ai2-releases-olmo-core-3-to-make-developing-large-mixture-of-experts-llms-more-efficient/"
  },
  "original_language": "en",
  "account": "Seattle-based AI research organization Allen Institute for AI introduced Olmo-core 3, a framework designed to simplify the development of large language models on Thursday. The framework improves the training process of mixture-of-experts large language models, enabling them to scale to a trillion parameters while maintaining computational efficiency. Unlike dense models that utilize the entire model for computations, mixture-of-experts models allocate computation across specialized portions of the model for each token generated. Olmo-core 3 facilitates the expansion of expert pools from eight to 128 while still activating only four experts per token, allowing LLMs to scale to over one trillion parameters. Benchmarks demonstrate that Olmo-core 3 processes 52,000 tokens per second on Nvidia B3000 GPUs for a 47-billion parameter model, a 2.7 times increase in throughput compared to Nvidia's established Megatron-core training architecture. The architecture employs expert parallelism, layer splitting, and distributed optimization to reduce memory overhead as models scale. Additionally, Ai2 supports MXFP8, a number format that can decrease computation and data movement between GPUs. The new training architecture and improved efficiency are part of Allen Institute's goal to provide researchers with tools to create and train larger models. The framework, along with related systems, is currently available on GitHub for developers and the open-source community.",
  "summary": "Seattle-based artificial intelligence research firm Allen Institute for AI announced a development framework for large language models Thursday that significantly improves how mixture-of-experts large language models are trained. The new framework, Olmo-core 3, allows MoE training to reach the trillion-parameter scale while keeping costs low by preserving computational efficiency.…",
  "key_points": [],
  "editors_take": null,
  "illustration": null,
  "coverage": {
    "outlets": 2,
    "also_reported_by": [
      {
        "outlet": "TechRadar",
        "title": "'It is possible that threat actors are finding it more accessible or efficient to use LLMs and AI tools': Google warns that AI explosion will lead to more dangerous and advanced security threats",
        "url": "https://urgent.news/2026/10/01/it-is-possible-that-threat-actors-are-finding-it-more-accessible-or",
        "published": "2026-10-01T13:15:00.000Z"
      }
    ]
  },
  "ai_generated": true,
  "disclaimer": "Summaries, key points and the editor’s take are written by software from other outlets’ reporting and may contain errors — always check the linked original."
}