Ai2 releases Olmo-core 3 to make developing large mixture-of-experts LLMs more efficient

Covered by 3 sources
Read the full article
The Allen Institute for AI released Olmo-core 3, a framework that improves training efficiency for large mixture-of-experts language models, enabling scaling to over one trillion parameters with lower costs. Olmo-core 3 achieves 2.7 times higher throughput than Nvidia's Megatron-core by using expert parallelism and distributed optimization across GPUs.

Covered by 3 sources


More in Machine Learning

Machine Learning2 min read

Google rolls out new Gemini AI model but restricts access over safety concerns

Covered by 11 sources
Machine Learning3 min read

Robotics startup FieldAI is set to raise $700 million at a $10 billion valuation

Covered by 2 sources
Machine Learning2 min read

Sean Parker is rebuilding Stability AI around music

TechCrunch