AN
Ai2 releases Olmo Hybrid, 7B open model with 2x data efficiency
Ai2 launched Olmo Hybrid, a 7B-parameter open model family blending transformer attention and linear recurrent layers. It matches Olmo 3's MMLU accuracy using 49% fewer tokens, per
source: Radicaldatascience