Biggerbrain
Collection
The collection of the biggerbrain lineup of SLMs. • 3 items • Updated
This model uses a Recurrent transformer(looping the central 5 layers), and soft averaged MoE architecture to achieve new levels of reasoning(reletive to the Biggerbrain lineup). This model used an improved training pipeline*, and a substantial parameter increase over V2, and a great boost to intelligence, focus, and overall usefullness.