News
AI Summary
28 Sept 202617 Rabiʻ II 1448 AH
NaiveAI Open-Weights Naive-N0.5-Flash: 309B MoE, Hybrid SWA–DSA 1M Context, MIT

NaiveAI Open-Weights Naive-N0.5-Flash: 309B MoE, Hybrid SWA–DSA 1M Context, MIT

NaiveAI announced the launch of the Naive-N0.5-Flash model on Hugging Face, aimed at coding and AI research. The model features approximately 309 billion parameters, with 15.5 billion active, and is MIT-licensed. It employs a hybrid architecture of Sliding-Window Attention and DeepSeek Sparse Attention, enabling it to handle contexts of up to one million tokens.

Follow these topics

Sign in to follow the topics that matter to you

Sign in to follow

This summary is generated with AI and receives periodic editorial review. Refer to the original source for full details.

0
0 reading now

Insight Score

Rate to unlock

Sign in to react, rate, and save. Sign In