456B parameter MoE model (45.9B active) with up to 4M token inference context using hybrid Lightning+Softmax attention. Open-weight.
No takes yet.
Try another filter or be the first to share your honest opinion.