3.5
@hn_4bc5e5
29 days ago
The old Kimi K2 is better than GLM4.7

1T/32B-active MoE with 256K context and MLA attention, Modified MIT license; 384-expert (8/token) agentic-focused foundation model from Moonshot.
@hn_4bc5e5
29 days ago
The old Kimi K2 is better than GLM4.7