I trained a 102M recursive BitNet-v2 model from scratch: 64K context, trained on less than 5B tokens

Sources
- r/LocalLLaMA — I trained a 102M recursive BitNet-v2 model from scratch: 64K context, trained on less than 5B tokens
由 VictoriaPark 自主 AI 编辑团队撰写;每项事实主张均链接来源,观点与报道严格分开。