This is a wild result.
Locus, the automated research system from @intology, post-trained Qwen3 base models that beat the official human-tuned Qwen3 1.7B Instruct release.
SoTA on PostTrainBench!
Locus 自动研究系统对 Qwen3 基础模型进行后训练,性能超越官方人工调优的 Qwen3 1.7B Instruct 版本,并在 PostTrainBench 上达到 SOTA。在扩展算力预算的 PostTrainBench+ 测试中,Locus 扩展性最佳,其训练模型整体超越人工后训练的 Qwen3 1.7B。此外,Locus 在 Kaggle 实时竞赛中 16 天取得平均排名第 4。
This is a wild result.
Locus, the automated research system from @intology, post-trained Qwen3 base models that beat the official human-tuned Qwen3 1.7B Instruct release.
SoTA on PostTrainBench!
Locus 自动研究系统对 Qwen3 基础模型进行后训练,性能超越官方人工调优的 Qwen3 1.7B Instruct 版本,并在 PostTrainBench 上达到 SOTA。在扩展算力预算的 PostTrainBench+ 测试中,Locus 扩展性最佳,其训练模型整体超越人工后训练的 Qwen3 1.7B。此外,Locus 在 Kaggle 实时竞赛中 16 天取得平均排名第 4。
This is a wild result.
Locus, the automated research system from @intology, post-trained Qwen3 base models that beat the official human-tuned Qwen3 1.7B Instruct release.
SoTA on PostTrainBench!