观点 / 方法普通
slop code bench 汇总 Fable 5.1、Sol 与 GLM 5.3 不同推理档位的编码评测结果
原始标题:many of the runs are incomplete, so will keep updating, but slop code bench coming together for Fabl…
内容摘要
Dex Horthy 发文称 slop code bench 正在成形,对比 Fable 5.1、Sol 的 med/high/xhigh 档位以及 GLM 5.3 的 med/high 档位。他表示许多 run 尚未完成,会持续更新,并感谢 GOrlanski 团队搭建该基准。
内容分类AI 观点与方法
内容层级普通情报
发布时间(北京时间)
本站收录时间(北京时间)
信息来源X:Dex Horthy(HumanLayer)(@dexhorthy)
站内情报编号intel-8bb4c628ecad023b211443ba