AI圈报
教程 / 实战普通

We ran HealthBench on our health AI's safety layer. It scored lower than the bare model.

信息来源:DEV Community·

内容摘要

A 150-conversation HealthBench subset, a safety layer that costs points, one real dose leak we found while measuring, and what moved the score.
内容分类AI 教程与实战
内容层级普通情报
发布时间(北京时间)
本站收录时间(北京时间)
信息来源DEV Community
站内情报编号intel-1ef7afa4039d04c8c084b4e2