decision.host
首页 / 评估 / Phishing gradient

Phishing gradient

Decision Index retrieval 1 个模型参评

Judge generated emails as phishing across four difficulty tiers, from obvious to misleading.

评估集信息

所属套件
Decision Index
分组
retrieval
决策类型
混合(Choice / Noul / Score)
样本量
800 个 case
主指标
accuracy
优化方向
越高越好
参评模型
1 个
最佳分
94.6
平均分
94.6
是否计入指数
未计入
上游数据集
Phishing difficulty gradient
数据许可证
见上游原始数据集

这个评估集测什么

Judge generated emails as phishing across four difficulty tiers, from obvious to misleading.

在该评估集上的模型成绩(1 个)

#模型得分 对比其他指标模态官方链接
1 Jev 1.13
TypeSafe AI · Decision Index
94.6 文本 官网 ↗ 官方文档 ↗