• 文献检索
  • 文档翻译
  • 深度研究
  • 学术资讯
  • Suppr Zotero 插件Zotero 插件
  • 邀请有礼
  • 套餐&价格
  • 历史记录
应用&插件
Suppr Zotero 插件Zotero 插件浏览器插件Mac 客户端Windows 客户端微信小程序
定价
高级版会员购买积分包购买API积分包
服务
文献检索文档翻译深度研究API 文档MCP 服务
关于我们
关于 Suppr公司介绍联系我们用户协议隐私条款
关注我们

Suppr 超能文献

核心技术专利:CN118964589B侵权必究
粤ICP备2023148730 号-1Suppr @ 2026

文献检索

告别复杂PubMed语法,用中文像聊天一样搜索,搜遍4000万医学文献。AI智能推荐,让科研检索更轻松。

立即免费搜索

文件翻译

保留排版,准确专业,支持PDF/Word/PPT等文件格式,支持 12+语言互译。

免费翻译文档

深度研究

AI帮你快速写综述,25分钟生成高质量综述,智能提取关键信息,辅助科研写作。

立即免费体验

作为多结论分类结果量表性能指标的交叉熵和对数似然比代价

Cross entropy and log likelihood ratio cost as performance measures for multi-conclusion categorical outcomes scales.

作者信息

Warren Eric M, Handley John C, Sheets H David

机构信息

SEP Forensic Consultants, Memphis, Tennessee, USA.

Simon Business School, University of Rochester, Rochester, New York, USA.

出版信息

J Forensic Sci. 2025 Mar;70(2):589-606. doi: 10.1111/1556-4029.15686. Epub 2024 Dec 10.

DOI:10.1111/1556-4029.15686
PMID:39655364
Abstract

The inconclusive category in forensics reporting is the appropriate response in many cases, but it poses challenges in estimating an "error rate". We discuss the use of a class of information-theoretic measures related to cross entropy as an alternative set of metrics that allows for performance evaluation of results presented using multi-category reporting scales. This paper shows how this class of performance metrics, and in particular the log likelihood ratio cost, which is already in use with likelihood ratio forensic reporting methods and in machine learning communities, can be readily adapted for use with the widely used multiple category conclusions scales. Bayesian credible intervals on these metrics can be estimated using numerical methods. The application of these metrics to published test results is shown. It is demonstrated, using these test results, that reducing the number of categories used in a proficiency test from five or six to three increases the cross entropy, indicating that the higher number of categories was justified, as it they increased the level of agreement with ground truth.

摘要

在法医报告中,不确定类别在许多情况下是适当的回应,但它在估计“错误率”方面带来了挑战。我们讨论使用一类与交叉熵相关的信息论度量作为一组替代指标,以评估使用多类别报告量表呈现的结果的性能。本文展示了这类性能指标,特别是对数似然比成本,如何能够很容易地适用于广泛使用的多类别结论量表,对数似然比成本已在似然比法医报告方法和机器学习领域中使用。这些指标的贝叶斯可信区间可以使用数值方法进行估计。展示了这些指标在已发表测试结果中的应用。利用这些测试结果表明,将能力验证中使用的类别数量从五六个减少到三个会增加交叉熵,这表明较多的类别数量是合理的,因为它们提高了与真实情况的一致程度。

相似文献

1
Cross entropy and log likelihood ratio cost as performance measures for multi-conclusion categorical outcomes scales.作为多结论分类结果量表性能指标的交叉熵和对数似然比代价
J Forensic Sci. 2025 Mar;70(2):589-606. doi: 10.1111/1556-4029.15686. Epub 2024 Dec 10.
2
The inconclusive category, entropy, and forensic firearm identification.不确定类别、熵与法医枪支鉴定
Forensic Sci Int. 2023 Aug;349:111741. doi: 10.1016/j.forsciint.2023.111741. Epub 2023 Jun 1.
3
Folic acid supplementation and malaria susceptibility and severity among people taking antifolate antimalarial drugs in endemic areas.在流行地区,服用抗叶酸抗疟药物的人群中,叶酸补充剂与疟疾易感性和严重程度的关系。
Cochrane Database Syst Rev. 2022 Feb 1;2(2022):CD014217. doi: 10.1002/14651858.CD014217.
4
Two heads are better than one: Dual systems obtain better performance in facial comparison.两个脑袋总比一个好:双重系统在面部比较中表现更好。
Forensic Sci Int. 2023 Dec;353:111879. doi: 10.1016/j.forsciint.2023.111879. Epub 2023 Nov 4.
5
Measuring the validity and reliability of forensic likelihood-ratio systems.
Sci Justice. 2011 Sep;51(3):91-8. doi: 10.1016/j.scijus.2011.03.002. Epub 2011 Apr 14.
6
Classification of multi-feature fusion ultrasound images of breast tumor within category 4 using convolutional neural networks.使用卷积神经网络对 4 类乳腺肿瘤多特征融合超声图像进行分类。
Med Phys. 2024 Jun;51(6):4243-4257. doi: 10.1002/mp.16946. Epub 2024 Mar 4.
7
Impact of summer programmes on the outcomes of disadvantaged or 'at risk' young people: A systematic review.暑期项目对处境不利或“有风险”的年轻人的影响:一项系统综述。
Campbell Syst Rev. 2024 Jun 13;20(2):e1406. doi: 10.1002/cl2.1406. eCollection 2024 Jun.
8
Information-theoretical assessment of the performance of likelihood ratio computation methods.似然比计算方法性能的信息论评估。
J Forensic Sci. 2013 Nov;58(6):1503-18. doi: 10.1111/1556-4029.12233. Epub 2013 Jul 23.
9
Rank order entropy: why one metric is not enough.秩次熵:为何一种度量指标并不够。
J Chem Inf Model. 2011 Sep 26;51(9):2302-19. doi: 10.1021/ci200170k. Epub 2011 Aug 29.
10
Automated face recognition in forensic science: Review and perspectives.自动人脸识别在法医学中的应用:综述与展望。
Forensic Sci Int. 2020 Feb;307:110124. doi: 10.1016/j.forsciint.2019.110124. Epub 2019 Dec 23.

引用本文的文献

1
Quantifying the strength of palmprint comparisons: Majority identifications with surprisingly low value.量化掌纹比对的强度:多数认定的价值低得出奇。
Forensic Sci Int Synerg. 2025 Jul 23;11:100628. doi: 10.1016/j.fsisyn.2025.100628. eCollection 2025 Dec.