展示 HN:huby 的 AI 产品评估方法
2 分•作者: dd-sharma•大约 4 小时前
市面上有不少人工智能的基准测试,但我认为它们在评估人工智能产品时并不实用。几个月前,我们开始着手设计一种独立评估人工智能产品的方法论。经过大量的研究,我们在几类不同的产品上测试了该方法论并进行了改进。
如果您对这个领域感兴趣,能否请您审阅该方法论并提供反馈意见?该方法论的结构简介如下:它是一个三层框架,首先是六个高级评估类别(质量、安全、隐私与安全、用例与定价、可持续性与生态系统,以及影响与伦理)。
该方法论可在 huby.ai/methodology 查阅。
提前致谢。
查看原文
There are a number of AI benchmarks but I feel they aren't practical when it comes to evaluating AI products. Several months back we started working on designing a methodology to independently evaluate AI products. After conducting a good bit of research, we tested the methodology on a few different categories of products and improvised it.
If this is an area of interest to you, may I request you to review this methodology and help us with feedback. A quick blurb on the structure of the methodology - It's a 3-tier framework starting with 6 high level evaluation categories (Quality, Security, Privacy & safety, Use cases & pricing, Sustainability & ecosystem, and Impact & ethics).
The methodology is available at huby.ai/methodology.
Thanks in advance.