RCreddit.com·
暂不在当前实时榜单
Best practices when running a benchmark on online models [D]
A developer is seeking best practices for benchmarking online models, particularly for low-resource languages, while preventing data leakage. The concern is that input data used for predictions might be used for training by API providers like Google and OpenAI, even with paid accounts. The developer is questioning the trustworthiness of these providers' assurances regarding data usage and is looking for established methods to evaluate online models without compromising input data.
This post highlights the specific challenge of data leakage when benchmarking online models, unlike local models where this is not an issue.
时间与来源
时间显示为 UTC
显示时区:UTC
本地时区尚不可用,暂时显示 UTC。
收录当时偏移:UTC+02026年10月8日 16:00 UTC
- 收录
- 2026年10月8日 16:00
- 来源类型
- 开发者社区
讨论趋势
→ 平稳
百分比基于采集到的讨论信号,不代表新增评论数或独立参与人数。曲线仅用于同一话题在不同时段的比较。
本站未收录正文。
前往源站阅读 →