Hacktoberfest 2026:维护者为十月标记出来的 issue,仍然开放、适合新手。 浏览 Hacktoberfest issue

Compute metrics for all data once

未关闭
#338 0 条评论 0 个 reaction 已指派 0 人 在 GitHub 查看

还没有人认领这个 Issue。

评估

难度
3/5
预计耗时
1-2 天
新手友好度
48/100
Issue 类型
重构
描述清晰度
基本清楚
活跃度
停滞
技术栈
python

调研方向

从 rsmtool/analyzer.py 第 833 行附近开始,然后跟踪 run_training_analyses 和 compute_correlations_by_group,以了解所有数据的相关性在哪里生成。确认子组分析不再重新计算这些相关性,并且其输出会与之前计算的 all_data 结果拼接,同时将 all_data 保持在顶部。

由索引模型根据 Issue 内容生成。

描述

enhancement

Currently the correlations for All data are computed multiple times: first for all data in run_training_analyses and then during the analysis for each subgroup under compute_correlations_by_group.

For various reasons it would be helpful to keep the all_data on top of the subgroup analyses, but we should remove the computation from *_by_group (https://github.com/EducationalTestingService/rsmtool/blob/07f91264795dc8fb5de9518cb65ff4e1685836ef/rsmtool/analyzer.py#L833) and simply concatenate the output of _by_group with all_data computed previously.

主要语言
Python
星标
71
派生
21
PR 合并指标
30 天内没有已合并 PR

贡献指南

这个仓库没有索引到贡献指南

从这里开始

  1. 先读完整个 Issue,再读项目的贡献指南。
  2. 在 Issue 下留言说明你要接手 —— 这能避免两个人做同样的事。
  3. Fork 仓库,在一个分支上完成修改。
  4. 提交 Pull Request,并在描述里引用这个 Issue 编号。

EducationalTestingService/rsmtool 的其他 Issue

查看 EducationalTestingService/rsmtool 的全部 Issue

相似的 Issue

更多 Python Issue

把新 issue 发到你的邮箱

精选适合新手参与的 GitHub issue 摘要。