Compute metrics for all data once
还没有人认领这个 Issue。
评估
- 难度
- 3/5
- 预计耗时
- 1-2 天
- 新手友好度
- 48/100
- Issue 类型
- 重构
- 描述清晰度
- 基本清楚
- 活跃度
- 停滞
- 技术栈
- python
- 领域
- data, machine-learning
调研方向
从 rsmtool/analyzer.py 第 833 行附近开始,然后跟踪 run_training_analyses 和 compute_correlations_by_group,以了解所有数据的相关性在哪里生成。确认子组分析不再重新计算这些相关性,并且其输出会与之前计算的 all_data 结果拼接,同时将 all_data 保持在顶部。
由索引模型根据 Issue 内容生成。
描述
Currently the correlations for All data are computed multiple times: first for all data in run_training_analyses and then during the analysis for each subgroup under compute_correlations_by_group.
For various reasons it would be helpful to keep the all_data on top of the subgroup analyses, but we should remove the computation from *_by_group (https://github.com/EducationalTestingService/rsmtool/blob/07f91264795dc8fb5de9518cb65ff4e1685836ef/rsmtool/analyzer.py#L833) and simply concatenate the output of _by_group with all_data computed previously.
- 主要语言
- Python
- 星标
- 71
- 派生
- 21
- PR 合并指标
- 30 天内没有已合并 PR
贡献指南
这个仓库没有索引到贡献指南
从这里开始
- 先读完整个 Issue,再读项目的贡献指南。
- 在 Issue 下留言说明你要接手 —— 这能避免两个人做同样的事。
- Fork 仓库,在一个分支上完成修改。
- 提交 Pull Request,并在描述里引用这个 Issue 编号。
EducationalTestingService/rsmtool 的其他 Issue
-
难度 5/5 一周以上 新手友好度 25/100
EducationalTestingService/rsmtool#689 · 5 条评论 ·
-
enhancement
难度 4/5 3-5 天 新手友好度 35/100
-
EducationalTestingService/rsmtool#557 · 已指派 1 人 ·
-
enhancement
难度 5/5 一周以上 新手友好度 35/100
-
good first issue help wanted
难度 4/5 3-5 天 新手友好度 30/100
查看 EducationalTestingService/rsmtool 的全部 Issue
相似的 Issue
-
bug
难度 2/5 1-3 小时 新手友好度 75/100
stephrobert/dsoxlab#238 ·
-
难度 2/5 1-3 小时 新手友好度 75/100
-
难度 2/5 1-3 小时 新手友好度 75/100
sublimehq/package_control#1780 ·
-
难度 2/5 1-3 小时 新手友好度 65/100
-
难度 2/5 1-3 小时 新手友好度 70/100
nwg-piotr/nwg-displays#145 ·