Add GEM Response Generation to Full Benchmark
还没有人认领这个 Issue。
评估
- 难度
- 5/5
- 预计耗时
- 一周以上
- 新手友好度
- 15/100
- Issue 类型
- 功能
- 描述清晰度
- 需要澄清
- 活跃度
- 停滞
- 技术栈
- python
调研方向
issue 中没有确定任何文件、测试或入口点。首先定位完整的 benchmark 流程及其现有的 Schema-Guided Dialog 评估,包括 shuffle challenge set。完成的标准是:响应生成已纳入完整的 benchmark,并且其结果得到一致的评估。
由索引模型根据 Issue 内容生成。
描述
Response generation in Schema-Guided Dialog (including shuffle challenge set)
- 主要语言
- Python
- 星标
- 42
- 派生
- 24
- PR 合并指标
- 30 天内没有已合并 PR
环境准备
- 没有 Dockerfile 或 Docker Compose 文件
- 没有 Pull Request 模板
- 阅读贡献指南
从这里开始
- 先读完整个 Issue,再读项目的贡献指南。
- 在 Issue 下留言说明你要接手 —— 这能避免两个人做同样的事。
- Fork 仓库,在一个分支上完成修改。
- 提交 Pull Request,并在描述里引用这个 Issue 编号。
bigscience-workshop/evaluation 的其他 Issue
-
难度 4/5 3-5 天 新手友好度 25/100
-
Refactor task template to merge multilingual.json and english.json可能重新可做 @wilsonyhlee 于 1861 天前认领,目前没有进行中的 PR。 未关闭
bigscience-workshop/evaluation#64 · 已指派 1 人 ·
-
难度 5/5 一周以上 新手友好度 20/100
bigscience-workshop/evaluation#61 · 2 条评论 · 3 个 reaction ·
-
Setup testing可能重新可做 @tttyuntian 于 1877 天前认领,目前没有进行中的 PR。 未关闭engineering
bigscience-workshop/evaluation#57 · 2 条评论 · 已指派 3 人 ·
-
Start overleaf for benchmark tech report可能重新可做 @arunraja-hub 于 1873 天前认领,目前没有进行中的 PR。 未关闭documentation
bigscience-workshop/evaluation#54 · 1 条评论 · 3 个 reaction · 已指派 2 人 ·
查看 bigscience-workshop/evaluation 的全部 Issue
相似的 Issue
-
docs(types): update the collection binding note now that typed collections shipped in pycubrid 1.9.0未关闭documentation priority: low size: S
难度 2/5 1-3 小时 新手友好度 75/100
cubrid-lab/sqlalchemy-cubrid#768 ·
维护者通常 1 天内回复
-
bug help wanted
难度 2/5 1-3 小时 新手友好度 75/100
维护者通常 1 天内回复
-
documentation
难度 1/5 1 小时以内 新手友好度 65/100
ansys/pydpf-core#3547 ·
维护者通常 1 天内回复
-
core
难度 2/5 1-3 小时 新手友好度 70/100
vectorize-io/hindsight#5457 ·
维护者通常 1 天内回复
-
[Bug]: LangChain drops OpenAI Responses text blocks from session recording可能已有人在做 @ktz03 今天认领。 未关闭
难度 2/5 1-3 小时 新手友好度 72/100
volcengine/OpenViking#5806 ·
维护者通常 1 天内回复