Hacktoberfest 2026:维护者为十月标记出来的 issue,仍然开放、适合新手。 浏览 Hacktoberfest issue

'Run this as a task on <model>' becomes a reading helper; chat misreports it

未关闭
#1,569 0 条评论 0 个 reaction 已指派 0 人 在 GitHub 查看

维护者通常 1 天内回复

还没有人认领这个 Issue。

评估

难度
4/5
预计耗时
3-5 天
新手友好度
68/100
Issue 类型
缺陷
描述清晰度
描述清楚
活跃度
活跃
技术栈
go
领域
cli, testing-qa

调研方向

Start with internal/session/readhandoff.go, especially the hand-off admission around lines 150-160 and sweepTitle around lines 324-329, then reproduce the message using the isolated profile and Python repo described. Compare the resulting tasks.json and usage.jsonl with the chat report; done means the request produces a named task proposal or clear refusal, avoids unbranched edits, and reports the task's actual model and usage.

由索引模型根据 Issue 内容生成。

描述

area:session bug sev:critical

Seen on: dev 837b2b0.

Behaviour

Run this as a task on glm 5.3 flash: implement parse and humanize … made no proposal and no crew line. A ◆ reading: Run this as a task on glm 5.3 flash … helper ran on z-ai/glm-5.3 ($0.032) and edited durations.py in place (uncommitted, no branch). The chat then said Task ran on glm 5.3 flash … Time: 11 seconds. Cost: $0.055 (estimate $0.203), while glm-5.3-flash spent only $0.003 in that workspace. The same words in another window made a real task on kimi-k3.

A request for a task on a named model should produce a task proposal (or a clear refusal), and the chat's report of model, time and cost must match what actually ran.

Replication

  1. Build dev 837b2b0 (git checkout 837b2b0 && make build, binary bin/codeaf), or install the dev build with curl -fsSL https://agentfield.ai/get/devaf | bash.
  2. Use an isolated profile: export HOME=$(mktemp -d), export OPENROUTER_API_KEY, and keep the default model (~deepseek/deepseek-v4-flash-latest, crew on auto).
  3. On a busy machine set task.max_load to 0 (/settings, Tasks) so the busy-machine gate does not hold tasks.
  4. Make a small Python repo: R=$(mktemp -d) && cd "$R" && git init -q && printf 'def parse(s):\n raise NotImplementedError\n\ndef humanize(n):\n raise NotImplementedError\n' > durations.py && git add -A && git commit -qm init.
  5. codeaf, then send: Run this as a task on glm 5.3 flash: implement parse and humanize in durations.py (parse "1h30m" to seconds, humanize seconds back) and add tests.
  6. Look for a ◆ reading: row instead of a proposal card; read git status (edits in place, no branch) and compare the chat's Cost: line with codeaf logs for the workspace.

Evidence

  • Screen: ◆ reading: Run this as a task on glm 5.3 flash …; chat reply Task ran on glm 5.3 flash … Time: 11 seconds. Cost: $0.055 (estimate $0.203).
  • The session's tasks.json: node 1 titled reading: …, model z-ai/glm-5.3; usage.jsonl shows glm-5.3-flash at $0.003.
  • internal/session/readhandoff.go:150-160 (hand-off admission) and :324-329 (sweepTitle gives reading: <first line>).

Guessed cause

A guess from reading the code, not a confirmed diagnosis. As in the reading-helper issue filed beside this one: the read hand-off fires before propose_task, the quick task runs on the reader's model with a write-capable belt, and the chat's summary is composed by the model rather than from the helper's real model and cost.

Acceptance

  • e2e: the message above produces a task proposal whose card names glm-5.3-flash (or the chat refuses in one line); no file changes outside a task branch.
  • e2e: any cost or model the chat quotes for a task matches the task's usage rows.

Found while writing the public docs; manual text differences are in #1545.

🤖 Generated with Claude Code

主要语言
Go
星标
115
派生
14
平均合并
9 小时 44 分钟
30 天内合并 PR
766

环境准备

我们还没有检查这个项目的环境配置文件。先看它的 README,通用步骤见我们的新手贡献指南。

从这里开始

  1. 先读完整个 Issue,再读项目的贡献指南。
  2. 在 Issue 下留言说明你要接手 —— 这能避免两个人做同样的事。
  3. Fork 仓库,在一个分支上完成修改。
  4. 提交 Pull Request,并在描述里引用这个 Issue 编号。

Agent-Field/CodeAF 的其他 Issue

查看 Agent-Field/CodeAF 的全部 Issue

相似的 Issue

更多 Go Issue

把新 issue 发到你的邮箱

精选适合新手参与的 GitHub issue 摘要。