'Run this as a task on <model>' becomes a reading helper; chat misreports it
メンテナーはふだん 1 日以内に返信
まだ誰も着手していません。
評価
- 難易度
- 4/5
- 見積もり時間
- 3〜5日
- 初心者へのやさしさ
- 68/100
- issue の種類
- バグ
- 明瞭さ
- 明確に書かれている
- 活発さ
- 活発
- 技術スタック
- go
- 領域
- cli, testing-qa
調査の方向性
Start with internal/session/readhandoff.go, especially the hand-off admission around lines 150-160 and sweepTitle around lines 324-329, then reproduce the message using the isolated profile and Python repo described. Compare the resulting tasks.json and usage.jsonl with the chat report; done means the request produces a named task proposal or clear refusal, avoids unbranched edits, and reports the task's actual model and usage.
索引モデルが issue の本文から書いたものです。
説明
Seen on: dev 837b2b0.
Behaviour
Run this as a task on glm 5.3 flash: implement parse and humanize … made no proposal and no crew line. A ◆ reading: Run this as a task on glm 5.3 flash … helper ran on z-ai/glm-5.3 ($0.032) and edited durations.py in place (uncommitted, no branch). The chat then said Task ran on glm 5.3 flash … Time: 11 seconds. Cost: $0.055 (estimate $0.203), while glm-5.3-flash spent only $0.003 in that workspace. The same words in another window made a real task on kimi-k3.
A request for a task on a named model should produce a task proposal (or a clear refusal), and the chat's report of model, time and cost must match what actually ran.
Replication
- Build dev 837b2b0 (
git checkout 837b2b0 && make build, binarybin/codeaf), or install the dev build withcurl -fsSL https://agentfield.ai/get/devaf | bash. - Use an isolated profile:
export HOME=$(mktemp -d), exportOPENROUTER_API_KEY, and keep the default model (~deepseek/deepseek-v4-flash-latest, crew on auto). - On a busy machine set
task.max_loadto0(/settings, Tasks) so the busy-machine gate does not hold tasks. - Make a small Python repo:
R=$(mktemp -d) && cd "$R" && git init -q && printf 'def parse(s):\n raise NotImplementedError\n\ndef humanize(n):\n raise NotImplementedError\n' > durations.py && git add -A && git commit -qm init. codeaf, then send:Run this as a task on glm 5.3 flash: implement parse and humanize in durations.py (parse "1h30m" to seconds, humanize seconds back) and add tests.- Look for a
◆ reading:row instead of a proposal card; readgit status(edits in place, no branch) and compare the chat'sCost:line withcodeaf logsfor the workspace.
Evidence
- Screen:
◆ reading: Run this as a task on glm 5.3 flash …; chat replyTask ran on glm 5.3 flash … Time: 11 seconds. Cost: $0.055 (estimate $0.203). - The session's tasks.json: node 1 titled
reading: …, modelz-ai/glm-5.3; usage.jsonl shows glm-5.3-flash at $0.003. internal/session/readhandoff.go:150-160(hand-off admission) and:324-329(sweepTitlegivesreading: <first line>).
Guessed cause
A guess from reading the code, not a confirmed diagnosis. As in the reading-helper issue filed beside this one: the read hand-off fires before propose_task, the quick task runs on the reader's model with a write-capable belt, and the chat's summary is composed by the model rather than from the helper's real model and cost.
Acceptance
- e2e: the message above produces a task proposal whose card names glm-5.3-flash (or the chat refuses in one line); no file changes outside a task branch.
- e2e: any cost or model the chat quotes for a task matches the task's usage rows.
Found while writing the public docs; manual text differences are in #1545.
🤖 Generated with Claude Code
- 主要言語
- Go
- スター
- 115
- フォーク
- 14
- 平均マージ
- 9時間 37分
- マージ済み PR(30日)
- 755
環境構築
このプロジェクトの環境構築ファイルはまだ確認していません。まず README を読み、一般的な手順ははじめてのコントリビューションガイドを参照してください。
はじめの一歩
- issue を最後まで読み、次にプロジェクトのコントリビューションガイドを読みます。
- 着手することを issue にコメントします — 二人が同じ作業をするのを防げます。
- リポジトリをフォークし、ブランチを切って変更します。
- issue 番号を参照したプルリクエストを送ります。
Agent-Field/CodeAF のほかの issue
-
area:chat bug sev:papercut
難易度 2/5 1〜3時間 初心者へのやさしさ 86/100
Agent-Field/CodeAF#1592 ·
メンテナーはふだん 1 日以内に返信
-
area:headless bug sev:critical
難易度 2/5 1〜3時間 初心者へのやさしさ 88/100
Agent-Field/CodeAF#1566 · 担当者 1 名 ·
メンテナーはふだん 1 日以内に返信
-
area:chat feature
難易度 2/5 1〜3時間 初心者へのやさしさ 88/100
Agent-Field/CodeAF#1510 ·
メンテナーはふだん 1 日以内に返信
-
area:tests bug
難易度 2/5 1〜3時間 初心者へのやさしさ 82/100
Agent-Field/CodeAF#1489 ·
メンテナーはふだん 1 日以内に返信
-
area:chat bug good first issue sev:papercut
難易度 2/5 1〜3時間 初心者へのやさしさ 78/100
Agent-Field/CodeAF#1470 ·
メンテナーはふだん 1 日以内に返信
Agent-Field/CodeAF の issue をすべて見る
似ている issue
-
難易度 2/5 1〜3時間 初心者へのやさしさ 88/100
rossoctl/context-guru#346 ·
メンテナーはふだん 1 日以内に返信
-
難易度 1/5 1時間未満 初心者へのやさしさ 90/100
prime-radiant-inc/evener#2883 ·
メンテナーはふだん 1 日以内に返信
-
難易度 2/5 1〜3時間 初心者へのやさしさ 88/100
gravitational/teleport#69805 ·
メンテナーはふだん 11 日以内に返信
-
難易度 2/5 1〜3時間 初心者へのやさしさ 72/100
メンテナーはふだん 1 日以内に返信
-
Under Poisson sampling, the `PLDAccountant` composes the inner event both before and after samplingオープン
難易度 2/5 半日 初心者へのやさしさ 78/100
google/differential-privacy#496 ·