'Run this as a task on <model>' becomes a reading helper; chat misreports it
Les mainteneurs répondent en général sous 1 jour
Personne n'a encore pris cette issue.
Évaluation
- Difficulté
- 4/5
- Temps estimé
- 3-5 jours
- Accessibilité débutants
- 68/100
- Type d'issue
- Bug
- Clarté
- Clairement spécifiée
- Activité
- Active
- Stack technique
- go
- Domaine
- cli, testing-qa
Piste de recherche
Start with internal/session/readhandoff.go, especially the hand-off admission around lines 150-160 and sweepTitle around lines 324-329, then reproduce the message using the isolated profile and Python repo described. Compare the resulting tasks.json and usage.jsonl with the chat report; done means the request produces a named task proposal or clear refusal, avoids unbranched edits, and reports the task's actual model and usage.
Rédigé par le modèle d'indexation à partir du texte de l'issue.
Description
Seen on: dev 837b2b0.
Behaviour
Run this as a task on glm 5.3 flash: implement parse and humanize … made no proposal and no crew line. A ◆ reading: Run this as a task on glm 5.3 flash … helper ran on z-ai/glm-5.3 ($0.032) and edited durations.py in place (uncommitted, no branch). The chat then said Task ran on glm 5.3 flash … Time: 11 seconds. Cost: $0.055 (estimate $0.203), while glm-5.3-flash spent only $0.003 in that workspace. The same words in another window made a real task on kimi-k3.
A request for a task on a named model should produce a task proposal (or a clear refusal), and the chat's report of model, time and cost must match what actually ran.
Replication
- Build dev 837b2b0 (
git checkout 837b2b0 && make build, binarybin/codeaf), or install the dev build withcurl -fsSL https://agentfield.ai/get/devaf | bash. - Use an isolated profile:
export HOME=$(mktemp -d), exportOPENROUTER_API_KEY, and keep the default model (~deepseek/deepseek-v4-flash-latest, crew on auto). - On a busy machine set
task.max_loadto0(/settings, Tasks) so the busy-machine gate does not hold tasks. - Make a small Python repo:
R=$(mktemp -d) && cd "$R" && git init -q && printf 'def parse(s):\n raise NotImplementedError\n\ndef humanize(n):\n raise NotImplementedError\n' > durations.py && git add -A && git commit -qm init. codeaf, then send:Run this as a task on glm 5.3 flash: implement parse and humanize in durations.py (parse "1h30m" to seconds, humanize seconds back) and add tests.- Look for a
◆ reading:row instead of a proposal card; readgit status(edits in place, no branch) and compare the chat'sCost:line withcodeaf logsfor the workspace.
Evidence
- Screen:
◆ reading: Run this as a task on glm 5.3 flash …; chat replyTask ran on glm 5.3 flash … Time: 11 seconds. Cost: $0.055 (estimate $0.203). - The session's tasks.json: node 1 titled
reading: …, modelz-ai/glm-5.3; usage.jsonl shows glm-5.3-flash at $0.003. internal/session/readhandoff.go:150-160(hand-off admission) and:324-329(sweepTitlegivesreading: <first line>).
Guessed cause
A guess from reading the code, not a confirmed diagnosis. As in the reading-helper issue filed beside this one: the read hand-off fires before propose_task, the quick task runs on the reader's model with a write-capable belt, and the chat's summary is composed by the model rather than from the helper's real model and cost.
Acceptance
- e2e: the message above produces a task proposal whose card names glm-5.3-flash (or the chat refuses in one line); no file changes outside a task branch.
- e2e: any cost or model the chat quotes for a task matches the task's usage rows.
Found while writing the public docs; manual text differences are in #1545.
🤖 Generated with Claude Code
- Langage dominant
- Go
- Étoiles
- 115
- Forks
- 14
- Merge moyen
- 9 h 44 min
- PR mergées (30 j)
- 775
Préparer son environnement
Nous n'avons pas encore vérifié les fichiers d'installation de ce projet. Commencez par son README, et consultez notre guide de la première contribution pour les étapes générales.
Par où commencer
- Lisez l'issue en entier, puis le guide de contribution du projet.
- Signalez en commentaire que vous la prenez — cela évite que deux personnes fassent le même travail.
- Forkez le dépôt et travaillez sur une branche.
- Ouvrez une pull request qui référence le numéro de l'issue.
Autres issues de Agent-Field/CodeAF
-
Difficulté 2/5 1-3 heures Accessibilité débutants 88/100
Agent-Field/CodeAF#1679 ·
Les mainteneurs répondent en général sous 1 jour
-
Difficulté 2/5 1-3 heures Accessibilité débutants 85/100
Agent-Field/CodeAF#1678 ·
Les mainteneurs répondent en général sous 1 jour
-
area:chat bug sev:papercut
Difficulté 2/5 1-3 heures Accessibilité débutants 86/100
Agent-Field/CodeAF#1592 ·
Les mainteneurs répondent en général sous 1 jour
-
codeaf do "" runs a paid job with an empty briefPeut-être pris @santoshkumarradha l’a pris il y a 2 jours. Ouvertearea:headless bug sev:critical
Difficulté 2/5 1-3 heures Accessibilité débutants 88/100
Agent-Field/CodeAF#1566 · 1 personne assignée ·
Les mainteneurs répondent en général sous 1 jour
-
area:chat feature
Difficulté 2/5 1-3 heures Accessibilité débutants 88/100
Agent-Field/CodeAF#1510 ·
Les mainteneurs répondent en général sous 1 jour
Toutes les issues de Agent-Field/CodeAF
Issues similaires
-
automation models
Difficulté 2/5 1-3 heures Accessibilité débutants 78/100
Les mainteneurs répondent en général sous 1 jour
-
Difficulté 2/5 1-3 heures Accessibilité débutants 72/100
txn2/mcp-data-platform#1984 ·
Les mainteneurs répondent en général sous 1 jour
-
agentic-workflows
Difficulté 2/5 1-3 heures Accessibilité débutants 68/100
Les mainteneurs répondent en général sous 1 jour
-
Difficulté 2/5 1-3 heures Accessibilité débutants 88/100
Les mainteneurs répondent en général sous 1 jour
-
Fix broken Code of Conduct linksOuvertekind/docs prio/P2
Difficulté 1/5 Moins d'une heure Accessibilité débutants 95/100
agent-substrate/substrate#1986 ·
Les mainteneurs répondent en général sous 1 jour