Trend-tracking jq recipe prints evaluation_completed=false as 'n/a', because jq's // operator treats false like null (F-W1-DOCS-TRAIN-16)
Nadie ha tomado este issue todavía.
Evaluación
- Dificultad
- 1/5
- Tiempo estimado
- 1-3 horas
- Aptitud para principiantes
- 80/100
- Tipo de issue
- Documentación
- Claridad
- Bien especificado
- Estado de actividad
- Activo
- Stack tecnológico
- python
- Área
- documentation
Línea de trabajo
La vista de jq defectuosa está en la línea 47 de docs/usermanuals/en/evaluation/trend-tracking.md, y la copia en turco en la misma línea de docs/usermanuals/tr/evaluation/trend-tracking.md necesita el mismo cambio. Ejecuta la expresión documentada con jq 1.7.1 contra las tres entradas {"evaluation_completed": false}, {"evaluation_completed": true} y {} para ver la salida actual. Está terminado cuando los tres casos imprimen false, true y n/a, y ambas páginas de idioma se actualizan en el mismo cambio.
Escrito por el modelo de indexación a partir del texto del issue.
Descripción
Wave 4 · group W4 Docs · highest severity Low · 1 finding(s) · effort docs fix
Summary
The page explains that evaluation_completed: false marks a safety run that produced no usable evidence, then gives a jq view that renders the field with .evaluation_completed // "n/a". jq's alternative operator falls back on false as well as null, so exactly the rows the view is meant to expose print as completed=n/a, indistinguishable from older rows that lack the field.
Primary finding
F-W1-DOCS-TRAIN-16 — The trend query hides an explicit evaluation failure as missing telemetry
severity Low · confidence high · reproduced with jq 1.7.1
What happens. docs/usermanuals/en/evaluation/trend-tracking.md:47 and the Turkish mirror at the same line. Rows with evaluation_completed: false do occur: the safety orchestrator sets it when the classifier abstains (forgelm/safety/_orchestrator.py:359-361) and _append_trend_entry writes it to safety_trend.jsonl (forgelm/safety/_results.py:168).
Where.
docs/usermanuals/en/evaluation/trend-tracking.md:47docs/usermanuals/tr/evaluation/trend-tracking.md:47forgelm/safety/_orchestrator.py:359-361forgelm/safety/_results.py:168
Evidence (re-verified on 2026-10-09).
- With jq 1.7.1 the documented expression prints
completed=n/afor{"evaluation_completed": false},completed=truefortrue, andcompleted=n/awhen the field is absent.
Recommended fix. Render the field without the alternative operator, for example completed=\(if .evaluation_completed == null then "n/a" else .evaluation_completed end), which prints true, false and n/a for the three cases.
Acceptance criteria
- The jq view prints
true,falseandn/afor true, false and missing values. - Documentation listed above is corrected in English and in its Turkish mirror in the same PR.
- The finding is resolved or explicitly closed with a reason.
Found by the September 2026 independent review of ForgeLM against commit cab2947 and re-verified against development at f94595f on 2026-10-09. Code links point at cab2947; line numbers may have moved since. Finding IDs (F-W1-…) are stable references for this backlog.
- Lenguaje dominante
- Python
- Estrellas
- 9
- Forks
- 1
- Merge medio
- 4 h 3 min
- PR fusionados (30 d)
- 3
Preparar el entorno
- Incluye un Dockerfile o un archivo de Docker Compose
- Tiene una plantilla de pull request
- Leer la guía de contribución
Primeros pasos
- Lee el issue completo y luego la guía de contribución del proyecto.
- Comenta en el issue que vas a ocuparte — evita que dos personas hagan lo mismo.
- Haz un fork del repositorio y trabaja en una rama.
- Abre un pull request que haga referencia al número del issue.
Más de HodeTech/ForgeLM
-
area: dev-tooling bug severity: low source: roadmap wave: 4
Dificultad 2/5 1-3 horas Aptitud para principiantes 70/100
-
area: site bug good first issue severity: low source: review-2026-09 wave: 4
Dificultad 2/5 1-3 horas Aptitud para principiantes 66/100
-
documentation good first issue severity: medium source: review-2026-09 wave: 3
Dificultad 2/5 1-3 horas Aptitud para principiantes 75/100
-
documentation severity: medium source: review-2026-09 wave: 3
Dificultad 2/5 1-3 horas Aptitud para principiantes 72/100
-
documentation good first issue severity: medium source: review-2026-09 wave: 3
Dificultad 2/5 1-3 horas Aptitud para principiantes 68/100
Todos los issues de HodeTech/ForgeLM
Issues similares
-
Dificultad 2/5 1-3 horas Aptitud para principiantes 72/100
NousResearch/hermes-agent#136483 ·
Los mantenedores suelen responder en 1 día
-
Dificultad 1/5 Menos de una hora Aptitud para principiantes 88/100
Los mantenedores suelen responder en 1 día
-
[BUG] LazyStackedTensorDictStore zeroes the last byte of a new key set on the last elementPosiblemente ocupada @peterdsharpe la tomó hoy. Abiertobug
Dificultad 2/5 1-3 horas Aptitud para principiantes 78/100
pytorch/tensordict#2307 ·
Los mantenedores suelen responder en 1 día
-
Dificultad 2/5 1-3 horas Aptitud para principiantes 78/100
Los mantenedores suelen responder en 1 día
-
GrokModel.generate/a_generate pass an OpenAI-style list-of-dicts to xai_sdk.chat.user(), so every call crashes with a protobuf TypeError before any network I/OPosiblemente ocupada @Christian-Sidak la tomó hoy. Abierto
Dificultad 2/5 1-3 horas Aptitud para principiantes 70/100
confident-ai/deepeval#3436 · 1 comentario ·
Los mantenedores suelen responder en 1 día