Flattened observations disagree with explicitly declared task specifications
Dieses Issue hat noch niemand übernommen.
Bewertung
- Schwierigkeit
- 3/5
- Geschätzter Aufwand
- 1-2 Tage
- Anfängerfreundlichkeit
- 74/100
- Issue-Typ
- Bug
- Klarheit
- Größtenteils klar
- Aktivitätsstatus
- Aktiv
- Tech-Stack
- numpy, python
- Bereich
- backend-api-design
Rechercherichtung
Start at control.Environment.observation_spec and compare the explicit-spec path with the existing fallback flattening path used when no specification is declared. Check that flat_observation preserves shapes and NumPy dtype promotion without materializing observations or calling get_observation, while leaving the caller's spec and mapping type unchanged. Run the regression against reset and step behavior to confirm the declared and inferred paths agree.
Vom Indexierungsmodell aus dem Issue-Text verfasst.
Beschreibung
Reproduction
On current main, an environment using flat_observation=True flattens the values returned by reset and step, but returns an unflattened specification when the task implements observation_spec.
For example, a task declaring a float32 position array of shape (2,) and an int32 scalar mode produces one float64 observations array of shape (3,). Environment.observation_spec() instead returns the original two specifications.
The fallback path for tasks without an explicit specification already flattens correctly. Agents that allocate or validate observations from the environment specification therefore behave differently depending on whether the task supplies its own spec.
Expected behavior
Flatten the declared shapes and apply NumPy's concatenation dtype promotion when flattening is enabled, without materializing observation arrays or calling get_observation. Preserve the caller's spec, mapping type, normal unflattened behavior, and the existing inference path.
Reproduced with actual control.Environment calls and a small native MuJoCo task. The new regression fails on current main at a04e3e4cf56c12117d2294bb090f9acec21e5c67; no renderer, trained policies or external service is required.
- Vorherrschende Sprache
- Python
- Sterne
- 4.7k
- Forks
- 768
- PR-Merge-Kennzahlen
- Keine gemergten PRs in 30 T.
Entwicklungsumgebung
- Kein Dockerfile und keine Docker-Compose-Datei
- Keine Pull-Request-Vorlage
- Beitragsleitfaden lesen
Erste Schritte
- Lesen Sie das ganze Issue und danach den Beitragsleitfaden des Projekts.
- Schreiben Sie ins Issue, dass Sie es übernehmen — das erspart doppelte Arbeit.
- Forken Sie das Repository und arbeiten Sie in einem Branch.
- Öffnen Sie einen Pull Request, der die Issue-Nummer nennt.
Mehr aus google-deepmind/dm_control
-
Schwierigkeit 2/5 1-3 Stunden Anfängerfreundlichkeit 85/100
google-deepmind/dm_control#559 ·
-
Schwierigkeit 3/5 1-2 Tage Anfängerfreundlichkeit 75/100
google-deepmind/dm_control#560 ·
-
Schwierigkeit 3/5 1-2 Tage Anfängerfreundlichkeit 76/100
google-deepmind/dm_control#552 ·
-
Schwierigkeit 3/5 1-2 Tage Anfängerfreundlichkeit 45/100
google-deepmind/dm_control#540 · 1 Kommentar ·
-
Schwierigkeit 3/5 1-2 Tage Anfängerfreundlichkeit 55/100
google-deepmind/dm_control#537 · 1 Kommentar ·
Alle Issues in google-deepmind/dm_control
Ähnliche Issues
-
Schwierigkeit 2/5 1-3 Stunden Anfängerfreundlichkeit 85/100
mozilla/bedrock#17413 · 1 Reaktion ·
Maintainer antworten meist innerhalb von 2 Tagen
-
instance instance add
Schwierigkeit 2/5 1-3 Stunden Anfängerfreundlichkeit 68/100
searxng/searx-instances#943 · 1 Kommentar ·
-
Schwierigkeit 2/5 1-3 Stunden Anfängerfreundlichkeit 68/100
Maintainer antworten meist innerhalb von 1 Tag
-
bug tools
Schwierigkeit 2/5 1-3 Stunden Anfängerfreundlichkeit 88/100
Maintainer antworten meist innerhalb von 1 Tag
-
bug
Schwierigkeit 2/5 1-3 Stunden Anfängerfreundlichkeit 86/100
lance-format/lance#9655 ·
Maintainer antworten meist innerhalb von 2 Tagen