Fully lazy imperative function evaluation
Maintainer antworten meist innerhalb von 1 Tag
Dieses Issue hat noch niemand übernommen.
Bewertung
- Schwierigkeit
- 5/5
- Geschätzter Aufwand
- Über eine Woche
- Anfängerfreundlichkeit
- 25/100
Rechercherichtung
Lies die verlinkte Implementierung von cumsum/cumprod mit axis=None und verfolge die Lazyexpr-Berechnungsmechanik, LazyUDFs und die Chunk-Verarbeitung von matmul. Sieh dir Issue #441 zum Kontext der Fancy-Indexierung an. Erledigt ist die Aufgabe, wenn Funktionen die angeforderten Ausgabe-Chunks verzögert berechnen, einschließlich verschachtelter matmul-/sum-Ausdrücke, ohne Temporärdaten für das vollständige Ergebnis zu erzeugen, und dabei die Präferenz für numexpr beibehalten sowie wiederholte Reduktionsberechnungen vermeiden.
Vom Indexierungsmodell aus dem Issue-Text verfasst.
Beschreibung
In order to make function evaluation properly lazy, all functions (e.g. matmul) should be implemented as LazyUDFs.
Secondly, the lazyexpr machinery of compute should loop over chunks of the result. Each function must then decide what slices of the operands are necessary to form the corresponding result chunk (as matmul currently does internally). Then, upon evaluation, although the expression evaluates term-by-term, it does not compute the full result for e..g matmul before proceeding to the next term, but only for the necessary chunk(s) of output. Thus there is a higher chance of cache hits (not the case currently for eager execution of linalg and reductions).
Something in this spirit has been implemented for cumsum/cumprod when axis=None (see this code ):
# Special case for cumulative operations with axis = None
if reduce_args["axis"] is None and reduce_op in {ReduceOp.CUMULATIVE_PROD, ReduceOp.CUMULATIVE_SUM}:
# res_out_ is just None, out set to all 0s (sum) or 1s (prod)
out, res_out_ = convert_none_out(dtype, reduce_op, reduced_shape)
# reduced_shape is just one-element tuple
chunklen = out.chunks[0] if hasattr(out, "chunks") else chunks[-1]
carry = 0
for cidx in range(0, reduced_shape[0] // chunklen):
slice_starts = np.unravel_index(cidx * chunklen, shape)
slice_stops = np.unravel_index((cidx + 1) * chunklen, shape)
cslice = tuple(
slice(start, stop) for start, stop in zip(slice_starts, slice_stops, strict=True)
)
_get_chunk_operands(operands, cslice, chunk_operands, shape)
result, _ = _get_result(expression, chunk_operands, ne_args, where)
result = np.require(result, requirements="C")
if reduce_op == ReduceOp.CUMULATIVE_SUM:
res = np.cumulative_sum(result, axis=None) + carry
else:
res = np.cumulative_prod(result, axis=None) * carry
carry = res[-1]
out[cidx * chunklen + include_initial : (cidx + 1) * chunklen + include_initial] = res
It should also be implemented for fancy-indexing with slice (see https://github.com/Blosc/python-blosc2/issues/441).
Should be possible to handle even something like "matmul(sum(a,axis=1), b)" by passing the desired slice from matmul->sum, which treats the asked-for-slice as a desired output slice and handles accordingly.
This would avoid large in-memory temporaries for example when calculating from eagerly executed linear algebra functions e.g. in "matmul(a, b) + b".
Problems:
1 - reductions could be a problem since for example "sum(a) + a" for the chunks of output would recalculate the scalar "sum(a)" for each chunk of output. This could be avoided by using some kind of result cache for reductions.
2 - naturally, numexpr would always be faster since it compiles the expression into bytecode. Thus we should make sure that, when possible, we still use numexpr preferentially (for most elementwise funcs).
- Vorherrschende Sprache
- Python
- Sterne
- 211
- Forks
- 66
- Ø Merge
- 2 T. 1 Std.
- Gemergte PRs (30 T.)
- 15
Entwicklungsumgebung
- Kein Dockerfile und keine Docker-Compose-Datei
- Keine Pull-Request-Vorlage
- Beitragsleitfaden lesen
Erste Schritte
- Lesen Sie das ganze Issue und danach den Beitragsleitfaden des Projekts.
- Schreiben Sie ins Issue, dass Sie es übernehmen — das erspart doppelte Arbeit.
- Forken Sie das Repository und arbeiten Sie in einem Branch.
- Öffnen Sie einen Pull Request, der die Issue-Nummer nennt.
Mehr aus Blosc/python-blosc2
-
Schwierigkeit 2/5 1-3 Stunden Anfängerfreundlichkeit 68/100
Blosc/python-blosc2#731 ·
Maintainer antworten meist innerhalb von 1 Tag
-
documentation sustain-2026
Schwierigkeit 2/5 1-3 Stunden Anfängerfreundlichkeit 75/100
Blosc/python-blosc2#720 · 3 Kommentare ·
Maintainer antworten meist innerhalb von 1 Tag
-
sustain-2026
Schwierigkeit 2/5 1-3 Stunden Anfängerfreundlichkeit 75/100
Blosc/python-blosc2#717 ·
Maintainer antworten meist innerhalb von 1 Tag
-
documentation sustain-2026
Schwierigkeit 2/5 1-3 Stunden Anfängerfreundlichkeit 80/100
Blosc/python-blosc2#714 ·
Maintainer antworten meist innerhalb von 1 Tag
-
documentation sustain-2026
Schwierigkeit 2/5 1-3 Stunden Anfängerfreundlichkeit 76/100
Blosc/python-blosc2#710 ·
Maintainer antworten meist innerhalb von 1 Tag
Alle Issues in Blosc/python-blosc2
Ähnliche Issues
-
Schwierigkeit 2/5 1-3 Stunden Anfängerfreundlichkeit 85/100
Maintainer antworten meist innerhalb von 3 Tagen
-
Negation with "not" and "no" is ignored during sentiment analysisEvtl. vergeben @vivek-3728 hat das heute übernommen. Offen
Schwierigkeit 2/5 1-3 Stunden Anfängerfreundlichkeit 68/100
techcsispit/mess-mood#11 · 1 Kommentar ·
-
changelog investigate
Schwierigkeit 2/5 1-3 Stunden Anfängerfreundlichkeit 68/100
ramnes/notion-sdk-py#408 ·
-
good first issue
Schwierigkeit 2/5 1-3 Stunden Anfängerfreundlichkeit 83/100
btclib-org/btclib-wallet#267 ·
Maintainer antworten meist innerhalb von 1 Tag
-
good first issue tech-debt
Schwierigkeit 2/5 1-3 Stunden Anfängerfreundlichkeit 85/100
knnmelprop/YAADO#111 ·