Hacktoberfest 2026: le issue che i maintainer hanno segnato per ottobre, aperte e adatte ai principianti. Sfoglia le issue Hacktoberfest

datalake_fdw: NUMERIC/DECIMAL in the Parquet format layer

Aperta
#1,988 1 commento 0 reazioni 1 assegnatario Vedi su GitHub

@MisterRaindrop ci sta già lavorando.

Dal 13/9/2026.

Valutazione

Questa issue non è ancora stata valutata.

Descrizione

datalake
Summary

contrib/datalake_fdw's Parquet format layer (#1951) refuses numeric columns at CREATE TABLE ... USING iceberg. DECIMAL is the most common column type in real lake tables, so this is the first type to add.

What has to be decided
  • Storage form. Parquet stores DECIMAL four ways (INT32, INT64, FIXED_LEN_BYTE_ARRAY, BYTE_ARRAY). Iceberg's spec fixes the writer to decimal(P,S) with P <= 38 backed by fixed-length bytes; the reader has to accept all four.
  • Unconstrained numeric. atttypmod = -1 has no precision or scale. Computing ((typmod - 4) >> 16) & 65535 without checking gives precision 65535 / scale 65531, so bare numeric columns silently match nothing (the trap in lithium-tech/tea validate.cpp:59-61). Iceberg needs P and S, so the likely answer is to refuse bare numeric at CREATE TABLE and require numeric(P,S).
  • Precision above 38. PostgreSQL allows up to 1000; decimal128 caps at 38. Refuse at CREATE TABLE.
  • Conversion. NumericVar <-> two's-complement int128, the way tea's bridge.cpp:269-277 and numeric_var.cpp do it, with NaN/Infinity refused on write.
Where
  • format/format_types.h / arrow_support.cpp: dl_format_type_refusal() decides what CREATE TABLE accepts; the mapping to arrow::decimal128(P, S) goes next to the other types.
  • arrow_builder.cpp / arrow_decode.c: append and decode. The decoder reads the Arrow C data interface directly; the format string is d:P,S.
  • Tests in test/automation/sqlrepo/smoke/format_parquet/, plus CREATE TABLE refusals in iceberg_am_reject.sql.

Deferred from #1951 on purpose: the framework there is settled, the type work is separate.

Lingua principale
C
Stelle
1.4k
Fork
248
Merge medio
4g 10h
PR unite (30g)
40

Guida per i contributori

Apri la guida per i contributori

Come iniziare

  1. Leggi tutta la issue e poi la guida ai contributi del progetto.
  2. Commenta sulla issue per dire che te ne occupi tu — evita che due persone facciano lo stesso lavoro.
  3. Fai un fork del repository e lavora su un branch.
  4. Apri una pull request che faccia riferimento al numero della issue.

Altre issue di apache/cloudberry

Tutte le issue di apache/cloudberry

Issue simili

Altre issue su C

Ricevi le nuove issue nella tua casella

Un breve riepilogo di issue GitHub adatte ai principianti.