apache/sedona

feat: add to_sedonadb() method

オープン

#2,511 opened on 2025/11/19

 (1 件のコメント) (0 件のリアクション) (0 人の担当者)Scala (693 件のフォーク)batch import
help wanted

Repository metrics

Stars
 (1,953 個のスター)
PR merge metrics
 (平均マージ 1d 3h) (30d で 35 merged PRs)

説明

It would be nice to have an interface that converts a SedonaSpark DataFrame to a SedonaDB DataFrame easily. Here is a current solution that works:

import sedona.db
sd = sedona.db.connect()

df = sd.create_data_frame(dataframe_to_arrow(spark_df))

This could be nice:

spark_df.to_sedonadb()

But maybe we'd have to do this:

spark_df.to_sedonadb(sd)

This would allow for cool spatial workflows, like this:

  • Read an Iceberg table with SedonaSpark and perform big data operations with a filtering operation at the end to make the data small enough to fit on a single machine
  • Convert the SedonaSpark DataFrame to SedonaDB
  • Use a library that's compatible with SedonaDB, like lonboard, to create a graph

Let me know what you think!

コントリビューターガイド