Hacktoberfest 2026: the issues maintainers tagged for October, open and beginner-friendly. Browse Hacktoberfest issues

Add KeyValueGroupedDataset

Open
#293 1 comment 0 reactions 0 assignees View on GitHub

Nobody has claimed this yet.

Assessment

Difficulty
5/5
Estimated time
Over a week
Newbie friendliness
25/100
Issue type
Feature
Clarity
Needs clarification
Activity status
Stale
Tech stack
scala, spark
Domain
data

Research direction

Start by reading issue #163 and the Spark KeyValueGroupedDataset API, including the stated groupByKey signature. The issue names no files or tests; done means integrating KeyValueGroupedDataset into Frameless sufficiently to support the planned groupByKey goal.

Written by the indexing model from the issue text.

Description

I saw that there is a plan to implement (https://github.com/typelevel/frameless/issues/163)
KeyValueGroupedDataset<K,T> groupByKey(scala.Function1<T,K> func, Encoder evidence3)

However, to reach this goal, first, we need to integrate KeyValueGroupedDataset

Dominant language
Scala
Stars
895
Forks
135
Avg merge
1d 16h
Merged PRs (30d)
3

Contributor guide

Open the contributing guide

First steps

  1. Read the whole issue, then the project's contributing guide.
  2. Comment on the issue to say you are picking it up — it saves two people doing the same work.
  3. Fork the repository and make your change on a branch.
  4. Open a pull request that references the issue number.

More from typelevel/frameless

All issues in typelevel/frameless

Similar issues

More Scala issues

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.