Hacktoberfest 2026: the issues maintainers tagged for October, open and beginner-friendly. Browse Hacktoberfest issues

data set switching in pandas section

Open
#69 5 comments 0 reactions 0 assignees View on GitHub

Nobody has claimed this yet.

Assessment

Difficulty
4/5
Estimated time
3-5 days
Newbie friendliness
35/100
Issue type
Documentation
Clarity
Needs clarification
Activity status
Stale
Tech stack
pandas, python

Research direction

The issue names SAFI_results.csv, SF7577.tab, the OpenRefine lesson, and the first two parts of the pandas lesson; start by comparing how each dataset is introduced and used. Review the existing discussion for the reason for switching contexts, then decide whether one dataset supports both sections. Done when the rationale is documented or the lesson consistently uses one dataset and its examples remain valid.

Written by the indexing model from the issue text.

Description

The SAFI_results.csv dataset is used in the openrefine lesson and in much of the analysis in this lesson, but just for the first two parts of pandas, it uses the SF7577.tab dataset.

Is there a reason to switch contexts? Would it be better to use one dataset throughout?

Dominant language
No language data
Stars
37
Forks
70
Avg merge
1m
Merged PRs (30d)
1

Contributor guide

Open the contributing guide

First steps

  1. Read the whole issue, then the project's contributing guide.
  2. Comment on the issue to say you are picking it up — it saves two people doing the same work.
  3. Fork the repository and make your change on a branch.
  4. Open a pull request that references the issue number.

More from datacarpentry/python-socialsci

All issues in datacarpentry/python-socialsci

Similar issues

More Data Engineering issues

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.