Add dataset: diorisis_ancient_greek_corpus
Nobody has claimed this yet.
Assessment
- Difficulty
- 2/5
- Estimated time
- 1-3 hours
- Newbie friendliness
- 68/100
- Issue type
- Feature
- Clarity
- Mostly clear
- Activity status
- Active
- Domain
- data
Research direction
Start with the linked Figshare dataset and inspect existing dataset entries in the repository to find the expected contribution format and location. Done means the Diorisis Ancient Greek Corpus is added with its URL, description, text modality, and Creative Commons Attribution 4.0 licence, with any repository checks passing.
Written by the indexing model from the issue text.
Description
A URL for this dataset
https://figshare.com/articles/dataset/The_Diorisis_Ancient_Greek_Corpus/6187256/1
Dataset description
This corpus consists of 820 texts spanning between the beginnings of the Ancient Greek literary tradition (Homer) to the fifth century AD. The texts are sourced from the Perseus Canonical Greek Lit Repository, "The Little Sailing" digital library, and the Bibliotheca Augustana digital library.
The dataset includes annotations for PoS
Dataset modality
Text
Dataset licence
Creative Commons Attribution 4.0 International
Other licence
No response
How can you access this data
As a download from a repository/website
Confirm the dataset has an open licence
- To the best of my knowledge, this dataset is accessible via an open licence
Contact details for data custodian
No response
- Dominant language
- No language data
- Stars
- 91
- Forks
- 8
- PR merge metrics
- No merged PRs in 30d
Contributor guide
No contributing guide indexed for this repository
First steps
- Read the whole issue, then the project's contributing guide.
- Comment on the issue to say you are picking it up — it saves two people doing the same work.
- Fork the repository and make your change on a branch.
- Open a pull request that references the issue number.
More from bigscience-workshop/lam
-
dataset good first issue
Difficulty 2/5 1-3 hours Newbie friendliness 68/100
bigscience-workshop/lam#86 · 1 comment ·
-
candidate-dataset
Difficulty 2/5 1-3 hours Newbie friendliness 35/100
bigscience-workshop/lam#92 · 1 comment ·
-
dataset
Difficulty 2/5 1-3 hours Newbie friendliness 35/100
bigscience-workshop/lam#87 ·
-
dataset
Difficulty 4/5 3-5 days Newbie friendliness 42/100
bigscience-workshop/lam#85 · 1 reaction ·
-
Add dataset: TexBiG Opendataset
Difficulty 3/5 1-2 days Newbie friendliness 35/100
bigscience-workshop/lam#84 ·
All issues in bigscience-workshop/lam
Similar issues
-
Difficulty 2/5 1-3 hours Newbie friendliness 86/100
-
Difficulty 2/5 1-3 hours Newbie friendliness 68/100
cisagov/cyhy-reports#149 · 3 comments ·
-
Difficulty 2/5 1-3 hours Newbie friendliness 88/100
-
Difficulty 2/5 1-3 hours Newbie friendliness 78/100
-
Difficulty 1/5 Under an hour Newbie friendliness 90/100
open-compass/VLMEvalKit#1698 ·