9 comments (9 comments)0 reactions (0 reactions)1 assignee (1 assignee)Python514 forks (514 forks)auto 404
good first issuehelp wantednew-task
Repository metrics
- Stars
- 2,496 stars (2,496 stars)
- PR merge metrics
- PR metrics pending (PR metrics pending)
Description
Evaluation short description
Bias/hallucination eval
Evaluation metadata
Provide all available
Contributor guide
- Research direction
- Implement the PHARE evaluation benchmark by adding the dataset from Hugging Face (giskardai/phare), implementing the metric and task configuration in the lighteval framework, and creating an evaluation example.
- Tech stack
- python
- Domain
- machine learning
- Issue type
- Feature
- Prerequisites
- Python