fava-uw/fava-data

Dataset

15

stars

5

commits

2

linked in READMEs

Dec 1, 2024

updated

README

FAVA Datasets

FAVA datasets include: annotation data and training data.

Dataset Details

Annotation Data

The annotation dataset includes 460 annotated passages identifying and editing errors using our hallucination taxonomy. This dataset was used for the fine-grained error detection task, using the annotated passages as the gold passages.

Training Data

The training data includes 35k training instances of erroneous input and corrected output pairs using our synthetic data generation pipeline.

Contributors

abhika-m

4 commits

akariasai

1 commits

fava-uw/fava-data

Dataset

15

stars

5

commits

2

linked in READMEs

Dec 1, 2024

updated

README

FAVA Datasets

FAVA datasets include: annotation data and training data.

Dataset Details

Annotation Data

The annotation dataset includes 460 annotated passages identifying and editing errors using our hallucination taxonomy. This dataset was used for the fine-grained error detection task, using the annotated passages as the gold passages.

Training Data

The training data includes 35k training instances of erroneous input and corrected output pairs using our synthetic data generation pipeline.

Contributors

abhika-m

4 commits

akariasai

1 commits