LazarusNLP/stsb_mt_id

Dataset

Machine Translated Indonesian STS-B

3

11 commits

4 linked in READMEs

updated Jan 6, 2024

See the code

README

Machine Translated Indonesian STS-B

We believe that a synthetic baseline is better than no baseline. Therefore, we followed approached done in the Thai Sentence Vector Benchmark project and translated the STS-B test set to Indonesian via Google Translate API. This dataset will be used to evaluate our model's Spearman correlation score on the translated test set.

You can find the latest STS results that we achieved on this dataset in Indonesian Sentence Embeddings.

Contributors

w11wo

11 commits

LazarusNLP/stsb_mt_id

Dataset

Machine Translated Indonesian STS-B

3

11 commits

4 linked in READMEs

updated Jan 6, 2024

See the code

README

Machine Translated Indonesian STS-B

We believe that a synthetic baseline is better than no baseline. Therefore, we followed approached done in the Thai Sentence Vector Benchmark project and translated the STS-B test set to Indonesian via Google Translate API. This dataset will be used to evaluate our model's Spearman correlation score on the translated test set.

You can find the latest STS results that we achieved on this dataset in Indonesian Sentence Embeddings.

Contributors

w11wo

11 commits