EhimeNLP/AcademicRoBERTa

Model

Model description

3

6 commits

1 linked in READMEs

updated Aug 6, 2024

See the code

README

Model description

We pretrained a RoBERTa-based Japanese masked language model on paper abstracts from the academic database CiNii Articles.
A Japanese Masked Language Model for Academic Domain

Vocabulary

The vocabulary consists of 32000 tokens including subwords induced by the unigram language model of sentencepiece.


license: apache-2.0
language:ja

endpoints_compatible
fill-mask
roberta
safetensors
transformers

Contributors

EhimeNLP

6 commits

EhimeNLP/AcademicRoBERTa

Model

Model description

3

6 commits

1 linked in READMEs

updated Aug 6, 2024

See the code

README

Model description

We pretrained a RoBERTa-based Japanese masked language model on paper abstracts from the academic database CiNii Articles.
A Japanese Masked Language Model for Academic Domain

Vocabulary

The vocabulary consists of 32000 tokens including subwords induced by the unigram language model of sentencepiece.


license: apache-2.0
language:ja

endpoints_compatible
fill-mask
roberta
safetensors
transformers

Contributors

EhimeNLP

6 commits