studio-ousia/luke-japanese-base

Model

luke-japanese

5

7 commits

2 linked in READMEs

updated Nov 9, 2022

See the code

README

luke-japanese

luke-japanese is the Japanese version of LUKE (Language Understanding with Knowledge-based Embeddings), a pre-trained knowledge-enhanced contextualized representation of words and entities. LUKE treats words and entities in a given text as independent tokens, and outputs contextualized representations of them. Please refer to our GitHub repository for more details and updates.

This model contains Wikipedia entity embeddings which are not used in general NLP tasks. Please use the lite version for tasks that do not use Wikipedia entities as inputs.

luke-japaneseは、単語とエンティティの知識拡張型訓練済みTransformerモデルLUKEの日本語版です。LUKEは単語とエンティティを独立したトークンとして扱い、これらの文脈を考慮した表現を出力します。詳細については、GitHub リポジトリを参照してください。

このモデルは、通常のNLPタスクでは使われないWikipediaエンティティのエンベディングを含んでいます。単語の入力のみを使うタスクには、lite versionを使用してください。

Experimental results on JGLUE

The experimental results evaluated on the dev set of JGLUE are shown as follows:

ModelMARC-jaJSTSJNLIJCommonsenseQA
accPearson/Spearmanaccacc
LUKE Japanese base0.9650.916/0.8770.9120.842
Baselines:
Tohoku BERT base0.9580.909/0.8680.8990.808
NICT BERT base0.9580.910/0.8710.9020.823
Waseda RoBERTa base0.9620.913/0.8730.8950.840
XLM RoBERTa base0.9610.877/0.8310.8930.687

The baseline scores are obtained from here.

Citation

@inproceedings{yamada2020luke,
  title={LUKE: Deep Contextualized Entity Representations with Entity-aware Self-attention},
  author={Ikuya Yamada and Akari Asai and Hiroyuki Shindo and Hideaki Takeda and Yuji Matsumoto},
  booktitle={EMNLP},
  year={2020}
}
endpoints_compatible
entity typing
fill-mask
luke
named entity recognition
pytorch
question answering
relation classification
transformers

studio-ousia/luke-japanese-base

Model

luke-japanese

5

7 commits

2 linked in READMEs

updated Nov 9, 2022

See the code

README

luke-japanese

luke-japanese is the Japanese version of LUKE (Language Understanding with Knowledge-based Embeddings), a pre-trained knowledge-enhanced contextualized representation of words and entities. LUKE treats words and entities in a given text as independent tokens, and outputs contextualized representations of them. Please refer to our GitHub repository for more details and updates.

This model contains Wikipedia entity embeddings which are not used in general NLP tasks. Please use the lite version for tasks that do not use Wikipedia entities as inputs.

luke-japaneseは、単語とエンティティの知識拡張型訓練済みTransformerモデルLUKEの日本語版です。LUKEは単語とエンティティを独立したトークンとして扱い、これらの文脈を考慮した表現を出力します。詳細については、GitHub リポジトリを参照してください。

このモデルは、通常のNLPタスクでは使われないWikipediaエンティティのエンベディングを含んでいます。単語の入力のみを使うタスクには、lite versionを使用してください。

Experimental results on JGLUE

The experimental results evaluated on the dev set of JGLUE are shown as follows:

ModelMARC-jaJSTSJNLIJCommonsenseQA
accPearson/Spearmanaccacc
LUKE Japanese base0.9650.916/0.8770.9120.842
Baselines:
Tohoku BERT base0.9580.909/0.8680.8990.808
NICT BERT base0.9580.910/0.8710.9020.823
Waseda RoBERTa base0.9620.913/0.8730.8950.840
XLM RoBERTa base0.9610.877/0.8310.8930.687

The baseline scores are obtained from here.

Citation

@inproceedings{yamada2020luke,
  title={LUKE: Deep Contextualized Entity Representations with Entity-aware Self-attention},
  author={Ikuya Yamada and Akari Asai and Hiroyuki Shindo and Hideaki Takeda and Yuji Matsumoto},
  booktitle={EMNLP},
  year={2020}
}
endpoints_compatible
entity typing
fill-mask
luke
named entity recognition
pytorch
question answering
relation classification
transformers