hfl/chinese-alpaca-2-13b-gguf

Model

10

stars

20

commits

1

linked in READMEs

Jan 24, 2024

updated

endpoints_compatible
gguf
Browse cluster: Chinese LLM Model Distribution

README

Chinese-Alpaca-2-13B-GGUF

This repository contains the GGUF-v3 models (llama.cpp compatible) for Chinese-Alpaca-2-13B.

Performance

Metric: PPL, lower is better

Quantoriginalimatrix (-im)
Q2_K13.7636 +/- 0.1944620.6803 +/- 0.31594
Q3_K9.5388 +/- 0.130789.1016 +/- 0.12565
Q4_09.1694 +/- 0.12668-
Q4_K8.6633 +/- 0.119578.6377 +/- 0.11932
Q5_08.6745 +/- 0.12020-
Q5_K8.5161 +/- 0.117968.5210 +/- 0.11803
Q6_K8.4943 +/- 0.117598.5011 +/- 0.11775
Q8_08.4595 +/- 0.11718-
F168.4550 +/- 0.11713-

The model with -im suffix is generated with important matrix, which has generally better performance (not always though).

Others

For Hugging Face version, please see: https://huggingface.co/hfl/chinese-alpaca-2-13b

Please refer to https://github.com/ymcui/Chinese-LLaMA-Alpaca-2/ for more details.

Contributors

hfl-rc

20 commits

hfl/chinese-alpaca-2-13b-gguf

Model

10

stars

20

commits

1

linked in READMEs

Jan 24, 2024

updated

endpoints_compatible
gguf
Browse cluster: Chinese LLM Model Distribution

README

Chinese-Alpaca-2-13B-GGUF

This repository contains the GGUF-v3 models (llama.cpp compatible) for Chinese-Alpaca-2-13B.

Performance

Metric: PPL, lower is better

Quantoriginalimatrix (-im)
Q2_K13.7636 +/- 0.1944620.6803 +/- 0.31594
Q3_K9.5388 +/- 0.130789.1016 +/- 0.12565
Q4_09.1694 +/- 0.12668-
Q4_K8.6633 +/- 0.119578.6377 +/- 0.11932
Q5_08.6745 +/- 0.12020-
Q5_K8.5161 +/- 0.117968.5210 +/- 0.11803
Q6_K8.4943 +/- 0.117598.5011 +/- 0.11775
Q8_08.4595 +/- 0.11718-
F168.4550 +/- 0.11713-

The model with -im suffix is generated with important matrix, which has generally better performance (not always though).

Others

For Hugging Face version, please see: https://huggingface.co/hfl/chinese-alpaca-2-13b

Please refer to https://github.com/ymcui/Chinese-LLaMA-Alpaca-2/ for more details.

Contributors

hfl-rc

20 commits