radi-cho/Meta-Llama-3-8B-FLUTE

Model

| | Wiki | C4 | PIQA | ARC-E | ARC-C | HellaSwag | Wino | Avg. |

0

4 commits

2 linked in READMEs

updated Jul 20, 2024

See the code

README

WikiC4PIQAARC-EARC-CHellaSwagWinoAvg.
Unquantized6.19.279.980.150.460.272.868.6
W4G646.119.3879.3379.7949.7459.2273.9568.41
W3G647.1311.0678.7876.2244.3756.6970.3265.28

Revisions available in this repository:

  • main (W4G64, scales learned);
  • nfl_w3g64 (W3G64, scales learned);
  • nf_w4g64 (W4G64, scales not learned);
  • nf_w3g64 (W3G64, scales not learned);

Evaluations are provided for models with learned scales.
Benchmark scores (zero-shot) are computed with lm-evaluation-harness.

endpoints_compatible
llama
safetensors
text-generation
text-generation-inference
transformers

radi-cho/Meta-Llama-3-8B-FLUTE

Model

| | Wiki | C4 | PIQA | ARC-E | ARC-C | HellaSwag | Wino | Avg. |

0

4 commits

2 linked in READMEs

updated Jul 20, 2024

See the code

README

WikiC4PIQAARC-EARC-CHellaSwagWinoAvg.
Unquantized6.19.279.980.150.460.272.868.6
W4G646.119.3879.3379.7949.7459.2273.9568.41
W3G647.1311.0678.7876.2244.3756.6970.3265.28

Revisions available in this repository:

  • main (W4G64, scales learned);
  • nfl_w3g64 (W3G64, scales learned);
  • nf_w4g64 (W4G64, scales not learned);
  • nf_w3g64 (W3G64, scales not learned);

Evaluations are provided for models with learned scales.
Benchmark scores (zero-shot) are computed with lm-evaluation-harness.

endpoints_compatible
llama
safetensors
text-generation
text-generation-inference
transformers