uukuguy/CollectiveCognition-v1.1-Mistral-7B-dare-0.85

Model

2

stars

3

commits

1

linked in READMEs

Nov 24, 2023

updated

endpoints_compatible
mistral
pytorch
text-generation
text-generation-inference
transformers

README

Experiment for DARE(Drop and REscale), most of the delta parameters can be directly set to zeros without affecting the capabilities of SFT LMs and larger models can tolerate a higher proportion of discarded parameters.

weight_mask_rate: 0.85 / use_weight_rescale: True / mask_stratery: random / scaling_coefficient: 1.0

ModelAverageARCHellaSwagMMLUTruthfulQAWinograndeGSM8KDROP
Intel/neural-chat-7b-v3-159.0666.2183.6462.3759.6578.1419.5643.84
migtissera/SynthIA-7B-v1.357.1162.1283.4562.6551.3778.8517.5943.76
bhenrym14/mistral-7b-platypus-fp1656.8963.0584.1564.1145.0778.5317.3645.92
jondurbin/airoboros-m-7b-3.1.256.2461.8683.5161.9153.7577.5813.8741.2
uukuguy/speechless-code-mistral-orca-7b-v1.055.3359.6482.2561.3348.4577.518.2649.89
teknium/CollectiveCognition-v1.1-Mistral-7B53.8762.1284.1762.3557.6275.3715.6219.85
Open-Orca/Mistral-7B-SlimOrca53.3462.5483.8662.7754.2377.4321.3811.2
uukuguy/speechless-mistral-dolphin-orca-platypus-samantha-7b53.3464.3384.463.7252.5278.3721.388.66
ehartford/dolphin-2.2.1-mistral-7b53.0663.4883.8663.2853.1778.3721.088.19
teknium/CollectiveCognition-v1-Mistral-7B52.5562.3785.562.7654.4877.5817.897.22
HuggingFaceH4/zephyr-7b-alpha52.461.0184.0461.3957.978.6114.039.82
ehartford/samantha-1.2-mistral-7b52.1664.0885.0863.9150.478.5316.986.13

Contributors

uukuguy

3 commits

uukuguy/CollectiveCognition-v1.1-Mistral-7B-dare-0.85

Model

2

stars

3

commits

1

linked in READMEs

Nov 24, 2023

updated

endpoints_compatible
mistral
pytorch
text-generation
text-generation-inference
transformers

README

Experiment for DARE(Drop and REscale), most of the delta parameters can be directly set to zeros without affecting the capabilities of SFT LMs and larger models can tolerate a higher proportion of discarded parameters.

weight_mask_rate: 0.85 / use_weight_rescale: True / mask_stratery: random / scaling_coefficient: 1.0

ModelAverageARCHellaSwagMMLUTruthfulQAWinograndeGSM8KDROP
Intel/neural-chat-7b-v3-159.0666.2183.6462.3759.6578.1419.5643.84
migtissera/SynthIA-7B-v1.357.1162.1283.4562.6551.3778.8517.5943.76
bhenrym14/mistral-7b-platypus-fp1656.8963.0584.1564.1145.0778.5317.3645.92
jondurbin/airoboros-m-7b-3.1.256.2461.8683.5161.9153.7577.5813.8741.2
uukuguy/speechless-code-mistral-orca-7b-v1.055.3359.6482.2561.3348.4577.518.2649.89
teknium/CollectiveCognition-v1.1-Mistral-7B53.8762.1284.1762.3557.6275.3715.6219.85
Open-Orca/Mistral-7B-SlimOrca53.3462.5483.8662.7754.2377.4321.3811.2
uukuguy/speechless-mistral-dolphin-orca-platypus-samantha-7b53.3464.3384.463.7252.5278.3721.388.66
ehartford/dolphin-2.2.1-mistral-7b53.0663.4883.8663.2853.1778.3721.088.19
teknium/CollectiveCognition-v1-Mistral-7B52.5562.3785.562.7654.4877.5817.897.22
HuggingFaceH4/zephyr-7b-alpha52.461.0184.0461.3957.978.6114.039.82
ehartford/samantha-1.2-mistral-7b52.1664.0885.0863.9150.478.5316.986.13

Contributors

uukuguy

3 commits