D_AU - Source files for GGUF, EXL2, AWQ, GPTQ, HQQ etc etc
Collection
Safetensor source files (by David_AU) to create different quants. Source files will be uploaded after GGUFs are uploaded.
•
30 items
•
Updated
•
1
This repo contains the full precision source code, in "safe tensors" format to generate GGUFs, GPTQ, EXL2, AWQ, HQQ and other formats. The source code can also be used directly.
For full information about this model, including:
Please go to:
[ https://huggingface.co/DavidAU/L3-SthenoMaidBlackroot-12.2B-V1-INSTRUCT-ULTRA-F32-GGUF ]
This is a merge of pre-trained language models created using mergekit.
This model was merged using the passthrough merge method.
The following models were included in the merge:
The following YAML configuration was used to produce this model:
slices:
- sources:
- model: G:/7B/Meta-Llama-3-8B-Instruct
layer_range: [0, 12]
- sources:
- model: G:/7B/L3-SthenoMaidBlackroot-8B-V1
layer_range: [6, 19]
parameters:
scale:
- filter: o_proj
value: 1
- filter: down_proj
value: 1
- value: 1
- sources:
- model: G:/7B/Meta-Llama-3-8B-Instruct
layer_range: [12, 18]
parameters:
scale:
- filter: o_proj
value: .5
- filter: down_proj
value: .5
- value: 1
- sources:
- model: G:/7B/Meta-Llama-3-8B-Instruct
layer_range: [18, 25]
parameters:
scale:
- filter: o_proj
value: .75
- filter: down_proj
value: .75
- value: 1
- sources:
- model: G:/7B/L3-SthenoMaidBlackroot-8B-V1
layer_range: [19, 32]
parameters:
scale:
- filter: o_proj
value: 1
- filter: down_proj
value: 1
- value: 1
merge_method: passthrough
dtype: float32
Detailed results can be found here
Metric | Value |
---|---|
Avg. | 18.80 |
IFEval (0-Shot) | 43.03 |
BBH (3-Shot) | 26.13 |
MATH Lvl 5 (4-Shot) | 4.08 |
GPQA (0-shot) | 4.25 |
MuSR (0-shot) | 9.62 |
MMLU-PRO (5-shot) | 25.69 |