File size: 940 Bytes
c03e4f3
 
 
 
 
cf05ca1
 
ab65cc6
 
 
 
 
121bb1d
 
 
 
f0ca39a
1
2
3
4
5
6
7
8
9
10
11
12
13
14
15
16
17
18
---
tags:
- fp8
base_model: mgoin/Nemotron-4-340B-Base-hf
---

```
lm_eval --model vllm --model_args pretrained=/home/mgoin/code/Nemotron-4-340B-Base-hf-FP8,tensor_parallel_size=8,distributed_executor_backend="ray",max_model_len=4096,gpu_memory_utilization=0.6 --tasks truthfulqa_mc2 --num_fewshot 0 --batch_size 16
vllm (pretrained=/home/mgoin/code/Nemotron-4-340B-Base-hf-FP8,tensor_parallel_size=8,distributed_executor_backend=ray,max_model_len=4096,gpu_memory_utilization=0.6), gen_kwargs: (None), limit: None, num_fewshot: 0, batch_size: 16
|    Tasks     |Version|Filter|n-shot|Metric|   |Value |   |Stderr|
|--------------|------:|------|-----:|------|---|-----:|---|-----:|
|truthfulqa_mc2|      2|none  |     0|acc   |↑  |0.4869|±  |0.0142|
```

The [original paper](https://arxiv.org/pdf/2407.14679) evals:

![image/png](https://cdn-uploads.huggingface.co/production/uploads/60466e4b4f40b01b66151416/ub0wv8pkZVWmh3zEGie04.png)