Avg. Total Time
31.37s
Avg. TTFT
28.85s
Avg. Prefill TPS
111.71
Avg. Gen TPS
21.85
Context Size
32768
Quantization
r64
Engine
aphrodite
Creation Method
Merge
Model Type
Llama70B
Chat Template
Llama 3
Reasoning
No
Vision
No
Parameters
70B
Added At
3/16/2025
No Model Read Me file available.