Avg. Total Time
30.92s
Avg. TTFT
24.04s
Avg. Prefill TPS
314.12
Avg. Gen TPS
19.15
Context Size
32768
Quantization
r64
Engine
aphrodite
Creation Method
LoRA Finetune
Model Type
Llama70B
Chat Template
Llama 3
Reasoning
No
Vision
No
Parameters
70B
Added At
7/2/2025
No Model Read Me file available.