Jae-Won Chung
New leaderboard prototype
b10121d
raw
history blame contribute delete
321 Bytes
{
"Model": "bigcode/starcoder2-7b",
"GPU": "NVIDIA H100 80GB HBM3",
"TP": 1,
"PP": 1,
"Energy/req (J)": 13.47550300902141,
"Avg TPOT (s)": 0.08442795061697184,
"Token tput (tok/s)": 1863.4226776197434,
"Avg Output Tokens": 90.10365853658537,
"Avg BS (reqs)": 252.87239583333334,
"Max BS (reqs)": 256
}