nm-testing/TinyLlama-1.1B-Chat-v1.0-sparse2of4_fp8_dynamic-e2e 0.7B • Updated about 19 hours ago • 57
nm-testing/Qwen2.5-0.5B-W8A8_tensor_weight_static_per_tensor_act-e2e 0.6B • Updated about 19 hours ago • 37
nm-testing/TinyLlama-1.1B-Chat-v1.0-W8A8_tensor_weight_static_per_tensor_act-e2e 1B • Updated about 19 hours ago • 59
nm-testing/TinyLlama-1.1B-Chat-v1.0-W8A8_channel_weight_static_per_tensor-e2e 1B • Updated about 19 hours ago • 57
nm-testing/TinyLlama-1.1B-Chat-v1.0-kv_cache_default_gptq_tinyllama-e2e 0.3B • Updated 2 days ago • 48