AIIC AI Intelligence Centre

SOURCE-LINKED INTELLIGENCE

llama3.1-8b — Server — C845A_M8_RTXPro4500_32GBx8_TRT

MLCommons MLPerf · observation · Publication date unknown

18186.2 Tokens/s; MLPerf v6.1, closed.

Read original source ↗ Open in workspace

recordType
benchmark
region
Global
value
18186.2
unit
Tokens/s
dimensions
{"division":"closed","location":"./closed/Cisco/results/C845A_M8_RTXPro4500_32GBx8_TRT/llama3.1-8b/Server/performance/run_1","metric":"Performance_Result","model":"llama3.1-8b","precision":"fp4","release":"v6.1","scenario":"Server","submissionId":null,"submitter":"Cisco","suite":"datacenter","system":"C845A_M8_RTXPro4500_32GBx8_TRT","unit":"Tokens/s"}

Evidence & attribution

MLCommons, MLPerf Inference v6.1, published submission results. Metadata adapted; scenarios and units remain separate.

License: Apache-2.0

First collected: 2026-09-19T22:50:59.123Z. This is not the publication date.