Share your thoughts, 1 month free Claude Pro on usSee more
WorkDL logo mark

A100

Benchmarks

Task NameDataset NameSOTA ResultTrend
GPU Inference ThroughputA100-SXM4-40GB
Throughput (tok/s)5,081
8
LLM GenerationA100 80GB (inference)
Maximum Batch Size128
6
Inference EfficiencyA100 80 GB GPU
Latency (s)0.019
5
Model DiscoveryA100
vit_tiny Discovery Rate5
4
Showing 4 of 4 rows