Compare runs
cohort sha256:5a6850cbeaf93698… · source-native profile
ranking-v1
no ranks: not comparable
Comparison identity
operation
kernelbench-l2-14-gemm-divide-sum-scaling
workload
batch_size = 1024 · input_size = 8192 · fp32 · 28368d54
protocol
KernelBench timing scripts · 547725b0
environment
NVIDIA H100 · 2750ae38
correctness policy
52ff5e65
metric
latency mean (ns)
Context · 10 fields›
architecture
sm_90
CUDA
unknown
driver
unknown
framework
unknown
samples
100
evidence
reported
license
MIT
source
available
status
passed
observed
2026-03-05
No winner: these runs didn't measure the same thing. The rows above show what differs. Why comparable?
Add run ›
Export ›
Markdown CSV JSON
PyTorch eager · sha256:0e8e45d6974d8… · observed 2026-03-05