Recursive
Decoder layer fused attention MLP · suite of 16 cases · mean latency · run 01a0352a-cd93…
●passedReported evidence
Primary measurement
1.16msmean
Rank 2 in its comparison group · source-native comparison · observed 2026-06-08
SOL 0.5171 · 1.06× reference · 12/16 cases faster
Identity
implementationRecursive
projectRecursive
revisionsubmission-6789
workloadsuite of 16 cases · mean latency
comparison keysha256:3b70da03f9f1f5cd…
sourceNVIDIA SOL-ExecBench
external idsubmission/6789
sha256:b74b60099e7e0a799c9e75e…
Correctness
Marked passed by the source; the correctness policy was not published.
Workload
definition comparatorsol_execbench_eval
descriptionCorrectness must pass on every case in the SOL evaluation stack.
Measurements
latency · mean1.16 ms
Protocol
harnessSOL-ExecBench evaluation stack v1.0
timerharness_reported
primaryStatisticmean
correctnessReferencedefinition reference implementation
comparabilityFamilysol_execbench_leaderboard
Environment
gpuNVIDIA B200 (sm_100)
Artifacts
No artifacts published with this run.
Replications and notes
No attestations yet.
Community attestations. They never change the evidence level; only a KernelIndex-controlled rerun does.
Add a reproduction or note
Canonical manifest
Show manifestHide manifest
{
"run": {
"kind": "BenchmarkRun",
"spec": {
"status": "passed",
"timing": {
"latencyNs": {
"mean": 1156517
},
"primaryStatistic": "mean"
},
"observedAt": "2026-06-08T02:31:27.259Z",
"sourceNative": {
"source": "sol-execbench",
"metrics": {
"sol_score": 0.517096,
"latency_ms": 1.156517,
"avg_speedup": 1.0633,
"fast_1_count": 12,
"fast_1_total": 16
},
"benchmark": "019_decoder_layer_fused_attention_mlp",
"externalId": "6789"
},
"protocolDigest": "sha256:2964c5c4ed9cb0dcdc0f74674035bc71bf5e66620939e38441060f76434ed3e5",
"workloadDigest": "sha256:2c30d18c531f135a364f6d25d07af057f57534c83cb81a830062e8d02f00a973",
"environmentDigest": "sha256:5296c820e166474e590abb1beaf7728cdfe0a4bde9bf55f8966a40f4f5401621",
"implementationDigest": "sha256:9019761a8234d967b42ccec0758660583c9ff296a7b330880c0024c2d251ed3a"
},
"metadata": {
"name": "sol-submission-6789",
"title": "Recursive · 019_decoder_layer_fused_attention_mlp · B200",
"labels": {
"worker": "B200-ext-prod-2527"
}
},
"apiVersion": "kernelindex.dev/v1alpha1"
},
"protocol": {
"kind": "BenchmarkProtocol",
"spec": {
"harness": {
"name": "SOL-ExecBench evaluation stack",
"version": "v1.0"
},
"correctness": {
"reference": "definition reference implementation",
"comparator": "sol_execbench_eval"
},
"measurement": {
"timer": "harness_reported",
"primaryStatistic": "mean"
},
"comparability": {
"notes": "Suite-mean latency and SOL-Score as published by the SOL-ExecBench leaderboard. Comparable only within one definition, GPU type, and evaluation stack version.",
"family": "sol_execbench_leaderboard"
}
},
"metadata": {
"name": "sol-execbench-leaderboard-v1-0",
"title": "SOL-ExecBench evaluation stack v1.0"
},
"apiVersion": "kernelindex.dev/v1alpha1"
},
"environment": {
"kind": "ExecutionEnvironment",
"spec": {
"hardware": {
"vendor": "nvidia",
"product": "NVIDIA B200",
"architecture": "sm_100"
},
"software": {
"libraries": {
"evaluation_stack": "v1.0"
}
}
},
"metadata": {
"name": "sol-execbench-b200-v1-0",
"title": "SOL-ExecBench NVIDIA B200, stack v1.0"
},
"apiVersion": "kernelindex.dev/v1alpha1"
}
}Cite this record (permalink, digest, access date)
Report an issue with this run
Published 2026-08-24 · NVIDIA SOL-ExecBenchAll results for Decoder layer fused attention MLP →JSON