Skip to content
KernelIndex
Search⌘K

GEMM n5120 k3072

0 eligible runs
gemm

General matrix multiply (GEMM) C = A @ B.T. Captured from Llama 3.2 3B attn.qkv_proj (fused q+k+v: 24*128 + 8*128 + 8*128 = 5120).

Current records

No published measurement for the selected workload.

Implementations

Implementation
Runtime
Best latency
Evidence
Availability

Semantics

Inputs and outputs
afp16 [m, k]
bfp16 [n, k]
cfp16 [m, n]
Axes and behavior
kconstant = 3072
mvariable
nconstant = 5120
determinismunspecified
constraintsNo mutation or aliasing
Identity
aliasgemm_n5120_k3072modelllama-3-2-3bsha25662efe03c183d…
No source imports for this operation yet.How records are decidedJSON