Skip to content
KernelIndex
Search⌘K

GEMM n10240 k8192

0 eligible runs
gemm

General matrix multiply (GEMM) C = A @ B.T. Captured from Llama 3.1/3.3 70B attn.qkv_proj (fused q+k+v: 64*128 + 8*128 + 8*128 = 10240).

Current records

No published measurement for the selected workload.

Implementations

Implementation
Runtime
Best latency
Evidence
Availability

Semantics

Inputs and outputs
afp16 [m, k]
bfp16 [n, k]
cfp16 [m, n]
Axes and behavior
kconstant = 8192
mvariable
nconstant = 10240
determinismunspecified
constraintsNo mutation or aliasing
Identity
aliasgemm_n10240_k8192modelllama-3-1-70bsha25656c33b537560…
No source imports for this operation yet.How records are decidedJSON