Skip to content
KernelIndex
Search⌘K

Top p sampling from probs v129280

0 eligible runs
sampling

Top-p (nucleus) sampling from probabilities with vocab_size=129280. Filters probabilities using cumulative probability threshold, then samples from the filtered distribution. Captured from DeepSeek V3/R1.

Current records

No published measurement for the selected workload.

Implementations

Implementation
Runtime
Best latency
Evidence
Availability
cuda
—
—
Apache-2.0 · source
triton
—
—
Apache-2.0 · source
cuda
—
—
Apache-2.0 · source
cuda
—
—
Apache-2.0 · source
triton
—
—
Apache-2.0 · source
cuda
—
—
Apache-2.0 · source
triton
—
—
Apache-2.0 · source

Semantics

Inputs and outputs
probsfp32 [batch_size, vocab_size]
top_pfp32 [batch_size]
samplesint64 [batch_size]
Axes and behavior
batch_sizevariable
vocab_sizeconstant = 129280
determinismunspecified
constraintsNo mutation or aliasing
Identity
aliastop_p_sampling_from_probs_v129280modeldeepseek-v3modeldeepseek-r1sha256f1b91d34068a…
No source imports for this operation yet.How records are decidedJSON