CONNECTED THINKING

Knowledge atlas

Follow people, viewpoints and their original evidence.

1 people · 1 sources · 1 viewpoints

IN CONTEXT

KernelBench GPU kernel evaluation

Choose a viewpoint. Follow it back to the conversation.

KernelBench measures correctness and speed of generated GPU kernels

IN CONTEXTKernelBench GPU kernel evaluation1 viewpoints
2026-07-04

Equal sectors are reading positions, not rankings.

Showing 1–1 of 1 viewpoints · Newest sources first

1 / 1

Selected viewpoint

KernelBench measures correctness and speed of generated GPU kernels

KernelBench evaluates LLMs on 250 PyTorch tasks to assess their ability to generate correct and fast GPU kernels, using fast_p—the percentage of generated kernels that are both correct and faster than the baseline—as its metric.

These are individual perspectives, not a measure of consensus. Source material stays in its original language.

Supporting evidence

Harness Engineering for Self-Improvement

Original excerpt

KernelBench : evaluate correctness and speed for generated GPU kernels. 250 PyTorch tasks to evaluate whether LLM can write fast and correct kernels.
Context

The evaluation metric fast_p = the percentage of generated kernels that are correct and faster than baseline.

Publication dates describe the sources, not changes in belief.