KernelGenBench evaluates LLM-generated Triton kernels across 210 operators and six heterogeneous hardware platforms. The benchmark covers both multi-source operators and multi-chip performance portability to assess the utility of agentic kernel generation.
HOW THIS AFFECTS YOU
●
builderYou can rigorously evaluate the performance and portability of LLM-generated code for specialized hardware acceleration.
●
researcherYou can benchmark the efficiency of agentic frameworks in specialized low-level programming tasks.