kernelbench.com

KernelBench cuda · RTX PRO 6000

GLM-5.2 Fused MoE Qwen 3.8 Max

in run contendeddid not score

audit verdict: contamination

The submitted kernel is a genuine live-input CUDA/PTX implementation and the archived correctness result is PASS, but this benchmark cell is contaminated and must not be published. The transcript deliberately discovers completed foreign runs for this exact problem, reads a prior run's result/check/benchmark data, skims that run's solution.py structure, and then reads lines 90-186 of the foreign solution's Model/forward implementation for design insight. The final 700-line implementation is substantially different from that 186-line prior implementation and does real MoE work, so there is no reward-hack or fake-compute finding in the submitted source itself. Nevertheless, direct foreign-solution and performance-artifact access invalidates the cell as an independent model benchmark, hence verdict=contaminated and contamination=contaminated. The observed peak_fraction 0.0924 is retained only as in-run, contended provenance; publish_grade is false.

harnessor-fable

20260803_194356_or-fable_qwen_qwen3.8-max_01_glm52_fused_moe