MLC ships TIRx Harness: open compiler harness for agentic GPU kernels
MLC’s TIRx Harness pairs a thin PTX-level compiler with a kernel zoo, sync/race diagnostics, and a remote benchmark server so coding agents can iterate on GPU kernels without measurement noise. On Blackwell, agents hit family geo-mean speedups from 1.33× to 6.84× vs baselines (KDA forward 2.94× over FlashKDA; backward 6.84× over FLA).