TIRx Harness is an open compiler environment for agent-written GPU kernels. It combines a small compiler foundation, hardware and kernel references, analysis tools, and a benchmark server. The project gives an agent more than a source file and a stopwatch.
The design matters because kernel tuning is a chain of decisions. A compiler can lower an idea differently than expected. A race can hide across ordinary test runs. Another job on the same GPU can move a timing. If the agent treats each result as ground truth, the next edit can optimize noise or preserve a bug.
The MLC team reports geometric-mean speedups between 1.33x and 6.84x across the workload families it evaluated. Those are project results on named workloads and hardware, not a general promise about generated kernels. The stronger lesson is in the test setup. The harness separates analysis, correctness policy, controlled execution, and retained candidates.