I did read the card before asking, and the bench block is why I asked.
Every comparison row in the Nightmedia table is labelled [base, non heretic]. Qwen3.6-27B, 35B-A3B, Qwen3.5-27B, all non heretic. So the table has fine tune plus merge plus heretic on one side and nothing but plain bases on the other. The arm I am asking about is not on the card.
By cheap I meant compute, not quality. No training, no new stage. Base plus heretic is a checkpoint that already exists in your pipeline, and your own card says testing and benching was done at each stage. I am not asking you to publish every stage. One row.
And the 2 to 4 point cost you just named is the reason to publish it, because it argues your way, not mine. ARC-C goes 0.647 to 0.711, so plus 6.4 points end to end. If heretic costs you 2 to 4 on the way through, the fine tune stages are carrying 8.4 to 10.4 and paying a tax on top. Right now the table hands that credit to a stack, and the stage that did the work is invisible inside it.
Your framing of heretic tuned versus heretic base is the same subtraction from the other side, and it works just as well. Either one needs the base plus heretic number to exist.
Still curious about per item outputs on the two quants, since mxfp8 and mxfp4 answered the same 1172 ARC-C items and a paired test is free once you have them.
Which is closer to hand, base plus heretic, or the per item dumps?