BUT it could be a lot, lot more. How to tell if they’re a major part of the solutionFalsifiable tests or Areas to ImproveDoes a natively-trained RLM at scale beat a scaffolded frontier model? If a Prime Intellect-trained 100B native RLM outperforms Opus 5 in Prime Agent, the axis-of-scale claim is real and this becomes a training-paradigm story rather than a harness story. Prime Intellect will actually train native RLMs. What happens if you train a native RLM at frontier scale ?