With feature parity achieved, we gave Prime Agent a runtime benchmark suite and let it hillclimb its own performance.
The agents profiled bottlenecks and tested candidate changes against the current build. Improvements that passed parity checks and independent review became the baseline for the next experiment.
The loop moved work off the startup and render paths and released memory after large sessions loaded.
