Agent Harnesses Start Improving Themselves
Agent improvement is moving from model retraining into the executable scaffold around the model. FlowEvo compiles successful workflows into persistent skills, Hierarchical Self-Improvement rewrites task-specific harnesses from environment feedback, and separate work shows that runtime budget awareness can improve tool-use scaling without changing model weights. This creates a distinct engineering layer for harness evaluation, controlled evolution, rollback and compatibility.