Context complexity: what is the Big-O of an agent API?
TLDR;
- small, local models are a good instrument for measuring what a harness costs an agent.
- an API is only as good as its outputs’ specificity, parsimony, and truthfulness.
- obvious code repair is cheaper done in your tools than elicited from the model through another retry, provided the tool discloses what it changed.