
One frozen open-source model, sixteen agent harnesses, two task types, 150 graded runs. A controlled experiment on whether harness complexity pays — plus a structured catalog of every common agent pattern, mapped to the framework you already use.
Check out more technical deep dives on AI systems, or connect with me to discuss your AI initiatives.