Structured Context Experiments: Graphs, Notes, and Scaffolding for Small Models

Overview Six experiments, eleven measured conditions, one conclusion: structured context is the bottleneck — not raw model capability. When a small model gets the right structure at the right time, it approaches frontier performance. When the structure is irrelevant or conflicts with existing knowledge, it becomes noise or causes regression. This page documents four experiments from the Boole Agent research line — one-shot structured note-taking with tool-based retrieval, synthetic novel-knowledge testing, and scaffolded step-by-step reasoning. The earlier multi-agent benchmark and code generation experiments are documented on their own pages. ...

July 1, 2026 · 12 min · Sheraz Mahmood