Complete workflow comparison
18–27%Fewer tool calls
3.4–4.7%Less returned text
30 synthetic runs, a separate experiment Method
MEASURED, NOT ASSUMED
Less repeated reading. Essential facts retained. Every number has a checkable basis.
Compared with reading the full source corpus
This compares input material, not total billing savings or model quality.
This measures input material, not total billing savings or model accuracy. Full JSON includes extra metadata. The baseline is a full-source read, not optimized search. 150 synthetic tasks · 39 active tasks · o200k_base · 5,000-token packet budget
30 synthetic runs, a separate experiment Method
Actual costs still depend on models, caching, output and rework.

Scan with WeChat to chat with AWR developers.