An experimental AI platform called Sue was developed to optimize reasoning control — determining when to reason, how much reasoning is required, what evidence to gather, and when to avoid model calls entirely. Early benchmarking shows a strict quality score of 62.5% on the original benchmark, with further improvements observed after