Why a general language model cannot write the workflow for a specialised business process, and what Opus adds so that it can: a graph of how the work is actually done, and a model trained on it.
On hospital medical coding the two Opus models scored more than twice the general-purpose models on most measures, and beat the best of them, Claude 3.5 Sonnet, by 38 and 29 per cent on average.