long context

DeepSeek V4.1-Flash Makes AI Cost a Cache Problem
DeepSeek’s new open-weight model promises to shrink the memory bill for long-running agents. The engineering claim is meaningful, but buyers still have to price the whole workflow—not just the token line item.