This release is severe financial dumping. While American companies (OpenAI, Anthropic) are trying to recoup their giant expenditures on NVIDIA servers, Chinese developers are offering comparable B2B inference quality for mere pennies. This shatters the unit economics of Western hyperscalers. The launch of V4-Flash means that small and medium language models (SLMs) are definitively turning into a commodity, where the only competitive advantage remains the price per million tokens. Corporate clients gain the ability to radically cut OPEX on automation, abandoning expensive American APIs.
Source: DeepSeek / Artificial Analysis / Reuters
PricingDeepSeekB2BChinaMacroeconomics