DeepSeek API output speed-up: supports 500 concurrency by default, enterprises can apply for larger scale
DeepSeek announced that the API has completed output speedup and service expansion. By default, it supports 500 concurrency online at the same time. Enterprise users can apply for a larger concurrency scale online. It will be promoted simultaneously with the V4-Pro price reduction strategy.
DeepSeek API output speed-up: supports 500 concurrency by default, combined with price reduction to create a combination punch
DeepSeek announced that the API has completed output speedup and service expansion. By default, it supports 500 concurrency online at the same time. Enterprise users who need greater concurrency can apply online.
This expansion and the permanent price reduction of V4-Pro API announced the day before (to 1/4 of the original price, cache hit is only 0.025 yuan/million Tokens) form a combination of "price reduction + volume expansion". The default support of 500 concurrency means that small and medium-sized development teams can obtain production-level API throughput without additional applications, while the enterprise-level online application channel for larger concurrency reserves flexible space for medium and large customers. In conjunction with the V4-Pro output announced the day before, the price dropped to 6 yuan/million Tokens, DeepSeek is simultaneously removing the dual barriers of cost and performance for large-scale deployment of AI applications.
Reviews