DeepSeek Launches V4.1-Flash and Prepares for Shanghai IPO
DeepSeek has released the DeepSeek-V4.1-Flash model, featuring native multimodality and a Mixture of Experts architecture with 552 billion parameters. The new model offers improved performance, lower costs, and reduced KV-cache requirements compared to the V4-Pro. DeepSeek has begun routing V4-Pro requests to the new model and plans to collaborate with the open-source community for inference support. Additionally, the company has reportedly hired CITIC Securities to prepare for an IPO on the Shanghai Stock Exchange's STAR Market.
Summaries are written by AI from the original article. Not investment advice.