deepseek发布v4.1flash模型成本大幅降低性能超越v4pro-品玩

硅星人
 来自北京

品玩9月10日讯,据DeepSeek 官方消息,DeepSeek发布V4.1 Flash模型,具备多模态视觉理解能力。552B参数MoE架构,采用Causal-Encoder-Decoder结构,输入激活仅8B、输出激活16B,成本低于同尺寸模型,超越V4 Pro。

KV Cache缩小437倍,HBM、SSD需求降至1/4、1/8,降低Agent任务成本。API调用名改为deepseek-flash,V4 Flash系列及V4 Pro路由至V4.1 Flash。腾讯WorkBuddy、CodeBuddy与OpenCode已接入。模型已开源,新价格9月10日生效。

“特别声明:以上作品内容(包括在内的视频、图片或音频)为凤凰网旗下自媒体平台“大风号”用户上传并发布,本平台仅提供信息存储空间服务。

Notice: The content above (including the videos, pictures and audios if any) is uploaded and posted by the user of Dafeng Hao, which is a social media platform and merely provides information storage space services.”

热点新闻