DeepSeek 计划于北京时间 2026 年 9 月 10 日前后正式发布 V4.1 Flash 模型。经内部、外部多方测试,V4.1 Flash 在性能、费用、速度、总用时等各项指标上已全面超越 V4 Pro。
秉持着对用户负责的态度,在 V4.1 Flash 正式上线之后、V4.1 Pro 上线之前,我们会将对 V4 Pro 的请求全部路由到 V4.1 Flash,并按 V4.1 Flash 单价计费。
but we might have to pause new Pro subscriptions for a bit if this continues.
GPT-6 Astra usage in Codex does not incur additional long-context multipliers above 272K input tokens. Codex does not charge for cache writes.
9 月 3 日至 9 月 20 日,GLM Coding Plan 推出夜间畅蹬活动,每晚 23 点至次日 9 点, 调用套餐中的 GLM-5.3-Flash,在 ZCode 中免费使用,在其他套餐支持的 Agent 中使用额度翻倍, 进入指定时间段自动生效,覆盖所有付费套餐用户。