Add deepseek-v4.1-flash to alibaba-cn provider - #7396
Merged
rekram1-node merged 1 commit intoSep 18, 2026
Merged
rekram1-node merged 1 commit into
rekram1-node merged 1 commit into
Conversation
Contributor
Action items
|
This file contains hidden or bidirectional Unicode text that may be interpreted or compiled differently than what appears below. To review, open the file in an editor that reveals hidden Unicode characters.
Learn more about bidirectional Unicode characters
Sign up for free
to join this conversation on GitHub.
Already have an account?
Sign in to comment
Add this suggestion to a batch that can be applied as a single commit.This suggestion is invalid because no changes were made to the code.Suggestions cannot be applied while the pull request is closed.Suggestions cannot be applied while viewing a subset of changes.Only one suggestion per line can be applied in a batch.Add this suggestion to a batch that can be applied as a single commit.Applying suggestions on deleted lines is not supported.You must change the existing code in this line in order to create a valid suggestion.Outdated suggestions cannot be applied.This suggestion has been applied or marked resolved.Suggestions cannot be applied from pending reviews.Suggestions cannot be applied on multi-line comments.Suggestions cannot be applied while the pull request is queued to merge.Suggestion cannot be applied right now. Please check back later.
Adds DeepSeek V4.1 Flash to the Alibaba Bailian pay-as-you-go provider (
alibaba-cn, Beijing), using the existing lab entrymodels/deepseek/deepseek-v4.1-flash.tomlviabase_model.Sources (accessed 2026-09-18):
Host-specific deltas vs the lab entry:
enable_thinking = true|falsetoggles thinking (the deepseek doc classifies the v4 family as hybrid-thinking models on this host), andreasoning_effortacceptslow/high/maxwith defaulthigh— the doc's capability table explicitly listslowas supported fordeepseek-v4.1-flash(only v4.1-flash, v4-flash-0731 and v4-pro-0813 supportlow).reasoning_contentstreams in deltas.deepseek-v4-flash-0731/deepseek-v4-pro-0813record the peak price), this entry records the peak price converted to USD: 0.278 / 1.111 / 0.014 (cache read).max_tokensis 393_216 shared with thinking, but the model's max output is 384_000 per the lab entry.Not added to
alibaba(international) — the international gateway's model list prices in a different rate card (Singapore peak CNY 2.188/8.75); it can be added separately once pricing is confirmed. Also not added to token-plan / coding-plan catalogs — the Beijing pay-as-you-go catalog is what the model market page lists.Related: #7148 added v4.1-flash to
alibaba-token-plan(Singapore) earlier; this fills the Beijing pay-as-you-go gap.bun validatepasses.