Commit Graph

89 Commits

Author SHA1 Message Date
Johnson-LYS
de98a1ad1e feat(compute): add GLM-5.3 Flash to Coding Plan (#83)
- add GLM-5.3 Flash to the Z.AI Coding Plan preset
- declare multimodal and always-on reasoning capabilities
- add regression coverage and bump preset data version
2026-08-28 14:30:03 +08:00
Johnson-LYS
d0e9bf5085 feat(compute): add DeepSeek V4 variants and Qwen3.8 Flash (#82)
- add DeepSeek V4 Flash Vision Exp provider and model specifications
- add DeepSeek V4 Flash 0731 and Qwen3.8 Flash model specifications
- add regression coverage and bump preset data version
2026-08-27 17:43:36 +08:00
c46cf07d2f feat(compute): 补 GLM-5.3 系列规格并标注 Ox Alpha 真身 (#81)
* feat(compute): 补 GLM-5.3 系列规格并标注 Ox Alpha 真身

Ox Alpha 是智谱 GLM-5.3-Flash 的匿名发布别名(canonical_slug
z-ai/glm-5.3-flash-20260826,2026-08-26 揭晓后已从 OpenRouter 列表下架)。

改动:
- model-specs/zhipu.json 新增 glm-5.3 与 glm-5.3-flash 正式规格。两者均
  强制思考(z.ai 文档:thinking.type 仅接受 enabled;OpenRouter models API:
  reasoning.mandatory=true,supported_efforts=[max,high,low]),故
  routing.reasoning.supportedModes 不含 off。按 #79 的教训只用 exact 匹配,
  避免 glm-5.3* 误吞 glm-5.3-flash。
- model-specs/stealth.json 的 ox-alpha 用 spec.extra.modelOrigin 标注真身。
  条目保留:云端算力可能仍以旧别名下发该模型。
- schemas/model-spec.schema.json 为 extra.thinkingOnly / thinkingDefault
  补正式定义与 description。此前这两个键无任何说明,客户端因此从未消费,
  用户选「关闭思考」即触发上游 400(见 desirecore#2307)。
- scripts/validate.mjs 新增静默失效键名巡检:extra 是开放对象,写错键名
  既不报错也无告警。当前巡出 17 处扁平 extra.reasoningEffort。刻意只告警
  不失败——存量取值需逐个核实各自接入面实际接受哪些 effort。

* docs(schema): 写明 extra.reasoning 与 thinkingOnly 的职责边界

provider schema 的 extra.reasoning 此前只有子字段 description、对象本身没有,
维护者看不出它与 model-spec 的 thinkingOnly 分别回答什么问题——issue
desirecore#2307 的误解正源于此。补上对象级说明:本键答「接入面接受哪些 effort
值」且只能写在 provider model;thinkingOnly 答「off 能不能用」、不声明深度档位。
2026-08-27 15:21:25 +08:00
fa2f10dd0d feat: 添加万相 3.0 视频模型配置 (#80) 2026-08-25 08:35:44 -04:00
Johnson-LYS
61b5046117 fix(compute): split suffixed model specs (#79) 2026-08-25 11:17:02 +08:00
Johnson-LYS
b8f8f0ae58 fix(compute): add explicit Ox Alpha alias (#78) 2026-08-21 17:01:38 +08:00
Johnson-LYS
e627aef16a feat(compute): add Ox Alpha model (#77) 2026-08-21 16:44:37 +08:00
ad8c515938 refactor: consolidate smart routing into model specs (#75) 2026-08-10 19:13:08 +08:00
Johnson-LYS
ed0aeccf29 feat: add qwen3.8 max preview (#67) 2026-08-10 17:10:47 +08:00
d6f939af52 feat: 下发智能路由模型目录 (#74) 2026-08-10 17:03:06 +08:00
78ed93f127 feat(compute): Claude 系列 defaultEffort 统一为 xhigh,presetDataVersion → 87 (#73)
Anthropic 对 Fable 5 / Opus 5 / Sonnet 5 的建议是编程与智能体场景用 xhigh
(Claude Code 自身默认即 xhigh),而此前沿用的是 API 默认值 high。DesireCore
是智能体操作系统,主力场景就是编程与长程智能体任务,故对齐到 xhigh。

两个 Anthropic provider 的 fable-5 / opus-5 / sonnet-5 同步调整,
extra.reasoning.defaultEffort 与顶层文档性 extra.defaultEffort 保持一致。
Haiku 4.5 不支持 effort 参数,无 reasoning 声明,不受影响。

注意:这会抬高默认 token 消耗;用户仍可在模型选择器按会话下调档位。
OpenAI Codex 的 per-model defaultEffort(low/medium/medium)是按模型分别调过的,
不属于本次对齐范围。
2026-08-08 18:52:31 +08:00
7832ae26a9 feat(compute): 下线 Opus 4.8 / Opus 4.7 / Sonnet 4.6,presetDataVersion → 86 (#72)
Anthropic provider 保留 Fable 5 / Opus 5 / Sonnet 5 / Haiku 4.5;Claude 订阅
provider 保留 Fable 5 / Opus 5 / Sonnet 5 / Haiku 4.5。

三个模型同时写入 provider.tombstones —— 仅从 models 数组删除只会触发 deprecated
软降级(保留本地数据),写入 tombstones 才是真删除,且客户端 resolver 的
resolveMappingModel() 只对 tombstone 模型执行「回退到同 serviceType 的其他可用
模型」,未 tombstone 的缺失模型会原样透传。因此写 tombstones 才能让存量会话平滑
迁移到 Opus 5,而不是继续指向一个已消失的模型名。

model-specs/anthropic.json 中三者的规格条目保留:规格库按 model_name 匹配任意
上游(含 OpenRouter 风格 vendor 前缀),与 provider 清单解耦——该文件本来就有
claude-sonnet-4-5 这类无 provider 条目的规格。删除会让经其他网关访问这些模型的
用户丢失上下文窗口/能力元数据。

用户侧 user-added / synced / ollama-discovery 来源的同名模型不受 tombstones 影响。
2026-08-08 18:44:57 +08:00
a752fddd0b feat(compute): 新增 Claude Opus 5 + 补齐 Claude 系列 effort 档位,presetDataVersion → 85 (#71)
Opus 5 是 Anthropic Opus 系列当前旗舰(1M 上下文 / 128K 输出 / $5·$25 per MTok,
与 Opus 4.8 同价),此前 config-center 完全缺失,用户在模型选择器里看不到它。

变更:
- compute/providers/anthropic.json:新增 claude-opus-5(排在 fable-5 之后、
  opus-4-8 之前);serverSideWebSearch.fallbackPriority 顺延重排
- compute/providers/anthropic-claude.json:Claude 订阅 provider 同步新增 claude-opus-5
- compute/model-specs/anthropic.json:补 claude-opus-5 与 claude-sonnet-5 规格
  (sonnet-5 此前只有 provider 条目、没有规格)
- 补 extra.reasoning.supportedEfforts:客户端 parseReasoningEffortConfig 只读
  extra.reasoning,原有的顶层 extra.defaultEffort 不被消费,导致 Claude 模型的
  effort 档位在 UI 上始终不可选。按各模型实际支持声明档位(sonnet-4-6 无 xhigh)
- __tests__/validate.test.mjs:WebSearch 允许名单登记 claude-opus-5
- manifest.json:presetDataVersion 84 → 85
2026-08-08 17:59:09 +08:00
6befa8ebb6 feat(compute): 重新发布 Claude 订阅 provider(anthropic-claude),presetDataVersion → 84 (#70)
#47/#48/#49 曾因 credentialSource 硬 enum 毒丸老客户端而被 #50 回滚。毒丸根因已由
客户端 desirecore#1021 系统性修复(credentialSource 改开放 string + 合并 salvage +
能力门控降级),claude-oauth 检测器与 SDK 网关由 desirecore#1008 提供,
requiredClientVersion 运行时门槛由 desirecore#1038 提供,三者同在 v10.0.83 发布。

本次重新发布相较 #47 的变化:
- 声明 requiredClientVersion = 10.0.83,低版本客户端优雅门控为「需更新客户端」
- displayName 去掉冗余 (订阅) 后缀,对齐 #58 的 Codex 命名约定
- 文件名 anthropic-claude.json 与 _index.json basename 一致(#48 的修正一并纳入)
2026-08-08 17:12:10 +08:00
89b292dc12 fix: correct MiMo V2.5 Pro multimodal capability (#69) 2026-07-28 09:47:26 +08:00
130b8f9ef9 feat(compute): 添加 qwen3.8-max-preview 模型 (#66)
- 新增 Qwen3.8-Max Preview 到 dashscope provider 和 model-specs
- 100万上下文,支持视觉理解(serviceType 含 vision)、Agent、推理
- supportsReasoning: true
- 定价:输入 ¥12/M token,输出 ¥36/M token
- presetDataVersion 81 → 82
2026-07-25 15:31:40 +08:00
2ddf3045d6 fix: account for native search request pricing (#65) 2026-07-24 12:28:07 +08:00
96842d2080 feat: configure tiered web search providers (#64) 2026-07-24 00:30:47 +08:00
Johnson-LYS
e59adbb166 feat: add Kimi K3 model spec (#62)
Co-authored-by: DesireCore CI <ci@desirecore.test>
2026-07-20 20:00:19 +08:00
Johnson-LYS
29ef625f32 fix(compute): match MiMo V2.5 ASR exactly (#61)
Co-authored-by: DesireCore CI <ci@desirecore.test>
2026-07-20 17:29:05 +08:00
523e667b40 fix(compute): enforce reasoning effort boundaries (#60) 2026-07-19 21:47:10 +08:00
55999ce633 feat(compute): declare GPT-5.6 reasoning efforts (#59)
Declare model-level reasoning effort capabilities for GPT-5.6 Sol, Terra, and Luna, extend validation schema, and bump preset data version.
2026-07-19 20:58:00 +08:00
d5d0dfafbd fix(compute): Codex displayName 去掉冗余 (订阅) 后缀 (#58)
* fix(compute): Codex 模型 displayName 去掉多余的"(订阅)"后缀

整个 openai-codex provider 就是 ChatGPT 订阅接入,displayName 上
逐个标注"(订阅)"是冗余信息,用户在 provider 层面已经清楚。

* chore(manifest): bump presetDataVersion 74 -> 75
2026-07-19 17:46:40 +08:00
97bde1fe00 feat(compute): Codex 补充 GPT-5.6 家族并下线三个已停用 -codex 模型 (#57)
* feat(compute): Codex 订阅接入补充 GPT-5.6 家族并下线三个已停用 -codex 模型

新增 gpt-5.6-sol/terra/luna(2026-07-09 发布)、gpt-5.4-mini、
gpt-5.3-codex-spark(ChatGPT Pro 专属研究预览);将已从 ChatGPT 订阅接入
下线的 gpt-5.3-codex/gpt-5.2-codex/gpt-5.1-codex-mini 移入 tombstones。
presetDataVersion 73 -> 74。

* fix(compute): 补全 Luna 的 reasoning 与 Spark 的 tool_use 能力标签

采纳 Codex review 意见:GPT-5.6 Luna 官方文档确认支持 reasoning token;
GPT-5.3 Codex Spark 官方公告确认可做代码编辑、按需运行测试,属于 tool_use。
2026-07-19 15:30:16 +08:00
ce1a936284 fix: restore provider-based pricing currencies (#56)
Restore MiniMax domestic CNY pricing and enforce provider-based currency classification.
2026-07-13 16:16:03 +08:00
5a9b9c87c4 fix: align DeepSeek V4 model specs with official profile
Align the shared DeepSeek V4 Pro and Flash specifications with the official 1M context, 384K output, and high/max reasoning profile. Consolidate duplicate specs, add regression coverage, and bump presetDataVersion to 72.
2026-07-10 11:47:40 +08:00
Johnson-LYS
57937a93b7 feat: add Seedance video generation models to volcengine (#52)
- Add 6 Seedance models: 2.0/2.0-fast/2.0-mini/1.5-pro/1.0-pro/1.0-pro-fast
- doubao-seedance-2.0: 4-15s 2K video, native synced audio, 9 images+3 videos+3 audio refs
- Data source: Volcengine Ark model list + seedance2-video.com version history
- Bump presetDataVersion 70→71

Co-authored-by: DesireCore CI <ci@desirecore.test>
2026-07-09 17:29:47 +08:00
6ed4b1a66a revert(compute): 回滚 Claude 订阅 provider(#47/#48/#49),presetDataVersion → 70 (#50)
* Revert "chore(compute): bump presetDataVersion to 69 (Claude provider now loads) (#49)"

This reverts commit 717ae69635.

* Revert "fix(compute): rename anthropic-claude provider file to match _index basename (#48)"

This reverts commit a1ec69ee71.

* Revert "feat(compute): 新增 Claude 订阅 provider(anthropic-claude) (#47)"

This reverts commit 1261460734.

* chore(compute): bump presetDataVersion to 70(推送回滚后的干净数据给已同步 v68/v69 的客户端)
2026-07-08 23:19:11 +08:00
717ae69635 chore(compute): bump presetDataVersion to 69 (Claude provider now loads) (#49)
PR #47 bumped to 68 but the provider file was misnamed and never loaded; the
rename (#48) kept v68, so clients that already synced v68 will not re-merge.
Bump to 69 so existing clients pick up the now-loadable anthropic-claude provider.
2026-07-07 22:52:02 +08:00
1261460734 feat(compute): 新增 Claude 订阅 provider(anthropic-claude) (#47)
* feat(compute): 新增 Claude 订阅 provider(anthropic-claude,credentialSource=claude-oauth)

* chore(schemas): credentialSource enum 增加 claude-oauth
2026-07-07 21:29:37 +08:00
c67dfc2143 chore: refresh vendor model presets
Refresh vendor model presets and bump config data version.
2026-07-07 16:31:24 +08:00
f0e14a8b5e 更新 DeepSeek 官方模型配置
更新 DeepSeek 官方模型配置
2026-07-07 15:47:06 +08:00
66113b06cf 更新 Qwen 模型配置 (#41)
更新 Qwen 模型配置
2026-07-07 15:41:24 +08:00
Johnson-LYS
1f51224901 Revert "Revert "feat(compute): 新增 ChatGPT 订阅 (Codex) Provider 预设 (#33)" (#36)" (#37)
This reverts commit d85810da79.
2026-07-04 11:31:02 +08:00
Johnson-LYS
e73b607706 feat: add HappyHorse (快乐马) video generation models (#39)
- Add happyhorse.json with 6 models: 1.1/1.0 × T2V/I2V/R2V
- T2V: text-to-video, I2V: image-to-video, R2V: reference-to-video (up to 9 images)
- Provider: HappyHorse via Alibaba Cloud Bailian
- Register in _index.json (now 22 providers)
- Bump presetDataVersion 62→63
2026-07-03 13:37:51 +08:00
Johnson-LYS
51935c0d54 chore(internal-testing): 退役内置 provider;bump presetDataVersion to 61 (#26)
* chore(internal-testing): 退役内置 provider;bump presetDataVersion to 52

internal-testing 一直作为内测阶段的硬编码 provider 存在,现在客户端已切换到
登录后绑定 desirecore-cloud(NewAPI 网关)方案,internal-testing 不再需要。

从此 commit 起:
- compute/providers/internal-testing.json 删除
- _index.json 移除 internal-testing 条目
- presetDataVersion 51 → 52,触发客户端重新合并预置

客户端需配套实现 retired-providers 清理逻辑:在 mergeBuiltinIfNewer 之后
执行一次性清理,移除本地 compute.json 中残留的 provider-internal-testing-001
+ secrets.json#internal-testing。

* chore: bump preset data version to 61

* feat: 增量同步 16 个厂商旗舰模型 (首批精选)

首次从 OpenRouter 增量同步,覆盖 13 家厂商的最新旗舰:

概要
---
- 新增 .sync-watermark 水位线文件 (1782276303 / 2026-06-24)
- presetDataVersion: 61 → 62
- 校验 70/70 全部通过

新增模型
---
OpenAI:     gpt-image-2 (image_gen), gpt-5.5, gpt-5.4
Anthropic:  claude-fable-5, claude-opus-4.8
Google:     gemini-3.5-flash, gemini-3.1-flash-image (Nano Banana 2)
DeepSeek:   deepseek-v4-pro
Qwen:       qwen3.7-max, qwen3.7-plus
MiniMax:    MiniMax-M3
Moonshot:   kimi-k2.7-code
Zhipu:      glm-5.2
xAI:        grok-4.3
Mistral:    mistral-medium-3.5
Tencent:    hy3-preview
Cohere:     north-mini-code

后续机制
---
每日 10:00 心跳检查 OpenRouter 新模型,对比 .sync-watermark
增量入库,有更新才提 PR,无更新静默。
2026-06-29 20:52:13 +08:00
Johnson-LYS
d85810da79 Revert "feat(compute): 新增 ChatGPT 订阅 (Codex) Provider 预设 (#33)" (#36)
This reverts commit d8e458c3ca.
2026-06-29 16:21:09 +08:00
Johnson-LYS
d8e458c3ca feat(compute): 新增 ChatGPT 订阅 (Codex) Provider 预设 (#33)
自动检测客户端本地 Codex CLI 的 ChatGPT 订阅授权(~/.codex/auth.json),
以用户订阅额度调用 gpt-5.x/codex 模型,无需 API Key。

- compute/providers/openai-codex.json:accessMode=coding-plan +
  credentialSource=codex-cli + apiFormat=openai-codex-responses,5 个模型,
  enabled=false(由客户端检测器按本地凭证自动启用)
- schemas/provider.schema.json:frozen baseline 新增 credentialSource 字段
- manifest.json:presetDataVersion 59 → 60

⚠️ 发布顺序:本数据引入老客户端不识别的 credentialSource 字段。必须先发布带
compute.json 韧性(容忍未知字段)的客户端并等自动更新铺开,再合并本 PR——
否则更早的客户端合并时会因严格 schema 校验失败而停收所有预设更新。
2026-06-29 16:17:40 +08:00
xyx
65a8bc4af5 feat(coding-plan): 新增百度千帆 Coding Plan 配置(presetDataVersion 58→59) (#32)
- 新增 baidu-coding.json:10 个模型(qianfan-code-latest 自动路由、
  deepseek-v3.2/v4-flash、ernie-4.5-turbo、glm-5/5.1、kimi-k2.5/k2.6、
  minimax-m2.5/m2.7)
- baseUrl: https://qianfan.baidubce.com/v2/coding
- accessMode: coding-plan
- 更新 _index.json 添加 baidu-coding 条目
2026-06-12 14:30:12 +08:00
Johnson-LYS
e84edec964 feat(model-specs): 新增模型规格库——跨 provider 模型参数统一维护(presetDataVersion 54→58)
* feat(model-specs): 新增模型规格库与 schema 契约

- compute/model-specs/:按厂商维护模型内在参数(上下文窗口/最大输出/能力/serviceType/默认温度,不含价签)
- schemas/model-spec.schema.json:Draft-07 契约,spec 允许 null(新文件不影响老客户端 frozen 契约)
- scripts/validate.mjs:pickSchemaKey 纳入 model-specs 校验
- manifest.presetDataVersion 54→55

* feat(model-specs): 新增小米 MiMo 系列模型规格;bump presetDataVersion 55→56

* feat(model-specs): 补全全量模型规格;presetDataVersion 56→57

* feat(model-specs): 新增 releasedAt/retiredAt 时间戳字段;补充 mimo 退役日期
2026-06-01 19:45:14 +08:00
9633df0219 chore(internal-testing): 默认模型改为 MiMo V2.5 Pro(小米 Pro);bump presetDataVersion to 54 (#29) 2026-05-30 15:05:04 +08:00
xyx
161eb04d39 fix(deepseek): 更新 DeepSeek 模型至 V4 系列,修正价格与参数 (#28)
- 新增 deepseek-v4-flash(主力)和 deepseek-v4-pro(旗舰)
- 上下文窗口 128K → 1M,最大输出 8K/64K → 384K
- 价格更新:V4 Flash 输入 1 元/M 输出 2 元/M;V4 Pro 输入 3 元/M 输出 6 元/M
- 保留 deepseek-chat / deepseek-reasoner 作为旧别名(标记 2026-07-24 弃用)
- bump presetDataVersion to 52

Co-authored-by: Yige <a@wyr.me>
2026-05-28 14:50:28 +08:00
xyx
55f948a725 fix: mark internal testing models as tool-capable (#27) 2026-05-28 14:49:07 +08:00
e17a00d48b chore(internal-testing): 新增 Qwen3.7 Max 并设为默认模型;bump presetDataVersion to 51 (#25) 2026-05-23 10:24:58 +08:00
cc3f0b53da chore(internal-testing): 回退 DeepSeek V4 Pro/Flash 视觉与 OCR 能力(实测不支持);bump presetDataVersion to 50 (#24) 2026-05-20 13:37:45 +08:00
b8019c4a30 chore(internal-testing): DeepSeek V4 Pro/Flash 新增视觉(图像理解)与 OCR 能力;bump presetDataVersion to 49 (#23) 2026-05-20 12:35:23 +08:00
72163bac46 chore(internal-testing): 默认模型改为 DeepSeek V4 Pro;bump presetDataVersion to 48 (#22) 2026-05-18 22:37:08 +08:00
fffa2b1980 chore(internal-testing): 新增 DeepSeek V4 Pro/Flash;bump presetDataVersion to 47 (#21)
- deepseek-v4-pro:1.6T/49B MoE 旗舰,1M 上下文,思考型(reasoning),省略温度参数
- deepseek-v4-flash:284B/13B MoE 高速版,1M 上下文,serviceType=[chat, fast]
- 价格统一保持 0,遵循内测专用 provider 惯例
- manifest: presetDataVersion 46→47,updatedAt 2026-05-12
2026-05-12 11:29:19 +08:00
xyx
00e148af8e feat: 算力配置全面审计 — 6 家供应商 + Token/Coding Plan 对齐官方文档 (#20)
Providers:
- moonshot: 移除到期模型 kimi-k2/k2-thinking,修正 k2.6 maxOutputTokens 32768、k2.5 contextWindow 262144
- tencent: 修正 hunyuan-t1-latest 上下文/输出/价格,新增 hunyuan-t1-vision/turbos-vision
- volcengine: doubao-seed-2.0-lite/mini 新增 audio_understanding/video_understanding
- internal-testing: services 扩展为全部 18 种类型,新增 mediaBaseUrl

Coding Plans:
- dashscope-token-plan: baseUrl 修正为 coding.dashscope.aliyuncs.com,移除非 Coding Plan 模型,新增 kimi-k2.5/qwen3.5-plus/qwen3-coder-plus/glm-4.7
- minimax-coding: 修正 usageTracking 端点,services 扩展为 7 种含媒体类型
- moonshot-coding: maxOutputTokens 修正为 32768
- zhipu-coding: 新增 glm-5.1/glm-5-turbo/glm-4.5-air,修正 glm-5 contextWindow 200000
- volcengine-coding: 新增 glm-5.1/kimi-k2.6/minimax-m2.7
- tencent-token: 新增腾讯云 Token Plan(10 个模型)

presetDataVersion: 45 → 46
2026-05-10 00:58:56 +08:00
xyx
82ad03d9eb feat: 新增小米 MiMo Provider 配置 (#19)
- 添加 xiaomi.json:MiMo-V2.5-Pro(chat/reasoning)+ TTS 系列(tts/voicedesign/voiceclone)
- 更新 _index.json 加载顺序
- presetDataVersion 42 → 43
2026-05-09 15:04:55 +08:00