Aiden Cline
debfd6339c
[vercel/interfaze] Add reasoning options
2026-06-13 16:04:27 -05:00
Aiden Cline
1ddcd9bc32
Merge pull request #2286 from anomalyco/fix/remove-glm-5.2-standard-apis
...
fix: remove GLM-5.2 from standard Z.AI APIs
2026-06-13 15:41:30 -05:00
Aiden Cline
7a9b4a167c
Merge pull request #2206 from anomalyco/feat/berget-reasoning-options-wave3
...
[berget] Add reasoning options
2026-06-13 15:40:33 -05:00
Aiden Cline
33d88e41df
Merge pull request #2203 from anomalyco/feat/the-grid-ai-reasoning-options-wave3
...
[the-grid-ai] Add reasoning options
2026-06-13 15:40:11 -05:00
Aiden Cline
300a46cfdc
Merge pull request #2201 from anomalyco/feat/neuralwatt-reasoning-options-wave3
...
[neuralwatt] Add reasoning options
2026-06-13 15:36:53 -05:00
Aiden Cline
0e3c00595b
[neuralwatt] Correct reasoning controls
2026-06-13 15:35:31 -05:00
Aiden Cline
1c9b748976
fix: remove GLM-5.2 from standard Z.AI APIs
2026-06-13 15:33:40 -05:00
Aiden Cline
cf4ccf92ae
Merge pull request #2285 from niushuai1991/add-glm-5.2-config
...
Add GLM-5.2 model config to zai and zhipuai providers
2026-06-13 15:31:44 -05:00
Aiden Cline
a4a9d29c99
[berget] Correct remaining reasoning efforts
2026-06-13 15:30:02 -05:00
Aiden Cline
4670f99aac
Merge pull request #2283 from anomalyco/fix/venice-sync-last-updated
...
fix(venice): preserve model update dates
2026-06-13 15:28:40 -05:00
Aiden Cline
3839c610a6
[berget] Remove unsupported GLM efforts
2026-06-13 15:28:06 -05:00
Aiden Cline
b1095231a5
Merge pull request #2215 from anomalyco/feat/dinference-reasoning-options-wave3
...
[dinference] Mark reasoning controls fixed
2026-06-13 15:27:56 -05:00
Aiden Cline
e8333b63a1
Merge pull request #2202 from anomalyco/feat/regolo-ai-reasoning-options-wave3
...
[regolo-ai] Add reasoning options
2026-06-13 15:27:32 -05:00
Aiden Cline
704ffc7371
Merge pull request #2229 from anomalyco/feat/anyapi-reasoning-options-wave4
...
[anyapi] Add reasoning options
2026-06-13 15:26:41 -05:00
Aiden Cline
4fdda4f262
[anyapi] Correct Claude reasoning options
2026-06-13 15:20:11 -05:00
Aiden Cline
41c243713f
[berget] Correct Kimi reasoning option
2026-06-13 15:16:12 -05:00
Aiden Cline
36024f1bd2
Merge pull request #2218 from anomalyco/feat/gmicloud-reasoning-options-wave3
...
[gmicloud] Add reasoning options
2026-06-13 15:14:50 -05:00
Aiden Cline
bfb0d3722a
[gmicloud] Correct Claude reasoning options
2026-06-13 15:12:38 -05:00
Niu Shuai
3877dcf8b6
feat: add GLM-5.2 model config to zai and zhipuai providers
2026-06-14 04:10:00 +08:00
Aiden Cline
d171755a90
Merge pull request #2220 from anomalyco/feat/azure-reasoning-options-wave3
...
[azure] Backfill DeepSeek V4 reasoning options
2026-06-13 14:39:42 -05:00
Aiden Cline
5394af4d63
Merge pull request #2189 from anomalyco/feat/upstage-reasoning-options-wave3
...
[upstage] Add reasoning options
2026-06-13 14:39:01 -05:00
Aiden Cline
b78f6fb53e
Merge pull request #2191 from anomalyco/feat/tencent-coding-plan-reasoning-options-wave3
...
[tencent-coding-plan] Add reasoning options
2026-06-13 14:36:46 -05:00
Aiden Cline
9e3c7d41ff
Merge pull request #2200 from anomalyco/feat/modelscope-reasoning-options-wave3
...
[modelscope] Add reasoning options
2026-06-13 14:36:35 -05:00
Aiden Cline
e8b1862dc2
Merge pull request #2195 from anomalyco/feat/minimax-cn-coding-plan-reasoning-options-wave3
...
[minimax-cn-coding-plan] Complete reasoning options
2026-06-13 14:36:20 -05:00
Aiden Cline
63da7583b5
Merge pull request #2188 from anomalyco/feat/moonshotai-reasoning-options-wave3
...
[moonshotai] Complete reasoning options
2026-06-13 14:36:03 -05:00
Aiden Cline
e5b1221baa
fix(venice): preserve model update dates
2026-06-13 14:31:04 -05:00
Aiden Cline
43e502a0bc
Merge pull request #2282 from anomalyco/fix/mimo-reasoning-options-audit
...
fix MiMo reasoning options across providers
2026-06-13 14:16:58 -05:00
Aiden Cline
b19a423f31
fix MiMo reasoning options across providers
2026-06-13 14:13:33 -05:00
Aiden Cline
77e04a5f1a
Merge pull request #2184 from anomalyco/feat/nova-reasoning-options-wave3
...
[nova] Add reasoning options
2026-06-13 13:32:23 -05:00
Aiden Cline
2e1a8245c6
Merge pull request #2181 from anomalyco/feat/cloudferro-sherlock-reasoning-options-wave3
...
[cloudferro-sherlock] Add reasoning options
2026-06-13 13:32:14 -05:00
Aiden Cline
b197346558
Merge pull request #2183 from anomalyco/feat/drun-reasoning-options-wave3
...
[drun] Add reasoning options
2026-06-13 13:32:03 -05:00
Aiden Cline
a04024bd21
Merge pull request #2185 from anomalyco/feat/moark-reasoning-options-wave3
...
[moark] Add reasoning options
2026-06-13 13:31:53 -05:00
Aiden Cline
0186f9e638
Merge pull request #2187 from anomalyco/feat/poolside-reasoning-options-wave3
...
[poolside] Add reasoning options
2026-06-13 13:31:35 -05:00
Aiden Cline
deb664d9c6
Merge pull request #2186 from anomalyco/feat/lucidquery-reasoning-options-wave3
...
[lucidquery] Add reasoning options
2026-06-13 13:31:26 -05:00
Aiden Cline
8ade756d9b
Merge pull request #2182 from anomalyco/feat/inception-reasoning-options-wave3
...
[inception] Add reasoning options
2026-06-13 13:31:14 -05:00
Aiden Cline
fc109cc2c2
Merge pull request #2221 from anomalyco/feat/xiaomi-token-plan-cn-reasoning-options-wave3
...
[xiaomi-token-plan] Add reasoning toggles
2026-06-13 13:30:47 -05:00
Aiden Cline
0427b955c6
Merge pull request #2222 from anomalyco/feat/ambient-reasoning-options-wave3
...
[ambient] Add reasoning options
2026-06-13 13:23:26 -05:00
Aiden Cline
e1c887294a
Merge pull request #2225 from anomalyco/feat/abacus-reasoning-options-wave4
...
[abacus] Add reasoning options
2026-06-13 13:23:10 -05:00
Aiden Cline
67cbb6e5f4
Merge pull request #2226 from anomalyco/feat/cloudflare-ai-gateway-reasoning-options-wave4
...
[cloudflare-ai-gateway] Add reasoning options
2026-06-13 13:18:42 -05:00
Aiden Cline
3f183e2962
Merge pull request #2227 from anomalyco/feat/chutes-reasoning-options-wave4
...
[chutes] Add reasoning options
2026-06-13 13:18:23 -05:00
Aiden Cline
b2ada02538
Merge pull request #2238 from anomalyco/feat/digitalocean-reasoning-options-wave4
...
[digitalocean] Add reasoning options
2026-06-13 13:16:48 -05:00
Aiden Cline
6dcd5d65a0
Merge pull request #2244 from anomalyco/feat/huggingface-reasoning-options-wave4
...
[huggingface] Add reasoning options
2026-06-13 13:14:23 -05:00
Aiden Cline
9f2265f81f
Merge pull request #2281 from anomalyco/automation/sync-models-openrouter
...
chore(sync): update OpenRouter model catalog
2026-06-13 13:13:02 -05:00
Aiden Cline
b0e4f07b73
Merge pull request #2224 from anomalyco/feat/auriko-reasoning-options-wave4
...
[auriko] Add reasoning options
2026-06-13 13:12:39 -05:00
Aiden Cline
503b07ddfa
Merge pull request #2219 from anomalyco/feat/mistral-reasoning-options-wave3
...
[mistral] Mark Magistral reasoning controls fixed
2026-06-13 13:11:11 -05:00
Aiden Cline
d13507ee94
Merge pull request #2217 from anomalyco/feat/lilac-reasoning-options-wave3
...
[lilac] Add reasoning options
2026-06-13 13:10:58 -05:00
github-actions[bot]
1f391a1908
chore(sync): update OpenRouter model catalog
2026-06-13 17:46:55 +00:00
Aiden Cline
f31519bf52
Merge pull request #2280 from CodeAnimal/az-deepseek-v4
...
Correct Azure DeepSeek V4 model prices
2026-06-13 11:25:26 -05:00
Aiden Cline
d56e8387ba
Merge pull request #2274 from anomalyco/automation/sync-models-openrouter
...
chore(sync): update OpenRouter model catalog
2026-06-13 11:25:15 -05:00
CodeAnimal
0132340a2c
Correct Azure DeepSeek V4 model prices
2026-06-13 17:23:18 +01:00
Aiden Cline
ead650e6b0
Merge pull request #2277 from ririnto/feat/zai-coding-plan-glm-5-2-anthropic
...
Add ZAI Coding Plan GLM-5.2 with Anthropic provider model
2026-06-13 11:19:37 -05:00
Aiden Cline
c3beda0405
fix(zai-coding-plan): use default GLM-5.2 provider
2026-06-13 11:17:47 -05:00
Aiden Cline
814770c1c7
Merge dev and reuse Kimi K2.7 metadata
2026-06-13 11:09:49 -05:00
Aiden Cline
e3d49ae6e2
Merge pull request #2279 from mathiasloh/feat_add_ollama_cloud_kimi_k27_code
...
feat(ollama-cloud): add kimi-k2.7-code
2026-06-13 11:07:44 -05:00
Aiden Cline
fdabca87a0
Merge pull request #2275 from jubalm/add-zai-glm-5-2
...
Add ZAI Coding Plan GLM-5.2
2026-06-13 11:07:27 -05:00
github-actions[bot]
acd9fa80ad
chore(sync): update OpenRouter model catalog
2026-06-13 15:49:42 +00:00
mathias.loh
00c5b75ed9
feat(ollama-cloud): add kimi-k2.7-code
2026-06-13 22:42:54 +08:00
ririnto
d210e45dd5
refactor(zai-coding-plan): split GLM-5.2 metadata into model + base_model reference
...
Move provider-agnostic facts (name, family, dates, capability flags,
limit, modalities) into models/zhipuai/glm-5.2.toml and reference it
via base_model in the provider TOML, which now keeps only provider-
specific fields (reasoning_options, interleaved, cost, per-model
anthropic [provider] override). Per README wrapper-provider guidance.
2026-06-13 22:39:29 +09:00
ririnto
c9e026b5de
feat(zai-coding-plan): add GLM-5.2 model metadata
2026-06-13 18:04:46 +09:00
Jubal Mabaquiao
3be537dad4
Add ZAI Coding Plan GLM-5.2
2026-06-13 16:45:57 +08:00
Aiden Cline
63a7bc11a9
Merge pull request #2216 from anomalyco/feat/inceptron-reasoning-options-wave3
...
[inceptron] Add reasoning options
2026-06-13 00:30:24 -05:00
Aiden Cline
0d2db5359d
Merge pull request #2258 from anomalyco/feat/alibaba-coding-plan-cn-reasoning-options-final
...
[alibaba-coding-plan-cn] Add reasoning options
2026-06-13 00:27:45 -05:00
Aiden Cline
a2d9f79f7e
Merge pull request #2255 from anomalyco/feat/alibaba-coding-plan-reasoning-options-final
...
[alibaba-coding-plan] Add reasoning options
2026-06-13 00:27:30 -05:00
Aiden Cline
e4d366cf4f
Merge pull request #2256 from anomalyco/feat/evroc-reasoning-options-final
...
[evroc] Add reasoning options
2026-06-13 00:27:15 -05:00
Aiden Cline
9631b93843
Merge pull request #2260 from anomalyco/feat/alibaba-token-plan-cn-reasoning-options-final
...
feat(alibaba-token-plan-cn): add reasoning options
2026-06-13 00:27:01 -05:00
Aiden Cline
5424fbd2de
Merge pull request #2270 from anomalyco/automation/sync-models-openrouter
...
chore(sync): update OpenRouter model catalog
2026-06-13 00:26:42 -05:00
Aiden Cline
73cbf85a43
Merge pull request #2272 from anomalyco/automation/sync-models-vercel
...
chore(sync): update Vercel AI Gateway model catalog
2026-06-13 00:26:27 -05:00
Aiden Cline
097afb7e31
Merge pull request #2273 from zainhas/dev
...
[Together AI] add minimax m3
2026-06-13 00:26:12 -05:00
Zain Hasan
2c5ebe8219
[Together AI] add minimax m3
2026-06-12 22:12:45 -07:00
Frank
3a7b0ba2dc
update zen models
2026-06-13 01:00:30 -04:00
github-actions[bot]
e6c0abd6a9
chore(sync): update Vercel AI Gateway model catalog
2026-06-13 03:26:15 +00:00
github-actions[bot]
5288464365
chore(sync): update OpenRouter model catalog
2026-06-13 03:26:13 +00:00
Aiden Cline
7900fcd5a6
Merge pull request #2271 from shzdehmd/dev
...
feat(fireworks-ai): adding Kimi K2.7 Code, Qwen 3.7 Plus, and Minimax M3; removing deprecated models
2026-06-12 21:51:13 -05:00
Ahmad Shahzad
2df2c13060
feat(fireworks-ai): add K2.7 Code variants, Qwen 3.7 Plus, Minimax M3; remove deprecated models
...
- Add Kimi K2.7 Code (accounts/fireworks/models/kimi-k2p7-code)
- Add Kimi K2.7 Code Fast (accounts/fireworks/routers/kimi-k2p7-code-fast)
- Add Qwen 3.7 Plus (accounts/fireworks/models/qwen3p7-plus)
- Add Minimax M3 (accounts/fireworks/models/minimax-m3)
- Remove deprecated Kimi K2.5, Minimax M2.5, and Qwen 3.6 Plus
2026-06-13 07:48:41 +05:00
Aiden Cline
f846124b1f
Merge pull request #2257 from anomalyco/feat/google-reasoning-options-final
...
[google] Complete reasoning options
2026-06-12 17:47:42 -05:00
Aiden Cline
cc81c1843f
Merge pull request #2259 from anomalyco/feat/freemodel-reasoning-options-final
...
[freemodel] Add reasoning options
2026-06-12 17:47:16 -05:00
Aiden Cline
c7823958d4
[freemodel] Add Anthropic reasoning controls
2026-06-12 17:42:25 -05:00
Aiden Cline
0178741953
[freemodel] Add OpenAI reasoning efforts
2026-06-12 17:34:32 -05:00
Aiden Cline
75db7752f5
Merge pull request #2156 from anomalyco/automation/sync-models-openrouter
...
chore(sync): update OpenRouter model catalog
2026-06-12 17:31:09 -05:00
Aiden Cline
48d1bc2ee6
Merge pull request #2265 from jcraftsman/update-umans-ai-provider
...
Update Umans AI Coding Plan + add Umans AI (pay-per-token) provider
2026-06-12 17:30:09 -05:00
Aiden Cline
d147b31d97
Merge pull request #2266 from anomalyco/automation/sync-models-vercel
...
chore(sync): update Vercel AI Gateway model catalog
2026-06-12 17:29:37 -05:00
github-actions[bot]
b6fea3d470
chore(sync): update Vercel AI Gateway model catalog
2026-06-12 21:54:31 +00:00
github-actions[bot]
d3fe5e3f4e
chore(sync): update OpenRouter model catalog
2026-06-12 21:54:30 +00:00
Frank
1280fc51bc
update go models
2026-06-12 16:24:50 -04:00
Aiden Cline
0e6590cc7a
Merge pull request #2269 from anomalyco/fix/kimi-family-sync
...
fix(sync): normalize Kimi model families
2026-06-12 14:13:38 -05:00
Aiden Cline
78cb28fe8e
fix(sync): normalize Kimi model families
2026-06-12 13:50:14 -05:00
Aiden Cline
fcd71901fc
Merge pull request #2267 from anomalyco/automation/sync-models-cloudflare-workers-ai
...
chore(sync): update Cloudflare Workers AI model catalog
2026-06-12 13:40:14 -05:00
github-actions[bot]
5ffe02911a
chore(sync): update Cloudflare Workers AI model catalog
2026-06-12 18:04:11 +00:00
Aiden Cline
add04f88a2
Merge pull request #2268 from leszek3737/zenmux/k2.7
...
Add Kimi K2.7 Code and version free model configuration
2026-06-12 12:32:38 -05:00
wassel alazhar
9fdad9f497
Add temperature = false for Kimi and Flash models (locked by Umans)
2026-06-12 19:28:31 +02:00
Leszek
1545b64010
[zenmux] Add Kimi K2.7 Code (Free) model configuration
2026-06-12 19:26:12 +02:00
Adrian Rogala
0382da663b
Merge branch 'anomalyco:dev' into zenmux/k2.7
2026-06-12 19:22:21 +02:00
Leszek
0a933025aa
[zenmux] Add kimi-k2.7-code model configuration
2026-06-12 19:20:59 +02:00
wassel alazhar
735e5f7f08
Remove umans-flash-beta from Coding Plan (deprecated, sunset 2026-06-07)
2026-06-12 19:16:05 +02:00
wassel alazhar
a414765865
Update Umans AI Coding Plan and add Umans AI (pay-per-token) provider
...
Umans AI Coding Plan (subscription):
- Add umans-kimi-k2.7 (Kimi K2.7 Code) model
- Add umans-flash-beta (deprecated alias for umans-flash)
- Fix umans-glm-5.1: add vision modality (via handoff) and interleaved reasoning
- Fix umans-flash: add interleaved reasoning field
- Fix umans-qwen3.6-35b-a3b: add interleaved reasoning field
- Add reasoning_options to all models
- Add explicit name field to models using base_model
Umans AI (new provider - pay-per-token for orgs):
- New provider for organization service-account usage
- Per-token pricing from the org billing page:
- Kimi K2.6: /bin/bash.95/.00 (input/output), /bin/bash.20 cache read
- Kimi K2.7 Code: /bin/bash.95/.00, /bin/bash.19 cache read
- GLM 5.1: .40/.40, /bin/bash.29 cache read
- Umans Flash (Qwen3.6-35B-A3B): /bin/bash.15/.00, /bin/bash.05 cache read
- Umans Coder: routes to Kimi K2.6 rates
- Same endpoint (api.code.umans.ai), different billing model
2026-06-12 19:08:15 +02:00
Aiden Cline
32f2c6de2a
Merge pull request #2263 from anomalyco/feat/kimi-k2.7-code
...
feat(models): add Kimi K2.7 Code
2026-06-12 11:36:40 -05:00
Aiden Cline
bd6fc4b145
feat(models): add Kimi K2.7 Code
2026-06-12 11:35:17 -05:00
Aiden Cline
a0957ebea8
Merge pull request #2237 from anomalyco/feat/crof-reasoning-options-wave4
...
[crof] Add reasoning options
2026-06-12 11:17:12 -05:00
Aiden Cline
8b50fa5029
[crof] Add DeepSeek V4 reasoning efforts
2026-06-12 11:01:44 -05:00
Aiden Cline
d6f87b6e64
Merge pull request #2233 from anomalyco/feat/gitlab-reasoning-options-wave4
...
[gitlab] Add reasoning options
2026-06-12 10:51:47 -05:00
Aiden Cline
a66ef9ca19
Merge pull request #2209 from anomalyco/feat/meganova-reasoning-options-wave3
...
[meganova] Add reasoning options
2026-06-12 10:30:41 -05:00
Aiden Cline
a48f0722e8
Merge pull request #2240 from anomalyco/feat/fastrouter-reasoning-options-wave4
...
[fastrouter] Add reasoning options
2026-06-12 10:28:15 -05:00
Aiden Cline
aed6a48d73
Merge pull request #2241 from anomalyco/feat/neon-reasoning-options-wave4
...
[neon] Add reasoning options
2026-06-12 10:27:49 -05:00
Aiden Cline
1b24a035da
Merge pull request #2245 from anomalyco/feat/helicone-reasoning-options-wave4
...
[helicone] Add reasoning options
2026-06-12 10:27:20 -05:00
Aiden Cline
560dceadf3
Merge pull request #2261 from anomalyco/feat/alibaba-token-plan-reasoning-options-final
...
[alibaba-token-plan] Add reasoning options
2026-06-12 10:26:40 -05:00
Aiden Cline
f3607ba954
Merge pull request #2262 from anomalyco/feat/lmstudio-reasoning-options-final
...
[lmstudio] Add GPT-OSS reasoning options
2026-06-12 10:26:18 -05:00
Aiden Cline
6b90987895
Merge pull request #2248 from anomalyco/automation/sync-models-xai
...
chore(sync): update xAI model catalog
2026-06-12 10:26:05 -05:00
github-actions[bot]
ec6b7af421
chore(sync): update xAI model catalog
2026-06-12 15:25:05 +00:00
Aiden Cline
25c2cf9224
[lmstudio] Add GPT-OSS reasoning options
2026-06-12 10:22:51 -05:00
Aiden Cline
a4cdec5316
[alibaba-token-plan] Add reasoning options
2026-06-12 10:21:30 -05:00
Aiden Cline
90f36e0550
[freemodel] Add reasoning options
2026-06-12 10:21:27 -05:00
Aiden Cline
f15aad08ef
feat(alibaba-token-plan-cn): add reasoning options
2026-06-12 10:21:26 -05:00
Aiden Cline
063429a4ae
[digitalocean] Cover remaining reasoning options
2026-06-12 10:20:56 -05:00
Aiden Cline
b21c08dc44
[alibaba-coding-plan-cn] Add reasoning options
2026-06-12 10:20:42 -05:00
Aiden Cline
592da7f45c
[google] Complete reasoning options
2026-06-12 10:20:36 -05:00
Aiden Cline
717892fd72
[evroc] Add reasoning options
2026-06-12 10:20:26 -05:00
Aiden Cline
deef87451d
[alibaba-coding-plan] Add reasoning options
2026-06-12 10:20:23 -05:00
Aiden Cline
cd3c1f22f7
Merge remote-tracking branch 'origin/dev' into feat/digitalocean-reasoning-options-wave4
2026-06-12 10:19:56 -05:00
Aiden Cline
3f9980168e
Merge pull request #2251 from houtanb/dev
...
Fix release dates
2026-06-12 09:51:58 -05:00
Jack
6f900fa761
Merge pull request #2180 from anomalyco/fix/opencode-go-minimax-m3-pricing
...
fix(opencode-go): update MiniMax M3 pricing
2026-06-12 21:09:49 +08:00
Houtan Bastani
61a153e6bb
Update Mistral Large 2411 release date
...
Set mistral-large-2411 release_date and last_updated metadata to 2024-11-18.
Evidence:
https://github.com/mistralai/platform-docs-public/blob/main/src/schema/models/models/mistral-large-2-1-24-11.ts#L9
2026-06-12 11:35:43 +02:00
Houtan Bastani
5a8f9d44d5
Update Gemini 2.5 Flash and Gemini 2.5 Pro release dates
...
Set Gemini 2.5 Flash and Gemini 2.5 Pro release_date and last_updated metadata to 2025-06-17.
Evidence:
https://ai.google.dev/gemini-api/docs/deprecations#gemini-2.5-flash-models
https://ai.google.dev/gemini-api/docs/deprecations#gemini-2.5-pro-models
2026-06-12 11:35:33 +02:00
Aiden Cline
629ce9b9f7
Merge pull request #2243 from anomalyco/feat/opencode-go-reasoning-options-wave4
...
[opencode-go] Add reasoning options
2026-06-12 00:07:39 -05:00
Aiden Cline
074022b5ad
[huggingface] Correct fixed reasoning controls
2026-06-11 23:58:31 -05:00
Aiden Cline
c562522e4f
[opencode-go] Normalize DeepSeek V4 efforts
2026-06-11 23:57:47 -05:00
Aiden Cline
dd051fd682
[helicone] Add reasoning options
2026-06-11 23:56:41 -05:00
Aiden Cline
074a288d15
[huggingface] Add reasoning options
2026-06-11 23:56:40 -05:00
Aiden Cline
2d0173d177
[opencode-go] Add reasoning options
2026-06-11 23:56:33 -05:00
Aiden Cline
df75f7c438
[neon] Add reasoning options
2026-06-11 23:56:10 -05:00
Aiden Cline
ffe754d81c
[fastrouter] Add reasoning options
2026-06-11 23:56:04 -05:00
Aiden Cline
c67d3e84ff
[digitalocean] Add reasoning options
2026-06-11 23:55:32 -05:00
Aiden Cline
2f5d53cdf3
[crof] Add reasoning options
2026-06-11 23:54:59 -05:00
Aiden Cline
08b5fd9061
[gitlab] Add reasoning options
2026-06-11 23:54:10 -05:00
Aiden Cline
5aeaaa8d47
[anyapi] Add reasoning options
2026-06-11 23:52:40 -05:00
Aiden Cline
12280f413b
[chutes] Add reasoning options
2026-06-11 23:52:31 -05:00
Aiden Cline
1a772fd297
Merge pull request #2223 from anomalyco/feat/google-vertex-reasoning-options-wave3
...
[google-vertex] Add MaaS reasoning options
2026-06-11 23:52:15 -05:00
Aiden Cline
38c708714f
[cloudflare-ai-gateway] Add reasoning options
2026-06-11 23:52:04 -05:00
Aiden Cline
13abc41ac9
[abacus] Add reasoning options
2026-06-11 23:51:42 -05:00
Aiden Cline
cf6a5e104d
[auriko] Add reasoning options
2026-06-11 23:51:41 -05:00
Aiden Cline
128e0fbd8a
[ambient] Add reasoning options
2026-06-11 23:50:16 -05:00
Aiden Cline
c63920050a
[google-vertex] Add MaaS reasoning options
2026-06-11 23:50:13 -05:00
Aiden Cline
24d2a74f4d
[xiaomi-token-plan] Add reasoning toggles
2026-06-11 23:50:05 -05:00
Aiden Cline
e62b47a005
[azure] Backfill DeepSeek V4 reasoning options
2026-06-11 23:49:52 -05:00
Aiden Cline
eac6d1aaed
[mistral] Mark Magistral reasoning controls fixed
2026-06-11 23:49:47 -05:00
Aiden Cline
a914876008
[gmicloud] Add reasoning options
2026-06-11 23:49:46 -05:00
Aiden Cline
794203e12c
[lilac] Add reasoning options
2026-06-11 23:49:45 -05:00
Aiden Cline
2624678dbe
[dinference] Mark reasoning controls fixed
2026-06-11 23:49:42 -05:00
Aiden Cline
2d18ebb301
[inceptron] Mark reasoning controls fixed
2026-06-11 23:49:40 -05:00
Aiden Cline
df45483742
Merge pull request #2211 from anomalyco/feat/wafer.ai-reasoning-options-wave3
...
[wafer.ai] Add reasoning options
2026-06-11 23:49:22 -05:00
Aiden Cline
4ecdcb8f2b
Merge pull request #2196 from anomalyco/feat/hpc-ai-reasoning-options-wave3
...
[hpc-ai] Add reasoning options
2026-06-11 23:48:29 -05:00
Aiden Cline
4644def1ae
[hpc-ai] Remove unsupported reasoning efforts
2026-06-11 23:45:03 -05:00
Aiden Cline
f9091af5ee
Merge pull request #2198 from anomalyco/feat/mixlayer-reasoning-options-wave3
...
[mixlayer] Add reasoning options
2026-06-11 23:41:20 -05:00
Aiden Cline
7022a48bfc
Merge pull request #2205 from anomalyco/feat/vultr-reasoning-options-wave3
...
[vultr] Add reasoning options
2026-06-11 23:40:45 -05:00
Aiden Cline
dda69225d8
Merge pull request #2190 from anomalyco/feat/perplexity-reasoning-options-wave3
...
[perplexity] Add reasoning options
2026-06-11 23:38:49 -05:00
Aiden Cline
2e020b1dd2
Merge pull request #2194 from anomalyco/feat/minimax-coding-plan-reasoning-options-wave3
...
[minimax-coding-plan] Complete reasoning options
2026-06-11 23:38:32 -05:00
Aiden Cline
67252bda30
Merge pull request #2197 from anomalyco/feat/submodel-reasoning-options-wave3
...
[submodel] Add reasoning options
2026-06-11 23:38:20 -05:00
Aiden Cline
ac1566f622
Merge pull request #2193 from anomalyco/feat/minimax-reasoning-options-wave3
...
[minimax] Complete reasoning options
2026-06-11 23:38:02 -05:00
Aiden Cline
48837609aa
Merge pull request #2199 from anomalyco/feat/v0-reasoning-options-wave3
...
[v0] Add reasoning options
2026-06-11 23:37:49 -05:00
Aiden Cline
c2a0ff023a
Merge pull request #2138 from martinmose/add-zeldoc-provider
...
feat(provider): add zeldoc provider
2026-06-11 23:28:07 -05:00
Aiden Cline
282821e7b6
Merge pull request #2208 from anomalyco/feat/qihang-ai-reasoning-options-wave3
...
[qihang-ai] Add reasoning options
2026-06-11 23:25:07 -05:00
Aiden Cline
1b8e53bcbd
Merge pull request #2192 from anomalyco/feat/minimax-cn-reasoning-options-wave3
...
[minimax-cn] Complete reasoning options
2026-06-11 23:14:09 -05:00
Aiden Cline
f671147f71
Merge pull request #2207 from anomalyco/feat/scaleway-reasoning-options-wave3
...
[scaleway] Add reasoning options
2026-06-11 23:13:54 -05:00
Aiden Cline
e23e759601
Merge pull request #2204 from anomalyco/feat/clarifai-reasoning-options-wave3
...
[clarifai] Add reasoning options
2026-06-11 23:13:40 -05:00
Aiden Cline
4b4546f3e3
Merge pull request #2213 from anomalyco/feat/friendli-reasoning-options-wave3
...
[friendli] Add reasoning options
2026-06-11 23:13:24 -05:00
Aiden Cline
040f0ca995
[wafer.ai] Add DeepSeek V4 effort controls
2026-06-11 23:12:29 -05:00
Aiden Cline
9acac34889
[friendli] Remove unsupported reasoning efforts
2026-06-11 23:11:37 -05:00
Aiden Cline
78ad91f773
Merge pull request #2212 from anomalyco/feat/iflowcn-reasoning-options-wave3
...
[iflowcn] Add reasoning options
2026-06-11 23:10:34 -05:00
Aiden Cline
4b0aeef538
Merge pull request #2214 from anomalyco/feat/io-net-reasoning-options-wave3
...
[io-net] Add reasoning options
2026-06-11 23:09:58 -05:00
Aiden Cline
856787cb84
[friendli] Add reasoning options
2026-06-11 23:08:36 -05:00
Aiden Cline
084f0bb4ba
[iflowcn] Add reasoning options
2026-06-11 23:08:36 -05:00
Aiden Cline
ee1301fb63
[io-net] Add reasoning options
2026-06-11 23:08:36 -05:00
Aiden Cline
502957b778
[wafer.ai] Add reasoning options
2026-06-11 23:08:29 -05:00
Aiden Cline
3fba77ee56
[the-grid-ai] Add reasoning options
2026-06-11 23:07:50 -05:00
Aiden Cline
0a6286e468
[regolo-ai] Add reasoning options
2026-06-11 23:07:50 -05:00
Aiden Cline
8af96ce933
[berget] Add reasoning options
2026-06-11 23:07:50 -05:00
Aiden Cline
4a056bd1ef
[meganova] Add reasoning options
2026-06-11 23:07:50 -05:00
Aiden Cline
9f11a93d06
[vultr] Add reasoning options
2026-06-11 23:07:50 -05:00
Aiden Cline
be8d8a2ec2
[qihang-ai] Add reasoning options
2026-06-11 23:07:50 -05:00
Aiden Cline
d4193dbad6
[scaleway] Add reasoning options
2026-06-11 23:07:50 -05:00
Aiden Cline
fb6b0985ce
[clarifai] Add reasoning options
2026-06-11 23:07:50 -05:00
Aiden Cline
506ca0f7c3
[neuralwatt] Add reasoning options
2026-06-11 23:07:49 -05:00
Aiden Cline
4cafdb31ff
[tencent-coding-plan] Add reasoning options
2026-06-11 23:07:20 -05:00
Aiden Cline
d97aedf7a1
[modelscope] Add reasoning options
2026-06-11 23:07:20 -05:00
Aiden Cline
404bba4d1f
[minimax-cn-coding-plan] Add reasoning options
2026-06-11 23:07:20 -05:00
Aiden Cline
2bb4fe287e
[hpc-ai] Add reasoning options
2026-06-11 23:07:20 -05:00
Aiden Cline
b58c396106
[mixlayer] Add reasoning options
2026-06-11 23:07:20 -05:00
Aiden Cline
575f078887
[minimax-coding-plan] Add reasoning options
2026-06-11 23:07:20 -05:00
Aiden Cline
acf194ec01
[submodel] Add reasoning options
2026-06-11 23:07:20 -05:00
Aiden Cline
c80ff1ad2c
[minimax] Add reasoning options
2026-06-11 23:07:20 -05:00
Aiden Cline
a41f6f5913
[v0] Add reasoning options
2026-06-11 23:07:20 -05:00
Aiden Cline
8dffe03a9c
[minimax-cn] Add reasoning options
2026-06-11 23:07:20 -05:00
Aiden Cline
9ea761af1a
[perplexity] Add reasoning options
2026-06-11 23:07:07 -05:00
Aiden Cline
c10ba76860
[upstage] Add reasoning options
2026-06-11 23:06:33 -05:00
Aiden Cline
30a019056f
[moonshotai] Add reasoning options
2026-06-11 23:06:33 -05:00
Aiden Cline
d7c4b11163
[nova] Add reasoning options
2026-06-11 23:06:33 -05:00
Aiden Cline
1018ddf139
[cloudferro-sherlock] Add reasoning options
2026-06-11 23:06:33 -05:00
Aiden Cline
a5e572bcf6
[drun] Add reasoning options
2026-06-11 23:06:33 -05:00
Aiden Cline
2d52062693
[moark] Add reasoning options
2026-06-11 23:06:33 -05:00
Aiden Cline
6f3b739418
[poolside] Add reasoning options
2026-06-11 23:06:33 -05:00
Aiden Cline
f52cc7e895
[lucidquery] Add reasoning options
2026-06-11 23:06:33 -05:00
Aiden Cline
473d12c44b
[inception] Add reasoning options
2026-06-11 23:06:33 -05:00
Jack
9401952857
chore(opencode-go): preserve MiniMax M3 file ending
2026-06-12 12:01:40 +08:00
Jack
623f5e8ea4
fix(opencode-go): update MiniMax M3 pricing
2026-06-12 11:59:03 +08:00
Aiden Cline
526b4cd543
Merge pull request #2179 from anomalyco/feat/deepseek-reasoning-options-wave2
...
[deepseek] Complete reasoning options
2026-06-11 22:55:29 -05:00
Aiden Cline
b7d165296c
Merge pull request #2172 from anomalyco/feat/privatemode-ai-reasoning-options
...
[privatemode-ai] Add reasoning options
2026-06-11 22:55:21 -05:00
Aiden Cline
655f757925
Merge pull request #2177 from anomalyco/feat/llmtr-reasoning-options
...
[llmtr] Complete reasoning options
2026-06-11 22:55:07 -05:00
Aiden Cline
cce129dfd9
Merge pull request #2173 from anomalyco/feat/zai-coding-plan-reasoning-options
...
[zai-coding-plan] Complete reasoning options
2026-06-11 22:54:57 -05:00
Aiden Cline
548f3b5869
Merge pull request #2175 from anomalyco/feat/zhipuai-coding-plan-reasoning-options
...
[zhipuai-coding-plan] Complete reasoning options
2026-06-11 22:54:51 -05:00
Aiden Cline
f349a9dde6
Merge pull request #2176 from anomalyco/feat/kuae-cloud-reasoning-options
...
[kuae-cloud-coding-plan] Add reasoning options
2026-06-11 22:54:34 -05:00
Aiden Cline
05c1041160
[deepseek] Mark reasoner controls fixed
2026-06-11 22:54:19 -05:00
Aiden Cline
215c3dfbc8
Merge pull request #2174 from anomalyco/feat/kimi-for-coding-reasoning-options
...
[kimi-for-coding] Complete reasoning options
2026-06-11 22:54:16 -05:00
Aiden Cline
ddc6ad75a2
Merge pull request #2178 from anomalyco/feat/firepass-reasoning-options
...
[firepass] Add reasoning options
2026-06-11 22:54:05 -05:00
Aiden Cline
757bee7de8
Merge pull request #2167 from anomalyco/feat/claudinio-reasoning-options
...
[claudinio] Add reasoning options
2026-06-11 22:53:39 -05:00
Aiden Cline
5ee32a47ae
Merge pull request #2170 from anomalyco/feat/bailing-reasoning-options
...
[bailing] Add reasoning options
2026-06-11 22:53:30 -05:00
Aiden Cline
2c22f4b488
[privatemode-ai] Add reasoning options
2026-06-11 22:51:42 -05:00
Aiden Cline
097296f6c6
[llmtr] Add reasoning options
2026-06-11 22:51:42 -05:00
Aiden Cline
2559ed64c5
[zai-coding-plan] Add reasoning options
2026-06-11 22:51:42 -05:00
Aiden Cline
c56ac404d3
[zhipuai-coding-plan] Add reasoning options
2026-06-11 22:51:42 -05:00
Aiden Cline
4681e30afb
[kuae-cloud-coding-plan] Add reasoning options
2026-06-11 22:51:42 -05:00
Aiden Cline
94cd023a88
[kimi-for-coding] Add reasoning options
2026-06-11 22:51:42 -05:00
Aiden Cline
b5c33b6347
[deepseek] Add reasoning options
2026-06-11 22:51:41 -05:00
Aiden Cline
2014d883e4
[firepass] Add reasoning options
2026-06-11 22:51:41 -05:00
Aiden Cline
2f749b7cf9
[claudinio] Add reasoning options
2026-06-11 22:51:31 -05:00
Aiden Cline
d7ab976e3c
[bailing] Add reasoning options
2026-06-11 22:51:31 -05:00
Aiden Cline
32066b7856
Merge pull request #2131 from anomalyco/feat/ovhcloud-reasoning-audit
...
feat(ovhcloud): add reasoning options
2026-06-11 22:42:30 -05:00
Aiden Cline
3ded2fae72
Merge pull request #2130 from anomalyco/audit/nebius-models
...
[nebius] Audit reasoning controls
2026-06-11 22:38:33 -05:00
Aiden Cline
9771180f83
[nebius] Correct reasoning controls
2026-06-11 22:31:54 -05:00
Aiden Cline
fa65113f37
Merge pull request #2162 from mikeyp/chore/update-digitalocean-models
...
Add anthropic-claude-fable-5 and nemotron-3-ultra-550b to DigitalOcean
2026-06-11 22:19:22 -05:00
Aiden Cline
1ab5784118
Merge pull request #2161 from andrelandgraf/feat/add-neon-provider
...
Add Neon provider
2026-06-11 22:17:07 -05:00
Mike Prasuhn
411bc157e6
Add anthropic-claude-fable-5 and nemotron-3-ultra-550b to DigitalOcean
2026-06-11 23:09:08 -04:00
Andre Landgraf
58ab76b8f3
Add Neon provider
...
Neon serves the same Databricks-backed model catalog through its
branch-scoped AI Gateway via an OpenAI-compatible endpoint, so this mirrors
the `databricks` provider's models.
- `api` uses the branch-scoped `NEON_AI_GATEWAY_BASE_URL` + the unified MLflow
OpenAI-compatible route; `NEON_AI_GATEWAY_TOKEN` is the bearer key. Both are
emitted by `neonctl env pull`.
- Model ids drop the `databricks-` prefix (the gateway accepts the bare ids),
so models resolve as `neon/claude-haiku-4-5`, `neon/gpt-5-nano`, etc.
2026-06-11 19:52:37 -07:00
Martin Mose Facondini
8f2607fcdb
fix(zeldoc): make logo black
2026-06-12 00:12:49 +02:00
Martin Mose Facondini
23e07a27d2
fix(zeldoc): correct z-code model fields
2026-06-12 00:12:44 +02:00
Aiden Cline
37e8e0cf95
Merge pull request #2086 from anomalyco/feat/openrouter-reasoning-options
...
feat(openrouter): add reasoning options
2026-06-11 16:21:15 -05:00
Aiden Cline
0e0fe311ab
Merge dev into feat/openrouter-reasoning-options
2026-06-11 16:14:33 -05:00
Aiden Cline
3d763d081e
fix(openrouter): document Claude effort mapping
2026-06-11 16:14:03 -05:00
Aiden Cline
ebbd3416e2
Merge pull request #2154 from anomalyco/automation/sync-models-openrouter
...
chore(sync): update OpenRouter model catalog
2026-06-11 16:07:40 -05:00
Aiden Cline
a3509f9a3d
Merge pull request #2155 from anomalyco/automation/sync-models-vercel
...
chore(sync): update Vercel AI Gateway model catalog
2026-06-11 16:07:24 -05:00
github-actions[bot]
71d00e7dfc
chore(sync): update Vercel AI Gateway model catalog
2026-06-11 21:06:24 +00:00
github-actions[bot]
26f7b6f6b6
chore(sync): update OpenRouter model catalog
2026-06-11 21:06:21 +00:00
Aiden Cline
270009bd93
fix(openrouter): correct reasoning controls
2026-06-11 15:03:25 -05:00
Aiden Cline
91cc389af3
fix(openrouter): correct Claude reasoning controls
2026-06-11 14:57:34 -05:00
Aiden Cline
1c7b7e8247
Merge dev into feat/openrouter-reasoning-options
2026-06-11 14:27:05 -05:00
Aiden Cline
a382026ce9
Merge pull request #2153 from anomalyco/automation/sync-models-vercel
...
chore(sync): update Vercel AI Gateway model catalog
2026-06-11 14:21:25 -05:00
github-actions[bot]
1fc7261c14
chore(sync): update Vercel AI Gateway model catalog
2026-06-11 19:15:03 +00:00
Aiden Cline
765dae9f61
Merge pull request #2134 from anomalyco/audit/deepinfra-reasoning
...
fix(deepinfra): reconcile reasoning controls
2026-06-11 13:02:22 -05:00
Aiden Cline
02cd80a2ab
Merge pull request #2032 from anthraxx/alibaba-qwen3.7-plus
...
add Qwen3.7 Plus model configuration to Alibana coding plan
2026-06-11 12:57:58 -05:00
Aiden Cline
fcbd02fa84
Merge pull request #2148 from anomalyco/automation/sync-models-venice
...
chore(sync): update Venice model catalog
2026-06-11 12:56:57 -05:00
Aiden Cline
e07a0bab97
Merge pull request #2132 from anomalyco/chore/fireworks-reasoning
...
fix(fireworks-ai): reconcile reasoning controls
2026-06-11 12:55:12 -05:00
Aiden Cline
ad166f7448
fix(fireworks-ai): verify reasoning toggles
2026-06-11 12:33:12 -05:00
github-actions[bot]
660350c5fa
chore(sync): update Venice model catalog
2026-06-11 17:29:06 +00:00
Aiden Cline
54d94bdc79
fix(deepinfra): restore Kimi K2.5 toggle
2026-06-11 12:28:40 -05:00
Aiden Cline
912ee1fe86
fix(deepinfra): restore V4 effort enum
2026-06-11 12:23:47 -05:00
Aiden Cline
9379be8911
Merge pull request #2151 from anomalyco/fix/togetherai-required-reasoning-options
...
[togetherai] Require reasoning options metadata
2026-06-11 12:21:18 -05:00
Aiden Cline
1ccb247f7f
[togetherai] Require reasoning options metadata
2026-06-11 12:05:22 -05:00
Aiden Cline
7d5469898d
[nebius] Complete reasoning option coverage
2026-06-11 12:04:48 -05:00
Aiden Cline
ec3c4ed8ea
fix(fireworks-ai): mark unresolved reasoning controls
2026-06-11 12:04:47 -05:00
Aiden Cline
2b42408582
fix(deepinfra): mark unresolved reasoning controls
2026-06-11 12:04:46 -05:00
Aiden Cline
8521822a96
Merge pull request #2135 from anomalyco/audit/togetherai-reasoning-20260610
...
Audit Together AI models and reasoning controls
2026-06-11 12:01:15 -05:00
Aiden Cline
6b62d03ac0
chore(deepinfra): remove provider test
2026-06-11 12:00:46 -05:00
Aiden Cline
9df50e0ccc
chore(ovhcloud): remove test changes
2026-06-11 12:00:37 -05:00
Aiden Cline
9b3d25aae8
[nebius] Remove provider matrix test
2026-06-11 12:00:36 -05:00
Aiden Cline
ea4d10b219
chore(fireworks-ai): remove catalog test
2026-06-11 12:00:36 -05:00
Aiden Cline
d3d163dd95
Remove Together provider matrix test
2026-06-11 12:00:35 -05:00
Aiden Cline
78f8fb92fc
Merge pull request #2129 from anomalyco/audit/cerebras-models-20260610
...
[cerebras] Refresh public model catalog
2026-06-11 12:00:23 -05:00
Aiden Cline
62b4296b3d
Delete packages/core/test/cerebras.test.ts
2026-06-11 12:00:10 -05:00
Aiden Cline
12057804f5
Merge pull request #2143 from Nindaleth/feature/ghcp-fable
...
feat(github-copilot): add Claude Fable 5 model
2026-06-11 11:48:42 -05:00
Aiden Cline
127290ebec
Merge pull request #2142 from anomalyco/automation/sync-models-openrouter
...
chore(sync): update OpenRouter model catalog
2026-06-11 11:45:55 -05:00
Aiden Cline
5abbfce7d6
Merge pull request #2120 from davidfierro/feat/snowflake-cortex-models
...
feat(snowflake-cortex): add missing but officially supported models
2026-06-11 11:45:36 -05:00
Aiden Cline
6ff6ab4028
Merge pull request #2145 from Omee11/feat/token-plan-cn
...
feat(alibaba-token-plan-cn): add Alibaba Token Plan (China) provider
2026-06-11 11:36:16 -05:00
Aiden Cline
0f3e0d7337
Merge pull request #2146 from Omee11/feat/token-plan-qwen3.7-plus
...
feat(alibaba-token-plan): add qwen3.7-plus
2026-06-11 10:51:52 -05:00
github-actions[bot]
8381089b11
chore(sync): update OpenRouter model catalog
2026-06-11 15:28:27 +00:00
Oliver Mee
98bed4baf6
feat(alibaba-token-plan-cn): add Alibaba Token Plan (China) provider
2026-06-11 18:14:20 +08:00
Oliver Mee
e55ab2daf8
feat(alibaba-token-plan): add qwen3.7-plus
2026-06-11 18:14:20 +08:00
Radek Liska
c60784fab5
feat(github-copilot): add Claude Fable 5 model
2026-06-11 09:15:02 +02:00
Aiden Cline
21581d8f4c
test(sync): cover factored reasoning overrides
2026-06-10 23:20:55 -05:00
Aiden Cline
20bfe37a53
fix(sync): resolve changed canonical base
2026-06-10 23:19:32 -05:00
Aiden Cline
fc4a781c2f
fix(ovhcloud): complete reasoning controls
2026-06-10 23:16:08 -05:00
Aiden Cline
28674e1af4
[cerebras] Assert complete resolved models
2026-06-10 23:16:01 -05:00
Aiden Cline
1e76995f8e
Test resolved Together provider matrix
2026-06-10 23:15:34 -05:00
Aiden Cline
0e371b761d
fix(sync): resolve reasoning before preservation
2026-06-10 23:14:13 -05:00
Aiden Cline
7a5bf4f56e
[cerebras] Test resolved model matrix
2026-06-10 23:13:39 -05:00
Aiden Cline
4c5b17b1db
[nebius] Test generated provider matrix
2026-06-10 23:13:35 -05:00
Aiden Cline
c4650219c3
fix(sync): drop stale reasoning options
2026-06-10 23:13:06 -05:00
Aiden Cline
9fb0474d1c
feat(ovhcloud): add reasoning options
2026-06-10 23:13:06 -05:00
Aiden Cline
8807bb0069
Correct Together reasoning and pricing metadata
2026-06-10 23:12:12 -05:00
Aiden Cline
56426cc834
fix(deepinfra): preserve cache pricing
2026-06-10 23:11:28 -05:00
Aiden Cline
18c35709c5
Merge pull request #2139 from anomalyco/fix/venice-sync-models
...
Venice: fix synced model metadata
2026-06-10 20:14:38 -05:00
Aiden Cline
b618a32341
fix(deepinfra): verify R1 reasoning controls
2026-06-10 20:13:26 -05:00
Aiden Cline
f8229eba2e
fix(deepinfra): narrow reasoning controls
2026-06-10 20:07:25 -05:00
Aiden Cline
6799ff1078
[nebius] Correct verified reasoning controls
2026-06-10 20:02:58 -05:00
Aiden Cline
236ff0e39b
fix(fireworks-ai): add Qwen reasoning budget
2026-06-10 20:02:44 -05:00
Aiden Cline
cb4fd81c59
Correct Together Qwen reasoning metadata
2026-06-10 19:51:01 -05:00
Aiden Cline
79e47be952
fix(deepinfra): use model-specific reasoning controls
2026-06-10 19:50:58 -05:00
Aiden Cline
03c161f038
[cerebras] Correct GLM reasoning option
2026-06-10 19:50:19 -05:00
Aiden Cline
38a2f09999
[nebius] Reconcile model lifecycle evidence
2026-06-10 19:49:42 -05:00
Aiden Cline
1f77766834
fix(fireworks-ai): remove unverified toggles
2026-06-10 19:48:27 -05:00
Aiden Cline
da1032a1cb
[venice] Fix synced model metadata
2026-06-10 19:46:52 -05:00
Aiden Cline
55848d41c6
Merge pull request #2123 from BlockListed/cortecs-add-claude-opus-4-8
...
add claude opus 4.8 to cortecs
2026-06-10 19:36:26 -05:00
BlockListed
e8304a0b0f
add claude opus 4.8 to cortecs
2026-06-10 23:48:02 +02:00
Martin Mose Facondini
23b4754d23
rename agentic-coding model to z-code
2026-06-10 23:06:51 +02:00
Martin Mose Facondini
20ffc3909f
add zeldoc provider with agentic-coding model
2026-06-10 23:06:21 +02:00
Aiden Cline
09c7f864f2
Audit Together AI model catalog and reasoning
2026-06-10 16:05:11 -05:00
Aiden Cline
c71d0c8065
fix(deepinfra): reconcile reasoning controls
2026-06-10 16:05:01 -05:00
Aiden Cline
5a3e0cacea
test(fireworks-ai): lock reasoning controls
2026-06-10 16:04:34 -05:00
Aiden Cline
f8ccb57731
[nebius] Audit reasoning controls
2026-06-10 16:04:09 -05:00
Aiden Cline
c0b03ed655
[cerebras] Refresh public model catalog
2026-06-10 16:03:52 -05:00
Aiden Cline
63feee7eca
Merge pull request #2128 from anomalyco/chore/close-stale-pull-requests
...
Automate stale pull request cleanup
2026-06-10 15:55:45 -05:00
Aiden Cline
233d636579
Merge pull request #2118 from dpuyosa/feat/venice-base-model
...
Venice: Update generation script to use base_model and reasoning_options
2026-06-10 15:55:01 -05:00
Aiden Cline
337d50d90f
Automate stale pull request cleanup
2026-06-10 15:54:40 -05:00
Aiden Cline
dfb3f2a421
Merge pull request #2062 from knowhycodata/add-llmtr-provider
...
feat: add LLMTR provider
2026-06-10 15:53:35 -05:00
Aiden Cline
a87f44dc06
[venice] Remove stale generated metadata
2026-06-10 14:59:51 -05:00
Aiden Cline
19cb771778
[venice] Generate metadata for new models
2026-06-10 14:40:06 -05:00
Aiden Cline
1f1a82c280
[venice] Reconcile synced model overrides
2026-06-10 14:24:32 -05:00
Aiden Cline
d9b2cd9076
Merge remote-tracking branch 'refs/remotes/contributor/feat/venice-base-model' into feat/venice-base-model
2026-06-10 14:24:21 -05:00
Aiden Cline
0ce7da24b6
[venice] Factor all models through metadata
2026-06-10 14:23:52 -05:00
Aiden Cline
44dabca8de
Merge pull request #2093 from jatingomnet/fastrouter_model_update
...
feat(sync): sync FastRouter model catalog
2026-06-10 14:22:20 -05:00
Aiden Cline
62c62ec36f
Merge pull request #2113 from anomalyco/automation/sync-models-openrouter
...
chore(sync): update OpenRouter model catalog
2026-06-10 14:21:33 -05:00
Aiden Cline
c3dcc2730f
Merge pull request #2114 from anomalyco/automation/sync-models-vercel
...
chore(sync): update Vercel AI Gateway model catalog
2026-06-10 14:21:16 -05:00
Aiden Cline
4d90460f96
Merge pull request #2127 from leszek3737/zenmux/step-free
...
feat(zenmux): remove of Step 3.5 Flash (Free) model and add Step 3.7 Flash (Free)
2026-06-10 14:16:59 -05:00
github-actions[bot]
87b37096fc
chore(sync): update OpenRouter model catalog
2026-06-10 19:10:45 +00:00
github-actions[bot]
1ea8f58f13
chore(sync): update Vercel AI Gateway model catalog
2026-06-10 19:10:44 +00:00
Aiden Cline
9a13c66393
Merge pull request #2126 from vglafirov/feat/gitlab-claude-fable-5
...
feat(gitlab): add Claude Fable 5 model
2026-06-10 11:57:34 -05:00
Leszek
f8a551f850
Remove of Step 3.5 Flash (Free) model and add Step 3.7 Flash (Free)
2026-06-10 18:55:15 +02:00
Vladimir Glafirov
1751dc2671
feat(gitlab): add Claude Fable 5 model
2026-06-10 18:50:26 +02:00
Aiden Cline
d333702d19
Merge pull request #2115 from pawelsierant/dev
...
Add Claude Fable 5 support for Azure
2026-06-10 11:40:16 -05:00
Aiden Cline
8f323c3a78
Merge pull request #2125 from Nichokas/add-freemodel-provider
...
feat(providers): add Claude Fable 5 to FreeModel
2026-06-10 11:40:03 -05:00
dpuyosa
8d3c8cfe21
Merge branch 'dev' into feat/venice-base-model
2026-06-10 18:39:23 +02:00
Aiden Cline
1f8adf44e7
Merge pull request #2091 from dpuyosa/feat/venice-models
...
Venice: Add tencent-hy3-preview and update minimax-m27
2026-06-10 11:13:10 -05:00
Aiden Cline
4201586665
[venice] Remove models missing from API
2026-06-10 11:08:15 -05:00
Nichokas
28d6a54256
feat: add Claude Fable 5 to FreeModel
2026-06-10 18:06:54 +02:00
Aiden Cline
33bceb35a5
Merge remote-tracking branch 'origin/dev' into feat/venice-base-model
...
# Conflicts:
# providers/venice/models/claude-fable-5.toml
2026-06-10 11:00:48 -05:00
Aiden Cline
25c3d6cd23
[venice] Migrate generator to sync runner
2026-06-10 10:58:12 -05:00
David Fierro Iglesias
5dcd077370
feat(snowflake-cortex): add officially supported models
2026-06-10 16:29:41 +02:00
Aiden Cline
987ca2d2f8
Merge pull request #2117 from dpuyosa/feat/venice-claude-fable-5
...
Venice: Add claude-fable-5 model
2026-06-10 09:28:58 -05:00
dpuyosa
07d2e3e23f
[venice] Update models with the new generation script version
...
- Use new `base_model` and `reasoning_options`
2026-06-10 13:37:13 +02:00
dpuyosa
5d5a421b35
[venice] Add claude-fable-5 model
...
- Add new provider model configuration inheriting from anthropic base
- Enable structured_output and define cost/modalities
- New file: providers/venice/models/claude-fable-5.toml
2026-06-10 13:15:23 +02:00
dpuyosa
63f56867b9
[venice] Add tencent-hy3-preview and update minimax-m27
...
- Inherit base_model metadata for both models
- Add reasoning_options with effort levels
- Remove redundant fields now provided by base
2026-06-10 13:06:17 +02:00
dpuyosa
d4bf232f91
[venice] Add base_model + reasoning_options to generator
...
- Derive open_weights from base model metadata when present
- Remove open_weights from baseModelOverrides and formatBaseModelToml
- Add temperature comparison in detectChanges for provider models
2026-06-10 12:59:38 +02:00
dpuyosa
0b27d6034d
[venice] Add base_model + reasoning_options to generator
...
- Add base_model lookup via models/ metadata directory
- Support new reasoning field (reasoning_options effort), audio pricing, and full TOML formatting
- Preserve existing fields and emit minimal override TOMLs when base_model present
- Update change detection and formatting for base_model mode
2026-06-10 12:42:39 +02:00
Frank
57caaf88a2
update zen models
2026-06-10 03:55:32 -04:00
Pawel Sierant
3b3d7ac3aa
Add Claude Fable 5 support for Azure
2026-06-10 08:37:24 +02:00
jatin.go
015679c420
chore(fastrouter): use base_model for grok-build-0.1 and sarvam models
...
Addresses PR #2093 review feedback to use base_model inheritance where a
canonical models/ entry exists or can be added.
- providers/fastrouter/models/x-ai/grok-build-0.1.toml: switch to
base_model = "xai/grok-build-0.1" (canonical already existed); drop
duplicated/conflicting facts.
- models/sarvam/sarvam-30b.toml, models/sarvam/sarvam-105b.toml: add new
canonical metadata so multiple sarvam-hosting providers can share it.
- providers/fastrouter/models/sarvam/sarvam-30b.toml,
providers/fastrouter/models/sarvam/sarvam-105b.toml: switch to
base_model with only [cost] override.
bun validate exits 0.
2026-06-10 11:22:29 +05:30
Aiden Cline
de6034494d
Merge pull request #2112 from anomalyco/feat/groq-reasoning-options
...
fix(groq): reconcile model catalog and reasoning options
2026-06-10 00:29:38 -05:00
Aiden Cline
feb387982b
fix(groq): reconcile active model catalog
2026-06-10 00:14:39 -05:00
Aiden Cline
3beb135e23
feat(groq): add reasoning options
2026-06-10 00:08:17 -05:00
Aiden Cline
eb2dc1750e
Merge pull request #2111 from anomalyco/feat/nvidia-reasoning-options
...
feat(nvidia): add reasoning options
2026-06-09 23:44:39 -05:00
Aiden Cline
0fa6f6a983
feat(schema): support unbounded reasoning budgets
2026-06-09 23:20:21 -05:00
Aiden Cline
aba6cae853
fix(nvidia): narrow reasoning controls
2026-06-09 23:16:13 -05:00
Aiden Cline
83e2a3437f
feat(nvidia): add reasoning options
2026-06-09 20:48:07 -05:00
Aiden Cline
f3d8034335
Merge pull request #2109 from anomalyco/fix/cloudflare-sync-reasoning-options
...
fix(sync): preserve reasoning options
2026-06-09 19:54:25 -05:00
Aiden Cline
42fbb1d9ca
Merge pull request #2067 from CodeAnimal/az-deepseek-v4
...
Azure DeepSeek-V4-Pro and DeepSeek-V4-Flash
2026-06-09 19:53:46 -05:00
Aiden Cline
431df4b758
fix(sync): preserve authored reasoning options
2026-06-09 19:51:11 -05:00
Aiden Cline
e9f798225c
Merge pull request #2090 from coder-wangbin/fix/qwen3.7-plus-params
...
fix(alibaba/qwen3.7-plus): correct max output to 64K and tier size to 256K
2026-06-09 19:47:27 -05:00
Aiden Cline
c2a85735d3
fix(cloudflare): preserve reasoning options during sync
2026-06-09 19:47:13 -05:00
Aiden Cline
8bc3d0b602
Merge pull request #2095 from anomalyco/automation/sync-models-openrouter
...
chore(sync): update OpenRouter model catalog
2026-06-09 19:45:58 -05:00
Aiden Cline
bcfcaccbd3
Merge pull request #2077 from anomalyco/feat/novita-reasoning-options
...
feat(novita-ai): add reasoning options
2026-06-09 19:45:42 -05:00
github-actions[bot]
e0f2b8a542
chore(sync): update OpenRouter model catalog
2026-06-09 23:45:17 +00:00
Aiden Cline
e0c0f0202d
fix(novita-ai): omit unusable V4 none effort
2026-06-09 17:32:35 -05:00
Aiden Cline
3bc08e0991
Merge pull request #2100 from kites262/feat/add-mimo-v25-pro-ultraspeed
...
feat(xiaomi): add mimo-v2.5-pro-ultraspeed
2026-06-09 17:28:51 -05:00
Aiden Cline
f63d764e53
Merge pull request #2108 from leszek3737/zenmux/claude-fable-5
...
Fead(zenmux): Add Claude Fable 5 support for Zenmux
2026-06-09 17:28:24 -05:00
Leszek
8a24eecc9a
Add Claude Fable 5 for Zenmux
2026-06-09 23:13:00 +02:00
Frank
fb13347f59
update zen model
2026-06-09 15:40:00 -04:00
Aiden Cline
d06fc9dfc8
fix(novita-ai): add DeepSeek V4 reasoning efforts
2026-06-09 13:47:34 -05:00
Aiden Cline
b1867ba865
Merge pull request #2104 from helloimalastair/cloudflare-aig-fable-5
...
Add Claude Fable 5 for Cloudflare AI Gateway
2026-06-09 13:46:31 -05:00
Aiden Cline
e5a8bf7e59
Merge pull request #2105 from anomalyco/automation/sync-models-vercel
...
chore(sync): update Vercel AI Gateway model catalog
2026-06-09 13:46:21 -05:00
Aiden Cline
a93d3b37f4
Merge pull request #2106 from unexge/push-tkvuvvvrvmlq
...
Add Claude Fable 5 for Amazon Bedrock
2026-06-09 13:46:06 -05:00
Aiden Cline
c08584c555
Merge pull request #2107 from vercel/fix-fable-5-reasoning-options
...
fix(vercel): declare claude-fable-5 reasoning as effort-only
2026-06-09 13:45:50 -05:00
R-Taneja
62fb2b4437
fix(vercel): declare claude-fable-5 reasoning as effort-only
...
claude-fable-5 rejects thinking.type=enabled (budget_tokens) and requires
thinking.type=adaptive + output_config.effort. Without reasoning_options,
consumers like OpenCode fall back to the legacy budget_tokens control and
the API errors. Mirror claude-opus-4-8, which is also effort-only.
2026-06-09 11:40:16 -07:00
Burak Varli
ab1b725656
Add Claude Fable 5 for Amazon Bedrock
...
Add us/eu/global cross-region inference profiles and set the knowledge
cutoff on the shared base model.
2026-06-09 18:39:45 +00:00
github-actions[bot]
6c710128ef
chore(sync): update Vercel AI Gateway model catalog
2026-06-09 17:56:48 +00:00
helloimalastair
ef15783481
add claude fable 5 for cloudflare ai gateway
2026-06-09 10:32:45 -07:00
Rohan Taneja
ea3976505e
Merge pull request #2103 from vercel/update-vercel-models-claude-fable-5
...
Add Claude Fable 5 (Vercel AI Gateway)
2026-06-09 10:32:16 -07:00
Aiden Cline
b2780222da
Merge pull request #2102 from anomalyco/add-anthropic-claude-fable-5
...
Add Claude Fable 5
2026-06-09 12:30:18 -05:00
R-Taneja
303758f170
chore(vercel): add claude-fable-5 model definition
...
Generated from the Vercel AI Gateway API (bun run vercel:generate --new-only).
2026-06-09 10:30:10 -07:00
Aiden Cline
259aff58eb
Add Claude Fable 5
2026-06-09 12:15:45 -05:00
Frank
22f6dd1b9d
update zen models
2026-06-09 12:08:10 -04:00
Aiden Cline
7c6727ffc9
Merge pull request #2094 from tomscohere/cohere-north-mini-code-1.0
...
[cohere] Add Cohere North-Mini-Code-1.0 and Cohere Command A+
2026-06-09 10:55:42 -05:00
kites262
7a25c3fe9c
feat(xiaomi): add mimo-v2.5-pro-ultraspeed
2026-06-09 23:52:25 +08:00
Aiden Cline
57e0020109
Merge pull request #2089 from RISHIKREDDYL/dev
...
fix(azure-cognitive-services): remove broken symlinks for retired xAI models
2026-06-09 10:48:29 -05:00
Aiden Cline
9b7dbfea77
refactor(cohere): use base models
2026-06-09 09:53:44 -05:00
Aiden Cline
8648cb4778
Merge pull request #2084 from anomalyco/feat/cloudflare-workers-ai-reasoning-options
...
feat(cloudflare-workers-ai): add reasoning options
2026-06-09 09:52:13 -05:00
tomscohere
ca1d731028
Update name
2026-06-09 14:21:49 +00:00
tomscohere
f0928f55be
Update North mini code to Cohere provider
2026-06-09 14:10:39 +00:00
tomscohere
8fbda849d7
Fix A+ last_updated
2026-06-09 11:55:42 +00:00
tomscohere
938ef111c4
Restore package lock
2026-06-09 11:50:56 +00:00
tomscohere
7b65a7c6de
Add North-Mini-Code and Cohere CMDA+
2026-06-09 11:50:07 +00:00
jatin.go
eaffc05654
chore(fastrouter): drop unrequested models from prior sync
...
Removes ~117 model TOMLs introduced by the merged-in big sync commit and
keeps only the 32 explicitly-requested new models. Also reverts the two
pricing changes (deepseek-r1-distill-llama-70b, z-ai/glm-5) and restores
the two previously-deleted files (moonshotai/kimi-k2.toml, z-ai/glm-4.5)
to their original pre-sync state.
bun validate exits 0.
2026-06-09 16:28:15 +05:30
jatin.go
c8adf1e849
Merge branch 'fastrouter_model_update' of https://github.com/jatingomnet/models.dev into fastrouter_model_update
2026-06-09 16:25:01 +05:30
jatin.go
8a8912c39d
feat(sync): add new FastRouter models
...
Adds 32 new model TOMLs matching the latest fastrouter.ai/models listing.
- Anthropic: claude-opus-4.8, claude-sonnet-4.6
- xAI: grok-4.3, grok-build-0.1
- OpenAI: gpt-5.5, gpt-5.5-pro, gpt-5.4-mini, gpt-5.4-nano,
gpt-5.3-codex, gpt-image-2, gpt-realtime-1.5
- Google: gemini-3.5-flash, gemini-3.1-pro-preview, gemma-4-31b-it,
gemini-3.1-flash-image-preview, gemini-3-pro-image-preview,
imagen-4.0-fast, imagen-4.0-ultra, veo3.1, veo3.1-fast, veo3.1-lite
- DeepSeek: deepseek-v4-pro
- MoonshotAI: kimi-k2.6
- Z.AI: glm-5.1
- MiniMax: minimax-m2.7, minimax-m2.7-highspeed
- Sarvam: sarvam-105b, sarvam-30b
- ByteDance: seedance-2
- Alibaba: wanx/wan-v2-6
- Leonardo.AI: lucid-origin, lucid-realism
Uses base_model inheritance where canonical models/ entries exist;
self-contained TOMLs otherwise. bun validate exits 0.
2026-06-09 16:20:29 +05:30
jatin.go
c6641ba93d
feat(sync): sync FastRouter model catalog
...
Adds ~117 new model TOMLs, removes 1 stale entry, and updates 2 pricing
files to match the current fastrouter.ai/models listing.
- Remove moonshotai/kimi-k2 (replaced by kimi-k2.5 and kimi-k2.6)
- Fix z-ai/glm-4.5 (missing .toml extension); convert to base_model ref
- Update pricing: deepseek-r1-distill-llama-70b, z-ai/glm-5
- Add Anthropic claude-opus-4.5 through claude-3-5-haiku-20241022
- Add OpenAI gpt-5.x/4.x/3.5, o-series, realtime, image, sora, embeddings
- Add Google gemini-3.x/gemma-4, imagen-4, veo2/veo3/veo3.1 families
- Add xAI grok-4.x/3.x/2, DeepSeek v3.x/v4-pro/R1 variants
- Add Qwen, MoonshotAI, MiniMax, Perplexity, Meta, Mistral, Z.AI, Sarvam
- Add FLUX, ByteDance seedream/seedance, Leonardo AI, Kling, Runway,
Pika, Pollo, Vidu, Wanx video/image models and ace-step audio
Uses base_model inheritance where canonical models/ entries exist;
self-contained TOMLs otherwise. bun validate exits 0.
2026-06-09 15:40:27 +05:30
dpuyosa
fc7493b27b
[venice] Add tencent-hy3-preview and update minimax-m27
...
- Add new tencent-hy3-preview model with cost/limit/modality config
- Update minimax-m27 last_updated and cache_read pricing
2026-06-09 11:56:16 +02:00
wangbin
d4d0b483b5
fix(alibaba/qwen3.7-plus): correct output to 64K and tier size to 256K
...
- output: 16,384 → 65,536 (official max output is 64K)
- tier.size: 128,000 → 256,000 (matches qwen3.6-plus tier threshold)
Verified against official spec:
https://bailian.console.aliyun.com/cn-beijing/?tab=model#/model-market/detail/qwen3.7-plus
Context: 1M | Max Output: 64K | Modalities: text + image + video
2026-06-09 16:57:14 +08:00
CodeAnimal
f83b8ea4a1
Introduce base_model and other corrections based on feedback
2026-06-09 09:53:54 +01:00
Ubuntu
b66908a347
fix(azure-cognitive-services): remove broken symlinks for retired xAI models
2026-06-09 11:54:19 +05:30
Aiden Cline
37b1d0ac95
fix(cloudflare-workers-ai): preserve schema compatibility
2026-06-08 23:19:03 -05:00
Aiden Cline
f674b240c3
Merge pull request #2087 from anomalyco/automation/sync-models-openrouter
...
chore(sync): update OpenRouter model catalog
2026-06-08 23:03:09 -05:00
Aiden Cline
b233bbdb81
Merge pull request #2075 from anomalyco/feat/azure-reasoning-options
...
feat(azure): add reasoning options
2026-06-08 23:02:50 -05:00
Aiden Cline
2817f22c8a
fix(azure): expose Kimi reasoning toggles
2026-06-08 23:02:12 -05:00
Aiden Cline
ec05c9d5f1
Merge pull request #2088 from anomalyco/feat/baseten-model-sync
...
feat(sync): add Baseten model sync
2026-06-08 22:53:56 -05:00
Aiden Cline
f96381a34d
feat(sync): add Baseten model sync
2026-06-08 22:51:22 -05:00
Aiden Cline
aeecf3b66f
fix(novita-ai): add GPT OSS reasoning efforts
2026-06-08 22:41:16 -05:00
Aiden Cline
a963e6ae00
Merge pull request #2078 from anomalyco/feat/baseten-reasoning-options
...
feat(baseten): add reasoning options
2026-06-08 22:28:13 -05:00
github-actions[bot]
075fd9263e
chore(sync): update OpenRouter model catalog
2026-06-09 03:25:49 +00:00
Aiden Cline
f4bea4e831
Merge pull request #2081 from anomalyco/feat/vertex-reasoning-options
...
feat(google-vertex): add reasoning options
2026-06-08 22:13:28 -05:00
Aiden Cline
dcae17ea26
fix(google-vertex): retain latest Gemini aliases
2026-06-08 22:12:00 -05:00
Aiden Cline
ffedd884f7
fix(google-vertex): remove retired models
2026-06-08 22:06:17 -05:00
Aiden Cline
e885955e2f
Merge pull request #2079 from anomalyco/feat/ollama-cloud-reasoning-options
...
feat(ollama-cloud): add reasoning options
2026-06-08 21:57:41 -05:00
Aiden Cline
bddf7e070c
Merge pull request #2083 from anomalyco/feat/deepinfra-reasoning-options
...
feat(deepinfra): add reasoning options
2026-06-08 21:56:56 -05:00
Aiden Cline
80c55103dd
fix(deepinfra): restore DeepSeek V4 effort controls
2026-06-08 21:13:16 -05:00
Aiden Cline
9a8efd2f2f
fix(ollama-cloud): expose MiniMax M3 reasoning controls
2026-06-08 21:00:49 -05:00
Aiden Cline
919ee8da23
fix(ollama-cloud): expose DeepSeek max reasoning
2026-06-08 20:49:29 -05:00
Aiden Cline
71f76d7a1b
Merge pull request #2080 from anomalyco/feat/fireworks-reasoning-options
...
feat(fireworks-ai): add reasoning options
2026-06-08 20:46:48 -05:00
Aiden Cline
b6e8a23d76
Merge pull request #2085 from anomalyco/feat/xai-reasoning-options
...
feat(xai): add reasoning options
2026-06-08 20:36:44 -05:00
Aiden Cline
f1fb54c7ba
test(xai): reflect language model sync fields
2026-06-08 20:34:29 -05:00
Aiden Cline
9d6cfa4c3f
fix(xai): preserve reasoning options during sync
2026-06-08 20:30:35 -05:00
Aiden Cline
5e7769cb71
fix(openrouter): expose Claude Opus effort
2026-06-08 20:22:10 -05:00
Aiden Cline
add3daacea
feat(openrouter): add reasoning options
2026-06-08 20:17:16 -05:00
Aiden Cline
83faa5efe4
feat(xai): add reasoning options
2026-06-08 20:17:02 -05:00
Aiden Cline
bf9b74e973
feat(cloudflare-workers-ai): add reasoning options
2026-06-08 20:16:58 -05:00
Aiden Cline
077a047eb0
feat(deepinfra): add reasoning options
2026-06-08 20:16:49 -05:00
Aiden Cline
e696b33e0f
feat(google-vertex): add reasoning options
2026-06-08 20:16:48 -05:00
Aiden Cline
d1f12dd63a
feat(fireworks-ai): add reasoning options
2026-06-08 20:16:45 -05:00
Aiden Cline
30406be8f4
feat(ollama-cloud): add reasoning options
2026-06-08 20:16:42 -05:00
Aiden Cline
e4768a2d76
feat(baseten): add reasoning options
2026-06-08 20:16:41 -05:00
Aiden Cline
c347e8b438
feat(novita-ai): add reasoning options
2026-06-08 20:16:40 -05:00
Aiden Cline
5bb6b2aaf8
Merge pull request #2076 from anomalyco/feat/bedrock-reasoning-options
...
feat(amazon-bedrock): add reasoning options
2026-06-08 19:36:54 -05:00
Aiden Cline
468e2ad4ca
feat(amazon-bedrock): add reasoning options
2026-06-08 19:28:00 -05:00
Aiden Cline
04ba6ca94e
feat(azure): add reasoning options
2026-06-08 17:53:46 -05:00
Aiden Cline
0d7d13bb38
Merge pull request #2074 from anomalyco/feat/openai-reasoning-options
...
feat(openai): add reasoning options
2026-06-08 17:28:22 -05:00
Aiden Cline
6cc99b6a97
feat(openai): add reasoning options
2026-06-08 17:17:21 -05:00
Aiden Cline
fe7927f2dd
Merge pull request #2061 from Astro-Han/add-qwen3.7-plus-coding-plan-cn
...
feat: add qwen3.7-plus to alibaba-coding-plan-cn provider
2026-06-08 16:43:46 -05:00
Aiden Cline
cef703a91b
Merge pull request #2072 from anomalyco/automation/sync-models-openrouter
...
chore(sync): update OpenRouter model catalog
2026-06-08 16:43:09 -05:00
Aiden Cline
3cdf2181f5
Merge pull request #2073 from anomalyco/fix/vercel-sync-all-model-types
...
fix(vercel): sync all gateway model types
2026-06-08 16:42:45 -05:00
Aiden Cline
b0e1ed9338
fix(vercel): sync all gateway model types
2026-06-08 16:32:25 -05:00
github-actions[bot]
d061339e5d
chore(sync): update OpenRouter model catalog
2026-06-08 21:05:13 +00:00
Aiden Cline
26362a4ce3
Merge pull request #2068 from anomalyco/automation/sync-models-openrouter
...
chore(sync): update OpenRouter model catalog
2026-06-08 15:55:27 -05:00
Aiden Cline
2b0e0b7c95
Merge pull request #2047 from leszek3737/zenmux-7.06
...
feat(zenmux): add 11 new model definitions
2026-06-08 15:53:19 -05:00
github-actions[bot]
6a1ef22dcd
chore(sync): update OpenRouter model catalog
2026-06-08 19:58:10 +00:00
Leszek
9a576170d4
fix(zenmux): correct Claude Opus 4.8 base model ID
2026-06-08 21:40:32 +02:00
CodeAnimal
b821602bbd
Add DeepSeek-V4-Flash to Azure provider
2026-06-08 17:21:48 +01:00
CodeAnimal
74bf471580
Add DeepSeek-V4-Pro to Azure provider
2026-06-08 17:21:36 +01:00
knowhy
e043cc6da1
refactor(llmtr): use base_model for qwen3-6-35b
2026-06-08 18:19:40 +03:00
knowhy
6c523206aa
feat(llmtr): add logo
2026-06-08 18:19:39 +03:00
Aiden Cline
bccfdc7b87
Merge pull request #2050 from hgraca/nvidia
...
Add NVIDIA/NVIDIA Nemotron 3 Ultra
2026-06-08 09:55:16 -05:00
Aiden Cline
7217d11f79
fix(nvidia): use base model for nemotron ultra
2026-06-08 09:41:07 -05:00
Aiden Cline
ca035f8d4c
Merge pull request #2060 from nathannli/dev
...
fix(cerebras): deprecate llama3.1-8b
2026-06-08 09:21:54 -05:00
Aiden Cline
047af97a74
Merge pull request #2063 from anomalyco/automation/sync-models-openrouter
...
chore(sync): update OpenRouter model catalog
2026-06-08 09:18:59 -05:00
Aiden Cline
3b05a8e199
Merge pull request #2066 from dpuyosa/feat/venice-nemotron
...
Venice: Update gemma pricing and add nemotron model
2026-06-08 09:18:46 -05:00
github-actions[bot]
67880dc681
chore(sync): update OpenRouter model catalog
2026-06-08 13:46:24 +00:00
dpuyosa
dcf7f6d0a3
[venice] Update gemma pricing and add nemotron model
...
- Update google-gemma-4-31b-it cost/last_updated
- Add nvidia-nemotron-3-ultra-550b-a55b.toml with pricing/limits
2026-06-08 13:30:36 +02:00
Jack
209ce771e9
update minimax-m3 price in go
2026-06-08 19:11:47 +08:00
knowhy
5ccbf972a5
feat(llmtr): add models/sincap.toml
2026-06-08 08:30:28 +03:00
knowhy
c178001cca
feat(llmtr): add models/magibu-11b-v8.toml
2026-06-08 08:30:27 +03:00
knowhy
69a833ef58
feat(llmtr): add models/trendyol-7b.toml
2026-06-08 08:30:26 +03:00
knowhy
33c79f65b1
feat(llmtr): add models/medgemma-4b.toml
2026-06-08 08:30:25 +03:00
knowhy
c4fba0747f
feat(llmtr): add models/qwen3-6-35b.toml
2026-06-08 08:30:24 +03:00
knowhy
daf5f684b8
feat(llmtr): add models/gemma-4.toml
2026-06-08 08:30:23 +03:00
knowhy
fa17e02dc2
feat(llmtr): add provider.toml
2026-06-08 08:30:22 +03:00
Yuhan Lei
4dbd6b6b13
fix: use base_model format instead of full definition
2026-06-08 10:56:26 +08:00
Yuhan Lei
ff9199d68c
feat: add qwen3.7-plus to alibaba-coding-plan-cn provider
...
Qwen3.7 Plus is now available on Alibaba Cloud Coding Plan (China).
2026-06-08 10:52:50 +08:00
Nathan Li
e46f128b3a
fix(cerebras): deprecate llama3.1-8b
2026-06-07 21:28:51 -04:00
Aiden Cline
b5a8387a39
Merge pull request #2053 from smakosh/add-llmgateway-minimax-m3-qwen37-plus
...
feat(llmgateway): add MiniMax M3 and Qwen3.7 Plus
2026-06-07 20:12:32 -05:00
Aiden Cline
f55137b165
Merge pull request #2056 from anomalyco/automation/sync-models-openrouter
...
chore(sync): update OpenRouter model catalog
2026-06-07 20:12:15 -05:00
Aiden Cline
08607a1ea0
Merge pull request #2057 from anomalyco/automation/sync-models-vercel
...
chore(sync): update Vercel AI Gateway model catalog
2026-06-07 20:12:07 -05:00
Aiden Cline
4cabf9f282
Merge pull request #2055 from anomalyco/automation/sync-models-cloudflare-workers-ai
...
chore(sync): update Cloudflare Workers AI model catalog
2026-06-07 20:11:55 -05:00
github-actions[bot]
d603abe1d9
chore(sync): update Cloudflare Workers AI model catalog
2026-06-07 23:38:41 +00:00
github-actions[bot]
aedf097709
chore(sync): update OpenRouter model catalog
2026-06-07 23:38:39 +00:00
github-actions[bot]
dafb6a02a2
chore(sync): update Vercel AI Gateway model catalog
2026-06-07 23:38:37 +00:00
Levente Polyak
15f015fd4a
add Qwen3.7 Plus model configuration to Alibana coding plan
...
Coding-plan models are at a fixed monthly fee.
Link: https://modelstudio.console.alibabacloud.com/eu-central-1?tab=doc#/doc/?type=model&url=3005961
2026-06-07 21:54:13 +02:00
Leszek
c02cf9ee89
refactor(zenmux): centralize model definitions and simplify provider configs
...
This refactors Zenmux model configurations by:
- Moving comprehensive model properties (e.g., limits, modalities) from `providers/zenmux/models/` to the shared `models/` directory.
- Introducing `base_model` references in `providers/zenmux/models/` files, which now primarily specify provider-specific attributes like `cost`.
- Updating parameters for `qwen3.7-plus`, `gpt-5.5-instant`, and `step-3.7-flash` during this reorganization.
2026-06-07 20:22:01 +02:00
Aiden Cline
f112360043
Merge pull request #2052 from anomalyco/fix/google-sync-preserve-base-models
...
fix(google): compact sync and add Vertex TTS
2026-06-07 12:49:43 -05:00
Aiden Cline
e6aca83545
feat(google-vertex): add Gemini 2.5 TTS models
2026-06-07 12:38:21 -05:00
smakosh
257a4fd79c
fix(llmgateway): use base_model inheritance for MiniMax M3 and Qwen3.7 Plus
...
Upstream renamed the inheritance keyword from [extends].from to
base_model. Switch both new llmgateway entries to base_model so they
inherit the shared canonical metadata and only override llmgateway cost.
Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com >
2026-06-07 18:29:09 +01:00
Aiden Cline
7feb291963
fix(google): remove unsupported TTS model ID
2026-06-07 12:28:33 -05:00
smakosh
3e8bec11b5
Merge remote-tracking branch 'upstream/dev' into add-llmgateway-minimax-m3-qwen37-plus
...
# Conflicts:
# providers/alibaba/models/qwen3.7-plus.toml
# providers/minimax/models/MiniMax-M3.toml
2026-06-07 18:21:42 +01:00
smakosh
d322c49fc5
feat(llmgateway): add MiniMax M3 and Qwen3.7 Plus
...
Add two new text models from the LLM Gateway catalog
(https://api.llmgateway.io/v1/models ), each as an llmgateway entry
extending a canonical provider model:
- minimax-m3 -> minimax/MiniMax-M3
- qwen3.7-plus -> alibaba/qwen3.7-plus
Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com >
2026-06-07 18:20:10 +01:00
Aiden Cline
a067d458a9
Merge pull request #2030 from Nichokas/add-freemodel-provider
...
feat(freemodel): add FreeModel.dev provider (Anthropic + OpenAI formats)
2026-06-07 12:17:29 -05:00
Aiden Cline
9c0aaa8482
chore(google): sync image context limit
2026-06-07 12:17:23 -05:00
Aiden Cline
317334dc33
fix(google): preserve synced base models
2026-06-07 12:16:58 -05:00
Aiden Cline
134906e37c
Merge pull request #2051 from anomalyco/refactor/vercel-shared-sync
...
chore(sync): update Vercel model catalog
2026-06-07 12:06:18 -05:00
Aiden Cline
f473f985d9
Merge pull request #2049 from anomalyco/automation/sync-models-openrouter
...
chore(sync): update OpenRouter model catalog
2026-06-07 12:05:21 -05:00
Frank
b8f06211b0
update zen models
2026-06-07 12:48:22 -04:00
github-actions[bot]
8b742acb8b
chore(sync): update OpenRouter model catalog
2026-06-07 16:46:19 +00:00
Aiden Cline
33be44552c
Merge pull request #2042 from anomalyco/chore/sync-vercel-catalog
...
chore(sync): update Vercel model catalog
2026-06-07 11:45:03 -05:00
Herberto Graca
0645480326
Add NVIDIA/NVIDIA Nemotron 3 Ultra
2026-06-07 16:47:31 +02:00
Leszek
8d2f754dd6
feat(zenmux): add 11 new model definitions
...
New models: claude-opus-4.8, gemini-3.1-flash-lite, gemini-3.5-flash, ring-2.6-1t, minimax-m3, gpt-5.5-instant, qwen3.7-max, qwen3.7-plus, step-3.7-flash, grok-4.3, grok-build-0.1
2026-06-07 13:56:12 +02:00
Nichokas
72bba2a4a3
fix: Update logo to comply with the guidelines
2026-06-07 10:36:11 +02:00
Aiden Cline
3cfa5e6583
Merge pull request #2041 from anomalyco/refactor/vercel-shared-sync
...
refactor(sync): migrate Vercel to shared runner
2026-06-07 00:22:31 -05:00
Aiden Cline
4a01190179
Merge pull request #2043 from anomalyco/automation/sync-models-cloudflare-workers-ai
...
chore(sync): update Cloudflare Workers AI model catalog
2026-06-07 00:04:35 -05:00
Aiden Cline
a00c4b4feb
Merge pull request #2044 from anomalyco/automation/sync-models-openrouter
...
chore(sync): update OpenRouter model catalog
2026-06-07 00:04:04 -05:00
github-actions[bot]
b32b126540
chore(sync): update Cloudflare Workers AI model catalog
2026-06-07 03:26:37 +00:00
github-actions[bot]
4f296dcf55
chore(sync): update OpenRouter model catalog
2026-06-07 03:26:36 +00:00
Aiden Cline
285ca1a865
chore(sync): update Vercel model catalog
2026-06-06 19:49:46 -05:00
Aiden Cline
55b630c1f6
fix(vercel): inherit model update dates
2026-06-06 19:49:09 -05:00
Aiden Cline
20756bbba8
fix(vercel): resolve Alibaba metadata links
2026-06-06 18:56:46 -05:00
Aiden Cline
9260d4d112
fix(sync): match canonical metadata casing
2026-06-06 18:56:10 -05:00
Aiden Cline
022e6350c7
fix(vercel): factor canonical model metadata
2026-06-06 18:55:48 -05:00
Aiden Cline
2fc28466e8
fix(sync): report retained Vercel models
2026-06-06 18:51:55 -05:00
Aiden Cline
a700d92235
refactor(sync): migrate Vercel to shared runner
2026-06-06 18:49:36 -05:00
Aiden Cline
95cfda05be
Merge pull request #2028 from anomalyco/automation/sync-models-openrouter
...
chore(sync): update OpenRouter model catalog
2026-06-06 18:44:57 -05:00
Aiden Cline
0f1783dad9
Merge pull request #2040 from anomalyco/fix/sync-retain-curated-base-models
...
fix(sync): retain curated base model links
2026-06-06 18:42:53 -05:00
Aiden Cline
7dfc7867f1
fix(sync): retain curated base model links
2026-06-06 18:41:11 -05:00
Aiden Cline
62ac8db1b8
Merge pull request #2039 from anomalyco/automation/sync-models-cloudflare-workers-ai
...
chore(sync): update Cloudflare Workers AI model catalog
2026-06-06 18:40:34 -05:00
github-actions[bot]
877da3f734
chore(sync): update OpenRouter model catalog
2026-06-06 23:38:33 +00:00
github-actions[bot]
62e9bdadb6
chore(sync): update Cloudflare Workers AI model catalog
2026-06-06 23:38:33 +00:00
Nichokas
14709d66d7
feat(freemodel): add provider logo
...
Addresses review feedback on #2030 — provider was missing a logo.svg.
2026-06-07 01:36:32 +02:00
Aiden Cline
1dd2250d1e
Merge pull request #2038 from anomalyco/fix/cloudflare-sync-base-models
...
fix(sync): preserve Cloudflare base models
2026-06-06 18:20:59 -05:00
Aiden Cline
5bb15e680d
fix(sync): preserve Cloudflare base models
2026-06-06 18:15:48 -05:00
Aiden Cline
caa3521f13
update logos
2026-06-06 18:00:19 -05:00
Aiden Cline
d996411611
Merge pull request #2037 from licat2023/fix/deepseek-cache-pricing
...
fix(models.dev): correct deepseek-chat/reasoner cache_read pricing
2026-06-06 17:51:44 -05:00
licat2023
68b75c7e5f
fix(models.dev): correct deepseek-reasoner cache_read pricing (0.028 to 0.0028)
...
Matches DeepSeek V4 Flash pricing as deepseek-reasoner is a deprecated alias.
Ref: https://api-docs.deepseek.com/quick_start/pricing
2026-06-07 05:01:20 +08:00
licat2023
b780073b8d
fix(models.dev): correct deepseek-chat cache_read pricing (0.028 to 0.0028)
...
Matches DeepSeek V4 Flash pricing as deepseek-chat is a deprecated alias.
Ref: https://api-docs.deepseek.com/quick_start/pricing
2026-06-07 05:01:18 +08:00
Nichokas
88abd70b24
refactor(freemodel): merge into one provider with per-model hosts
...
opencode exposes freemodel as a single provider behind one login. Move the
four GPT models out of the separate `freemodel-codex` provider and into
`freemodel`, giving each a per-model `[provider]` override
(`@ai-sdk/openai-compatible`, https://api.freemodel.dev/v1 ) so the Claude
models keep the provider default (`@ai-sdk/anthropic`, cc.freemodel.dev)
and the GPT models route to the OpenAI host. Removes `freemodel-codex`.
Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com >
2026-06-06 15:32:29 +02:00
Adam
e524b01b9a
fix(models): add nvidia nemotron bases
2026-06-06 06:38:27 -05:00
Adam
f753ddacdb
fix(web): search provider models
2026-06-06 05:04:20 -05:00
Nichokas
84ed973170
feat(freemodel): add FreeModel.dev provider (Anthropic + OpenAI formats)
...
Adds two provider entries for freemodel.dev, a gateway exposing two model
sets depending on the API format:
- freemodel: Anthropic-format endpoint (cc.freemodel.dev) serving Claude
models, via @ai-sdk/anthropic
- freemodel-codex: OpenAI-compatible endpoint (api.freemodel.dev) serving
GPT/Codex models, via @ai-sdk/openai-compatible
Models inherit metadata via base_model and are priced at the providers'
standard rates; freemodel additionally charges cache_write at the input
rate for the OpenAI models.
2026-06-05 20:55:11 +02:00
Aiden Cline
a6538b3644
Merge pull request #2012 from mvanhorn/fix/1848-evroc-qwen3-embedding-output-limit
...
fix: correct evroc Qwen3-Embedding-8B limit.output to 4096
2026-06-05 11:07:54 -05:00
Adam
d70426e098
feat(web): improve search palette
2026-06-05 10:49:32 -05:00
Aiden Cline
a4895edaac
Merge pull request #2011 from Suat-B/dev
...
Add GPT 5.5 model to Xpersona provider
2026-06-05 09:34:10 -05:00
Adam
724ee7c132
feat(web): add search palette
2026-06-05 09:31:32 -05:00
Aiden Cline
b9684cc0e8
Merge pull request #2027 from shzdehmd/dev
...
feat(fireworks-ai): add kimi-k2p6-fast model
2026-06-05 09:02:26 -05:00
Aiden Cline
c22a18f293
Merge pull request #2017 from anomalyco/automation/sync-models-openrouter
...
chore(sync): update OpenRouter model catalog
2026-06-05 08:55:58 -05:00
Ahmad Shahzad
173d5c6add
feat(fireworks-ai): add kimi-k2p6-fast model
...
Fireworks AI is standardizing their naming convention to just "Fast"
for Fast/Turbo offerings (e.g., GLM 5.1 Fast). Adding kimi-k2p6-fast
to match this pattern while retaining kimi-k2p6-turbo for backward
compatibility with existing user workflows.
2026-06-05 18:52:13 +05:00
Aiden Cline
d50aa210b7
fix: copilot modes
2026-06-05 08:34:41 -05:00
github-actions[bot]
7447979e06
chore(sync): update OpenRouter model catalog
2026-06-05 13:23:27 +00:00
Jack
189bd4139b
Merge pull request #2023 from anomalyco/fix/opencode-go-qwen-plus-pricing-limits
...
feat(opencode-go): update Qwen Plus pricing and limits
2026-06-05 18:51:59 +08:00
Jack
350704d762
feat(opencode-go): update Qwen Plus pricing and limits
2026-06-05 18:32:20 +08:00
Aiden Cline
75c0d2cad3
Merge pull request #2020 from anomalyco/fix/google-missing-gemini-models
...
fix(google): add missing Gemini models
2026-06-05 00:48:27 -05:00
Aiden Cline
c525d2a0fe
fix(google): add missing Gemini models
2026-06-05 00:46:54 -05:00
Aiden Cline
0dc4a5f947
Merge pull request #2019 from anomalyco/fix/openrouter-inherit-input-limits
...
fix(openrouter): inherit input limits when context matches
2026-06-05 00:40:16 -05:00
Aiden Cline
ad3e1e7d02
fix(openrouter): inherit input limits when context matches
2026-06-05 00:30:17 -05:00
Aiden Cline
efc8827451
fix(bedrock): remove legacy model extends
2026-06-04 22:19:06 -05:00
Aiden Cline
947990c838
Merge pull request #1978 from anomalyco/feat/amazon-bedrock-openai-mantle-models
...
feat(amazon-bedrock): add OpenAI Mantle models
2026-06-04 22:15:31 -05:00
Aiden Cline
90a383e04e
Merge pull request #2013 from Sawyerb/dev
...
Removed mercury-coder-small
2026-06-04 21:29:10 -05:00
Adam
9342481774
feat(web): redesign model-centric navigation ( #2014 )
2026-06-04 19:17:05 -05:00
Adam
c7e827b320
fix(sync): use zhipuai metadata ( #2010 )
2026-06-04 18:30:39 -05:00
Sawyer
74dc5f8b46
removed mercury-coder-small
2026-06-04 16:19:20 -07:00
Matt Van Horn
dd74f0c51c
fix: correct evroc Qwen3-Embedding-8B output limit to 4096
2026-06-04 16:05:07 -07:00
Suat-B
a7ab4866a2
Add trailing newline to Xpersona GPT-5.5 model
2026-06-04 17:31:20 -05:00
Suat-B
3f0137ae71
Use base model metadata for Xpersona GPT-5.5
2026-06-04 17:29:54 -05:00
Suat-B
b3ba5f7727
Add GPT 5.5 model listed as official Xpersona offering
2026-06-04 17:17:55 -05:00
Aiden Cline
4bbb2c486f
Merge pull request #1986 from anomalyco/feat/mistral-reasoning-options
...
feat(mistral): add reasoning effort options
2026-06-04 14:24:57 -05:00
Aiden Cline
5fe04aee59
fix(mistral): correct medium latest alias
2026-06-04 14:13:20 -05:00
Aiden Cline
3e5716e147
fix(mistral): add verified reasoning effort models
2026-06-04 13:48:44 -05:00
Aiden Cline
c4aac4aea7
Merge pull request #2006 from anomalyco/automation/sync-models-openrouter
...
chore(sync): update OpenRouter model catalog
2026-06-04 13:08:50 -05:00
Adam
69d1cb7773
feat(web): remove benchmarks column
2026-06-04 13:03:45 -05:00
Adam
32d509dd3b
Normalize base-model inheritance ( #2007 )
2026-06-04 12:53:42 -05:00
github-actions[bot]
6c89d5fef7
chore(sync): update OpenRouter model catalog
2026-06-04 17:20:46 +00:00
Aiden Cline
8e6d393c01
Merge pull request #1984 from anomalyco/feat/sarvam-reasoning-options
...
feat(sarvam): add reasoning options
2026-06-04 11:59:21 -05:00
Aiden Cline
ee8104f0a1
Merge pull request #2005 from zainhas/dev
...
[Together AI] add nemotron 3 ultra
2026-06-04 11:51:16 -05:00
Aiden Cline
10155a62eb
Merge pull request #2004 from anomalyco/fix/sync-reasoning-options
...
fix(sync): preserve reasoning options
2026-06-04 11:51:03 -05:00
Zain Hasan
afa81334ce
[Together AI] add nemotron 3 ultra
2026-06-04 09:49:59 -07:00
Aiden Cline
cc956258d6
fix(sync): preserve reasoning options
2026-06-04 11:49:16 -05:00
Aiden Cline
dfcf5ba1cf
Merge pull request #1999 from houtanb/fix-together-deepseek-models
...
fix(together): DeepSeek-R1, DeepSeek-V3 name & release date
2026-06-04 11:40:33 -05:00
Aiden Cline
637f7e3eec
Merge pull request #1997 from dpuyosa/update-models
...
Venice: Add Qwen 3.7 Plus and update models
2026-06-04 11:40:20 -05:00
Aiden Cline
a4b39711da
Merge pull request #1996 from anomalyco/automation/sync-models-openrouter
...
chore(sync): update OpenRouter model catalog
2026-06-04 11:39:34 -05:00
Aiden Cline
8f5d9e7b82
Merge pull request #1998 from chenkuilinckl-ai/fix/qwen3.7-plus-vision-1m-context
...
fix(alibaba): qwen3.7-plus GA adds vision (image+video) and 1M context
2026-06-04 11:39:04 -05:00
Aiden Cline
7beeab6cef
Merge pull request #2003 from BlockListed/cortecs-gpt-5-4
...
add gpt 5.4 to cortecs
2026-06-04 11:37:46 -05:00
BlockListed
0a03c18424
add gpt 5.4 to cortecs
2026-06-04 18:13:57 +02:00
Adam
c8a52d19c1
feat(models): add coding benchmarks and weights ( #2000 )
...
* feat(models): add coding benchmarks and weights
* feat(models): add more coding benchmarks
* feat(models): add agent benchmark scores
* feat(models): normalize benchmark metadata
2026-06-04 11:09:53 -05:00
Adam
859bb31ffa
refactor(chutes): use base_model for tee wrappers ( #2002 )
2026-06-04 11:05:45 -05:00
Adam
a672fe4b08
feat(web): surface model metadata links ( #2001 )
2026-06-04 11:03:53 -05:00
Frank
72fc81a4da
update zen models
2026-06-04 11:31:06 -04:00
github-actions[bot]
2f77d0afac
chore(sync): update OpenRouter model catalog
2026-06-04 15:25:27 +00:00
Houtan Bastani
03333d9aa4
fix(together): DeepSeek-R1, DeepSeek-V3 name & release date
...
See timestamps:
* "Release DeepSeek-R1": https://github.com/deepseek-ai/DeepSeek-R1/commit/23807ced51627276434655dd9f27725354818974
* "Release DeepSeek-V3": https://github.com/deepseek-ai/DeepSeek-V3/commit/4c2fdb8f55e049553b9f4f1a3241f86d739c8cf8
2026-06-04 13:16:14 +02:00
chenkuilinckl-ai
977e2f0155
fix(alibaba): qwen3.7-plus GA adds vision (image+video) and 1M context
2026-06-04 17:29:42 +08:00
dpuyosa
963e9a6bab
[venice] Add Qwen 3.7 Plus and update models
...
- Added qwen-3-7-plus
- Update google-gemma-4-31b-it cache_read
- Update minimax-m3 output limit, modalities
2026-06-04 10:01:57 +02:00
Aiden Cline
9e8cad9ac9
Merge pull request #1988 from anomalyco/feat/stepfun-reasoning-options
...
feat(stepfun): add reasoning effort options
2026-06-04 00:38:32 -05:00
Aiden Cline
f6c1036a86
Merge pull request #1985 from anomalyco/feat/xiaomi-reasoning-options
...
feat(xiaomi): add reasoning toggles
2026-06-04 00:37:58 -05:00
Aiden Cline
1926832a9d
fix(sarvam): expose null reasoning effort
2026-06-04 00:17:23 -05:00
Aiden Cline
cb22ad5625
Merge pull request #1992 from anomalyco/fix/sync-base-model-output
...
fix(sync): preserve base model output
2026-06-04 00:02:46 -05:00
Aiden Cline
909db75087
fix(sync): preserve base model output
2026-06-04 00:00:46 -05:00
Aiden Cline
b551552f14
Merge pull request #1983 from anomalyco/feat/cohere-reasoning-options
...
feat(cohere): add reasoning options
2026-06-03 23:34:36 -05:00
Aiden Cline
0919062b40
Merge pull request #1989 from anomalyco/feat/google-gemini-reasoning-options
...
feat(google): add Gemini reasoning options
2026-06-03 23:05:44 -05:00
Aiden Cline
8662c63313
fix(google): retain deprecated Gemini reasoning metadata
2026-06-03 23:03:08 -05:00
Aiden Cline
36b808691a
Merge pull request #1991 from shzdehmd/dev
...
chore(fireworks): update qwen3p6-plus limits to 262K context / 65K output
2026-06-03 22:21:52 -05:00
Ahmad Shahzad
c8feebb7cd
chore(fireworks): update qwen3p6-plus limits to 262K context / 65K output
2026-06-04 07:05:29 +05:00
Aiden Cline
259801fba5
fix(google): exclude unavailable Gemini 3 Pro preview
2026-06-03 18:28:25 -05:00
Aiden Cline
ce35e18561
feat(stepfun): add reasoning effort options
2026-06-03 17:49:21 -05:00
Aiden Cline
133a0b0126
feat(google): add Gemini reasoning options
2026-06-03 17:49:09 -05:00
Aiden Cline
19d26ba611
feat(mistral): add reasoning effort options
2026-06-03 17:48:39 -05:00
Aiden Cline
99406ae7df
feat(xiaomi): add reasoning toggles
2026-06-03 17:48:22 -05:00
Aiden Cline
4f25170be2
feat(cohere): add reasoning options
2026-06-03 17:48:05 -05:00
Aiden Cline
6ddf935238
feat(sarvam): add reasoning options
2026-06-03 17:48:00 -05:00
Aiden Cline
6ae56b00a8
Merge pull request #1981 from anomalyco/feat/glm-coding-plan-reasoning-toggle
...
feat(glm): add Zhipu and coding plan reasoning toggles
2026-06-03 17:31:39 -05:00
Aiden Cline
2cb0d28e17
feat(zhipuai): add reasoning toggles
2026-06-03 17:27:57 -05:00
Aiden Cline
03d90aeacc
Merge pull request #1982 from anomalyco/fix/sync-model-catalog-matrix
...
fix(sync): restore model catalog workflow
2026-06-03 17:26:22 -05:00
Aiden Cline
eb7dbead75
fix(sync): restore model catalog workflow
2026-06-03 17:20:01 -05:00
Aiden Cline
36cbbfc577
feat(glm): add coding plan reasoning toggles
2026-06-03 16:56:42 -05:00
Aiden Cline
d84e883ede
Merge pull request #1980 from anomalyco/feat/deepseek-reasoning-options
...
feat(deepseek): add reasoning options
2026-06-03 16:09:22 -05:00
Aiden Cline
dd9d6ff54e
feat(deepseek): add reasoning options
2026-06-03 16:08:09 -05:00
Aiden Cline
d7e19c7627
Merge pull request #1979 from anomalyco/feat/moonshot-reasoning-toggle
...
feat(moonshot): add reasoning toggle options
2026-06-03 15:46:39 -05:00
Aiden Cline
ca3eb39fbd
feat(moonshot): add reasoning toggle options
2026-06-03 15:45:15 -05:00
Aiden Cline
d136e7b036
Merge pull request #1955 from eliasaronson/chore/mark-deprecated-models
...
chore: mark retired models as deprecated
2026-06-03 15:25:27 -05:00
Adam
f6c6f04367
feat(models): add model metadata ( #1974 )
...
* feat(models): add model metadata
* feat(models): rename model metadata namespaces
2026-06-03 15:13:54 -05:00
Aiden Cline
1e9e4bbdab
Merge pull request #1977 from Ardakilic/feat/nano-gpt-20260603
...
chore: sync nano-gpt models: 20260603
2026-06-03 15:02:42 -05:00
Aiden Cline
fbaf50d323
feat(amazon-bedrock): add OpenAI Mantle models
2026-06-03 14:58:05 -05:00
Arda Kılıçdağı
efb553d287
chore: sync nano-gpt models: 20260603
2026-06-03 21:31:28 +03:00
Jack
d9b83ca9cd
Merge pull request #1975 from anomalyco/update/opencode-go-qwen3.7-plus
...
feat(opencode-go): add Qwen3.7 Plus model
2026-06-04 01:30:11 +08:00
Aiden Cline
25592e361a
Merge pull request #1976 from jerome-benoit/feat/sap-ai-core-gpt-5.5
...
feat(sap-ai-core): add GPT-5.5
2026-06-03 12:28:07 -05:00
Jérôme Benoit
012928800c
fix(sap-ai-core): align GPT-5.4 release_date with upstream
...
SAP AI Core routes to OpenAI gpt-5.4; release_date should reflect
the actual model release (2026-03-05) rather than the SAP catalog
availability date (2026-04-27).
2026-06-03 19:09:54 +02:00
Jérôme Benoit
ee60c2e4d4
feat(sap-ai-core): add GPT-5.5
...
SAP AI Core routes to OpenAI gpt-5.5; specs mirror the canonical
provider/openai/gpt-5.5 with the established sap-ai-core wrapper
adjustments (lowercase name, drop [[cost.tiers]], drop
[experimental.modes.fast]).
2026-06-03 19:06:07 +02:00
Jack
7e76cde0b1
fix(opencode-go): correct Qwen3.7 Plus dates
2026-06-04 00:59:42 +08:00
Jack
bbd2479e12
feat(opencode-go): add Qwen3.7 Plus model
2026-06-04 00:57:42 +08:00
Aiden Cline
eeb17ccab4
Merge pull request #1971 from coder-wangbin/feat/alibaba-cn-qwen3.7-plus
...
feat: add Qwen3.7 Plus model for alibaba-cn provider
2026-06-03 10:00:39 -05:00
Aiden Cline
46ffeee012
Merge pull request #1972 from Phosmachina/feature/update-deepinfra-deepseek-and-mimo
...
Feature/update deepinfra deepseek and mimo
2026-06-03 09:56:37 -05:00
Aiden Cline
8c1e45007d
Merge pull request #1967 from anomalyco/automation/sync-models-openrouter
...
chore(sync): update OpenRouter model catalog
2026-06-03 09:54:14 -05:00
Frank
364628cb93
update zen models
2026-06-03 10:32:41 -04:00
github-actions[bot]
50f5d40b4e
chore(sync): update OpenRouter model catalog
2026-06-03 14:02:08 +00:00
Michel Paronnaud
707be517c6
chore(deepinfra): adjust prices and limit for DeepSeek models
2026-06-03 12:40:49 +02:00
Michel Paronnaud
657598e609
fix(deepinfra): path for mimo models
2026-06-03 12:28:08 +02:00
wangbin
0999a475ba
feat: add Qwen3.7 Plus model for alibaba-cn provider
...
- Add base definition in providers/alibaba/models/qwen3.7-plus.toml
- Add extends reference in providers/alibaba-cn/models/qwen3.7-plus.toml
- Release date: 2026-06-02
- Context: 131K tokens, Output: 16K tokens
- Pricing: $0.50/$3.00 per 1M tokens (input/output)
- Supports reasoning and tool calling
2026-06-03 17:40:12 +08:00
Frank
e1ea0254ed
update zen models
2026-06-02 22:51:19 -04:00
Aiden Cline
7fad18b054
fix
2026-06-02 16:36:29 -05:00
Aiden Cline
710bbc5375
Merge pull request #1961 from anomalyco/feat/minimax-m3-reasoning-toggle
...
feat(minimax): add M3 reasoning toggle
2026-06-02 16:34:19 -05:00
Aiden Cline
0b9fa62153
Merge pull request #1963 from nicholasgriffintn/open-mistral-nemo
...
fix: update mistral nemo
2026-06-02 14:34:54 -05:00
Aiden Cline
126a481a70
Merge pull request #1965 from nicholasgriffintn/update-devstral-models
...
chore: update devstral models
2026-06-02 14:34:38 -05:00
Aiden Cline
b080cf1fe4
Merge pull request #1952 from anomalyco/automation/sync-models-xai
...
chore(sync): update xAI model catalog
2026-06-02 14:03:02 -05:00
Nicholas Griffin
042002d3a3
chore: update devstral models
2026-06-02 19:47:17 +01:00
Nicholas Griffin
a192b51e84
chore: undo
2026-06-02 19:46:13 +01:00
Nicholas Griffin
b870c796ef
chore: undo
2026-06-02 19:45:25 +01:00
Nicholas Griffin
35e4981aa9
chore: update
2026-06-02 19:40:02 +01:00
Frank
970b660070
sync
2026-06-02 14:24:55 -04:00
Nicholas Griffin
bb03a8c207
fix: update mistral nemo
2026-06-02 19:22:36 +01:00
github-actions[bot]
b6dbb68c32
chore(sync): update xAI model catalog
2026-06-02 18:21:19 +00:00
Aiden Cline
b34a7c21f3
feat(minimax): add M3 reasoning toggle
2026-06-02 12:43:17 -05:00
Aiden Cline
4b169b8cb6
Merge pull request #1960 from anomalyco/feat/zai-reasoning-toggle
...
feat(zai): add reasoning toggle options
2026-06-02 12:38:15 -05:00
Aiden Cline
f8ecf4daec
feat(zai): add reasoning toggle options
2026-06-02 12:02:54 -05:00
Frank
eb68be3fbc
stats
2026-06-02 12:38:49 -04:00
Aiden Cline
55fb058cc5
tweak: update available options
2026-06-02 11:35:26 -05:00
Aiden Cline
5545a98866
tweak: handle missing omits gracefully
2026-06-02 11:25:55 -05:00
Aiden Cline
1a7b103bba
fix: omit
2026-06-02 11:22:52 -05:00
Aiden Cline
f341858398
Merge pull request #1896 from stevenyeung/add/alibaba-token-plan
...
feat: add Alibaba Token Plan provider with 15 models
2026-06-02 11:20:42 -05:00
Aiden Cline
f2d1e42582
Merge pull request #1951 from anomalyco/automation/sync-models-openrouter
...
chore(sync): update OpenRouter model catalog
2026-06-02 11:03:27 -05:00
Aiden Cline
2f9d2d8d78
Merge pull request #1957 from kapelame/fix/minimax-m3-prune
...
fix(minimax): correct MiniMax-M3 pricing and max output
2026-06-02 10:53:58 -05:00
Aiden Cline
bf6b0e32d7
Merge pull request #1959 from Tavernari/feat/claudinio-add-audio-video
...
feat(claudinio): add audio and video input modalities
2026-06-02 10:53:25 -05:00
github-actions[bot]
5b4d2ea2e4
chore(sync): update OpenRouter model catalog
2026-06-02 15:41:53 +00:00
Victor Carvalho Tavernari
ee4badab7b
fix(claudinio): update cache_read price to 0.150 per 1M tokens
2026-06-02 16:34:05 +01:00
Victor Carvalho Tavernari
6c05a6b519
Merge branch 'dev' into feat/claudinio-add-audio-video
2026-06-02 16:32:12 +01:00
Victor Carvalho Tavernari
91ae9a48e3
feat(claudinio): add audio and video input modalities
2026-06-02 16:24:06 +01:00
kapelame
85d02711bd
fix(minimax): correct MiniMax-M3 pricing and max output
...
The M3 entries added in #1940 copied M2.7's cost values. Correct them to
the official M3 pricing and limits:
- minimax / minimax-cn (pay-as-you-go): input 0.30 -> 0.60,
output 1.20 -> 2.40, cache_read 0.06 -> 0.12, and remove cache_write
(M3 has no active prompt-cache-write tier).
- max output 131072 -> 128000 across all four providers.
- coding-plan variants keep their subscription-plan zero pricing; only
max output is corrected.
Context (512K), modalities, and the other flags are unchanged.
2026-06-02 21:03:45 +08:00
Elias H Aronsson
1ba404612d
chore: mark retired models as deprecated
...
Add status = "deprecated" to models that are past their provider's
shutdown/retirement date (no longer served by the public API).
Google Gemini (4):
gemini-2.0-flash, gemini-2.0-flash-lite, gemini-3-pro-preview,
gemini-3.1-flash-lite-preview
Anthropic Claude (7):
claude-3-sonnet-20240229, claude-3-5-sonnet-20240620,
claude-3-5-sonnet-20241022, claude-3-opus-20240229,
claude-3-7-sonnet-20250219, claude-3-5-haiku-20241022,
claude-3-haiku-20240307
OpenAI (2):
o1-preview, o1-mini
Sources:
https://ai.google.dev/gemini-api/docs/deprecations
https://platform.claude.com/docs/en/about-claude/model-deprecations
https://developers.openai.com/api/docs/deprecations
2026-06-02 10:06:18 +02:00
Aiden Cline
f91dd4ad0b
Merge pull request #1950 from anomalyco/fix/reasoning-options-inheritance
...
fix: do not inherit reasoning options
2026-06-01 23:24:13 -05:00
Aiden Cline
449b926f40
fix: do not inherit reasoning options
2026-06-01 23:22:04 -05:00
Aiden Cline
2546ffe570
Merge pull request #1940 from matstrange/add-minimax-m3
...
Add MiniMax-M3 model (#1933 )
2026-06-01 22:57:45 -05:00
Hex Agent
f33ff9ba78
Add MiniMax-M3 model to 5 providers
...
MiniMax-M3 is MiniMax's new frontier multimodal coding model: 1M context
window (512K minimum on ollama-cloud), native text/image/video input,
tool calling, reasoning, and open weights.
Adds the model to all five providers where it should be available:
- minimax (pay-as-you-go)
- minimax-cn (pay-as-you-go, China)
- minimax-coding-plan (token plan subscription)
- minimax-cn-coding-plan (token plan subscription, China)
- ollama-cloud
Closes #1933 .
Notes for reviewers:
- Cost fields on minimax/minimax-cn match M2.7; M3 docs state the
pricing is unchanged from M2.7.
- The 1M/512K context divergence on ollama-cloud is intentional —
ollama advertises 1M with a 512K minimum, so 512K is the safe floor
that won't surprise opencode users with mid-request rejections.
- output = 131072 is inherited from the existing M2.7 files; MiniMax's
published M3 docs only advertise the 1M input context, not a separate
output cap.
- The minimax-coding-plan variant has been verified end-to-end in
opencode against the MiniMax token plan API.
2026-06-01 22:54:17 -05:00
Aiden Cline
98bf80cd77
Merge pull request #1949 from anomalyco/fix/github-copilot-context-limits-complete
...
fix(github-copilot): preserve model-specific limits
2026-06-01 22:49:00 -05:00
Aiden Cline
68660cad83
Merge pull request #1938 from anyapi-ai/dev
...
Add AnyAPI provider
2026-06-01 22:47:26 -05:00
Aiden Cline
78e6c1a2dd
fix(github-copilot): preserve model-specific limits
2026-06-01 22:46:56 -05:00
Aiden Cline
062ba1fbb7
Merge pull request #1939 from anomalyco/automation/sync-models-openrouter
...
chore(sync): update OpenRouter model catalog
2026-06-01 22:46:18 -05:00
Aiden Cline
477ccffee8
Merge pull request #1948 from anomalyco/feat/anthropic-reasoning-options
...
feat(anthropic): add reasoning options
2026-06-01 22:41:23 -05:00
github-actions[bot]
6ce2f88eac
chore(sync): update OpenRouter model catalog
2026-06-02 03:26:49 +00:00
Aiden Cline
4501589a20
feat(anthropic): add reasoning options
2026-06-01 22:17:07 -05:00
Aiden Cline
0b643efd17
Merge pull request #1947 from LYY/fix/github-copilot-context-limits
...
fix(github-copilot): correct limits for 8 models with verified CAPI data (partial, refs #1946 )
2026-06-01 22:16:07 -05:00
LYY
8a0ea15aca
fix(github-copilot): correct limits for models with verified CAPI data
...
The github-copilot model stubs use [extends] and inherit the upstream
base model's context/input/output limits (~1M), which do not match the
limits GitHub Copilot actually enforces via api.githubcopilot.com/models.
Override the limit fields for the 8 models whose live CAPI values were
verified first-hand. Models that were not enabled on the test account
(no CAPI data available) are intentionally left unchanged.
Refs anomalyco/models.dev#1946
2026-06-02 11:08:51 +08:00
Aiden Cline
b3676fd77d
Merge pull request #1934 from yukoba/github-copilot-2026-06
...
June 2026 changes of GitHub Copilot
2026-06-01 16:10:45 -05:00
Aiden Cline
f97a02078c
Merge pull request #1932 from eliasto/ovhcloud/update-models-qwen
...
feat(ovhcloud): Add new Qwen models and sync mode
2026-06-01 14:27:58 -05:00
Aiden Cline
2a7153fcef
Merge pull request #1943 from peculiarnewbie/fix/crof-models-update
...
fix: update crof.ai models to match current API
2026-06-01 13:41:23 -05:00
Aiden Cline
3396f15854
Merge pull request #1942 from anomalyco/feat/reasoning-options-schema
...
feat: add reasoning options schema
2026-06-01 13:41:10 -05:00
Aiden Cline
ccfd99b5e9
Merge pull request #1941 from KTibow/fix-llmgateway-deepseek-v3-2-price
...
fix(llmgateway): update model pricing
2026-06-01 13:40:20 -05:00
bolt
ef6eb88216
fix: update crof.ai models to match current API
2026-06-02 00:52:14 +07:00
Aiden Cline
94e128244a
feat: add reasoning options schema
2026-06-01 12:48:11 -05:00
KTibow
ec6d07d41c
fix(llmgateway): use weighted pricing selection
2026-06-01 10:41:23 -07:00
KTibow
4efbe58af8
fix(llmgateway): update model pricing
2026-06-01 10:24:05 -07:00
Christina
3b55102a45
Add AnyAPI provider with 30 models
...
Adds AnyAPI (https://anyapi.ai ) as a new provider. Models reuse existing
canonical entries through `extends`. Cost fields are omitted as AnyAPI uses
a credit-based pricing system. Validated locally with `bun validate`.
Models (30):
- openai: gpt-5.4, gpt-5.2, gpt-5.1, gpt-5, gpt-5-mini, gpt-4.1, gpt-4.1-mini, o4-mini, o3, o3-mini
- anthropic: claude-opus-4-7, claude-opus-4-6, claude-sonnet-4-6, claude-sonnet-4-5, claude-haiku-4-5
- google: gemini-2.5-pro, gemini-2.5-flash, gemini-2.5-flash-lite, gemini-3-pro-preview, gemini-3-flash-preview
- deepseek: deepseek-v4-pro, deepseek-v4-flash, deepseek-chat, deepseek-r1
- mistralai: mistral-large-2512, devstral-2512
- perplexity: sonar-pro, sonar-reasoning-pro
- cohere: command-r-plus-08-2024
- xai: grok-4.3
2026-06-01 17:14:24 +02:00
Aiden Cline
a4fe8fc36b
fmt
2026-06-01 09:53:16 -05:00
Aiden Cline
b24b44ca0a
Merge pull request #1920 from anomalyco/automation/sync-models-openrouter
...
chore(sync): update OpenRouter model catalog
2026-06-01 09:39:43 -05:00
Aiden Cline
466d664bf9
Merge pull request #1930 from dpuyosa/dev
...
Venice: Add MiniMax M3 model
2026-06-01 09:39:27 -05:00
Aiden Cline
340ef131e9
Merge pull request #1936 from JoshuaDietz/dev
...
feat(ollama-cloud): add minimax m3
2026-06-01 09:38:47 -05:00
Aiden Cline
6ce69a706a
Merge pull request #1937 from KTibow/fix/hpc-ai-pricing
...
fix: update HPC-AI model pricing
2026-06-01 09:38:34 -05:00
KTibow
3505674a1b
fix: update HPC-AI model pricing
2026-06-01 07:29:43 -07:00
github-actions[bot]
cfeaef8f77
chore(sync): update OpenRouter model catalog
2026-06-01 14:25:41 +00:00
Joshua Dietz
f73a5d571b
feat(ollama-cloud): add minimax m3
2026-06-01 16:14:49 +02:00
Yu Kobayashi
62d31abff8
June 2026 changes of GitHub Copilot
2026-06-01 22:03:44 +09:00
Elias TOURNEUX
ee71ea7181
feat(ovhcloud): Add qwen 3.6 27b
2026-06-01 11:57:58 +02:00
Elias TOURNEUX
b038f40b64
feat(ovhcloud): Add OVHcloud sync mode
2026-06-01 10:28:46 +02:00
Elias TOURNEUX
4b7bdd9692
feat(ovhcloud): Add new Qwen models
2026-06-01 10:16:54 +02:00
dpuyosa
b73ef1a634
[venice] Add MiniMax M3 model
...
- Add minimax-m3.toml with cost, limits, modalities
- Enable text/image input and text output support
- Set 500k context, 32k output, cache pricing
2026-06-01 09:38:12 +02:00
Aiden Cline
c11b5bb00b
Merge pull request #1927 from Jercik/codex/update-wafer-price-cuts
...
fix: update Wafer.ai pricing
2026-06-01 00:01:03 -05:00
Łukasz Jerciński
01bc1785b4
fix: update Wafer price cuts
2026-06-01 06:25:20 +02:00
Aiden Cline
36b5847fa4
Merge pull request #1925 from anomalyco/fix/vercel-claude-opus-4-1-symlink
...
fix vercel claude opus 4.1 symlink
2026-05-31 22:12:27 -05:00
Aiden Cline
040beaf818
fix vercel claude opus 4.1 symlink
2026-05-31 22:11:05 -05:00
Frank
9f8e1a3858
update zen models
2026-05-31 22:07:52 -04:00
Jack
ce7a1c0218
update context limit of M3 in Go
2026-06-01 09:45:46 +08:00
Jack
55e6c7a2ee
add image,video modalities to M3
2026-06-01 08:58:02 +08:00
Frank
8bacf496fe
update zen models
2026-05-31 19:49:41 -04:00
Frank
ac12cf1f43
update zen models
2026-05-31 19:47:33 -04:00
Frank
06c7ea2303
update zen models
2026-05-31 19:47:04 -04:00
Frank
92b6e9da53
update zen models
2026-05-31 19:43:09 -04:00
Frank
660ff026a0
update zen models
2026-05-31 19:40:57 -04:00
Aiden Cline
9a910a1fb3
Merge pull request #1917 from markgibaud/fix/eu-opus-bedrock-pricing
...
fix(bedrock): correct EU cross-region pricing for Claude Opus and Sonnet models
2026-05-31 17:37:45 -05:00
Aiden Cline
7dc5f787b0
Merge pull request #1915 from anomalyco/automation/sync-models-openrouter
...
chore(sync): update OpenRouter model catalog
2026-05-31 17:10:10 -05:00
Aiden Cline
c25ddce07c
Merge pull request #1916 from anomalyco/automation/sync-models-cloudflare-workers-ai
...
chore(sync): update Cloudflare Workers AI model catalog
2026-05-31 17:05:57 -05:00
github-actions[bot]
e73a82e44e
chore(sync): update OpenRouter model catalog
2026-05-31 21:36:49 +00:00
github-actions[bot]
7de9710b5a
chore(sync): update Cloudflare Workers AI model catalog
2026-05-31 21:36:49 +00:00
Frank
d1a8dbcdde
update zen models
2026-05-31 14:19:49 -04:00
markgibaud
556ef84d11
fix(bedrock): correct EU cross-region pricing for Claude Opus and Sonnet models
...
AWS Bedrock EU (Europe/London) cross-region inference has a 10% premium
over the base Anthropic pricing. The EU models were incorrectly using
the same pricing as the US/base models.
Affected models:
- eu.anthropic.claude-opus-4-6-v1
- eu.anthropic.claude-opus-4-7
- eu.anthropic.claude-opus-4-8
- eu.anthropic.claude-sonnet-4-5-20250929-v1:0
- eu.anthropic.claude-sonnet-4-6
Corrected Opus per 1M token prices:
- Input: $5.00 -> $5.50
- Output: $25.00 -> $27.50
- Cache read: $0.50 -> $0.55
- Cache write: $6.25 -> $6.875
Corrected Sonnet per 1M token prices:
- Input: $3.00 -> $3.30
- Output: $15.00 -> $16.50
- Cache read: $0.30 -> $0.33
- Cache write: $3.75 -> $4.125
Source: AWS Bedrock pricing page, Europe (London) region
2026-05-31 11:07:18 +01:00
Aiden Cline
f2020553ea
Merge pull request #1898 from Suat-B/codex/xpersona-frieren-1
...
Rename Xpersona Frieren display name
2026-05-30 16:44:39 -05:00
Aiden Cline
ca0a7e17cb
Merge pull request #1910 from anomalyco/automation/sync-models-openrouter
...
chore(sync): update OpenRouter model catalog
2026-05-30 16:44:22 -05:00
github-actions[bot]
d2468eafb2
chore(sync): update OpenRouter model catalog
2026-05-30 21:37:16 +00:00
Aiden Cline
96d0f9beff
Merge pull request #1913 from orangeclk/add-glm-4.6v-zhipuai-coding-plan
...
feat(zhipuai-coding-plan): add glm-4.6v symlink from zai
2026-05-30 11:52:50 -05:00
Aiden Cline
7989be9f34
Merge pull request #1912 from Jercik/codex/update-wafer-provider-models
...
feat: update Wafer provider models
2026-05-30 11:52:41 -05:00
Łukasz Jerciński
f30e44f4b5
feat: update Wafer provider models
2026-05-30 14:53:02 +02:00
OrangeCLK
1383177893
feat(zhipuai-coding-plan): add glm-4.6v symlink from zai
2026-05-30 18:42:09 +08:00
Aiden Cline
2e58165af9
Merge pull request #1902 from kameshsampath/feat/provider/snowflake-cortex
...
feat(snowflake-cortex): add Snowflake Cortex provider
2026-05-29 23:55:55 -05:00
Kamesh Sampath
f3466affc0
feat(snowflake-cortex): add Snowflake Cortex provider
...
Adds the snowflake-cortex provider which exposes Snowflake's Cortex
REST API (OpenAI Chat Completions-compatible endpoint) to opencode.
Provider details:
- npm: @ai-sdk/openai-compatible
- API: https://${SNOWFLAKE_ACCOUNT}.snowflakecomputing.com/api/v2/cortex/v1
- Auth: SNOWFLAKE_ACCOUNT + SNOWFLAKE_CORTEX_PAT (Programmatic Access Token)
Models (11, all with tool_call support):
- Anthropic: claude-opus-4-7 (beta/preview), claude-sonnet-4-6,
claude-sonnet-4-5, claude-haiku-4-5
- OpenAI: openai-gpt-5.4 (beta), openai-gpt-5.2, openai-gpt-5.1,
openai-gpt-5 (beta), openai-gpt-5-mini (beta), openai-gpt-5-nano (beta),
openai-gpt-4.1
Models without tool_call support (deepseek-r1, llama3.1-70b,
snowflake-llama-3.3-70b, mistral-large2) are excluded per Snowflake docs:
"Tool calling is supported for OpenAI and Claude models only."
Preview/not-GA models are marked with status = "beta".
Cost fields are intentionally omitted on all models. The Cortex REST API
is billed in USD (AI_INFERENCE service type) at rates defined in the
Snowflake Service Consumption Table.
Closes #1895
2026-05-30 09:04:43 +05:30
Aiden Cline
3697c99297
Merge pull request #1901 from anomalyco/automation/sync-models-openrouter
...
chore(sync): update OpenRouter model catalog
2026-05-29 20:50:33 -05:00
Aiden Cline
796dda2fd1
Merge pull request #1904 from sdnts/dev
...
feat(cloudflare-ai-gateway): add opus 4.8
2026-05-29 20:50:11 -05:00
Aiden Cline
1d363cbd7f
Merge pull request #1908 from anomalyco/fix/unique-provider-names
...
Ensure provider names are unique
2026-05-29 20:49:54 -05:00
Aiden Cline
3758725979
Ensure provider names are unique
...
Rename the China StepFun provider to "StepFun (China)" so it no longer
collides with the international "StepFun" provider, following the existing
Alibaba / Alibaba (China) naming convention.
Add a validation check in generate() that throws when two providers share
a (case-insensitive) name, preventing future duplicates.
Fixes #1906
2026-05-29 20:48:09 -05:00
github-actions[bot]
c039e82f12
chore(sync): update OpenRouter model catalog
2026-05-30 01:16:42 +00:00
Siddhant
1e90242cdb
feat(cloudflare-ai-gateway): add opus 4.8
2026-05-29 13:55:09 -04:00
Aiden Cline
277ac8577e
Merge pull request #1897 from mikeyp/update-digitalocean-models
...
Add deepseek-4-flash and claude-opus-4.8 for DigitalOcean
2026-05-29 10:20:26 -05:00
Aiden Cline
6e5f35546d
Merge pull request #1900 from ceyhanmolla/add-step-3.7-flash-v2
...
Add Step 3.7 Flash model to NVIDIA provider
2026-05-29 10:20:02 -05:00
ceyhanmolla
aa69da14d6
Add Step 3.7 Flash model to NVIDIA provider
...
StepFun AI's Step 3.7 Flash - sparse MoE multimodal reasoning model:
- 198B total params, ~11B active per token
- 256K context window with sliding window attention
- Text + image input, text output
- Reasoning, tool calling, and attachment support
- Apache 2.0 license
2026-05-29 12:55:13 +02:00
Jack
c7af5ba9be
update model in Go
2026-05-29 16:20:16 +08:00
Suat-B
d6e9e6735f
Rename Xpersona Frieren model display name
2026-05-29 03:00:21 -05:00
Mike Prasuhn
a987719410
Add deepseek-4-flash and claude-opus-4.8
2026-05-29 03:45:50 -04:00
Steven Yeung
3b792029c3
feat(alibaba-token-plan): add provider and 15 model TOMLs
...
Add Alibaba Cloud Model Studio Token Plan (Team Edition) provider with:
- Provider config (Singapore region, OpenAI-compatible endpoint)
- 10 models using extends pattern (zero-cost overrides from canonical providers)
- 5 models with full definitions (no canonical source available)
- Image generation models (qwen-image, wan2.7) with output=0 per convention
- deepseek-v3.2 with structured_output and corrected release date
Models: qwen3.7-max, qwen3.6-flash, qwen3.6-plus, kimi-k2.5, kimi-k2.6,
glm-5, glm-5.1, MiniMax-M2.5, deepseek-v4-pro, deepseek-v4-flash,
deepseek-v3.2, qwen-image-2.0, qwen-image-2.0-pro, wan2.7-image, wan2.7-image-pro
2026-05-29 15:12:27 +08:00
Aiden Cline
9528528f69
Merge pull request #1890 from anomalyco/automation/sync-models-openrouter
...
chore(sync): update OpenRouter model catalog
2026-05-28 23:42:03 -05:00
github-actions[bot]
fbad40d863
chore(sync): update OpenRouter model catalog
2026-05-29 03:26:05 +00:00
Aiden Cline
84559b598b
Merge pull request #1894 from anomalyco/add-siliconflow-cn-deepseek-v4-pro
...
feat(siliconflow-cn): add deepseek-ai/DeepSeek-V4-Pro
2026-05-28 19:10:18 -05:00
Aiden Cline
9d45730db9
Merge pull request #1889 from fhennerkes/dev
...
poe: add Claude-Opus-4.8 model
2026-05-28 19:10:06 -05:00
Aiden Cline
791089e4aa
Merge pull request #1891 from ticoombs/dev
...
feat(copilot): update all copilot models, add claude-opus-4.8
2026-05-28 19:09:52 -05:00
Aiden Cline
fcc4387187
feat(siliconflow-cn): add deepseek-ai/DeepSeek-V4-Pro
2026-05-28 19:09:15 -05:00
Tim C
6aa32ba566
fix(copilot): update all context,input,output with correct limits
2026-05-29 08:56:19 +10:00
Tim C
459a563d2b
feat(copilot): add claude-opus-4.8
2026-05-29 08:52:42 +10:00
Aiden Cline
739e5a7c8e
Merge pull request #1877 from aakash-gupte/add-merge-gateway-provider
...
Add Merge Gateway provider
2026-05-28 16:55:40 -05:00
Aiden Cline
efc87afa3a
Merge pull request #1888 from smakosh/add-llmgateway-opus-4-8
...
feat(llmgateway): add Claude Opus 4.8
2026-05-28 16:54:46 -05:00
Aiden Cline
f616aa6a0b
Merge pull request #1887 from dpuyosa/dev
...
Venice: Add Claude Opus 4.8 models
2026-05-28 16:54:35 -05:00
fhennerkes
f77e75a567
poe: add Claude-Opus-4.8 model
...
Add new Anthropic model from Poe API (released 2026-05-28).
Uses extends format inheriting from anthropic/claude-opus-4-8
with Poe-specific overrides (name format, markup pricing,
slightly different context limit).
Co-Authored-By: Claude Opus 4.7 <noreply@anthropic.com >
2026-05-28 14:39:52 -07:00
smakosh
651bd4d91a
feat(llmgateway): add Claude Opus 4.8
...
Extends anthropic/claude-opus-4-8, omitting the fast mode which LLM
Gateway does not expose, matching the existing 4.6/4.7 entries.
Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com >
2026-05-28 22:07:40 +01:00
dpuyosa
30b9170106
[venice] Add Claude Opus 4.8 models
...
- Add claude-opus-4-8.toml (standard pricing)
- Add claude-opus-4-8-fast.toml (fast variant)
- Both: 1M context, image input, Opus 4.8 family
2026-05-28 22:54:39 +02:00
Aiden Cline
81e42c653c
Merge pull request #1886 from vglafirov/add-gitlab-opus-4-8
...
feat: add GitLab Agentic Chat Opus 4.8
2026-05-28 15:25:06 -05:00
Vladimir Glafirov
11f2a38594
feat: add GitLab Agentic Chat Opus 4.8
2026-05-28 22:12:12 +02:00
Aiden Cline
6621be978f
Merge pull request #1880 from bas3line/update-routing-run-base-url
...
Update routing.run API base URL
2026-05-28 14:46:09 -05:00
Aiden Cline
c8d72661d6
Merge pull request #1885 from alaviss/bedrock-opus-4-8
...
feat: add cross-region inference entries for Bedrock for Opus 4.8
2026-05-28 14:45:43 -05:00
Hiếu Lê
21b7cacc33
feat: add cross-region inference entries for Bedrock for Opus 4.8
2026-05-28 12:38:19 -07:00
Aiden Cline
ac1dc14439
Merge pull request #1884 from calebboyd/update-vercel-models-latest
...
feat: add vercel ai gateway opus 4.8
2026-05-28 14:31:03 -05:00
Aiden Cline
63acb799da
Merge pull request #1883 from anomalyco/automation/sync-models-openrouter
...
chore(sync): update OpenRouter model catalog
2026-05-28 14:30:56 -05:00
github-actions[bot]
2e6ebb6750
chore(sync): update OpenRouter model catalog
2026-05-28 19:11:47 +00:00
calebboyd
d8c2e15632
feat: add vercel ai gateway opus 4.8
2026-05-28 13:33:37 -05:00
Frank
8a70160200
update zen model
2026-05-28 14:01:35 -04:00
Aiden Cline
7ea1d4e15c
Merge pull request #1882 from anomalyco/add-opus-4.8
...
feat: add opus 4.8
2026-05-28 12:04:51 -05:00
Aiden Cline
1c71b1bf17
feat: add opus 4.8
2026-05-28 12:04:14 -05:00
Aiden Cline
acc3126194
Merge pull request #1879 from Ardakilic/chore/ci-fork-ensurance
...
CI Scheduled fork upstream ensurance
2026-05-28 11:27:47 -05:00
Aiden Cline
1f67addc01
Merge pull request #1878 from anomalyco/automation/sync-models-openrouter
...
chore(sync): update OpenRouter model catalog
2026-05-28 11:27:07 -05:00
github-actions[bot]
bb011e71a7
chore(sync): update OpenRouter model catalog
2026-05-28 15:28:44 +00:00
bas3line
3df181411f
fix(routing-run): update API base URL
2026-05-28 08:58:33 +05:30
Arda Kilicdagi
a1f588463c
chore: prevent forks to run cronjobbed sync commands
2026-05-28 06:25:45 +04:00
Arda Kilicdagi
cb414c6886
chore: prevent forks to run cronjobbed sync commands
2026-05-28 05:31:19 +04:00
Arda Kilicdagi
4095e819c9
chore: prevent forks to run cronjobbed sync commands
...
chore: prevent forks to run cronjobbed sync commands
2026-05-28 05:26:15 +04:00
Aakash Gupte
9367cc1825
Add Merge Gateway logo
...
Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com >
2026-05-27 17:22:44 -04:00
Aakash Gupte
96b9a04c4b
Add Merge Gateway provider
...
Merge Gateway (https://merge.dev ) is an LLM gateway exposing an
OpenAI/Anthropic-compatible API across many providers, using the
published `merge-gateway-ai-sdk-provider` npm package. Models are
defined via `extends` from existing canonical entries.
Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com >
2026-05-27 14:32:21 -04:00
Jack
d9de217dc4
update zen models
2026-05-28 01:46:38 +08:00
Aiden Cline
64ea80d416
Merge pull request #1869 from anomalyco/fix/xiaomi-token-plan-models
...
fix Xiaomi Token Plan model catalog
2026-05-27 12:02:22 -05:00
Aiden Cline
d3859517d1
Merge pull request #1870 from anomalyco/automation/sync-models-openrouter
...
chore(sync): update OpenRouter model catalog
2026-05-27 12:01:59 -05:00
Aiden Cline
985a144c5f
Merge pull request #1866 from tarikko/patch-1
...
Update Mistral Small model details in TOML file
2026-05-27 11:43:52 -05:00
Jack
5b05070eb4
Merge pull request #1875 from anomalyco/update/opencode-go-mimo-v2-5-pricing
...
update opencode go MiMo V2.5 pricing
2026-05-28 00:39:53 +08:00
Aiden Cline
dad245c5e1
Merge pull request #1871 from oskarkocol/chore/update-groq-cache-prices
...
chore: update groq cache pricing
2026-05-27 11:39:22 -05:00
Aiden Cline
bcff9bd6b2
Merge pull request #1874 from Alex-yang00/codex/sync-novita-models
...
Sync NovitaAI models
2026-05-27 11:39:03 -05:00
Aiden Cline
561bdaa546
Merge pull request #1876 from sebastiand-cerebras/deprecate-cerebras-qwen235b-llama8b-20260527
...
Remove deprecated Cerebras Qwen 3 235B model
2026-05-27 11:38:26 -05:00
Jack
f4f210986c
fix opencode go MiMo pricing conflict resolution
2026-05-28 00:36:48 +08:00
Jack
f4206f4eaa
Merge branch 'dev' into update/opencode-go-mimo-v2-5-pricing
2026-05-28 00:34:02 +08:00
Seb Duerr
52e5b67bc4
Remove deprecated Cerebras Qwen 3 235B model
2026-05-27 09:13:50 -07:00
Jack
d11b0e448c
update opencode go MiMo V2.5 pricing
2026-05-27 23:54:58 +08:00
github-actions[bot]
591745690e
chore(sync): update OpenRouter model catalog
2026-05-27 15:28:06 +00:00
Codex
5a4bc9b383
Add selected NovitaAI models
2026-05-27 21:30:38 +08:00
Tarik
a14171bcd4
Update mistral-small.toml
...
the only difference is the price so I removed the other fields
2026-05-27 14:00:33 +01:00
oskar
0741c5a53c
chore: update kimi cache rate
2026-05-27 13:09:20 +07:00
oskar
7a1ec8737b
chore: update groq cache pricing
2026-05-27 13:05:24 +07:00
Aiden Cline
7c37c92c2d
fix Xiaomi Token Plan model catalog
2026-05-27 00:39:17 -05:00
Aiden Cline
ec4ec6d441
Merge pull request #1867 from anomalyco/automation/sync-models-openrouter
...
chore(sync): update OpenRouter model catalog
2026-05-27 00:00:17 -05:00
Aiden Cline
d96da5ece6
Merge pull request #1868 from Ardakilic/chore/sync-kilo-models-20260527-1
...
Sync Kilo models with upstream gateway
2026-05-26 23:59:26 -05:00
github-actions[bot]
0342c79d03
chore(sync): update OpenRouter model catalog
2026-05-27 03:26:29 +00:00
Arda Kilicdagi
37136fc2f3
chore: sync upstream kilo api gateway models
2026-05-27 02:27:53 +04:00
Tarik
131eca9c5e
Update Mistral Small model configuration
...
use [extends] syntax
2026-05-26 20:44:22 +01:00
Tarik
3e38d46c4c
Update Mistral Small model details in TOML file
...
The details on helicone website are false,
accurate details are pulled from deepinfra website
https://www.helicone.ai/model/mistral-small
https://deepinfra.com/mistralai/Mistral-Small-3.2-24B-Instruct-2506
2026-05-26 19:34:07 +01:00
Aiden Cline
989939773b
Merge pull request #1863 from anomalyco/automation/sync-models-openrouter
...
chore(sync): update OpenRouter model catalog
2026-05-26 13:04:26 -05:00
Frank
4f1d5c511a
update go models
2026-05-26 13:54:19 -04:00
github-actions[bot]
34a21e07f2
chore(sync): update OpenRouter model catalog
2026-05-26 17:26:26 +00:00
Aiden Cline
979f5da0ca
Merge pull request #1864 from oskarkocol/update/openai-cache-rates
...
chore: update openai cache rates
2026-05-26 12:06:15 -05:00
oskar
73b357e2ab
update cache rates
2026-05-26 23:28:11 +07:00
Aiden Cline
97c71f1663
Merge pull request #1862 from anomalyco/automation/sync-models-openrouter
...
chore(sync): update OpenRouter model catalog
2026-05-26 09:11:54 -05:00
github-actions[bot]
59c270bc85
chore(sync): update OpenRouter model catalog
2026-05-26 13:23:20 +00:00
Aiden Cline
a4c18f88ed
Merge pull request #1860 from bas3line/sync-routing-run-models-2
...
Sync routing.run model catalog
2026-05-25 23:44:24 -05:00
bas3line
3343a585ac
fix(routing-run): sync model catalog
2026-05-26 09:51:14 +05:30
Aiden Cline
f23db95550
Merge pull request #1858 from smakosh/fix/llmgateway-qwen3.7-max-id
...
fix(llmgateway): correct Qwen3.7 Max model id to qwen3.7-max
2026-05-25 17:22:56 -05:00
Aiden Cline
556d9a9045
Merge pull request #1859 from Suat-B/update-xpersona-frieren-coder-limits
...
Update Xpersona Frieren Coder limits
2026-05-25 17:22:46 -05:00
Frank
49991c8f8f
update zen models
2026-05-25 17:56:52 -04:00
Suat-B
20c1e810ce
Update Xpersona Frieren Coder limits
2026-05-25 13:23:59 -05:00
smakosh
15d08b4540
fix(llmgateway): correct Qwen3.7 Max model id to qwen3.7-max
...
Rename qwen37-max.toml to qwen3.7-max.toml so the model id matches
the canonical alibaba/qwen3.7-max definition (name: Qwen3.7 Max).
Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com >
2026-05-25 18:48:12 +01:00
Aiden Cline
14bbc303e5
Merge pull request #1852 from anomalyco/automation/sync-models-openrouter
...
chore(sync): update OpenRouter model catalog
2026-05-25 09:49:26 -05:00
Aiden Cline
02da63ed48
Merge pull request #1853 from dpuyosa/chore/venice-pricing
...
Venice: Update pricing and limits for 4 models
2026-05-25 09:49:09 -05:00
github-actions[bot]
07daef8c54
chore(sync): update OpenRouter model catalog
2026-05-25 13:27:22 +00:00
dpuyosa
b6ebe8696c
[venice] Update pricing, limits, and last_updated for 4 models
...
- Reduce input/output/cache prices for gemini-3-5-flash, google-gemma-4-31b-it, and qwen-3-7-max
- Lower max output tokens from 65,536 to 16,384 for qwen3-5-35b-a3b
- Bump last_updated to 2026-05-25 for all 4 models
2026-05-25 10:07:32 +02:00
Aiden Cline
ad654e71da
Merge pull request #1850 from huxeon/dev
...
fix: modify the deepseek v4 flash/pro price
2026-05-25 00:18:18 -05:00
Aiden Cline
508f4d48e1
Merge pull request #1847 from ceyhanmolla/poolside/laguna-direct
...
Add Poolside provider with Laguna M.1 and XS.2 models
2026-05-24 23:26:34 -05:00
opencode-agent[bot]
cd7c70b4fe
revert: remove opencode-go deepseek-v4-pro price changes
...
Keep only the deepseek provider price updates as intended.
2026-05-25 04:26:14 +00:00
Aiden Cline
10e752ea84
Merge pull request #1845 from yukoba/vultr
...
Update Vultr models
2026-05-24 23:26:12 -05:00
huxeon
4cdb4c700b
fix: modify the deepseek v4 flash/pro price in provider deepseek and opencode-go
2026-05-24 13:33:56 +08:00
Aiden Cline
d497a446eb
Merge pull request #1849 from technoabsurdist/add-wafer-ai-qwen3.6-35b-a3b-and-kimi-k2.6
...
providers/wafer.ai: add Qwen3.6-35B-A3B and Kimi-K2.6
2026-05-23 16:41:24 -05:00
Emilio Andere
745cf557b3
feat(wafer.ai): add Qwen3.6-35B-A3B and Kimi-K2.6
...
Both models are public serverless on pass.wafer.ai/v1/models but were
missing from the wafer.ai provider in models.dev, so OpenCode and other
tools that pull from the registry could not discover them.
- Qwen3.6 35B A3B: compact MoE, 32K context, vision-capable, $0.19/M in,
$1.25/M out (NVFP4 on AMD MI355X — see wafer.ai/blog/qwen36-mi355x).
- Kimi K2.6: 1T sparse MoE, 262K context, vision-capable, $1.10/M in,
$4.80/M out (NVFP4 on Blackwell — see wafer.ai/blog/kimi-k26-nvfp4).
Co-authored-by: Cursor <cursoragent@cursor.com >
2026-05-23 13:26:18 -04:00
Yu Kobayashi
db295d6842
Refactor to use the extends syntax
2026-05-24 01:20:31 +09:00
ceyhanmolla
47c0eed41e
Add Poolside provider with Laguna M.1 and XS.2 models
2026-05-23 17:09:33 +02:00
Yu Kobayashi
e690980857
Update Vultr models
2026-05-23 16:25:11 +09:00
Aiden Cline
f5f7d1a167
Merge pull request #1840 from anomalyco/automation/sync-models-openrouter
...
chore(sync): update OpenRouter model catalog
2026-05-22 16:58:23 -05:00
Aiden Cline
cfbbb55c4d
Merge pull request #1839 from anthraxx/alibaba-qwen3.7-plus
...
Alibaba: Add Qwen 3.7 Max and 3.6 Flash to all regions and plans
2026-05-22 16:58:08 -05:00
github-actions[bot]
1956762cdf
chore(sync): update OpenRouter model catalog
2026-05-22 21:41:35 +00:00
Levente Polyak
28571433f7
feat(alibaba): add Qwen3.7 Max model configuration to all regions
...
Link: https://bailian.console.alibabacloud.com/cn-beijing?tab=model#/model-market/detail/qwen3.7-max?serviceSite=asia-pacific-china
2026-05-22 19:00:29 +02:00
Levente Polyak
aef5e48bae
feat(alibaba): add Qwen3.6 Flash model configuration to all regions
...
Link: https://bailian.console.alibabacloud.com/cn-beijing?tab=model#/model-market/detail/qwen3.6-flash?serviceSite=asia-pacific-china
2026-05-22 18:54:22 +02:00
Aiden Cline
8ba19639a7
Merge pull request #1825 from shzdehmd/dev
...
update(fireworks): sync models and pricing with current offerings
2026-05-22 11:45:45 -05:00
Aiden Cline
0f9b4c9edc
Merge pull request #1834 from monotykamary/chore/update-neuralwatt-qwen3.6-pricing
...
fix(neuralwatt): update Qwen3.6 pricing to match API
2026-05-22 09:13:02 -05:00
Aiden Cline
2a9b6256dc
Merge pull request #1836 from PierreLeGuen/nearai-provider
...
Add current NEAR AI Cloud models
2026-05-22 09:12:36 -05:00
Aiden Cline
51633fe106
Merge pull request #1838 from Quentinchampenois/fix/update-models-scaleway
...
fix: update Scaleway provider models list
2026-05-22 09:12:26 -05:00
Aiden Cline
169b5c4331
Merge pull request #1833 from NicoAvanzDev/add-copilot-gemini-3-5-flash
...
[GitHub Copilot] add Gemini 3.5 Flash
2026-05-22 09:09:25 -05:00
Aiden Cline
9d0f0d6d56
Merge pull request #1837 from fydrah/fix/google-vertex-gemini-3.5-flash
...
fix: missing google-vertex gemini 3.5 flash extend
2026-05-22 09:07:34 -05:00
Aiden Cline
1d2af0c97b
Merge pull request #1835 from dpuyosa/feat/venice-models
...
Venice: Add Gemini 3.5 Flash and Qwen 3.7 Max, update Grok Build 0.1 pricing
2026-05-22 09:07:03 -05:00
Aiden Cline
cc14ae4370
Merge pull request #1832 from anomalyco/automation/sync-models-openrouter
...
chore(sync): update OpenRouter model catalog
2026-05-22 09:06:16 -05:00
github-actions[bot]
ad9a3448d2
chore(sync): update OpenRouter model catalog
2026-05-22 14:03:32 +00:00
Quentin Champenois
c6f9cf58e4
fix(scaleway): add new models gemma-4-26b-a4b-it and qwen3.6-35b-a3b
2026-05-22 15:43:11 +02:00
Quentin Champenois
bba5809471
fix(scaleway): Extends existing models
2026-05-22 15:42:36 +02:00
Quentin Champenois
c0852f1b4f
fix(scaleway): Clear removed models from list
2026-05-22 15:11:44 +02:00
Flavien Hardy
0fc3a3f635
fix: missing google-vertex gemini 3.5 flash extend
2026-05-22 08:56:44 -04:00
Pierre LE GUEN
68936dc062
Add current NEAR AI Cloud models
2026-05-22 09:27:57 +00:00
dpuyosa
ade9760060
[venice] Add Gemini 3.5 Flash and Qwen 3.7 Max, update Grok Build 0.1 pricing
...
- Add Gemini 3.5 Flash model (1M context, multimodal input)
- Add Qwen 3.7 Max model (1M context, text-only)
- Update Grok Build 0.1 cost tiers and pricing
2026-05-22 10:57:38 +02:00
Tom X Nguyen
2a8b90b197
fix(neuralwatt): update Qwen3.6 pricing to match API
...
Updates the per-token cost for Qwen3.6-35B-A3B and qwen3.6-35b-fast from /bin/bash.05//bin/bash.10 to /bin/bash.29/.15 (input/output per million tokens), matching the actual Neuralwatt API pricing as reflected in pi-neuralwatt-provider commit f634286.
2026-05-22 15:19:51 +07:00
NicoAvanzDev
8569f0dfef
[GitHub Copilot] add Gemini 3.5 Flash
2026-05-22 07:57:10 +00:00
Aiden Cline
fc98ceb72e
fix sync workflow force lease
2026-05-21 23:52:34 -05:00
Aiden Cline
be7c5afc94
Merge pull request #1831 from anomalyco/fix-vertex-sonnet-4-6-limits
...
Fix Vertex Sonnet 4.6 token limits
2026-05-21 23:38:18 -05:00
Aiden Cline
57e62c43b0
fix vertex sonnet 4.6 limits
2026-05-21 23:37:30 -05:00
Aiden Cline
0898c35c9f
Merge pull request #1830 from zainhas/dev
...
[Together AI] add Qwen3.7 max
2026-05-21 21:00:51 -05:00
Zain Hasan
46b23fb313
Merge branch 'anomalyco:dev' into dev
2026-05-21 17:47:35 -07:00
Zain Hasan
05fedc76cc
[Together AI] add Qwen3.7
2026-05-21 17:47:19 -07:00
Aiden Cline
2738f81d1a
Merge pull request #1828 from anomalyco/refactor/sync-core-layout
...
refactor: move sync implementation into core src
2026-05-21 18:10:09 -05:00
Aiden Cline
89b834086a
refactor: move sync implementation into core src
2026-05-21 18:06:25 -05:00
Aiden Cline
1ab2ff8163
Merge pull request #1826 from smakosh/add-llmgateway-models
...
feat: add new LLM Gateway text models
2026-05-21 17:58:29 -05:00
Frank
9468676683
update zen models
2026-05-21 18:42:36 -04:00
Claude
6cdd2f054b
Merge upstream/dev into add-llmgateway-models; resolve gemini-3.5-flash conflict
...
# Conflicts:
# providers/google/models/gemini-3.5-flash.toml
2026-05-21 21:48:48 +00:00
Aiden Cline
b13abc9141
Merge pull request #1827 from anomalyco/update-xai-pricing
...
fix xAI long-context pricing
2026-05-21 16:45:43 -05:00
Aiden Cline
e5ba264751
fix xAI long-context pricing
2026-05-21 16:41:36 -05:00
smakosh
a7811fb522
refactor: use extends for gemini and qwen models
...
Add canonical google/gemini-3.5-flash and alibaba/qwen3.7-max defs and
have the llmgateway entries extend them, per PR review.
Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com >
2026-05-21 23:38:57 +02:00
smakosh
605fae75d9
feat: add new LLM Gateway text models
...
Add Grok 4.20 (reasoning/non-reasoning), Gemini 3.5 Flash, and Qwen3.7 Max to the llmgateway provider.
Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com >
2026-05-21 23:04:47 +02:00
Ahmad Shahzad
bcab0885bc
update(fireworks): sync models and pricing with current offerings
...
Removed — 11 deprecated models:
- deepseek-v3p1
- deepseek-v3p2
- glm-4p5
- glm-4p5-air
- glm-4p7
- glm-5
- kimi-k2-instruct
- kimi-k2-thinking
- minimax-m2p1
- routers/kimi-k2p5-turbo
Added — 2 new Turbo (routers) models:
- routers/glm-5p1-fast
- routers/kimi-k2p6-turbo
Modified — pricing fixes:
- deepseek-v4-pro: cache_read 0.15 → 0.145
- gpt-oss-120b: added cache_read = 0.015
- gpt-oss-20b: input 0.05 → 0.07, output 0.20 → 0.30, added cache_read = 0.035
- minimax-m2p7: cache_read 0.03 → 0.06
2026-05-22 01:55:31 +05:00
Aiden Cline
26b05268ae
Merge pull request #1824 from anomalyco/fix/vercel-gemini-35-flash
...
Add new Vercel AI Gateway models
2026-05-21 15:30:35 -05:00
Aiden Cline
1aee13d2e5
Add new Vercel AI Gateway models
2026-05-21 13:21:25 -05:00
Aiden Cline
0a924e6bb2
Merge pull request #1822 from anomalyco/automation/sync-models-xai
...
chore(sync): update xAI model catalog
2026-05-21 13:05:20 -05:00
Aiden Cline
9769b2b11d
Merge pull request #1823 from anomalyco/automation/sync-models-openrouter
...
chore(sync): update OpenRouter model catalog
2026-05-21 13:05:13 -05:00
github-actions[bot]
1b0db099cf
chore(sync): update OpenRouter model catalog
2026-05-21 17:56:13 +00:00
github-actions[bot]
16ed78587c
chore(sync): update xAI model catalog
2026-05-21 17:56:11 +00:00
Frank
0b88965165
update zen models
2026-05-21 13:41:55 -04:00
Aiden Cline
acc704ce39
Merge pull request #1797 from arnavchachra/add-crof-provider
...
add crof.ai provider with 21 models
2026-05-21 11:43:02 -05:00
Aiden Cline
51ad3b264e
Merge pull request #1821 from anomalyco/sync-provider-ci
...
chore: automate provider sync jobs
2026-05-21 11:23:49 -05:00
Aiden Cline
146b6c7084
Merge pull request #1819 from Inceptron-Software/add_inceptron_provider
...
Add Inceptron provider
2026-05-21 11:21:27 -05:00
Aiden Cline
0e3cbe3c64
chore: automate provider sync jobs
2026-05-21 11:21:12 -05:00
Aiden Cline
604d4d66a4
Merge pull request #1820 from Suat-B/codex/xpersona-image-input-20260521
...
Add image input modality to Xpersona model
2026-05-21 11:00:49 -05:00
SuatB
f5090028b8
Add image input modality to Xpersona model
2026-05-21 09:46:28 -05:00
Frank
4bad8faf29
update zen models
2026-05-21 09:05:13 -04:00
Oskar Gustafsson
0df2ccf586
Add Inceptron provider
2026-05-21 09:24:43 +02:00
Aiden Cline
bafdc00b45
Merge pull request #1812 from anomalyco/openrouter-extends-sync
...
Sync OpenRouter models with extends
2026-05-20 21:07:21 -05:00
Aiden Cline
49840c013b
Merge pull request #1814 from neonn0d/feat/stepfun-ai
...
feat(stepfun-ai): add international StepFun platform
2026-05-20 20:58:37 -05:00
Aiden Cline
eccae0b54e
sync openrouter models with extends
2026-05-20 20:32:11 -05:00
Aiden Cline
4cca29405f
Merge pull request #1817 from dpuyosa/dev
...
Venice: Remove Grok 4.1 Fast and add Grok Build 0.1
2026-05-20 20:26:45 -05:00
Aiden Cline
e40d9dd338
Merge pull request #1818 from anomalyco/cloudflare-sync-env
...
chore(sync): isolate cloudflare credentials
2026-05-20 20:26:22 -05:00
Aiden Cline
6a74991397
chore(sync): isolate cloudflare credentials
2026-05-20 20:19:47 -05:00
dpuyosa
035999cb58
[venice] Replace Grok 4.1 Fast with Grok Build 0.1
...
- Remove deprecated grok-41-fast model entry
- Add grok-build-0-1 with 200K token tiered pricing
- Update context to 256K and output limit to 65,536
2026-05-21 02:36:19 +02:00
Frank
cec56bf1bc
update zen models
2026-05-20 19:43:25 -04:00
Aiden Cline
85f0cdcb2f
Merge pull request #1816 from anomalyco/xai-sync
...
Add PDF input modality to Grok models
2026-05-20 18:12:47 -05:00
Aiden Cline
ef80d4df4e
Infer PDF modality for xAI image models
2026-05-20 18:12:11 -05:00
Aiden Cline
af0ef00109
Update xAI Grok PDF modalities
2026-05-20 18:06:33 -05:00
Aiden Cline
5fdcea6b36
Merge pull request #1815 from anomalyco/cloudflare-ai-gateway
...
chore(sync): add cloudflare workers ai sync
2026-05-20 18:02:28 -05:00
Aiden Cline
31e56480b4
chore(sync): add cloudflare workers ai sync
2026-05-20 16:55:29 -05:00
Aiden Cline
92a621594e
Merge pull request #1813 from anomalyco/sync-xai
...
chore(sync): add xai model sync
2026-05-20 16:02:38 -05:00
Aiden Cline
d2db353ceb
chore: ignore sync reports
2026-05-20 16:01:51 -05:00
Aiden Cline
900ae509d2
Merge pull request #1808 from ajussak/scaleway
...
Added Mistral Medium 3.5 128B from Scaleway
2026-05-20 15:58:00 -05:00
neo
9d60164243
feat(stepfun-ai): add international StepFun platform
...
StepFun runs two separate platforms with distinct accounts/keys:
platform.stepfun.com (China, already covered by providers/stepfun) and
platform.stepfun.ai (international). Keys are not interchangeable
across the two — .ai keys are rejected by api.stepfun.com as
invalid_api_key.
Stepfun's own opencode integration guide instructs users to point at
https://api.stepfun.ai/step_plan/v1 . This adds providers/stepfun-ai
for that endpoint, symlinking the shared chat models. Follows the
moonshotai / moonshotai-cn pattern.
2026-05-20 20:43:47 +02:00
Adrien Jussak
cd3e99025f
Update Mistral Medium 3.5 128B model configuration to extend from mistral-medium-2604 and adjust context window size.
2026-05-20 20:43:45 +02:00
Aiden Cline
1098981eb6
chore(sync): add xai model sync
2026-05-20 13:23:02 -05:00
Aiden Cline
27a151cf53
Merge pull request #1811 from anomalyco/xai-grok-build-model
...
Add xAI Grok Build model
2026-05-20 13:05:06 -05:00
Aiden Cline
41ff42ab7f
Merge pull request #1810 from fhennerkes/dev
...
poe: add Gemini-3.5-Flash model
2026-05-20 13:04:47 -05:00
Aiden Cline
adf1cbdecd
add xai grok build model
2026-05-20 13:04:13 -05:00
Frank
10ddc78ce0
update zen models
2026-05-20 14:02:01 -04:00
fhennerkes
7ae897e440
poe: add Gemini-3.5-Flash model
...
Add new Google model from Poe API (released 2026-05-19).
Uses extends format inheriting from google/gemini-3.5-flash with
Poe-specific overrides (name format, no temperature, markup pricing,
limited input modalities).
Co-Authored-By: Claude Opus 4.7 <noreply@anthropic.com >
2026-05-20 10:52:15 -07:00
Aiden Cline
02e452c2e8
Merge pull request #1805 from anomalyco/sync-google
...
sync google models
2026-05-20 10:53:57 -05:00
Aiden Cline
f05b63fff5
Merge pull request #1807 from anomalyco/automation/sync-models-aggregators
...
chore(sync): update aggregator model catalogs
2026-05-20 10:34:30 -05:00
Adrien Jussak
e277d60236
Add Mistral Medium 3.5 128B to Scaleway
2026-05-20 16:14:37 +02:00
github-actions[bot]
b01277a737
chore(sync): update aggregator model catalogs
2026-05-20 09:25:37 +00:00
Frank
a5da5aa429
update zen models
2026-05-20 04:14:28 -04:00
Aiden Cline
5ceac8a58b
updates
2026-05-19 23:43:01 -05:00
Aiden Cline
11e1d5623a
Merge pull request #1806 from Cahl-Dee/grid-model-updates-2026-05
...
Grid model updates 2026-05
2026-05-19 23:42:26 -05:00
Aiden Cline
f107afc57c
sync
2026-05-19 22:56:50 -05:00
Carl DiClementi
e789d7c1f2
Merge branch 'anomalyco:dev' into grid-model-updates-2026-05
2026-05-19 16:06:23 -05:00
Cahl-Dee
899668ad49
added new code and agent models and updated existing text models
2026-05-19 16:04:58 -05:00
Aiden Cline
8c677f0134
sync google models
2026-05-19 15:58:06 -05:00
Aiden Cline
462c7877d9
add gemini 3.5 flash
2026-05-19 15:43:38 -05:00
Aiden Cline
55871f9dca
Merge pull request #1804 from Adanlink/dev
...
feat: add deepseek-v4-flash to the fireworks-ai provider
2026-05-19 15:29:33 -05:00
Aiden Cline
4f7194a3c8
test
2026-05-19 15:09:11 -05:00
Aiden Cline
f65f0148da
add sync guide
2026-05-19 15:08:44 -05:00
Adán
14952f8855
Rename deepseek-v4-flash to deepseek-v4-flash.toml
2026-05-19 18:56:43 +02:00
Adán
d7c6d3ad12
Add deepseek-v4-flash model configuration
2026-05-19 18:53:48 +02:00
Aiden Cline
356bc79d08
Merge pull request #1637 from elvexai/fix/amazon-bedrock-kimi-token-limits
...
fix: Token limits for Amazon Bedrock Kimi K2 models
2026-05-19 09:42:30 -05:00
Aiden Cline
a89b1ed726
Merge pull request #1801 from bas3line/sync-routing-run-models
...
Sync routing.run model catalog
2026-05-19 09:41:38 -05:00
bas3line
a998576773
fix(routing-run): match live model metadata
2026-05-19 10:38:58 +05:30
bas3line
fbe842bbea
fix(routing-run): expose reasoning metadata
2026-05-19 08:39:51 +05:30
bas3line
6c0c3d1b10
fix(routing-run): sync model catalog
2026-05-19 07:37:53 +05:30
Aiden Cline
db0a7cf611
Merge pull request #1798 from anomalyco/rework-sync-logic
...
sync: centralize aggregator model updates
2026-05-18 20:12:50 -05:00
Aiden Cline
d775e37e3b
Merge pull request #1800 from jerome-benoit/feat/sap-ai-core-gpt-5.4
...
feat(sap-ai-core): add GPT-5.4
2026-05-18 20:12:21 -05:00
Jérôme Benoit
36753063d9
feat(sap-ai-core): add GPT-5.4
...
Add gpt-5.4 with availability date from official SAP source.
Drop [[cost.tiers]] from gemini-2.5-pro pending SAP-side tiered
pricing confirmation; sap-ai-core now declares no per-model tiers
(SAP Note 3437766 is login-gated and authoritative for capacity
unit conversion rates).
2026-05-19 02:58:29 +02:00
Aiden Cline
5ee955297a
sync: drop vercel catalog updates
2026-05-18 19:07:14 -05:00
Aiden Cline
1b77511903
Merge pull request #1799 from vglafirov/add-gitlab-gpt-5-5
...
feat(gitlab): add Agentic Chat (GPT-5.5) model
2026-05-18 15:24:18 -05:00
Aiden Cline
8896ead7bf
sync: fix vercel pricing tiers
2026-05-18 14:52:40 -05:00
Vladimir Glafirov
eb96594d47
feat(gitlab): add Agentic Chat (GPT-5.5) model
...
Adds duo-chat-gpt-5-5 to the GitLab provider. The GitLab AI Gateway
proxies this model to OpenAI's gpt-5.5-2026-04-23 backend with a
1.05M token context window (922k input + 128k output).
Source: gitlab-org/modelops/applied-ml/code-suggestions/ai-assist
models.yml (gpt_5_5 entry with proxy_provider: openai).
The gitlab-ai-provider npm package exposes this model id starting in
v6.7.0.
2026-05-18 20:44:34 +02:00
Aiden Cline
327332efe3
Merge pull request #1794 from bas3line/add-routing-run-provider
...
Add routing.run provider
2026-05-18 12:32:34 -05:00
Aiden Cline
5020951745
sync: refresh openrouter after dev merge
2026-05-18 12:30:29 -05:00
Aiden Cline
cb6f97774e
Merge remote-tracking branch 'origin/dev' into rework-sync-logic
2026-05-18 12:29:28 -05:00
Aiden Cline
7f8b493b0c
Merge pull request #1795 from delafthi/delafthi/lxxqxzktnozv
...
fix(providers/novita-ai): use lowercase model names
2026-05-18 12:28:32 -05:00
Aiden Cline
d65a862533
sync: centralize aggregator model updates
2026-05-18 12:12:15 -05:00
arnavchachra
9420048dfe
fix crof model limits and reasoning flag to match Crof API
2026-05-18 21:51:19 +05:30
arnavchachra
8db6c27634
add crof provider with 21 models
2026-05-18 21:40:11 +05:30
Victor Navarro
8e710e19ea
bring back old bick-pickle
...
Added interleaved section with reasoning_content field and removed provider section.
2026-05-18 11:44:14 +02:00
Frank
36c6896e97
update zen models
2026-05-17 22:58:06 -04:00
Aiden Cline
a8be548a5d
Merge pull request #1416 from Luew2/add-lilac-provider
...
Add Lilac provider
2026-05-17 19:23:07 -05:00
Luew2
8c2fae4ab0
Keep exact Lilac Gemma model name
2026-05-17 17:19:53 -07:00
Luew2
5e7fad350d
Align Lilac Gemma display name
2026-05-17 17:19:04 -07:00
Luew2
feb85ef2c9
Align Lilac provider with registry conventions
2026-05-17 17:15:40 -07:00
Luew2
d4161ebf24
Follow models.dev conventions for Lilac provider
2026-05-17 17:11:03 -07:00
Luew2
b91ab02e2b
Add Lilac MiniMax M2.7 model
2026-05-17 17:07:44 -07:00
Luew2
dd09d07f75
Update Lilac Kimi model to K2.6
2026-05-17 17:07:44 -07:00
Luew2
4515f85d47
Add Lilac cache pricing
2026-05-17 17:07:44 -07:00
Luew2
97572240e1
Add Gemma 4 31B IT model
...
Adds google/gemma-4-31b-it to the Lilac provider ($0.11/M input,
$0.35/M output, 262K context, native multimodal with image/video).
Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com >
2026-05-17 17:07:44 -07:00
Luew2
80cc04e9e5
Fix Kimi K2.5 output limit to 262,144 tokens
2026-05-17 17:07:44 -07:00
Luew2
d532ebb89b
Use purple gradient for Lilac logo (brand colors #6451dc → #b6a6f9)
2026-05-17 17:07:44 -07:00
Luew2
9bae3e887f
Replace placeholder logo with Lilac icon mark (currentColor)
2026-05-17 17:07:44 -07:00
Luew2
0be69bf872
Add Lilac provider
...
Add Lilac as an OpenAI-compatible provider serving:
- z-ai/glm-5.1: Z.ai's flagship agentic model (754B MoE, 202.8K context)
- moonshotai/kimi-k2.5: Moonshot AI's multimodal reasoning model (1T MoE, 262K context)
API: https://api.getlilac.com/v1
Docs: https://docs.getlilac.com
2026-05-17 17:07:44 -07:00
Thierry Delafontaine
0ed38cbecf
fix(providers/novita-ai): use lowercase model names
...
Mixed case naming causes conflicts on case-insensitive filesystems like macOS.
2026-05-17 21:17:46 +02:00
bas3line
e65703382a
feat: add routing.run provider
2026-05-17 20:27:56 +05:30
Aiden Cline
748754c99b
Merge pull request #1792 from monotykamary/fix/neuralwatt-context-limits
...
fix(neuralwatt): sync context window and output limits with upstream API
2026-05-16 13:14:05 -05:00
Tom X Nguyen
ec9c12d0fc
fix(neuralwatt): sync context window and output limits with upstream API
...
Updates all 14 neuralwatt model TOML files with corrected context window
and max output token values as reported by the Neuralwatt API:
- Devstral-Small-2-24B-Instruct-2512: 262,144 -> 262,128
- GLM-5/GLM-5.1 variants: 200,000 -> 202,736
- GPT-OSS-20B: 16,384 -> 16,368
- Kimi-K2.5/K2.6 variants: 262,144 -> 262,128
- MiniMax-M2.5: 196,608 -> 196,592
- Qwen3.5-397B variants: 262,144 -> 262,128
- Qwen3.6-35B variants: 131,072 -> 131,056
Also fixes the README: moves kimi-k2.6-fast from 'Reasoning Models' to
'Fast Variants' and removes incorrect claim that fast variants support
reasoning.
2026-05-16 23:16:34 +07:00
Aiden Cline
ac81822c89
Merge pull request #1777 from berget-ai/feat/berget-kimi-k2.6
...
feat: add Kimi K2.6 to berget.ai
2026-05-16 06:09:00 -05:00
Aiden Cline
d32ed764bd
Merge pull request #1791 from anomalyco/automation/sync-openrouter-models
...
Sync OpenRouter models
2026-05-16 06:08:09 -05:00
Christian Landgren
45fb951c42
feat: add Kimi K2.6 to berget.ai
...
Add Moonshot AI Kimi K2.6 model to berget.ai provider catalog.
- 262K context window
- 16K output tokens
- Text input/output
- Supports: reasoning, structured output, tool calling
- Pricing: /bin/zsh.83/M input, .85/M output (EUR-based)
- Open weights
2026-05-16 12:46:35 +02:00
github-actions[bot]
260d79b2d5
Sync OpenRouter models
2026-05-16 08:54:41 +00:00
Aiden Cline
746b9caf79
Merge pull request #1785 from jerome-benoit/feat/sap-ai-core-opus-4-7
...
feat(sap-ai-core): add Claude Opus 4.7 and sync model specs
2026-05-15 23:24:47 -05:00
Aiden Cline
8362b55503
Merge pull request #1786 from Ardakilic/chore/kilo-sync-20260516
...
providers(kilo): sync upstream
2026-05-15 23:24:34 -05:00
Aiden Cline
dde3953a9f
Merge pull request #1787 from Suat-B/codex/xpersona-www-api-url
...
Fix Xpersona API base URL
2026-05-15 23:24:06 -05:00
Aiden Cline
0a1695212c
Merge pull request #1788 from Jaaneek/xai-may-15-2026-retirement
...
xai: drop models retired May 15, 2026 + add Grok Imagine models
2026-05-15 23:23:52 -05:00
Jaaneek
89fbb6bb69
xai: drop models retired May 15, 2026 + add Grok Imagine models
2026-05-16 01:57:15 +01:00
SuatB
005fe0fb5a
Fix Xpersona provider API URL
2026-05-15 18:37:41 -05:00
Frank
e283875ce7
update zen models
2026-05-15 17:24:53 -04:00
Arda Kilicdagi
3598019251
providers(kilo): sync upstream
2026-05-16 01:11:25 +04:00
Jérôme Benoit
e9ad8b0a3f
feat(sap-ai-core): add Claude Opus 4.7 and sync model specs
2026-05-15 22:16:07 +02:00
Aiden Cline
0ee78eeda5
sync: openrouter models
2026-05-15 10:20:36 -05:00
Aiden Cline
0a5b33e518
Merge pull request #1778 from zhenjunchen-png/add-orcarouter
...
feat: add OrcaRouter as a new provider
2026-05-15 10:16:29 -05:00
Aiden Cline
22416dda64
Merge pull request #1783 from anomalyco/sync-openrouter
...
add sync script for openrouter, sync openrouter models
2026-05-15 10:12:02 -05:00
Aiden Cline
eba2702e3a
fix: families
2026-05-15 10:03:39 -05:00
Aiden Cline
c2c5cc8f21
add sync script for openrouter, sync openrouter models
2026-05-15 10:00:32 -05:00
zhenjun.chen
022b1b9946
feat(orcarouter): expand to 80 chat models and add brand logo
...
Adds 55 additional upstream-mirrored models alongside the existing 25,
covering the full OrcaRouter chat catalog as exposed by
https://www.orcarouter.ai/api/pricing (text-only chat — TTS, embeddings,
video, and image generation are filtered out).
Per-namespace upstream mappings used by [extends]:
OrcaRouter ns -> models.dev provider
---------------- + ---------------
openai -> openai
anthropic -> anthropic (dot version -> dash, e.g. opus-4.7 -> opus-4-7)
google -> google
deepseek -> deepseek
qwen -> alibaba
grok -> xai
kimi -> moonshotai
minimax -> minimax (minimax-m2.7 -> MiniMax-M2.7)
z-ai -> zai
OrcaRouter-specific aliases (dated snapshots like gpt-5-2025-08-07,
search-preview variants, qwen3-vl-* visual variants) are excluded from
v1 because their upstream canonical files do not yet exist in models.dev.
Also adds providers/orcarouter/logo.svg.
2026-05-15 14:48:31 +08:00
Aiden Cline
8269e04222
Merge pull request #1782 from Suat-B/codex/xpersona-limits-logo
...
Update Xpersona limits, cutoff, and logo
2026-05-15 00:29:32 -05:00
Aiden Cline
a75cf2ed1c
Merge pull request #1552 from aredridel/as/add-umans
...
feat(models): add umans.ai coding plan
2026-05-14 22:40:39 -05:00
Aria Stewart
7b00aafa79
feat(models): umans.ai definitions
2026-05-14 23:39:30 -04:00
Suat-B
e54e0fc8b5
Update Xpersona limits, cutoff, and logo
2026-05-15 03:22:16 +00:00
Aiden Cline
38611e75fa
Merge pull request #1781 from dpuyosa/feat/add-venice-claude-opus-4-7-fast-model
...
Venice: Add Claude Opus 4.7 Fast model
2026-05-14 22:08:14 -05:00
dpuyosa
e62c1e973e
[venice] Add Claude Opus 4.7 Fast model
...
- New pricing with 36/180 input/output per million tokens
- 1M context window with 128K output limit
- Text+image input, text-only output
2026-05-15 01:32:24 +02:00
Aiden Cline
c2b3c601e4
Merge pull request #1724 from isaachuangGMICLOUD/feat/add-gmicloud-provider
...
providers(gmicloud): add GMI Cloud provider
2026-05-14 17:26:50 -05:00
Aiden Cline
9ff1d36a21
Merge pull request #1476 from Vect0rM/feat/add-atomic-chat-provider
...
feat: add Atomic Chat provider
2026-05-14 17:18:21 -05:00
Aiden Cline
e0f4042ad1
Merge pull request #1776 from kapelame/docs/minimax-token-plan-rename
...
providers(minimax): rename Coding Plan → Token Plan in display labels
2026-05-14 17:16:28 -05:00
Frank
14736ba4b6
update zen models
2026-05-14 17:15:53 -04:00
Aiden Cline
9351d68731
Merge pull request #1772 from nearai/add-nearai
...
Add NEAR AI Cloud provider
2026-05-14 10:30:28 -05:00
Aiden Cline
7a5a1d2aff
Merge pull request #1779 from NameIsHiki/siliconflow-deepseek-v4
...
feat(siliconflow): add DeepSeek v4 models
2026-05-14 10:29:52 -05:00
zhenjun.chen
699284ce91
chore(orcarouter): drop oversize logo, fall back to models.dev default
...
The previously committed logo is ~100KB; existing wrapper-provider logos
(openrouter, llmgateway, kilo, aihubmix, ambient) are all 0.3-6KB and use
`currentColor`. Falling back to the default logo per README:
> If we don't have a provider's logo, a default logo is served instead.
A properly-sized currentColor logo will follow in a separate PR.
2026-05-14 21:37:48 +08:00
Hiki
21ce5c3ac8
Create deepseek-v4-flash.toml
2026-05-14 15:26:37 +02:00
Hiki
3485cf52d0
Create deepseek-v4-pro.toml
2026-05-14 15:22:53 +02:00
Frank
99e8f25c78
update zen models
2026-05-14 08:59:42 -04:00
zhenjun.chen
7102978cb4
feat: add OrcaRouter provider
...
OrcaRouter is an OpenAI-compatible meta-router aggregating 150+ LLMs
(OpenAI, Anthropic, Google, xAI, DeepSeek, Qwen, Kimi, MiniMax, ...)
behind a single API key, with a virtual orcarouter/auto smart-routing
entry that picks an upstream per request.
This initial scope covers 26 models (1 AUTO router + 25 upstream mirrors
using [extends]). Pricing computed from https://www.orcarouter.ai/api/pricing
on 2026-05-14: input = model_ratio * $2, output = model_ratio *
completion_ratio * $2 (USD per 1M tokens).
Disclosure: I'm an engineer on the OrcaRouter team.
2026-05-14 20:55:47 +08:00
kapelame
530f60c69c
providers(minimax): rename Coding Plan → Token Plan in display labels
...
The product was renamed from "Coding Plan" to "Token Plan" when its
scope expanded beyond coding to cover all MiniMax modalities (text,
speech, video, music, image). Per
https://platform.minimax.io/docs/token-plan/intro :
"Token Plan extends upon our former Coding Plan."
Updates display name and doc URL for the two affected provider
catalog entries. Provider IDs (minimax-coding-plan,
minimax-cn-coding-plan) are intentionally unchanged for backward
compatibility — anyone with these IDs in opencode.json or
elsewhere keeps working. Old /coding-plan/* URLs still 307-redirect
to the new /token-plan/* paths upstream.
Region disambiguation stays as the URL in parens (matching the
existing minimax / minimax-cn naming convention) — no "China" word
added, since the URL already conveys the region cleanly in the
provider picker.
2026-05-14 19:18:03 +08:00
Misha Skvortsov
2415c5be21
fix(atomic-chat): update logo.svg with new design
...
Replaces the existing logo.svg file with an updated design for the Atomic Chat provider. This change enhances the visual branding of the application.
2026-05-14 10:58:36 +03:00
Aiden Cline
85aba468cd
Merge pull request #1767 from Suat-B/codex/xpersona-provider-20260513
...
Add Xpersona provider
2026-05-13 23:20:20 -05:00
Aiden Cline
d1ec1ba777
Merge pull request #1773 from ambient-gregory/dev
...
feat: add Ambient provider with GLM-5.1 and Kimi K2.6
2026-05-13 19:20:20 -05:00
Gregory
0f94bf16ec
fix(ambient): shrink logo display size to match other providers
2026-05-13 19:33:26 -04:00
Aiden Cline
506e8f48a9
Merge pull request #1770 from EriDeLee/dev
...
chore(aihubmix): sync model catalog
2026-05-13 17:43:30 -05:00
Aiden Cline
3480bc5992
Merge pull request #1775 from michaelnchin/fix/amazon-bedrock-gpt-oss-tokens
...
fix: Output tokens for Bedrock GPT-OSS models
2026-05-13 17:36:42 -05:00
Michael Chin
fde97814ef
fix: Output tokens for Bedrock GPT-OSS models
2026-05-13 14:45:17 -07:00
Gregory
ff7eddcb70
feat: add Ambient provider with GLM-5.1 and Kimi K2.6
...
Adds the Ambient inference provider (api.ambient.xyz) with an initial
catalog of GLM-5.1 and Kimi K2.6, plus a generator script that pulls
from /v1/models so pricing and limits stay in sync with the upstream API.
Run `bun run ambient:generate` to refresh model TOMLs.
2026-05-13 11:30:45 -04:00
Evrard-Nil Daillet
5cbab85b8d
Add nearai logo.svg from cloud.near.ai favicon
2026-05-13 16:32:08 +02:00
Evrard-Nil Daillet
6f9820de9f
Add NEAR AI Cloud provider
...
Adds nearai as an OpenAI-compatible provider at https://cloud-api.near.ai/v1
serving 33 models. First-party mirrors (anthropic/openai/google) use `extends`;
NEAR-hosted open-weight models (Qwen, GLM-5.1-FP8, gpt-oss, whisper, FLUX) have
full definitions.
Pricing and context limits sourced from cloud-api.near.ai/v1/models.
2026-05-13 16:32:08 +02:00
EriDeLee
cdfb429098
chore(aihubmix): sync model catalog
2026-05-13 21:55:32 +08:00
Victor Navarro
1c2546af8a
perf: virtualize models table and other improvements
2026-05-13 12:29:27 +02:00
vimtor
2288a1626b
trim search index to essential fields and remove dead code
2026-05-13 12:21:52 +02:00
vimtor
b3ecfc3d70
improve row scanning
2026-05-13 11:20:42 +02:00
Suat-B
30b3e677fd
Add Xpersona provider
2026-05-13 00:39:09 -05:00
Suat-B
3b37eee86e
Add Xpersona provider
2026-05-13 00:39:08 -05:00
Suat-B
71f069670e
Add Xpersona provider
2026-05-13 00:39:07 -05:00
Aiden Cline
f401672689
Merge pull request #1766 from michaelnchin/fix/amazon-bedrock-structured-output-05122026-2
...
fix: add structured_output=True for more supported Bedrock models
2026-05-12 23:24:39 -05:00
Michael Chin
2a0d86a034
update structured_output for more Bedrock models
2026-05-12 20:52:36 -07:00
Aiden Cline
d9439cdf2f
Merge pull request #1762 from zxyaction/feat/add-auriko-provider
...
feat: add Auriko provider with 15 models
2026-05-12 22:16:02 -05:00
Aiden Cline
3d443d568d
Merge pull request #1765 from michaelnchin/fix/amazon-bedrock-structured-output-05122026
...
fix: update structured_output for Bedrock Claude 4.x models
2026-05-12 22:15:51 -05:00
Michael Chin
a76c8fe9dd
fix: update structured_output for Bedrock Claude 4.x models
2026-05-12 20:07:12 -07:00
Aiden Cline
d08e8d6cc1
Merge pull request #1763 from Tavernari/feat/add-claudinio-provider
...
feat: add Claudinio provider
2026-05-12 19:13:54 -05:00
Victor Carvalho Tavernari
4c06e44047
fix: use currentColor in logo SVG per contributing guidelines
2026-05-12 23:59:03 +01:00
Aiden Cline
5e344ded49
Merge pull request #1755 from anomalyco/correct-context-tracking
...
feat: add new context pricing tiers
2026-05-12 17:40:48 -05:00
Aiden Cline
458b7f4d1a
use Venice context tier thresholds
2026-05-12 17:39:40 -05:00
Aiden Cline
baf4432140
Merge pull request #1759 from NameIsHiki/deepinfra-xiaomi-mimo-models
...
feat(deepinfra): add Xiaomi MiMo v2.5 and v2.5 Pro
2026-05-12 17:10:59 -05:00
Aiden Cline
bbf72ea4e0
Merge pull request #1764 from anomalyco/add-anthropic-opus-4-7-fast-mode
...
Add fast mode for Anthropic Opus 4.7
2026-05-12 17:10:49 -05:00
Aiden Cline
8f9adc7567
fix generated tier change detection
2026-05-12 17:01:35 -05:00
Aiden Cline
addaf1c036
add fast mode for anthropic opus 4.7
2026-05-12 17:01:05 -05:00
Victor Carvalho Tavernari
55d16a58b6
feat: add claudinio provider (OpenAI-compatible, 256K ctx, $0.50/$2.00 per MTok)
2026-05-12 22:33:45 +01:00
Hiki
72a4deab66
Update mimo-v2.5.toml
2026-05-12 23:23:53 +02:00
Hiki
8dd829a187
Update mimo-v2.5-pro.toml
2026-05-12 23:23:01 +02:00
Aiden Cline
a671cc05d5
align cost tiers with model schema
2026-05-12 16:01:49 -05:00
Aiden Cline
656c6f08a7
Merge pull request #1761 from Sewer56/deprecate-wafer-models
...
providers/wafer.ai: Remove DeepSeek-V4-Pro and MiniMax-M2.7 models
2026-05-12 15:51:10 -05:00
Aiden Cline
8979741a32
Merge pull request #1760 from Ardakilic/fix/kilo/kimik26
...
Fix: Kimi k2.6 definition on Kilo Gateway
2026-05-12 15:50:53 -05:00
Frank
82851b9a3d
update zen models
2026-05-12 16:44:41 -04:00
zxy_action
ae511892d7
feat: add Auriko provider with 15 models
...
All models use [extends] to inherit from canonical definitions,
overriding only Auriko-specific pricing. Omits remove cost tiers
and features Auriko doesn't carry.
Models: claude-opus-4-{6,7}, claude-sonnet-4-6, deepseek-v4-{pro,flash},
gemini-{2.5-pro,2.5-flash,3.1-pro-preview}, grok-4.3, kimi-k2.{5,6},
minimax-m2-7{,-highspeed}, glm-5.1, qwen-3.6-plus
2026-05-12 13:34:34 -07:00
Sewer56
9312418242
Changed: Remove DSv4 Pro & MiniMax M2.7 from models.dev
2026-05-12 20:16:08 +01:00
vimtor
9828a0177d
improve empty row
2026-05-12 19:49:03 +02:00
Arda Kılıçdağı
122627a852
fix: Kimi k2.6 definition on Kilo Gateway
2026-05-12 20:35:58 +03:00
vimtor
4dee8d0d34
minor improvements
2026-05-12 19:10:21 +02:00
Hiki
6883e793ce
Create mimo-v2.5-pro.toml
2026-05-12 18:22:41 +02:00
Hiki
e69064709b
Update mimo-v2.5.toml
2026-05-12 18:19:52 +02:00
Hiki
bd8e582b96
Create mimo-v2.5.toml
2026-05-12 18:03:09 +02:00
Shoubhit Dash
21945db90f
Merge pull request #1662 from anomalyco/nxl/add-sarvam-provider
...
provider(sarvam): add chat models
2026-05-12 13:31:59 +05:30
Aiden Cline
1771e02be8
Merge pull request #1660 from Alex-wuhu/feat/novita-ai-sync-models
...
provider(novita-ai): sync latest models
2026-05-11 23:15:44 -05:00
Aiden Cline
bb08fc26e9
Merge pull request #1706 from rohita5l/rohit/addDatabricks
...
Add Databricks as a provider
2026-05-11 19:28:34 -05:00
Aiden Cline
2e015de42d
preserve generated tier thresholds
2026-05-11 17:07:19 -05:00
Rohit Agrawal
914a3d9d18
fix: restore bun.lock to use default registry instead of Databricks npm proxy
...
Co-Authored-By: Claude Sonnet 4.6 <noreply@anthropic.com >
2026-05-11 17:54:26 -04:00
Rohit Agrawal
bab01dd9ab
refactor: move databricks generate to packages/core/script following repo conventions
...
Addresses review feedback by removing AI SDK dependencies from package.json
and aligning with the Vercel/Helicone/Wandb pattern.
- Move generate-databricks.ts to packages/core/script/
- Add databricks:generate to root scripts
- Remove smoke test and runtime filtering (catalog should reflect what the
upstream API exposes; AI SDK compatibility is a downstream concern)
- Add --dry-run and --new-only flags
- Merge with existing TOMLs instead of nuking them; warn about orphans
- Restore databricks-gemini-3-pro and databricks-gemini-3-1-pro
- Drop @ai-sdk/openai-compatible, ai, zod from root dependencies
Co-Authored-By: Claude Sonnet 4.6 <noreply@anthropic.com >
2026-05-11 17:54:04 -04:00
Rohit Agrawal
bdaae956af
feat: add AI SDK compatibility test to generate script, remove incompatible models
...
Generate script now smoke-tests each model with streamText after writing TOMLs
and removes any that return empty responses (incompatible with @ai-sdk/openai-compatible).
Removes databricks-gemini-3-pro and databricks-gemini-3-1-pro which return content
as array with thoughtSignature that the AI SDK cannot parse.
Also adds test-databricks.ts for standalone smoke testing.
Co-Authored-By: Claude Sonnet 4.6 <noreply@anthropic.com >
2026-05-11 17:53:44 -04:00
Rohit Agrawal
6f118145c0
fix: inline gpt-oss model metadata instead of invalid extends path
...
openrouter models in subdirectories can't use extends (schema requires
provider/model format); resolve() now inlines the source TOML content directly.
Co-Authored-By: Claude Sonnet 4.6 <noreply@anthropic.com >
2026-05-11 17:53:44 -04:00
Rohit Agrawal
386eaed119
add databricks
2026-05-11 17:53:44 -04:00
Aiden Cline
151e9c9071
fix tiered cost generation
2026-05-11 16:43:32 -05:00
Isaac
ba99f1edce
Add GMI Cloud GLM models
2026-05-11 14:40:03 -07:00
Aiden Cline
b96b074a3b
fix long-context cost tier omissions
2026-05-11 16:31:07 -05:00
Aiden Cline
2593e131a1
wip
2026-05-11 16:11:19 -05:00
Aiden Cline
4aebbe5ca3
Merge pull request #1754 from BruceMacD/brucemacd/fix-ollama-kimi-k2-6-model-id
...
fix ollama cloud kimi k2.6 model id
2026-05-11 14:57:03 -05:00
Bruce MacDonald
b98495e9c8
fix ollama cloud kimi k2.6 model id
2026-05-11 12:40:17 -07:00
Frank
bc95b42ccd
update zen models
2026-05-11 11:26:08 -04:00
Aiden Cline
6139fb8c69
Merge pull request #1598 from mugnimaestra/feat/chutes-generate-script
...
feat(chutes): add API-driven model generator script
2026-05-11 09:30:38 -05:00
Aiden Cline
3070758007
Merge pull request #1678 from 5kahoisaac/chore/nvidia-models
...
Sync NVIDIA endpoint model catalog
2026-05-11 09:29:54 -05:00
Frank
5525e83de4
update zen models
2026-05-11 10:00:09 -04:00
Frank
359fd879b8
update zen models
2026-05-10 03:54:03 -04:00
Frank
01b5a1a656
update zen models
2026-05-10 02:52:44 -04:00
Frank
b1958be099
update zen models
2026-05-10 02:42:44 -04:00
Aiden Cline
08aa068523
Temporarily remove kiro provider and models
2026-05-10 01:19:47 -05:00
Aiden Cline
f31ad0b02f
Merge pull request #1738 from mattiacerutti/chore/remove-gh-copilot-deprecated
...
chore(copilot): mark deprecated models
2026-05-09 15:42:19 -05:00
Aiden Cline
585aa7fa1b
Merge pull request #1741 from EriDeLee/dev
...
Update aihubmix models
2026-05-09 15:42:03 -05:00
Aiden Cline
c42a327b3e
Merge pull request #1745 from mads-digitial-solutions/patch-1
...
Update Google provider docs url from pricing page to models page
2026-05-09 15:41:51 -05:00
mads-digitial-solutions
92ebbfb5c4
Update provider.toml
...
Update Google provider docs URL from the pricing page to the models page
2026-05-09 19:49:19 +01:00
Aiden Cline
535fe8c971
Merge pull request #1744 from OpeOginni/fix/bedrock-model-ids
...
chore(bedrock): Getting rid of legacy Amazon Bedrock model offerings
2026-05-09 13:48:52 -05:00
OpeOginni
a3b4bfc16c
fix(bedrock): remove uneeded model configurations
2026-05-09 20:35:34 +02:00
OpeOginni
d0fcd6f11f
fix(bedrock): align models with current docs
2026-05-09 20:29:11 +02:00
OpeOginni
e55cd54218
fix(bedrock): remove legacy model entries
2026-05-09 20:21:33 +02:00
OpeOginni
0d73b82b9f
fix(bedrock): restore regional model IDs
2026-05-09 20:18:59 +02:00
Aiden Cline
83c7e2b63f
Merge pull request #1742 from Adam8234/add-firepass-provider
...
feat: add Fireworks (Firepass) provider
2026-05-09 12:45:09 -05:00
Adam
83ae4cf813
feat: add Fireworks (Firepass) provider
...
Adds the Fireworks AI Firepass subscription provider.
- Provider uses a dedicated FIREPASS_API_KEY
- Uses @ai-sdk/openai-compatible SDK
- Includes Kimi K2.6 Turbo (accounts/fireworks/routers/kimi-k2p6-turbo)
- Zero per-token cost since covered by subscription
2026-05-09 12:22:13 -05:00
EriDeLee
77eae6eef7
Update aihubmix models
2026-05-09 20:53:57 +08:00
Aiden Cline
8cbf6ed10e
Merge pull request #1736 from vercel/update-vercel-models-1778258030
...
Update Vercel models
2026-05-08 21:58:40 -05:00
Mattia Cerutti
06d87e4411
chore(copilot): remove deprecated models
2026-05-09 00:19:50 +02:00
Frank
2cb3832618
update zen models
2026-05-08 17:11:05 -04:00
vimtor
950a0446d4
bring back svg loading
2026-05-08 20:18:12 +02:00
vimtor
41fbfb1a17
change sst mention
2026-05-08 20:13:09 +02:00
vimtor
cf5045a90b
move copy button next to name
2026-05-08 20:12:39 +02:00
vimtor
ef739220de
lock virtualized table column widths to prevent scroll jitter
2026-05-08 19:07:25 +02:00
github-actions[bot]
df960d1a90
chore(vercel): update Vercel model definitions
...
Auto-generated by weekly workflow from Vercel AI Gateway API.
Co-Authored-By: github-actions[bot] <41898282+github-actions[bot]@users.noreply.github.com>
2026-05-08 16:33:52 +00:00
vimtor
7294efee15
move row-render.ts to shared.ts
2026-05-08 18:30:14 +02:00
vimtor
1e537147eb
remove table row tuple optimization
2026-05-08 18:24:18 +02:00
Aiden Cline
8f2f83ef61
Merge pull request #1735 from slpdy/dev
...
Create DeepSeek-V4-Pro.toml
2026-05-08 10:59:33 -05:00
Aiden Cline
b133426465
Merge pull request #1734 from oskarkocol/chore/update-novita-deepseek-prices
...
chore: update novita deepseek-v4-pro prices
2026-05-08 10:59:25 -05:00
Aiden Cline
dafff5a770
Merge pull request #1730 from smakosh/add-llmgateway-models
...
Add new LLM Gateway text models (gpt-5.5, grok-4-3, gemini-3.1-flash-lite, qwen3.6, MiMo v2)
2026-05-08 10:58:20 -05:00
smakosh
dd894f077f
Merge remote-tracking branch 'upstream/dev' into add-llmgateway-models
...
# Conflicts:
# providers/google/models/gemini-3.1-flash-lite.toml
2026-05-08 17:46:52 +02:00
smakosh
91590874e7
Revert "fix(models): use canonical entries for qwen3.6-max-preview and grok-4.3"
...
This reverts commit 70ac6fccda .
2026-05-08 17:43:07 +02:00
vimtor
e9f4cecc54
extract shared row rendering module
2026-05-08 17:29:04 +02:00
vimtor
fb1ac09883
virtualize model table
2026-05-08 17:07:28 +02:00
Jj
a436236146
Create DeepSeek-V4-Pro.toml
...
Added DeepSeek-v4-Pro model to Nebius provider
2026-05-08 11:03:16 +01:00
oskar
1415b4be97
chore: update novita deepseek prices
2026-05-08 14:21:27 +07:00
Aiden Cline
1437da86e7
Merge pull request #1685 from 8023/dev
...
Add kimi-k2.6/deepseek-v4 model and EmpirioLabs AI integration for poe.com
2026-05-07 22:38:51 -05:00
Aiden Cline
e7d57885d1
Merge pull request #1733 from chl-0537/feature/add-tencent
...
add model by openrouter
2026-05-07 22:21:07 -05:00
Aiden Cline
8157916515
fix: attachment
2026-05-07 22:12:11 -05:00
Aiden Cline
2e5b87a9c2
Merge pull request #1731 from mikeyp/chore/update-digitalocean-models
...
Add script to generate/update Digitalocean models
2026-05-07 16:42:09 -05:00
Aiden Cline
92e19432d3
add google gemini 3.1 flash lite
2026-05-07 15:44:25 -05:00
smakosh
70ac6fccda
fix(models): use canonical entries for qwen3.6-max-preview and grok-4.3
...
Apply the canonical TOML provided by the LLM Gateway team for the
Qwen3.6 Max Preview and Grok 4.3 parent definitions, replacing the
upstream-merged variants whose dates and pricing did not match.
Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com >
2026-05-07 22:34:01 +02:00
smakosh
dc3283417d
Merge remote-tracking branch 'upstream/dev' into add-llmgateway-models
...
# Conflicts:
# providers/alibaba/models/qwen3.6-max-preview.toml
# providers/llmgateway/models/qwen3.6-max-preview.toml
2026-05-07 22:28:23 +02:00
smakosh
34fd6673e5
chore(llmgateway): add new text models from llmgateway catalog
...
Add gemini-3.1-flash-lite, grok-4-3, gpt-5.5, gpt-5.5-pro, qwen3.6
and MiMo v2 models that exist in llmgateway.io but were missing
from models.dev. Adds parent definitions for grok-4-3,
gemini-3.1-flash-lite, and qwen3.6-max-preview where they did not
already exist.
Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com >
2026-05-07 22:20:52 +02:00
Mike Prasuhn
5df9293314
Add script to generate/update Digitalocean models
2026-05-07 15:28:07 -04:00
Aiden Cline
a81a9559d7
Merge pull request #1726 from Ardakilic/chore/sync-kilo-models
...
Chore: Sync Kilo models with upstream gateway
2026-05-07 13:04:51 -05:00
Aiden Cline
4197bf57a6
Merge pull request #1725 from dpuyosa/chore/venice-grok-costs
...
Venice: Update grok-4-20 pricing
2026-05-07 13:04:26 -05:00
Aiden Cline
d0ac772507
Merge pull request #1717 from juls0730/refactor/mimo-extends/token-plan
...
refactor(xiaomi-token-plan): extends xiaomi base provider for xiaomi token plans
2026-05-07 13:04:05 -05:00
Aiden Cline
1afaa053e7
Merge pull request #1727 from sergeykonkin/update-nebius-models-2026-05
...
chore: update nebius provider models
2026-05-07 13:03:46 -05:00
Aiden Cline
9b77ce1c9e
Merge pull request #1729 from Sewer56/add-minimax-m27-wafer
...
Added: MiniMax-M2.7 model for wafer.ai
2026-05-07 12:59:49 -05:00
Aiden Cline
68dc6d1820
Merge pull request #1728 from arafatkatze/codex/openrouter-qwen-cache-pricing
...
Add OpenRouter Qwen cache pricing
2026-05-07 12:59:29 -05:00
Arafatkatze
a1eb5eece3
Add OpenRouter Qwen cache pricing
2026-05-07 10:29:48 -07:00
Sewer56
f7ec2c517f
Added: wafer.ai/MiniMax-M2.7 model
2026-05-07 18:12:50 +01:00
Sergey Konkin
f98e8ec793
chore: update nebius provider models
2026-05-07 15:37:18 +02:00
Arda Kilicdagi
d210066793
chore: sync kilo gw models
2026-05-07 14:17:15 +03:00
Frank
06908cbf36
update zen models
2026-05-07 04:47:22 -04:00
dpuyosa
b0614d2088
[venice] Update grok-4-20 pricing
...
- Lower grok-4-20 and multi-agent input/output costs to latest Venice pricing
2026-05-07 10:29:48 +02:00
mickalchen
32c1c45c52
add openrouter model
2026-05-07 11:22:11 +08:00
Frank
7c033f27e6
update zen models
2026-05-06 23:00:15 -04:00
mickalchen
d648e63499
Merge remote-tracking branch 'origin/dev' into feature/add-tencent
2026-05-07 10:54:56 +08:00
Zoe
1353f965b9
refactor(xiaomi-token-plan): extends xiaomi base provider for xiaomi token plan
2026-05-06 18:53:49 -05:00
Isaac Huang
175082d43f
Add GMI Cloud provider
2026-05-06 15:12:07 -07:00
Aiden Cline
bba0a9c3f4
Merge pull request #1312 from NachoFLizaur/feat/kiro-provider
...
feat: add Kiro provider with 12 models
2026-05-06 12:16:47 -05:00
Frank
4bdb195178
update zen models
2026-05-06 12:56:58 -04:00
Aiden Cline
12706b7652
Merge pull request #1716 from juls0730/refactor/mimo-extends/zenmux
...
refactor(zenmux): extends mimo models from xiaomi provider
2026-05-06 10:45:01 -05:00
Aiden Cline
d8c76d0c67
Merge pull request #1715 from juls0730/refactor/mimo-extends/qiniu-ai
...
refactor(qiniu-ai): extends mimo models from xiaomi provider
2026-05-06 10:30:46 -05:00
Aiden Cline
3963dd13d5
Merge pull request #1714 from juls0730/refactor/mimo-extends/openrouter
...
refactor(openrouter): extends mimo models from xiaomi provider
2026-05-06 10:30:32 -05:00
Aiden Cline
8749a56efa
Merge pull request #1711 from juls0730/refactor/mimo-extends/kilo
...
refactor(kilo/xiaomi): extends mimo models from xiaomi provider
2026-05-06 10:29:31 -05:00
Aiden Cline
25de2ee27d
Merge pull request #1718 from juls0730/refactor/mimo-extends/xiaomi
...
refactor(xiaomi): round xiaomi models to powers of 2 & fix small errors
2026-05-06 10:27:36 -05:00
Aiden Cline
438df7f03c
Merge pull request #1723 from CloudFerro/fix/cloudferro-sherlock/minimax-m2.5
...
fix: cloudferro sherlock - update context values for minimax m2.5
2026-05-06 10:25:54 -05:00
Aiden Cline
7d18558aa3
Merge pull request #1722 from dpuyosa/chore/venice-model-updates
...
Venice: Remove deprecated models, enable reasoning on gpt-oss-120b
2026-05-06 10:25:36 -05:00
Jan Szypulski
c82f736fcc
fix: cloudferro sherlock - update context values for minimax m2.5
2026-05-06 10:52:39 +02:00
dpuyosa
6a436805b3
[venice] Remove deprecated models, enable reasoning on gpt-oss-120b
...
- Remove kimi-k2-thinking, qwen3-coder-480b-a35b-instruct, and venice-uncensored models
- Enable reasoning capability on openai-gpt-oss-120b
2026-05-06 09:44:24 +02:00
Jack
ce7823f073
Merge pull request #1720 from anomalyco/fix/opencode-go-kimi-k26-pricing
...
fix(opencode-go): restore kimi k2.6 pricing
2026-05-06 12:33:24 +08:00
Jack
033efdb7d4
fix(opencode-go): restore kimi k2.6 pricing
2026-05-06 12:32:06 +08:00
Alex-wuhu
0f1855c0a7
provider(novita-ai): use extends for kimi k2.6
2026-05-06 10:41:46 +08:00
Zoe
bd9e0c2677
refactor(xiaomi): round xiaomi models to powers of 2 & fix small errors
2026-05-05 18:40:05 -05:00
Zoe
38545d63a5
refactor(zenmux): extends mimo models from xiaomi provider
2026-05-05 18:17:21 -05:00
Zoe
014be328d6
refactor(qiniu-ai): extends mimo models from xiaomi provider
2026-05-05 18:05:39 -05:00
Zoe
c139147540
refactor(openrouter): extends mimo models from xiaomi provider
2026-05-05 17:49:39 -05:00
Zoe
59cd93cafc
refactor(kilo/xiaomi): extends mimo models from xiaomi provider
2026-05-05 17:13:06 -05:00
Aiden Cline
e91db96d83
Merge pull request #1710 from Spherrrical/add-digitalocean-kimi-2-6
...
feat(digitalocean): add kimi-k2.6 model
2026-05-05 14:07:50 -05:00
Spherrrical
d70dd8dcdc
feat(digitalocean): add kimi-k2.6 model
2026-05-05 12:05:16 -07:00
Frank
b18e681457
update deepseek flash on deepinfra
2026-05-05 14:24:03 -04:00
Aiden Cline
b73a6a2130
Merge pull request #1709 from xiaomochn/fix/xiaomi-mimo-v2.5-modalities
...
fix(xiaomi): swap modalities for MiMo-V2.5 and MiMo-V2.5-Pro
2026-05-05 11:36:43 -05:00
Aiden Cline
153c1cc420
Merge pull request #1614 from Yashwanth-Kumar-26/patch-1
...
Add Qwen 3.6 27B model configuration
2026-05-05 11:22:43 -05:00
Aiden Cline
ca0b30569e
update google vertex to include all anthropic models
2026-05-05 11:12:23 -05:00
xiaomochn
8345bfbd06
fix(xiaomi): correct modalities for MiMo V2.5 models across providers
...
Issues fixed:
1. MiMo-V2.5 and MiMo-V2.5-Pro had their modalities swapped in the
xiaomi canonical source (affects OpenRouter/ZenMux via extends)
2. Removed 'pdf' from V2.5 models — not a supported input modality
3. Fixed vercel provider: V2.5-Pro incorrectly marked as multimodal
4. Fixed opencode-go and vercel V2.5: added missing 'video', removed pdf
Correct modalities:
- MiMo-V2.5: input = ["text", "image", "audio", "video"]
- MiMo-V2.5-Pro: input = ["text"]
Affected providers: xiaomi, opencode-go, vercel, openrouter (extends),
zenmux (extends)
Fixes #1708
2026-05-05 22:46:44 +08:00
Shoubhit Dash
28c0d9ce23
fix(sarvam): correct output limits
2026-05-05 15:30:47 +05:30
Aiden Cline
c16f3da694
Merge pull request #1707 from deathbeam/revert-1664-fix/glm-qwen-ollama-output-limit
...
Revert "fix(ollama): set glm-5.1 and qwen3.5:397b output limits to match context"
2026-05-04 23:51:53 -05:00
Tomas Slusny
6aa6e55d60
fix(ollam): use correct output limit for qwen3.5:397b
...
{"error":"max_tokens (262144) exceeds model's maximum output tokens (65536) for model qwen3.5:397b (ref: 8554a681-e6a8-45d7-9fdd-433785eb6c67)"}
Signed-off-by: Tomas Slusny <slusnucky@gmail.com >
2026-05-05 01:31:20 +02:00
Tomas Slusny
9efaf754a5
Revert "fix(ollama): set glm-5.1 and qwen3.5:397b output limits to match context"
2026-05-05 01:04:11 +02:00
Aiden Cline
104e4bdc1f
Merge pull request #1684 from TheBaconWizard/add-clarifai-kimi-k2.6
...
provider(clarifai): add Kimi-K2.6 (moonshotai/chat-completion)
2026-05-04 12:00:33 -05:00
Aiden Cline
db16c113f7
Merge pull request #1704 from stylings/stylings/add-openrouter-grok-4-3
...
feat: add OpenRouter Grok 4.3
2026-05-04 12:00:08 -05:00
Alex
64a122a13d
fix: update OpenRouter Grok 4.3 file
2026-05-04 12:44:07 -04:00
Alex
fcc6521d1d
fix: simplify OpenRouter Grok 4.3 file
2026-05-04 12:41:56 -04:00
Alex
c623b4a55f
feat: add OpenRouter Grok 4.3
2026-05-04 12:36:33 -04:00
Aiden Cline
1d730fea16
Merge pull request #1696 from cgilly2fast/dev
...
chore(frogbot): convert firmware provider to frogbot
2026-05-04 10:26:36 -05:00
Aiden Cline
5457215e29
Merge pull request #1650 from PedroACosta/feat/add-dinference-models
...
feat(dinference): add GLM-5.1 and MiniMax-M2.5 models
2026-05-04 10:26:02 -05:00
Aiden Cline
b3ab45990f
Merge pull request #1697 from rocuevas9511/feat/add-deepinfra-gemma4
...
add gemma4 26b a4b and 31b to deepinfra
2026-05-04 10:25:31 -05:00
Aiden Cline
d906a07e31
Merge pull request #1703 from dpuyosa/feat/venice-grok-4-3
...
Venice: Add Grok 4.3 model configuration
2026-05-04 10:25:16 -05:00
dpuyosa
4b1f6edd52
[venice] Add Grok 4.3 model configuration
...
- Add Grok 4.3 model with 1M context and 32K output
- Configure standard and >200K cost tiers
- Enable text+image input with text output modalities
2026-05-04 09:54:01 +02:00
Aiden Cline
a92a2cfe3d
Merge pull request #1702 from langyo/fix/glm-5v-turbo-naming
...
fix: use proper GLM family casing for GLM-5V-Turbo
2026-05-03 16:53:44 -05:00
Aiden Cline
70891f58e5
Merge pull request #1701 from kaeltrn/add-perplexity-agent-opus-4-7-gpt-5-5
...
Add Claude Opus 4.7 and GPT-5.5 models for perplexity-agent
2026-05-03 16:53:26 -05:00
Aiden Cline
1600c827fc
Merge pull request #1698 from tim-mcdonald/add-kimi-k2.6-nvidia
...
Add Kimi K2.6 model for NVIDIA provider
2026-05-03 16:53:01 -05:00
Aiden Cline
5851cdc136
Merge pull request #1700 from JDinABox/dev
...
Add Synthetic Kimi-K2.6 model configuration
2026-05-03 16:52:46 -05:00
langyo
4fd0e38c58
fix: use proper GLM family casing for GLM-5V-Turbo
...
- Rename glm-5v-turbo to GLM-5V-Turbo in zai, zhipuai, and 302ai providers
- Add GLM-5V-Turbo back to zhipuai-coding-plan (removed in #1589 )
Ref: #1589
2026-05-04 00:51:54 +08:00
Pedro
dd685ea42c
refactor(dinference): use extends format for GLM and MiniMax models
2026-05-03 17:54:18 +02:00
kaeltrn
7835298241
Add Claude Opus 4.7 and GPT-5.5 models for perplexity-agent
2026-05-03 20:09:14 +07:00
Isaac Ng
c5fbcc2c9b
📦 CHORE: remove senera
2026-05-03 16:17:37 +08:00
Isaac Ng
ab2eb51b4e
chore(nvidia): align Nemotron endpoint slugs
...
Replace stale NVIDIA Nemotron entries with the live Build catalog slugs so the local provider catalog matches current free and partner endpoints.
2026-05-03 15:47:19 +08:00
Isaac Ng
8e19ec580c
📦 CHORE: sync latest nvidia model
2026-05-03 15:19:17 +08:00
Isaac Ng
3aecc94c46
chore(nvidia): sync endpoint model catalog
...
Update NVIDIA model TOMLs to match the live Build endpoint list by removing stale entries and adding missing ones.
This keeps the provider catalog aligned with the current free and partner endpoint inventory.
2026-05-03 14:47:30 +08:00
JD Crawford
2ac7912ee7
use extends format
2026-05-03 01:04:53 -04:00
JD Crawford
9ac5b5f625
Add Synthetic Kimi-K2.6 model configuration
2026-05-03 00:17:37 -04:00
Aiden Cline
8c4d9f4696
Merge pull request #1699 from cfal/qwen3.6-max-preview
...
providers/alibaba/models/qwen3.6-max-preview.toml: add qwen 3.6 max
2026-05-02 21:35:02 -05:00
cfal
c4bb0b4b11
providers/alibaba/models/qwen3.6-max-preview.toml: add qwen 3.6 max
2026-05-03 09:25:31 +08:00
Tim McDonald
db4d03c171
Add Kimi K2.6 model for NVIDIA provider
2026-05-02 16:09:37 -06:00
rocuevas9511
60d1d4df77
add gemma4 26b a4b and 31b to deepinfra
2026-05-02 14:26:19 -06:00
Colby Gilbert
31654fc2ef
chore(frogbot): convert firmware provider to frogbot
2026-05-02 12:36:20 -07:00
Aiden Cline
01e56b1e1e
Merge pull request #1682 from varunrandery/poolside/laguna-openrouter
...
Add Poolside Laguna series models (OpenRouter)
2026-05-02 14:30:20 -05:00
Aiden Cline
4a7d9275d6
Merge pull request #1695 from hgraca/cortecs
...
Cortecs
2026-05-02 14:04:00 -05:00
Aiden Cline
978fe11e0c
Merge pull request #1694 from cgilly2fast/dev
...
feat(firmware): add deepseek v4, remove gemini 3 pro add gpt 5.4 min,…
2026-05-02 14:03:14 -05:00
Herberto Graca
0b4198955f
refactor(cortecs): use extends for models with canonical bases
...
Convert 7 Cortecs models to extends format, inheriting from their
canonical provider definitions (deepseek, llama, mistral, alibaba).
Reduces duplication by ~55 lines while preserving Cortecs-specific
cost overrides.
2026-05-02 20:56:11 +02:00
Herberto Graca
66b568b681
Add deepseek-v4-pro model for cortecs
2026-05-02 20:56:11 +02:00
Herberto Graca
589fddeed7
Add codestral-2508 model for cortecs
2026-05-02 20:56:10 +02:00
Herberto Graca
6f75266d42
Add deepseek-v3.2 model for cortecs
2026-05-02 20:56:10 +02:00
Herberto Graca
0f03115ee0
Add deepseek-r1-0528 model for cortecs
2026-05-02 20:56:09 +02:00
Herberto Graca
ae2564edc2
Add mixtral-8x7B-instruct-v0.1 model for cortecs
2026-05-02 20:56:09 +02:00
Herberto Graca
629856cf41
Add hermes-4-70b model for cortecs
2026-05-02 20:56:08 +02:00
Herberto Graca
567bc13ab4
Add llama-3.3-70b-instruct model for cortecs
2026-05-02 20:56:08 +02:00
Herberto Graca
5eac173bc5
Add qwen3-235b-a22b-instruct-2507 model for cortecs
2026-05-02 20:56:07 +02:00
Herberto Graca
cb0d640afd
Add qwen3-coder-30b-a3b-instruct model for cortecs
2026-05-02 20:56:07 +02:00
Herberto Graca
b5e27e652f
Add nemotron-3-super-120b-a12b model for cortecs
2026-05-02 20:56:06 +02:00
Herberto Graca
650c27411f
Add qwen3.5-122b-a10b model for cortecs
2026-05-02 20:56:06 +02:00
Herberto Graca
98c55325bf
Add mistral-large-2512 model for cortecs
2026-05-02 20:56:05 +02:00
Herberto Graca
2fdd33b8a7
Add qwen3.5-397b-a17b model for cortecs
2026-05-02 20:55:57 +02:00
Colby Gilbert
e4dea0a3d8
feat(firmware): grok 4.3
2026-05-02 10:53:24 -07:00
Colby Gilbert
b4b3622bfd
feat(firmware): add deepseek v4, remove gemini 3 pro add gpt 5.4 min, add gpt 5.4 nano, add gpt 5.5, and minimax m2.7
2026-05-02 09:23:42 -07:00
Aiden Cline
3b35b5598a
Merge pull request #1693 from hgraca/add-deepseek-v4-flash-cortecs
...
Add deepseek-v4-flash model for cortecs
2026-05-02 10:46:57 -05:00
Aiden Cline
d2e16bab34
Merge pull request #1610 from monotykamary/feat/neuralwatt-provider
...
feat: add neuralwatt provider with 14 models
2026-05-02 10:45:50 -05:00
Tom X Nguyen
051dc6236a
fix(neuralwatt): sync model capabilities with provider API
2026-05-02 20:42:00 +07:00
Herberto Graca
5b9dee1f3d
Add deepseek-v4-flash model for cortecs
2026-05-02 12:50:59 +02:00
8023
426dfe7be5
fix validate error
2026-05-02 12:15:50 +08:00
8023
fddfb083d1
add EmpirioLabs AI and deepseek v4
2026-05-02 11:14:51 +08:00
8023
4eccfaba87
Add Kimi-K2.6 model configuration file
2026-05-02 11:08:32 +08:00
Jeff Lim
ea46d016d3
fix(clarifai): drop redundant cost override for Kimi-K2.6
...
Clarifai's published pricing ($0.95 input, $4.00 output) matches the
canonical moonshotai/kimi-k2.6, so the explicit [cost] block was just
duplicating upstream. Inherit it via extends instead.
Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com >
2026-05-01 18:20:39 -07:00
Jeff Lim
1bd1449ec5
provider(clarifai): add Kimi-K2.6 (moonshotai/chat-completion)
...
Extends moonshotai/kimi-k2.6 with Clarifai-specific cost and modalities
(text+image only on Clarifai; cache pricing not exposed).
Model URL: https://clarifai.com/moonshotai/chat-completion/models/Kimi-K2_6
Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com >
2026-05-01 18:14:19 -07:00
Varun Randery
7d334a51c9
Add Laguna models
2026-05-01 23:35:13 +01:00
Aiden Cline
10ae0b5a1d
Merge pull request #1664 from fernandoenzo/fix/glm-qwen-ollama-output-limit
...
fix(ollama): set glm-5.1 and qwen3.5:397b output limits to match context
2026-05-01 15:50:22 -05:00
Aiden Cline
d25cce170d
Merge pull request #1665 from Ardakilic/chore/cleanup-kilo-provider
...
Chore: Sync Kilo Gateway provider models with upstream API changes and add Owl Alpha model.
2026-05-01 15:49:53 -05:00
Aiden Cline
3e097a5d89
Merge pull request #1671 from hgraca/add-qwen-2.5-72b-instruct-cortecs
...
Add qwen-2.5-72b-instruct model for cortecs
2026-05-01 15:49:28 -05:00
Aiden Cline
074b38eb98
Merge pull request #1663 from fernandoenzo/fix/minimax-m2.7-ollama-context-output-limit
...
fix(ollama): set minimax-m2.7 context and output limits to match Ollama API
2026-05-01 14:24:35 -05:00
Aiden Cline
d474922588
Merge pull request #1679 from Spherrrical/add-digitalocean-deepseek-v4-pro
...
feat(digitalocean): add deepseek-v4-pro model
2026-05-01 13:46:50 -05:00
Spherrrical
56ddb9017c
feat(digitalocean): add deepseek-v4-pro model
2026-05-01 11:20:09 -07:00
Rohan Taneja
a565bbebc0
Merge pull request #1677 from vercel/update-vercel-models-1777652649
2026-05-01 10:47:39 -07:00
github-actions[bot]
5b97fca592
chore(vercel): update Vercel model definitions
...
Auto-generated by weekly workflow from Vercel AI Gateway API.
Co-Authored-By: github-actions[bot] <41898282+github-actions[bot]@users.noreply.github.com>
2026-05-01 16:24:16 +00:00
Arda Kilicdagi
aec3f94081
feat: owl alpha, chore: sync kilo code upstream
...
feat: owl alpha, chore: sync kilo code upstream
2026-05-01 14:09:51 +03:00
Fernando Guarini
e5c63d671e
fix(ollama): set glm-5.1 and qwen3.5:397b output limits to match context
2026-05-01 11:29:49 +02:00
Fernando Guarini
22e87ab31b
fix(ollama): set minimax-m2.7 context and output limits to match Ollama API
2026-05-01 11:29:41 +02:00
Shoubhit Dash
90515f1913
provider(sarvam): add chat models
2026-05-01 14:41:47 +05:30
Herberto Graca
97d93676bc
Add qwen-2.5-72b-instruct model for cortecs
2026-05-01 09:36:16 +02:00
Alex-wuhu
6c3c4a721c
provider(novita-ai): sync latest models
2026-05-01 14:19:17 +08:00
Aiden Cline
692fbd0f19
Merge pull request #1628 from fanweixiao/dev
...
provider(vivgrid): remove GLM-5, add GPT-5.5 and DeepSeek-v4-Pro model
2026-04-30 23:57:52 -05:00
C.C. Fan
c3bd263a67
provider(vivgrid): add deepseek-v4-pro model
...
Co-Authored-By: Claude Sonnet 4.6 <noreply@anthropic.com >
2026-05-01 11:46:39 +08:00
Aiden Cline
9b9685c297
Merge pull request #1658 from v1gnesh/dev
...
Create grok-4.3.toml
2026-04-30 22:46:08 -05:00
C.C. Fan
b3f063da79
provider(vivgrid): use extends for gpt-5.5 instead of duplicating fields
...
Co-Authored-By: Claude Sonnet 4.6 <noreply@anthropic.com >
2026-05-01 11:45:13 +08:00
C.C.
a33d0aea81
Merge branch 'anomalyco:dev' into dev
2026-05-01 11:38:42 +08:00
v1gnesh
ca380a136e
Create grok-4.3.toml
2026-05-01 08:17:12 +05:30
Arda Kılıçdağı
dfa186a768
feat: owl alpha
2026-05-01 01:47:25 +03:00
Aiden Cline
d63ffa53e8
Merge pull request #1654 from zainhas/dev
...
[Together AI] add qwen 3.6 plus
2026-04-30 16:34:28 -05:00
Aiden Cline
284def86ef
Merge pull request #1653 from smakosh/feat/llmgateway-add-gpt-5-5-and-qwen3-6
...
feat(llmgateway): add gpt-5.5, gpt-5.5-pro, qwen3.6-35b-a3b, qwen3.6-plus, qwen3.6-max-preview
2026-04-30 16:34:18 -05:00
Siddharth Dhulipalla
c19936565d
Remove Fire Pass from Fireworks Kimi K2.5 Turbo description ( #1655 )
2026-04-30 17:28:36 -04:00
Zain Hasan
73b872e141
output 500_000
2026-04-30 14:22:41 -07:00
Zain Hasan
7373bb5878
[Together AI] add qwen 3.6 plus
2026-04-30 14:21:40 -07:00
smakosh
d11c151c25
feat(llmgateway): add gpt-5.5, gpt-5.5-pro, qwen3.6-35b-a3b, qwen3.6-plus, qwen3.6-max-preview
...
Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com >
2026-04-30 22:05:40 +02:00
Aiden Cline
f3f4fea66c
Merge pull request #1645 from stylings/stylings/add-mistral-medium-3-5
...
feat: add Mistral Medium 3.5
2026-04-30 14:37:16 -05:00
Aiden Cline
90116256a6
Merge pull request #1652 from Spherrrical/add-digitalocean-provider
...
feat(digitalocean): sync model catalog
2026-04-30 14:32:07 -05:00
Spherrrical
5008df8bcc
feat(digitalocean): sync model catalog with /v1/models API
...
Add 16 new models (anthropic, openai, alibaba, deepseek, google,
meta, mistral, nvidia, baai, intfloat, fal-hosted) to match the
current DigitalOcean Gradient AI Platform catalog, and rename
openai-gpt-5-2-pro to openai-gpt-5.2-pro to match the API id.
2026-04-30 12:08:57 -07:00
Alex
696aa80dec
revert(openrouter): remove Mistral Medium 3.5 stub
2026-04-30 14:59:56 -04:00
Aiden Cline
7d00d863dc
Merge pull request #1647 from dpuyosa/chore/venice-kimi-pricing
...
Venice: Update Kimi K2.5 and K2.6 pricing and dates
2026-04-30 11:29:08 -05:00
Aiden Cline
ad9eb83b8c
Merge pull request #1648 from Snat3r/patch-1
...
Fix casing in model name MiniMax m2.7 foor cortects provider
2026-04-30 11:28:55 -05:00
Aiden Cline
467da9a82c
Merge pull request #1649 from berget-ai/feat/berget-mistral-medium-3.5
...
feat: add Mistral Medium 3.5 128B to berget.ai
2026-04-30 11:28:43 -05:00
Pedro
22786bcf4b
feat(dinference): add GLM-5.1 and MiniMax-M2.5 models
2026-04-30 13:43:44 +02:00
Christian Landgren
c8d258b7cb
feat: add Mistral Medium 3.5 128B to berget.ai
2026-04-30 12:31:28 +02:00
Snat3r
70309829ca
Fix casing in model name and update output limit
2026-04-30 11:59:44 +02:00
dpuyosa
71e00f193b
[venice] Update Kimi K2.5 and K2.6 pricing and dates
...
- Bump kimi-k2-5 cache_read from 0.11 to 0.22
- Bump kimi-k2-6 input from 0.7448 to 0.85 and cache_read from 0.1463 to 0.22
- Update last_updated dates to 2026-04-30
2026-04-30 11:18:54 +02:00
Alex
d44e724170
feat(openrouter): add Mistral Medium 3.5
2026-04-29 18:58:26 -04:00
Alex
414695db9f
feat(mistral): add Mistral Medium 3.5
2026-04-29 18:44:28 -04:00
Aiden Cline
e8b5a27723
Merge pull request #1642 from Sewer56/add-wafer-deepseek-v4-pro
...
feat(wafer.ai): add DeepSeek V4 Pro
2026-04-29 17:14:17 -05:00
Aiden Cline
6dc9d805db
Merge pull request #1643 from Sawyerb/patch-1
...
Delete providers/vercel/models/inception/mercury-coder-small.toml
2026-04-29 17:14:00 -05:00
Aiden Cline
bdf56c9111
add kimi k2.6 to azure cognitive services
2026-04-29 17:13:19 -05:00
Sewer56
293cc3e82f
feat(wafer.ai): add DeepSeek V4 Pro
2026-04-29 23:12:04 +01:00
Sawyer Birnbaum
7f5bd4f231
Delete providers/vercel/models/inception/mercury-coder-small.toml
...
Mercury Coder Small has been deprecated. People should use Mercury Edit 2 instead.
2026-04-29 14:50:35 -07:00
Aiden Cline
c4826babc5
Merge pull request #1497 from Lydanne/fix/302ai-models
...
Update 302ai model metadata and add GPT-5.4 configs
2026-04-29 14:26:01 -05:00
Rohan Taneja
91e8bb985e
Merge pull request #1641 from vercel/update-vercel-models-1777480545
...
Update Vercel models
2026-04-29 11:29:45 -07:00
Aiden Cline
56723051d4
Merge pull request #1615 from deaquino/dev
...
Add Qwen3.5-9B model configuration file to OVHCloud
2026-04-29 13:09:19 -05:00
Aiden Cline
0bb3e55c08
Merge pull request #1640 from dpuyosa/chore/venice-update-pricing
...
Venice: Update DeepSeek v4 and Qwen 3.6 model pricing and metadata
2026-04-29 13:02:40 -05:00
github-actions[bot]
080b7a328b
chore(vercel): update Vercel model definitions
...
Auto-generated by weekly workflow from Vercel AI Gateway API.
Co-Authored-By: github-actions[bot] <41898282+github-actions[bot]@users.noreply.github.com>
2026-04-29 16:35:47 +00:00
Misha Skvortsov
c59c4aae6c
feat(atomic-chat): add curated initial model list
...
Re-introduces a small curated list of models that ship preconfigured
in Atomic Chat, so opencode users get a working `models.dev` entry
out of the box instead of an empty `models: {}`.
Models (ids match the normalized form returned by Atomic Chat's
/v1/models endpoint, i.e. dots replaced with underscores):
- gemma-4-E4B-it-IQ4_XS
- gemma-4-E4B-it-MLX-4bit
- Qwen3_5-9B-Q4_K_M
- Qwen3_5-9B-MLX-4bit
- Meta-Llama-3_1-8B-Instruct-GGUF
Qwen 3.5 9B dates are taken from the verified providers/venice entry
for the same base model; quantization does not change release dates.
Made-with: Cursor
2026-04-29 17:52:56 +03:00
Frank
c5d696583e
update zen models
2026-04-29 09:38:00 -04:00
Mike Sukmanowsky
2cb5a99b98
fix: add model card links for Kimi K2 and Kimi K2.5
2026-04-29 09:29:36 -04:00
dpuyosa
70490d892f
[venice] Update DeepSeek v4 and Qwen 3.6 model pricing and metadata
...
- Reduce DeepSeek v4 Flash/Pro input and output pricing
- Add cache_read pricing for DeepSeek v4 models
- Fix Qwen 3.6 27B model name formatting
2026-04-29 09:39:11 +02:00
Aiden Cline
f858a85ba6
Merge pull request #1625 from xinrui-z/fix-aihubmix-2026-04-28
...
fix: sync AIHubMix models (2026-04-28)
2026-04-28 23:03:28 -05:00
Aiden Cline
44e4c92882
Merge pull request #1620 from kill74/add-zai-coding-plan-glm-5v-turbo
...
Add GLM-5V-Turbo to Z.ai coding plan
2026-04-28 19:28:48 -05:00
Aiden Cline
43eafcb258
Merge pull request #1581 from xiaojiezj/zenmux_0425
...
feat: add models for zenmux provider
2026-04-28 19:27:38 -05:00
Aiden Cline
b382ac7af9
Merge branch 'dev' into zenmux_0425
2026-04-28 19:08:09 -05:00
Tom X Nguyen
ca21110644
fix: sync neuralwatt models with updated API pricing and capabilities
...
The Neuralwatt API now returns accurate pricing and capabilities,
eliminating the need for manual patches (patch.json is now empty).
Changes:
- Update pricing for all 14 models from API (significant changes for
GLM, GPT-OSS, Qwen, and MiniMax models)
- Devstral Small 2 now supports image input (vision)
- kimi-k2.5-fast now supports image input (vision)
- kimi-k2.6-fast now supports reasoning + image input (was non-reasoning)
- Qwen3.6-35B-A3B now supports reasoning (was non-reasoning)
- GLM models context window: 202,752 → 200,000
- Rename fast variant model IDs to match API (dropped org prefix):
zai-org/glm-5-fast → glm-5-fast
zai-org/glm-5.1-fast → glm-5.1-fast
moonshotai/kimi-k2.5-fast → kimi-k2.5-fast
moonshotai/kimi-k2.6-fast → kimi-k2.6-fast
Qwen/qwen3.5-397b-fast → qwen3.5-397b-fast
Qwen/qwen3.6-35b-fast → qwen3.6-35b-fast
2026-04-29 07:05:23 +07:00
Mike Sukmanowsky
0c2e47e8ba
Fix token limits for Amazon Bedrock Kimi K2 models
...
Correct context and output limits for moonshot.kimi-k2-thinking and
moonshotai.kimi-k2.5 on Amazon Bedrock:
- context: 256_000 → 262_143
- output: 256_000 → 16_000
2026-04-28 17:45:59 -04:00
Aiden Cline
6a0704574b
Merge pull request #1621 from YuzhongHuangCS/dev
...
feat(wandb): Add GLM-5.1
2026-04-28 15:51:43 -05:00
Aiden Cline
1cb1341516
Merge pull request #1635 from stylings/feat/nemotron-3-nano-omni
...
feat: add Nemotron 3 Nano Omni model
2026-04-28 15:04:07 -05:00
Alex
bc47e95427
fix: rename Nemotron Omni metadata
2026-04-28 15:20:24 -04:00
Aiden Cline
b071e8add8
Merge pull request #1619 from fernandoenzo/fix/deepseek-v4-pro-ollama-cloud
...
fix(ollama-cloud): correct deepseek-v4-pro model config
2026-04-28 14:00:42 -05:00
Aiden Cline
23e527753e
Merge pull request #1623 from itsnebulalol/dev
...
feat: add gpt-5.5 pro on openai and openrouter
2026-04-28 14:00:34 -05:00
Aiden Cline
e81c045ed0
Merge pull request #1636 from dsingal0/feat/openrouter-deepseek-v4
...
feat(baseten): add DeepSeek V4 Pro
2026-04-28 13:49:11 -05:00
Dhruv Singal
0c602ca936
feat(baseten): update DeepSeek V4 Pro pricing
2026-04-28 11:45:02 -07:00
Dhruv Singal
9866f84989
feat(baseten): add DeepSeek V4 Pro
2026-04-28 11:38:22 -07:00
Alex
6b397ffe37
fix: align nvidia output limit
2026-04-28 14:35:31 -04:00
Alex
bc21596889
fix: drop openrouter provider prefix
2026-04-28 14:22:04 -04:00
Alex
20abba3190
feat: add Nemotron 3 Nano Omni
2026-04-28 14:17:22 -04:00
Dominic Frye
5881bf98a0
fix: enable pdf input modality for gpt-5.5 pro
2026-04-28 13:33:31 -04:00
Aiden Cline
595f7d028c
Merge pull request #1632 from rocuevas9511/feat/deepinfra-deepseek-v4-pro
...
feat: add DeepSeek-V4-Pro to deepinfra
2026-04-28 12:07:22 -05:00
rocuevas9511
c81dec9c5d
feat: add DeepSeek-V4-Pro to deepinfra
2026-04-28 10:56:44 -06:00
Guiii
4d45ed25d2
Use extended GLM-5V-Turbo config
...
Removed various fields and added extends section.
2026-04-28 17:49:34 +01:00
Yuzhong Huang
03cf48de53
use extends instead
2026-04-28 09:19:26 -07:00
Aiden Cline
0d3a284395
Merge pull request #1622 from eduqr/feat/fireworks-ai-deepseek-v4-pro
...
feat(fireworks-ai): add deepseek-v4-pro
2026-04-28 10:52:54 -05:00
Aiden Cline
5b1bb0fc80
Merge pull request #1624 from shelvick/add-azure-kimi-k2-6
...
Add Kimi K2.6 to Azure
2026-04-28 10:38:07 -05:00
Aiden Cline
332ebb8811
Merge pull request #1627 from ceoAppsknight/kilo/add-mimo-models
...
Add Kilo Mimo v2.5 models
2026-04-28 10:37:40 -05:00
Aiden Cline
2111813bd4
Merge pull request #1629 from ndeybach/PR-azure-5.4-limits
...
fix(azure): correct GPT-5.4 series limits and cleanup
2026-04-28 10:37:01 -05:00
Nils DEYBACH
f5b8521af6
fix: use extends and not symlinks
2026-04-28 17:34:44 +02:00
Nils DEYBACH
79481cff40
fix(azure): update GPT-5.4 metadata
...
Use `extends` for Azure GPT-5.4 variants and keep Azure-specific overrides for
PDF input and omitted fast mode.
Validated with `bun validate`.
Azure runtime manual probing confirmed GPT-5.4 uses the documented 1.05M context /
922K input / 128K output limits.
2026-04-28 14:00:20 +02:00
Nils DEYBACH
40dc356d4c
fix(azure): correct GPT-5.4 and GPT-5.4 Pro limits (and convert to extend)
...
Correct Azure GPT-5.4 and GPT-5.4 Pro limits to `1_050_000` context,
`922_000` input, and `128_000` output based on Azure runtime results and
Microsoft Learn docs. Mini and Nano already matched and are unchanged.
The limits were tested directly (see script at : https://github.com/ndeybach/Azure_endpoint_limit_test_script )
2026-04-28 12:33:27 +02:00
C.C. Fan
f929fe89e7
provider(vivgrid): remove GLM-5, add GPT-5.5 model
2026-04-28 16:44:59 +08:00
Syed Assadullah Shah
cbd245d454
add Kilo Mimo v2.5 models
2026-04-28 13:04:36 +05:00
xinrui
e5289e9b3a
fix: sync AIHubMix models (2026-04-28)
2026-04-28 11:33:25 +08:00
Scott Helvick
bb623f3ff9
Add Kimi K2.6 to Azure
2026-04-28 02:28:46 +00:00
Dominic Frye
37fffafafc
feat: add gpt-5.5 pro on openai and openrouter
2026-04-27 22:24:32 -04:00
eduqr
d329310745
feat(fireworks-ai): add deepseek-v4-pro
2026-04-27 21:05:55 -05:00
Yuzhong Huang
8016a6c45a
Add GLM-5.1 to wandb provider
2026-04-27 17:35:38 -07:00
Guilherme Sales
1eeaa0b756
Add GLM-5V-Turbo to Z.ai coding plan
2026-04-28 00:47:15 +01:00
Frank
dd3533b4e0
update zen models
2026-04-27 19:31:24 -04:00
Fernando Guarini
3a5867834f
fix(ollama-cloud): correct deepseek-v4-pro model config
...
- Remove fields that don't belong in ollama-cloud: temperature, structured_output, knowledge, interleaved
- Set output = context (1048576) per ollama-cloud convention
- Set name to lowercase per ollama-cloud convention
- Reorder fields to match existing ollama-cloud model files
2026-04-28 00:24:54 +02:00
Aiden Cline
1e83bca7a3
Merge pull request #1617 from JoshuaDietz/dev
...
feat(ollama cloud): add deepseek v4 pro
2026-04-27 16:42:56 -05:00
Aiden Cline
cd8853f88b
Merge pull request #1616 from fhennerkes/dev
...
poe: add GPT-5.5 and GPT-5.5-Pro models
2026-04-27 16:08:36 -05:00
Joshua Dietz
23b290c6a8
fix(ollama cloud): fix model name
...
Model name was inconsistent with naming schema of flash model on ollama cloud
2026-04-27 21:47:30 +02:00
fhennerkes
4e7849cee7
poe: reduce omits in gpt-5.5 extends configs
...
Inherit family, knowledge, and structured_output from base models
instead of omitting them. Only omit fields that genuinely don't
apply to Poe (provider-specific pricing tiers, different context
limits, opencode-specific provider config).
Co-Authored-By: Claude Sonnet 4.6 <noreply@anthropic.com >
2026-04-27 12:44:49 -07:00
Joshua Dietz
e3e63a7247
feat(ollama cloud): add deepseek v4 pro
2026-04-27 21:42:10 +02:00
Aiden Cline
fb297153e4
Merge pull request #1572 from YoshiTabletopGamer/qwen3.5-3.6-alibaba-open
...
[alibaba] Add remaining open Qwen 3.5 and 3.6 models, fix Qwen-3.5 397B-A17B
2026-04-27 14:36:51 -05:00
fhennerkes
e8dd06e0ce
poe: use extends format for gpt-5.5-pro
...
Address PR review comment to use extends format. Inherit from
opencode/gpt-5.5-pro since no openai/gpt-5.5-pro base exists yet.
Co-Authored-By: Claude Sonnet 4.6 <noreply@anthropic.com >
2026-04-27 10:49:36 -07:00
fhennerkes
bdee3d438b
poe: use extends format for gpt-5.5
...
Address PR review comment to use extends format and inherit from
openai/gpt-5.5 base model.
Co-Authored-By: Claude Sonnet 4.6 <noreply@anthropic.com >
2026-04-27 10:36:10 -07:00
fhennerkes
b410bc3ef2
poe: add GPT-5.5 and GPT-5.5-Pro models
2026-04-27 10:34:52 -07:00
Aiden Cline
870e3d1d26
Merge pull request #1612 from hanouticelina/add-deepseek-v4-for-huggingface
...
feat(huggingface): add DeepSeek V4 Pro
2026-04-27 11:20:12 -05:00
Aiden Cline
5f21f3b603
Merge pull request #1586 from fernandoenzo/add-ollama-cloud-deepseek-v4-flash
...
feat(ollama-cloud): add deepseek-v4-flash model
2026-04-27 10:42:53 -05:00
Celina Hanouti
27dad05e25
extend deepseek/deepseek-v4-pro
2026-04-27 16:38:36 +01:00
Aiden Cline
2ed88cbcc0
Merge pull request #1604 from ndeybach/PR-gpt-5.5
...
feat(azure): add GPT-5.5 model metadata
2026-04-27 10:20:32 -05:00
Nils DEYBACH
4d199c932e
fix: base azure-cognitive-services model not on azure
...
extend of extend does not seem to be supported
2026-04-27 17:04:20 +02:00
Jaime de Aquino
2d142f920c
Add Qwen3.5-9B model configuration file
2026-04-27 16:46:04 +02:00
Celina Hanouti
47dea9e551
fix
2026-04-27 15:40:35 +01:00
Celina Hanouti
0d25c3dcac
use extends
2026-04-27 15:37:31 +01:00
Aiden Cline
3d7f9256cb
Merge pull request #1583 from abliteration-ai/codex/add-abliteration-provider
...
Add abliteration.ai provider
2026-04-27 09:29:26 -05:00
Aiden Cline
6aa1ebd4be
Merge pull request #1595 from Contraboi/contra/add-openrouter-nano-banana-2
...
feat(openrouter): add Gemini 3.1 flash image preview (Nano Banana 2)
2026-04-27 09:28:23 -05:00
Aiden Cline
f9ebebaffd
Merge pull request #1601 from shikbupt/alibaba-deepseek
...
add alibaba-cn deepseek-v4
2026-04-27 09:28:08 -05:00
sk
7b3fe83c09
use extend format
2026-04-27 21:48:26 +08:00
Yashwanth Kumar
0a06b3efc2
Update Qwen model configuration in TOML file
2026-04-27 16:40:28 +05:30
Yashwanth Kumar
90dcbbbcc5
Add Qwen3.6 27B model configuration
2026-04-27 16:29:56 +05:30
Yashwanth Kumar
b17f5fd8ae
Delete providers/openrouter/models/qwen/qwen-3.6-27b.toml
2026-04-27 16:28:19 +05:30
Yashwanth Kumar
68691ac3f9
Update providers/openrouter/models/qwen/qwen-3.6-27b.toml
...
Co-authored-by: Copilot <175728472+Copilot@users.noreply.github.com >
2026-04-27 16:27:56 +05:30
Yashwanth Kumar
78fe905fcb
Add Qwen 3.6 27B model configuration
2026-04-27 16:22:24 +05:30
Nils DEYBACH
6b488802cf
fix: parsing error and add context over price
...
- adds the context over X price capability from parent (awaiting refactor to be correct on exact limit threashold)
- fix parsing since anything must be before extends.
2026-04-27 12:27:47 +02:00
Jack
4d0505b70e
Merge pull request #1611 from anomalyco/fix/opencode-go-deepseek-v4-flash-cache-read-20260427
...
fix(opencode-go): update deepseek v4 flash cache pricing in Go
2026-04-27 17:34:52 +08:00
Jack
b729923bd9
fix(opencode-go): correct deepseek v4 flash cache pricing
2026-04-27 17:31:41 +08:00
Celina Hanouti
34c7aa7dfe
update context limit
2026-04-27 09:33:57 +01:00
Celina Hanouti
09b4d3548b
add support for DeepSeek V4 Pro for Hugging Face provider
2026-04-27 09:31:55 +01:00
xiaojie.zj
f1cad8fdc0
feat: add zenmux models
2026-04-27 16:27:46 +08:00
Tom X Nguyen
b2f7f57f26
feat: add neuralwatt provider with 14 models
...
Add Neuralwatt as an OpenAI-compatible inference provider with
energy-aware GPU optimization. Includes 14 models across 6
sub-providers (Mistral, ZAI, OpenAI, Moonshot, MiniMax, Qwen).
Models include reasoning variants (Kimi K2.5/K2.6, GLM 5.1 FP8,
MiniMax M2.5, Qwen3.5 397B, GPT OSS 20B) and fast non-reasoning
variants (Kimi K2.5/K2.6 Fast, GLM 5/5.1 Fast, Qwen3.5/3.6 Fast),
plus Devstral Small 2 and Qwen3.6 35B A3B.
Logo derived from official Neuralwatt favicon (currentColor variant).
Pricing sourced from Neuralwatt's published rates.
2026-04-27 15:01:56 +07:00
Aiden Cline
925d4eba1f
Merge pull request #1536 from philipmat/add-openrouter-pareto-code-router
...
Adds support for openrouter/pareto-code
2026-04-26 23:52:04 -05:00
Aiden Cline
bc1e4b870b
Merge pull request #1608 from Alex-wuhu/dev
...
add deepseek-v4, qwen3.6 on novita
2026-04-26 23:16:56 -05:00
Alex-wuhu
ef913f9645
refactor: use extends format for novita deepseek v4 models
...
Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com >
2026-04-27 12:00:04 +08:00
Alex-wuhu
cec747ade3
fix: add cache_read pricing for deepseek v4 models
2026-04-27 11:09:15 +08:00
Aiden Cline
2ced32d52f
Merge pull request #1596 from NathanDrake2406/add-cf-ai-gateway-gpt-5.5
...
feat(cloudflare-ai-gateway): add openai/gpt-5.5
2026-04-26 21:57:13 -05:00
Alex-wuhu
92cc5a79df
feat: add deepseek-v4, qwen3.6 on novita
2026-04-27 10:49:13 +08:00
Aiden Cline
4cf8661f92
Merge pull request #1602 from shikbupt/alibaba-qwen3.6-max
...
add alibaba-cn qwen3.6 max
2026-04-26 17:48:14 -05:00
Aiden Cline
ebe431c0dc
Merge pull request #1606 from LightAndy1/dev
...
Add gemini-3.1-flash-preview for google-vertex
2026-04-26 17:30:05 -05:00
LightAndy
8b2c5f30a0
✏️ Fix typo
2026-04-26 21:48:54 +03:00
Nils DEYBACH
0fc92e1c47
fix: align limit on base 5.5 model
...
now that limits were fixed in base, we align azure on it
2026-04-26 20:48:00 +02:00
Nils DEYBACH
d73ca9024a
Merge remote-tracking branch 'upstream/dev' into PR-gpt-5.5
2026-04-26 20:45:20 +02:00
LightAndy
bff48f9fde
Merge branch 'anomalyco:dev' into dev
2026-04-26 21:43:13 +03:00
Aiden Cline
c5c803f415
fix: ensure openai gpt-5.5 limits are exact
2026-04-26 13:38:56 -05:00
Nils DEYBACH
cd69509b48
fix: simplify by extending the azure model from opeani
2026-04-26 20:04:27 +02:00
Aiden Cline
83d15fd756
Merge pull request #1574 from juls0730/dev
...
feat: add mimo v2.5/pro to xiaomi and openrouter providers
2026-04-26 13:55:57 -04:00
LightAndy
b3c09451d9
Add gemini-3.1-flash-preview model configuration
2026-04-26 20:27:17 +03:00
Frank
d98bdb5eff
Merge pull request #1560 from TigerBeanst/patch-1
...
fix: opencode go mimo-v2.5 context limit to 1,000,000
2026-04-26 13:16:33 -04:00
Frank
ea205913ce
update zen models
2026-04-26 12:49:37 -04:00
Frank
3532801639
update zen models
2026-04-26 11:48:32 -04:00
Nils DEYBACH
1bb50141c2
feat(azure): add GPT-5.5 model metadata
...
## Summary
Adds GPT-5.5 metadata for:
- Azure
- Azure Cognitive Services
The Azure Cognitive Services entry mirrors the existing local convention of full TOML model definitions.
## Sources
- Microsoft Learn lists `gpt-5.5` for Azure OpenAI / Microsoft Foundry with version `2026-04-24`, `1,050,000` context, `922,000` input, `128,000` output, structured outputs, tools, image input, and December 2025 training data.
- Microsoft’s Azure GPT-5.5 announcement lists pricing at `$5.00` input, `$0.50` cached input, and `$30.00` output per 1M tokens.
- Azure Responses API docs list PDF input support for vision-capable models and include `gpt-5.5` version `2026-04-24`.
## Notes
This intentionally does not add `gpt-5.5-pro`, since Azure Learn currently lists `gpt-5.5` but not `gpt-5.5-pro` in the Azure model catalog.
This also intentionally omits OpenAI-specific `context_over_200k` and `experimental.modes.fast` metadata because the Azure sources confirm the standard pricing and limits, but not those OpenAI-specific fields.
2026-04-26 17:45:25 +02:00
sk
fa8bdd63fb
add alibaba-cn qwen3.6 max
2026-04-26 19:39:41 +08:00
sk
56f64577ce
add alibaba-cn deepseek-v4
2026-04-26 19:23:04 +08:00
Nathan Nguyen
f726af5767
refactor(cloudflare-ai-gateway): use [extends] for openai/gpt-5.5
...
The model entry duplicated every field from providers/openai/models/gpt-5.5.toml,
so any future change to the upstream OpenAI definition would silently drift here.
Switch to the `[extends] from = "openai/gpt-5.5"` form already used by sibling
providers (openrouter, requesty), omitting `experimental.modes.fast` since the
gateway does not surface the OpenAI priority service tier. Validation output is
byte-identical to the prior expanded form.
2026-04-26 13:33:39 +10:00
Zoe
4838e3cb9b
feat: add mimo v2.5/pro to xiaomi and openrouter providers
2026-04-25 21:32:24 -05:00
Muhammad Mugni Hadi
dbe92646c3
chore(chutes): add header comments to generated TOML files
...
Each generated TOML now includes a comment noting which fields are
auto-managed vs manually overridable on re-run.
2026-04-26 06:52:21 +07:00
Muhammad Mugni Hadi
4717c67054
feat(chutes): add API-driven model generator script
...
Add generate-chutes.ts that fetches models from https://llm.chutes.ai/v1/models
and generates/updates TOML files, following the same pattern as generate-vercel.ts.
Supports --dry-run, --new-only, and --keep-orphans flags. Auto-deletes TOML files
for models no longer in the API (with empty directory cleanup).
Preserves manually-set fields (family, knowledge, interleaved, status) when merging
with API data. Also syncs current models from the API.
2026-04-26 06:51:16 +07:00
Nathan Nguyen
d4c77c14fd
feat(cloudflare-ai-gateway): add openai/gpt-5.5
...
Mirrors the existing direct openai/gpt-5.5 entry under the
cloudflare-ai-gateway provider so opencode and other consumers can
route GPT-5.5 traffic through Cloudflare AI Gateway without hitting
ProviderModelNotFoundError.
Pricing, limits, modalities, and dates copied from
providers/openai/models/gpt-5.5.toml; provider stanza follows the
sibling gpt-5.4 entry (npm = "ai-gateway-provider").
2026-04-26 05:22:29 +10:00
Selmir Nedzibi
96b3d65307
feat(openrouter): add Gemini 3.1 flash image preview (Nano Banana 2)
2026-04-25 21:14:44 +02:00
Aiden Cline
b491c29cf9
Merge pull request #1573 from zainhas/dev
...
[Together AI] add deepseek-v4
2026-04-25 13:47:14 -04:00
Aiden Cline
d937abd849
Merge pull request #1539 from manascb1344/fix-xiaomi-provider-ids
...
feat: add MiMo-V2.5 and MiMo-V2.5-Pro to xiaomi-token-plan providers
2026-04-25 13:33:42 -04:00
Aiden Cline
df52175b0c
Merge pull request #1580 from LeGazeon/add-nvidia-deepseek-v4-pro/flash
...
Add NVIDIA DeepSeek-V4 models
2026-04-25 13:32:41 -04:00
Aiden Cline
bee8339c07
Merge pull request #1589 from MiyakoMeow/feat/restrict-zai-zhipuai-coding-plan-models
...
rm: unavailable models in zai/zhipuai coding plan
2026-04-25 13:30:39 -04:00
Aiden Cline
181bf96fa3
Merge pull request #1585 from saju01/add-copilot-gpt-5.5
...
feat(github-copilot): add gpt-5.5
2026-04-25 13:30:16 -04:00
Aiden Cline
9d49d2fd52
Merge pull request #1587 from smakosh/claude/rebase-add-llmgateway-models-yqKLn
...
feat(llmgateway): add deepseek-v4-pro, deepseek-v4-flash, kimi-k2.6
2026-04-25 13:29:39 -04:00
Aiden Cline
648776aa85
Merge pull request #1590 from dpuyosa/feat/venice-models
...
Venice: Add GPT-5.5 and Qwen3.6 model configs
2026-04-25 13:29:03 -04:00
Aiden Cline
421cb099b0
Merge pull request #1591 from dpuyosa/fix/venice-deepseek-family
...
Venice: Fix DeepSeek V4 Flash family classification
2026-04-25 13:28:54 -04:00
Aiden Cline
f458b19994
Merge pull request #1592 from MiyakoMeow/feat/deepseek-1m-context
...
fix(deepseek): all has 1M context / 384k output / adjusted price
2026-04-25 13:28:45 -04:00
MiyakoMeow
d347093b03
feat(deepseek): 1M context / 384k output
2026-04-25 18:58:12 +08:00
MiyakoMeow
3328712262
feat: restrict zai/zhipuai coding plan models to glm-5.1, glm-5-turbo, glm-4.7, glm-4.5-air only
...
Based on official documentation:
- ZAI DevPack Coding Plan: https://docs.z.ai/devpack/overview
- Zhipu AI BigModel Coding Plan: https://docs.bigmodel.cn/cn/coding-plan/overview
Both providers only officially support the following GLM models for coding plans:
- glm-5.1
- glm-5-turbo
- glm-4.7
- glm-4.5-air
Removed unsupported models from zai-coding-plan:
- glm-4.5, glm-4.5-flash, glm-4.5v
- glm-4.6, glm-4.6v
- glm-4.7-flash, glm-4.7-flashx
- glm-5, glm-5v-turbo
Removed unsupported models from zhipuai-coding-plan:
- glm-4.5, glm-4.5-flash, glm-4.5v
- glm-4.6, glm-4.6v, glm-4.6v-flash
- glm-4.7-flash, glm-4.7-flashx
- glm-5, glm-5v-turbo
2026-04-25 18:49:10 +08:00
dpuyosa
60edc1b52d
[venice] Add GPT-5.5 and Qwen3.6 model configs
...
- Add OpenAI GPT-5.5 with 1M context window and tiered pricing
- Add OpenAI GPT-5.5 Pro with premium pricing and 128K output limit
- Add Qwen3.6 27B with text, image, and video input modalities
2026-04-25 12:41:31 +02:00
dpuyosa
eee44cd080
[venice] Fix DeepSeek V4 Flash family classification
...
- Correct family from "deepseek" to "deepseek-flash" for accurate model categorization
2026-04-25 12:36:06 +02:00
smakosh
048a3235e8
feat(llmgateway): add deepseek-v4-pro, deepseek-v4-flash, kimi-k2.6
2026-04-25 12:19:40 +02:00
Fernando Guarini
1db03ec1e6
feat(ollama-cloud): add deepseek-v4-flash model
2026-04-25 11:24:12 +02:00
Saju Sarangdharan
6d283349ad
feat(github-copilot): add gpt-5.5
...
GitHub Copilot now serves gpt-5.5 (verified via GET https://api.githubcopilot.com/models with a Copilot Enterprise token). Adding the catalog row so downstream consumers (e.g. pi-ai) can route requests.
2026-04-25 10:31:21 +02:00
Abliteration.ai
7e07302ecd
add abliteration.ai provider
2026-04-24 22:46:54 -07:00
LeGazeon
8bc407a617
chore: remove deepseek-v4-pro config (duplicated by #1578 )
...
The Pro model configuration was already added via #1578 which
has been merged. Removing the duplicate from this branch to
keep only the Flash variant.
2026-04-25 13:17:12 +08:00
LeGazeon
5305d2bae2
refactor: extend flash config from deepseek base
...
Remove duplicated fields by inheriting common settings
from providers/deepseek base config via [extends].
This addresses the review comment in #1580
2026-04-25 13:11:26 +08:00
Aiden Cline
fee96c27b9
Merge pull request #1578 from panwar-stack/dev
...
feat(nvidia): add DeepSeek V4 model
2026-04-25 00:41:03 -04:00
Aiden Cline
66520adbc6
Merge pull request #1577 from ezShroom/dev
...
add openrouter gpt-5.5
2026-04-25 00:40:34 -04:00
Zain Hasan
40714995cc
Add interleaved section to DeepSeek-V4-Pro.toml
2026-04-24 18:59:45 -07:00
LeGazeon
66c4896003
Add NVIDIA DeepSeek-V4 models
...
Add model entries for DeepSeek V4 Pro and DeepSeek V4 Flash to the NVIDIA NIM provider.
## Changes
- Added `providers/nvidia/deepseek-v4-pro.toml`
- Added `providers/nvidia/deepseek-v4-flash.toml`
## Data Sources
- NVIDIA NIM Model Cards:
- DeepSeek V4 Pro: https://build.nvidia.com/deepseek-ai/deepseek-v4-pro/modelcard
- DeepSeek V4 Flash: https://build.nvidia.com/deepseek-ai/deepseek-v4-flash/modelcard
2026-04-25 09:57:43 +08:00
panwar-stack
31091f3d4e
Rename deepseek-v4.toml to deepseek-v4-pro.toml
2026-04-24 17:09:12 -07:00
panwar-stack
5fe512c1b1
Follow extends pattern
...
Follow extends pattern
2026-04-24 17:08:49 -07:00
panwar-stack
63efa131e7
feat(nvidia): add DeepSeek V4 model
...
add DeepSeek V4 model
2026-04-24 17:05:12 -07:00
Shroom
29cd503070
Add gpt-5.5.toml configuration file
2026-04-25 00:08:04 +01:00
Rohan Taneja
a9b704c656
Merge pull request #1575 from vercel/update-vercel-models-1777063875
2026-04-24 15:29:39 -07:00
Aiden Cline
0a88e412e5
Merge pull request #1576 from dsingal0/feat/openrouter-deepseek-v4
...
Add OpenRouter DeepSeek V4 models
2026-04-24 17:42:48 -04:00
Dhruv Singal
b46e29ccd5
fix(openrouter): use DeepSeek reasoning content field
2026-04-24 14:06:05 -07:00
Jerilyn Zheng
a181b770d6
Update kimi-k2.6.toml
2026-04-24 13:55:27 -07:00
Jerilyn Zheng
fa71201f20
Update deepseek-v4-pro.toml
2026-04-24 13:54:56 -07:00
Jerilyn Zheng
3ac17aefb6
Enable open_weights in deepseek-v4-flash configuration
2026-04-24 13:54:34 -07:00
Jerilyn Zheng
1f2ceb91a5
Update qwen-3.6-max-preview.toml
2026-04-24 13:53:57 -07:00
github-actions[bot]
d82681900c
chore(vercel): update Vercel model definitions
...
Auto-generated by weekly workflow from Vercel AI Gateway API.
Co-Authored-By: github-actions[bot] <41898282+github-actions[bot]@users.noreply.github.com>
2026-04-24 20:51:22 +00:00
Zain Hasan
7e296f4cbd
ds recommend 384,000
2026-04-24 12:39:09 -07:00
Zain Hasan
f202ff51e4
Reduce output limit from 512000 to 300000
2026-04-24 12:13:53 -07:00
Zain Hasan
4085a0b536
Merge branch 'anomalyco:dev' into dev
2026-04-24 12:12:28 -07:00
Zain Hasan
935fdeca65
[Together AI] Add deepseekv4 pro
2026-04-24 12:08:33 -07:00
Frank
cef8828fbe
update zen models
2026-04-24 14:51:59 -04:00
Frank
9277a23a29
update zen models
2026-04-24 14:50:50 -04:00
Frank
d7bfb16b0f
update zen models
2026-04-24 14:42:25 -04:00
Frank
37ae6fe7c8
update zen models
2026-04-24 12:11:33 -04:00
Aiden Cline
87ce527476
Merge pull request #1571 from cgilly2fast/dev
...
fix(firmware): glm 5.1 name
2026-04-24 11:49:51 -04:00
Aiden Cline
34bed30fa1
Merge pull request #1567 from dsingal0/feat/openrouter-deepseek-v4
...
feat(openrouter): add DeepSeek V4 Pro and V4 Flash
2026-04-24 11:49:35 -04:00
Dhruv Singal
8a0c3cb75f
refactor(openrouter): extend official DeepSeek V4 Pro/Flash
...
Use [extends] from deepseek/ with OpenRouter-specific overrides
(attachment, interleaved reasoning_details, limits).
Made-with: Cursor
2026-04-24 08:43:38 -07:00
Colby Gilbert
f65460ac16
fix(firmware): glm 5.1 name
2026-04-24 08:40:19 -07:00
YoshiTabletopGamer
8ea92aed9a
[alibaba] Add remaining open Qwen 3.5 models, fix Qwen-3.5 397B-A17B, add open Qwen 3.6 models
...
- Added Qwen 3.5 122B-A10B
- Added Qwen 3.5 27B
- Added Qwen 3.5 35B-A3B
- Fixed Qwen 3.5 397B-A17B (see below)
- Added Qwen 3.6 27b
- Added Qwen 3.6-35B-A3B
I was not able to find a reliable source for the knowledge cutoff of any of these models.
2025-04 was already set as the cutoff for Qwen 3, and Qwen 3.5 is newer.
All data is from the ModelStudio webpage.
It seems to not include audio, but the ModelStudio page clearly has an audio symbol and the model is capable of this.
And I found no data for a price for reasoning tokens in particular, unlike what was in the file for Qwen 3.5 397B-A17B.
The models are all capable of structured output.
2026-04-24 12:38:52 -03:00
Dhruv Singal
f340d82fc3
feat(openrouter): add DeepSeek V4 Pro and V4 Flash
...
Add model configs aligned with OpenRouter pricing and limits
(1M context, 384K max output, cache read rates from provider page).
Made-with: Cursor
2026-04-24 08:25:17 -07:00
Frank
c7431ae24c
update zen models
2026-04-24 10:53:10 -04:00
Frank
3d1888b7b5
update zen models
2026-04-24 10:24:34 -04:00
Misha Skvortsov
16a8fa5c20
improve(atomic-chat): drop hardcoded model list per maintainer feedback
...
Made-with: Cursor
2026-04-24 17:20:26 +03:00
Aiden Cline
dcd37ccdbb
add deepseek v4 flash
2026-04-24 08:34:03 -04:00
Aiden Cline
d18c3f910c
Merge pull request #1562 from dpuyosa/update/venice-kimi-pricing
...
Venice: Update kimi-k2-6 pricing
2026-04-24 08:08:23 -04:00
Aiden Cline
2cec5a492c
Merge pull request #1563 from dpuyosa/feat/venice-deepseek-v4
...
Venice: Add DeepSeek V4 Flash and Pro models
2026-04-24 08:08:13 -04:00
dpuyosa
61dd0ec489
[venice] Add DeepSeek V4 Flash and Pro models
...
- Add DeepSeek V4 Flash with 1M context, reasoning, and tool support
- Add DeepSeek V4 Pro with 1M context, reasoning, and tool support
- Set pricing and interleaved reasoning_content field for both
2026-04-24 12:21:07 +02:00
dpuyosa
c7758204b5
[venice] Update kimi-k2-6 pricing
...
- Update input, output, and cache_read costs to current rates
- Update last_updated timestamp to 2026-04-24
2026-04-24 12:17:51 +02:00
manascb1344
ed91520aa2
feat: add MiMo-V2.5 and MiMo-V2.5-Pro to xiaomi-token-plan providers
2026-04-24 15:38:55 +05:30
Frank
3e82669a82
Merge pull request #1561 from wenbindu/dev
...
add deepseek new moels
2026-04-24 03:04:52 -04:00
Frank
1cc0c9c074
sync
2026-04-24 03:03:06 -04:00
TigerBeanst
d73d7f6453
fix: opencode go mimo-v2.5 context limit to 1,000,000
...
https://platform.xiaomimimo.com/docs/pricing
2026-04-24 12:48:34 +08:00
Aiden Cline
afb59f86ee
Merge pull request #1557 from seffhunnn/dev
...
feat: add AU Sonnet and Opus models for Amazon Bedrock
2026-04-24 00:30:51 -04:00
wenbindu
05242f68d4
add deepseek new moel
2026-04-24 12:06:19 +08:00
Mohd Saif
c1b029dcc1
feat: add AU Opus model for Amazon Bedrock
2026-04-24 03:26:28 +05:30
Mohd Saif
bc2dd5137a
feat: add AU Sonnet model for Amazon Bedrock
2026-04-24 03:25:38 +05:30
Aiden Cline
99ec4900c7
Merge pull request #1555 from brentdurksen/add-azure-claude-sonnet-4-6
...
feat(azure): add Claude Sonnet 4.6 model
2026-04-23 17:33:44 -04:00
Brent Durksen
a8c124ac9e
refactor: use extends to inherit from anthropic/claude-sonnet-4-6
2026-04-23 15:16:38 -06:00
Aiden Cline
0d20a363a9
Merge pull request #1556 from fhennerkes/dev
...
poe: add GPT-Image-2 model
2026-04-23 17:12:20 -04:00
fhennerkes
3ac613678b
poe: add GPT-Image-2 model
2026-04-23 12:38:22 -07:00
Brent Durksen
e2ead1b4e6
feat(azure): add Claude Sonnet 4.6 model
2026-04-23 13:37:50 -06:00
Aiden Cline
be53c33588
Merge pull request #1550 from BlockListed/cortecs-kimi-k2.6
...
Add kimi k2.6 to cortecs
2026-04-23 15:26:58 -04:00
Aiden Cline
55cf5fa310
Merge pull request #1554 from mattyatea/add-gpt-5-5
...
[codex] Add GPT-5.5
2026-04-23 15:17:20 -04:00
mattyatea
3e0fe362f2
add gpt-5.5 model
2026-04-24 04:13:51 +09:00
BlockListed
89d06ae31f
add kimi k2.6 to cortecs
2026-04-23 19:49:18 +02:00
Aiden Cline
c994b116ae
Merge pull request #1542 from u007/patch-1
...
Add Chutes: Kimi K2.6 TEE
2026-04-23 12:47:49 -04:00
Aiden Cline
833e8f7a66
Merge pull request #1548 from fernandoenzo/fix/gemma4-ollama-output-limit
...
fix(ollama): set gemma4:31b output limit to match context
2026-04-23 12:46:00 -04:00
Aiden Cline
32bd1427fb
Merge pull request #1545 from Alex-wuhu/dev
...
Add deepseek, gemma, ling, llama, kimi on NovitaAI
2026-04-23 12:36:06 -04:00
Frank
ae7672b87e
update zen models
2026-04-23 11:11:30 -04:00
Fernando Guarini
9b27cc5a76
fix(ollama): set gemma4:31b output limit to match context
...
Ollama does not impose official output limits. The existing convention for Gemma models on Ollama (gemma3:4b, gemma3:12b, gemma3:27b) is to set output equal to context. gemma4:31b was the only exception with output=8192 vs context=262144.
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-openagent )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-04-23 11:42:30 +02:00
Alex-wuhu
7014e3e318
feat: add missing Novita AI model configurations
...
Add 6 models served by Novita:
- deepseek/deepseek-r1-distill-qwen-14b
- deepseek/deepseek-r1-distill-qwen-32b
- google/gemma-3-12b-it
- inclusionai/ling-2.6-1t
- meta-llama/llama-3.2-3b-instruct
- moonshotai/kimi-k2.6
Capabilities, pricing, context, and modalities sourced from Novita's
/v1/models API; family slugs and release dates aligned with existing
same-model entries in the repo.
2026-04-23 14:58:43 +08:00
Frank
e0e153e8d6
update zen models
2026-04-23 02:51:28 -04:00
mickalchen
a0fbef93e4
revert
2026-04-23 14:41:05 +08:00
mickalchen
e4a89229be
add model by openrouter
2026-04-23 14:23:06 +08:00
mickalchen
9ad15e97cd
add model by openrouter
2026-04-23 14:18:51 +08:00
James
e3b4df4e87
Update Kimi-K2.6-TEE.toml
...
fix reasoning
2026-04-23 13:53:37 +08:00
Aiden Cline
e3e2066c83
Merge pull request #1544 from GodTamIt/deepinfra/kimi-k2.6
...
deepinfra: Add Kimi-K2.6 support
2026-04-23 00:40:50 -04:00
Aiden Cline
208dcd12a2
Merge pull request #1533 from qychen2001/dev
...
Add Kimi-K2.6 and Qwen3.6-35B-A3B, update Kimi-K2.5 config for siliconflow and siliconflow-cn
2026-04-23 00:40:40 -04:00
Aiden Cline
d166444aa1
Merge pull request #1541 from zainhas/dev
...
[Together AI] add Kimi k2.6 support
2026-04-23 00:40:15 -04:00
Aiden Cline
cd35f96e73
Merge pull request #1528 from seffhunnn/dev
...
Fix incorrect model ID for Gemma 4 26B (Google provider)
2026-04-23 00:39:12 -04:00
Christopher Tam
2b0c21e0b4
deepinfra: Add Kimi-K2.6 support
2026-04-22 23:39:42 -04:00
James
787b0bf9e9
Update Kimi K2.5 TEE to Kimi K2.6 TEE
2026-04-23 10:58:39 +08:00
Zain Hasan
77fbf02ff6
remove interleaved
2026-04-22 16:19:10 -07:00
Zain Hasan
89973fd122
[Together AI] add Kimi k2.6 support
2026-04-22 16:16:08 -07:00
Frank
e458a9f5b8
Merge pull request #1540 from dsingal0/feat/baseten-kimi-k2.6
...
feat(baseten): add Kimi K2.6
2026-04-22 16:59:51 -04:00
Dhruv Singal
61d95afae2
feat(baseten): add Kimi K2.6
2026-04-22 13:53:46 -07:00
Jack
4a2df5e008
Merge pull request #1537 from anomalyco/feat/opencode-go-mimo-v2.5
...
Feat/opencode go mimo-v2.5-pro & mimo-v2.5
2026-04-23 00:51:49 +08:00
Aiden Cline
3db907f3ce
Merge pull request #758 from regolo-ai/dev
...
Add Regolo-ai Provider
2026-04-22 12:33:02 -04:00
Jack
70d8f9cc6e
update mimo v2 output limits to 128k
2026-04-22 23:32:16 +08:00
Jack
b9b354ada0
update mimo v2.5 output limits to 128k
2026-04-22 23:17:40 +08:00
Jack
95632ad376
remove mimo-v2.5-omni (renamed to mimo-v2.5)
2026-04-22 23:15:00 +08:00
Philip M
7153506989
Adds support for openrouter/pareto-code
...
The Pareto Router is a way to have OpenRouter always pick a strong coding model for your needs without committing to a specific one. You express a single min_coding_score preference between 0 and 1, and the router routes your request to a coding model that meets that bar.
The Pareto Router is tuned for coding use cases. Under the hood it keeps a curated shortlist of strong coding models currently available on OpenRouter. The exact shortlist and selection logic evolve over time as new models land and benchmarks shift.
2026-04-22 10:14:46 -05:00
Jack
c73bab2e7e
providers(opencode-go): rename mimo-v2.5-omni to mimo-v2.5
2026-04-22 23:14:27 +08:00
Jack
a783dc808d
providers(opencode-go): add mimo v2.5 models and separate v2 families
2026-04-22 23:07:40 +08:00
Daniele Scasciafratte
58e72802bc
feat(models): update
2026-04-22 16:05:48 +02:00
Mohd Saif
19233d93c4
fix: remove unnecessary id field
2026-04-22 15:19:47 +05:30
QiyuanChen
7eea45e078
feat(siliconflow-cn): add Kimi-K2.6, Qwen3.6-35B-A3B and update Kimi-K2.5 config
2026-04-22 13:38:34 +08:00
QiyuanChen
001ec226f6
feat(siliconflow): add Kimi-K2.6 and update Kimi-K2.5 config
2026-04-22 13:37:59 +08:00
Jack
32461d5b44
Merge pull request #1532 from chl-0537/feature/add-tencent
...
Remove tencent token plan
2026-04-22 13:12:06 +08:00
mickalchen
9f078294c0
Remove tencent token plan
2026-04-22 13:07:03 +08:00
Aiden Cline
c885ed49cd
Merge pull request #1531 from zhiyuan1024/zhiyuan/alibaba-cn_kimi-k2.6
...
feat(alibaba-cn): add Kimi K2.6 model configuration
2026-04-21 23:41:29 -04:00
Aiden Cline
2fc434062f
Merge pull request #1529 from compumike/compumike/fix-openrouter-openai-gpt-5.4-pricing
...
Fix pricing for openrouter/openai gpt-5.4-[mini,nano] off by 10^6
2026-04-21 23:40:51 -04:00
Aiden Cline
dbcb7e6d69
Merge pull request #1530 from cgilly2fast/dev
...
feat(firmware): kimi k2.6 model
2026-04-21 23:40:22 -04:00
Zhiyuan Hou
b58392fc62
feat(alibaba-cn): add Kimi K2.6 model configuration
...
Signed-off-by: Zhiyuan Hou <zhiyuan2048@outlook.com >
2026-04-22 10:37:37 +08:00
Frank
a4818c90ca
update zen models
2026-04-21 20:18:35 -04:00
Colby Gilbert
49fbdba49f
feat(firmware): kimi k2.6 model
2026-04-21 17:14:46 -07:00
Mike Robbins
f2dd4da7f9
Fix pricing for openrouter/openai gpt-5.4-[mini,nano] off by 10^6
2026-04-21 18:41:02 -04:00
Frank
a3ed215038
update zen models
2026-04-21 17:43:38 -04:00
Mohd Saif
d45df0530b
fix: correct Gemma 4 26B model ID for Google provider
...
Updated model ID from gemma-4-26b-it to gemma-4-26b-a4b-it to match actual Gemini API. Also added missing id field and renamed the file accordingly.
2026-04-22 01:51:23 +05:30
Jack
990531258b
Merge pull request #1527 from anomalyco/feat/opencode-go-kimi-k2.6-3x-name
...
providers(opencode-go): rename kimi k2.6
2026-04-21 22:59:26 +08:00
Jack
cd2c9e3b62
providers(opencode-go): rename kimi k2.6
2026-04-21 22:54:53 +08:00
Aiden Cline
a114991278
Merge pull request #1505 from rocuevas9511/feat/deepinfra-qwen-3.5-35b
...
feat: add Qwen 3.5 35B A3B to deepinfra
2026-04-21 10:02:27 -04:00
Aiden Cline
5ff2035cee
Merge pull request #1520 from Marenz/add-deepinfra-qwen3.6-35b-a3b
...
Add Qwen3.6-35B-A3B to Deep Infra
2026-04-21 10:01:56 -04:00
Aiden Cline
a5993cc140
Merge pull request #1515 from llc1123/chore/zenmux-update
...
providers(zenmux): add support for kimi k2.6
2026-04-21 10:00:06 -04:00
Aiden Cline
aebe4b6cd0
Merge pull request #1517 from otterDeveloper/kimi2.6-pull
...
add Firework's kimi k2.6
2026-04-21 09:59:54 -04:00
Aiden Cline
214adb1154
Merge pull request #1522 from ceoAppsknight/kilo/kimi-k2.6
...
Added kilo/kimi-k2.6
2026-04-21 09:59:33 -04:00
Aiden Cline
b301c1f8b6
Merge pull request #1523 from sk0x0y/feature/nanogpt-kimi-k2.6-qwen-3.6
...
feat(nano-gpt): add Kimi K2.6 and Qwen 3.6 models
2026-04-21 09:58:45 -04:00
Aiden Cline
b81c8b385b
Merge pull request #1507 from rocuevas9511/feat/deepinfra-qwen-3.5-397b
...
feat: add Qwen 3.5 397B A17B to deepinfra
2026-04-21 09:58:28 -04:00
rocuevas9511
1caa438b3e
fix: remove id field (per Marenz feedback)
2026-04-21 07:35:52 -06:00
rocuevas9511
673bc92f4d
fix: remove id field (per Marenz feedback)
2026-04-21 07:35:38 -06:00
Jack
0159eaa158
Merge pull request #1526 from anomalyco/feat/moonshotai-cn-kimi-k2.6
...
providers(moonshotai-cn): add kimi k2.6
2026-04-21 20:39:05 +08:00
Jack
efdc7b9a54
providers(moonshotai-cn): add kimi k2.6
2026-04-21 20:26:19 +08:00
Jack
b08721206d
Merge pull request #1525 from anomalyco/feat/moonshotai-kimi-k2.6
...
providers(moonshotai): add kimi k2.6
2026-04-21 19:35:13 +08:00
Jack
fe0d4cd9fa
providers(moonshotai): add kimi k2.6
2026-04-21 19:33:06 +08:00
sk0x0y
738ad6ed70
feat(nano-gpt): add Kimi K2.6 and Qwen 3.6 model family
2026-04-21 19:22:38 +09:00
Syed Assadullah Shah
19b66fbaf2
Added kilo/kimi-k2.6
2026-04-21 15:19:17 +05:00
Mathias L. Baumann
c416836497
Add Qwen3.6-35B-A3B to Deep Infra
...
35B-total / 3B-active MoE (256 experts, 8 routed + 1 shared).
262K native context, vision + video input, thinking mode, tool calls.
Apache 2.0, $0.20 in / $1.00 out per 1M tokens.
2026-04-21 11:58:04 +02:00
rocuevas9511
059cc4b91f
fix: update model id to match DeepInfra API
2026-04-21 00:33:20 -06:00
rocuevas9511
116bd328a6
fix: update model id to match DeepInfra API
2026-04-21 00:31:22 -06:00
Frank
aa30ce3ef2
update zen models
2026-04-21 02:12:27 -04:00
Frank
9533a47906
update zen models
2026-04-21 01:20:47 -04:00
Miguel Medina
ad18f780f0
add firework's kimi 2.6
2026-04-20 23:12:11 -06:00
粒粒橙
a48519e557
providers(zenmux): add support for kimi k2.6
2026-04-21 10:13:22 +08:00
Aiden Cline
23f5e74392
Merge pull request #1514 from mfbalestra/add/kimi-k2.6-ollama-cloud
...
providers/ollama-cloud: add kimi-k2.6:cloud
2026-04-20 21:47:51 -04:00
Aiden Cline
7a50ea28a1
Merge pull request #1510 from dpuyosa/feat/venice-add-kimi-k2-6
...
Venice: Add Kimi K2.6 model configuration
2026-04-20 21:45:49 -04:00
mfbalestra
5944f94197
providers(ollama-cloud): add kimi-k2.6:cloud
2026-04-20 22:45:01 -03:00
Aiden Cline
98732b7d76
Merge pull request #1512 from SomeoneWithOptions/dev
...
add kimi-K2.6 for OpenRouter provider
2026-04-20 21:44:39 -04:00
SomeoneWithOptions
c6412d7e59
add kimi-K2.6 for OpenRouter provider
2026-04-20 18:59:49 -05:00
dpuyosa
b5a7a6e974
[venice] Add Kimi K2.6 model configuration
...
- Add new model definition for Venice provider
- Include cost, limits, and modality specs
- Enable reasoning, tool calling, and image input
2026-04-21 00:40:19 +02:00
Aiden Cline
ba7c3d7b0b
Merge pull request #1509 from kostiak/patch-1
...
Add support for Kimi-K2.6 in Kimi For Coding provider
2026-04-20 18:02:04 -04:00
Aiden Cline
53a9a2a36d
Merge pull request #1508 from hanouticelina/add-kimi-k2.6-modeling
...
feat(huggingface): add Kimi K2.6
2026-04-20 18:01:32 -04:00
kostiak
3b6bca9b90
Add support for Kimi-K2.6 for Kimi For Coding provider
2026-04-21 00:24:51 +03:00
Celina Hanouti
f5a048060f
add support for Kimi-K2.6 for Hugging Face provider
2026-04-20 21:41:14 +01:00
rocuevas9511
9e6178a5e3
fix: update cost for Qwen 3.5 397B A17B
2026-04-20 13:38:45 -06:00
rocuevas9511
e4150361b3
fix: update cost for Qwen 3.5 35B A3B
2026-04-20 13:38:06 -06:00
rocuevas9511
8de0fc059d
add Qwen 3.5 397B A17B to deepinfra
2026-04-20 13:34:56 -06:00
rocuevas9511
054733884f
add Qwen 3.5 35B A3B to deepinfra
2026-04-20 13:32:54 -06:00
Aiden Cline
3d09981eda
Merge pull request #1499 from rovo89/patch-1
...
[google] Fix cache_read cost in gemini-2.5-flash model
2026-04-20 14:43:53 -04:00
Aiden Cline
6877af7770
Merge pull request #1501 from mchenco/kimi-k2.6
...
Add Kimi K2.6 to Workers AI and AI Gateway
2026-04-20 14:41:58 -04:00
Nacho F. Lizaur
802985f76c
feat: update kiro provider to use kiro-acp-ai-provider, add opus 4.7
2026-04-20 20:13:43 +02:00
mchen
794993fd48
Add Kimi K2.6 to Workers AI and AI Gateway
2026-04-20 13:54:31 -04:00
Jack
00b53a422a
separate opencode-go kimi k2 families
2026-04-21 01:00:50 +08:00
Jack
2ccecb6011
Merge pull request #1500 from chl-0537/feature/add-tencent
...
feat: rename model
2026-04-20 22:49:09 +08:00
mickalchen
9b7e3abf00
rename model
2026-04-20 22:02:37 +08:00
Robert Vollmer
27d6a3d503
[google] Fix cache_read cost in gemini-2.5-flash model
...
https://ai.google.dev/gemini-api/docs/pricing#gemini-2.5-flash
There's no cache_read_audio, is there?
2026-04-20 15:02:54 +02:00
Lyda
6beb1f8be2
feat(302ai): standardize Claude model metadata and capabilities
...
- Add family field for all Claude models (claude-haiku, claude-opus, claude-sonnet)
- Standardize knowledge cutoff dates to full date format (YYYY-MM-DD)
- Enable reasoning capability for Claude Opus 4.x and Sonnet 4.x series models
- Add PDF input modality support for claude-opus-4-1-20250805
- Update claude-opus-4-7 context limit to 1,000,000 tokens
2026-04-20 16:10:52 +08:00
Lyda
e0c1124fe2
feat(302ai): update GPT model capabilities and specifications
...
- Add structured_output capability for GPT-4.1, GPT-4o, and GPT-5 series models
- Enable reasoning capability and disable temperature for GPT-5 series models
- Update context limits: GPT-4.1 series to 1,047,576 tokens, GPT-5.4 series to 1,050,000 tokens
- Add input token limits for GPT-5 series models (272,000 or 922,000 tokens)
- Update knowledge cutoffs across GPT-5 series (2024-05-30 to 2025-08-31)
- Add PDF input modality support for GPT-4
2026-04-20 15:58:32 +08:00
Lyda
bf4ceb7baa
feat(302ai): update GLM model capabilities and knowledge cutoffs
...
- Enable reasoning capability for GLM-4.5-air, GLM-4.5, GLM-4.5V, and GLM-4.6V models
- Update knowledge cutoff to 2025-04 for GLM-4.5, GLM-4.5V, GLM-4.6, GLM-4.6V, and GLM-4.7
- Add video input modality support for GLM-4.5V and GLM-4.6V
- Add structured_output capability for GLM-5-turbo and GLM-5.1
- Add interleaved reasoning_content field for GLM-4.7, GLM-5, GLM-5-turbo, GLM-5.1, and GLM-5V-turbo
2026-04-20 15:48:29 +08:00
Aiden Cline
1a41934e55
Merge pull request #1448 from Lydanne/dev
...
feat(302ai): supplement commonly missing models
2026-04-19 22:19:18 -05:00
Aiden Cline
aeb4caec9f
Merge pull request #1488 from Sewer56/add-wafer-provider
...
Add wafer.ai provider
2026-04-19 22:19:06 -05:00
Aiden Cline
ccb8dcc65f
Merge pull request #1493 from rovo89/patch-1
...
Add context_over_200k for gemini-2.5-pro and adjust cache_read costs
2026-04-19 22:17:45 -05:00
Lyda
435ec1df7b
feat(302ai): add claude-opus-4-7 model
2026-04-20 11:14:27 +08:00
Robert Vollmer
7c9d609143
Add context_over_200k for gemini-2.5-pro and adjust cache_read costs
...
https://ai.google.dev/gemini-api/docs/pricing#gemini-2.5-pro
https://cloud.google.com/vertex-ai/generative-ai/pricing#gemini-models-2.5 (rounds 0.125 to 0.13)
2026-04-20 00:09:11 +02:00
Aiden Cline
add7947164
Merge pull request #1492 from dpuyosa/feat/venice-add-gemma4-uncensored
...
Venice: Add Gemma 4 and Venice Uncensored 1.2 models
2026-04-19 16:52:28 -05:00
Aiden Cline
dd0c1af12a
Merge pull request #1491 from dpuyosa/chore/venice-pricing-update
...
Venice: Update pricing for Grok 4.20 and Qwen3.5 9B
2026-04-19 16:52:15 -05:00
Aiden Cline
0c93cc03be
Merge pull request #1485 from anomalyco/more-extends-cases
...
migrate more providers to extends format
2026-04-19 16:51:58 -05:00
Aiden Cline
62e25b73f8
Merge branch 'dev' into more-extends-cases
2026-04-19 16:47:05 -05:00
Aiden Cline
17093e0031
Merge pull request #1489 from berget-ai/update/berget-prices-gemma4
...
chore: update berget.ai models - prices and Gemma 4
2026-04-19 16:42:06 -05:00
Aiden Cline
ae542978f0
Merge pull request #1490 from BlockListed/cortecs-add-claude-opus-4-7
...
add claude opus 4.7
2026-04-19 16:41:36 -05:00
dpuyosa
7f16117bba
[venice] Add Gemma 4 and Venice Uncensored 1.2 models
...
- Add Gemma 4 Uncensored with 256K context, image support
- Add Venice Uncensored 1.2 with 128K context, image support
- Both models support tool calls and structured output
2026-04-19 23:41:17 +02:00
dpuyosa
2b96a2d3d6
[venice-models] Update pricing for Grok 4.20 and Qwen3.5 9B
...
- Update cache_read pricing for Grok 4.20 context_over_200k (0.23 → 0.45)
- Update input cost for Qwen3.5 9B (0.05 → 0.1)
2026-04-19 23:38:20 +02:00
BlockListed
9799a841c6
add claude opus 4.7
...
yes this model id is correct, cortecs is weird.
2026-04-19 23:13:58 +02:00
Christian Landgren
71c59b4235
chore: update berget.ai models - prices and Gemma 4
...
- Add Google Gemma 4 31B Instruct model
- Update prices for existing models (EUR to USD conversion)
- Remove non-coding models (bge-reranker, multilingual-e5 embeddings, kb-whisper)
- Remove deprecated Llama-3.1-8B-Instruct
Updated models:
- GLM-4.7: 0.77/2.75 USD/M (was 0.7/2.3)
- Llama-3.3-70B: 0.99/0.99 USD/M (was 0.9/0.9)
- Mistral-Small-3.2: 0.33/0.33 USD/M (was 0.3/0.3)
- GPT-OSS-120B: 0.44/0.99 USD/M (was 0.3/0.9)
New models:
- Gemma-4-31B-it: 0.275/0.55 USD/M
Removed models (not relevant for coding):
- BAAI/bge-reranker-v2-m3 (reranker)
- intfloat/multilingual-e5-large/* (embeddings)
- KBLab/kb-whisper-large (speech-to-text)
- meta-llama/Llama-3.1-8B-Instruct (deprecated)
2026-04-19 12:19:32 +02:00
Sewer56
0568b412aa
Add wafer.ai provider with GLM-5.1 and Qwen3.5-397B-A17B models
2026-04-19 03:02:04 +01:00
Aiden Cline
812612465a
Merge pull request #1484 from smakosh/feat/llmgateway-new-models
...
feat: update LLM Gateway to 182 models
2026-04-18 18:37:26 -05:00
Aiden Cline
ac8a79dd74
Merge pull request #1487 from WJQSERVER/add/nvidia(nim)-z-ai-glm-5.1
...
Add Z.AI GLM-5.1 to NVIDIA(NIM)
2026-04-18 18:36:55 -05:00
smakosh
9460d981d4
fix: replace broken extends with concrete glm-4.6v-flash def
...
zhipuai/glm-4.6v-flash.toml is a symlink to zai/models/glm-4.6v-flash.toml
which does not exist, causing validate to fail with 'Unable to resolve
extends.from'. Inline the concrete definition instead.
2026-04-18 13:36:52 +02:00
smakosh
91db9d81ec
feat: update LLM Gateway to 182 models
...
- Uses extends to reference canonical providers where
possible (116 models), keeping 66 full definitions
- Only includes active text/chat models
- Adds new models: Claude Opus 4.7, Grok 4 Fast,
Kimi K2, Mimo V2, GLM 5.1, Qwen 3 Coder, and more
Co-Authored-By: Claude Opus 4.6 (1M context) <noreply@anthropic.com >
2026-04-18 13:20:41 +02:00
Jack
d8d5f08df6
Merge pull request #1486 from chl-0537/feature/add-tencent
...
feat: add new provider and model
2026-04-18 17:30:53 +08:00
WJQSERVER
11d5443ab1
follow the nim modelcard change context length to 131072
...
https://build.nvidia.com/z-ai/glm-5.1/modelcard
Other Properties Related to Input: Supports multi-turn conversations, tool calling, system prompts, and extended agentic sessions. Input context length: 131,072 tokens.
2026-04-18 16:48:29 +08:00
wjqserver
90558e9eed
add glm-5.1
2026-04-18 16:41:02 +08:00
Aiden Cline
cb7d258e33
migrate more providers to extends format
2026-04-17 23:02:34 -05:00
Frank
2af43dc4f8
update zen models
2026-04-17 19:08:05 -04:00
Aiden Cline
93ddb6b131
Merge pull request #1482 from sopial42/ovhcloud/update-models-clean
...
chore(ovhcloud): remove 3 models no longer available in AI Endpoints
2026-04-17 16:50:53 -05:00
Aiden Cline
04bf671f18
Merge pull request #1481 from Spherrrical/add-digitalocean-provider
...
feat(provider): add DigitalOcean provider
2026-04-17 16:49:44 -05:00
Aiden Cline
1f3ba4ba21
Merge pull request #1483 from anomalyco/add-extends-support
...
feat: add extends support
2026-04-17 16:49:23 -05:00
Aiden Cline
305bdb6cdc
Merge branch 'dev' into add-extends-support
2026-04-17 16:22:41 -05:00
Aiden Cline
2e5b4b44ae
update some modes
2026-04-17 16:22:14 -05:00
aadhondt
bb1c08dd41
chore(ovhcloud): remove 3 models no longer available in AI Endpoints
...
- deepseek-r1-distill-llama-70b
- mixtral-8x7b-instruct-v0.1
- qwen2.5-coder-32b-instruct
Co-Authored-By: Claude Sonnet 4.6 <noreply@anthropic.com >
2026-04-17 22:38:35 +02:00
Spherrrical
86d0f14397
feat(digitalocean): add DigitalOcean Gradient AI Platform provider
...
Adds the DigitalOcean provider with 46 models (Anthropic, OpenAI,
Arcee, fal, and DO-hosted open-source/embedding models) served via
the OpenAI-compatible endpoint at https://inference.do-ai.run/v1 .
2026-04-17 13:07:47 -07:00
Aiden Cline
f93c1a8998
Merge pull request #1478 from Kaspazza/dev
...
Add Github Copilot Claude Opus 4.7
2026-04-17 15:01:04 -05:00
Aiden Cline
27a6758eb7
new gen script
2026-04-17 14:57:57 -05:00
Aiden Cline
9dbafb81fa
add script
2026-04-17 14:57:41 -05:00
Aiden Cline
419f8a3a20
add migration checker script
2026-04-17 14:57:32 -05:00
Aiden Cline
785a091073
Merge pull request #1480 from nicocasaisd/openai/remove-deprecated-codex-mini-latest
...
chore(openai): remove deprecated model codex-mini-latest
2026-04-17 14:52:34 -05:00
nicocasaisd
ad2409dce1
chore(openai): remove deprecated model codex-mini-latest
2026-04-17 15:35:43 -03:00
kaspazza
63450ab67d
Add Github Copilot Claude Opus 4.7
2026-04-17 20:19:36 +02:00
Aiden Cline
4002cb6739
model
2026-04-17 13:11:51 -05:00
Aiden Cline
72fa27a91e
Merge pull request #1474 from vglafirov/add-gitlab-duo-chat-opus-4-7
...
feat(gitlab): add duo-chat-opus-4-7 model definition
2026-04-17 12:37:36 -05:00
Aiden Cline
ddb3a0ff05
update agents.md
2026-04-17 12:14:04 -05:00
Aiden Cline
96c12042bc
remeda
2026-04-17 12:13:54 -05:00
Misha Skvortsov
7336b3619c
atomic-chat: add provider with initial blessed models
...
Adds Atomic Chat as a local OpenAI-compatible provider at
http://127.0.0.1:1337/v1 . Includes logo and three curated models:
- unsloth/Qwen3.5-9B-IQ4_XS (id: Qwen3_5-9B-IQ4_XS)
- unsloth/gemma-4-E4B-it-IQ4_XS (id: gemma-4-E4B-it-IQ4_XS)
- unsloth/MiniMax-M2.5-UD-TQ1_0 (id: MiniMax-M2_5-UD-TQ1_0)
Model ids match the normalized form returned by Atomic Chat's
/v1/models endpoint (dots replaced with underscores).
Made-with: Cursor
2026-04-17 13:03:53 +03:00
mickalchen
73b81ac027
add tencent provider
2026-04-17 17:27:35 +08:00
Vladimir Glafirov
9c5839a414
feat(gitlab): add duo-chat-opus-4-7 model definition
2026-04-17 08:54:07 +02:00
Aiden Cline
721464bc3c
Merge pull request #1469 from GrahamCampbell/ops-4-7-fixes
...
Corrected and normalized claude opus 4.7 knowledge cut-off dates
2026-04-16 22:30:13 -05:00
Aiden Cline
b6b45a9d25
Merge pull request #1463 from GrahamCampbell/claude-4-6
...
Correct Anthropic Claude 4.6 model knowledge cut-off dates
2026-04-16 21:36:19 -05:00
Aiden Cline
7b8f98bb23
Merge pull request #1471 from cfbender/fix/openrouter-opus-4-7
...
feat: add openrouter opus 4.7
2026-04-16 21:35:52 -05:00
Aiden Cline
92ac48b07f
Merge pull request #1473 from fhennerkes/dev
...
Poe: add Claude-Opus-4.7
2026-04-16 20:57:21 -05:00
fhennerkes
36a455ce4a
Merge branch 'anomalyco:dev' into dev
2026-04-16 18:11:32 -07:00
fhennerkes
0bf5c60319
poe: add Claude-Opus-4.7 model
...
Add new Anthropic model from Poe API (released 2026-04-15):
- Reasoning support
- 1M context window with 128K output
- Cost: $4.3/M input, $21/M output, $0.43/M cache read, $5.4/M cache write
- Modalities: text, image, pdf
2026-04-16 18:08:02 -07:00
Kit Langton
b123711494
Merge pull request #1472 from elithrar/patch-4
...
cloudflare: add opus 4.7
2026-04-16 19:38:46 -04:00
Matt Silverlock
832064c1d1
cloudflare: add opus 4.7
2026-04-16 18:51:54 -04:00
Cody Bender
c16e1c817b
fix: add openrouter opus 4.7
2026-04-16 18:38:36 -04:00
Aiden Cline
8aaf31711b
Merge pull request #1468 from heimoshuiyu/fix/opus-4-7-temperature
...
fix: set temperature=false for Claude Opus 4.7
2026-04-16 14:34:52 -05:00
Graham Campbell
ca6acf0b3e
Corrected and normalized claude opus 4.7 knowledge cut-off dates
2026-04-16 20:33:49 +01:00
heimoshuiyu
660a672647
fix: set temperature=false for firmware and venice Opus 4.7
2026-04-17 03:23:11 +08:00
heimoshuiyu
147cb3138a
fix: set temperature=false for Claude Opus 4.7 across all providers
2026-04-17 03:22:33 +08:00
Aiden Cline
5c1fb729fd
Merge pull request #1462 from dpuyosa/dev
...
Venice: Add Claude Opus 4.7 and remove deprecated models
2026-04-16 14:10:03 -05:00
Aiden Cline
0d34900078
Merge pull request #1464 from cgilly2fast/dev
...
feat(firmware): opus 4.7 remove old claude models
2026-04-16 14:09:39 -05:00
Aiden Cline
4bc71918f5
Merge pull request #1467 from vercel/update-vercel-models-20260416-1812
...
Update Vercel models
2026-04-16 13:34:40 -05:00
Jerilyn Zheng
6f62522505
Set temperature to false in claude-opus-4.7 configuration
...
Changed temperature setting from true to false.
2026-04-16 11:23:29 -07:00
Aiden Cline
7680a1d169
Merge pull request #1466 from vercel/fix-reranking-type-upstream
...
fix(vercel): accept reranking model type from API
2026-04-16 13:22:29 -05:00
github-actions[bot]
bdc15a57ce
chore(vercel): update Vercel model definitions
...
Auto-generated by weekly workflow from Vercel AI Gateway API.
Co-Authored-By: github-actions[bot] <41898282+github-actions[bot]@users.noreply.github.com>
2026-04-16 18:12:52 +00:00
R-Taneja
c7d324fed8
fix(vercel): accept reranking model type from API
...
The Vercel AI Gateway API now returns models with type "reranking",
which caused the generate-vercel script to fail schema validation.
Add "reranking" to the ModelType enum and skip these models in the
main loop, matching the existing pattern for image/video types that
OpenCode does not consume.
2026-04-16 11:05:26 -07:00
Colby Gilbert
863bada2cf
feat(firmware): opus 4.7 remove old claude models
2026-04-16 10:37:52 -07:00
dpuyosa
28c4af0631
[venice] Add Claude Opus 4.7 and update Qwen models
...
- Add Claude Opus 4.7 with 1M context, 128K output, multimodal support
- Update Qwen 3.5 35B and 397B to open_weights=true and refresh last_updated
- Remove deprecated models: Grok Code Fast 1, Mercury Edit 2, MiniMax M2.1
2026-04-16 18:30:45 +02:00
Aiden Cline
38f6b7dfc5
Merge pull request #1456 from Snat3r/dev
...
Add GML5.1.toml model to cortecs provider
2026-04-16 11:16:51 -05:00
Aiden Cline
1ba010f66e
Merge pull request #1461 from llc1123/chore/zenmux-update
...
chore(zenmux): add claude-opus-4.7
2026-04-16 11:16:41 -05:00
粒粒橙
132a5ade8d
chore(zenmux): add claude-opus-4.7
2026-04-17 00:02:58 +08:00
Graham Campbell
2ef2b10847
Correct anthropic 4.6 knowledge cut-off dates
2026-04-16 16:52:15 +01:00
Frank
87e1dcb70f
update zen modles
2026-04-16 11:31:20 -04:00
Aiden Cline
43d2e058d9
Merge pull request #1449 from shikbupt/alibaba-glm5.1
...
add alibaba-cn glm5.1
2026-04-16 10:25:58 -05:00
Aiden Cline
0a6c2ca49d
Merge pull request #1459 from itsnebulalol/dev
...
feat: add anthropic claude opus 4.7 models
2026-04-16 10:25:04 -05:00
Dominic Frye
2f954fc838
feat: add anthropic claude opus 4.7 models
2026-04-16 11:19:19 -04:00
Frank
91b7851971
update zen models
2026-04-16 04:51:30 -04:00
Jack
80e23ab90e
Merge pull request #1266 from lioZ129/feature/add-hpc-ai-provider
...
feat: add HPC-AI model provider support
2026-04-16 15:09:44 +08:00
Snat3r
6e574da2e4
Update input modalities in minimax-M2.7.toml
2026-04-16 08:21:20 +02:00
Snat3r
5b68565783
Add MiniMax-M2.7 model configuration file cortecs
2026-04-16 08:19:51 +02:00
lioZ129
65dafb56a3
add [cost] and glm5.1 support
2026-04-16 14:13:27 +08:00
Snat3r
32c0c88600
Update context and output limits in glm-5.1.toml
2026-04-16 08:08:09 +02:00
Snat3r
d911b6f610
Add GLM-5.1 model configuration file
2026-04-16 08:06:37 +02:00
Aiden Cline
5cf28a566f
Merge pull request #1424 from WJQSERVER/feat/nvidia-minimax-m2.7
...
Add MiniMax M2.7 to NVIDIA(NIM)
2026-04-15 20:19:01 -05:00
Aiden Cline
52cdb783a1
Merge pull request #1452 from wwth8819/dev
...
For aihubmix add GPT-5.4 \ GPT-5.4-mini, remove Incorrect value from old models, update cost
2026-04-15 20:18:24 -05:00
wwth8819
4a14ce5ae6
Remove temperature setting from gpt-5.2-codex.toml
...
Removed the temperature setting from the configuration.
2026-04-16 03:12:14 +08:00
wwth8819
f17c352027
Update cost values in coding-glm-4.7.toml
2026-04-16 03:11:16 +08:00
wwth8819
46ec19ab02
add gpt-5.4-mini gpt-5.4
2026-04-16 03:09:26 +08:00
Frank
f12aae094e
update zen models
2026-04-15 10:55:02 -04:00
sk
377d0f1c8d
add alibaba-cn glm5.1
2026-04-15 22:02:52 +08:00
WJQSERVER
884b799012
Merge branch 'dev' into feat/nvidia-minimax-m2.7
2026-04-15 21:53:54 +08:00
Lyda
ac7e35af4e
feat(302ai): supplement commonly missing models
2026-04-15 17:32:18 +08:00
Frank
6f04d267cf
update zen models
2026-04-15 02:17:31 -04:00
Aiden Cline
a0b89e739b
Merge pull request #1425 from ceyhanmolla/add-minimax-m2.7-nvidia
...
feat(nvidia): add MiniMax-M2.7
2026-04-14 22:59:38 -05:00
Frank
0ce000a521
update go models
2026-04-14 23:09:06 -04:00
Frank
4b7dda6cc6
update go models
2026-04-14 22:50:12 -04:00
Aiden Cline
7220310828
Merge pull request #1447 from Sawyerb/dev
...
Removed deprecated models and added ME2 to all relevant providers.
2026-04-14 21:49:22 -05:00
Aiden Cline
15746b9845
Merge pull request #1444 from wwth8819/dev
...
Add glm-5.1, coding-glm-5.1 TO AiHubMix
2026-04-14 21:49:06 -05:00
Aiden Cline
e9be42b4bc
Merge pull request #1443 from teodortomas/add-minimax-m2.7
...
add minimax-m2p7 to fireworks-ai provider
2026-04-14 17:11:09 -05:00
Aiden Cline
c0d21d802f
Merge pull request #1429 from Lee-Si-Yoon/fix/cache-read-friendli
...
fix: cache read costs for friendliAI models
2026-04-14 17:10:56 -05:00
Aiden Cline
4937952a52
Merge pull request #1437 from Ardakilic/dev
...
feat: kilo gateway: elephant alpha
2026-04-14 17:10:37 -05:00
Aiden Cline
1bb5deadea
Merge pull request #1439 from fhennerkes/dev
...
poe: update models with pricing, deprecations, and display name fixes
2026-04-14 17:10:26 -05:00
Aiden Cline
c2ad18c87d
Merge pull request #1445 from cantalupo555/chore/openrouter-remove-deprecated-free-models
...
chore(openrouter): remove deprecated free-tier models no longer available via API
2026-04-14 17:10:01 -05:00
Sawyer
e4b0a53e26
Removed deprecated models and added ME2 to all relevant providers.
2026-04-14 12:33:35 -07:00
cantalupo555
61ea2e2093
chore(openrouter): remove deprecated free-tier models no longer available via API
2026-04-14 08:11:45 -03:00
wwth8819
162fd8b72b
Add configuration for Coding-GLM-5.1 model
2026-04-14 17:35:23 +08:00
wwth8819
b8fae048db
Add GLM-5.1 model configuration file
2026-04-14 17:32:59 +08:00
Teodor Tomáš
ca6cf794c1
add minimax-m2p7 to fireworks-ai provider
2026-04-14 11:06:47 +02:00
fhennerkes
152b401976
poe: update models with pricing, deprecations, and display name fixes
...
Mark 11 models no longer on the Poe API as deprecated:
- anthropic: claude-sonnet-3.5, claude-sonnet-3.5-june
- cerebras: qwen3-235b-2507-cs, qwen3-32b-cs, llama-3.3-70b-cs
- google: gemini-3-pro, gemini-deep-research
- novita: glm-4.7
- openai: chatgpt-4o-latest, gpt-4-classic-0314, gpt-4-classic
Update existing models with latest API data:
- Cerebras (gpt-oss-120b-cs, llama-3.1-8b-cs): Add pricing and correct context (128K)
- kimi-k2.5: Add pricing, fix display name, temperature/reasoning from API
- gpt-4o: Fix output formatting (8_192)
- gpt-5.1-codex-max: Fix display name
Co-Authored-By: Claude Sonnet 4.6 <noreply@anthropic.com >
2026-04-13 16:55:18 -07:00
Nacho F. Lizaur
8f340e1eb2
feat: update kiro provider to use kiro-ai-provider npm package
2026-04-13 23:12:11 +02:00
Aiden Cline
d99a1581df
Merge pull request #1435 from cantalupo555/feat/add-openrouter-elephant-alpha
...
feat(openrouter): add Elephant Alpha model
2026-04-13 15:01:55 -05:00
Nacho F. Lizaur
62da5b0cbd
feat: enable reasoning on Kiro Claude models
2026-04-13 19:49:25 +02:00
Nacho F. Lizaur
4e49abc10c
feat: add Kiro provider
2026-04-13 19:49:25 +02:00
Arda Kilicdagi
21cf1fb1e9
feat: kilo gateway: elephant alpha
2026-04-13 20:39:54 +03:00
cantalupo555
970a00c3ad
feat(openrouter): add Elephant Alpha model
2026-04-13 13:43:56 -03:00
cantalupo555
d611fc2ef7
feat(core): add elephant model family
2026-04-13 13:35:10 -03:00
Aiden Cline
7d7711878d
Merge pull request #1434 from cantalupo555/chore/openrouter-step-3.5-flash-free
...
chore(openrouter): remove Step 3.5 Flash free-tier model
2026-04-13 10:45:25 -05:00
cantalupo555
fc5d5ec67f
chore(openrouter): remove deprecated step-3.5-flash free-tier variant
...
- Remove the free-tier variant of Step 3.5 Flash as it is no longer needed
2026-04-13 11:51:10 -03:00
Aiden Cline
bd5050a15f
Merge pull request #1423 from spiffytech/dev
...
Add Synthetic support for GLM-5.1
2026-04-13 09:18:49 -05:00
Aiden Cline
ac67c648a2
Merge pull request #1422 from hanouticelina/add-minimax-2.7-huggingface
...
feat(huggingface): add MiniMax-M2.7
2026-04-13 09:18:35 -05:00
Aiden Cline
f58038fc21
Merge pull request #1421 from zainhas/dev
...
[Together AI]add minimax m2.7
2026-04-13 09:18:14 -05:00
Aiden Cline
3ea7d56b96
Merge pull request #1428 from dpuyosa/feat/venice-glm-5-models
...
Venice: Add Z-AI GLM-5 Turbo and GLM-5V Turbo models
2026-04-13 09:17:45 -05:00
Aiden Cline
e23657368e
Merge pull request #1426 from dpuyosa/venice/update-model-configs-0412
...
Venice: Update model configs and pricing
2026-04-13 09:17:33 -05:00
Aiden Cline
d59ad3ffa3
Merge pull request #1427 from dpuyosa/refactor/venice-model-naming
...
Venice: Update model naming convention
2026-04-13 09:17:08 -05:00
siyoon
dcbb415960
fix: add cache_read parameter to cost section
2026-04-13 16:05:36 +09:00
dpuyosa
e302956dc4
[venice] Update model naming convention
...
- Rename files to use dashes instead of dots
- Rename opus/sonnet 45 to 4-5 format
- Remove beta suffix from Grok 4.20 models
- Update family from grok-beta to grok
- Remove knowledge field from Claude models
- Update output limits
2026-04-13 01:10:06 +02:00
dpuyosa
5d100a8243
[venice] Add Z-AI GLM-5 Turbo and GLM-5V Turbo models
...
- Add Z-AI GLM-5 Turbo (text-only, reasoning, tool_call)
- Add Z-AI GLM-5V Turbo (vision, reasoning, tool_call)
2026-04-13 01:05:47 +02:00
dpuyosa
9ff52ee539
[venice] Update model configs and pricing
...
- Update last_updated dates for 5 models
- Set open_weights to false for 4 models
- Add context_over_200k cache pricing for qwen-3-6-plus
2026-04-13 00:59:21 +02:00
ceyhanmolla
8114522671
feat(nvidia): add MiniMax-M2.7
2026-04-12 16:56:04 +02:00
wjqserver
5e534d84bc
feat: add NVIDIA MiniMax M2.7 model
2026-04-12 22:51:46 +08:00
spiffytech
1f79cd8fd0
Added Synthetic support for GLM-5.1
2026-04-12 08:46:10 -04:00
Celina Hanouti
462b822e60
add minimax M2.7 for hugging face provider
2026-04-12 11:06:30 +02:00
Zain Hasan
8021e1afd0
Merge branch 'anomalyco:dev' into dev
2026-04-11 22:23:20 -07:00
Zain Hasan
304d85fa43
[Together AI]add minimax m2.7
2026-04-11 22:14:05 -07:00
Aiden Cline
f07262370f
Merge pull request #1418 from Ardakilic/dev
...
Kilo Gateway: Sync model list with upstream (2026-04-11)
2026-04-11 16:48:46 -05:00
Aiden Cline
7cfaa0393e
Merge pull request #1419 from nicopujia/add-openrouter-deepseek-r1
...
Add OpenRouter support for DeepSeek R1
2026-04-11 16:46:58 -05:00
Arda Kilicdagi
c93de9f54f
chore: sync all kilo gw models
2026-04-11 03:31:00 +03:00
Aiden Cline
2b17d9efc2
Merge pull request #1414 from Ardakilic/dev
...
Feat: Kilo Gateway: Add MiniMax M2.7
2026-04-10 13:31:50 -05:00
Arda Kilicdagi
8d8521e0e1
feat: kilo gateway: MiniMax M2.7
2026-04-10 20:42:45 +03:00
Aiden Cline
c9f22d26b9
Merge pull request #1413 from Ardakilic/dev
...
Kilo Gateway: Add GLM 5.1, Remove MiniMax M2.5 Free
2026-04-10 10:33:18 -05:00
Aiden Cline
61e8ea2a2e
Merge pull request #1410 from gjtiquia/dev
...
Poe: add GLM-5 model
2026-04-10 10:26:57 -05:00
Aiden Cline
7d2cd9818f
Merge pull request #1411 from Alex-wuhu/dev
...
novita-ai: add 8 new models and remove 2 deprecated models
2026-04-10 10:26:45 -05:00
Arda Kilicdagi
385b4fbcf6
feat: kilo gateway: glm-5.1
2026-04-10 15:11:37 +03:00
Alex-wuhu
2f0890c9a9
Add new model configurations for Gemma, MiniMax, Qwen, and GLM
2026-04-10 16:28:34 +08:00
GJ Tiquia
9d377b768f
poe: GLM-5 temperature set to true
2026-04-10 15:57:25 +08:00
GJ Tiquia
aa4f3f550a
poe: add GLM-5 model
2026-04-10 14:13:03 +08:00
Aiden Cline
f82d6fc61a
Merge pull request #1404 from nicopujia/add-openrouter-qwen3.5-flash-02-23
...
Add OpenRouter support for Qwen3.5 Flash 2026-02-23
2026-04-09 22:37:05 -05:00
Aiden Cline
4222b040b7
Merge pull request #1407 from line72/deepinfra-glm-5.1
...
[DeepInfra] Add GLM 5.1
2026-04-09 22:36:53 -05:00
Aiden Cline
d2e4174103
Merge pull request #1397 from qychen2001/dev
...
Add GLM-5.1 and GLM-5V-Turbo model configurations for siliconflow
2026-04-09 22:35:52 -05:00
Aiden Cline
f30b5fc754
Merge pull request #1408 from nanai10a/dev
...
Add MiniMax M2.5 (free) configuration file
2026-04-09 22:35:37 -05:00
Aiden Cline
10239c95e2
Merge pull request #1401 from dpuyosa/venice-open-weights-fix
...
Venice: Remove private field fallback for open weights
2026-04-09 20:05:03 -05:00
Aiden Cline
59831a9a0e
Merge pull request #1399 from dpuyosa/venice-model-updates-2026-04-09
...
Venice: Update model configs with refreshed pricing and limits
2026-04-09 20:04:55 -05:00
Aiden Cline
c7552d0e00
Merge pull request #1396 from shelvick/add-azure-grok-4-20
...
Add Grok 4.20 reasoning and non-reasoning to Azure
2026-04-09 20:04:04 -05:00
Aiden Cline
76d53c9e96
Merge pull request #1382 from cgilly2fast/dev
...
feat(firmware): add zai 5.1 and qwen 3.6 plus
2026-04-09 20:03:50 -05:00
Aiden Cline
61573f666a
Merge pull request #1395 from zainhas/dev
...
[Together AI] add GLM-5.1 + Gemma 4 31B it
2026-04-09 20:03:13 -05:00
Aiden Cline
1d30640400
Merge branch 'dev' into dev
2026-04-09 20:02:59 -05:00
Aiden Cline
1f1eafe173
Merge pull request #1406 from riccardogiorato/dev
...
Update GLM to version 5.1 for together provider
2026-04-09 20:02:08 -05:00
Aiden Cline
a75c0f0fe9
Merge pull request #1400 from dpuyosa/venice-add-new-models
...
Venice: Add 4 new AI models
2026-04-09 17:34:55 -05:00
Marcus Dillavou
57c6d817d5
DeepInfra: Add GLM 5.1
2026-04-09 15:50:04 -05:00
Riccardo Giorato
3d1d77da76
Merge pull request #2 from riccardogiorato/orchestrator/add-glm-5-1-together-r8k9f
...
add GLM-5.1 for together provider
2026-04-09 22:42:54 +02:00
orchestrator-build[bot]
e7cfed72ff
fix GLM-5.1 open_weights to true
2026-04-09 20:40:20 +00:00
orchestrator-build[bot]
b748364a20
fix GLM-5.1 pricing for together provider
2026-04-09 20:39:49 +00:00
orchestrator-build[bot]
30439f4409
replace GLM-5 with GLM-5.1 for together provider
2026-04-09 20:37:34 +00:00
orchestrator-build[bot]
7baad3cc22
add GLM-5.1 for together provider
2026-04-09 20:35:51 +00:00
Nicolás Pujia
564992885b
Add OpenRouter support for DeepSeek R1
2026-04-09 11:54:16 -03:00
Nicolás Pujia
57a53889db
Add OpenRouter support for Qwen3.5 Flash 2026-02-23
2026-04-09 11:53:53 -03:00
Nanai Jua
0f019a5e9a
Add MiniMax M2.5 (free) configuration file
...
https://openrouter.ai/provider/open-inference
2026-04-09 18:11:40 +09:00
dpuyosa
d6fec11252
[venice] Remove private field fallback for open weights
...
- Rely solely on modelSource for open weights detection
- Remove privacy field fallback per new ZDR policies
2026-04-09 11:03:03 +02:00
dpuyosa
7604313114
[venice] Add 4 new AI models
...
- Add Mercury 2 (reasoning model)
- Add Mistral Small 4 (multimodal)
- Add Nemotron Cascade 2 30B A3B
- Add Qwen 3.5 397B (multimodal)
2026-04-09 10:31:24 +02:00
dpuyosa
c3a2b1a1db
[venice] Update model configs with refreshed pricing and limits
...
- Update last_updated dates to 2026-04-09 across 6 models
- Adjust Grok pricing to reflect current rates
- Add context_over_200k pricing for Qwen 3.6 Plus
- Correct Gemma 4 output limits from 12288 to 8192
- Rename Qwen 3.6 Plus to "Uncensored" variant
2026-04-09 10:17:05 +02:00
QiyuanChen
68294f3bd6
Add GLM-5V-Turbo model configuration for siliconflow provider
2026-04-09 11:18:25 +08:00
QiyuanChen
ad10ce6232
Add GLM-5.1 model configuration files for siliconflow provider
2026-04-09 11:08:53 +08:00
Scott Helvick
5a09420d65
Add Grok 4.20 reasoning and non-reasoning to Azure
2026-04-09 02:30:55 +00:00
Zain Hasan
f1b3177ff7
add gemma 4 31b instruct
2026-04-08 18:49:38 -07:00
Zain Hasan
a7d0152fd1
[Together AI] add GLM-5.1
2026-04-08 17:55:11 -07:00
Aiden Cline
46c6aef51b
Merge pull request #1393 from spiffytech/dev
...
Add Synthetic support for GLM-5 and Nemotron 3 Super
2026-04-08 16:01:12 -05:00
Aiden Cline
7c34bf01b5
Merge pull request #1394 from spiffytech/ollama-changes
...
Add Ollama Cloud support for Gemma 4. Updated properties on Gemini 3 Flash
2026-04-08 16:00:12 -05:00
spiffytech
5ce20c0978
Added Ollama Cloud support for Gemma 4. Updated properties on Gemini 3 Flash.
2026-04-08 15:03:29 -04:00
spiffytech
6055551b33
Added Synthetic support for GLM-5 and Nemotron 3 Super
2026-04-08 14:57:51 -04:00
Aiden Cline
a96094c059
Merge pull request #1390 from GoGoris/add-qwen3-coder-next-cortecs
...
Add qwen3-coder-next model for cortecs
2026-04-08 11:26:42 -05:00
Aiden Cline
39e86eb055
Merge pull request #1384 from dpuyosa/feat/add-venice-claude-opus-4-6-fast-glm-5-1
...
Venice: Add Claude Opus 4.6 Fast and GLM 5.1 models
2026-04-08 11:25:56 -05:00
Aiden Cline
119f421437
Merge pull request #1387 from cantalupo555/feat/add-gemma-4-free-openrouter
...
feat(openrouter): add Gemma 4 free models
2026-04-08 11:25:30 -05:00
Aiden Cline
6ab4c1be04
Merge pull request #1386 from cantalupo555/chore/remove-qwen3.6-plus-free-openrouter
...
chore(openrouter): remove discontinued Qwen3.6 Plus free
2026-04-08 11:25:15 -05:00
Aiden Cline
f1eaa4bd9d
Merge pull request #1392 from Solidsilver/feat/fireworks-add-glm-5-1-qwen-3-6-plus
...
feat(fireworks): add GLM 5.1 and Qwen 3.6 Plus models
2026-04-08 11:24:52 -05:00
Aiden Cline
2f4693b9cf
Merge pull request #1391 from CassiusXiang/fix/openrouter-qwen3.6-plus
...
fix(openrouter): replace discontinued qwen3.6-plus free with paid model
2026-04-08 11:24:41 -05:00
Luke M
5c867d5fba
feat(fireworks): add GLM 5.1 and Qwen 3.6 Plus models
2026-04-08 09:06:50 -07:00
XiangChang
682a486f09
fix(openrouter): replace discontinued qwen3.6-plus free with paid model
2026-04-08 22:31:52 +08:00
Steven Goris
ea10c5674b
Add qwen3-coder-next model for cortecs
2026-04-08 15:14:31 +02:00
cantalupo555
5829a9174c
feat(openrouter): add Gemma 4 26B A4B free
...
- Add google/gemma-4-26b-a4b-it:free (MoE, 256K context, multimodal, reasoning)
2026-04-08 08:32:07 -03:00
cantalupo555
4717e5a902
feat(openrouter): add Gemma 4 31B free
...
- Add google/gemma-4-31b-it:free (256K context, multimodal, reasoning)
2026-04-08 08:31:55 -03:00
cantalupo555
20f225d6a6
chore(openrouter): remove discontinued Qwen3.6 Plus free
...
- Model no longer available on OpenRouter API
2026-04-08 08:16:47 -03:00
dpuyosa
04ada2c06e
[venice] Add Claude Opus 4.6 Fast and GLM 5.1 models
...
- Add claude-opus-4-6-fast model with 1M context
- Add zai-org-glm-5-1 model with reasoning and tool_call
2026-04-08 11:40:46 +02:00
Frank
23fd440f57
update zen models
2026-04-08 02:20:44 -04:00
Colby Gilbert
a4c09d58c3
feat(firmware): add zai 5.1 and qwen 3.6 plus
2026-04-07 22:18:54 -07:00
Aiden Cline
82924aa6f2
Merge pull request #1376 from mugnimaestra/feat/add-glm-5.1-tee-chutes
...
feat(chutes): add zai-org/GLM-5.1-TEE model
2026-04-07 23:55:23 -05:00
Aiden Cline
c8f0b6d573
Merge pull request #1377 from mchenco/mchen/update-cf-workers-ai-models
...
update cloudflare-workers-ai: add gemma-4, remove non-LLMs, fix metadata
2026-04-07 23:55:07 -05:00
Aiden Cline
09c3cb3e0a
Merge pull request #1379 from zhongruan0522/dev
...
add GLM-5.1 to zhipuai and zai providers
2026-04-07 23:54:32 -05:00
Aiden Cline
555662ec80
Merge pull request #1381 from llc1123/chore/zenmux-update
...
chore(zenmux): add GLM-5.1 model configuration
2026-04-07 23:54:20 -05:00
Aiden Cline
fc63cc19c3
feat: add experimental modes to models to express things like "fast" that induce additional price changes
2026-04-07 23:47:19 -05:00
Aiden Cline
26d2f3e8e9
Merge pull request #1378 from friendliai/minpeter/add-friendli-glm-5.1
...
feat(friendli): add GLM-5.1 and remove deprecated models
2026-04-07 22:30:44 -05:00
粒粒橙
92374dd74d
chore(zenmux): add GLM-5.1 model configuration
2026-04-08 11:30:14 +08:00
minpeter
ee4de44fbe
fix(friendli): align model display names with cross-provider majority convention
2026-04-08 11:39:05 +09:00
zhongruan0522
5513af5b6c
add GLM-5.1 to zhipuai and zai providers
2026-04-08 02:34:12 +00:00
minpeter
755509be95
feat(friendli): add GLM-5.1 and remove deprecated models
2026-04-08 11:31:50 +09:00
mchen
392b0988f2
update cloudflare-workers-ai: add gemma-4, remove non-LLMs, fix metadata
...
- Add gemma-4-27b-a4b-it (multimodal, reasoning, tool calling)
- Remove non-LLM models: embeddings, TTS, translation, sentiment analysis
- Remove redundant models: Llama 2/3.x variants, Qwen, Mistral, DeepSeek, Gemma 3
- Update metadata: tool_call, reasoning, open_weights, attachment for remaining models
- Final models: gemma-4, llama-4-scout, kimi-k2.5, nemotron-3, gpt-oss-20b/120b, glm-4.7-flash
2026-04-07 21:00:44 -04:00
Muhammad Mugni Hadi
7f1e94f571
feat(chutes): add zai-org/GLM-5.1-TEE model
2026-04-08 06:40:30 +07:00
Aiden Cline
61c596874c
feat: add new provider.body and provider.headers support
2026-04-07 17:36:40 -05:00
Frank
ca0e64451d
update zen models
2026-04-07 18:01:31 -04:00
Aiden Cline
d2870fcfeb
Merge pull request #1374 from fhennerkes/dev
...
Poe: adding Gemma-4-31B (free model)
2026-04-07 16:48:52 -05:00
fhennerkes
9c8e1e0fa0
Merge branch 'anomalyco:dev' into dev
2026-04-07 14:02:06 -07:00
Frank
30d42207cb
update zen models
2026-04-07 16:48:49 -04:00
fhennerkes
c450a6ffaf
poe: add Gemma-4-31B model
...
Add new Google model from Poe API (released 2026-04-02):
- Free during preview
- 262K context, 8K output
- Modalities: text, image
Co-Authored-By: Claude Sonnet 4.6 <noreply@anthropic.com >
2026-04-07 13:04:03 -07:00
Frank
e889bd7c3b
update zen models
2026-04-07 13:43:40 -04:00
Aiden Cline
dcd3803dea
Merge pull request #1359 from rdbisme/dev
...
Add missing Qwen3 Coder Next model to Amazon Bedrock
2026-04-07 12:41:41 -05:00
Aiden Cline
05b22d8f8e
Merge pull request #1372 from JoshuaDietz/dev
...
feat(ollama-cloud): add GLM-5.1
2026-04-07 12:30:52 -05:00
Aiden Cline
5c1a00bbe8
Merge pull request #1373 from cantalupo555/feat/openrouter-glm-5.1
...
feat(openrouter): add z-ai/glm-5.1 model
2026-04-07 12:30:12 -05:00
cantalupo555
27cbd6557d
feat(openrouter): add z-ai/glm-5.1 model
...
- Add GLM-5.1 with 202K context, reasoning, tool call, and structured output
- Pricing: $1.40/M input, $4.40/M output, $0.26/M cache read
2026-04-07 14:02:55 -03:00
JoshuaDietz
8114fa2f13
Merge branch 'anomalyco:dev' into dev
2026-04-07 19:01:23 +02:00
Joshua Dietz
5dc325bf09
feat(ollama-cloud): add GLM-5.1
...
unsure about temperature=true which is not set for GLM-5 but is set for GLM-5.1 huggingface.
2026-04-07 19:00:54 +02:00
Aiden Cline
d27ce785fe
Merge pull request #1370 from gary149/feat/huggingface-glm-5.1
...
feat(huggingface): add GLM-5.1
2026-04-07 11:55:12 -05:00
Aiden Cline
f47f9e8414
Merge pull request #1225 from mixlayer/add_mixlayer
...
New provider: Mixlayer
2026-04-07 11:46:19 -05:00
Victor Muštar
df41a38c6f
feat(huggingface): add GLM-5.1
2026-04-07 18:33:53 +02:00
Aiden Cline
5b37d05f82
Merge pull request #1367 from cantalupo555/feat/stepfun-step-3.5-flash-2603
...
feat(stepfun): add step-3.5-flash-2603 model
2026-04-07 11:21:02 -05:00
Aiden Cline
99e046916c
Merge pull request #1360 from jonathancaevans/update-kimi-k2p5-turbo-name
...
Update Kimi K2.5 Turbo display name for Firepass clarity
2026-04-07 11:09:02 -05:00
Aiden Cline
98baf7eaca
Merge pull request #1364 from seffhunnn/fix-openrouter-glm-5-turbo-web
...
fix: correct glm-5-turbo pricing and context for openrouter
2026-04-07 11:08:29 -05:00
Aiden Cline
462d7fa620
Merge pull request #1365 from dpuyosa/feat/venice-add-qwen-3-6-plus
...
Venice: Add Qwen 3.6 Plus model
2026-04-07 11:08:15 -05:00
cantalupo555
7eea1dec18
feat(stepfun): add step-3.5-flash-2603 model
...
- Add Step 3.5 Flash 2603 model optimized for agent workflows
- Released April 2, 2026, same pricing as step-3.5-flash
2026-04-07 12:01:26 -03:00
dpuyosa
10652fb8dc
feat(venice): add Qwen 3.6 Plus model
...
- Add Qwen 3.6 Plus with 1M context window
- Support text, image, and video input modalities
- Enable reasoning, tool calling, and structured output
2026-04-07 13:41:42 +02:00
Mohd Saif
816f9bb585
fix: correct glm-5-turbo pricing and context
2026-04-07 14:45:31 +05:30
Jonathan Evans
03376e9986
Update Kimi K2.5 Turbo display name
...
Add (firepass) suffix to clarify this is the Firepass router endpoint.
Follow-up to #1256
2026-04-06 13:10:01 -07:00
Ruben Di Battista
91d73da942
Enable reasoning capability for Qwen3 Coder Next model
2026-04-06 21:51:16 +02:00
Ruben Di Battista
db7e4ff9ba
Add missing Qwen3 Coder Next model to Amazon Bedrock
2026-04-06 21:07:36 +02:00
Aiden Cline
d6145d1479
Merge pull request #1354 from llc1123/chore/zenmux-updates
...
zenmux: remove deprecated models and add Agnes 1.5 entries
2026-04-06 08:27:09 -07:00
粒粒橙
ec314aa0f1
fix(zenmux): add image support for agnes-1.5-lite
2026-04-06 14:28:22 +08:00
粒粒橙
188c36696e
zenmux: remove deprecated models and add models from sapiens-ai
2026-04-06 14:08:08 +08:00
Aiden Cline
2b5f3f961d
Merge pull request #1340 from battall/patch-1
...
fix: google/gemma-4 -it suffixes
2026-04-05 21:10:43 -07:00
Aiden Cline
bf8ce0b8f8
Merge pull request #1342 from seffhunnn/fix-deepseek-context-window
...
fix: correct deepseek-chat context window to 131072
2026-04-05 21:06:53 -07:00
Aiden Cline
5e3eb74da2
Merge pull request #1341 from u1630022/feat-openrouter-trinity-large-thinking
...
add trinity large thinking to openrouter provider
2026-04-05 21:06:19 -07:00
Aiden Cline
542620b288
Merge pull request #1343 from spyridonas/patch-1
...
Fix capitalization in model name
2026-04-05 21:05:15 -07:00
Aiden Cline
f6576ffb0f
Merge pull request #1344 from dpuyosa/venice/gemma4-trinity
...
Venice: Add Gemma 4 and Arcee Trinity models, enable GLM 4.6 reasoning
2026-04-05 21:05:00 -07:00
Aiden Cline
88649810a6
Merge pull request #1346 from GHagui/add-gemma-4-openrouter
...
feat(openrouter): add Gemma 4 26B A4B and Gemma 4 31B models
2026-04-05 21:04:42 -07:00
Aiden Cline
1367d90f32
Merge pull request #1353 from cyberofficial/vultr
...
[Vultr] Update Inference Models
2026-04-05 21:03:58 -07:00
Cyber Official
322b1154be
Update Inference Models
...
Vultr Updated Inference API endpoint with different versions of models, this commit adds in new models and corrects some information
2026-04-05 17:12:13 -04:00
Gabriel Hagui
7336867f65
Apply suggestions from code review
...
Co-authored-by: Copilot <175728472+Copilot@users.noreply.github.com >
2026-04-04 21:17:55 -03:00
GHagui
caaf4d8b0a
Merge branch 'add-gemma-4-openrouter' of https://github.com/GHagui/models.dev into add-gemma-4-openrouter
2026-04-05 00:09:19 +00:00
GHagui
08c2b928b6
fix(openrouter) Normalize formatting in Gemma 4 26B A4B and Gemma 4 31B TOML files
2026-04-05 00:03:44 +00:00
Gabriel Hagui
3a4b5e50cb
Add files via upload
...
Fixing CRLF to LF
2026-04-04 20:51:25 -03:00
Gabriel Hagui
3244ef835e
feat(openrouter) Add Gemma 4 26B A4B and Gemma 4 31B
2026-04-04 20:43:15 -03:00
GHagui
346b12f42f
feat(openrouter) Add Gemma 4 26B A4B and Gemma 4 31B
2026-04-04 23:38:56 +00:00
dpuyosa
116d91a0cc
[venice] Add Gemma 4 and Arcee Trinity models, enable GLM 4.6 reasoning
...
- Add Google Gemma 4 26B A4B and 31B instruct models with multimodal support
- Add Arcee Trinity Large Thinking reasoning model
- Enable reasoning capability for GLM 4.6
- Reduce Qwen3 5-9B output limit to 32K
2026-04-05 00:08:47 +02:00
Spyros Sakellaropoulos
778036c53c
Fix capitalization in model name
2026-04-05 00:10:17 +03:00
Mohd Saif Ansari
e86cc85dd4
fix: update deepseek-chat context window
2026-04-05 01:59:40 +05:30
Eavan Pattie
a6f030fe1c
add trinity large thinking to openrouter provider
...
* fixes trinity-large-thinking erroneously marked as not open_weight in
vercel provider
2026-04-04 22:44:25 +03:00
Aiden Cline
1eb0b8c8e1
Merge pull request #1338 from anthraxx/alibaba-qwen3.6-plus
...
feat(alibaba): add Qwen3.6 Plus model configuration to all regions
2026-04-04 11:55:23 -07:00
Aiden Cline
1bc0b9d81f
Merge pull request #1335 from branchgrove/dev
...
Add google-vertex DeepSeek V3.2 model
2026-04-04 11:51:37 -07:00
Battal Doğukan Hazar
eca166ed4d
fix: -it suffix for google/gemma-4
2026-04-04 21:48:31 +03:00
Aiden Cline
e64404b173
Merge pull request #1336 from WJQSERVER/dev
...
Add Google Gemma 4 31B IT to NVIDIA(NIM) provider
2026-04-04 11:48:19 -07:00
Battal Doğukan Hazar
e9c8425b32
fix: google/gemma-4 -it suffix
2026-04-04 21:47:27 +03:00
Aiden Cline
8618de3429
Merge pull request #1333 from riccardogiorato/dev
...
Remove deprecated models from TogetherAI provider
2026-04-04 11:28:46 -07:00
Aiden Cline
6e64316225
Merge pull request #1337 from fanweixiao/vivgrd/gpt-5.4
...
provider(vivgrid): add GPT-5.3 Codex, GPT-5.4 Mini, and GPT-5.4 Nano models
2026-04-04 11:28:33 -07:00
Aiden Cline
a36d032e93
Merge pull request #1339 from cantalupo555/remove/qwen3.6-plus-preview-free
...
chore(openrouter): remove discontinued Qwen3.6 Plus Preview free
2026-04-04 11:28:18 -07:00
Aiden Cline
e74dce023f
Merge pull request #1315 from seffhunnn/fix-pdf-modalities
...
fix: add missing pdf modality to supported OpenAI models
2026-04-04 11:28:07 -07:00
cantalupo555
47f9b2b910
chore(openrouter): remove discontinued Qwen3.6 Plus Preview free
...
- Remove qwen3.6-plus-preview:free model after Qwen3.6 Plus free replaced it
2026-04-04 15:08:11 -03:00
Levente Polyak
6ba0af61d6
feat(alibaba): add Qwen3.6 Plus model configuration to all regions
...
- Add missing regions
- Add coding-plan variants
- Use pricing from model info page
Link: https://bailian.console.alibabacloud.com/cn-beijing?tab=model#/model-market/detail/qwen3.6-plus
2026-04-04 19:52:15 +02:00
C.C. Fan
24d4a9b8dd
provider(vivgrid): add GPT-5.3 Codex, GPT-5.4 Mini, and GPT-5.4 Nano models
...
Co-Authored-By: Claude Sonnet 4.6 <noreply@anthropic.com >
2026-04-04 23:07:49 +08:00
wjqserver
d5b420768a
feat: add Google Gemma 4 31B IT to NVIDIA provider
2026-04-04 22:26:39 +08:00
Riccardo Giorato
66ada85298
more deprecations
2026-04-04 14:18:08 +02:00
Elias Lundgren
2ffb1d9181
Add google-vertex DeepSeek V3.2 model
2026-04-04 13:57:42 +02:00
Riccardo Giorato
04768ff141
Merge pull request #1 from riccardogiorato/orchestrator/remove-deprecated-models-g5nx
...
Remove deprecated models from togetherai provider
2026-04-03 21:51:02 +02:00
orchestrator-dev[bot]
9560d51908
Remove deprecated models from togetherai provider
2026-04-03 19:49:40 +00:00
Aiden Cline
6a41e31306
Merge pull request #1326 from michaelnchin/fix/amazon-bedrock-structured-output
...
fix: set correct structured output values for Amazon Bedrock models
2026-04-03 14:03:05 -05:00
Aiden Cline
133c529ebf
Merge pull request #1327 from llc1123/chore/zenmux-new-models
...
zenmux: add KAT-Coder-Pro-V2, Qwen3.6-Plus, and GLM 5V Turbo
2026-04-03 14:02:40 -05:00
Aiden Cline
0a6b828e42
Merge pull request #1331 from zhongruan0522/feat/xiaomi-token-plan
...
feat: add Xiaomi Token Plan providers (cn/sgp/ams)
2026-04-03 14:00:46 -05:00
Aiden Cline
406f2f66c6
Merge pull request #1323 from Pxys-io/fix-xiaomi-mimo-cache-pricing
...
fix(openrouter): add missing cache_read pricing for xiaomi/mimo-v2-pro and xiaomi/mimo-v2-omni
2026-04-03 13:59:51 -05:00
Aiden Cline
948ce76d8d
Merge pull request #1325 from DEAN-Cherry/dev
...
revert: alibaba-cn MiniMax-M2.5 to MiniMax-M2.7
2026-04-03 13:47:53 -05:00
Aiden Cline
e3dd89ba2e
Merge pull request #1332 from vercel/update-vercel-models-20260403-1639
...
Update Vercel models
2026-04-03 13:47:33 -05:00
Aiden Cline
7e5ae3bb06
Merge pull request #1329 from Track07-cda/alibaba-cn-qwen3.6
...
feat(alibaba-cn): add Qwen3.6 Plus model configuration
2026-04-03 13:47:19 -05:00
Aiden Cline
5392c185d2
Merge pull request #1328 from sadoclaw/add-gemma-4-models
...
Add Gemma 4 26B and 31B models
2026-04-03 13:47:00 -05:00
github-actions[bot]
72613f5dbf
chore(vercel): update Vercel model definitions
...
Auto-generated by weekly workflow from Vercel AI Gateway API.
Co-Authored-By: github-actions[bot] <41898282+github-actions[bot]@users.noreply.github.com>
2026-04-03 16:39:32 +00:00
zhongruan0522
01098f81a9
feat: add Xiaomi Token Plan providers (cn/sgp/ams)
2026-04-03 11:22:04 +00:00
Track07-cda
65e798c8f2
feat(alibaba-cn): add Qwen3.6 Plus model configuration
2026-04-03 15:58:53 +08:00
sadoclaw
28a75176c4
Add Gemma 4 26B and 31B models
2026-04-03 08:34:41 +03:00
粒粒橙
61e63db8e3
zenmux: add new models
2026-04-03 13:09:53 +08:00
Michael Chin
95cd6196fa
set correct structured_output values for Amazon Bedrock models
2026-04-02 21:15:18 -07:00
Bryan
19c1f2ebe2
revert: alibaba-cn MiniMax-M2.5 to MiniMax-M2.7
...
Revert PR #1002 - MiniMax-M2.5 is no longer available, now using MiniMax-M2.7
2026-04-03 12:09:05 +08:00
pxys-io
4d25bb5622
fix(openrouter): add missing cache_read pricing for xiaomi/mimo-v2-pro and xiaomi/mimo-v2-omni
...
OpenRouter charges bash.20/M cache_read tokens for mimo-v2-pro and
bash.08/M for mimo-v2-omni, but these were missing from the cost section.
Source: https://openrouter.ai/api/v1/models
Co-authored-by: Qwen-Coder <qwen-coder@alibabacloud.com >
2026-04-03 04:08:55 +02:00
Aiden Cline
8845bf3f3b
Merge pull request #1321 from BlockListed/cortecs-add-glm-5
...
Add glm-5 to cortecs
2026-04-02 19:28:50 -05:00
Aiden Cline
9b5bcde109
Merge pull request #1322 from fhennerkes/dev
...
poe: add GPT-5.3-Codex-Spark and Kimi-K2.5-FW models
2026-04-02 19:28:41 -05:00
Frank
fa75002f19
update zen models
2026-04-02 19:01:00 -04:00
fhennerkes
69b6f3e94a
poe: add GPT-5.3-Codex-Spark and Kimi-K2.5-FW models
...
Add 2 new free models
2026-04-02 15:11:34 -07:00
BlockListed
26f6b602fc
add glm-5 to cortecs
2026-04-02 21:45:20 +02:00
Aiden Cline
287c69acaf
Merge pull request #1314 from dpark01/add-kimi-k2-thinking-vertex
...
feat(google-vertex): add moonshotai/kimi-k2-thinking-maas model
2026-04-02 11:30:08 -05:00
Aiden Cline
2b46c3aef2
Merge pull request #1320 from cantalupo555/feature/qwen3.6-plus-free
...
feat: add Qwen3.6 Plus free on OpenRouter
2026-04-02 11:29:46 -05:00
cantalupo555
a9b7faa409
feat(openrouter): add qwen3.6-plus free model configuration
...
- Add Qwen3.6 Plus (free) provider configuration
- $0 pricing with 1M context window
- Multimodal input support (text, image, video)
- Full capabilities: reasoning, tool calls, structured output
- Attachment enabled for vision modality
2026-04-02 13:23:09 -03:00
Jack
ad3305bc08
Merge branch 'dev' of github.com:anomalyco/models.dev into dev
2026-04-03 00:12:23 +08:00
Jack
0198228bbb
Add MiMo V2 Pro/Omni models and family entries
...
Register two new MiMo V2 models and update model family values. Added "mimo-pro" and "mimo-omni" to ModelFamilyValues in packages/core/src/family.ts, and added corresponding TOML model descriptors under providers/opencode-go/models: mimo-v2-pro.toml and mimo-v2-omni.toml. The Pro model includes very large context (1,048,576) and tiered costs for >200k context, while the Omni model exposes multimodal input (text, image, audio, pdf) and a large 262,144 context. Both files set metadata (release_date, last_updated, knowledge cutoff, open_weights) and define interleaved reasoning field, costs, limits, and modalities.
2026-04-03 00:11:58 +08:00
Aiden Cline
165bc7df94
Merge pull request #1316 from NIKU-SINGH/remove-claude-3-7-sonnet-latest
...
Remove invalid claude-3-7-sonnet-latest model entry
2026-04-02 10:51:44 -05:00
NIKU-SINGH
36956dc4f4
Remove invalid claude-3-7-sonnet-latest model entry
...
Anthropic's API does not accept claude-3-7-sonnet-latest as a model ID
(returns 404). The versioned alias claude-3-7-sonnet-20250219 should be
used instead.
Co-Authored-By: Claude Sonnet 4.6 <noreply@anthropic.com >
2026-04-02 16:27:15 +05:30
Mohd Saif Ansari
5ec371653b
fix: add missing pdf modality to supported OpenAI models
2026-04-02 12:40:50 +05:30
Frank
bb62fc43c1
update zen models
2026-04-01 22:59:53 -04:00
Frank
c1465447ab
update zen models
2026-04-01 17:53:17 -04:00
Daniel Park
471e9ec492
feat(google-vertex): add moonshotai/kimi-k2-thinking-maas model
2026-04-01 16:24:14 -04:00
Aiden Cline
57c3d38c21
Merge pull request #1313 from zhongruan0522/dev
...
Add GLM-5V-Turbo
2026-04-01 13:08:38 -05:00
Aiden Cline
d19d2508b0
Merge pull request #1273 from Cahl-Dee/add-the-grid-ai
...
feat: add thegrid.ai provider and associated models
2026-04-01 12:25:35 -05:00
zhongruan0522
d0d2fffcfb
Add GLM-5V-Turbo
2026-04-01 16:30:38 +00:00
Aiden Cline
07f48b1f2a
Merge pull request #1311 from Moniet/feat/add-gpt-image-models
...
feat(openai): add openai gpt-image models
2026-04-01 10:52:12 -05:00
Moniet
a33e15d095
feat(openai): add openai gpt-image models
2026-04-01 17:03:36 +05:30
Aiden Cline
b64c2100f4
Merge pull request #1306 from dinhkim/feat/add-openrouter-glm-5-turbo
...
feat: add GLM-5-Turbo model in OpenRouter AI provider
2026-03-31 23:11:14 -05:00
Kim Truong
7ba7237633
feat: add GLM-5-Turbo model in OpenRouter AI provider
2026-03-31 23:18:02 +07:00
Aiden Cline
6e1ca23e6c
Merge pull request #1305 from marcelarie/dev
...
Update synthetic.new model: Qwen3.5-397B
2026-03-31 10:43:23 -05:00
Aiden Cline
8ffb4ed5a7
Merge pull request #1301 from xinrui-z/fix/aihubmix-zod-validation-provider
...
fix(aihubmix): zod-validation-error
2026-03-31 10:43:12 -05:00
Xinrui
6b12398083
Replace @ai-sdk/openai-compatible with the official aihubmix/ai-sdk-provider package and remove the hardcoded api URL, as the dedicated package handles schema validation and endpoint configuration internally.
2026-03-31 23:29:10 +08:00
Aiden Cline
751745f200
Merge pull request #1304 from dpuyosa/add-gpt-54-mini-venice
...
Venice: Add GPT-5.4 Mini and remove discontinued models
2026-03-31 10:19:03 -05:00
marcelarie
593308b596
Merge branch 'dev' of github.com:marcelarie/models.dev into dev
2026-03-31 13:05:29 +02:00
marcelarie
54386f35b7
Added: Missing synthetic.new Qwen3.5-397B model
2026-03-31 13:02:19 +02:00
dpuyosa
73a83971e8
[venice] Add GPT-5.4 Mini and remove discontinued models
...
- Add GPT-5.4 Mini with reasoning and tool_call
- Update Aion 2.0 with reasoning capability
- Remove discontinued mistral-31-24b and qwen3-4b
2026-03-31 09:49:09 +02:00
Xinrui
b5d8da29bb
fix(aihubmix): switch to dedicated provider package to resolve Zod validation error
...
Replace @ai-sdk/openai-compatible with the official aihubmix/ai-sdk-provider
package and remove the hardcoded api URL, as the dedicated package handles
schema validation and endpoint configuration internally.
2026-03-31 12:37:56 +08:00
Aiden Cline
798538f9ae
Merge pull request #1254 from llc1123/dev
...
feat(zenmux): route models through protocol-specific SDKs
2026-03-30 18:50:02 -05:00
Aiden Cline
d4ce566f27
Merge pull request #1296 from sylviezhang37/update-vercel-models-20260330-1655
...
Update Vercel models
2026-03-30 18:49:47 -05:00
Aiden Cline
2042e3dd71
Merge pull request #1298 from cantalupo555/feature/qwen3.6-plus-preview-free
...
feat: add Qwen3.6 Plus Preview free on OpenRouter
2026-03-30 18:49:34 -05:00
Aiden Cline
08b577b6c9
Merge pull request #1299 from cyberofficial/vultr
...
Remove discontinued Vultr models
2026-03-30 18:49:24 -05:00
Cyber Official
01c3d44f99
Remove discontinued Vultr models
...
Vultr no longer supports these models on Serverless Inference:
- DeepSeek R1 Distill Llama 70B
- DeepSeek R1 Distill Qwen 32B
- GPT OSS 120B
- Llama 3.1 Nemotron Ultra 253B v1
- NVIDIA Nemotron 3 Super 120B A12B NVFP4
Remaining active models:
- MiniMax-M2.5: $0.30/M in, $1.20/M out
- DeepSeek-V3.2: $0.55/M in, $1.65/M out
- Kimi-K2.5: $0.55/M in, $2.75/M out
- GLM-5-FP8: $0.85/M in, $3.10/M out
2026-03-30 17:39:00 -04:00
cantalupo555
1695bfb958
feat: add Qwen3.6 Plus Preview free on OpenRouter
...
- Add free variant of Qwen3.6 Plus Preview to OpenRouter provider
- 1M context, 65K max output, /bin/bash.00 pricing
- Text-only modality (OpenRouter API reports text->text)
- Source: OpenRouter /api/v1/models API
---
Co-Authored-By: opencode https://opencode.ai
2026-03-30 18:08:54 -03:00
Frank
bad8bedd25
update zen models
2026-03-30 16:50:10 -04:00
github-actions[bot]
ad26e881ce
chore(vercel): update Vercel model definitions
...
Auto-generated by weekly workflow from Vercel AI Gateway API.
Co-Authored-By: github-actions[bot] <41898282+github-actions[bot]@users.noreply.github.com>
2026-03-30 16:55:28 +00:00
Aiden Cline
348932f00c
Merge pull request #1290 from zhongruan0522/dev
...
feat: add gpt-5.3-chat-latest model to openai
2026-03-29 22:25:52 -05:00
Aiden Cline
d564c80b78
Merge pull request #1289 from pedrxd/mistral-add-mistral-small-4
...
Adding mistral small 4
2026-03-29 22:23:57 -05:00
Aiden Cline
cba38e4707
Merge pull request #1287 from aeonzh/patch-1
...
Use correct name for Nemotron 3 Super (free) on OpenRouter
2026-03-29 12:35:18 -05:00
Aiden Cline
7ce9e95d62
Merge pull request #1294 from sk0x0y/add-glm-5.1-nanogpt
...
feat(nano-gpt): add glm-5.1 and glm-5.1:thinking models
2026-03-29 12:34:31 -05:00
sk0x0y
a05b2a6be7
add glm-5.1 model to nano-gpt provider
2026-03-29 23:12:38 +09:00
zhongruan0522
17f41b72ad
feat: add gpt-5.3-chat-latest model to openai
2026-03-29 11:18:44 +00:00
Pedro Ruiz
16cc382572
feat(mistral): Adding mistral small 4
2026-03-29 09:46:58 +02:00
Aiden Cline
3d456e3798
Merge pull request #1286 from khda-tech/dev
...
Add gemma family for google provider
2026-03-28 20:07:16 -05:00
Aiden Cline
0223ab3107
Merge pull request #1288 from cgilly2fast/dev
...
fix(firmware): incorrect model name for glm-5
2026-03-28 20:06:57 -05:00
Colby Gilbert
28ad50f6e9
fix(firmware): incorrect model name for glm-5
2026-03-27 22:26:19 -07:00
Zheng He Hu
bf8fb378a4
Rename nemotron-3-super-120b-a12b-free.toml to nemotron-3-super-120b-a12b:free.toml
2026-03-28 02:54:35 +01:00
Aiden Cline
b74242fdbf
Merge pull request #1278 from fhennerkes/dev
...
poe: add Grok-4.20-Multi-Agent and DeepSeek-V3.2 models
2026-03-27 15:48:08 -05:00
Aiden Cline
8131cc947c
Update providers/poe/models/novita/deepseek-v3.2.toml
...
Co-authored-by: Oleg Voronkovich <oleg-voronkovich@yandex.ru >
2026-03-27 15:21:16 -05:00
Khrulev Danil
95a73581cc
Add gemma family for google provider
2026-03-27 21:39:52 +03:00
Zack Angelo
39ee133c98
mixlayer: adhere to logo standards, remove fill and size attributes
2026-03-27 09:23:09 -07:00
Aiden Cline
c5e4e2c740
Merge pull request #1276 from voronkovich/feat-update-groq
...
feat(groq): Update Groq models
2026-03-27 10:51:40 -05:00
Aiden Cline
357c3021fb
Merge pull request #1281 from zhongruan0522/dev
...
Added support for Zhipu AI's official CodingPlan GLM-5.1 model.
2026-03-27 09:49:49 -05:00
Aiden Cline
6afb0fea06
Merge pull request #1280 from dpuyosa/dev
...
Venice: Add Aion 2.0, update DeepSeek V3.2, remove Gemini 3 Pro Preview
2026-03-27 09:46:58 -05:00
阮
dd781c4c15
feat: add glm-5.1 model to zai-coding-plan and zhipuai-coding-plan
2026-03-27 11:51:14 +00:00
dpuyosa
04e3b7c008
[venice] Add Aion 2.0, update DeepSeek V3.2, remove Gemini 3 Pro Preview
...
- feat(venice): add Aion 2.0 model
- fix(venice): enable tool_call and structured_output on DeepSeek V3.2
- fix(venice): remove deprecated Gemini 3 Pro Preview
2026-03-27 09:19:28 +01:00
Oleg Voronkovich
ca40cb538d
Updates
2026-03-27 00:37:40 +03:00
Aiden Cline
f03de60559
Merge pull request #1259 from smakosh/llmgateway-models
...
feat: update LLM Gateway to 204 models
2026-03-26 15:25:22 -05:00
fhennerkes
0226a37a51
poe: add Grok-4.20-Multi-Agent and DeepSeek-V3.2 models
2026-03-26 12:18:13 -07:00
smakosh
c2ba40d7d5
fix: logo format, remove README, minimax weights
...
- Normalize logo to 24x24, viewBox 0 0 40 40, currentColor
- Remove README.md (other providers don't have one)
- Mark all MiniMax models as open_weights = true
Co-Authored-By: Claude Opus 4.6 (1M context) <noreply@anthropic.com >
2026-03-26 20:02:58 +01:00
smakosh
d77bee4f29
fix: correct reasoning, vision, tools flags
...
The export script only checked the first active
provider for capabilities. Now checks all providers
and uses model ID patterns for reasoning detection.
Co-Authored-By: Claude Opus 4.6 (1M context) <noreply@anthropic.com >
2026-03-26 19:55:28 +01:00
Oleg Voronkovich
5ec0ac2807
feat(groq): Update Groq models
2026-03-26 18:59:09 +03:00
Aiden Cline
62015086c6
Merge pull request #1274 from petit-blaireau/copilot/add-openrouter-mistral-small-4
...
Add Mistral Small 4 for OpenRouter
2026-03-25 19:47:51 -05:00
copilot-swe-agent[bot]
2d86bafb20
feat(openrouter): add Mistral Small 4 (mistral-small-2603)
...
Co-authored-by: petit-blaireau <1893252+petit-blaireau@users.noreply.github.com >
Agent-Logs-Url: https://github.com/petit-blaireau/models.dev/sessions/871bbf50-3ef0-4926-afd1-e95c79f6bd57
2026-03-26 00:03:38 +00:00
Cahl-Dee
bf4e5aab17
remove family property and add open_weight
2026-03-25 16:57:02 -05:00
Cahl-Dee
1590791225
adding thegrid.ai provider and associated models
2026-03-25 16:09:54 -05:00
Aiden Cline
047f3356d6
Merge pull request #1265 from MiyakoMeow/add-glm-4.7-flashx
...
Add glm-4.7-flashx to ZAI/ZhipuAI
2026-03-25 16:09:24 -05:00
Aiden Cline
1394d2ca4d
Merge pull request #1270 from NachoFLizaur/feat/bedrock-add-nemotron-super-3-120b
...
feat(amazon-bedrock): add NVIDIA Nemotron 3 Super 120B
2026-03-25 15:05:31 -05:00
Aiden Cline
5b73677b33
Merge pull request #1272 from NachoFLizaur/fix/bedrock-minimax-m2.5-glm-5-limits
...
fix(amazon-bedrock): correct MiniMax M2.5 and GLM-5 context/output limits
2026-03-25 15:05:18 -05:00
Nacho F. Lizaur
780db038c4
fix(amazon-bedrock): correct MiniMax M2.5 and GLM-5 context/output limits
2026-03-25 16:49:31 +01:00
Nacho F. Lizaur
29c1249d64
feat(amazon-bedrock): add NVIDIA Nemotron 3 Super 120B
2026-03-25 15:44:56 +01:00
lioZ129
1b0633b172
Update providers/hpc-ai/models/moonshotai/kimi-k2.5.toml
...
Co-authored-by: Copilot <175728472+Copilot@users.noreply.github.com >
2026-03-25 17:15:16 +08:00
lioZ129
aed0ee3bb7
Update providers/hpc-ai/models/moonshotai/kimi-k2.5.toml
...
Co-authored-by: Copilot <175728472+Copilot@users.noreply.github.com >
2026-03-25 17:15:03 +08:00
Contributor
9fa7856fc6
feat: add HPC-AI model provider support
2026-03-25 16:55:45 +08:00
MiyakoMeow
8b80a2b34b
Add glm-4.7-flashx to ZAI/ZhipuAI
2026-03-25 06:13:03 +08:00
Aiden Cline
897aa53905
Merge pull request #1257 from fhennerkes/dev
...
poe: add GPT-5.4-Nano and GPT-5.4-Mini models
2026-03-24 15:19:13 -05:00
Aiden Cline
5c36e54a43
Merge pull request #1260 from Happily-Coding/dev
...
Add MiniMax 2.5 to siliconflow
2026-03-24 15:18:39 -05:00
Aiden Cline
a38f9373ab
Merge pull request #1261 from cyberofficial/vultr
...
[Vultr] Delete Qwen2.5-Coder-32B-Instruct.toml
2026-03-24 10:14:38 -05:00
Aiden Cline
fb72c181f8
Merge pull request #1263 from fanweixiao/vivgrd/gpt-5.4
...
provider(vivgrid): add gpt-5.4 and upgrade gemini-3 to gemini-3.1
2026-03-24 10:14:20 -05:00
C.C. Fan
601300c7e7
provider(vivgrid): add gpt-5.4 and upgrade gemini-3 to gemini-3.1
2026-03-24 21:04:59 +08:00
Cyber Official
8d9e966867
Delete Qwen2.5-Coder-32B-Instruct.toml
...
Model no longer offered
2026-03-24 02:48:38 -04:00
UrielS
cf8f12b375
Add MiniMax 2.5 to siliconflow
...
Add MiniMax 2.5 to siliconflow
2026-03-24 01:14:36 -03:00
UrielS
116e35a32a
Add MiniMax 2.5 to siliconflow
2026-03-24 01:13:37 -03:00
Aiden Cline
87a02b897a
Merge pull request #1258 from vglafirov/feat/gitlab-gpt-5-4-models
...
feat(gitlab): add GPT-5.4, GPT-5.4 Mini, GPT-5.4 Nano, and GPT-5.3 Codex models
2026-03-23 22:00:07 -05:00
Vladimir Glafirov
849a529a42
fix(gitlab): use unscoped gitlab-ai-provider npm package name
2026-03-23 23:52:13 +01:00
smakosh
c2f0bd08e6
feat: update LLM Gateway models to 204
...
Regenerated model exports from latest LLM Gateway
source. Adds 66 new models including Claude 4.6,
GPT-5.x, Gemini 3.1, Grok 4, and more. Removes
deprecated model aliases.
Co-Authored-By: Claude Opus 4.6 (1M context) <noreply@anthropic.com >
2026-03-23 22:28:46 +01:00
Vladimir Glafirov
5257d0919a
feat(gitlab): add GPT-5.4, GPT-5.4 Mini, GPT-5.4 Nano, and GPT-5.3 Codex models
2026-03-23 21:41:11 +01:00
fhennerkes
379ab2757f
poe: add GPT-5.4-Nano and GPT-5.4-Mini models
...
Add 2 new OpenAI models from Poe API:
GPT-5.4-Nano (released 2026-03-11):
- Reasoning support, 400K context, 128K output
- Cost: $0.18/M input, $1.1/M output
- Modalities: text, image
GPT-5.4-Mini (released 2026-03-12):
- Reasoning support, 400K context, 128K output
- Cost: $0.68/M input, $4/M output
- Modalities: text, image
Co-Authored-By: Claude Sonnet 4.6 <noreply@anthropic.com >
2026-03-23 13:13:10 -07:00
Aiden Cline
9838c55e29
Merge pull request #1256 from jonathancaevans/add-kimi-k2p5-turbo-router
...
Add Fireworks Kimi K2.5 Turbo router
2026-03-23 15:08:32 -05:00
Aiden Cline
4235bc6432
Merge pull request #1249 from tobwen/cleanup/openrouter-deprecated-models
...
chore(openrouter): remove deprecated and unavailable models
2026-03-23 15:07:27 -05:00
Jonathan Evans
29c3e9cbf4
Add Kimi K2.5 Turbo router for Fireworks
...
- Model ID: accounts/fireworks/routers/kimi-k2p5-turbo
- Pricing set to 0 (handled at subscription layer)
- Supports text and image input, text output
- Includes reasoning capabilities
2026-03-23 16:01:08 -04:00
Aiden Cline
b0c1f37aad
Merge pull request #1255 from sylviezhang37/update-vercel-models-20260323-1954
...
Update Vercel models
2026-03-23 15:00:14 -05:00
github-actions[bot]
bd07f155f5
chore(vercel): update Vercel model definitions
...
Auto-generated by weekly workflow from Vercel AI Gateway API.
Co-Authored-By: github-actions[bot] <41898282+github-actions[bot]@users.noreply.github.com>
2026-03-23 19:54:44 +00:00
粒粒橙
0f95849d38
fix(zenmux): align model metadata with runtime support
2026-03-23 23:06:36 +08:00
粒粒橙
8dc90a4097
feat(zenmux): route models through protocol-specific SDKs
2026-03-23 21:48:00 +08:00
Aiden Cline
0a19559e4e
Merge pull request #1252 from tarun1793/add-glm-5-fastrouter
...
Add GLM-5 model to fastrouter provider
2026-03-23 08:13:45 -05:00
Aiden Cline
f27f83bdf9
Merge pull request #1253 from jacksonwilliamsva/add-bedrock-minimax-m2.5-glm-5
...
feat(amazon-bedrock): add MiniMax M2.5 and GLM-5 models
2026-03-23 08:13:35 -05:00
Aiden Cline
7b050ec719
Merge pull request #1244 from BlockListed/cortecs-add-claude-models
...
Add more claude models to cortecs
2026-03-23 08:13:27 -05:00
Aiden Cline
443cfd7703
Merge pull request #1247 from wojons/dev
...
Add Nemotron 3 Super model to OpenRouter and NVIDIA providers
2026-03-23 08:11:12 -05:00
Aiden Cline
6d648bf471
Merge pull request #1250 from dsingal0/dev
...
correct model name for nemotron super 3 on baseten
2026-03-23 08:10:45 -05:00
Jackson Williams
e1a83a6812
feat(amazon-bedrock): add MiniMax M2.5 and GLM-5 models
...
Add two newly available Amazon Bedrock models:
- minimax.minimax-m2.5: 1M context, /bin/bash.30/.20 per 1M tokens
- zai.glm-5: 200K context, .00/.20 per 1M tokens
Both models were added to Amazon Bedrock on March 18, 2026.
Specs sourced from AWS Bedrock pricing page and vendor documentation.
2026-03-23 11:41:21 +11:00
Tarun
b8606ed0e4
use latest price from fastrouter
2026-03-22 23:47:13 +00:00
Tarun
69c4600842
Override fastrouter glm-5 with zai glm-5 values
2026-03-22 23:43:09 +00:00
Tarun
5eb4a369f3
Add GLM-5 model to fastrouter provider
2026-03-22 23:31:53 +00:00
Dhruv Singal
17df1b49ce
Update Baseten Nemotron model name
2026-03-22 13:04:40 -07:00
Dhruv Singal
f18bc9b95b
Update Baseten Nemotron display name
2026-03-22 13:00:49 -07:00
Dhruv Singal
485c37e862
Rename Baseten Nemotron model to match API ID
2026-03-22 12:57:10 -07:00
tobwen
0f092e3f62
chore(openrouter): remove expired/revealed/ended endpoints
2026-03-22 12:28:08 +00:00
tobwen
a7bcd7e632
chore(openrouter): remove models without endpoints
2026-03-22 12:27:59 +00:00
Alexis Okuwa
6d081af472
Add Nemotron 3 Super model to OpenRouter and NVIDIA providers
2026-03-22 05:45:00 -05:00
Aiden Cline
8ee9ea1d96
Merge pull request #1243 from v1gnesh/dev
...
Update Grok 4.2 model names
2026-03-21 11:56:37 -05:00
Aiden Cline
c71a365320
Merge pull request #1245 from Daltonganger/add-nanogpt-minimax-m2-7
...
Add NanoGPT MiniMax M2.7 model metadata
2026-03-21 11:55:15 -05:00
Ruben Beuker
bb0e828b77
add NanoGPT MiniMax M2.7 model metadata
2026-03-21 14:58:35 +01:00
BlockListed
58cb222125
add more claude models to cortecs
2026-03-21 09:18:00 +01:00
v1gnesh
33f67289ee
Rename model and remove beta status
2026-03-21 11:29:13 +05:30
v1gnesh
ad34d7948c
Update model name and status in TOML file
2026-03-21 11:28:22 +05:30
v1gnesh
7132293513
Add grok-4.20-0309-non-reasoning.toml file
2026-03-21 11:27:53 +05:30
Aiden Cline
495bc263e7
Merge pull request #1241 from BlockListed/add-minimax-2.5-cortecs
...
Add minimax M2.5 to cortecs
2026-03-20 15:55:28 -05:00
Aiden Cline
3811a45efe
Merge pull request #1235 from anomalyco/github-sync
...
sync github copilot limits
2026-03-20 15:55:09 -05:00
BlockListed
c4d04d2ed9
add minimax m2.5 to cortecs
2026-03-20 21:51:30 +01:00
Aiden Cline
20fcbbc336
Merge pull request #1240 from sylviezhang37/update-vercel-models-20260320-1642
...
Update Vercel models
2026-03-20 13:08:30 -05:00
github-actions[bot]
482b8ed69d
chore(vercel): update Vercel model definitions
...
Auto-generated by weekly workflow from Vercel AI Gateway API.
Co-Authored-By: github-actions[bot] <41898282+github-actions[bot]@users.noreply.github.com>
2026-03-20 16:42:15 +00:00
Aiden Cline
b77e95b84e
Merge pull request #1236 from LYY/update/zenmux-sync
...
Sync ZenMux models with latest website data
2026-03-20 10:23:40 -05:00
Aiden Cline
463006f80c
Merge pull request #1237 from vincentbernat/fix/scaleway-qwen3.5
...
fix(scaleway): set the correct family for Qwen 3.5 for Scaleway
2026-03-20 10:23:17 -05:00
Vincent Bernat
45b5abb875
fix(scaleway): set the correct family for Qwen 3.5 for Scaleway
2026-03-20 06:32:51 +01:00
LYY
4255e420ec
Sync zenmux models with website
...
Add 17 new models found on zenmux.ai:
- google/gemini-3-pro-image-preview
- google/gemini-3.1-flash-lite-preview
- minimax/minimax-m2.7
- minimax/minimax-m2.7-highspeed
- openai/gpt-5.3-chat
- openai/gpt-5.3-codex
- openai/gpt-5.4
- openai/gpt-5.4-mini
- openai/gpt-5.4-nano
- openai/gpt-5.4-pro
- qwen/qwen3.5-flash
- qwen/qwen3.5-plus
- volcengine/doubao-seed-2.0-code
- x-ai/grok-4.2-fast
- x-ai/grok-4.2-fast-non-reasoning
- xiaomi/mimo-v2-omni
- xiaomi/mimo-v2-pro
- z-ai/glm-5-turbo
Mark anthropic/claude-3.5-sonnet as deprecated (not found on website)
2026-03-20 12:46:16 +08:00
Aiden Cline
cda892e0d7
sync github copilot limits
2026-03-19 21:54:42 -05:00
Aiden Cline
098ff4f5bf
Merge pull request #1227 from Verizane/dev
...
add OpenRouter models for gpt-5.4 mini and gpt-5.4 nano
2026-03-19 21:23:15 -05:00
Aiden Cline
0f70b8959f
Merge pull request #1234 from mchenco/dev
...
Add Workers AI models: kimi-k2.5, nemotron-3-120b-a12b, glm-4.7-flash
2026-03-19 15:11:19 -05:00
mchen
b8e6d58e5b
add workers-ai models: kimi-k2.5, nemotron-3-120b-a12b, glm-4.7-flash
2026-03-19 14:59:17 -04:00
Roman Koslowski
a855001a7e
apply changes from review
2026-03-19 17:20:26 +01:00
Aiden Cline
ac760b2268
Merge pull request #1230 from SamizuHM/feature/zhipuai-coding-plan-add-glm-5-turbo
...
zhipuai-coding-plan: Add glm-5-turbo.toml and replace symlink
2026-03-19 10:42:43 -05:00
Aiden Cline
d4a5ea7ae7
Merge pull request #1226 from spiffytech/dev
...
Add Ollama Cloud support for Minimax M2.7
2026-03-19 10:41:47 -05:00
Aiden Cline
434ed89ba2
Merge pull request #1228 from dpuyosa/minimax_m2_7
...
Venice: Add MiniMax M2.7 and update DeepSeek V3.2 pricing
2026-03-19 10:41:16 -05:00
Aiden Cline
6d7719a62a
Merge pull request #1229 from 0b1000/dev
...
Xiaomi: Add MiMo-V2-Pro and MiMo-V2-Omni
2026-03-19 10:41:06 -05:00
Aiden Cline
93637039ef
Merge pull request #1231 from ariane-emory/feat/feat/add-xiaomi-mimo-v2-pro-and-omni
...
feat: add the Xiaomi MiMo V2 Pro and Xiaomi MiMo V2 Omni models to the OpenRouter provide
2026-03-19 10:40:44 -05:00
Ariane Emory
9c95f796c0
Merge remote-tracking branch 'upstream/dev' into feat/feat/add-xiaomi-mimo-v2-pro
2026-03-19 11:22:43 -04:00
Ariane Emory
e8650b6073
feat: add xiaomi mimo-v2-pro and mimo-v2-omni models to openrouter
2026-03-19 11:18:46 -04:00
SamizuHM
23c2be6ff7
feat(zhipuai-coding-plan): add glm-5-turbo.toml and replace glm-5-turbo with symlink
2026-03-19 18:09:17 +08:00
Frank
913a63dbe6
update zen models
2026-03-19 00:33:45 -04:00
0b1000
503087e99b
Merge branch 'anomalyco:dev' into dev
2026-03-19 12:28:38 +08:00
0b1000
48150f09d3
Xiaomi: Add MiMo-V2-Pro and MiMo-V2-Omni
2026-03-19 12:27:00 +08:00
Aiden Cline
5fef681657
Disable tool_call in grok model configuration
2026-03-18 23:09:30 -05:00
Frank
123054ae0c
update zen models
2026-03-18 20:45:44 -04:00
Frank
03060d154b
update zen models
2026-03-18 20:37:47 -04:00
dpuyosa
5c9b8108e0
Update minimax-m27.toml
2026-03-19 01:02:24 +01:00
dpuyosa
c8084681f9
[venice] Add MiniMax M2.7 and update DeepSeek V3.2 pricing
...
- Add MiniMax M2.7 model with reasoning and tool_call support
- Update DeepSeek V3.2 pricing (input: $0.33, output: $0.48, cache: $0.16)
2026-03-19 00:58:50 +01:00
Roman Koslowski
352ab4ae1b
add gpt-5.4 mini and gpt-5.4 nano
2026-03-18 22:16:55 +01:00
spiffytech
cf0b416b15
Added Ollama Cloud support for Minimax M2.7
2026-03-18 16:15:07 -04:00
Aiden Cline
38339a2a90
Merge pull request #1224 from APonce911/minimax-m2.7-openrouter
...
add MiniMax M2.7 to OpenRouter
2026-03-18 14:10:13 -05:00
Aiden Cline
ff9040bf52
Update minimax-m2.7.toml
2026-03-18 14:09:26 -05:00
Aiden Cline
3039804af4
Delete providers/opencode/models/minimax-m2.7.toml
2026-03-18 14:08:55 -05:00
Frank
7a4ad7bec8
update go models
2026-03-18 14:40:25 -04:00
Zack Angelo
7a2ec5ab95
New provider: Mixlayer
2026-03-18 11:19:35 -07:00
airton
721cc122bc
add MiniMax M2.7 to OpenRouter and OpenCode
2026-03-18 18:57:09 +01:00
Aiden Cline
0527f019af
Merge pull request #1221 from sergical/fix/bedrock-claude-4-6-context-window-and-pricing
...
fix(amazon-bedrock): set Claude Sonnet 4.6 and Opus 4.6 context window to 1M
2026-03-18 12:17:57 -05:00
Aiden Cline
c89371de50
Merge pull request #1223 from sylviezhang37/update-vercel-models-20260318-1659
...
Update Vercel models
2026-03-18 12:17:22 -05:00
Sylvie Zhang
6d6d4220d8
Enable open_weights in minimax-m2.7.toml
2026-03-18 10:12:37 -07:00
Sylvie Zhang
8b984eeec1
Enable open_weights in minimax-m2.7-highspeed model
2026-03-18 10:12:21 -07:00
github-actions[bot]
586027c8f1
chore(vercel): update Vercel model definitions
...
Auto-generated by weekly workflow from Vercel AI Gateway API.
Co-Authored-By: github-actions[bot] <41898282+github-actions[bot]@users.noreply.github.com>
2026-03-18 16:59:24 +00:00
Sergiy Dybskiy
343b5f87ef
fix(amazon-bedrock): set Claude Sonnet 4.6 and Opus 4.6 context window to 1M
...
Both models support a 1M token context window natively on Bedrock via the
Converse API with no beta headers required. Verified empirically via the
AWS CLI (bedrock-runtime converse): 950K tokens succeeds, >1M returns
'prompt is too long: N tokens > 1000000 maximum'.
The AWS Bedrock pricing page confirms long context pricing for these two
models is identical to standard pricing (no surcharge), so the
[cost.context_over_200k] section is removed as it was incorrect.
2026-03-18 12:20:38 -04:00
Aiden Cline
955b773ee5
Merge pull request #1218 from pomidornijfrukt/azure/5.4-mini-nano
...
Add GPT-5.4 Mini and Nano models for Azure providers
2026-03-18 10:31:13 -05:00
Aiden Cline
98559071f0
Merge pull request #1217 from cgilly2fast/dev
...
chore(firmware): update base url and docs url
2026-03-18 10:30:44 -05:00
eCube-cachy
0660308816
add: GPT-5.4 Mini and Nano model configurations for Azure providers
2026-03-18 15:17:52 +02:00
Jack
380f9dd8eb
Merge pull request #1216 from no1wudi/dev
...
Add MiniMax M2.7 and M2.7-highspeed models to 4 official providers
2026-03-18 16:29:59 +08:00
Jack
1cfdab1b18
update MiniMax-M2.7 cache_read to 0.06
2026-03-18 16:27:53 +08:00
Colby Gilbert
75a981f957
chore(firmware): update base url and docs url
2026-03-18 00:41:05 -07:00
Huang Qi
7fadbcadc8
Add MiniMax M2.7 and M2.7-highspeed models to 4 official providers
2026-03-18 15:21:06 +08:00
Frank
38f9092292
update zen models
2026-03-18 02:30:18 -04:00
Aiden Cline
92149b9eaa
rm nonexistant github model
2026-03-17 21:41:51 -05:00
Aiden Cline
b614f0e69c
Merge pull request #1214 from luisrudge/dev
...
Add GPT-5.4 mini and nano to GitHub Copilot provider
2026-03-17 20:13:46 -05:00
Luís Rudge
67d6dac5c5
Add GPT-5.4 mini and nano to GitHub Copilot provider
2026-03-17 18:44:38 -06:00
Aiden Cline
7d3cc61a48
Merge pull request #1207 from PedroACosta/feat/add-dinference-provider
...
feat(providers): add dinference provider
2026-03-17 14:51:31 -05:00
Aiden Cline
f02ea6c4d2
Merge pull request #1115 from skywalker512/feat/add-tencent-coding-plan
...
feat: add Tencent Coding Plan provider
2026-03-17 14:51:19 -05:00
Aiden Cline
0cb50eeece
Merge pull request #1208 from scwgoire/march-update
...
Scaleway 26-03 model updates
2026-03-17 14:48:12 -05:00
Aiden Cline
878311d2e0
Merge pull request #1210 from dm-cohere/dm/fix-update-cohere-model-capabilities
...
fix(models): update cohere model capabilities
2026-03-17 14:32:27 -05:00
Aiden Cline
a0e89f65d6
Merge pull request #1206 from 0b1000/dev
...
Rename minimax-m2.5.toml to MiniMax-M2.5.toml
2026-03-17 14:32:19 -05:00
Aiden Cline
74099b7c9c
Merge pull request #1213 from smrdotgg/add-openai-gpt-5-4-mini-and-nano
...
Add OpenAI GPT-5.4 mini and nano
2026-03-17 14:31:24 -05:00
Aiden Cline
ec522435c3
Merge pull request #1211 from sylviezhang37/update-vercel-models-20260317-1807
...
Update Vercel models
2026-03-17 14:30:24 -05:00
smr
d839cd37d4
Add OpenAI GPT-5.4 mini and nano
...
Capture the newly released mini and nano model metadata so models.dev reflects OpenAI's latest GPT-5.4 lineup with current pricing, limits, and knowledge cutoff.
2026-03-17 22:09:13 +03:00
github-actions[bot]
ecb6ef7f93
chore(vercel): update Vercel model definitions
...
Auto-generated by weekly workflow from Vercel AI Gateway API.
Co-Authored-By: github-actions[bot] <41898282+github-actions[bot]@users.noreply.github.com>
2026-03-17 18:07:06 +00:00
Deirdre Meehan
8b4d341054
fix: cohere models on non-cohere providers
2026-03-17 16:51:24 +00:00
Deirdre Meehan
62f4a28308
fix: cohere provider models
2026-03-17 16:44:55 +00:00
Pedro
2fb8ef0dc8
feat(providers): add dinference provider
2026-03-17 14:13:36 +01:00
Gregoire de Turckheim
96968e2bf8
feat: Scaleway 26-03 model updates
2026-03-17 12:15:52 +01:00
0b1000
ee9d7879ce
Rename minimax-m2.5.toml to MiniMax-M2.5.toml
2026-03-17 14:50:35 +08:00
Frank
71283512a6
update zen models
2026-03-17 02:21:13 -04:00
Frank
cd4afd7e7c
update zen models
2026-03-17 02:19:17 -04:00
Aiden Cline
1239d0190b
Merge pull request #1204 from cyberofficial/vultr
...
VULTR: Updated Vultr model pricing to reflect current serverless inference rates
2026-03-16 16:10:39 -05:00
Aiden Cline
491bf6ccba
Merge pull request #1202 from RaviTharuma/fix/chutes-pricing-update-2026-03
...
fix(chutes): update pricing and limits from live API
2026-03-16 16:10:25 -05:00
Cyber Official
c993d0c121
Updated Vultr model pricing to reflect current serverless inference rates
...
Updated Vultr model pricing to reflect current serverless inference rates
This commit updates the cost configuration for all Vultr models to align with their latest pricing tiers:
**Cost Reductions:**
- DeepSeek-R1-Distill-Qwen-32B: Input $0.55→$0.30, Output $2.75→$0.30 (73% reduction)
- NVIDIA-Nemotron-3-Super-120B-A12B-NVFP4: Input $0.55→$0.20, Output $2.75→$0.80 (64% input, 71% output reduction)
- Qwen2.5-Coder-32B-Instruct: Input $0.55→$0.20, Output $2.75→$0.60 (64% input, 78% output reduction)
- gpt-oss-120b: Input $0.55→$0.15, Output $2.75→$0.60 (73% input, 78% output reduction)
- MiniMax-M2.5: Input $0.55→$0.30, Output $2.75→$1.20 (45% input, 56% output reduction)
**Cost Adjustments:**
- DeepSeek-R1-Distill-Llama-70B: Input $0.55→$2.00, Output $2.75→$2.00 (significant increase)
- DeepSeek-V3.2: Output $2.75→$1.65 (40% reduction)
- Llama-3.1-Nemotron-Ultra-253B-v1: Output $2.75→$1.80 (35% reduction)
- GLM-5-FP8: Input $0.55→$0.85, Output $2.75→$3.10 (55% input, 13% output increase)
2026-03-16 13:58:28 -04:00
Aiden Cline
e55c39a83d
Merge pull request #1141 from sk0x0y/feature/nanogpt-confirmed-suffix2-fixes
...
fix(nano-gpt): rename confirmed 2-suffix model ids
2026-03-16 10:57:44 -05:00
Aiden Cline
6dea000e25
Merge pull request #1148 from sk0x0y/feature/nanogpt-bundled-confirmed-suffix2-fixes
...
fix(nano-gpt): rename bundled confirmed 2-suffix model ids
2026-03-16 10:57:30 -05:00
Aiden Cline
3f7a757b3f
Merge pull request #1142 from sk0x0y/feature/nanogpt-more-confirmed-suffix2-fixes
...
fix(nano-gpt): rename more confirmed 2-suffix model ids
2026-03-16 10:56:36 -05:00
Aiden Cline
c693fd71e2
Merge pull request #1194 from cyberofficial/vultr
...
Update Vultr model list with 10 new models and updated pricing
2026-03-16 10:55:19 -05:00
Aiden Cline
54e04e288a
Merge pull request #1198 from amritbanerjee/add-glm-5-turbo
...
Add GLM-5-Turbo model support
2026-03-16 10:47:08 -05:00
Aiden Cline
462a179eee
Merge pull request #1203 from jerome-benoit/fix/sap-ai-core-model-specs
...
fix(sap-ai-core): align model specs with official sources
2026-03-16 10:46:40 -05:00
Aiden Cline
95db59034d
Merge pull request #1201 from dpuyosa/venice-new-models
...
Venice: Add new provider models
2026-03-16 10:45:59 -05:00
Aiden Cline
74dcc74e32
Merge pull request #1200 from dpuyosa/venice/pricing-update
...
Venice: Update model pricing for 7 models
2026-03-16 10:45:47 -05:00
Jérôme Benoit
57975f5f25
fix(sap-ai-core): align model specs with official sources
2026-03-16 13:59:06 +01:00
Ravi Tharuma
ad7b063747
fix(chutes): update pricing and limits from live API
...
Synced 6 Chutes model definitions against the live API at
https://llm.chutes.ai/v1/models (queried 2026-03-16).
Models updated:
- deepseek-ai/DeepSeek-V3.2-TEE: cost 0.25/0.38→0.28/0.42, cache 0.125→0.14, context 163840→131072
- zai-org/GLM-5-TEE: cost 0.75/2.5→0.95/3.15, added cache_read 0.475
- zai-org/GLM-4.6-TEE: cost 0.35/1.5→0.4/1.7, added cache_read 0.2
- zai-org/GLM-4.6V: added cache_read 0.15
- MiniMaxAI/MiniMax-M2.5-TEE: cost 0.15/0.6→0.3/1.1, added cache_read 0.15
- Qwen/Qwen3.5-397B-A17B-TEE: cost 0.3/1.2→0.39/2.34, cache 0.15→0.195
2026-03-16 11:42:29 +01:00
dpuyosa
f76e9f0551
[venice] Add new provider models
...
- Add mistral-small-3.2-24b-instruct, qwen3-5-9b, venice-uncensored-role-play, zai-org-glm-4.6
2026-03-16 09:37:41 +01:00
dpuyosa
d70a49b36f
[venice] Update model pricing for 7 models
...
- Remove context_over_200k pricing from Claude models
- Update Grok cache_read pricing from 0.5 to 0.25
- Update Kimi, MiniMax input/output pricing
2026-03-16 09:05:37 +01:00
amrit
3487135f9f
Add GLM-5-Turbo model support
2026-03-16 12:14:50 +11:00
Aiden Cline
458a66c766
Merge pull request #1197 from kesku/update-perplexity-agent-models
...
Update Perplexity Agent API models
2026-03-15 10:59:23 -05:00
Frank
d3a84dc7ec
update zen models
2026-03-15 10:59:52 -04:00
Kesku
ae61b25583
update perplexity-agent: add gpt-5.4 & nemotron, remove gemini-3-pro
2026-03-15 03:46:50 +00:00
Aiden Cline
74be576eda
Merge pull request #1178 from Sewer56/change-synthetic-endpoint
...
Add OpenAI and Anthropic compatible endpoints
2026-03-14 20:55:30 -05:00
Aiden Cline
164df2cda0
Merge pull request #1191 from Alcatraz-Zhang/update/kilo-models
...
Sync Kilo model definitions with latest gateway catalog
2026-03-14 20:54:45 -05:00
Cyber Official
2cd7908369
Update Vultr model list with 10 new models and updated pricing
...
- Updated pricing to $0.55/M input tokens, $2.75/M output tokens
- Updated context limits to safe floor values from official testing
- Added accurate output token limits from official model documentation
- Added 5 new models: MiniMax M2.5, DeepSeek V3.2, GLM-5 FP8, Llama 3.1 Nemotron Ultra 253B, NVIDIA Nemotron 3 Super 120B A12B NVFP4
- Updated existing models: DeepSeek R1 Distill variants, GPT OSS 120B, Kimi K2.5, Qwen2.5 Coder 32B
Model specifications:
- MiniMax M2.5: 196K context, 4,096 output
- Qwen2.5-Coder-32B: 15K context, 256 output (notable low default)
- DeepSeek R1 Distill Llama 70B: 130K context, 4,096 output
- DeepSeek R1 Distill Qwen 32B: 130K context, 4,096 output
- DeepSeek V3.2: 163K context, 4,096 output
- Kimi K2.5: 261K context, 32,768 output (high output limit)
- GPT OSS 120B: 130K context, 8,192 output
- GLM-5 FP8: 202K context, 131,072 output (exceptionally high)
- Llama 3.1 Nemotron Ultra 253B: 32K context, 4,096 output
- NVIDIA Nemotron 3 Super 120B A12B NVFP4: 260K context, 8,192 output
All models set to text-only (no vision support) as confirmed.
2026-03-14 19:47:01 -04:00
Alcatraz-Zhang
cc667340f5
Sync Kilo model definitions with latest gateway catalog
...
Refresh the Kilo provider catalog so models.dev matches the current gateway inventory, pricing, and availability.
2026-03-15 04:35:38 +08:00
Sewer56
f2cfc1435d
Changed: Synthetic to use newer openai endpoint
2026-03-14 17:09:44 +00:00
Aiden Cline
35bb8cca47
Merge pull request #1172 from bigfluffycookie/add-deepinfra-llama-models
...
Add deepinfra llama models
2026-03-14 10:55:13 -05:00
Aiden Cline
3468a410e1
Merge pull request #1177 from ar27111994/dev
...
Add Grok 4.1 Fast configurations for reasoning and non-reasoning
2026-03-14 10:54:57 -05:00
Aiden Cline
b1b5e3c5cd
Merge pull request #1174 from dacbd/patch-1
...
fix(wandb): fix k2.5 settings
2026-03-14 10:54:35 -05:00
Aiden Cline
97f03ec672
Merge pull request #1175 from dacbd/patch-2
...
chore(docs): add note for manual testing with opencode
2026-03-14 10:54:22 -05:00
BigFluffyCookie
9b516924aa
Add limit output for llama models
2026-03-14 11:49:57 +01:00
Ahmed Rehan
929a39600b
feat(models): add Grok 4.1 Fast (Reasoning and Non-Reasoning) configurations
2026-03-14 14:27:24 +05:00
Daniel Barnes
a87d8bb8cc
chore(docs): add note for manual testing with opencode
2026-03-14 13:42:57 +09:00
Daniel Barnes
574139eb49
fix(wandb): fix k2.5 settings
2026-03-14 13:07:16 +09:00
Aiden Cline
1e3bc38b31
Merge pull request #1137 from mcowger/mcowger/correct-gemini-flash-lite-pricing
...
Fix incorrect pricing for gemini-3.1-flash-lite-preview
2026-03-13 18:41:41 -05:00
Aiden Cline
8916fe9874
Merge pull request #1171 from stephenkuhn214/dev
...
fix(amazon-bedrock): Remove deprecated and add missing models
2026-03-13 18:26:25 -05:00
BigFluffyCookie
5d956b41a6
Rename llama models to remove "Meta" prefix
2026-03-13 23:15:27 +01:00
BigFluffyCookie
42a7a14f69
Add Meta Llama models to DeepInfra provider
2026-03-13 22:53:39 +01:00
Stephen Kuhn
f24ee000d7
fix(amazon-bedrock): update and add models
...
- Remove 19 deprecated/EOL models
- Add 7 new models: DeepSeek V3.2, Llama 3.1 405B, Magistral Small 1.2, Ministral 3 3B, Mistral Large 3, Pixtral Large, NVIDIA Nemotron Nano 3 30B
- Fix Devstral 2 123B: correct name, family, and open_weights
- Set accurate Bedrock launch dates for all new models
2026-03-13 16:02:04 -04:00
Aiden Cline
7196b1fb2c
Merge pull request #1170 from anomalyco/revert-1166-fix/update-gpt53-codex-spark-preview
...
Revert "fix(openai): rename gpt-5.3-codex-spark to gpt-5.3-codex-spark-preview"
2026-03-13 14:31:33 -05:00
Aiden Cline
f6c0d5a29d
Revert "fix(openai): rename gpt-5.3-codex-spark to gpt-5.3-codex-spark-preview"
2026-03-13 14:30:58 -05:00
Aiden Cline
ee63449aa5
sonnet 4.6 and opus 4.6 1M context
2026-03-13 14:27:55 -05:00
Aiden Cline
92aa44ec00
Merge pull request #1166 from rluisr/fix/update-gpt53-codex-spark-preview
...
fix(openai): rename gpt-5.3-codex-spark to gpt-5.3-codex-spark-preview
2026-03-13 14:18:41 -05:00
Aiden Cline
477284535c
Rename model from 'GPT-5.3 Codex Spark Preview' to 'GPT-5.3 Codex Spark'
2026-03-13 14:17:44 -05:00
Aiden Cline
304233bdda
Merge pull request #1169 from mdrxy/mdrxy/anthropic-token-limits
...
Update Claude 4.6 context/pricing
2026-03-13 14:13:40 -05:00
Aiden Cline
25d782ee2c
Reduce context limit from 1,000,000 to 200,000
2026-03-13 14:13:30 -05:00
Aiden Cline
0f63393d51
Update context limit in claude-opus-4-6.toml
2026-03-13 14:12:56 -05:00
rluisr
e780eefce2
fix(openai): rename gpt-5.3-codex-spark to gpt-5.3-codex-spark-preview
...
The OpenAI API expects model ID 'gpt-5.3-codex-spark-preview', not
'gpt-5.3-codex-spark'. Rename model files in both openai and opencode
providers so the generated model ID matches the actual API.
2026-03-14 03:59:03 +09:00
Aiden Cline
a79585fa83
Merge pull request #1163 from micuintus/feature/Kimi2.5-fast
...
feat(nebius): add Kimi-K2.5-fast model
2026-03-13 13:14:38 -05:00
Aiden Cline
00801f74f2
Merge pull request #1164 from butyess/dev
...
Openrouter models: gemini 3.1 flash lite preview, grok 4.20 beta models.
2026-03-13 13:14:22 -05:00
Aiden Cline
185f6731ee
Merge pull request #1162 from dpuyosa/feature/venice-grok-4-20-beta
...
Venice: Add Grok 4.20 Beta models
2026-03-13 12:53:28 -05:00
Aiden Cline
d291b0575c
Merge pull request #1167 from sylviezhang37/update-vercel-models-20260313-1639
...
Update Vercel models
2026-03-13 12:53:11 -05:00
Mason Daugherty
382d9f3e7d
Update Claude 4.6 context/pricing
2026-03-13 13:53:04 -04:00
Aiden Cline
e64f5fe963
Merge pull request #1168 from mdrxy/mdrxy/update-baseten
...
Update Baseten models
2026-03-13 12:51:56 -05:00
Mason Daugherty
ea57ddfe7e
Update Baseten models
2026-03-13 13:48:41 -04:00
github-actions[bot]
29463d7fa8
chore(vercel): update Vercel model definitions
...
Auto-generated by weekly workflow from Vercel AI Gateway API.
Co-Authored-By: github-actions[bot] <41898282+github-actions[bot]@users.noreply.github.com>
2026-03-13 16:39:30 +00:00
Jack
bcc8db49ee
Merge pull request #1165 from anomalyco/chore/openrouter-alpha-reasoning-details-20260313
...
feat(openrouter): add interleaved reasoning details for alpha models
2026-03-13 22:25:25 +08:00
Jack
c8521d70f3
feat(openrouter): add interleaved reasoning details for alpha models
2026-03-13 22:20:54 +08:00
Federico Masi
490cd249e4
Openrouter models: gemini 3.1 flash lite preview, grok 4.20 beta models.
2026-03-13 15:11:12 +01:00
Michael Voigt
dbc636f5f3
feat(nebius): add Kimi-K2.5-fast model
2026-03-13 12:35:13 +01:00
Michael Voigt
9a32f671a1
fix(nebius): lowercase model ID for Nemotron-3-Super-120B-A12B
...
The filename must match the API casing (lowercase) to avoid 'model does not exist' errors.
2026-03-13 12:35:08 +01:00
dpuyosa
856d925eda
[venice] Add Grok 4.20 Beta models
...
- Add Grok 4.20 Beta model configuration (2M context, 128K output)
- Add Grok 4.20 Multi-Agent Beta model configuration
2026-03-13 10:48:38 +01:00
Aiden Cline
066a425917
Merge pull request #1158 from micuintus/feature/Nebius_Nemotron-3-Super-120b-a12b
...
feat(nebius): Add support for Nemotron-3-Super-120B-A12B
2026-03-12 22:20:20 -05:00
Aiden Cline
6df7f20cdc
Merge pull request #1156 from dsingal0/dev
...
added nemotron super on baseten
2026-03-12 22:20:06 -05:00
Aiden Cline
78bb47b90e
Merge pull request #1151 from dacbd/dacbd
...
fix(wandb): update models
2026-03-12 22:19:43 -05:00
Aiden Cline
c121d86419
Merge pull request #1160 from kreatoo/dev
...
feat: add zai-org/glm-4.7 and zai-org/glm-4.7-flash to NanoGPT
2026-03-12 22:11:18 -05:00
Aiden Cline
ab148eeb14
Merge pull request #1161 from Grin1024/dev
...
Add Claude Opus 4.6 and Sonnet 4.6 models to RequestY provider
2026-03-12 22:11:07 -05:00
lihui
49d196d326
Add Claude Opus 4.6 and Sonnet 4.6 models to RequestY provider
2026-03-13 09:00:54 +08:00
Kreato
8899b390ef
feat: add zai-org/glm-4.7 and zai-org/glm-4.7-flash to NanoGPT
2026-03-13 00:27:09 +03:00
Michael Voigt
5217f62ddf
fix(nebius): Follow context updates for Kimi 2.5 and GLM-5
2026-03-12 20:22:48 +01:00
Michael Voigt
55eaff9af1
feat(nebius): Add support for Nemotron-3-Super-120B-A12B
2026-03-12 20:22:21 +01:00
Dhruv Singal
7557c06ac0
update output length
2026-03-12 09:41:25 -07:00
Dhruv Singal
e85d820121
fix input output
2026-03-12 08:29:01 -07:00
Dhruv Singal
499d3a39ef
remove cache pricing
2026-03-12 08:21:22 -07:00
Dhruv Singal
b9b38d6e33
added nemotron super on baseten
2026-03-12 08:18:46 -07:00
Aiden Cline
ca24ac14fa
Merge pull request #1153 from dpuyosa/dev
...
Venice: Update model output token limits
2026-03-12 10:08:46 -05:00
Aiden Cline
822546fc67
Merge pull request #1155 from spiffytech/dev
...
Add Ollama Cloud support for Nemotron 3 Super
2026-03-12 10:08:31 -05:00
Aiden Cline
4555195b71
Merge pull request #1152 from v1gnesh/dev
...
Update grok-4.20 model defs
2026-03-12 10:08:15 -05:00
spiffytech
5eae8effc6
Added Ollama Cloud support for Nemotron 3 Super
2026-03-12 09:28:47 -04:00
dpuyosa
c1801aef87
[venice] Normalize model output token limits
...
- Update output limits to standard values across all models
2026-03-12 10:08:39 +01:00
v1gnesh
5e6464b272
Update grok-4.20-beta-reasoning
2026-03-12 10:27:40 +05:30
v1gnesh
e1a4f23332
Update grok-4.20-beta-non-reasoning
2026-03-12 10:26:03 +05:30
v1gnesh
753e1f9f0c
grok-multi-agent-beta update
2026-03-12 10:23:57 +05:30
Daniel Barnes
123ecd2ba5
docs url
2026-03-12 13:27:56 +09:00
Daniel Barnes
f15cda9fcb
remove old
2026-03-12 13:26:08 +09:00
Daniel Barnes
0205debbd3
fix values
2026-03-12 13:22:29 +09:00
Daniel Barnes
0059766509
number formating
2026-03-12 13:17:22 +09:00
Daniel Barnes
be81b02916
additional model files
2026-03-12 13:02:17 +09:00
Daniel Barnes
2dab141166
initial script & model updates
2026-03-12 13:01:35 +09:00
Aiden Cline
45aa49af25
tweak: azure kimi k2.5
2026-03-11 22:35:20 -05:00
Aiden Cline
781fad3ad4
Merge pull request #1150 from cau1k/5.4-family
...
feat(azure): add 5.4/pro families
2026-03-11 22:14:08 -05:00
cau1k
99d2ffcfdd
feat(azure): add 5.4/pro families
2026-03-11 20:59:11 -04:00
Aiden Cline
381d7cc19d
Merge pull request #1149 from ariane-emory/fear/add-march-or-stealth-models
...
Add OpenRouter stealth models: Hunter Alpha and Healer Alpha
2026-03-11 18:07:50 -05:00
Ariane Emory
7482e22458
Fix family field to use 'alpha' for stealth models
2026-03-11 18:49:32 -04:00
Ariane Emory
f5e6a402e6
Add OpenRouter stealth models: Hunter Alpha and Healer Alpha
2026-03-11 18:41:58 -04:00
Aiden Cline
9265852852
tweak: adjust some gh limits to align better w/ api
2026-03-11 15:23:44 -05:00
Aiden Cline
dc98a32996
Merge pull request #1018 from Sewer56/add-synthetic-missing-models
...
Update synthetic.new models: promote MiniMax-M2.5, add GLM-4.7-Flash
2026-03-11 14:55:50 -05:00
Aiden Cline
56c39ae0f6
Merge pull request #1140 from sk0x0y/feature/nanogpt-thudm-id-fixes
...
fix(nano-gpt): rename THUDM 2 ids to canonical THUDM ids
2026-03-11 14:55:07 -05:00
Aiden Cline
b1f43a7595
Merge pull request #1147 from msadiks/fix/alibaba-coding-minimax
...
fix: alibaba-coding-plan MiniMax-M2.5 context window
2026-03-11 14:54:37 -05:00
Matt Cowger
fed8bcae19
Merge branch 'dev' into mcowger/correct-gemini-flash-lite-pricing
2026-03-11 12:23:42 -07:00
sk0x0y
fb95150d02
fix(nano-gpt): rename VongolaChouko model id
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-12 04:20:39 +09:00
sk0x0y
a7c9a240b4
fix(nano-gpt): rename Steelskull model ids
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-12 04:20:39 +09:00
sk0x0y
6432a4a3e6
fix(nano-gpt): rename Sao10K model ids
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-12 04:20:38 +09:00
sk0x0y
f2e4a249fe
fix(nano-gpt): rename NeverSleep model ids
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-12 04:20:38 +09:00
sk0x0y
8667a6eed8
fix(nano-gpt): rename MarinaraSpaghetti model id
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-12 04:19:56 +09:00
sk0x0y
429554397a
fix(nano-gpt): rename LatitudeGames model id
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-12 04:19:56 +09:00
sk0x0y
a64e6ad0ac
fix(nano-gpt): rename LLM360 model id
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-12 04:19:56 +09:00
sk0x0y
d68d79888c
fix(nano-gpt): rename Infermatic model id
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-12 04:19:56 +09:00
sk0x0y
6c52905c6a
fix(nano-gpt): rename Gryphe model id
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-12 04:19:55 +09:00
sk0x0y
62410b8f26
fix(nano-gpt): rename GalrionSoftworks model id
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-12 04:19:55 +09:00
sk0x0y
50ce68ccab
fix(nano-gpt): rename Envoid model ids
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-12 04:19:20 +09:00
sk0x0y
d1c6a6b873
fix(nano-gpt): rename EVA-UNIT-01 model ids
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-12 04:19:20 +09:00
Frank
7193b068a5
update zen models
2026-03-11 13:52:50 -04:00
Sadik
79a8a06bd7
fix MiniMax-M2.5 context window
2026-03-11 20:50:33 +03:00
Aiden Cline
b60c03e11c
Merge pull request #1139 from zainhas/dev
...
[Together AI] add prompt caching pricing for MiniMax m2.5
2026-03-11 12:31:56 -05:00
Aiden Cline
15cf98d57b
Merge pull request #1146 from gotjoshua/patch-1
...
Rename step-3-5-flash.toml to step-3.5-flash.toml
2026-03-11 12:31:39 -05:00
Aiden Cline
b2ee6c407b
Merge pull request #1144 from micuintus/feature/update-nebius-changes
...
Feat: update Nebius changes
2026-03-11 12:31:29 -05:00
gotjoshua
96a14a06e7
Rename step-3-5-flash.toml to step-3.5-flash.toml
...
on nvidia it is 3.5 not 3-5
2026-03-11 11:41:36 +00:00
Michael Voigt
adc358606d
fix(nebius): update model context limits per API
2026-03-11 11:33:14 +01:00
Michael Voigt
63d52adf6f
feat(nebius): add GLM-5 model
2026-03-11 11:33:14 +01:00
sk0x0y
9a31387766
fix(nano-gpt): rename Salesforce model id
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-11 18:17:41 +09:00
sk0x0y
735157b837
fix(nano-gpt): rename ReadyArt model ids
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-11 18:17:41 +09:00
sk0x0y
d75b46fb37
fix(nano-gpt): rename Doctor-Shotgun model id
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-11 18:17:41 +09:00
sk0x0y
cc555f8482
fix(nano-gpt): rename CrucibleLab model id
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-11 18:17:13 +09:00
sk0x0y
7fbbcf2b49
fix(nano-gpt): rename MiniMaxAI model id
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-11 16:04:28 +09:00
sk0x0y
14c8ec8ca5
fix(nano-gpt): rename Tongyi-Zhiwen model id
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-11 16:04:28 +09:00
sk0x0y
72568bbdb3
fix(nano-gpt): rename Alibaba-NLP model id
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-11 16:03:57 +09:00
sk0x0y
c2225b715f
fix(nano-gpt): rename THUDM GLM-Z1 rumination id
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-11 15:49:20 +09:00
sk0x0y
7ce25e3742
fix(nano-gpt): rename THUDM GLM-Z1 model ids
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-11 15:49:20 +09:00
sk0x0y
427868604b
fix(nano-gpt): rename THUDM GLM-4 model ids
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-11 15:49:20 +09:00
Zain Hasan
247cd801a8
add prompt caching pricing for MiniMax m2.5
2026-03-10 22:54:42 -07:00
Aiden Cline
1aa2ee22b1
Merge pull request #1134 from sk0x0y/feature/nanogpt-catalog-fixes
...
fix(nano-gpt): correct TEE path ids and add missing canonical entries
2026-03-10 22:02:52 -05:00
Aiden Cline
0f57233eff
Merge pull request #1105 from sylviezhang37/add-vercel-input-context-and-new-models
...
feat(vercel): add input context calculation + new models
2026-03-10 22:01:52 -05:00
Aiden Cline
73a78eebfc
Merge pull request #1138 from mugnimaestra/feat/add-glm-5-turbo-chutes
...
feat: add GLM-5-Turbo to Chutes provider listings
2026-03-10 22:01:08 -05:00
Sylvie Zhang
3a6789b819
Merge branch 'dev' into add-vercel-input-context-and-new-models
2026-03-10 17:44:14 -07:00
Sylvie Zhang
f7c505e140
remove context from gemini models
2026-03-10 17:43:08 -07:00
Sylvie Zhang
6bb36806d6
only calc input context for openai models
2026-03-10 17:40:46 -07:00
Sylvie Zhang
20a404eb88
revert non openai changes
2026-03-10 17:38:46 -07:00
Muhammad Mugni Hadi
65ecb5cd4a
feat: add GLM-5-Turbo to Chutes provider listings
2026-03-11 05:26:11 +07:00
Matt Cowger
56062a9129
Fix incorrect pricing
2026-03-10 14:57:44 -07:00
Aiden Cline
d3d9c580d4
Merge pull request #1135 from gitpush-gitpaid/fix/gpt-5-4-pdf-input-modalities
...
Added PDF to input modalities for GPT-5.4
2026-03-10 13:53:42 -05:00
gitpush-gitpaid
ef98d8a9cb
Updated GPT-5.4 PDF input modalities
2026-03-10 13:59:29 -04:00
sk0x0y
9d17752b88
fix(nano-gpt): add missing GLM 5 thinking model
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-11 00:56:22 +09:00
sk0x0y
b5a838fe8b
fix(nano-gpt): add missing TEE qwen3.5 model
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-11 00:56:22 +09:00
sk0x0y
aa1ac39ee6
fix(nano-gpt): rename TEE gemma and minimax ids
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-11 00:56:22 +09:00
sk0x0y
4bc17ccf96
fix(nano-gpt): rename TEE oss and llama ids
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-11 00:56:22 +09:00
sk0x0y
08c1899bfe
fix(nano-gpt): rename TEE deepseek model ids
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-11 00:56:02 +09:00
sk0x0y
ad50e4a5ed
fix(nano-gpt): rename TEE qwen model ids
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-11 00:56:02 +09:00
sk0x0y
730915a123
fix(nano-gpt): rename TEE kimi model ids
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-11 00:56:02 +09:00
sk0x0y
6f12d18cb8
fix(nano-gpt): rename TEE glm model ids
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-11 00:56:02 +09:00
Aiden Cline
bd8774db99
Merge pull request #1132 from sk0x0y/feature/nanogpt-model-sync
...
feat(nano-gpt): add text and image models
2026-03-10 10:31:50 -05:00
Aiden Cline
88fbea52a4
Merge pull request #1133 from anomalyco/fix-model
...
fix: bedrock devstral
2026-03-10 10:31:08 -05:00
Aiden Cline
70e5d9b34b
fix: bedrock devstral
2026-03-10 10:30:20 -05:00
Aiden Cline
edb6ef0d71
Merge pull request #1129 from Grin1024/dev
...
feat: add GPT-5 series models to requesty provider
2026-03-10 10:29:30 -05:00
Aiden Cline
df1280ed8b
add families to some bedrock models
2026-03-10 10:12:13 -05:00
sk0x0y
898b3c18b7
feat(nano-gpt): add image models
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-10 22:13:48 +09:00
sk0x0y
6316e543ef
feat(nano-gpt): add text models
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-10 22:13:48 +09:00
Aiden Cline
64d9a97f9d
Merge pull request #1128 from JWahle/dev
...
chore: Updated abacus model definitions
2026-03-10 07:40:01 -05:00
Aiden Cline
c62a2a3fc1
Merge pull request #1130 from Mingholy/fix/alibaba-coding-plan-model-limits
...
fix: update model limits for alibaba-coding-plan providers
2026-03-10 07:39:48 -05:00
Aiden Cline
cc1937a177
Merge pull request #1131 from janszypulski/cloudferro-sherlock-fix-minimax-model-id
...
fix minimax-m2.5 model id - wrong file path
2026-03-10 07:39:34 -05:00
Jan Szypulski
9d3a88863d
fix minimax-m2.5 model id - wrong file path
2026-03-10 11:21:55 +01:00
mingholy.lmh
b9123e26e0
fix: update model limits for alibaba-coding-plan providers
...
- Add MiniMax-M2.5 to alibaba-coding-plan-cn
- Update qwen3-max output limit (65536 -> 32768)
- Update qwen3-coder-plus context limit (1048576 -> 1000000)
- Update MiniMax-M2.5 limits per ref.json (context: 196608, output: 24576)
Co-authored-by: Qwen-Coder <qwen-coder@alibabacloud.com >
2026-03-10 15:45:01 +08:00
lihui
933e450104
feat: add GPT-5 series models to requesty provider
...
Add missing OpenAI GPT-5 series models to requesty provider:
- GPT-5 Chat, Codex, Image, Pro
- GPT-5.1 Chat, Codex, Codex-Max, Codex-Mini
- GPT-5.2 Chat, Codex, Pro
- GPT-5.3 Codex
- GPT-5.4, GPT-5.4 Pro
2026-03-10 14:53:28 +08:00
JWahle
2ba6383e70
chore: Updated abacus model definitions
...
Added: gpt-5.4.toml
Removed: gemini-3-pro-preview.toml
2026-03-10 05:10:45 +01:00
Aiden Cline
65ed6ac5dd
Merge pull request #1126 from mcowger/feature/gemini-3.1-flash-lite-vercel
...
feat: add gemini-3.1-flash-lite-preview to vercel gateway provider
2026-03-09 21:43:33 -05:00
Aiden Cline
e7d04aec7a
Merge pull request #1127 from anomalyco/add-shape
...
feat: add 'shape' field to provider so models can specify if they use responses vs completions apis (use only if model only supports 1 of)
2026-03-09 21:43:04 -05:00
Aiden Cline
37fe334aed
feat: add 'shape' field to provider so models can specify if they use responses vs completions apis (use only if model only supports 1 of)
2026-03-09 21:42:23 -05:00
Matt Cowger
ce8fc9e4f0
feat: add gemini-3.1-flash-lite-preview to vercel gateway provider
2026-03-09 19:32:43 -07:00
Aiden Cline
be8eb8ba54
fix name
2026-03-09 20:01:02 -05:00
Aiden Cline
7c625b3b82
Merge pull request #945 from Daltonganger/feat/nano-gpt-sync-models-api
...
sync nano-gpt models with live API catalog
2026-03-09 20:00:06 -05:00
Aiden Cline
33700d27dc
Merge pull request #1032 from propilideno/feature/new_gpt_5.3_codex_and_missing_structured_output_attr
...
Add gpt-5.3-codex (Azure) and fill missing structured output flags
2026-03-09 19:40:20 -05:00
Aiden Cline
e5c300a5e5
fix
2026-03-09 19:38:30 -05:00
Aiden Cline
e5e9175c5d
Merge branch 'dev' into feature/new_gpt_5.3_codex_and_missing_structured_output_attr
2026-03-09 19:37:40 -05:00
Aiden Cline
a9f79d6794
Merge pull request #1123 from dpuyosa/feature/venice-gpt54-multimodal
...
Venice: Add GPT-5.4 Pro and enable multimodal inputs for GPT-5.4 & Qwen3.5
2026-03-09 18:19:52 -05:00
Aiden Cline
fd4c4a8f28
Merge pull request #1038 from muldercw/add-clarifai-model-provider
...
Add Clarifai Model Provider
2026-03-09 18:19:14 -05:00
dpuyosa
f9b5385868
[venice] Add GPT-5.4 Pro and enable multimodal inputs
...
- Add GPT-5.4 Pro model
- Enable attachment/image input for GPT-5.4
- Enable attachment/image/video input for Qwen3.5 35B A3B
2026-03-09 22:52:54 +01:00
Aiden Cline
b2f7a72410
Merge pull request #1110 from fhennerkes/dev
...
poe: add GPT-5.4, GPT-5.4-Pro, and GPT-5.3-Instant models
2026-03-09 14:10:12 -05:00
Aiden Cline
7b5d9aa645
Merge pull request #1025 from liuchang-reolink/dev
...
add qwen3.5-397b-a17b and step-3-5-flash for nvidia
2026-03-09 14:05:08 -05:00
Aiden Cline
6e0040dbfd
Merge pull request #1089 from Krule/krule/update_gitlab_anthropic_context_size
...
feat(gitlab): update context limit to 1M for Claude Sonnet and Opus 4.6
2026-03-09 14:03:49 -05:00
Aiden Cline
943ad8481b
Merge pull request #1121 from illusion77/fix/chutes-mimo-v2-flash-context-16709
...
fix(chutes): correct MiMo-V2-Flash context window and capabilities
2026-03-09 14:02:51 -05:00
Aiden Cline
7f1b6fb0eb
Merge pull request #1122 from riccardogiorato/dev
...
remove deprecated kimi models from together.ai
2026-03-09 14:02:36 -05:00
Riccardo Giorato
23eff95e5d
remove deprecated kimi from together.ai
2026-03-09 17:30:40 +01:00
illusion77
ddbd396205
fix(chutes): correct MiMo-V2-Flash context window and capabilities
...
The chutes provider had incorrect metadata for MiMo-V2-Flash:
context 32K → 262K, output 8K → 32K, reasoning and tool_call enabled.
Fixes anomalyco/opencode#16709
2026-03-09 10:57:57 -05:00
Aiden Cline
f3ee1a530b
Merge pull request #1120 from stephenkuhn214/dev
...
Add Amazon-Bedrock Devstral 2 123B model
2026-03-09 09:35:46 -05:00
Aiden Cline
9c51b65440
Merge pull request #1119 from cgilly2fast/dev
...
fix(firmware): proper 5.3 codex model id
2026-03-09 09:30:35 -05:00
Frank
353aeb4998
update zen models
2026-03-09 10:08:55 -04:00
Frank
11991fecb5
update zen models
2026-03-09 10:03:13 -04:00
stephenkuhn214
1b4599773d
Create mistral.devstral-2-123b
2026-03-09 08:58:19 -04:00
Colby Gilbert
78e1a3b0c9
fix(firmware): proper 5.3 codex model id
2026-03-08 21:58:27 -07:00
Sewer56
7a02946620
Update synthetic models: promote MiniMax-M2.5, add GLM-4.7-Flash, remove deprecated Qwen3.5
2026-03-08 22:56:31 +00:00
Aiden Cline
44686797c8
Merge pull request #1118 from shelvick/add-azure-gpt-5.3-chat
...
Add GPT-5.3 Chat to Azure
2026-03-08 16:41:10 -05:00
Aiden Cline
065cec8431
fix: input limit for context
2026-03-08 16:40:38 -05:00
Scott Helvick
f491c2bec9
Add GPT-5.3 Chat to Azure
2026-03-08 21:20:27 +00:00
Aiden Cline
cf1ac3053f
Merge pull request #1081 from djmaze/fix/nebius-model-casing
...
fix(nebius): correct model ID casing to match Token Factory API
2026-03-08 14:26:52 -05:00
Aiden Cline
6be1e929fc
Merge pull request #1114 from v1gnesh/dev
...
add grok 4.2 experimentals
2026-03-08 14:25:12 -05:00
Aiden Cline
49524827e2
Merge pull request #1113 from shelvick/add-vertex-glm-5
...
Fix GLM-5 context window size on Google Vertex
2026-03-08 10:31:05 -05:00
Aiden Cline
5ab5d389fc
Merge pull request #1112 from cau1k/feat/az-5.4
...
feat(azure): add gpt-5.4/5.4-pro
2026-03-08 10:30:54 -05:00
Aiden Cline
d5367ed978
Merge pull request #1116 from xiaojiezj/xj_dev_0308
...
fix: Adjust the logo for ZenMux
2026-03-08 10:30:18 -05:00
Aiden Cline
5069faa25b
Merge pull request #1117 from kailiu42/feat/siliconflow-cn
...
feat(siliconflow-cn): add Qwen3.5 model family
2026-03-08 10:29:48 -05:00
Kai Liu
b279f33d9b
feat(siliconflow-cn): add Qwen3.5 model family
...
New models:
- Qwen/Qwen3.5-4B
- Qwen/Qwen3.5-9B
- Qwen/Qwen3.5-27B
- Qwen/Qwen3.5-35B-A3B
- Qwen/Qwen3.5-122B-A10B
- Qwen/Qwen3.5-397B-A17B
Signed-off-by: Kai Liu <kraml.liu@gmail.com >
2026-03-08 20:07:40 +08:00
xiaojie.zj
fe8249d706
fix: Adjust the logo
2026-03-08 16:36:01 +08:00
skywalker512
236af40da3
feat: add Tencent Coding Plan provider
...
Add support for Tencent Coding Plan with 8 models:
- Auto (tc-code-latest)
- Hunyuan 2.0 Instruct
- Hunyuan 2.0 Think
- Hunyuan-T1
- Hunyuan-TurboS
- MiniMax-M2.5
- Kimi-K2.5
- GLM-5
Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com >
2026-03-08 15:34:36 +08:00
v1gnesh
a24a23d57b
add grok 4.2 experimentals
2026-03-08 07:42:09 +05:30
Scott Helvick
7e4773d9b5
Fix GLM-5 context window size on Google Vertex
...
Correct the context limit from 204800 to 202752 tokens.
2026-03-07 22:20:56 +00:00
zero
9f937f3fc5
Merge branch 'dev' into feat/az-5.4
2026-03-07 17:20:29 -05:00
cau1k
8cbdbc1102
add day cutoff
2026-03-07 17:19:15 -05:00
cau1k
f1ca3b0015
feat(azure-cognitive-services): symlink 5.4/pro from azure provider
...
;
2026-03-07 17:02:26 -05:00
cau1k
35c757bad5
feat(azure): add 5.4/pro
2026-03-07 17:00:43 -05:00
fhennerkes
78781f3901
poe: add GPT-5.4, GPT-5.4-Pro, and GPT-5.3-Instant models
2026-03-07 13:58:35 -08:00
Sylvie Zhang
26465319d6
Merge branch 'dev' into add-vercel-input-context-and-new-models
2026-03-07 11:39:53 -08:00
fhennerkes
1371cbf9de
poe: add GPT-5.4, GPT-5.4-Pro, and GPT-5.3-Instant models
2026-03-07 09:48:17 -08:00
Aiden Cline
559ccd6966
Merge pull request #1024 from yinxulai/feat/qiniu-ai
...
feat(qiniu-ai): add new model configurations
2026-03-07 11:26:01 -06:00
Aiden Cline
2691cb4e8d
Merge pull request #1083 from samzong/feat/add-drun-provider
...
feat: add d.run(China) provider (OpenAI-compatible)
2026-03-07 11:25:12 -06:00
Aiden Cline
83ed1f0125
Merge pull request #1015 from RioPlay/dev
...
add: newer MiniMax, GLM, and Kimi models to DeepInfra
2026-03-07 11:24:53 -06:00
Aiden Cline
869f831466
Merge branch 'dev' into dev
2026-03-07 11:23:37 -06:00
Aiden Cline
5a673af2ae
Merge pull request #1061 from JonasGao/dev
...
Add Qwen3.5 Flash & GLM-5 & M2.5 models to alibaba-cn
2026-03-07 11:22:56 -06:00
Aiden Cline
9b8543a074
Add interleaved section to minimax-m2.5.toml
2026-03-07 11:21:54 -06:00
Aiden Cline
b3fb902331
Merge pull request #1030 from Mingholy/feat/alibaba-coding-plan-cn
...
feat(alibaba-coding-plan-cn): add Coding Plan provider for China region
2026-03-07 11:21:26 -06:00
Aiden Cline
adb0c0b305
Merge pull request #1062 from viitana/bump-deepseek-details
...
feat: [deepseek]: update official DeepSeek model details
2026-03-07 11:21:21 -06:00
Aiden Cline
ed01410d82
Merge pull request #1088 from mcowger/feature/gemini-3.1-flash-lite
...
feat: add gemini-3.1-flash-lite-preview model
2026-03-07 11:13:45 -06:00
Aiden Cline
7e23b780cc
Merge pull request #1077 from evroc-oss/evroc/correct-model-config
...
fix(evroc): correct model config
2026-03-07 11:13:15 -06:00
Aiden Cline
c46b652c8e
Merge pull request #1076 from jerome-benoit/feat/add-sonar-deep-research-sap-ai-core
...
feat(sap-ai-core): add Perplexity Sonar Deep Research model
2026-03-07 11:11:29 -06:00
Aiden Cline
ddb74e9b09
Merge pull request #1063 from dpuyosa/fix/models-pricing-limits-update
...
Venice: Update model pricing and limits
2026-03-07 11:10:37 -06:00
Aiden Cline
face36ecb8
Merge pull request #1064 from BlockListed/fix-cortecs-models
...
Fix Cortecs models
2026-03-07 11:10:03 -06:00
Aiden Cline
6f170651b3
Merge pull request #1075 from Track07-cda/alibaba-cn-third-party-models
...
Add third party providers' models to alibaba-cn provider
2026-03-07 11:09:40 -06:00
Aiden Cline
47dbe45dd5
Merge pull request #1066 from dpuyosa/feat/add-qwen3-5-35b-a3b
...
Venice: Add Qwen 3.5 35B A3B model
2026-03-07 11:09:10 -06:00
Aiden Cline
0ee43b64b3
Merge branch 'dev' into alibaba-cn-third-party-models
2026-03-07 11:08:37 -06:00
Aiden Cline
6130a1f74e
Merge pull request #1068 from MauroDruwel/dev
...
NVIDIA: Add MiniMax M2.5 model and remove MiniMax M2
2026-03-07 11:07:17 -06:00
Aiden Cline
c0c82a5f04
Merge pull request #1072 from sylviezhang37/update-vercel-models-20260302-1656
...
Update Vercel models
2026-03-07 11:06:17 -06:00
Aiden Cline
b0ba8b14d5
Merge pull request #1092 from janszypulski/cloudferro-sherlock-add-minimax-2.5
...
add MiniMaxAI/MiniMax-M2.5 to CloudFerro Sherlock
2026-03-07 11:02:11 -06:00
Aiden Cline
4780f9ddc1
Merge pull request #1109 from dinhkim/feat/add-cf-glm-4.7-flash
...
feat: add GLM-4.7-Flash to the Cloudflare Workers AI provider
2026-03-07 11:01:50 -06:00
Aiden Cline
27e02de632
Merge pull request #1078 from SomeoneWithOptions/dev
...
add gpt 5.3 codex for openrouter and Mercury models
2026-03-07 11:01:41 -06:00
Aiden Cline
f22c827045
Merge branch 'dev' into dev
2026-03-07 11:01:17 -06:00
Aiden Cline
cfc4585ed7
Merge pull request #1107 from Rinuuri/deepinfra-glm5
...
Add deepinfra GLM-5 model
2026-03-07 10:59:16 -06:00
Aiden Cline
fa07bc2088
Merge pull request #1039 from rholak/add-abacus-models
...
Add sonnet 4.6 and opus 4.6 to abacus model list
2026-03-07 10:59:00 -06:00
Aiden Cline
497b1daaf2
Merge pull request #1103 from dpuyosa/feat/venice-add-gpt-models
...
Venice: Add OpenAI GPT-4o, GPT-4o Mini, GPT-5.4 models
2026-03-07 10:58:44 -06:00
Aiden Cline
442afa8c7e
Merge pull request #1060 from yanismiraoui/inception/mercury2
...
Add Inception Mercury 2 and Mercury Edit models
2026-03-07 10:57:38 -06:00
Aiden Cline
4bd0c387fe
Merge pull request #1044 from shrwnsan/feat/openrouter-routers
...
feat(openrouter/free): add free router
2026-03-07 10:57:24 -06:00
Aiden Cline
b8c0c1d3a1
Merge pull request #1053 from laiiihz/update-xiaomi-models
...
Update Xiaomi models metadata
2026-03-07 10:57:17 -06:00
Aiden Cline
5c6c3e5a32
Merge pull request #1055 from shantanugoel/gemini-3.1-flash-image-preview
...
Add Gemini 3.1 Flash Image Preview
2026-03-07 10:57:06 -06:00
Aiden Cline
53d3cca3a0
Merge pull request #1052 from spiffytech/dev
...
Improve Ollama Cloud generator. Remove Gemini 3 Pro from Ollama Cloud.
2026-03-07 10:56:45 -06:00
Aiden Cline
105970c173
Merge pull request #1049 from heimoshuiyu/fix/glm-5-open-weights
...
fix: mark GLM-5 as open weights
2026-03-07 10:56:31 -06:00
Aiden Cline
788ee04034
Merge pull request #1045 from xinrui-z/aihubmix-add-models
...
aihubmix add models
2026-03-07 10:56:03 -06:00
Aiden Cline
6626db4044
Merge pull request #1098 from JWahle/dev
...
chore: updated abacus model definitions
2026-03-07 10:55:31 -06:00
Aiden Cline
8902640664
Merge pull request #1023 from PandaSt0rm/add-alibaba-coding-plan
...
Add Alibaba Coding Plan provider and model configs
2026-03-07 10:53:34 -06:00
Kim Truong
cab247ddf8
update context to match Cloudflare doc
2026-03-07 23:50:06 +07:00
Kim Truong
c1a42fa0a0
feat: add GLM-4.7-Flash mode in Cloudflare Workers AI provider
2026-03-07 23:45:52 +07:00
Aiden Cline
35023bba5a
Merge pull request #1001 from ItsWendell/feat/bedrock-bearer-token
...
Add AWS_BEARER_TOKEN_BEDROCK to Amazon Bedrock provider env
2026-03-07 09:52:21 -06:00
Aiden Cline
604e49792b
Merge pull request #1002 from DEAN-Cherry/feat/add-minimax-m2.5
...
models: alibaba-cn: add MiniMax-M2.5
2026-03-07 09:51:36 -06:00
Aiden Cline
0ca77b0cda
Merge branch 'dev' into dev
2026-03-07 09:50:26 -06:00
Aiden Cline
ea9505a40f
Merge pull request #1004 from BlockListed/cortecs-models
...
Add Cortecs AI models
2026-03-07 09:50:07 -06:00
Aiden Cline
990b8d7308
Merge pull request #1005 from cgilly2fast/dev
...
feat(firmware): gemini 3.1 pro, sonnet reasoning
2026-03-07 09:49:54 -06:00
Aiden Cline
ec173e86d4
Merge pull request #996 from fhennerkes/dev
...
poe: add Gemini-3.1-Pro, GPT-5.3-Codex and Gemini 3.1 Flash Lite
2026-03-07 09:47:52 -06:00
Aiden Cline
4a6e92a7c9
Merge pull request #997 from xiaojiezj/zenmux_dev_0221
...
feat: add Gemini 3.1 Pro Preview for ZenMux provider
2026-03-07 09:47:37 -06:00
Aiden Cline
f0f686bdf5
Merge pull request #999 from mikalsande/mistral_latest
...
Append (latest) to Mistral models that refer to the latest version.
2026-03-07 09:46:40 -06:00
Aiden Cline
35ff0c2629
Merge pull request #995 from Phoen1xCode/dev
...
fix(zenmux:minimax): remove duplicated prefix & feat(zenmux:openai): add GPT-5.2-Pro model
2026-03-07 09:45:07 -06:00
Aiden Cline
f99e9e89df
Merge pull request #1090 from litvix-whale/feat/add-minimax-m2-5
...
feat(provider): add MiniMax M2.5 for DeepInfra
2026-03-07 09:41:28 -06:00
Armin Pašalić
09722ac264
Merge branch 'anomalyco:dev' into krule/update_gitlab_anthropic_context_size
2026-03-07 13:17:10 +01:00
Rinuuri
fa67d00aeb
Update GLM-5.toml
2026-03-06 21:23:42 +00:00
Rinuuri
ddd2dd73ed
Adding deepinfra GLM-5
2026-03-07 00:03:29 +03:00
fhennerkes
d7929fd00b
Merge branch 'anomalyco:dev' into dev
2026-03-06 12:00:20 -08:00
Frank
06e7d4db42
Merge pull request #1014 from NachoFLizaur/fix/bedrock-opus-4-6-context-window
...
fix(amazon-bedrock): correct Claude Opus 4.6 context window from 1M to 200K
2026-03-06 11:25:37 -05:00
Sylvie Zhang
7a11ef241d
update more models
2026-03-06 08:24:49 -08:00
Sylvie Zhang
145862315d
add input calculation + new models
2026-03-06 08:07:11 -08:00
dpuyosa
d871710ba4
[venice] Add OpenAI GPT-4o, GPT-4o Mini, GPT-5.4 models
...
- Add gpt-4o-2024-11-20 model configuration
- Add gpt-4o-mini-2024-07-18 model configuration
- Add gpt-5.4 model configuration with reasoning capability
2026-03-06 09:53:06 +01:00
Colby Gilbert
16486087c6
Merge branch 'anomalyco:dev' into dev
2026-03-05 21:38:25 -08:00
Frank
2939af9330
Merge pull request #1100 from sachnun/feat/github-copilot-gpt-5-4
...
feat(provider): add gpt-5.4 for GitHub Copilot
2026-03-05 23:33:57 -05:00
sachnun
7c68dab3bb
feat(provider): add gpt-5.4 for GitHub Copilot
2026-03-06 11:18:11 +07:00
Mike Soylu
caceb0b310
openrouter openai models ( #1099 )
2026-03-05 22:26:58 -05:00
Frank
7a0d3be1e7
Update zen models
2026-03-05 18:55:49 -05:00
ShivamB25
e11ad7c01a
feat(openai): add GPT-5.4 and GPT-5.4 Pro model specs ( #1095 )
2026-03-05 18:50:22 -05:00
Matt Silverlock
d30fa82e4c
Cloudflare: add gpt-5.4.toml ( #1096 )
2026-03-05 18:50:10 -05:00
Rishi Vhavle
771102a960
feat: add gpt-5.3-codex to github-copilot provider ( #1097 )
2026-03-05 18:49:56 -05:00
JWahle
30f98b15ef
chore: updated abacus model definitions
...
Added: GPT-5 Codex, GPT-5.1/5.2/5.3 Codex, GPT-5.3 Chat, Gemini 3.1 Flash Lite/Pro Preview, Claude Opus/Sonnet 4.6, Kimi K2.5, GLM-5
Removed: Gemini 2.0 Flash 001, Gemini 2.0 Pro Exp, Meta-Llama 3.1 70B Instruct
Updated pricing: DeepSeek V3.1, GLM-4.7, GPT-5.2 Chat Latest, o3-pro, Route LLM
2026-03-06 00:46:19 +01:00
Colby Gilbert
6f7ab479fb
feat(firmware): gpt 5.4
2026-03-05 13:23:08 -08:00
Colby Gilbert
a4efbcd5ce
Merge branch 'anomalyco:dev' into dev
2026-03-05 13:15:49 -08:00
Frank
bcbfba03bd
update zen models
2026-03-05 15:51:33 -05:00
Frank
bdb5dac941
update zen models
2026-03-05 15:50:03 -05:00
Frank
1538bdcedb
update zen models
2026-03-05 13:31:27 -05:00
SomeoneWithOptions
e5211f3105
add inception mercury models for openrouter
2026-03-05 12:13:52 -05:00
Andres Castellanos
4a2209dbd4
Merge branch 'anomalyco:dev' into dev
2026-03-05 11:51:38 -05:00
Jan Szypulski
900014fe52
add MiniMax-M2.5
2026-03-05 14:58:19 +01:00
Kyrylo Lytvishko
5bbaf3c3f2
feat(provider): add MiniMax M2.5 for DeepInfra
2026-03-05 14:04:11 +02:00
Armin Pasalic
7dd0a26ff4
feat(gitlab): update context limit to 1M for Sonnet and Opus 4.6
2026-03-05 12:01:02 +01:00
Matt Cowger
3b7e0f02f1
feat: add gemini-3.1-flash-lite-preview model
2026-03-04 13:19:21 -08:00
samzong
e67f921ea3
feat: add official d.run logo
2026-03-04 13:42:01 +08:00
samzong
f5411eeeda
feat: add D.Run (China) provider with minimax-m25, deepseek-r1, deepseek-v3
2026-03-04 13:33:04 +08:00
Frank
0d83ab8909
Merge pull request #1082 from kesku/kesku/add-ppl-agent-api
...
Add Perplexity Agent API provider
2026-03-03 23:03:49 -05:00
Kesku
26a629debc
add models
2026-03-03 23:19:01 +00:00
Kesku
4b3319561b
set up provider
2026-03-03 23:10:43 +00:00
Ubuntu
b89ce0d986
fix(nebius): correct model ID casing to match Token Factory API
...
Fix lowercase model ID bug that caused "The model does not exist" errors.
- qwen/ → Qwen/ directory
- Fixed model file casing to match API exactly across all providers
2026-03-03 22:11:57 +00:00
fhennerkes
1c01f8172b
poe: add Gemini-3.1-Flash-Lite and update gpt-4o-mini context
...
Add new Gemini 3.1 Flash Lite model
Update gpt-4o-mini context window: 128K → 124,096
2026-03-03 11:24:20 -08:00
SomeoneWithOptions
d76040c514
add gpt 5.3 codex for openrouter
2026-03-03 12:53:39 -05:00
Simon Rygård
feffa8119f
fix(evroc): correct modality config
2026-03-03 16:52:07 +01:00
Simon Rygård
59c6e5df62
fix(evroc): correct tool call config
2026-03-03 16:51:47 +01:00
Jérôme Benoit
fe2204d42c
feat(sap-ai-core): add Perplexity Sonar Deep Research model
2026-03-03 14:56:18 +01:00
Track07-cda
07cc5335ac
Add third party providers' models to alibaba-cn provider
...
- Add `MiniMax/MiniMax-M2.5` and `kimi/kimi-k2.5` to the `alibaba-cn`
provider.
- Update `kimi-k2.5` to include video modality and adjust release/update
dates.
- Add several `siliconflow/deepseek` models to the `alibaba-cn`
provider.
2026-03-03 16:46:32 +08:00
github-actions[bot]
fefbb90a29
chore(vercel): update Vercel model definitions
...
Auto-generated by weekly workflow from Vercel AI Gateway API.
Co-Authored-By: github-actions[bot] <41898282+github-actions[bot]@users.noreply.github.com>
2026-03-02 16:56:48 +00:00
Frank
fec48b83d3
update zen models
2026-03-01 13:23:08 -05:00
Mauro Druwel
c30bbe7718
Add knowledge
2026-03-01 08:53:22 +01:00
Mauro Druwel
1fba668f0f
Add minimax-m2.5 to nvidia-nim and remove deprecated minimax-m2 from nvidia-nim
2026-03-01 08:52:33 +01:00
Aiden Cline
33ec088bda
Merge pull request #1008 from friendliai/feat/friendli-minimax-m2.5
...
add friendli minimax m2.5 model config
2026-03-01 07:54:08 +05:00
Aiden Cline
add7f9a914
Merge pull request #1065 from friendliai/minpeter/remove-exaone-models
...
Remove all EXAONE models
2026-03-01 07:53:42 +05:00
dpuyosa
369fa2de6d
[venice] Add Qwen 3.5 35B A3B model
...
- Add new model configuration for Qwen 3.5 35B A3B
- Includes cost, limits, and capabilities (reasoning, tool_call, structured_output)
2026-02-28 21:07:28 +01:00
minpeter
c8732e7e74
Remove all EXAONE models
...
Remove LGAI-EXAONE model definitions (EXAONE-4.0.1-32B, K-EXAONE-236B-A23B)
and related family references from core packages.
Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com >
2026-03-01 04:48:08 +09:00
Jonas
f00f9f3c11
Add Qwen3.5 Flash & GLM-5 & M2.5 models to alibaba-cn provider
2026-02-28 23:34:06 +08:00
BlockListed
c87ca238de
fix cortecs models
...
the should have periods not a p as a decimal separator
2026-02-28 15:16:59 +01:00
BlockListed
102f55aeef
add glm 4.7 flash model to cortecs
2026-02-28 15:13:53 +01:00
BlockListed
2767754a02
add kimi K2.5 model to cortecs
2026-02-28 15:13:53 +01:00
dpuyosa
4289b59a04
[models] Update model pricing and limits
...
- Update Claude Sonnet 4-6 pricing and output limit
- Update Grok 41 Fast pricing and context/limits
2026-02-28 14:24:16 +01:00
Atte Viitanen
b986313f42
feat: [deepseek]: update official deepseek model details
2026-02-28 13:22:26 +02:00
yanismiraoui
28cfd4cab6
naming mercury 2 and mercury edit for inception provider
2026-02-27 17:45:01 -08:00
yanismiraoui
b93e62fc3a
Add Inception Mercury 2 and Mercury Edit models
2026-02-27 17:41:00 -08:00
Aiden Cline
e23b5ab010
Merge pull request #1026 from ryot/venice
...
Venice: Add GPT-5.3 Codex
2026-02-28 06:30:13 +05:00
Aiden Cline
b5f6024868
Merge pull request #1020 from jerome-benoit/feat/sap-ai-core-add-models
...
feat(sap-ai-core): Add GPT-4.1, Gemini 2.5 Flash Lite, Perplexity Sonar, and Claude 4.6 models
2026-02-28 06:29:28 +05:00
Aiden Cline
07db15e984
Merge pull request #1029 from dpuyosa/veniceScript
...
Venice: Remove interactive API key prompt & use new maxCompletionTokens field
2026-02-28 06:28:33 +05:00
Aiden Cline
74abf8851a
Merge pull request #1042 from SomeoneWithOptions/dev
...
add gemini 3.1 pro preview custom tools for openrouter
2026-02-28 06:27:56 +05:00
Aiden Cline
7e13ecdfd9
Merge pull request #1056 from xezpeleta/fix/azure-gpt-5-3-codex
...
fix(azure): add gpt-5.3-codex model
2026-02-28 06:27:34 +05:00
Aiden Cline
6ad2c28b2d
Merge pull request #1048 from dpuyosa/feat/add-venice-models
...
Venice: Add NVIDIA Nemotron 3 Nano and Qwen 3 Coder Turbo models
2026-02-28 06:27:20 +05:00
Frank
a124036692
update zen models
2026-02-27 16:16:37 -05:00
Xabi Ezpeleta
d37d362cc8
fix(azure): add gpt-5.3-codex model
2026-02-27 16:41:11 +01:00
Shantanu Goel
c387f94c8e
Add Gemini 3.1 Flash Image Preview
2026-02-27 20:03:41 +05:30
laiiihz
45457c34d8
update xiaomi models detail
2026-02-27 14:56:16 +08:00
spiffytech
44774ec3d6
Ollama Cloud removed support for Gemini 3 Pro
2026-02-26 17:18:34 -05:00
spiffytech
c8fdcf80dd
Updated Ollama Cloud generator to delete old models, only write out files if they changed
2026-02-26 17:18:33 -05:00
fhennerkes
9d33b6409c
Merge branch 'anomalyco:dev' into dev
2026-02-26 12:04:52 -08:00
Matt Silverlock
c76586a174
Cloudflare: add codex models to AI Gateway ( #1050 )
...
* add gpt-5.2-codex
* add gpt-5.3-codex
* Update gpt-5.2-codex.toml
* Update gpt-5.3-codex.toml
2026-02-26 14:41:12 -05:00
Jérôme Benoit
2a267614aa
feat(sap-ai-core): add Claude Opus 4.6 and Sonnet 4.6 models
2026-02-26 17:58:02 +01:00
PandaSt0rm
aac62378b2
Update MiniMax-M2.5 guidance per Alibaba docs
2026-02-26 17:20:58 +02:00
David Hill
56cc5f71bf
fix(ui): opencode zen logo update
2026-02-26 11:09:25 +00:00
David Hill
df2c87d32a
fix(ui): opencode go logo
2026-02-26 11:09:13 +00:00
heimoshuiyu
ff41c2b6c3
fix: mark GLM-5 as open weights
...
GLM-5 is an open-source model, but several provider config files
incorrectly had open_weights set to false. This commit corrects
all GLM-5 configurations to properly reflect its open-source status.
Affected providers:
- zhipuai
- zhipuai-coding-plan
- zai
- zai-coding-plan
- zenmux
- vercel
- siliconflow
- siliconflow-cn
- meganova
2026-02-26 18:43:28 +08:00
dpuyosa
1d137e2f1f
[venice] Add NVIDIA Nemotron 3 Nano and Qwen 3 Coder models
...
- Add NVIDIA Nemotron 3 Nano 30B A3B model configuration
- Add Qwen 3 Coder 480B A35B Instruct Turbo model configuration
2026-02-26 10:53:00 +01:00
dpuyosa
16720bcd1a
[venice] Use maxCompletionTokens for output limit
...
- Add optional maxCompletionTokens field to model spec schema
- Use maxCompletionTokens when calculating output token limit instead of checking existing limit
2026-02-26 10:24:43 +01:00
Xinrui
1feaf76749
aihubmix add models
2026-02-26 16:12:51 +08:00
shrwnsan
080ef5cc9e
fix(openrouter): remove auto router and add missing limit.input
...
- Remove auto router (cost varies, doesn't fit schema)
- Add limit.input = 200_000 to free.toml (schema requirement)
OpenRouter's auto router has 'pricing varied' - it charges based on the
routed model. This doesn't fit the numeric cost schema required by
models.dev, so we're removing it. The free router is retained as it
genuinely costs $0.
2026-02-26 14:40:38 +08:00
Ryo Tulman
f8121c8dc3
Update Venice GPT 5.3 Codex output limit
2026-02-26 00:32:13 -06:00
shrwnsan
d2d5c5a7cc
feat: add openrouter free and auto routers
2026-02-26 10:51:55 +08:00
SomeoneWithOptions
09d9e91d83
add gemini 3.1 pro preview custom tools for openrouter
2026-02-25 15:13:43 -05:00
Robert Holak
930d6a94b8
Add sonnet 4.6 and opus 4.6 to abacus model list
2026-02-25 12:31:06 -06:00
mulder
b9217aff8e
Add Clarifai Model Provider
...
Add Clarifai as a new provider with 11 models:
- GPT OSS 20B, GPT OSS 120B High Throughput
- Ministral 3 14B/3B Reasoning 2512
- Qwen3 Coder 30B, Qwen3 30B Instruct/Thinking 2507
- MiniMax-M2.5 High Throughput
- Trinity Mini, DeepSeek OCR, MM Poly 8B
Also adds 'mm-poly' family to family.ts for the Clarifai multimodal model.
2026-02-25 12:28:40 -05:00
Lucas Almeida
c240bce614
fix: adding missing structured_output parameter
2026-02-25 11:19:54 -03:00
Lucas Almeida
843a1d182a
feat: adding gpt-5.3-codex for Azure Foundry
2026-02-25 11:09:10 -03:00
PandaSt0rm
443c06ca03
fix MiniMax M2.5 modalities in Alibaba Coding Plan
...
- set MiniMax-M2.5 input modalities to text-only
- keep output modality as text
- validate with bun validate
2026-02-25 13:10:58 +02:00
PandaSt0rm
84466021ca
add MiniMax M2.5 to Alibaba Coding Plan and align third-party limits
...
- add MiniMax-M2.5 model config under providers/alibaba-coding-plan/models
- update GLM-4.7 limits to 202,752 context / 16,384 output
- update GLM-5 limits to 202,752 context / 16,384 output
- update Kimi K2.5 output limit to 32,768
- validate with bun validate
2026-02-25 13:05:18 +02:00
mingholy.lmh
b995e90cf5
fix: update context and output limits for alibaba-coding-plan-cn models
...
Update model limits:
- qwen3-coder-plus: context 1_048_576 → 1_000_000
- glm-5: output 131_072 → 16_384
- glm-4.7: output 131_072 → 16_384
- kimi-k2.5: output 65_536 → 32_768
Co-authored-by: Qwen-Coder <qwen-coder@alibabacloud.com >
2026-02-25 17:44:49 +08:00
dpuyosa
291e2eefe9
[venice] Remove interactive API key prompt
...
- Remove readline import and promptForApiKey function
- Remove prompt fallback, rely on CLI arg or env var only
- Update README to reflect change
2026-02-25 09:51:50 +01:00
Sewer56
0428299773
Added: Qwen3.5-397B natively supports image, MM2.5 No Image as it was a mistake.
2026-02-25 08:10:42 +00:00
Frank
96e9537b34
update zen models
2026-02-25 01:05:35 -05:00
Ryo Tulman
09dc7060ac
Venice: Add GPT-5.3 Codex
2026-02-24 23:37:03 -06:00
Colby Gilbert
a463717783
chore(firmware): remove gpt-5
2026-02-24 20:33:49 -08:00
Colby Gilbert
6d721dd32d
feat(firmware): gpt-5.3-codex
2026-02-24 20:32:17 -08:00
liuchang-reolink
3ae513785a
add step-3-5-flash for nvidia
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-02-25 12:05:19 +08:00
Colby Gilbert
e8d667b628
Merge branch 'anomalyco:dev' into dev
2026-02-24 20:01:26 -08:00
liuchang-reolink
7616a65e63
add qwen3.5-397b-a17b for nvidia
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-02-25 11:40:15 +08:00
yinxulai
a5b3da12c1
chore(qiniu-ai): update provider config
2026-02-25 10:20:55 +08:00
yinxulai
a05fbc3604
feat(qiniu-ai): add new model configurations
2026-02-25 10:16:40 +08:00
Aiden Cline
2189030e57
Merge pull request #1022 from armishra/feat/add-minimax-m2.5-baseten
...
feat(provider): Add MiniMax-M2.5 for baseten
2026-02-24 17:28:04 -06:00
Aiden Cline
c7ecc08442
Merge pull request #1019 from dpuyosa/venice
...
Venice: Update gemini-3-1-pro-preview config
2026-02-24 17:27:46 -06:00
Aiden Cline
830046e45e
Merge pull request #1021 from sylviezhang37/update-vercel-models-20260224-2134
...
Update Vercel models
2026-02-24 17:27:19 -06:00
Aiden Cline
c7b26477b9
Update cache_read value in gemini-3.1-pro-preview.toml
2026-02-25 04:26:56 +05:00
Aiden Cline
9e60f516fa
Update cost input and output values in TOML file
2026-02-25 04:26:18 +05:00
PandaSt0rm
7ccbb58c5f
add Alibaba Coding Plan provider and model configs
2026-02-25 01:10:42 +02:00
Archit Mishra
da906a0816
feat(provider): Add MiniMax-M2.5 for baseten
2026-02-24 14:33:41 -08:00
github-actions[bot]
36c9f82905
chore(vercel): update Vercel model definitions
...
Auto-generated by weekly workflow from Vercel AI Gateway API.
Co-Authored-By: github-actions[bot] <41898282+github-actions[bot]@users.noreply.github.com>
2026-02-24 21:34:20 +00:00
Frank
978214e143
update zen models
2026-02-24 15:23:24 -05:00
Jérôme Benoit
6aa69460a1
feat(sap-ai-core): add GPT-4.1, Gemini 2.5 Flash Lite, and Sonar models
...
Add 5 new model definitions for SAP AI Core provider:
- gpt-4.1: OpenAI GPT-4.1 (1M context, 32K output)
- gpt-4.1-mini: OpenAI GPT-4.1 Mini (1M context, 32K output)
- gemini-2.5-flash-lite: Google Gemini 2.5 Flash Lite (1M context, 65K output)
- sonar: Perplexity Sonar (128K context, 4K output)
- sonar-pro: Perplexity Sonar Pro (200K context, 8K output)
All specs verified against official provider documentation.
2026-02-24 20:50:10 +01:00
fhennerkes
0f16fcf231
poe: add GPT-5.3-Codex model
2026-02-24 11:39:47 -08:00
fhennerkes
e583c700f9
Merge branch 'anomalyco:dev' into dev
2026-02-24 11:36:20 -08:00
dpuyosa
96d278932c
[venice] Update gemini-3-1-pro-preview config
...
- Reduce output token limit from 250K to 65K
2026-02-24 11:27:18 +01:00
Sewer56
eee3303df0
Add missing synthetic.new models
...
Add configuration for hf:Qwen/Qwen3.5-397B-A17B and hf:MiniMaxAI/MiniMax-M2.5
to the synthetic provider, based on API specs from synthetic.new.
Note: API reports image support but these models may not natively support
images (likely rerouted/proxied through vision-capable infrastructure).
2026-02-24 09:19:57 +00:00
RioPlay
838416044f
add: newer MiniMax, GLM, and Kimi models to DeepInfra
2026-02-23 22:21:28 -06:00
Frank
51441f47d9
update zen models
2026-02-23 15:08:28 -05:00
Nacho F. Lizaur
7fc2c6154d
fix(amazon-bedrock): correct Claude Opus 4.6 context window from 1M to 200K
2026-02-23 20:10:58 +01:00
Colby Gilbert
1439781a76
feat(firmware): add deepseek 3.2, glm 5, kimi k2.5, minimax m2.5
2026-02-22 21:29:51 -08:00
minpeter
8fc0d87742
add friendli minimax m2.5 model config
2026-02-23 13:18:17 +09:00
Colby Gilbert
f660955784
feat(firmware): add grok models
2026-02-22 15:37:53 -08:00
Colby Gilbert
eb11c327b8
feat(firmware): gemini 3.1 pro, sonnet reasoning
2026-02-21 23:43:25 -08:00
Bryan Nie
0dfde60c14
models: alibaba-cn: add MiniMax-M2.5
2026-02-22 01:09:24 +08:00
Wendell Misiedjan
bab7727bad
Add AWS_BEARER_TOKEN_BEDROCK to Amazon Bedrock provider env
...
The @ai-sdk/amazon-bedrock package supports Bearer token authentication
via the AWS_BEARER_TOKEN_BEDROCK environment variable as an alternative
to IAM SigV4 auth. This uses Bedrock API keys for simplified access.
2026-02-21 16:27:45 +01:00
Mikal Sande
4b4a2364c6
Append (latest) to Mistral models that refer to the latest version.
2026-02-21 09:19:13 +01:00
Frank
c36b8e9433
update zen models
2026-02-20 23:20:24 -05:00
xiaojie.zj
5b8e983e7c
feat: add Gemini 3.1 Pro Preview for ZenMux provider
2026-02-21 10:39:30 +08:00
Frank
0d2a52dd9d
update zen models
2026-02-20 20:41:52 -05:00
Frank
b667ab78ac
update zen models
2026-02-20 20:19:33 -05:00
fhennerkes
e2da96cde4
poe: add Gemini-3.1-Pro and update Claude Sonnet 4.6
2026-02-20 11:28:18 -08:00
Jake Jia
9d042ac986
Update providers/zenmux/models/openai/gpt-5.2-pro.toml
...
Co-authored-by: Copilot <175728472+Copilot@users.noreply.github.com >
2026-02-21 01:22:04 +08:00
Phoen1xCode
7192dc0ba8
feat(openai): add GPT-5.2-Pro model via zenmux provider
...
Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com >
2026-02-21 01:07:58 +08:00
Phoen1xCode
05ea56a12a
fix(minimax): remove duplicated provider prefix from MiniMax M2.5 Lightning name
...
Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com >
2026-02-21 01:07:47 +08:00
Aiden Cline
beb449a417
Merge pull request #994 from davidfph/fix/qwen3.5-release-date
...
fix(qwen): update Qwen3.5 release_date and last_updated to 2026-02-16
2026-02-20 10:29:15 -06:00
Aiden Cline
8f76f9b217
Merge pull request #988 from MeganovaAI/fix-meganova-logo
...
Update Meganova logo to official brand icon
2026-02-20 10:29:02 -06:00
Aiden Cline
18ddcde669
fix: azure & cognitive model distinctions
2026-02-20 10:28:22 -06:00
David Fu
ff36ded35f
fix(qwen): update Qwen3.5 release_date and last_updated to 2026-02-16
2026-02-20 20:33:08 +08:00
Aiden Cline
b0b8074a94
Merge pull request #990 from kailiu42/dev
...
models: siliconflow-cn: add new models
2026-02-20 03:01:45 -06:00
Kai Liu
4c30a522f7
models: siliconflow-cn: add new models
...
New models per the latest list: https://cloud.siliconflow.cn/me/models
- Pro/MiniMaxAI/MiniMax-M2.5
- deepseek-ai/DeepSeek-OCR
- PaddlePaddle/PaddleOCR-VL
- PaddlePaddle/PaddleOCR-VL-1.5
Signed-off-by: Kai Liu <kraml.liu@gmail.com >
2026-02-20 16:26:26 +08:00
Aiden Cline
2d63d713de
Merge pull request #992 from anomalyco/fix-azure-models
...
fix: ensure that anthropic models on azure providers have correct urls
2026-02-20 02:19:13 -06:00
Aiden Cline
ac9d0af8e3
fixes
2026-02-20 02:13:23 -06:00
Aiden Cline
a1ad90a9b1
Merge pull request #991 from zainhas/dev
...
[Together AI] add qwen3.5
2026-02-20 01:02:23 -06:00
Zain Hasan
8a6e0dd917
add qwen3.5
2026-02-19 21:38:46 -08:00
Aiden Cline
2da10b739c
Merge pull request #989 from propilideno/fix/adding_missing_azure_foundry_model
...
Add missing GPT-5.2 metadata for Azure Cognitive Services
2026-02-19 18:47:56 -06:00
Aiden Cline
c17e0b9d0f
Merge pull request #987 from dpuyosa/venice
...
Venice: Add Gemini 3.1 Pro Preview and update model configs
2026-02-19 18:47:48 -06:00
Lucas Almeida
877a1175f4
chore: replacing by symbolic link like the other ones
2026-02-19 21:21:05 -03:00
Boqian
1bc83abeeb
Update Meganova logo to official brand icon
2026-02-19 18:57:17 -05:00
dpuyosa
70caedba86
[venice] Add Gemini 3.1 Pro Preview and update model configs
...
- Add new Gemini 3.1 Pro Preview model configuration
- Update Claude Sonnet 4.6 release dates
- Enable open_weights for MiniMax M25
2026-02-19 22:54:59 +01:00
Aiden Cline
60c90a27a0
Merge pull request #985 from sylviezhang37/update-vercel-model-gen-script
...
feat(provider): exclude image/video models
2026-02-19 15:49:25 -06:00
Aiden Cline
2bd0d5446e
Merge pull request #986 from riasvdv/add-gemini-3.1-pro
...
Add Gemini 3.1 Pro Preview to copilot models
2026-02-19 15:49:12 -06:00
Aiden Cline
5f135517b1
Remove audio and video from input modalities
2026-02-19 15:48:42 -06:00
Aiden Cline
e2af7819b4
Rename gemini-3.5-pro-preview.toml to gemini-3.1-pro-preview.toml
2026-02-19 15:47:25 -06:00
Rias
ca6c251b3a
Add Gemini 3.1 Pro Preview to copilot models
2026-02-19 22:43:26 +01:00
Sylvie Zhang
7b1b590d10
exclude image/video gen models
2026-02-19 13:24:11 -08:00
Aiden Cline
5097a1e954
Merge pull request #966 from mhkok/mkok/feat/add-evroc-provider
...
add evroc provider + models
2026-02-19 14:15:51 -06:00
Aiden Cline
05959a83b6
Update font family in Kimi-K2.5 configuration
2026-02-19 14:15:07 -06:00
Aiden Cline
1492e067a4
Merge pull request #976 from too-green/patch-2
...
Add Qwen3 Coder Next model for openrouter
2026-02-19 14:01:07 -06:00
Aiden Cline
41c81535c3
fix: zen
2026-02-19 12:54:29 -06:00
Aiden Cline
829756fc41
Merge pull request #983 from mdrxy/mdrxy/fix-gemini-3
...
fix Gemini 3.1 model names
2026-02-19 12:34:07 -06:00
Mason Daugherty
e6ef906c41
fix
2026-02-19 13:21:52 -05:00
Aiden Cline
0f84db6bc6
Merge pull request #975 from xiaojiezj/zenmux_dev_0219
...
feat: Add new models for ZenMux provider
2026-02-19 11:33:49 -06:00
Aiden Cline
4bb6d52a7c
Merge pull request #979 from hanouticelina/fix-interleaved-for-hf-provider
...
Fix Hugging Face interleaved `reasoning field: reasoning_details` -> `reasoning_content`
2026-02-19 11:33:34 -06:00
Aiden Cline
bdd0194e73
Merge pull request #981 from mdrxy/mdrxy/add-gemini-3.1
...
add gemini 3.1 to google/openrouter
2026-02-19 11:33:15 -06:00
Frank
41a9502628
update zen models
2026-02-19 11:51:37 -05:00
Mason Daugherty
384e747129
add gemini 3.1 to google/openrouter
2026-02-19 11:24:30 -05:00
Frank
e4bb5ceac6
update zen models
2026-02-19 10:16:51 -05:00
Frank
c830964c3f
update zen models
2026-02-19 09:37:07 -05:00
Celina Hanouti
782b6277ae
Fix Hugging Face interleaved reasoning field
2026-02-19 15:27:49 +01:00
Frank
e63d48ae9c
update zen models
2026-02-19 07:42:52 -05:00
Matthijs Kok
029522aa96
fix family names
2026-02-19 08:48:43 +01:00
Ahmed
482ed2e833
Add Qwen3 Coder Next model for openrouter
...
Added model configuration for Qwen3 Coder Next
2026-02-19 12:36:55 +05:00
Aiden Cline
c6635aa7c3
Merge pull request #968 from MeganovaAI/add-meganova-provider
...
Add Meganova as a provider
2026-02-18 23:44:03 -06:00
Aiden Cline
a6ffef7e4f
Merge pull request #969 from SomeoneWithOptions/dev
...
add claude sonnet 4.6 on openrouter
2026-02-18 23:40:02 -06:00
Aiden Cline
fda9bb5335
Merge pull request #973 from sylviezhang37/update-vercel-models-20260219-0026
...
Update Vercel models
2026-02-18 23:38:56 -06:00
Aiden Cline
98d6901697
Update input cost value in qwen3.5-plus.toml
2026-02-18 23:38:49 -06:00
Aiden Cline
13ae499d7c
tweak values
2026-02-18 23:38:18 -06:00
xiaojie.zj
f21d205d6e
feat: 增加Claude Sonnet 4.6/Doubao-Seed-2.0-lite/Doubao-Seed-2.0-mini/Doubao-Seed-2.0-pro模型
2026-02-19 11:10:28 +08:00
Lucas Almeida
c82b08d778
fix: adding missing gpt-5.2 model on azure foundry
2026-02-18 23:38:10 -03:00
Sylvie Zhang
ea612760cf
Delete providers/vercel/models/recraft/recraft-v4.toml
2026-02-18 16:34:47 -08:00
Sylvie Zhang
48a64f0834
Delete providers/vercel/models/recraft/recraft-v4-pro.toml
2026-02-18 16:34:35 -08:00
github-actions[bot]
040e7fff4e
chore(vercel): update Vercel model definitions
...
Auto-generated by weekly workflow from Vercel AI Gateway API.
Co-Authored-By: github-actions[bot] <41898282+github-actions[bot]@users.noreply.github.com>
2026-02-19 00:26:59 +00:00
SomeoneWithOptions
a6d14928b6
added claude sonnet 4.6 on openrouter
2026-02-18 18:49:56 -05:00
Boqian
92d9e89690
Set reasoning=false for DeepSeek V3 series
...
V3-0324, V3.1, V3.2, V3.2-Exp are chat models, not reasoning models.
Only DeepSeek-R1 is a reasoning model.
Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com >
2026-02-18 16:41:07 -05:00
Boqian
2d83b96bb1
Fix interleaved reasoning_content based on Meganova API testing
...
Tested each model with include_reasoning=true against the live API.
Added [interleaved] to: GLM-4.6, MiniMax-M2.1, MiniMax-M2.5, Kimi-K2.5
Removed [interleaved] from: DeepSeek-V3.1, V3.2, V3.2-Exp, MiMo-V2-Flash
Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com >
2026-02-18 16:35:51 -05:00
Boqian
6e02385a6b
Add interleaved reasoning_content to DeepSeek V3.1, V3.2, V3.2-Exp
...
These models support interleaved reasoning output, matching how other
providers (deepinfra, baseten, chutes) configure them.
Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com >
2026-02-18 16:29:41 -05:00
Boqian
a2f8234c8e
Update pricing and context limits from Meganova API
...
Use actual pricing from https://api.meganova.ai/v1/models instead of
reference data from other providers.
Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com >
2026-02-18 16:25:43 -05:00
Aiden Cline
2f43c70397
Merge pull request #967 from nicolasgere/dev
...
feat(provider): Add glm-5 for baseten
2026-02-18 15:19:47 -06:00
Boqian
006cb53c1e
Add Meganova as a provider with 19 open-weight models
...
Adds Meganova AI (https://api.meganova.ai/v1 ) as an OpenAI-compatible provider
with curated open-weight models including DeepSeek, GLM, Qwen, Kimi, MiniMax,
MiMo, Llama, and Mistral families.
Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com >
2026-02-18 16:17:51 -05:00
nicolasgere
d2832f9d1d
Update GLM-5.toml
2026-02-18 16:15:22 -05:00
nicolasgere
73c0fcc19e
Rename GLM-5 to GLM-5.toml
2026-02-18 16:14:26 -05:00
nicolasgere
04a1fe8043
Create GLM-5
2026-02-18 16:14:09 -05:00
Matthijs Kok
7fa0eeebf9
add evroc provider + models
2026-02-18 19:57:44 +01:00
Aiden Cline
cb72e1780f
Merge pull request #962 from aldosch/add-sonnet-4-6-vercel
...
add sonnet 4.6 to vercel ai gateway
2026-02-18 12:19:00 -06:00
Aiden Cline
7ee7e88346
Merge pull request #960 from fa-sharp/patch-1
...
fix: OpenRouter output modalities for image-only models
2026-02-18 12:18:27 -06:00
Aiden Cline
4cc230ba6c
Merge pull request #963 from vglafirov/gitlab/add-sonnet-4-6
...
feat(gitlab): add Claude Sonnet 4.6 model
2026-02-18 12:17:57 -06:00
Aiden Cline
65d6355ffa
Merge pull request #964 from xinrui-z/aihubmix-add-claude-4-6
...
aihubmix: add models
2026-02-18 12:17:48 -06:00
Xinrui
1b6157c0ee
aihubmix: add models
2026-02-18 22:55:03 +08:00
Vladimir Glafirov
5e25c0838e
feat(gitlab): add Claude Sonnet 4.6 model
2026-02-18 13:52:26 -01:00
aldosch
b06ab95c9e
add sonnet 4.6 to vercel ai gateway
2026-02-19 00:38:35 +11:00
Daltonganger
33f4e77ee4
fix(nano-gpt): normalize family enums for model validation
2026-02-18 12:53:40 +01:00
Daltonganger
51ef2e51ae
finalize nano-gpt model sync and release metadata
2026-02-18 12:44:36 +01:00
farshad
9d873cc764
add back trailing newline
2026-02-18 02:18:40 -05:00
farshad
d0acb0a35d
fix output modalities for black forest flux models
2026-02-18 01:52:31 -05:00
farshad
697c191fea
Update output modalities in seedream-4.5.toml
2026-02-18 01:45:06 -05:00
Aiden Cline
7d9ff92ffd
fix: glm 5 maas
2026-02-17 23:41:48 -06:00
Aiden Cline
03a2aee0de
Merge pull request #716 from bluet/feat/google-vertex-openai
...
feat: add google-vertex-openai provider for Vertex AI partner models
2026-02-17 23:38:08 -06:00
Aiden Cline
4eb19459dd
Merge pull request #958 from mugnimaestra/feat/add-qwen3.5-397b-a17b-tee-chutes
...
feat: add Qwen3.5 397B A17B TEE to Chutes provider listings
2026-02-17 23:27:18 -06:00
Aiden Cline
b74d7eb495
Merge pull request #957 from dpuyosa/veniceScript
...
Venice: Update generate-venice script to include context_over_200k cost
2026-02-17 23:26:28 -06:00
Aiden Cline
60aa5aca14
Merge pull request #956 from dpuyosa/venice
...
Venice: Add Claude Sonnet 4.6 and GLM 4.7 Flash Heretic models
2026-02-17 20:22:44 -06:00
Muhammad Mugni Hadi
67235a4e39
feat: add Qwen3.5 397B A17B TEE to Chutes provider listings
2026-02-18 09:09:45 +07:00
dpuyosa
ff03fd6906
[venice] Add Claude Sonnet 4.6 and GLM 4.7 models
...
- Add Claude Sonnet 4.6 model with context_over_200k pricing
- Add GLM 4.7 Flash Heretic model (open weights)
- Add context_over_200k pricing tier to Claude Opus 4.6
2026-02-18 02:02:19 +01:00
dpuyosa
423b177c2b
Update generate-venice script to include context_over_200k cost
2026-02-18 01:55:52 +01:00
Aiden Cline
8c263109c5
Merge pull request #955 from maahir30/open-router-structured-output
...
Add structured output support for OpenRouter models
2026-02-17 18:52:02 -06:00
Maahir Sachdev
1bab7e8438
update open router models
2026-02-17 16:42:52 -08:00
Aiden Cline
7a163dbc60
Merge pull request #953 from mongrelion/dev
...
feat: add github copilot claude sonnet 4.6 model
2026-02-17 18:30:54 -06:00
Aiden Cline
c94d025aa7
fixes
2026-02-17 18:23:01 -06:00
Aiden Cline
55e8be17b5
Merge pull request #949 from cgilly2fast/dev
...
feat(firmware): add sonnet 4.6
2026-02-17 18:21:16 -06:00
Aiden Cline
56096012b2
Merge pull request #948 from elithrar/patch-1
...
add Sonnet 4.6 model config
2026-02-17 17:53:20 -06:00
Aiden Cline
9b8a22d756
Merge pull request #725 from janszypulski/add-provider-cloudferro-sherlock
...
Add provider - Cloudferro Sherlock
2026-02-17 17:50:14 -06:00
Aiden Cline
5a778c6e93
Merge pull request #682 from the-lazy-me/add-qihang-provider
...
feat: add QiHang provider with 7 models
2026-02-17 17:48:51 -06:00
Aiden Cline
4d08659acf
Merge pull request #651 from yinxulai/feat/qiniu-ai
...
feat: add Qiniu AI provider configuration
2026-02-17 17:46:05 -06:00
Aiden Cline
128c9ec469
Merge pull request #308 from d-oit/feature/perplexity-sonar-deep-research
...
Feature/perplexity sonar deep research
2026-02-17 17:37:20 -06:00
Carlos León
dc11781324
feat: add github copilot claude sonnet 4.6 model
...
Model list sourced from GitHub Settings page showing currently available models. Specifications cross-referenced with Anthropic provider implementation.
2026-02-18 00:22:22 +01:00
Colby Gilbert
9d0b37bea3
feat(firmware): add sonnet 4.6
2026-02-17 15:03:21 -08:00
Matt Silverlock
c8d09fe349
add Sonnet 4.6 model config
2026-02-17 17:27:51 -05:00
Aiden Cline
1a22b93fc2
Merge pull request #947 from fhennerkes/dev
...
poe: add Claude-Sonnet-4.6 and update XAI models
2026-02-17 15:56:11 -06:00
Aiden Cline
3918131cb8
Merge pull request #946 from monotykamary/remove-fireworks-deprecated-models-2026-02-12
...
chore(fireworks-ai): remove deprecated serverless models
2026-02-17 15:56:00 -06:00
fhennerkes
a1d9c5134c
poe: add Claude-Sonnet-4.6 and update XAI models
2026-02-17 13:38:54 -08:00
Ruben Beuker
20abb5b8df
preserve curated release dates for key nano-gpt models
...
Keep existing curated release and last-updated values for models where NanoGPT API uses the generic created timestamp baseline.
2026-02-17 22:09:03 +01:00
Tom X Nguyen
dc36ed54ae
chore(fireworks-ai): remove deprecated serverless models
...
Remove 6 Fireworks serverless models deprecated on February 12, 2026:
- glm-4.6 (migrate to glm-4.7)
- deepseek-r1-0528 (migrate to deepseek-v3.2 or deepseek-v3.1)
- deepseek-v3-0324 (migrate to deepseek-v3.2 or deepseek-v3.1)
- qwen3-235b-a22b (migrate to kimi-k2-instruct-0905)
- qwen3-coder-480b-a35b-instruct (migrate to kimi-k2-instruct-0905)
- minimax-m2 (migrate to MiniMax-M2.1)
See: https://fireworks.ai/models?modelTypes=Serverless
2026-02-18 04:04:51 +07:00
Ruben Beuker
8cb462f29b
sync nano-gpt models with live API catalog
...
Refresh NanoGPT model files to match the current /api/v1/models output, remove stale entries, and add newly available models while preserving path-based IDs.
Also ignore local TokenSpeed sqlite artifacts so private monitoring data is not shown or committed.
2026-02-17 22:01:54 +01:00
Frank
89486ec705
update zen models
2026-02-17 14:12:30 -05:00
Aiden Cline
f313f802ee
Merge pull request #940 from nitishxyz/add-claude-sonnet-4-6
...
feat(models): add Claude Sonnet 4.6 model configurations
2026-02-17 13:12:10 -06:00
nitishxyz
128615ddd7
feat(models): add Claude Sonnet 4.6 model configurations
...
- Add Claude Sonnet 4.6 to Anthropic provider with full capabilities
- Add regional variants (US, EU, Global) for Amazon Bedrock provider
- Add Google Vertex Anthropic provider configuration
- Define pricing, context limits (200k tokens), and modalities
Co-authored-by: ottocode-io[bot] <261994719+ottocode-io[bot]@users.noreply.github.com>
2026-02-18 00:01:02 +05:30
Aiden Cline
756fb772c1
Merge pull request #939 from Nomadcxx/fix/kilo-npm-provider
...
fix(kilo): use @ai-sdk/openai-compatible instead of opencode-kilo-auth
2026-02-17 11:29:52 -06:00
Nomadcxx
e86f0afd87
fix(kilo): use @ai-sdk/openai-compatible npm package
...
The npm field pointed to opencode-kilo-auth which causes
ProviderInitError when loading Kilo models.
Switched to @ai-sdk/openai-compatible (already bundled in OpenCode)
and added api field for the gateway endpoint.
2026-02-18 04:21:52 +11:00
Aiden Cline
29c5e28a43
Merge pull request #791 from samsja/add-intellect-3
...
Add Intellect 3 model from Prime Intellect
2026-02-17 10:56:34 -06:00
Aiden Cline
ea414b1500
Merge pull request #935 from ConceptCodes/feat/add-glm-flashx-model
...
feat: add GLM-4.7-FlashX model configuration
2026-02-17 10:34:45 -06:00
Aiden Cline
4556fe8b5b
Merge pull request #937 from gary149/feat/huggingface-qwen3.5-m2.5-coder-next
...
feat(huggingface): add Qwen3.5-397B, MiniMax-M2.5, Qwen3-Coder-Next
2026-02-17 10:34:33 -06:00
Aiden Cline
8af23aeba5
Merge pull request #938 from spiffytech/dev
...
Add Ollama Cloud support for Qwen 3.5
2026-02-17 10:34:18 -06:00
Aiden Cline
9bfe1203c6
ci
2026-02-17 10:34:02 -06:00
spiffytech
a619966e22
Added Ollama Cloud support for Qwen 3.5
2026-02-17 10:00:15 -05:00
Victor Muštar
f48d55e1aa
chore: remove accidentally committed skill file
2026-02-17 10:31:27 +01:00
Victor Muštar
fe0ddcb666
feat(huggingface): add Qwen3.5-397B, MiniMax-M2.5, Qwen3-Coder-Next
2026-02-17 10:31:18 +01:00
Frank
af1e1d1f51
update zen models
2026-02-17 02:08:24 -05:00
Aiden Cline
774a9f40b0
Merge pull request #933 from too-green/patch-1
...
Fix the display name of GLM-4.7-Flash
2026-02-17 00:23:26 -06:00
Aiden Cline
4f01ffb017
Merge pull request #934 from PandaSt0rm/add-minimax-m2.5-highspeed-models
...
Add MiniMax-M2.5-highspeed models for official MiniMax providers
2026-02-17 00:23:10 -06:00
Aiden Cline
7d768260cf
Merge pull request #936 from Alex-wuhu/dev
...
add Qwen3.5-397B-A17B for novita
2026-02-17 00:22:30 -06:00
Alex-wuhu
467d269522
add Qwen3.5-397B-A17B for novita
2026-02-17 13:22:04 +08:00
concept
5ec496d6e1
feat: add GLM-4.7-FlashX model configuration
2026-02-16 21:31:16 -06:00
PandaSt0rm
45ca42f95a
add MiniMax-M2.5-highspeed models
2026-02-17 03:53:05 +02:00
Ahmed
f19ebce14c
Rename model to GLM-4.7-Flash
...
Both GLM 4.7 and GLM 4.7 Flash had been named to the same "GLM 4.7"
2026-02-17 05:03:41 +05:00
Aiden Cline
85f5340eeb
Merge pull request #931 from rifandyzv/dev
...
Add Qwen3.5 models for alibaba & alibaba-cn provider
2026-02-16 16:06:32 -06:00
Aiden Cline
39f06e82e7
Merge pull request #932 from cantalupo555/feat/add-openrouter-qwen3.5-plus-and-397b-a17b
...
feat: add Qwen3.5 models on OpenRouter
2026-02-16 16:05:56 -06:00
cantalupo555
4e7725d244
feat: add Qwen3.5 models on OpenRouter
2026-02-16 17:47:45 -03:00
Aiden Cline
7fe64bc498
Revert "Add image and video to input modalities"
...
This reverts commit 76e84a8b06 .
2026-02-16 12:12:29 -06:00
rifandyzv
f84a4a5cf9
feat: add Qwen3.5 models for alibaba & alibaba-cn provider
2026-02-17 01:28:16 +08:00
Aiden Cline
96f60c3329
Merge pull request #928 from Daltonganger/feat/kilo-provider-models
...
Add Kilo Gateway provider and import Kilo models
2026-02-16 11:08:11 -06:00
Aiden Cline
f67b9bdef7
Merge pull request #929 from Daltonganger/feat/nano-gpt-qwen35-models
...
Add four Qwen3.5 models for NanoGPT
2026-02-16 11:05:23 -06:00
Frank
76e84a8b06
Add image and video to input modalities
2026-02-16 12:00:43 -05:00
Daltonganger
cec16274c7
Add NanoGPT Qwen3.5 model variants
2026-02-16 17:09:17 +01:00
Daltonganger
17094722ea
Add Kilo provider and import Kilo model catalog
2026-02-16 17:00:15 +01:00
Matthew (BlueT) Lien
e3e230e3b3
fix: add api base URL template to partner model [provider] overrides
...
Add the api field with env-var template URL to all partner models so
opencode's loadBaseURL() can resolve the OpenAI-compatible endpoint.
Uses GOOGLE_VERTEX_PROJECT (not GOOGLE_CLOUD_PROJECT) because
googleVertexVars() resolves it through the full fallback chain
(GOOGLE_VERTEX_PROJECT → options.project → GOOGLE_CLOUD_PROJECT →
GCP_PROJECT → GCLOUD_PROJECT).
2026-02-16 21:51:55 +08:00
Aiden Cline
4666f36f3e
Merge pull request #926 from zainhas/dev
...
[Together AI] add minimax M2.5
2026-02-15 23:57:45 -06:00
Aiden Cline
5a2bcd704e
Merge pull request #900 from conglinyizhi/dev
...
feat: Add StepFun provider support
2026-02-15 23:57:35 -06:00
Zain Hasan
05861fd7fd
add minimax M2.5
2026-02-15 21:48:37 -08:00
Aiden Cline
e37bb8ae68
Merge pull request #923 from juls0730/dev
...
Fix cerebras/zai-gml-4.7 pricing
2026-02-15 20:03:47 -06:00
Aiden Cline
495e8006df
Merge pull request #905 from shelvick/add-vertex-glm-5
...
Add GLM-5 to Google Vertex AI
2026-02-15 20:03:34 -06:00
Aiden Cline
ad8dde798d
Merge pull request #924 from cgilly2fast/dev
...
feat(firmware): add reason to anthropic models
2026-02-15 20:03:23 -06:00
Aiden Cline
ab4fa333e3
Merge pull request #925 from 8dazo/dev
...
feat: add MiniMax M2.5 to Chutes provider listings
2026-02-15 20:03:12 -06:00
8dazo
6255298cd1
minimax model update
2026-02-16 06:31:01 +05:30
Colby Gilbert
3614087be6
Merge branch 'anomalyco:dev' into dev
2026-02-15 15:55:32 -08:00
Colby Gilbert
4495cb3569
feat(firmware): add reason to anthropic models
2026-02-15 15:55:02 -08:00
juls0730
8444d9293d
Fix cerebras/zai-gml-4.7 pricing
...
Prices from https://inference-docs.cerebras.ai/models/zai-glm-47#z-ai-glm-4-7
2026-02-15 17:53:57 -06:00
Aiden Cline
c1d36715ee
Merge pull request #914 from 8dazo/dev
...
feat: add Z-AI GLM-5 to Chutes provider listings
2026-02-15 15:45:37 -06:00
Aiden Cline
860e610b73
Merge pull request #922 from cgilly2fast/dev
...
chore: remove unsupported models
2026-02-15 15:45:28 -06:00
Colby Gilbert
beb84e769a
chore: remove unsupported models
2026-02-15 13:38:52 -08:00
Aiden Cline
0408546681
Merge pull request #921 from zerone0x/feat/add-bedrock-deepseek-v3.2
...
feat(amazon-bedrock): add DeepSeek V3.2
2026-02-15 15:29:26 -06:00
Aiden Cline
97a040bc5e
Merge pull request #915 from fanweixiao/dev
...
add glm-5, gpt-5-mini, deepseek-v3.2 models for vivgrid provider
2026-02-15 15:29:08 -06:00
Aiden Cline
f68786b892
Merge pull request #920 from anomalyco/revert-912-add-github-copilot-gpt-5-3-codex
...
Revert "feat: add GitHub Copilot GPT-5.3 Codex"
2026-02-15 08:42:51 -06:00
Clawdbot
816c3d96b9
feat(amazon-bedrock): add DeepSeek V3.2
...
Add DeepSeek V3.2 model to Amazon Bedrock provider.
Model ID: deepseek.v3.2-v1:0
Pricing (US regions): $0.62/1M input, $1.85/1M output
Ref: https://aws.amazon.com/about-aws/whats-new/2026/02/amazon-bedrock-adds-support-six-open-weights-models/
2026-02-15 09:19:18 +01:00
Aiden Cline
bac557c176
Revert "feat: add github copilot gpt-5.3-codex model ( #912 )"
...
This reverts commit 08db483d58 .
2026-02-14 18:31:30 -06:00
Aiden Cline
97e81f356e
Merge pull request #908 from hsnyus-09/feature/add-aurora-alpha
...
feat(openrouter): add aurora-alpha model definition
2026-02-14 17:34:22 -06:00
Matthew (BlueT) Lien
4c361218de
feat: add Vertex AI partner models with openai-compatible overrides
...
Add DeepSeek V3.1, Llama 4 Maverick, Llama 3.3 70B, and Qwen3 235B as
partner models under google-vertex provider. Update GLM-4.7 with
corrected specs from official Google Cloud docs.
Each partner model uses [provider] npm override to @ai-sdk/openai-compatible
since these models are served via Google's OpenAI-compatible endpoint,
while staying consolidated under the google-vertex provider per
maintainer feedback.
All specs (context windows, output limits, pricing, modalities)
verified against official Google Cloud documentation:
- cloud.google.com/vertex-ai/generative-ai/pricing
- cloud.google.com/vertex-ai/generative-ai/docs/maas/*
Changes:
- Update zai-org/glm-4.7-maas: fix context=200K, output=128K, add pdf
modality, correct release_date, add structured_output, add [provider]
- Add deepseek-ai/deepseek-v3.1-maas ($0.60/$1.70, 163K context)
- Add meta/llama-4-maverick-17b-128e-instruct-maas (vision, 524K ctx)
- Add meta/llama-3.3-70b-instruct-maas ($0.72/$0.72, 128K context)
- Add qwen/qwen3-235b-a22b-instruct-2507-maas ($0.22/$0.88, 262K ctx)
2026-02-15 06:35:17 +08:00
Anjul Garg
08db483d58
feat: add github copilot gpt-5.3-codex model ( #912 )
2026-02-14 14:27:49 -05:00
YuSung Han
e0c14d7883
Remove redundant lines in aurora-alpha.toml
2026-02-15 03:50:02 +09:00
Aiden Cline
e457c7f1dd
Merge pull request #916 from arshadbarves/add-nvidia-glm5
...
Add GLM5 model to nvidia provider
2026-02-14 11:53:09 -06:00
Arshad Barves
c86b97226c
Update providers/nvidia/models/z-ai/glm5.toml
...
Co-authored-by: Copilot <175728472+Copilot@users.noreply.github.com >
2026-02-14 17:36:22 +05:30
Test User
8ec101e8f1
Add GLM5 model to nvidia provider
2026-02-14 15:32:11 +05:30
C.C. Fan
cc1feddc11
add glm-5, gpt-5-mini, deepseek-v3.2 models
2026-02-14 17:08:07 +08:00
8dazo
95fd20e66a
update name
2026-02-14 13:54:20 +05:30
8dazo
8c322d46dd
Chutes Model Listings update
2026-02-14 13:49:06 +05:30
conglinyizhi
da7a26b7b1
fix: 修复 StepFun provider 文档链接
...
将 doc 字段从 https://platform.stepfun.com/docs
改为 https://platform.stepfun.com/docs/zh/overview/concept
2026-02-14 15:07:40 +08:00
Aiden Cline
5a00835470
Merge pull request #913 from kavin-kr/patch-1
...
Update release date for nova-2-pro-v1 model
2026-02-14 00:50:06 -06:00
Frank
b0c0a91926
update zen models
2026-02-14 00:52:08 -05:00
Kavin
8b2d995801
Update release date for nova-2-pro-v1 model
2026-02-13 23:46:55 -06:00
Aiden Cline
3f4b804ca1
Merge pull request #911 from monotykamary/feat/add-minimax-m2.5
...
feat: add fireworks minimax m2.5 model and fix m2.1 cache pricing
2026-02-13 19:00:46 -06:00
Tom X Nguyen
133e605c23
feat: add minimax m2.5 model and fix m2.1 cache pricing
2026-02-14 07:53:15 +07:00
Aiden Cline
f3cff10e78
Merge pull request #910 from keenborder786/fix/gpt_5_2_pro
...
fix: gpt 5-2-pro does not support structured output
2026-02-13 17:53:00 -06:00
keenborder786
4732ca7f77
fix: gpt 5-2-pro does not support structured output
2026-02-14 04:50:19 +05:00
Aiden Cline
06f79b4142
Merge pull request #909 from juls0730/dev
...
Add all missing cohere models offered by the cohere api
2026-02-13 16:47:11 -06:00
Aiden Cline
69106c6f36
Merge pull request #907 from Algowary/dev
...
Chutes Model Listings update
2026-02-13 16:44:43 -06:00
Zoe
dc1d8b78d0
Add all missing cohere models offered by the cohere api
...
This commit adds all the models offered by the official cohere api
that are not yet available in the models.dev repo, excluding the
rerank and embed models.
2026-02-13 16:43:57 -06:00
hsnyus-09
4383829304
feat(openrouter): add aurora-alpha model definition
2026-02-14 06:17:25 +09:00
Algowarry
610713b805
Merge branch 'dev' of https://github.com/Algowary/models.dev into dev
2026-02-13 16:08:26 -05:00
Algowarry
07e3eec2cf
Chutes Model Inventory Update
...
Updating the models available from the provider chutes.ai
2026-02-13 16:01:46 -05:00
Algowarry
ccf8a3e82a
Merge branch 'dev' of https://github.com/Algowary/models.dev into dev
2026-02-13 15:07:16 -05:00
Algowarry
8e1d34323a
Chutes Model Update
...
Model inventory and stats update
2026-02-13 14:38:07 -05:00
Aiden Cline
5f6d36a463
Merge pull request #787 from elithrar/fix/cloudflare-ai-gateway-provider-package
...
use official ai-gateway-provider package for Cloudflare AI Gateway
2026-02-13 12:44:45 -06:00
Scott Helvick
84e43b1912
Add GLM-5 to Google Vertex AI
2026-02-13 17:57:08 +00:00
Aiden Cline
a9f14cbae3
Merge pull request #903 from micuintus/dev
...
fix(nebius): correct model ID casing to match Token Factory API
2026-02-13 10:39:51 -06:00
Aiden Cline
97baff037b
Merge pull request #896 from zainhas/dev
...
[Together AI] Add GLM-5
2026-02-13 10:38:06 -06:00
Aiden Cline
b149bd83ad
Merge pull request #902 from qychen2001/dev
...
chore(siliconflow): update siliconflow/siliconflow-cn models
2026-02-13 10:25:04 -06:00
Aiden Cline
30ef1d1b51
Merge pull request #898 from 888-wzk/feature/chenger_20260128
...
feat(models): Added minimax model profile
2026-02-13 10:24:36 -06:00
Aiden Cline
a62ccfd392
Merge pull request #901 from dpuyosa/venice
...
Venice: Add MiniMax M2.5 model configuration
2026-02-13 10:24:18 -06:00
QiyuanChen
23e195a0fa
feat(models): Add interleaved reasoning_content field to GLM-4.7 and GLM-5 configurations for zai-org and Pro
2026-02-13 23:35:56 +08:00
Aiden Cline
d1c5bc811e
Merge pull request #899 from niushuai1991/feature/kuae-cloud-coding-plan
...
add provider: kuae cloud coding plan
2026-02-13 09:23:33 -06:00
Michael Voigt
ad7e8047b9
fix(nebius): correct model ID casing to match Token Factory API
...
Fix lowercase model ID bug that caused "The model does not exist" errors.
- qwen/ → Qwen/ directory
- Fixed model file casing to match API exactly:
- google/gemma-* → lowercase (gemma-2-2b-it, etc.)
- meta-llama/*-Fast → lowercase fast suffix
- nvidia/Llama-3_1-* → underscore instead of dot
- nvidia/NVIDIA-* → uppercase NVIDIA prefix
- black-forest-labs/flux-* → all lowercase
- BAAI/bge-* → all lowercase
- All Qwen models → proper casing
* Remove outdated models not in API:
- deepseek-ai/DeepSeek-V3
- meta-llama/Llama-3.1-405B-Instruct
- zai-org/GLM-4.7
* Add new model:
- moonshotai/Kimi-K2.5 (262K context, multimodal)
Fixes: https://github.com/anomalyco/opencode/issues/12461
and: https://ideas.nebius.com/en/p/token-factory-api-lowercase-model-ids
Note: The changes made and verified with actual Nebius API access
2026-02-13 14:28:48 +01:00
QiyuanChen
73393e9e41
chore(models): Remove Qwen3-30B-A3B and DeepSeek-R1-Distill-Qwen-7B model configuration files from siliconflow and siliconflow-cn
2026-02-13 20:11:19 +08:00
QiyuanChen
3e5566ab9d
chore(models): Remove GLM-4.1V-9B-Thinking model configuration files from siliconflow and siliconflow-cn
2026-02-13 20:09:14 +08:00
QiyuanChen
f5096e4b54
chore(models): Remove Kimi-Dev-72B model configuration files from siliconflow and siliconflow-cn
2026-02-13 20:08:14 +08:00
QiyuanChen
1d6e26574f
chore(models): Remove MiniMaxAI/MiniMax-M1-80k and MiniMax-M2 model configuration files
2026-02-13 20:07:15 +08:00
QiyuanChen
57b0608e70
feat(models): Introduce Step-3.5-Flash model configuration and remove deprecated Step-3 model files
2026-02-13 20:06:05 +08:00
QiyuanChen
88ed698a69
feat(models): Enable structured_output in GLM-4.7 and GLM-5 configurations for zai-org and Pro
2026-02-13 20:04:27 +08:00
QiyuanChen
286c43f2cd
feat(glm-5): Add new GLM-5 model configuration files for zai-org and Pro
2026-02-13 20:00:10 +08:00
dpuyosa
04d82741fa
[venice] Add MiniMax M2.5 model configuration
...
- Modalities: text input/output
- Context window: 198K tokens
- Max output: 32K tokens
- Pricing: $0.40/M input, $1.60/M output, $0.04/M cache read
2026-02-13 09:44:22 +01:00
conglinyizhi
58c595b95f
feat: Add StepFun provider support
...
- Add StepFun(阶跃星辰) as a new provider with OpenAI-compatible API
- Support step-3.5-flash (256K context, reasoning model)
- Support step-2-16k (1T parameters, 16K context)
- Support step-1-32k (100B parameters, 32K context)
Pricing based on official StepFun documentation (converted from CNY to USD):
- step-3.5-flash: bash.096 input / bash.288 output / bash.019 cache
- step-2-16k: .21 input / 6.44 output / .04 cache
- step-1-32k: .05 input / .59 output / bash.41 cache
Note: Logo not included as it is optional per contributing guidelines.
A default logo will be served by models.dev API instead.
All model definitions follow the official schema.
Fixes anomalyco/opencode#11760
Fixes anomalyco/opencode#11960
StepFun API: https://api.stepfun.com/v1
Documentation: https://platform.stepfun.com/docs/zh/pricing/details
2026-02-13 16:28:07 +08:00
城二
58de85c2e8
feat(minimax): Add interleaved configuration
...
- Add the reasoning_content field configuration to the minimax model.
- Update the configuration files for m2.5 and m2.5-lightning.
2026-02-13 16:20:45 +08:00
城二
f87ecffbf0
feat(minimax): Update m2.5 model name and price
...
- Change the model name from "lightning" to "highspeed"
- Adjust the input/output and cache read/write prices
2026-02-13 16:18:22 +08:00
niushuai1991
6dfb2f9c83
add kuae cloud coding plan
2026-02-13 15:02:52 +08:00
城二
ee8c1bce7d
feat(models): Added minimax model profile
2026-02-13 14:15:44 +08:00
Zain Hasan
a2dd10d09d
Update output limit in GLM-5 configuration
2026-02-12 22:12:56 -08:00
Zain Hasan
f6cfc2ebd2
try remove reasoning
2026-02-12 21:54:02 -08:00
Zain Hasan
3192856cc3
finx glm 5 settings
2026-02-12 21:46:13 -08:00
Aiden Cline
5507f42604
Merge pull request #874 from 888-wzk/feature/chenger_20260128
...
feat(z-ai): New glm-5 model configuration file
2026-02-12 22:56:36 -06:00
Aiden Cline
995aabf33f
Merge pull request #894 from fhennerkes/dev
...
Poe: fix formatting, naming and update outputs
2026-02-12 22:56:08 -06:00
城二
fc5c3613eb
feat(glm-5): Add reasoning_content field
2026-02-13 11:37:33 +08:00
fhennerkes
72f10a52e1
poe: update model names to use display_name
2026-02-12 19:08:49 -08:00
fhennerkes
e944012d95
poe: small fixes (formatting and reasoning)
2026-02-12 18:55:19 -08:00
Aiden Cline
e117f37d4e
Merge pull request #892 from pat-baseten/add-kimi-2.5-baseten
...
Add Kimi K2.5 model for Baseten
2026-02-12 17:45:46 -06:00
Pat
b0d71629fe
Add Kimi K2.5 model for Baseten
2026-02-12 16:39:56 -06:00
Aiden Cline
aa5e8634b2
Merge pull request #890 from cfal/fireworks-glm-5
...
fireworks: add GLM-5
2026-02-12 16:15:18 -06:00
Aiden Cline
7acba1db3f
Merge pull request #891 from lucianjon/feat/openrouter-minimax-m2.5
...
feat(openrouter/minimax): add minimax-m2.5
2026-02-12 16:15:08 -06:00
Aiden Cline
c5095973e1
Merge pull request #886 from brentdurksen/dev
...
feat(amazon-bedrock): add Writer Palmyra X4 and X5 models
2026-02-12 16:14:58 -06:00
Aiden Cline
5b8797cf89
Merge pull request #889 from Daltonganger/feat/nano-gpt-add-minimax-m2.5-official
...
feat(nano-gpt): add MiniMax M2.5 route alongside official variant
2026-02-12 16:14:36 -06:00
Daltonganger
f1317184b5
Enable reasoning and add interleaved field in TOML
2026-02-12 23:07:12 +01:00
Lucian Jones
7ba286c7f2
feat(openrouter/minimax): add minimax-m2.5
2026-02-13 10:59:30 +13:00
cfal
dec532b3b0
providers/fireworks-ai/models/accounts/fireworks/models/glm-5.toml: add GLM-5 to fireworks
2026-02-13 01:45:51 +04:00
Aiden Cline
ccff680988
Merge pull request #864 from sylviezhang37/vercel-model-file-gen-script
...
feat(provider): Vercel model file generation and update script
2026-02-12 15:37:17 -06:00
Aiden Cline
1b63e4670e
Merge pull request #887 from PandaSt0rm/add-minimax-m2-5-support
...
Add MiniMax-M2.5 across minimax and coding-plan providers
2026-02-12 15:36:56 -06:00
Aiden Cline
d5cbd6fb5d
Merge pull request #888 from spiffytech/dev
...
Add Ollama Cloud support for Minimax 2.5
2026-02-12 15:36:14 -06:00
Ruben Beuker
e8f2f6b14f
feat(nano-gpt): add MiniMax M2.5 route and align official variant
2026-02-12 22:35:23 +01:00
spiffytech
e92fe6e9d7
Added Ollama Cloud support for Minimax 2.5
2026-02-12 16:24:11 -05:00
PandaSt0rm
7a32f17911
add MiniMax-M2.5 configs across minimax providers
2026-02-12 23:22:52 +02:00
Brent Durksen
57db1db84f
feat(amazon-bedrock): add Writer Palmyra X4 and X5 models
...
Add two new Writer AI models to the Amazon Bedrock provider:
- writer.palmyra-x4-v1:0 (Palmyra X4): 128K context, 8K output,
reasoning and tool calling, $2.50/$10 per M tokens (input/output)
- writer.palmyra-x5-v1:0 (Palmyra X5): 1M context, 8K output,
reasoning and tool calling, $0.60/$6 per M tokens (input/output)
Both models support text-only input/output modalities and are
closed-weight.
Also adds the 'palmyra' family to the ModelFamilyValues enum in
packages/core/src/family.ts to support validation.
2026-02-12 13:43:32 -07:00
Aiden Cline
ba91bb6612
Merge pull request #883 from ryanskidmore/ryanskidmore/cloudflare-ai-gateway-bump-opus-4-6-limits
...
cloudflare-ai-gateway: bump Opus 4.6 output limit to 128k
2026-02-12 13:01:38 -06:00
Ryan Skidmore
9761d0ef87
cloudflare-ai-gateway: bump Opus 4.6 output limit to 128k
2026-02-12 12:39:00 -06:00
Aiden Cline
98be9a2078
fix: family
2026-02-12 12:27:16 -06:00
Dax Raad
4aa17d26cb
feat(openai): add gpt-5.3-codex-spark model
2026-02-12 13:24:43 -05:00
Aiden Cline
7f96ee576a
Merge pull request #880 from Daltonganger/feat/nano-gpt-glm5-original-models
...
feat(nano-gpt): add GLM 5 original model variants
2026-02-12 12:18:46 -06:00
Aiden Cline
bd5ce80e56
Merge pull request #882 from Daltonganger/feat/nano-gpt-add-minimax-m2.5-official
...
feat(nano-gpt): add MiniMax M2.5 Official model
2026-02-12 12:18:37 -06:00
Daltonganger
ed2af4ad45
feat(nano-gpt): add MiniMax M2.5 Official model
2026-02-12 18:07:52 +01:00
Aiden Cline
ac0868c886
Merge pull request #881 from Alex-wuhu/dev
...
add minmax-2.5 on novita
2026-02-12 10:32:50 -06:00
Aiden Cline
fdd13245cc
Revert "feat(github-copilot): add gpt-5.3-codex model ( #857 )"
...
This reverts commit 27abb8a570 .
2026-02-12 10:32:15 -06:00
Alex
37c77c58ad
Merge branch 'anomalyco:dev' into dev
2026-02-13 00:27:45 +08:00
Alex-wuhu
62ee8129e6
add minimax-m2.5 on novita
2026-02-13 00:23:06 +08:00
Daltonganger
4bc6f07570
fix(nano-gpt): correct GLM-5 dates to 2026-02-11
2026-02-12 17:22:18 +01:00
Aiden Cline
bc0336c8ec
Merge pull request #878 from cantalupo555/feat/add-openrouter-stepfun-step-3.5-flash
...
feat: add StepFun Step 3.5 Flash on OpenRouter
2026-02-12 10:13:17 -06:00
Aiden Cline
dd78db4dc6
Merge pull request #879 from amankalra172/add-stackit-provider
...
fix: reorganize STACKIT models with organization prefixes
2026-02-12 10:12:48 -06:00
Daltonganger
4fd32c741d
refactor(nano-gpt): consolidate z-ai GLM models under zai-org
2026-02-12 17:11:50 +01:00
Frank
c78ca7c132
update zen models
2026-02-12 11:05:41 -05:00
Alex
2a99397516
add GLM5 on novita ( #877 )
2026-02-12 11:03:42 -05:00
Frank
554440be4f
update zen models
2026-02-12 11:01:51 -05:00
Daltonganger
eb52a76d11
feat(nano-gpt): add GLM 5 original model variants
2026-02-12 16:48:59 +01:00
amankalra172
c9a7f6c814
fix: reorganize STACKIT models with organization prefixes and correct pricing
...
- Move models to organization subfolders (Qwen/, cortecs/, google/, etc.)
- Update pricing from EUR to USD (1.09 conversion rate)
- Fix GPT-OSS context limit to 131K tokens
- Add architectural family classifications
- Verify tool_call settings for all models
2026-02-12 14:00:45 +01:00
cantalupo555
e640802d34
feat: add StepFun Step 3.5 Flash (free) on OpenRouter
2026-02-12 08:44:15 -03:00
cantalupo555
c226863912
feat: add StepFun Step 3.5 Flash on OpenRouter
2026-02-12 08:42:37 -03:00
Alex-wuhu
8bcd634743
add GLM5 on novita
2026-02-12 16:26:05 +08:00
Aiden Cline
812cd1763a
Merge pull request #873 from juls0730/dev
...
Create cerebras/llama3.1-8b.toml
2026-02-12 00:40:48 -06:00
城二
11f4ae568e
feat(z-ai): New glm-5 model configuration file
2026-02-12 11:23:22 +08:00
juls0730
9f1629a26a
Create cerebras/llama3.1-8b.toml
2026-02-11 21:01:08 -06:00
Yunfei He
27abb8a570
feat(github-copilot): add gpt-5.3-codex model ( #857 )
...
* feat(github-copilot): add gpt-5.3-codex model
* fix(github-copilot): align gpt-5.3-codex release metadata
2026-02-11 21:53:36 -05:00
Aiden Cline
2aa4a2290e
Merge pull request #868 from dpuyosa/venice
...
Venice: Add GLM-5 model
2026-02-11 19:49:16 -06:00
Aiden Cline
c58b36c605
Merge pull request #872 from Track07-cda/openrouter-glm5
...
OpenRouter: Add GLM-5 and remove Pony Alpha
2026-02-11 19:49:07 -06:00
Aiden Cline
e91dbd1fc4
Merge pull request #870 from spiffytech/dev
...
Add Ollama Cloud support for GLM-5
2026-02-11 19:39:23 -06:00
Track07-cda
8924ee3092
feat(openrouter): add GLM-5 and remove Pony Alpha
...
Add the Z-AI GLM-5 model definition to the OpenRouter provider and
remove the deprecated Pony Alpha model.
2026-02-12 09:38:12 +08:00
spiffytech
fc5b6533d1
Added Ollama Cloud support for GLM-5
2026-02-11 19:57:46 -05:00
Aiden Cline
7a760e3a4f
Merge pull request #871 from Kunde21/synthetic_k2_5_nvfp4
...
Synthetic: Add Kimi -2 5 in NVFP4 remove GLM-4.5
2026-02-11 18:53:56 -06:00
Chad Kunde
38835801b1
synthetic: deprecate GLM-4.5
...
Model removed from models list as of 12 Feb 2026
2026-02-12 07:27:24 +07:00
Chad Kunde
81103438a3
synthetic: Add NVFP4 variant of Kimi K2.5
2026-02-12 07:25:41 +07:00
dpuyosa
0b0b36eb45
[venice] Add GLM-5 model with 198K context window
...
- Add ZAI-ORG GLM-5 model configuration to Venice provider
- Supports reasoning, tool calls, structured output
- Text in/out: 198K context, 49.5K output tokens
2026-02-11 22:53:23 +01:00
Aiden Cline
b18b73f0a0
Revert "fix: openai advertises a 400k context window, that is just the sum of max input + max output, so real context window is 272k"
...
This reverts commit ea276d57a7 .
2026-02-11 15:24:43 -06:00
Aiden Cline
3cd48b273a
Merge pull request #867 from AnishShah1803/nano-gpt/add-glm-5-models
...
Add GLM 5 to NanoGPT models list
2026-02-11 14:47:57 -06:00
twisted
890992b8ef
update release date
2026-02-11 20:37:59 +00:00
twisted
a843d84d77
Add GLM 5 to NanoGPT models list
2026-02-11 20:35:41 +00:00
Aiden Cline
d89897d07e
Merge pull request #866 from Sczr0/dev
...
Update pricing for ZAI GLM-5
2026-02-11 14:35:23 -06:00
Aiden Cline
1b26792073
fix zai
2026-02-11 14:34:43 -06:00
Aiden Cline
6a0da0a91d
Revert "Fixed ZAI GLM-5 pricing to free (0 cost)"
...
This reverts commit 79d1222e3c .
2026-02-11 14:33:39 -06:00
opencode-agent[bot]
79d1222e3c
Fixed ZAI GLM-5 pricing to free (0 cost)
...
Co-authored-by: rekram1-node <rekram1-node@users.noreply.github.com >
2026-02-11 20:14:03 +00:00
Sylvie Zhang
1ad45b44c2
add readme
2026-02-11 12:13:38 -08:00
弦塔_
45ba8066df
Update cost parameters in glm-5.toml
2026-02-12 04:07:57 +08:00
弦塔_
4861f14a65
Update cost parameters in glm-5.toml
2026-02-12 04:07:29 +08:00
弦塔_
c83b4e22bd
Update cost parameters in glm-5.toml
2026-02-12 03:56:40 +08:00
Sylvie Zhang
44c1ed5aeb
additional data cleaning logic
2026-02-11 11:48:32 -08:00
Sylvie Zhang
28c09d83a0
add fallback logic
2026-02-11 11:48:32 -08:00
Sylvie Zhang
ea41cbc4ba
draft script
2026-02-11 11:48:32 -08:00
Aiden Cline
c893ac5f9d
Merge pull request #863 from AnishShah1803/nano-gpt/update-Kimi-K2-5-models
...
Add Kimi K2.5 models to NanoGPT provider
2026-02-11 13:41:43 -06:00
twisted
8e10faf38a
set reasoning to true for kimi k2.5
2026-02-11 19:33:12 +00:00
Aiden Cline
4c3a17fbe8
Merge pull request #859 from friendliai/minpeter/add-glm5-friendli
...
Add zai-org/GLM-5 model to Friendli provider
2026-02-11 13:14:58 -06:00
twisted
f7d997e4f0
fix last_updated
2026-02-11 19:08:48 +00:00
twisted
6961c57c86
make open_weights set to true
2026-02-11 19:07:35 +00:00
minpeter
e44307c27c
Add interleaved reasoning_content to GLM-4.7 and MiniMax-M2.1
...
Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com >
2026-02-12 04:01:29 +09:00
minpeter
b9f7907150
Merge remote-tracking branch 'origin/dev' into minpeter/add-glm5-friendli
2026-02-12 04:00:31 +09:00
minpeter
bf0ee2a3eb
Add interleaved reasoning_content field for GLM-5
...
GLM models use interleaved reasoning via the reasoning_content field with OpenAI-compatible providers.
Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com >
2026-02-12 03:58:58 +09:00
Aiden Cline
b9a73edcdc
Merge pull request #862 from hanouticelina/feat/huggingface-glm-5
...
feat(huggingface): add GLM-5 for Hugging Face provider
2026-02-11 12:45:43 -06:00
Aiden Cline
7ff3be2dab
fix: github copilot model discrepencies
2026-02-11 12:40:36 -06:00
twisted
406f82e09f
Add Kimi K2.5 models to NanoGPT provider
2026-02-11 18:40:09 +00:00
Celina Hanouti
c8fa26624d
add GLM-5 for hugging face provider
2026-02-11 19:39:35 +01:00
Aiden Cline
ae31005ef3
Merge pull request #830 from amankalra172/add-stackit-provider
...
feat: add STACKIT provider with 8 AI models
2026-02-11 12:20:47 -06:00
Aiden Cline
0aa7c9f3c2
Merge pull request #852 from zainhas/dev
...
[Together AI] update output token length to match context length
2026-02-11 12:20:09 -06:00
Aiden Cline
40be65f301
Merge pull request #853 from captain1379/feat/jiekou
...
Add new models for Jiekou.AI
2026-02-11 12:19:40 -06:00
Aiden Cline
49ab2c0a48
feat: add glm 5 to zai, zhipuai, and zai coding plan
2026-02-11 12:17:49 -06:00
minpeter
7fd96a6c9d
Add zai-org/GLM-5 model to Friendli provider
...
Add GLM-5 model configuration with reasoning, tool calling, and structured output support. Update family pattern inference in generate script to recognize GLM-5 models.
Co-Authored-By: Claude Haiku 4.5 <noreply@anthropic.com >
2026-02-12 03:07:25 +09:00
Aiden Cline
fc4027fe98
Merge pull request #856 from josetorres1/add-bedrock-zai-minimax-models
...
Add GLM 4.7 Family and MiniMax M2.1 to Amazon Bedrock
2026-02-11 11:25:21 -06:00
Aiden Cline
b260564060
Merge pull request #855 from dihan-dff-user/dev
...
Add ZAI coding plan GLM-5 model
2026-02-11 11:24:47 -06:00
opencode-agent[bot]
ecd5927bed
Removed knowledge field from GLM-5 config
...
Co-authored-by: rekram1-node <rekram1-node@users.noreply.github.com >
2026-02-11 17:23:31 +00:00
Jose Torres
490381f370
Add GLM 4.7 family and MiniMax M2.1 to Amazon Bedrock provider
...
- Add GLM-4.7 (zai.glm-4.7): /bin/zsh.60/.20 per 1M tokens
- Add GLM-4.7-Flash (zai.glm-4.7-flash): /bin/zsh.07//bin/zsh.40 per 1M tokens
- Add MiniMax M2.1 (minimax.minimax-m2.1): /bin/zsh.30/.20 per 1M tokens
Pricing sources:
- AWS Bedrock pricing page: https://aws.amazon.com/bedrock/pricing/
- Model IDs confirmed via AWS console/CLI
Related to GH issue #835
2026-02-11 09:51:18 -06:00
dihan
621687e457
Add ZAI coding plan GLM-5 model
2026-02-11 20:38:40 +05:30
captain1379
d5a3bfad90
feat: add new models for Jiekou.AI
...
- Introduced `claude-opus-4-6`, `qwen3-coder-next` and `gpt-5.1` models with detailed configurations.
- Removed deprecated `qwen2.5-vl-72b-instruct` model.
- Implemented a new script for generating model configurations.
2026-02-11 15:19:40 +08:00
Zain Hasan
310bc174fa
update output token length to match context length
2026-02-10 23:05:23 -08:00
Aiden Cline
9fb1233073
Merge pull request #848 from BlockListed/cortecs-glm-models
...
Add supported z.ai GLM models to cortecs
2026-02-10 20:04:21 -06:00
Aiden Cline
029f13545b
Merge pull request #849 from BlockListed/cortecs-minimax-models
...
Add MiniMax models to cortecs
2026-02-10 20:04:12 -06:00
Aiden Cline
31b1acba5f
Merge pull request #850 from cgilly2fast/dev
...
feat(firmware): kimi and glm models
2026-02-10 20:04:00 -06:00
Aiden Cline
120881916e
Merge pull request #851 from anomalyco/fix-models
...
fix: openai advertises a 400k context window, that is just the sum of max input + max output, so real context window is 272k
2026-02-10 20:03:48 -06:00
Aiden Cline
ea276d57a7
fix: openai advertises a 400k context window, that is just the sum of max input + max output, so real context window is 272k
2026-02-10 19:02:33 -06:00
Colby Gilbert
05d940ab8d
feat(firmware): kimi and glm models
2026-02-10 14:25:46 -08:00
BlockListed
7d250ea857
cortecs add minimax models
2026-02-10 23:13:39 +01:00
BlockListed
82d72e85f8
add supported z.ai GLM models to cortecs
2026-02-10 22:58:22 +01:00
Aiden Cline
995934de32
Merge pull request #847 from rubenandre/add-bedrock-moonshotai-kimi-k2.5
...
add moonshotai kimi K2.5 to amazon-bedrock provider
2026-02-10 10:43:13 -06:00
Aiden Cline
d10392573e
Merge pull request #845 from cgilly2fast/dev
...
fix(firmware): remove unsupported model and fix name of gpt oss 20b
2026-02-10 10:04:15 -06:00
Aiden Cline
21c3c1b8e1
Merge pull request #846 from dpuyosa/venice
...
Venice: Enable reasoning for GLM-4.7-Flash model
2026-02-10 10:04:06 -06:00
Rúben Silva
4680aaedc3
add moonshotai kimi K2.5 to amazon-bedrock provider
2026-02-10 15:21:50 +00:00
dpuyosa
a9f3ad978d
[venice] Enable reasoning for GLM-4.7-Flash model
2026-02-10 10:11:48 +01:00
Colby Gilbert
de75687ea1
fix(firmware): remove unsupported model and fix name of spt oss 20b
2026-02-09 23:01:40 -08:00
Aiden Cline
539cc930c4
Merge pull request #843 from cgilly2fast/dev
...
chore: clean up firmware available models
2026-02-09 18:40:24 -06:00
Colby Gilbert
31f3c63acc
fix(firmware): remove reasoning from anthropic and deepseek models
2026-02-09 16:07:44 -08:00
Colby Gilbert
b019787ad8
chore: clean up firmware available models
2026-02-09 16:03:09 -08:00
Aiden Cline
e41fca18a2
Merge pull request #842 from riccardogiorato/dev
...
fix: Increase output limit to match context for kimi K2.5 on Together
2026-02-09 17:12:07 -06:00
Riccardo Giorato
74163f7314
Increase output limit to match context
...
Update providers/togetherai/models/moonshotai/Kimi-K2.5.toml to set [limit].output from 32_768 to 262_144. This aligns the output token limit with the context size (262_144) to avoid premature truncation and allow full-length responses.
2026-02-09 22:39:34 +01:00
Aiden Cline
7b763695fd
Merge pull request #839 from shelvick/add-azure-kimi-k2.5
...
Add Azure Kimi-K2.5 model
2026-02-09 14:15:33 -06:00
Aiden Cline
686b47d01e
Merge pull request #840 from shelvick/add-azure-claude-opus-4-6
...
Add Azure Claude Opus 4.6 model
2026-02-09 14:15:16 -06:00
Scott Helvick
46f0726d7f
Add Azure Claude Opus 4.6 model
2026-02-09 20:07:31 +00:00
Scott Helvick
3c14600fc6
Add Azure Kimi-K2.5 model
2026-02-09 19:49:02 +00:00
Aiden Cline
721c025af1
Merge pull request #836 from PeppeRu96/feat/add-deepinfra-claude
...
feat: add DeepInfra Claude Opus 4 and Claude Sonnet 3.7 (latest) models
2026-02-09 12:29:34 -06:00
Aiden Cline
c591f9b213
Merge pull request #837 from PeppeRu96/feat/add-deepinfra-deepseek
...
feat: add DeepInfra DeepSeek models
2026-02-09 12:23:02 -06:00
Aiden Cline
57580b28d3
Merge pull request #765 from captain1379/feat/jiekou
...
feat: add Jiekou.AI provider
2026-02-09 12:22:21 -06:00
Giuseppe Ruggeri
11e92f093f
fix: fix price for DeepInfra DeepSeek-V3.2
2026-02-09 14:32:34 +01:00
Giuseppe Ruggeri
17cf21ba46
feat: add DeepInfra DeepSeek models
2026-02-09 14:29:02 +01:00
Giuseppe Ruggeri
60a3f09b8e
fix: update deepinfra/claude-3-7-sonnet-latest family field
2026-02-09 14:02:52 +01:00
Giuseppe Ruggeri
6f907bce35
feat: add DeepInfra Claude Opus 4 and Claude Sonnet 3.7 (latest) models
2026-02-09 13:56:16 +01:00
Frank
1f20d47ef5
update zen models
2026-02-08 21:43:43 -05:00
Aiden Cline
d5c23c9c95
Merge pull request #827 from modpotato/dev
...
fix: rename glm 5 stealth from 'Stealth' to 'Pony Alpha' + remove status
2026-02-08 14:03:59 -06:00
Aiden Cline
125abf1a21
Merge pull request #829 from 888-wzk/feature/chenger_20260128
...
fix: Update model configurations to adjust reasoning and interleaved …
2026-02-08 14:03:44 -06:00
Aiden Cline
38ccea666f
Merge pull request #831 from spiffytech/dev
...
Add Ollama Cloud support for qwen3-coder-next
2026-02-08 14:03:28 -06:00
Frank
42ca5faeb8
sync
2026-02-08 14:21:34 -05:00
spiffytech
bcd9e3dba1
Added Ollama Cloud support for qwen3-coder-next
2026-02-08 13:36:54 -05:00
amankalra172
7504dc2947
feat: add STACKIT provider with 8 AI models
...
Add STACKIT as a new provider with complete model specifications:
Chat Models:
- Llama 3.1 8B Instruct FP8
- Llama 3.3 70B Instruct FP8
- GPT-OSS 120B
- Mistral Nemo Instruct 2407 FP8
- Gemma 3 27B (multimodal)
- Qwen3-VL 235B (vision-language)
Embedding Models:
- E5 Mistral 7B
- Qwen3-VL Embedding 8B (multimodal)
All models include:
- Proper schema compliance (attachment, reasoning, tool_call, etc.)
- Pricing in USD per million tokens
- Context limits and modalities
- Official STACKIT logo with currentColor support
STACKIT is a German sovereign cloud provider offering OpenAI-compatible
AI model serving with open-source models.
2026-02-08 12:35:51 +01:00
城二
67bddb6b61
Merge branch 'dev' of https://github.com/888-wzk/models.dev into feature/chenger_20260128
2026-02-08 11:23:44 +08:00
城二
77330e78c6
fix: Update model configurations to adjust reasoning and interleaved fields
2026-02-08 11:21:58 +08:00
mod
e55a05f6a6
Merge branch 'anomalyco:dev' into dev
2026-02-07 00:43:00 -05:00
mod
e5ce677899
fix: rename glm 5 stealth from 'Stealth' to 'Pony Alpha'
2026-02-07 00:42:50 -05:00
Aiden Cline
e1747322ad
Merge pull request #826 from modpotato/dev
...
add pony alpha (glm 5 stealth)
2026-02-06 23:21:17 -06:00
Aiden Cline
9303c7be2e
Merge pull request #825 from cantalupo555/feat/add-openrouter-mimo-v2-flash
...
feat: add Xiaomi MiMo-V2-Flash on OpenRouter
2026-02-06 16:43:34 -06:00
John Doe
204eb52c0d
feat: pony alpha (glm 5 demo) on openrouter
2026-02-06 21:06:54 +00:00
John Doe
a2abd136f5
feat: pony alpha (glm 5 demo) on openrouter
2026-02-06 21:02:40 +00:00
Aiden Cline
ea6e487e77
fix: change anthropic default to 200k instead of 1M since not everyone can access the 1M
2026-02-06 13:46:09 -06:00
Aiden Cline
1033ee450c
Merge pull request #821 from 888-wzk/feature/chenger_20260128
...
Added Claude Opus 4.6 model configuration file
2026-02-06 11:00:49 -06:00
Aiden Cline
de8e46b2ab
Merge pull request #823 from dpuyosa/venice
...
Venice: Tweak model generation script
2026-02-06 11:00:37 -06:00
Aiden Cline
8181d97317
Merge pull request #824 from vglafirov/feat/gitlab-opus-4-6
...
feat(gitlab): add Claude Opus 4.6 model (duo-chat-opus-4-6)
2026-02-06 11:00:11 -06:00
Vladimir Glafirov
a5c9640163
feat(gitlab): add Claude Opus 4.6 model (duo-chat-opus-4-6)
...
Add the newly released Claude Opus 4.6 model for GitLab Duo Agentic Chat.
Related:
- AI Gateway MR: https://gitlab.com/gitlab-org/modelops/applied-ml/code-suggestions/ai-assist/-/merge_requests/4492
2026-02-06 17:04:17 +01:00
cantalupo555
c220f2a790
feat: add Xiaomi MiMo-V2-Flash on OpenRouter
2026-02-06 12:52:35 -03:00
Frank
c88c849e5a
Merge pull request #822 from imdevarsh/imdevarsh/openrouter-opus-4.6
...
feat(openrouter): add claude opus 4.6 to openrouter models list
2026-02-06 10:07:38 -05:00
dpuyosa
75ff468a9a
[venice] Refactor model generation with privacy field
...
- Add optional privacy field to ModelSpec schema
- Use privacy field to determine open_weights capability
- Preserve existing output token limit when smaller than proposed
2026-02-06 13:18:47 +01:00
Devarsh
5b9186f6a9
feat(openrouter): add claude opus 4.6 to openrouter models list
2026-02-06 18:06:33 +13:00
城二
974713311b
feat: Added Claude Opus 4.6 model configuration file
2026-02-06 11:11:19 +08:00
Aiden Cline
2d143f96d1
Merge pull request #813 from cgilly2fast/dev
...
feat: add opus 4.6 to firmware provider
2026-02-05 16:25:47 -06:00
Aiden Cline
4c4cd139f8
Merge pull request #812 from fhennerkes/dev
...
poe: add Claude Opus 4.6 model
2026-02-05 16:25:29 -06:00
Aiden Cline
22688b1260
Merge pull request #816 from markusylisiurunen/add-eu-opus-4.6
...
Add the missing EU variant back for Opus 4.6 on AWS Bedrock
2026-02-05 16:23:54 -06:00
Aiden Cline
13631caba9
Merge pull request #817 from dpuyosa/venice
...
Venice: Add Claude Opus 4.6 and GLM 4.7 models
2026-02-05 16:21:41 -06:00
dpuyosa
7e901e93bf
[venice] Add Claude Opus 4.6 and GLM 4.7 models
...
- Add claude-opus-4.6 model configuration for Venice provider
- Add zai-org-glm-4.7-flash model configuration for Venice provider
2026-02-05 22:44:27 +01:00
Markus Ylisiurunen
ace9626163
also fix pricing for opus 4.5
2026-02-05 23:25:25 +02:00
Markus Ylisiurunen
2bd869959b
fix pricing
2026-02-05 23:12:40 +02:00
Markus Ylisiurunen
efecbc137c
Add EU variant for Opus 4.6
2026-02-05 23:03:01 +02:00
Colby Gilbert
cae3f84930
Merge branch 'anomalyco:dev' into dev
2026-02-05 12:49:38 -08:00
Colby Gilbert
37cdea639f
feat: add opus 4.6
2026-02-05 12:49:19 -08:00
Ryan Vogel
2c67792e4b
Merge pull request #811 from anomalyco/add-claude-opus-4-6
...
Fix Claude Opus 4.6 model IDs and remove incorrect variants
2026-02-05 15:49:03 -05:00
Ryan Vogel
24addedada
Fix Vertex AI model ID to claude-opus-4-6@default
2026-02-05 15:46:07 -05:00
Ryan Vogel
dc9f404dbb
Condense AGENTS.md model configuration section
2026-02-05 15:44:00 -05:00
Ryan Vogel
db2212ab9f
Update AGENTS.md with model configuration learnings
2026-02-05 15:42:47 -05:00
Ryan Vogel
e2777a44ed
Fix Vertex AI model ID: claude-opus-4-6@default -> claude-opus-4-6
2026-02-05 15:41:53 -05:00
fhennerkes
c0f0394f67
poe: add Claude Opus 4.6 model
2026-02-05 12:39:00 -08:00
Ryan Vogel
a762f47461
Fix model IDs: remove unannounced dated alias, remove EU Bedrock, fix Bedrock ID (v1:0 -> v1), fix Vertex ID (@20260205 -> @default)
...
Fixes #809
2026-02-05 15:38:28 -05:00
Aiden Cline
768f841f79
Merge pull request #781 from Dagnan/add-glm-4.7-flash-deepinfra
...
feat(deepinfra): add GLM-4.7-Flash model
2026-02-05 14:34:27 -06:00
Aiden Cline
00f239a852
Merge pull request #770 from jerilynzheng/feat/add-vercel-models-jan-30
...
vercel: add new models and interleaved support
2026-02-05 14:19:20 -06:00
Michel Pigassou
69f72041e1
Added missing interleaved/reasoning_content for GLM 4.7-Flash
2026-02-05 21:16:25 +01:00
Aiden Cline
43e98540ec
fix: output limit for opus 4.6 on gh copilot
2026-02-05 14:15:56 -06:00
jerilynzheng
f0854ab7b8
vercel: add interleaved = true for confirmed models
...
Add interleaved reasoning support to models confirmed by other providers:
- Claude: 3.7-sonnet, haiku-4.5, opus-4/4.1/4.5/4.6, sonnet-4/4.5
- DeepSeek: R1, V3.2-thinking
- MiniMax: M2, M2.1
- Kimi: K2-thinking, K2-thinking-turbo, K2.5
- GLM: 4.5, 4.6, 4.7, 4.7-flashx
Co-Authored-By: Claude Opus 4.5 <noreply@anthropic.com >
2026-02-05 12:14:02 -08:00
Aiden Cline
a05a47097f
Merge pull request #799 from iamanishx/deepinfra-kimi
...
feat: added support for kimi k2.5 (deepinfra)
2026-02-05 14:08:08 -06:00
jerilynzheng
ed96ac7c74
fix: update Claude Opus 4.6 knowledge cutoff to 2025-05
...
Co-Authored-By: Claude Opus 4.5 <noreply@anthropic.com >
2026-02-05 12:03:40 -08:00
jerilynzheng
3985ee8556
vercel: add Claude Opus 4.6
...
Add anthropic/claude-opus-4.6 from Vercel AI Gateway:
- 1M context window, 128K output
- $5.00/$25.00 per 1M tokens (input/output)
- Supports vision, reasoning, and tool use
Co-Authored-By: Claude Opus 4.5 <noreply@anthropic.com >
2026-02-05 12:03:11 -08:00
Aiden Cline
53c150347f
Merge pull request #808 from stickyburn/chutes-qwen3-coder-next
...
chore: add qwen3-coder-next for chutes.ai
2026-02-05 14:01:46 -06:00
Aiden Cline
1b4b15d73b
Merge pull request #807 from smrdotgg/patch-1
...
Fix release and last updated dates for GPT-5.3 Codex
2026-02-05 14:01:37 -06:00
stickyburn
535b1afe4f
chore: add qwen3 coder next for chutes.ai
2026-02-05 14:41:59 -05:00
Mathis
e239c4d51e
Add configuration for Claude Opus 4.6 model ( #806 )
2026-02-05 14:24:07 -05:00
imanishx
3dddd7b5e3
fix: interleaved opt added
...
Signed-off-by: imanishx <manishbiswal754@gmail.com >
2026-02-05 19:05:25 +00:00
Semere Tereffe
941487c8d4
Fix release and last updated dates for GPT-5.3 Codex
2026-02-05 21:52:29 +03:00
Ryan Vogel
ce3bbe64a3
Update context window to 1M tokens for all Claude Opus 4.6 models
2026-02-05 13:39:48 -05:00
Aiden Cline
183dd5c4f3
Merge pull request #803 from anomalyco/add-claude-opus-4-6
...
Add Claude Opus 4.6 model
2026-02-05 12:39:29 -06:00
Ryan Vogel
b4733df6b7
Merge branch 'dev' into add-claude-opus-4-6
2026-02-05 13:38:54 -05:00
Ryan Vogel
921ec8f8fc
Add cost.context_over_200k long context pricing to all Claude Opus 4.6 models
2026-02-05 13:36:52 -05:00
Aiden Cline
6c28f3974c
Merge pull request #804 from rexdotsh/feat/add-anthropic-opus-4-6
...
feat: add opus 4.6
2026-02-05 12:34:34 -06:00
Aiden Cline
8b3a813da9
Merge pull request #802 from dmmulroy/cloudflare-opus-4-6
...
cloudflare-ai-gateway: add claude opus 4.6
2026-02-05 12:34:00 -06:00
rexdotsh
10297c8490
feat: add opus 4.6
2026-02-05 23:58:13 +05:30
Ryan Vogel
61c8d50df2
Add Claude Opus 4.6 model across Anthropic, Bedrock, and Vertex AI providers
2026-02-05 13:26:28 -05:00
Dax Raad
4faf622d1e
add gpt-5.3-codex.toml
2026-02-05 13:26:24 -05:00
Dillon Mulroy
89db412884
cloudflare-ai-gateway: add claude opus 4.6
2026-02-05 13:24:36 -05:00
Frank
107b285e1c
update zen models
2026-02-05 13:14:06 -05:00
Frank
8bd58cf186
update zen models
2026-02-05 13:04:56 -05:00
Aiden Cline
9573a8fccc
Merge pull request #796 from manascb1344/nebius-token-factory-models
...
feat: add Nebius Token Factory models
2026-02-05 11:45:09 -06:00
Aiden Cline
0153e6408a
Merge pull request #792 from Alex-wuhu/dev
...
feat: add deepseek OCR model configuration and Qwen3 Coder Next model…
2026-02-05 11:44:28 -06:00
massaindustries
6ecd9ec509
add qwen-next-coder-2
2026-02-05 09:11:08 +00:00
massaindustries
bb42d0b855
add qwen-next-coder
2026-02-05 09:09:30 +00:00
imanishx
8cb19035ff
feat: added tomal for kini k2.4 (deepinfra)
...
Signed-off-by: imanishx <manishbiswal754@gmail.com >
2026-02-05 09:08:33 +00:00
captain1379
7c0e1e142f
fix: removed old models
2026-02-05 14:02:59 +08:00
captain1379
7aeca69c4e
fix: fix logo
2026-02-05 13:44:38 +08:00
Aiden Cline
cfde47ca60
Revert "Update Amazon Bedrock models to add cross-region inference and remove deprecated models"
...
This reverts commit bc58036964 .
2026-02-04 12:11:29 -06:00
Aiden Cline
b01c07a3d0
Merge pull request #793 from zainhas/patch-1
...
[fix] Rename model to 'Qwen3 Coder Next FP8'
2026-02-04 10:33:22 -06:00
Aiden Cline
444c3071ee
Merge pull request #795 from riccardogiorato/dev
...
remove wrongly typed Kimi-K2-5.toml
2026-02-04 10:32:27 -06:00
manascb1344
42a79c717d
feat(nebius): update Meta-Llama, NVIDIA models and mark deprecated
...
- Update Llama-3.3-70B-Instruct (Base & Fast) with new pricing
- Mark Llama-3.1-405B-Instruct as deprecated (no longer available)
- Update Llama-3.1-Nemotron-Ultra-253B-v1 with new pricing
- Mark DeepSeek-V3 as deprecated (replaced by V3.2 and V3-0324)
2026-02-04 20:18:37 +05:30
manascb1344
578df73ffb
feat(nebius): update Z.ai, OpenAI, Moonshot AI, and NousResearch models
...
- Update GLM-4.5 and GLM-4.5-Air with new pricing
- Update gpt-oss-120b and gpt-oss-20b with new pricing and features
- Update Kimi-K2-Instruct with new pricing and multimodal support
- Update Hermes-4-405B and Hermes-4-70B with new pricing
2026-02-04 20:17:51 +05:30
manascb1344
2639e20a97
feat(nebius): add new models from Z.ai, Moonshot AI, Meta, and NVIDIA
...
- Add GLM-4.7 and GLM-4.7-FP8 (Z.ai)
- Add Kimi-K2-Thinking (Moonshot AI)
- Add Llama-Guard-3-8B, Meta-Llama-3.1-8B-Instruct (Base & Fast) (Meta)
- Add Nemotron-Nano-V2-12b and NVIDIA-Nemotron-3-Nano-30B-A3B (NVIDIA)
2026-02-04 20:17:14 +05:30
manascb1344
ca6206b78e
feat(nebius): add Qwen models to Token Factory
...
- Add Qwen3-Next-80B-A3B-Thinking
- Add Qwen3-30B-A3B-Thinking-2507 and Qwen3-30B-A3B-Instruct-2507
- Add Qwen3-Coder-30B-A3B-Instruct
- Add Qwen3-32B (Base & Fast)
- Add Qwen2.5-Coder-7B-fast
- Add Qwen2.5-VL-72B-Instruct
- Add Qwen3-Embedding-8B
2026-02-04 20:16:46 +05:30
manascb1344
51fe42982f
feat(nebius): add DeepSeek models to Token Factory
...
- Add DeepSeek-V3.2, DeepSeek-V3-0324 (Base & Fast), DeepSeek-R1-0528 (Base & Fast)
- These are new models available on Nebius Token Factory
2026-02-04 20:16:21 +05:30
manascb1344
4c78ea9f36
feat(nebius): add new providers for Nebius Token Factory
...
- Add MiniMaxAI provider with MiniMax-M2.1 model
- Add PrimeIntellect provider with INTELLECT-3 model
- Add black-forest-labs provider with FLUX.1-schnell and FLUX.1-dev
- Add BAAI provider with bge-multilingual-gemma2 and BGE-ICL
- Add intfloat provider with e5-mistral-7b-instruct
- Add Google provider with Gemma-2-2b-it, Gemma-2-9b-it-fast, Gemma-3-27b-it, and Gemma-3-27b-it-fast
2026-02-04 20:16:01 +05:30
Riccardo Giorato
9092f0b106
Delete Kimi-K2-5.toml
2026-02-04 11:22:49 +01:00
Zain Hasan
1acd3c199a
Rename model to 'Qwen3 Coder Next FP8'
2026-02-04 01:57:58 -08:00
Alex-wuhu
7deb00a333
feat: add deepseek OCR model configuration and Qwen3 Coder Next model configuration
2026-02-04 16:56:28 +08:00
samsja
f180f49df5
Add Intellect 3 model from Prime Intellect
2026-02-03 23:58:12 -08:00
Aiden Cline
59f13d1c0a
feat: make all openrouter models use openrouter sdk
2026-02-03 23:15:41 -06:00
Aiden Cline
ce6950074f
Revert "Add Bedrock cross-region inference profiles and update validation"
...
This reverts commit 89f62005cc .
2026-02-03 23:08:05 -06:00
Aiden Cline
b2b0f612f4
Merge pull request #788 from zainhas/dev
...
[Together AI] add qwen3 coder next
2026-02-03 22:48:52 -06:00
Aiden Cline
d1e92ce8ad
Merge pull request #790 from anomalyco/update-cf-workers
...
fix: update cf workers ai
2026-02-03 22:48:41 -06:00
Aiden Cline
20c81eb600
fix: update cf workers ai
2026-02-03 22:47:12 -06:00
Frank
6934bf2c66
Merge pull request #789 from qychen2001/dev
...
feat(models): add Kimi-K2.5 model support
2026-02-03 22:55:46 -05:00
QiyuanChen
83503944ba
feat(models): add Kimi-K2.5 model support
...
Add support for Moonshot AI's Kimi-K2.5 model with reasoning capabilities,
structured output, and multi-modal support (text/image input, text output).
Configured with a large context window of 262,000 tokens for both input
and output. Added to both SiliconFlow and SiliconFlow CN providers.
2026-02-04 11:45:28 +08:00
Zain Hasan
536ac44708
add qwen3 coder next
2026-02-03 14:54:42 -08:00
Aiden Cline
02f7969d53
Merge pull request #786 from unexge/push-lpupkorvtnuw
...
Update Amazon Bedrock models to add cross-region inference and remove deprecated models
2026-02-03 15:30:30 -06:00
Matt Silverlock
0ba8852f91
use official ai-gateway-provider package for Cloudflare AI Gateway
2026-02-03 15:41:11 -05:00
Burak Varlı
89f62005cc
Add Bedrock cross-region inference profiles and update validation
...
- Add Nova models for Global, US, EU, and APAC regions
- Add Llama 3.1/3.2 cross-region profiles for US and EU
- Add Claude Sonnet 4/3.7 APAC profiles
- Add Claude Sonnet 4.5/3.7/3.5 US Gov profiles
- Update validate-bedrock to include ap-southeast-1 region
- Skip us-gov models in validation (requires GovCloud access)
2026-02-03 20:10:40 +00:00
Aiden Cline
5afc754db3
Merge pull request #697 from berget-ai/feat/add-berget-ai-provider
...
feat: add Berget.AI provider
2026-02-03 12:15:36 -06:00
Aiden Cline
2fdfeecfc8
Merge pull request #784 from bendews/patch-1
...
Increase Github Copilot GPT 4.1 context limit from 64k to 128k
2026-02-03 09:24:29 -06:00
Aiden Cline
7bf852e19f
Merge pull request #785 from thePrnvBot/chore--updating-free-openrouter-models
...
fix: Update models tool call to false
2026-02-03 09:24:07 -06:00
Burak Varlı
bc58036964
Update Amazon Bedrock models to add cross-region inference and remove deprecated models
...
This change adds a new script to validate all Amazon Bedrock models by making a simple inference request using model identifiers.
As a result of that script, made some changes to make sure all model identifiers are usable via Amazon Bedrock:
- Added cross-region inference for various models including DeepSeek, Llama, Amazon Nova
- Removed some reprecated/EoL'd models including Amazon Titan, Claude v2, Cohere Command Light
2026-02-03 13:43:07 +00:00
thePrnvBot
613843b529
fix: update nousresearch model tool call to false
2026-02-03 16:12:39 +04:00
thePrnvBot
836b07aeaf
fix: update cognitivecomputation model tool call to false
2026-02-03 16:12:10 +04:00
thePrnvBot
e4f8c752ac
fix: update allenai model tool call to false
2026-02-03 16:11:48 +04:00
thePrnvBot
fa04882d5f
fix: update liquid models tool call to false
2026-02-03 16:11:36 +04:00
thePrnvBot
e69df42539
fix: update llama model tool call to false
2026-02-03 16:11:18 +04:00
thePrnvBot
64b7e989eb
fix: update tng-r1t-chimera :free tool call to false
2026-02-03 16:10:53 +04:00
Ben Dews
abf1259e58
Increase context limit from 64k to 128k
2026-02-03 21:22:15 +10:00
Aiden Cline
93fe136ec1
Merge pull request #777 from xiaojiezj/zenmux_dev
...
feat: ““Replace the chat-completion protocol in the Zenmux provider with the Anthropic protocol, and replace the model.”
2026-02-02 20:50:32 -06:00
Aiden Cline
be098329c5
Merge pull request #782 from fhennerkes/dev
...
poe: model update 2/2/26
2026-02-02 20:49:42 -06:00
fhennerkes
884e901c3f
poe: model update 2/2/26
2026-02-02 18:22:49 -08:00
Michel Pigassou
81ddc26ed0
feat(deepinfra): add GLM-4.7-Flash model
...
Add zai-org/GLM-4.7-Flash to DeepInfra provider
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-02-02 22:04:36 +01:00
Aiden Cline
e35973bdf9
Merge pull request #775 from Track07-cda/alibaba-cn_kimi
...
feat(alibaba-cn): add Kimi K2 Thinking and K2.5 models to alibaba-cn provider
2026-02-02 10:52:51 -06:00
Aiden Cline
f94d9dda7f
Merge pull request #607 from unexge/push-svvwrlmunkkt
...
Add cross-region inference profiles for Claude 4.x family models in Amazon Bedrock
2026-02-02 10:20:13 -06:00
xiaojie.zj
d354b55137
feat: “Replace the chat-completion protocol with the Anthropic protocol, and replace the model.”
2026-02-02 18:03:42 +08:00
Track07-cda
e30f361b88
feat(alibaba-cn): add reasoning content support for Kimi models
...
Add interleaved reasoning_content field to Kimi K2 Thinking and K2.5.
Also correct the display name for Moonshot Kimi K2.5.
2026-02-02 16:28:02 +08:00
Track07-cda
efa90f6fbf
feat(alibaba-cn): add Kimi K2 Thinking and K2.5 models
...
Add new model definitions for Moonshot Kimi K2 Thinking and K2.5.
Update Moonshot Kimi K2 Instruct metadata including open weights
status and output token limits.
2026-02-02 14:13:06 +08:00
Aiden Cline
93d03d87c1
Merge pull request #772 from thePrnvBot/chore--updating-free-openrouter-models
...
feat: add free openrouter models
2026-01-31 21:41:49 -06:00
Aiden Cline
67b31f3371
Merge pull request #773 from ccurme/cc/gpt-5.2-structured-output
...
fix: add structured_output to gpt-5.1 and 5.2
2026-01-31 20:58:12 -06:00
Aiden Cline
0513b73b17
fix: correct model id
2026-01-31 20:39:40 -06:00
Chester Curme
866974df3a
add structured_output to gpt-5.1 and 5.2
2026-01-31 21:39:31 -05:00
thePrnvBot
fd07fe7953
feat: add free qwen models to openrouter provider
2026-01-31 20:25:30 +04:00
thePrnvBot
d176299fdf
feat: add free nemotron models to openrouter provider
2026-01-31 20:24:53 +04:00
thePrnvBot
6193824e92
feat: add gpt oss free models to openrouter
2026-01-31 20:23:25 +04:00
thePrnvBot
b85b481fd0
feat: add hermes 3 llama 3.1 405b free model
2026-01-31 20:22:54 +04:00
thePrnvBot
e33d225a0b
chore: update deepseek r1 0528 free tool call to false
2026-01-31 20:22:02 +04:00
thePrnvBot
3c9e76cf89
feat: add tng-r1t-chimera free model
2026-01-31 20:21:22 +04:00
thePrnvBot
42e7f67b67
feat: add dolphin mistral 24b venice edition
2026-01-31 20:20:53 +04:00
thePrnvBot
0e46820a00
feat: add seedream model
2026-01-31 20:20:16 +04:00
thePrnvBot
e7dd66e51a
feat: add free meta llama models
2026-01-31 20:19:28 +04:00
thePrnvBot
ea1d856847
feat: add free black forest lab models
2026-01-31 20:18:46 +04:00
thePrnvBot
26fd570130
feat: added free liquid, sourceful and allenai models
2026-01-31 20:17:30 +04:00
Aiden Cline
008c521304
Merge branch 'dev' into feat/add-vercel-models-jan-30
2026-01-30 16:35:40 -06:00
Aiden Cline
c6870e97c5
Merge pull request #717 from MichaelYochpaz/fix-vertex-anthropic-npm-import
...
fix(google-vertex-anthropic): Fix incorrect NPM package used for Anthropic models used through Vertex
2026-01-30 15:57:19 -06:00
jerilynzheng
e7e8af6934
vercel: add 5 new models from Vercel AI Gateway
...
Add new models:
- alibaba/qwen3-max-thinking: Qwen 3 Max with reasoning
- arcee-ai/trinity-large-preview: Trinity 400B MoE model
- moonshotai/kimi-k2.5: Kimi K2.5 with vision and reasoning
- openai/gpt-4o-mini-search-preview: GPT-4o Mini search variant
- zai/glm-4.7-flashx: GLM 4.7 Flash lightweight model
Co-Authored-By: Claude Sonnet 4.5 <noreply@anthropic.com >
2026-01-30 13:38:07 -08:00
Aiden Cline
665a9fe17d
Merge pull request #767 from remorses/model-schema
...
add model-schema.json endpoint for model autocomplete
2026-01-30 13:32:54 -06:00
Aiden Cline
98114d5721
fix or glm flash
2026-01-30 13:21:57 -06:00
Aiden Cline
2deedb5fa5
Merge pull request #762 from davidcharbonnier/dev
...
Add GLM 4.7 Flash model on Openrouter
2026-01-30 13:20:33 -06:00
Aiden Cline
43cb68a64e
Merge pull request #768 from amazon-nova-api/nova-provider
...
Add nova as a model provider
2026-01-30 13:19:57 -06:00
Adnan Hajar
ee89f7ee6b
Add nova as a model provider
2026-01-30 14:01:14 -05:00
Tommy D. Rossi
6f45f3949f
add model-schema.json endpoint for model autocomplete
2026-01-30 15:51:04 +01:00
David Charbonnier
ebafef01d7
feat: add glm 4.7 flash model on openrouter
2026-01-30 09:06:48 -05:00
captain1379
77dccbe959
feat: add Jiekou.AI provider
...
Add Jiekou.AI as a new LLM provider with 102 models including:
- DeepSeek (V3, R1, OCR)
- Qwen (Qwen3, Qwen2.5)
- Claude (Opus, Sonnet, Haiku)
- GPT models (GPT-5.x, GPT-4.x, GPT-OSS)
- Gemini (Pro, Flash)
- GLM (4.5, 4.7)
- Kimi (K2, K2.5)
- Llama (3.x, 4.x)
- And more...
Jiekou.AI is an OpenAI-compatible API provider.
Co-Authored-By: Claude (pa/claude-opus-4-5-20251101) <noreply@anthropic.com >
2026-01-30 18:44:11 +08:00
Frank
8b2b4b40a1
update zen models
2026-01-30 00:52:31 -05:00
Frank
96da5d8331
update zen models
2026-01-29 16:46:15 -05:00
Frank
0146cb114e
update zen models
2026-01-29 16:38:42 -05:00
Aiden Cline
21177b3f6b
Merge pull request #760 from cgilly2fast/dev
...
feat(firmware): add kimi models and clean up model names
2026-01-29 15:02:50 -06:00
Colby Gilbert
522c486815
fix: wrong name for kimi k2.5
2026-01-29 12:16:15 -08:00
Colby Gilbert
85ef0fe0a8
feat: add kimi models
2026-01-29 10:33:25 -08:00
Colby Gilbert
70681d3398
chore: rename glm and gpt oss models
2026-01-29 10:33:17 -08:00
Frank
c2a6830fde
sync
2026-01-29 12:38:07 -05:00
Frank
4b9631cb89
update zen models
2026-01-29 12:35:06 -05:00
Aiden Cline
9efb6c1a73
Merge pull request #742 from 888-wzk/feature/chenger_20260128
...
feat(models): Add configuration files for the Kimi K2.5, GPT-5.2-Codex, Qwen3-Max-Thinking, and GLM 4.7 FlashX models.
2026-01-29 10:45:16 -06:00
Aiden Cline
9b5adb8230
Merge pull request #751 from cravenceiling/fix/openrouter-google-gemma-models
...
add and fix some google gemma models from openrouter
2026-01-29 10:44:52 -06:00
Aiden Cline
9105b7ba75
Merge pull request #753 from otterDeveloper/patch-1
...
fireworks: Raise Kimi K2.5 max output
2026-01-29 10:44:40 -06:00
Aiden Cline
a0dc4149cd
Merge pull request #754 from fanweixiao/dev
...
fix(vivgrid): set npm for gemini-3 models for vivgrid provider
2026-01-29 10:43:38 -06:00
Aiden Cline
ccb98b9597
Merge pull request #755 from friendliai/minpeter/add-minimax-friendli-model
...
feat(friendli): add MiniMax M2.1 model and update Qwen3
2026-01-29 10:42:14 -06:00
Aiden Cline
c33581d89d
Merge pull request #757 from s-scheck/feature/adjust-pricing-of-devstral-2512
...
feat: adjust pricing of devstral-2512 hosted by mistral
2026-01-29 10:41:58 -06:00
Aiden Cline
ab1a8c21db
Merge pull request #759 from FrancoStino/patch-5
...
Delete providers/nvidia/models/z-ai/glm-4.7.toml
2026-01-29 10:41:45 -06:00
Davide Ladisa
f969e060c8
Delete providers/nvidia/models/z-ai/glm-4.7.toml
...
Duplicate
https://github.com/anomalyco/models.dev/blob/dev/providers/nvidia/models/z-ai/glm4.7.toml
2026-01-29 17:17:59 +01:00
massaindustries
a46966efef
add-regolo-02
2026-01-29 15:14:02 +00:00
massaindustries
19ace4436f
add-regolo-01
2026-01-29 12:36:44 +00:00
Sinan Scheck
66823bcd6e
chore: adjust pricing of devstral-2512 hosted by mistral
2026-01-29 11:45:20 +01:00
minpeter
01f391f1fe
Add MiniMax M2.1 model and update Qwen3 date
...
Co-Authored-By: Claude Opus 4.5 <noreply@anthropic.com >
2026-01-29 18:55:00 +09:00
minpeter
1d2ddf9b1d
Update friendli provider model configs
...
Remove outdated Qwen3 and Llama 4 model configurations.
Reorganize meta-llama models into subdirectory.
Upgrade GLM model to 4.7 with increased context limits (202,752 tokens).
2026-01-29 18:42:08 +09:00
C.C. Fan
a34c2a390b
fix(vivgrid): set npm for gemini-3 models for vivgrid provider
2026-01-29 16:43:32 +08:00
Frank
0c0d917719
sync
2026-01-29 03:16:34 -05:00
otterDeveloper
e5dcee15bb
fireworks: increase kimi-k2p5 max output
2026-01-29 02:04:24 -06:00
城二
0894de420e
feat(models): Add family and interleaved fields to Kimi K2.5 and GLM 4.7 FlashX configuration files
2026-01-29 10:36:04 +08:00
Aiden Cline
abfad44131
Merge pull request #734 from riccardogiorato/dev
...
[together.ai] add four new models: Qwen3-235B, Qwen3-Next-80B, Kimi-K2-Instruct, GLM 4.7
2026-01-28 20:13:26 -06:00
Aiden Cline
cc3566a9a2
Merge pull request #747 from monotykamary/update-kimi-k2.5-model
...
feat(synthetic): update Kimi K2.5 model configuration
2026-01-28 20:12:59 -06:00
Aiden Cline
71b90fe6ce
Merge pull request #748 from dpuyosa/venice
...
Venice: Adjust limit context/output token limits for all models
2026-01-28 20:12:42 -06:00
Aiden Cline
c854148259
Merge pull request #749 from esafak/chore/kimi-k2.5
...
chore: disable `temperature` in `moonshotai/kimi-k2.5`
2026-01-28 20:12:30 -06:00
Aiden Cline
980b64e860
Merge pull request #750 from esafak/feat/zai-glm-4.7-flash
...
feat: Add `zai-coding-plan/glm-4.7-flash`
2026-01-28 20:12:12 -06:00
cravenceiling
632fc61b95
add to the limit section
2026-01-28 19:33:02 -05:00
cravenceiling
5a76cee029
add and fix some google gemma models from openrouter
2026-01-28 19:07:09 -05:00
Emre Şafak
0971bd00f3
feat: Add zai-coding-plan/glm-4.7-flash.
...
* Create the file `providers/zai-coding-plan/models/glm-4.7-flash.toml` to define the new model.
* Set the model name to `GLM-4.7-Flash`.
* Configure model capabilities including reasoning, tool call, and knowledge cutoff of `2025-04`.
* Define context limit as `200_000` tokens.
* Set input and output costs to zero.
2026-01-28 18:44:46 -05:00
Emre Şafak
94fca7064d
chore: disable temperature in moonshotai/kimi-k2.5
2026-01-28 18:11:32 -05:00
dpuyosa
40ad678ae7
[venice] Adjust limit context/output token limits for all models
...
- All models limit context/output was reduced by 2.4%
2026-01-28 18:34:59 +01:00
Tom X Nguyen
e18738adda
feat(synthetic): update Kimi K2.5 model configuration
...
Update model config with corrected values:
- max_output: 65_536 (from 32_768)
- cost.input: 0.55, cost.output: 2.19
- modalities.input: [text, image]
- Add interleaved section for reasoning_content
- Use underscores for large numbers
2026-01-28 23:39:59 +07:00
Aiden Cline
0b003c9c18
add new arcee models
2026-01-28 10:50:59 -05:00
Aiden Cline
cb5035a8ee
Merge pull request #743 from zainhas/dev
...
add kimi k2.5
2026-01-28 10:30:17 -05:00
Aiden Cline
01e3069901
Merge pull request #745 from reissbaker/kk25
...
Add Kimi K2.5 for Synthetic
2026-01-28 10:29:59 -05:00
Aiden Cline
7b577ad7a8
Merge pull request #737 from cravenceiling/add-google-gemma-3-27b-it-free
...
feat: add google-gemma-3-27b-it:free model
2026-01-28 10:29:21 -05:00
Matt Baker
55e72aa8db
Correct output tokens
2026-01-28 02:08:58 -08:00
Matt Baker
a8fc1d2e44
Add Kimi K2.5 for Synthetic
2026-01-28 02:07:52 -08:00
Riccardo Giorato
101b5042bd
Create Kimi-K2-5.toml
2026-01-28 10:57:37 +01:00
Riccardo Giorato
7b835b2f29
Merge remote-tracking branch 'upstream/dev' into dev
2026-01-28 10:50:29 +01:00
Aiden Cline
9658a500f7
Merge pull request #740 from thatoddmailbox/dev
...
Fix Kimi pricing for fireworks-ai
2026-01-28 02:24:39 -05:00
Aiden Cline
7030a7c77f
Merge pull request #741 from Alex-wuhu/dev
...
feat(models): add Kimi K2.5 and GLM-4.7-Flash model
2026-01-28 02:24:12 -05:00
Frank
2be2a8c109
Update zai models
2026-01-28 01:49:08 -05:00
Frank
465335102f
Update kimi-k2.5.toml
2026-01-28 01:38:18 -05:00
城二
82ddcad9f8
feat(models): Add configuration files for the Kimi K2.5, GPT-5.2-Codex, Qwen3-Max-Thinking, and GLM 4.7 FlashX models.
2026-01-28 14:31:21 +08:00
Zain Hasan
9acd2ffa60
add kimi k2.5
2026-01-27 22:30:24 -08:00
Alex-wuhu
b3d2cfdc34
feat(models): add Kimi K2.5 and GLM-4.7-Flash model
2026-01-28 14:25:53 +08:00
Alex Studer
eead89fd8e
fix kimi pricing for fireworks-ai
2026-01-28 01:20:01 -05:00
Aiden Cline
48de510380
Merge pull request #635 from mthezi/feature/add-302ai-provider
...
feat: add 302ai provider
2026-01-27 22:01:19 -05:00
Aiden Cline
52332705ca
Merge pull request #736 from alissonlauffer/chore/update-chutes-kimi-k2.5
...
feat(chutes): update Kimi K2.5 TEE model capabilities
2026-01-27 22:00:07 -05:00
Aiden Cline
08e5d0f830
Update Kimi-K2.5-TEE.toml configuration settings
2026-01-27 21:59:39 -05:00
Aiden Cline
0b7f253ee0
Merge pull request #739 from xinrui-z/feat/aihubmix-add-models
...
feat(models): add kimi-k2.5, coding-glm-4.7, glm-4.6v, and qwen3-max
2026-01-27 21:54:59 -05:00
Xinrui
bfe953d2f0
feat(models): add kimi-k2.5, coding-glm-4.7, glm-4.6v, and qwen3-max
2026-01-28 10:48:40 +08:00
cravenceiling
641fa6f2e7
feat: add google-gemma-3-27b-it:free model
2026-01-27 19:34:05 -05:00
Alisson Lauffer
c344db1bf8
feat(chutes): update Kimi K2.5 TEE model capabilities
...
Enable reasoning, tool calling, and multimodal input support for the
Kimi K2.5 TEE model. Increase context limit from 32k to 262k tokens and
output limit from 8k to 65k tokens. Add support for image and video
inputs alongside text. Configure interleaved reasoning content field.
2026-01-27 21:05:33 -03:00
Aiden Cline
36c6206d32
Merge pull request #732 from mmealman/add_fireworks_k2p5
...
Added Kimi K2.5 to FireworksAI.
2026-01-27 17:54:31 -05:00
Aiden Cline
4f6a59d7be
Merge pull request #729 from gary149/feat/huggingface-kimi-k2.5
...
feat(huggingface): add Kimi-K2.5 model
2026-01-27 17:54:15 -05:00
Aiden Cline
f5b8e3fe83
Merge pull request #735 from spiffytech/dev
...
Add Kimi K2.5 to Ollama Cloud
2026-01-27 17:53:31 -05:00
Aiden Cline
07c70ca9f2
Merge pull request #731 from arguiot/add-vercel-kimi-k2.5
...
Add Kimi K2.5 to Vercel provider
2026-01-27 17:53:22 -05:00
Aiden Cline
d35ad7ec49
Merge pull request #727 from ProlowN/dev
...
fix : removed duplicate kimi k2.5 model from venice
2026-01-27 17:53:09 -05:00
Aiden Cline
7c57f4ce15
Merge pull request #733 from dpuyosa/dev
...
Venice: Add interleaved thinking to k2.5
2026-01-27 17:52:44 -05:00
spiffytech
83eeb304a5
Add Kimi K2.5 to Ollama Cloud
2026-01-27 16:08:16 -05:00
Aiden Cline
a13f101e0c
Merge pull request #638 from jerome-benoit/feature/sap-ai-core-updates
...
fix(sap-ai-core): use working provider fork for stable OpenCode integration
2026-01-27 15:32:40 -05:00
Riccardo Giorato
ec6101a629
Merge remote-tracking branch 'upstream/dev' into dev
2026-01-27 21:00:55 +01:00
Riccardo Giorato
231313aad0
Add four new models: Qwen3-235B, Qwen3-Next-80B, Kimi-K2-Instruct, GLM-4.7
2026-01-27 21:00:44 +01:00
dpuyosa
b9793731e6
Add interleaved thinking to k2.5
2026-01-27 20:32:38 +01:00
Frank
22edc4d92d
update zen model
2026-01-27 14:12:42 -05:00
Frank
3f62b2dd5a
update moonshot models
2026-01-27 14:05:13 -05:00
Mark Mealman
d57592dba3
Added Kimi K2.5 to FireworksAI.
2026-01-27 14:00:19 -05:00
Frank
e2b43f180c
Merge pull request #730 from esafak/moonshotai/kimi-k2.5
...
chore: add `moonshotai/kimi-k2.5` model
2026-01-27 13:59:59 -05:00
Arthur Guiot
0acff9cf7c
add Kimi K2.5 to Vercel provider
2026-01-27 10:50:37 -08:00
Emre Şafak
563c43f004
add moonshotai/kimi-k2.5 model
2026-01-27 13:46:30 -05:00
Victor Muštar
e53bb9c7ad
feat(huggingface): add Kimi-K2.5 model
2026-01-27 18:38:18 +01:00
Frank
1522bc4a9a
update zen models
2026-01-27 12:34:50 -05:00
Frank
c28701d579
update zen models
2026-01-27 12:34:29 -05:00
Magnus
eb5bff1f6a
fix : removed duplicate kimi k2.5 model from venice
2026-01-27 18:01:37 +01:00
Frank
15b4b02e6e
update zen models
2026-01-27 12:00:36 -05:00
Jan Szypulski
22d6a24c7a
fix: llama 3.3 last update
2026-01-27 18:00:11 +01:00
Jan Szypulski
c7bc5b7c98
fix: corrected logo color and size
2026-01-27 17:59:55 +01:00
Aiden Cline
b1910161d4
Merge pull request #726 from ProlowN/dev
...
Added kimi k2.5 to Venice AI
2026-01-27 11:45:18 -05:00
Magnus
13e48c2ca0
fix/ wrong output size
2026-01-27 17:44:27 +01:00
Magnus
516cfe355d
fix/ wrong family name
2026-01-27 17:09:38 +01:00
Magnus
068eacd6b3
Added kimi k2.5 to Venice AI
2026-01-27 17:06:02 +01:00
Aiden Cline
336e43494b
Merge pull request #719 from Jakey-Jakey/dev
...
add-kimi-k2.5 from OpenRouter
2026-01-27 11:05:16 -05:00
Aiden Cline
7fc046f833
Merge pull request #720 from kassieclaire/add-kimi-k2p5-model
...
feat(providers): add Kimi K2.5 model
2026-01-27 11:04:43 -05:00
Jan Szypulski
2c63a024b3
delete unrecognized model family
2026-01-27 17:04:36 +01:00
Aiden Cline
a572cf8a1a
Merge pull request #721 from matthusby/dev
...
[Chutes] Add new model configs and update pricing for several models
2026-01-27 11:03:54 -05:00
Aiden Cline
e335f919f2
Merge pull request #722 from FrancoStino/dev
...
feat(providers): Add NVIDIA models: Kimi K2.5 and GLM-4.7
2026-01-27 11:03:38 -05:00
Aiden Cline
4e883ea026
Merge branch 'dev' into dev
2026-01-27 11:02:10 -05:00
Aiden Cline
3dfb74d1ea
Merge pull request #723 from arshadbarves/feat/nvidia-kimi-k2.5
...
feat(nvidia): add Kimi K2.5 multimodal model
2026-01-27 11:01:42 -05:00
Aiden Cline
edb551b275
Merge pull request #724 from dpuyosa/dev
...
Venice: Add Kimi K2.5 model configuration
2026-01-27 11:01:29 -05:00
Jan Szypulski
2f34ee47ee
add cloudferro logo
2026-01-27 16:48:42 +01:00
Jan Szypulski
f6cb6631b8
add cloudferro sherlock models
2026-01-27 16:48:28 +01:00
dpuyosa
d95d22e89c
[venice] Add Kimi K2.5 model configuration
...
- Add new Kimi K2.5 model with 262K context support
- Include pricing for input, output, and cache_read operations
- Enable reasoning, tool calling, and structured output capabilities
- Support text and image input with text output
2026-01-27 16:23:08 +01:00
Davide Ladisa
dc771f54df
Update knowledge and release dates in kimi-k2.5.toml
2026-01-27 15:52:43 +01:00
Arshad Barves
27b99e9ccf
feat(nvidia): add Kimi K2.5 multimodal model
...
Add Kimi K2.5, a 1T parameter multimodal MoE model by Moonshot AI
with support for text, image, and video inputs.
Key features:
- 256K context window (262,144 tokens)
- Native multimodal support (text, image, video)
- Interleaved reasoning with reasoning_content field
- Tool calling and temperature control
- Open weights available
Model ID: moonshotai/kimi-k2.5
Provider: NVIDIA NIM
Validation: ✅ Passes bun validate
2026-01-27 20:00:42 +05:30
Davide Ladisa
c04069b5a3
Merge pull request #102 from FrancoStino/add-nvidia-models-kimi-glm
...
Add NVIDIA models: Kimi K2.5 and GLM-4.7
2026-01-27 15:01:43 +01:00
Davide Ladisa
af08a750b2
Add GLM-4.7 with correct filename and family field
2026-01-27 15:01:13 +01:00
Davide Ladisa
ed4270cc8f
Remove old glm4_7.toml to rename to glm-4.7.toml
2026-01-27 15:01:03 +01:00
Davide Ladisa
dae3873284
Fix GLM-4.7 release date to December 2025 and update knowledge cutoff
2026-01-27 14:59:23 +01:00
Davide Ladisa
6daad4c4eb
Update knowledge cutoff dates to more accurate values
2026-01-27 14:57:07 +01:00
Davide Ladisa
5302dc4452
Add NVIDIA models: Kimi K2.5 and GLM-4.7
2026-01-27 14:53:55 +01:00
Matt Husby
7d26d504ec
Add new model configs and update pricing for several models
2026-01-27 07:58:35 -05:00
kassieclaire
63f116b27f
fix: remove interleaved reasoning for kimi-k2.5
2026-01-27 07:05:35 -05:00
kassieclaire
bf6582bb83
fix: update knowledge cutoff to 2025-01 for kimi-k2.5
2026-01-27 06:25:46 -05:00
Kassie Povinelli
965f5365bc
Update providers/kimi-for-coding/models/k2p5.toml
...
checked docs for kimi-for-coding plan, still shows up as this lower value, so going with it for now -- keep an eye on the docs in case they update the information
Co-authored-by: Copilot <175728472+Copilot@users.noreply.github.com >
2026-01-27 06:19:39 -05:00
kassieclaire
3b5717f6b9
feat(providers): add Kimi K2.5 model
2026-01-27 06:12:01 -05:00
Jakey-Jakey
cae7be4925
Add 'video' modality to input options
2026-01-27 04:01:00 -05:00
Jakey-Jakey
d218bb7fe8
Add cache_read cost to kimi-k2.5 configuration
2026-01-27 03:56:18 -05:00
Jakey-Jakey
9f5b80af72
Add knowledge parameter with value '2025-01'
2026-01-27 03:55:37 -05:00
Jakey-Jakey
6c928db808
Remove knowledge field from kimi-k2.5.toml
...
Remove knowledge field from configuration.
2026-01-27 03:54:38 -05:00
Jakey-Jakey
63b1cfb9fc
Add provider section to kimi-k2.5.toml
2026-01-27 03:53:43 -05:00
Jakey-Jakey
ade61a5018
Add files via upload
2026-01-27 03:45:43 -05:00
Aiden Cline
e768c2afbd
Merge pull request #713 from qychen2001/dev
...
feat(providers): update siliconflow-cn model catalog
2026-01-26 21:01:47 -05:00
Aiden Cline
31e503a516
Merge pull request #715 from dpuyosa/dev
...
Venice: Add cache_read to GLM 4.7
2026-01-26 21:00:43 -05:00
Frank
98a455cb0f
sync
2026-01-26 18:24:50 -05:00
Michael Yochpaz
f1d2e47772
fix(google-vertex-anthropic): use @ai-sdk/google-vertex/anthropic npm package
...
The google-vertex-anthropic provider requires the `/anthropic` subpath import for thinking/reasoning to work correctly with Claude models on Vertex AI.
2026-01-26 22:09:05 +00:00
dpuyosa
4d82211cea
Add cache_read to GLM 4.7
2026-01-26 22:59:16 +01:00
mthezi
4e0a2d34b4
refactor(models): update family names for various models to improve consistency
2026-01-26 14:47:20 +08:00
⌞L⌝
effa34d17b
Merge branch 'anomalyco:dev' into feature/add-302ai-provider
2026-01-26 14:29:57 +08:00
QiyuanChen
ed59411f9e
feat(providers): update siliconflow-cn model catalog
...
Add new Pro tier models for deepseek-ai and moonshotai, including DeepSeek-R1, DeepSeek-V3 series, and Kimi-K2-Thinking models with reasoning capabilities. Remove older Qwen, Kimi-K2, and other legacy model configurations.
2026-01-26 12:57:56 +08:00
Aiden Cline
1286f6449c
Merge pull request #710 from hsyysy/dev
...
feat(provider): add DeepSeek-V3.2 for Nvidia
2026-01-25 22:56:41 -05:00
Aiden Cline
f92551d3ac
Merge pull request #709 from fanweixiao/feat/add-vivgrid-models
...
add gpt-5.1-codex-max, gpt-5.2-codex and more models for vivgrid provider
2026-01-25 22:56:31 -05:00
Aiden Cline
8c502a36b9
Delete pnpm-lock.yaml
2026-01-25 21:29:19 -05:00
Aiden Cline
6e40a4744a
Merge pull request #712 from xinrui-z/fix/aihubmix-provider-invalid-type
...
fix(provider): correct invalid type in provider.toml
2026-01-25 21:28:54 -05:00
Xinrui
5ac346644a
fix(provider): correct invalid type in provider.toml
2026-01-26 10:12:08 +08:00
Thomas Young
c03332fb3d
feat(provider): add DeepSeek-V3.2 for Nvidia
2026-01-25 20:10:23 +08:00
C.C. Fan
acb8319afb
add gpt-5.1-codex-max, gpt-5.2-codex, gemini-3-pro-preview and gemini-3-flash-preview for vivgrid provider
2026-01-25 16:37:27 +08:00
Aiden Cline
568f5319be
Merge pull request #708 from jsdtxm/feat/add-glm-4.7
...
feat(provider): add Pro/zai-org/GLM-4.7 for SiliconFlow-CN
2026-01-24 23:39:40 -05:00
lazy
2b331310b7
fix(qihang-ai): rename provider and fix logo to match standards
...
- Rename provider from qihang to qihang-ai
- Update logo to use standard size (24x24) and currentColor
Co-Authored-By: Claude Sonnet 4.5 <noreply@anthropic.com >
2026-01-25 12:34:48 +08:00
xiamin
0fb16d2cb8
feat(provider): add Pro/zai-org/GLM-4.7 for SiliconFlow-CN
2026-01-25 12:16:12 +08:00
Aiden Cline
18e555c2af
Merge pull request #656 from fchange/feat/new-provider
...
feat: add moark provider
2026-01-24 23:15:07 -05:00
Aiden Cline
1938af666f
Merge pull request #707 from arshadbarves/fix/nvidia-glm4.7-model-id
...
fix(nvidia): correct model ID for GLM-4.7 (z-ai/glm4.7)
2026-01-24 23:12:22 -05:00
Arshad Barves
fea35d7bb4
fix(nvidia): correct model ID for GLM-4.7 (z-ai/glm4.7)
...
Rename model file from glm-4.7.toml to glm4.7.toml to generate the
correct model ID z-ai/glm4.7 (without dot) as per NVIDIA API specification.
The model ID is derived from the file path, so the filename must match
the exact model identifier used by the provider's API.
- Renamed: providers/nvidia/models/z-ai/glm-4.7.toml → glm4.7.toml
- Model ID: z-ai/glm-4.7 → z-ai/glm4.7
- Validation: ✅ Passes bun validate
2026-01-25 09:34:50 +05:30
Aiden Cline
53b821523b
Merge pull request #700 from vglafirov/feat/gitlab-gpt-5-2
...
feat(gitlab): add GPT-5.2 model definition (duo-chat-gpt-5-2)
2026-01-24 12:50:30 -05:00
Aiden Cline
813b2d57b3
Merge pull request #704 from jsdtxm/feat/add-minimax-m2-1
...
feat(provider): add MiniMax M2.1 for SiliconFlow
2026-01-24 12:50:18 -05:00
xiamin
ee5c39bb18
fix: move MiniMax-M2.1 config
2026-01-24 22:17:29 +08:00
xiamin
ec2bf4bf7c
feat(provider): add MiniMax M2.1 for SiliconFlow-CN
2026-01-24 16:16:03 +08:00
xiamin
d2d6bc2d6c
chore: remove MiniMax-M2.1.toml symlink
2026-01-24 16:15:15 +08:00
xiamin
a4978b8b1c
feat(provider): add MiniMax M2.1 for SiliconFlow
2026-01-24 16:08:32 +08:00
Frank
545bf83089
update zen models
2026-01-23 23:19:34 -05:00
Vladimir Glafirov
5651a0efe1
feat(gitlab): add GPT-5.2 model definition (duo-chat-gpt-5-2)
2026-01-23 16:10:47 +01:00
Frank
b5fc3e3f54
update zen models
2026-01-23 01:19:18 -05:00
Frank
c8f6d7ace2
update zen models
2026-01-23 01:12:50 -05:00
Frank
4dd2e77ad1
update zen models
2026-01-23 01:05:55 -05:00
Aiden Cline
e5e859ac63
fix context limit for copilto gpt-4.1
2026-01-22 19:37:27 -06:00
Luca Steeb
92269282eb
fix: use correct family for gemma and gpt-oss models
...
- Gemma models now use "gemma" family instead of "gemini"
- GPT OSS models now use "gpt-oss" family instead of "gpt"
Co-Authored-By: Claude Opus 4.5 <noreply@anthropic.com >
2026-01-23 01:36:29 +00:00
Luca Steeb
6b9b340fbc
fix: map llmgateway family to auto
...
Co-Authored-By: Claude Opus 4.5 <noreply@anthropic.com >
2026-01-23 01:28:38 +00:00
Luca Steeb
e8a6793654
fix: use valid models.dev family enum values
...
Maps internal family names to valid models.dev families:
- moonshot → kimi
- bytedance → seed
- zai → glm
- nvidia → nemotron
Co-Authored-By: Claude Opus 4.5 <noreply@anthropic.com >
2026-01-23 01:27:15 +00:00
Luca Steeb
b22ff136a8
fix: add required output limit to all models
...
models.dev schema requires limit.output field.
Defaults to 16384 when not specified.
Co-Authored-By: Claude Opus 4.5 <noreply@anthropic.com >
2026-01-23 01:25:53 +00:00
Luca Steeb
e28d4e387f
chore: trigger CI
2026-01-23 01:21:06 +00:00
Luca Steeb
09b5dd4d84
refactor: remove scripts/ dir, link to repo script
...
Removes empty generate.ts file and scripts/ directory.
README now links to llmgateway repo for regeneration.
Co-Authored-By: Claude Opus 4.5 <noreply@anthropic.com >
2026-01-23 01:18:53 +00:00
Luca Steeb
24e575a86e
refactor: flatten model structure to models/ directory
...
Removes provider subdirectories, exports all models directly
to models/ folder for simpler structure.
Co-Authored-By: Claude Opus 4.5 <noreply@anthropic.com >
2026-01-23 01:15:27 +00:00
Luca Steeb
3549abec36
feat: add LLM Gateway provider with 153 models
...
Add LLM Gateway (llmgateway.io) as a new provider with all supported models
organized by upstream provider subdirectory.
LLM Gateway is an OpenAI-compatible API gateway that provides unified
access to 40+ LLM providers through a single API endpoint.
Directory structure:
providers/llmgateway/
├── provider.toml
├── README.md
├── scripts/
│ └── generate.ts
└── models/
├── anthropic/ (16 models)
├── openai/ (28 models)
├── google/ (19 models)
├── zai/ (17 models - GLM, CogView)
├── alibaba/ (27 models - Qwen)
├── meta/ (12 models - Llama)
├── xai/ (9 models - Grok)
├── deepseek/ (5 models)
├── bytedance/ (6 models - Seed)
├── moonshot/ (4 models - Kimi)
├── mistral/ (3 models)
├── perplexity/ (3 models - Sonar)
├── minimax/ (1 model)
├── nvidia/ (1 model)
└── llmgateway/ (2 models - auto, custom)
Co-Authored-By: Claude Opus 4.5 <noreply@anthropic.com >
2026-01-23 01:10:53 +00:00
Christian Landgren
eb98dd4305
fix: address Copilot review comments
...
- Change Mistral family from 'mistral' to 'mistral-small' for consistency
- Fix Llama 3.3 70B knowledge date from '2024-12' to '2023-12'
- Set tool_call to false for KB-Whisper-Large (speech-to-text models don't support tool calling)
2026-01-23 01:10:35 +01:00
Christian Landgren
cb8d8e8698
chore: remove Qwen3 32B model
2026-01-23 01:05:12 +01:00
Christian Landgren
3cb9a1cd3d
feat: add Berget.AI provider
...
Add Berget.AI as an OpenAI-compatible provider with base URL api.berget.ai/v1.
Models included:
- Text: Llama 3.3 70B, Qwen3 32B, GPT-OSS-120B, GLM 4.7, Mistral Small 3.2 24B
- Embedding: Multilingual-E5-large-instruct, Multilingual-E5-large
- Rerank: bge-reranker-v2-m3
- Speech-to-Text: KB-Whisper-Large
2026-01-23 01:01:40 +01:00
Aiden Cline
830a03e46b
Merge pull request #696 from vglafirov/feat/gitlab-openai-models
...
fix: increase output token limit for GitLab Claude models to 64k
2026-01-22 15:20:31 -08:00
Vladimir Glafirov
b21c8870a5
fix: align GitLab Claude models with native Anthropic model capabilities
...
Updated to match native Anthropic model definitions:
- attachment: false → true (supports image/pdf attachments)
- reasoning: false → true (supports extended thinking)
- modalities.input: ["text"] → ["text", "image", "pdf"]
- Added knowledge cutoff dates from native models
2026-01-23 00:18:27 +01:00
Vladimir Glafirov
4bcba6a482
fix: increase output token limit for GitLab Claude models to 64k
...
The output limit was set to 4,096 tokens which caused tool calls with
large content (like file generation) to be truncated mid-JSON.
Updated to match standard Anthropic model limits:
- duo-chat-opus-4-5: 4,096 → 64,000
- duo-chat-sonnet-4-5: 4,096 → 64,000
- duo-chat-haiku-4-5: 4,096 → 64,000
2026-01-23 00:02:49 +01:00
Aiden Cline
c67ccd8def
Merge pull request #694 from cgilly2fast/dev
...
chore: remove deepseek-coder for firmware provider
2026-01-22 11:51:38 -08:00
Colby Gilbert
7aa00eb8dc
chore: remove deepseek-coder for firmware provider
2026-01-22 11:48:50 -08:00
Aiden Cline
d799a6ae6e
Merge pull request #692 from vglafirov/feat/gitlab-openai-models
...
feat(gitlab): add OpenAI GPT-5 model definitions
2026-01-22 08:52:22 -08:00
Vladimir Glafirov
a770639c25
feat(gitlab): add OpenAI GPT-5 model definitions
...
Add GitLab Duo model definitions for OpenAI GPT-5 family:
- duo-chat-gpt-5-1: GPT-5.1 flagship model
- duo-chat-gpt-5-mini: GPT-5 Mini (cost-effective)
- duo-chat-gpt-5-codex: GPT-5 Codex (agentic coding)
- duo-chat-gpt-5-2-codex: GPT-5.2 Codex
2026-01-22 17:43:59 +01:00
Jan Szypulski
d934e26168
add cloudferro sherlock as provider
2026-01-22 16:11:45 +01:00
mthezi
ea20440d0d
fix: update model family name for gpt-4.1-nano
2026-01-22 13:51:54 +08:00
Aiden Cline
eef424f296
Merge pull request #686 from zhzy0077/nvidia-patch
...
Add nvidia 2 new models.
2026-01-21 16:23:25 -08:00
Aiden Cline
05415ee2ec
Merge pull request #685 from spiffytech/dev
...
Remove duplicate GLM-4.7 model file
2026-01-21 16:20:01 -08:00
Aiden Cline
23e99a093a
Merge pull request #687 from eliasto/ovhcloud/update-models
...
Update OVHcloud AI Endpoints models
2026-01-21 16:19:52 -08:00
Aiden Cline
02df983581
Merge pull request #688 from gitpush-gitpaid/dev
...
Added PDF to input modalities for gpt 5.2 codex
2026-01-21 16:19:36 -08:00
Aiden Cline
3d102d3bd9
Add 'pdf' to input modalities in gpt-5.2-codex.toml
2026-01-21 18:19:14 -06:00
gitpush-gitpaid
6307a2c223
added PDF to input modalities for gpt 5.2 codex
2026-01-21 18:30:18 -05:00
Aiden Cline
d79ae1d684
chore: kill deprecated copilot models from list
2026-01-21 16:59:11 -06:00
Elias TOURNEUX
67d192dd9c
Update OVHcloud AI Endpoints models
2026-01-21 08:17:47 -05:00
lazy
74cb010892
feat(qihang): add Gemini 2.5 Flash and GPT-5.2 models
2026-01-21 15:19:30 +08:00
zhzy0077
10acfc848d
Add nvidia 2 new models.
2026-01-21 08:39:12 +08:00
spiffytech
5943a24d41
Remove duplicate GLM-4.7 model file
2026-01-20 14:02:19 -05:00
Aiden Cline
a52b64222e
Merge pull request #684 from sebastiand-cerebras/final-removal-of-glm4_6
...
Remove deprecated zai-glm-4.6 model (Jan 20, 2026)
2026-01-20 10:17:47 -08:00
Seb Duerr
a767bf0a6d
Remove deprecated zai-glm-4.6 model (Jan 20, 2026)
...
Thank you for your patience and understanding with our timeline adjustments! I truly appreciate your team's responsiveness and flexibility in working with us on this deprecation.
As of January 20, 2026, the zai-glm-4.6 model has been officially deprecated.
2026-01-20 09:56:33 -08:00
Aiden Cline
b131f86a1f
Merge pull request #666 from spiffytech/dev
...
Update Ollama Cloud models. Add generator for model files.
2026-01-20 08:06:40 -08:00
Aiden Cline
c84e382bbe
Merge pull request #679 from WSQS/dev
...
feat: add GLM-4.7-Flash for zhipuai provider
2026-01-20 08:03:16 -08:00
Aiden Cline
9de5f304fe
Merge pull request #683 from nickdowse/dev
...
Fix: Fix incorrect OpenAI, Gemini prices
2026-01-20 08:03:07 -08:00
Aiden Cline
8a854771d7
Merge pull request #677 from ivivek/dev
...
feat: add GLM-4.7 to google-vertex
2026-01-20 08:02:57 -08:00
Aiden Cline
b190cdaecc
Merge pull request #678 from dpuyosa/UpdateModel
...
Venice: Update provider package
2026-01-20 08:02:47 -08:00
Aiden Cline
5712350b30
Merge pull request #680 from cgilly2fast/cgilly2fast/firmware-provider
...
feat: add cerebras glm 4.7 and gpt OSS, clean up claude model ids
2026-01-20 08:02:12 -08:00
Nick Dowse
b933688a77
Fix incorrect openai, gemini prices
2026-01-20 10:06:03 -05:00
dpuyosa
64f034bb72
Update interleaved field to reasoning_content
...
- Change field value in claude-sonnet-45, gemini-3-flash-preview, qwen3-235b-a22b-thinking-2507, and zai-org-glm-4.7 configs
2026-01-20 15:34:42 +01:00
Frank
72de414c2f
Merge pull request #681 from tars90percent/minimax-provider-names
...
Add MiniMax coding plan providers
2026-01-20 09:11:13 -05:00
Frank
bc6698d98b
sync
2026-01-20 09:10:16 -05:00
lazy
b465cec21a
feat: add QiHang provider with 7 models
...
- Add QiHang provider configuration (OpenAI-compatible API)
- API endpoint: https://api.qhaigc.net/v1
- Add 7 models:
- gpt-5.2-codex (/bin/zsh.14/.14)
- gpt-5-mini (/bin/zsh.04//bin/zsh.29)
- claude-opus-4-5-20251101 (/bin/zsh.71/.57)
- claude-sonnet-4-5-20250929 (/bin/zsh.43/.14)
- claude-haiku-4-5-20251001 (/bin/zsh.14//bin/zsh.71)
- gemini-3-flash-preview (/bin/zsh.07//bin/zsh.43)
- gemini-3-pro-preview (/bin/zsh.57/.43)
- All configurations validated with bun validate
2026-01-20 16:45:35 +08:00
tars90percent
e0fcf8f638
Add MiniMax coding plan providers
2026-01-20 13:42:43 +08:00
Colby Gilbert
36a6197da9
feat: add cerebras glm 4.7 and gpt OSS, clean up claude model ids
2026-01-19 21:34:34 -08:00
WSQS
f6d82c43a7
feat: add GLM-4.7-Flash for zhipuai
2026-01-20 10:49:21 +08:00
dpuyosa
44b8ed5871
Comment-out 'api' for validation script
2026-01-20 01:30:13 +01:00
dpuyosa
c0d9ec4777
Update Venice provider package:
...
- Replace @ai-sdk/openai-compatible with venice-ai-sdk-provider
- Fix cache_control limitations
- Add Venice-specific features
2026-01-20 01:11:34 +01:00
spiffytech
2ae1e23591
Update Ollama Cloud models. Add generator for model files.
2026-01-19 17:29:38 -05:00
Vivek K
1ff1405664
feat: add GLM-4.7 to google-vertex
2026-01-20 00:43:34 +05:30
Aiden Cline
1c32145339
Merge pull request #674 from zerone0x/add/gpt-5.1-codex-max
...
feat(openrouter): add openai/gpt-5.1-codex-max model
2026-01-19 09:52:54 -08:00
Aiden Cline
fbebe356b5
Merge pull request #675 from ElecTwix/glm-4.7-flash
...
feat: add glm-4.7-flash model
2026-01-19 09:52:25 -08:00
ElecTwix
e694f0136f
feat: add glm-4.7-flash model
2026-01-19 20:44:29 +03:00
zerone0x
5062058b6a
feat(openrouter): add openai/gpt-5.1-codex-max model
...
Add GPT-5.1-Codex-Max model to OpenRouter provider. This model is available
in OpenRouter's API but was missing from models.dev.
Pricing sourced from OpenRouter API.
Co-Authored-By: Claude <noreply@anthropic.com >
2026-01-20 01:25:22 +08:00
Aiden Cline
7b132f2cd8
Merge pull request #673 from gary149/feat/huggingface-glm-4.7-flash
...
feat(huggingface): add GLM-4.7-Flash model
2026-01-19 08:49:44 -08:00
Victor Muštar
e319a707fd
feat(huggingface): add GLM-4.7-Flash model
2026-01-19 17:36:59 +01:00
Aiden Cline
627ac7bcf1
Merge pull request #671 from uniquename/ollama/glm-4.7
...
feat: add Ollama GLM-4.7 model configuration file
2026-01-19 07:38:50 -08:00
Aiden Cline
89408c71e0
Merge pull request #669 from dpuyosa/UpdateModel
...
Venice: Replace vision models glm4.6v -> qwen3-vl
2026-01-19 07:38:29 -08:00
Aiden Cline
0651768fd9
Merge pull request #672 from sebastiand-cerebras/add-glm4_6-deprecation-notice
...
Re-add zai-glm-4.6 temporarily until Jan 20, 2026
2026-01-19 07:38:00 -08:00
Seb Duerr
c8ce0db2b1
Re-add zai-glm-4.6 temporarily until Jan 20, 2026
...
Thanks for the incredibly fast merge! We appreciate the efficiency, though we need to temporarily re-add GLM 4.6. The model will be officially deprecated on January 20, 2026. Our apologies for any confusion - we should have been clearer about the timeline in the original PR.
2026-01-19 07:14:34 -08:00
User
c3b177ed7a
feat: add Ollama GLM-4.7 model configuration file
2026-01-19 12:45:35 +00:00
dpuyosa
1ce4dc41c1
Update model configurations:
...
- Add qwen3-vl-235b-a22b model
- Remove deprecated zai-org-glm-4.6v model
2026-01-19 10:49:32 +01:00
Aiden Cline
438e834043
add input field to more openai models
2026-01-19 00:59:44 -06:00
Jérôme Benoit
189aa03281
Apply suggestion from @jerome-benoit
2026-01-19 04:02:10 +01:00
Aiden Cline
5f293ca6ce
Merge pull request #665 from sebastiand-cerebras/removal_of_glm4_6
...
Remove deprecated zai-glm-4.6 model from Cerebras provider
2026-01-17 22:50:25 -08:00
Aiden Cline
490a03f2d9
Remove deprecated zai-glm-4.6 model from Cerebras provider
2026-01-17 22:49:51 -08:00
Aiden Cline
cbd215cbc1
Merge pull request #652 from hueyexe/dev
...
feat: add GPT 5.2 Codex to Azure and Azure Cognitive Services
2026-01-17 22:48:21 -08:00
Aiden Cline
1b49c365a0
fix: restore gpt-5.2-codex.toml as symlink to fix validation CI
2026-01-18 00:46:56 -06:00
Aiden Cline
51d2a4f2b6
Merge pull request #664 from jerome-benoit/feat/sap-ai-core-claude-4.5-opus
...
feat(sap-ai-core): add Claude 4.5 Opus and align pricing
2026-01-17 19:02:07 -08:00
Seb Duerr
afba23ed62
Remove deprecated zai-glm-4.6 model from Cerebras provider
2026-01-17 18:37:18 -08:00
Jérôme Benoit
9e25ca1521
feat(sap-ai-core): add Claude 4.5 Opus and align pricing
...
- Add Claude 4.5 Opus model with official Anthropic pricing
- Align cache pricing for Claude 3 Sonnet, Gemini 2.5 models, and GPT-5 Mini with official pricing
2026-01-18 01:08:12 +01:00
Aiden Cline
971e8734ae
Merge pull request #659 from KagurazakaNyaa/dev
...
Update SiliconFlow model list
2026-01-16 20:35:02 -08:00
Aiden Cline
ef7731ec13
Merge pull request #661 from cgilly2fast/cgilly2fast/firmware-provider
...
fix: make gpt-nano and mini calculate as 0 price
2026-01-16 20:32:32 -08:00
Colby Gilbert
e2d8670828
chore: update firmware provider docs url
2026-01-16 17:04:53 -08:00
神楽坂·喵
ecfab1717a
Merge branch 'anomalyco:dev' into dev
2026-01-17 08:21:48 +08:00
KagurazakaNyaa
ac037c57ad
fix pangu family
2026-01-17 08:20:30 +08:00
KagurazakaNyaa
a886d60715
fix kat family
2026-01-17 08:11:32 +08:00
Colby Gilbert
44e980e810
fix: make gpt-nano and mini calculate as 0 price
2026-01-16 15:42:37 -08:00
Aiden Cline
d1f3ddfe44
Merge pull request #660 from jerilynzheng/feat/vercel-models-update-2
...
vercel: add new models from Vercel AI Gateway
2026-01-16 12:59:21 -08:00
Aiden Cline
f0192d8759
Update gpt-5.2-codex.toml
2026-01-16 14:52:45 -06:00
jerilynzheng
070fd89c38
vercel: add new models from Vercel AI Gateway
...
- Add bytedance/seed-1.8 (multimodal with reasoning)
- Add openai/gpt-5.2-codex (agentic coding)
- Add recraft/recraft-v2 and recraft-v3 (image generation)
- Add recraft to model family schema
Co-Authored-By: Claude Opus 4.5 <noreply@anthropic.com >
2026-01-16 12:11:47 -08:00
神楽坂·喵
0ae941ee4d
Merge branch 'anomalyco:dev' into dev
2026-01-17 01:03:43 +08:00
KagurazakaNyaa
3a4cc60e04
update siliconflow model list
2026-01-17 01:02:33 +08:00
Aiden Cline
a243d9ba84
Merge pull request #658 from litvix-whale/feat/add-minimax-m2-1
...
feat(provider): add MiniMax M2.1 for DeepInfra
2026-01-16 08:15:39 -08:00
Kyrylo Lytvishko
5be35e44e4
feat(provider): add MiniMax M2.1 for DeepInfra
2026-01-16 18:04:00 +02:00
yinxulai
faa78aa42b
feat: add new Qiniu AI models - Claude 3.5/3.7/4.0/4.1/4.5 series, Gemini 2.0/2.5/3.0 series, GPT-5/5.2, Grok 4/4.1 series, and Kling v2-6
2026-01-16 17:46:45 +08:00
Aiden Cline
433008fef0
fix: more abacus things - fix model ids
2026-01-16 00:18:21 -06:00
franco
bfb6bb315a
feat: add moark provider
2026-01-16 10:26:06 +08:00
Aiden Cline
7f49452691
Merge pull request #653 from dpuyosa/UpdateModel
...
Venice: Update generate script & add new models (sonnet 4.5, gpt 5.2 codex)
2026-01-15 12:57:31 -08:00
Aiden Cline
6b793ad28e
rm raptor mini model
2026-01-15 12:42:06 -06:00
mthezi
53d77d2f2b
chore: update output limits for various models
2026-01-15 18:42:04 +08:00
dpuyosa
a8436a1e8a
Add new model configurations:
...
- Add claude-sonnet-45 model configuration
- Add openai-gpt-52-codex model configuration
2026-01-15 10:18:29 +01:00
dpuyosa
2ef222a882
Updated model configurations:
...
- Changed family from 'llama' to 'hermes' in hermes-3-llama-3.1-405b
- Changed family from 'glm' to 'glmv' in zai-org-glm-4.6v
- Added interleaved reasoning_details field in zai-org-glm-4.6v
2026-01-15 10:16:49 +01:00
dpuyosa
da6e0354da
Updated family inference logic:
...
- Refactored family inference to use ModelFamilyValues and subsequence matching algorithm
2026-01-15 10:14:09 +01:00
Aiden Cline
5aa046c596
fix: abacus provider
2026-01-14 23:58:31 -06:00
hueyexe
f54b8d8c6d
Add gpt 5.2 codex to azure cognitive services
2026-01-15 16:00:15 +11:00
hueyexe
a2d657f75b
Add gpt 5.2 codex to azure
2026-01-15 15:58:42 +11:00
yinxulai
79636dec83
fix: add required date fields and default output limits for Qiniu AI models
2026-01-15 10:49:52 +08:00
yinxulai
f9983aae19
feat: add Qiniu AI model definitions
...
- Add 49 OpenAI-compatible model definitions
- Models filtered from Qiniu API with OpenAI protocol support
- Include models from DeepSeek, Qwen, Kimi, GLM, Doubao, MiniMax, etc.
- No pricing information included (aggregation platform)
2026-01-15 10:38:15 +08:00
Aiden Cline
b9411cb00c
feat: add Qiniu AI provider configuration
2026-01-15 10:06:30 +08:00
Aiden Cline
5a329d79bc
Merge pull request #650 from cgilly2fast/cgilly2fast/firmware-provider
...
refactor: simplify model ids so sub agents work
2026-01-14 15:18:26 -08:00
Colby Gilbert
1e9ee75804
refactor: simplify model ids so sub agents work
2026-01-14 15:07:36 -08:00
Aiden Cline
64e82beb55
Merge pull request #645 from TheEpTic/dev
...
chore: Add gpt-5.2-codex to GitHub Copilot provider
2026-01-14 14:59:58 -08:00
Frank
256bab07a3
update zen models
2026-01-14 16:27:53 -05:00
Frank
78fd2e0fa0
update zen models
2026-01-14 16:18:51 -05:00
Aiden Cline
969430c25e
Merge pull request #647 from KonarkRajMisra/dev
...
Add GPT-5.2-Codex to OpenRouter
2026-01-14 12:39:08 -08:00
Aiden Cline
c4b43c090d
Merge pull request #646 from brandon93s/52-input
...
chore(openai): gpt-5.2-codex input limit
2026-01-14 12:38:52 -08:00
Konark Misra
bfb92b5e46
Add GPT-5.2-Codex to OpenRouter
2026-01-14 12:15:39 -08:00
TheEpTic
b0e5b914c8
Fix context size
2026-01-14 20:03:52 +00:00
Brandon Smith
66e5d76e05
input
2026-01-14 13:53:57 -06:00
TheEpTic
f1f27989d8
Add gpt-5.2-codex to GitHub Copilot provider
2026-01-14 19:41:09 +00:00
Aiden Cline
949f9b9909
Merge pull request #623 from cyhhao/add-gpt-5-2-codex
...
feat: add gpt-5.2-codex model
2026-01-14 11:25:43 -08:00
Aiden Cline
6a614ab0ac
Update model family name in gpt-5.2-codex.toml
2026-01-14 13:24:47 -06:00
Aiden Cline
664079661d
Merge pull request #641 from liyishuai/iflow-cleanup
...
chore(iflowcn): cleanup models
2026-01-14 07:46:53 -08:00
Aiden Cline
58e2fd8462
Merge pull request #642 from brandon93s/openai-codex-input-limit
...
openai: codex input context limit
2026-01-14 07:31:37 -08:00
Aiden Cline
25eda4cc82
Merge pull request #612 from Alex-wuhu/dev
...
add LLM Provider : novita ai
2026-01-14 07:30:55 -08:00
Alex-wuhu
c60ec95e75
Update model family names for consistency and clarity
2026-01-14 23:04:51 +08:00
Alex
952de0d081
Merge branch 'anomalyco:dev' into dev
2026-01-14 23:00:39 +08:00
Brandon Smith
453f16ce42
add input limit for codex models
2026-01-14 08:32:48 -06:00
Alex-wuhu
01f338231e
Update LLM info
2026-01-14 19:10:34 +08:00
Yishuai Li
ce48f4ee7b
chore(iflowcn): cleanup models
...
Signed-off-by: Yishuai Li <yishuai.li@pingcap.com >
2026-01-14 16:53:58 +08:00
Aiden Cline
db79e08e38
Merge pull request #636 from Eric-Guo/patch-1
...
Using CN in API key, so it won't loading both siliconflow-cn and siliconflow
2026-01-13 21:36:47 -08:00
Aiden Cline
71cf624135
Merge pull request #640 from fanweixiao/dev
...
feat(provider): Add configuration for GPT-5.1 Codex Max model to Vivgrid provider
2026-01-13 21:36:36 -08:00
Aiden Cline
9f7c0cec79
Merge pull request #639 from anomalyco/update-model-families
...
Update model families
2026-01-13 21:36:21 -08:00
C.C.
399b469927
Add configuration for GPT-5.1 Codex Max model
2026-01-14 02:55:08 +00:00
Aiden Cline
bcb7182f67
tweak
2026-01-13 20:07:18 -06:00
Aiden Cline
1008f394ee
wip
2026-01-13 18:33:02 -06:00
Aiden Cline
1a96ad9764
Merge pull request #637 from dpuyosa/UpdateModel
...
Venice: Updated llama-3.2-3b model configuration
2026-01-13 15:00:59 -08:00
Jérôme Benoit
0ada0ed52e
fix(sap-ai-core): use temporary fork for stable OpenCode integration
2026-01-13 19:52:14 +01:00
dpuyosa
4c39b53744
Updated llama-3.2-3b model configuration:
...
- Removed structured_output property
2026-01-13 13:52:50 +01:00
Eric Guo
d84aff0e75
Using CN in API key, so it won't loading both siliconflow-cn and siliconflow
2026-01-13 20:16:29 +08:00
mthezi
c9aefb0af1
feat: add 302ai provider
2026-01-13 14:21:52 +08:00
Aiden Cline
6e6d31f803
Merge pull request #634 from serithemage/feat/upstage-solar-pro3
...
upstage: add solar-pro3 model
2026-01-12 16:54:34 -08:00
Dohyun Jung
ad0e98f745
upstage: add solar-pro3 model
...
Add Solar Pro 3 model with 128K context window.
Pricing is estimated based on Solar Pro 2 (official pricing not yet published).
Co-Authored-By: Claude Opus 4.5 <noreply@anthropic.com >
2026-01-13 09:04:13 +09:00
Aiden Cline
c76336dcb5
Merge pull request #631 from msanft/msanft/ci/fix
...
ci: fix deploy workflow
2026-01-12 14:45:32 -08:00
Aiden Cline
a752e16754
Merge pull request #633 from serithemage/fix/upstage-api-url
...
upstage: fix API base URL
2026-01-12 14:44:35 -08:00
Dohyun Jung
e5dea00090
upstage: fix API base URL
...
Change API URL from https://api.upstage.ai to https://api.upstage.ai/v1/solar
to match the correct endpoint for OpenAI-compatible API access.
Reference: https://console.upstage.ai/docs/models/solar-pro-2
Co-Authored-By: Claude Opus 4.5 <noreply@anthropic.com >
2026-01-13 07:00:44 +09:00
Moritz Sanft
af10061a5c
ci: fix deploy workflow
2026-01-12 10:17:33 +01:00
Aiden Cline
5204232521
Merge pull request #620 from matthusby/dev
...
Update all chutes models with data for the api, and fix formatting on bunch of them.
2026-01-11 11:17:49 -08:00
Aiden Cline
d180bd1f8b
Merge pull request #619 from cgilly2fast/cgilly2fast/firmware-provider
...
feat: add firmware provider models
2026-01-11 11:16:53 -08:00
Aiden Cline
3cf7f9f3c2
Merge pull request #626 from davidcharbonnier/dev
...
Add Qwen3 Coder 30B A3B Instruct to Openrouter
2026-01-11 11:15:09 -08:00
Aiden Cline
5771ec68d0
Merge pull request #627 from jerilynzheng/feat/vercel-models-update
...
feat: add new models from Vercel AI Gateway
2026-01-10 22:30:56 -08:00
jerilynzheng
bbbe1c10dd
vercel: add family field to new models
...
Add family field to 99 new models following existing provider patterns:
- OpenAI: gpt-5, gpt-5.1, gpt-5.2, gpt-oss, o3, text-embedding, codex
- Google: gemini-flash, gemini-pro, gemini-embedding, imagen-4
- Anthropic: claude-sonnet
- xAI: grok
- Meta: llama-3.1, llama-3.2
- Mistral: devstral-small, devstral-medium, ministral, mistral-large
- DeepSeek: deepseek-v3
- Alibaba: qwen-max, qwen-coder, qwen-embedding, qwen3-*
- Others: kimi-k2, minimax-m2.1, glm-4.x, voyage, flux, etc.
Co-Authored-By: Claude Opus 4.5 <noreply@anthropic.com >
2026-01-10 22:18:14 -08:00
jerilynzheng
e69e3fe493
vercel: add new models from Vercel AI Gateway
...
- Add 105+ new models from Vercel AI Gateway API
- Update pricing and limits synced from API
- Remove deprecated models (grok-2, mistral-large, etc.)
- Rename claude-4.5-sonnet -> claude-sonnet-4.5 to match API
New models include:
- GPT-5.x series (gpt-5, gpt-5.1, gpt-5.2, codex variants)
- Gemini 2.5/3.x with image generation support
- Grok 4.x series
- GLM 4.5-4.7 series
- Llama 3.x/4.x series
- DeepSeek v3.x series
- Qwen3 series
- Various embedding models (voyage, text-embedding)
- Image generation (Flux, Imagen 4.0)
Co-Authored-By: Claude Opus 4.5 <noreply@anthropic.com >
2026-01-10 22:07:24 -08:00
David Charbonnier
d467cd4a1e
feat: add qwen3 coder 30b a3b instruct model to openrouter provider
2026-01-10 18:27:02 -05:00
Aiden Cline
b86678775c
Merge pull request #622 from shelvick/fix-opus-4.5-cache-pricing
...
Fix Claude Opus 4.5 cache pricing (3x too high)
2026-01-10 14:34:20 -08:00
Aiden Cline
650ece42a3
Merge pull request #625 from shelvick/fix-azure-deepseek-pricing
...
Fix Azure DeepSeek V3.2 pricing
2026-01-10 14:33:59 -08:00
Scott Helvick
188a869ea9
Fix Azure DeepSeek V3.2 pricing
...
Corrected pricing to match Azure AI Foundry rates:
- Input: $0.28 → $0.58 per 1M tokens
- Output: $0.42 → $1.68 per 1M tokens
- Removed cache_read (Azure doesn't offer prompt caching for third-party models)
Fixes #624
2026-01-10 22:04:34 +00:00
cyhhao
94310d742d
Add gpt-5.2-codex model
2026-01-11 01:55:50 +08:00
Scott Helvick
ab8d9dc081
Fix Claude Opus 4.5 cache pricing (3x too high)
...
Anthropic reduced Opus 4.5 cache pricing. Updated:
- Amazon Bedrock (regional and global)
- Azure
- Helicone (also fixed floating point precision)
cache_read: 1.50 → 0.50
cache_write: 18.75 → 6.25
2026-01-10 17:02:07 +00:00
Matt Husby
74fb94a6d1
Update all chutes models with data for the api, and fix formatting for a bunch of them.
2026-01-09 21:23:32 -05:00
Colby Gilbert
930a70dbac
update firmware logo
2026-01-09 10:50:23 -08:00
Aiden Cline
0480d3cd23
Merge pull request #601 from xinrui-z/aihubmix-free-model
...
aihubmix: add free models
2026-01-09 09:53:21 -08:00
Aiden Cline
0a7cab6773
Merge pull request #602 from msanft/msanft/privatemode-ai
...
Add privatemode.ai provider
2026-01-09 09:52:40 -08:00
Colby Gilbert
b232abe303
add firmware provider models
2026-01-09 09:15:58 -08:00
Alex-wuhu
a50c04d060
Update minimax-m2.1.toml
2026-01-09 13:31:49 +08:00
Aiden Cline
21bd51da8f
bump sst version
2026-01-08 23:20:19 -06:00
Aiden Cline
4f9aea44a6
Merge pull request #606 from qychen2001/add/siliconflow-models-2025-01-06
...
Add 3 new SiliconFlow models (GLM-4.7, GLM-4.6V, DeepSeek-V3.2)
2026-01-08 19:37:49 -08:00
Aiden Cline
524fd462fa
Delete providers/siliconflow-cn/models/zai-org/GLM-4.7.toml
2026-01-08 21:36:53 -06:00
Aiden Cline
43d7b9b558
Delete providers/siliconflow-cn/models/zai-org/GLM-4.6V.toml
2026-01-08 21:36:38 -06:00
Aiden Cline
8ac502e533
Delete providers/siliconflow-cn/models/deepseek-ai/DeepSeek-V3.2.toml
2026-01-08 21:36:20 -06:00
opencode-agent[bot]
b51d257ee2
Added 3 SiliconFlow CN models via symlinks
...
Co-authored-by: rekram1-node <rekram1-node@users.noreply.github.com >
2026-01-09 03:31:46 +00:00
Aiden Cline
1c23f5bb42
Merge pull request #610 from fanweixiao/dev
...
feat(provider): add vivgrid provider
2026-01-08 19:29:04 -08:00
Aiden Cline
9e6a1a7e18
Merge pull request #615 from friendliai/update-freindli-model-list-250108
...
Update the list of friendli provider models
2026-01-08 19:28:20 -08:00
Aiden Cline
457b824af9
Merge pull request #618 from dpuyosa/UpdateModel
...
Venice: Add cache_write pricing support and update model configurations:
2026-01-08 16:07:16 -08:00
dpuyosa
4604b25070
Add cache_write pricing support and update model configurations:
...
- Updated generate-venice.ts to handle cache_write
- Updated claude-opus-45.toml with cache_write pricing
- Removed deprecated zai-org-glm-4.6.toml model
2026-01-09 00:57:18 +01:00
Aiden Cline
3445f7f2b8
Merge pull request #616 from vglafirov/dev
...
feat: added GitLab Duo Agentic models
2026-01-08 15:44:32 -08:00
Vladimir Glafirov
ab9f44ca5f
Updated gitlab logo
2026-01-08 20:11:01 +01:00
Aaron Iker
e4bf0b52e4
Merge pull request #617 from anomalyco/provider-logo-adjustments
...
feat: Small provider logo adjustments
2026-01-08 12:40:10 +01:00
Aaron Iker
8acc313424
fix: friendli logo size
2026-01-08 12:36:37 +01:00
Aaron Iker
fe52c00b19
feat: abacus logo adjustment
2026-01-08 12:36:20 +01:00
Vladimir Glafirov
ee5767efaa
feat: added GitLab Duo Agentic models
2026-01-08 09:38:48 +01:00
minpeter
16389046d1
Add K EXAONE 236B A23B model configuration
...
The model configuration has been added to the provider's models
directory. The file includes necessary metadata such as name, family,
supported features, release date, cost, limits, and modalities.
2026-01-08 13:01:24 +09:00
minpeter
430c89cf0f
Remove DeepSeek R1 0528 configuration
...
Deleted the provider configuration file for DeepSeek R1 0528 as it has
been deprecated or is no longer supported.
2026-01-08 13:01:19 +09:00
C.C.
1ad46ad294
fix: vivgrid logo size and color
2026-01-08 09:54:37 +08:00
Aiden Cline
caf7fc09a8
Merge pull request #614 from gary149/feat/huggingface-model-updates
...
feat(huggingface): add 5 new models, remove 5 deprecated
2026-01-07 13:06:15 -08:00
Victor Muštar
7cf962bea9
feat(huggingface): add 5 new models, remove 5 deprecated
2026-01-07 22:00:18 +01:00
Aiden Cline
012ace7de6
Merge pull request #608 from Algowary/dev
...
Chutes Models Update
2026-01-07 08:54:50 -08:00
Aiden Cline
224e4c0a59
Merge pull request #611 from dpuyosa/UpdateModel
...
Venice: Updated GLM 4.7 model configuration
2026-01-07 08:54:04 -08:00
Aiden Cline
49afb24047
Merge pull request #613 from scwgoire/scw-devstral2
...
feat(scaleway): add devstral 2 123B to Scaleway catalog
2026-01-07 08:53:33 -08:00
Gregoire de Turckheim
909d0f63a1
feat(scaleway): add devstral 2 123B to Scaleway catalog
2026-01-07 16:07:12 +01:00
Alex-wuhu
cc2619dd5f
add LLM Provider : novita ai
2026-01-07 19:27:42 +08:00
Xinrui
38ffcec5d0
AIHubMix: Update model
2026-01-07 17:35:58 +08:00
Xinrui
788c911a40
AIHubMix: Update model
2026-01-07 17:35:37 +08:00
dpuyosa
e195692307
Updated GLM 4.7 model configuration:
...
- Enabled reasoning capability
- Reduced input cost from 0.85 to 0.55
- Reduced output cost from 2.75 to 2.65
- Increased context limit from 131_072 to 202_752
- Increased output limit from 32_768 to 50_688
- Added interleaved reasoning_details field
2026-01-07 09:55:11 +01:00
C.C.
003d5ea41d
feat(provider): add vivgrid provider
2026-01-07 08:42:55 +08:00
Aiden Cline
33bb01c66e
Merge pull request #609 from sebastiand-cerebras/adding_new_glm_47_model
...
feat(cerebras): add zai-glm-4.7 model
2026-01-06 16:35:34 -08:00
Seb Duerr
84366efe77
revert: remove interleaved flag
...
Remove the temporary interleaved field from the model schema and the Cerebras zai-glm-4.7 definition.
2026-01-06 16:22:24 -08:00
Seb Duerr
a70ca6a4ce
Merge branch 'dev' into adding_new_glm_47_model
2026-01-06 18:10:41 -06:00
Seb Duerr
8b0a6ce497
feat(schema): add model interleaved flag
...
- Add optional field to model schema\n- Set for Cerebras zai-glm-4.7
2026-01-06 16:08:15 -08:00
Seb Duerr
391197e8b9
feat(cerebras): add zai-glm-4.7 model
...
Adds a models.dev definition for Z.ai GLM 4.7 under the Cerebras provider.
2026-01-06 15:55:24 -08:00
Algowarry
a52699b069
Chutes Models Update
...
Updated the models attributed to the Chutes.ai provider with accurate info derived from the API.
2026-01-06 14:51:33 -05:00
Burak Varlı
f14775e355
Add cross-region inference profiles for Claude 4.x family models in Amazon Bedrock
...
Amazon Bedrock requires usage of cross-region inference for some models, especially the latest models including all Claude 4.x family.
This change creates model files for all Claude 4.x models for cross-region inference profiles for Global, US and EU.
2026-01-06 11:41:49 +00:00
Xinrui
00d7bf8b2b
add model
2026-01-06 19:05:33 +08:00
QiyuanChen
e150d03dc5
Add 3 new SiliconFlow models (GLM-4.7, GLM-4.6V, DeepSeek-V3.2)
2026-01-06 17:05:53 +08:00
Aiden Cline
7b984aaeec
Merge pull request #605 from anomalyco/opencode/issue604-20260106043944
...
Made api optional for @ai-sdk/openai
2026-01-05 20:56:56 -08:00
opencode-agent[bot]
40bcce25de
Made api optional for @ai-sdk/openai
...
Co-authored-by: rekram1-node <rekram1-node@users.noreply.github.com >
2026-01-06 04:41:52 +00:00
Moritz Sanft
cded6124ac
Add privatemode.ai provider
2026-01-05 11:25:08 +01:00
Xinrui
b0b0dac51b
aihubmix: add free models
2026-01-05 17:26:31 +08:00
Xinrui
b34b1f418c
aihubmix: add free models
2026-01-05 17:25:03 +08:00
Aiden Cline
840fe7fef6
Merge pull request #600 from xiaojiezj/zenmux_dev
...
feat: Update the model configuration file based on ZenMux’s model list
2026-01-04 23:27:39 -08:00
xiaojie.zj
8f070f49b3
fix: fix validate
2026-01-05 15:23:24 +08:00
Aiden Cline
fa61b458a0
Merge pull request #599 from billycao/billy/remove-Kimi-K2-Instruct
...
fix: Remove Kimi-K2-Instruct for provider Synthetic
2026-01-04 22:49:06 -08:00
xiaojie.zj
6ac98b8a8f
feat: Update the model configuration file based on ZenMux’s model list
2026-01-05 14:18:23 +08:00
Billy Cao
b643aa45b0
Remove Kimi-K2-Instruct for provider Synthetic
2026-01-04 20:06:15 -08:00
Aiden Cline
65b43b44c1
Merge pull request #598 from ishaksebsib/feat/groq-provider
...
Feat: update Groq provider with latest pricing, limits, and capabilities
2026-01-04 11:16:54 -08:00
ishaksebsib
7a7fe98987
groq: update structured output capability for models that support it
2026-01-04 20:48:46 +03:00
ishaksebsib
f94e0257c2
groq: update input and output cost/limit
2026-01-04 20:42:33 +03:00
ishaksebsib
0df9997e71
groq: update status for deprecated models
2026-01-04 20:41:21 +03:00
Aiden Cline
d1710b2b08
Merge pull request #597 from xinrui-z/aihubmix-add-minimax-m2.1-and-glm-4.7
...
aihubmix: add glm-4.7 and minimax-m2.1
2026-01-03 23:15:14 -08:00
Xinrui
b568ddd33d
aihubmix: add glm-4.7 and minimax-m2.1
2026-01-04 14:50:47 +08:00
Aiden Cline
4b96508663
Merge pull request #595 from dpuyosa/UpdateModel
...
Venice: Fixed modalities for grok-code-fast-1 and minimax-m21
2026-01-02 09:56:24 -08:00
Aiden Cline
4b02644ff9
Merge pull request #486 from xinrui-z/update/aihubmix
...
aihubmix: add models
2026-01-02 09:55:55 -08:00
Xinrui
c35f968adc
aihubmix: add models
2026-01-02 21:48:49 +08:00
dpuyosa
75957fb866
Update model configurations for grok-code-fast-1 and minimax-m21:
...
- Set attachment to false for both models
- Update last_updated to 2026-01-02 for both models
- Remove image input modality from grok-code-fast-1
- Remove image input modality from minimax-m21
- Add interleaved reasoning_content field to minimax-m21
- Add family field to minimax-m21
2026-01-02 02:59:42 +01:00
Aiden Cline
07da33848a
Merge pull request #582 from mark182es/fix/update-chutes-models-20251229
...
feat: add new Chutes TEE providers and fix model configuration fields
2026-01-01 16:05:23 -08:00
Aiden Cline
5d2d213aa4
Merge pull request #592 from janspoerer/provider-abacus
...
Added Abacus as a provider
2026-01-01 11:57:35 -08:00
Jan Spoerer
46728af37c
Transformed the Abacus logo into a matching format, color, size to the other logos
2026-01-01 20:52:26 +01:00
Jan Spoerer
8ce390ce26
Added Abacus svg
2026-01-01 20:48:26 +01:00
Aiden Cline
53a61d9e31
Merge pull request #593 from fhennerkes/dev
...
Poe: update 1/1/26
2026-01-01 10:46:49 -08:00
fhennerkes
4be86ce87f
Merge poe-pricing-sync into dev (selective)
...
Added new models:
- Cerebras: gpt-oss-120b-cs, zai-glm-4.6-cs
- Google: gemini-3-flash
- Novita: glm-4.6v, glm-4.7, kat-coder-pro, minimax-m2.1
- OpenAI: gpt-image-1.5
Updated pricing and configuration:
- Google Gemini 2.5 Flash/Pro cache pricing corrections
- Google Nano Banana models pricing updates
- xAI Grok models: added image input support
2026-01-01 10:34:26 -08:00
github-actions
989e70553e
chore: sync Poe pricing
2026-01-01 18:01:20 +00:00
fhennerkes
1b78419770
chore: restore Poe pricing sync automation
2026-01-01 09:59:25 -08:00
Jan Spoerer
3a7a4e9654
Added Abacus as a provider
2025-12-31 15:14:36 +01:00
Aiden Cline
7ea8fba795
Merge pull request #589 from dpuyosa/UpdateModel
...
Venice: Added new models Grok Code Fast 1 and Minimax M2.1
2025-12-30 14:40:21 -08:00
dpuyosa
3a61532edb
Updated model configurations and added new models:
...
- gemini-3-flash-preview: updated last_updated and added cache_read cost
- grok-code-fast-1: added new model configuration
- kimi-k2-thinking: updated release_date, last_updated, and cache_read cost
- minimax-m21: added new model configuration
2025-12-30 22:06:38 +01:00
Aiden Cline
f4069f92d3
Merge pull request #588 from jerome-benoit/fix/sap-ai-core-claude-haiku-4.5
...
fix(sap-ai-core): rename anthropic--claude-haiku-4.5 to anthropic--cl…
2025-12-30 11:34:55 -08:00
Jérôme Benoit
b9588df9ed
fix(sap-ai-core): rename anthropic--claude-haiku-4.5 to anthropic--claude-4.5-haiku
...
Signed-off-by: Jérôme Benoit <jerome.benoit@piment-noir.org >
2025-12-30 20:03:13 +01:00
Aiden Cline
a9aaa0f9ae
Merge pull request #587 from cravenceiling/refactor/siliconflow-model-ids
...
refactor siliconflow model ids
2025-12-30 10:31:34 -08:00
Aiden Cline
f5428a81a8
fix: some model dates
2025-12-30 12:23:25 -06:00
cravenceiling
0a55c3d1cd
refactor siliconflow model ids
...
* Change the model file structure to use foldes for ids containing `/`
* Update the models and file structure in the `providers/siliconflow-cn` directory
2025-12-30 09:57:37 -05:00
Aiden Cline
59ce57823c
Merge pull request #586 from mounta11n/patch-1
...
Fix typo from 4B to 8B
2025-12-29 20:03:53 -08:00
Yazan Agha-Schrader
eb85004c3e
Fix typo from 4B to 8B
2025-12-30 04:31:27 +01:00
Aiden Cline
53afc6aefb
Merge pull request #584 from jerome-benoit/feat/sap-ai-core-claude-updates
...
feat(sap-ai-core): add Claude Haiku 4.5 and cache pricing for Claude …
2025-12-29 14:57:14 -08:00
Frank
54d0d65ed5
update zen models
2025-12-29 16:56:56 -05:00
Aiden Cline
eaa25b5268
Merge pull request #579 from wojons/dev
...
Modify cost parameters in MiniMax-M2.1.toml
2025-12-29 13:17:06 -08:00
Aiden Cline
001367833d
Merge pull request #583 from dpuyosa/UpdateModel
...
Venice: Update/fix 'release_date' for many models
2025-12-29 13:16:40 -08:00
Jérôme Benoit
a24b564ee7
feat(sap-ai-core): add Claude Haiku 4.5 and cache pricing for Claude models
2025-12-29 22:01:37 +01:00
dpuyosa
d5d1b36cb8
Update/fix 'release_date' for many models
2025-12-29 20:26:17 +01:00
Marco
b1d82a1946
fix: add missing fields to all chutes models
...
Add structured_output field
2025-12-29 18:14:15 +01:00
Marco
b19dceffdd
feat: add new TEE providers, update context sizes and pricing
...
New TEE providers:
- MiniMaxAI: M2.1-TEE
- NousResearch: Hermes-4-405B-FP8-TEE
- Qwen: Qwen2.5-VL-72B-Instruct-TEE, Qwen3-235B-A22B-Instruct-2507-TEE, Qwen3-Coder-480B-A35B-Instruct-FP8-TEE
- deepseek-ai: DeepSeek-R1-0528-TEE, DeepSeek-R1-TEE, DeepSeek-V3-0324-TEE, DeepSeek-V3.1-TEE, DeepSeek-V3.1-Terminus-TEE, DeepSeek-V3.2-TEE
- moonshotai: Kimi-K2-Thinking-TEE
- openai: GPT-OSS-120B-TEE
- zai-org: GLM-4.5-TEE, GLM-4.7-TEE
Updates:
- Fixed context window sizes for multiple models
- Updated pricing for all affected providers
- Added NVIDIA Nemotron 3 Nano 30B model
2025-12-29 17:48:04 +01:00
Frank
057361ad5e
update zen models
2025-12-29 10:23:30 -05:00
Aiden Cline
d2939fa5ad
rm perplexity deprecated model
2025-12-28 17:57:11 -06:00
Alexis Okuwa
8952013668
Add interleaved section with reasoning_content field
2025-12-28 15:09:31 -08:00
Alexis Okuwa
e033fc186d
Modify cost parameters in MiniMax-M2.1.toml
...
Updated cost parameters for input and output.
2025-12-27 18:47:30 -08:00
Aiden Cline
9250fbe2bc
Merge pull request #567 from b3nw/feat/add-nano-gpt-models
...
Feat: Add Nano-GPT provider models
2025-12-26 22:49:12 -08:00
Ben
e6ab0814e2
Feat: Add Nano-GPT provider models
2025-12-27 05:25:59 +00:00
Aiden Cline
4bd8337131
Merge pull request #574 from friendliai/feat/add-friendli-provider
...
fix: correct friendli model file structure for API IDs with slashes
2025-12-26 20:54:00 -08:00
Aiden Cline
cb2c762dec
Merge pull request #575 from otterDeveloper/fireworks-pull-2
...
add Firework's MiniMax-M2.1
2025-12-26 20:53:40 -08:00
Miguel Medina
8bf3ff5fca
feat: add minimax 2.1
...
https://app.fireworks.ai/models/fireworks/minimax-m2p1
2025-12-26 22:34:05 -06:00
Miguel Medina
dfb7e0087a
fix: update context and pricing
...
obtained from https://app.fireworks.ai/models/fireworks/minimax-m2
2025-12-26 22:27:20 -06:00
minpeter
7a55ce44e7
fix: restructure friendli model files to match API IDs with slashes
...
- Change model file structure to use directories for IDs containing '/'
- Update generate-friendli.ts to create directory structure instead of replacing '/' with '-'
- Fixes 404 errors caused by model ID mismatch (e.g., Qwen/Qwen3-30B-A3B vs Qwen-Qwen3-30B-A3B)
2025-12-26 04:19:24 +09:00
Aiden Cline
006d0208f0
Merge pull request #558 from friendliai/feat/add-friendli-provider
...
Add Friendli serverless endpoints provider
2025-12-24 22:23:33 -08:00
Frank
7ac941d483
update zen mdoels
2025-12-24 14:34:23 -05:00
Aiden Cline
31ec6424b2
Merge pull request #570 from dpuyosa/UpdateModel
...
Venice: Update zai-org-glm-4.7
2025-12-24 09:15:49 -08:00
dpuyosa
0279094eff
Update glm 4.7 data
2025-12-24 17:41:52 +01:00
Aiden Cline
889dabce96
Merge pull request #569 from b3nw/feat/update-nvidia-models
...
Feat/update nvidia models
2025-12-24 07:33:12 -08:00
Aiden Cline
b68935b403
Merge pull request #566 from otterDeveloper/fireworks-pull-1
...
Add recent fireworks models
2025-12-24 07:32:55 -08:00
Aiden Cline
1ba3c3cc4a
Merge pull request #568 from M16X/deepinfra-glm-4.7
...
Rename glm-4.7.toml to GLM-4.7.toml
2025-12-24 07:32:23 -08:00
b3nw
31feccceda
Merge branch 'sst:dev' into feat/update-nvidia-models
2025-12-24 08:11:26 -06:00
Ben
ea4063510f
feat(nvidia): model update
2025-12-24 14:10:40 +00:00
minpeter
e5c683d830
Update Friendli logo to use currentColor in SVG
2025-12-24 15:07:05 +09:00
Nazar
07bbf78e9c
Rename glm-4.7.toml to GLM-4.7.toml
2025-12-24 11:32:41 +05:30
Miguel Medina
bc6981debc
fix: increase output token limit
...
16_384 seems to be the ui limit
2025-12-23 23:01:31 -06:00
Miguel Medina
4eb339327e
fix: document interleaved thinking
2025-12-23 22:58:13 -06:00
Aiden Cline
ef30eb2bca
Merge pull request #565 from M16X/deepinfra-glm-4.7
...
Add GLM-4.7 for DeepInfra
2025-12-23 20:26:16 -08:00
Nazar
9aa2515fd5
deepinfra: add interleaved reasoning for glm-4.7
2025-12-24 09:45:44 +05:30
Aiden Cline
02420741cf
Merge pull request #563 from dpuyosa/UpdateModel
...
Venice: Add cost.cache_read to kimi-k2-thinking
2025-12-23 20:11:57 -08:00
Aiden Cline
df4fa098b2
Merge pull request #564 from superhighfives/cgleason/fix-cloudflare-workers-pricing
...
Adds missing pricing information for Workers in Cloudflare AI Gateway
2025-12-23 20:11:47 -08:00
Miguel Medina
e5069994ba
feat: document fireworks.ai models
2025-12-23 22:03:53 -06:00
Nazar
067df451a5
[deepinfra] glm-4.7: update knowledge cut off time
2025-12-24 08:31:27 +05:30
Nazar
3c9f132359
deepinfra: remove cache_write for glm-4.7
2025-12-24 08:29:00 +05:30
Nazar
6fd63482ae
deepinfra: add docs on output limit
2025-12-24 08:24:54 +05:30
Nazar
c2a81337f5
deepinfra: deprecate glm-4.5
2025-12-24 08:20:09 +05:30
Nazar
450053341e
deepinfra: add glm-4.7
2025-12-24 08:12:50 +05:30
Charlie Gleason
1c4e1ec6b6
Adds missing pricing information
2025-12-23 18:20:48 -08:00
dpuyosa
e4a102b68f
Add cost.cache_read to kimi-k2-thinking
2025-12-24 03:15:58 +01:00
Aiden Cline
336e4583ba
fix: filename
2025-12-23 18:54:03 -06:00
Frank
2abd51895b
update zen models
2025-12-23 19:18:53 -05:00
Aiden Cline
3a9d8af6a5
Merge pull request #562 from dsingal0/dev
...
add GLM 4.7 for baseten provider
2025-12-23 15:34:45 -08:00
Dhruv Singal
a66d418cfa
Rename glm-4.7.toml to GLM-4.7.toml
2025-12-23 14:48:38 -08:00
Dhruv Singal
cc6079297c
add glm 4.7
2025-12-23 14:48:06 -08:00
Aiden Cline
09d5d80a8a
Merge pull request #561 from KevinPoorDeveloper/dev
...
Add GLM 4.7 for Provider Venice.ai
2025-12-23 14:17:14 -08:00
Kevin
cb4a5843c5
Add GLM 4.7 for Provider Venice.ai
...
New model toml
2025-12-23 13:41:02 -08:00
Aiden Cline
d54bc052eb
Merge pull request #560 from superhighfives/cgleason/fix-workers-model-ids
...
Fix Workers AI models in Cloudflare AI Gateway
2025-12-23 12:17:22 -08:00
Aiden Cline
e174f5400f
fix: properly set interleaved setting for glm 4.7
2025-12-23 14:09:20 -06:00
Charlie Gleason
821f350e9a
Add model families
2025-12-23 10:58:08 -08:00
Aiden Cline
663b10c8b3
Merge pull request #557 from no1wudi/mini
...
feat: add MiniMax-M2.1 model configuration
2025-12-23 10:43:10 -08:00
Charlie Gleason
f6ff799d21
Update models
2025-12-23 10:22:16 -08:00
Charlie Gleason
a15c161d2d
Fix Workers AI model references and model names
2025-12-23 10:18:43 -08:00
minpeter
6a97400dd7
Update reasoning and open_weights for Friendli models
2025-12-23 19:46:11 +09:00
minpeter
506ab5646c
Add Friendli provider with 11 models
2025-12-23 19:17:00 +09:00
minpeter
0c4b20a7ad
Refactor API key argument parsing to remove redundant null checks
2025-12-23 19:15:14 +09:00
minpeter
e47641da01
Fix TypeScript type errors in generator scripts
...
- Fix possible undefined array access in generate-friendli.ts
- Fix undefined type assignments in generate-venice.ts
- Use safe array access with .at() and nullish coalescing operators
🤖 Generated with [Claude Code](https://claude.com/claude-code )
Co-Authored-By: Claude Sonnet 4.5 <noreply@anthropic.com >
2025-12-23 19:06:40 +09:00
minpeter
6d6743331c
Add Friendli serverless endpoints provider
...
- Add provider configuration for Friendli serverless endpoints
- Implement auto-generation script (generate-friendli.ts)
- Add 11 models: Llama, Qwen, DeepSeek-R1, EXAONE, GLM
- Handle TOKEN-based pricing (4 models) and SECOND-based pricing (7 models)
- Auto-infer model families and open_weights status
🤖 Generated with [Claude Code](https://claude.com/claude-code )
Co-Authored-By: Claude Sonnet 4.5 <noreply@anthropic.com >
2025-12-23 19:02:25 +09:00
Huang Qi
88471de736
feat: add MiniMax-M2.1 model configuration
...
Add new MiniMax-M2.1 reasoning model to both global and Chinese providers
* Supports reasoning capabilities, temperature settings, and tool calls
* Includes open weights support with context limit of 196,608 tokens
* Pricing: $0.30/input and $1.20/output per million tokens
2025-12-23 14:44:41 +08:00
Aiden Cline
a090e9a69a
Merge pull request #556 from InduwaraSMPN/dev-copy
...
Adds configuration for Openrouter/MiniMax M2.1 model
2025-12-22 22:10:19 -08:00
InduwaraSMPN
63b1f5e6a9
Adds configuration for MiniMax M2.1 model
...
Introduces support for MiniMax M2.1 with detailed model parameters,
cost estimates, and modality specifications to enable integration and
usage with the provider ecosystem.
2025-12-23 11:37:45 +05:30
Aiden Cline
0abcb5b3cf
Merge pull request #554 from dpuyosa/VeniceUpdate
...
Venice: Update Autogenerate Script
2025-12-22 20:49:09 -08:00
Aiden Cline
f0ae336593
Merge pull request #555 from no1wudi/glm
...
fix: set costs to 0 for zai-coding-plan provider
2025-12-22 20:48:41 -08:00
Huang Qi
f4b3e08d2a
fix: set costs to 0 for zai-coding-plan provider
...
Updated cost configuration for GLM-4.7 model to reflect subscription-based
billing rather than token-based pricing. Since zai-coding-plan uses a
subscription model, all per-token costs should be zero.
- Changed input cost from 0.6 to 0
- Changed output cost from 2.2 to 0
- Changed cache_read cost from 0.11 to 0
2025-12-23 12:38:18 +08:00
Frank
87ccab8807
update zen models
2025-12-22 19:43:48 -05:00
dpuyosa
e571d678ac
Update cost.cache_read for models that support cache
2025-12-23 01:30:24 +01:00
dpuyosa
d6650a05df
Update generate-venice.ts to include new cache feature (cache_read cost)
2025-12-23 01:28:41 +01:00
Frank
4be4bea336
update zen models
2025-12-22 19:11:48 -05:00
Aiden Cline
840be0e2b8
Merge pull request #551 from reissbaker/glm-4.7
...
Add Synthetic's GLM-4.7 hosting
2025-12-22 15:54:53 -08:00
Matt Baker
484a603366
Add interleaved thinking setting
2025-12-22 15:45:40 -08:00
Frank
b604eabecf
update zen models
2025-12-22 18:13:41 -05:00
Frank
fb4abefa68
update zen models
2025-12-22 17:49:12 -05:00
Aiden Cline
e3e89e4fbf
Merge pull request #552 from titouv/dev
...
Add OpenRouter Z.AI GLM 4.7
2025-12-22 14:33:28 -08:00
Aiden Cline
e7bc32be6b
fix
2025-12-22 16:31:59 -06:00
Aiden Cline
e469476732
Merge pull request #553 from superhighfives/cgleason/fix-model-mapping
...
Fix model mapping for Cloudflare AI Gateway.
2025-12-22 14:24:35 -08:00
Charlie Gleason
0d448bf0a2
Updated Cloudflare AI Gateway models
2025-12-22 14:11:24 -08:00
Titouan V
d408d60f49
openrouter glm 4.7 add interleaved reasoning_content
2025-12-22 21:54:21 +00:00
Titouan V
ea0250433c
feat: add openrouter z.ai glm 4.7
2025-12-22 21:32:24 +00:00
Matt Baker
e22dd2f3de
Add Synthetic's GLM-4.7 hosting
2025-12-22 13:27:22 -08:00
Frank
39e5930b3b
update zen models
2025-12-22 12:02:12 -05:00
Aiden Cline
a6093eafe6
Merge pull request #548 from no1wudi/glm
...
feat: add glm-4.7 model with interleaved thinking
2025-12-22 08:07:41 -08:00
Huang Qi
26778bd4cf
feat: add glm-4.7 model with interleaved thinking
...
Add GLM-4.7 model configuration to zai, zai-coding-plan, zhipuai,
and zhipuai-coding-plan providers. The model features interleaved
reasoning capability with reasoning_content field.
* Created glm-4.7.toml in zai and zai-coding-plan providers
* Added symlinks in zhipuai and zhipuai-coding-plan providers
* Configured with same pricing and limits as glm-4.6
* Supports reasoning, tool_call, and temperature settings
2025-12-22 23:12:48 +08:00
Aiden Cline
970c13f8d7
Merge pull request #546 from dpuyosa/RemoveDeprecated
...
Venice: Remove deprecated model devstral-2-2512
2025-12-21 20:00:38 -08:00
Aaron Iker
4654d1f039
Merge pull request #547 from sst/visually-align-logo-weights
...
fix: Visually align provider logos
2025-12-21 23:14:01 +01:00
Aaron Iker
b2b7f7e99a
fix: align colors, subtle fills
2025-12-21 23:09:03 +01:00
Aaron Iker
9a19c74ef9
fix: remaining provider logos
2025-12-21 22:42:41 +01:00
Aaron Iker
6fafd70000
fix: reorder some provider logos
2025-12-21 22:38:00 +01:00
dpuyosa
1e64e546cc
Remove deprecated model devstral-2-2512
2025-12-21 22:25:02 +01:00
Aaron Iker
d35af033d0
fix: visually align provider logos
2025-12-21 22:08:18 +01:00
Aiden Cline
9264f8bead
deepinfra: add minimax m2 & kimi k2 thinking
2025-12-20 23:39:17 -06:00
Aiden Cline
cf2c8aecd3
Merge pull request #545 from JaviMaligno/oss-agent/issue-528-add-nvidia-nemotron-3-nano
...
fix: Add Nvidia Nemotron 3 Nano
2025-12-20 19:14:43 -08:00
Aiden Cline
469026c193
Delete validation_output.json
2025-12-20 17:30:42 -06:00
Javier
13f09f8b90
fix: Add Nvidia Nemotron 3 Nano ( #528 )
...
Fixes #528
---
Changes prepared with assistance from OSS-Agent
2025-12-21 00:05:05 +01:00
Aiden Cline
620c92a5ed
Merge pull request #544 from dpuyosa/RemoveDeprecated
...
Venice: Remove deprecated model qwen3-235b
2025-12-20 12:35:39 -08:00
Aiden Cline
6b6e733a72
Revert "tweak: update ollama logo, add ollama local"
...
This reverts commit 8ff1ca4747 .
2025-12-20 14:28:03 -06:00
Aiden Cline
be1f2f9bc8
Revert "fix: validation err"
...
This reverts commit 91e7dac265 .
2025-12-20 14:27:59 -06:00
Aiden Cline
bf6f1ac1c9
Revert "fix: env"
...
This reverts commit 4105e28730 .
2025-12-20 14:27:57 -06:00
Aiden Cline
b01ccb1504
Revert "fix: handle empty dir"
...
This reverts commit 62ff421d08 .
2025-12-20 14:27:54 -06:00
Aiden Cline
b07097c62b
Revert "fix: dir check"
...
This reverts commit 4e69c0f724 .
2025-12-20 14:27:53 -06:00
Aiden Cline
4e69c0f724
fix: dir check
2025-12-20 14:01:01 -06:00
Aiden Cline
62ff421d08
fix: handle empty dir
2025-12-20 13:57:27 -06:00
dpuyosa
d9f270fe31
Remove deprecated model qwen3-235b
2025-12-20 20:52:31 +01:00
Aiden Cline
4105e28730
fix: env
2025-12-20 13:03:11 -06:00
Aiden Cline
91e7dac265
fix: validation err
2025-12-20 12:42:39 -06:00
Aiden Cline
8ff1ca4747
tweak: update ollama logo, add ollama local
2025-12-20 12:40:47 -06:00
Aiden Cline
3f4d29af7b
revert venice ai npm change
2025-12-20 12:07:18 -06:00
Aiden Cline
7a9c0a9591
Revert "Merge pull request #543 from sst/revert-536-VeniceUpdate"
...
This reverts commit 16f9f608de , reversing
changes made to 9b0ae67d59 .
2025-12-20 12:06:09 -06:00
Aiden Cline
16f9f608de
Merge pull request #543 from sst/revert-536-VeniceUpdate
...
Revert "Venice Autogenerate Script"
2025-12-20 08:50:00 -08:00
Charlie Gleason
eb31887b3d
Fix model mapping
2025-12-15 16:14:58 -08:00
Xinrui
a6b28915fe
aihubmix: add models claude-opus-4.5, gpt-5.1-codex-max, coding-glm-4.6-free
2025-12-09 19:18:08 +08:00
Dominik Oswald
c35c42fc34
Add Sonar Deep Research model configuration
...
- Introduce TOML configuration for Perplexity Sonar Deep Research model
- Include token pricing, request fees, and model limits
- Follow OpenCode AI schema conventions for model definitions
2025-10-17 13:12:19 +02:00
Dominik Oswald
8abedde07c
Add Perplexity Sonar Deep Research model configuration
...
- Introduce TOML configuration for Perplexity Sonar Deep Research model
- Include token pricing, request fees, and model limits
2025-10-17 13:10:38 +02:00