Compare commits

...

1196 Commits

Author SHA1 Message Date
Aiden Cline cda892e0d7 sync github copilot limits 2026-03-19 21:54:42 -05:00
Aiden Cline 098ff4f5bf Merge pull request #1227 from Verizane/dev
add OpenRouter models for gpt-5.4 mini and gpt-5.4 nano
2026-03-19 21:23:15 -05:00
Aiden Cline 0f70b8959f Merge pull request #1234 from mchenco/dev
Add Workers AI models: kimi-k2.5, nemotron-3-120b-a12b, glm-4.7-flash
2026-03-19 15:11:19 -05:00
mchen b8e6d58e5b add workers-ai models: kimi-k2.5, nemotron-3-120b-a12b, glm-4.7-flash 2026-03-19 14:59:17 -04:00
Roman Koslowski a855001a7e apply changes from review 2026-03-19 17:20:26 +01:00
Aiden Cline ac760b2268 Merge pull request #1230 from SamizuHM/feature/zhipuai-coding-plan-add-glm-5-turbo
zhipuai-coding-plan: Add glm-5-turbo.toml and replace symlink
2026-03-19 10:42:43 -05:00
Aiden Cline d4a5ea7ae7 Merge pull request #1226 from spiffytech/dev
Add Ollama Cloud support for Minimax M2.7
2026-03-19 10:41:47 -05:00
Aiden Cline 434ed89ba2 Merge pull request #1228 from dpuyosa/minimax_m2_7
Venice: Add MiniMax M2.7 and update DeepSeek V3.2 pricing
2026-03-19 10:41:16 -05:00
Aiden Cline 6d7719a62a Merge pull request #1229 from 0b1000/dev
Xiaomi: Add MiMo-V2-Pro and MiMo-V2-Omni
2026-03-19 10:41:06 -05:00
Aiden Cline 93637039ef Merge pull request #1231 from ariane-emory/feat/feat/add-xiaomi-mimo-v2-pro-and-omni
feat: add the Xiaomi MiMo V2 Pro and Xiaomi MiMo V2 Omni models to the OpenRouter provide
2026-03-19 10:40:44 -05:00
Ariane Emory 9c95f796c0 Merge remote-tracking branch 'upstream/dev' into feat/feat/add-xiaomi-mimo-v2-pro 2026-03-19 11:22:43 -04:00
Ariane Emory e8650b6073 feat: add xiaomi mimo-v2-pro and mimo-v2-omni models to openrouter 2026-03-19 11:18:46 -04:00
SamizuHM 23c2be6ff7 feat(zhipuai-coding-plan): add glm-5-turbo.toml and replace glm-5-turbo with symlink 2026-03-19 18:09:17 +08:00
Frank 913a63dbe6 update zen models 2026-03-19 00:33:45 -04:00
0b1000 503087e99b Merge branch 'anomalyco:dev' into dev 2026-03-19 12:28:38 +08:00
0b1000 48150f09d3 Xiaomi: Add MiMo-V2-Pro and MiMo-V2-Omni 2026-03-19 12:27:00 +08:00
Aiden Cline 5fef681657 Disable tool_call in grok model configuration 2026-03-18 23:09:30 -05:00
Frank 123054ae0c update zen models 2026-03-18 20:45:44 -04:00
Frank 03060d154b update zen models 2026-03-18 20:37:47 -04:00
dpuyosa 5c9b8108e0 Update minimax-m27.toml 2026-03-19 01:02:24 +01:00
dpuyosa c8084681f9 [venice] Add MiniMax M2.7 and update DeepSeek V3.2 pricing
- Add MiniMax M2.7 model with reasoning and tool_call support
 - Update DeepSeek V3.2 pricing (input: $0.33, output: $0.48, cache: $0.16)
2026-03-19 00:58:50 +01:00
Roman Koslowski 352ab4ae1b add gpt-5.4 mini and gpt-5.4 nano 2026-03-18 22:16:55 +01:00
spiffytech cf0b416b15 Added Ollama Cloud support for Minimax M2.7 2026-03-18 16:15:07 -04:00
Aiden Cline 38339a2a90 Merge pull request #1224 from APonce911/minimax-m2.7-openrouter
add MiniMax M2.7 to OpenRouter
2026-03-18 14:10:13 -05:00
Aiden Cline ff9040bf52 Update minimax-m2.7.toml 2026-03-18 14:09:26 -05:00
Aiden Cline 3039804af4 Delete providers/opencode/models/minimax-m2.7.toml 2026-03-18 14:08:55 -05:00
Frank 7a4ad7bec8 update go models 2026-03-18 14:40:25 -04:00
airton 721cc122bc add MiniMax M2.7 to OpenRouter and OpenCode 2026-03-18 18:57:09 +01:00
Aiden Cline 0527f019af Merge pull request #1221 from sergical/fix/bedrock-claude-4-6-context-window-and-pricing
fix(amazon-bedrock): set Claude Sonnet 4.6 and Opus 4.6 context window to 1M
2026-03-18 12:17:57 -05:00
Aiden Cline c89371de50 Merge pull request #1223 from sylviezhang37/update-vercel-models-20260318-1659
Update Vercel models
2026-03-18 12:17:22 -05:00
Sylvie Zhang 6d6d4220d8 Enable open_weights in minimax-m2.7.toml 2026-03-18 10:12:37 -07:00
Sylvie Zhang 8b984eeec1 Enable open_weights in minimax-m2.7-highspeed model 2026-03-18 10:12:21 -07:00
github-actions[bot] 586027c8f1 chore(vercel): update Vercel model definitions
Auto-generated by weekly workflow from Vercel AI Gateway API.

Co-Authored-By: github-actions[bot] <41898282+github-actions[bot]@users.noreply.github.com>
2026-03-18 16:59:24 +00:00
Sergiy Dybskiy 343b5f87ef fix(amazon-bedrock): set Claude Sonnet 4.6 and Opus 4.6 context window to 1M
Both models support a 1M token context window natively on Bedrock via the
Converse API with no beta headers required. Verified empirically via the
AWS CLI (bedrock-runtime converse): 950K tokens succeeds, >1M returns
'prompt is too long: N tokens > 1000000 maximum'.

The AWS Bedrock pricing page confirms long context pricing for these two
models is identical to standard pricing (no surcharge), so the
[cost.context_over_200k] section is removed as it was incorrect.
2026-03-18 12:20:38 -04:00
Aiden Cline 955b773ee5 Merge pull request #1218 from pomidornijfrukt/azure/5.4-mini-nano
Add GPT-5.4 Mini and Nano models for Azure providers
2026-03-18 10:31:13 -05:00
Aiden Cline 98559071f0 Merge pull request #1217 from cgilly2fast/dev
chore(firmware): update base url and docs url
2026-03-18 10:30:44 -05:00
eCube-cachy 0660308816 add: GPT-5.4 Mini and Nano model configurations for Azure providers 2026-03-18 15:17:52 +02:00
Jack 380f9dd8eb Merge pull request #1216 from no1wudi/dev
Add MiniMax M2.7 and M2.7-highspeed models to 4 official providers
2026-03-18 16:29:59 +08:00
Jack 1cfdab1b18 update MiniMax-M2.7 cache_read to 0.06 2026-03-18 16:27:53 +08:00
Colby Gilbert 75a981f957 chore(firmware): update base url and docs url 2026-03-18 00:41:05 -07:00
Huang Qi 7fadbcadc8 Add MiniMax M2.7 and M2.7-highspeed models to 4 official providers 2026-03-18 15:21:06 +08:00
Frank 38f9092292 update zen models 2026-03-18 02:30:18 -04:00
Aiden Cline 92149b9eaa rm nonexistant github model 2026-03-17 21:41:51 -05:00
Aiden Cline b614f0e69c Merge pull request #1214 from luisrudge/dev
Add GPT-5.4 mini and nano to GitHub Copilot provider
2026-03-17 20:13:46 -05:00
Luís Rudge 67d6dac5c5 Add GPT-5.4 mini and nano to GitHub Copilot provider 2026-03-17 18:44:38 -06:00
Aiden Cline 7d3cc61a48 Merge pull request #1207 from PedroACosta/feat/add-dinference-provider
feat(providers): add dinference provider
2026-03-17 14:51:31 -05:00
Aiden Cline f02ea6c4d2 Merge pull request #1115 from skywalker512/feat/add-tencent-coding-plan
feat: add Tencent Coding Plan provider
2026-03-17 14:51:19 -05:00
Aiden Cline 0cb50eeece Merge pull request #1208 from scwgoire/march-update
Scaleway 26-03 model updates
2026-03-17 14:48:12 -05:00
Aiden Cline 878311d2e0 Merge pull request #1210 from dm-cohere/dm/fix-update-cohere-model-capabilities
fix(models): update cohere model capabilities
2026-03-17 14:32:27 -05:00
Aiden Cline a0e89f65d6 Merge pull request #1206 from 0b1000/dev
Rename minimax-m2.5.toml to MiniMax-M2.5.toml
2026-03-17 14:32:19 -05:00
Aiden Cline 74099b7c9c Merge pull request #1213 from smrdotgg/add-openai-gpt-5-4-mini-and-nano
Add OpenAI GPT-5.4 mini and nano
2026-03-17 14:31:24 -05:00
Aiden Cline ec522435c3 Merge pull request #1211 from sylviezhang37/update-vercel-models-20260317-1807
Update Vercel models
2026-03-17 14:30:24 -05:00
smr d839cd37d4 Add OpenAI GPT-5.4 mini and nano
Capture the newly released mini and nano model metadata so models.dev reflects OpenAI's latest GPT-5.4 lineup with current pricing, limits, and knowledge cutoff.
2026-03-17 22:09:13 +03:00
github-actions[bot] ecb6ef7f93 chore(vercel): update Vercel model definitions
Auto-generated by weekly workflow from Vercel AI Gateway API.

Co-Authored-By: github-actions[bot] <41898282+github-actions[bot]@users.noreply.github.com>
2026-03-17 18:07:06 +00:00
Deirdre Meehan 8b4d341054 fix: cohere models on non-cohere providers 2026-03-17 16:51:24 +00:00
Deirdre Meehan 62f4a28308 fix: cohere provider models 2026-03-17 16:44:55 +00:00
Pedro 2fb8ef0dc8 feat(providers): add dinference provider 2026-03-17 14:13:36 +01:00
Gregoire de Turckheim 96968e2bf8 feat: Scaleway 26-03 model updates 2026-03-17 12:15:52 +01:00
0b1000 ee9d7879ce Rename minimax-m2.5.toml to MiniMax-M2.5.toml 2026-03-17 14:50:35 +08:00
Frank 71283512a6 update zen models 2026-03-17 02:21:13 -04:00
Frank cd4afd7e7c update zen models 2026-03-17 02:19:17 -04:00
Aiden Cline 1239d0190b Merge pull request #1204 from cyberofficial/vultr
VULTR: Updated Vultr model pricing to reflect current serverless inference rates
2026-03-16 16:10:39 -05:00
Aiden Cline 491bf6ccba Merge pull request #1202 from RaviTharuma/fix/chutes-pricing-update-2026-03
fix(chutes): update pricing and limits from live API
2026-03-16 16:10:25 -05:00
Cyber Official c993d0c121 Updated Vultr model pricing to reflect current serverless inference rates
Updated Vultr model pricing to reflect current serverless inference rates

This commit updates the cost configuration for all Vultr models to align with their latest pricing tiers:

**Cost Reductions:**
- DeepSeek-R1-Distill-Qwen-32B: Input $0.55→$0.30, Output $2.75→$0.30 (73% reduction)
- NVIDIA-Nemotron-3-Super-120B-A12B-NVFP4: Input $0.55→$0.20, Output $2.75→$0.80 (64% input, 71% output reduction)
- Qwen2.5-Coder-32B-Instruct: Input $0.55→$0.20, Output $2.75→$0.60 (64% input, 78% output reduction)
- gpt-oss-120b: Input $0.55→$0.15, Output $2.75→$0.60 (73% input, 78% output reduction)
- MiniMax-M2.5: Input $0.55→$0.30, Output $2.75→$1.20 (45% input, 56% output reduction)

**Cost Adjustments:**
- DeepSeek-R1-Distill-Llama-70B: Input $0.55→$2.00, Output $2.75→$2.00 (significant increase)
- DeepSeek-V3.2: Output $2.75→$1.65 (40% reduction)
- Llama-3.1-Nemotron-Ultra-253B-v1: Output $2.75→$1.80 (35% reduction)
- GLM-5-FP8: Input $0.55→$0.85, Output $2.75→$3.10 (55% input, 13% output increase)
2026-03-16 13:58:28 -04:00
Aiden Cline e55c39a83d Merge pull request #1141 from sk0x0y/feature/nanogpt-confirmed-suffix2-fixes
fix(nano-gpt): rename confirmed 2-suffix model ids
2026-03-16 10:57:44 -05:00
Aiden Cline 6dea000e25 Merge pull request #1148 from sk0x0y/feature/nanogpt-bundled-confirmed-suffix2-fixes
fix(nano-gpt): rename bundled confirmed 2-suffix model ids
2026-03-16 10:57:30 -05:00
Aiden Cline 3f7a757b3f Merge pull request #1142 from sk0x0y/feature/nanogpt-more-confirmed-suffix2-fixes
fix(nano-gpt): rename more confirmed 2-suffix model ids
2026-03-16 10:56:36 -05:00
Aiden Cline c693fd71e2 Merge pull request #1194 from cyberofficial/vultr
Update Vultr model list with 10 new models and updated pricing
2026-03-16 10:55:19 -05:00
Aiden Cline 54e04e288a Merge pull request #1198 from amritbanerjee/add-glm-5-turbo
Add GLM-5-Turbo model support
2026-03-16 10:47:08 -05:00
Aiden Cline 462a179eee Merge pull request #1203 from jerome-benoit/fix/sap-ai-core-model-specs
fix(sap-ai-core): align model specs with official sources
2026-03-16 10:46:40 -05:00
Aiden Cline 95db59034d Merge pull request #1201 from dpuyosa/venice-new-models
Venice: Add new provider models
2026-03-16 10:45:59 -05:00
Aiden Cline 74dcc74e32 Merge pull request #1200 from dpuyosa/venice/pricing-update
Venice: Update model pricing for 7 models
2026-03-16 10:45:47 -05:00
Jérôme Benoit 57975f5f25 fix(sap-ai-core): align model specs with official sources 2026-03-16 13:59:06 +01:00
Ravi Tharuma ad7b063747 fix(chutes): update pricing and limits from live API
Synced 6 Chutes model definitions against the live API at
https://llm.chutes.ai/v1/models (queried 2026-03-16).

Models updated:
- deepseek-ai/DeepSeek-V3.2-TEE: cost 0.25/0.38→0.28/0.42, cache 0.125→0.14, context 163840→131072
- zai-org/GLM-5-TEE: cost 0.75/2.5→0.95/3.15, added cache_read 0.475
- zai-org/GLM-4.6-TEE: cost 0.35/1.5→0.4/1.7, added cache_read 0.2
- zai-org/GLM-4.6V: added cache_read 0.15
- MiniMaxAI/MiniMax-M2.5-TEE: cost 0.15/0.6→0.3/1.1, added cache_read 0.15
- Qwen/Qwen3.5-397B-A17B-TEE: cost 0.3/1.2→0.39/2.34, cache 0.15→0.195
2026-03-16 11:42:29 +01:00
dpuyosa f76e9f0551 [venice] Add new provider models
- Add mistral-small-3.2-24b-instruct, qwen3-5-9b, venice-uncensored-role-play, zai-org-glm-4.6
2026-03-16 09:37:41 +01:00
dpuyosa d70a49b36f [venice] Update model pricing for 7 models
- Remove context_over_200k pricing from Claude models
- Update Grok cache_read pricing from 0.5 to 0.25
- Update Kimi, MiniMax input/output pricing
2026-03-16 09:05:37 +01:00
amrit 3487135f9f Add GLM-5-Turbo model support 2026-03-16 12:14:50 +11:00
Aiden Cline 458a66c766 Merge pull request #1197 from kesku/update-perplexity-agent-models
Update Perplexity Agent API models
2026-03-15 10:59:23 -05:00
Frank d3a84dc7ec update zen models 2026-03-15 10:59:52 -04:00
Kesku ae61b25583 update perplexity-agent: add gpt-5.4 & nemotron, remove gemini-3-pro 2026-03-15 03:46:50 +00:00
Aiden Cline 74be576eda Merge pull request #1178 from Sewer56/change-synthetic-endpoint
Add OpenAI and Anthropic compatible endpoints
2026-03-14 20:55:30 -05:00
Aiden Cline 164df2cda0 Merge pull request #1191 from Alcatraz-Zhang/update/kilo-models
Sync Kilo model definitions with latest gateway catalog
2026-03-14 20:54:45 -05:00
Cyber Official 2cd7908369 Update Vultr model list with 10 new models and updated pricing
- Updated pricing to $0.55/M input tokens, $2.75/M output tokens
- Updated context limits to safe floor values from official testing
- Added accurate output token limits from official model documentation
- Added 5 new models: MiniMax M2.5, DeepSeek V3.2, GLM-5 FP8, Llama 3.1 Nemotron Ultra 253B, NVIDIA Nemotron 3 Super 120B A12B NVFP4
- Updated existing models: DeepSeek R1 Distill variants, GPT OSS 120B, Kimi K2.5, Qwen2.5 Coder 32B

Model specifications:
- MiniMax M2.5: 196K context, 4,096 output
- Qwen2.5-Coder-32B: 15K context, 256 output (notable low default)
- DeepSeek R1 Distill Llama 70B: 130K context, 4,096 output
- DeepSeek R1 Distill Qwen 32B: 130K context, 4,096 output
- DeepSeek V3.2: 163K context, 4,096 output
- Kimi K2.5: 261K context, 32,768 output (high output limit)
- GPT OSS 120B: 130K context, 8,192 output
- GLM-5 FP8: 202K context, 131,072 output (exceptionally high)
- Llama 3.1 Nemotron Ultra 253B: 32K context, 4,096 output
- NVIDIA Nemotron 3 Super 120B A12B NVFP4: 260K context, 8,192 output

All models set to text-only (no vision support) as confirmed.
2026-03-14 19:47:01 -04:00
Alcatraz-Zhang cc667340f5 Sync Kilo model definitions with latest gateway catalog
Refresh the Kilo provider catalog so models.dev matches the current gateway inventory, pricing, and availability.
2026-03-15 04:35:38 +08:00
Sewer56 f2cfc1435d Changed: Synthetic to use newer openai endpoint 2026-03-14 17:09:44 +00:00
Aiden Cline 35bb8cca47 Merge pull request #1172 from bigfluffycookie/add-deepinfra-llama-models
Add deepinfra llama models
2026-03-14 10:55:13 -05:00
Aiden Cline 3468a410e1 Merge pull request #1177 from ar27111994/dev
Add Grok 4.1 Fast configurations for reasoning and non-reasoning
2026-03-14 10:54:57 -05:00
Aiden Cline b1b5e3c5cd Merge pull request #1174 from dacbd/patch-1
fix(wandb): fix k2.5 settings
2026-03-14 10:54:35 -05:00
Aiden Cline 97f03ec672 Merge pull request #1175 from dacbd/patch-2
chore(docs): add note for manual testing with opencode
2026-03-14 10:54:22 -05:00
BigFluffyCookie 9b516924aa Add limit output for llama models 2026-03-14 11:49:57 +01:00
Ahmed Rehan 929a39600b feat(models): add Grok 4.1 Fast (Reasoning and Non-Reasoning) configurations 2026-03-14 14:27:24 +05:00
Daniel Barnes a87d8bb8cc chore(docs): add note for manual testing with opencode 2026-03-14 13:42:57 +09:00
Daniel Barnes 574139eb49 fix(wandb): fix k2.5 settings 2026-03-14 13:07:16 +09:00
Aiden Cline 1e3bc38b31 Merge pull request #1137 from mcowger/mcowger/correct-gemini-flash-lite-pricing
Fix incorrect pricing for gemini-3.1-flash-lite-preview
2026-03-13 18:41:41 -05:00
Aiden Cline 8916fe9874 Merge pull request #1171 from stephenkuhn214/dev
fix(amazon-bedrock): Remove deprecated and add missing models
2026-03-13 18:26:25 -05:00
BigFluffyCookie 5d956b41a6 Rename llama models to remove "Meta" prefix 2026-03-13 23:15:27 +01:00
BigFluffyCookie 42a7a14f69 Add Meta Llama models to DeepInfra provider 2026-03-13 22:53:39 +01:00
Stephen Kuhn f24ee000d7 fix(amazon-bedrock): update and add models
- Remove 19 deprecated/EOL models
- Add 7 new models: DeepSeek V3.2, Llama 3.1 405B, Magistral Small 1.2, Ministral 3 3B, Mistral Large 3, Pixtral Large, NVIDIA Nemotron Nano 3 30B
- Fix Devstral 2 123B: correct name, family, and open_weights
- Set accurate Bedrock launch dates for all new models
2026-03-13 16:02:04 -04:00
Aiden Cline 7196b1fb2c Merge pull request #1170 from anomalyco/revert-1166-fix/update-gpt53-codex-spark-preview
Revert "fix(openai): rename gpt-5.3-codex-spark to gpt-5.3-codex-spark-preview"
2026-03-13 14:31:33 -05:00
Aiden Cline f6c0d5a29d Revert "fix(openai): rename gpt-5.3-codex-spark to gpt-5.3-codex-spark-preview" 2026-03-13 14:30:58 -05:00
Aiden Cline ee63449aa5 sonnet 4.6 and opus 4.6 1M context 2026-03-13 14:27:55 -05:00
Aiden Cline 92aa44ec00 Merge pull request #1166 from rluisr/fix/update-gpt53-codex-spark-preview
fix(openai): rename gpt-5.3-codex-spark to gpt-5.3-codex-spark-preview
2026-03-13 14:18:41 -05:00
Aiden Cline 477284535c Rename model from 'GPT-5.3 Codex Spark Preview' to 'GPT-5.3 Codex Spark' 2026-03-13 14:17:44 -05:00
Aiden Cline 304233bdda Merge pull request #1169 from mdrxy/mdrxy/anthropic-token-limits
Update Claude 4.6 context/pricing
2026-03-13 14:13:40 -05:00
Aiden Cline 25d782ee2c Reduce context limit from 1,000,000 to 200,000 2026-03-13 14:13:30 -05:00
Aiden Cline 0f63393d51 Update context limit in claude-opus-4-6.toml 2026-03-13 14:12:56 -05:00
rluisr e780eefce2 fix(openai): rename gpt-5.3-codex-spark to gpt-5.3-codex-spark-preview
The OpenAI API expects model ID 'gpt-5.3-codex-spark-preview', not
'gpt-5.3-codex-spark'. Rename model files in both openai and opencode
providers so the generated model ID matches the actual API.
2026-03-14 03:59:03 +09:00
Aiden Cline a79585fa83 Merge pull request #1163 from micuintus/feature/Kimi2.5-fast
feat(nebius): add Kimi-K2.5-fast model
2026-03-13 13:14:38 -05:00
Aiden Cline 00801f74f2 Merge pull request #1164 from butyess/dev
Openrouter models: gemini 3.1 flash lite preview, grok 4.20 beta models.
2026-03-13 13:14:22 -05:00
Aiden Cline 185f6731ee Merge pull request #1162 from dpuyosa/feature/venice-grok-4-20-beta
Venice: Add Grok 4.20 Beta models
2026-03-13 12:53:28 -05:00
Aiden Cline d291b0575c Merge pull request #1167 from sylviezhang37/update-vercel-models-20260313-1639
Update Vercel models
2026-03-13 12:53:11 -05:00
Mason Daugherty 382d9f3e7d Update Claude 4.6 context/pricing 2026-03-13 13:53:04 -04:00
Aiden Cline e64f5fe963 Merge pull request #1168 from mdrxy/mdrxy/update-baseten
Update Baseten models
2026-03-13 12:51:56 -05:00
Mason Daugherty ea57ddfe7e Update Baseten models 2026-03-13 13:48:41 -04:00
github-actions[bot] 29463d7fa8 chore(vercel): update Vercel model definitions
Auto-generated by weekly workflow from Vercel AI Gateway API.

Co-Authored-By: github-actions[bot] <41898282+github-actions[bot]@users.noreply.github.com>
2026-03-13 16:39:30 +00:00
Jack bcc8db49ee Merge pull request #1165 from anomalyco/chore/openrouter-alpha-reasoning-details-20260313
feat(openrouter): add interleaved reasoning details for alpha models
2026-03-13 22:25:25 +08:00
Jack c8521d70f3 feat(openrouter): add interleaved reasoning details for alpha models 2026-03-13 22:20:54 +08:00
Federico Masi 490cd249e4 Openrouter models: gemini 3.1 flash lite preview, grok 4.20 beta models. 2026-03-13 15:11:12 +01:00
Michael Voigt dbc636f5f3 feat(nebius): add Kimi-K2.5-fast model 2026-03-13 12:35:13 +01:00
Michael Voigt 9a32f671a1 fix(nebius): lowercase model ID for Nemotron-3-Super-120B-A12B
The filename must match the API casing (lowercase) to avoid 'model does not exist' errors.
2026-03-13 12:35:08 +01:00
dpuyosa 856d925eda [venice] Add Grok 4.20 Beta models
- Add Grok 4.20 Beta model configuration (2M context, 128K output)
- Add Grok 4.20 Multi-Agent Beta model configuration
2026-03-13 10:48:38 +01:00
Aiden Cline 066a425917 Merge pull request #1158 from micuintus/feature/Nebius_Nemotron-3-Super-120b-a12b
feat(nebius): Add support for Nemotron-3-Super-120B-A12B
2026-03-12 22:20:20 -05:00
Aiden Cline 6df7f20cdc Merge pull request #1156 from dsingal0/dev
added nemotron super on baseten
2026-03-12 22:20:06 -05:00
Aiden Cline 78bb47b90e Merge pull request #1151 from dacbd/dacbd
fix(wandb): update models
2026-03-12 22:19:43 -05:00
Aiden Cline c121d86419 Merge pull request #1160 from kreatoo/dev
feat: add zai-org/glm-4.7 and zai-org/glm-4.7-flash to NanoGPT
2026-03-12 22:11:18 -05:00
Aiden Cline ab148eeb14 Merge pull request #1161 from Grin1024/dev
Add Claude Opus 4.6 and Sonnet 4.6 models to RequestY provider
2026-03-12 22:11:07 -05:00
lihui 49d196d326 Add Claude Opus 4.6 and Sonnet 4.6 models to RequestY provider 2026-03-13 09:00:54 +08:00
Kreato 8899b390ef feat: add zai-org/glm-4.7 and zai-org/glm-4.7-flash to NanoGPT 2026-03-13 00:27:09 +03:00
Michael Voigt 5217f62ddf fix(nebius): Follow context updates for Kimi 2.5 and GLM-5 2026-03-12 20:22:48 +01:00
Michael Voigt 55eaff9af1 feat(nebius): Add support for Nemotron-3-Super-120B-A12B 2026-03-12 20:22:21 +01:00
Dhruv Singal 7557c06ac0 update output length 2026-03-12 09:41:25 -07:00
Dhruv Singal e85d820121 fix input output 2026-03-12 08:29:01 -07:00
Dhruv Singal 499d3a39ef remove cache pricing 2026-03-12 08:21:22 -07:00
Dhruv Singal b9b38d6e33 added nemotron super on baseten 2026-03-12 08:18:46 -07:00
Aiden Cline ca24ac14fa Merge pull request #1153 from dpuyosa/dev
Venice: Update model output token limits
2026-03-12 10:08:46 -05:00
Aiden Cline 822546fc67 Merge pull request #1155 from spiffytech/dev
Add Ollama Cloud support for Nemotron 3 Super
2026-03-12 10:08:31 -05:00
Aiden Cline 4555195b71 Merge pull request #1152 from v1gnesh/dev
Update grok-4.20 model defs
2026-03-12 10:08:15 -05:00
spiffytech 5eae8effc6 Added Ollama Cloud support for Nemotron 3 Super 2026-03-12 09:28:47 -04:00
dpuyosa c1801aef87 [venice] Normalize model output token limits
- Update output limits to standard values across all models
2026-03-12 10:08:39 +01:00
v1gnesh 5e6464b272 Update grok-4.20-beta-reasoning 2026-03-12 10:27:40 +05:30
v1gnesh e1a4f23332 Update grok-4.20-beta-non-reasoning 2026-03-12 10:26:03 +05:30
v1gnesh 753e1f9f0c grok-multi-agent-beta update 2026-03-12 10:23:57 +05:30
Daniel Barnes 123ecd2ba5 docs url 2026-03-12 13:27:56 +09:00
Daniel Barnes f15cda9fcb remove old 2026-03-12 13:26:08 +09:00
Daniel Barnes 0205debbd3 fix values 2026-03-12 13:22:29 +09:00
Daniel Barnes 0059766509 number formating 2026-03-12 13:17:22 +09:00
Daniel Barnes be81b02916 additional model files 2026-03-12 13:02:17 +09:00
Daniel Barnes 2dab141166 initial script & model updates 2026-03-12 13:01:35 +09:00
Aiden Cline 45aa49af25 tweak: azure kimi k2.5 2026-03-11 22:35:20 -05:00
Aiden Cline 781fad3ad4 Merge pull request #1150 from cau1k/5.4-family
feat(azure): add 5.4/pro families
2026-03-11 22:14:08 -05:00
cau1k 99d2ffcfdd feat(azure): add 5.4/pro families 2026-03-11 20:59:11 -04:00
Aiden Cline 381d7cc19d Merge pull request #1149 from ariane-emory/fear/add-march-or-stealth-models
Add OpenRouter stealth models: Hunter Alpha and Healer Alpha
2026-03-11 18:07:50 -05:00
Ariane Emory 7482e22458 Fix family field to use 'alpha' for stealth models 2026-03-11 18:49:32 -04:00
Ariane Emory f5e6a402e6 Add OpenRouter stealth models: Hunter Alpha and Healer Alpha 2026-03-11 18:41:58 -04:00
Aiden Cline 9265852852 tweak: adjust some gh limits to align better w/ api 2026-03-11 15:23:44 -05:00
Aiden Cline dc98a32996 Merge pull request #1018 from Sewer56/add-synthetic-missing-models
Update synthetic.new models: promote MiniMax-M2.5, add GLM-4.7-Flash
2026-03-11 14:55:50 -05:00
Aiden Cline 56c39ae0f6 Merge pull request #1140 from sk0x0y/feature/nanogpt-thudm-id-fixes
fix(nano-gpt): rename THUDM 2 ids to canonical THUDM ids
2026-03-11 14:55:07 -05:00
Aiden Cline b1f43a7595 Merge pull request #1147 from msadiks/fix/alibaba-coding-minimax
fix: alibaba-coding-plan MiniMax-M2.5 context window
2026-03-11 14:54:37 -05:00
Matt Cowger fed8bcae19 Merge branch 'dev' into mcowger/correct-gemini-flash-lite-pricing 2026-03-11 12:23:42 -07:00
sk0x0y fb95150d02 fix(nano-gpt): rename VongolaChouko model id
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode)

Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai>
2026-03-12 04:20:39 +09:00
sk0x0y a7c9a240b4 fix(nano-gpt): rename Steelskull model ids
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode)

Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai>
2026-03-12 04:20:39 +09:00
sk0x0y 6432a4a3e6 fix(nano-gpt): rename Sao10K model ids
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode)

Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai>
2026-03-12 04:20:38 +09:00
sk0x0y f2e4a249fe fix(nano-gpt): rename NeverSleep model ids
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode)

Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai>
2026-03-12 04:20:38 +09:00
sk0x0y 8667a6eed8 fix(nano-gpt): rename MarinaraSpaghetti model id
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode)

Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai>
2026-03-12 04:19:56 +09:00
sk0x0y 429554397a fix(nano-gpt): rename LatitudeGames model id
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode)

Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai>
2026-03-12 04:19:56 +09:00
sk0x0y a64e6ad0ac fix(nano-gpt): rename LLM360 model id
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode)

Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai>
2026-03-12 04:19:56 +09:00
sk0x0y d68d79888c fix(nano-gpt): rename Infermatic model id
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode)

Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai>
2026-03-12 04:19:56 +09:00
sk0x0y 6c52905c6a fix(nano-gpt): rename Gryphe model id
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode)

Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai>
2026-03-12 04:19:55 +09:00
sk0x0y 62410b8f26 fix(nano-gpt): rename GalrionSoftworks model id
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode)

Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai>
2026-03-12 04:19:55 +09:00
sk0x0y 50ce68ccab fix(nano-gpt): rename Envoid model ids
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode)

Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai>
2026-03-12 04:19:20 +09:00
sk0x0y d1c6a6b873 fix(nano-gpt): rename EVA-UNIT-01 model ids
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode)

Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai>
2026-03-12 04:19:20 +09:00
Frank 7193b068a5 update zen models 2026-03-11 13:52:50 -04:00
Sadik 79a8a06bd7 fix MiniMax-M2.5 context window 2026-03-11 20:50:33 +03:00
Aiden Cline b60c03e11c Merge pull request #1139 from zainhas/dev
[Together AI] add prompt caching pricing for MiniMax m2.5
2026-03-11 12:31:56 -05:00
Aiden Cline 15cf98d57b Merge pull request #1146 from gotjoshua/patch-1
Rename step-3-5-flash.toml to step-3.5-flash.toml
2026-03-11 12:31:39 -05:00
Aiden Cline b2ee6c407b Merge pull request #1144 from micuintus/feature/update-nebius-changes
Feat: update Nebius changes
2026-03-11 12:31:29 -05:00
gotjoshua 96a14a06e7 Rename step-3-5-flash.toml to step-3.5-flash.toml
on nvidia it is 3.5 not 3-5
2026-03-11 11:41:36 +00:00
Michael Voigt adc358606d fix(nebius): update model context limits per API 2026-03-11 11:33:14 +01:00
Michael Voigt 63d52adf6f feat(nebius): add GLM-5 model 2026-03-11 11:33:14 +01:00
sk0x0y 9a31387766 fix(nano-gpt): rename Salesforce model id
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode)

Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai>
2026-03-11 18:17:41 +09:00
sk0x0y 735157b837 fix(nano-gpt): rename ReadyArt model ids
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode)

Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai>
2026-03-11 18:17:41 +09:00
sk0x0y d75b46fb37 fix(nano-gpt): rename Doctor-Shotgun model id
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode)

Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai>
2026-03-11 18:17:41 +09:00
sk0x0y cc555f8482 fix(nano-gpt): rename CrucibleLab model id
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode)

Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai>
2026-03-11 18:17:13 +09:00
sk0x0y 7fbbcf2b49 fix(nano-gpt): rename MiniMaxAI model id
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode)

Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai>
2026-03-11 16:04:28 +09:00
sk0x0y 14c8ec8ca5 fix(nano-gpt): rename Tongyi-Zhiwen model id
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode)

Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai>
2026-03-11 16:04:28 +09:00
sk0x0y 72568bbdb3 fix(nano-gpt): rename Alibaba-NLP model id
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode)

Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai>
2026-03-11 16:03:57 +09:00
sk0x0y c2225b715f fix(nano-gpt): rename THUDM GLM-Z1 rumination id
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode)

Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai>
2026-03-11 15:49:20 +09:00
sk0x0y 7ce25e3742 fix(nano-gpt): rename THUDM GLM-Z1 model ids
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode)

Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai>
2026-03-11 15:49:20 +09:00
sk0x0y 427868604b fix(nano-gpt): rename THUDM GLM-4 model ids
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode)

Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai>
2026-03-11 15:49:20 +09:00
Zain Hasan 247cd801a8 add prompt caching pricing for MiniMax m2.5 2026-03-10 22:54:42 -07:00
Aiden Cline 1aa2ee22b1 Merge pull request #1134 from sk0x0y/feature/nanogpt-catalog-fixes
fix(nano-gpt): correct TEE path ids and add missing canonical entries
2026-03-10 22:02:52 -05:00
Aiden Cline 0f57233eff Merge pull request #1105 from sylviezhang37/add-vercel-input-context-and-new-models
feat(vercel): add input context calculation + new models
2026-03-10 22:01:52 -05:00
Aiden Cline 73a78eebfc Merge pull request #1138 from mugnimaestra/feat/add-glm-5-turbo-chutes
feat: add GLM-5-Turbo to Chutes provider listings
2026-03-10 22:01:08 -05:00
Sylvie Zhang 3a6789b819 Merge branch 'dev' into add-vercel-input-context-and-new-models 2026-03-10 17:44:14 -07:00
Sylvie Zhang f7c505e140 remove context from gemini models 2026-03-10 17:43:08 -07:00
Sylvie Zhang 6bb36806d6 only calc input context for openai models 2026-03-10 17:40:46 -07:00
Sylvie Zhang 20a404eb88 revert non openai changes 2026-03-10 17:38:46 -07:00
Muhammad Mugni Hadi 65ecb5cd4a feat: add GLM-5-Turbo to Chutes provider listings 2026-03-11 05:26:11 +07:00
Matt Cowger 56062a9129 Fix incorrect pricing 2026-03-10 14:57:44 -07:00
Aiden Cline d3d9c580d4 Merge pull request #1135 from gitpush-gitpaid/fix/gpt-5-4-pdf-input-modalities
Added PDF to input modalities for GPT-5.4
2026-03-10 13:53:42 -05:00
gitpush-gitpaid ef98d8a9cb Updated GPT-5.4 PDF input modalities 2026-03-10 13:59:29 -04:00
sk0x0y 9d17752b88 fix(nano-gpt): add missing GLM 5 thinking model
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode)

Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai>
2026-03-11 00:56:22 +09:00
sk0x0y b5a838fe8b fix(nano-gpt): add missing TEE qwen3.5 model
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode)

Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai>
2026-03-11 00:56:22 +09:00
sk0x0y aa1ac39ee6 fix(nano-gpt): rename TEE gemma and minimax ids
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode)

Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai>
2026-03-11 00:56:22 +09:00
sk0x0y 4bc17ccf96 fix(nano-gpt): rename TEE oss and llama ids
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode)

Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai>
2026-03-11 00:56:22 +09:00
sk0x0y 08c1899bfe fix(nano-gpt): rename TEE deepseek model ids
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode)

Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai>
2026-03-11 00:56:02 +09:00
sk0x0y ad50e4a5ed fix(nano-gpt): rename TEE qwen model ids
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode)

Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai>
2026-03-11 00:56:02 +09:00
sk0x0y 730915a123 fix(nano-gpt): rename TEE kimi model ids
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode)

Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai>
2026-03-11 00:56:02 +09:00
sk0x0y 6f12d18cb8 fix(nano-gpt): rename TEE glm model ids
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode)

Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai>
2026-03-11 00:56:02 +09:00
Aiden Cline bd8774db99 Merge pull request #1132 from sk0x0y/feature/nanogpt-model-sync
feat(nano-gpt): add text and image models
2026-03-10 10:31:50 -05:00
Aiden Cline 88fbea52a4 Merge pull request #1133 from anomalyco/fix-model
fix: bedrock devstral
2026-03-10 10:31:08 -05:00
Aiden Cline 70e5d9b34b fix: bedrock devstral 2026-03-10 10:30:20 -05:00
Aiden Cline edb6ef0d71 Merge pull request #1129 from Grin1024/dev
feat: add GPT-5 series models to requesty provider
2026-03-10 10:29:30 -05:00
Aiden Cline df1280ed8b add families to some bedrock models 2026-03-10 10:12:13 -05:00
sk0x0y 898b3c18b7 feat(nano-gpt): add image models
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode)

Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai>
2026-03-10 22:13:48 +09:00
sk0x0y 6316e543ef feat(nano-gpt): add text models
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode)

Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai>
2026-03-10 22:13:48 +09:00
Aiden Cline 64d9a97f9d Merge pull request #1128 from JWahle/dev
chore: Updated abacus model definitions
2026-03-10 07:40:01 -05:00
Aiden Cline c62a2a3fc1 Merge pull request #1130 from Mingholy/fix/alibaba-coding-plan-model-limits
fix: update model limits for alibaba-coding-plan providers
2026-03-10 07:39:48 -05:00
Aiden Cline cc1937a177 Merge pull request #1131 from janszypulski/cloudferro-sherlock-fix-minimax-model-id
fix minimax-m2.5 model id - wrong file path
2026-03-10 07:39:34 -05:00
Jan Szypulski 9d3a88863d fix minimax-m2.5 model id - wrong file path 2026-03-10 11:21:55 +01:00
mingholy.lmh b9123e26e0 fix: update model limits for alibaba-coding-plan providers
- Add MiniMax-M2.5 to alibaba-coding-plan-cn
- Update qwen3-max output limit (65536 -> 32768)
- Update qwen3-coder-plus context limit (1048576 -> 1000000)
- Update MiniMax-M2.5 limits per ref.json (context: 196608, output: 24576)

Co-authored-by: Qwen-Coder <qwen-coder@alibabacloud.com>
2026-03-10 15:45:01 +08:00
lihui 933e450104 feat: add GPT-5 series models to requesty provider
Add missing OpenAI GPT-5 series models to requesty provider:
- GPT-5 Chat, Codex, Image, Pro
- GPT-5.1 Chat, Codex, Codex-Max, Codex-Mini
- GPT-5.2 Chat, Codex, Pro
- GPT-5.3 Codex
- GPT-5.4, GPT-5.4 Pro
2026-03-10 14:53:28 +08:00
JWahle 2ba6383e70 chore: Updated abacus model definitions
Added: gpt-5.4.toml
Removed: gemini-3-pro-preview.toml
2026-03-10 05:10:45 +01:00
Aiden Cline 65ed6ac5dd Merge pull request #1126 from mcowger/feature/gemini-3.1-flash-lite-vercel
feat: add gemini-3.1-flash-lite-preview to vercel gateway provider
2026-03-09 21:43:33 -05:00
Aiden Cline e7d04aec7a Merge pull request #1127 from anomalyco/add-shape
feat: add 'shape' field to provider so models can specify if they use responses vs completions apis (use only if model only supports 1 of)
2026-03-09 21:43:04 -05:00
Aiden Cline 37fe334aed feat: add 'shape' field to provider so models can specify if they use responses vs completions apis (use only if model only supports 1 of) 2026-03-09 21:42:23 -05:00
Matt Cowger ce8fc9e4f0 feat: add gemini-3.1-flash-lite-preview to vercel gateway provider 2026-03-09 19:32:43 -07:00
Aiden Cline be8eb8ba54 fix name 2026-03-09 20:01:02 -05:00
Aiden Cline 7c625b3b82 Merge pull request #945 from Daltonganger/feat/nano-gpt-sync-models-api
sync nano-gpt models with live API catalog
2026-03-09 20:00:06 -05:00
Aiden Cline 33700d27dc Merge pull request #1032 from propilideno/feature/new_gpt_5.3_codex_and_missing_structured_output_attr
Add gpt-5.3-codex (Azure) and fill missing structured output flags
2026-03-09 19:40:20 -05:00
Aiden Cline e5c300a5e5 fix 2026-03-09 19:38:30 -05:00
Aiden Cline e5e9175c5d Merge branch 'dev' into feature/new_gpt_5.3_codex_and_missing_structured_output_attr 2026-03-09 19:37:40 -05:00
Aiden Cline a9f79d6794 Merge pull request #1123 from dpuyosa/feature/venice-gpt54-multimodal
Venice: Add GPT-5.4 Pro and enable multimodal inputs for GPT-5.4 & Qwen3.5
2026-03-09 18:19:52 -05:00
Aiden Cline fd4c4a8f28 Merge pull request #1038 from muldercw/add-clarifai-model-provider
Add Clarifai Model Provider
2026-03-09 18:19:14 -05:00
dpuyosa f9b5385868 [venice] Add GPT-5.4 Pro and enable multimodal inputs
- Add GPT-5.4 Pro model
- Enable attachment/image input for GPT-5.4
- Enable attachment/image/video input for Qwen3.5 35B A3B
2026-03-09 22:52:54 +01:00
Aiden Cline b2f7a72410 Merge pull request #1110 from fhennerkes/dev
poe: add GPT-5.4, GPT-5.4-Pro, and GPT-5.3-Instant models
2026-03-09 14:10:12 -05:00
Aiden Cline 7b5d9aa645 Merge pull request #1025 from liuchang-reolink/dev
add qwen3.5-397b-a17b and step-3-5-flash for nvidia
2026-03-09 14:05:08 -05:00
Aiden Cline 6e0040dbfd Merge pull request #1089 from Krule/krule/update_gitlab_anthropic_context_size
feat(gitlab): update context limit to 1M for Claude Sonnet and Opus 4.6
2026-03-09 14:03:49 -05:00
Aiden Cline 943ad8481b Merge pull request #1121 from illusion77/fix/chutes-mimo-v2-flash-context-16709
fix(chutes): correct MiMo-V2-Flash context window and capabilities
2026-03-09 14:02:51 -05:00
Aiden Cline 7f1b6fb0eb Merge pull request #1122 from riccardogiorato/dev
remove deprecated kimi models from together.ai
2026-03-09 14:02:36 -05:00
Riccardo Giorato 23eff95e5d remove deprecated kimi from together.ai 2026-03-09 17:30:40 +01:00
illusion77 ddbd396205 fix(chutes): correct MiMo-V2-Flash context window and capabilities
The chutes provider had incorrect metadata for MiMo-V2-Flash:
context 32K → 262K, output 8K → 32K, reasoning and tool_call enabled.

Fixes anomalyco/opencode#16709
2026-03-09 10:57:57 -05:00
Aiden Cline f3ee1a530b Merge pull request #1120 from stephenkuhn214/dev
Add Amazon-Bedrock Devstral 2 123B model
2026-03-09 09:35:46 -05:00
Aiden Cline 9c51b65440 Merge pull request #1119 from cgilly2fast/dev
fix(firmware): proper 5.3 codex model id
2026-03-09 09:30:35 -05:00
Frank 353aeb4998 update zen models 2026-03-09 10:08:55 -04:00
Frank 11991fecb5 update zen models 2026-03-09 10:03:13 -04:00
stephenkuhn214 1b4599773d Create mistral.devstral-2-123b 2026-03-09 08:58:19 -04:00
Colby Gilbert 78e1a3b0c9 fix(firmware): proper 5.3 codex model id 2026-03-08 21:58:27 -07:00
Sewer56 7a02946620 Update synthetic models: promote MiniMax-M2.5, add GLM-4.7-Flash, remove deprecated Qwen3.5 2026-03-08 22:56:31 +00:00
Aiden Cline 44686797c8 Merge pull request #1118 from shelvick/add-azure-gpt-5.3-chat
Add GPT-5.3 Chat to Azure
2026-03-08 16:41:10 -05:00
Aiden Cline 065cec8431 fix: input limit for context 2026-03-08 16:40:38 -05:00
Scott Helvick f491c2bec9 Add GPT-5.3 Chat to Azure 2026-03-08 21:20:27 +00:00
Aiden Cline cf1ac3053f Merge pull request #1081 from djmaze/fix/nebius-model-casing
fix(nebius): correct model ID casing to match Token Factory API
2026-03-08 14:26:52 -05:00
Aiden Cline 6be1e929fc Merge pull request #1114 from v1gnesh/dev
add grok 4.2 experimentals
2026-03-08 14:25:12 -05:00
Aiden Cline 49524827e2 Merge pull request #1113 from shelvick/add-vertex-glm-5
Fix GLM-5 context window size on Google Vertex
2026-03-08 10:31:05 -05:00
Aiden Cline 5ab5d389fc Merge pull request #1112 from cau1k/feat/az-5.4
feat(azure): add gpt-5.4/5.4-pro
2026-03-08 10:30:54 -05:00
Aiden Cline d5367ed978 Merge pull request #1116 from xiaojiezj/xj_dev_0308
fix: Adjust the logo  for ZenMux
2026-03-08 10:30:18 -05:00
Aiden Cline 5069faa25b Merge pull request #1117 from kailiu42/feat/siliconflow-cn
feat(siliconflow-cn): add Qwen3.5 model family
2026-03-08 10:29:48 -05:00
Kai Liu b279f33d9b feat(siliconflow-cn): add Qwen3.5 model family
New models:

- Qwen/Qwen3.5-4B
- Qwen/Qwen3.5-9B
- Qwen/Qwen3.5-27B
- Qwen/Qwen3.5-35B-A3B
- Qwen/Qwen3.5-122B-A10B
- Qwen/Qwen3.5-397B-A17B

Signed-off-by: Kai Liu <kraml.liu@gmail.com>
2026-03-08 20:07:40 +08:00
xiaojie.zj fe8249d706 fix: Adjust the logo 2026-03-08 16:36:01 +08:00
skywalker512 236af40da3 feat: add Tencent Coding Plan provider
Add support for Tencent Coding Plan with 8 models:
- Auto (tc-code-latest)
- Hunyuan 2.0 Instruct
- Hunyuan 2.0 Think
- Hunyuan-T1
- Hunyuan-TurboS
- MiniMax-M2.5
- Kimi-K2.5
- GLM-5

Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>
2026-03-08 15:34:36 +08:00
v1gnesh a24a23d57b add grok 4.2 experimentals 2026-03-08 07:42:09 +05:30
Scott Helvick 7e4773d9b5 Fix GLM-5 context window size on Google Vertex
Correct the context limit from 204800 to 202752 tokens.
2026-03-07 22:20:56 +00:00
zero 9f937f3fc5 Merge branch 'dev' into feat/az-5.4 2026-03-07 17:20:29 -05:00
cau1k 8cbdbc1102 add day cutoff 2026-03-07 17:19:15 -05:00
cau1k f1ca3b0015 feat(azure-cognitive-services): symlink 5.4/pro from azure provider
;
2026-03-07 17:02:26 -05:00
cau1k 35c757bad5 feat(azure): add 5.4/pro 2026-03-07 17:00:43 -05:00
fhennerkes 78781f3901 poe: add GPT-5.4, GPT-5.4-Pro, and GPT-5.3-Instant models 2026-03-07 13:58:35 -08:00
Sylvie Zhang 26465319d6 Merge branch 'dev' into add-vercel-input-context-and-new-models 2026-03-07 11:39:53 -08:00
fhennerkes 1371cbf9de poe: add GPT-5.4, GPT-5.4-Pro, and GPT-5.3-Instant models 2026-03-07 09:48:17 -08:00
Aiden Cline 559ccd6966 Merge pull request #1024 from yinxulai/feat/qiniu-ai
feat(qiniu-ai): add new model configurations
2026-03-07 11:26:01 -06:00
Aiden Cline 2691cb4e8d Merge pull request #1083 from samzong/feat/add-drun-provider
feat: add d.run(China) provider (OpenAI-compatible)
2026-03-07 11:25:12 -06:00
Aiden Cline 83ed1f0125 Merge pull request #1015 from RioPlay/dev
add: newer MiniMax, GLM, and Kimi models to DeepInfra
2026-03-07 11:24:53 -06:00
Aiden Cline 869f831466 Merge branch 'dev' into dev 2026-03-07 11:23:37 -06:00
Aiden Cline 5a673af2ae Merge pull request #1061 from JonasGao/dev
Add Qwen3.5 Flash & GLM-5 & M2.5 models to alibaba-cn
2026-03-07 11:22:56 -06:00
Aiden Cline 9b8543a074 Add interleaved section to minimax-m2.5.toml 2026-03-07 11:21:54 -06:00
Aiden Cline b3fb902331 Merge pull request #1030 from Mingholy/feat/alibaba-coding-plan-cn
feat(alibaba-coding-plan-cn): add Coding Plan provider for China region
2026-03-07 11:21:26 -06:00
Aiden Cline adb0c0b305 Merge pull request #1062 from viitana/bump-deepseek-details
feat: [deepseek]: update official DeepSeek model details
2026-03-07 11:21:21 -06:00
Aiden Cline ed01410d82 Merge pull request #1088 from mcowger/feature/gemini-3.1-flash-lite
feat: add gemini-3.1-flash-lite-preview model
2026-03-07 11:13:45 -06:00
Aiden Cline 7e23b780cc Merge pull request #1077 from evroc-oss/evroc/correct-model-config
fix(evroc): correct model config
2026-03-07 11:13:15 -06:00
Aiden Cline c46b652c8e Merge pull request #1076 from jerome-benoit/feat/add-sonar-deep-research-sap-ai-core
feat(sap-ai-core): add Perplexity Sonar Deep Research model
2026-03-07 11:11:29 -06:00
Aiden Cline ddb74e9b09 Merge pull request #1063 from dpuyosa/fix/models-pricing-limits-update
Venice: Update model pricing and limits
2026-03-07 11:10:37 -06:00
Aiden Cline face36ecb8 Merge pull request #1064 from BlockListed/fix-cortecs-models
Fix Cortecs models
2026-03-07 11:10:03 -06:00
Aiden Cline 6f170651b3 Merge pull request #1075 from Track07-cda/alibaba-cn-third-party-models
Add third party providers' models to alibaba-cn provider
2026-03-07 11:09:40 -06:00
Aiden Cline 47dbe45dd5 Merge pull request #1066 from dpuyosa/feat/add-qwen3-5-35b-a3b
Venice: Add Qwen 3.5 35B A3B model
2026-03-07 11:09:10 -06:00
Aiden Cline 0ee43b64b3 Merge branch 'dev' into alibaba-cn-third-party-models 2026-03-07 11:08:37 -06:00
Aiden Cline 6130a1f74e Merge pull request #1068 from MauroDruwel/dev
NVIDIA: Add MiniMax M2.5 model and remove MiniMax M2
2026-03-07 11:07:17 -06:00
Aiden Cline c0c82a5f04 Merge pull request #1072 from sylviezhang37/update-vercel-models-20260302-1656
Update Vercel models
2026-03-07 11:06:17 -06:00
Aiden Cline b0ba8b14d5 Merge pull request #1092 from janszypulski/cloudferro-sherlock-add-minimax-2.5
add MiniMaxAI/MiniMax-M2.5 to CloudFerro Sherlock
2026-03-07 11:02:11 -06:00
Aiden Cline 4780f9ddc1 Merge pull request #1109 from dinhkim/feat/add-cf-glm-4.7-flash
feat: add GLM-4.7-Flash to the Cloudflare Workers AI provider
2026-03-07 11:01:50 -06:00
Aiden Cline 27e02de632 Merge pull request #1078 from SomeoneWithOptions/dev
add gpt 5.3 codex for openrouter and Mercury models
2026-03-07 11:01:41 -06:00
Aiden Cline f22c827045 Merge branch 'dev' into dev 2026-03-07 11:01:17 -06:00
Aiden Cline cfc4585ed7 Merge pull request #1107 from Rinuuri/deepinfra-glm5
Add deepinfra GLM-5 model
2026-03-07 10:59:16 -06:00
Aiden Cline fa07bc2088 Merge pull request #1039 from rholak/add-abacus-models
Add sonnet 4.6 and opus 4.6 to abacus model list
2026-03-07 10:59:00 -06:00
Aiden Cline 497b1daaf2 Merge pull request #1103 from dpuyosa/feat/venice-add-gpt-models
Venice: Add OpenAI GPT-4o, GPT-4o Mini, GPT-5.4 models
2026-03-07 10:58:44 -06:00
Aiden Cline 442afa8c7e Merge pull request #1060 from yanismiraoui/inception/mercury2
Add Inception Mercury 2 and Mercury Edit models
2026-03-07 10:57:38 -06:00
Aiden Cline 4bd0c387fe Merge pull request #1044 from shrwnsan/feat/openrouter-routers
feat(openrouter/free): add free router
2026-03-07 10:57:24 -06:00
Aiden Cline b8c0c1d3a1 Merge pull request #1053 from laiiihz/update-xiaomi-models
Update Xiaomi models metadata
2026-03-07 10:57:17 -06:00
Aiden Cline 5c6c3e5a32 Merge pull request #1055 from shantanugoel/gemini-3.1-flash-image-preview
Add Gemini 3.1 Flash Image Preview
2026-03-07 10:57:06 -06:00
Aiden Cline 53d3cca3a0 Merge pull request #1052 from spiffytech/dev
Improve Ollama Cloud generator. Remove Gemini 3 Pro from Ollama Cloud.
2026-03-07 10:56:45 -06:00
Aiden Cline 105970c173 Merge pull request #1049 from heimoshuiyu/fix/glm-5-open-weights
fix: mark GLM-5 as open weights
2026-03-07 10:56:31 -06:00
Aiden Cline 788ee04034 Merge pull request #1045 from xinrui-z/aihubmix-add-models
aihubmix add models
2026-03-07 10:56:03 -06:00
Aiden Cline 6626db4044 Merge pull request #1098 from JWahle/dev
chore: updated abacus model definitions
2026-03-07 10:55:31 -06:00
Aiden Cline 8902640664 Merge pull request #1023 from PandaSt0rm/add-alibaba-coding-plan
Add Alibaba Coding Plan provider and model configs
2026-03-07 10:53:34 -06:00
Kim Truong cab247ddf8 update context to match Cloudflare doc 2026-03-07 23:50:06 +07:00
Kim Truong c1a42fa0a0 feat: add GLM-4.7-Flash mode in Cloudflare Workers AI provider 2026-03-07 23:45:52 +07:00
Aiden Cline 35023bba5a Merge pull request #1001 from ItsWendell/feat/bedrock-bearer-token
Add AWS_BEARER_TOKEN_BEDROCK to Amazon Bedrock provider env
2026-03-07 09:52:21 -06:00
Aiden Cline 604e49792b Merge pull request #1002 from DEAN-Cherry/feat/add-minimax-m2.5
models: alibaba-cn: add MiniMax-M2.5
2026-03-07 09:51:36 -06:00
Aiden Cline 0ca77b0cda Merge branch 'dev' into dev 2026-03-07 09:50:26 -06:00
Aiden Cline ea9505a40f Merge pull request #1004 from BlockListed/cortecs-models
Add Cortecs AI models
2026-03-07 09:50:07 -06:00
Aiden Cline 990b8d7308 Merge pull request #1005 from cgilly2fast/dev
feat(firmware): gemini 3.1 pro, sonnet reasoning
2026-03-07 09:49:54 -06:00
Aiden Cline ec173e86d4 Merge pull request #996 from fhennerkes/dev
poe: add Gemini-3.1-Pro, GPT-5.3-Codex and Gemini 3.1 Flash Lite
2026-03-07 09:47:52 -06:00
Aiden Cline 4a6e92a7c9 Merge pull request #997 from xiaojiezj/zenmux_dev_0221
feat: add Gemini 3.1 Pro Preview for ZenMux provider
2026-03-07 09:47:37 -06:00
Aiden Cline f0f686bdf5 Merge pull request #999 from mikalsande/mistral_latest
Append (latest) to Mistral models that refer to the latest version.
2026-03-07 09:46:40 -06:00
Aiden Cline 35ff0c2629 Merge pull request #995 from Phoen1xCode/dev
fix(zenmux:minimax): remove duplicated prefix & feat(zenmux:openai): add GPT-5.2-Pro model
2026-03-07 09:45:07 -06:00
Aiden Cline f99e9e89df Merge pull request #1090 from litvix-whale/feat/add-minimax-m2-5
feat(provider): add MiniMax M2.5 for DeepInfra
2026-03-07 09:41:28 -06:00
Armin Pašalić 09722ac264 Merge branch 'anomalyco:dev' into krule/update_gitlab_anthropic_context_size 2026-03-07 13:17:10 +01:00
Rinuuri fa67d00aeb Update GLM-5.toml 2026-03-06 21:23:42 +00:00
Rinuuri ddd2dd73ed Adding deepinfra GLM-5 2026-03-07 00:03:29 +03:00
fhennerkes d7929fd00b Merge branch 'anomalyco:dev' into dev 2026-03-06 12:00:20 -08:00
Frank 06e7d4db42 Merge pull request #1014 from NachoFLizaur/fix/bedrock-opus-4-6-context-window
fix(amazon-bedrock): correct Claude Opus 4.6 context window from 1M to 200K
2026-03-06 11:25:37 -05:00
Sylvie Zhang 7a11ef241d update more models 2026-03-06 08:24:49 -08:00
Sylvie Zhang 145862315d add input calculation + new models 2026-03-06 08:07:11 -08:00
dpuyosa d871710ba4 [venice] Add OpenAI GPT-4o, GPT-4o Mini, GPT-5.4 models
- Add gpt-4o-2024-11-20 model configuration
- Add gpt-4o-mini-2024-07-18 model configuration
- Add gpt-5.4 model configuration with reasoning capability
2026-03-06 09:53:06 +01:00
Colby Gilbert 16486087c6 Merge branch 'anomalyco:dev' into dev 2026-03-05 21:38:25 -08:00
Frank 2939af9330 Merge pull request #1100 from sachnun/feat/github-copilot-gpt-5-4
feat(provider): add gpt-5.4 for GitHub Copilot
2026-03-05 23:33:57 -05:00
sachnun 7c68dab3bb feat(provider): add gpt-5.4 for GitHub Copilot 2026-03-06 11:18:11 +07:00
Mike Soylu caceb0b310 openrouter openai models (#1099) 2026-03-05 22:26:58 -05:00
Frank 7a0d3be1e7 Update zen models 2026-03-05 18:55:49 -05:00
ShivamB25 e11ad7c01a feat(openai): add GPT-5.4 and GPT-5.4 Pro model specs (#1095) 2026-03-05 18:50:22 -05:00
Matt Silverlock d30fa82e4c Cloudflare: add gpt-5.4.toml (#1096) 2026-03-05 18:50:10 -05:00
Rishi Vhavle 771102a960 feat: add gpt-5.3-codex to github-copilot provider (#1097) 2026-03-05 18:49:56 -05:00
JWahle 30f98b15ef chore: updated abacus model definitions
Added: GPT-5 Codex, GPT-5.1/5.2/5.3 Codex, GPT-5.3 Chat, Gemini 3.1 Flash Lite/Pro Preview, Claude Opus/Sonnet 4.6, Kimi K2.5, GLM-5
Removed: Gemini 2.0 Flash 001, Gemini 2.0 Pro Exp, Meta-Llama 3.1 70B Instruct
Updated pricing: DeepSeek V3.1, GLM-4.7, GPT-5.2 Chat Latest, o3-pro, Route LLM
2026-03-06 00:46:19 +01:00
Colby Gilbert 6f7ab479fb feat(firmware): gpt 5.4 2026-03-05 13:23:08 -08:00
Colby Gilbert a4efbcd5ce Merge branch 'anomalyco:dev' into dev 2026-03-05 13:15:49 -08:00
Frank bcbfba03bd update zen models 2026-03-05 15:51:33 -05:00
Frank bdb5dac941 update zen models 2026-03-05 15:50:03 -05:00
Frank 1538bdcedb update zen models 2026-03-05 13:31:27 -05:00
SomeoneWithOptions e5211f3105 add inception mercury models for openrouter 2026-03-05 12:13:52 -05:00
Andres Castellanos 4a2209dbd4 Merge branch 'anomalyco:dev' into dev 2026-03-05 11:51:38 -05:00
Jan Szypulski 900014fe52 add MiniMax-M2.5 2026-03-05 14:58:19 +01:00
Kyrylo Lytvishko 5bbaf3c3f2 feat(provider): add MiniMax M2.5 for DeepInfra 2026-03-05 14:04:11 +02:00
Armin Pasalic 7dd0a26ff4 feat(gitlab): update context limit to 1M for Sonnet and Opus 4.6 2026-03-05 12:01:02 +01:00
Matt Cowger 3b7e0f02f1 feat: add gemini-3.1-flash-lite-preview model 2026-03-04 13:19:21 -08:00
samzong e67f921ea3 feat: add official d.run logo 2026-03-04 13:42:01 +08:00
samzong f5411eeeda feat: add D.Run (China) provider with minimax-m25, deepseek-r1, deepseek-v3 2026-03-04 13:33:04 +08:00
Frank 0d83ab8909 Merge pull request #1082 from kesku/kesku/add-ppl-agent-api
Add Perplexity Agent API provider
2026-03-03 23:03:49 -05:00
Kesku 26a629debc add models 2026-03-03 23:19:01 +00:00
Kesku 4b3319561b set up provider 2026-03-03 23:10:43 +00:00
Ubuntu b89ce0d986 fix(nebius): correct model ID casing to match Token Factory API
Fix lowercase model ID bug that caused "The model does not exist" errors.

- qwen/ → Qwen/ directory
- Fixed model file casing to match API exactly across all providers
2026-03-03 22:11:57 +00:00
fhennerkes 1c01f8172b poe: add Gemini-3.1-Flash-Lite and update gpt-4o-mini context
Add new Gemini 3.1 Flash Lite model
Update gpt-4o-mini context window: 128K → 124,096
2026-03-03 11:24:20 -08:00
SomeoneWithOptions d76040c514 add gpt 5.3 codex for openrouter 2026-03-03 12:53:39 -05:00
Simon Rygård feffa8119f fix(evroc): correct modality config 2026-03-03 16:52:07 +01:00
Simon Rygård 59c6e5df62 fix(evroc): correct tool call config 2026-03-03 16:51:47 +01:00
Jérôme Benoit fe2204d42c feat(sap-ai-core): add Perplexity Sonar Deep Research model 2026-03-03 14:56:18 +01:00
Track07-cda 07cc5335ac Add third party providers' models to alibaba-cn provider
- Add `MiniMax/MiniMax-M2.5` and `kimi/kimi-k2.5` to the `alibaba-cn`
  provider.
- Update `kimi-k2.5` to include video modality and adjust release/update
  dates.
- Add several `siliconflow/deepseek` models to the `alibaba-cn`
  provider.
2026-03-03 16:46:32 +08:00
github-actions[bot] fefbb90a29 chore(vercel): update Vercel model definitions
Auto-generated by weekly workflow from Vercel AI Gateway API.

Co-Authored-By: github-actions[bot] <41898282+github-actions[bot]@users.noreply.github.com>
2026-03-02 16:56:48 +00:00
Frank fec48b83d3 update zen models 2026-03-01 13:23:08 -05:00
Mauro Druwel c30bbe7718 Add knowledge 2026-03-01 08:53:22 +01:00
Mauro Druwel 1fba668f0f Add minimax-m2.5 to nvidia-nim and remove deprecated minimax-m2 from nvidia-nim 2026-03-01 08:52:33 +01:00
Aiden Cline 33ec088bda Merge pull request #1008 from friendliai/feat/friendli-minimax-m2.5
add friendli minimax m2.5 model config
2026-03-01 07:54:08 +05:00
Aiden Cline add7f9a914 Merge pull request #1065 from friendliai/minpeter/remove-exaone-models
Remove all EXAONE models
2026-03-01 07:53:42 +05:00
dpuyosa 369fa2de6d [venice] Add Qwen 3.5 35B A3B model
- Add new model configuration for Qwen 3.5 35B A3B
- Includes cost, limits, and capabilities (reasoning, tool_call, structured_output)
2026-02-28 21:07:28 +01:00
minpeter c8732e7e74 Remove all EXAONE models
Remove LGAI-EXAONE model definitions (EXAONE-4.0.1-32B, K-EXAONE-236B-A23B)
and related family references from core packages.

Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>
2026-03-01 04:48:08 +09:00
Jonas f00f9f3c11 Add Qwen3.5 Flash & GLM-5 & M2.5 models to alibaba-cn provider 2026-02-28 23:34:06 +08:00
BlockListed c87ca238de fix cortecs models
the should have periods not a p as a decimal separator
2026-02-28 15:16:59 +01:00
BlockListed 102f55aeef add glm 4.7 flash model to cortecs 2026-02-28 15:13:53 +01:00
BlockListed 2767754a02 add kimi K2.5 model to cortecs 2026-02-28 15:13:53 +01:00
dpuyosa 4289b59a04 [models] Update model pricing and limits
- Update Claude Sonnet 4-6 pricing and output limit
- Update Grok 41 Fast pricing and context/limits
2026-02-28 14:24:16 +01:00
Atte Viitanen b986313f42 feat: [deepseek]: update official deepseek model details 2026-02-28 13:22:26 +02:00
yanismiraoui 28cfd4cab6 naming mercury 2 and mercury edit for inception provider 2026-02-27 17:45:01 -08:00
yanismiraoui b93e62fc3a Add Inception Mercury 2 and Mercury Edit models 2026-02-27 17:41:00 -08:00
Aiden Cline e23b5ab010 Merge pull request #1026 from ryot/venice
Venice: Add GPT-5.3 Codex
2026-02-28 06:30:13 +05:00
Aiden Cline b5f6024868 Merge pull request #1020 from jerome-benoit/feat/sap-ai-core-add-models
feat(sap-ai-core): Add GPT-4.1, Gemini 2.5 Flash Lite, Perplexity Sonar, and Claude 4.6 models
2026-02-28 06:29:28 +05:00
Aiden Cline 07db15e984 Merge pull request #1029 from dpuyosa/veniceScript
Venice: Remove interactive API key prompt & use new maxCompletionTokens field
2026-02-28 06:28:33 +05:00
Aiden Cline 74abf8851a Merge pull request #1042 from SomeoneWithOptions/dev
add gemini 3.1 pro preview custom tools for openrouter
2026-02-28 06:27:56 +05:00
Aiden Cline 7e13ecdfd9 Merge pull request #1056 from xezpeleta/fix/azure-gpt-5-3-codex
fix(azure): add gpt-5.3-codex model
2026-02-28 06:27:34 +05:00
Aiden Cline 6ad2c28b2d Merge pull request #1048 from dpuyosa/feat/add-venice-models
Venice: Add NVIDIA Nemotron 3 Nano and Qwen 3 Coder Turbo models
2026-02-28 06:27:20 +05:00
Frank a124036692 update zen models 2026-02-27 16:16:37 -05:00
Xabi Ezpeleta d37d362cc8 fix(azure): add gpt-5.3-codex model 2026-02-27 16:41:11 +01:00
Shantanu Goel c387f94c8e Add Gemini 3.1 Flash Image Preview 2026-02-27 20:03:41 +05:30
laiiihz 45457c34d8 update xiaomi models detail 2026-02-27 14:56:16 +08:00
spiffytech 44774ec3d6 Ollama Cloud removed support for Gemini 3 Pro 2026-02-26 17:18:34 -05:00
spiffytech c8fdcf80dd Updated Ollama Cloud generator to delete old models, only write out files if they changed 2026-02-26 17:18:33 -05:00
fhennerkes 9d33b6409c Merge branch 'anomalyco:dev' into dev 2026-02-26 12:04:52 -08:00
Matt Silverlock c76586a174 Cloudflare: add codex models to AI Gateway (#1050)
* add gpt-5.2-codex

* add gpt-5.3-codex

* Update gpt-5.2-codex.toml

* Update gpt-5.3-codex.toml
2026-02-26 14:41:12 -05:00
Jérôme Benoit 2a267614aa feat(sap-ai-core): add Claude Opus 4.6 and Sonnet 4.6 models 2026-02-26 17:58:02 +01:00
PandaSt0rm aac62378b2 Update MiniMax-M2.5 guidance per Alibaba docs 2026-02-26 17:20:58 +02:00
David Hill 56cc5f71bf fix(ui): opencode zen logo update 2026-02-26 11:09:25 +00:00
David Hill df2c87d32a fix(ui): opencode go logo 2026-02-26 11:09:13 +00:00
heimoshuiyu ff41c2b6c3 fix: mark GLM-5 as open weights
GLM-5 is an open-source model, but several provider config files
incorrectly had open_weights set to false. This commit corrects
all GLM-5 configurations to properly reflect its open-source status.

Affected providers:
- zhipuai
- zhipuai-coding-plan
- zai
- zai-coding-plan
- zenmux
- vercel
- siliconflow
- siliconflow-cn
- meganova
2026-02-26 18:43:28 +08:00
dpuyosa 1d137e2f1f [venice] Add NVIDIA Nemotron 3 Nano and Qwen 3 Coder models
- Add NVIDIA Nemotron 3 Nano 30B A3B model configuration
- Add Qwen 3 Coder 480B A35B Instruct Turbo model configuration
2026-02-26 10:53:00 +01:00
dpuyosa 16720bcd1a [venice] Use maxCompletionTokens for output limit
- Add optional maxCompletionTokens field to model spec schema
- Use maxCompletionTokens when calculating output token limit instead of checking existing limit
2026-02-26 10:24:43 +01:00
Xinrui 1feaf76749 aihubmix add models 2026-02-26 16:12:51 +08:00
shrwnsan 080ef5cc9e fix(openrouter): remove auto router and add missing limit.input
- Remove auto router (cost varies, doesn't fit schema)
- Add limit.input = 200_000 to free.toml (schema requirement)

OpenRouter's auto router has 'pricing varied' - it charges based on the
routed model. This doesn't fit the numeric cost schema required by
models.dev, so we're removing it. The free router is retained as it
genuinely costs $0.
2026-02-26 14:40:38 +08:00
Ryo Tulman f8121c8dc3 Update Venice GPT 5.3 Codex output limit 2026-02-26 00:32:13 -06:00
shrwnsan d2d5c5a7cc feat: add openrouter free and auto routers 2026-02-26 10:51:55 +08:00
SomeoneWithOptions 09d9e91d83 add gemini 3.1 pro preview custom tools for openrouter 2026-02-25 15:13:43 -05:00
Robert Holak 930d6a94b8 Add sonnet 4.6 and opus 4.6 to abacus model list 2026-02-25 12:31:06 -06:00
mulder b9217aff8e Add Clarifai Model Provider
Add Clarifai as a new provider with 11 models:
- GPT OSS 20B, GPT OSS 120B High Throughput
- Ministral 3 14B/3B Reasoning 2512
- Qwen3 Coder 30B, Qwen3 30B Instruct/Thinking 2507
- MiniMax-M2.5 High Throughput
- Trinity Mini, DeepSeek OCR, MM Poly 8B

Also adds 'mm-poly' family to family.ts for the Clarifai multimodal model.
2026-02-25 12:28:40 -05:00
Lucas Almeida c240bce614 fix: adding missing structured_output parameter 2026-02-25 11:19:54 -03:00
Lucas Almeida 843a1d182a feat: adding gpt-5.3-codex for Azure Foundry 2026-02-25 11:09:10 -03:00
PandaSt0rm 443c06ca03 fix MiniMax M2.5 modalities in Alibaba Coding Plan
- set MiniMax-M2.5 input modalities to text-only
- keep output modality as text
- validate with bun validate
2026-02-25 13:10:58 +02:00
PandaSt0rm 84466021ca add MiniMax M2.5 to Alibaba Coding Plan and align third-party limits
- add MiniMax-M2.5 model config under providers/alibaba-coding-plan/models
- update GLM-4.7 limits to 202,752 context / 16,384 output
- update GLM-5 limits to 202,752 context / 16,384 output
- update Kimi K2.5 output limit to 32,768
- validate with bun validate
2026-02-25 13:05:18 +02:00
mingholy.lmh b995e90cf5 fix: update context and output limits for alibaba-coding-plan-cn models
Update model limits:
- qwen3-coder-plus: context 1_048_576 → 1_000_000
- glm-5: output 131_072 → 16_384
- glm-4.7: output 131_072 → 16_384
- kimi-k2.5: output 65_536 → 32_768

Co-authored-by: Qwen-Coder <qwen-coder@alibabacloud.com>
2026-02-25 17:44:49 +08:00
dpuyosa 291e2eefe9 [venice] Remove interactive API key prompt
- Remove readline import and promptForApiKey function
- Remove prompt fallback, rely on CLI arg or env var only
- Update README to reflect change
2026-02-25 09:51:50 +01:00
Sewer56 0428299773 Added: Qwen3.5-397B natively supports image, MM2.5 No Image as it was a mistake. 2026-02-25 08:10:42 +00:00
Frank 96e9537b34 update zen models 2026-02-25 01:05:35 -05:00
Ryo Tulman 09dc7060ac Venice: Add GPT-5.3 Codex 2026-02-24 23:37:03 -06:00
Colby Gilbert a463717783 chore(firmware): remove gpt-5 2026-02-24 20:33:49 -08:00
Colby Gilbert 6d721dd32d feat(firmware): gpt-5.3-codex 2026-02-24 20:32:17 -08:00
liuchang-reolink 3ae513785a add step-3-5-flash for nvidia
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode)

Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai>
2026-02-25 12:05:19 +08:00
Colby Gilbert e8d667b628 Merge branch 'anomalyco:dev' into dev 2026-02-24 20:01:26 -08:00
liuchang-reolink 7616a65e63 add qwen3.5-397b-a17b for nvidia
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode)

Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai>
2026-02-25 11:40:15 +08:00
yinxulai a5b3da12c1 chore(qiniu-ai): update provider config 2026-02-25 10:20:55 +08:00
yinxulai a05fbc3604 feat(qiniu-ai): add new model configurations 2026-02-25 10:16:40 +08:00
Aiden Cline 2189030e57 Merge pull request #1022 from armishra/feat/add-minimax-m2.5-baseten
feat(provider): Add MiniMax-M2.5 for baseten
2026-02-24 17:28:04 -06:00
Aiden Cline c7ecc08442 Merge pull request #1019 from dpuyosa/venice
Venice: Update gemini-3-1-pro-preview config
2026-02-24 17:27:46 -06:00
Aiden Cline 830046e45e Merge pull request #1021 from sylviezhang37/update-vercel-models-20260224-2134
Update Vercel models
2026-02-24 17:27:19 -06:00
Aiden Cline c7b26477b9 Update cache_read value in gemini-3.1-pro-preview.toml 2026-02-25 04:26:56 +05:00
Aiden Cline 9e60f516fa Update cost input and output values in TOML file 2026-02-25 04:26:18 +05:00
PandaSt0rm 7ccbb58c5f add Alibaba Coding Plan provider and model configs 2026-02-25 01:10:42 +02:00
Archit Mishra da906a0816 feat(provider): Add MiniMax-M2.5 for baseten 2026-02-24 14:33:41 -08:00
github-actions[bot] 36c9f82905 chore(vercel): update Vercel model definitions
Auto-generated by weekly workflow from Vercel AI Gateway API.

Co-Authored-By: github-actions[bot] <41898282+github-actions[bot]@users.noreply.github.com>
2026-02-24 21:34:20 +00:00
Frank 978214e143 update zen models 2026-02-24 15:23:24 -05:00
Jérôme Benoit 6aa69460a1 feat(sap-ai-core): add GPT-4.1, Gemini 2.5 Flash Lite, and Sonar models
Add 5 new model definitions for SAP AI Core provider:
- gpt-4.1: OpenAI GPT-4.1 (1M context, 32K output)
- gpt-4.1-mini: OpenAI GPT-4.1 Mini (1M context, 32K output)
- gemini-2.5-flash-lite: Google Gemini 2.5 Flash Lite (1M context, 65K output)
- sonar: Perplexity Sonar (128K context, 4K output)
- sonar-pro: Perplexity Sonar Pro (200K context, 8K output)

All specs verified against official provider documentation.
2026-02-24 20:50:10 +01:00
fhennerkes 0f16fcf231 poe: add GPT-5.3-Codex model 2026-02-24 11:39:47 -08:00
fhennerkes e583c700f9 Merge branch 'anomalyco:dev' into dev 2026-02-24 11:36:20 -08:00
dpuyosa 96d278932c [venice] Update gemini-3-1-pro-preview config
- Reduce output token limit from 250K to 65K
2026-02-24 11:27:18 +01:00
Sewer56 eee3303df0 Add missing synthetic.new models
Add configuration for hf:Qwen/Qwen3.5-397B-A17B and hf:MiniMaxAI/MiniMax-M2.5
to the synthetic provider, based on API specs from synthetic.new.

Note: API reports image support but these models may not natively support
images (likely rerouted/proxied through vision-capable infrastructure).
2026-02-24 09:19:57 +00:00
RioPlay 838416044f add: newer MiniMax, GLM, and Kimi models to DeepInfra 2026-02-23 22:21:28 -06:00
Frank 51441f47d9 update zen models 2026-02-23 15:08:28 -05:00
Nacho F. Lizaur 7fc2c6154d fix(amazon-bedrock): correct Claude Opus 4.6 context window from 1M to 200K 2026-02-23 20:10:58 +01:00
Colby Gilbert 1439781a76 feat(firmware): add deepseek 3.2, glm 5, kimi k2.5, minimax m2.5 2026-02-22 21:29:51 -08:00
minpeter 8fc0d87742 add friendli minimax m2.5 model config 2026-02-23 13:18:17 +09:00
Colby Gilbert f660955784 feat(firmware): add grok models 2026-02-22 15:37:53 -08:00
Colby Gilbert eb11c327b8 feat(firmware): gemini 3.1 pro, sonnet reasoning 2026-02-21 23:43:25 -08:00
Bryan Nie 0dfde60c14 models: alibaba-cn: add MiniMax-M2.5 2026-02-22 01:09:24 +08:00
Wendell Misiedjan bab7727bad Add AWS_BEARER_TOKEN_BEDROCK to Amazon Bedrock provider env
The @ai-sdk/amazon-bedrock package supports Bearer token authentication
via the AWS_BEARER_TOKEN_BEDROCK environment variable as an alternative
to IAM SigV4 auth. This uses Bedrock API keys for simplified access.
2026-02-21 16:27:45 +01:00
Mikal Sande 4b4a2364c6 Append (latest) to Mistral models that refer to the latest version. 2026-02-21 09:19:13 +01:00
Frank c36b8e9433 update zen models 2026-02-20 23:20:24 -05:00
xiaojie.zj 5b8e983e7c feat: add Gemini 3.1 Pro Preview for ZenMux provider 2026-02-21 10:39:30 +08:00
Frank 0d2a52dd9d update zen models 2026-02-20 20:41:52 -05:00
Frank b667ab78ac update zen models 2026-02-20 20:19:33 -05:00
fhennerkes e2da96cde4 poe: add Gemini-3.1-Pro and update Claude Sonnet 4.6 2026-02-20 11:28:18 -08:00
Jake Jia 9d042ac986 Update providers/zenmux/models/openai/gpt-5.2-pro.toml
Co-authored-by: Copilot <175728472+Copilot@users.noreply.github.com>
2026-02-21 01:22:04 +08:00
Phoen1xCode 7192dc0ba8 feat(openai): add GPT-5.2-Pro model via zenmux provider
Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>
2026-02-21 01:07:58 +08:00
Phoen1xCode 05ea56a12a fix(minimax): remove duplicated provider prefix from MiniMax M2.5 Lightning name
Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>
2026-02-21 01:07:47 +08:00
Aiden Cline beb449a417 Merge pull request #994 from davidfph/fix/qwen3.5-release-date
fix(qwen): update Qwen3.5 release_date and last_updated to 2026-02-16
2026-02-20 10:29:15 -06:00
Aiden Cline 8f76f9b217 Merge pull request #988 from MeganovaAI/fix-meganova-logo
Update Meganova logo to official brand icon
2026-02-20 10:29:02 -06:00
Aiden Cline 18ddcde669 fix: azure & cognitive model distinctions 2026-02-20 10:28:22 -06:00
David Fu ff36ded35f fix(qwen): update Qwen3.5 release_date and last_updated to 2026-02-16 2026-02-20 20:33:08 +08:00
Aiden Cline b0b8074a94 Merge pull request #990 from kailiu42/dev
models: siliconflow-cn: add new models
2026-02-20 03:01:45 -06:00
Kai Liu 4c30a522f7 models: siliconflow-cn: add new models
New models per the latest list: https://cloud.siliconflow.cn/me/models

- Pro/MiniMaxAI/MiniMax-M2.5
- deepseek-ai/DeepSeek-OCR
- PaddlePaddle/PaddleOCR-VL
- PaddlePaddle/PaddleOCR-VL-1.5

Signed-off-by: Kai Liu <kraml.liu@gmail.com>
2026-02-20 16:26:26 +08:00
Aiden Cline 2d63d713de Merge pull request #992 from anomalyco/fix-azure-models
fix: ensure that anthropic models on azure providers have correct urls
2026-02-20 02:19:13 -06:00
Aiden Cline ac9d0af8e3 fixes 2026-02-20 02:13:23 -06:00
Aiden Cline a1ad90a9b1 Merge pull request #991 from zainhas/dev
[Together AI] add qwen3.5
2026-02-20 01:02:23 -06:00
Zain Hasan 8a6e0dd917 add qwen3.5 2026-02-19 21:38:46 -08:00
Aiden Cline 2da10b739c Merge pull request #989 from propilideno/fix/adding_missing_azure_foundry_model
Add missing GPT-5.2 metadata for Azure Cognitive Services
2026-02-19 18:47:56 -06:00
Aiden Cline c17e0b9d0f Merge pull request #987 from dpuyosa/venice
Venice: Add Gemini 3.1 Pro Preview and update model configs
2026-02-19 18:47:48 -06:00
Lucas Almeida 877a1175f4 chore: replacing by symbolic link like the other ones 2026-02-19 21:21:05 -03:00
Boqian 1bc83abeeb Update Meganova logo to official brand icon 2026-02-19 18:57:17 -05:00
dpuyosa 70caedba86 [venice] Add Gemini 3.1 Pro Preview and update model configs
- Add new Gemini 3.1 Pro Preview model configuration
- Update Claude Sonnet 4.6 release dates
- Enable open_weights for MiniMax M25
2026-02-19 22:54:59 +01:00
Aiden Cline 60c90a27a0 Merge pull request #985 from sylviezhang37/update-vercel-model-gen-script
feat(provider): exclude image/video models
2026-02-19 15:49:25 -06:00
Aiden Cline 2bd0d5446e Merge pull request #986 from riasvdv/add-gemini-3.1-pro
Add Gemini 3.1 Pro Preview to copilot models
2026-02-19 15:49:12 -06:00
Aiden Cline 5f135517b1 Remove audio and video from input modalities 2026-02-19 15:48:42 -06:00
Aiden Cline e2af7819b4 Rename gemini-3.5-pro-preview.toml to gemini-3.1-pro-preview.toml 2026-02-19 15:47:25 -06:00
Rias ca6c251b3a Add Gemini 3.1 Pro Preview to copilot models 2026-02-19 22:43:26 +01:00
Sylvie Zhang 7b1b590d10 exclude image/video gen models 2026-02-19 13:24:11 -08:00
Aiden Cline 5097a1e954 Merge pull request #966 from mhkok/mkok/feat/add-evroc-provider
add evroc provider + models
2026-02-19 14:15:51 -06:00
Aiden Cline 05959a83b6 Update font family in Kimi-K2.5 configuration 2026-02-19 14:15:07 -06:00
Aiden Cline 1492e067a4 Merge pull request #976 from too-green/patch-2
Add Qwen3 Coder Next model for openrouter
2026-02-19 14:01:07 -06:00
Aiden Cline 41c81535c3 fix: zen 2026-02-19 12:54:29 -06:00
Aiden Cline 829756fc41 Merge pull request #983 from mdrxy/mdrxy/fix-gemini-3
fix Gemini 3.1 model names
2026-02-19 12:34:07 -06:00
Mason Daugherty e6ef906c41 fix 2026-02-19 13:21:52 -05:00
Aiden Cline 0f84db6bc6 Merge pull request #975 from xiaojiezj/zenmux_dev_0219
feat:  Add new models for ZenMux provider
2026-02-19 11:33:49 -06:00
Aiden Cline 4bb6d52a7c Merge pull request #979 from hanouticelina/fix-interleaved-for-hf-provider
Fix Hugging Face interleaved `reasoning field: reasoning_details` -> `reasoning_content`
2026-02-19 11:33:34 -06:00
Aiden Cline bdd0194e73 Merge pull request #981 from mdrxy/mdrxy/add-gemini-3.1
add gemini 3.1 to google/openrouter
2026-02-19 11:33:15 -06:00
Frank 41a9502628 update zen models 2026-02-19 11:51:37 -05:00
Mason Daugherty 384e747129 add gemini 3.1 to google/openrouter 2026-02-19 11:24:30 -05:00
Frank e4bb5ceac6 update zen models 2026-02-19 10:16:51 -05:00
Frank c830964c3f update zen models 2026-02-19 09:37:07 -05:00
Celina Hanouti 782b6277ae Fix Hugging Face interleaved reasoning field 2026-02-19 15:27:49 +01:00
Frank e63d48ae9c update zen models 2026-02-19 07:42:52 -05:00
Matthijs Kok 029522aa96 fix family names 2026-02-19 08:48:43 +01:00
Ahmed 482ed2e833 Add Qwen3 Coder Next model for openrouter
Added model configuration for Qwen3 Coder Next
2026-02-19 12:36:55 +05:00
Aiden Cline c6635aa7c3 Merge pull request #968 from MeganovaAI/add-meganova-provider
Add Meganova as a provider
2026-02-18 23:44:03 -06:00
Aiden Cline a6ffef7e4f Merge pull request #969 from SomeoneWithOptions/dev
add claude sonnet 4.6 on openrouter
2026-02-18 23:40:02 -06:00
Aiden Cline fda9bb5335 Merge pull request #973 from sylviezhang37/update-vercel-models-20260219-0026
Update Vercel models
2026-02-18 23:38:56 -06:00
Aiden Cline 98d6901697 Update input cost value in qwen3.5-plus.toml 2026-02-18 23:38:49 -06:00
Aiden Cline 13ae499d7c tweak values 2026-02-18 23:38:18 -06:00
xiaojie.zj f21d205d6e feat: 增加Claude Sonnet 4.6/Doubao-Seed-2.0-lite/Doubao-Seed-2.0-mini/Doubao-Seed-2.0-pro模型 2026-02-19 11:10:28 +08:00
Lucas Almeida c82b08d778 fix: adding missing gpt-5.2 model on azure foundry 2026-02-18 23:38:10 -03:00
Sylvie Zhang ea612760cf Delete providers/vercel/models/recraft/recraft-v4.toml 2026-02-18 16:34:47 -08:00
Sylvie Zhang 48a64f0834 Delete providers/vercel/models/recraft/recraft-v4-pro.toml 2026-02-18 16:34:35 -08:00
github-actions[bot] 040e7fff4e chore(vercel): update Vercel model definitions
Auto-generated by weekly workflow from Vercel AI Gateway API.

Co-Authored-By: github-actions[bot] <41898282+github-actions[bot]@users.noreply.github.com>
2026-02-19 00:26:59 +00:00
SomeoneWithOptions a6d14928b6 added claude sonnet 4.6 on openrouter 2026-02-18 18:49:56 -05:00
Boqian 92d9e89690 Set reasoning=false for DeepSeek V3 series
V3-0324, V3.1, V3.2, V3.2-Exp are chat models, not reasoning models.
Only DeepSeek-R1 is a reasoning model.

Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>
2026-02-18 16:41:07 -05:00
Boqian 2d83b96bb1 Fix interleaved reasoning_content based on Meganova API testing
Tested each model with include_reasoning=true against the live API.

Added [interleaved] to: GLM-4.6, MiniMax-M2.1, MiniMax-M2.5, Kimi-K2.5
Removed [interleaved] from: DeepSeek-V3.1, V3.2, V3.2-Exp, MiMo-V2-Flash

Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>
2026-02-18 16:35:51 -05:00
Boqian 6e02385a6b Add interleaved reasoning_content to DeepSeek V3.1, V3.2, V3.2-Exp
These models support interleaved reasoning output, matching how other
providers (deepinfra, baseten, chutes) configure them.

Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>
2026-02-18 16:29:41 -05:00
Boqian a2f8234c8e Update pricing and context limits from Meganova API
Use actual pricing from https://api.meganova.ai/v1/models instead of
reference data from other providers.

Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>
2026-02-18 16:25:43 -05:00
Aiden Cline 2f43c70397 Merge pull request #967 from nicolasgere/dev
feat(provider): Add glm-5 for baseten
2026-02-18 15:19:47 -06:00
Boqian 006cb53c1e Add Meganova as a provider with 19 open-weight models
Adds Meganova AI (https://api.meganova.ai/v1) as an OpenAI-compatible provider
with curated open-weight models including DeepSeek, GLM, Qwen, Kimi, MiniMax,
MiMo, Llama, and Mistral families.

Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>
2026-02-18 16:17:51 -05:00
nicolasgere d2832f9d1d Update GLM-5.toml 2026-02-18 16:15:22 -05:00
nicolasgere 73c0fcc19e Rename GLM-5 to GLM-5.toml 2026-02-18 16:14:26 -05:00
nicolasgere 04a1fe8043 Create GLM-5 2026-02-18 16:14:09 -05:00
Matthijs Kok 7fa0eeebf9 add evroc provider + models 2026-02-18 19:57:44 +01:00
Aiden Cline cb72e1780f Merge pull request #962 from aldosch/add-sonnet-4-6-vercel
add sonnet 4.6 to vercel ai gateway
2026-02-18 12:19:00 -06:00
Aiden Cline 7ee7e88346 Merge pull request #960 from fa-sharp/patch-1
fix: OpenRouter output modalities for image-only models
2026-02-18 12:18:27 -06:00
Aiden Cline 4cc230ba6c Merge pull request #963 from vglafirov/gitlab/add-sonnet-4-6
feat(gitlab): add Claude Sonnet 4.6 model
2026-02-18 12:17:57 -06:00
Aiden Cline 65d6355ffa Merge pull request #964 from xinrui-z/aihubmix-add-claude-4-6
aihubmix: add models
2026-02-18 12:17:48 -06:00
Xinrui 1b6157c0ee aihubmix: add models 2026-02-18 22:55:03 +08:00
Vladimir Glafirov 5e25c0838e feat(gitlab): add Claude Sonnet 4.6 model 2026-02-18 13:52:26 -01:00
aldosch b06ab95c9e add sonnet 4.6 to vercel ai gateway 2026-02-19 00:38:35 +11:00
Daltonganger 33f4e77ee4 fix(nano-gpt): normalize family enums for model validation 2026-02-18 12:53:40 +01:00
Daltonganger 51ef2e51ae finalize nano-gpt model sync and release metadata 2026-02-18 12:44:36 +01:00
farshad 9d873cc764 add back trailing newline 2026-02-18 02:18:40 -05:00
farshad d0acb0a35d fix output modalities for black forest flux models 2026-02-18 01:52:31 -05:00
farshad 697c191fea Update output modalities in seedream-4.5.toml 2026-02-18 01:45:06 -05:00
Aiden Cline 7d9ff92ffd fix: glm 5 maas 2026-02-17 23:41:48 -06:00
Aiden Cline 03a2aee0de Merge pull request #716 from bluet/feat/google-vertex-openai
feat: add google-vertex-openai provider for Vertex AI partner models
2026-02-17 23:38:08 -06:00
Aiden Cline 4eb19459dd Merge pull request #958 from mugnimaestra/feat/add-qwen3.5-397b-a17b-tee-chutes
feat: add Qwen3.5 397B A17B TEE to Chutes provider listings
2026-02-17 23:27:18 -06:00
Aiden Cline b74d7eb495 Merge pull request #957 from dpuyosa/veniceScript
Venice: Update generate-venice script to include context_over_200k cost
2026-02-17 23:26:28 -06:00
Aiden Cline 60aa5aca14 Merge pull request #956 from dpuyosa/venice
Venice: Add Claude Sonnet 4.6 and GLM 4.7 Flash Heretic models
2026-02-17 20:22:44 -06:00
Muhammad Mugni Hadi 67235a4e39 feat: add Qwen3.5 397B A17B TEE to Chutes provider listings 2026-02-18 09:09:45 +07:00
dpuyosa ff03fd6906 [venice] Add Claude Sonnet 4.6 and GLM 4.7 models
- Add Claude Sonnet 4.6 model with context_over_200k pricing
- Add GLM 4.7 Flash Heretic model (open weights)
- Add context_over_200k pricing tier to Claude Opus 4.6
2026-02-18 02:02:19 +01:00
dpuyosa 423b177c2b Update generate-venice script to include context_over_200k cost 2026-02-18 01:55:52 +01:00
Aiden Cline 8c263109c5 Merge pull request #955 from maahir30/open-router-structured-output
Add structured output support for OpenRouter models
2026-02-17 18:52:02 -06:00
Maahir Sachdev 1bab7e8438 update open router models 2026-02-17 16:42:52 -08:00
Aiden Cline 7a163dbc60 Merge pull request #953 from mongrelion/dev
feat: add github copilot claude sonnet 4.6 model
2026-02-17 18:30:54 -06:00
Aiden Cline c94d025aa7 fixes 2026-02-17 18:23:01 -06:00
Aiden Cline 55e8be17b5 Merge pull request #949 from cgilly2fast/dev
feat(firmware): add sonnet 4.6
2026-02-17 18:21:16 -06:00
Aiden Cline 56096012b2 Merge pull request #948 from elithrar/patch-1
add Sonnet 4.6 model config
2026-02-17 17:53:20 -06:00
Aiden Cline 9b8a22d756 Merge pull request #725 from janszypulski/add-provider-cloudferro-sherlock
Add provider - Cloudferro Sherlock
2026-02-17 17:50:14 -06:00
Aiden Cline 5a778c6e93 Merge pull request #682 from the-lazy-me/add-qihang-provider
feat: add QiHang provider with 7 models
2026-02-17 17:48:51 -06:00
Aiden Cline 4d08659acf Merge pull request #651 from yinxulai/feat/qiniu-ai
feat: add Qiniu AI provider configuration
2026-02-17 17:46:05 -06:00
Aiden Cline 128c9ec469 Merge pull request #308 from d-oit/feature/perplexity-sonar-deep-research
Feature/perplexity sonar deep research
2026-02-17 17:37:20 -06:00
Carlos León dc11781324 feat: add github copilot claude sonnet 4.6 model
Model list sourced from GitHub Settings page showing currently available models. Specifications cross-referenced with Anthropic provider implementation.
2026-02-18 00:22:22 +01:00
Colby Gilbert 9d0b37bea3 feat(firmware): add sonnet 4.6 2026-02-17 15:03:21 -08:00
Matt Silverlock c8d09fe349 add Sonnet 4.6 model config 2026-02-17 17:27:51 -05:00
Aiden Cline 1a22b93fc2 Merge pull request #947 from fhennerkes/dev
poe: add Claude-Sonnet-4.6 and update XAI models
2026-02-17 15:56:11 -06:00
Aiden Cline 3918131cb8 Merge pull request #946 from monotykamary/remove-fireworks-deprecated-models-2026-02-12
chore(fireworks-ai): remove deprecated serverless models
2026-02-17 15:56:00 -06:00
fhennerkes a1d9c5134c poe: add Claude-Sonnet-4.6 and update XAI models 2026-02-17 13:38:54 -08:00
Ruben Beuker 20abb5b8df preserve curated release dates for key nano-gpt models
Keep existing curated release and last-updated values for models where NanoGPT API uses the generic created timestamp baseline.
2026-02-17 22:09:03 +01:00
Tom X Nguyen dc36ed54ae chore(fireworks-ai): remove deprecated serverless models
Remove 6 Fireworks serverless models deprecated on February 12, 2026:
- glm-4.6 (migrate to glm-4.7)
- deepseek-r1-0528 (migrate to deepseek-v3.2 or deepseek-v3.1)
- deepseek-v3-0324 (migrate to deepseek-v3.2 or deepseek-v3.1)
- qwen3-235b-a22b (migrate to kimi-k2-instruct-0905)
- qwen3-coder-480b-a35b-instruct (migrate to kimi-k2-instruct-0905)
- minimax-m2 (migrate to MiniMax-M2.1)

See: https://fireworks.ai/models?modelTypes=Serverless
2026-02-18 04:04:51 +07:00
Ruben Beuker 8cb462f29b sync nano-gpt models with live API catalog
Refresh NanoGPT model files to match the current /api/v1/models output, remove stale entries, and add newly available models while preserving path-based IDs.

Also ignore local TokenSpeed sqlite artifacts so private monitoring data is not shown or committed.
2026-02-17 22:01:54 +01:00
Frank 89486ec705 update zen models 2026-02-17 14:12:30 -05:00
Aiden Cline f313f802ee Merge pull request #940 from nitishxyz/add-claude-sonnet-4-6
feat(models): add Claude Sonnet 4.6 model configurations
2026-02-17 13:12:10 -06:00
nitishxyz 128615ddd7 feat(models): add Claude Sonnet 4.6 model configurations
- Add Claude Sonnet 4.6 to Anthropic provider with full capabilities
- Add regional variants (US, EU, Global) for Amazon Bedrock provider
- Add Google Vertex Anthropic provider configuration
- Define pricing, context limits (200k tokens), and modalities

Co-authored-by: ottocode-io[bot] <261994719+ottocode-io[bot]@users.noreply.github.com>
2026-02-18 00:01:02 +05:30
Aiden Cline 756fb772c1 Merge pull request #939 from Nomadcxx/fix/kilo-npm-provider
fix(kilo): use @ai-sdk/openai-compatible instead of opencode-kilo-auth
2026-02-17 11:29:52 -06:00
Nomadcxx e86f0afd87 fix(kilo): use @ai-sdk/openai-compatible npm package
The npm field pointed to opencode-kilo-auth which causes
ProviderInitError when loading Kilo models.

Switched to @ai-sdk/openai-compatible (already bundled in OpenCode)
and added api field for the gateway endpoint.
2026-02-18 04:21:52 +11:00
Aiden Cline 29c5e28a43 Merge pull request #791 from samsja/add-intellect-3
Add Intellect 3 model from Prime Intellect
2026-02-17 10:56:34 -06:00
Aiden Cline ea414b1500 Merge pull request #935 from ConceptCodes/feat/add-glm-flashx-model
feat: add GLM-4.7-FlashX model configuration
2026-02-17 10:34:45 -06:00
Aiden Cline 4556fe8b5b Merge pull request #937 from gary149/feat/huggingface-qwen3.5-m2.5-coder-next
feat(huggingface): add Qwen3.5-397B, MiniMax-M2.5, Qwen3-Coder-Next
2026-02-17 10:34:33 -06:00
Aiden Cline 8af23aeba5 Merge pull request #938 from spiffytech/dev
Add Ollama Cloud support for Qwen 3.5
2026-02-17 10:34:18 -06:00
Aiden Cline 9bfe1203c6 ci 2026-02-17 10:34:02 -06:00
spiffytech a619966e22 Added Ollama Cloud support for Qwen 3.5 2026-02-17 10:00:15 -05:00
Victor Muštar f48d55e1aa chore: remove accidentally committed skill file 2026-02-17 10:31:27 +01:00
Victor Muštar fe0ddcb666 feat(huggingface): add Qwen3.5-397B, MiniMax-M2.5, Qwen3-Coder-Next 2026-02-17 10:31:18 +01:00
Frank af1e1d1f51 update zen models 2026-02-17 02:08:24 -05:00
Aiden Cline 774a9f40b0 Merge pull request #933 from too-green/patch-1
Fix the display name of GLM-4.7-Flash
2026-02-17 00:23:26 -06:00
Aiden Cline 4f01ffb017 Merge pull request #934 from PandaSt0rm/add-minimax-m2.5-highspeed-models
Add MiniMax-M2.5-highspeed models for official MiniMax providers
2026-02-17 00:23:10 -06:00
Aiden Cline 7d768260cf Merge pull request #936 from Alex-wuhu/dev
add Qwen3.5-397B-A17B for novita
2026-02-17 00:22:30 -06:00
Alex-wuhu 467d269522 add Qwen3.5-397B-A17B for novita 2026-02-17 13:22:04 +08:00
concept 5ec496d6e1 feat: add GLM-4.7-FlashX model configuration 2026-02-16 21:31:16 -06:00
PandaSt0rm 45ca42f95a add MiniMax-M2.5-highspeed models 2026-02-17 03:53:05 +02:00
Ahmed f19ebce14c Rename model to GLM-4.7-Flash
Both GLM 4.7 and GLM 4.7 Flash had been named to the same "GLM 4.7"
2026-02-17 05:03:41 +05:00
Aiden Cline 85f5340eeb Merge pull request #931 from rifandyzv/dev
Add Qwen3.5 models for alibaba & alibaba-cn provider
2026-02-16 16:06:32 -06:00
Aiden Cline 39f06e82e7 Merge pull request #932 from cantalupo555/feat/add-openrouter-qwen3.5-plus-and-397b-a17b
feat: add Qwen3.5 models on OpenRouter
2026-02-16 16:05:56 -06:00
cantalupo555 4e7725d244 feat: add Qwen3.5 models on OpenRouter 2026-02-16 17:47:45 -03:00
Aiden Cline 7fe64bc498 Revert "Add image and video to input modalities"
This reverts commit 76e84a8b06.
2026-02-16 12:12:29 -06:00
rifandyzv f84a4a5cf9 feat: add Qwen3.5 models for alibaba & alibaba-cn provider 2026-02-17 01:28:16 +08:00
Aiden Cline 96f60c3329 Merge pull request #928 from Daltonganger/feat/kilo-provider-models
Add Kilo Gateway provider and import Kilo models
2026-02-16 11:08:11 -06:00
Aiden Cline f67b9bdef7 Merge pull request #929 from Daltonganger/feat/nano-gpt-qwen35-models
Add four Qwen3.5 models for NanoGPT
2026-02-16 11:05:23 -06:00
Frank 76e84a8b06 Add image and video to input modalities 2026-02-16 12:00:43 -05:00
Daltonganger cec16274c7 Add NanoGPT Qwen3.5 model variants 2026-02-16 17:09:17 +01:00
Daltonganger 17094722ea Add Kilo provider and import Kilo model catalog 2026-02-16 17:00:15 +01:00
Matthew (BlueT) Lien e3e230e3b3 fix: add api base URL template to partner model [provider] overrides
Add the api field with env-var template URL to all partner models so
opencode's loadBaseURL() can resolve the OpenAI-compatible endpoint.

Uses GOOGLE_VERTEX_PROJECT (not GOOGLE_CLOUD_PROJECT) because
googleVertexVars() resolves it through the full fallback chain
(GOOGLE_VERTEX_PROJECT → options.project → GOOGLE_CLOUD_PROJECT →
GCP_PROJECT → GCLOUD_PROJECT).
2026-02-16 21:51:55 +08:00
Aiden Cline 4666f36f3e Merge pull request #926 from zainhas/dev
[Together AI] add minimax M2.5
2026-02-15 23:57:45 -06:00
Aiden Cline 5a2bcd704e Merge pull request #900 from conglinyizhi/dev
feat: Add StepFun provider support
2026-02-15 23:57:35 -06:00
Zain Hasan 05861fd7fd add minimax M2.5 2026-02-15 21:48:37 -08:00
Aiden Cline e37bb8ae68 Merge pull request #923 from juls0730/dev
Fix cerebras/zai-gml-4.7 pricing
2026-02-15 20:03:47 -06:00
Aiden Cline 495e8006df Merge pull request #905 from shelvick/add-vertex-glm-5
Add GLM-5 to Google Vertex AI
2026-02-15 20:03:34 -06:00
Aiden Cline ad8dde798d Merge pull request #924 from cgilly2fast/dev
feat(firmware): add reason to anthropic models
2026-02-15 20:03:23 -06:00
Aiden Cline ab4fa333e3 Merge pull request #925 from 8dazo/dev
feat: add MiniMax M2.5 to Chutes provider listings
2026-02-15 20:03:12 -06:00
8dazo 6255298cd1 minimax model update 2026-02-16 06:31:01 +05:30
Colby Gilbert 3614087be6 Merge branch 'anomalyco:dev' into dev 2026-02-15 15:55:32 -08:00
Colby Gilbert 4495cb3569 feat(firmware): add reason to anthropic models 2026-02-15 15:55:02 -08:00
juls0730 8444d9293d Fix cerebras/zai-gml-4.7 pricing
Prices from https://inference-docs.cerebras.ai/models/zai-glm-47#z-ai-glm-4-7
2026-02-15 17:53:57 -06:00
Aiden Cline c1d36715ee Merge pull request #914 from 8dazo/dev
feat: add Z-AI GLM-5 to Chutes provider listings
2026-02-15 15:45:37 -06:00
Aiden Cline 860e610b73 Merge pull request #922 from cgilly2fast/dev
chore: remove unsupported models
2026-02-15 15:45:28 -06:00
Colby Gilbert beb84e769a chore: remove unsupported models 2026-02-15 13:38:52 -08:00
Aiden Cline 0408546681 Merge pull request #921 from zerone0x/feat/add-bedrock-deepseek-v3.2
feat(amazon-bedrock): add DeepSeek V3.2
2026-02-15 15:29:26 -06:00
Aiden Cline 97a040bc5e Merge pull request #915 from fanweixiao/dev
add glm-5, gpt-5-mini, deepseek-v3.2 models for vivgrid provider
2026-02-15 15:29:08 -06:00
Aiden Cline f68786b892 Merge pull request #920 from anomalyco/revert-912-add-github-copilot-gpt-5-3-codex
Revert "feat: add GitHub Copilot GPT-5.3 Codex"
2026-02-15 08:42:51 -06:00
Clawdbot 816c3d96b9 feat(amazon-bedrock): add DeepSeek V3.2
Add DeepSeek V3.2 model to Amazon Bedrock provider.

Model ID: deepseek.v3.2-v1:0
Pricing (US regions): $0.62/1M input, $1.85/1M output

Ref: https://aws.amazon.com/about-aws/whats-new/2026/02/amazon-bedrock-adds-support-six-open-weights-models/
2026-02-15 09:19:18 +01:00
Aiden Cline bac557c176 Revert "feat: add github copilot gpt-5.3-codex model (#912)"
This reverts commit 08db483d58.
2026-02-14 18:31:30 -06:00
Aiden Cline 97e81f356e Merge pull request #908 from hsnyus-09/feature/add-aurora-alpha
feat(openrouter): add aurora-alpha model definition
2026-02-14 17:34:22 -06:00
Matthew (BlueT) Lien 4c361218de feat: add Vertex AI partner models with openai-compatible overrides
Add DeepSeek V3.1, Llama 4 Maverick, Llama 3.3 70B, and Qwen3 235B as
partner models under google-vertex provider. Update GLM-4.7 with
corrected specs from official Google Cloud docs.

Each partner model uses [provider] npm override to @ai-sdk/openai-compatible
since these models are served via Google's OpenAI-compatible endpoint,
while staying consolidated under the google-vertex provider per
maintainer feedback.

All specs (context windows, output limits, pricing, modalities)
verified against official Google Cloud documentation:
- cloud.google.com/vertex-ai/generative-ai/pricing
- cloud.google.com/vertex-ai/generative-ai/docs/maas/*

Changes:
- Update zai-org/glm-4.7-maas: fix context=200K, output=128K, add pdf
  modality, correct release_date, add structured_output, add [provider]
- Add deepseek-ai/deepseek-v3.1-maas ($0.60/$1.70, 163K context)
- Add meta/llama-4-maverick-17b-128e-instruct-maas (vision, 524K ctx)
- Add meta/llama-3.3-70b-instruct-maas ($0.72/$0.72, 128K context)
- Add qwen/qwen3-235b-a22b-instruct-2507-maas ($0.22/$0.88, 262K ctx)
2026-02-15 06:35:17 +08:00
Anjul Garg 08db483d58 feat: add github copilot gpt-5.3-codex model (#912) 2026-02-14 14:27:49 -05:00
YuSung Han e0c14d7883 Remove redundant lines in aurora-alpha.toml 2026-02-15 03:50:02 +09:00
Aiden Cline e457c7f1dd Merge pull request #916 from arshadbarves/add-nvidia-glm5
Add GLM5 model to nvidia provider
2026-02-14 11:53:09 -06:00
Arshad Barves c86b97226c Update providers/nvidia/models/z-ai/glm5.toml
Co-authored-by: Copilot <175728472+Copilot@users.noreply.github.com>
2026-02-14 17:36:22 +05:30
Test User 8ec101e8f1 Add GLM5 model to nvidia provider 2026-02-14 15:32:11 +05:30
C.C. Fan cc1feddc11 add glm-5, gpt-5-mini, deepseek-v3.2 models 2026-02-14 17:08:07 +08:00
8dazo 95fd20e66a update name 2026-02-14 13:54:20 +05:30
8dazo 8c322d46dd Chutes Model Listings update 2026-02-14 13:49:06 +05:30
conglinyizhi da7a26b7b1 fix: 修复 StepFun provider 文档链接
将 doc 字段从 https://platform.stepfun.com/docs
改为 https://platform.stepfun.com/docs/zh/overview/concept
2026-02-14 15:07:40 +08:00
Aiden Cline 5a00835470 Merge pull request #913 from kavin-kr/patch-1
Update release date for nova-2-pro-v1 model
2026-02-14 00:50:06 -06:00
Frank b0c0a91926 update zen models 2026-02-14 00:52:08 -05:00
Kavin 8b2d995801 Update release date for nova-2-pro-v1 model 2026-02-13 23:46:55 -06:00
Aiden Cline 3f4b804ca1 Merge pull request #911 from monotykamary/feat/add-minimax-m2.5
feat: add fireworks minimax m2.5 model and fix m2.1 cache pricing
2026-02-13 19:00:46 -06:00
Tom X Nguyen 133e605c23 feat: add minimax m2.5 model and fix m2.1 cache pricing 2026-02-14 07:53:15 +07:00
Aiden Cline f3cff10e78 Merge pull request #910 from keenborder786/fix/gpt_5_2_pro
fix: gpt 5-2-pro does not support structured output
2026-02-13 17:53:00 -06:00
keenborder786 4732ca7f77 fix: gpt 5-2-pro does not support structured output 2026-02-14 04:50:19 +05:00
Aiden Cline 06f79b4142 Merge pull request #909 from juls0730/dev
Add all missing cohere models offered by the cohere api
2026-02-13 16:47:11 -06:00
Aiden Cline 69106c6f36 Merge pull request #907 from Algowary/dev
Chutes Model Listings update
2026-02-13 16:44:43 -06:00
Zoe dc1d8b78d0 Add all missing cohere models offered by the cohere api
This commit adds all the models offered by the official cohere api
that are not yet available in the models.dev repo, excluding the
rerank and embed models.
2026-02-13 16:43:57 -06:00
hsnyus-09 4383829304 feat(openrouter): add aurora-alpha model definition 2026-02-14 06:17:25 +09:00
Algowarry 610713b805 Merge branch 'dev' of https://github.com/Algowary/models.dev into dev 2026-02-13 16:08:26 -05:00
Algowarry 07e3eec2cf Chutes Model Inventory Update
Updating the models available from the provider chutes.ai
2026-02-13 16:01:46 -05:00
Algowarry ccf8a3e82a Merge branch 'dev' of https://github.com/Algowary/models.dev into dev 2026-02-13 15:07:16 -05:00
Algowarry 8e1d34323a Chutes Model Update
Model inventory and stats update
2026-02-13 14:38:07 -05:00
Aiden Cline 5f6d36a463 Merge pull request #787 from elithrar/fix/cloudflare-ai-gateway-provider-package
use official ai-gateway-provider package for Cloudflare AI Gateway
2026-02-13 12:44:45 -06:00
Scott Helvick 84e43b1912 Add GLM-5 to Google Vertex AI 2026-02-13 17:57:08 +00:00
Aiden Cline a9f14cbae3 Merge pull request #903 from micuintus/dev
fix(nebius): correct model ID casing to match Token Factory API
2026-02-13 10:39:51 -06:00
Aiden Cline 97baff037b Merge pull request #896 from zainhas/dev
[Together AI] Add GLM-5
2026-02-13 10:38:06 -06:00
Aiden Cline b149bd83ad Merge pull request #902 from qychen2001/dev
chore(siliconflow): update siliconflow/siliconflow-cn models
2026-02-13 10:25:04 -06:00
Aiden Cline 30ef1d1b51 Merge pull request #898 from 888-wzk/feature/chenger_20260128
feat(models): Added minimax model profile
2026-02-13 10:24:36 -06:00
Aiden Cline a62ccfd392 Merge pull request #901 from dpuyosa/venice
Venice: Add MiniMax M2.5 model configuration
2026-02-13 10:24:18 -06:00
QiyuanChen 23e195a0fa feat(models): Add interleaved reasoning_content field to GLM-4.7 and GLM-5 configurations for zai-org and Pro 2026-02-13 23:35:56 +08:00
Aiden Cline d1c5bc811e Merge pull request #899 from niushuai1991/feature/kuae-cloud-coding-plan
add provider: kuae cloud coding plan
2026-02-13 09:23:33 -06:00
Michael Voigt ad7e8047b9 fix(nebius): correct model ID casing to match Token Factory API
Fix lowercase model ID bug that caused "The model does not exist" errors.

- qwen/ → Qwen/ directory
- Fixed model file casing to match API exactly:
  - google/gemma-* → lowercase (gemma-2-2b-it, etc.)
  - meta-llama/*-Fast → lowercase fast suffix
  - nvidia/Llama-3_1-* → underscore instead of dot
  - nvidia/NVIDIA-* → uppercase NVIDIA prefix
  - black-forest-labs/flux-* → all lowercase
  - BAAI/bge-* → all lowercase
  - All Qwen models → proper casing

* Remove outdated models not in API:
- deepseek-ai/DeepSeek-V3
- meta-llama/Llama-3.1-405B-Instruct
- zai-org/GLM-4.7

* Add new model:
- moonshotai/Kimi-K2.5 (262K context, multimodal)

Fixes: https://github.com/anomalyco/opencode/issues/12461
and: https://ideas.nebius.com/en/p/token-factory-api-lowercase-model-ids

Note: The changes made and verified with actual Nebius API access
2026-02-13 14:28:48 +01:00
QiyuanChen 73393e9e41 chore(models): Remove Qwen3-30B-A3B and DeepSeek-R1-Distill-Qwen-7B model configuration files from siliconflow and siliconflow-cn 2026-02-13 20:11:19 +08:00
QiyuanChen 3e5566ab9d chore(models): Remove GLM-4.1V-9B-Thinking model configuration files from siliconflow and siliconflow-cn 2026-02-13 20:09:14 +08:00
QiyuanChen f5096e4b54 chore(models): Remove Kimi-Dev-72B model configuration files from siliconflow and siliconflow-cn 2026-02-13 20:08:14 +08:00
QiyuanChen 1d6e26574f chore(models): Remove MiniMaxAI/MiniMax-M1-80k and MiniMax-M2 model configuration files 2026-02-13 20:07:15 +08:00
QiyuanChen 57b0608e70 feat(models): Introduce Step-3.5-Flash model configuration and remove deprecated Step-3 model files 2026-02-13 20:06:05 +08:00
QiyuanChen 88ed698a69 feat(models): Enable structured_output in GLM-4.7 and GLM-5 configurations for zai-org and Pro 2026-02-13 20:04:27 +08:00
QiyuanChen 286c43f2cd feat(glm-5): Add new GLM-5 model configuration files for zai-org and Pro 2026-02-13 20:00:10 +08:00
dpuyosa 04d82741fa [venice] Add MiniMax M2.5 model configuration
- Modalities: text input/output
- Context window: 198K tokens
- Max output: 32K tokens
- Pricing: $0.40/M input, $1.60/M output, $0.04/M cache read
2026-02-13 09:44:22 +01:00
conglinyizhi 58c595b95f feat: Add StepFun provider support
- Add StepFun(阶跃星辰) as a new provider with OpenAI-compatible API
- Support step-3.5-flash (256K context, reasoning model)
- Support step-2-16k (1T parameters, 16K context)
- Support step-1-32k (100B parameters, 32K context)

Pricing based on official StepFun documentation (converted from CNY to USD):
- step-3.5-flash: bash.096 input / bash.288 output / bash.019 cache
- step-2-16k: .21 input / 6.44 output / .04 cache
- step-1-32k: .05 input / .59 output / bash.41 cache

Note: Logo not included as it is optional per contributing guidelines.
A default logo will be served by models.dev API instead.

All model definitions follow the official schema.

Fixes anomalyco/opencode#11760
Fixes anomalyco/opencode#11960

StepFun API: https://api.stepfun.com/v1
Documentation: https://platform.stepfun.com/docs/zh/pricing/details
2026-02-13 16:28:07 +08:00
城二 58de85c2e8 feat(minimax): Add interleaved configuration
- Add the reasoning_content field configuration to the minimax model.
- Update the configuration files for m2.5 and m2.5-lightning.
2026-02-13 16:20:45 +08:00
城二 f87ecffbf0 feat(minimax): Update m2.5 model name and price
- Change the model name from "lightning" to "highspeed"
- Adjust the input/output and cache read/write prices
2026-02-13 16:18:22 +08:00
niushuai1991 6dfb2f9c83 add kuae cloud coding plan 2026-02-13 15:02:52 +08:00
城二 ee8c1bce7d feat(models): Added minimax model profile 2026-02-13 14:15:44 +08:00
Zain Hasan a2dd10d09d Update output limit in GLM-5 configuration 2026-02-12 22:12:56 -08:00
Zain Hasan f6cfc2ebd2 try remove reasoning 2026-02-12 21:54:02 -08:00
Zain Hasan 3192856cc3 finx glm 5 settings 2026-02-12 21:46:13 -08:00
Aiden Cline 5507f42604 Merge pull request #874 from 888-wzk/feature/chenger_20260128
feat(z-ai): New glm-5 model configuration file
2026-02-12 22:56:36 -06:00
Aiden Cline 995aabf33f Merge pull request #894 from fhennerkes/dev
Poe: fix formatting, naming and update outputs
2026-02-12 22:56:08 -06:00
城二 fc5c3613eb feat(glm-5): Add reasoning_content field 2026-02-13 11:37:33 +08:00
fhennerkes 72f10a52e1 poe: update model names to use display_name 2026-02-12 19:08:49 -08:00
fhennerkes e944012d95 poe: small fixes (formatting and reasoning) 2026-02-12 18:55:19 -08:00
Aiden Cline e117f37d4e Merge pull request #892 from pat-baseten/add-kimi-2.5-baseten
Add Kimi K2.5 model for Baseten
2026-02-12 17:45:46 -06:00
Pat b0d71629fe Add Kimi K2.5 model for Baseten 2026-02-12 16:39:56 -06:00
Aiden Cline aa5e8634b2 Merge pull request #890 from cfal/fireworks-glm-5
fireworks: add GLM-5
2026-02-12 16:15:18 -06:00
Aiden Cline 7acba1db3f Merge pull request #891 from lucianjon/feat/openrouter-minimax-m2.5
feat(openrouter/minimax): add minimax-m2.5
2026-02-12 16:15:08 -06:00
Aiden Cline c5095973e1 Merge pull request #886 from brentdurksen/dev
feat(amazon-bedrock): add Writer Palmyra X4 and X5 models
2026-02-12 16:14:58 -06:00
Aiden Cline 5b8797cf89 Merge pull request #889 from Daltonganger/feat/nano-gpt-add-minimax-m2.5-official
feat(nano-gpt): add MiniMax M2.5 route alongside official variant
2026-02-12 16:14:36 -06:00
Daltonganger f1317184b5 Enable reasoning and add interleaved field in TOML 2026-02-12 23:07:12 +01:00
Lucian Jones 7ba286c7f2 feat(openrouter/minimax): add minimax-m2.5 2026-02-13 10:59:30 +13:00
cfal dec532b3b0 providers/fireworks-ai/models/accounts/fireworks/models/glm-5.toml: add GLM-5 to fireworks 2026-02-13 01:45:51 +04:00
Aiden Cline ccff680988 Merge pull request #864 from sylviezhang37/vercel-model-file-gen-script
feat(provider): Vercel model file generation and update script
2026-02-12 15:37:17 -06:00
Aiden Cline 1b63e4670e Merge pull request #887 from PandaSt0rm/add-minimax-m2-5-support
Add MiniMax-M2.5 across minimax and coding-plan providers
2026-02-12 15:36:56 -06:00
Aiden Cline d5cbd6fb5d Merge pull request #888 from spiffytech/dev
Add Ollama Cloud support for Minimax 2.5
2026-02-12 15:36:14 -06:00
Ruben Beuker e8f2f6b14f feat(nano-gpt): add MiniMax M2.5 route and align official variant 2026-02-12 22:35:23 +01:00
spiffytech e92fe6e9d7 Added Ollama Cloud support for Minimax 2.5 2026-02-12 16:24:11 -05:00
PandaSt0rm 7a32f17911 add MiniMax-M2.5 configs across minimax providers 2026-02-12 23:22:52 +02:00
Brent Durksen 57db1db84f feat(amazon-bedrock): add Writer Palmyra X4 and X5 models
Add two new Writer AI models to the Amazon Bedrock provider:

- writer.palmyra-x4-v1:0 (Palmyra X4): 128K context, 8K output,
  reasoning and tool calling, $2.50/$10 per M tokens (input/output)
- writer.palmyra-x5-v1:0 (Palmyra X5): 1M context, 8K output,
  reasoning and tool calling, $0.60/$6 per M tokens (input/output)

Both models support text-only input/output modalities and are
closed-weight.

Also adds the 'palmyra' family to the ModelFamilyValues enum in
packages/core/src/family.ts to support validation.
2026-02-12 13:43:32 -07:00
Aiden Cline ba91bb6612 Merge pull request #883 from ryanskidmore/ryanskidmore/cloudflare-ai-gateway-bump-opus-4-6-limits
cloudflare-ai-gateway: bump Opus 4.6 output limit to 128k
2026-02-12 13:01:38 -06:00
Ryan Skidmore 9761d0ef87 cloudflare-ai-gateway: bump Opus 4.6 output limit to 128k 2026-02-12 12:39:00 -06:00
Aiden Cline 98be9a2078 fix: family 2026-02-12 12:27:16 -06:00
Dax Raad 4aa17d26cb feat(openai): add gpt-5.3-codex-spark model 2026-02-12 13:24:43 -05:00
Aiden Cline 7f96ee576a Merge pull request #880 from Daltonganger/feat/nano-gpt-glm5-original-models
feat(nano-gpt): add GLM 5 original model variants
2026-02-12 12:18:46 -06:00
Aiden Cline bd5ce80e56 Merge pull request #882 from Daltonganger/feat/nano-gpt-add-minimax-m2.5-official
feat(nano-gpt): add MiniMax M2.5 Official model
2026-02-12 12:18:37 -06:00
Daltonganger ed2af4ad45 feat(nano-gpt): add MiniMax M2.5 Official model 2026-02-12 18:07:52 +01:00
Aiden Cline ac0868c886 Merge pull request #881 from Alex-wuhu/dev
add minmax-2.5 on novita
2026-02-12 10:32:50 -06:00
Aiden Cline fdd13245cc Revert "feat(github-copilot): add gpt-5.3-codex model (#857)"
This reverts commit 27abb8a570.
2026-02-12 10:32:15 -06:00
Alex 37c77c58ad Merge branch 'anomalyco:dev' into dev 2026-02-13 00:27:45 +08:00
Alex-wuhu 62ee8129e6 add minimax-m2.5 on novita 2026-02-13 00:23:06 +08:00
Daltonganger 4bc6f07570 fix(nano-gpt): correct GLM-5 dates to 2026-02-11 2026-02-12 17:22:18 +01:00
Aiden Cline bc0336c8ec Merge pull request #878 from cantalupo555/feat/add-openrouter-stepfun-step-3.5-flash
feat: add StepFun Step 3.5 Flash on OpenRouter
2026-02-12 10:13:17 -06:00
Aiden Cline dd78db4dc6 Merge pull request #879 from amankalra172/add-stackit-provider
fix: reorganize STACKIT models with organization prefixes
2026-02-12 10:12:48 -06:00
Daltonganger 4fd32c741d refactor(nano-gpt): consolidate z-ai GLM models under zai-org 2026-02-12 17:11:50 +01:00
Frank c78ca7c132 update zen models 2026-02-12 11:05:41 -05:00
Alex 2a99397516 add GLM5 on novita (#877) 2026-02-12 11:03:42 -05:00
Frank 554440be4f update zen models 2026-02-12 11:01:51 -05:00
Daltonganger eb52a76d11 feat(nano-gpt): add GLM 5 original model variants 2026-02-12 16:48:59 +01:00
amankalra172 c9a7f6c814 fix: reorganize STACKIT models with organization prefixes and correct pricing
- Move models to organization subfolders (Qwen/, cortecs/, google/, etc.)
- Update pricing from EUR to USD (1.09 conversion rate)
- Fix GPT-OSS context limit to 131K tokens
- Add architectural family classifications
- Verify tool_call settings for all models
2026-02-12 14:00:45 +01:00
cantalupo555 e640802d34 feat: add StepFun Step 3.5 Flash (free) on OpenRouter 2026-02-12 08:44:15 -03:00
cantalupo555 c226863912 feat: add StepFun Step 3.5 Flash on OpenRouter 2026-02-12 08:42:37 -03:00
Alex-wuhu 8bcd634743 add GLM5 on novita 2026-02-12 16:26:05 +08:00
Aiden Cline 812cd1763a Merge pull request #873 from juls0730/dev
Create cerebras/llama3.1-8b.toml
2026-02-12 00:40:48 -06:00
城二 11f4ae568e feat(z-ai): New glm-5 model configuration file 2026-02-12 11:23:22 +08:00
juls0730 9f1629a26a Create cerebras/llama3.1-8b.toml 2026-02-11 21:01:08 -06:00
Yunfei He 27abb8a570 feat(github-copilot): add gpt-5.3-codex model (#857)
* feat(github-copilot): add gpt-5.3-codex model

* fix(github-copilot): align gpt-5.3-codex release metadata
2026-02-11 21:53:36 -05:00
Aiden Cline 2aa4a2290e Merge pull request #868 from dpuyosa/venice
Venice: Add GLM-5 model
2026-02-11 19:49:16 -06:00
Aiden Cline c58b36c605 Merge pull request #872 from Track07-cda/openrouter-glm5
OpenRouter: Add GLM-5 and remove Pony Alpha
2026-02-11 19:49:07 -06:00
Aiden Cline e91dbd1fc4 Merge pull request #870 from spiffytech/dev
Add Ollama Cloud support for GLM-5
2026-02-11 19:39:23 -06:00
Track07-cda 8924ee3092 feat(openrouter): add GLM-5 and remove Pony Alpha
Add the Z-AI GLM-5 model definition to the OpenRouter provider and
remove the deprecated Pony Alpha model.
2026-02-12 09:38:12 +08:00
spiffytech fc5b6533d1 Added Ollama Cloud support for GLM-5 2026-02-11 19:57:46 -05:00
Aiden Cline 7a760e3a4f Merge pull request #871 from Kunde21/synthetic_k2_5_nvfp4
Synthetic: Add Kimi -2 5 in NVFP4 remove GLM-4.5
2026-02-11 18:53:56 -06:00
Chad Kunde 38835801b1 synthetic: deprecate GLM-4.5
Model removed from models list as of 12 Feb 2026
2026-02-12 07:27:24 +07:00
Chad Kunde 81103438a3 synthetic: Add NVFP4 variant of Kimi K2.5 2026-02-12 07:25:41 +07:00
dpuyosa 0b0b36eb45 [venice] Add GLM-5 model with 198K context window
- Add ZAI-ORG GLM-5 model configuration to Venice provider
- Supports reasoning, tool calls, structured output
- Text in/out: 198K context, 49.5K output tokens
2026-02-11 22:53:23 +01:00
Aiden Cline b18b73f0a0 Revert "fix: openai advertises a 400k context window, that is just the sum of max input + max output, so real context window is 272k"
This reverts commit ea276d57a7.
2026-02-11 15:24:43 -06:00
Aiden Cline 3cd48b273a Merge pull request #867 from AnishShah1803/nano-gpt/add-glm-5-models
Add GLM 5 to NanoGPT models list
2026-02-11 14:47:57 -06:00
twisted 890992b8ef update release date 2026-02-11 20:37:59 +00:00
twisted a843d84d77 Add GLM 5 to NanoGPT models list 2026-02-11 20:35:41 +00:00
Aiden Cline d89897d07e Merge pull request #866 from Sczr0/dev
Update pricing for ZAI GLM-5
2026-02-11 14:35:23 -06:00
Aiden Cline 1b26792073 fix zai 2026-02-11 14:34:43 -06:00
Aiden Cline 6a0da0a91d Revert "Fixed ZAI GLM-5 pricing to free (0 cost)"
This reverts commit 79d1222e3c.
2026-02-11 14:33:39 -06:00
opencode-agent[bot] 79d1222e3c Fixed ZAI GLM-5 pricing to free (0 cost)
Co-authored-by: rekram1-node <rekram1-node@users.noreply.github.com>
2026-02-11 20:14:03 +00:00
Sylvie Zhang 1ad45b44c2 add readme 2026-02-11 12:13:38 -08:00
弦塔_ 45ba8066df Update cost parameters in glm-5.toml 2026-02-12 04:07:57 +08:00
弦塔_ 4861f14a65 Update cost parameters in glm-5.toml 2026-02-12 04:07:29 +08:00
弦塔_ c83b4e22bd Update cost parameters in glm-5.toml 2026-02-12 03:56:40 +08:00
Sylvie Zhang 44c1ed5aeb additional data cleaning logic 2026-02-11 11:48:32 -08:00
Sylvie Zhang 28c09d83a0 add fallback logic 2026-02-11 11:48:32 -08:00
Sylvie Zhang ea41cbc4ba draft script 2026-02-11 11:48:32 -08:00
Aiden Cline c893ac5f9d Merge pull request #863 from AnishShah1803/nano-gpt/update-Kimi-K2-5-models
Add Kimi K2.5 models to NanoGPT provider
2026-02-11 13:41:43 -06:00
twisted 8e10faf38a set reasoning to true for kimi k2.5 2026-02-11 19:33:12 +00:00
Aiden Cline 4c3a17fbe8 Merge pull request #859 from friendliai/minpeter/add-glm5-friendli
Add zai-org/GLM-5 model to Friendli provider
2026-02-11 13:14:58 -06:00
twisted f7d997e4f0 fix last_updated 2026-02-11 19:08:48 +00:00
twisted 6961c57c86 make open_weights set to true 2026-02-11 19:07:35 +00:00
minpeter e44307c27c Add interleaved reasoning_content to GLM-4.7 and MiniMax-M2.1
Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>
2026-02-12 04:01:29 +09:00
minpeter b9f7907150 Merge remote-tracking branch 'origin/dev' into minpeter/add-glm5-friendli 2026-02-12 04:00:31 +09:00
minpeter bf0ee2a3eb Add interleaved reasoning_content field for GLM-5
GLM models use interleaved reasoning via the reasoning_content field with OpenAI-compatible providers.

Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>
2026-02-12 03:58:58 +09:00
Aiden Cline b9a73edcdc Merge pull request #862 from hanouticelina/feat/huggingface-glm-5
feat(huggingface): add GLM-5 for Hugging Face provider
2026-02-11 12:45:43 -06:00
Aiden Cline 7ff3be2dab fix: github copilot model discrepencies 2026-02-11 12:40:36 -06:00
twisted 406f82e09f Add Kimi K2.5 models to NanoGPT provider 2026-02-11 18:40:09 +00:00
Celina Hanouti c8fa26624d add GLM-5 for hugging face provider 2026-02-11 19:39:35 +01:00
Aiden Cline ae31005ef3 Merge pull request #830 from amankalra172/add-stackit-provider
feat: add STACKIT provider with 8 AI models
2026-02-11 12:20:47 -06:00
Aiden Cline 0aa7c9f3c2 Merge pull request #852 from zainhas/dev
[Together AI] update output token length to match context length
2026-02-11 12:20:09 -06:00
Aiden Cline 40be65f301 Merge pull request #853 from captain1379/feat/jiekou
Add new models for Jiekou.AI
2026-02-11 12:19:40 -06:00
Aiden Cline 49ab2c0a48 feat: add glm 5 to zai, zhipuai, and zai coding plan 2026-02-11 12:17:49 -06:00
minpeter 7fd96a6c9d Add zai-org/GLM-5 model to Friendli provider
Add GLM-5 model configuration with reasoning, tool calling, and structured output support. Update family pattern inference in generate script to recognize GLM-5 models.

Co-Authored-By: Claude Haiku 4.5 <noreply@anthropic.com>
2026-02-12 03:07:25 +09:00
Aiden Cline fc4027fe98 Merge pull request #856 from josetorres1/add-bedrock-zai-minimax-models
Add GLM 4.7 Family and MiniMax M2.1 to Amazon Bedrock
2026-02-11 11:25:21 -06:00
Aiden Cline b260564060 Merge pull request #855 from dihan-dff-user/dev
Add ZAI coding plan GLM-5 model
2026-02-11 11:24:47 -06:00
opencode-agent[bot] ecd5927bed Removed knowledge field from GLM-5 config
Co-authored-by: rekram1-node <rekram1-node@users.noreply.github.com>
2026-02-11 17:23:31 +00:00
Jose Torres 490381f370 Add GLM 4.7 family and MiniMax M2.1 to Amazon Bedrock provider
- Add GLM-4.7 (zai.glm-4.7): /bin/zsh.60/.20 per 1M tokens
- Add GLM-4.7-Flash (zai.glm-4.7-flash): /bin/zsh.07//bin/zsh.40 per 1M tokens
- Add MiniMax M2.1 (minimax.minimax-m2.1): /bin/zsh.30/.20 per 1M tokens

Pricing sources:
- AWS Bedrock pricing page: https://aws.amazon.com/bedrock/pricing/
- Model IDs confirmed via AWS console/CLI

Related to GH issue #835
2026-02-11 09:51:18 -06:00
dihan 621687e457 Add ZAI coding plan GLM-5 model 2026-02-11 20:38:40 +05:30
captain1379 d5a3bfad90 feat: add new models for Jiekou.AI
- Introduced `claude-opus-4-6`, `qwen3-coder-next` and `gpt-5.1` models with detailed configurations.
- Removed deprecated `qwen2.5-vl-72b-instruct` model.
- Implemented a new script for generating model configurations.
2026-02-11 15:19:40 +08:00
Zain Hasan 310bc174fa update output token length to match context length 2026-02-10 23:05:23 -08:00
Aiden Cline 9fb1233073 Merge pull request #848 from BlockListed/cortecs-glm-models
Add supported z.ai GLM models to cortecs
2026-02-10 20:04:21 -06:00
Aiden Cline 029f13545b Merge pull request #849 from BlockListed/cortecs-minimax-models
Add MiniMax models to cortecs
2026-02-10 20:04:12 -06:00
Aiden Cline 31b1acba5f Merge pull request #850 from cgilly2fast/dev
feat(firmware): kimi and glm models
2026-02-10 20:04:00 -06:00
Aiden Cline 120881916e Merge pull request #851 from anomalyco/fix-models
fix: openai advertises a 400k context window, that is just the sum of max input + max output, so real context window is 272k
2026-02-10 20:03:48 -06:00
Aiden Cline ea276d57a7 fix: openai advertises a 400k context window, that is just the sum of max input + max output, so real context window is 272k 2026-02-10 19:02:33 -06:00
Colby Gilbert 05d940ab8d feat(firmware): kimi and glm models 2026-02-10 14:25:46 -08:00
BlockListed 7d250ea857 cortecs add minimax models 2026-02-10 23:13:39 +01:00
BlockListed 82d72e85f8 add supported z.ai GLM models to cortecs 2026-02-10 22:58:22 +01:00
Aiden Cline 995934de32 Merge pull request #847 from rubenandre/add-bedrock-moonshotai-kimi-k2.5
add moonshotai kimi K2.5 to amazon-bedrock provider
2026-02-10 10:43:13 -06:00
Aiden Cline d10392573e Merge pull request #845 from cgilly2fast/dev
fix(firmware): remove unsupported model and fix name of gpt oss 20b
2026-02-10 10:04:15 -06:00
Aiden Cline 21c3c1b8e1 Merge pull request #846 from dpuyosa/venice
Venice: Enable reasoning for GLM-4.7-Flash model
2026-02-10 10:04:06 -06:00
Rúben Silva 4680aaedc3 add moonshotai kimi K2.5 to amazon-bedrock provider 2026-02-10 15:21:50 +00:00
dpuyosa a9f3ad978d [venice] Enable reasoning for GLM-4.7-Flash model 2026-02-10 10:11:48 +01:00
Colby Gilbert de75687ea1 fix(firmware): remove unsupported model and fix name of spt oss 20b 2026-02-09 23:01:40 -08:00
Aiden Cline 539cc930c4 Merge pull request #843 from cgilly2fast/dev
chore: clean up firmware available models
2026-02-09 18:40:24 -06:00
Colby Gilbert 31f3c63acc fix(firmware): remove reasoning from anthropic and deepseek models 2026-02-09 16:07:44 -08:00
Colby Gilbert b019787ad8 chore: clean up firmware available models 2026-02-09 16:03:09 -08:00
Aiden Cline e41fca18a2 Merge pull request #842 from riccardogiorato/dev
fix: Increase output limit to match context for kimi K2.5 on Together
2026-02-09 17:12:07 -06:00
Riccardo Giorato 74163f7314 Increase output limit to match context
Update providers/togetherai/models/moonshotai/Kimi-K2.5.toml to set [limit].output from 32_768 to 262_144. This aligns the output token limit with the context size (262_144) to avoid premature truncation and allow full-length responses.
2026-02-09 22:39:34 +01:00
Aiden Cline 7b763695fd Merge pull request #839 from shelvick/add-azure-kimi-k2.5
Add Azure Kimi-K2.5 model
2026-02-09 14:15:33 -06:00
Aiden Cline 686b47d01e Merge pull request #840 from shelvick/add-azure-claude-opus-4-6
Add Azure Claude Opus 4.6 model
2026-02-09 14:15:16 -06:00
Scott Helvick 46f0726d7f Add Azure Claude Opus 4.6 model 2026-02-09 20:07:31 +00:00
Scott Helvick 3c14600fc6 Add Azure Kimi-K2.5 model 2026-02-09 19:49:02 +00:00
Aiden Cline 721c025af1 Merge pull request #836 from PeppeRu96/feat/add-deepinfra-claude
feat: add DeepInfra Claude Opus 4 and Claude Sonnet 3.7 (latest) models
2026-02-09 12:29:34 -06:00
Aiden Cline c591f9b213 Merge pull request #837 from PeppeRu96/feat/add-deepinfra-deepseek
feat: add DeepInfra DeepSeek models
2026-02-09 12:23:02 -06:00
Aiden Cline 57580b28d3 Merge pull request #765 from captain1379/feat/jiekou
feat: add Jiekou.AI provider
2026-02-09 12:22:21 -06:00
Giuseppe Ruggeri 11e92f093f fix: fix price for DeepInfra DeepSeek-V3.2 2026-02-09 14:32:34 +01:00
Giuseppe Ruggeri 17cf21ba46 feat: add DeepInfra DeepSeek models 2026-02-09 14:29:02 +01:00
Giuseppe Ruggeri 60a3f09b8e fix: update deepinfra/claude-3-7-sonnet-latest family field 2026-02-09 14:02:52 +01:00
Giuseppe Ruggeri 6f907bce35 feat: add DeepInfra Claude Opus 4 and Claude Sonnet 3.7 (latest) models 2026-02-09 13:56:16 +01:00
Frank 1f20d47ef5 update zen models 2026-02-08 21:43:43 -05:00
Aiden Cline d5c23c9c95 Merge pull request #827 from modpotato/dev
fix: rename glm 5 stealth from 'Stealth' to 'Pony Alpha' + remove status
2026-02-08 14:03:59 -06:00
Aiden Cline 125abf1a21 Merge pull request #829 from 888-wzk/feature/chenger_20260128
fix: Update model configurations to adjust reasoning and interleaved …
2026-02-08 14:03:44 -06:00
Aiden Cline 38ccea666f Merge pull request #831 from spiffytech/dev
Add Ollama Cloud support for qwen3-coder-next
2026-02-08 14:03:28 -06:00
Frank 42ca5faeb8 sync 2026-02-08 14:21:34 -05:00
spiffytech bcd9e3dba1 Added Ollama Cloud support for qwen3-coder-next 2026-02-08 13:36:54 -05:00
amankalra172 7504dc2947 feat: add STACKIT provider with 8 AI models
Add STACKIT as a new provider with complete model specifications:

Chat Models:
- Llama 3.1 8B Instruct FP8
- Llama 3.3 70B Instruct FP8
- GPT-OSS 120B
- Mistral Nemo Instruct 2407 FP8
- Gemma 3 27B (multimodal)
- Qwen3-VL 235B (vision-language)

Embedding Models:
- E5 Mistral 7B
- Qwen3-VL Embedding 8B (multimodal)

All models include:
- Proper schema compliance (attachment, reasoning, tool_call, etc.)
- Pricing in USD per million tokens
- Context limits and modalities
- Official STACKIT logo with currentColor support

STACKIT is a German sovereign cloud provider offering OpenAI-compatible
AI model serving with open-source models.
2026-02-08 12:35:51 +01:00
城二 67bddb6b61 Merge branch 'dev' of https://github.com/888-wzk/models.dev into feature/chenger_20260128 2026-02-08 11:23:44 +08:00
城二 77330e78c6 fix: Update model configurations to adjust reasoning and interleaved fields 2026-02-08 11:21:58 +08:00
mod e55a05f6a6 Merge branch 'anomalyco:dev' into dev 2026-02-07 00:43:00 -05:00
mod e5ce677899 fix: rename glm 5 stealth from 'Stealth' to 'Pony Alpha' 2026-02-07 00:42:50 -05:00
Aiden Cline e1747322ad Merge pull request #826 from modpotato/dev
add pony alpha (glm 5 stealth)
2026-02-06 23:21:17 -06:00
Aiden Cline 9303c7be2e Merge pull request #825 from cantalupo555/feat/add-openrouter-mimo-v2-flash
feat: add Xiaomi MiMo-V2-Flash on OpenRouter
2026-02-06 16:43:34 -06:00
John Doe 204eb52c0d feat: pony alpha (glm 5 demo) on openrouter 2026-02-06 21:06:54 +00:00
John Doe a2abd136f5 feat: pony alpha (glm 5 demo) on openrouter 2026-02-06 21:02:40 +00:00
Aiden Cline ea6e487e77 fix: change anthropic default to 200k instead of 1M since not everyone can access the 1M 2026-02-06 13:46:09 -06:00
Aiden Cline 1033ee450c Merge pull request #821 from 888-wzk/feature/chenger_20260128
Added Claude Opus 4.6 model configuration file
2026-02-06 11:00:49 -06:00
Aiden Cline de8e46b2ab Merge pull request #823 from dpuyosa/venice
Venice: Tweak model generation script
2026-02-06 11:00:37 -06:00
Aiden Cline 8181d97317 Merge pull request #824 from vglafirov/feat/gitlab-opus-4-6
feat(gitlab): add Claude Opus 4.6 model (duo-chat-opus-4-6)
2026-02-06 11:00:11 -06:00
Vladimir Glafirov a5c9640163 feat(gitlab): add Claude Opus 4.6 model (duo-chat-opus-4-6)
Add the newly released Claude Opus 4.6 model for GitLab Duo Agentic Chat.

Related:
- AI Gateway MR: https://gitlab.com/gitlab-org/modelops/applied-ml/code-suggestions/ai-assist/-/merge_requests/4492
2026-02-06 17:04:17 +01:00
cantalupo555 c220f2a790 feat: add Xiaomi MiMo-V2-Flash on OpenRouter 2026-02-06 12:52:35 -03:00
Frank c88c849e5a Merge pull request #822 from imdevarsh/imdevarsh/openrouter-opus-4.6
feat(openrouter): add claude opus 4.6 to openrouter models list
2026-02-06 10:07:38 -05:00
dpuyosa 75ff468a9a [venice] Refactor model generation with privacy field
- Add optional privacy field to ModelSpec schema
- Use privacy field to determine open_weights capability
- Preserve existing output token limit when smaller than proposed
2026-02-06 13:18:47 +01:00
Devarsh 5b9186f6a9 feat(openrouter): add claude opus 4.6 to openrouter models list 2026-02-06 18:06:33 +13:00
城二 974713311b feat: Added Claude Opus 4.6 model configuration file 2026-02-06 11:11:19 +08:00
Aiden Cline 2d143f96d1 Merge pull request #813 from cgilly2fast/dev
feat: add opus 4.6 to firmware provider
2026-02-05 16:25:47 -06:00
Aiden Cline 4c4cd139f8 Merge pull request #812 from fhennerkes/dev
poe: add Claude Opus 4.6 model
2026-02-05 16:25:29 -06:00
Aiden Cline 22688b1260 Merge pull request #816 from markusylisiurunen/add-eu-opus-4.6
Add the missing EU variant back for Opus 4.6 on AWS Bedrock
2026-02-05 16:23:54 -06:00
Aiden Cline 13631caba9 Merge pull request #817 from dpuyosa/venice
Venice: Add Claude Opus 4.6 and GLM 4.7 models
2026-02-05 16:21:41 -06:00
dpuyosa 7e901e93bf [venice] Add Claude Opus 4.6 and GLM 4.7 models
- Add claude-opus-4.6 model configuration for Venice provider
- Add zai-org-glm-4.7-flash model configuration for Venice provider
2026-02-05 22:44:27 +01:00
Markus Ylisiurunen ace9626163 also fix pricing for opus 4.5 2026-02-05 23:25:25 +02:00
Markus Ylisiurunen 2bd869959b fix pricing 2026-02-05 23:12:40 +02:00
Markus Ylisiurunen efecbc137c Add EU variant for Opus 4.6 2026-02-05 23:03:01 +02:00
Colby Gilbert cae3f84930 Merge branch 'anomalyco:dev' into dev 2026-02-05 12:49:38 -08:00
Colby Gilbert 37cdea639f feat: add opus 4.6 2026-02-05 12:49:19 -08:00
Ryan Vogel 2c67792e4b Merge pull request #811 from anomalyco/add-claude-opus-4-6
Fix Claude Opus 4.6 model IDs and remove incorrect variants
2026-02-05 15:49:03 -05:00
Ryan Vogel 24addedada Fix Vertex AI model ID to claude-opus-4-6@default 2026-02-05 15:46:07 -05:00
Ryan Vogel dc9f404dbb Condense AGENTS.md model configuration section 2026-02-05 15:44:00 -05:00
Ryan Vogel db2212ab9f Update AGENTS.md with model configuration learnings 2026-02-05 15:42:47 -05:00
Ryan Vogel e2777a44ed Fix Vertex AI model ID: claude-opus-4-6@default -> claude-opus-4-6 2026-02-05 15:41:53 -05:00
fhennerkes c0f0394f67 poe: add Claude Opus 4.6 model 2026-02-05 12:39:00 -08:00
Ryan Vogel a762f47461 Fix model IDs: remove unannounced dated alias, remove EU Bedrock, fix Bedrock ID (v1:0 -> v1), fix Vertex ID (@20260205 -> @default)
Fixes #809
2026-02-05 15:38:28 -05:00
Aiden Cline 768f841f79 Merge pull request #781 from Dagnan/add-glm-4.7-flash-deepinfra
feat(deepinfra): add GLM-4.7-Flash model
2026-02-05 14:34:27 -06:00
Aiden Cline 00f239a852 Merge pull request #770 from jerilynzheng/feat/add-vercel-models-jan-30
vercel: add new models and interleaved support
2026-02-05 14:19:20 -06:00
Michel Pigassou 69f72041e1 Added missing interleaved/reasoning_content for GLM 4.7-Flash 2026-02-05 21:16:25 +01:00
Aiden Cline 43e98540ec fix: output limit for opus 4.6 on gh copilot 2026-02-05 14:15:56 -06:00
jerilynzheng f0854ab7b8 vercel: add interleaved = true for confirmed models
Add interleaved reasoning support to models confirmed by other providers:
- Claude: 3.7-sonnet, haiku-4.5, opus-4/4.1/4.5/4.6, sonnet-4/4.5
- DeepSeek: R1, V3.2-thinking
- MiniMax: M2, M2.1
- Kimi: K2-thinking, K2-thinking-turbo, K2.5
- GLM: 4.5, 4.6, 4.7, 4.7-flashx

Co-Authored-By: Claude Opus 4.5 <noreply@anthropic.com>
2026-02-05 12:14:02 -08:00
Aiden Cline a05a47097f Merge pull request #799 from iamanishx/deepinfra-kimi
feat: added support for kimi k2.5 (deepinfra)
2026-02-05 14:08:08 -06:00
jerilynzheng ed96ac7c74 fix: update Claude Opus 4.6 knowledge cutoff to 2025-05
Co-Authored-By: Claude Opus 4.5 <noreply@anthropic.com>
2026-02-05 12:03:40 -08:00
jerilynzheng 3985ee8556 vercel: add Claude Opus 4.6
Add anthropic/claude-opus-4.6 from Vercel AI Gateway:
- 1M context window, 128K output
- $5.00/$25.00 per 1M tokens (input/output)
- Supports vision, reasoning, and tool use

Co-Authored-By: Claude Opus 4.5 <noreply@anthropic.com>
2026-02-05 12:03:11 -08:00
Aiden Cline 53c150347f Merge pull request #808 from stickyburn/chutes-qwen3-coder-next
chore: add qwen3-coder-next for chutes.ai
2026-02-05 14:01:46 -06:00
Aiden Cline 1b4b15d73b Merge pull request #807 from smrdotgg/patch-1
Fix release and last updated dates for GPT-5.3 Codex
2026-02-05 14:01:37 -06:00
stickyburn 535b1afe4f chore: add qwen3 coder next for chutes.ai 2026-02-05 14:41:59 -05:00
Mathis e239c4d51e Add configuration for Claude Opus 4.6 model (#806) 2026-02-05 14:24:07 -05:00
imanishx 3dddd7b5e3 fix: interleaved opt added
Signed-off-by: imanishx <manishbiswal754@gmail.com>
2026-02-05 19:05:25 +00:00
Semere Tereffe 941487c8d4 Fix release and last updated dates for GPT-5.3 Codex 2026-02-05 21:52:29 +03:00
Ryan Vogel ce3bbe64a3 Update context window to 1M tokens for all Claude Opus 4.6 models 2026-02-05 13:39:48 -05:00
Aiden Cline 183dd5c4f3 Merge pull request #803 from anomalyco/add-claude-opus-4-6
Add Claude Opus 4.6 model
2026-02-05 12:39:29 -06:00
Ryan Vogel b4733df6b7 Merge branch 'dev' into add-claude-opus-4-6 2026-02-05 13:38:54 -05:00
Ryan Vogel 921ec8f8fc Add cost.context_over_200k long context pricing to all Claude Opus 4.6 models 2026-02-05 13:36:52 -05:00
Aiden Cline 6c28f3974c Merge pull request #804 from rexdotsh/feat/add-anthropic-opus-4-6
feat: add opus 4.6
2026-02-05 12:34:34 -06:00
Aiden Cline 8b3a813da9 Merge pull request #802 from dmmulroy/cloudflare-opus-4-6
cloudflare-ai-gateway: add claude opus 4.6
2026-02-05 12:34:00 -06:00
rexdotsh 10297c8490 feat: add opus 4.6 2026-02-05 23:58:13 +05:30
Ryan Vogel 61c8d50df2 Add Claude Opus 4.6 model across Anthropic, Bedrock, and Vertex AI providers 2026-02-05 13:26:28 -05:00
Dax Raad 4faf622d1e add gpt-5.3-codex.toml 2026-02-05 13:26:24 -05:00
Dillon Mulroy 89db412884 cloudflare-ai-gateway: add claude opus 4.6 2026-02-05 13:24:36 -05:00
Frank 107b285e1c update zen models 2026-02-05 13:14:06 -05:00
Frank 8bd58cf186 update zen models 2026-02-05 13:04:56 -05:00
Aiden Cline 9573a8fccc Merge pull request #796 from manascb1344/nebius-token-factory-models
feat: add Nebius Token Factory models
2026-02-05 11:45:09 -06:00
Aiden Cline 0153e6408a Merge pull request #792 from Alex-wuhu/dev
feat: add deepseek OCR model configuration and Qwen3 Coder Next model…
2026-02-05 11:44:28 -06:00
imanishx 8cb19035ff feat: added tomal for kini k2.4 (deepinfra)
Signed-off-by: imanishx <manishbiswal754@gmail.com>
2026-02-05 09:08:33 +00:00
captain1379 7c0e1e142f fix: removed old models 2026-02-05 14:02:59 +08:00
captain1379 7aeca69c4e fix: fix logo 2026-02-05 13:44:38 +08:00
Aiden Cline cfde47ca60 Revert "Update Amazon Bedrock models to add cross-region inference and remove deprecated models"
This reverts commit bc58036964.
2026-02-04 12:11:29 -06:00
Aiden Cline b01c07a3d0 Merge pull request #793 from zainhas/patch-1
[fix] Rename model to 'Qwen3 Coder Next FP8'
2026-02-04 10:33:22 -06:00
Aiden Cline 444c3071ee Merge pull request #795 from riccardogiorato/dev
remove wrongly typed Kimi-K2-5.toml
2026-02-04 10:32:27 -06:00
manascb1344 42a79c717d feat(nebius): update Meta-Llama, NVIDIA models and mark deprecated
- Update Llama-3.3-70B-Instruct (Base & Fast) with new pricing
- Mark Llama-3.1-405B-Instruct as deprecated (no longer available)
- Update Llama-3.1-Nemotron-Ultra-253B-v1 with new pricing
- Mark DeepSeek-V3 as deprecated (replaced by V3.2 and V3-0324)
2026-02-04 20:18:37 +05:30
manascb1344 578df73ffb feat(nebius): update Z.ai, OpenAI, Moonshot AI, and NousResearch models
- Update GLM-4.5 and GLM-4.5-Air with new pricing
- Update gpt-oss-120b and gpt-oss-20b with new pricing and features
- Update Kimi-K2-Instruct with new pricing and multimodal support
- Update Hermes-4-405B and Hermes-4-70B with new pricing
2026-02-04 20:17:51 +05:30
manascb1344 2639e20a97 feat(nebius): add new models from Z.ai, Moonshot AI, Meta, and NVIDIA
- Add GLM-4.7 and GLM-4.7-FP8 (Z.ai)
- Add Kimi-K2-Thinking (Moonshot AI)
- Add Llama-Guard-3-8B, Meta-Llama-3.1-8B-Instruct (Base & Fast) (Meta)
- Add Nemotron-Nano-V2-12b and NVIDIA-Nemotron-3-Nano-30B-A3B (NVIDIA)
2026-02-04 20:17:14 +05:30
manascb1344 ca6206b78e feat(nebius): add Qwen models to Token Factory
- Add Qwen3-Next-80B-A3B-Thinking
- Add Qwen3-30B-A3B-Thinking-2507 and Qwen3-30B-A3B-Instruct-2507
- Add Qwen3-Coder-30B-A3B-Instruct
- Add Qwen3-32B (Base & Fast)
- Add Qwen2.5-Coder-7B-fast
- Add Qwen2.5-VL-72B-Instruct
- Add Qwen3-Embedding-8B
2026-02-04 20:16:46 +05:30
manascb1344 51fe42982f feat(nebius): add DeepSeek models to Token Factory
- Add DeepSeek-V3.2, DeepSeek-V3-0324 (Base & Fast), DeepSeek-R1-0528 (Base & Fast)
- These are new models available on Nebius Token Factory
2026-02-04 20:16:21 +05:30
manascb1344 4c78ea9f36 feat(nebius): add new providers for Nebius Token Factory
- Add MiniMaxAI provider with MiniMax-M2.1 model
- Add PrimeIntellect provider with INTELLECT-3 model
- Add black-forest-labs provider with FLUX.1-schnell and FLUX.1-dev
- Add BAAI provider with bge-multilingual-gemma2 and BGE-ICL
- Add intfloat provider with e5-mistral-7b-instruct
- Add Google provider with Gemma-2-2b-it, Gemma-2-9b-it-fast, Gemma-3-27b-it, and Gemma-3-27b-it-fast
2026-02-04 20:16:01 +05:30
Riccardo Giorato 9092f0b106 Delete Kimi-K2-5.toml 2026-02-04 11:22:49 +01:00
Zain Hasan 1acd3c199a Rename model to 'Qwen3 Coder Next FP8' 2026-02-04 01:57:58 -08:00
Alex-wuhu 7deb00a333 feat: add deepseek OCR model configuration and Qwen3 Coder Next model configuration 2026-02-04 16:56:28 +08:00
samsja f180f49df5 Add Intellect 3 model from Prime Intellect 2026-02-03 23:58:12 -08:00
Aiden Cline 59f13d1c0a feat: make all openrouter models use openrouter sdk 2026-02-03 23:15:41 -06:00
Aiden Cline ce6950074f Revert "Add Bedrock cross-region inference profiles and update validation"
This reverts commit 89f62005cc.
2026-02-03 23:08:05 -06:00
Aiden Cline b2b0f612f4 Merge pull request #788 from zainhas/dev
[Together AI] add qwen3 coder next
2026-02-03 22:48:52 -06:00
Aiden Cline d1e92ce8ad Merge pull request #790 from anomalyco/update-cf-workers
fix: update cf workers ai
2026-02-03 22:48:41 -06:00
Aiden Cline 20c81eb600 fix: update cf workers ai 2026-02-03 22:47:12 -06:00
Frank 6934bf2c66 Merge pull request #789 from qychen2001/dev
feat(models): add Kimi-K2.5 model support
2026-02-03 22:55:46 -05:00
QiyuanChen 83503944ba feat(models): add Kimi-K2.5 model support
Add support for Moonshot AI's Kimi-K2.5 model with reasoning capabilities,
structured output, and multi-modal support (text/image input, text output).
Configured with a large context window of 262,000 tokens for both input
and output. Added to both SiliconFlow and SiliconFlow CN providers.
2026-02-04 11:45:28 +08:00
Zain Hasan 536ac44708 add qwen3 coder next 2026-02-03 14:54:42 -08:00
Aiden Cline 02f7969d53 Merge pull request #786 from unexge/push-lpupkorvtnuw
Update Amazon Bedrock models to add cross-region inference and remove deprecated models
2026-02-03 15:30:30 -06:00
Matt Silverlock 0ba8852f91 use official ai-gateway-provider package for Cloudflare AI Gateway 2026-02-03 15:41:11 -05:00
Burak Varlı 89f62005cc Add Bedrock cross-region inference profiles and update validation
- Add Nova models for Global, US, EU, and APAC regions
- Add Llama 3.1/3.2 cross-region profiles for US and EU
- Add Claude Sonnet 4/3.7 APAC profiles
- Add Claude Sonnet 4.5/3.7/3.5 US Gov profiles
- Update validate-bedrock to include ap-southeast-1 region
- Skip us-gov models in validation (requires GovCloud access)
2026-02-03 20:10:40 +00:00
Aiden Cline 5afc754db3 Merge pull request #697 from berget-ai/feat/add-berget-ai-provider
feat: add Berget.AI provider
2026-02-03 12:15:36 -06:00
Aiden Cline 2fdfeecfc8 Merge pull request #784 from bendews/patch-1
Increase Github Copilot GPT 4.1 context limit from 64k to 128k
2026-02-03 09:24:29 -06:00
Aiden Cline 7bf852e19f Merge pull request #785 from thePrnvBot/chore--updating-free-openrouter-models
fix: Update models tool call to false
2026-02-03 09:24:07 -06:00
Burak Varlı bc58036964 Update Amazon Bedrock models to add cross-region inference and remove deprecated models
This change adds a new script to validate all Amazon Bedrock models by making a simple inference request using model identifiers.
As a result of that script, made some changes to make sure all model identifiers are usable via Amazon Bedrock:
- Added cross-region inference for various models including DeepSeek, Llama, Amazon Nova
- Removed some reprecated/EoL'd models including Amazon Titan, Claude v2, Cohere Command Light
2026-02-03 13:43:07 +00:00
thePrnvBot 613843b529 fix: update nousresearch model tool call to false 2026-02-03 16:12:39 +04:00
thePrnvBot 836b07aeaf fix: update cognitivecomputation model tool call to false 2026-02-03 16:12:10 +04:00
thePrnvBot e4f8c752ac fix: update allenai model tool call to false 2026-02-03 16:11:48 +04:00
thePrnvBot fa04882d5f fix: update liquid models tool call to false 2026-02-03 16:11:36 +04:00
thePrnvBot e69df42539 fix: update llama model tool call to false 2026-02-03 16:11:18 +04:00
thePrnvBot 64b7e989eb fix: update tng-r1t-chimera :free tool call to false 2026-02-03 16:10:53 +04:00
Ben Dews abf1259e58 Increase context limit from 64k to 128k 2026-02-03 21:22:15 +10:00
Aiden Cline 93fe136ec1 Merge pull request #777 from xiaojiezj/zenmux_dev
feat: ““Replace the chat-completion protocol in the Zenmux provider with the Anthropic protocol, and replace the model.”
2026-02-02 20:50:32 -06:00
Aiden Cline be098329c5 Merge pull request #782 from fhennerkes/dev
poe: model update 2/2/26
2026-02-02 20:49:42 -06:00
fhennerkes 884e901c3f poe: model update 2/2/26 2026-02-02 18:22:49 -08:00
Michel Pigassou 81ddc26ed0 feat(deepinfra): add GLM-4.7-Flash model
Add zai-org/GLM-4.7-Flash to DeepInfra provider

Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode)

Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai>
2026-02-02 22:04:36 +01:00
Aiden Cline e35973bdf9 Merge pull request #775 from Track07-cda/alibaba-cn_kimi
feat(alibaba-cn): add Kimi K2 Thinking and K2.5 models to alibaba-cn provider
2026-02-02 10:52:51 -06:00
Aiden Cline f94d9dda7f Merge pull request #607 from unexge/push-svvwrlmunkkt
Add cross-region inference profiles for Claude 4.x family models in Amazon Bedrock
2026-02-02 10:20:13 -06:00
xiaojie.zj d354b55137 feat: “Replace the chat-completion protocol with the Anthropic protocol, and replace the model.” 2026-02-02 18:03:42 +08:00
Track07-cda e30f361b88 feat(alibaba-cn): add reasoning content support for Kimi models
Add interleaved reasoning_content field to Kimi K2 Thinking and K2.5.
Also correct the display name for Moonshot Kimi K2.5.
2026-02-02 16:28:02 +08:00
Track07-cda efa90f6fbf feat(alibaba-cn): add Kimi K2 Thinking and K2.5 models
Add new model definitions for Moonshot Kimi K2 Thinking and K2.5.
Update Moonshot Kimi K2 Instruct metadata including open weights
status and output token limits.
2026-02-02 14:13:06 +08:00
Aiden Cline 93d03d87c1 Merge pull request #772 from thePrnvBot/chore--updating-free-openrouter-models
feat: add free openrouter models
2026-01-31 21:41:49 -06:00
Aiden Cline 67b31f3371 Merge pull request #773 from ccurme/cc/gpt-5.2-structured-output
fix: add structured_output to gpt-5.1 and 5.2
2026-01-31 20:58:12 -06:00
Aiden Cline 0513b73b17 fix: correct model id 2026-01-31 20:39:40 -06:00
Chester Curme 866974df3a add structured_output to gpt-5.1 and 5.2 2026-01-31 21:39:31 -05:00
thePrnvBot fd07fe7953 feat: add free qwen models to openrouter provider 2026-01-31 20:25:30 +04:00
thePrnvBot d176299fdf feat: add free nemotron models to openrouter provider 2026-01-31 20:24:53 +04:00
thePrnvBot 6193824e92 feat: add gpt oss free models to openrouter 2026-01-31 20:23:25 +04:00
thePrnvBot b85b481fd0 feat: add hermes 3 llama 3.1 405b free model 2026-01-31 20:22:54 +04:00
thePrnvBot e33d225a0b chore: update deepseek r1 0528 free tool call to false 2026-01-31 20:22:02 +04:00
thePrnvBot 3c9e76cf89 feat: add tng-r1t-chimera free model 2026-01-31 20:21:22 +04:00
thePrnvBot 42e7f67b67 feat: add dolphin mistral 24b venice edition 2026-01-31 20:20:53 +04:00
thePrnvBot 0e46820a00 feat: add seedream model 2026-01-31 20:20:16 +04:00
thePrnvBot e7dd66e51a feat: add free meta llama models 2026-01-31 20:19:28 +04:00
thePrnvBot ea1d856847 feat: add free black forest lab models 2026-01-31 20:18:46 +04:00
thePrnvBot 26fd570130 feat: added free liquid, sourceful and allenai models 2026-01-31 20:17:30 +04:00
Aiden Cline 008c521304 Merge branch 'dev' into feat/add-vercel-models-jan-30 2026-01-30 16:35:40 -06:00
Aiden Cline c6870e97c5 Merge pull request #717 from MichaelYochpaz/fix-vertex-anthropic-npm-import
fix(google-vertex-anthropic): Fix incorrect NPM package used for Anthropic models used through Vertex
2026-01-30 15:57:19 -06:00
jerilynzheng e7e8af6934 vercel: add 5 new models from Vercel AI Gateway
Add new models:
- alibaba/qwen3-max-thinking: Qwen 3 Max with reasoning
- arcee-ai/trinity-large-preview: Trinity 400B MoE model
- moonshotai/kimi-k2.5: Kimi K2.5 with vision and reasoning
- openai/gpt-4o-mini-search-preview: GPT-4o Mini search variant
- zai/glm-4.7-flashx: GLM 4.7 Flash lightweight model

Co-Authored-By: Claude Sonnet 4.5 <noreply@anthropic.com>
2026-01-30 13:38:07 -08:00
Aiden Cline 665a9fe17d Merge pull request #767 from remorses/model-schema
add model-schema.json endpoint for model autocomplete
2026-01-30 13:32:54 -06:00
Aiden Cline 98114d5721 fix or glm flash 2026-01-30 13:21:57 -06:00
Aiden Cline 2deedb5fa5 Merge pull request #762 from davidcharbonnier/dev
Add GLM 4.7 Flash model on Openrouter
2026-01-30 13:20:33 -06:00
Aiden Cline 43cb68a64e Merge pull request #768 from amazon-nova-api/nova-provider
Add nova as a model provider
2026-01-30 13:19:57 -06:00
Adnan Hajar ee89f7ee6b Add nova as a model provider 2026-01-30 14:01:14 -05:00
Tommy D. Rossi 6f45f3949f add model-schema.json endpoint for model autocomplete 2026-01-30 15:51:04 +01:00
David Charbonnier ebafef01d7 feat: add glm 4.7 flash model on openrouter 2026-01-30 09:06:48 -05:00
captain1379 77dccbe959 feat: add Jiekou.AI provider
Add Jiekou.AI as a new LLM provider with 102 models including:
- DeepSeek (V3, R1, OCR)
- Qwen (Qwen3, Qwen2.5)
- Claude (Opus, Sonnet, Haiku)
- GPT models (GPT-5.x, GPT-4.x, GPT-OSS)
- Gemini (Pro, Flash)
- GLM (4.5, 4.7)
- Kimi (K2, K2.5)
- Llama (3.x, 4.x)
- And more...

Jiekou.AI is an OpenAI-compatible API provider.

Co-Authored-By: Claude (pa/claude-opus-4-5-20251101) <noreply@anthropic.com>
2026-01-30 18:44:11 +08:00
Frank 8b2b4b40a1 update zen models 2026-01-30 00:52:31 -05:00
Frank 96da5d8331 update zen models 2026-01-29 16:46:15 -05:00
Frank 0146cb114e update zen models 2026-01-29 16:38:42 -05:00
Aiden Cline 21177b3f6b Merge pull request #760 from cgilly2fast/dev
feat(firmware): add kimi models and clean up model names
2026-01-29 15:02:50 -06:00
Colby Gilbert 522c486815 fix: wrong name for kimi k2.5 2026-01-29 12:16:15 -08:00
Colby Gilbert 85ef0fe0a8 feat: add kimi models 2026-01-29 10:33:25 -08:00
Colby Gilbert 70681d3398 chore: rename glm and gpt oss models 2026-01-29 10:33:17 -08:00
Frank c2a6830fde sync 2026-01-29 12:38:07 -05:00
Frank 4b9631cb89 update zen models 2026-01-29 12:35:06 -05:00
Aiden Cline 9efb6c1a73 Merge pull request #742 from 888-wzk/feature/chenger_20260128
feat(models): Add configuration files for the Kimi K2.5, GPT-5.2-Codex, Qwen3-Max-Thinking, and GLM 4.7 FlashX models.
2026-01-29 10:45:16 -06:00
Aiden Cline 9b5adb8230 Merge pull request #751 from cravenceiling/fix/openrouter-google-gemma-models
add and fix some google gemma models from openrouter
2026-01-29 10:44:52 -06:00
Aiden Cline 9105b7ba75 Merge pull request #753 from otterDeveloper/patch-1
fireworks: Raise Kimi K2.5 max output
2026-01-29 10:44:40 -06:00
Aiden Cline a0dc4149cd Merge pull request #754 from fanweixiao/dev
fix(vivgrid): set npm for gemini-3 models for vivgrid provider
2026-01-29 10:43:38 -06:00
Aiden Cline ccb98b9597 Merge pull request #755 from friendliai/minpeter/add-minimax-friendli-model
feat(friendli): add MiniMax M2.1 model and update Qwen3
2026-01-29 10:42:14 -06:00
Aiden Cline c33581d89d Merge pull request #757 from s-scheck/feature/adjust-pricing-of-devstral-2512
feat: adjust pricing of devstral-2512 hosted by mistral
2026-01-29 10:41:58 -06:00
Aiden Cline ab1a8c21db Merge pull request #759 from FrancoStino/patch-5
Delete providers/nvidia/models/z-ai/glm-4.7.toml
2026-01-29 10:41:45 -06:00
Davide Ladisa f969e060c8 Delete providers/nvidia/models/z-ai/glm-4.7.toml
Duplicate

https://github.com/anomalyco/models.dev/blob/dev/providers/nvidia/models/z-ai/glm4.7.toml
2026-01-29 17:17:59 +01:00
Sinan Scheck 66823bcd6e chore: adjust pricing of devstral-2512 hosted by mistral 2026-01-29 11:45:20 +01:00
minpeter 01f391f1fe Add MiniMax M2.1 model and update Qwen3 date
Co-Authored-By: Claude Opus 4.5 <noreply@anthropic.com>
2026-01-29 18:55:00 +09:00
minpeter 1d2ddf9b1d Update friendli provider model configs
Remove outdated Qwen3 and Llama 4 model configurations.
Reorganize meta-llama models into subdirectory.
Upgrade GLM model to 4.7 with increased context limits (202,752 tokens).
2026-01-29 18:42:08 +09:00
C.C. Fan a34c2a390b fix(vivgrid): set npm for gemini-3 models for vivgrid provider 2026-01-29 16:43:32 +08:00
Frank 0c0d917719 sync 2026-01-29 03:16:34 -05:00
otterDeveloper e5dcee15bb fireworks: increase kimi-k2p5 max output 2026-01-29 02:04:24 -06:00
城二 0894de420e feat(models): Add family and interleaved fields to Kimi K2.5 and GLM 4.7 FlashX configuration files 2026-01-29 10:36:04 +08:00
Aiden Cline abfad44131 Merge pull request #734 from riccardogiorato/dev
[together.ai] add four new models: Qwen3-235B, Qwen3-Next-80B, Kimi-K2-Instruct, GLM 4.7
2026-01-28 20:13:26 -06:00
Aiden Cline cc3566a9a2 Merge pull request #747 from monotykamary/update-kimi-k2.5-model
feat(synthetic): update Kimi K2.5 model configuration
2026-01-28 20:12:59 -06:00
Aiden Cline 71b90fe6ce Merge pull request #748 from dpuyosa/venice
Venice: Adjust limit context/output token limits for all models
2026-01-28 20:12:42 -06:00
Aiden Cline c854148259 Merge pull request #749 from esafak/chore/kimi-k2.5
chore: disable `temperature` in `moonshotai/kimi-k2.5`
2026-01-28 20:12:30 -06:00
Aiden Cline 980b64e860 Merge pull request #750 from esafak/feat/zai-glm-4.7-flash
feat: Add `zai-coding-plan/glm-4.7-flash`
2026-01-28 20:12:12 -06:00
cravenceiling 632fc61b95 add to the limit section 2026-01-28 19:33:02 -05:00
cravenceiling 5a76cee029 add and fix some google gemma models from openrouter 2026-01-28 19:07:09 -05:00
Emre Şafak 0971bd00f3 feat: Add zai-coding-plan/glm-4.7-flash.
* Create the file `providers/zai-coding-plan/models/glm-4.7-flash.toml` to define the new model.
* Set the model name to `GLM-4.7-Flash`.
* Configure model capabilities including reasoning, tool call, and knowledge cutoff of `2025-04`.
* Define context limit as `200_000` tokens.
* Set input and output costs to zero.
2026-01-28 18:44:46 -05:00
Emre Şafak 94fca7064d chore: disable temperature in moonshotai/kimi-k2.5 2026-01-28 18:11:32 -05:00
dpuyosa 40ad678ae7 [venice] Adjust limit context/output token limits for all models
- All models limit context/output was reduced by 2.4%
2026-01-28 18:34:59 +01:00
Tom X Nguyen e18738adda feat(synthetic): update Kimi K2.5 model configuration
Update model config with corrected values:
- max_output: 65_536 (from 32_768)
- cost.input: 0.55, cost.output: 2.19
- modalities.input: [text, image]
- Add interleaved section for reasoning_content
- Use underscores for large numbers
2026-01-28 23:39:59 +07:00
Aiden Cline 0b003c9c18 add new arcee models 2026-01-28 10:50:59 -05:00
Aiden Cline cb5035a8ee Merge pull request #743 from zainhas/dev
add kimi k2.5
2026-01-28 10:30:17 -05:00
Aiden Cline 01e3069901 Merge pull request #745 from reissbaker/kk25
Add Kimi K2.5 for Synthetic
2026-01-28 10:29:59 -05:00
Aiden Cline 7b577ad7a8 Merge pull request #737 from cravenceiling/add-google-gemma-3-27b-it-free
feat: add google-gemma-3-27b-it:free model
2026-01-28 10:29:21 -05:00
Matt Baker 55e72aa8db Correct output tokens 2026-01-28 02:08:58 -08:00
Matt Baker a8fc1d2e44 Add Kimi K2.5 for Synthetic 2026-01-28 02:07:52 -08:00
Riccardo Giorato 101b5042bd Create Kimi-K2-5.toml 2026-01-28 10:57:37 +01:00
Riccardo Giorato 7b835b2f29 Merge remote-tracking branch 'upstream/dev' into dev 2026-01-28 10:50:29 +01:00
Aiden Cline 9658a500f7 Merge pull request #740 from thatoddmailbox/dev
Fix Kimi pricing for fireworks-ai
2026-01-28 02:24:39 -05:00
Aiden Cline 7030a7c77f Merge pull request #741 from Alex-wuhu/dev
feat(models): add Kimi K2.5 and GLM-4.7-Flash model
2026-01-28 02:24:12 -05:00
Frank 2be2a8c109 Update zai models 2026-01-28 01:49:08 -05:00
Frank 465335102f Update kimi-k2.5.toml 2026-01-28 01:38:18 -05:00
城二 82ddcad9f8 feat(models): Add configuration files for the Kimi K2.5, GPT-5.2-Codex, Qwen3-Max-Thinking, and GLM 4.7 FlashX models. 2026-01-28 14:31:21 +08:00
Zain Hasan 9acd2ffa60 add kimi k2.5 2026-01-27 22:30:24 -08:00
Alex-wuhu b3d2cfdc34 feat(models): add Kimi K2.5 and GLM-4.7-Flash model 2026-01-28 14:25:53 +08:00
Alex Studer eead89fd8e fix kimi pricing for fireworks-ai 2026-01-28 01:20:01 -05:00
Aiden Cline 48de510380 Merge pull request #635 from mthezi/feature/add-302ai-provider
feat: add 302ai provider
2026-01-27 22:01:19 -05:00
Aiden Cline 52332705ca Merge pull request #736 from alissonlauffer/chore/update-chutes-kimi-k2.5
feat(chutes): update Kimi K2.5 TEE model capabilities
2026-01-27 22:00:07 -05:00
Aiden Cline 08e5d0f830 Update Kimi-K2.5-TEE.toml configuration settings 2026-01-27 21:59:39 -05:00
Aiden Cline 0b7f253ee0 Merge pull request #739 from xinrui-z/feat/aihubmix-add-models
feat(models): add kimi-k2.5, coding-glm-4.7, glm-4.6v, and qwen3-max
2026-01-27 21:54:59 -05:00
Xinrui bfe953d2f0 feat(models): add kimi-k2.5, coding-glm-4.7, glm-4.6v, and qwen3-max 2026-01-28 10:48:40 +08:00
cravenceiling 641fa6f2e7 feat: add google-gemma-3-27b-it:free model 2026-01-27 19:34:05 -05:00
Alisson Lauffer c344db1bf8 feat(chutes): update Kimi K2.5 TEE model capabilities
Enable reasoning, tool calling, and multimodal input support for the
Kimi K2.5 TEE model. Increase context limit from 32k to 262k tokens and
output limit from 8k to 65k tokens. Add support for image and video
inputs alongside text. Configure interleaved reasoning content field.
2026-01-27 21:05:33 -03:00
Aiden Cline 36c6206d32 Merge pull request #732 from mmealman/add_fireworks_k2p5
Added Kimi K2.5 to FireworksAI.
2026-01-27 17:54:31 -05:00
Aiden Cline 4f6a59d7be Merge pull request #729 from gary149/feat/huggingface-kimi-k2.5
feat(huggingface): add Kimi-K2.5 model
2026-01-27 17:54:15 -05:00
Aiden Cline f5b8e3fe83 Merge pull request #735 from spiffytech/dev
Add Kimi K2.5 to Ollama Cloud
2026-01-27 17:53:31 -05:00
Aiden Cline 07c70ca9f2 Merge pull request #731 from arguiot/add-vercel-kimi-k2.5
Add Kimi K2.5 to Vercel provider
2026-01-27 17:53:22 -05:00
Aiden Cline d35ad7ec49 Merge pull request #727 from ProlowN/dev
fix : removed duplicate kimi k2.5 model from venice
2026-01-27 17:53:09 -05:00
Aiden Cline 7c57f4ce15 Merge pull request #733 from dpuyosa/dev
Venice: Add interleaved thinking to k2.5
2026-01-27 17:52:44 -05:00
spiffytech 83eeb304a5 Add Kimi K2.5 to Ollama Cloud 2026-01-27 16:08:16 -05:00
Aiden Cline a13f101e0c Merge pull request #638 from jerome-benoit/feature/sap-ai-core-updates
fix(sap-ai-core): use working provider fork for stable OpenCode integration
2026-01-27 15:32:40 -05:00
Riccardo Giorato ec6101a629 Merge remote-tracking branch 'upstream/dev' into dev 2026-01-27 21:00:55 +01:00
Riccardo Giorato 231313aad0 Add four new models: Qwen3-235B, Qwen3-Next-80B, Kimi-K2-Instruct, GLM-4.7 2026-01-27 21:00:44 +01:00
dpuyosa b9793731e6 Add interleaved thinking to k2.5 2026-01-27 20:32:38 +01:00
Frank 22edc4d92d update zen model 2026-01-27 14:12:42 -05:00
Frank 3f62b2dd5a update moonshot models 2026-01-27 14:05:13 -05:00
Mark Mealman d57592dba3 Added Kimi K2.5 to FireworksAI. 2026-01-27 14:00:19 -05:00
Frank e2b43f180c Merge pull request #730 from esafak/moonshotai/kimi-k2.5
chore: add `moonshotai/kimi-k2.5` model
2026-01-27 13:59:59 -05:00
Arthur Guiot 0acff9cf7c add Kimi K2.5 to Vercel provider 2026-01-27 10:50:37 -08:00
Emre Şafak 563c43f004 add moonshotai/kimi-k2.5 model 2026-01-27 13:46:30 -05:00
Victor Muštar e53bb9c7ad feat(huggingface): add Kimi-K2.5 model 2026-01-27 18:38:18 +01:00
Frank 1522bc4a9a update zen models 2026-01-27 12:34:50 -05:00
Frank c28701d579 update zen models 2026-01-27 12:34:29 -05:00
Magnus eb5bff1f6a fix : removed duplicate kimi k2.5 model from venice 2026-01-27 18:01:37 +01:00
Frank 15b4b02e6e update zen models 2026-01-27 12:00:36 -05:00
Jan Szypulski 22d6a24c7a fix: llama 3.3 last update 2026-01-27 18:00:11 +01:00
Jan Szypulski c7bc5b7c98 fix: corrected logo color and size 2026-01-27 17:59:55 +01:00
Aiden Cline b1910161d4 Merge pull request #726 from ProlowN/dev
Added kimi k2.5 to Venice AI
2026-01-27 11:45:18 -05:00
Magnus 13e48c2ca0 fix/ wrong output size 2026-01-27 17:44:27 +01:00
Magnus 516cfe355d fix/ wrong family name 2026-01-27 17:09:38 +01:00
Magnus 068eacd6b3 Added kimi k2.5 to Venice AI 2026-01-27 17:06:02 +01:00
Aiden Cline 336e43494b Merge pull request #719 from Jakey-Jakey/dev
add-kimi-k2.5 from OpenRouter
2026-01-27 11:05:16 -05:00
Aiden Cline 7fc046f833 Merge pull request #720 from kassieclaire/add-kimi-k2p5-model
feat(providers): add Kimi K2.5 model
2026-01-27 11:04:43 -05:00
Jan Szypulski 2c63a024b3 delete unrecognized model family 2026-01-27 17:04:36 +01:00
Aiden Cline a572cf8a1a Merge pull request #721 from matthusby/dev
[Chutes] Add new model configs and update pricing for several models
2026-01-27 11:03:54 -05:00
Aiden Cline e335f919f2 Merge pull request #722 from FrancoStino/dev
feat(providers): Add NVIDIA models: Kimi K2.5 and GLM-4.7
2026-01-27 11:03:38 -05:00
Aiden Cline 4e883ea026 Merge branch 'dev' into dev 2026-01-27 11:02:10 -05:00
Aiden Cline 3dfb74d1ea Merge pull request #723 from arshadbarves/feat/nvidia-kimi-k2.5
feat(nvidia): add Kimi K2.5 multimodal model
2026-01-27 11:01:42 -05:00
Aiden Cline edb551b275 Merge pull request #724 from dpuyosa/dev
Venice: Add Kimi K2.5 model configuration
2026-01-27 11:01:29 -05:00
Jan Szypulski 2f34ee47ee add cloudferro logo 2026-01-27 16:48:42 +01:00
Jan Szypulski f6cb6631b8 add cloudferro sherlock models 2026-01-27 16:48:28 +01:00
dpuyosa d95d22e89c [venice] Add Kimi K2.5 model configuration
- Add new Kimi K2.5 model with 262K context support
- Include pricing for input, output, and cache_read operations
- Enable reasoning, tool calling, and structured output capabilities
- Support text and image input with text output
2026-01-27 16:23:08 +01:00
Davide Ladisa dc771f54df Update knowledge and release dates in kimi-k2.5.toml 2026-01-27 15:52:43 +01:00
Arshad Barves 27b99e9ccf feat(nvidia): add Kimi K2.5 multimodal model
Add Kimi K2.5, a 1T parameter multimodal MoE model by Moonshot AI
with support for text, image, and video inputs.

Key features:
- 256K context window (262,144 tokens)
- Native multimodal support (text, image, video)
- Interleaved reasoning with reasoning_content field
- Tool calling and temperature control
- Open weights available

Model ID: moonshotai/kimi-k2.5
Provider: NVIDIA NIM
Validation:  Passes bun validate
2026-01-27 20:00:42 +05:30
Davide Ladisa c04069b5a3 Merge pull request #102 from FrancoStino/add-nvidia-models-kimi-glm
Add NVIDIA models: Kimi K2.5 and GLM-4.7
2026-01-27 15:01:43 +01:00
Davide Ladisa af08a750b2 Add GLM-4.7 with correct filename and family field 2026-01-27 15:01:13 +01:00
Davide Ladisa ed4270cc8f Remove old glm4_7.toml to rename to glm-4.7.toml 2026-01-27 15:01:03 +01:00
Davide Ladisa dae3873284 Fix GLM-4.7 release date to December 2025 and update knowledge cutoff 2026-01-27 14:59:23 +01:00
Davide Ladisa 6daad4c4eb Update knowledge cutoff dates to more accurate values 2026-01-27 14:57:07 +01:00
Davide Ladisa 5302dc4452 Add NVIDIA models: Kimi K2.5 and GLM-4.7 2026-01-27 14:53:55 +01:00
Matt Husby 7d26d504ec Add new model configs and update pricing for several models 2026-01-27 07:58:35 -05:00
kassieclaire 63f116b27f fix: remove interleaved reasoning for kimi-k2.5 2026-01-27 07:05:35 -05:00
kassieclaire bf6582bb83 fix: update knowledge cutoff to 2025-01 for kimi-k2.5 2026-01-27 06:25:46 -05:00
Kassie Povinelli 965f5365bc Update providers/kimi-for-coding/models/k2p5.toml
checked docs for kimi-for-coding plan, still shows up as this lower value, so going with it for now -- keep an eye on the docs in case they update the information

Co-authored-by: Copilot <175728472+Copilot@users.noreply.github.com>
2026-01-27 06:19:39 -05:00
kassieclaire 3b5717f6b9 feat(providers): add Kimi K2.5 model 2026-01-27 06:12:01 -05:00
Jakey-Jakey cae7be4925 Add 'video' modality to input options 2026-01-27 04:01:00 -05:00
Jakey-Jakey d218bb7fe8 Add cache_read cost to kimi-k2.5 configuration 2026-01-27 03:56:18 -05:00
Jakey-Jakey 9f5b80af72 Add knowledge parameter with value '2025-01' 2026-01-27 03:55:37 -05:00
Jakey-Jakey 6c928db808 Remove knowledge field from kimi-k2.5.toml
Remove knowledge field from configuration.
2026-01-27 03:54:38 -05:00
Jakey-Jakey 63b1cfb9fc Add provider section to kimi-k2.5.toml 2026-01-27 03:53:43 -05:00
Jakey-Jakey ade61a5018 Add files via upload 2026-01-27 03:45:43 -05:00
Aiden Cline e768c2afbd Merge pull request #713 from qychen2001/dev
feat(providers): update siliconflow-cn model catalog
2026-01-26 21:01:47 -05:00
Aiden Cline 31e503a516 Merge pull request #715 from dpuyosa/dev
Venice: Add cache_read to GLM 4.7
2026-01-26 21:00:43 -05:00
Frank 98a455cb0f sync 2026-01-26 18:24:50 -05:00
Michael Yochpaz f1d2e47772 fix(google-vertex-anthropic): use @ai-sdk/google-vertex/anthropic npm package
The google-vertex-anthropic provider requires the `/anthropic` subpath import for thinking/reasoning to work correctly with Claude models on Vertex AI.
2026-01-26 22:09:05 +00:00
dpuyosa 4d82211cea Add cache_read to GLM 4.7 2026-01-26 22:59:16 +01:00
mthezi 4e0a2d34b4 refactor(models): update family names for various models to improve consistency 2026-01-26 14:47:20 +08:00
⌞L⌝ effa34d17b Merge branch 'anomalyco:dev' into feature/add-302ai-provider 2026-01-26 14:29:57 +08:00
QiyuanChen ed59411f9e feat(providers): update siliconflow-cn model catalog
Add new Pro tier models for deepseek-ai and moonshotai, including DeepSeek-R1, DeepSeek-V3 series, and Kimi-K2-Thinking models with reasoning capabilities. Remove older Qwen, Kimi-K2, and other legacy model configurations.
2026-01-26 12:57:56 +08:00
Aiden Cline 1286f6449c Merge pull request #710 from hsyysy/dev
feat(provider): add DeepSeek-V3.2 for Nvidia
2026-01-25 22:56:41 -05:00
Aiden Cline f92551d3ac Merge pull request #709 from fanweixiao/feat/add-vivgrid-models
add gpt-5.1-codex-max, gpt-5.2-codex and more models for vivgrid provider
2026-01-25 22:56:31 -05:00
Aiden Cline 8c502a36b9 Delete pnpm-lock.yaml 2026-01-25 21:29:19 -05:00
Aiden Cline 6e40a4744a Merge pull request #712 from xinrui-z/fix/aihubmix-provider-invalid-type
fix(provider): correct invalid type in provider.toml
2026-01-25 21:28:54 -05:00
Xinrui 5ac346644a fix(provider): correct invalid type in provider.toml 2026-01-26 10:12:08 +08:00
Thomas Young c03332fb3d feat(provider): add DeepSeek-V3.2 for Nvidia 2026-01-25 20:10:23 +08:00
C.C. Fan acb8319afb add gpt-5.1-codex-max, gpt-5.2-codex, gemini-3-pro-preview and gemini-3-flash-preview for vivgrid provider 2026-01-25 16:37:27 +08:00
Aiden Cline 568f5319be Merge pull request #708 from jsdtxm/feat/add-glm-4.7
feat(provider): add Pro/zai-org/GLM-4.7 for SiliconFlow-CN
2026-01-24 23:39:40 -05:00
lazy 2b331310b7 fix(qihang-ai): rename provider and fix logo to match standards
- Rename provider from qihang to qihang-ai
- Update logo to use standard size (24x24) and currentColor

Co-Authored-By: Claude Sonnet 4.5 <noreply@anthropic.com>
2026-01-25 12:34:48 +08:00
xiamin 0fb16d2cb8 feat(provider): add Pro/zai-org/GLM-4.7 for SiliconFlow-CN 2026-01-25 12:16:12 +08:00
Aiden Cline 18e555c2af Merge pull request #656 from fchange/feat/new-provider
feat: add moark provider
2026-01-24 23:15:07 -05:00
Aiden Cline 1938af666f Merge pull request #707 from arshadbarves/fix/nvidia-glm4.7-model-id
fix(nvidia): correct model ID for GLM-4.7 (z-ai/glm4.7)
2026-01-24 23:12:22 -05:00
Arshad Barves fea35d7bb4 fix(nvidia): correct model ID for GLM-4.7 (z-ai/glm4.7)
Rename model file from glm-4.7.toml to glm4.7.toml to generate the
correct model ID z-ai/glm4.7 (without dot) as per NVIDIA API specification.

The model ID is derived from the file path, so the filename must match
the exact model identifier used by the provider's API.

- Renamed: providers/nvidia/models/z-ai/glm-4.7.toml → glm4.7.toml
- Model ID: z-ai/glm-4.7 → z-ai/glm4.7
- Validation:  Passes bun validate
2026-01-25 09:34:50 +05:30
Aiden Cline 53b821523b Merge pull request #700 from vglafirov/feat/gitlab-gpt-5-2
feat(gitlab): add GPT-5.2 model definition (duo-chat-gpt-5-2)
2026-01-24 12:50:30 -05:00
Aiden Cline 813b2d57b3 Merge pull request #704 from jsdtxm/feat/add-minimax-m2-1
feat(provider): add MiniMax M2.1 for SiliconFlow
2026-01-24 12:50:18 -05:00
xiamin ee5c39bb18 fix: move MiniMax-M2.1 config 2026-01-24 22:17:29 +08:00
xiamin ec2bf4bf7c feat(provider): add MiniMax M2.1 for SiliconFlow-CN 2026-01-24 16:16:03 +08:00
xiamin d2d6bc2d6c chore: remove MiniMax-M2.1.toml symlink 2026-01-24 16:15:15 +08:00
xiamin a4978b8b1c feat(provider): add MiniMax M2.1 for SiliconFlow 2026-01-24 16:08:32 +08:00
Frank 545bf83089 update zen models 2026-01-23 23:19:34 -05:00
Vladimir Glafirov 5651a0efe1 feat(gitlab): add GPT-5.2 model definition (duo-chat-gpt-5-2) 2026-01-23 16:10:47 +01:00
Frank b5fc3e3f54 update zen models 2026-01-23 01:19:18 -05:00
Frank c8f6d7ace2 update zen models 2026-01-23 01:12:50 -05:00
Frank 4dd2e77ad1 update zen models 2026-01-23 01:05:55 -05:00
Aiden Cline e5e859ac63 fix context limit for copilto gpt-4.1 2026-01-22 19:37:27 -06:00
Christian Landgren eb98dd4305 fix: address Copilot review comments
- Change Mistral family from 'mistral' to 'mistral-small' for consistency
- Fix Llama 3.3 70B knowledge date from '2024-12' to '2023-12'
- Set tool_call to false for KB-Whisper-Large (speech-to-text models don't support tool calling)
2026-01-23 01:10:35 +01:00
Christian Landgren cb8d8e8698 chore: remove Qwen3 32B model 2026-01-23 01:05:12 +01:00
Christian Landgren 3cb9a1cd3d feat: add Berget.AI provider
Add Berget.AI as an OpenAI-compatible provider with base URL api.berget.ai/v1.

Models included:
- Text: Llama 3.3 70B, Qwen3 32B, GPT-OSS-120B, GLM 4.7, Mistral Small 3.2 24B
- Embedding: Multilingual-E5-large-instruct, Multilingual-E5-large
- Rerank: bge-reranker-v2-m3
- Speech-to-Text: KB-Whisper-Large
2026-01-23 01:01:40 +01:00
Aiden Cline 830a03e46b Merge pull request #696 from vglafirov/feat/gitlab-openai-models
fix: increase output token limit for GitLab Claude models to 64k
2026-01-22 15:20:31 -08:00
Vladimir Glafirov b21c8870a5 fix: align GitLab Claude models with native Anthropic model capabilities
Updated to match native Anthropic model definitions:
- attachment: false → true (supports image/pdf attachments)
- reasoning: false → true (supports extended thinking)
- modalities.input: ["text"] → ["text", "image", "pdf"]
- Added knowledge cutoff dates from native models
2026-01-23 00:18:27 +01:00
Vladimir Glafirov 4bcba6a482 fix: increase output token limit for GitLab Claude models to 64k
The output limit was set to 4,096 tokens which caused tool calls with
large content (like file generation) to be truncated mid-JSON.

Updated to match standard Anthropic model limits:
- duo-chat-opus-4-5: 4,096 → 64,000
- duo-chat-sonnet-4-5: 4,096 → 64,000
- duo-chat-haiku-4-5: 4,096 → 64,000
2026-01-23 00:02:49 +01:00
Aiden Cline c67ccd8def Merge pull request #694 from cgilly2fast/dev
chore: remove deepseek-coder for firmware provider
2026-01-22 11:51:38 -08:00
Colby Gilbert 7aa00eb8dc chore: remove deepseek-coder for firmware provider 2026-01-22 11:48:50 -08:00
Aiden Cline d799a6ae6e Merge pull request #692 from vglafirov/feat/gitlab-openai-models
feat(gitlab): add OpenAI GPT-5 model definitions
2026-01-22 08:52:22 -08:00
Vladimir Glafirov a770639c25 feat(gitlab): add OpenAI GPT-5 model definitions
Add GitLab Duo model definitions for OpenAI GPT-5 family:
- duo-chat-gpt-5-1: GPT-5.1 flagship model
- duo-chat-gpt-5-mini: GPT-5 Mini (cost-effective)
- duo-chat-gpt-5-codex: GPT-5 Codex (agentic coding)
- duo-chat-gpt-5-2-codex: GPT-5.2 Codex
2026-01-22 17:43:59 +01:00
Jan Szypulski d934e26168 add cloudferro sherlock as provider 2026-01-22 16:11:45 +01:00
mthezi ea20440d0d fix: update model family name for gpt-4.1-nano 2026-01-22 13:51:54 +08:00
Aiden Cline eef424f296 Merge pull request #686 from zhzy0077/nvidia-patch
Add nvidia 2 new models.
2026-01-21 16:23:25 -08:00
Aiden Cline 05415ee2ec Merge pull request #685 from spiffytech/dev
Remove duplicate GLM-4.7 model file
2026-01-21 16:20:01 -08:00
Aiden Cline 23e99a093a Merge pull request #687 from eliasto/ovhcloud/update-models
Update OVHcloud AI Endpoints models
2026-01-21 16:19:52 -08:00
Aiden Cline 02df983581 Merge pull request #688 from gitpush-gitpaid/dev
Added PDF to input modalities for gpt 5.2 codex
2026-01-21 16:19:36 -08:00
Aiden Cline 3d102d3bd9 Add 'pdf' to input modalities in gpt-5.2-codex.toml 2026-01-21 18:19:14 -06:00
gitpush-gitpaid 6307a2c223 added PDF to input modalities for gpt 5.2 codex 2026-01-21 18:30:18 -05:00
Aiden Cline d79ae1d684 chore: kill deprecated copilot models from list 2026-01-21 16:59:11 -06:00
Elias TOURNEUX 67d192dd9c Update OVHcloud AI Endpoints models 2026-01-21 08:17:47 -05:00
lazy 74cb010892 feat(qihang): add Gemini 2.5 Flash and GPT-5.2 models 2026-01-21 15:19:30 +08:00
zhzy0077 10acfc848d Add nvidia 2 new models. 2026-01-21 08:39:12 +08:00
spiffytech 5943a24d41 Remove duplicate GLM-4.7 model file 2026-01-20 14:02:19 -05:00
Aiden Cline a52b64222e Merge pull request #684 from sebastiand-cerebras/final-removal-of-glm4_6
Remove deprecated zai-glm-4.6 model (Jan 20, 2026)
2026-01-20 10:17:47 -08:00
Seb Duerr a767bf0a6d Remove deprecated zai-glm-4.6 model (Jan 20, 2026)
Thank you for your patience and understanding with our timeline adjustments! I truly appreciate your team's responsiveness and flexibility in working with us on this deprecation.

As of January 20, 2026, the zai-glm-4.6 model has been officially deprecated.
2026-01-20 09:56:33 -08:00
Aiden Cline b131f86a1f Merge pull request #666 from spiffytech/dev
Update Ollama Cloud models. Add generator for model files.
2026-01-20 08:06:40 -08:00
Aiden Cline c84e382bbe Merge pull request #679 from WSQS/dev
feat: add GLM-4.7-Flash for zhipuai provider
2026-01-20 08:03:16 -08:00
Aiden Cline 9de5f304fe Merge pull request #683 from nickdowse/dev
Fix: Fix incorrect OpenAI, Gemini prices
2026-01-20 08:03:07 -08:00
Aiden Cline 8a854771d7 Merge pull request #677 from ivivek/dev
feat: add GLM-4.7 to google-vertex
2026-01-20 08:02:57 -08:00
Aiden Cline b190cdaecc Merge pull request #678 from dpuyosa/UpdateModel
Venice: Update provider package
2026-01-20 08:02:47 -08:00
Aiden Cline 5712350b30 Merge pull request #680 from cgilly2fast/cgilly2fast/firmware-provider
feat: add cerebras glm 4.7 and gpt OSS, clean up claude model ids
2026-01-20 08:02:12 -08:00
Nick Dowse b933688a77 Fix incorrect openai, gemini prices 2026-01-20 10:06:03 -05:00
dpuyosa 64f034bb72 Update interleaved field to reasoning_content
- Change field value in claude-sonnet-45, gemini-3-flash-preview, qwen3-235b-a22b-thinking-2507, and zai-org-glm-4.7 configs
2026-01-20 15:34:42 +01:00
Frank 72de414c2f Merge pull request #681 from tars90percent/minimax-provider-names
Add MiniMax coding plan providers
2026-01-20 09:11:13 -05:00
Frank bc6698d98b sync 2026-01-20 09:10:16 -05:00
lazy b465cec21a feat: add QiHang provider with 7 models
- Add QiHang provider configuration (OpenAI-compatible API)
- API endpoint: https://api.qhaigc.net/v1
- Add 7 models:
  - gpt-5.2-codex (/bin/zsh.14/.14)
  - gpt-5-mini (/bin/zsh.04//bin/zsh.29)
  - claude-opus-4-5-20251101 (/bin/zsh.71/.57)
  - claude-sonnet-4-5-20250929 (/bin/zsh.43/.14)
  - claude-haiku-4-5-20251001 (/bin/zsh.14//bin/zsh.71)
  - gemini-3-flash-preview (/bin/zsh.07//bin/zsh.43)
  - gemini-3-pro-preview (/bin/zsh.57/.43)
- All configurations validated with bun validate
2026-01-20 16:45:35 +08:00
tars90percent e0fcf8f638 Add MiniMax coding plan providers 2026-01-20 13:42:43 +08:00
Colby Gilbert 36a6197da9 feat: add cerebras glm 4.7 and gpt OSS, clean up claude model ids 2026-01-19 21:34:34 -08:00
WSQS f6d82c43a7 feat: add GLM-4.7-Flash for zhipuai 2026-01-20 10:49:21 +08:00
dpuyosa 44b8ed5871 Comment-out 'api' for validation script 2026-01-20 01:30:13 +01:00
dpuyosa c0d9ec4777 Update Venice provider package:
- Replace @ai-sdk/openai-compatible with venice-ai-sdk-provider
- Fix cache_control limitations
- Add Venice-specific features
2026-01-20 01:11:34 +01:00
spiffytech 2ae1e23591 Update Ollama Cloud models. Add generator for model files. 2026-01-19 17:29:38 -05:00
Vivek K 1ff1405664 feat: add GLM-4.7 to google-vertex 2026-01-20 00:43:34 +05:30
Aiden Cline 1c32145339 Merge pull request #674 from zerone0x/add/gpt-5.1-codex-max
feat(openrouter): add openai/gpt-5.1-codex-max model
2026-01-19 09:52:54 -08:00
Aiden Cline fbebe356b5 Merge pull request #675 from ElecTwix/glm-4.7-flash
feat: add glm-4.7-flash model
2026-01-19 09:52:25 -08:00
ElecTwix e694f0136f feat: add glm-4.7-flash model 2026-01-19 20:44:29 +03:00
zerone0x 5062058b6a feat(openrouter): add openai/gpt-5.1-codex-max model
Add GPT-5.1-Codex-Max model to OpenRouter provider. This model is available
in OpenRouter's API but was missing from models.dev.

Pricing sourced from OpenRouter API.

Co-Authored-By: Claude <noreply@anthropic.com>
2026-01-20 01:25:22 +08:00
Aiden Cline 7b132f2cd8 Merge pull request #673 from gary149/feat/huggingface-glm-4.7-flash
feat(huggingface): add GLM-4.7-Flash model
2026-01-19 08:49:44 -08:00
Victor Muštar e319a707fd feat(huggingface): add GLM-4.7-Flash model 2026-01-19 17:36:59 +01:00
Aiden Cline 627ac7bcf1 Merge pull request #671 from uniquename/ollama/glm-4.7
feat: add Ollama GLM-4.7 model configuration file
2026-01-19 07:38:50 -08:00
Aiden Cline 89408c71e0 Merge pull request #669 from dpuyosa/UpdateModel
Venice: Replace vision models glm4.6v -> qwen3-vl
2026-01-19 07:38:29 -08:00
Aiden Cline 0651768fd9 Merge pull request #672 from sebastiand-cerebras/add-glm4_6-deprecation-notice
Re-add zai-glm-4.6 temporarily until Jan 20, 2026
2026-01-19 07:38:00 -08:00
Seb Duerr c8ce0db2b1 Re-add zai-glm-4.6 temporarily until Jan 20, 2026
Thanks for the incredibly fast merge! We appreciate the efficiency, though we need to temporarily re-add GLM 4.6. The model will be officially deprecated on January 20, 2026. Our apologies for any confusion - we should have been clearer about the timeline in the original PR.
2026-01-19 07:14:34 -08:00
User c3b177ed7a feat: add Ollama GLM-4.7 model configuration file 2026-01-19 12:45:35 +00:00
dpuyosa 1ce4dc41c1 Update model configurations:
- Add qwen3-vl-235b-a22b model
- Remove deprecated zai-org-glm-4.6v model
2026-01-19 10:49:32 +01:00
Aiden Cline 438e834043 add input field to more openai models 2026-01-19 00:59:44 -06:00
Jérôme Benoit 189aa03281 Apply suggestion from @jerome-benoit 2026-01-19 04:02:10 +01:00
Aiden Cline 5f293ca6ce Merge pull request #665 from sebastiand-cerebras/removal_of_glm4_6
Remove deprecated zai-glm-4.6 model from Cerebras provider
2026-01-17 22:50:25 -08:00
Aiden Cline 490a03f2d9 Remove deprecated zai-glm-4.6 model from Cerebras provider 2026-01-17 22:49:51 -08:00
Aiden Cline cbd215cbc1 Merge pull request #652 from hueyexe/dev
feat: add GPT 5.2 Codex to Azure and Azure Cognitive Services
2026-01-17 22:48:21 -08:00
Aiden Cline 1b49c365a0 fix: restore gpt-5.2-codex.toml as symlink to fix validation CI 2026-01-18 00:46:56 -06:00
Aiden Cline 51d2a4f2b6 Merge pull request #664 from jerome-benoit/feat/sap-ai-core-claude-4.5-opus
feat(sap-ai-core): add Claude 4.5 Opus and align pricing
2026-01-17 19:02:07 -08:00
Seb Duerr afba23ed62 Remove deprecated zai-glm-4.6 model from Cerebras provider 2026-01-17 18:37:18 -08:00
Jérôme Benoit 9e25ca1521 feat(sap-ai-core): add Claude 4.5 Opus and align pricing
- Add Claude 4.5 Opus model with official Anthropic pricing
- Align cache pricing for Claude 3 Sonnet, Gemini 2.5 models, and GPT-5 Mini with official pricing
2026-01-18 01:08:12 +01:00
Aiden Cline 971e8734ae Merge pull request #659 from KagurazakaNyaa/dev
Update SiliconFlow model list
2026-01-16 20:35:02 -08:00
Aiden Cline ef7731ec13 Merge pull request #661 from cgilly2fast/cgilly2fast/firmware-provider
fix: make gpt-nano and mini calculate as 0 price
2026-01-16 20:32:32 -08:00
Colby Gilbert e2d8670828 chore: update firmware provider docs url 2026-01-16 17:04:53 -08:00
神楽坂·喵 ecfab1717a Merge branch 'anomalyco:dev' into dev 2026-01-17 08:21:48 +08:00
KagurazakaNyaa ac037c57ad fix pangu family 2026-01-17 08:20:30 +08:00
KagurazakaNyaa a886d60715 fix kat family 2026-01-17 08:11:32 +08:00
Colby Gilbert 44e980e810 fix: make gpt-nano and mini calculate as 0 price 2026-01-16 15:42:37 -08:00
Aiden Cline d1f3ddfe44 Merge pull request #660 from jerilynzheng/feat/vercel-models-update-2
vercel: add new models from Vercel AI Gateway
2026-01-16 12:59:21 -08:00
Aiden Cline f0192d8759 Update gpt-5.2-codex.toml 2026-01-16 14:52:45 -06:00
jerilynzheng 070fd89c38 vercel: add new models from Vercel AI Gateway
- Add bytedance/seed-1.8 (multimodal with reasoning)
- Add openai/gpt-5.2-codex (agentic coding)
- Add recraft/recraft-v2 and recraft-v3 (image generation)
- Add recraft to model family schema

Co-Authored-By: Claude Opus 4.5 <noreply@anthropic.com>
2026-01-16 12:11:47 -08:00
神楽坂·喵 0ae941ee4d Merge branch 'anomalyco:dev' into dev 2026-01-17 01:03:43 +08:00
KagurazakaNyaa 3a4cc60e04 update siliconflow model list 2026-01-17 01:02:33 +08:00
Aiden Cline a243d9ba84 Merge pull request #658 from litvix-whale/feat/add-minimax-m2-1
feat(provider): add MiniMax M2.1 for DeepInfra
2026-01-16 08:15:39 -08:00
Kyrylo Lytvishko 5be35e44e4 feat(provider): add MiniMax M2.1 for DeepInfra 2026-01-16 18:04:00 +02:00
yinxulai faa78aa42b feat: add new Qiniu AI models - Claude 3.5/3.7/4.0/4.1/4.5 series, Gemini 2.0/2.5/3.0 series, GPT-5/5.2, Grok 4/4.1 series, and Kling v2-6 2026-01-16 17:46:45 +08:00
Aiden Cline 433008fef0 fix: more abacus things - fix model ids 2026-01-16 00:18:21 -06:00
franco bfb6bb315a feat: add moark provider 2026-01-16 10:26:06 +08:00
Aiden Cline 7f49452691 Merge pull request #653 from dpuyosa/UpdateModel
Venice: Update generate script & add new models (sonnet 4.5, gpt 5.2 codex)
2026-01-15 12:57:31 -08:00
Aiden Cline 6b793ad28e rm raptor mini model 2026-01-15 12:42:06 -06:00
mthezi 53d77d2f2b chore: update output limits for various models 2026-01-15 18:42:04 +08:00
dpuyosa a8436a1e8a Add new model configurations:
- Add claude-sonnet-45 model configuration
 - Add openai-gpt-52-codex model configuration
2026-01-15 10:18:29 +01:00
dpuyosa 2ef222a882 Updated model configurations:
- Changed family from 'llama' to 'hermes' in hermes-3-llama-3.1-405b
 - Changed family from 'glm' to 'glmv' in zai-org-glm-4.6v
 - Added interleaved reasoning_details field in zai-org-glm-4.6v
2026-01-15 10:16:49 +01:00
dpuyosa da6e0354da Updated family inference logic:
- Refactored family inference to use ModelFamilyValues and subsequence matching algorithm
2026-01-15 10:14:09 +01:00
Aiden Cline 5aa046c596 fix: abacus provider 2026-01-14 23:58:31 -06:00
hueyexe f54b8d8c6d Add gpt 5.2 codex to azure cognitive services 2026-01-15 16:00:15 +11:00
hueyexe a2d657f75b Add gpt 5.2 codex to azure 2026-01-15 15:58:42 +11:00
yinxulai 79636dec83 fix: add required date fields and default output limits for Qiniu AI models 2026-01-15 10:49:52 +08:00
yinxulai f9983aae19 feat: add Qiniu AI model definitions
- Add 49 OpenAI-compatible model definitions
- Models filtered from Qiniu API with OpenAI protocol support
- Include models from DeepSeek, Qwen, Kimi, GLM, Doubao, MiniMax, etc.
- No pricing information included (aggregation platform)
2026-01-15 10:38:15 +08:00
Aiden Cline b9411cb00c feat: add Qiniu AI provider configuration 2026-01-15 10:06:30 +08:00
Aiden Cline 5a329d79bc Merge pull request #650 from cgilly2fast/cgilly2fast/firmware-provider
refactor: simplify model ids so sub agents work
2026-01-14 15:18:26 -08:00
Colby Gilbert 1e9ee75804 refactor: simplify model ids so sub agents work 2026-01-14 15:07:36 -08:00
Aiden Cline 64e82beb55 Merge pull request #645 from TheEpTic/dev
chore: Add gpt-5.2-codex to GitHub Copilot provider
2026-01-14 14:59:58 -08:00
Frank 256bab07a3 update zen models 2026-01-14 16:27:53 -05:00
Frank 78fd2e0fa0 update zen models 2026-01-14 16:18:51 -05:00
Aiden Cline 969430c25e Merge pull request #647 from KonarkRajMisra/dev
Add GPT-5.2-Codex to OpenRouter
2026-01-14 12:39:08 -08:00
Aiden Cline c4b43c090d Merge pull request #646 from brandon93s/52-input
chore(openai): gpt-5.2-codex input limit
2026-01-14 12:38:52 -08:00
Konark Misra bfb92b5e46 Add GPT-5.2-Codex to OpenRouter 2026-01-14 12:15:39 -08:00
TheEpTic b0e5b914c8 Fix context size 2026-01-14 20:03:52 +00:00
Brandon Smith 66e5d76e05 input 2026-01-14 13:53:57 -06:00
TheEpTic f1f27989d8 Add gpt-5.2-codex to GitHub Copilot provider 2026-01-14 19:41:09 +00:00
Aiden Cline 949f9b9909 Merge pull request #623 from cyhhao/add-gpt-5-2-codex
feat: add gpt-5.2-codex model
2026-01-14 11:25:43 -08:00
Aiden Cline 6a614ab0ac Update model family name in gpt-5.2-codex.toml 2026-01-14 13:24:47 -06:00
Aiden Cline 664079661d Merge pull request #641 from liyishuai/iflow-cleanup
chore(iflowcn): cleanup models
2026-01-14 07:46:53 -08:00
Aiden Cline 58e2fd8462 Merge pull request #642 from brandon93s/openai-codex-input-limit
openai: codex input context limit
2026-01-14 07:31:37 -08:00
Aiden Cline 25eda4cc82 Merge pull request #612 from Alex-wuhu/dev
add LLM Provider : novita ai
2026-01-14 07:30:55 -08:00
Alex-wuhu c60ec95e75 Update model family names for consistency and clarity 2026-01-14 23:04:51 +08:00
Alex 952de0d081 Merge branch 'anomalyco:dev' into dev 2026-01-14 23:00:39 +08:00
Brandon Smith 453f16ce42 add input limit for codex models 2026-01-14 08:32:48 -06:00
Alex-wuhu 01f338231e Update LLM info 2026-01-14 19:10:34 +08:00
Yishuai Li ce48f4ee7b chore(iflowcn): cleanup models
Signed-off-by: Yishuai Li <yishuai.li@pingcap.com>
2026-01-14 16:53:58 +08:00
Aiden Cline db79e08e38 Merge pull request #636 from Eric-Guo/patch-1
Using CN in API key, so it won't loading both siliconflow-cn and siliconflow
2026-01-13 21:36:47 -08:00
Aiden Cline 71cf624135 Merge pull request #640 from fanweixiao/dev
feat(provider): Add configuration for GPT-5.1 Codex Max model to Vivgrid provider
2026-01-13 21:36:36 -08:00
Aiden Cline 9f7c0cec79 Merge pull request #639 from anomalyco/update-model-families
Update model families
2026-01-13 21:36:21 -08:00
C.C. 399b469927 Add configuration for GPT-5.1 Codex Max model 2026-01-14 02:55:08 +00:00
Aiden Cline 1a96ad9764 Merge pull request #637 from dpuyosa/UpdateModel
Venice: Updated llama-3.2-3b model configuration
2026-01-13 15:00:59 -08:00
Jérôme Benoit 0ada0ed52e fix(sap-ai-core): use temporary fork for stable OpenCode integration 2026-01-13 19:52:14 +01:00
dpuyosa 4c39b53744 Updated llama-3.2-3b model configuration:
- Removed structured_output property
2026-01-13 13:52:50 +01:00
Eric Guo d84aff0e75 Using CN in API key, so it won't loading both siliconflow-cn and siliconflow 2026-01-13 20:16:29 +08:00
mthezi c9aefb0af1 feat: add 302ai provider 2026-01-13 14:21:52 +08:00
cyhhao 94310d742d Add gpt-5.2-codex model 2026-01-11 01:55:50 +08:00
Alex-wuhu a50c04d060 Update minimax-m2.1.toml 2026-01-09 13:31:49 +08:00
Alex-wuhu cc2619dd5f add LLM Provider : novita ai 2026-01-07 19:27:42 +08:00
Burak Varlı f14775e355 Add cross-region inference profiles for Claude 4.x family models in Amazon Bedrock
Amazon Bedrock requires usage of cross-region inference for some models, especially the latest models including all Claude 4.x family.
This change creates model files for all Claude 4.x models for cross-region inference profiles for Global, US and EU.
2026-01-06 11:41:49 +00:00
Dominik Oswald c35c42fc34 Add Sonar Deep Research model configuration
- Introduce TOML configuration for Perplexity Sonar Deep Research model
- Include token pricing, request fees, and model limits
- Follow OpenCode AI schema conventions for model definitions
2025-10-17 13:12:19 +02:00
Dominik Oswald 8abedde07c Add Perplexity Sonar Deep Research model configuration
- Introduce TOML configuration for Perplexity Sonar Deep Research model
- Include token pricing, request fees, and model limits
2025-10-17 13:10:38 +02:00
2590 changed files with 44099 additions and 3670 deletions
+5 -3
View File
@@ -7,8 +7,10 @@ on:
jobs:
opencode:
if: |
contains(github.event.comment.body, '/oc') ||
contains(github.event.comment.body, '/opencode')
contains(github.event.comment.body, ' /oc') ||
startsWith(github.event.comment.body, '/oc') ||
contains(github.event.comment.body, ' /opencode') ||
startsWith(github.event.comment.body, '/opencode')
runs-on: ubuntu-latest
permissions:
contents: read
@@ -22,4 +24,4 @@ jobs:
env:
ANTHROPIC_API_KEY: ${{ secrets.ANTHROPIC_API_KEY }}
with:
model: anthropic/claude-sonnet-4-20250514
model: anthropic/claude-sonnet-4-20250514
+4
View File
@@ -1,5 +1,9 @@
.env
.sst
.idea
dist
.DS_Store
node_modules
data/tokenspeed-monitor.sqlite
data/tokenspeed-monitor.sqlite-shm
data/tokenspeed-monitor.sqlite-wal
+29 -1
View File
@@ -26,4 +26,32 @@
- Use `export interface` for API types, `export const Schema = z.object()` for validation
- Prefix unused variables with underscore or use `_` for ignored parameters
- Handle undefined values explicitly in comparisons and sorting
- Use optional chaining (`?.`) and nullish coalescing (`??`) for safe property access
- Use optional chaining (`?.`) and nullish coalescing (`??`) for safe property access
## Model Configuration
- Model `id` is **auto-injected** from filename (minus `.toml`) — never put `id` in TOML files
- Same model is duplicated across provider directories with no cross-referencing
- Schema uses `.strict()` — extra fields cause validation errors
### Bedrock Naming Patterns
- Dated models: `-v1:0` suffix (`anthropic.claude-3-5-sonnet-20241022-v1:0.toml`)
- Latest/undated models: bare `-v1` (`anthropic.claude-opus-4-6-v1.toml`)
- Region prefixes: `us.`, `eu.`, `global.` (default has no prefix)
### Vertex AI Naming Patterns
- Dated models: `@YYYYMMDD` (`claude-opus-4-5@20251101.toml`)
- Latest/undated models: `@default` (`claude-opus-4-6@default.toml`)
### Cost Schema
- `cost.context_over_200k` is a nested `Cost` object for >200K token pricing
- Cache pricing ratios: standard models use 10%/125% (read/write), regional variants may use 30%/375%
### Required vs Optional Fields
| Field | Required? | Notes |
|-------|-----------|-------|
| `name`, `release_date`, `last_updated` | Yes | Human-readable metadata |
| `attachment`, `reasoning`, `tool_call`, `open_weights` | Yes | Boolean capabilities |
| `cost`, `limit`, `modalities` | Yes | Objects with their own required fields |
| `family`, `knowledge`, `temperature`, `structured_output` | No | Optional metadata |
| `status` | No | Use for `"alpha"`, `"beta"`, `"deprecated"` lifecycle |
+12 -1
View File
@@ -109,7 +109,7 @@ output_audio = 10.00 # Cost per million audio output tokens (USD)
[limit]
context = 400_000 # Maximum context window (tokens)
context = 272_000 # Maximum input tokens
input = 272_000 # Maximum input tokens
output = 8_192 # Maximum output tokens
[modalities]
@@ -199,6 +199,17 @@ $ bun run dev
And it'll open the frontend at http://localhost:3000
### Manual testing with opencode
You can manually check provider changes with opencode by:
```bash
$ bun install
$ cd packages/web
$ bun run build
$ OPENCODE_MODELS_PATH="dist/_api.json" opencode
```
### Questions?
Open an issue if you need help or have questions about contributing.
+3 -1
View File
@@ -17,7 +17,9 @@
"scripts": {
"validate": "bun ./packages/core/script/validate.ts",
"helicone:generate": "bun ./packages/core/script/generate-helicone.ts",
"venice:generate": "bun ./packages/core/script/generate-venice.ts"
"venice:generate": "bun ./packages/core/script/generate-venice.ts",
"vercel:generate": "bun ./packages/core/script/generate-vercel.ts",
"wandb:generate": "bun ./packages/core/script/generate-wandb.ts"
},
"dependencies": {
"@cloudflare/workers-types": "^4.20250801.0",
+1 -1
View File
@@ -48,8 +48,8 @@ const familyPatterns: [RegExp, string][] = [
[/llama-4/i, "llama-4"],
[/qwen3/i, "qwen3"],
[/deepseek-r1/i, "deepseek-r1"],
[/exaone/i, "exaone"],
[/glm-4/i, "glm-4"],
[/glm-5/i, "glm"],
];
function inferFamily(modelId: string, modelName: string): string | undefined {
+237
View File
@@ -0,0 +1,237 @@
#!/usr/bin/env bun
/**
* Generates model files from the data in Ollama Cloud's API.
*
* Ollama Cloud does not provide some data fields, such as release date or
* knowledge cutoff. The `family` field provided by Ollama Cloud may not match
* the values in family.ts. We expect that when TOML validaton fails, the
* maintainer will manually source those data points (such as from other
* provider TOML files, or from the internet at large). This script preserves
* those fields when overwriting Ollama Cloud's TOML files.
*/
import { z } from "zod";
import path from "node:path";
import type { Model } from "../src/schema";
import type { ModelFamily } from "../src/family";
const modelsDir = path.join(
import.meta.dirname,
"..",
"..",
"..",
"providers",
"ollama-cloud",
"models"
);
function modelFileName(modelName: string): string {
return modelName + ".toml";
}
type OllamaModel = Omit<Model, "id"> & {
limit: Model["limit"] & { output?: number };
};
type ComparableModel = Pick<Model,
| "name"
| "attachment"
| "reasoning"
| "tool_call"
| "knowledge"
| "open_weights"
| "modalities"
> & {
limit: Pick<Model["limit"], "context">;
};
function normalizeForComparison(model: Omit<Model, "id">): ComparableModel {
return {
name: model.name,
attachment: model.attachment,
reasoning: model.reasoning,
tool_call: model.tool_call,
knowledge: model.knowledge,
open_weights: model.open_weights,
limit: { context: model.limit.context },
modalities: model.modalities,
};
}
const OllamaTagsResponse = z.object({
models: z.array(
z.object({
name: z.string(),
})
),
});
type OllamaTagsResponse = z.infer<typeof OllamaTagsResponse>;
const OllamaModelDetails = z.object({
modified_at: z.string(),
details: z.object({
parent_model: z.string(),
format: z.string(),
family: z.string(),
families: z.array(z.string()).nullable(),
parameter_size: z.string().transform(Number),
quantization_level: z.string(),
}),
model_info: z.record(z.union([z.string(), z.number()])),
capabilities: z.array(z.enum(["thinking", "completion", "tools", "vision"])),
});
type OllamaModelDetails = z.infer<typeof OllamaModelDetails>;
function generateToml(modelName: string, model: OllamaModel): string {
const lines: string[] = [];
lines.push(`name = "${modelName}"`);
lines.push(`family = "${model.family}"`);
lines.push(`attachment = ${model.attachment}`);
lines.push(`reasoning = ${model.reasoning}`);
lines.push(`tool_call = ${model.tool_call}`);
if (model.release_date) {
lines.push(`release_date = "${model.release_date}"`);
}
if (model.knowledge) {
lines.push(`knowledge = "${model.knowledge}"`);
}
lines.push(`last_updated = "${model.last_updated}"`);
lines.push(`open_weights = ${model.open_weights}`);
lines.push("");
lines.push("[limit]");
lines.push(`context = ${model.limit.context}`);
if (model.limit.output !== undefined) {
lines.push(`output = ${model.limit.output}`);
}
lines.push("");
lines.push("[modalities]");
lines.push(`input = ${JSON.stringify(model.modalities.input)}`);
lines.push(`output = ${JSON.stringify(model.modalities.output)}`);
return lines.join("\n") + "\n";
}
const tagsResponse = await fetch("https://ollama.com/api/tags");
if (!tagsResponse.ok) {
console.error(
`Failed to fetch tags: ${tagsResponse.status} ${tagsResponse.statusText}`
);
process.exit(1);
}
const tagsJson = await tagsResponse.json();
const tagsParsed = OllamaTagsResponse.safeParse(tagsJson);
if (!tagsParsed.success) {
console.error("Invalid tags response:", tagsParsed.error.errors);
process.exit(1);
}
const tagsData: OllamaTagsResponse = tagsParsed.data;
const modelNames = tagsData.models.map((m) => m.name);
console.log(`Fetching details for ${modelNames.length} models...`);
const modelsData: Array<{ name: string; data: OllamaModelDetails }> = [];
for (const modelName of modelNames) {
const showResponse = await fetch("https://ollama.com/api/show", {
method: "POST",
headers: { "Content-Type": "application/json" },
body: JSON.stringify({ model: modelName }),
});
if (!showResponse.ok) {
console.error(
`Failed to fetch details for ${modelName}: ${showResponse.status} ${showResponse.statusText}`
);
process.exit(1);
}
const showJson = await showResponse.json();
const showParsed = OllamaModelDetails.safeParse(showJson);
if (!showParsed.success) {
console.error(
`Invalid response for ${modelName}:`,
showParsed.error.errors
);
process.exit(1);
}
modelsData.push({ name: modelName, data: showParsed.data });
}
console.log(`Fetched all models. Syncing files...`);
const existingFiles = Array.from(new Bun.Glob("*.toml").scanSync(modelsDir));
const existingModelNames = new Set(existingFiles.map((f) => f.replace(/\.toml$/, "")));
const apiModelNames = new Set(modelNames);
let deleted = 0;
for (const existingName of existingModelNames) {
if (!apiModelNames.has(existingName)) {
const filePath = path.join(modelsDir, modelFileName(existingName));
await Bun.file(filePath).delete();
console.log(`Deleted: ${modelFileName(existingName)}`);
deleted++;
}
}
let created = 0;
let skipped = 0;
for (const { name, data } of modelsData) {
const fileName = modelFileName(name);
const filePath = path.join(modelsDir, fileName);
let existingData: Omit<Model, "id"> | null = null;
try {
const existingToml = await Bun.file(filePath).text();
existingData = Bun.TOML.parse(existingToml) as Omit<Model, "id">;
} catch {
// File doesn't exist
}
const family = existingData?.family ?? (data.details.family as ModelFamily);
const contextLength =
(data.model_info[`${data.details.family}.context_length`] as number) ?? 0;
const ollamaModel: OllamaModel = {
name,
family,
attachment: data.capabilities.includes("vision"),
reasoning: data.capabilities.includes("thinking"),
tool_call: data.capabilities.includes("tools"),
release_date: existingData?.release_date,
knowledge: existingData?.knowledge,
last_updated: new Date().toISOString().slice(0, 10),
open_weights: true,
modalities: {
input: data.capabilities.includes("vision")
? ["text", "image"]
: ["text"],
output: ["text"],
},
limit: {
context: contextLength,
output: existingData?.limit.output,
},
};
if (existingData) {
const normalizedExisting = normalizeForComparison(existingData);
const normalizedIncoming = normalizeForComparison(ollamaModel);
if (Bun.deepEquals(normalizedExisting, normalizedIncoming)) {
console.log(`Skipped (no changes): ${fileName}`);
skipped++;
continue;
}
}
await Bun.write(filePath, generateToml(name, ollamaModel));
console.log(`Created: ${fileName}`);
created++;
}
console.log(`\nDone. Created: ${created}, Skipped: ${skipped}, Deleted: ${deleted}`);
+87 -64
View File
@@ -3,29 +3,11 @@
import { z } from "zod";
import path from "node:path";
import { readdir } from "node:fs/promises";
import * as readline from "node:readline";
import { ModelFamilyValues } from "../src/family.js";
// Venice API endpoint
const API_ENDPOINT = "https://api.venice.ai/api/v1/models?type=text";
async function promptForApiKey(): Promise<string | null> {
const rl = readline.createInterface({
input: process.stdin,
output: process.stdout,
});
return new Promise((resolve) => {
rl.question(
"Enter Venice API key to include alpha models (or press Enter to skip): ",
(answer) => {
rl.close();
const trimmed = answer.trim();
resolve(trimmed.length > 0 ? trimmed : null);
},
);
});
}
// Zod schemas for API response validation
const Capabilities = z
.object({
@@ -42,12 +24,25 @@ const Capabilities = z
})
.passthrough();
const PricingTier = z.object({ usd: z.number(), diem: z.number().optional() }).passthrough();
const ExtendedPricing = z
.object({
context_token_threshold: z.number(),
input: PricingTier,
output: PricingTier,
cache_input: PricingTier.optional(),
cache_write: PricingTier.optional(),
})
.passthrough();
const Pricing = z
.object({
input: z.object({ usd: z.number(), diem: z.number().optional() }).passthrough(),
output: z.object({ usd: z.number(), diem: z.number().optional() }).passthrough(),
cache_input: z.object({ usd: z.number(), diem: z.number().optional() }).passthrough().optional(),
cache_write: z.object({ usd: z.number(), diem: z.number().optional() }).passthrough().optional(),
input: PricingTier,
output: PricingTier,
cache_input: PricingTier.optional(),
cache_write: PricingTier.optional(),
extended: ExtendedPricing.optional(),
})
.passthrough();
@@ -55,11 +50,13 @@ const ModelSpec = z
.object({
pricing: Pricing.optional(),
availableContextTokens: z.number(),
maxCompletionTokens: z.number().optional(),
capabilities: Capabilities,
constraints: z.any().optional(),
name: z.string(),
modelSource: z.string().optional(),
offline: z.boolean().optional(),
privacy: z.string().optional(),
traits: z.array(z.string()).optional(),
})
.passthrough();
@@ -83,31 +80,35 @@ const VeniceResponse = z
})
.passthrough();
// Family inference patterns
const familyPatterns: [RegExp, string][] = [
[/^llama-3\.3/i, "llama-3.3"],
[/^llama-3\.2/i, "llama-3.2"],
[/^qwen3/i, "qwen3"],
[/^deepseek/i, "deepseek"],
[/^mistral/i, "mistral"],
[/^devstral/i, "devstral"],
[/^gemini/i, "gemini"],
[/^grok/i, "grok"],
[/^claude/i, "claude"],
[/^hermes/i, "hermes"],
[/^google-gemma/i, "gemma"],
[/^kimi/i, "kimi"],
[/glm-4.6/i, "glm-4.6"],
[/^venice/i, "venice-uncensored"],
[/^openai-gpt/i, "openai-gpt"],
];
function matchesFamily(target: string, family: string): boolean {
const targetLower = target.toLowerCase();
const familyLower = family.toLowerCase();
let familyIdx = 0;
for (let i = 0; i < targetLower.length && familyIdx < familyLower.length; i++) {
if (targetLower[i] === familyLower[familyIdx]) {
familyIdx++;
}
}
return familyIdx === familyLower.length;
}
function inferFamily(modelId: string, modelName: string): string | undefined {
for (const [pattern, family] of familyPatterns) {
if (pattern.test(modelId) || pattern.test(modelName)) {
const sortedFamilies = [...ModelFamilyValues].sort((a, b) => b.length - a.length);
for (const family of sortedFamilies) {
if (matchesFamily(modelId, family)) {
return family;
}
}
for (const family of sortedFamilies) {
if (matchesFamily(modelName, family)) {
return family;
}
}
return undefined;
}
@@ -156,6 +157,12 @@ interface ExistingModel {
reasoning?: number;
cache_read?: number;
cache_write?: number;
context_over_200k?: {
input?: number;
output?: number;
cache_read?: number;
cache_write?: number;
};
};
limit?: {
context?: number;
@@ -207,6 +214,12 @@ interface MergedModel {
output: number;
cache_read?: number;
cache_write?: number;
context_over_200k?: {
input: number;
output: number;
cache_read?: number;
cache_write?: number;
};
};
limit: {
context: number;
@@ -226,29 +239,24 @@ function mergeModel(
const caps = spec.capabilities;
const contextTokens = spec.availableContextTokens;
const outputTokens = Math.floor(contextTokens / 4);
const outputTokens = spec.maxCompletionTokens ?? Math.floor(contextTokens / 4);
// Determine open_weights from modelSource
const openWeights = spec.modelSource
? spec.modelSource.toLowerCase().includes("huggingface")
: false;
: spec.privacy === "private";
// Build input modalities from API (no auto-PDF)
const inputModalities = buildInputModalities(caps);
// Check if existing has PDF in modalities - preserve it
if (existing?.modalities?.input?.includes("pdf") && !inputModalities.includes("pdf")) {
inputModalities.push("pdf");
}
// Determine attachment based on vision/audio/video support
const attachment =
caps.supportsVision === true ||
caps.supportsAudioInput === true ||
caps.supportsVideoInput === true;
const merged: MergedModel = {
// Always from API
name: spec.name,
attachment,
reasoning: caps.supportsReasoning === true,
@@ -280,18 +288,21 @@ function mergeModel(
...(spec.pricing.cache_input && { cache_read: spec.pricing.cache_input.usd }),
...(spec.pricing.cache_write && { cache_write: spec.pricing.cache_write.usd }),
};
}
// Preserve from existing OR infer
if (existing?.family) {
merged.family = existing.family;
} else {
const inferred = inferFamily(apiModel.id, spec.name);
if (inferred) {
merged.family = inferred;
// Extended pricing maps to context_over_200k
if (spec.pricing.extended) {
merged.cost.context_over_200k = {
input: spec.pricing.extended.input.usd,
output: spec.pricing.extended.output.usd,
...(spec.pricing.extended.cache_input && { cache_read: spec.pricing.extended.cache_input.usd }),
...(spec.pricing.extended.cache_write && { cache_write: spec.pricing.extended.cache_write.usd }),
};
}
}
const inferred = inferFamily(apiModel.id, spec.name);
merged.family = inferred ?? existing?.family;
// Preserve manual fields from existing
if (existing?.knowledge) {
merged.knowledge = existing.knowledge;
@@ -354,6 +365,19 @@ function formatToml(model: MergedModel): string {
if (model.cost.cache_write !== undefined) {
lines.push(`cache_write = ${model.cost.cache_write}`);
}
if (model.cost.context_over_200k) {
lines.push("");
lines.push(`[cost.context_over_200k]`);
lines.push(`input = ${model.cost.context_over_200k.input}`);
lines.push(`output = ${model.cost.context_over_200k.output}`);
if (model.cost.context_over_200k.cache_read !== undefined) {
lines.push(`cache_read = ${model.cost.context_over_200k.cache_read}`);
}
if (model.cost.context_over_200k.cache_write !== undefined) {
lines.push(`cache_write = ${model.cost.context_over_200k.cache_write}`);
}
}
}
// Limit section
@@ -416,6 +440,10 @@ function detectChanges(
compare("cost.output", existing.cost?.output, merged.cost?.output);
compare("cost.cache_read", existing.cost?.cache_read, merged.cost?.cache_read);
compare("cost.cache_write", existing.cost?.cache_write, merged.cost?.cache_write);
compare("cost.context_over_200k.input", existing.cost?.context_over_200k?.input, merged.cost?.context_over_200k?.input);
compare("cost.context_over_200k.output", existing.cost?.context_over_200k?.output, merged.cost?.context_over_200k?.output);
compare("cost.context_over_200k.cache_read", existing.cost?.context_over_200k?.cache_read, merged.cost?.context_over_200k?.cache_read);
compare("cost.context_over_200k.cache_write", existing.cost?.context_over_200k?.cache_write, merged.cost?.context_over_200k?.cache_write);
compare("limit.context", existing.limit?.context, merged.limit.context);
compare("limit.output", existing.limit?.output, merged.limit.output);
compare("modalities.input", existing.modalities?.input, merged.modalities.input);
@@ -437,7 +465,7 @@ async function main() {
"models",
);
// Check for API key from CLI argument, environment, or prompt
// Check for API key from CLI argument or environment variable
let apiKey: string | null = null;
// Check CLI args for --api-key=xxx or --api-key xxx
@@ -456,11 +484,6 @@ async function main() {
apiKey = process.env.VENICE_API_KEY ?? null;
}
// Prompt if still no key
if (!apiKey) {
apiKey = await promptForApiKey();
}
const includeAlpha = apiKey !== null;
if (dryRun) {
+584
View File
@@ -0,0 +1,584 @@
#!/usr/bin/env bun
/**
* Generates Vercel model TOML files from the AI Gateway API.
*
* Flags:
* --dry-run: Preview changes without writing files
* --new-only: Only create new models, skip updating existing ones
*/
import { z } from "zod";
import path from "node:path";
import { mkdir } from "node:fs/promises";
import { ModelFamilyValues } from "../src/family.js";
const API_ENDPOINT = "https://ai-gateway.vercel.sh/v1/models";
enum ModelType {
Language = "language",
Embedding = "embedding",
Image = "image",
Video = "video",
}
enum SkipZeroFields {
LimitContext = "limit.context",
LimitInput = "limit.input",
LimitOutput = "limit.output",
}
const PricingTier = z.object({
cost: z.string(),
min: z.number(),
max: z.number().optional(),
});
const Pricing = z.object({
input: z.string().optional(),
output: z.string().optional(),
input_cache_read: z.string().optional(),
input_cache_write: z.string().optional(),
input_tiers: z.array(PricingTier).optional(),
output_tiers: z.array(PricingTier).optional(),
input_cache_read_tiers: z.array(PricingTier).optional(),
input_cache_write_tiers: z.array(PricingTier).optional(),
}).passthrough();
const VercelModel = z.object({
id: z.string(),
name: z.string(),
created: z.number(),
released: z.number().optional(),
context_window: z.number(),
max_tokens: z.number(),
type: z.nativeEnum(ModelType),
tags: z.array(z.string()).optional().default([]),
pricing: Pricing.optional(),
}).passthrough();
const VercelResponse = z.object({
data: z.array(VercelModel),
}).passthrough();
interface ExistingModel {
name?: string;
family?: string;
attachment?: boolean;
reasoning?: boolean;
tool_call?: boolean;
structured_output?: boolean;
temperature?: boolean;
knowledge?: string;
release_date?: string;
last_updated?: string;
open_weights?: boolean;
interleaved?: boolean | { field: string };
status?: string;
cost?: {
input?: number;
output?: number;
cache_read?: number;
cache_write?: number;
};
limit?: {
context?: number;
input?: number;
output?: number;
};
modalities?: {
input?: string[];
output?: string[];
};
}
interface MergedModel {
name: string;
family?: string;
attachment: boolean;
reasoning: boolean;
tool_call: boolean;
structured_output?: boolean;
temperature: boolean;
knowledge?: string;
release_date: string;
last_updated: string;
open_weights: boolean;
interleaved?: boolean | { field: string };
status?: string;
cost?: {
input: number;
output: number;
cache_read?: number;
cache_write?: number;
};
limit: {
context: number;
input?: number;
output: number;
};
modalities: {
input: string[];
output: string[];
};
}
interface Changes {
field: string;
oldValue: string;
newValue: string;
}
function timestampToDate(timestamp: number): string {
const date = new Date(timestamp * 1000);
return date.toISOString().slice(0, 10);
}
function getTodayDate(): string {
return new Date().toISOString().slice(0, 10);
}
// Number utilities
function formatNumber(n: number): string {
if (n >= 1000) {
return n.toString().replace(/\B(?=(\d{3})+(?!\d))/g, "_");
}
return n.toString();
}
function isSubstring(target: string, family: string): boolean {
return target.toLowerCase().includes(family.toLowerCase());
}
function matchesFamily(target: string, family: string): boolean {
const targetLower = target.toLowerCase();
const familyLower = family.toLowerCase();
let familyIdx = 0;
for (let i = 0; i < targetLower.length && familyIdx < familyLower.length; i++) {
if (targetLower[i] === familyLower[familyIdx]) {
familyIdx++;
}
}
return familyIdx === familyLower.length;
}
function inferFamily(modelId: string, modelName: string): string | undefined {
const sortedFamilies = [...ModelFamilyValues].sort((a, b) => b.length - a.length);
// First pass: try exact substring matches
for (const family of sortedFamilies) {
if (isSubstring(modelId, family)) {
return family;
}
}
for (const family of sortedFamilies) {
if (isSubstring(modelName, family)) {
return family;
}
}
// Second pass: fall back to subsequence matching
for (const family of sortedFamilies) {
if (matchesFamily(modelId, family)) {
return family;
}
}
for (const family of sortedFamilies) {
if (matchesFamily(modelName, family)) {
return family;
}
}
return undefined;
}
function buildInputModalities(tags: string[]): string[] {
const mods: string[] = ["text"];
const tagSet = new Set(tags);
if (tagSet.has("vision")) mods.push("image");
if (tagSet.has("file-input")) mods.push("pdf");
return mods;
}
function buildOutputModalities(modelType: ModelType, tags: string[]): string[] {
const mods: string[] = ["text"];
const tagSet = new Set(tags);
if (modelType === ModelType.Image || tagSet.has("image-generation")) {
mods.push("image");
} else if (modelType === ModelType.Video) {
mods.push("video");
}
return mods;
}
async function loadExistingModel(filePath: string): Promise<ExistingModel | null> {
try {
const file = Bun.file(filePath);
if (!(await file.exists())) {
return null;
}
const toml = await import(filePath, { with: { type: "toml" } }).then(
(mod) => mod.default,
);
return toml as ExistingModel;
} catch (e) {
console.warn(`Warning: Failed to parse existing file ${filePath}:`, e);
return null;
}
}
function isOpenAIModel(modelId: string): boolean {
return modelId.startsWith("openai/");
}
function mergeModel(
apiModel: z.infer<typeof VercelModel>,
existing: ExistingModel | null,
): MergedModel {
const tagSet = new Set(apiModel.tags);
const inputModalities = buildInputModalities(apiModel.tags);
const outputModalities = buildOutputModalities(apiModel.type, apiModel.tags);
// Preserve existing values when available (previously manually specified)
const name = existing?.name ?? apiModel.name;
const attachment = existing?.attachment ?? (tagSet.has("vision") || tagSet.has("file-input"));
const reasoning = existing?.reasoning ?? tagSet.has("reasoning");
const toolCall = existing?.tool_call ?? tagSet.has("tool-use");
const openWeights = existing?.open_weights ?? false;
const family = existing?.family ?? inferFamily(apiModel.id, apiModel.name);
const structuredOutput = existing?.structured_output;
const knowledge = existing?.knowledge;
const interleaved = existing?.interleaved;
const status = existing?.status;
// Release date: use API, fallback to existing, then today
const releaseDate = apiModel.released
? timestampToDate(apiModel.released)
: (existing?.release_date ?? getTodayDate());
// Preserve existing limits if API returns 0 (indicates missing/invalid data)
const contextLimit = apiModel.context_window > 0
? apiModel.context_window
: (existing?.limit?.context ?? 0);
const outputLimit = apiModel.max_tokens > 0
? apiModel.max_tokens
: (existing?.limit?.output ?? 0);
const merged: MergedModel = {
name,
family,
attachment,
reasoning,
tool_call: toolCall,
temperature: true,
release_date: releaseDate,
last_updated: getTodayDate(),
open_weights: openWeights,
...(structuredOutput !== undefined && { structured_output: structuredOutput }),
...(knowledge && { knowledge }),
...(interleaved !== undefined && { interleaved }),
...(status && { status }),
limit: {
context: contextLimit,
...(isOpenAIModel(apiModel.id) && contextLimit > outputLimit && { input: contextLimit - outputLimit }),
output: outputLimit,
},
modalities: {
input: inputModalities,
output: outputModalities,
},
};
if (apiModel.pricing) {
const inputPrice = apiModel.pricing.input_tiers?.[0]?.cost ?? apiModel.pricing.input;
const outputPrice = apiModel.pricing.output_tiers?.[0]?.cost ?? apiModel.pricing.output;
const cacheReadPrice = apiModel.pricing.input_cache_read_tiers?.[0]?.cost ?? apiModel.pricing.input_cache_read;
const cacheWritePrice = apiModel.pricing.input_cache_write_tiers?.[0]?.cost ?? apiModel.pricing.input_cache_write;
if (inputPrice && outputPrice) {
merged.cost = {
input: parseFloat(inputPrice) * 1_000_000,
output: parseFloat(outputPrice) * 1_000_000,
...(cacheReadPrice && {
cache_read: parseFloat(cacheReadPrice) * 1_000_000,
}),
...(cacheWritePrice && {
cache_write: parseFloat(cacheWritePrice) * 1_000_000,
}),
};
}
}
return merged;
}
function formatToml(model: MergedModel): string {
const lines: string[] = [];
lines.push(`name = "${model.name.replace(/"/g, '\\"')}"`);
if (model.family) {
lines.push(`family = "${model.family}"`);
}
lines.push(`attachment = ${model.attachment}`);
lines.push(`reasoning = ${model.reasoning}`);
lines.push(`tool_call = ${model.tool_call}`);
if (model.structured_output !== undefined) {
lines.push(`structured_output = ${model.structured_output}`);
}
lines.push(`temperature = ${model.temperature}`);
if (model.knowledge) {
lines.push(`knowledge = "${model.knowledge}"`);
}
lines.push(`release_date = "${model.release_date}"`);
lines.push(`last_updated = "${model.last_updated}"`);
lines.push(`open_weights = ${model.open_weights}`);
if (model.status) {
lines.push(`status = "${model.status}"`);
}
if (model.interleaved !== undefined) {
lines.push("");
if (model.interleaved === true) {
lines.push(`interleaved = true`);
} else if (typeof model.interleaved === "object") {
lines.push(`[interleaved]`);
lines.push(`field = "${model.interleaved.field}"`);
}
}
if (model.cost) {
lines.push("");
lines.push(`[cost]`);
lines.push(`input = ${model.cost.input}`);
lines.push(`output = ${model.cost.output}`);
if (model.cost.cache_read !== undefined) {
lines.push(`cache_read = ${model.cost.cache_read}`);
}
if (model.cost.cache_write !== undefined) {
lines.push(`cache_write = ${model.cost.cache_write}`);
}
}
lines.push("");
lines.push(`[limit]`);
lines.push(`context = ${formatNumber(model.limit.context)}`);
if (model.limit.input !== undefined) {
lines.push(`input = ${formatNumber(model.limit.input)}`);
}
lines.push(`output = ${formatNumber(model.limit.output)}`);
lines.push("");
lines.push(`[modalities]`);
lines.push(`input = [${model.modalities.input.map((m) => `"${m}"`).join(", ")}]`);
lines.push(`output = [${model.modalities.output.map((m) => `"${m}"`).join(", ")}]`);
return lines.join("\n") + "\n";
}
function detectChanges(
existing: ExistingModel | null,
merged: MergedModel,
): Changes[] {
if (!existing) return [];
const changes: Changes[] = [];
const EPSILON = 0.001; // price diff to ignore (per million tokens)
const shouldSkipZero = (field: string, oldVal: unknown, newVal: unknown): boolean => {
if (!Object.values(SkipZeroFields).includes(field as SkipZeroFields)) {
return false;
}
return (typeof oldVal === "number" && oldVal === 0) || (typeof newVal === "number" && newVal === 0);
};
const formatValue = (val: unknown): string => {
if (typeof val === "number") return formatNumber(val);
if (Array.isArray(val)) return `[${val.join(", ")}]`;
if (val === undefined) return "(none)";
return String(val);
};
const isMaterialPriceDiff = (oldPrice: unknown, newPrice: unknown): boolean => {
// 0 → undefined is not material (cost removed)
if (oldPrice === 0 && newPrice === undefined) return false;
if (oldPrice !== undefined && newPrice !== undefined) {
return Math.abs((oldPrice as number) - (newPrice as number)) > EPSILON;
}
return oldPrice !== newPrice;
};
const compare = (field: string, oldVal: unknown, newVal: unknown) => {
if (shouldSkipZero(field, oldVal, newVal)) return;
const isDiff = field.startsWith("cost.")
? isMaterialPriceDiff(oldVal, newVal)
: JSON.stringify(oldVal) !== JSON.stringify(newVal);
if (isDiff) {
changes.push({
field,
oldValue: formatValue(oldVal),
newValue: formatValue(newVal),
});
}
};
compare("name", existing.name, merged.name);
compare("family", existing.family, merged.family);
compare("attachment", existing.attachment, merged.attachment);
compare("reasoning", existing.reasoning, merged.reasoning);
compare("tool_call", existing.tool_call, merged.tool_call);
compare("structured_output", existing.structured_output, merged.structured_output);
compare("open_weights", existing.open_weights, merged.open_weights);
compare("release_date", existing.release_date, merged.release_date);
compare("cost.input", existing.cost?.input, merged.cost?.input);
compare("cost.output", existing.cost?.output, merged.cost?.output);
compare("cost.cache_read", existing.cost?.cache_read, merged.cost?.cache_read);
compare("cost.cache_write", existing.cost?.cache_write, merged.cost?.cache_write);
compare("limit.context", existing.limit?.context, merged.limit.context);
compare("limit.input", existing.limit?.input, merged.limit.input);
compare("limit.output", existing.limit?.output, merged.limit.output);
compare("modalities.input", existing.modalities?.input, merged.modalities.input);
return changes;
}
async function main() {
const args = process.argv.slice(2);
const dryRun = args.includes("--dry-run");
const newOnly = args.includes("--new-only");
const modelsDir = path.join(
import.meta.dirname,
"..",
"..",
"..",
"providers",
"vercel",
"models",
);
console.log(`${dryRun ? "[DRY RUN] " : ""}${newOnly ? "[NEW ONLY] " : ""}Fetching Vercel models from API...`);
const res = await fetch(API_ENDPOINT);
if (!res.ok) {
console.error(`Failed to fetch API: ${res.status} ${res.statusText}`);
process.exit(1);
}
const json = await res.json();
const parsed = VercelResponse.safeParse(json);
if (!parsed.success) {
console.error("Invalid API response:", parsed.error.errors);
process.exit(1);
}
const apiModels = parsed.data.data;
const existingFiles = new Set<string>();
try {
for await (const file of new Bun.Glob("**/*.toml").scan({
cwd: modelsDir,
absolute: false,
})) {
existingFiles.add(file);
}
} catch {
}
console.log(`Found ${apiModels.length} models in API, ${existingFiles.size} existing files\n`);
const apiModelIds = new Set<string>();
let created = 0;
let updated = 0;
let unchanged = 0;
for (const apiModel of apiModels) {
// Skip these since OpenCode does not support image / video generation yet
if (apiModel.type === ModelType.Image || apiModel.type === ModelType.Video) {
continue;
}
const relativePath = `${apiModel.id}.toml`;
const filePath = path.join(modelsDir, relativePath);
const dirPath = path.dirname(filePath);
apiModelIds.add(relativePath);
const existing = await loadExistingModel(filePath);
const merged = mergeModel(apiModel, existing);
const tomlContent = formatToml(merged);
if (existing === null) {
created++;
if (dryRun) {
console.log(`[DRY RUN] Would create: ${relativePath}`);
console.log(` name = "${merged.name}"`);
if (merged.family) {
console.log(` family = "${merged.family}" (inferred)`);
}
console.log("");
} else {
await mkdir(dirPath, { recursive: true });
await Bun.write(filePath, tomlContent);
console.log(`Created: ${relativePath}`);
}
} else {
if (newOnly) {
unchanged++;
continue;
}
const changes = detectChanges(existing, merged);
if (changes.length > 0) {
updated++;
if (dryRun) {
console.log(`[DRY RUN] Would update: ${relativePath}`);
} else {
await mkdir(dirPath, { recursive: true });
await Bun.write(filePath, tomlContent);
console.log(`Updated: ${relativePath}`);
}
for (const change of changes) {
console.log(` ${change.field}: ${change.oldValue}${change.newValue}`);
}
console.log("");
} else {
unchanged++;
}
}
}
const orphaned: string[] = [];
for (const file of existingFiles) {
if (!apiModelIds.has(file)) {
orphaned.push(file);
console.log(`Warning: Orphaned file (not in API): ${file}`);
}
}
console.log("");
if (dryRun) {
console.log(
`Summary: ${created} would be created, ${updated} would be updated, ${unchanged} unchanged, ${orphaned.length} orphaned`,
);
} else {
console.log(
`Summary: ${created} created, ${updated} updated, ${unchanged} unchanged, ${orphaned.length} orphaned`,
);
}
}
await main();
+525
View File
@@ -0,0 +1,525 @@
#!/usr/bin/env bun
import path from "node:path";
import { mkdir } from "node:fs/promises";
import { z } from "zod";
import { ModelFamilyValues } from "../src/family.js";
const API_ENDPOINT = "https://trace.wandb.ai/inference/analysis/artificialanalysis/models";
const Pricing = z
.object({
prompt: z.string().optional(),
completion: z.string().optional(),
image: z.string().optional(),
request: z.string().optional(),
input_cache_reads: z.string().optional(),
input_cache_writes: z.string().optional(),
})
.passthrough();
const WandbModel = z
.object({
id: z.string(),
name: z.string(),
created: z.number(),
input_modalities: z.array(z.string()),
output_modalities: z.array(z.string()),
context_length: z.number(),
max_output_length: z.number(),
pricing: Pricing.optional(),
supported_sampling_parameters: z.array(z.string()).default([]),
supported_features: z.array(z.string()).default([]),
})
.passthrough();
const WandbResponse = z
.object({
data: z.array(WandbModel),
})
.strict();
interface ExistingModel {
name?: string;
family?: string;
attachment?: boolean;
reasoning?: boolean;
tool_call?: boolean;
structured_output?: boolean;
temperature?: boolean;
knowledge?: string;
release_date?: string;
last_updated?: string;
open_weights?: boolean;
interleaved?: boolean | { field: string };
status?: string;
cost?: {
input?: number;
output?: number;
cache_read?: number;
cache_write?: number;
};
limit?: {
context?: number;
input?: number;
output?: number;
};
modalities?: {
input?: string[];
output?: string[];
};
}
interface MergedModel {
name: string;
family?: string;
attachment: boolean;
reasoning: boolean;
tool_call: boolean;
structured_output?: boolean;
temperature: boolean;
knowledge?: string;
release_date: string;
last_updated: string;
open_weights: boolean;
interleaved?: boolean | { field: string };
status?: string;
cost?: {
input: number;
output: number;
cache_read?: number;
cache_write?: number;
};
limit: {
context: number;
output: number;
};
modalities: {
input: Array<"text" | "audio" | "image" | "video" | "pdf">;
output: Array<"text" | "audio" | "image" | "video" | "pdf">;
};
}
interface Changes {
field: string;
oldValue: string;
newValue: string;
}
type SupportedModality = "text" | "audio" | "image" | "video" | "pdf";
const modalityMap: Record<string, SupportedModality | undefined> = {
text: "text",
image: "image",
audio: "audio",
video: "video",
pdf: "pdf",
file: "pdf",
files: "pdf",
};
const openWeightsPrefixes = new Set([
"deepseek-ai/",
"meta-llama/",
"microsoft/",
"MiniMaxAI/",
"moonshotai/",
"nvidia/",
"OpenPipe/",
"Qwen/",
"zai-org/",
]);
function timestampToDate(timestamp: number): string {
return new Date(timestamp * 1000).toISOString().slice(0, 10);
}
function getTodayDate(): string {
return new Date().toISOString().slice(0, 10);
}
function formatNumber(n: number): string {
if (n >= 1000) {
return n.toString().replace(/\B(?=(\d{3})+(?!\d))/g, "_");
}
return n.toString();
}
function formatDecimal(n: number): string {
return Number(n.toFixed(6)).toString();
}
function priceToPerMillion(value: string): number {
return Number((parseFloat(value) * 1_000_000).toFixed(6));
}
function isSubstring(target: string, family: string): boolean {
return target.toLowerCase().includes(family.toLowerCase());
}
function matchesFamily(target: string, family: string): boolean {
const targetLower = target.toLowerCase();
const familyLower = family.toLowerCase();
let familyIdx = 0;
for (let i = 0; i < targetLower.length && familyIdx < familyLower.length; i++) {
if (targetLower[i] === familyLower[familyIdx]) {
familyIdx++;
}
}
return familyIdx === familyLower.length;
}
function inferFamily(modelId: string, modelName: string): string | undefined {
const sortedFamilies = [...ModelFamilyValues].sort((a, b) => b.length - a.length);
for (const family of sortedFamilies) {
if (isSubstring(modelId, family) || isSubstring(modelName, family)) {
return family;
}
}
for (const family of sortedFamilies) {
if (matchesFamily(modelId, family) || matchesFamily(modelName, family)) {
return family;
}
}
return undefined;
}
function normalizeName(apiModel: z.infer<typeof WandbModel>): string {
const stripped = apiModel.name.replace(/^[^:]+:\s*/, "").trim();
return stripped || path.basename(apiModel.id);
}
function inferReasoning(apiModel: z.infer<typeof WandbModel>): boolean {
const text = `${apiModel.id} ${apiModel.name}`.toLowerCase();
return text.includes("thinking") || /\br1\b/.test(text) || text.includes("reasoning");
}
function inferOpenWeights(modelId: string): boolean {
for (const prefix of openWeightsPrefixes) {
if (modelId.startsWith(prefix)) {
return true;
}
}
return false;
}
function normalizeModalities(values: string[]): SupportedModality[] {
const normalized = values
.map((value) => modalityMap[value.toLowerCase()])
.filter((value): value is SupportedModality => value !== undefined);
return [...new Set(normalized)];
}
async function loadExistingModel(filePath: string): Promise<ExistingModel | null> {
try {
const file = Bun.file(filePath);
if (!(await file.exists())) {
return null;
}
const toml = await import(filePath, { with: { type: "toml" } }).then((mod) => mod.default);
return toml as ExistingModel;
} catch (cause) {
console.warn(`Warning: Failed to parse existing file ${filePath}:`, cause);
return null;
}
}
function mergeModel(
apiModel: z.infer<typeof WandbModel>,
existing: ExistingModel | null,
): MergedModel {
const featureSet = new Set(apiModel.supported_features);
const samplingSet = new Set(apiModel.supported_sampling_parameters);
const inputModalities = normalizeModalities(apiModel.input_modalities);
const outputModalities = normalizeModalities(apiModel.output_modalities);
const merged: MergedModel = {
name: existing?.name ?? normalizeName(apiModel),
family: existing?.family ?? inferFamily(apiModel.id, apiModel.name),
attachment: existing?.attachment ?? inputModalities.some((m) => m !== "text"),
reasoning: existing?.reasoning ?? inferReasoning(apiModel),
tool_call: existing?.tool_call ?? featureSet.has("tools"),
temperature: existing?.temperature ?? samplingSet.has("temperature"),
release_date: existing?.release_date ?? timestampToDate(apiModel.created),
last_updated: getTodayDate(),
open_weights: existing?.open_weights ?? inferOpenWeights(apiModel.id),
...(existing?.structured_output !== undefined
? { structured_output: existing.structured_output }
: featureSet.has("structured_outputs")
? { structured_output: true }
: {}),
...(existing?.knowledge ? { knowledge: existing.knowledge } : {}),
...(existing?.interleaved !== undefined ? { interleaved: existing.interleaved } : {}),
...(existing?.status ? { status: existing.status } : {}),
limit: {
context: apiModel.context_length > 0 ? apiModel.context_length : (existing?.limit?.context ?? 0),
output: apiModel.max_output_length > 0
? apiModel.max_output_length
: (existing?.limit?.output ?? 0),
},
modalities: {
input: inputModalities.length > 0
? inputModalities
: ((existing?.modalities?.input as SupportedModality[] | undefined) ?? ["text"]),
output: outputModalities.length > 0
? outputModalities
: ((existing?.modalities?.output as SupportedModality[] | undefined) ?? ["text"]),
},
};
const prompt = apiModel.pricing?.prompt;
const completion = apiModel.pricing?.completion;
const cacheRead = apiModel.pricing?.input_cache_reads;
const cacheWrite = apiModel.pricing?.input_cache_writes;
if (prompt && completion) {
merged.cost = {
input: priceToPerMillion(prompt),
output: priceToPerMillion(completion),
...(cacheRead && parseFloat(cacheRead) > 0
? { cache_read: priceToPerMillion(cacheRead) }
: {}),
...(cacheWrite && parseFloat(cacheWrite) > 0
? { cache_write: priceToPerMillion(cacheWrite) }
: {}),
};
} else if (existing?.cost?.input !== undefined && existing.cost.output !== undefined) {
merged.cost = {
input: existing.cost.input,
output: existing.cost.output,
...(existing.cost.cache_read !== undefined ? { cache_read: existing.cost.cache_read } : {}),
...(existing.cost.cache_write !== undefined ? { cache_write: existing.cost.cache_write } : {}),
};
}
return merged;
}
function formatToml(model: MergedModel): string {
const lines: string[] = [];
lines.push(`name = "${model.name.replace(/"/g, '\\"')}"`);
if (model.family) {
lines.push(`family = "${model.family}"`);
}
lines.push(`release_date = "${model.release_date}"`);
lines.push(`last_updated = "${model.last_updated}"`);
lines.push(`attachment = ${model.attachment}`);
lines.push(`reasoning = ${model.reasoning}`);
if (model.structured_output !== undefined) {
lines.push(`structured_output = ${model.structured_output}`);
}
lines.push(`temperature = ${model.temperature}`);
lines.push(`tool_call = ${model.tool_call}`);
if (model.knowledge) {
lines.push(`knowledge = "${model.knowledge}"`);
}
lines.push(`open_weights = ${model.open_weights}`);
if (model.status) {
lines.push(`status = "${model.status}"`);
}
if (model.interleaved !== undefined) {
lines.push("");
if (model.interleaved === true) {
lines.push("interleaved = true");
} else {
lines.push("[interleaved]");
lines.push(`field = "${model.interleaved.field}"`);
}
}
if (model.cost) {
lines.push("");
lines.push("[cost]");
lines.push(`input = ${formatDecimal(model.cost.input)}`);
lines.push(`output = ${formatDecimal(model.cost.output)}`);
if (model.cost.cache_read !== undefined) {
lines.push(`cache_read = ${formatDecimal(model.cost.cache_read)}`);
}
if (model.cost.cache_write !== undefined) {
lines.push(`cache_write = ${formatDecimal(model.cost.cache_write)}`);
}
}
lines.push("");
lines.push("[limit]");
lines.push(`context = ${formatNumber(model.limit.context)}`);
lines.push(`output = ${formatNumber(model.limit.output)}`);
lines.push("");
lines.push("[modalities]");
lines.push(`input = [${model.modalities.input.map((m) => `"${m}"`).join(", ")}]`);
lines.push(`output = [${model.modalities.output.map((m) => `"${m}"`).join(", ")}]`);
return `${lines.join("\n")}\n`;
}
function detectChanges(existing: ExistingModel | null, merged: MergedModel): Changes[] {
if (!existing) {
return [];
}
const changes: Changes[] = [];
const epsilon = 0.001;
const formatValue = (value: unknown): string => {
if (typeof value === "number") return formatNumber(value);
if (Array.isArray(value)) return `[${value.join(", ")}]`;
if (value === undefined) return "(none)";
return String(value);
};
const compare = (field: string, oldValue: unknown, newValue: unknown) => {
const changed = field.startsWith("cost.")
? (
oldValue === undefined && newValue === undefined
? false
: oldValue === undefined || newValue === undefined
? true
: Math.abs((oldValue as number) - (newValue as number)) > epsilon
)
: JSON.stringify(oldValue) !== JSON.stringify(newValue);
if (changed) {
changes.push({
field,
oldValue: formatValue(oldValue),
newValue: formatValue(newValue),
});
}
};
compare("name", existing.name, merged.name);
compare("family", existing.family, merged.family);
compare("release_date", existing.release_date, merged.release_date);
compare("attachment", existing.attachment, merged.attachment);
compare("reasoning", existing.reasoning, merged.reasoning);
compare("structured_output", existing.structured_output, merged.structured_output);
compare("temperature", existing.temperature, merged.temperature);
compare("tool_call", existing.tool_call, merged.tool_call);
compare("open_weights", existing.open_weights, merged.open_weights);
compare("cost.input", existing.cost?.input, merged.cost?.input);
compare("cost.output", existing.cost?.output, merged.cost?.output);
compare("cost.cache_read", existing.cost?.cache_read, merged.cost?.cache_read);
compare("cost.cache_write", existing.cost?.cache_write, merged.cost?.cache_write);
compare("limit.context", existing.limit?.context, merged.limit.context);
compare("limit.output", existing.limit?.output, merged.limit.output);
compare("modalities.input", existing.modalities?.input, merged.modalities.input);
compare("modalities.output", existing.modalities?.output, merged.modalities.output);
return changes;
}
async function main() {
const args = process.argv.slice(2);
const dryRun = args.includes("--dry-run");
const newOnly = args.includes("--new-only");
const modelsDir = path.join(import.meta.dirname, "..", "..", "..", "providers", "wandb", "models");
console.log(`${dryRun ? "[DRY RUN] " : ""}${newOnly ? "[NEW ONLY] " : ""}Fetching WandB models from API...`);
const res = await fetch(API_ENDPOINT);
if (!res.ok) {
console.error(`Failed to fetch API: ${res.status} ${res.statusText}`);
process.exit(1);
}
const json = await res.json();
const parsed = WandbResponse.safeParse(json);
if (!parsed.success) {
console.error("Invalid API response:", parsed.error.errors);
process.exit(1);
}
const apiModels = parsed.data.data;
const existingFiles = new Set<string>();
for await (const file of new Bun.Glob("**/*.toml").scan({ cwd: modelsDir, absolute: false })) {
existingFiles.add(file);
}
console.log(`Found ${apiModels.length} models in API, ${existingFiles.size} existing files\n`);
const apiModelIds = new Set<string>();
let created = 0;
let updated = 0;
let unchanged = 0;
for (const apiModel of apiModels) {
const relativePath = `${apiModel.id}.toml`;
const filePath = path.join(modelsDir, relativePath);
const dirPath = path.dirname(filePath);
apiModelIds.add(relativePath);
const existing = await loadExistingModel(filePath);
const merged = mergeModel(apiModel, existing);
const tomlContent = formatToml(merged);
if (existing === null) {
created++;
if (dryRun) {
console.log(`[DRY RUN] Would create: ${relativePath}`);
console.log(` name = "${merged.name}"`);
if (merged.family) {
console.log(` family = "${merged.family}"`);
}
console.log("");
} else {
await mkdir(dirPath, { recursive: true });
await Bun.write(filePath, tomlContent);
console.log(`Created: ${relativePath}`);
}
continue;
}
if (newOnly) {
unchanged++;
continue;
}
const changes = detectChanges(existing, merged);
if (changes.length === 0) {
unchanged++;
continue;
}
updated++;
if (dryRun) {
console.log(`[DRY RUN] Would update: ${relativePath}`);
} else {
await mkdir(dirPath, { recursive: true });
await Bun.write(filePath, tomlContent);
console.log(`Updated: ${relativePath}`);
}
for (const change of changes) {
console.log(` ${change.field}: ${change.oldValue}${change.newValue}`);
}
console.log("");
}
const orphaned = [...existingFiles].filter((file) => !apiModelIds.has(file));
for (const file of orphaned) {
console.log(`Warning: Orphaned file (not in API): ${file}`);
}
console.log("");
console.log(
dryRun
? `Summary: ${created} would be created, ${updated} would be updated, ${unchanged} unchanged, ${orphaned.length} orphaned`
: `Summary: ${created} created, ${updated} updated, ${unchanged} unchanged, ${orphaned.length} orphaned`,
);
}
await main();
+41 -5
View File
@@ -1,9 +1,14 @@
import { z } from "zod";
export const ModelFamilyValues = [
// Arcee
"trinity",
"trinity-mini",
// OpenAI/GPT style
"gpt",
"gpt-codex",
"gpt-codex-spark",
"gpt-codex-mini",
"gpt-pro",
"gpt-mini",
@@ -51,6 +56,7 @@ export const ModelFamilyValues = [
// Moonshot Kimi
"kimi",
"kimi-free",
"kimi-thinking",
// Mistral family
@@ -90,6 +96,7 @@ export const ModelFamilyValues = [
// NVIDIA Nemotron
"nemotron",
"nemotron-free",
// AWS Titan
"titan",
@@ -97,6 +104,9 @@ export const ModelFamilyValues = [
// MiniMax
"minimax",
"minimax-m2.5",
"minimax-m2.7",
"minimax-free",
// Hunyuan
"hunyuan",
@@ -121,9 +131,6 @@ export const ModelFamilyValues = [
"solar-mini",
"solar-pro",
// Exaone
"exaone",
// Step (StepFun)
"step",
@@ -140,6 +147,7 @@ export const ModelFamilyValues = [
"dall-e",
"flux",
"imagen",
"recraft",
"stable-diffusion",
"ideogram",
"dreamshaper",
@@ -177,6 +185,9 @@ export const ModelFamilyValues = [
// Sherlock
"sherlock",
// Pony
"pony",
// Mercury
"mercury",
@@ -185,6 +196,12 @@ export const ModelFamilyValues = [
// Mimo
"mimo",
"mimo-pro-free",
"mimo-omni-free",
"mimo-flash-free",
// Clarifai
"mm-poly",
// Longcat
"longcat",
@@ -218,6 +235,12 @@ export const ModelFamilyValues = [
// Falcon
"falcon",
// Baichuan
"baichuan",
// Skywork
"skywork",
// BART
"bart",
@@ -270,8 +293,6 @@ export const ModelFamilyValues = [
// NeMo
"nemoretriever",
// Nano Banana
"nano-banana",
@@ -338,6 +359,21 @@ export const ModelFamilyValues = [
// Neural Chat
"neural-chat",
// Pangu (Ascend Tribe)
"pangu",
// LiquidAI
"liquid",
// Sourceful
"sourceful",
// AllenAI
"allenai",
// Writer
"palmyra",
] as const;
export const ModelFamily = z.enum(ModelFamilyValues);
+1
View File
@@ -73,6 +73,7 @@ export const Model = z
.object({
npm: z.string().optional(),
api: z.string().optional(),
shape: z.enum(["responses", "completions"]).optional(),
})
.optional(),
})
+40
View File
@@ -11,6 +11,7 @@ export default {
): Promise<Response> {
const url = new URL(request.url);
const ip = request.headers.get("cf-connecting-ip") || "unknown";
const country = request.headers.get("cf-ipcountry") || "unknown";
const agent = request.headers.get("user-agent") || "unknown";
if (agent.includes("opencode") || agent.includes("bun")) {
ctx.waitUntil(
@@ -26,6 +27,7 @@ export default {
properties: {
$process_person_profile: false,
user_agent: agent,
country,
path: url.pathname,
},
}),
@@ -33,6 +35,44 @@ export default {
);
}
if (url.pathname === "/model-schema.json") {
const apiUrl = new URL(url);
apiUrl.pathname = "/_api.json";
const apiResponse = await env.ASSETS.fetch(
new Request(apiUrl.toString(), request),
);
const providers = (await apiResponse.json()) as Record<
string,
{ models: Record<string, unknown> }
>;
const modelIds: string[] = [];
for (const [providerId, provider] of Object.entries(providers)) {
for (const modelId of Object.keys(provider.models)) {
modelIds.push(`${providerId}/${modelId}`);
}
}
const schema = {
$schema: "https://json-schema.org/draft/2020-12/schema",
$id: "https://models.dev/model-schema.json",
$defs: {
Model: {
type: "string",
enum: modelIds.sort(),
description: "AI model identifier in provider/model format",
},
},
};
return new Response(JSON.stringify(schema, null, 2), {
headers: {
"Content-Type": "application/json",
"Cache-Control": "public, max-age=3600",
},
});
}
if (url.pathname === "/api.json") {
url.pathname = "/_api.json";
} else if (
+7
View File
@@ -0,0 +1,7 @@
<svg version="1.1" xmlns="http://www.w3.org/2000/svg" style="display: block;" viewBox="0 0 2048 2048" width="1046" height="1046" preserveAspectRatio="none">
<path transform="translate(0,0)" fill="rgb(156,155,155)" d="M 388.193 1682.23 C 380.878 1674.23 349.014 1650.79 338.487 1642.13 C 148.161 1485.91 28.1107 1260.14 5.01063 1015 C -18.4493 771.014 55.9168 527.694 211.767 338.509 C 367.861 148.455 593.42 28.6444 838.288 5.7197 C 1092.75 -18.7181 1345.95 63.524 1537.48 232.828 C 1567.76 259.726 1596.32 288.496 1623 318.968 C 1631.78 329.058 1640.32 339.36 1648.6 349.865 C 1654.15 356.824 1662.64 368.871 1669.4 374.06 C 1866.4 518.124 1998.49 734.204 2036.88 975.222 C 2041.98 1007.66 2045.11 1039.38 2046.62 1072.12 C 2046.85 1077.09 2047.16 1082.22 2048 1087.12 L 2048 1155.79 L 2047.85 1156.71 C 2045.71 1170.93 2045.27 1196.57 2043.86 1212.22 C 2040.62 1246.59 2035.38 1280.75 2028.17 1314.51 C 1981.27 1534.16 1856.1 1729.25 1675.99 1863.43 C 1479.61 2010.02 1232.99 2072.49 990.498 2037.07 C 801.767 2009.86 626.135 1924.74 487.842 1793.46 C 455.956 1763.37 418.158 1722.72 392.022 1687.5 C 390.729 1685.76 389.452 1684 388.193 1682.23 z"/>
<path transform="translate(0,0)" fill="rgb(117,116,116)" d="M 1669.4 374.06 C 1866.4 518.124 1998.49 734.204 2036.88 975.222 C 2041.98 1007.66 2045.11 1039.38 2046.62 1072.12 C 2046.85 1077.09 2047.16 1082.22 2048 1087.12 L 2048 1155.79 L 2047.85 1156.71 C 2045.71 1170.93 2045.27 1196.57 2043.86 1212.22 C 2040.62 1246.59 2035.38 1280.75 2028.17 1314.51 C 1981.27 1534.16 1856.1 1729.25 1675.99 1863.43 C 1479.61 2010.02 1232.99 2072.49 990.498 2037.07 C 801.767 2009.86 626.135 1924.74 487.842 1793.46 C 455.956 1763.37 418.158 1722.72 392.022 1687.5 C 390.729 1685.76 389.452 1684 388.193 1682.23 C 394.373 1684.07 421.092 1702.62 428.047 1707.15 C 439.611 1714.6 451.324 1721.83 463.177 1728.82 C 490.564 1744.99 528.003 1763.59 557.361 1776.32 C 782.899 1873.92 1037.96 1878.01 1266.51 1787.69 C 1495.36 1696.62 1678.57 1518.24 1775.72 1291.9 C 1877.79 1053.06 1875.1 782.375 1768.31 545.611 C 1753.03 511.993 1733.8 474.565 1714.15 443.177 C 1706.81 431.355 1699.2 419.704 1691.32 408.231 C 1685.67 400.055 1672.56 382.889 1669.4 374.06 z"/>
<path transform="translate(0,0)" fill="rgb(254,254,254)" d="M 907.581 300.401 C 923.66 299.084 947.483 300.532 962.897 302.556 C 1041.39 313.102 1112.41 354.577 1160.17 417.755 C 1202.42 473.452 1228.57 555 1218.78 624.854 C 1256.03 619.819 1285.79 618.643 1323.38 625.442 C 1401.3 639.582 1470.3 684.357 1514.95 749.754 C 1560.04 815.479 1576.85 896.56 1561.6 974.792 C 1546.59 1052.29 1501.28 1120.59 1435.71 1164.55 C 1366.19 1211.67 1287.89 1223.85 1206.7 1207.99 L 1208.16 1225.92 C 1213.72 1304.44 1187.72 1381.94 1135.93 1441.23 C 1080.14 1505.71 1008.69 1536.8 924.661 1542.84 C 910.37 1543.02 898.883 1543.12 884.551 1541.84 C 806.009 1534.5 733.606 1496.24 683.294 1435.48 C 630.495 1371.76 609.152 1293.49 617.004 1211.82 C 577.182 1218.28 543.704 1219.15 503.598 1211.31 C 425.862 1195.87 357.514 1150.02 313.752 1083.94 C 269.997 1017.95 254.398 937.224 270.419 859.682 C 286.452 782.387 332.522 714.62 398.502 671.28 C 468.674 625.191 546.96 613.555 628.149 630.32 C 625.459 583.013 627.036 545.389 643.173 499.662 C 684.204 383.394 785.737 308.984 907.581 300.401 z"/>
<path transform="translate(0,0)" fill="rgb(156,155,155)" d="M 907.659 407.406 C 928.022 404.422 959.425 408.947 978.901 414.989 C 1028.11 429.972 1069.15 464.261 1092.65 510.025 C 1116.27 555.792 1120.24 609.2 1103.65 657.957 C 1098.74 672.384 1090.33 686.336 1086.3 699.186 C 1082.14 712.53 1083.57 726.994 1090.26 739.265 C 1100.24 757.657 1118.84 768.259 1139.57 767.119 C 1155.78 766.227 1164.63 759.273 1178.02 751.678 C 1221.56 726.903 1273.23 720.668 1321.41 734.376 C 1370.6 748.209 1412.18 781.215 1436.82 825.985 C 1461.35 870.569 1467.02 923.112 1452.58 971.905 C 1438.06 1020.82 1404.54 1061.88 1359.51 1085.9 C 1280.31 1128.12 1181.39 1108.75 1124.87 1039.74 C 1113.7 1026.12 1105.83 1013.42 1087.4 1008.03 C 1041 993.798 1000.36 1044.75 1026.36 1086.49 C 1041.95 1111.53 1062.76 1127.17 1078.18 1153.03 C 1090.26 1175.57 1097.84 1197.66 1100.84 1223.2 C 1107.04 1273.73 1092.65 1324.64 1060.9 1364.44 C 1029.42 1404.28 983.238 1429.79 932.752 1435.23 C 882.189 1440.86 831.491 1425.87 792.121 1393.65 C 752.588 1361.54 727.605 1314.91 722.781 1264.21 C 718.894 1225.82 727.653 1167.83 758.478 1141.35 C 771.141 1130.47 781.921 1118.81 792.404 1105.84 C 869.373 1010.12 879.456 876.886 817.771 770.669 C 809.316 756.055 799.574 742.224 788.661 729.342 C 782.235 721.815 773.191 712.791 767.461 705.187 C 749.008 680.699 737.483 646.859 734.444 616.659 C 729.188 565.849 744.584 515.06 777.168 475.721 C 810.289 435.451 856.055 412.394 907.659 407.406 z"/>
<path transform="translate(0,0)" fill="rgb(156,155,155)" d="M 554.228 729.711 C 659.104 725.931 747.227 807.802 751.164 912.672 C 755.1 1017.54 673.362 1105.79 568.498 1109.88 C 463.411 1113.98 374.937 1032.03 370.992 926.942 C 367.047 821.849 449.129 733.498 554.228 729.711 z"/>
</svg>

After

Width:  |  Height:  |  Size: 4.9 KiB

+21
View File
@@ -0,0 +1,21 @@
name = "MiniMax-M1"
family = "minimax"
release_date = "2025-06-16"
last_updated = "2025-06-16"
attachment = false
reasoning = false
temperature = true
tool_call = true
open_weights = false
[cost]
input = 0.132
output = 1.254
[limit]
context = 1_000_000
output = 128_000
[modalities]
input = ["text"]
output = ["text"]
+20
View File
@@ -0,0 +1,20 @@
name = "MiniMax-M2.1"
release_date = "2025-12-19"
last_updated = "2025-12-19"
attachment = false
reasoning = false
temperature = true
tool_call = true
open_weights = false
[cost]
input = 0.300
output = 1.200
[limit]
context = 1_000_000
output = 131_072
[modalities]
input = ["text"]
output = ["text"]
+20
View File
@@ -0,0 +1,20 @@
name = "MiniMax-M2"
release_date = "2025-10-26"
last_updated = "2025-10-26"
attachment = false
reasoning = false
temperature = true
tool_call = true
open_weights = false
[cost]
input = 0.330
output = 1.320
[limit]
context = 1_000_000
output = 128_000
[modalities]
input = ["text"]
output = ["text"]
@@ -0,0 +1,22 @@
name = "chatgpt-4o-latest"
family = "gpt"
release_date = "2024-08-08"
last_updated = "2024-08-08"
attachment = true
reasoning = false
temperature = true
tool_call = false
open_weights = false
knowledge = "2023-09"
[cost]
input = 5.000
output = 15.000
[limit]
context = 128_000
output = 16_384
[modalities]
input = ["text", "image"]
output = ["text"]
@@ -0,0 +1,21 @@
name = "claude-haiku-4-5-20251001"
release_date = "2025-10-16"
last_updated = "2025-10-16"
attachment = true
reasoning = false
temperature = true
tool_call = true
open_weights = false
knowledge = "2025-03"
[cost]
input = 1.000
output = 5.000
[limit]
context = 200_000
output = 64_000
[modalities]
input = ["text", "image"]
output = ["text"]
@@ -0,0 +1,21 @@
name = "claude-opus-4-1-20250805-thinking"
release_date = "2025-05-27"
last_updated = "2025-05-27"
attachment = true
reasoning = true
temperature = true
tool_call = true
open_weights = false
knowledge = "2025-03"
[cost]
input = 15.000
output = 75.000
[limit]
context = 200_000
output = 32_000
[modalities]
input = ["text", "image"]
output = ["text"]
@@ -0,0 +1,21 @@
name = "claude-opus-4-1-20250805"
release_date = "2025-08-05"
last_updated = "2025-08-05"
attachment = true
reasoning = false
temperature = true
tool_call = true
open_weights = false
knowledge = "2025-03"
[cost]
input = 15.000
output = 75.000
[limit]
context = 200_000
output = 32_000
[modalities]
input = ["text", "image"]
output = ["text"]
@@ -0,0 +1,21 @@
name = "claude-opus-4-5-20251101-thinking"
release_date = "2025-11-25"
last_updated = "2025-11-25"
attachment = true
reasoning = true
temperature = true
tool_call = true
open_weights = false
knowledge = "2025-03"
[cost]
input = 5.000
output = 25.000
[limit]
context = 200_000
output = 64_000
[modalities]
input = ["text", "image"]
output = ["text"]
@@ -1,22 +1,21 @@
name = "stepfun-ai/step3"
family = "step"
release_date = "2025-08-06"
name = "claude-opus-4-5-20251101"
release_date = "2025-11-25"
last_updated = "2025-11-25"
attachment = true
reasoning = false
temperature = true
tool_call = true
structured_output = true
open_weights = false
knowledge = "2025-03"
[cost]
input = 0.57
output = 1.42
input = 5.000
output = 25.000
[limit]
context = 66_000
output = 66_000
context = 200_000
output = 64_000
[modalities]
input = ["text", "image"]
output = ["text"]
output = ["text"]
@@ -0,0 +1,21 @@
name = "claude-sonnet-4-5-20250929-thinking"
release_date = "2025-09-30"
last_updated = "2025-09-30"
attachment = true
reasoning = true
temperature = true
tool_call = true
open_weights = false
knowledge = "2025-03"
[cost]
input = 3.000
output = 15.000
[limit]
context = 200_000
output = 64_000
[modalities]
input = ["text", "image"]
output = ["text"]
@@ -0,0 +1,21 @@
name = "claude-sonnet-4-5-20250929"
release_date = "2025-09-29"
last_updated = "2025-09-29"
attachment = true
reasoning = false
temperature = true
tool_call = true
open_weights = false
knowledge = "2025-03"
[cost]
input = 3.000
output = 15.000
[limit]
context = 200_000
output = 64_000
[modalities]
input = ["text", "image"]
output = ["text"]
+22
View File
@@ -0,0 +1,22 @@
name = "Deepseek-Chat"
family = "deepseek"
release_date = "2024-11-29"
last_updated = "2024-11-29"
attachment = false
reasoning = false
temperature = true
tool_call = true
open_weights = false
knowledge = "2024-07"
[cost]
input = 0.290
output = 0.430
[limit]
context = 128_000
output = 8_192
[modalities]
input = ["text"]
output = ["text"]
@@ -1,24 +1,21 @@
name = "DeepSeek Reasoner"
name = "Deepseek-Reasoner"
family = "deepseek-thinking"
release_date = "2025-01-20"
last_updated = "2025-09-29"
attachment = true
last_updated = "2025-01-20"
attachment = false
reasoning = true
temperature = true
knowledge = "2024-07"
tool_call = true
open_weights = false
[interleaved]
field = "reasoning_content"
knowledge = "2024-07"
[cost]
input = 0
output = 0
input = 0.290
output = 0.430
[limit]
context = 128_000
output = 65_536
output = 128_000
[modalities]
input = ["text"]
@@ -0,0 +1,21 @@
name = "DeepSeek-V3.2-Thinking"
release_date = "2025-12-01"
last_updated = "2025-12-01"
attachment = false
reasoning = true
temperature = true
tool_call = true
open_weights = false
knowledge = "2024-12"
[cost]
input = 0.290
output = 0.430
[limit]
context = 128_000
output = 128_000
[modalities]
input = ["text"]
output = ["text"]
+21
View File
@@ -0,0 +1,21 @@
name = "deepseek-v3.2"
release_date = "2025-12-01"
last_updated = "2025-12-01"
attachment = false
reasoning = false
temperature = true
tool_call = true
open_weights = false
knowledge = "2024-12"
[cost]
input = 0.290
output = 0.430
[limit]
context = 128_000
output = 8_192
[modalities]
input = ["text"]
output = ["text"]
@@ -0,0 +1,20 @@
name = "doubao-seed-1-6-thinking-250715"
release_date = "2025-07-15"
last_updated = "2025-07-15"
attachment = true
reasoning = true
temperature = true
tool_call = true
open_weights = false
[cost]
input = 0.121
output = 1.210
[limit]
context = 256_000
output = 16_000
[modalities]
input = ["text", "image"]
output = ["text"]
@@ -0,0 +1,20 @@
name = "doubao-seed-1-6-vision-250815"
release_date = "2025-09-30"
last_updated = "2025-09-30"
attachment = true
reasoning = false
temperature = true
tool_call = true
open_weights = false
[cost]
input = 0.114
output = 1.143
[limit]
context = 256_000
output = 32_000
[modalities]
input = ["text", "image"]
output = ["text"]
@@ -0,0 +1,20 @@
name = "doubao-seed-1-8-251215"
release_date = "2025-12-18"
last_updated = "2025-12-18"
attachment = true
reasoning = false
temperature = true
tool_call = true
open_weights = false
[cost]
input = 0.114
output = 0.286
[limit]
context = 224_000
output = 64_000
[modalities]
input = ["text", "image"]
output = ["text"]
@@ -0,0 +1,20 @@
name = "doubao-seed-code-preview-251028"
release_date = "2025-11-11"
last_updated = "2025-11-11"
attachment = true
reasoning = false
temperature = true
tool_call = true
open_weights = false
[cost]
input = 0.170
output = 1.140
[limit]
context = 256_000
output = 32_000
[modalities]
input = ["text", "image"]
output = ["text"]
@@ -0,0 +1,22 @@
name = "gemini-2.0-flash-lite"
family = "gemini-flash-lite"
release_date = "2025-06-16"
last_updated = "2025-06-16"
attachment = true
reasoning = false
temperature = true
tool_call = false
open_weights = false
knowledge = "2024-11"
[cost]
input = 0.075
output = 0.300
[limit]
context = 2_000_000
output = 8_192
[modalities]
input = ["text", "image"]
output = ["text"]
@@ -0,0 +1,21 @@
name = "gemini-2.5-flash-image"
release_date = "2025-10-08"
last_updated = "2025-10-08"
attachment = true
reasoning = false
temperature = true
tool_call = false
open_weights = false
knowledge = "2025-01"
[cost]
input = 0.300
output = 30.000
[limit]
context = 32_768
output = 32_768
[modalities]
input = ["text", "image"]
output = ["text"]
@@ -0,0 +1,21 @@
name = "gemini-2.5-flash-lite-preview-09-2025"
release_date = "2025-09-26"
last_updated = "2025-09-26"
attachment = true
reasoning = false
temperature = true
tool_call = true
open_weights = false
knowledge = "2025-01"
[cost]
input = 0.100
output = 0.400
[limit]
context = 1_000_000
output = 65_536
[modalities]
input = ["text", "image"]
output = ["text"]
@@ -0,0 +1,22 @@
name = "gemini-2.5-flash-nothink"
family = "gemini-flash"
release_date = "2025-06-24"
last_updated = "2025-06-24"
attachment = true
reasoning = false
temperature = true
tool_call = true
open_weights = false
knowledge = "2025-01"
[cost]
input = 0.300
output = 2.500
[limit]
context = 1_000_000
output = 65_536
[modalities]
input = ["text", "image"]
output = ["text"]
@@ -0,0 +1,21 @@
name = "gemini-2.5-flash-preview-09-2025"
release_date = "2025-09-26"
last_updated = "2025-09-26"
attachment = true
reasoning = false
temperature = true
tool_call = true
open_weights = false
knowledge = "2025-01"
[cost]
input = 0.300
output = 2.500
[limit]
context = 1_000_000
output = 65_536
[modalities]
input = ["text", "image"]
output = ["text"]
@@ -0,0 +1,22 @@
name = "gemini-2.5-flash"
family = "gemini-flash"
release_date = "2025-06-17"
last_updated = "2025-06-17"
attachment = true
reasoning = false
temperature = true
tool_call = true
open_weights = false
knowledge = "2025-01"
[cost]
input = 0.300
output = 2.500
[limit]
context = 1_000_000
output = 65_536
[modalities]
input = ["text", "image"]
output = ["text"]
@@ -0,0 +1,22 @@
name = "gemini-2.5-pro"
family = "gemini-pro"
release_date = "2025-06-17"
last_updated = "2025-06-17"
attachment = true
reasoning = false
temperature = true
tool_call = true
open_weights = false
knowledge = "2025-01"
[cost]
input = 1.250
output = 10.000
[limit]
context = 1_000_000
output = 65_536
[modalities]
input = ["text", "image"]
output = ["text"]
@@ -0,0 +1,21 @@
name = "gemini-3-flash-preview"
release_date = "2025-12-18"
last_updated = "2025-12-18"
attachment = true
reasoning = false
temperature = true
tool_call = true
open_weights = false
knowledge = "2025-06"
[cost]
input = 0.500
output = 3.000
[limit]
context = 1_000_000
output = 65_536
[modalities]
input = ["text", "image"]
output = ["text"]
@@ -0,0 +1,21 @@
name = "gemini-3-pro-image-preview"
release_date = "2025-11-20"
last_updated = "2025-11-20"
attachment = true
reasoning = false
temperature = true
tool_call = false
open_weights = false
knowledge = "2025-06"
[cost]
input = 2.000
output = 120.000
[limit]
context = 32_768
output = 64_000
[modalities]
input = ["text", "image"]
output = ["text"]
@@ -0,0 +1,21 @@
name = "gemini-3-pro-preview"
release_date = "2025-11-19"
last_updated = "2025-11-19"
attachment = true
reasoning = false
temperature = true
tool_call = true
open_weights = false
knowledge = "2025-06"
[cost]
input = 2.000
output = 12.000
[limit]
context = 1_000_000
output = 64_000
[modalities]
input = ["text", "image"]
output = ["text"]
@@ -1,20 +1,20 @@
name = "Mistral Large (24.02)"
family = "mistral-large"
release_date = "2024-12-01"
last_updated = "2024-12-01"
name = "GLM-4.5"
release_date = "2025-07-29"
last_updated = "2025-07-29"
attachment = false
reasoning = false
temperature = true
tool_call = true
open_weights = false
knowledge = "2024-10"
[cost]
input = 0.50
output = 1.50
input = 0.286
output = 1.142
[limit]
context = 128_000
output = 4_096
output = 98_304
[modalities]
input = ["text"]
+21
View File
@@ -0,0 +1,21 @@
name = "GLM-4.5V"
release_date = "2025-07-29"
last_updated = "2025-07-29"
attachment = true
reasoning = false
temperature = true
tool_call = true
open_weights = false
knowledge = "2024-10"
[cost]
input = 0.290
output = 0.860
[limit]
context = 64_000
output = 16_384
[modalities]
input = ["text", "image"]
output = ["text"]
+21
View File
@@ -0,0 +1,21 @@
name = "glm-4.6"
release_date = "2025-09-30"
last_updated = "2025-09-30"
attachment = false
reasoning = false
temperature = true
tool_call = true
open_weights = false
knowledge = "2025-03"
[cost]
input = 0.286
output = 1.142
[limit]
context = 200_000
output = 131_072
[modalities]
input = ["text"]
output = ["text"]
+21
View File
@@ -0,0 +1,21 @@
name = "GLM-4.6V"
release_date = "2025-12-08"
last_updated = "2025-12-08"
attachment = true
reasoning = false
temperature = true
tool_call = true
open_weights = false
knowledge = "2025-03"
[cost]
input = 0.145
output = 0.430
[limit]
context = 128_000
output = 32_768
[modalities]
input = ["text", "image"]
output = ["text"]
+21
View File
@@ -0,0 +1,21 @@
name = "glm-4.7"
release_date = "2025-12-22"
last_updated = "2025-12-22"
attachment = false
reasoning = false
temperature = true
tool_call = true
open_weights = false
knowledge = "2025-06"
[cost]
input = 0.286
output = 1.142
[limit]
context = 200_000
output = 131_072
[modalities]
input = ["text"]
output = ["text"]
+22
View File
@@ -0,0 +1,22 @@
name = "gpt-4.1-mini"
family = "gpt-mini"
release_date = "2025-04-14"
last_updated = "2025-04-14"
attachment = true
reasoning = false
temperature = true
tool_call = true
open_weights = false
knowledge = "2024-04"
[cost]
input = 0.400
output = 1.600
[limit]
context = 1_000_000
output = 32_768
[modalities]
input = ["text", "image"]
output = ["text"]
+22
View File
@@ -0,0 +1,22 @@
name = "gpt-4.1-nano"
family = "gpt-nano"
release_date = "2025-04-14"
last_updated = "2025-04-14"
attachment = true
reasoning = false
temperature = true
tool_call = true
open_weights = false
knowledge = "2024-04"
[cost]
input = 0.100
output = 0.400
[limit]
context = 1_000_000
output = 32_768
[modalities]
input = ["text", "image"]
output = ["text"]
+22
View File
@@ -0,0 +1,22 @@
name = "gpt-4.1"
family = "gpt"
release_date = "2025-04-14"
last_updated = "2025-04-14"
attachment = true
reasoning = false
temperature = true
tool_call = true
open_weights = false
knowledge = "2024-04"
[cost]
input = 2.000
output = 8.000
[limit]
context = 1_000_000
output = 32_768
[modalities]
input = ["text", "image"]
output = ["text"]
+22
View File
@@ -0,0 +1,22 @@
name = "gpt-4o"
family = "gpt"
release_date = "2024-05-13"
last_updated = "2024-05-13"
attachment = true
reasoning = false
temperature = true
tool_call = true
open_weights = false
knowledge = "2023-09"
[cost]
input = 2.500
output = 10.000
[limit]
context = 128_000
output = 16_384
[modalities]
input = ["text", "image"]
output = ["text"]
+21
View File
@@ -0,0 +1,21 @@
name = "gpt-5-mini"
release_date = "2025-08-08"
last_updated = "2025-08-08"
attachment = true
reasoning = false
temperature = true
tool_call = true
open_weights = false
knowledge = "2024-10"
[cost]
input = 0.250
output = 2.000
[limit]
context = 400_000
output = 128_000
[modalities]
input = ["text", "image"]
output = ["text"]
+21
View File
@@ -0,0 +1,21 @@
name = "gpt-5-pro"
release_date = "2025-10-08"
last_updated = "2025-10-08"
attachment = true
reasoning = false
temperature = true
tool_call = true
open_weights = false
knowledge = "2024-10"
[cost]
input = 15.000
output = 120.000
[limit]
context = 400_000
output = 272_000
[modalities]
input = ["text", "image"]
output = ["text"]
@@ -0,0 +1,21 @@
name = "gpt-5-thinking"
release_date = "2025-08-08"
last_updated = "2025-08-08"
attachment = true
reasoning = true
temperature = true
tool_call = true
open_weights = false
knowledge = "2024-10"
[cost]
input = 1.250
output = 10.000
[limit]
context = 400_000
output = 128_000
[modalities]
input = ["text", "image"]
output = ["text"]
@@ -0,0 +1,21 @@
name = "gpt-5.1-chat-latest"
release_date = "2025-11-14"
last_updated = "2025-11-14"
attachment = true
reasoning = false
temperature = true
tool_call = true
open_weights = false
knowledge = "2024-10"
[cost]
input = 1.250
output = 10.000
[limit]
context = 128_000
output = 16_384
[modalities]
input = ["text", "image"]
output = ["text"]
+21
View File
@@ -0,0 +1,21 @@
name = "gpt-5.1"
release_date = "2025-11-14"
last_updated = "2025-11-14"
attachment = true
reasoning = false
temperature = true
tool_call = true
open_weights = false
knowledge = "2024-10"
[cost]
input = 1.250
output = 10.000
[limit]
context = 400_000
output = 128_000
[modalities]
input = ["text", "image"]
output = ["text"]
@@ -0,0 +1,21 @@
name = "gpt-5.2-chat-latest"
release_date = "2025-12-12"
last_updated = "2025-12-12"
attachment = true
reasoning = false
temperature = true
tool_call = true
open_weights = false
knowledge = "2024-10"
[cost]
input = 1.750
output = 14.000
[limit]
context = 128_000
output = 16_384
[modalities]
input = ["text", "image"]
output = ["text"]
+21
View File
@@ -0,0 +1,21 @@
name = "gpt-5.2"
release_date = "2025-12-12"
last_updated = "2025-12-12"
attachment = true
reasoning = false
temperature = true
tool_call = true
open_weights = false
knowledge = "2024-10"
[cost]
input = 1.750
output = 14.000
[limit]
context = 400_000
output = 128_000
[modalities]
input = ["text", "image"]
output = ["text"]
+21
View File
@@ -0,0 +1,21 @@
name = "gpt-5"
release_date = "2025-08-08"
last_updated = "2025-08-08"
attachment = true
reasoning = false
temperature = true
tool_call = true
open_weights = false
knowledge = "2024-10"
[cost]
input = 1.250
output = 10.000
[limit]
context = 400_000
output = 128_000
[modalities]
input = ["text", "image"]
output = ["text"]
@@ -0,0 +1,21 @@
name = "grok-4-1-fast-non-reasoning"
release_date = "2025-11-20"
last_updated = "2025-11-20"
attachment = true
reasoning = false
temperature = true
tool_call = true
open_weights = false
knowledge = "2025-06"
[cost]
input = 0.200
output = 0.500
[limit]
context = 2_000_000
output = 30_000
[modalities]
input = ["text", "image"]
output = ["text"]
@@ -0,0 +1,21 @@
name = "grok-4-1-fast-reasoning"
release_date = "2025-11-20"
last_updated = "2025-11-20"
attachment = true
reasoning = true
temperature = true
tool_call = true
open_weights = false
knowledge = "2025-06"
[cost]
input = 0.200
output = 0.500
[limit]
context = 2_000_000
output = 30_000
[modalities]
input = ["text", "image"]
output = ["text"]
@@ -0,0 +1,21 @@
name = "grok-4-fast-non-reasoning"
release_date = "2025-09-23"
last_updated = "2025-09-23"
attachment = true
reasoning = false
temperature = true
tool_call = true
open_weights = false
knowledge = "2025-06"
[cost]
input = 0.200
output = 0.500
[limit]
context = 2_000_000
output = 30_000
[modalities]
input = ["text", "image"]
output = ["text"]
@@ -1,17 +1,16 @@
name = "Grok 4 Fast (Reasoning)"
family = "grok"
release_date = "2025-09-19"
last_updated = "2025-09-19"
name = "grok-4-fast-reasoning"
release_date = "2025-09-23"
last_updated = "2025-09-23"
attachment = true
reasoning = true
temperature = true
knowledge = "2025-07"
tool_call = true
open_weights = false
knowledge = "2025-06"
[cost]
input = 0
output = 0
input = 0.200
output = 0.500
[limit]
context = 2_000_000
@@ -1,18 +1,21 @@
name = "Gemini 3 Pro Preview"
family = "gemini-pro"
name = "grok-4.1"
release_date = "2025-11-18"
last_updated = "2025-11-18"
attachment = true
reasoning = false
tool_call = true
structured_output = true
temperature = true
tool_call = true
open_weights = false
knowledge = "2025-06"
[cost]
input = 2.000
output = 10.000
[limit]
context = 1_000_000
context = 200_000
output = 64_000
[modalities]
input = ["text", "image", "audio", "video"]
output = ["text"]
input = ["text", "image"]
output = ["text"]
@@ -0,0 +1,21 @@
name = "kimi-k2-0905-preview"
release_date = "2025-09-05"
last_updated = "2025-09-05"
attachment = false
reasoning = false
temperature = true
tool_call = true
open_weights = false
knowledge = "2025-06"
[cost]
input = 0.632
output = 2.530
[limit]
context = 262_144
output = 262_144
[modalities]
input = ["text"]
output = ["text"]
@@ -0,0 +1,21 @@
name = "kimi-k2-thinking-turbo"
release_date = "2025-09-05"
last_updated = "2025-09-05"
attachment = false
reasoning = true
temperature = true
tool_call = true
open_weights = false
knowledge = "2025-06"
[cost]
input = 1.265
output = 9.119
[limit]
context = 262_144
output = 262_144
[modalities]
input = ["text"]
output = ["text"]
@@ -0,0 +1,21 @@
name = "kimi-k2-thinking"
release_date = "2025-09-05"
last_updated = "2025-09-05"
attachment = false
reasoning = true
temperature = true
tool_call = true
open_weights = false
knowledge = "2025-06"
[cost]
input = 0.575
output = 2.300
[limit]
context = 262_144
output = 262_144
[modalities]
input = ["text"]
output = ["text"]
@@ -0,0 +1,21 @@
name = "ministral-14b-2512"
release_date = "2025-12-16"
last_updated = "2025-12-16"
attachment = true
reasoning = false
temperature = true
tool_call = true
open_weights = false
knowledge = "2024-12"
[cost]
input = 0.330
output = 0.330
[limit]
context = 128_000
output = 128_000
[modalities]
input = ["text", "image"]
output = ["text"]
@@ -0,0 +1,21 @@
name = "mistral-large-2512"
release_date = "2025-12-16"
last_updated = "2025-12-16"
attachment = true
reasoning = false
temperature = true
tool_call = true
open_weights = false
knowledge = "2024-12"
[cost]
input = 1.100
output = 3.300
[limit]
context = 128_000
output = 262_144
[modalities]
input = ["text", "image"]
output = ["text"]
+20
View File
@@ -0,0 +1,20 @@
name = "Qwen-Flash"
release_date = "2025-07-28"
last_updated = "2025-07-28"
attachment = false
reasoning = false
temperature = true
tool_call = true
open_weights = false
[cost]
input = 0.022
output = 0.220
[limit]
context = 1_000_000
output = 32_768
[modalities]
input = ["text"]
output = ["text"]
@@ -0,0 +1,22 @@
name = "Qwen-Max-Latest"
family = "qwen"
release_date = "2024-04-03"
last_updated = "2025-01-25"
attachment = false
reasoning = false
temperature = true
tool_call = true
open_weights = false
knowledge = "2024-11"
[cost]
input = 0.343
output = 1.372
[limit]
context = 131_072
output = 8_192
[modalities]
input = ["text"]
output = ["text"]
+22
View File
@@ -0,0 +1,22 @@
name = "Qwen-Plus"
family = "qwen"
release_date = "2024-07-23"
last_updated = "2024-07-23"
attachment = false
reasoning = false
temperature = true
tool_call = true
open_weights = false
knowledge = "2024-10"
[cost]
input = 0.120
output = 1.200
[limit]
context = 1_000_000
output = 32_768
[modalities]
input = ["text"]
output = ["text"]
@@ -0,0 +1,21 @@
name = "qwen3-235b-a22b-instruct-2507"
release_date = "2025-07-30"
last_updated = "2025-07-30"
attachment = false
reasoning = false
temperature = true
tool_call = true
open_weights = false
knowledge = "2025-04"
[cost]
input = 0.290
output = 1.143
[limit]
context = 128_000
output = 65_536
[modalities]
input = ["text"]
output = ["text"]
@@ -1,17 +1,17 @@
name = "Qwen3 235B-A22B"
name = "Qwen3-235B-A22B"
family = "qwen"
release_date = "2025-04-29"
last_updated = "2025-04-29"
attachment = false
reasoning = true
reasoning = false
temperature = true
knowledge = "2025-04"
tool_call = true
open_weights = true
open_weights = false
knowledge = "2025-04"
[cost]
input = 0.22
output = 0.88
input = 0.290
output = 2.860
[limit]
context = 128_000
+22
View File
@@ -0,0 +1,22 @@
name = "Qwen3-30B-A3B"
family = "qwen"
release_date = "2025-04-29"
last_updated = "2025-04-29"
attachment = false
reasoning = false
temperature = true
tool_call = true
open_weights = false
knowledge = "2025-04"
[cost]
input = 0.110
output = 1.080
[limit]
context = 128_000
output = 8_192
[modalities]
input = ["text"]
output = ["text"]
@@ -0,0 +1,21 @@
name = "qwen3-coder-480b-a35b-instruct"
release_date = "2025-07-23"
last_updated = "2025-07-23"
attachment = false
reasoning = false
temperature = true
tool_call = true
open_weights = false
knowledge = "2025-04"
[cost]
input = 0.860
output = 3.430
[limit]
context = 262_144
output = 65_536
[modalities]
input = ["text"]
output = ["text"]
@@ -0,0 +1,21 @@
name = "qwen3-max-2025-09-23"
release_date = "2025-09-24"
last_updated = "2025-09-24"
attachment = false
reasoning = false
temperature = true
tool_call = true
open_weights = false
knowledge = "2025-04"
[cost]
input = 0.860
output = 3.430
[limit]
context = 258_048
output = 65_536
[modalities]
input = ["text"]
output = ["text"]
+5
View File
@@ -0,0 +1,5 @@
name = "302.AI"
env = ["302AI_API_KEY"]
npm = "@ai-sdk/openai-compatible"
doc = "https://doc.302.ai"
api = "https://api.302.ai/v1"
@@ -0,0 +1,22 @@
name = "Claude Opus 4.6"
family = "claude-opus"
release_date = "2026-02-05"
last_updated = "2026-02-05"
attachment = true
reasoning = true
temperature = true
tool_call = true
knowledge = "2025-05"
open_weights = false
[cost]
input = 5.00
output = 25.00
[limit]
context = 200_000
output = 128_000
[modalities]
input = ["text", "image", "pdf"]
output = ["text"]
@@ -1,12 +1,12 @@
name = "Claude Sonnet 3"
name = "Claude Sonnet 4.6"
family = "claude-sonnet"
release_date = "2024-03-04"
last_updated = "2024-03-04"
release_date = "2026-02-17"
last_updated = "2026-02-17"
attachment = true
reasoning = false
reasoning = true
temperature = true
knowledge = "2023-08"
tool_call = true
knowledge = "2025-08"
open_weights = false
[cost]
@@ -15,7 +15,7 @@ output = 15.00
[limit]
context = 200_000
output = 4_096
output = 64_000
[modalities]
input = ["text", "image", "pdf"]
@@ -1,7 +1,7 @@
name = "DeepSeek V3.1"
family = "deepseek"
release_date = "2025-08-01"
last_updated = "2025-08-01"
release_date = "2025-01-20"
last_updated = "2025-01-20"
attachment = false
reasoning = true
temperature = true
@@ -0,0 +1,24 @@
name = "Gemini 3.1 Flash Lite Preview"
family = "gemini-flash"
release_date = "2026-03-01"
last_updated = "2026-03-01"
attachment = true
reasoning = true
temperature = true
tool_call = true
structured_output = true
open_weights = false
[cost]
input = 0.25
output = 1.50
cache_read = 0.025
cache_write = 1.00
[limit]
context = 1_048_576
output = 65_536
[modalities]
input = ["text", "image", "audio", "video", "pdf"]
output = ["text"]
@@ -0,0 +1,23 @@
name = "Gemini 3.1 Pro Preview"
family = "gemini-pro"
release_date = "2026-02-19"
last_updated = "2026-02-19"
attachment = true
reasoning = true
temperature = true
tool_call = true
structured_output = true
knowledge = "2025-01"
open_weights = false
[cost]
input = 2.00
output = 12.00
[limit]
context = 1_048_576
output = 65_536
[modalities]
input = ["text", "image", "video", "audio", "pdf"]
output = ["text"]
@@ -1,5 +1,5 @@
name = "GPT-5-Codex"
family = "gpt-codex"
name = "GPT-5 Codex"
family = "gpt"
release_date = "2025-09-15"
last_updated = "2025-09-15"
attachment = false
@@ -7,14 +7,16 @@ reasoning = true
temperature = false
knowledge = "2024-09-30"
tool_call = true
structured_output = true
open_weights = false
[cost]
input = 0
output = 0
input = 1.25
output = 10.00
[limit]
context = 128_000
context = 400_000
input = 272_000
output = 128_000
[modalities]
@@ -0,0 +1,24 @@
name = "GPT-5.1 Codex Max"
family = "gpt"
release_date = "2025-11-13"
last_updated = "2025-11-13"
attachment = true
reasoning = true
temperature = false
knowledge = "2024-09-30"
tool_call = true
structured_output = true
open_weights = false
[cost]
input = 1.25
output = 10.00
[limit]
context = 400_000
input = 272_000
output = 128_000
[modalities]
input = ["text", "image"]
output = ["text"]
@@ -0,0 +1,24 @@
name = "GPT-5.1 Codex"
family = "gpt"
release_date = "2025-11-13"
last_updated = "2025-11-13"
attachment = true
reasoning = true
temperature = false
knowledge = "2024-09-30"
tool_call = true
structured_output = true
open_weights = false
[cost]
input = 1.25
output = 10.00
[limit]
context = 400_000
input = 272_000
output = 128_000
[modalities]
input = ["text", "image"]
output = ["text"]
@@ -0,0 +1,22 @@
name = "GPT-5.2 Chat Latest"
family = "gpt"
release_date = "2026-01-01"
last_updated = "2026-01-01"
attachment = true
reasoning = true
temperature = true
knowledge = "2024-09-30"
tool_call = true
open_weights = false
[cost]
input = 1.75
output = 14.00
[limit]
context = 400_000
output = 128_000
[modalities]
input = ["text", "image"]
output = ["text"]
@@ -0,0 +1,24 @@
name = "GPT-5.2 Codex"
family = "gpt"
release_date = "2025-12-11"
last_updated = "2025-12-11"
attachment = true
reasoning = true
temperature = false
knowledge = "2025-08-31"
tool_call = true
structured_output = true
open_weights = false
[cost]
input = 1.75
output = 14.00
[limit]
context = 400_000
input = 272_000
output = 128_000
[modalities]
input = ["text", "image", "pdf"]
output = ["text"]
@@ -0,0 +1,21 @@
name = "GPT-5.3 Chat Latest"
family = "gpt"
release_date = "2026-03-01"
last_updated = "2026-03-01"
attachment = true
reasoning = true
temperature = true
tool_call = true
open_weights = false
[cost]
input = 1.75
output = 14.00
[limit]
context = 400_000
output = 128_000
[modalities]
input = ["text", "image"]
output = ["text"]
@@ -0,0 +1,24 @@
name = "GPT-5.3 Codex XHigh"
family = "gpt"
release_date = "2026-02-05"
last_updated = "2026-02-05"
attachment = true
reasoning = true
temperature = false
knowledge = "2025-08-31"
tool_call = true
structured_output = true
open_weights = false
[cost]
input = 1.75
output = 14.00
[limit]
context = 400_000
input = 272_000
output = 128_000
[modalities]
input = ["text", "image", "pdf"]
output = ["text"]
@@ -0,0 +1,24 @@
name = "GPT-5.3 Codex"
family = "gpt"
release_date = "2026-02-05"
last_updated = "2026-02-05"
attachment = true
reasoning = true
temperature = false
knowledge = "2025-08-31"
tool_call = true
structured_output = true
open_weights = false
[cost]
input = 1.75
output = 14.00
[limit]
context = 400_000
input = 272_000
output = 128_000
[modalities]
input = ["text", "image", "pdf"]
output = ["text"]

Some files were not shown because too many files have changed in this diff Show More