Aiden Cline
cda892e0d7
sync github copilot limits
2026-03-19 21:54:42 -05:00
Aiden Cline
098ff4f5bf
Merge pull request #1227 from Verizane/dev
...
add OpenRouter models for gpt-5.4 mini and gpt-5.4 nano
2026-03-19 21:23:15 -05:00
Aiden Cline
0f70b8959f
Merge pull request #1234 from mchenco/dev
...
Add Workers AI models: kimi-k2.5, nemotron-3-120b-a12b, glm-4.7-flash
2026-03-19 15:11:19 -05:00
mchen
b8e6d58e5b
add workers-ai models: kimi-k2.5, nemotron-3-120b-a12b, glm-4.7-flash
2026-03-19 14:59:17 -04:00
Roman Koslowski
a855001a7e
apply changes from review
2026-03-19 17:20:26 +01:00
Aiden Cline
ac760b2268
Merge pull request #1230 from SamizuHM/feature/zhipuai-coding-plan-add-glm-5-turbo
...
zhipuai-coding-plan: Add glm-5-turbo.toml and replace symlink
2026-03-19 10:42:43 -05:00
Aiden Cline
d4a5ea7ae7
Merge pull request #1226 from spiffytech/dev
...
Add Ollama Cloud support for Minimax M2.7
2026-03-19 10:41:47 -05:00
Aiden Cline
434ed89ba2
Merge pull request #1228 from dpuyosa/minimax_m2_7
...
Venice: Add MiniMax M2.7 and update DeepSeek V3.2 pricing
2026-03-19 10:41:16 -05:00
Aiden Cline
6d7719a62a
Merge pull request #1229 from 0b1000/dev
...
Xiaomi: Add MiMo-V2-Pro and MiMo-V2-Omni
2026-03-19 10:41:06 -05:00
Aiden Cline
93637039ef
Merge pull request #1231 from ariane-emory/feat/feat/add-xiaomi-mimo-v2-pro-and-omni
...
feat: add the Xiaomi MiMo V2 Pro and Xiaomi MiMo V2 Omni models to the OpenRouter provide
2026-03-19 10:40:44 -05:00
Ariane Emory
9c95f796c0
Merge remote-tracking branch 'upstream/dev' into feat/feat/add-xiaomi-mimo-v2-pro
2026-03-19 11:22:43 -04:00
Ariane Emory
e8650b6073
feat: add xiaomi mimo-v2-pro and mimo-v2-omni models to openrouter
2026-03-19 11:18:46 -04:00
SamizuHM
23c2be6ff7
feat(zhipuai-coding-plan): add glm-5-turbo.toml and replace glm-5-turbo with symlink
2026-03-19 18:09:17 +08:00
Frank
913a63dbe6
update zen models
2026-03-19 00:33:45 -04:00
0b1000
503087e99b
Merge branch 'anomalyco:dev' into dev
2026-03-19 12:28:38 +08:00
0b1000
48150f09d3
Xiaomi: Add MiMo-V2-Pro and MiMo-V2-Omni
2026-03-19 12:27:00 +08:00
Aiden Cline
5fef681657
Disable tool_call in grok model configuration
2026-03-18 23:09:30 -05:00
Frank
123054ae0c
update zen models
2026-03-18 20:45:44 -04:00
Frank
03060d154b
update zen models
2026-03-18 20:37:47 -04:00
dpuyosa
5c9b8108e0
Update minimax-m27.toml
2026-03-19 01:02:24 +01:00
dpuyosa
c8084681f9
[venice] Add MiniMax M2.7 and update DeepSeek V3.2 pricing
...
- Add MiniMax M2.7 model with reasoning and tool_call support
- Update DeepSeek V3.2 pricing (input: $0.33, output: $0.48, cache: $0.16)
2026-03-19 00:58:50 +01:00
Roman Koslowski
352ab4ae1b
add gpt-5.4 mini and gpt-5.4 nano
2026-03-18 22:16:55 +01:00
spiffytech
cf0b416b15
Added Ollama Cloud support for Minimax M2.7
2026-03-18 16:15:07 -04:00
Aiden Cline
38339a2a90
Merge pull request #1224 from APonce911/minimax-m2.7-openrouter
...
add MiniMax M2.7 to OpenRouter
2026-03-18 14:10:13 -05:00
Aiden Cline
ff9040bf52
Update minimax-m2.7.toml
2026-03-18 14:09:26 -05:00
Aiden Cline
3039804af4
Delete providers/opencode/models/minimax-m2.7.toml
2026-03-18 14:08:55 -05:00
Frank
7a4ad7bec8
update go models
2026-03-18 14:40:25 -04:00
airton
721cc122bc
add MiniMax M2.7 to OpenRouter and OpenCode
2026-03-18 18:57:09 +01:00
Aiden Cline
0527f019af
Merge pull request #1221 from sergical/fix/bedrock-claude-4-6-context-window-and-pricing
...
fix(amazon-bedrock): set Claude Sonnet 4.6 and Opus 4.6 context window to 1M
2026-03-18 12:17:57 -05:00
Aiden Cline
c89371de50
Merge pull request #1223 from sylviezhang37/update-vercel-models-20260318-1659
...
Update Vercel models
2026-03-18 12:17:22 -05:00
Sylvie Zhang
6d6d4220d8
Enable open_weights in minimax-m2.7.toml
2026-03-18 10:12:37 -07:00
Sylvie Zhang
8b984eeec1
Enable open_weights in minimax-m2.7-highspeed model
2026-03-18 10:12:21 -07:00
github-actions[bot]
586027c8f1
chore(vercel): update Vercel model definitions
...
Auto-generated by weekly workflow from Vercel AI Gateway API.
Co-Authored-By: github-actions[bot] <41898282+github-actions[bot]@users.noreply.github.com>
2026-03-18 16:59:24 +00:00
Sergiy Dybskiy
343b5f87ef
fix(amazon-bedrock): set Claude Sonnet 4.6 and Opus 4.6 context window to 1M
...
Both models support a 1M token context window natively on Bedrock via the
Converse API with no beta headers required. Verified empirically via the
AWS CLI (bedrock-runtime converse): 950K tokens succeeds, >1M returns
'prompt is too long: N tokens > 1000000 maximum'.
The AWS Bedrock pricing page confirms long context pricing for these two
models is identical to standard pricing (no surcharge), so the
[cost.context_over_200k] section is removed as it was incorrect.
2026-03-18 12:20:38 -04:00
Aiden Cline
955b773ee5
Merge pull request #1218 from pomidornijfrukt/azure/5.4-mini-nano
...
Add GPT-5.4 Mini and Nano models for Azure providers
2026-03-18 10:31:13 -05:00
Aiden Cline
98559071f0
Merge pull request #1217 from cgilly2fast/dev
...
chore(firmware): update base url and docs url
2026-03-18 10:30:44 -05:00
eCube-cachy
0660308816
add: GPT-5.4 Mini and Nano model configurations for Azure providers
2026-03-18 15:17:52 +02:00
Jack
380f9dd8eb
Merge pull request #1216 from no1wudi/dev
...
Add MiniMax M2.7 and M2.7-highspeed models to 4 official providers
2026-03-18 16:29:59 +08:00
Jack
1cfdab1b18
update MiniMax-M2.7 cache_read to 0.06
2026-03-18 16:27:53 +08:00
Colby Gilbert
75a981f957
chore(firmware): update base url and docs url
2026-03-18 00:41:05 -07:00
Huang Qi
7fadbcadc8
Add MiniMax M2.7 and M2.7-highspeed models to 4 official providers
2026-03-18 15:21:06 +08:00
Frank
38f9092292
update zen models
2026-03-18 02:30:18 -04:00
Aiden Cline
92149b9eaa
rm nonexistant github model
2026-03-17 21:41:51 -05:00
Aiden Cline
b614f0e69c
Merge pull request #1214 from luisrudge/dev
...
Add GPT-5.4 mini and nano to GitHub Copilot provider
2026-03-17 20:13:46 -05:00
Luís Rudge
67d6dac5c5
Add GPT-5.4 mini and nano to GitHub Copilot provider
2026-03-17 18:44:38 -06:00
Aiden Cline
7d3cc61a48
Merge pull request #1207 from PedroACosta/feat/add-dinference-provider
...
feat(providers): add dinference provider
2026-03-17 14:51:31 -05:00
Aiden Cline
f02ea6c4d2
Merge pull request #1115 from skywalker512/feat/add-tencent-coding-plan
...
feat: add Tencent Coding Plan provider
2026-03-17 14:51:19 -05:00
Aiden Cline
0cb50eeece
Merge pull request #1208 from scwgoire/march-update
...
Scaleway 26-03 model updates
2026-03-17 14:48:12 -05:00
Aiden Cline
878311d2e0
Merge pull request #1210 from dm-cohere/dm/fix-update-cohere-model-capabilities
...
fix(models): update cohere model capabilities
2026-03-17 14:32:27 -05:00
Aiden Cline
a0e89f65d6
Merge pull request #1206 from 0b1000/dev
...
Rename minimax-m2.5.toml to MiniMax-M2.5.toml
2026-03-17 14:32:19 -05:00
Aiden Cline
74099b7c9c
Merge pull request #1213 from smrdotgg/add-openai-gpt-5-4-mini-and-nano
...
Add OpenAI GPT-5.4 mini and nano
2026-03-17 14:31:24 -05:00
Aiden Cline
ec522435c3
Merge pull request #1211 from sylviezhang37/update-vercel-models-20260317-1807
...
Update Vercel models
2026-03-17 14:30:24 -05:00
smr
d839cd37d4
Add OpenAI GPT-5.4 mini and nano
...
Capture the newly released mini and nano model metadata so models.dev reflects OpenAI's latest GPT-5.4 lineup with current pricing, limits, and knowledge cutoff.
2026-03-17 22:09:13 +03:00
github-actions[bot]
ecb6ef7f93
chore(vercel): update Vercel model definitions
...
Auto-generated by weekly workflow from Vercel AI Gateway API.
Co-Authored-By: github-actions[bot] <41898282+github-actions[bot]@users.noreply.github.com>
2026-03-17 18:07:06 +00:00
Deirdre Meehan
8b4d341054
fix: cohere models on non-cohere providers
2026-03-17 16:51:24 +00:00
Deirdre Meehan
62f4a28308
fix: cohere provider models
2026-03-17 16:44:55 +00:00
Pedro
2fb8ef0dc8
feat(providers): add dinference provider
2026-03-17 14:13:36 +01:00
Gregoire de Turckheim
96968e2bf8
feat: Scaleway 26-03 model updates
2026-03-17 12:15:52 +01:00
0b1000
ee9d7879ce
Rename minimax-m2.5.toml to MiniMax-M2.5.toml
2026-03-17 14:50:35 +08:00
Frank
71283512a6
update zen models
2026-03-17 02:21:13 -04:00
Frank
cd4afd7e7c
update zen models
2026-03-17 02:19:17 -04:00
Aiden Cline
1239d0190b
Merge pull request #1204 from cyberofficial/vultr
...
VULTR: Updated Vultr model pricing to reflect current serverless inference rates
2026-03-16 16:10:39 -05:00
Aiden Cline
491bf6ccba
Merge pull request #1202 from RaviTharuma/fix/chutes-pricing-update-2026-03
...
fix(chutes): update pricing and limits from live API
2026-03-16 16:10:25 -05:00
Cyber Official
c993d0c121
Updated Vultr model pricing to reflect current serverless inference rates
...
Updated Vultr model pricing to reflect current serverless inference rates
This commit updates the cost configuration for all Vultr models to align with their latest pricing tiers:
**Cost Reductions:**
- DeepSeek-R1-Distill-Qwen-32B: Input $0.55→$0.30, Output $2.75→$0.30 (73% reduction)
- NVIDIA-Nemotron-3-Super-120B-A12B-NVFP4: Input $0.55→$0.20, Output $2.75→$0.80 (64% input, 71% output reduction)
- Qwen2.5-Coder-32B-Instruct: Input $0.55→$0.20, Output $2.75→$0.60 (64% input, 78% output reduction)
- gpt-oss-120b: Input $0.55→$0.15, Output $2.75→$0.60 (73% input, 78% output reduction)
- MiniMax-M2.5: Input $0.55→$0.30, Output $2.75→$1.20 (45% input, 56% output reduction)
**Cost Adjustments:**
- DeepSeek-R1-Distill-Llama-70B: Input $0.55→$2.00, Output $2.75→$2.00 (significant increase)
- DeepSeek-V3.2: Output $2.75→$1.65 (40% reduction)
- Llama-3.1-Nemotron-Ultra-253B-v1: Output $2.75→$1.80 (35% reduction)
- GLM-5-FP8: Input $0.55→$0.85, Output $2.75→$3.10 (55% input, 13% output increase)
2026-03-16 13:58:28 -04:00
Aiden Cline
e55c39a83d
Merge pull request #1141 from sk0x0y/feature/nanogpt-confirmed-suffix2-fixes
...
fix(nano-gpt): rename confirmed 2-suffix model ids
2026-03-16 10:57:44 -05:00
Aiden Cline
6dea000e25
Merge pull request #1148 from sk0x0y/feature/nanogpt-bundled-confirmed-suffix2-fixes
...
fix(nano-gpt): rename bundled confirmed 2-suffix model ids
2026-03-16 10:57:30 -05:00
Aiden Cline
3f7a757b3f
Merge pull request #1142 from sk0x0y/feature/nanogpt-more-confirmed-suffix2-fixes
...
fix(nano-gpt): rename more confirmed 2-suffix model ids
2026-03-16 10:56:36 -05:00
Aiden Cline
c693fd71e2
Merge pull request #1194 from cyberofficial/vultr
...
Update Vultr model list with 10 new models and updated pricing
2026-03-16 10:55:19 -05:00
Aiden Cline
54e04e288a
Merge pull request #1198 from amritbanerjee/add-glm-5-turbo
...
Add GLM-5-Turbo model support
2026-03-16 10:47:08 -05:00
Aiden Cline
462a179eee
Merge pull request #1203 from jerome-benoit/fix/sap-ai-core-model-specs
...
fix(sap-ai-core): align model specs with official sources
2026-03-16 10:46:40 -05:00
Aiden Cline
95db59034d
Merge pull request #1201 from dpuyosa/venice-new-models
...
Venice: Add new provider models
2026-03-16 10:45:59 -05:00
Aiden Cline
74dcc74e32
Merge pull request #1200 from dpuyosa/venice/pricing-update
...
Venice: Update model pricing for 7 models
2026-03-16 10:45:47 -05:00
Jérôme Benoit
57975f5f25
fix(sap-ai-core): align model specs with official sources
2026-03-16 13:59:06 +01:00
Ravi Tharuma
ad7b063747
fix(chutes): update pricing and limits from live API
...
Synced 6 Chutes model definitions against the live API at
https://llm.chutes.ai/v1/models (queried 2026-03-16).
Models updated:
- deepseek-ai/DeepSeek-V3.2-TEE: cost 0.25/0.38→0.28/0.42, cache 0.125→0.14, context 163840→131072
- zai-org/GLM-5-TEE: cost 0.75/2.5→0.95/3.15, added cache_read 0.475
- zai-org/GLM-4.6-TEE: cost 0.35/1.5→0.4/1.7, added cache_read 0.2
- zai-org/GLM-4.6V: added cache_read 0.15
- MiniMaxAI/MiniMax-M2.5-TEE: cost 0.15/0.6→0.3/1.1, added cache_read 0.15
- Qwen/Qwen3.5-397B-A17B-TEE: cost 0.3/1.2→0.39/2.34, cache 0.15→0.195
2026-03-16 11:42:29 +01:00
dpuyosa
f76e9f0551
[venice] Add new provider models
...
- Add mistral-small-3.2-24b-instruct, qwen3-5-9b, venice-uncensored-role-play, zai-org-glm-4.6
2026-03-16 09:37:41 +01:00
dpuyosa
d70a49b36f
[venice] Update model pricing for 7 models
...
- Remove context_over_200k pricing from Claude models
- Update Grok cache_read pricing from 0.5 to 0.25
- Update Kimi, MiniMax input/output pricing
2026-03-16 09:05:37 +01:00
amrit
3487135f9f
Add GLM-5-Turbo model support
2026-03-16 12:14:50 +11:00
Aiden Cline
458a66c766
Merge pull request #1197 from kesku/update-perplexity-agent-models
...
Update Perplexity Agent API models
2026-03-15 10:59:23 -05:00
Frank
d3a84dc7ec
update zen models
2026-03-15 10:59:52 -04:00
Kesku
ae61b25583
update perplexity-agent: add gpt-5.4 & nemotron, remove gemini-3-pro
2026-03-15 03:46:50 +00:00
Aiden Cline
74be576eda
Merge pull request #1178 from Sewer56/change-synthetic-endpoint
...
Add OpenAI and Anthropic compatible endpoints
2026-03-14 20:55:30 -05:00
Aiden Cline
164df2cda0
Merge pull request #1191 from Alcatraz-Zhang/update/kilo-models
...
Sync Kilo model definitions with latest gateway catalog
2026-03-14 20:54:45 -05:00
Cyber Official
2cd7908369
Update Vultr model list with 10 new models and updated pricing
...
- Updated pricing to $0.55/M input tokens, $2.75/M output tokens
- Updated context limits to safe floor values from official testing
- Added accurate output token limits from official model documentation
- Added 5 new models: MiniMax M2.5, DeepSeek V3.2, GLM-5 FP8, Llama 3.1 Nemotron Ultra 253B, NVIDIA Nemotron 3 Super 120B A12B NVFP4
- Updated existing models: DeepSeek R1 Distill variants, GPT OSS 120B, Kimi K2.5, Qwen2.5 Coder 32B
Model specifications:
- MiniMax M2.5: 196K context, 4,096 output
- Qwen2.5-Coder-32B: 15K context, 256 output (notable low default)
- DeepSeek R1 Distill Llama 70B: 130K context, 4,096 output
- DeepSeek R1 Distill Qwen 32B: 130K context, 4,096 output
- DeepSeek V3.2: 163K context, 4,096 output
- Kimi K2.5: 261K context, 32,768 output (high output limit)
- GPT OSS 120B: 130K context, 8,192 output
- GLM-5 FP8: 202K context, 131,072 output (exceptionally high)
- Llama 3.1 Nemotron Ultra 253B: 32K context, 4,096 output
- NVIDIA Nemotron 3 Super 120B A12B NVFP4: 260K context, 8,192 output
All models set to text-only (no vision support) as confirmed.
2026-03-14 19:47:01 -04:00
Alcatraz-Zhang
cc667340f5
Sync Kilo model definitions with latest gateway catalog
...
Refresh the Kilo provider catalog so models.dev matches the current gateway inventory, pricing, and availability.
2026-03-15 04:35:38 +08:00
Sewer56
f2cfc1435d
Changed: Synthetic to use newer openai endpoint
2026-03-14 17:09:44 +00:00
Aiden Cline
35bb8cca47
Merge pull request #1172 from bigfluffycookie/add-deepinfra-llama-models
...
Add deepinfra llama models
2026-03-14 10:55:13 -05:00
Aiden Cline
3468a410e1
Merge pull request #1177 from ar27111994/dev
...
Add Grok 4.1 Fast configurations for reasoning and non-reasoning
2026-03-14 10:54:57 -05:00
Aiden Cline
b1b5e3c5cd
Merge pull request #1174 from dacbd/patch-1
...
fix(wandb): fix k2.5 settings
2026-03-14 10:54:35 -05:00
Aiden Cline
97f03ec672
Merge pull request #1175 from dacbd/patch-2
...
chore(docs): add note for manual testing with opencode
2026-03-14 10:54:22 -05:00
BigFluffyCookie
9b516924aa
Add limit output for llama models
2026-03-14 11:49:57 +01:00
Ahmed Rehan
929a39600b
feat(models): add Grok 4.1 Fast (Reasoning and Non-Reasoning) configurations
2026-03-14 14:27:24 +05:00
Daniel Barnes
a87d8bb8cc
chore(docs): add note for manual testing with opencode
2026-03-14 13:42:57 +09:00
Daniel Barnes
574139eb49
fix(wandb): fix k2.5 settings
2026-03-14 13:07:16 +09:00
Aiden Cline
1e3bc38b31
Merge pull request #1137 from mcowger/mcowger/correct-gemini-flash-lite-pricing
...
Fix incorrect pricing for gemini-3.1-flash-lite-preview
2026-03-13 18:41:41 -05:00
Aiden Cline
8916fe9874
Merge pull request #1171 from stephenkuhn214/dev
...
fix(amazon-bedrock): Remove deprecated and add missing models
2026-03-13 18:26:25 -05:00
BigFluffyCookie
5d956b41a6
Rename llama models to remove "Meta" prefix
2026-03-13 23:15:27 +01:00
BigFluffyCookie
42a7a14f69
Add Meta Llama models to DeepInfra provider
2026-03-13 22:53:39 +01:00
Stephen Kuhn
f24ee000d7
fix(amazon-bedrock): update and add models
...
- Remove 19 deprecated/EOL models
- Add 7 new models: DeepSeek V3.2, Llama 3.1 405B, Magistral Small 1.2, Ministral 3 3B, Mistral Large 3, Pixtral Large, NVIDIA Nemotron Nano 3 30B
- Fix Devstral 2 123B: correct name, family, and open_weights
- Set accurate Bedrock launch dates for all new models
2026-03-13 16:02:04 -04:00
Aiden Cline
7196b1fb2c
Merge pull request #1170 from anomalyco/revert-1166-fix/update-gpt53-codex-spark-preview
...
Revert "fix(openai): rename gpt-5.3-codex-spark to gpt-5.3-codex-spark-preview"
2026-03-13 14:31:33 -05:00
Aiden Cline
f6c0d5a29d
Revert "fix(openai): rename gpt-5.3-codex-spark to gpt-5.3-codex-spark-preview"
2026-03-13 14:30:58 -05:00
Aiden Cline
ee63449aa5
sonnet 4.6 and opus 4.6 1M context
2026-03-13 14:27:55 -05:00
Aiden Cline
92aa44ec00
Merge pull request #1166 from rluisr/fix/update-gpt53-codex-spark-preview
...
fix(openai): rename gpt-5.3-codex-spark to gpt-5.3-codex-spark-preview
2026-03-13 14:18:41 -05:00
Aiden Cline
477284535c
Rename model from 'GPT-5.3 Codex Spark Preview' to 'GPT-5.3 Codex Spark'
2026-03-13 14:17:44 -05:00
Aiden Cline
304233bdda
Merge pull request #1169 from mdrxy/mdrxy/anthropic-token-limits
...
Update Claude 4.6 context/pricing
2026-03-13 14:13:40 -05:00
Aiden Cline
25d782ee2c
Reduce context limit from 1,000,000 to 200,000
2026-03-13 14:13:30 -05:00
Aiden Cline
0f63393d51
Update context limit in claude-opus-4-6.toml
2026-03-13 14:12:56 -05:00
rluisr
e780eefce2
fix(openai): rename gpt-5.3-codex-spark to gpt-5.3-codex-spark-preview
...
The OpenAI API expects model ID 'gpt-5.3-codex-spark-preview', not
'gpt-5.3-codex-spark'. Rename model files in both openai and opencode
providers so the generated model ID matches the actual API.
2026-03-14 03:59:03 +09:00
Aiden Cline
a79585fa83
Merge pull request #1163 from micuintus/feature/Kimi2.5-fast
...
feat(nebius): add Kimi-K2.5-fast model
2026-03-13 13:14:38 -05:00
Aiden Cline
00801f74f2
Merge pull request #1164 from butyess/dev
...
Openrouter models: gemini 3.1 flash lite preview, grok 4.20 beta models.
2026-03-13 13:14:22 -05:00
Aiden Cline
185f6731ee
Merge pull request #1162 from dpuyosa/feature/venice-grok-4-20-beta
...
Venice: Add Grok 4.20 Beta models
2026-03-13 12:53:28 -05:00
Aiden Cline
d291b0575c
Merge pull request #1167 from sylviezhang37/update-vercel-models-20260313-1639
...
Update Vercel models
2026-03-13 12:53:11 -05:00
Mason Daugherty
382d9f3e7d
Update Claude 4.6 context/pricing
2026-03-13 13:53:04 -04:00
Aiden Cline
e64f5fe963
Merge pull request #1168 from mdrxy/mdrxy/update-baseten
...
Update Baseten models
2026-03-13 12:51:56 -05:00
Mason Daugherty
ea57ddfe7e
Update Baseten models
2026-03-13 13:48:41 -04:00
github-actions[bot]
29463d7fa8
chore(vercel): update Vercel model definitions
...
Auto-generated by weekly workflow from Vercel AI Gateway API.
Co-Authored-By: github-actions[bot] <41898282+github-actions[bot]@users.noreply.github.com>
2026-03-13 16:39:30 +00:00
Jack
bcc8db49ee
Merge pull request #1165 from anomalyco/chore/openrouter-alpha-reasoning-details-20260313
...
feat(openrouter): add interleaved reasoning details for alpha models
2026-03-13 22:25:25 +08:00
Jack
c8521d70f3
feat(openrouter): add interleaved reasoning details for alpha models
2026-03-13 22:20:54 +08:00
Federico Masi
490cd249e4
Openrouter models: gemini 3.1 flash lite preview, grok 4.20 beta models.
2026-03-13 15:11:12 +01:00
Michael Voigt
dbc636f5f3
feat(nebius): add Kimi-K2.5-fast model
2026-03-13 12:35:13 +01:00
Michael Voigt
9a32f671a1
fix(nebius): lowercase model ID for Nemotron-3-Super-120B-A12B
...
The filename must match the API casing (lowercase) to avoid 'model does not exist' errors.
2026-03-13 12:35:08 +01:00
dpuyosa
856d925eda
[venice] Add Grok 4.20 Beta models
...
- Add Grok 4.20 Beta model configuration (2M context, 128K output)
- Add Grok 4.20 Multi-Agent Beta model configuration
2026-03-13 10:48:38 +01:00
Aiden Cline
066a425917
Merge pull request #1158 from micuintus/feature/Nebius_Nemotron-3-Super-120b-a12b
...
feat(nebius): Add support for Nemotron-3-Super-120B-A12B
2026-03-12 22:20:20 -05:00
Aiden Cline
6df7f20cdc
Merge pull request #1156 from dsingal0/dev
...
added nemotron super on baseten
2026-03-12 22:20:06 -05:00
Aiden Cline
78bb47b90e
Merge pull request #1151 from dacbd/dacbd
...
fix(wandb): update models
2026-03-12 22:19:43 -05:00
Aiden Cline
c121d86419
Merge pull request #1160 from kreatoo/dev
...
feat: add zai-org/glm-4.7 and zai-org/glm-4.7-flash to NanoGPT
2026-03-12 22:11:18 -05:00
Aiden Cline
ab148eeb14
Merge pull request #1161 from Grin1024/dev
...
Add Claude Opus 4.6 and Sonnet 4.6 models to RequestY provider
2026-03-12 22:11:07 -05:00
lihui
49d196d326
Add Claude Opus 4.6 and Sonnet 4.6 models to RequestY provider
2026-03-13 09:00:54 +08:00
Kreato
8899b390ef
feat: add zai-org/glm-4.7 and zai-org/glm-4.7-flash to NanoGPT
2026-03-13 00:27:09 +03:00
Michael Voigt
5217f62ddf
fix(nebius): Follow context updates for Kimi 2.5 and GLM-5
2026-03-12 20:22:48 +01:00
Michael Voigt
55eaff9af1
feat(nebius): Add support for Nemotron-3-Super-120B-A12B
2026-03-12 20:22:21 +01:00
Dhruv Singal
7557c06ac0
update output length
2026-03-12 09:41:25 -07:00
Dhruv Singal
e85d820121
fix input output
2026-03-12 08:29:01 -07:00
Dhruv Singal
499d3a39ef
remove cache pricing
2026-03-12 08:21:22 -07:00
Dhruv Singal
b9b38d6e33
added nemotron super on baseten
2026-03-12 08:18:46 -07:00
Aiden Cline
ca24ac14fa
Merge pull request #1153 from dpuyosa/dev
...
Venice: Update model output token limits
2026-03-12 10:08:46 -05:00
Aiden Cline
822546fc67
Merge pull request #1155 from spiffytech/dev
...
Add Ollama Cloud support for Nemotron 3 Super
2026-03-12 10:08:31 -05:00
Aiden Cline
4555195b71
Merge pull request #1152 from v1gnesh/dev
...
Update grok-4.20 model defs
2026-03-12 10:08:15 -05:00
spiffytech
5eae8effc6
Added Ollama Cloud support for Nemotron 3 Super
2026-03-12 09:28:47 -04:00
dpuyosa
c1801aef87
[venice] Normalize model output token limits
...
- Update output limits to standard values across all models
2026-03-12 10:08:39 +01:00
v1gnesh
5e6464b272
Update grok-4.20-beta-reasoning
2026-03-12 10:27:40 +05:30
v1gnesh
e1a4f23332
Update grok-4.20-beta-non-reasoning
2026-03-12 10:26:03 +05:30
v1gnesh
753e1f9f0c
grok-multi-agent-beta update
2026-03-12 10:23:57 +05:30
Daniel Barnes
123ecd2ba5
docs url
2026-03-12 13:27:56 +09:00
Daniel Barnes
f15cda9fcb
remove old
2026-03-12 13:26:08 +09:00
Daniel Barnes
0205debbd3
fix values
2026-03-12 13:22:29 +09:00
Daniel Barnes
0059766509
number formating
2026-03-12 13:17:22 +09:00
Daniel Barnes
be81b02916
additional model files
2026-03-12 13:02:17 +09:00
Daniel Barnes
2dab141166
initial script & model updates
2026-03-12 13:01:35 +09:00
Aiden Cline
45aa49af25
tweak: azure kimi k2.5
2026-03-11 22:35:20 -05:00
Aiden Cline
781fad3ad4
Merge pull request #1150 from cau1k/5.4-family
...
feat(azure): add 5.4/pro families
2026-03-11 22:14:08 -05:00
cau1k
99d2ffcfdd
feat(azure): add 5.4/pro families
2026-03-11 20:59:11 -04:00
Aiden Cline
381d7cc19d
Merge pull request #1149 from ariane-emory/fear/add-march-or-stealth-models
...
Add OpenRouter stealth models: Hunter Alpha and Healer Alpha
2026-03-11 18:07:50 -05:00
Ariane Emory
7482e22458
Fix family field to use 'alpha' for stealth models
2026-03-11 18:49:32 -04:00
Ariane Emory
f5e6a402e6
Add OpenRouter stealth models: Hunter Alpha and Healer Alpha
2026-03-11 18:41:58 -04:00
Aiden Cline
9265852852
tweak: adjust some gh limits to align better w/ api
2026-03-11 15:23:44 -05:00
Aiden Cline
dc98a32996
Merge pull request #1018 from Sewer56/add-synthetic-missing-models
...
Update synthetic.new models: promote MiniMax-M2.5, add GLM-4.7-Flash
2026-03-11 14:55:50 -05:00
Aiden Cline
56c39ae0f6
Merge pull request #1140 from sk0x0y/feature/nanogpt-thudm-id-fixes
...
fix(nano-gpt): rename THUDM 2 ids to canonical THUDM ids
2026-03-11 14:55:07 -05:00
Aiden Cline
b1f43a7595
Merge pull request #1147 from msadiks/fix/alibaba-coding-minimax
...
fix: alibaba-coding-plan MiniMax-M2.5 context window
2026-03-11 14:54:37 -05:00
Matt Cowger
fed8bcae19
Merge branch 'dev' into mcowger/correct-gemini-flash-lite-pricing
2026-03-11 12:23:42 -07:00
sk0x0y
fb95150d02
fix(nano-gpt): rename VongolaChouko model id
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-12 04:20:39 +09:00
sk0x0y
a7c9a240b4
fix(nano-gpt): rename Steelskull model ids
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-12 04:20:39 +09:00
sk0x0y
6432a4a3e6
fix(nano-gpt): rename Sao10K model ids
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-12 04:20:38 +09:00
sk0x0y
f2e4a249fe
fix(nano-gpt): rename NeverSleep model ids
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-12 04:20:38 +09:00
sk0x0y
8667a6eed8
fix(nano-gpt): rename MarinaraSpaghetti model id
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-12 04:19:56 +09:00
sk0x0y
429554397a
fix(nano-gpt): rename LatitudeGames model id
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-12 04:19:56 +09:00
sk0x0y
a64e6ad0ac
fix(nano-gpt): rename LLM360 model id
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-12 04:19:56 +09:00
sk0x0y
d68d79888c
fix(nano-gpt): rename Infermatic model id
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-12 04:19:56 +09:00
sk0x0y
6c52905c6a
fix(nano-gpt): rename Gryphe model id
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-12 04:19:55 +09:00
sk0x0y
62410b8f26
fix(nano-gpt): rename GalrionSoftworks model id
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-12 04:19:55 +09:00
sk0x0y
50ce68ccab
fix(nano-gpt): rename Envoid model ids
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-12 04:19:20 +09:00
sk0x0y
d1c6a6b873
fix(nano-gpt): rename EVA-UNIT-01 model ids
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-12 04:19:20 +09:00
Frank
7193b068a5
update zen models
2026-03-11 13:52:50 -04:00
Sadik
79a8a06bd7
fix MiniMax-M2.5 context window
2026-03-11 20:50:33 +03:00
Aiden Cline
b60c03e11c
Merge pull request #1139 from zainhas/dev
...
[Together AI] add prompt caching pricing for MiniMax m2.5
2026-03-11 12:31:56 -05:00
Aiden Cline
15cf98d57b
Merge pull request #1146 from gotjoshua/patch-1
...
Rename step-3-5-flash.toml to step-3.5-flash.toml
2026-03-11 12:31:39 -05:00
Aiden Cline
b2ee6c407b
Merge pull request #1144 from micuintus/feature/update-nebius-changes
...
Feat: update Nebius changes
2026-03-11 12:31:29 -05:00
gotjoshua
96a14a06e7
Rename step-3-5-flash.toml to step-3.5-flash.toml
...
on nvidia it is 3.5 not 3-5
2026-03-11 11:41:36 +00:00
Michael Voigt
adc358606d
fix(nebius): update model context limits per API
2026-03-11 11:33:14 +01:00
Michael Voigt
63d52adf6f
feat(nebius): add GLM-5 model
2026-03-11 11:33:14 +01:00
sk0x0y
9a31387766
fix(nano-gpt): rename Salesforce model id
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-11 18:17:41 +09:00
sk0x0y
735157b837
fix(nano-gpt): rename ReadyArt model ids
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-11 18:17:41 +09:00
sk0x0y
d75b46fb37
fix(nano-gpt): rename Doctor-Shotgun model id
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-11 18:17:41 +09:00
sk0x0y
cc555f8482
fix(nano-gpt): rename CrucibleLab model id
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-11 18:17:13 +09:00
sk0x0y
7fbbcf2b49
fix(nano-gpt): rename MiniMaxAI model id
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-11 16:04:28 +09:00
sk0x0y
14c8ec8ca5
fix(nano-gpt): rename Tongyi-Zhiwen model id
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-11 16:04:28 +09:00
sk0x0y
72568bbdb3
fix(nano-gpt): rename Alibaba-NLP model id
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-11 16:03:57 +09:00
sk0x0y
c2225b715f
fix(nano-gpt): rename THUDM GLM-Z1 rumination id
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-11 15:49:20 +09:00
sk0x0y
7ce25e3742
fix(nano-gpt): rename THUDM GLM-Z1 model ids
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-11 15:49:20 +09:00
sk0x0y
427868604b
fix(nano-gpt): rename THUDM GLM-4 model ids
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-11 15:49:20 +09:00
Zain Hasan
247cd801a8
add prompt caching pricing for MiniMax m2.5
2026-03-10 22:54:42 -07:00
Aiden Cline
1aa2ee22b1
Merge pull request #1134 from sk0x0y/feature/nanogpt-catalog-fixes
...
fix(nano-gpt): correct TEE path ids and add missing canonical entries
2026-03-10 22:02:52 -05:00
Aiden Cline
0f57233eff
Merge pull request #1105 from sylviezhang37/add-vercel-input-context-and-new-models
...
feat(vercel): add input context calculation + new models
2026-03-10 22:01:52 -05:00
Aiden Cline
73a78eebfc
Merge pull request #1138 from mugnimaestra/feat/add-glm-5-turbo-chutes
...
feat: add GLM-5-Turbo to Chutes provider listings
2026-03-10 22:01:08 -05:00
Sylvie Zhang
3a6789b819
Merge branch 'dev' into add-vercel-input-context-and-new-models
2026-03-10 17:44:14 -07:00
Sylvie Zhang
f7c505e140
remove context from gemini models
2026-03-10 17:43:08 -07:00
Sylvie Zhang
6bb36806d6
only calc input context for openai models
2026-03-10 17:40:46 -07:00
Sylvie Zhang
20a404eb88
revert non openai changes
2026-03-10 17:38:46 -07:00
Muhammad Mugni Hadi
65ecb5cd4a
feat: add GLM-5-Turbo to Chutes provider listings
2026-03-11 05:26:11 +07:00
Matt Cowger
56062a9129
Fix incorrect pricing
2026-03-10 14:57:44 -07:00
Aiden Cline
d3d9c580d4
Merge pull request #1135 from gitpush-gitpaid/fix/gpt-5-4-pdf-input-modalities
...
Added PDF to input modalities for GPT-5.4
2026-03-10 13:53:42 -05:00
gitpush-gitpaid
ef98d8a9cb
Updated GPT-5.4 PDF input modalities
2026-03-10 13:59:29 -04:00
sk0x0y
9d17752b88
fix(nano-gpt): add missing GLM 5 thinking model
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-11 00:56:22 +09:00
sk0x0y
b5a838fe8b
fix(nano-gpt): add missing TEE qwen3.5 model
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-11 00:56:22 +09:00
sk0x0y
aa1ac39ee6
fix(nano-gpt): rename TEE gemma and minimax ids
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-11 00:56:22 +09:00
sk0x0y
4bc17ccf96
fix(nano-gpt): rename TEE oss and llama ids
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-11 00:56:22 +09:00
sk0x0y
08c1899bfe
fix(nano-gpt): rename TEE deepseek model ids
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-11 00:56:02 +09:00
sk0x0y
ad50e4a5ed
fix(nano-gpt): rename TEE qwen model ids
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-11 00:56:02 +09:00
sk0x0y
730915a123
fix(nano-gpt): rename TEE kimi model ids
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-11 00:56:02 +09:00
sk0x0y
6f12d18cb8
fix(nano-gpt): rename TEE glm model ids
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-11 00:56:02 +09:00
Aiden Cline
bd8774db99
Merge pull request #1132 from sk0x0y/feature/nanogpt-model-sync
...
feat(nano-gpt): add text and image models
2026-03-10 10:31:50 -05:00
Aiden Cline
88fbea52a4
Merge pull request #1133 from anomalyco/fix-model
...
fix: bedrock devstral
2026-03-10 10:31:08 -05:00
Aiden Cline
70e5d9b34b
fix: bedrock devstral
2026-03-10 10:30:20 -05:00
Aiden Cline
edb6ef0d71
Merge pull request #1129 from Grin1024/dev
...
feat: add GPT-5 series models to requesty provider
2026-03-10 10:29:30 -05:00
Aiden Cline
df1280ed8b
add families to some bedrock models
2026-03-10 10:12:13 -05:00
sk0x0y
898b3c18b7
feat(nano-gpt): add image models
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-10 22:13:48 +09:00
sk0x0y
6316e543ef
feat(nano-gpt): add text models
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-10 22:13:48 +09:00
Aiden Cline
64d9a97f9d
Merge pull request #1128 from JWahle/dev
...
chore: Updated abacus model definitions
2026-03-10 07:40:01 -05:00
Aiden Cline
c62a2a3fc1
Merge pull request #1130 from Mingholy/fix/alibaba-coding-plan-model-limits
...
fix: update model limits for alibaba-coding-plan providers
2026-03-10 07:39:48 -05:00
Aiden Cline
cc1937a177
Merge pull request #1131 from janszypulski/cloudferro-sherlock-fix-minimax-model-id
...
fix minimax-m2.5 model id - wrong file path
2026-03-10 07:39:34 -05:00
Jan Szypulski
9d3a88863d
fix minimax-m2.5 model id - wrong file path
2026-03-10 11:21:55 +01:00
mingholy.lmh
b9123e26e0
fix: update model limits for alibaba-coding-plan providers
...
- Add MiniMax-M2.5 to alibaba-coding-plan-cn
- Update qwen3-max output limit (65536 -> 32768)
- Update qwen3-coder-plus context limit (1048576 -> 1000000)
- Update MiniMax-M2.5 limits per ref.json (context: 196608, output: 24576)
Co-authored-by: Qwen-Coder <qwen-coder@alibabacloud.com >
2026-03-10 15:45:01 +08:00
lihui
933e450104
feat: add GPT-5 series models to requesty provider
...
Add missing OpenAI GPT-5 series models to requesty provider:
- GPT-5 Chat, Codex, Image, Pro
- GPT-5.1 Chat, Codex, Codex-Max, Codex-Mini
- GPT-5.2 Chat, Codex, Pro
- GPT-5.3 Codex
- GPT-5.4, GPT-5.4 Pro
2026-03-10 14:53:28 +08:00
JWahle
2ba6383e70
chore: Updated abacus model definitions
...
Added: gpt-5.4.toml
Removed: gemini-3-pro-preview.toml
2026-03-10 05:10:45 +01:00
Aiden Cline
65ed6ac5dd
Merge pull request #1126 from mcowger/feature/gemini-3.1-flash-lite-vercel
...
feat: add gemini-3.1-flash-lite-preview to vercel gateway provider
2026-03-09 21:43:33 -05:00
Aiden Cline
e7d04aec7a
Merge pull request #1127 from anomalyco/add-shape
...
feat: add 'shape' field to provider so models can specify if they use responses vs completions apis (use only if model only supports 1 of)
2026-03-09 21:43:04 -05:00
Aiden Cline
37fe334aed
feat: add 'shape' field to provider so models can specify if they use responses vs completions apis (use only if model only supports 1 of)
2026-03-09 21:42:23 -05:00
Matt Cowger
ce8fc9e4f0
feat: add gemini-3.1-flash-lite-preview to vercel gateway provider
2026-03-09 19:32:43 -07:00
Aiden Cline
be8eb8ba54
fix name
2026-03-09 20:01:02 -05:00
Aiden Cline
7c625b3b82
Merge pull request #945 from Daltonganger/feat/nano-gpt-sync-models-api
...
sync nano-gpt models with live API catalog
2026-03-09 20:00:06 -05:00
Aiden Cline
33700d27dc
Merge pull request #1032 from propilideno/feature/new_gpt_5.3_codex_and_missing_structured_output_attr
...
Add gpt-5.3-codex (Azure) and fill missing structured output flags
2026-03-09 19:40:20 -05:00
Aiden Cline
e5c300a5e5
fix
2026-03-09 19:38:30 -05:00
Aiden Cline
e5e9175c5d
Merge branch 'dev' into feature/new_gpt_5.3_codex_and_missing_structured_output_attr
2026-03-09 19:37:40 -05:00
Aiden Cline
a9f79d6794
Merge pull request #1123 from dpuyosa/feature/venice-gpt54-multimodal
...
Venice: Add GPT-5.4 Pro and enable multimodal inputs for GPT-5.4 & Qwen3.5
2026-03-09 18:19:52 -05:00
Aiden Cline
fd4c4a8f28
Merge pull request #1038 from muldercw/add-clarifai-model-provider
...
Add Clarifai Model Provider
2026-03-09 18:19:14 -05:00
dpuyosa
f9b5385868
[venice] Add GPT-5.4 Pro and enable multimodal inputs
...
- Add GPT-5.4 Pro model
- Enable attachment/image input for GPT-5.4
- Enable attachment/image/video input for Qwen3.5 35B A3B
2026-03-09 22:52:54 +01:00
Aiden Cline
b2f7a72410
Merge pull request #1110 from fhennerkes/dev
...
poe: add GPT-5.4, GPT-5.4-Pro, and GPT-5.3-Instant models
2026-03-09 14:10:12 -05:00
Aiden Cline
7b5d9aa645
Merge pull request #1025 from liuchang-reolink/dev
...
add qwen3.5-397b-a17b and step-3-5-flash for nvidia
2026-03-09 14:05:08 -05:00
Aiden Cline
6e0040dbfd
Merge pull request #1089 from Krule/krule/update_gitlab_anthropic_context_size
...
feat(gitlab): update context limit to 1M for Claude Sonnet and Opus 4.6
2026-03-09 14:03:49 -05:00
Aiden Cline
943ad8481b
Merge pull request #1121 from illusion77/fix/chutes-mimo-v2-flash-context-16709
...
fix(chutes): correct MiMo-V2-Flash context window and capabilities
2026-03-09 14:02:51 -05:00
Aiden Cline
7f1b6fb0eb
Merge pull request #1122 from riccardogiorato/dev
...
remove deprecated kimi models from together.ai
2026-03-09 14:02:36 -05:00
Riccardo Giorato
23eff95e5d
remove deprecated kimi from together.ai
2026-03-09 17:30:40 +01:00
illusion77
ddbd396205
fix(chutes): correct MiMo-V2-Flash context window and capabilities
...
The chutes provider had incorrect metadata for MiMo-V2-Flash:
context 32K → 262K, output 8K → 32K, reasoning and tool_call enabled.
Fixes anomalyco/opencode#16709
2026-03-09 10:57:57 -05:00
Aiden Cline
f3ee1a530b
Merge pull request #1120 from stephenkuhn214/dev
...
Add Amazon-Bedrock Devstral 2 123B model
2026-03-09 09:35:46 -05:00
Aiden Cline
9c51b65440
Merge pull request #1119 from cgilly2fast/dev
...
fix(firmware): proper 5.3 codex model id
2026-03-09 09:30:35 -05:00
Frank
353aeb4998
update zen models
2026-03-09 10:08:55 -04:00
Frank
11991fecb5
update zen models
2026-03-09 10:03:13 -04:00
stephenkuhn214
1b4599773d
Create mistral.devstral-2-123b
2026-03-09 08:58:19 -04:00
Colby Gilbert
78e1a3b0c9
fix(firmware): proper 5.3 codex model id
2026-03-08 21:58:27 -07:00
Sewer56
7a02946620
Update synthetic models: promote MiniMax-M2.5, add GLM-4.7-Flash, remove deprecated Qwen3.5
2026-03-08 22:56:31 +00:00
Aiden Cline
44686797c8
Merge pull request #1118 from shelvick/add-azure-gpt-5.3-chat
...
Add GPT-5.3 Chat to Azure
2026-03-08 16:41:10 -05:00
Aiden Cline
065cec8431
fix: input limit for context
2026-03-08 16:40:38 -05:00
Scott Helvick
f491c2bec9
Add GPT-5.3 Chat to Azure
2026-03-08 21:20:27 +00:00
Aiden Cline
cf1ac3053f
Merge pull request #1081 from djmaze/fix/nebius-model-casing
...
fix(nebius): correct model ID casing to match Token Factory API
2026-03-08 14:26:52 -05:00
Aiden Cline
6be1e929fc
Merge pull request #1114 from v1gnesh/dev
...
add grok 4.2 experimentals
2026-03-08 14:25:12 -05:00
Aiden Cline
49524827e2
Merge pull request #1113 from shelvick/add-vertex-glm-5
...
Fix GLM-5 context window size on Google Vertex
2026-03-08 10:31:05 -05:00
Aiden Cline
5ab5d389fc
Merge pull request #1112 from cau1k/feat/az-5.4
...
feat(azure): add gpt-5.4/5.4-pro
2026-03-08 10:30:54 -05:00
Aiden Cline
d5367ed978
Merge pull request #1116 from xiaojiezj/xj_dev_0308
...
fix: Adjust the logo for ZenMux
2026-03-08 10:30:18 -05:00
Aiden Cline
5069faa25b
Merge pull request #1117 from kailiu42/feat/siliconflow-cn
...
feat(siliconflow-cn): add Qwen3.5 model family
2026-03-08 10:29:48 -05:00
Kai Liu
b279f33d9b
feat(siliconflow-cn): add Qwen3.5 model family
...
New models:
- Qwen/Qwen3.5-4B
- Qwen/Qwen3.5-9B
- Qwen/Qwen3.5-27B
- Qwen/Qwen3.5-35B-A3B
- Qwen/Qwen3.5-122B-A10B
- Qwen/Qwen3.5-397B-A17B
Signed-off-by: Kai Liu <kraml.liu@gmail.com >
2026-03-08 20:07:40 +08:00
xiaojie.zj
fe8249d706
fix: Adjust the logo
2026-03-08 16:36:01 +08:00
skywalker512
236af40da3
feat: add Tencent Coding Plan provider
...
Add support for Tencent Coding Plan with 8 models:
- Auto (tc-code-latest)
- Hunyuan 2.0 Instruct
- Hunyuan 2.0 Think
- Hunyuan-T1
- Hunyuan-TurboS
- MiniMax-M2.5
- Kimi-K2.5
- GLM-5
Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com >
2026-03-08 15:34:36 +08:00
v1gnesh
a24a23d57b
add grok 4.2 experimentals
2026-03-08 07:42:09 +05:30
Scott Helvick
7e4773d9b5
Fix GLM-5 context window size on Google Vertex
...
Correct the context limit from 204800 to 202752 tokens.
2026-03-07 22:20:56 +00:00
zero
9f937f3fc5
Merge branch 'dev' into feat/az-5.4
2026-03-07 17:20:29 -05:00
cau1k
8cbdbc1102
add day cutoff
2026-03-07 17:19:15 -05:00
cau1k
f1ca3b0015
feat(azure-cognitive-services): symlink 5.4/pro from azure provider
...
;
2026-03-07 17:02:26 -05:00
cau1k
35c757bad5
feat(azure): add 5.4/pro
2026-03-07 17:00:43 -05:00
fhennerkes
78781f3901
poe: add GPT-5.4, GPT-5.4-Pro, and GPT-5.3-Instant models
2026-03-07 13:58:35 -08:00
Sylvie Zhang
26465319d6
Merge branch 'dev' into add-vercel-input-context-and-new-models
2026-03-07 11:39:53 -08:00
fhennerkes
1371cbf9de
poe: add GPT-5.4, GPT-5.4-Pro, and GPT-5.3-Instant models
2026-03-07 09:48:17 -08:00
Aiden Cline
559ccd6966
Merge pull request #1024 from yinxulai/feat/qiniu-ai
...
feat(qiniu-ai): add new model configurations
2026-03-07 11:26:01 -06:00
Aiden Cline
2691cb4e8d
Merge pull request #1083 from samzong/feat/add-drun-provider
...
feat: add d.run(China) provider (OpenAI-compatible)
2026-03-07 11:25:12 -06:00
Aiden Cline
83ed1f0125
Merge pull request #1015 from RioPlay/dev
...
add: newer MiniMax, GLM, and Kimi models to DeepInfra
2026-03-07 11:24:53 -06:00
Aiden Cline
869f831466
Merge branch 'dev' into dev
2026-03-07 11:23:37 -06:00
Aiden Cline
5a673af2ae
Merge pull request #1061 from JonasGao/dev
...
Add Qwen3.5 Flash & GLM-5 & M2.5 models to alibaba-cn
2026-03-07 11:22:56 -06:00
Aiden Cline
9b8543a074
Add interleaved section to minimax-m2.5.toml
2026-03-07 11:21:54 -06:00
Aiden Cline
b3fb902331
Merge pull request #1030 from Mingholy/feat/alibaba-coding-plan-cn
...
feat(alibaba-coding-plan-cn): add Coding Plan provider for China region
2026-03-07 11:21:26 -06:00
Aiden Cline
adb0c0b305
Merge pull request #1062 from viitana/bump-deepseek-details
...
feat: [deepseek]: update official DeepSeek model details
2026-03-07 11:21:21 -06:00
Aiden Cline
ed01410d82
Merge pull request #1088 from mcowger/feature/gemini-3.1-flash-lite
...
feat: add gemini-3.1-flash-lite-preview model
2026-03-07 11:13:45 -06:00
Aiden Cline
7e23b780cc
Merge pull request #1077 from evroc-oss/evroc/correct-model-config
...
fix(evroc): correct model config
2026-03-07 11:13:15 -06:00
Aiden Cline
c46b652c8e
Merge pull request #1076 from jerome-benoit/feat/add-sonar-deep-research-sap-ai-core
...
feat(sap-ai-core): add Perplexity Sonar Deep Research model
2026-03-07 11:11:29 -06:00
Aiden Cline
ddb74e9b09
Merge pull request #1063 from dpuyosa/fix/models-pricing-limits-update
...
Venice: Update model pricing and limits
2026-03-07 11:10:37 -06:00
Aiden Cline
face36ecb8
Merge pull request #1064 from BlockListed/fix-cortecs-models
...
Fix Cortecs models
2026-03-07 11:10:03 -06:00
Aiden Cline
6f170651b3
Merge pull request #1075 from Track07-cda/alibaba-cn-third-party-models
...
Add third party providers' models to alibaba-cn provider
2026-03-07 11:09:40 -06:00
Aiden Cline
47dbe45dd5
Merge pull request #1066 from dpuyosa/feat/add-qwen3-5-35b-a3b
...
Venice: Add Qwen 3.5 35B A3B model
2026-03-07 11:09:10 -06:00
Aiden Cline
0ee43b64b3
Merge branch 'dev' into alibaba-cn-third-party-models
2026-03-07 11:08:37 -06:00
Aiden Cline
6130a1f74e
Merge pull request #1068 from MauroDruwel/dev
...
NVIDIA: Add MiniMax M2.5 model and remove MiniMax M2
2026-03-07 11:07:17 -06:00
Aiden Cline
c0c82a5f04
Merge pull request #1072 from sylviezhang37/update-vercel-models-20260302-1656
...
Update Vercel models
2026-03-07 11:06:17 -06:00
Aiden Cline
b0ba8b14d5
Merge pull request #1092 from janszypulski/cloudferro-sherlock-add-minimax-2.5
...
add MiniMaxAI/MiniMax-M2.5 to CloudFerro Sherlock
2026-03-07 11:02:11 -06:00
Aiden Cline
4780f9ddc1
Merge pull request #1109 from dinhkim/feat/add-cf-glm-4.7-flash
...
feat: add GLM-4.7-Flash to the Cloudflare Workers AI provider
2026-03-07 11:01:50 -06:00
Aiden Cline
27e02de632
Merge pull request #1078 from SomeoneWithOptions/dev
...
add gpt 5.3 codex for openrouter and Mercury models
2026-03-07 11:01:41 -06:00
Aiden Cline
f22c827045
Merge branch 'dev' into dev
2026-03-07 11:01:17 -06:00
Aiden Cline
cfc4585ed7
Merge pull request #1107 from Rinuuri/deepinfra-glm5
...
Add deepinfra GLM-5 model
2026-03-07 10:59:16 -06:00
Aiden Cline
fa07bc2088
Merge pull request #1039 from rholak/add-abacus-models
...
Add sonnet 4.6 and opus 4.6 to abacus model list
2026-03-07 10:59:00 -06:00
Aiden Cline
497b1daaf2
Merge pull request #1103 from dpuyosa/feat/venice-add-gpt-models
...
Venice: Add OpenAI GPT-4o, GPT-4o Mini, GPT-5.4 models
2026-03-07 10:58:44 -06:00
Aiden Cline
442afa8c7e
Merge pull request #1060 from yanismiraoui/inception/mercury2
...
Add Inception Mercury 2 and Mercury Edit models
2026-03-07 10:57:38 -06:00
Aiden Cline
4bd0c387fe
Merge pull request #1044 from shrwnsan/feat/openrouter-routers
...
feat(openrouter/free): add free router
2026-03-07 10:57:24 -06:00
Aiden Cline
b8c0c1d3a1
Merge pull request #1053 from laiiihz/update-xiaomi-models
...
Update Xiaomi models metadata
2026-03-07 10:57:17 -06:00
Aiden Cline
5c6c3e5a32
Merge pull request #1055 from shantanugoel/gemini-3.1-flash-image-preview
...
Add Gemini 3.1 Flash Image Preview
2026-03-07 10:57:06 -06:00
Aiden Cline
53d3cca3a0
Merge pull request #1052 from spiffytech/dev
...
Improve Ollama Cloud generator. Remove Gemini 3 Pro from Ollama Cloud.
2026-03-07 10:56:45 -06:00
Aiden Cline
105970c173
Merge pull request #1049 from heimoshuiyu/fix/glm-5-open-weights
...
fix: mark GLM-5 as open weights
2026-03-07 10:56:31 -06:00
Aiden Cline
788ee04034
Merge pull request #1045 from xinrui-z/aihubmix-add-models
...
aihubmix add models
2026-03-07 10:56:03 -06:00
Aiden Cline
6626db4044
Merge pull request #1098 from JWahle/dev
...
chore: updated abacus model definitions
2026-03-07 10:55:31 -06:00
Aiden Cline
8902640664
Merge pull request #1023 from PandaSt0rm/add-alibaba-coding-plan
...
Add Alibaba Coding Plan provider and model configs
2026-03-07 10:53:34 -06:00
Kim Truong
cab247ddf8
update context to match Cloudflare doc
2026-03-07 23:50:06 +07:00
Kim Truong
c1a42fa0a0
feat: add GLM-4.7-Flash mode in Cloudflare Workers AI provider
2026-03-07 23:45:52 +07:00
Aiden Cline
35023bba5a
Merge pull request #1001 from ItsWendell/feat/bedrock-bearer-token
...
Add AWS_BEARER_TOKEN_BEDROCK to Amazon Bedrock provider env
2026-03-07 09:52:21 -06:00
Aiden Cline
604e49792b
Merge pull request #1002 from DEAN-Cherry/feat/add-minimax-m2.5
...
models: alibaba-cn: add MiniMax-M2.5
2026-03-07 09:51:36 -06:00
Aiden Cline
0ca77b0cda
Merge branch 'dev' into dev
2026-03-07 09:50:26 -06:00
Aiden Cline
ea9505a40f
Merge pull request #1004 from BlockListed/cortecs-models
...
Add Cortecs AI models
2026-03-07 09:50:07 -06:00
Aiden Cline
990b8d7308
Merge pull request #1005 from cgilly2fast/dev
...
feat(firmware): gemini 3.1 pro, sonnet reasoning
2026-03-07 09:49:54 -06:00
Aiden Cline
ec173e86d4
Merge pull request #996 from fhennerkes/dev
...
poe: add Gemini-3.1-Pro, GPT-5.3-Codex and Gemini 3.1 Flash Lite
2026-03-07 09:47:52 -06:00
Aiden Cline
4a6e92a7c9
Merge pull request #997 from xiaojiezj/zenmux_dev_0221
...
feat: add Gemini 3.1 Pro Preview for ZenMux provider
2026-03-07 09:47:37 -06:00
Aiden Cline
f0f686bdf5
Merge pull request #999 from mikalsande/mistral_latest
...
Append (latest) to Mistral models that refer to the latest version.
2026-03-07 09:46:40 -06:00
Aiden Cline
35ff0c2629
Merge pull request #995 from Phoen1xCode/dev
...
fix(zenmux:minimax): remove duplicated prefix & feat(zenmux:openai): add GPT-5.2-Pro model
2026-03-07 09:45:07 -06:00
Aiden Cline
f99e9e89df
Merge pull request #1090 from litvix-whale/feat/add-minimax-m2-5
...
feat(provider): add MiniMax M2.5 for DeepInfra
2026-03-07 09:41:28 -06:00
Armin Pašalić
09722ac264
Merge branch 'anomalyco:dev' into krule/update_gitlab_anthropic_context_size
2026-03-07 13:17:10 +01:00
Rinuuri
fa67d00aeb
Update GLM-5.toml
2026-03-06 21:23:42 +00:00
Rinuuri
ddd2dd73ed
Adding deepinfra GLM-5
2026-03-07 00:03:29 +03:00
fhennerkes
d7929fd00b
Merge branch 'anomalyco:dev' into dev
2026-03-06 12:00:20 -08:00
Frank
06e7d4db42
Merge pull request #1014 from NachoFLizaur/fix/bedrock-opus-4-6-context-window
...
fix(amazon-bedrock): correct Claude Opus 4.6 context window from 1M to 200K
2026-03-06 11:25:37 -05:00
Sylvie Zhang
7a11ef241d
update more models
2026-03-06 08:24:49 -08:00
Sylvie Zhang
145862315d
add input calculation + new models
2026-03-06 08:07:11 -08:00
dpuyosa
d871710ba4
[venice] Add OpenAI GPT-4o, GPT-4o Mini, GPT-5.4 models
...
- Add gpt-4o-2024-11-20 model configuration
- Add gpt-4o-mini-2024-07-18 model configuration
- Add gpt-5.4 model configuration with reasoning capability
2026-03-06 09:53:06 +01:00
Colby Gilbert
16486087c6
Merge branch 'anomalyco:dev' into dev
2026-03-05 21:38:25 -08:00
Frank
2939af9330
Merge pull request #1100 from sachnun/feat/github-copilot-gpt-5-4
...
feat(provider): add gpt-5.4 for GitHub Copilot
2026-03-05 23:33:57 -05:00
sachnun
7c68dab3bb
feat(provider): add gpt-5.4 for GitHub Copilot
2026-03-06 11:18:11 +07:00
Mike Soylu
caceb0b310
openrouter openai models ( #1099 )
2026-03-05 22:26:58 -05:00
Frank
7a0d3be1e7
Update zen models
2026-03-05 18:55:49 -05:00
ShivamB25
e11ad7c01a
feat(openai): add GPT-5.4 and GPT-5.4 Pro model specs ( #1095 )
2026-03-05 18:50:22 -05:00
Matt Silverlock
d30fa82e4c
Cloudflare: add gpt-5.4.toml ( #1096 )
2026-03-05 18:50:10 -05:00
Rishi Vhavle
771102a960
feat: add gpt-5.3-codex to github-copilot provider ( #1097 )
2026-03-05 18:49:56 -05:00
JWahle
30f98b15ef
chore: updated abacus model definitions
...
Added: GPT-5 Codex, GPT-5.1/5.2/5.3 Codex, GPT-5.3 Chat, Gemini 3.1 Flash Lite/Pro Preview, Claude Opus/Sonnet 4.6, Kimi K2.5, GLM-5
Removed: Gemini 2.0 Flash 001, Gemini 2.0 Pro Exp, Meta-Llama 3.1 70B Instruct
Updated pricing: DeepSeek V3.1, GLM-4.7, GPT-5.2 Chat Latest, o3-pro, Route LLM
2026-03-06 00:46:19 +01:00
Colby Gilbert
6f7ab479fb
feat(firmware): gpt 5.4
2026-03-05 13:23:08 -08:00
Colby Gilbert
a4efbcd5ce
Merge branch 'anomalyco:dev' into dev
2026-03-05 13:15:49 -08:00
Frank
bcbfba03bd
update zen models
2026-03-05 15:51:33 -05:00
Frank
bdb5dac941
update zen models
2026-03-05 15:50:03 -05:00
Frank
1538bdcedb
update zen models
2026-03-05 13:31:27 -05:00
SomeoneWithOptions
e5211f3105
add inception mercury models for openrouter
2026-03-05 12:13:52 -05:00
Andres Castellanos
4a2209dbd4
Merge branch 'anomalyco:dev' into dev
2026-03-05 11:51:38 -05:00
Jan Szypulski
900014fe52
add MiniMax-M2.5
2026-03-05 14:58:19 +01:00
Kyrylo Lytvishko
5bbaf3c3f2
feat(provider): add MiniMax M2.5 for DeepInfra
2026-03-05 14:04:11 +02:00
Armin Pasalic
7dd0a26ff4
feat(gitlab): update context limit to 1M for Sonnet and Opus 4.6
2026-03-05 12:01:02 +01:00
Matt Cowger
3b7e0f02f1
feat: add gemini-3.1-flash-lite-preview model
2026-03-04 13:19:21 -08:00
samzong
e67f921ea3
feat: add official d.run logo
2026-03-04 13:42:01 +08:00
samzong
f5411eeeda
feat: add D.Run (China) provider with minimax-m25, deepseek-r1, deepseek-v3
2026-03-04 13:33:04 +08:00
Frank
0d83ab8909
Merge pull request #1082 from kesku/kesku/add-ppl-agent-api
...
Add Perplexity Agent API provider
2026-03-03 23:03:49 -05:00
Kesku
26a629debc
add models
2026-03-03 23:19:01 +00:00
Kesku
4b3319561b
set up provider
2026-03-03 23:10:43 +00:00
Ubuntu
b89ce0d986
fix(nebius): correct model ID casing to match Token Factory API
...
Fix lowercase model ID bug that caused "The model does not exist" errors.
- qwen/ → Qwen/ directory
- Fixed model file casing to match API exactly across all providers
2026-03-03 22:11:57 +00:00
fhennerkes
1c01f8172b
poe: add Gemini-3.1-Flash-Lite and update gpt-4o-mini context
...
Add new Gemini 3.1 Flash Lite model
Update gpt-4o-mini context window: 128K → 124,096
2026-03-03 11:24:20 -08:00
SomeoneWithOptions
d76040c514
add gpt 5.3 codex for openrouter
2026-03-03 12:53:39 -05:00
Simon Rygård
feffa8119f
fix(evroc): correct modality config
2026-03-03 16:52:07 +01:00
Simon Rygård
59c6e5df62
fix(evroc): correct tool call config
2026-03-03 16:51:47 +01:00
Jérôme Benoit
fe2204d42c
feat(sap-ai-core): add Perplexity Sonar Deep Research model
2026-03-03 14:56:18 +01:00
Track07-cda
07cc5335ac
Add third party providers' models to alibaba-cn provider
...
- Add `MiniMax/MiniMax-M2.5` and `kimi/kimi-k2.5` to the `alibaba-cn`
provider.
- Update `kimi-k2.5` to include video modality and adjust release/update
dates.
- Add several `siliconflow/deepseek` models to the `alibaba-cn`
provider.
2026-03-03 16:46:32 +08:00
github-actions[bot]
fefbb90a29
chore(vercel): update Vercel model definitions
...
Auto-generated by weekly workflow from Vercel AI Gateway API.
Co-Authored-By: github-actions[bot] <41898282+github-actions[bot]@users.noreply.github.com>
2026-03-02 16:56:48 +00:00
Frank
fec48b83d3
update zen models
2026-03-01 13:23:08 -05:00
Mauro Druwel
c30bbe7718
Add knowledge
2026-03-01 08:53:22 +01:00
Mauro Druwel
1fba668f0f
Add minimax-m2.5 to nvidia-nim and remove deprecated minimax-m2 from nvidia-nim
2026-03-01 08:52:33 +01:00
Aiden Cline
33ec088bda
Merge pull request #1008 from friendliai/feat/friendli-minimax-m2.5
...
add friendli minimax m2.5 model config
2026-03-01 07:54:08 +05:00
Aiden Cline
add7f9a914
Merge pull request #1065 from friendliai/minpeter/remove-exaone-models
...
Remove all EXAONE models
2026-03-01 07:53:42 +05:00
dpuyosa
369fa2de6d
[venice] Add Qwen 3.5 35B A3B model
...
- Add new model configuration for Qwen 3.5 35B A3B
- Includes cost, limits, and capabilities (reasoning, tool_call, structured_output)
2026-02-28 21:07:28 +01:00
minpeter
c8732e7e74
Remove all EXAONE models
...
Remove LGAI-EXAONE model definitions (EXAONE-4.0.1-32B, K-EXAONE-236B-A23B)
and related family references from core packages.
Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com >
2026-03-01 04:48:08 +09:00
Jonas
f00f9f3c11
Add Qwen3.5 Flash & GLM-5 & M2.5 models to alibaba-cn provider
2026-02-28 23:34:06 +08:00
BlockListed
c87ca238de
fix cortecs models
...
the should have periods not a p as a decimal separator
2026-02-28 15:16:59 +01:00
BlockListed
102f55aeef
add glm 4.7 flash model to cortecs
2026-02-28 15:13:53 +01:00
BlockListed
2767754a02
add kimi K2.5 model to cortecs
2026-02-28 15:13:53 +01:00
dpuyosa
4289b59a04
[models] Update model pricing and limits
...
- Update Claude Sonnet 4-6 pricing and output limit
- Update Grok 41 Fast pricing and context/limits
2026-02-28 14:24:16 +01:00
Atte Viitanen
b986313f42
feat: [deepseek]: update official deepseek model details
2026-02-28 13:22:26 +02:00
yanismiraoui
28cfd4cab6
naming mercury 2 and mercury edit for inception provider
2026-02-27 17:45:01 -08:00
yanismiraoui
b93e62fc3a
Add Inception Mercury 2 and Mercury Edit models
2026-02-27 17:41:00 -08:00
Aiden Cline
e23b5ab010
Merge pull request #1026 from ryot/venice
...
Venice: Add GPT-5.3 Codex
2026-02-28 06:30:13 +05:00
Aiden Cline
b5f6024868
Merge pull request #1020 from jerome-benoit/feat/sap-ai-core-add-models
...
feat(sap-ai-core): Add GPT-4.1, Gemini 2.5 Flash Lite, Perplexity Sonar, and Claude 4.6 models
2026-02-28 06:29:28 +05:00
Aiden Cline
07db15e984
Merge pull request #1029 from dpuyosa/veniceScript
...
Venice: Remove interactive API key prompt & use new maxCompletionTokens field
2026-02-28 06:28:33 +05:00
Aiden Cline
74abf8851a
Merge pull request #1042 from SomeoneWithOptions/dev
...
add gemini 3.1 pro preview custom tools for openrouter
2026-02-28 06:27:56 +05:00
Aiden Cline
7e13ecdfd9
Merge pull request #1056 from xezpeleta/fix/azure-gpt-5-3-codex
...
fix(azure): add gpt-5.3-codex model
2026-02-28 06:27:34 +05:00
Aiden Cline
6ad2c28b2d
Merge pull request #1048 from dpuyosa/feat/add-venice-models
...
Venice: Add NVIDIA Nemotron 3 Nano and Qwen 3 Coder Turbo models
2026-02-28 06:27:20 +05:00
Frank
a124036692
update zen models
2026-02-27 16:16:37 -05:00
Xabi Ezpeleta
d37d362cc8
fix(azure): add gpt-5.3-codex model
2026-02-27 16:41:11 +01:00
Shantanu Goel
c387f94c8e
Add Gemini 3.1 Flash Image Preview
2026-02-27 20:03:41 +05:30
laiiihz
45457c34d8
update xiaomi models detail
2026-02-27 14:56:16 +08:00
spiffytech
44774ec3d6
Ollama Cloud removed support for Gemini 3 Pro
2026-02-26 17:18:34 -05:00
spiffytech
c8fdcf80dd
Updated Ollama Cloud generator to delete old models, only write out files if they changed
2026-02-26 17:18:33 -05:00
fhennerkes
9d33b6409c
Merge branch 'anomalyco:dev' into dev
2026-02-26 12:04:52 -08:00
Matt Silverlock
c76586a174
Cloudflare: add codex models to AI Gateway ( #1050 )
...
* add gpt-5.2-codex
* add gpt-5.3-codex
* Update gpt-5.2-codex.toml
* Update gpt-5.3-codex.toml
2026-02-26 14:41:12 -05:00
Jérôme Benoit
2a267614aa
feat(sap-ai-core): add Claude Opus 4.6 and Sonnet 4.6 models
2026-02-26 17:58:02 +01:00
PandaSt0rm
aac62378b2
Update MiniMax-M2.5 guidance per Alibaba docs
2026-02-26 17:20:58 +02:00
David Hill
56cc5f71bf
fix(ui): opencode zen logo update
2026-02-26 11:09:25 +00:00
David Hill
df2c87d32a
fix(ui): opencode go logo
2026-02-26 11:09:13 +00:00
heimoshuiyu
ff41c2b6c3
fix: mark GLM-5 as open weights
...
GLM-5 is an open-source model, but several provider config files
incorrectly had open_weights set to false. This commit corrects
all GLM-5 configurations to properly reflect its open-source status.
Affected providers:
- zhipuai
- zhipuai-coding-plan
- zai
- zai-coding-plan
- zenmux
- vercel
- siliconflow
- siliconflow-cn
- meganova
2026-02-26 18:43:28 +08:00
dpuyosa
1d137e2f1f
[venice] Add NVIDIA Nemotron 3 Nano and Qwen 3 Coder models
...
- Add NVIDIA Nemotron 3 Nano 30B A3B model configuration
- Add Qwen 3 Coder 480B A35B Instruct Turbo model configuration
2026-02-26 10:53:00 +01:00
dpuyosa
16720bcd1a
[venice] Use maxCompletionTokens for output limit
...
- Add optional maxCompletionTokens field to model spec schema
- Use maxCompletionTokens when calculating output token limit instead of checking existing limit
2026-02-26 10:24:43 +01:00
Xinrui
1feaf76749
aihubmix add models
2026-02-26 16:12:51 +08:00
shrwnsan
080ef5cc9e
fix(openrouter): remove auto router and add missing limit.input
...
- Remove auto router (cost varies, doesn't fit schema)
- Add limit.input = 200_000 to free.toml (schema requirement)
OpenRouter's auto router has 'pricing varied' - it charges based on the
routed model. This doesn't fit the numeric cost schema required by
models.dev, so we're removing it. The free router is retained as it
genuinely costs $0.
2026-02-26 14:40:38 +08:00
Ryo Tulman
f8121c8dc3
Update Venice GPT 5.3 Codex output limit
2026-02-26 00:32:13 -06:00
shrwnsan
d2d5c5a7cc
feat: add openrouter free and auto routers
2026-02-26 10:51:55 +08:00
SomeoneWithOptions
09d9e91d83
add gemini 3.1 pro preview custom tools for openrouter
2026-02-25 15:13:43 -05:00
Robert Holak
930d6a94b8
Add sonnet 4.6 and opus 4.6 to abacus model list
2026-02-25 12:31:06 -06:00
mulder
b9217aff8e
Add Clarifai Model Provider
...
Add Clarifai as a new provider with 11 models:
- GPT OSS 20B, GPT OSS 120B High Throughput
- Ministral 3 14B/3B Reasoning 2512
- Qwen3 Coder 30B, Qwen3 30B Instruct/Thinking 2507
- MiniMax-M2.5 High Throughput
- Trinity Mini, DeepSeek OCR, MM Poly 8B
Also adds 'mm-poly' family to family.ts for the Clarifai multimodal model.
2026-02-25 12:28:40 -05:00
Lucas Almeida
c240bce614
fix: adding missing structured_output parameter
2026-02-25 11:19:54 -03:00
Lucas Almeida
843a1d182a
feat: adding gpt-5.3-codex for Azure Foundry
2026-02-25 11:09:10 -03:00
PandaSt0rm
443c06ca03
fix MiniMax M2.5 modalities in Alibaba Coding Plan
...
- set MiniMax-M2.5 input modalities to text-only
- keep output modality as text
- validate with bun validate
2026-02-25 13:10:58 +02:00
PandaSt0rm
84466021ca
add MiniMax M2.5 to Alibaba Coding Plan and align third-party limits
...
- add MiniMax-M2.5 model config under providers/alibaba-coding-plan/models
- update GLM-4.7 limits to 202,752 context / 16,384 output
- update GLM-5 limits to 202,752 context / 16,384 output
- update Kimi K2.5 output limit to 32,768
- validate with bun validate
2026-02-25 13:05:18 +02:00
mingholy.lmh
b995e90cf5
fix: update context and output limits for alibaba-coding-plan-cn models
...
Update model limits:
- qwen3-coder-plus: context 1_048_576 → 1_000_000
- glm-5: output 131_072 → 16_384
- glm-4.7: output 131_072 → 16_384
- kimi-k2.5: output 65_536 → 32_768
Co-authored-by: Qwen-Coder <qwen-coder@alibabacloud.com >
2026-02-25 17:44:49 +08:00
dpuyosa
291e2eefe9
[venice] Remove interactive API key prompt
...
- Remove readline import and promptForApiKey function
- Remove prompt fallback, rely on CLI arg or env var only
- Update README to reflect change
2026-02-25 09:51:50 +01:00
Sewer56
0428299773
Added: Qwen3.5-397B natively supports image, MM2.5 No Image as it was a mistake.
2026-02-25 08:10:42 +00:00
Frank
96e9537b34
update zen models
2026-02-25 01:05:35 -05:00
Ryo Tulman
09dc7060ac
Venice: Add GPT-5.3 Codex
2026-02-24 23:37:03 -06:00
Colby Gilbert
a463717783
chore(firmware): remove gpt-5
2026-02-24 20:33:49 -08:00
Colby Gilbert
6d721dd32d
feat(firmware): gpt-5.3-codex
2026-02-24 20:32:17 -08:00
liuchang-reolink
3ae513785a
add step-3-5-flash for nvidia
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-02-25 12:05:19 +08:00
Colby Gilbert
e8d667b628
Merge branch 'anomalyco:dev' into dev
2026-02-24 20:01:26 -08:00
liuchang-reolink
7616a65e63
add qwen3.5-397b-a17b for nvidia
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-02-25 11:40:15 +08:00
yinxulai
a5b3da12c1
chore(qiniu-ai): update provider config
2026-02-25 10:20:55 +08:00
yinxulai
a05fbc3604
feat(qiniu-ai): add new model configurations
2026-02-25 10:16:40 +08:00
Aiden Cline
2189030e57
Merge pull request #1022 from armishra/feat/add-minimax-m2.5-baseten
...
feat(provider): Add MiniMax-M2.5 for baseten
2026-02-24 17:28:04 -06:00
Aiden Cline
c7ecc08442
Merge pull request #1019 from dpuyosa/venice
...
Venice: Update gemini-3-1-pro-preview config
2026-02-24 17:27:46 -06:00
Aiden Cline
830046e45e
Merge pull request #1021 from sylviezhang37/update-vercel-models-20260224-2134
...
Update Vercel models
2026-02-24 17:27:19 -06:00
Aiden Cline
c7b26477b9
Update cache_read value in gemini-3.1-pro-preview.toml
2026-02-25 04:26:56 +05:00
Aiden Cline
9e60f516fa
Update cost input and output values in TOML file
2026-02-25 04:26:18 +05:00
PandaSt0rm
7ccbb58c5f
add Alibaba Coding Plan provider and model configs
2026-02-25 01:10:42 +02:00
Archit Mishra
da906a0816
feat(provider): Add MiniMax-M2.5 for baseten
2026-02-24 14:33:41 -08:00
github-actions[bot]
36c9f82905
chore(vercel): update Vercel model definitions
...
Auto-generated by weekly workflow from Vercel AI Gateway API.
Co-Authored-By: github-actions[bot] <41898282+github-actions[bot]@users.noreply.github.com>
2026-02-24 21:34:20 +00:00
Frank
978214e143
update zen models
2026-02-24 15:23:24 -05:00
Jérôme Benoit
6aa69460a1
feat(sap-ai-core): add GPT-4.1, Gemini 2.5 Flash Lite, and Sonar models
...
Add 5 new model definitions for SAP AI Core provider:
- gpt-4.1: OpenAI GPT-4.1 (1M context, 32K output)
- gpt-4.1-mini: OpenAI GPT-4.1 Mini (1M context, 32K output)
- gemini-2.5-flash-lite: Google Gemini 2.5 Flash Lite (1M context, 65K output)
- sonar: Perplexity Sonar (128K context, 4K output)
- sonar-pro: Perplexity Sonar Pro (200K context, 8K output)
All specs verified against official provider documentation.
2026-02-24 20:50:10 +01:00
fhennerkes
0f16fcf231
poe: add GPT-5.3-Codex model
2026-02-24 11:39:47 -08:00
fhennerkes
e583c700f9
Merge branch 'anomalyco:dev' into dev
2026-02-24 11:36:20 -08:00
dpuyosa
96d278932c
[venice] Update gemini-3-1-pro-preview config
...
- Reduce output token limit from 250K to 65K
2026-02-24 11:27:18 +01:00
Sewer56
eee3303df0
Add missing synthetic.new models
...
Add configuration for hf:Qwen/Qwen3.5-397B-A17B and hf:MiniMaxAI/MiniMax-M2.5
to the synthetic provider, based on API specs from synthetic.new.
Note: API reports image support but these models may not natively support
images (likely rerouted/proxied through vision-capable infrastructure).
2026-02-24 09:19:57 +00:00
RioPlay
838416044f
add: newer MiniMax, GLM, and Kimi models to DeepInfra
2026-02-23 22:21:28 -06:00
Frank
51441f47d9
update zen models
2026-02-23 15:08:28 -05:00
Nacho F. Lizaur
7fc2c6154d
fix(amazon-bedrock): correct Claude Opus 4.6 context window from 1M to 200K
2026-02-23 20:10:58 +01:00
Colby Gilbert
1439781a76
feat(firmware): add deepseek 3.2, glm 5, kimi k2.5, minimax m2.5
2026-02-22 21:29:51 -08:00
minpeter
8fc0d87742
add friendli minimax m2.5 model config
2026-02-23 13:18:17 +09:00
Colby Gilbert
f660955784
feat(firmware): add grok models
2026-02-22 15:37:53 -08:00
Colby Gilbert
eb11c327b8
feat(firmware): gemini 3.1 pro, sonnet reasoning
2026-02-21 23:43:25 -08:00
Bryan Nie
0dfde60c14
models: alibaba-cn: add MiniMax-M2.5
2026-02-22 01:09:24 +08:00
Wendell Misiedjan
bab7727bad
Add AWS_BEARER_TOKEN_BEDROCK to Amazon Bedrock provider env
...
The @ai-sdk/amazon-bedrock package supports Bearer token authentication
via the AWS_BEARER_TOKEN_BEDROCK environment variable as an alternative
to IAM SigV4 auth. This uses Bedrock API keys for simplified access.
2026-02-21 16:27:45 +01:00
Mikal Sande
4b4a2364c6
Append (latest) to Mistral models that refer to the latest version.
2026-02-21 09:19:13 +01:00
Frank
c36b8e9433
update zen models
2026-02-20 23:20:24 -05:00
xiaojie.zj
5b8e983e7c
feat: add Gemini 3.1 Pro Preview for ZenMux provider
2026-02-21 10:39:30 +08:00
Frank
0d2a52dd9d
update zen models
2026-02-20 20:41:52 -05:00
Frank
b667ab78ac
update zen models
2026-02-20 20:19:33 -05:00
fhennerkes
e2da96cde4
poe: add Gemini-3.1-Pro and update Claude Sonnet 4.6
2026-02-20 11:28:18 -08:00
Jake Jia
9d042ac986
Update providers/zenmux/models/openai/gpt-5.2-pro.toml
...
Co-authored-by: Copilot <175728472+Copilot@users.noreply.github.com >
2026-02-21 01:22:04 +08:00
Phoen1xCode
7192dc0ba8
feat(openai): add GPT-5.2-Pro model via zenmux provider
...
Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com >
2026-02-21 01:07:58 +08:00
Phoen1xCode
05ea56a12a
fix(minimax): remove duplicated provider prefix from MiniMax M2.5 Lightning name
...
Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com >
2026-02-21 01:07:47 +08:00
Aiden Cline
beb449a417
Merge pull request #994 from davidfph/fix/qwen3.5-release-date
...
fix(qwen): update Qwen3.5 release_date and last_updated to 2026-02-16
2026-02-20 10:29:15 -06:00
Aiden Cline
8f76f9b217
Merge pull request #988 from MeganovaAI/fix-meganova-logo
...
Update Meganova logo to official brand icon
2026-02-20 10:29:02 -06:00
Aiden Cline
18ddcde669
fix: azure & cognitive model distinctions
2026-02-20 10:28:22 -06:00
David Fu
ff36ded35f
fix(qwen): update Qwen3.5 release_date and last_updated to 2026-02-16
2026-02-20 20:33:08 +08:00
Aiden Cline
b0b8074a94
Merge pull request #990 from kailiu42/dev
...
models: siliconflow-cn: add new models
2026-02-20 03:01:45 -06:00
Kai Liu
4c30a522f7
models: siliconflow-cn: add new models
...
New models per the latest list: https://cloud.siliconflow.cn/me/models
- Pro/MiniMaxAI/MiniMax-M2.5
- deepseek-ai/DeepSeek-OCR
- PaddlePaddle/PaddleOCR-VL
- PaddlePaddle/PaddleOCR-VL-1.5
Signed-off-by: Kai Liu <kraml.liu@gmail.com >
2026-02-20 16:26:26 +08:00
Aiden Cline
2d63d713de
Merge pull request #992 from anomalyco/fix-azure-models
...
fix: ensure that anthropic models on azure providers have correct urls
2026-02-20 02:19:13 -06:00
Aiden Cline
ac9d0af8e3
fixes
2026-02-20 02:13:23 -06:00
Aiden Cline
a1ad90a9b1
Merge pull request #991 from zainhas/dev
...
[Together AI] add qwen3.5
2026-02-20 01:02:23 -06:00
Zain Hasan
8a6e0dd917
add qwen3.5
2026-02-19 21:38:46 -08:00
Aiden Cline
2da10b739c
Merge pull request #989 from propilideno/fix/adding_missing_azure_foundry_model
...
Add missing GPT-5.2 metadata for Azure Cognitive Services
2026-02-19 18:47:56 -06:00
Aiden Cline
c17e0b9d0f
Merge pull request #987 from dpuyosa/venice
...
Venice: Add Gemini 3.1 Pro Preview and update model configs
2026-02-19 18:47:48 -06:00
Lucas Almeida
877a1175f4
chore: replacing by symbolic link like the other ones
2026-02-19 21:21:05 -03:00
Boqian
1bc83abeeb
Update Meganova logo to official brand icon
2026-02-19 18:57:17 -05:00
dpuyosa
70caedba86
[venice] Add Gemini 3.1 Pro Preview and update model configs
...
- Add new Gemini 3.1 Pro Preview model configuration
- Update Claude Sonnet 4.6 release dates
- Enable open_weights for MiniMax M25
2026-02-19 22:54:59 +01:00
Aiden Cline
60c90a27a0
Merge pull request #985 from sylviezhang37/update-vercel-model-gen-script
...
feat(provider): exclude image/video models
2026-02-19 15:49:25 -06:00
Aiden Cline
2bd0d5446e
Merge pull request #986 from riasvdv/add-gemini-3.1-pro
...
Add Gemini 3.1 Pro Preview to copilot models
2026-02-19 15:49:12 -06:00
Aiden Cline
5f135517b1
Remove audio and video from input modalities
2026-02-19 15:48:42 -06:00
Aiden Cline
e2af7819b4
Rename gemini-3.5-pro-preview.toml to gemini-3.1-pro-preview.toml
2026-02-19 15:47:25 -06:00
Rias
ca6c251b3a
Add Gemini 3.1 Pro Preview to copilot models
2026-02-19 22:43:26 +01:00
Sylvie Zhang
7b1b590d10
exclude image/video gen models
2026-02-19 13:24:11 -08:00
Aiden Cline
5097a1e954
Merge pull request #966 from mhkok/mkok/feat/add-evroc-provider
...
add evroc provider + models
2026-02-19 14:15:51 -06:00
Aiden Cline
05959a83b6
Update font family in Kimi-K2.5 configuration
2026-02-19 14:15:07 -06:00
Aiden Cline
1492e067a4
Merge pull request #976 from too-green/patch-2
...
Add Qwen3 Coder Next model for openrouter
2026-02-19 14:01:07 -06:00
Aiden Cline
41c81535c3
fix: zen
2026-02-19 12:54:29 -06:00
Aiden Cline
829756fc41
Merge pull request #983 from mdrxy/mdrxy/fix-gemini-3
...
fix Gemini 3.1 model names
2026-02-19 12:34:07 -06:00
Mason Daugherty
e6ef906c41
fix
2026-02-19 13:21:52 -05:00
Aiden Cline
0f84db6bc6
Merge pull request #975 from xiaojiezj/zenmux_dev_0219
...
feat: Add new models for ZenMux provider
2026-02-19 11:33:49 -06:00
Aiden Cline
4bb6d52a7c
Merge pull request #979 from hanouticelina/fix-interleaved-for-hf-provider
...
Fix Hugging Face interleaved `reasoning field: reasoning_details` -> `reasoning_content`
2026-02-19 11:33:34 -06:00
Aiden Cline
bdd0194e73
Merge pull request #981 from mdrxy/mdrxy/add-gemini-3.1
...
add gemini 3.1 to google/openrouter
2026-02-19 11:33:15 -06:00
Frank
41a9502628
update zen models
2026-02-19 11:51:37 -05:00
Mason Daugherty
384e747129
add gemini 3.1 to google/openrouter
2026-02-19 11:24:30 -05:00
Frank
e4bb5ceac6
update zen models
2026-02-19 10:16:51 -05:00
Frank
c830964c3f
update zen models
2026-02-19 09:37:07 -05:00
Celina Hanouti
782b6277ae
Fix Hugging Face interleaved reasoning field
2026-02-19 15:27:49 +01:00
Frank
e63d48ae9c
update zen models
2026-02-19 07:42:52 -05:00
Matthijs Kok
029522aa96
fix family names
2026-02-19 08:48:43 +01:00
Ahmed
482ed2e833
Add Qwen3 Coder Next model for openrouter
...
Added model configuration for Qwen3 Coder Next
2026-02-19 12:36:55 +05:00
Aiden Cline
c6635aa7c3
Merge pull request #968 from MeganovaAI/add-meganova-provider
...
Add Meganova as a provider
2026-02-18 23:44:03 -06:00
Aiden Cline
a6ffef7e4f
Merge pull request #969 from SomeoneWithOptions/dev
...
add claude sonnet 4.6 on openrouter
2026-02-18 23:40:02 -06:00
Aiden Cline
fda9bb5335
Merge pull request #973 from sylviezhang37/update-vercel-models-20260219-0026
...
Update Vercel models
2026-02-18 23:38:56 -06:00
Aiden Cline
98d6901697
Update input cost value in qwen3.5-plus.toml
2026-02-18 23:38:49 -06:00
Aiden Cline
13ae499d7c
tweak values
2026-02-18 23:38:18 -06:00
xiaojie.zj
f21d205d6e
feat: 增加Claude Sonnet 4.6/Doubao-Seed-2.0-lite/Doubao-Seed-2.0-mini/Doubao-Seed-2.0-pro模型
2026-02-19 11:10:28 +08:00
Lucas Almeida
c82b08d778
fix: adding missing gpt-5.2 model on azure foundry
2026-02-18 23:38:10 -03:00
Sylvie Zhang
ea612760cf
Delete providers/vercel/models/recraft/recraft-v4.toml
2026-02-18 16:34:47 -08:00
Sylvie Zhang
48a64f0834
Delete providers/vercel/models/recraft/recraft-v4-pro.toml
2026-02-18 16:34:35 -08:00
github-actions[bot]
040e7fff4e
chore(vercel): update Vercel model definitions
...
Auto-generated by weekly workflow from Vercel AI Gateway API.
Co-Authored-By: github-actions[bot] <41898282+github-actions[bot]@users.noreply.github.com>
2026-02-19 00:26:59 +00:00
SomeoneWithOptions
a6d14928b6
added claude sonnet 4.6 on openrouter
2026-02-18 18:49:56 -05:00
Boqian
92d9e89690
Set reasoning=false for DeepSeek V3 series
...
V3-0324, V3.1, V3.2, V3.2-Exp are chat models, not reasoning models.
Only DeepSeek-R1 is a reasoning model.
Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com >
2026-02-18 16:41:07 -05:00
Boqian
2d83b96bb1
Fix interleaved reasoning_content based on Meganova API testing
...
Tested each model with include_reasoning=true against the live API.
Added [interleaved] to: GLM-4.6, MiniMax-M2.1, MiniMax-M2.5, Kimi-K2.5
Removed [interleaved] from: DeepSeek-V3.1, V3.2, V3.2-Exp, MiMo-V2-Flash
Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com >
2026-02-18 16:35:51 -05:00
Boqian
6e02385a6b
Add interleaved reasoning_content to DeepSeek V3.1, V3.2, V3.2-Exp
...
These models support interleaved reasoning output, matching how other
providers (deepinfra, baseten, chutes) configure them.
Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com >
2026-02-18 16:29:41 -05:00
Boqian
a2f8234c8e
Update pricing and context limits from Meganova API
...
Use actual pricing from https://api.meganova.ai/v1/models instead of
reference data from other providers.
Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com >
2026-02-18 16:25:43 -05:00
Aiden Cline
2f43c70397
Merge pull request #967 from nicolasgere/dev
...
feat(provider): Add glm-5 for baseten
2026-02-18 15:19:47 -06:00
Boqian
006cb53c1e
Add Meganova as a provider with 19 open-weight models
...
Adds Meganova AI (https://api.meganova.ai/v1 ) as an OpenAI-compatible provider
with curated open-weight models including DeepSeek, GLM, Qwen, Kimi, MiniMax,
MiMo, Llama, and Mistral families.
Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com >
2026-02-18 16:17:51 -05:00
nicolasgere
d2832f9d1d
Update GLM-5.toml
2026-02-18 16:15:22 -05:00
nicolasgere
73c0fcc19e
Rename GLM-5 to GLM-5.toml
2026-02-18 16:14:26 -05:00
nicolasgere
04a1fe8043
Create GLM-5
2026-02-18 16:14:09 -05:00
Matthijs Kok
7fa0eeebf9
add evroc provider + models
2026-02-18 19:57:44 +01:00
Aiden Cline
cb72e1780f
Merge pull request #962 from aldosch/add-sonnet-4-6-vercel
...
add sonnet 4.6 to vercel ai gateway
2026-02-18 12:19:00 -06:00
Aiden Cline
7ee7e88346
Merge pull request #960 from fa-sharp/patch-1
...
fix: OpenRouter output modalities for image-only models
2026-02-18 12:18:27 -06:00
Aiden Cline
4cc230ba6c
Merge pull request #963 from vglafirov/gitlab/add-sonnet-4-6
...
feat(gitlab): add Claude Sonnet 4.6 model
2026-02-18 12:17:57 -06:00
Aiden Cline
65d6355ffa
Merge pull request #964 from xinrui-z/aihubmix-add-claude-4-6
...
aihubmix: add models
2026-02-18 12:17:48 -06:00
Xinrui
1b6157c0ee
aihubmix: add models
2026-02-18 22:55:03 +08:00
Vladimir Glafirov
5e25c0838e
feat(gitlab): add Claude Sonnet 4.6 model
2026-02-18 13:52:26 -01:00
aldosch
b06ab95c9e
add sonnet 4.6 to vercel ai gateway
2026-02-19 00:38:35 +11:00
Daltonganger
33f4e77ee4
fix(nano-gpt): normalize family enums for model validation
2026-02-18 12:53:40 +01:00
Daltonganger
51ef2e51ae
finalize nano-gpt model sync and release metadata
2026-02-18 12:44:36 +01:00
farshad
9d873cc764
add back trailing newline
2026-02-18 02:18:40 -05:00
farshad
d0acb0a35d
fix output modalities for black forest flux models
2026-02-18 01:52:31 -05:00
farshad
697c191fea
Update output modalities in seedream-4.5.toml
2026-02-18 01:45:06 -05:00
Aiden Cline
7d9ff92ffd
fix: glm 5 maas
2026-02-17 23:41:48 -06:00
Aiden Cline
03a2aee0de
Merge pull request #716 from bluet/feat/google-vertex-openai
...
feat: add google-vertex-openai provider for Vertex AI partner models
2026-02-17 23:38:08 -06:00
Aiden Cline
4eb19459dd
Merge pull request #958 from mugnimaestra/feat/add-qwen3.5-397b-a17b-tee-chutes
...
feat: add Qwen3.5 397B A17B TEE to Chutes provider listings
2026-02-17 23:27:18 -06:00
Aiden Cline
b74d7eb495
Merge pull request #957 from dpuyosa/veniceScript
...
Venice: Update generate-venice script to include context_over_200k cost
2026-02-17 23:26:28 -06:00
Aiden Cline
60aa5aca14
Merge pull request #956 from dpuyosa/venice
...
Venice: Add Claude Sonnet 4.6 and GLM 4.7 Flash Heretic models
2026-02-17 20:22:44 -06:00
Muhammad Mugni Hadi
67235a4e39
feat: add Qwen3.5 397B A17B TEE to Chutes provider listings
2026-02-18 09:09:45 +07:00
dpuyosa
ff03fd6906
[venice] Add Claude Sonnet 4.6 and GLM 4.7 models
...
- Add Claude Sonnet 4.6 model with context_over_200k pricing
- Add GLM 4.7 Flash Heretic model (open weights)
- Add context_over_200k pricing tier to Claude Opus 4.6
2026-02-18 02:02:19 +01:00
dpuyosa
423b177c2b
Update generate-venice script to include context_over_200k cost
2026-02-18 01:55:52 +01:00
Aiden Cline
8c263109c5
Merge pull request #955 from maahir30/open-router-structured-output
...
Add structured output support for OpenRouter models
2026-02-17 18:52:02 -06:00
Maahir Sachdev
1bab7e8438
update open router models
2026-02-17 16:42:52 -08:00
Aiden Cline
7a163dbc60
Merge pull request #953 from mongrelion/dev
...
feat: add github copilot claude sonnet 4.6 model
2026-02-17 18:30:54 -06:00
Aiden Cline
c94d025aa7
fixes
2026-02-17 18:23:01 -06:00
Aiden Cline
55e8be17b5
Merge pull request #949 from cgilly2fast/dev
...
feat(firmware): add sonnet 4.6
2026-02-17 18:21:16 -06:00
Aiden Cline
56096012b2
Merge pull request #948 from elithrar/patch-1
...
add Sonnet 4.6 model config
2026-02-17 17:53:20 -06:00
Aiden Cline
9b8a22d756
Merge pull request #725 from janszypulski/add-provider-cloudferro-sherlock
...
Add provider - Cloudferro Sherlock
2026-02-17 17:50:14 -06:00
Aiden Cline
5a778c6e93
Merge pull request #682 from the-lazy-me/add-qihang-provider
...
feat: add QiHang provider with 7 models
2026-02-17 17:48:51 -06:00
Aiden Cline
4d08659acf
Merge pull request #651 from yinxulai/feat/qiniu-ai
...
feat: add Qiniu AI provider configuration
2026-02-17 17:46:05 -06:00
Aiden Cline
128c9ec469
Merge pull request #308 from d-oit/feature/perplexity-sonar-deep-research
...
Feature/perplexity sonar deep research
2026-02-17 17:37:20 -06:00
Carlos León
dc11781324
feat: add github copilot claude sonnet 4.6 model
...
Model list sourced from GitHub Settings page showing currently available models. Specifications cross-referenced with Anthropic provider implementation.
2026-02-18 00:22:22 +01:00
Colby Gilbert
9d0b37bea3
feat(firmware): add sonnet 4.6
2026-02-17 15:03:21 -08:00
Matt Silverlock
c8d09fe349
add Sonnet 4.6 model config
2026-02-17 17:27:51 -05:00
Aiden Cline
1a22b93fc2
Merge pull request #947 from fhennerkes/dev
...
poe: add Claude-Sonnet-4.6 and update XAI models
2026-02-17 15:56:11 -06:00
Aiden Cline
3918131cb8
Merge pull request #946 from monotykamary/remove-fireworks-deprecated-models-2026-02-12
...
chore(fireworks-ai): remove deprecated serverless models
2026-02-17 15:56:00 -06:00
fhennerkes
a1d9c5134c
poe: add Claude-Sonnet-4.6 and update XAI models
2026-02-17 13:38:54 -08:00
Ruben Beuker
20abb5b8df
preserve curated release dates for key nano-gpt models
...
Keep existing curated release and last-updated values for models where NanoGPT API uses the generic created timestamp baseline.
2026-02-17 22:09:03 +01:00
Tom X Nguyen
dc36ed54ae
chore(fireworks-ai): remove deprecated serverless models
...
Remove 6 Fireworks serverless models deprecated on February 12, 2026:
- glm-4.6 (migrate to glm-4.7)
- deepseek-r1-0528 (migrate to deepseek-v3.2 or deepseek-v3.1)
- deepseek-v3-0324 (migrate to deepseek-v3.2 or deepseek-v3.1)
- qwen3-235b-a22b (migrate to kimi-k2-instruct-0905)
- qwen3-coder-480b-a35b-instruct (migrate to kimi-k2-instruct-0905)
- minimax-m2 (migrate to MiniMax-M2.1)
See: https://fireworks.ai/models?modelTypes=Serverless
2026-02-18 04:04:51 +07:00
Ruben Beuker
8cb462f29b
sync nano-gpt models with live API catalog
...
Refresh NanoGPT model files to match the current /api/v1/models output, remove stale entries, and add newly available models while preserving path-based IDs.
Also ignore local TokenSpeed sqlite artifacts so private monitoring data is not shown or committed.
2026-02-17 22:01:54 +01:00
Frank
89486ec705
update zen models
2026-02-17 14:12:30 -05:00
Aiden Cline
f313f802ee
Merge pull request #940 from nitishxyz/add-claude-sonnet-4-6
...
feat(models): add Claude Sonnet 4.6 model configurations
2026-02-17 13:12:10 -06:00
nitishxyz
128615ddd7
feat(models): add Claude Sonnet 4.6 model configurations
...
- Add Claude Sonnet 4.6 to Anthropic provider with full capabilities
- Add regional variants (US, EU, Global) for Amazon Bedrock provider
- Add Google Vertex Anthropic provider configuration
- Define pricing, context limits (200k tokens), and modalities
Co-authored-by: ottocode-io[bot] <261994719+ottocode-io[bot]@users.noreply.github.com>
2026-02-18 00:01:02 +05:30
Aiden Cline
756fb772c1
Merge pull request #939 from Nomadcxx/fix/kilo-npm-provider
...
fix(kilo): use @ai-sdk/openai-compatible instead of opencode-kilo-auth
2026-02-17 11:29:52 -06:00
Nomadcxx
e86f0afd87
fix(kilo): use @ai-sdk/openai-compatible npm package
...
The npm field pointed to opencode-kilo-auth which causes
ProviderInitError when loading Kilo models.
Switched to @ai-sdk/openai-compatible (already bundled in OpenCode)
and added api field for the gateway endpoint.
2026-02-18 04:21:52 +11:00
Aiden Cline
29c5e28a43
Merge pull request #791 from samsja/add-intellect-3
...
Add Intellect 3 model from Prime Intellect
2026-02-17 10:56:34 -06:00
Aiden Cline
ea414b1500
Merge pull request #935 from ConceptCodes/feat/add-glm-flashx-model
...
feat: add GLM-4.7-FlashX model configuration
2026-02-17 10:34:45 -06:00
Aiden Cline
4556fe8b5b
Merge pull request #937 from gary149/feat/huggingface-qwen3.5-m2.5-coder-next
...
feat(huggingface): add Qwen3.5-397B, MiniMax-M2.5, Qwen3-Coder-Next
2026-02-17 10:34:33 -06:00
Aiden Cline
8af23aeba5
Merge pull request #938 from spiffytech/dev
...
Add Ollama Cloud support for Qwen 3.5
2026-02-17 10:34:18 -06:00
Aiden Cline
9bfe1203c6
ci
2026-02-17 10:34:02 -06:00
spiffytech
a619966e22
Added Ollama Cloud support for Qwen 3.5
2026-02-17 10:00:15 -05:00
Victor Muštar
f48d55e1aa
chore: remove accidentally committed skill file
2026-02-17 10:31:27 +01:00
Victor Muštar
fe0ddcb666
feat(huggingface): add Qwen3.5-397B, MiniMax-M2.5, Qwen3-Coder-Next
2026-02-17 10:31:18 +01:00
Frank
af1e1d1f51
update zen models
2026-02-17 02:08:24 -05:00
Aiden Cline
774a9f40b0
Merge pull request #933 from too-green/patch-1
...
Fix the display name of GLM-4.7-Flash
2026-02-17 00:23:26 -06:00
Aiden Cline
4f01ffb017
Merge pull request #934 from PandaSt0rm/add-minimax-m2.5-highspeed-models
...
Add MiniMax-M2.5-highspeed models for official MiniMax providers
2026-02-17 00:23:10 -06:00
Aiden Cline
7d768260cf
Merge pull request #936 from Alex-wuhu/dev
...
add Qwen3.5-397B-A17B for novita
2026-02-17 00:22:30 -06:00
Alex-wuhu
467d269522
add Qwen3.5-397B-A17B for novita
2026-02-17 13:22:04 +08:00
concept
5ec496d6e1
feat: add GLM-4.7-FlashX model configuration
2026-02-16 21:31:16 -06:00
PandaSt0rm
45ca42f95a
add MiniMax-M2.5-highspeed models
2026-02-17 03:53:05 +02:00
Ahmed
f19ebce14c
Rename model to GLM-4.7-Flash
...
Both GLM 4.7 and GLM 4.7 Flash had been named to the same "GLM 4.7"
2026-02-17 05:03:41 +05:00
Aiden Cline
85f5340eeb
Merge pull request #931 from rifandyzv/dev
...
Add Qwen3.5 models for alibaba & alibaba-cn provider
2026-02-16 16:06:32 -06:00
Aiden Cline
39f06e82e7
Merge pull request #932 from cantalupo555/feat/add-openrouter-qwen3.5-plus-and-397b-a17b
...
feat: add Qwen3.5 models on OpenRouter
2026-02-16 16:05:56 -06:00
cantalupo555
4e7725d244
feat: add Qwen3.5 models on OpenRouter
2026-02-16 17:47:45 -03:00
Aiden Cline
7fe64bc498
Revert "Add image and video to input modalities"
...
This reverts commit 76e84a8b06 .
2026-02-16 12:12:29 -06:00
rifandyzv
f84a4a5cf9
feat: add Qwen3.5 models for alibaba & alibaba-cn provider
2026-02-17 01:28:16 +08:00
Aiden Cline
96f60c3329
Merge pull request #928 from Daltonganger/feat/kilo-provider-models
...
Add Kilo Gateway provider and import Kilo models
2026-02-16 11:08:11 -06:00
Aiden Cline
f67b9bdef7
Merge pull request #929 from Daltonganger/feat/nano-gpt-qwen35-models
...
Add four Qwen3.5 models for NanoGPT
2026-02-16 11:05:23 -06:00
Frank
76e84a8b06
Add image and video to input modalities
2026-02-16 12:00:43 -05:00
Daltonganger
cec16274c7
Add NanoGPT Qwen3.5 model variants
2026-02-16 17:09:17 +01:00
Daltonganger
17094722ea
Add Kilo provider and import Kilo model catalog
2026-02-16 17:00:15 +01:00
Matthew (BlueT) Lien
e3e230e3b3
fix: add api base URL template to partner model [provider] overrides
...
Add the api field with env-var template URL to all partner models so
opencode's loadBaseURL() can resolve the OpenAI-compatible endpoint.
Uses GOOGLE_VERTEX_PROJECT (not GOOGLE_CLOUD_PROJECT) because
googleVertexVars() resolves it through the full fallback chain
(GOOGLE_VERTEX_PROJECT → options.project → GOOGLE_CLOUD_PROJECT →
GCP_PROJECT → GCLOUD_PROJECT).
2026-02-16 21:51:55 +08:00
Aiden Cline
4666f36f3e
Merge pull request #926 from zainhas/dev
...
[Together AI] add minimax M2.5
2026-02-15 23:57:45 -06:00
Aiden Cline
5a2bcd704e
Merge pull request #900 from conglinyizhi/dev
...
feat: Add StepFun provider support
2026-02-15 23:57:35 -06:00
Zain Hasan
05861fd7fd
add minimax M2.5
2026-02-15 21:48:37 -08:00
Aiden Cline
e37bb8ae68
Merge pull request #923 from juls0730/dev
...
Fix cerebras/zai-gml-4.7 pricing
2026-02-15 20:03:47 -06:00
Aiden Cline
495e8006df
Merge pull request #905 from shelvick/add-vertex-glm-5
...
Add GLM-5 to Google Vertex AI
2026-02-15 20:03:34 -06:00
Aiden Cline
ad8dde798d
Merge pull request #924 from cgilly2fast/dev
...
feat(firmware): add reason to anthropic models
2026-02-15 20:03:23 -06:00
Aiden Cline
ab4fa333e3
Merge pull request #925 from 8dazo/dev
...
feat: add MiniMax M2.5 to Chutes provider listings
2026-02-15 20:03:12 -06:00
8dazo
6255298cd1
minimax model update
2026-02-16 06:31:01 +05:30
Colby Gilbert
3614087be6
Merge branch 'anomalyco:dev' into dev
2026-02-15 15:55:32 -08:00
Colby Gilbert
4495cb3569
feat(firmware): add reason to anthropic models
2026-02-15 15:55:02 -08:00
juls0730
8444d9293d
Fix cerebras/zai-gml-4.7 pricing
...
Prices from https://inference-docs.cerebras.ai/models/zai-glm-47#z-ai-glm-4-7
2026-02-15 17:53:57 -06:00
Aiden Cline
c1d36715ee
Merge pull request #914 from 8dazo/dev
...
feat: add Z-AI GLM-5 to Chutes provider listings
2026-02-15 15:45:37 -06:00
Aiden Cline
860e610b73
Merge pull request #922 from cgilly2fast/dev
...
chore: remove unsupported models
2026-02-15 15:45:28 -06:00
Colby Gilbert
beb84e769a
chore: remove unsupported models
2026-02-15 13:38:52 -08:00
Aiden Cline
0408546681
Merge pull request #921 from zerone0x/feat/add-bedrock-deepseek-v3.2
...
feat(amazon-bedrock): add DeepSeek V3.2
2026-02-15 15:29:26 -06:00
Aiden Cline
97a040bc5e
Merge pull request #915 from fanweixiao/dev
...
add glm-5, gpt-5-mini, deepseek-v3.2 models for vivgrid provider
2026-02-15 15:29:08 -06:00
Aiden Cline
f68786b892
Merge pull request #920 from anomalyco/revert-912-add-github-copilot-gpt-5-3-codex
...
Revert "feat: add GitHub Copilot GPT-5.3 Codex"
2026-02-15 08:42:51 -06:00
Clawdbot
816c3d96b9
feat(amazon-bedrock): add DeepSeek V3.2
...
Add DeepSeek V3.2 model to Amazon Bedrock provider.
Model ID: deepseek.v3.2-v1:0
Pricing (US regions): $0.62/1M input, $1.85/1M output
Ref: https://aws.amazon.com/about-aws/whats-new/2026/02/amazon-bedrock-adds-support-six-open-weights-models/
2026-02-15 09:19:18 +01:00
Aiden Cline
bac557c176
Revert "feat: add github copilot gpt-5.3-codex model ( #912 )"
...
This reverts commit 08db483d58 .
2026-02-14 18:31:30 -06:00
Aiden Cline
97e81f356e
Merge pull request #908 from hsnyus-09/feature/add-aurora-alpha
...
feat(openrouter): add aurora-alpha model definition
2026-02-14 17:34:22 -06:00
Matthew (BlueT) Lien
4c361218de
feat: add Vertex AI partner models with openai-compatible overrides
...
Add DeepSeek V3.1, Llama 4 Maverick, Llama 3.3 70B, and Qwen3 235B as
partner models under google-vertex provider. Update GLM-4.7 with
corrected specs from official Google Cloud docs.
Each partner model uses [provider] npm override to @ai-sdk/openai-compatible
since these models are served via Google's OpenAI-compatible endpoint,
while staying consolidated under the google-vertex provider per
maintainer feedback.
All specs (context windows, output limits, pricing, modalities)
verified against official Google Cloud documentation:
- cloud.google.com/vertex-ai/generative-ai/pricing
- cloud.google.com/vertex-ai/generative-ai/docs/maas/*
Changes:
- Update zai-org/glm-4.7-maas: fix context=200K, output=128K, add pdf
modality, correct release_date, add structured_output, add [provider]
- Add deepseek-ai/deepseek-v3.1-maas ($0.60/$1.70, 163K context)
- Add meta/llama-4-maverick-17b-128e-instruct-maas (vision, 524K ctx)
- Add meta/llama-3.3-70b-instruct-maas ($0.72/$0.72, 128K context)
- Add qwen/qwen3-235b-a22b-instruct-2507-maas ($0.22/$0.88, 262K ctx)
2026-02-15 06:35:17 +08:00
Anjul Garg
08db483d58
feat: add github copilot gpt-5.3-codex model ( #912 )
2026-02-14 14:27:49 -05:00
YuSung Han
e0c14d7883
Remove redundant lines in aurora-alpha.toml
2026-02-15 03:50:02 +09:00
Aiden Cline
e457c7f1dd
Merge pull request #916 from arshadbarves/add-nvidia-glm5
...
Add GLM5 model to nvidia provider
2026-02-14 11:53:09 -06:00
Arshad Barves
c86b97226c
Update providers/nvidia/models/z-ai/glm5.toml
...
Co-authored-by: Copilot <175728472+Copilot@users.noreply.github.com >
2026-02-14 17:36:22 +05:30
Test User
8ec101e8f1
Add GLM5 model to nvidia provider
2026-02-14 15:32:11 +05:30
C.C. Fan
cc1feddc11
add glm-5, gpt-5-mini, deepseek-v3.2 models
2026-02-14 17:08:07 +08:00
8dazo
95fd20e66a
update name
2026-02-14 13:54:20 +05:30
8dazo
8c322d46dd
Chutes Model Listings update
2026-02-14 13:49:06 +05:30
conglinyizhi
da7a26b7b1
fix: 修复 StepFun provider 文档链接
...
将 doc 字段从 https://platform.stepfun.com/docs
改为 https://platform.stepfun.com/docs/zh/overview/concept
2026-02-14 15:07:40 +08:00
Aiden Cline
5a00835470
Merge pull request #913 from kavin-kr/patch-1
...
Update release date for nova-2-pro-v1 model
2026-02-14 00:50:06 -06:00
Frank
b0c0a91926
update zen models
2026-02-14 00:52:08 -05:00
Kavin
8b2d995801
Update release date for nova-2-pro-v1 model
2026-02-13 23:46:55 -06:00
Aiden Cline
3f4b804ca1
Merge pull request #911 from monotykamary/feat/add-minimax-m2.5
...
feat: add fireworks minimax m2.5 model and fix m2.1 cache pricing
2026-02-13 19:00:46 -06:00
Tom X Nguyen
133e605c23
feat: add minimax m2.5 model and fix m2.1 cache pricing
2026-02-14 07:53:15 +07:00
Aiden Cline
f3cff10e78
Merge pull request #910 from keenborder786/fix/gpt_5_2_pro
...
fix: gpt 5-2-pro does not support structured output
2026-02-13 17:53:00 -06:00
keenborder786
4732ca7f77
fix: gpt 5-2-pro does not support structured output
2026-02-14 04:50:19 +05:00
Aiden Cline
06f79b4142
Merge pull request #909 from juls0730/dev
...
Add all missing cohere models offered by the cohere api
2026-02-13 16:47:11 -06:00
Aiden Cline
69106c6f36
Merge pull request #907 from Algowary/dev
...
Chutes Model Listings update
2026-02-13 16:44:43 -06:00
Zoe
dc1d8b78d0
Add all missing cohere models offered by the cohere api
...
This commit adds all the models offered by the official cohere api
that are not yet available in the models.dev repo, excluding the
rerank and embed models.
2026-02-13 16:43:57 -06:00
hsnyus-09
4383829304
feat(openrouter): add aurora-alpha model definition
2026-02-14 06:17:25 +09:00
Algowarry
610713b805
Merge branch 'dev' of https://github.com/Algowary/models.dev into dev
2026-02-13 16:08:26 -05:00
Algowarry
07e3eec2cf
Chutes Model Inventory Update
...
Updating the models available from the provider chutes.ai
2026-02-13 16:01:46 -05:00
Algowarry
ccf8a3e82a
Merge branch 'dev' of https://github.com/Algowary/models.dev into dev
2026-02-13 15:07:16 -05:00
Algowarry
8e1d34323a
Chutes Model Update
...
Model inventory and stats update
2026-02-13 14:38:07 -05:00
Aiden Cline
5f6d36a463
Merge pull request #787 from elithrar/fix/cloudflare-ai-gateway-provider-package
...
use official ai-gateway-provider package for Cloudflare AI Gateway
2026-02-13 12:44:45 -06:00
Scott Helvick
84e43b1912
Add GLM-5 to Google Vertex AI
2026-02-13 17:57:08 +00:00
Aiden Cline
a9f14cbae3
Merge pull request #903 from micuintus/dev
...
fix(nebius): correct model ID casing to match Token Factory API
2026-02-13 10:39:51 -06:00
Aiden Cline
97baff037b
Merge pull request #896 from zainhas/dev
...
[Together AI] Add GLM-5
2026-02-13 10:38:06 -06:00
Aiden Cline
b149bd83ad
Merge pull request #902 from qychen2001/dev
...
chore(siliconflow): update siliconflow/siliconflow-cn models
2026-02-13 10:25:04 -06:00
Aiden Cline
30ef1d1b51
Merge pull request #898 from 888-wzk/feature/chenger_20260128
...
feat(models): Added minimax model profile
2026-02-13 10:24:36 -06:00
Aiden Cline
a62ccfd392
Merge pull request #901 from dpuyosa/venice
...
Venice: Add MiniMax M2.5 model configuration
2026-02-13 10:24:18 -06:00
QiyuanChen
23e195a0fa
feat(models): Add interleaved reasoning_content field to GLM-4.7 and GLM-5 configurations for zai-org and Pro
2026-02-13 23:35:56 +08:00
Aiden Cline
d1c5bc811e
Merge pull request #899 from niushuai1991/feature/kuae-cloud-coding-plan
...
add provider: kuae cloud coding plan
2026-02-13 09:23:33 -06:00
Michael Voigt
ad7e8047b9
fix(nebius): correct model ID casing to match Token Factory API
...
Fix lowercase model ID bug that caused "The model does not exist" errors.
- qwen/ → Qwen/ directory
- Fixed model file casing to match API exactly:
- google/gemma-* → lowercase (gemma-2-2b-it, etc.)
- meta-llama/*-Fast → lowercase fast suffix
- nvidia/Llama-3_1-* → underscore instead of dot
- nvidia/NVIDIA-* → uppercase NVIDIA prefix
- black-forest-labs/flux-* → all lowercase
- BAAI/bge-* → all lowercase
- All Qwen models → proper casing
* Remove outdated models not in API:
- deepseek-ai/DeepSeek-V3
- meta-llama/Llama-3.1-405B-Instruct
- zai-org/GLM-4.7
* Add new model:
- moonshotai/Kimi-K2.5 (262K context, multimodal)
Fixes: https://github.com/anomalyco/opencode/issues/12461
and: https://ideas.nebius.com/en/p/token-factory-api-lowercase-model-ids
Note: The changes made and verified with actual Nebius API access
2026-02-13 14:28:48 +01:00
QiyuanChen
73393e9e41
chore(models): Remove Qwen3-30B-A3B and DeepSeek-R1-Distill-Qwen-7B model configuration files from siliconflow and siliconflow-cn
2026-02-13 20:11:19 +08:00
QiyuanChen
3e5566ab9d
chore(models): Remove GLM-4.1V-9B-Thinking model configuration files from siliconflow and siliconflow-cn
2026-02-13 20:09:14 +08:00
QiyuanChen
f5096e4b54
chore(models): Remove Kimi-Dev-72B model configuration files from siliconflow and siliconflow-cn
2026-02-13 20:08:14 +08:00
QiyuanChen
1d6e26574f
chore(models): Remove MiniMaxAI/MiniMax-M1-80k and MiniMax-M2 model configuration files
2026-02-13 20:07:15 +08:00
QiyuanChen
57b0608e70
feat(models): Introduce Step-3.5-Flash model configuration and remove deprecated Step-3 model files
2026-02-13 20:06:05 +08:00
QiyuanChen
88ed698a69
feat(models): Enable structured_output in GLM-4.7 and GLM-5 configurations for zai-org and Pro
2026-02-13 20:04:27 +08:00
QiyuanChen
286c43f2cd
feat(glm-5): Add new GLM-5 model configuration files for zai-org and Pro
2026-02-13 20:00:10 +08:00
dpuyosa
04d82741fa
[venice] Add MiniMax M2.5 model configuration
...
- Modalities: text input/output
- Context window: 198K tokens
- Max output: 32K tokens
- Pricing: $0.40/M input, $1.60/M output, $0.04/M cache read
2026-02-13 09:44:22 +01:00
conglinyizhi
58c595b95f
feat: Add StepFun provider support
...
- Add StepFun(阶跃星辰) as a new provider with OpenAI-compatible API
- Support step-3.5-flash (256K context, reasoning model)
- Support step-2-16k (1T parameters, 16K context)
- Support step-1-32k (100B parameters, 32K context)
Pricing based on official StepFun documentation (converted from CNY to USD):
- step-3.5-flash: bash.096 input / bash.288 output / bash.019 cache
- step-2-16k: .21 input / 6.44 output / .04 cache
- step-1-32k: .05 input / .59 output / bash.41 cache
Note: Logo not included as it is optional per contributing guidelines.
A default logo will be served by models.dev API instead.
All model definitions follow the official schema.
Fixes anomalyco/opencode#11760
Fixes anomalyco/opencode#11960
StepFun API: https://api.stepfun.com/v1
Documentation: https://platform.stepfun.com/docs/zh/pricing/details
2026-02-13 16:28:07 +08:00
城二
58de85c2e8
feat(minimax): Add interleaved configuration
...
- Add the reasoning_content field configuration to the minimax model.
- Update the configuration files for m2.5 and m2.5-lightning.
2026-02-13 16:20:45 +08:00
城二
f87ecffbf0
feat(minimax): Update m2.5 model name and price
...
- Change the model name from "lightning" to "highspeed"
- Adjust the input/output and cache read/write prices
2026-02-13 16:18:22 +08:00
niushuai1991
6dfb2f9c83
add kuae cloud coding plan
2026-02-13 15:02:52 +08:00
城二
ee8c1bce7d
feat(models): Added minimax model profile
2026-02-13 14:15:44 +08:00
Zain Hasan
a2dd10d09d
Update output limit in GLM-5 configuration
2026-02-12 22:12:56 -08:00
Zain Hasan
f6cfc2ebd2
try remove reasoning
2026-02-12 21:54:02 -08:00
Zain Hasan
3192856cc3
finx glm 5 settings
2026-02-12 21:46:13 -08:00
Aiden Cline
5507f42604
Merge pull request #874 from 888-wzk/feature/chenger_20260128
...
feat(z-ai): New glm-5 model configuration file
2026-02-12 22:56:36 -06:00
Aiden Cline
995aabf33f
Merge pull request #894 from fhennerkes/dev
...
Poe: fix formatting, naming and update outputs
2026-02-12 22:56:08 -06:00
城二
fc5c3613eb
feat(glm-5): Add reasoning_content field
2026-02-13 11:37:33 +08:00
fhennerkes
72f10a52e1
poe: update model names to use display_name
2026-02-12 19:08:49 -08:00
fhennerkes
e944012d95
poe: small fixes (formatting and reasoning)
2026-02-12 18:55:19 -08:00
Aiden Cline
e117f37d4e
Merge pull request #892 from pat-baseten/add-kimi-2.5-baseten
...
Add Kimi K2.5 model for Baseten
2026-02-12 17:45:46 -06:00
Pat
b0d71629fe
Add Kimi K2.5 model for Baseten
2026-02-12 16:39:56 -06:00
Aiden Cline
aa5e8634b2
Merge pull request #890 from cfal/fireworks-glm-5
...
fireworks: add GLM-5
2026-02-12 16:15:18 -06:00
Aiden Cline
7acba1db3f
Merge pull request #891 from lucianjon/feat/openrouter-minimax-m2.5
...
feat(openrouter/minimax): add minimax-m2.5
2026-02-12 16:15:08 -06:00
Aiden Cline
c5095973e1
Merge pull request #886 from brentdurksen/dev
...
feat(amazon-bedrock): add Writer Palmyra X4 and X5 models
2026-02-12 16:14:58 -06:00
Aiden Cline
5b8797cf89
Merge pull request #889 from Daltonganger/feat/nano-gpt-add-minimax-m2.5-official
...
feat(nano-gpt): add MiniMax M2.5 route alongside official variant
2026-02-12 16:14:36 -06:00
Daltonganger
f1317184b5
Enable reasoning and add interleaved field in TOML
2026-02-12 23:07:12 +01:00
Lucian Jones
7ba286c7f2
feat(openrouter/minimax): add minimax-m2.5
2026-02-13 10:59:30 +13:00
cfal
dec532b3b0
providers/fireworks-ai/models/accounts/fireworks/models/glm-5.toml: add GLM-5 to fireworks
2026-02-13 01:45:51 +04:00
Aiden Cline
ccff680988
Merge pull request #864 from sylviezhang37/vercel-model-file-gen-script
...
feat(provider): Vercel model file generation and update script
2026-02-12 15:37:17 -06:00
Aiden Cline
1b63e4670e
Merge pull request #887 from PandaSt0rm/add-minimax-m2-5-support
...
Add MiniMax-M2.5 across minimax and coding-plan providers
2026-02-12 15:36:56 -06:00
Aiden Cline
d5cbd6fb5d
Merge pull request #888 from spiffytech/dev
...
Add Ollama Cloud support for Minimax 2.5
2026-02-12 15:36:14 -06:00
Ruben Beuker
e8f2f6b14f
feat(nano-gpt): add MiniMax M2.5 route and align official variant
2026-02-12 22:35:23 +01:00
spiffytech
e92fe6e9d7
Added Ollama Cloud support for Minimax 2.5
2026-02-12 16:24:11 -05:00
PandaSt0rm
7a32f17911
add MiniMax-M2.5 configs across minimax providers
2026-02-12 23:22:52 +02:00
Brent Durksen
57db1db84f
feat(amazon-bedrock): add Writer Palmyra X4 and X5 models
...
Add two new Writer AI models to the Amazon Bedrock provider:
- writer.palmyra-x4-v1:0 (Palmyra X4): 128K context, 8K output,
reasoning and tool calling, $2.50/$10 per M tokens (input/output)
- writer.palmyra-x5-v1:0 (Palmyra X5): 1M context, 8K output,
reasoning and tool calling, $0.60/$6 per M tokens (input/output)
Both models support text-only input/output modalities and are
closed-weight.
Also adds the 'palmyra' family to the ModelFamilyValues enum in
packages/core/src/family.ts to support validation.
2026-02-12 13:43:32 -07:00
Aiden Cline
ba91bb6612
Merge pull request #883 from ryanskidmore/ryanskidmore/cloudflare-ai-gateway-bump-opus-4-6-limits
...
cloudflare-ai-gateway: bump Opus 4.6 output limit to 128k
2026-02-12 13:01:38 -06:00
Ryan Skidmore
9761d0ef87
cloudflare-ai-gateway: bump Opus 4.6 output limit to 128k
2026-02-12 12:39:00 -06:00
Aiden Cline
98be9a2078
fix: family
2026-02-12 12:27:16 -06:00
Dax Raad
4aa17d26cb
feat(openai): add gpt-5.3-codex-spark model
2026-02-12 13:24:43 -05:00
Aiden Cline
7f96ee576a
Merge pull request #880 from Daltonganger/feat/nano-gpt-glm5-original-models
...
feat(nano-gpt): add GLM 5 original model variants
2026-02-12 12:18:46 -06:00
Aiden Cline
bd5ce80e56
Merge pull request #882 from Daltonganger/feat/nano-gpt-add-minimax-m2.5-official
...
feat(nano-gpt): add MiniMax M2.5 Official model
2026-02-12 12:18:37 -06:00
Daltonganger
ed2af4ad45
feat(nano-gpt): add MiniMax M2.5 Official model
2026-02-12 18:07:52 +01:00
Aiden Cline
ac0868c886
Merge pull request #881 from Alex-wuhu/dev
...
add minmax-2.5 on novita
2026-02-12 10:32:50 -06:00
Aiden Cline
fdd13245cc
Revert "feat(github-copilot): add gpt-5.3-codex model ( #857 )"
...
This reverts commit 27abb8a570 .
2026-02-12 10:32:15 -06:00
Alex
37c77c58ad
Merge branch 'anomalyco:dev' into dev
2026-02-13 00:27:45 +08:00
Alex-wuhu
62ee8129e6
add minimax-m2.5 on novita
2026-02-13 00:23:06 +08:00
Daltonganger
4bc6f07570
fix(nano-gpt): correct GLM-5 dates to 2026-02-11
2026-02-12 17:22:18 +01:00
Aiden Cline
bc0336c8ec
Merge pull request #878 from cantalupo555/feat/add-openrouter-stepfun-step-3.5-flash
...
feat: add StepFun Step 3.5 Flash on OpenRouter
2026-02-12 10:13:17 -06:00
Aiden Cline
dd78db4dc6
Merge pull request #879 from amankalra172/add-stackit-provider
...
fix: reorganize STACKIT models with organization prefixes
2026-02-12 10:12:48 -06:00
Daltonganger
4fd32c741d
refactor(nano-gpt): consolidate z-ai GLM models under zai-org
2026-02-12 17:11:50 +01:00
Frank
c78ca7c132
update zen models
2026-02-12 11:05:41 -05:00
Alex
2a99397516
add GLM5 on novita ( #877 )
2026-02-12 11:03:42 -05:00
Frank
554440be4f
update zen models
2026-02-12 11:01:51 -05:00
Daltonganger
eb52a76d11
feat(nano-gpt): add GLM 5 original model variants
2026-02-12 16:48:59 +01:00
amankalra172
c9a7f6c814
fix: reorganize STACKIT models with organization prefixes and correct pricing
...
- Move models to organization subfolders (Qwen/, cortecs/, google/, etc.)
- Update pricing from EUR to USD (1.09 conversion rate)
- Fix GPT-OSS context limit to 131K tokens
- Add architectural family classifications
- Verify tool_call settings for all models
2026-02-12 14:00:45 +01:00
cantalupo555
e640802d34
feat: add StepFun Step 3.5 Flash (free) on OpenRouter
2026-02-12 08:44:15 -03:00
cantalupo555
c226863912
feat: add StepFun Step 3.5 Flash on OpenRouter
2026-02-12 08:42:37 -03:00
Alex-wuhu
8bcd634743
add GLM5 on novita
2026-02-12 16:26:05 +08:00
Aiden Cline
812cd1763a
Merge pull request #873 from juls0730/dev
...
Create cerebras/llama3.1-8b.toml
2026-02-12 00:40:48 -06:00
城二
11f4ae568e
feat(z-ai): New glm-5 model configuration file
2026-02-12 11:23:22 +08:00
juls0730
9f1629a26a
Create cerebras/llama3.1-8b.toml
2026-02-11 21:01:08 -06:00
Yunfei He
27abb8a570
feat(github-copilot): add gpt-5.3-codex model ( #857 )
...
* feat(github-copilot): add gpt-5.3-codex model
* fix(github-copilot): align gpt-5.3-codex release metadata
2026-02-11 21:53:36 -05:00
Aiden Cline
2aa4a2290e
Merge pull request #868 from dpuyosa/venice
...
Venice: Add GLM-5 model
2026-02-11 19:49:16 -06:00
Aiden Cline
c58b36c605
Merge pull request #872 from Track07-cda/openrouter-glm5
...
OpenRouter: Add GLM-5 and remove Pony Alpha
2026-02-11 19:49:07 -06:00
Aiden Cline
e91dbd1fc4
Merge pull request #870 from spiffytech/dev
...
Add Ollama Cloud support for GLM-5
2026-02-11 19:39:23 -06:00
Track07-cda
8924ee3092
feat(openrouter): add GLM-5 and remove Pony Alpha
...
Add the Z-AI GLM-5 model definition to the OpenRouter provider and
remove the deprecated Pony Alpha model.
2026-02-12 09:38:12 +08:00
spiffytech
fc5b6533d1
Added Ollama Cloud support for GLM-5
2026-02-11 19:57:46 -05:00
Aiden Cline
7a760e3a4f
Merge pull request #871 from Kunde21/synthetic_k2_5_nvfp4
...
Synthetic: Add Kimi -2 5 in NVFP4 remove GLM-4.5
2026-02-11 18:53:56 -06:00
Chad Kunde
38835801b1
synthetic: deprecate GLM-4.5
...
Model removed from models list as of 12 Feb 2026
2026-02-12 07:27:24 +07:00
Chad Kunde
81103438a3
synthetic: Add NVFP4 variant of Kimi K2.5
2026-02-12 07:25:41 +07:00
dpuyosa
0b0b36eb45
[venice] Add GLM-5 model with 198K context window
...
- Add ZAI-ORG GLM-5 model configuration to Venice provider
- Supports reasoning, tool calls, structured output
- Text in/out: 198K context, 49.5K output tokens
2026-02-11 22:53:23 +01:00
Aiden Cline
b18b73f0a0
Revert "fix: openai advertises a 400k context window, that is just the sum of max input + max output, so real context window is 272k"
...
This reverts commit ea276d57a7 .
2026-02-11 15:24:43 -06:00
Aiden Cline
3cd48b273a
Merge pull request #867 from AnishShah1803/nano-gpt/add-glm-5-models
...
Add GLM 5 to NanoGPT models list
2026-02-11 14:47:57 -06:00
twisted
890992b8ef
update release date
2026-02-11 20:37:59 +00:00
twisted
a843d84d77
Add GLM 5 to NanoGPT models list
2026-02-11 20:35:41 +00:00
Aiden Cline
d89897d07e
Merge pull request #866 from Sczr0/dev
...
Update pricing for ZAI GLM-5
2026-02-11 14:35:23 -06:00
Aiden Cline
1b26792073
fix zai
2026-02-11 14:34:43 -06:00
Aiden Cline
6a0da0a91d
Revert "Fixed ZAI GLM-5 pricing to free (0 cost)"
...
This reverts commit 79d1222e3c .
2026-02-11 14:33:39 -06:00
opencode-agent[bot]
79d1222e3c
Fixed ZAI GLM-5 pricing to free (0 cost)
...
Co-authored-by: rekram1-node <rekram1-node@users.noreply.github.com >
2026-02-11 20:14:03 +00:00
Sylvie Zhang
1ad45b44c2
add readme
2026-02-11 12:13:38 -08:00
弦塔_
45ba8066df
Update cost parameters in glm-5.toml
2026-02-12 04:07:57 +08:00
弦塔_
4861f14a65
Update cost parameters in glm-5.toml
2026-02-12 04:07:29 +08:00
弦塔_
c83b4e22bd
Update cost parameters in glm-5.toml
2026-02-12 03:56:40 +08:00
Sylvie Zhang
44c1ed5aeb
additional data cleaning logic
2026-02-11 11:48:32 -08:00
Sylvie Zhang
28c09d83a0
add fallback logic
2026-02-11 11:48:32 -08:00
Sylvie Zhang
ea41cbc4ba
draft script
2026-02-11 11:48:32 -08:00
Aiden Cline
c893ac5f9d
Merge pull request #863 from AnishShah1803/nano-gpt/update-Kimi-K2-5-models
...
Add Kimi K2.5 models to NanoGPT provider
2026-02-11 13:41:43 -06:00
twisted
8e10faf38a
set reasoning to true for kimi k2.5
2026-02-11 19:33:12 +00:00
Aiden Cline
4c3a17fbe8
Merge pull request #859 from friendliai/minpeter/add-glm5-friendli
...
Add zai-org/GLM-5 model to Friendli provider
2026-02-11 13:14:58 -06:00
twisted
f7d997e4f0
fix last_updated
2026-02-11 19:08:48 +00:00
twisted
6961c57c86
make open_weights set to true
2026-02-11 19:07:35 +00:00
minpeter
e44307c27c
Add interleaved reasoning_content to GLM-4.7 and MiniMax-M2.1
...
Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com >
2026-02-12 04:01:29 +09:00
minpeter
b9f7907150
Merge remote-tracking branch 'origin/dev' into minpeter/add-glm5-friendli
2026-02-12 04:00:31 +09:00
minpeter
bf0ee2a3eb
Add interleaved reasoning_content field for GLM-5
...
GLM models use interleaved reasoning via the reasoning_content field with OpenAI-compatible providers.
Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com >
2026-02-12 03:58:58 +09:00
Aiden Cline
b9a73edcdc
Merge pull request #862 from hanouticelina/feat/huggingface-glm-5
...
feat(huggingface): add GLM-5 for Hugging Face provider
2026-02-11 12:45:43 -06:00
Aiden Cline
7ff3be2dab
fix: github copilot model discrepencies
2026-02-11 12:40:36 -06:00
twisted
406f82e09f
Add Kimi K2.5 models to NanoGPT provider
2026-02-11 18:40:09 +00:00
Celina Hanouti
c8fa26624d
add GLM-5 for hugging face provider
2026-02-11 19:39:35 +01:00
Aiden Cline
ae31005ef3
Merge pull request #830 from amankalra172/add-stackit-provider
...
feat: add STACKIT provider with 8 AI models
2026-02-11 12:20:47 -06:00
Aiden Cline
0aa7c9f3c2
Merge pull request #852 from zainhas/dev
...
[Together AI] update output token length to match context length
2026-02-11 12:20:09 -06:00
Aiden Cline
40be65f301
Merge pull request #853 from captain1379/feat/jiekou
...
Add new models for Jiekou.AI
2026-02-11 12:19:40 -06:00
Aiden Cline
49ab2c0a48
feat: add glm 5 to zai, zhipuai, and zai coding plan
2026-02-11 12:17:49 -06:00
minpeter
7fd96a6c9d
Add zai-org/GLM-5 model to Friendli provider
...
Add GLM-5 model configuration with reasoning, tool calling, and structured output support. Update family pattern inference in generate script to recognize GLM-5 models.
Co-Authored-By: Claude Haiku 4.5 <noreply@anthropic.com >
2026-02-12 03:07:25 +09:00
Aiden Cline
fc4027fe98
Merge pull request #856 from josetorres1/add-bedrock-zai-minimax-models
...
Add GLM 4.7 Family and MiniMax M2.1 to Amazon Bedrock
2026-02-11 11:25:21 -06:00
Aiden Cline
b260564060
Merge pull request #855 from dihan-dff-user/dev
...
Add ZAI coding plan GLM-5 model
2026-02-11 11:24:47 -06:00
opencode-agent[bot]
ecd5927bed
Removed knowledge field from GLM-5 config
...
Co-authored-by: rekram1-node <rekram1-node@users.noreply.github.com >
2026-02-11 17:23:31 +00:00
Jose Torres
490381f370
Add GLM 4.7 family and MiniMax M2.1 to Amazon Bedrock provider
...
- Add GLM-4.7 (zai.glm-4.7): /bin/zsh.60/.20 per 1M tokens
- Add GLM-4.7-Flash (zai.glm-4.7-flash): /bin/zsh.07//bin/zsh.40 per 1M tokens
- Add MiniMax M2.1 (minimax.minimax-m2.1): /bin/zsh.30/.20 per 1M tokens
Pricing sources:
- AWS Bedrock pricing page: https://aws.amazon.com/bedrock/pricing/
- Model IDs confirmed via AWS console/CLI
Related to GH issue #835
2026-02-11 09:51:18 -06:00
dihan
621687e457
Add ZAI coding plan GLM-5 model
2026-02-11 20:38:40 +05:30
captain1379
d5a3bfad90
feat: add new models for Jiekou.AI
...
- Introduced `claude-opus-4-6`, `qwen3-coder-next` and `gpt-5.1` models with detailed configurations.
- Removed deprecated `qwen2.5-vl-72b-instruct` model.
- Implemented a new script for generating model configurations.
2026-02-11 15:19:40 +08:00
Zain Hasan
310bc174fa
update output token length to match context length
2026-02-10 23:05:23 -08:00
Aiden Cline
9fb1233073
Merge pull request #848 from BlockListed/cortecs-glm-models
...
Add supported z.ai GLM models to cortecs
2026-02-10 20:04:21 -06:00
Aiden Cline
029f13545b
Merge pull request #849 from BlockListed/cortecs-minimax-models
...
Add MiniMax models to cortecs
2026-02-10 20:04:12 -06:00
Aiden Cline
31b1acba5f
Merge pull request #850 from cgilly2fast/dev
...
feat(firmware): kimi and glm models
2026-02-10 20:04:00 -06:00
Aiden Cline
120881916e
Merge pull request #851 from anomalyco/fix-models
...
fix: openai advertises a 400k context window, that is just the sum of max input + max output, so real context window is 272k
2026-02-10 20:03:48 -06:00
Aiden Cline
ea276d57a7
fix: openai advertises a 400k context window, that is just the sum of max input + max output, so real context window is 272k
2026-02-10 19:02:33 -06:00
Colby Gilbert
05d940ab8d
feat(firmware): kimi and glm models
2026-02-10 14:25:46 -08:00
BlockListed
7d250ea857
cortecs add minimax models
2026-02-10 23:13:39 +01:00
BlockListed
82d72e85f8
add supported z.ai GLM models to cortecs
2026-02-10 22:58:22 +01:00
Aiden Cline
995934de32
Merge pull request #847 from rubenandre/add-bedrock-moonshotai-kimi-k2.5
...
add moonshotai kimi K2.5 to amazon-bedrock provider
2026-02-10 10:43:13 -06:00
Aiden Cline
d10392573e
Merge pull request #845 from cgilly2fast/dev
...
fix(firmware): remove unsupported model and fix name of gpt oss 20b
2026-02-10 10:04:15 -06:00
Aiden Cline
21c3c1b8e1
Merge pull request #846 from dpuyosa/venice
...
Venice: Enable reasoning for GLM-4.7-Flash model
2026-02-10 10:04:06 -06:00
Rúben Silva
4680aaedc3
add moonshotai kimi K2.5 to amazon-bedrock provider
2026-02-10 15:21:50 +00:00
dpuyosa
a9f3ad978d
[venice] Enable reasoning for GLM-4.7-Flash model
2026-02-10 10:11:48 +01:00
Colby Gilbert
de75687ea1
fix(firmware): remove unsupported model and fix name of spt oss 20b
2026-02-09 23:01:40 -08:00
Aiden Cline
539cc930c4
Merge pull request #843 from cgilly2fast/dev
...
chore: clean up firmware available models
2026-02-09 18:40:24 -06:00
Colby Gilbert
31f3c63acc
fix(firmware): remove reasoning from anthropic and deepseek models
2026-02-09 16:07:44 -08:00
Colby Gilbert
b019787ad8
chore: clean up firmware available models
2026-02-09 16:03:09 -08:00
Aiden Cline
e41fca18a2
Merge pull request #842 from riccardogiorato/dev
...
fix: Increase output limit to match context for kimi K2.5 on Together
2026-02-09 17:12:07 -06:00
Riccardo Giorato
74163f7314
Increase output limit to match context
...
Update providers/togetherai/models/moonshotai/Kimi-K2.5.toml to set [limit].output from 32_768 to 262_144. This aligns the output token limit with the context size (262_144) to avoid premature truncation and allow full-length responses.
2026-02-09 22:39:34 +01:00
Aiden Cline
7b763695fd
Merge pull request #839 from shelvick/add-azure-kimi-k2.5
...
Add Azure Kimi-K2.5 model
2026-02-09 14:15:33 -06:00
Aiden Cline
686b47d01e
Merge pull request #840 from shelvick/add-azure-claude-opus-4-6
...
Add Azure Claude Opus 4.6 model
2026-02-09 14:15:16 -06:00
Scott Helvick
46f0726d7f
Add Azure Claude Opus 4.6 model
2026-02-09 20:07:31 +00:00
Scott Helvick
3c14600fc6
Add Azure Kimi-K2.5 model
2026-02-09 19:49:02 +00:00
Aiden Cline
721c025af1
Merge pull request #836 from PeppeRu96/feat/add-deepinfra-claude
...
feat: add DeepInfra Claude Opus 4 and Claude Sonnet 3.7 (latest) models
2026-02-09 12:29:34 -06:00
Aiden Cline
c591f9b213
Merge pull request #837 from PeppeRu96/feat/add-deepinfra-deepseek
...
feat: add DeepInfra DeepSeek models
2026-02-09 12:23:02 -06:00
Aiden Cline
57580b28d3
Merge pull request #765 from captain1379/feat/jiekou
...
feat: add Jiekou.AI provider
2026-02-09 12:22:21 -06:00
Giuseppe Ruggeri
11e92f093f
fix: fix price for DeepInfra DeepSeek-V3.2
2026-02-09 14:32:34 +01:00
Giuseppe Ruggeri
17cf21ba46
feat: add DeepInfra DeepSeek models
2026-02-09 14:29:02 +01:00
Giuseppe Ruggeri
60a3f09b8e
fix: update deepinfra/claude-3-7-sonnet-latest family field
2026-02-09 14:02:52 +01:00
Giuseppe Ruggeri
6f907bce35
feat: add DeepInfra Claude Opus 4 and Claude Sonnet 3.7 (latest) models
2026-02-09 13:56:16 +01:00
Frank
1f20d47ef5
update zen models
2026-02-08 21:43:43 -05:00
Aiden Cline
d5c23c9c95
Merge pull request #827 from modpotato/dev
...
fix: rename glm 5 stealth from 'Stealth' to 'Pony Alpha' + remove status
2026-02-08 14:03:59 -06:00
Aiden Cline
125abf1a21
Merge pull request #829 from 888-wzk/feature/chenger_20260128
...
fix: Update model configurations to adjust reasoning and interleaved …
2026-02-08 14:03:44 -06:00
Aiden Cline
38ccea666f
Merge pull request #831 from spiffytech/dev
...
Add Ollama Cloud support for qwen3-coder-next
2026-02-08 14:03:28 -06:00
Frank
42ca5faeb8
sync
2026-02-08 14:21:34 -05:00
spiffytech
bcd9e3dba1
Added Ollama Cloud support for qwen3-coder-next
2026-02-08 13:36:54 -05:00
amankalra172
7504dc2947
feat: add STACKIT provider with 8 AI models
...
Add STACKIT as a new provider with complete model specifications:
Chat Models:
- Llama 3.1 8B Instruct FP8
- Llama 3.3 70B Instruct FP8
- GPT-OSS 120B
- Mistral Nemo Instruct 2407 FP8
- Gemma 3 27B (multimodal)
- Qwen3-VL 235B (vision-language)
Embedding Models:
- E5 Mistral 7B
- Qwen3-VL Embedding 8B (multimodal)
All models include:
- Proper schema compliance (attachment, reasoning, tool_call, etc.)
- Pricing in USD per million tokens
- Context limits and modalities
- Official STACKIT logo with currentColor support
STACKIT is a German sovereign cloud provider offering OpenAI-compatible
AI model serving with open-source models.
2026-02-08 12:35:51 +01:00
城二
67bddb6b61
Merge branch 'dev' of https://github.com/888-wzk/models.dev into feature/chenger_20260128
2026-02-08 11:23:44 +08:00
城二
77330e78c6
fix: Update model configurations to adjust reasoning and interleaved fields
2026-02-08 11:21:58 +08:00
mod
e55a05f6a6
Merge branch 'anomalyco:dev' into dev
2026-02-07 00:43:00 -05:00
mod
e5ce677899
fix: rename glm 5 stealth from 'Stealth' to 'Pony Alpha'
2026-02-07 00:42:50 -05:00
Aiden Cline
e1747322ad
Merge pull request #826 from modpotato/dev
...
add pony alpha (glm 5 stealth)
2026-02-06 23:21:17 -06:00
Aiden Cline
9303c7be2e
Merge pull request #825 from cantalupo555/feat/add-openrouter-mimo-v2-flash
...
feat: add Xiaomi MiMo-V2-Flash on OpenRouter
2026-02-06 16:43:34 -06:00
John Doe
204eb52c0d
feat: pony alpha (glm 5 demo) on openrouter
2026-02-06 21:06:54 +00:00
John Doe
a2abd136f5
feat: pony alpha (glm 5 demo) on openrouter
2026-02-06 21:02:40 +00:00
Aiden Cline
ea6e487e77
fix: change anthropic default to 200k instead of 1M since not everyone can access the 1M
2026-02-06 13:46:09 -06:00
Aiden Cline
1033ee450c
Merge pull request #821 from 888-wzk/feature/chenger_20260128
...
Added Claude Opus 4.6 model configuration file
2026-02-06 11:00:49 -06:00
Aiden Cline
de8e46b2ab
Merge pull request #823 from dpuyosa/venice
...
Venice: Tweak model generation script
2026-02-06 11:00:37 -06:00
Aiden Cline
8181d97317
Merge pull request #824 from vglafirov/feat/gitlab-opus-4-6
...
feat(gitlab): add Claude Opus 4.6 model (duo-chat-opus-4-6)
2026-02-06 11:00:11 -06:00
Vladimir Glafirov
a5c9640163
feat(gitlab): add Claude Opus 4.6 model (duo-chat-opus-4-6)
...
Add the newly released Claude Opus 4.6 model for GitLab Duo Agentic Chat.
Related:
- AI Gateway MR: https://gitlab.com/gitlab-org/modelops/applied-ml/code-suggestions/ai-assist/-/merge_requests/4492
2026-02-06 17:04:17 +01:00
cantalupo555
c220f2a790
feat: add Xiaomi MiMo-V2-Flash on OpenRouter
2026-02-06 12:52:35 -03:00
Frank
c88c849e5a
Merge pull request #822 from imdevarsh/imdevarsh/openrouter-opus-4.6
...
feat(openrouter): add claude opus 4.6 to openrouter models list
2026-02-06 10:07:38 -05:00
dpuyosa
75ff468a9a
[venice] Refactor model generation with privacy field
...
- Add optional privacy field to ModelSpec schema
- Use privacy field to determine open_weights capability
- Preserve existing output token limit when smaller than proposed
2026-02-06 13:18:47 +01:00
Devarsh
5b9186f6a9
feat(openrouter): add claude opus 4.6 to openrouter models list
2026-02-06 18:06:33 +13:00
城二
974713311b
feat: Added Claude Opus 4.6 model configuration file
2026-02-06 11:11:19 +08:00
Aiden Cline
2d143f96d1
Merge pull request #813 from cgilly2fast/dev
...
feat: add opus 4.6 to firmware provider
2026-02-05 16:25:47 -06:00
Aiden Cline
4c4cd139f8
Merge pull request #812 from fhennerkes/dev
...
poe: add Claude Opus 4.6 model
2026-02-05 16:25:29 -06:00
Aiden Cline
22688b1260
Merge pull request #816 from markusylisiurunen/add-eu-opus-4.6
...
Add the missing EU variant back for Opus 4.6 on AWS Bedrock
2026-02-05 16:23:54 -06:00
Aiden Cline
13631caba9
Merge pull request #817 from dpuyosa/venice
...
Venice: Add Claude Opus 4.6 and GLM 4.7 models
2026-02-05 16:21:41 -06:00
dpuyosa
7e901e93bf
[venice] Add Claude Opus 4.6 and GLM 4.7 models
...
- Add claude-opus-4.6 model configuration for Venice provider
- Add zai-org-glm-4.7-flash model configuration for Venice provider
2026-02-05 22:44:27 +01:00
Markus Ylisiurunen
ace9626163
also fix pricing for opus 4.5
2026-02-05 23:25:25 +02:00
Markus Ylisiurunen
2bd869959b
fix pricing
2026-02-05 23:12:40 +02:00
Markus Ylisiurunen
efecbc137c
Add EU variant for Opus 4.6
2026-02-05 23:03:01 +02:00
Colby Gilbert
cae3f84930
Merge branch 'anomalyco:dev' into dev
2026-02-05 12:49:38 -08:00
Colby Gilbert
37cdea639f
feat: add opus 4.6
2026-02-05 12:49:19 -08:00
Ryan Vogel
2c67792e4b
Merge pull request #811 from anomalyco/add-claude-opus-4-6
...
Fix Claude Opus 4.6 model IDs and remove incorrect variants
2026-02-05 15:49:03 -05:00
Ryan Vogel
24addedada
Fix Vertex AI model ID to claude-opus-4-6@default
2026-02-05 15:46:07 -05:00
Ryan Vogel
dc9f404dbb
Condense AGENTS.md model configuration section
2026-02-05 15:44:00 -05:00
Ryan Vogel
db2212ab9f
Update AGENTS.md with model configuration learnings
2026-02-05 15:42:47 -05:00
Ryan Vogel
e2777a44ed
Fix Vertex AI model ID: claude-opus-4-6@default -> claude-opus-4-6
2026-02-05 15:41:53 -05:00
fhennerkes
c0f0394f67
poe: add Claude Opus 4.6 model
2026-02-05 12:39:00 -08:00
Ryan Vogel
a762f47461
Fix model IDs: remove unannounced dated alias, remove EU Bedrock, fix Bedrock ID (v1:0 -> v1), fix Vertex ID (@20260205 -> @default)
...
Fixes #809
2026-02-05 15:38:28 -05:00
Aiden Cline
768f841f79
Merge pull request #781 from Dagnan/add-glm-4.7-flash-deepinfra
...
feat(deepinfra): add GLM-4.7-Flash model
2026-02-05 14:34:27 -06:00
Aiden Cline
00f239a852
Merge pull request #770 from jerilynzheng/feat/add-vercel-models-jan-30
...
vercel: add new models and interleaved support
2026-02-05 14:19:20 -06:00
Michel Pigassou
69f72041e1
Added missing interleaved/reasoning_content for GLM 4.7-Flash
2026-02-05 21:16:25 +01:00
Aiden Cline
43e98540ec
fix: output limit for opus 4.6 on gh copilot
2026-02-05 14:15:56 -06:00
jerilynzheng
f0854ab7b8
vercel: add interleaved = true for confirmed models
...
Add interleaved reasoning support to models confirmed by other providers:
- Claude: 3.7-sonnet, haiku-4.5, opus-4/4.1/4.5/4.6, sonnet-4/4.5
- DeepSeek: R1, V3.2-thinking
- MiniMax: M2, M2.1
- Kimi: K2-thinking, K2-thinking-turbo, K2.5
- GLM: 4.5, 4.6, 4.7, 4.7-flashx
Co-Authored-By: Claude Opus 4.5 <noreply@anthropic.com >
2026-02-05 12:14:02 -08:00
Aiden Cline
a05a47097f
Merge pull request #799 from iamanishx/deepinfra-kimi
...
feat: added support for kimi k2.5 (deepinfra)
2026-02-05 14:08:08 -06:00
jerilynzheng
ed96ac7c74
fix: update Claude Opus 4.6 knowledge cutoff to 2025-05
...
Co-Authored-By: Claude Opus 4.5 <noreply@anthropic.com >
2026-02-05 12:03:40 -08:00
jerilynzheng
3985ee8556
vercel: add Claude Opus 4.6
...
Add anthropic/claude-opus-4.6 from Vercel AI Gateway:
- 1M context window, 128K output
- $5.00/$25.00 per 1M tokens (input/output)
- Supports vision, reasoning, and tool use
Co-Authored-By: Claude Opus 4.5 <noreply@anthropic.com >
2026-02-05 12:03:11 -08:00
Aiden Cline
53c150347f
Merge pull request #808 from stickyburn/chutes-qwen3-coder-next
...
chore: add qwen3-coder-next for chutes.ai
2026-02-05 14:01:46 -06:00
Aiden Cline
1b4b15d73b
Merge pull request #807 from smrdotgg/patch-1
...
Fix release and last updated dates for GPT-5.3 Codex
2026-02-05 14:01:37 -06:00
stickyburn
535b1afe4f
chore: add qwen3 coder next for chutes.ai
2026-02-05 14:41:59 -05:00
Mathis
e239c4d51e
Add configuration for Claude Opus 4.6 model ( #806 )
2026-02-05 14:24:07 -05:00
imanishx
3dddd7b5e3
fix: interleaved opt added
...
Signed-off-by: imanishx <manishbiswal754@gmail.com >
2026-02-05 19:05:25 +00:00
Semere Tereffe
941487c8d4
Fix release and last updated dates for GPT-5.3 Codex
2026-02-05 21:52:29 +03:00
Ryan Vogel
ce3bbe64a3
Update context window to 1M tokens for all Claude Opus 4.6 models
2026-02-05 13:39:48 -05:00
Aiden Cline
183dd5c4f3
Merge pull request #803 from anomalyco/add-claude-opus-4-6
...
Add Claude Opus 4.6 model
2026-02-05 12:39:29 -06:00
Ryan Vogel
b4733df6b7
Merge branch 'dev' into add-claude-opus-4-6
2026-02-05 13:38:54 -05:00
Ryan Vogel
921ec8f8fc
Add cost.context_over_200k long context pricing to all Claude Opus 4.6 models
2026-02-05 13:36:52 -05:00
Aiden Cline
6c28f3974c
Merge pull request #804 from rexdotsh/feat/add-anthropic-opus-4-6
...
feat: add opus 4.6
2026-02-05 12:34:34 -06:00
Aiden Cline
8b3a813da9
Merge pull request #802 from dmmulroy/cloudflare-opus-4-6
...
cloudflare-ai-gateway: add claude opus 4.6
2026-02-05 12:34:00 -06:00
rexdotsh
10297c8490
feat: add opus 4.6
2026-02-05 23:58:13 +05:30
Ryan Vogel
61c8d50df2
Add Claude Opus 4.6 model across Anthropic, Bedrock, and Vertex AI providers
2026-02-05 13:26:28 -05:00
Dax Raad
4faf622d1e
add gpt-5.3-codex.toml
2026-02-05 13:26:24 -05:00
Dillon Mulroy
89db412884
cloudflare-ai-gateway: add claude opus 4.6
2026-02-05 13:24:36 -05:00
Frank
107b285e1c
update zen models
2026-02-05 13:14:06 -05:00
Frank
8bd58cf186
update zen models
2026-02-05 13:04:56 -05:00
Aiden Cline
9573a8fccc
Merge pull request #796 from manascb1344/nebius-token-factory-models
...
feat: add Nebius Token Factory models
2026-02-05 11:45:09 -06:00
Aiden Cline
0153e6408a
Merge pull request #792 from Alex-wuhu/dev
...
feat: add deepseek OCR model configuration and Qwen3 Coder Next model…
2026-02-05 11:44:28 -06:00
imanishx
8cb19035ff
feat: added tomal for kini k2.4 (deepinfra)
...
Signed-off-by: imanishx <manishbiswal754@gmail.com >
2026-02-05 09:08:33 +00:00
captain1379
7c0e1e142f
fix: removed old models
2026-02-05 14:02:59 +08:00
captain1379
7aeca69c4e
fix: fix logo
2026-02-05 13:44:38 +08:00
Aiden Cline
cfde47ca60
Revert "Update Amazon Bedrock models to add cross-region inference and remove deprecated models"
...
This reverts commit bc58036964 .
2026-02-04 12:11:29 -06:00
Aiden Cline
b01c07a3d0
Merge pull request #793 from zainhas/patch-1
...
[fix] Rename model to 'Qwen3 Coder Next FP8'
2026-02-04 10:33:22 -06:00
Aiden Cline
444c3071ee
Merge pull request #795 from riccardogiorato/dev
...
remove wrongly typed Kimi-K2-5.toml
2026-02-04 10:32:27 -06:00
manascb1344
42a79c717d
feat(nebius): update Meta-Llama, NVIDIA models and mark deprecated
...
- Update Llama-3.3-70B-Instruct (Base & Fast) with new pricing
- Mark Llama-3.1-405B-Instruct as deprecated (no longer available)
- Update Llama-3.1-Nemotron-Ultra-253B-v1 with new pricing
- Mark DeepSeek-V3 as deprecated (replaced by V3.2 and V3-0324)
2026-02-04 20:18:37 +05:30
manascb1344
578df73ffb
feat(nebius): update Z.ai, OpenAI, Moonshot AI, and NousResearch models
...
- Update GLM-4.5 and GLM-4.5-Air with new pricing
- Update gpt-oss-120b and gpt-oss-20b with new pricing and features
- Update Kimi-K2-Instruct with new pricing and multimodal support
- Update Hermes-4-405B and Hermes-4-70B with new pricing
2026-02-04 20:17:51 +05:30
manascb1344
2639e20a97
feat(nebius): add new models from Z.ai, Moonshot AI, Meta, and NVIDIA
...
- Add GLM-4.7 and GLM-4.7-FP8 (Z.ai)
- Add Kimi-K2-Thinking (Moonshot AI)
- Add Llama-Guard-3-8B, Meta-Llama-3.1-8B-Instruct (Base & Fast) (Meta)
- Add Nemotron-Nano-V2-12b and NVIDIA-Nemotron-3-Nano-30B-A3B (NVIDIA)
2026-02-04 20:17:14 +05:30
manascb1344
ca6206b78e
feat(nebius): add Qwen models to Token Factory
...
- Add Qwen3-Next-80B-A3B-Thinking
- Add Qwen3-30B-A3B-Thinking-2507 and Qwen3-30B-A3B-Instruct-2507
- Add Qwen3-Coder-30B-A3B-Instruct
- Add Qwen3-32B (Base & Fast)
- Add Qwen2.5-Coder-7B-fast
- Add Qwen2.5-VL-72B-Instruct
- Add Qwen3-Embedding-8B
2026-02-04 20:16:46 +05:30
manascb1344
51fe42982f
feat(nebius): add DeepSeek models to Token Factory
...
- Add DeepSeek-V3.2, DeepSeek-V3-0324 (Base & Fast), DeepSeek-R1-0528 (Base & Fast)
- These are new models available on Nebius Token Factory
2026-02-04 20:16:21 +05:30
manascb1344
4c78ea9f36
feat(nebius): add new providers for Nebius Token Factory
...
- Add MiniMaxAI provider with MiniMax-M2.1 model
- Add PrimeIntellect provider with INTELLECT-3 model
- Add black-forest-labs provider with FLUX.1-schnell and FLUX.1-dev
- Add BAAI provider with bge-multilingual-gemma2 and BGE-ICL
- Add intfloat provider with e5-mistral-7b-instruct
- Add Google provider with Gemma-2-2b-it, Gemma-2-9b-it-fast, Gemma-3-27b-it, and Gemma-3-27b-it-fast
2026-02-04 20:16:01 +05:30
Riccardo Giorato
9092f0b106
Delete Kimi-K2-5.toml
2026-02-04 11:22:49 +01:00
Zain Hasan
1acd3c199a
Rename model to 'Qwen3 Coder Next FP8'
2026-02-04 01:57:58 -08:00
Alex-wuhu
7deb00a333
feat: add deepseek OCR model configuration and Qwen3 Coder Next model configuration
2026-02-04 16:56:28 +08:00
samsja
f180f49df5
Add Intellect 3 model from Prime Intellect
2026-02-03 23:58:12 -08:00
Aiden Cline
59f13d1c0a
feat: make all openrouter models use openrouter sdk
2026-02-03 23:15:41 -06:00
Aiden Cline
ce6950074f
Revert "Add Bedrock cross-region inference profiles and update validation"
...
This reverts commit 89f62005cc .
2026-02-03 23:08:05 -06:00
Aiden Cline
b2b0f612f4
Merge pull request #788 from zainhas/dev
...
[Together AI] add qwen3 coder next
2026-02-03 22:48:52 -06:00
Aiden Cline
d1e92ce8ad
Merge pull request #790 from anomalyco/update-cf-workers
...
fix: update cf workers ai
2026-02-03 22:48:41 -06:00
Aiden Cline
20c81eb600
fix: update cf workers ai
2026-02-03 22:47:12 -06:00
Frank
6934bf2c66
Merge pull request #789 from qychen2001/dev
...
feat(models): add Kimi-K2.5 model support
2026-02-03 22:55:46 -05:00
QiyuanChen
83503944ba
feat(models): add Kimi-K2.5 model support
...
Add support for Moonshot AI's Kimi-K2.5 model with reasoning capabilities,
structured output, and multi-modal support (text/image input, text output).
Configured with a large context window of 262,000 tokens for both input
and output. Added to both SiliconFlow and SiliconFlow CN providers.
2026-02-04 11:45:28 +08:00
Zain Hasan
536ac44708
add qwen3 coder next
2026-02-03 14:54:42 -08:00
Aiden Cline
02f7969d53
Merge pull request #786 from unexge/push-lpupkorvtnuw
...
Update Amazon Bedrock models to add cross-region inference and remove deprecated models
2026-02-03 15:30:30 -06:00
Matt Silverlock
0ba8852f91
use official ai-gateway-provider package for Cloudflare AI Gateway
2026-02-03 15:41:11 -05:00
Burak Varlı
89f62005cc
Add Bedrock cross-region inference profiles and update validation
...
- Add Nova models for Global, US, EU, and APAC regions
- Add Llama 3.1/3.2 cross-region profiles for US and EU
- Add Claude Sonnet 4/3.7 APAC profiles
- Add Claude Sonnet 4.5/3.7/3.5 US Gov profiles
- Update validate-bedrock to include ap-southeast-1 region
- Skip us-gov models in validation (requires GovCloud access)
2026-02-03 20:10:40 +00:00
Aiden Cline
5afc754db3
Merge pull request #697 from berget-ai/feat/add-berget-ai-provider
...
feat: add Berget.AI provider
2026-02-03 12:15:36 -06:00
Aiden Cline
2fdfeecfc8
Merge pull request #784 from bendews/patch-1
...
Increase Github Copilot GPT 4.1 context limit from 64k to 128k
2026-02-03 09:24:29 -06:00
Aiden Cline
7bf852e19f
Merge pull request #785 from thePrnvBot/chore--updating-free-openrouter-models
...
fix: Update models tool call to false
2026-02-03 09:24:07 -06:00
Burak Varlı
bc58036964
Update Amazon Bedrock models to add cross-region inference and remove deprecated models
...
This change adds a new script to validate all Amazon Bedrock models by making a simple inference request using model identifiers.
As a result of that script, made some changes to make sure all model identifiers are usable via Amazon Bedrock:
- Added cross-region inference for various models including DeepSeek, Llama, Amazon Nova
- Removed some reprecated/EoL'd models including Amazon Titan, Claude v2, Cohere Command Light
2026-02-03 13:43:07 +00:00
thePrnvBot
613843b529
fix: update nousresearch model tool call to false
2026-02-03 16:12:39 +04:00
thePrnvBot
836b07aeaf
fix: update cognitivecomputation model tool call to false
2026-02-03 16:12:10 +04:00
thePrnvBot
e4f8c752ac
fix: update allenai model tool call to false
2026-02-03 16:11:48 +04:00
thePrnvBot
fa04882d5f
fix: update liquid models tool call to false
2026-02-03 16:11:36 +04:00
thePrnvBot
e69df42539
fix: update llama model tool call to false
2026-02-03 16:11:18 +04:00
thePrnvBot
64b7e989eb
fix: update tng-r1t-chimera :free tool call to false
2026-02-03 16:10:53 +04:00
Ben Dews
abf1259e58
Increase context limit from 64k to 128k
2026-02-03 21:22:15 +10:00
Aiden Cline
93fe136ec1
Merge pull request #777 from xiaojiezj/zenmux_dev
...
feat: ““Replace the chat-completion protocol in the Zenmux provider with the Anthropic protocol, and replace the model.”
2026-02-02 20:50:32 -06:00
Aiden Cline
be098329c5
Merge pull request #782 from fhennerkes/dev
...
poe: model update 2/2/26
2026-02-02 20:49:42 -06:00
fhennerkes
884e901c3f
poe: model update 2/2/26
2026-02-02 18:22:49 -08:00
Michel Pigassou
81ddc26ed0
feat(deepinfra): add GLM-4.7-Flash model
...
Add zai-org/GLM-4.7-Flash to DeepInfra provider
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-02-02 22:04:36 +01:00
Aiden Cline
e35973bdf9
Merge pull request #775 from Track07-cda/alibaba-cn_kimi
...
feat(alibaba-cn): add Kimi K2 Thinking and K2.5 models to alibaba-cn provider
2026-02-02 10:52:51 -06:00
Aiden Cline
f94d9dda7f
Merge pull request #607 from unexge/push-svvwrlmunkkt
...
Add cross-region inference profiles for Claude 4.x family models in Amazon Bedrock
2026-02-02 10:20:13 -06:00
xiaojie.zj
d354b55137
feat: “Replace the chat-completion protocol with the Anthropic protocol, and replace the model.”
2026-02-02 18:03:42 +08:00
Track07-cda
e30f361b88
feat(alibaba-cn): add reasoning content support for Kimi models
...
Add interleaved reasoning_content field to Kimi K2 Thinking and K2.5.
Also correct the display name for Moonshot Kimi K2.5.
2026-02-02 16:28:02 +08:00
Track07-cda
efa90f6fbf
feat(alibaba-cn): add Kimi K2 Thinking and K2.5 models
...
Add new model definitions for Moonshot Kimi K2 Thinking and K2.5.
Update Moonshot Kimi K2 Instruct metadata including open weights
status and output token limits.
2026-02-02 14:13:06 +08:00
Aiden Cline
93d03d87c1
Merge pull request #772 from thePrnvBot/chore--updating-free-openrouter-models
...
feat: add free openrouter models
2026-01-31 21:41:49 -06:00
Aiden Cline
67b31f3371
Merge pull request #773 from ccurme/cc/gpt-5.2-structured-output
...
fix: add structured_output to gpt-5.1 and 5.2
2026-01-31 20:58:12 -06:00
Aiden Cline
0513b73b17
fix: correct model id
2026-01-31 20:39:40 -06:00
Chester Curme
866974df3a
add structured_output to gpt-5.1 and 5.2
2026-01-31 21:39:31 -05:00
thePrnvBot
fd07fe7953
feat: add free qwen models to openrouter provider
2026-01-31 20:25:30 +04:00
thePrnvBot
d176299fdf
feat: add free nemotron models to openrouter provider
2026-01-31 20:24:53 +04:00
thePrnvBot
6193824e92
feat: add gpt oss free models to openrouter
2026-01-31 20:23:25 +04:00
thePrnvBot
b85b481fd0
feat: add hermes 3 llama 3.1 405b free model
2026-01-31 20:22:54 +04:00
thePrnvBot
e33d225a0b
chore: update deepseek r1 0528 free tool call to false
2026-01-31 20:22:02 +04:00
thePrnvBot
3c9e76cf89
feat: add tng-r1t-chimera free model
2026-01-31 20:21:22 +04:00
thePrnvBot
42e7f67b67
feat: add dolphin mistral 24b venice edition
2026-01-31 20:20:53 +04:00
thePrnvBot
0e46820a00
feat: add seedream model
2026-01-31 20:20:16 +04:00
thePrnvBot
e7dd66e51a
feat: add free meta llama models
2026-01-31 20:19:28 +04:00
thePrnvBot
ea1d856847
feat: add free black forest lab models
2026-01-31 20:18:46 +04:00
thePrnvBot
26fd570130
feat: added free liquid, sourceful and allenai models
2026-01-31 20:17:30 +04:00
Aiden Cline
008c521304
Merge branch 'dev' into feat/add-vercel-models-jan-30
2026-01-30 16:35:40 -06:00
Aiden Cline
c6870e97c5
Merge pull request #717 from MichaelYochpaz/fix-vertex-anthropic-npm-import
...
fix(google-vertex-anthropic): Fix incorrect NPM package used for Anthropic models used through Vertex
2026-01-30 15:57:19 -06:00
jerilynzheng
e7e8af6934
vercel: add 5 new models from Vercel AI Gateway
...
Add new models:
- alibaba/qwen3-max-thinking: Qwen 3 Max with reasoning
- arcee-ai/trinity-large-preview: Trinity 400B MoE model
- moonshotai/kimi-k2.5: Kimi K2.5 with vision and reasoning
- openai/gpt-4o-mini-search-preview: GPT-4o Mini search variant
- zai/glm-4.7-flashx: GLM 4.7 Flash lightweight model
Co-Authored-By: Claude Sonnet 4.5 <noreply@anthropic.com >
2026-01-30 13:38:07 -08:00
Aiden Cline
665a9fe17d
Merge pull request #767 from remorses/model-schema
...
add model-schema.json endpoint for model autocomplete
2026-01-30 13:32:54 -06:00
Aiden Cline
98114d5721
fix or glm flash
2026-01-30 13:21:57 -06:00
Aiden Cline
2deedb5fa5
Merge pull request #762 from davidcharbonnier/dev
...
Add GLM 4.7 Flash model on Openrouter
2026-01-30 13:20:33 -06:00
Aiden Cline
43cb68a64e
Merge pull request #768 from amazon-nova-api/nova-provider
...
Add nova as a model provider
2026-01-30 13:19:57 -06:00
Adnan Hajar
ee89f7ee6b
Add nova as a model provider
2026-01-30 14:01:14 -05:00
Tommy D. Rossi
6f45f3949f
add model-schema.json endpoint for model autocomplete
2026-01-30 15:51:04 +01:00
David Charbonnier
ebafef01d7
feat: add glm 4.7 flash model on openrouter
2026-01-30 09:06:48 -05:00
captain1379
77dccbe959
feat: add Jiekou.AI provider
...
Add Jiekou.AI as a new LLM provider with 102 models including:
- DeepSeek (V3, R1, OCR)
- Qwen (Qwen3, Qwen2.5)
- Claude (Opus, Sonnet, Haiku)
- GPT models (GPT-5.x, GPT-4.x, GPT-OSS)
- Gemini (Pro, Flash)
- GLM (4.5, 4.7)
- Kimi (K2, K2.5)
- Llama (3.x, 4.x)
- And more...
Jiekou.AI is an OpenAI-compatible API provider.
Co-Authored-By: Claude (pa/claude-opus-4-5-20251101) <noreply@anthropic.com >
2026-01-30 18:44:11 +08:00
Frank
8b2b4b40a1
update zen models
2026-01-30 00:52:31 -05:00
Frank
96da5d8331
update zen models
2026-01-29 16:46:15 -05:00
Frank
0146cb114e
update zen models
2026-01-29 16:38:42 -05:00
Aiden Cline
21177b3f6b
Merge pull request #760 from cgilly2fast/dev
...
feat(firmware): add kimi models and clean up model names
2026-01-29 15:02:50 -06:00
Colby Gilbert
522c486815
fix: wrong name for kimi k2.5
2026-01-29 12:16:15 -08:00
Colby Gilbert
85ef0fe0a8
feat: add kimi models
2026-01-29 10:33:25 -08:00
Colby Gilbert
70681d3398
chore: rename glm and gpt oss models
2026-01-29 10:33:17 -08:00
Frank
c2a6830fde
sync
2026-01-29 12:38:07 -05:00
Frank
4b9631cb89
update zen models
2026-01-29 12:35:06 -05:00
Aiden Cline
9efb6c1a73
Merge pull request #742 from 888-wzk/feature/chenger_20260128
...
feat(models): Add configuration files for the Kimi K2.5, GPT-5.2-Codex, Qwen3-Max-Thinking, and GLM 4.7 FlashX models.
2026-01-29 10:45:16 -06:00
Aiden Cline
9b5adb8230
Merge pull request #751 from cravenceiling/fix/openrouter-google-gemma-models
...
add and fix some google gemma models from openrouter
2026-01-29 10:44:52 -06:00
Aiden Cline
9105b7ba75
Merge pull request #753 from otterDeveloper/patch-1
...
fireworks: Raise Kimi K2.5 max output
2026-01-29 10:44:40 -06:00
Aiden Cline
a0dc4149cd
Merge pull request #754 from fanweixiao/dev
...
fix(vivgrid): set npm for gemini-3 models for vivgrid provider
2026-01-29 10:43:38 -06:00
Aiden Cline
ccb98b9597
Merge pull request #755 from friendliai/minpeter/add-minimax-friendli-model
...
feat(friendli): add MiniMax M2.1 model and update Qwen3
2026-01-29 10:42:14 -06:00
Aiden Cline
c33581d89d
Merge pull request #757 from s-scheck/feature/adjust-pricing-of-devstral-2512
...
feat: adjust pricing of devstral-2512 hosted by mistral
2026-01-29 10:41:58 -06:00
Aiden Cline
ab1a8c21db
Merge pull request #759 from FrancoStino/patch-5
...
Delete providers/nvidia/models/z-ai/glm-4.7.toml
2026-01-29 10:41:45 -06:00
Davide Ladisa
f969e060c8
Delete providers/nvidia/models/z-ai/glm-4.7.toml
...
Duplicate
https://github.com/anomalyco/models.dev/blob/dev/providers/nvidia/models/z-ai/glm4.7.toml
2026-01-29 17:17:59 +01:00
Sinan Scheck
66823bcd6e
chore: adjust pricing of devstral-2512 hosted by mistral
2026-01-29 11:45:20 +01:00
minpeter
01f391f1fe
Add MiniMax M2.1 model and update Qwen3 date
...
Co-Authored-By: Claude Opus 4.5 <noreply@anthropic.com >
2026-01-29 18:55:00 +09:00
minpeter
1d2ddf9b1d
Update friendli provider model configs
...
Remove outdated Qwen3 and Llama 4 model configurations.
Reorganize meta-llama models into subdirectory.
Upgrade GLM model to 4.7 with increased context limits (202,752 tokens).
2026-01-29 18:42:08 +09:00
C.C. Fan
a34c2a390b
fix(vivgrid): set npm for gemini-3 models for vivgrid provider
2026-01-29 16:43:32 +08:00
Frank
0c0d917719
sync
2026-01-29 03:16:34 -05:00
otterDeveloper
e5dcee15bb
fireworks: increase kimi-k2p5 max output
2026-01-29 02:04:24 -06:00
城二
0894de420e
feat(models): Add family and interleaved fields to Kimi K2.5 and GLM 4.7 FlashX configuration files
2026-01-29 10:36:04 +08:00
Aiden Cline
abfad44131
Merge pull request #734 from riccardogiorato/dev
...
[together.ai] add four new models: Qwen3-235B, Qwen3-Next-80B, Kimi-K2-Instruct, GLM 4.7
2026-01-28 20:13:26 -06:00
Aiden Cline
cc3566a9a2
Merge pull request #747 from monotykamary/update-kimi-k2.5-model
...
feat(synthetic): update Kimi K2.5 model configuration
2026-01-28 20:12:59 -06:00
Aiden Cline
71b90fe6ce
Merge pull request #748 from dpuyosa/venice
...
Venice: Adjust limit context/output token limits for all models
2026-01-28 20:12:42 -06:00
Aiden Cline
c854148259
Merge pull request #749 from esafak/chore/kimi-k2.5
...
chore: disable `temperature` in `moonshotai/kimi-k2.5`
2026-01-28 20:12:30 -06:00
Aiden Cline
980b64e860
Merge pull request #750 from esafak/feat/zai-glm-4.7-flash
...
feat: Add `zai-coding-plan/glm-4.7-flash`
2026-01-28 20:12:12 -06:00
cravenceiling
632fc61b95
add to the limit section
2026-01-28 19:33:02 -05:00
cravenceiling
5a76cee029
add and fix some google gemma models from openrouter
2026-01-28 19:07:09 -05:00
Emre Şafak
0971bd00f3
feat: Add zai-coding-plan/glm-4.7-flash.
...
* Create the file `providers/zai-coding-plan/models/glm-4.7-flash.toml` to define the new model.
* Set the model name to `GLM-4.7-Flash`.
* Configure model capabilities including reasoning, tool call, and knowledge cutoff of `2025-04`.
* Define context limit as `200_000` tokens.
* Set input and output costs to zero.
2026-01-28 18:44:46 -05:00
Emre Şafak
94fca7064d
chore: disable temperature in moonshotai/kimi-k2.5
2026-01-28 18:11:32 -05:00
dpuyosa
40ad678ae7
[venice] Adjust limit context/output token limits for all models
...
- All models limit context/output was reduced by 2.4%
2026-01-28 18:34:59 +01:00
Tom X Nguyen
e18738adda
feat(synthetic): update Kimi K2.5 model configuration
...
Update model config with corrected values:
- max_output: 65_536 (from 32_768)
- cost.input: 0.55, cost.output: 2.19
- modalities.input: [text, image]
- Add interleaved section for reasoning_content
- Use underscores for large numbers
2026-01-28 23:39:59 +07:00
Aiden Cline
0b003c9c18
add new arcee models
2026-01-28 10:50:59 -05:00
Aiden Cline
cb5035a8ee
Merge pull request #743 from zainhas/dev
...
add kimi k2.5
2026-01-28 10:30:17 -05:00
Aiden Cline
01e3069901
Merge pull request #745 from reissbaker/kk25
...
Add Kimi K2.5 for Synthetic
2026-01-28 10:29:59 -05:00
Aiden Cline
7b577ad7a8
Merge pull request #737 from cravenceiling/add-google-gemma-3-27b-it-free
...
feat: add google-gemma-3-27b-it:free model
2026-01-28 10:29:21 -05:00
Matt Baker
55e72aa8db
Correct output tokens
2026-01-28 02:08:58 -08:00
Matt Baker
a8fc1d2e44
Add Kimi K2.5 for Synthetic
2026-01-28 02:07:52 -08:00
Riccardo Giorato
101b5042bd
Create Kimi-K2-5.toml
2026-01-28 10:57:37 +01:00
Riccardo Giorato
7b835b2f29
Merge remote-tracking branch 'upstream/dev' into dev
2026-01-28 10:50:29 +01:00
Aiden Cline
9658a500f7
Merge pull request #740 from thatoddmailbox/dev
...
Fix Kimi pricing for fireworks-ai
2026-01-28 02:24:39 -05:00
Aiden Cline
7030a7c77f
Merge pull request #741 from Alex-wuhu/dev
...
feat(models): add Kimi K2.5 and GLM-4.7-Flash model
2026-01-28 02:24:12 -05:00
Frank
2be2a8c109
Update zai models
2026-01-28 01:49:08 -05:00
Frank
465335102f
Update kimi-k2.5.toml
2026-01-28 01:38:18 -05:00
城二
82ddcad9f8
feat(models): Add configuration files for the Kimi K2.5, GPT-5.2-Codex, Qwen3-Max-Thinking, and GLM 4.7 FlashX models.
2026-01-28 14:31:21 +08:00
Zain Hasan
9acd2ffa60
add kimi k2.5
2026-01-27 22:30:24 -08:00
Alex-wuhu
b3d2cfdc34
feat(models): add Kimi K2.5 and GLM-4.7-Flash model
2026-01-28 14:25:53 +08:00
Alex Studer
eead89fd8e
fix kimi pricing for fireworks-ai
2026-01-28 01:20:01 -05:00
Aiden Cline
48de510380
Merge pull request #635 from mthezi/feature/add-302ai-provider
...
feat: add 302ai provider
2026-01-27 22:01:19 -05:00
Aiden Cline
52332705ca
Merge pull request #736 from alissonlauffer/chore/update-chutes-kimi-k2.5
...
feat(chutes): update Kimi K2.5 TEE model capabilities
2026-01-27 22:00:07 -05:00
Aiden Cline
08e5d0f830
Update Kimi-K2.5-TEE.toml configuration settings
2026-01-27 21:59:39 -05:00
Aiden Cline
0b7f253ee0
Merge pull request #739 from xinrui-z/feat/aihubmix-add-models
...
feat(models): add kimi-k2.5, coding-glm-4.7, glm-4.6v, and qwen3-max
2026-01-27 21:54:59 -05:00
Xinrui
bfe953d2f0
feat(models): add kimi-k2.5, coding-glm-4.7, glm-4.6v, and qwen3-max
2026-01-28 10:48:40 +08:00
cravenceiling
641fa6f2e7
feat: add google-gemma-3-27b-it:free model
2026-01-27 19:34:05 -05:00
Alisson Lauffer
c344db1bf8
feat(chutes): update Kimi K2.5 TEE model capabilities
...
Enable reasoning, tool calling, and multimodal input support for the
Kimi K2.5 TEE model. Increase context limit from 32k to 262k tokens and
output limit from 8k to 65k tokens. Add support for image and video
inputs alongside text. Configure interleaved reasoning content field.
2026-01-27 21:05:33 -03:00
Aiden Cline
36c6206d32
Merge pull request #732 from mmealman/add_fireworks_k2p5
...
Added Kimi K2.5 to FireworksAI.
2026-01-27 17:54:31 -05:00
Aiden Cline
4f6a59d7be
Merge pull request #729 from gary149/feat/huggingface-kimi-k2.5
...
feat(huggingface): add Kimi-K2.5 model
2026-01-27 17:54:15 -05:00
Aiden Cline
f5b8e3fe83
Merge pull request #735 from spiffytech/dev
...
Add Kimi K2.5 to Ollama Cloud
2026-01-27 17:53:31 -05:00
Aiden Cline
07c70ca9f2
Merge pull request #731 from arguiot/add-vercel-kimi-k2.5
...
Add Kimi K2.5 to Vercel provider
2026-01-27 17:53:22 -05:00
Aiden Cline
d35ad7ec49
Merge pull request #727 from ProlowN/dev
...
fix : removed duplicate kimi k2.5 model from venice
2026-01-27 17:53:09 -05:00
Aiden Cline
7c57f4ce15
Merge pull request #733 from dpuyosa/dev
...
Venice: Add interleaved thinking to k2.5
2026-01-27 17:52:44 -05:00
spiffytech
83eeb304a5
Add Kimi K2.5 to Ollama Cloud
2026-01-27 16:08:16 -05:00
Aiden Cline
a13f101e0c
Merge pull request #638 from jerome-benoit/feature/sap-ai-core-updates
...
fix(sap-ai-core): use working provider fork for stable OpenCode integration
2026-01-27 15:32:40 -05:00
Riccardo Giorato
ec6101a629
Merge remote-tracking branch 'upstream/dev' into dev
2026-01-27 21:00:55 +01:00
Riccardo Giorato
231313aad0
Add four new models: Qwen3-235B, Qwen3-Next-80B, Kimi-K2-Instruct, GLM-4.7
2026-01-27 21:00:44 +01:00
dpuyosa
b9793731e6
Add interleaved thinking to k2.5
2026-01-27 20:32:38 +01:00
Frank
22edc4d92d
update zen model
2026-01-27 14:12:42 -05:00
Frank
3f62b2dd5a
update moonshot models
2026-01-27 14:05:13 -05:00
Mark Mealman
d57592dba3
Added Kimi K2.5 to FireworksAI.
2026-01-27 14:00:19 -05:00
Frank
e2b43f180c
Merge pull request #730 from esafak/moonshotai/kimi-k2.5
...
chore: add `moonshotai/kimi-k2.5` model
2026-01-27 13:59:59 -05:00
Arthur Guiot
0acff9cf7c
add Kimi K2.5 to Vercel provider
2026-01-27 10:50:37 -08:00
Emre Şafak
563c43f004
add moonshotai/kimi-k2.5 model
2026-01-27 13:46:30 -05:00
Victor Muštar
e53bb9c7ad
feat(huggingface): add Kimi-K2.5 model
2026-01-27 18:38:18 +01:00
Frank
1522bc4a9a
update zen models
2026-01-27 12:34:50 -05:00
Frank
c28701d579
update zen models
2026-01-27 12:34:29 -05:00
Magnus
eb5bff1f6a
fix : removed duplicate kimi k2.5 model from venice
2026-01-27 18:01:37 +01:00
Frank
15b4b02e6e
update zen models
2026-01-27 12:00:36 -05:00
Jan Szypulski
22d6a24c7a
fix: llama 3.3 last update
2026-01-27 18:00:11 +01:00
Jan Szypulski
c7bc5b7c98
fix: corrected logo color and size
2026-01-27 17:59:55 +01:00
Aiden Cline
b1910161d4
Merge pull request #726 from ProlowN/dev
...
Added kimi k2.5 to Venice AI
2026-01-27 11:45:18 -05:00
Magnus
13e48c2ca0
fix/ wrong output size
2026-01-27 17:44:27 +01:00
Magnus
516cfe355d
fix/ wrong family name
2026-01-27 17:09:38 +01:00
Magnus
068eacd6b3
Added kimi k2.5 to Venice AI
2026-01-27 17:06:02 +01:00
Aiden Cline
336e43494b
Merge pull request #719 from Jakey-Jakey/dev
...
add-kimi-k2.5 from OpenRouter
2026-01-27 11:05:16 -05:00
Aiden Cline
7fc046f833
Merge pull request #720 from kassieclaire/add-kimi-k2p5-model
...
feat(providers): add Kimi K2.5 model
2026-01-27 11:04:43 -05:00
Jan Szypulski
2c63a024b3
delete unrecognized model family
2026-01-27 17:04:36 +01:00
Aiden Cline
a572cf8a1a
Merge pull request #721 from matthusby/dev
...
[Chutes] Add new model configs and update pricing for several models
2026-01-27 11:03:54 -05:00
Aiden Cline
e335f919f2
Merge pull request #722 from FrancoStino/dev
...
feat(providers): Add NVIDIA models: Kimi K2.5 and GLM-4.7
2026-01-27 11:03:38 -05:00
Aiden Cline
4e883ea026
Merge branch 'dev' into dev
2026-01-27 11:02:10 -05:00
Aiden Cline
3dfb74d1ea
Merge pull request #723 from arshadbarves/feat/nvidia-kimi-k2.5
...
feat(nvidia): add Kimi K2.5 multimodal model
2026-01-27 11:01:42 -05:00
Aiden Cline
edb551b275
Merge pull request #724 from dpuyosa/dev
...
Venice: Add Kimi K2.5 model configuration
2026-01-27 11:01:29 -05:00
Jan Szypulski
2f34ee47ee
add cloudferro logo
2026-01-27 16:48:42 +01:00
Jan Szypulski
f6cb6631b8
add cloudferro sherlock models
2026-01-27 16:48:28 +01:00
dpuyosa
d95d22e89c
[venice] Add Kimi K2.5 model configuration
...
- Add new Kimi K2.5 model with 262K context support
- Include pricing for input, output, and cache_read operations
- Enable reasoning, tool calling, and structured output capabilities
- Support text and image input with text output
2026-01-27 16:23:08 +01:00
Davide Ladisa
dc771f54df
Update knowledge and release dates in kimi-k2.5.toml
2026-01-27 15:52:43 +01:00
Arshad Barves
27b99e9ccf
feat(nvidia): add Kimi K2.5 multimodal model
...
Add Kimi K2.5, a 1T parameter multimodal MoE model by Moonshot AI
with support for text, image, and video inputs.
Key features:
- 256K context window (262,144 tokens)
- Native multimodal support (text, image, video)
- Interleaved reasoning with reasoning_content field
- Tool calling and temperature control
- Open weights available
Model ID: moonshotai/kimi-k2.5
Provider: NVIDIA NIM
Validation: ✅ Passes bun validate
2026-01-27 20:00:42 +05:30
Davide Ladisa
c04069b5a3
Merge pull request #102 from FrancoStino/add-nvidia-models-kimi-glm
...
Add NVIDIA models: Kimi K2.5 and GLM-4.7
2026-01-27 15:01:43 +01:00
Davide Ladisa
af08a750b2
Add GLM-4.7 with correct filename and family field
2026-01-27 15:01:13 +01:00
Davide Ladisa
ed4270cc8f
Remove old glm4_7.toml to rename to glm-4.7.toml
2026-01-27 15:01:03 +01:00
Davide Ladisa
dae3873284
Fix GLM-4.7 release date to December 2025 and update knowledge cutoff
2026-01-27 14:59:23 +01:00
Davide Ladisa
6daad4c4eb
Update knowledge cutoff dates to more accurate values
2026-01-27 14:57:07 +01:00
Davide Ladisa
5302dc4452
Add NVIDIA models: Kimi K2.5 and GLM-4.7
2026-01-27 14:53:55 +01:00
Matt Husby
7d26d504ec
Add new model configs and update pricing for several models
2026-01-27 07:58:35 -05:00
kassieclaire
63f116b27f
fix: remove interleaved reasoning for kimi-k2.5
2026-01-27 07:05:35 -05:00
kassieclaire
bf6582bb83
fix: update knowledge cutoff to 2025-01 for kimi-k2.5
2026-01-27 06:25:46 -05:00
Kassie Povinelli
965f5365bc
Update providers/kimi-for-coding/models/k2p5.toml
...
checked docs for kimi-for-coding plan, still shows up as this lower value, so going with it for now -- keep an eye on the docs in case they update the information
Co-authored-by: Copilot <175728472+Copilot@users.noreply.github.com >
2026-01-27 06:19:39 -05:00
kassieclaire
3b5717f6b9
feat(providers): add Kimi K2.5 model
2026-01-27 06:12:01 -05:00
Jakey-Jakey
cae7be4925
Add 'video' modality to input options
2026-01-27 04:01:00 -05:00
Jakey-Jakey
d218bb7fe8
Add cache_read cost to kimi-k2.5 configuration
2026-01-27 03:56:18 -05:00
Jakey-Jakey
9f5b80af72
Add knowledge parameter with value '2025-01'
2026-01-27 03:55:37 -05:00
Jakey-Jakey
6c928db808
Remove knowledge field from kimi-k2.5.toml
...
Remove knowledge field from configuration.
2026-01-27 03:54:38 -05:00
Jakey-Jakey
63b1cfb9fc
Add provider section to kimi-k2.5.toml
2026-01-27 03:53:43 -05:00
Jakey-Jakey
ade61a5018
Add files via upload
2026-01-27 03:45:43 -05:00
Aiden Cline
e768c2afbd
Merge pull request #713 from qychen2001/dev
...
feat(providers): update siliconflow-cn model catalog
2026-01-26 21:01:47 -05:00
Aiden Cline
31e503a516
Merge pull request #715 from dpuyosa/dev
...
Venice: Add cache_read to GLM 4.7
2026-01-26 21:00:43 -05:00
Frank
98a455cb0f
sync
2026-01-26 18:24:50 -05:00
Michael Yochpaz
f1d2e47772
fix(google-vertex-anthropic): use @ai-sdk/google-vertex/anthropic npm package
...
The google-vertex-anthropic provider requires the `/anthropic` subpath import for thinking/reasoning to work correctly with Claude models on Vertex AI.
2026-01-26 22:09:05 +00:00
dpuyosa
4d82211cea
Add cache_read to GLM 4.7
2026-01-26 22:59:16 +01:00
mthezi
4e0a2d34b4
refactor(models): update family names for various models to improve consistency
2026-01-26 14:47:20 +08:00
⌞L⌝
effa34d17b
Merge branch 'anomalyco:dev' into feature/add-302ai-provider
2026-01-26 14:29:57 +08:00
QiyuanChen
ed59411f9e
feat(providers): update siliconflow-cn model catalog
...
Add new Pro tier models for deepseek-ai and moonshotai, including DeepSeek-R1, DeepSeek-V3 series, and Kimi-K2-Thinking models with reasoning capabilities. Remove older Qwen, Kimi-K2, and other legacy model configurations.
2026-01-26 12:57:56 +08:00
Aiden Cline
1286f6449c
Merge pull request #710 from hsyysy/dev
...
feat(provider): add DeepSeek-V3.2 for Nvidia
2026-01-25 22:56:41 -05:00
Aiden Cline
f92551d3ac
Merge pull request #709 from fanweixiao/feat/add-vivgrid-models
...
add gpt-5.1-codex-max, gpt-5.2-codex and more models for vivgrid provider
2026-01-25 22:56:31 -05:00
Aiden Cline
8c502a36b9
Delete pnpm-lock.yaml
2026-01-25 21:29:19 -05:00
Aiden Cline
6e40a4744a
Merge pull request #712 from xinrui-z/fix/aihubmix-provider-invalid-type
...
fix(provider): correct invalid type in provider.toml
2026-01-25 21:28:54 -05:00
Xinrui
5ac346644a
fix(provider): correct invalid type in provider.toml
2026-01-26 10:12:08 +08:00
Thomas Young
c03332fb3d
feat(provider): add DeepSeek-V3.2 for Nvidia
2026-01-25 20:10:23 +08:00
C.C. Fan
acb8319afb
add gpt-5.1-codex-max, gpt-5.2-codex, gemini-3-pro-preview and gemini-3-flash-preview for vivgrid provider
2026-01-25 16:37:27 +08:00
Aiden Cline
568f5319be
Merge pull request #708 from jsdtxm/feat/add-glm-4.7
...
feat(provider): add Pro/zai-org/GLM-4.7 for SiliconFlow-CN
2026-01-24 23:39:40 -05:00
lazy
2b331310b7
fix(qihang-ai): rename provider and fix logo to match standards
...
- Rename provider from qihang to qihang-ai
- Update logo to use standard size (24x24) and currentColor
Co-Authored-By: Claude Sonnet 4.5 <noreply@anthropic.com >
2026-01-25 12:34:48 +08:00
xiamin
0fb16d2cb8
feat(provider): add Pro/zai-org/GLM-4.7 for SiliconFlow-CN
2026-01-25 12:16:12 +08:00
Aiden Cline
18e555c2af
Merge pull request #656 from fchange/feat/new-provider
...
feat: add moark provider
2026-01-24 23:15:07 -05:00
Aiden Cline
1938af666f
Merge pull request #707 from arshadbarves/fix/nvidia-glm4.7-model-id
...
fix(nvidia): correct model ID for GLM-4.7 (z-ai/glm4.7)
2026-01-24 23:12:22 -05:00
Arshad Barves
fea35d7bb4
fix(nvidia): correct model ID for GLM-4.7 (z-ai/glm4.7)
...
Rename model file from glm-4.7.toml to glm4.7.toml to generate the
correct model ID z-ai/glm4.7 (without dot) as per NVIDIA API specification.
The model ID is derived from the file path, so the filename must match
the exact model identifier used by the provider's API.
- Renamed: providers/nvidia/models/z-ai/glm-4.7.toml → glm4.7.toml
- Model ID: z-ai/glm-4.7 → z-ai/glm4.7
- Validation: ✅ Passes bun validate
2026-01-25 09:34:50 +05:30
Aiden Cline
53b821523b
Merge pull request #700 from vglafirov/feat/gitlab-gpt-5-2
...
feat(gitlab): add GPT-5.2 model definition (duo-chat-gpt-5-2)
2026-01-24 12:50:30 -05:00
Aiden Cline
813b2d57b3
Merge pull request #704 from jsdtxm/feat/add-minimax-m2-1
...
feat(provider): add MiniMax M2.1 for SiliconFlow
2026-01-24 12:50:18 -05:00
xiamin
ee5c39bb18
fix: move MiniMax-M2.1 config
2026-01-24 22:17:29 +08:00
xiamin
ec2bf4bf7c
feat(provider): add MiniMax M2.1 for SiliconFlow-CN
2026-01-24 16:16:03 +08:00
xiamin
d2d6bc2d6c
chore: remove MiniMax-M2.1.toml symlink
2026-01-24 16:15:15 +08:00
xiamin
a4978b8b1c
feat(provider): add MiniMax M2.1 for SiliconFlow
2026-01-24 16:08:32 +08:00
Frank
545bf83089
update zen models
2026-01-23 23:19:34 -05:00
Vladimir Glafirov
5651a0efe1
feat(gitlab): add GPT-5.2 model definition (duo-chat-gpt-5-2)
2026-01-23 16:10:47 +01:00
Frank
b5fc3e3f54
update zen models
2026-01-23 01:19:18 -05:00
Frank
c8f6d7ace2
update zen models
2026-01-23 01:12:50 -05:00
Frank
4dd2e77ad1
update zen models
2026-01-23 01:05:55 -05:00
Aiden Cline
e5e859ac63
fix context limit for copilto gpt-4.1
2026-01-22 19:37:27 -06:00
Christian Landgren
eb98dd4305
fix: address Copilot review comments
...
- Change Mistral family from 'mistral' to 'mistral-small' for consistency
- Fix Llama 3.3 70B knowledge date from '2024-12' to '2023-12'
- Set tool_call to false for KB-Whisper-Large (speech-to-text models don't support tool calling)
2026-01-23 01:10:35 +01:00
Christian Landgren
cb8d8e8698
chore: remove Qwen3 32B model
2026-01-23 01:05:12 +01:00
Christian Landgren
3cb9a1cd3d
feat: add Berget.AI provider
...
Add Berget.AI as an OpenAI-compatible provider with base URL api.berget.ai/v1.
Models included:
- Text: Llama 3.3 70B, Qwen3 32B, GPT-OSS-120B, GLM 4.7, Mistral Small 3.2 24B
- Embedding: Multilingual-E5-large-instruct, Multilingual-E5-large
- Rerank: bge-reranker-v2-m3
- Speech-to-Text: KB-Whisper-Large
2026-01-23 01:01:40 +01:00
Aiden Cline
830a03e46b
Merge pull request #696 from vglafirov/feat/gitlab-openai-models
...
fix: increase output token limit for GitLab Claude models to 64k
2026-01-22 15:20:31 -08:00
Vladimir Glafirov
b21c8870a5
fix: align GitLab Claude models with native Anthropic model capabilities
...
Updated to match native Anthropic model definitions:
- attachment: false → true (supports image/pdf attachments)
- reasoning: false → true (supports extended thinking)
- modalities.input: ["text"] → ["text", "image", "pdf"]
- Added knowledge cutoff dates from native models
2026-01-23 00:18:27 +01:00
Vladimir Glafirov
4bcba6a482
fix: increase output token limit for GitLab Claude models to 64k
...
The output limit was set to 4,096 tokens which caused tool calls with
large content (like file generation) to be truncated mid-JSON.
Updated to match standard Anthropic model limits:
- duo-chat-opus-4-5: 4,096 → 64,000
- duo-chat-sonnet-4-5: 4,096 → 64,000
- duo-chat-haiku-4-5: 4,096 → 64,000
2026-01-23 00:02:49 +01:00
Aiden Cline
c67ccd8def
Merge pull request #694 from cgilly2fast/dev
...
chore: remove deepseek-coder for firmware provider
2026-01-22 11:51:38 -08:00
Colby Gilbert
7aa00eb8dc
chore: remove deepseek-coder for firmware provider
2026-01-22 11:48:50 -08:00
Aiden Cline
d799a6ae6e
Merge pull request #692 from vglafirov/feat/gitlab-openai-models
...
feat(gitlab): add OpenAI GPT-5 model definitions
2026-01-22 08:52:22 -08:00
Vladimir Glafirov
a770639c25
feat(gitlab): add OpenAI GPT-5 model definitions
...
Add GitLab Duo model definitions for OpenAI GPT-5 family:
- duo-chat-gpt-5-1: GPT-5.1 flagship model
- duo-chat-gpt-5-mini: GPT-5 Mini (cost-effective)
- duo-chat-gpt-5-codex: GPT-5 Codex (agentic coding)
- duo-chat-gpt-5-2-codex: GPT-5.2 Codex
2026-01-22 17:43:59 +01:00
Jan Szypulski
d934e26168
add cloudferro sherlock as provider
2026-01-22 16:11:45 +01:00
mthezi
ea20440d0d
fix: update model family name for gpt-4.1-nano
2026-01-22 13:51:54 +08:00
Aiden Cline
eef424f296
Merge pull request #686 from zhzy0077/nvidia-patch
...
Add nvidia 2 new models.
2026-01-21 16:23:25 -08:00
Aiden Cline
05415ee2ec
Merge pull request #685 from spiffytech/dev
...
Remove duplicate GLM-4.7 model file
2026-01-21 16:20:01 -08:00
Aiden Cline
23e99a093a
Merge pull request #687 from eliasto/ovhcloud/update-models
...
Update OVHcloud AI Endpoints models
2026-01-21 16:19:52 -08:00
Aiden Cline
02df983581
Merge pull request #688 from gitpush-gitpaid/dev
...
Added PDF to input modalities for gpt 5.2 codex
2026-01-21 16:19:36 -08:00
Aiden Cline
3d102d3bd9
Add 'pdf' to input modalities in gpt-5.2-codex.toml
2026-01-21 18:19:14 -06:00
gitpush-gitpaid
6307a2c223
added PDF to input modalities for gpt 5.2 codex
2026-01-21 18:30:18 -05:00
Aiden Cline
d79ae1d684
chore: kill deprecated copilot models from list
2026-01-21 16:59:11 -06:00
Elias TOURNEUX
67d192dd9c
Update OVHcloud AI Endpoints models
2026-01-21 08:17:47 -05:00
lazy
74cb010892
feat(qihang): add Gemini 2.5 Flash and GPT-5.2 models
2026-01-21 15:19:30 +08:00
zhzy0077
10acfc848d
Add nvidia 2 new models.
2026-01-21 08:39:12 +08:00
spiffytech
5943a24d41
Remove duplicate GLM-4.7 model file
2026-01-20 14:02:19 -05:00
Aiden Cline
a52b64222e
Merge pull request #684 from sebastiand-cerebras/final-removal-of-glm4_6
...
Remove deprecated zai-glm-4.6 model (Jan 20, 2026)
2026-01-20 10:17:47 -08:00
Seb Duerr
a767bf0a6d
Remove deprecated zai-glm-4.6 model (Jan 20, 2026)
...
Thank you for your patience and understanding with our timeline adjustments! I truly appreciate your team's responsiveness and flexibility in working with us on this deprecation.
As of January 20, 2026, the zai-glm-4.6 model has been officially deprecated.
2026-01-20 09:56:33 -08:00
Aiden Cline
b131f86a1f
Merge pull request #666 from spiffytech/dev
...
Update Ollama Cloud models. Add generator for model files.
2026-01-20 08:06:40 -08:00
Aiden Cline
c84e382bbe
Merge pull request #679 from WSQS/dev
...
feat: add GLM-4.7-Flash for zhipuai provider
2026-01-20 08:03:16 -08:00
Aiden Cline
9de5f304fe
Merge pull request #683 from nickdowse/dev
...
Fix: Fix incorrect OpenAI, Gemini prices
2026-01-20 08:03:07 -08:00
Aiden Cline
8a854771d7
Merge pull request #677 from ivivek/dev
...
feat: add GLM-4.7 to google-vertex
2026-01-20 08:02:57 -08:00
Aiden Cline
b190cdaecc
Merge pull request #678 from dpuyosa/UpdateModel
...
Venice: Update provider package
2026-01-20 08:02:47 -08:00
Aiden Cline
5712350b30
Merge pull request #680 from cgilly2fast/cgilly2fast/firmware-provider
...
feat: add cerebras glm 4.7 and gpt OSS, clean up claude model ids
2026-01-20 08:02:12 -08:00
Nick Dowse
b933688a77
Fix incorrect openai, gemini prices
2026-01-20 10:06:03 -05:00
dpuyosa
64f034bb72
Update interleaved field to reasoning_content
...
- Change field value in claude-sonnet-45, gemini-3-flash-preview, qwen3-235b-a22b-thinking-2507, and zai-org-glm-4.7 configs
2026-01-20 15:34:42 +01:00
Frank
72de414c2f
Merge pull request #681 from tars90percent/minimax-provider-names
...
Add MiniMax coding plan providers
2026-01-20 09:11:13 -05:00
Frank
bc6698d98b
sync
2026-01-20 09:10:16 -05:00
lazy
b465cec21a
feat: add QiHang provider with 7 models
...
- Add QiHang provider configuration (OpenAI-compatible API)
- API endpoint: https://api.qhaigc.net/v1
- Add 7 models:
- gpt-5.2-codex (/bin/zsh.14/.14)
- gpt-5-mini (/bin/zsh.04//bin/zsh.29)
- claude-opus-4-5-20251101 (/bin/zsh.71/.57)
- claude-sonnet-4-5-20250929 (/bin/zsh.43/.14)
- claude-haiku-4-5-20251001 (/bin/zsh.14//bin/zsh.71)
- gemini-3-flash-preview (/bin/zsh.07//bin/zsh.43)
- gemini-3-pro-preview (/bin/zsh.57/.43)
- All configurations validated with bun validate
2026-01-20 16:45:35 +08:00
tars90percent
e0fcf8f638
Add MiniMax coding plan providers
2026-01-20 13:42:43 +08:00
Colby Gilbert
36a6197da9
feat: add cerebras glm 4.7 and gpt OSS, clean up claude model ids
2026-01-19 21:34:34 -08:00
WSQS
f6d82c43a7
feat: add GLM-4.7-Flash for zhipuai
2026-01-20 10:49:21 +08:00
dpuyosa
44b8ed5871
Comment-out 'api' for validation script
2026-01-20 01:30:13 +01:00
dpuyosa
c0d9ec4777
Update Venice provider package:
...
- Replace @ai-sdk/openai-compatible with venice-ai-sdk-provider
- Fix cache_control limitations
- Add Venice-specific features
2026-01-20 01:11:34 +01:00
spiffytech
2ae1e23591
Update Ollama Cloud models. Add generator for model files.
2026-01-19 17:29:38 -05:00
Vivek K
1ff1405664
feat: add GLM-4.7 to google-vertex
2026-01-20 00:43:34 +05:30
Aiden Cline
1c32145339
Merge pull request #674 from zerone0x/add/gpt-5.1-codex-max
...
feat(openrouter): add openai/gpt-5.1-codex-max model
2026-01-19 09:52:54 -08:00
Aiden Cline
fbebe356b5
Merge pull request #675 from ElecTwix/glm-4.7-flash
...
feat: add glm-4.7-flash model
2026-01-19 09:52:25 -08:00
ElecTwix
e694f0136f
feat: add glm-4.7-flash model
2026-01-19 20:44:29 +03:00
zerone0x
5062058b6a
feat(openrouter): add openai/gpt-5.1-codex-max model
...
Add GPT-5.1-Codex-Max model to OpenRouter provider. This model is available
in OpenRouter's API but was missing from models.dev.
Pricing sourced from OpenRouter API.
Co-Authored-By: Claude <noreply@anthropic.com >
2026-01-20 01:25:22 +08:00
Aiden Cline
7b132f2cd8
Merge pull request #673 from gary149/feat/huggingface-glm-4.7-flash
...
feat(huggingface): add GLM-4.7-Flash model
2026-01-19 08:49:44 -08:00
Victor Muštar
e319a707fd
feat(huggingface): add GLM-4.7-Flash model
2026-01-19 17:36:59 +01:00
Aiden Cline
627ac7bcf1
Merge pull request #671 from uniquename/ollama/glm-4.7
...
feat: add Ollama GLM-4.7 model configuration file
2026-01-19 07:38:50 -08:00
Aiden Cline
89408c71e0
Merge pull request #669 from dpuyosa/UpdateModel
...
Venice: Replace vision models glm4.6v -> qwen3-vl
2026-01-19 07:38:29 -08:00
Aiden Cline
0651768fd9
Merge pull request #672 from sebastiand-cerebras/add-glm4_6-deprecation-notice
...
Re-add zai-glm-4.6 temporarily until Jan 20, 2026
2026-01-19 07:38:00 -08:00
Seb Duerr
c8ce0db2b1
Re-add zai-glm-4.6 temporarily until Jan 20, 2026
...
Thanks for the incredibly fast merge! We appreciate the efficiency, though we need to temporarily re-add GLM 4.6. The model will be officially deprecated on January 20, 2026. Our apologies for any confusion - we should have been clearer about the timeline in the original PR.
2026-01-19 07:14:34 -08:00
User
c3b177ed7a
feat: add Ollama GLM-4.7 model configuration file
2026-01-19 12:45:35 +00:00
dpuyosa
1ce4dc41c1
Update model configurations:
...
- Add qwen3-vl-235b-a22b model
- Remove deprecated zai-org-glm-4.6v model
2026-01-19 10:49:32 +01:00
Aiden Cline
438e834043
add input field to more openai models
2026-01-19 00:59:44 -06:00
Jérôme Benoit
189aa03281
Apply suggestion from @jerome-benoit
2026-01-19 04:02:10 +01:00
Aiden Cline
5f293ca6ce
Merge pull request #665 from sebastiand-cerebras/removal_of_glm4_6
...
Remove deprecated zai-glm-4.6 model from Cerebras provider
2026-01-17 22:50:25 -08:00
Aiden Cline
490a03f2d9
Remove deprecated zai-glm-4.6 model from Cerebras provider
2026-01-17 22:49:51 -08:00
Aiden Cline
cbd215cbc1
Merge pull request #652 from hueyexe/dev
...
feat: add GPT 5.2 Codex to Azure and Azure Cognitive Services
2026-01-17 22:48:21 -08:00
Aiden Cline
1b49c365a0
fix: restore gpt-5.2-codex.toml as symlink to fix validation CI
2026-01-18 00:46:56 -06:00
Aiden Cline
51d2a4f2b6
Merge pull request #664 from jerome-benoit/feat/sap-ai-core-claude-4.5-opus
...
feat(sap-ai-core): add Claude 4.5 Opus and align pricing
2026-01-17 19:02:07 -08:00
Seb Duerr
afba23ed62
Remove deprecated zai-glm-4.6 model from Cerebras provider
2026-01-17 18:37:18 -08:00
Jérôme Benoit
9e25ca1521
feat(sap-ai-core): add Claude 4.5 Opus and align pricing
...
- Add Claude 4.5 Opus model with official Anthropic pricing
- Align cache pricing for Claude 3 Sonnet, Gemini 2.5 models, and GPT-5 Mini with official pricing
2026-01-18 01:08:12 +01:00
Aiden Cline
971e8734ae
Merge pull request #659 from KagurazakaNyaa/dev
...
Update SiliconFlow model list
2026-01-16 20:35:02 -08:00
Aiden Cline
ef7731ec13
Merge pull request #661 from cgilly2fast/cgilly2fast/firmware-provider
...
fix: make gpt-nano and mini calculate as 0 price
2026-01-16 20:32:32 -08:00
Colby Gilbert
e2d8670828
chore: update firmware provider docs url
2026-01-16 17:04:53 -08:00
神楽坂·喵
ecfab1717a
Merge branch 'anomalyco:dev' into dev
2026-01-17 08:21:48 +08:00
KagurazakaNyaa
ac037c57ad
fix pangu family
2026-01-17 08:20:30 +08:00
KagurazakaNyaa
a886d60715
fix kat family
2026-01-17 08:11:32 +08:00
Colby Gilbert
44e980e810
fix: make gpt-nano and mini calculate as 0 price
2026-01-16 15:42:37 -08:00
Aiden Cline
d1f3ddfe44
Merge pull request #660 from jerilynzheng/feat/vercel-models-update-2
...
vercel: add new models from Vercel AI Gateway
2026-01-16 12:59:21 -08:00
Aiden Cline
f0192d8759
Update gpt-5.2-codex.toml
2026-01-16 14:52:45 -06:00
jerilynzheng
070fd89c38
vercel: add new models from Vercel AI Gateway
...
- Add bytedance/seed-1.8 (multimodal with reasoning)
- Add openai/gpt-5.2-codex (agentic coding)
- Add recraft/recraft-v2 and recraft-v3 (image generation)
- Add recraft to model family schema
Co-Authored-By: Claude Opus 4.5 <noreply@anthropic.com >
2026-01-16 12:11:47 -08:00
神楽坂·喵
0ae941ee4d
Merge branch 'anomalyco:dev' into dev
2026-01-17 01:03:43 +08:00
KagurazakaNyaa
3a4cc60e04
update siliconflow model list
2026-01-17 01:02:33 +08:00
Aiden Cline
a243d9ba84
Merge pull request #658 from litvix-whale/feat/add-minimax-m2-1
...
feat(provider): add MiniMax M2.1 for DeepInfra
2026-01-16 08:15:39 -08:00
Kyrylo Lytvishko
5be35e44e4
feat(provider): add MiniMax M2.1 for DeepInfra
2026-01-16 18:04:00 +02:00
yinxulai
faa78aa42b
feat: add new Qiniu AI models - Claude 3.5/3.7/4.0/4.1/4.5 series, Gemini 2.0/2.5/3.0 series, GPT-5/5.2, Grok 4/4.1 series, and Kling v2-6
2026-01-16 17:46:45 +08:00
Aiden Cline
433008fef0
fix: more abacus things - fix model ids
2026-01-16 00:18:21 -06:00
franco
bfb6bb315a
feat: add moark provider
2026-01-16 10:26:06 +08:00
Aiden Cline
7f49452691
Merge pull request #653 from dpuyosa/UpdateModel
...
Venice: Update generate script & add new models (sonnet 4.5, gpt 5.2 codex)
2026-01-15 12:57:31 -08:00
Aiden Cline
6b793ad28e
rm raptor mini model
2026-01-15 12:42:06 -06:00
mthezi
53d77d2f2b
chore: update output limits for various models
2026-01-15 18:42:04 +08:00
dpuyosa
a8436a1e8a
Add new model configurations:
...
- Add claude-sonnet-45 model configuration
- Add openai-gpt-52-codex model configuration
2026-01-15 10:18:29 +01:00
dpuyosa
2ef222a882
Updated model configurations:
...
- Changed family from 'llama' to 'hermes' in hermes-3-llama-3.1-405b
- Changed family from 'glm' to 'glmv' in zai-org-glm-4.6v
- Added interleaved reasoning_details field in zai-org-glm-4.6v
2026-01-15 10:16:49 +01:00
dpuyosa
da6e0354da
Updated family inference logic:
...
- Refactored family inference to use ModelFamilyValues and subsequence matching algorithm
2026-01-15 10:14:09 +01:00
Aiden Cline
5aa046c596
fix: abacus provider
2026-01-14 23:58:31 -06:00
hueyexe
f54b8d8c6d
Add gpt 5.2 codex to azure cognitive services
2026-01-15 16:00:15 +11:00
hueyexe
a2d657f75b
Add gpt 5.2 codex to azure
2026-01-15 15:58:42 +11:00
yinxulai
79636dec83
fix: add required date fields and default output limits for Qiniu AI models
2026-01-15 10:49:52 +08:00
yinxulai
f9983aae19
feat: add Qiniu AI model definitions
...
- Add 49 OpenAI-compatible model definitions
- Models filtered from Qiniu API with OpenAI protocol support
- Include models from DeepSeek, Qwen, Kimi, GLM, Doubao, MiniMax, etc.
- No pricing information included (aggregation platform)
2026-01-15 10:38:15 +08:00
Aiden Cline
b9411cb00c
feat: add Qiniu AI provider configuration
2026-01-15 10:06:30 +08:00
Aiden Cline
5a329d79bc
Merge pull request #650 from cgilly2fast/cgilly2fast/firmware-provider
...
refactor: simplify model ids so sub agents work
2026-01-14 15:18:26 -08:00
Colby Gilbert
1e9ee75804
refactor: simplify model ids so sub agents work
2026-01-14 15:07:36 -08:00
Aiden Cline
64e82beb55
Merge pull request #645 from TheEpTic/dev
...
chore: Add gpt-5.2-codex to GitHub Copilot provider
2026-01-14 14:59:58 -08:00
Frank
256bab07a3
update zen models
2026-01-14 16:27:53 -05:00
Frank
78fd2e0fa0
update zen models
2026-01-14 16:18:51 -05:00
Aiden Cline
969430c25e
Merge pull request #647 from KonarkRajMisra/dev
...
Add GPT-5.2-Codex to OpenRouter
2026-01-14 12:39:08 -08:00
Aiden Cline
c4b43c090d
Merge pull request #646 from brandon93s/52-input
...
chore(openai): gpt-5.2-codex input limit
2026-01-14 12:38:52 -08:00
Konark Misra
bfb92b5e46
Add GPT-5.2-Codex to OpenRouter
2026-01-14 12:15:39 -08:00
TheEpTic
b0e5b914c8
Fix context size
2026-01-14 20:03:52 +00:00
Brandon Smith
66e5d76e05
input
2026-01-14 13:53:57 -06:00
TheEpTic
f1f27989d8
Add gpt-5.2-codex to GitHub Copilot provider
2026-01-14 19:41:09 +00:00
Aiden Cline
949f9b9909
Merge pull request #623 from cyhhao/add-gpt-5-2-codex
...
feat: add gpt-5.2-codex model
2026-01-14 11:25:43 -08:00
Aiden Cline
6a614ab0ac
Update model family name in gpt-5.2-codex.toml
2026-01-14 13:24:47 -06:00
Aiden Cline
664079661d
Merge pull request #641 from liyishuai/iflow-cleanup
...
chore(iflowcn): cleanup models
2026-01-14 07:46:53 -08:00
Aiden Cline
58e2fd8462
Merge pull request #642 from brandon93s/openai-codex-input-limit
...
openai: codex input context limit
2026-01-14 07:31:37 -08:00
Aiden Cline
25eda4cc82
Merge pull request #612 from Alex-wuhu/dev
...
add LLM Provider : novita ai
2026-01-14 07:30:55 -08:00
Alex-wuhu
c60ec95e75
Update model family names for consistency and clarity
2026-01-14 23:04:51 +08:00
Alex
952de0d081
Merge branch 'anomalyco:dev' into dev
2026-01-14 23:00:39 +08:00
Brandon Smith
453f16ce42
add input limit for codex models
2026-01-14 08:32:48 -06:00
Alex-wuhu
01f338231e
Update LLM info
2026-01-14 19:10:34 +08:00
Yishuai Li
ce48f4ee7b
chore(iflowcn): cleanup models
...
Signed-off-by: Yishuai Li <yishuai.li@pingcap.com >
2026-01-14 16:53:58 +08:00
Aiden Cline
db79e08e38
Merge pull request #636 from Eric-Guo/patch-1
...
Using CN in API key, so it won't loading both siliconflow-cn and siliconflow
2026-01-13 21:36:47 -08:00
Aiden Cline
71cf624135
Merge pull request #640 from fanweixiao/dev
...
feat(provider): Add configuration for GPT-5.1 Codex Max model to Vivgrid provider
2026-01-13 21:36:36 -08:00
Aiden Cline
9f7c0cec79
Merge pull request #639 from anomalyco/update-model-families
...
Update model families
2026-01-13 21:36:21 -08:00
C.C.
399b469927
Add configuration for GPT-5.1 Codex Max model
2026-01-14 02:55:08 +00:00
Aiden Cline
1a96ad9764
Merge pull request #637 from dpuyosa/UpdateModel
...
Venice: Updated llama-3.2-3b model configuration
2026-01-13 15:00:59 -08:00
Jérôme Benoit
0ada0ed52e
fix(sap-ai-core): use temporary fork for stable OpenCode integration
2026-01-13 19:52:14 +01:00
dpuyosa
4c39b53744
Updated llama-3.2-3b model configuration:
...
- Removed structured_output property
2026-01-13 13:52:50 +01:00
Eric Guo
d84aff0e75
Using CN in API key, so it won't loading both siliconflow-cn and siliconflow
2026-01-13 20:16:29 +08:00
mthezi
c9aefb0af1
feat: add 302ai provider
2026-01-13 14:21:52 +08:00
cyhhao
94310d742d
Add gpt-5.2-codex model
2026-01-11 01:55:50 +08:00
Alex-wuhu
a50c04d060
Update minimax-m2.1.toml
2026-01-09 13:31:49 +08:00
Alex-wuhu
cc2619dd5f
add LLM Provider : novita ai
2026-01-07 19:27:42 +08:00
Burak Varlı
f14775e355
Add cross-region inference profiles for Claude 4.x family models in Amazon Bedrock
...
Amazon Bedrock requires usage of cross-region inference for some models, especially the latest models including all Claude 4.x family.
This change creates model files for all Claude 4.x models for cross-region inference profiles for Global, US and EU.
2026-01-06 11:41:49 +00:00
Dominik Oswald
c35c42fc34
Add Sonar Deep Research model configuration
...
- Introduce TOML configuration for Perplexity Sonar Deep Research model
- Include token pricing, request fees, and model limits
- Follow OpenCode AI schema conventions for model definitions
2025-10-17 13:12:19 +02:00
Dominik Oswald
8abedde07c
Add Perplexity Sonar Deep Research model configuration
...
- Introduce TOML configuration for Perplexity Sonar Deep Research model
- Include token pricing, request fees, and model limits
2025-10-17 13:10:38 +02:00