Aiden Cline
cda892e0d7
sync github copilot limits
2026-03-19 21:54:42 -05:00
Aiden Cline
098ff4f5bf
Merge pull request #1227 from Verizane/dev
...
add OpenRouter models for gpt-5.4 mini and gpt-5.4 nano
2026-03-19 21:23:15 -05:00
Aiden Cline
0f70b8959f
Merge pull request #1234 from mchenco/dev
...
Add Workers AI models: kimi-k2.5, nemotron-3-120b-a12b, glm-4.7-flash
2026-03-19 15:11:19 -05:00
mchen
b8e6d58e5b
add workers-ai models: kimi-k2.5, nemotron-3-120b-a12b, glm-4.7-flash
2026-03-19 14:59:17 -04:00
Roman Koslowski
a855001a7e
apply changes from review
2026-03-19 17:20:26 +01:00
Aiden Cline
ac760b2268
Merge pull request #1230 from SamizuHM/feature/zhipuai-coding-plan-add-glm-5-turbo
...
zhipuai-coding-plan: Add glm-5-turbo.toml and replace symlink
2026-03-19 10:42:43 -05:00
Aiden Cline
d4a5ea7ae7
Merge pull request #1226 from spiffytech/dev
...
Add Ollama Cloud support for Minimax M2.7
2026-03-19 10:41:47 -05:00
Aiden Cline
434ed89ba2
Merge pull request #1228 from dpuyosa/minimax_m2_7
...
Venice: Add MiniMax M2.7 and update DeepSeek V3.2 pricing
2026-03-19 10:41:16 -05:00
Aiden Cline
6d7719a62a
Merge pull request #1229 from 0b1000/dev
...
Xiaomi: Add MiMo-V2-Pro and MiMo-V2-Omni
2026-03-19 10:41:06 -05:00
Aiden Cline
93637039ef
Merge pull request #1231 from ariane-emory/feat/feat/add-xiaomi-mimo-v2-pro-and-omni
...
feat: add the Xiaomi MiMo V2 Pro and Xiaomi MiMo V2 Omni models to the OpenRouter provide
2026-03-19 10:40:44 -05:00
Ariane Emory
9c95f796c0
Merge remote-tracking branch 'upstream/dev' into feat/feat/add-xiaomi-mimo-v2-pro
2026-03-19 11:22:43 -04:00
Ariane Emory
e8650b6073
feat: add xiaomi mimo-v2-pro and mimo-v2-omni models to openrouter
2026-03-19 11:18:46 -04:00
SamizuHM
23c2be6ff7
feat(zhipuai-coding-plan): add glm-5-turbo.toml and replace glm-5-turbo with symlink
2026-03-19 18:09:17 +08:00
Frank
913a63dbe6
update zen models
2026-03-19 00:33:45 -04:00
0b1000
503087e99b
Merge branch 'anomalyco:dev' into dev
2026-03-19 12:28:38 +08:00
0b1000
48150f09d3
Xiaomi: Add MiMo-V2-Pro and MiMo-V2-Omni
2026-03-19 12:27:00 +08:00
Aiden Cline
5fef681657
Disable tool_call in grok model configuration
2026-03-18 23:09:30 -05:00
Frank
123054ae0c
update zen models
2026-03-18 20:45:44 -04:00
Frank
03060d154b
update zen models
2026-03-18 20:37:47 -04:00
dpuyosa
5c9b8108e0
Update minimax-m27.toml
2026-03-19 01:02:24 +01:00
dpuyosa
c8084681f9
[venice] Add MiniMax M2.7 and update DeepSeek V3.2 pricing
...
- Add MiniMax M2.7 model with reasoning and tool_call support
- Update DeepSeek V3.2 pricing (input: $0.33, output: $0.48, cache: $0.16)
2026-03-19 00:58:50 +01:00
Roman Koslowski
352ab4ae1b
add gpt-5.4 mini and gpt-5.4 nano
2026-03-18 22:16:55 +01:00
spiffytech
cf0b416b15
Added Ollama Cloud support for Minimax M2.7
2026-03-18 16:15:07 -04:00
Aiden Cline
38339a2a90
Merge pull request #1224 from APonce911/minimax-m2.7-openrouter
...
add MiniMax M2.7 to OpenRouter
2026-03-18 14:10:13 -05:00
Aiden Cline
ff9040bf52
Update minimax-m2.7.toml
2026-03-18 14:09:26 -05:00
Aiden Cline
3039804af4
Delete providers/opencode/models/minimax-m2.7.toml
2026-03-18 14:08:55 -05:00
Frank
7a4ad7bec8
update go models
2026-03-18 14:40:25 -04:00
airton
721cc122bc
add MiniMax M2.7 to OpenRouter and OpenCode
2026-03-18 18:57:09 +01:00
Aiden Cline
0527f019af
Merge pull request #1221 from sergical/fix/bedrock-claude-4-6-context-window-and-pricing
...
fix(amazon-bedrock): set Claude Sonnet 4.6 and Opus 4.6 context window to 1M
2026-03-18 12:17:57 -05:00
Aiden Cline
c89371de50
Merge pull request #1223 from sylviezhang37/update-vercel-models-20260318-1659
...
Update Vercel models
2026-03-18 12:17:22 -05:00
Sylvie Zhang
6d6d4220d8
Enable open_weights in minimax-m2.7.toml
2026-03-18 10:12:37 -07:00
Sylvie Zhang
8b984eeec1
Enable open_weights in minimax-m2.7-highspeed model
2026-03-18 10:12:21 -07:00
github-actions[bot]
586027c8f1
chore(vercel): update Vercel model definitions
...
Auto-generated by weekly workflow from Vercel AI Gateway API.
Co-Authored-By: github-actions[bot] <41898282+github-actions[bot]@users.noreply.github.com>
2026-03-18 16:59:24 +00:00
Sergiy Dybskiy
343b5f87ef
fix(amazon-bedrock): set Claude Sonnet 4.6 and Opus 4.6 context window to 1M
...
Both models support a 1M token context window natively on Bedrock via the
Converse API with no beta headers required. Verified empirically via the
AWS CLI (bedrock-runtime converse): 950K tokens succeeds, >1M returns
'prompt is too long: N tokens > 1000000 maximum'.
The AWS Bedrock pricing page confirms long context pricing for these two
models is identical to standard pricing (no surcharge), so the
[cost.context_over_200k] section is removed as it was incorrect.
2026-03-18 12:20:38 -04:00
Aiden Cline
955b773ee5
Merge pull request #1218 from pomidornijfrukt/azure/5.4-mini-nano
...
Add GPT-5.4 Mini and Nano models for Azure providers
2026-03-18 10:31:13 -05:00
Aiden Cline
98559071f0
Merge pull request #1217 from cgilly2fast/dev
...
chore(firmware): update base url and docs url
2026-03-18 10:30:44 -05:00
eCube-cachy
0660308816
add: GPT-5.4 Mini and Nano model configurations for Azure providers
2026-03-18 15:17:52 +02:00
Jack
380f9dd8eb
Merge pull request #1216 from no1wudi/dev
...
Add MiniMax M2.7 and M2.7-highspeed models to 4 official providers
2026-03-18 16:29:59 +08:00
Jack
1cfdab1b18
update MiniMax-M2.7 cache_read to 0.06
2026-03-18 16:27:53 +08:00
Colby Gilbert
75a981f957
chore(firmware): update base url and docs url
2026-03-18 00:41:05 -07:00
Huang Qi
7fadbcadc8
Add MiniMax M2.7 and M2.7-highspeed models to 4 official providers
2026-03-18 15:21:06 +08:00
Frank
38f9092292
update zen models
2026-03-18 02:30:18 -04:00
Aiden Cline
92149b9eaa
rm nonexistant github model
2026-03-17 21:41:51 -05:00
Aiden Cline
b614f0e69c
Merge pull request #1214 from luisrudge/dev
...
Add GPT-5.4 mini and nano to GitHub Copilot provider
2026-03-17 20:13:46 -05:00
Luís Rudge
67d6dac5c5
Add GPT-5.4 mini and nano to GitHub Copilot provider
2026-03-17 18:44:38 -06:00
Aiden Cline
7d3cc61a48
Merge pull request #1207 from PedroACosta/feat/add-dinference-provider
...
feat(providers): add dinference provider
2026-03-17 14:51:31 -05:00
Aiden Cline
f02ea6c4d2
Merge pull request #1115 from skywalker512/feat/add-tencent-coding-plan
...
feat: add Tencent Coding Plan provider
2026-03-17 14:51:19 -05:00
Aiden Cline
0cb50eeece
Merge pull request #1208 from scwgoire/march-update
...
Scaleway 26-03 model updates
2026-03-17 14:48:12 -05:00
Aiden Cline
878311d2e0
Merge pull request #1210 from dm-cohere/dm/fix-update-cohere-model-capabilities
...
fix(models): update cohere model capabilities
2026-03-17 14:32:27 -05:00
Aiden Cline
a0e89f65d6
Merge pull request #1206 from 0b1000/dev
...
Rename minimax-m2.5.toml to MiniMax-M2.5.toml
2026-03-17 14:32:19 -05:00
Aiden Cline
74099b7c9c
Merge pull request #1213 from smrdotgg/add-openai-gpt-5-4-mini-and-nano
...
Add OpenAI GPT-5.4 mini and nano
2026-03-17 14:31:24 -05:00
Aiden Cline
ec522435c3
Merge pull request #1211 from sylviezhang37/update-vercel-models-20260317-1807
...
Update Vercel models
2026-03-17 14:30:24 -05:00
smr
d839cd37d4
Add OpenAI GPT-5.4 mini and nano
...
Capture the newly released mini and nano model metadata so models.dev reflects OpenAI's latest GPT-5.4 lineup with current pricing, limits, and knowledge cutoff.
2026-03-17 22:09:13 +03:00
github-actions[bot]
ecb6ef7f93
chore(vercel): update Vercel model definitions
...
Auto-generated by weekly workflow from Vercel AI Gateway API.
Co-Authored-By: github-actions[bot] <41898282+github-actions[bot]@users.noreply.github.com>
2026-03-17 18:07:06 +00:00
Deirdre Meehan
8b4d341054
fix: cohere models on non-cohere providers
2026-03-17 16:51:24 +00:00
Deirdre Meehan
62f4a28308
fix: cohere provider models
2026-03-17 16:44:55 +00:00
Pedro
2fb8ef0dc8
feat(providers): add dinference provider
2026-03-17 14:13:36 +01:00
Gregoire de Turckheim
96968e2bf8
feat: Scaleway 26-03 model updates
2026-03-17 12:15:52 +01:00
0b1000
ee9d7879ce
Rename minimax-m2.5.toml to MiniMax-M2.5.toml
2026-03-17 14:50:35 +08:00
Frank
71283512a6
update zen models
2026-03-17 02:21:13 -04:00
Frank
cd4afd7e7c
update zen models
2026-03-17 02:19:17 -04:00
Aiden Cline
1239d0190b
Merge pull request #1204 from cyberofficial/vultr
...
VULTR: Updated Vultr model pricing to reflect current serverless inference rates
2026-03-16 16:10:39 -05:00
Aiden Cline
491bf6ccba
Merge pull request #1202 from RaviTharuma/fix/chutes-pricing-update-2026-03
...
fix(chutes): update pricing and limits from live API
2026-03-16 16:10:25 -05:00
Cyber Official
c993d0c121
Updated Vultr model pricing to reflect current serverless inference rates
...
Updated Vultr model pricing to reflect current serverless inference rates
This commit updates the cost configuration for all Vultr models to align with their latest pricing tiers:
**Cost Reductions:**
- DeepSeek-R1-Distill-Qwen-32B: Input $0.55→$0.30, Output $2.75→$0.30 (73% reduction)
- NVIDIA-Nemotron-3-Super-120B-A12B-NVFP4: Input $0.55→$0.20, Output $2.75→$0.80 (64% input, 71% output reduction)
- Qwen2.5-Coder-32B-Instruct: Input $0.55→$0.20, Output $2.75→$0.60 (64% input, 78% output reduction)
- gpt-oss-120b: Input $0.55→$0.15, Output $2.75→$0.60 (73% input, 78% output reduction)
- MiniMax-M2.5: Input $0.55→$0.30, Output $2.75→$1.20 (45% input, 56% output reduction)
**Cost Adjustments:**
- DeepSeek-R1-Distill-Llama-70B: Input $0.55→$2.00, Output $2.75→$2.00 (significant increase)
- DeepSeek-V3.2: Output $2.75→$1.65 (40% reduction)
- Llama-3.1-Nemotron-Ultra-253B-v1: Output $2.75→$1.80 (35% reduction)
- GLM-5-FP8: Input $0.55→$0.85, Output $2.75→$3.10 (55% input, 13% output increase)
2026-03-16 13:58:28 -04:00
Aiden Cline
e55c39a83d
Merge pull request #1141 from sk0x0y/feature/nanogpt-confirmed-suffix2-fixes
...
fix(nano-gpt): rename confirmed 2-suffix model ids
2026-03-16 10:57:44 -05:00
Aiden Cline
6dea000e25
Merge pull request #1148 from sk0x0y/feature/nanogpt-bundled-confirmed-suffix2-fixes
...
fix(nano-gpt): rename bundled confirmed 2-suffix model ids
2026-03-16 10:57:30 -05:00
Aiden Cline
3f7a757b3f
Merge pull request #1142 from sk0x0y/feature/nanogpt-more-confirmed-suffix2-fixes
...
fix(nano-gpt): rename more confirmed 2-suffix model ids
2026-03-16 10:56:36 -05:00
Aiden Cline
c693fd71e2
Merge pull request #1194 from cyberofficial/vultr
...
Update Vultr model list with 10 new models and updated pricing
2026-03-16 10:55:19 -05:00
Aiden Cline
54e04e288a
Merge pull request #1198 from amritbanerjee/add-glm-5-turbo
...
Add GLM-5-Turbo model support
2026-03-16 10:47:08 -05:00
Aiden Cline
462a179eee
Merge pull request #1203 from jerome-benoit/fix/sap-ai-core-model-specs
...
fix(sap-ai-core): align model specs with official sources
2026-03-16 10:46:40 -05:00
Aiden Cline
95db59034d
Merge pull request #1201 from dpuyosa/venice-new-models
...
Venice: Add new provider models
2026-03-16 10:45:59 -05:00
Aiden Cline
74dcc74e32
Merge pull request #1200 from dpuyosa/venice/pricing-update
...
Venice: Update model pricing for 7 models
2026-03-16 10:45:47 -05:00
Jérôme Benoit
57975f5f25
fix(sap-ai-core): align model specs with official sources
2026-03-16 13:59:06 +01:00
Ravi Tharuma
ad7b063747
fix(chutes): update pricing and limits from live API
...
Synced 6 Chutes model definitions against the live API at
https://llm.chutes.ai/v1/models (queried 2026-03-16).
Models updated:
- deepseek-ai/DeepSeek-V3.2-TEE: cost 0.25/0.38→0.28/0.42, cache 0.125→0.14, context 163840→131072
- zai-org/GLM-5-TEE: cost 0.75/2.5→0.95/3.15, added cache_read 0.475
- zai-org/GLM-4.6-TEE: cost 0.35/1.5→0.4/1.7, added cache_read 0.2
- zai-org/GLM-4.6V: added cache_read 0.15
- MiniMaxAI/MiniMax-M2.5-TEE: cost 0.15/0.6→0.3/1.1, added cache_read 0.15
- Qwen/Qwen3.5-397B-A17B-TEE: cost 0.3/1.2→0.39/2.34, cache 0.15→0.195
2026-03-16 11:42:29 +01:00
dpuyosa
f76e9f0551
[venice] Add new provider models
...
- Add mistral-small-3.2-24b-instruct, qwen3-5-9b, venice-uncensored-role-play, zai-org-glm-4.6
2026-03-16 09:37:41 +01:00
dpuyosa
d70a49b36f
[venice] Update model pricing for 7 models
...
- Remove context_over_200k pricing from Claude models
- Update Grok cache_read pricing from 0.5 to 0.25
- Update Kimi, MiniMax input/output pricing
2026-03-16 09:05:37 +01:00
amrit
3487135f9f
Add GLM-5-Turbo model support
2026-03-16 12:14:50 +11:00
Aiden Cline
458a66c766
Merge pull request #1197 from kesku/update-perplexity-agent-models
...
Update Perplexity Agent API models
2026-03-15 10:59:23 -05:00
Frank
d3a84dc7ec
update zen models
2026-03-15 10:59:52 -04:00
Kesku
ae61b25583
update perplexity-agent: add gpt-5.4 & nemotron, remove gemini-3-pro
2026-03-15 03:46:50 +00:00
Aiden Cline
74be576eda
Merge pull request #1178 from Sewer56/change-synthetic-endpoint
...
Add OpenAI and Anthropic compatible endpoints
2026-03-14 20:55:30 -05:00
Aiden Cline
164df2cda0
Merge pull request #1191 from Alcatraz-Zhang/update/kilo-models
...
Sync Kilo model definitions with latest gateway catalog
2026-03-14 20:54:45 -05:00
Cyber Official
2cd7908369
Update Vultr model list with 10 new models and updated pricing
...
- Updated pricing to $0.55/M input tokens, $2.75/M output tokens
- Updated context limits to safe floor values from official testing
- Added accurate output token limits from official model documentation
- Added 5 new models: MiniMax M2.5, DeepSeek V3.2, GLM-5 FP8, Llama 3.1 Nemotron Ultra 253B, NVIDIA Nemotron 3 Super 120B A12B NVFP4
- Updated existing models: DeepSeek R1 Distill variants, GPT OSS 120B, Kimi K2.5, Qwen2.5 Coder 32B
Model specifications:
- MiniMax M2.5: 196K context, 4,096 output
- Qwen2.5-Coder-32B: 15K context, 256 output (notable low default)
- DeepSeek R1 Distill Llama 70B: 130K context, 4,096 output
- DeepSeek R1 Distill Qwen 32B: 130K context, 4,096 output
- DeepSeek V3.2: 163K context, 4,096 output
- Kimi K2.5: 261K context, 32,768 output (high output limit)
- GPT OSS 120B: 130K context, 8,192 output
- GLM-5 FP8: 202K context, 131,072 output (exceptionally high)
- Llama 3.1 Nemotron Ultra 253B: 32K context, 4,096 output
- NVIDIA Nemotron 3 Super 120B A12B NVFP4: 260K context, 8,192 output
All models set to text-only (no vision support) as confirmed.
2026-03-14 19:47:01 -04:00
Alcatraz-Zhang
cc667340f5
Sync Kilo model definitions with latest gateway catalog
...
Refresh the Kilo provider catalog so models.dev matches the current gateway inventory, pricing, and availability.
2026-03-15 04:35:38 +08:00
Sewer56
f2cfc1435d
Changed: Synthetic to use newer openai endpoint
2026-03-14 17:09:44 +00:00
Aiden Cline
35bb8cca47
Merge pull request #1172 from bigfluffycookie/add-deepinfra-llama-models
...
Add deepinfra llama models
2026-03-14 10:55:13 -05:00
Aiden Cline
3468a410e1
Merge pull request #1177 from ar27111994/dev
...
Add Grok 4.1 Fast configurations for reasoning and non-reasoning
2026-03-14 10:54:57 -05:00
Aiden Cline
b1b5e3c5cd
Merge pull request #1174 from dacbd/patch-1
...
fix(wandb): fix k2.5 settings
2026-03-14 10:54:35 -05:00
Aiden Cline
97f03ec672
Merge pull request #1175 from dacbd/patch-2
...
chore(docs): add note for manual testing with opencode
2026-03-14 10:54:22 -05:00
BigFluffyCookie
9b516924aa
Add limit output for llama models
2026-03-14 11:49:57 +01:00
Ahmed Rehan
929a39600b
feat(models): add Grok 4.1 Fast (Reasoning and Non-Reasoning) configurations
2026-03-14 14:27:24 +05:00
Daniel Barnes
a87d8bb8cc
chore(docs): add note for manual testing with opencode
2026-03-14 13:42:57 +09:00
Daniel Barnes
574139eb49
fix(wandb): fix k2.5 settings
2026-03-14 13:07:16 +09:00
Aiden Cline
1e3bc38b31
Merge pull request #1137 from mcowger/mcowger/correct-gemini-flash-lite-pricing
...
Fix incorrect pricing for gemini-3.1-flash-lite-preview
2026-03-13 18:41:41 -05:00
Aiden Cline
8916fe9874
Merge pull request #1171 from stephenkuhn214/dev
...
fix(amazon-bedrock): Remove deprecated and add missing models
2026-03-13 18:26:25 -05:00
BigFluffyCookie
5d956b41a6
Rename llama models to remove "Meta" prefix
2026-03-13 23:15:27 +01:00
BigFluffyCookie
42a7a14f69
Add Meta Llama models to DeepInfra provider
2026-03-13 22:53:39 +01:00
Stephen Kuhn
f24ee000d7
fix(amazon-bedrock): update and add models
...
- Remove 19 deprecated/EOL models
- Add 7 new models: DeepSeek V3.2, Llama 3.1 405B, Magistral Small 1.2, Ministral 3 3B, Mistral Large 3, Pixtral Large, NVIDIA Nemotron Nano 3 30B
- Fix Devstral 2 123B: correct name, family, and open_weights
- Set accurate Bedrock launch dates for all new models
2026-03-13 16:02:04 -04:00
Aiden Cline
7196b1fb2c
Merge pull request #1170 from anomalyco/revert-1166-fix/update-gpt53-codex-spark-preview
...
Revert "fix(openai): rename gpt-5.3-codex-spark to gpt-5.3-codex-spark-preview"
2026-03-13 14:31:33 -05:00
Aiden Cline
f6c0d5a29d
Revert "fix(openai): rename gpt-5.3-codex-spark to gpt-5.3-codex-spark-preview"
2026-03-13 14:30:58 -05:00
Aiden Cline
ee63449aa5
sonnet 4.6 and opus 4.6 1M context
2026-03-13 14:27:55 -05:00
Aiden Cline
92aa44ec00
Merge pull request #1166 from rluisr/fix/update-gpt53-codex-spark-preview
...
fix(openai): rename gpt-5.3-codex-spark to gpt-5.3-codex-spark-preview
2026-03-13 14:18:41 -05:00
Aiden Cline
477284535c
Rename model from 'GPT-5.3 Codex Spark Preview' to 'GPT-5.3 Codex Spark'
2026-03-13 14:17:44 -05:00
Aiden Cline
304233bdda
Merge pull request #1169 from mdrxy/mdrxy/anthropic-token-limits
...
Update Claude 4.6 context/pricing
2026-03-13 14:13:40 -05:00
Aiden Cline
25d782ee2c
Reduce context limit from 1,000,000 to 200,000
2026-03-13 14:13:30 -05:00
Aiden Cline
0f63393d51
Update context limit in claude-opus-4-6.toml
2026-03-13 14:12:56 -05:00
rluisr
e780eefce2
fix(openai): rename gpt-5.3-codex-spark to gpt-5.3-codex-spark-preview
...
The OpenAI API expects model ID 'gpt-5.3-codex-spark-preview', not
'gpt-5.3-codex-spark'. Rename model files in both openai and opencode
providers so the generated model ID matches the actual API.
2026-03-14 03:59:03 +09:00
Aiden Cline
a79585fa83
Merge pull request #1163 from micuintus/feature/Kimi2.5-fast
...
feat(nebius): add Kimi-K2.5-fast model
2026-03-13 13:14:38 -05:00
Aiden Cline
00801f74f2
Merge pull request #1164 from butyess/dev
...
Openrouter models: gemini 3.1 flash lite preview, grok 4.20 beta models.
2026-03-13 13:14:22 -05:00
Aiden Cline
185f6731ee
Merge pull request #1162 from dpuyosa/feature/venice-grok-4-20-beta
...
Venice: Add Grok 4.20 Beta models
2026-03-13 12:53:28 -05:00
Aiden Cline
d291b0575c
Merge pull request #1167 from sylviezhang37/update-vercel-models-20260313-1639
...
Update Vercel models
2026-03-13 12:53:11 -05:00
Mason Daugherty
382d9f3e7d
Update Claude 4.6 context/pricing
2026-03-13 13:53:04 -04:00
Aiden Cline
e64f5fe963
Merge pull request #1168 from mdrxy/mdrxy/update-baseten
...
Update Baseten models
2026-03-13 12:51:56 -05:00
Mason Daugherty
ea57ddfe7e
Update Baseten models
2026-03-13 13:48:41 -04:00
github-actions[bot]
29463d7fa8
chore(vercel): update Vercel model definitions
...
Auto-generated by weekly workflow from Vercel AI Gateway API.
Co-Authored-By: github-actions[bot] <41898282+github-actions[bot]@users.noreply.github.com>
2026-03-13 16:39:30 +00:00
Jack
bcc8db49ee
Merge pull request #1165 from anomalyco/chore/openrouter-alpha-reasoning-details-20260313
...
feat(openrouter): add interleaved reasoning details for alpha models
2026-03-13 22:25:25 +08:00
Jack
c8521d70f3
feat(openrouter): add interleaved reasoning details for alpha models
2026-03-13 22:20:54 +08:00
Federico Masi
490cd249e4
Openrouter models: gemini 3.1 flash lite preview, grok 4.20 beta models.
2026-03-13 15:11:12 +01:00
Michael Voigt
dbc636f5f3
feat(nebius): add Kimi-K2.5-fast model
2026-03-13 12:35:13 +01:00
Michael Voigt
9a32f671a1
fix(nebius): lowercase model ID for Nemotron-3-Super-120B-A12B
...
The filename must match the API casing (lowercase) to avoid 'model does not exist' errors.
2026-03-13 12:35:08 +01:00
dpuyosa
856d925eda
[venice] Add Grok 4.20 Beta models
...
- Add Grok 4.20 Beta model configuration (2M context, 128K output)
- Add Grok 4.20 Multi-Agent Beta model configuration
2026-03-13 10:48:38 +01:00
Aiden Cline
066a425917
Merge pull request #1158 from micuintus/feature/Nebius_Nemotron-3-Super-120b-a12b
...
feat(nebius): Add support for Nemotron-3-Super-120B-A12B
2026-03-12 22:20:20 -05:00
Aiden Cline
6df7f20cdc
Merge pull request #1156 from dsingal0/dev
...
added nemotron super on baseten
2026-03-12 22:20:06 -05:00
Aiden Cline
78bb47b90e
Merge pull request #1151 from dacbd/dacbd
...
fix(wandb): update models
2026-03-12 22:19:43 -05:00
Aiden Cline
c121d86419
Merge pull request #1160 from kreatoo/dev
...
feat: add zai-org/glm-4.7 and zai-org/glm-4.7-flash to NanoGPT
2026-03-12 22:11:18 -05:00
Aiden Cline
ab148eeb14
Merge pull request #1161 from Grin1024/dev
...
Add Claude Opus 4.6 and Sonnet 4.6 models to RequestY provider
2026-03-12 22:11:07 -05:00
lihui
49d196d326
Add Claude Opus 4.6 and Sonnet 4.6 models to RequestY provider
2026-03-13 09:00:54 +08:00
Kreato
8899b390ef
feat: add zai-org/glm-4.7 and zai-org/glm-4.7-flash to NanoGPT
2026-03-13 00:27:09 +03:00
Michael Voigt
5217f62ddf
fix(nebius): Follow context updates for Kimi 2.5 and GLM-5
2026-03-12 20:22:48 +01:00
Michael Voigt
55eaff9af1
feat(nebius): Add support for Nemotron-3-Super-120B-A12B
2026-03-12 20:22:21 +01:00
Dhruv Singal
7557c06ac0
update output length
2026-03-12 09:41:25 -07:00
Dhruv Singal
e85d820121
fix input output
2026-03-12 08:29:01 -07:00
Dhruv Singal
499d3a39ef
remove cache pricing
2026-03-12 08:21:22 -07:00
Dhruv Singal
b9b38d6e33
added nemotron super on baseten
2026-03-12 08:18:46 -07:00
Aiden Cline
ca24ac14fa
Merge pull request #1153 from dpuyosa/dev
...
Venice: Update model output token limits
2026-03-12 10:08:46 -05:00
Aiden Cline
822546fc67
Merge pull request #1155 from spiffytech/dev
...
Add Ollama Cloud support for Nemotron 3 Super
2026-03-12 10:08:31 -05:00
Aiden Cline
4555195b71
Merge pull request #1152 from v1gnesh/dev
...
Update grok-4.20 model defs
2026-03-12 10:08:15 -05:00
spiffytech
5eae8effc6
Added Ollama Cloud support for Nemotron 3 Super
2026-03-12 09:28:47 -04:00
dpuyosa
c1801aef87
[venice] Normalize model output token limits
...
- Update output limits to standard values across all models
2026-03-12 10:08:39 +01:00
v1gnesh
5e6464b272
Update grok-4.20-beta-reasoning
2026-03-12 10:27:40 +05:30
v1gnesh
e1a4f23332
Update grok-4.20-beta-non-reasoning
2026-03-12 10:26:03 +05:30
v1gnesh
753e1f9f0c
grok-multi-agent-beta update
2026-03-12 10:23:57 +05:30
Daniel Barnes
123ecd2ba5
docs url
2026-03-12 13:27:56 +09:00
Daniel Barnes
f15cda9fcb
remove old
2026-03-12 13:26:08 +09:00
Daniel Barnes
0205debbd3
fix values
2026-03-12 13:22:29 +09:00
Daniel Barnes
0059766509
number formating
2026-03-12 13:17:22 +09:00
Daniel Barnes
be81b02916
additional model files
2026-03-12 13:02:17 +09:00
Daniel Barnes
2dab141166
initial script & model updates
2026-03-12 13:01:35 +09:00
Aiden Cline
45aa49af25
tweak: azure kimi k2.5
2026-03-11 22:35:20 -05:00
Aiden Cline
781fad3ad4
Merge pull request #1150 from cau1k/5.4-family
...
feat(azure): add 5.4/pro families
2026-03-11 22:14:08 -05:00
cau1k
99d2ffcfdd
feat(azure): add 5.4/pro families
2026-03-11 20:59:11 -04:00
Aiden Cline
381d7cc19d
Merge pull request #1149 from ariane-emory/fear/add-march-or-stealth-models
...
Add OpenRouter stealth models: Hunter Alpha and Healer Alpha
2026-03-11 18:07:50 -05:00
Ariane Emory
7482e22458
Fix family field to use 'alpha' for stealth models
2026-03-11 18:49:32 -04:00
Ariane Emory
f5e6a402e6
Add OpenRouter stealth models: Hunter Alpha and Healer Alpha
2026-03-11 18:41:58 -04:00
Aiden Cline
9265852852
tweak: adjust some gh limits to align better w/ api
2026-03-11 15:23:44 -05:00
Aiden Cline
dc98a32996
Merge pull request #1018 from Sewer56/add-synthetic-missing-models
...
Update synthetic.new models: promote MiniMax-M2.5, add GLM-4.7-Flash
2026-03-11 14:55:50 -05:00
Aiden Cline
56c39ae0f6
Merge pull request #1140 from sk0x0y/feature/nanogpt-thudm-id-fixes
...
fix(nano-gpt): rename THUDM 2 ids to canonical THUDM ids
2026-03-11 14:55:07 -05:00
Aiden Cline
b1f43a7595
Merge pull request #1147 from msadiks/fix/alibaba-coding-minimax
...
fix: alibaba-coding-plan MiniMax-M2.5 context window
2026-03-11 14:54:37 -05:00
Matt Cowger
fed8bcae19
Merge branch 'dev' into mcowger/correct-gemini-flash-lite-pricing
2026-03-11 12:23:42 -07:00
sk0x0y
fb95150d02
fix(nano-gpt): rename VongolaChouko model id
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-12 04:20:39 +09:00
sk0x0y
a7c9a240b4
fix(nano-gpt): rename Steelskull model ids
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-12 04:20:39 +09:00
sk0x0y
6432a4a3e6
fix(nano-gpt): rename Sao10K model ids
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-12 04:20:38 +09:00
sk0x0y
f2e4a249fe
fix(nano-gpt): rename NeverSleep model ids
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-12 04:20:38 +09:00
sk0x0y
8667a6eed8
fix(nano-gpt): rename MarinaraSpaghetti model id
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-12 04:19:56 +09:00
sk0x0y
429554397a
fix(nano-gpt): rename LatitudeGames model id
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-12 04:19:56 +09:00
sk0x0y
a64e6ad0ac
fix(nano-gpt): rename LLM360 model id
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-12 04:19:56 +09:00
sk0x0y
d68d79888c
fix(nano-gpt): rename Infermatic model id
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-12 04:19:56 +09:00
sk0x0y
6c52905c6a
fix(nano-gpt): rename Gryphe model id
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-12 04:19:55 +09:00
sk0x0y
62410b8f26
fix(nano-gpt): rename GalrionSoftworks model id
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-12 04:19:55 +09:00
sk0x0y
50ce68ccab
fix(nano-gpt): rename Envoid model ids
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-12 04:19:20 +09:00
sk0x0y
d1c6a6b873
fix(nano-gpt): rename EVA-UNIT-01 model ids
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-12 04:19:20 +09:00
Frank
7193b068a5
update zen models
2026-03-11 13:52:50 -04:00
Sadik
79a8a06bd7
fix MiniMax-M2.5 context window
2026-03-11 20:50:33 +03:00
Aiden Cline
b60c03e11c
Merge pull request #1139 from zainhas/dev
...
[Together AI] add prompt caching pricing for MiniMax m2.5
2026-03-11 12:31:56 -05:00
Aiden Cline
15cf98d57b
Merge pull request #1146 from gotjoshua/patch-1
...
Rename step-3-5-flash.toml to step-3.5-flash.toml
2026-03-11 12:31:39 -05:00
Aiden Cline
b2ee6c407b
Merge pull request #1144 from micuintus/feature/update-nebius-changes
...
Feat: update Nebius changes
2026-03-11 12:31:29 -05:00
gotjoshua
96a14a06e7
Rename step-3-5-flash.toml to step-3.5-flash.toml
...
on nvidia it is 3.5 not 3-5
2026-03-11 11:41:36 +00:00
Michael Voigt
adc358606d
fix(nebius): update model context limits per API
2026-03-11 11:33:14 +01:00
Michael Voigt
63d52adf6f
feat(nebius): add GLM-5 model
2026-03-11 11:33:14 +01:00
sk0x0y
9a31387766
fix(nano-gpt): rename Salesforce model id
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-11 18:17:41 +09:00
sk0x0y
735157b837
fix(nano-gpt): rename ReadyArt model ids
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-11 18:17:41 +09:00
sk0x0y
d75b46fb37
fix(nano-gpt): rename Doctor-Shotgun model id
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-11 18:17:41 +09:00
sk0x0y
cc555f8482
fix(nano-gpt): rename CrucibleLab model id
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-11 18:17:13 +09:00
sk0x0y
7fbbcf2b49
fix(nano-gpt): rename MiniMaxAI model id
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-11 16:04:28 +09:00
sk0x0y
14c8ec8ca5
fix(nano-gpt): rename Tongyi-Zhiwen model id
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-11 16:04:28 +09:00
sk0x0y
72568bbdb3
fix(nano-gpt): rename Alibaba-NLP model id
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-11 16:03:57 +09:00
sk0x0y
c2225b715f
fix(nano-gpt): rename THUDM GLM-Z1 rumination id
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-11 15:49:20 +09:00
sk0x0y
7ce25e3742
fix(nano-gpt): rename THUDM GLM-Z1 model ids
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-11 15:49:20 +09:00
sk0x0y
427868604b
fix(nano-gpt): rename THUDM GLM-4 model ids
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-11 15:49:20 +09:00
Zain Hasan
247cd801a8
add prompt caching pricing for MiniMax m2.5
2026-03-10 22:54:42 -07:00
Aiden Cline
1aa2ee22b1
Merge pull request #1134 from sk0x0y/feature/nanogpt-catalog-fixes
...
fix(nano-gpt): correct TEE path ids and add missing canonical entries
2026-03-10 22:02:52 -05:00
Aiden Cline
0f57233eff
Merge pull request #1105 from sylviezhang37/add-vercel-input-context-and-new-models
...
feat(vercel): add input context calculation + new models
2026-03-10 22:01:52 -05:00
Aiden Cline
73a78eebfc
Merge pull request #1138 from mugnimaestra/feat/add-glm-5-turbo-chutes
...
feat: add GLM-5-Turbo to Chutes provider listings
2026-03-10 22:01:08 -05:00
Sylvie Zhang
3a6789b819
Merge branch 'dev' into add-vercel-input-context-and-new-models
2026-03-10 17:44:14 -07:00
Sylvie Zhang
f7c505e140
remove context from gemini models
2026-03-10 17:43:08 -07:00
Sylvie Zhang
6bb36806d6
only calc input context for openai models
2026-03-10 17:40:46 -07:00
Sylvie Zhang
20a404eb88
revert non openai changes
2026-03-10 17:38:46 -07:00
Muhammad Mugni Hadi
65ecb5cd4a
feat: add GLM-5-Turbo to Chutes provider listings
2026-03-11 05:26:11 +07:00
Matt Cowger
56062a9129
Fix incorrect pricing
2026-03-10 14:57:44 -07:00
Aiden Cline
d3d9c580d4
Merge pull request #1135 from gitpush-gitpaid/fix/gpt-5-4-pdf-input-modalities
...
Added PDF to input modalities for GPT-5.4
2026-03-10 13:53:42 -05:00
gitpush-gitpaid
ef98d8a9cb
Updated GPT-5.4 PDF input modalities
2026-03-10 13:59:29 -04:00
sk0x0y
9d17752b88
fix(nano-gpt): add missing GLM 5 thinking model
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-11 00:56:22 +09:00
sk0x0y
b5a838fe8b
fix(nano-gpt): add missing TEE qwen3.5 model
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-11 00:56:22 +09:00
sk0x0y
aa1ac39ee6
fix(nano-gpt): rename TEE gemma and minimax ids
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-11 00:56:22 +09:00
sk0x0y
4bc17ccf96
fix(nano-gpt): rename TEE oss and llama ids
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-11 00:56:22 +09:00
sk0x0y
08c1899bfe
fix(nano-gpt): rename TEE deepseek model ids
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-11 00:56:02 +09:00
sk0x0y
ad50e4a5ed
fix(nano-gpt): rename TEE qwen model ids
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-11 00:56:02 +09:00
sk0x0y
730915a123
fix(nano-gpt): rename TEE kimi model ids
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-11 00:56:02 +09:00
sk0x0y
6f12d18cb8
fix(nano-gpt): rename TEE glm model ids
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-11 00:56:02 +09:00
Aiden Cline
bd8774db99
Merge pull request #1132 from sk0x0y/feature/nanogpt-model-sync
...
feat(nano-gpt): add text and image models
2026-03-10 10:31:50 -05:00
Aiden Cline
88fbea52a4
Merge pull request #1133 from anomalyco/fix-model
...
fix: bedrock devstral
2026-03-10 10:31:08 -05:00
Aiden Cline
70e5d9b34b
fix: bedrock devstral
2026-03-10 10:30:20 -05:00
Aiden Cline
edb6ef0d71
Merge pull request #1129 from Grin1024/dev
...
feat: add GPT-5 series models to requesty provider
2026-03-10 10:29:30 -05:00
Aiden Cline
df1280ed8b
add families to some bedrock models
2026-03-10 10:12:13 -05:00
sk0x0y
898b3c18b7
feat(nano-gpt): add image models
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-10 22:13:48 +09:00
sk0x0y
6316e543ef
feat(nano-gpt): add text models
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-03-10 22:13:48 +09:00
Aiden Cline
64d9a97f9d
Merge pull request #1128 from JWahle/dev
...
chore: Updated abacus model definitions
2026-03-10 07:40:01 -05:00
Aiden Cline
c62a2a3fc1
Merge pull request #1130 from Mingholy/fix/alibaba-coding-plan-model-limits
...
fix: update model limits for alibaba-coding-plan providers
2026-03-10 07:39:48 -05:00
Aiden Cline
cc1937a177
Merge pull request #1131 from janszypulski/cloudferro-sherlock-fix-minimax-model-id
...
fix minimax-m2.5 model id - wrong file path
2026-03-10 07:39:34 -05:00
Jan Szypulski
9d3a88863d
fix minimax-m2.5 model id - wrong file path
2026-03-10 11:21:55 +01:00
mingholy.lmh
b9123e26e0
fix: update model limits for alibaba-coding-plan providers
...
- Add MiniMax-M2.5 to alibaba-coding-plan-cn
- Update qwen3-max output limit (65536 -> 32768)
- Update qwen3-coder-plus context limit (1048576 -> 1000000)
- Update MiniMax-M2.5 limits per ref.json (context: 196608, output: 24576)
Co-authored-by: Qwen-Coder <qwen-coder@alibabacloud.com >
2026-03-10 15:45:01 +08:00
lihui
933e450104
feat: add GPT-5 series models to requesty provider
...
Add missing OpenAI GPT-5 series models to requesty provider:
- GPT-5 Chat, Codex, Image, Pro
- GPT-5.1 Chat, Codex, Codex-Max, Codex-Mini
- GPT-5.2 Chat, Codex, Pro
- GPT-5.3 Codex
- GPT-5.4, GPT-5.4 Pro
2026-03-10 14:53:28 +08:00
JWahle
2ba6383e70
chore: Updated abacus model definitions
...
Added: gpt-5.4.toml
Removed: gemini-3-pro-preview.toml
2026-03-10 05:10:45 +01:00
Aiden Cline
65ed6ac5dd
Merge pull request #1126 from mcowger/feature/gemini-3.1-flash-lite-vercel
...
feat: add gemini-3.1-flash-lite-preview to vercel gateway provider
2026-03-09 21:43:33 -05:00
Aiden Cline
e7d04aec7a
Merge pull request #1127 from anomalyco/add-shape
...
feat: add 'shape' field to provider so models can specify if they use responses vs completions apis (use only if model only supports 1 of)
2026-03-09 21:43:04 -05:00
Aiden Cline
37fe334aed
feat: add 'shape' field to provider so models can specify if they use responses vs completions apis (use only if model only supports 1 of)
2026-03-09 21:42:23 -05:00
Matt Cowger
ce8fc9e4f0
feat: add gemini-3.1-flash-lite-preview to vercel gateway provider
2026-03-09 19:32:43 -07:00
Aiden Cline
be8eb8ba54
fix name
2026-03-09 20:01:02 -05:00
Aiden Cline
7c625b3b82
Merge pull request #945 from Daltonganger/feat/nano-gpt-sync-models-api
...
sync nano-gpt models with live API catalog
2026-03-09 20:00:06 -05:00
Aiden Cline
33700d27dc
Merge pull request #1032 from propilideno/feature/new_gpt_5.3_codex_and_missing_structured_output_attr
...
Add gpt-5.3-codex (Azure) and fill missing structured output flags
2026-03-09 19:40:20 -05:00
Aiden Cline
e5c300a5e5
fix
2026-03-09 19:38:30 -05:00
Aiden Cline
e5e9175c5d
Merge branch 'dev' into feature/new_gpt_5.3_codex_and_missing_structured_output_attr
2026-03-09 19:37:40 -05:00
Aiden Cline
a9f79d6794
Merge pull request #1123 from dpuyosa/feature/venice-gpt54-multimodal
...
Venice: Add GPT-5.4 Pro and enable multimodal inputs for GPT-5.4 & Qwen3.5
2026-03-09 18:19:52 -05:00
Aiden Cline
fd4c4a8f28
Merge pull request #1038 from muldercw/add-clarifai-model-provider
...
Add Clarifai Model Provider
2026-03-09 18:19:14 -05:00
dpuyosa
f9b5385868
[venice] Add GPT-5.4 Pro and enable multimodal inputs
...
- Add GPT-5.4 Pro model
- Enable attachment/image input for GPT-5.4
- Enable attachment/image/video input for Qwen3.5 35B A3B
2026-03-09 22:52:54 +01:00
Aiden Cline
b2f7a72410
Merge pull request #1110 from fhennerkes/dev
...
poe: add GPT-5.4, GPT-5.4-Pro, and GPT-5.3-Instant models
2026-03-09 14:10:12 -05:00
Aiden Cline
7b5d9aa645
Merge pull request #1025 from liuchang-reolink/dev
...
add qwen3.5-397b-a17b and step-3-5-flash for nvidia
2026-03-09 14:05:08 -05:00
Aiden Cline
6e0040dbfd
Merge pull request #1089 from Krule/krule/update_gitlab_anthropic_context_size
...
feat(gitlab): update context limit to 1M for Claude Sonnet and Opus 4.6
2026-03-09 14:03:49 -05:00
Aiden Cline
943ad8481b
Merge pull request #1121 from illusion77/fix/chutes-mimo-v2-flash-context-16709
...
fix(chutes): correct MiMo-V2-Flash context window and capabilities
2026-03-09 14:02:51 -05:00
Aiden Cline
7f1b6fb0eb
Merge pull request #1122 from riccardogiorato/dev
...
remove deprecated kimi models from together.ai
2026-03-09 14:02:36 -05:00
Riccardo Giorato
23eff95e5d
remove deprecated kimi from together.ai
2026-03-09 17:30:40 +01:00
illusion77
ddbd396205
fix(chutes): correct MiMo-V2-Flash context window and capabilities
...
The chutes provider had incorrect metadata for MiMo-V2-Flash:
context 32K → 262K, output 8K → 32K, reasoning and tool_call enabled.
Fixes anomalyco/opencode#16709
2026-03-09 10:57:57 -05:00
Aiden Cline
f3ee1a530b
Merge pull request #1120 from stephenkuhn214/dev
...
Add Amazon-Bedrock Devstral 2 123B model
2026-03-09 09:35:46 -05:00
Aiden Cline
9c51b65440
Merge pull request #1119 from cgilly2fast/dev
...
fix(firmware): proper 5.3 codex model id
2026-03-09 09:30:35 -05:00
Frank
353aeb4998
update zen models
2026-03-09 10:08:55 -04:00
Frank
11991fecb5
update zen models
2026-03-09 10:03:13 -04:00
stephenkuhn214
1b4599773d
Create mistral.devstral-2-123b
2026-03-09 08:58:19 -04:00
Colby Gilbert
78e1a3b0c9
fix(firmware): proper 5.3 codex model id
2026-03-08 21:58:27 -07:00
Sewer56
7a02946620
Update synthetic models: promote MiniMax-M2.5, add GLM-4.7-Flash, remove deprecated Qwen3.5
2026-03-08 22:56:31 +00:00
Aiden Cline
44686797c8
Merge pull request #1118 from shelvick/add-azure-gpt-5.3-chat
...
Add GPT-5.3 Chat to Azure
2026-03-08 16:41:10 -05:00
Aiden Cline
065cec8431
fix: input limit for context
2026-03-08 16:40:38 -05:00
Scott Helvick
f491c2bec9
Add GPT-5.3 Chat to Azure
2026-03-08 21:20:27 +00:00
Aiden Cline
cf1ac3053f
Merge pull request #1081 from djmaze/fix/nebius-model-casing
...
fix(nebius): correct model ID casing to match Token Factory API
2026-03-08 14:26:52 -05:00
Aiden Cline
6be1e929fc
Merge pull request #1114 from v1gnesh/dev
...
add grok 4.2 experimentals
2026-03-08 14:25:12 -05:00
Aiden Cline
49524827e2
Merge pull request #1113 from shelvick/add-vertex-glm-5
...
Fix GLM-5 context window size on Google Vertex
2026-03-08 10:31:05 -05:00
Aiden Cline
5ab5d389fc
Merge pull request #1112 from cau1k/feat/az-5.4
...
feat(azure): add gpt-5.4/5.4-pro
2026-03-08 10:30:54 -05:00
Aiden Cline
d5367ed978
Merge pull request #1116 from xiaojiezj/xj_dev_0308
...
fix: Adjust the logo for ZenMux
2026-03-08 10:30:18 -05:00
Aiden Cline
5069faa25b
Merge pull request #1117 from kailiu42/feat/siliconflow-cn
...
feat(siliconflow-cn): add Qwen3.5 model family
2026-03-08 10:29:48 -05:00
Kai Liu
b279f33d9b
feat(siliconflow-cn): add Qwen3.5 model family
...
New models:
- Qwen/Qwen3.5-4B
- Qwen/Qwen3.5-9B
- Qwen/Qwen3.5-27B
- Qwen/Qwen3.5-35B-A3B
- Qwen/Qwen3.5-122B-A10B
- Qwen/Qwen3.5-397B-A17B
Signed-off-by: Kai Liu <kraml.liu@gmail.com >
2026-03-08 20:07:40 +08:00
xiaojie.zj
fe8249d706
fix: Adjust the logo
2026-03-08 16:36:01 +08:00
skywalker512
236af40da3
feat: add Tencent Coding Plan provider
...
Add support for Tencent Coding Plan with 8 models:
- Auto (tc-code-latest)
- Hunyuan 2.0 Instruct
- Hunyuan 2.0 Think
- Hunyuan-T1
- Hunyuan-TurboS
- MiniMax-M2.5
- Kimi-K2.5
- GLM-5
Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com >
2026-03-08 15:34:36 +08:00
v1gnesh
a24a23d57b
add grok 4.2 experimentals
2026-03-08 07:42:09 +05:30
Scott Helvick
7e4773d9b5
Fix GLM-5 context window size on Google Vertex
...
Correct the context limit from 204800 to 202752 tokens.
2026-03-07 22:20:56 +00:00
zero
9f937f3fc5
Merge branch 'dev' into feat/az-5.4
2026-03-07 17:20:29 -05:00
cau1k
8cbdbc1102
add day cutoff
2026-03-07 17:19:15 -05:00
cau1k
f1ca3b0015
feat(azure-cognitive-services): symlink 5.4/pro from azure provider
...
;
2026-03-07 17:02:26 -05:00
cau1k
35c757bad5
feat(azure): add 5.4/pro
2026-03-07 17:00:43 -05:00
fhennerkes
78781f3901
poe: add GPT-5.4, GPT-5.4-Pro, and GPT-5.3-Instant models
2026-03-07 13:58:35 -08:00
Sylvie Zhang
26465319d6
Merge branch 'dev' into add-vercel-input-context-and-new-models
2026-03-07 11:39:53 -08:00
fhennerkes
1371cbf9de
poe: add GPT-5.4, GPT-5.4-Pro, and GPT-5.3-Instant models
2026-03-07 09:48:17 -08:00
Aiden Cline
559ccd6966
Merge pull request #1024 from yinxulai/feat/qiniu-ai
...
feat(qiniu-ai): add new model configurations
2026-03-07 11:26:01 -06:00
Aiden Cline
2691cb4e8d
Merge pull request #1083 from samzong/feat/add-drun-provider
...
feat: add d.run(China) provider (OpenAI-compatible)
2026-03-07 11:25:12 -06:00
Aiden Cline
83ed1f0125
Merge pull request #1015 from RioPlay/dev
...
add: newer MiniMax, GLM, and Kimi models to DeepInfra
2026-03-07 11:24:53 -06:00
Aiden Cline
869f831466
Merge branch 'dev' into dev
2026-03-07 11:23:37 -06:00
Aiden Cline
5a673af2ae
Merge pull request #1061 from JonasGao/dev
...
Add Qwen3.5 Flash & GLM-5 & M2.5 models to alibaba-cn
2026-03-07 11:22:56 -06:00
Aiden Cline
9b8543a074
Add interleaved section to minimax-m2.5.toml
2026-03-07 11:21:54 -06:00
Aiden Cline
b3fb902331
Merge pull request #1030 from Mingholy/feat/alibaba-coding-plan-cn
...
feat(alibaba-coding-plan-cn): add Coding Plan provider for China region
2026-03-07 11:21:26 -06:00
Aiden Cline
adb0c0b305
Merge pull request #1062 from viitana/bump-deepseek-details
...
feat: [deepseek]: update official DeepSeek model details
2026-03-07 11:21:21 -06:00
Aiden Cline
ed01410d82
Merge pull request #1088 from mcowger/feature/gemini-3.1-flash-lite
...
feat: add gemini-3.1-flash-lite-preview model
2026-03-07 11:13:45 -06:00
Aiden Cline
7e23b780cc
Merge pull request #1077 from evroc-oss/evroc/correct-model-config
...
fix(evroc): correct model config
2026-03-07 11:13:15 -06:00
Aiden Cline
c46b652c8e
Merge pull request #1076 from jerome-benoit/feat/add-sonar-deep-research-sap-ai-core
...
feat(sap-ai-core): add Perplexity Sonar Deep Research model
2026-03-07 11:11:29 -06:00
Aiden Cline
ddb74e9b09
Merge pull request #1063 from dpuyosa/fix/models-pricing-limits-update
...
Venice: Update model pricing and limits
2026-03-07 11:10:37 -06:00
Aiden Cline
face36ecb8
Merge pull request #1064 from BlockListed/fix-cortecs-models
...
Fix Cortecs models
2026-03-07 11:10:03 -06:00
Aiden Cline
6f170651b3
Merge pull request #1075 from Track07-cda/alibaba-cn-third-party-models
...
Add third party providers' models to alibaba-cn provider
2026-03-07 11:09:40 -06:00
Aiden Cline
47dbe45dd5
Merge pull request #1066 from dpuyosa/feat/add-qwen3-5-35b-a3b
...
Venice: Add Qwen 3.5 35B A3B model
2026-03-07 11:09:10 -06:00
Aiden Cline
0ee43b64b3
Merge branch 'dev' into alibaba-cn-third-party-models
2026-03-07 11:08:37 -06:00
Aiden Cline
6130a1f74e
Merge pull request #1068 from MauroDruwel/dev
...
NVIDIA: Add MiniMax M2.5 model and remove MiniMax M2
2026-03-07 11:07:17 -06:00
Aiden Cline
c0c82a5f04
Merge pull request #1072 from sylviezhang37/update-vercel-models-20260302-1656
...
Update Vercel models
2026-03-07 11:06:17 -06:00
Aiden Cline
b0ba8b14d5
Merge pull request #1092 from janszypulski/cloudferro-sherlock-add-minimax-2.5
...
add MiniMaxAI/MiniMax-M2.5 to CloudFerro Sherlock
2026-03-07 11:02:11 -06:00
Aiden Cline
4780f9ddc1
Merge pull request #1109 from dinhkim/feat/add-cf-glm-4.7-flash
...
feat: add GLM-4.7-Flash to the Cloudflare Workers AI provider
2026-03-07 11:01:50 -06:00
Aiden Cline
27e02de632
Merge pull request #1078 from SomeoneWithOptions/dev
...
add gpt 5.3 codex for openrouter and Mercury models
2026-03-07 11:01:41 -06:00
Aiden Cline
f22c827045
Merge branch 'dev' into dev
2026-03-07 11:01:17 -06:00
Aiden Cline
cfc4585ed7
Merge pull request #1107 from Rinuuri/deepinfra-glm5
...
Add deepinfra GLM-5 model
2026-03-07 10:59:16 -06:00
Aiden Cline
fa07bc2088
Merge pull request #1039 from rholak/add-abacus-models
...
Add sonnet 4.6 and opus 4.6 to abacus model list
2026-03-07 10:59:00 -06:00
Aiden Cline
497b1daaf2
Merge pull request #1103 from dpuyosa/feat/venice-add-gpt-models
...
Venice: Add OpenAI GPT-4o, GPT-4o Mini, GPT-5.4 models
2026-03-07 10:58:44 -06:00
Aiden Cline
442afa8c7e
Merge pull request #1060 from yanismiraoui/inception/mercury2
...
Add Inception Mercury 2 and Mercury Edit models
2026-03-07 10:57:38 -06:00
Aiden Cline
4bd0c387fe
Merge pull request #1044 from shrwnsan/feat/openrouter-routers
...
feat(openrouter/free): add free router
2026-03-07 10:57:24 -06:00
Aiden Cline
b8c0c1d3a1
Merge pull request #1053 from laiiihz/update-xiaomi-models
...
Update Xiaomi models metadata
2026-03-07 10:57:17 -06:00
Aiden Cline
5c6c3e5a32
Merge pull request #1055 from shantanugoel/gemini-3.1-flash-image-preview
...
Add Gemini 3.1 Flash Image Preview
2026-03-07 10:57:06 -06:00
Aiden Cline
53d3cca3a0
Merge pull request #1052 from spiffytech/dev
...
Improve Ollama Cloud generator. Remove Gemini 3 Pro from Ollama Cloud.
2026-03-07 10:56:45 -06:00
Aiden Cline
105970c173
Merge pull request #1049 from heimoshuiyu/fix/glm-5-open-weights
...
fix: mark GLM-5 as open weights
2026-03-07 10:56:31 -06:00
Aiden Cline
788ee04034
Merge pull request #1045 from xinrui-z/aihubmix-add-models
...
aihubmix add models
2026-03-07 10:56:03 -06:00
Aiden Cline
6626db4044
Merge pull request #1098 from JWahle/dev
...
chore: updated abacus model definitions
2026-03-07 10:55:31 -06:00
Aiden Cline
8902640664
Merge pull request #1023 from PandaSt0rm/add-alibaba-coding-plan
...
Add Alibaba Coding Plan provider and model configs
2026-03-07 10:53:34 -06:00
Kim Truong
cab247ddf8
update context to match Cloudflare doc
2026-03-07 23:50:06 +07:00
Kim Truong
c1a42fa0a0
feat: add GLM-4.7-Flash mode in Cloudflare Workers AI provider
2026-03-07 23:45:52 +07:00
Aiden Cline
35023bba5a
Merge pull request #1001 from ItsWendell/feat/bedrock-bearer-token
...
Add AWS_BEARER_TOKEN_BEDROCK to Amazon Bedrock provider env
2026-03-07 09:52:21 -06:00
Aiden Cline
604e49792b
Merge pull request #1002 from DEAN-Cherry/feat/add-minimax-m2.5
...
models: alibaba-cn: add MiniMax-M2.5
2026-03-07 09:51:36 -06:00
Aiden Cline
0ca77b0cda
Merge branch 'dev' into dev
2026-03-07 09:50:26 -06:00
Aiden Cline
ea9505a40f
Merge pull request #1004 from BlockListed/cortecs-models
...
Add Cortecs AI models
2026-03-07 09:50:07 -06:00
Aiden Cline
990b8d7308
Merge pull request #1005 from cgilly2fast/dev
...
feat(firmware): gemini 3.1 pro, sonnet reasoning
2026-03-07 09:49:54 -06:00
Aiden Cline
ec173e86d4
Merge pull request #996 from fhennerkes/dev
...
poe: add Gemini-3.1-Pro, GPT-5.3-Codex and Gemini 3.1 Flash Lite
2026-03-07 09:47:52 -06:00
Aiden Cline
4a6e92a7c9
Merge pull request #997 from xiaojiezj/zenmux_dev_0221
...
feat: add Gemini 3.1 Pro Preview for ZenMux provider
2026-03-07 09:47:37 -06:00
Aiden Cline
f0f686bdf5
Merge pull request #999 from mikalsande/mistral_latest
...
Append (latest) to Mistral models that refer to the latest version.
2026-03-07 09:46:40 -06:00
Aiden Cline
35ff0c2629
Merge pull request #995 from Phoen1xCode/dev
...
fix(zenmux:minimax): remove duplicated prefix & feat(zenmux:openai): add GPT-5.2-Pro model
2026-03-07 09:45:07 -06:00
Aiden Cline
f99e9e89df
Merge pull request #1090 from litvix-whale/feat/add-minimax-m2-5
...
feat(provider): add MiniMax M2.5 for DeepInfra
2026-03-07 09:41:28 -06:00
Armin Pašalić
09722ac264
Merge branch 'anomalyco:dev' into krule/update_gitlab_anthropic_context_size
2026-03-07 13:17:10 +01:00
Rinuuri
fa67d00aeb
Update GLM-5.toml
2026-03-06 21:23:42 +00:00
Rinuuri
ddd2dd73ed
Adding deepinfra GLM-5
2026-03-07 00:03:29 +03:00
fhennerkes
d7929fd00b
Merge branch 'anomalyco:dev' into dev
2026-03-06 12:00:20 -08:00
Frank
06e7d4db42
Merge pull request #1014 from NachoFLizaur/fix/bedrock-opus-4-6-context-window
...
fix(amazon-bedrock): correct Claude Opus 4.6 context window from 1M to 200K
2026-03-06 11:25:37 -05:00
Sylvie Zhang
7a11ef241d
update more models
2026-03-06 08:24:49 -08:00
Sylvie Zhang
145862315d
add input calculation + new models
2026-03-06 08:07:11 -08:00
dpuyosa
d871710ba4
[venice] Add OpenAI GPT-4o, GPT-4o Mini, GPT-5.4 models
...
- Add gpt-4o-2024-11-20 model configuration
- Add gpt-4o-mini-2024-07-18 model configuration
- Add gpt-5.4 model configuration with reasoning capability
2026-03-06 09:53:06 +01:00
Colby Gilbert
16486087c6
Merge branch 'anomalyco:dev' into dev
2026-03-05 21:38:25 -08:00
Frank
2939af9330
Merge pull request #1100 from sachnun/feat/github-copilot-gpt-5-4
...
feat(provider): add gpt-5.4 for GitHub Copilot
2026-03-05 23:33:57 -05:00
sachnun
7c68dab3bb
feat(provider): add gpt-5.4 for GitHub Copilot
2026-03-06 11:18:11 +07:00
Mike Soylu
caceb0b310
openrouter openai models ( #1099 )
2026-03-05 22:26:58 -05:00
Frank
7a0d3be1e7
Update zen models
2026-03-05 18:55:49 -05:00
ShivamB25
e11ad7c01a
feat(openai): add GPT-5.4 and GPT-5.4 Pro model specs ( #1095 )
2026-03-05 18:50:22 -05:00
Matt Silverlock
d30fa82e4c
Cloudflare: add gpt-5.4.toml ( #1096 )
2026-03-05 18:50:10 -05:00
Rishi Vhavle
771102a960
feat: add gpt-5.3-codex to github-copilot provider ( #1097 )
2026-03-05 18:49:56 -05:00
JWahle
30f98b15ef
chore: updated abacus model definitions
...
Added: GPT-5 Codex, GPT-5.1/5.2/5.3 Codex, GPT-5.3 Chat, Gemini 3.1 Flash Lite/Pro Preview, Claude Opus/Sonnet 4.6, Kimi K2.5, GLM-5
Removed: Gemini 2.0 Flash 001, Gemini 2.0 Pro Exp, Meta-Llama 3.1 70B Instruct
Updated pricing: DeepSeek V3.1, GLM-4.7, GPT-5.2 Chat Latest, o3-pro, Route LLM
2026-03-06 00:46:19 +01:00
Colby Gilbert
6f7ab479fb
feat(firmware): gpt 5.4
2026-03-05 13:23:08 -08:00
Colby Gilbert
a4efbcd5ce
Merge branch 'anomalyco:dev' into dev
2026-03-05 13:15:49 -08:00
Frank
bcbfba03bd
update zen models
2026-03-05 15:51:33 -05:00
Frank
bdb5dac941
update zen models
2026-03-05 15:50:03 -05:00
Frank
1538bdcedb
update zen models
2026-03-05 13:31:27 -05:00
SomeoneWithOptions
e5211f3105
add inception mercury models for openrouter
2026-03-05 12:13:52 -05:00
Andres Castellanos
4a2209dbd4
Merge branch 'anomalyco:dev' into dev
2026-03-05 11:51:38 -05:00
Jan Szypulski
900014fe52
add MiniMax-M2.5
2026-03-05 14:58:19 +01:00
Kyrylo Lytvishko
5bbaf3c3f2
feat(provider): add MiniMax M2.5 for DeepInfra
2026-03-05 14:04:11 +02:00
Armin Pasalic
7dd0a26ff4
feat(gitlab): update context limit to 1M for Sonnet and Opus 4.6
2026-03-05 12:01:02 +01:00
Matt Cowger
3b7e0f02f1
feat: add gemini-3.1-flash-lite-preview model
2026-03-04 13:19:21 -08:00
samzong
e67f921ea3
feat: add official d.run logo
2026-03-04 13:42:01 +08:00
samzong
f5411eeeda
feat: add D.Run (China) provider with minimax-m25, deepseek-r1, deepseek-v3
2026-03-04 13:33:04 +08:00
Frank
0d83ab8909
Merge pull request #1082 from kesku/kesku/add-ppl-agent-api
...
Add Perplexity Agent API provider
2026-03-03 23:03:49 -05:00
Kesku
26a629debc
add models
2026-03-03 23:19:01 +00:00
Kesku
4b3319561b
set up provider
2026-03-03 23:10:43 +00:00
Ubuntu
b89ce0d986
fix(nebius): correct model ID casing to match Token Factory API
...
Fix lowercase model ID bug that caused "The model does not exist" errors.
- qwen/ → Qwen/ directory
- Fixed model file casing to match API exactly across all providers
2026-03-03 22:11:57 +00:00
fhennerkes
1c01f8172b
poe: add Gemini-3.1-Flash-Lite and update gpt-4o-mini context
...
Add new Gemini 3.1 Flash Lite model
Update gpt-4o-mini context window: 128K → 124,096
2026-03-03 11:24:20 -08:00
SomeoneWithOptions
d76040c514
add gpt 5.3 codex for openrouter
2026-03-03 12:53:39 -05:00
Simon Rygård
feffa8119f
fix(evroc): correct modality config
2026-03-03 16:52:07 +01:00
Simon Rygård
59c6e5df62
fix(evroc): correct tool call config
2026-03-03 16:51:47 +01:00
Jérôme Benoit
fe2204d42c
feat(sap-ai-core): add Perplexity Sonar Deep Research model
2026-03-03 14:56:18 +01:00
Track07-cda
07cc5335ac
Add third party providers' models to alibaba-cn provider
...
- Add `MiniMax/MiniMax-M2.5` and `kimi/kimi-k2.5` to the `alibaba-cn`
provider.
- Update `kimi-k2.5` to include video modality and adjust release/update
dates.
- Add several `siliconflow/deepseek` models to the `alibaba-cn`
provider.
2026-03-03 16:46:32 +08:00
github-actions[bot]
fefbb90a29
chore(vercel): update Vercel model definitions
...
Auto-generated by weekly workflow from Vercel AI Gateway API.
Co-Authored-By: github-actions[bot] <41898282+github-actions[bot]@users.noreply.github.com>
2026-03-02 16:56:48 +00:00
Frank
fec48b83d3
update zen models
2026-03-01 13:23:08 -05:00
Mauro Druwel
c30bbe7718
Add knowledge
2026-03-01 08:53:22 +01:00
Mauro Druwel
1fba668f0f
Add minimax-m2.5 to nvidia-nim and remove deprecated minimax-m2 from nvidia-nim
2026-03-01 08:52:33 +01:00
Aiden Cline
33ec088bda
Merge pull request #1008 from friendliai/feat/friendli-minimax-m2.5
...
add friendli minimax m2.5 model config
2026-03-01 07:54:08 +05:00
Aiden Cline
add7f9a914
Merge pull request #1065 from friendliai/minpeter/remove-exaone-models
...
Remove all EXAONE models
2026-03-01 07:53:42 +05:00
dpuyosa
369fa2de6d
[venice] Add Qwen 3.5 35B A3B model
...
- Add new model configuration for Qwen 3.5 35B A3B
- Includes cost, limits, and capabilities (reasoning, tool_call, structured_output)
2026-02-28 21:07:28 +01:00
minpeter
c8732e7e74
Remove all EXAONE models
...
Remove LGAI-EXAONE model definitions (EXAONE-4.0.1-32B, K-EXAONE-236B-A23B)
and related family references from core packages.
Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com >
2026-03-01 04:48:08 +09:00
Jonas
f00f9f3c11
Add Qwen3.5 Flash & GLM-5 & M2.5 models to alibaba-cn provider
2026-02-28 23:34:06 +08:00
BlockListed
c87ca238de
fix cortecs models
...
the should have periods not a p as a decimal separator
2026-02-28 15:16:59 +01:00
BlockListed
102f55aeef
add glm 4.7 flash model to cortecs
2026-02-28 15:13:53 +01:00
BlockListed
2767754a02
add kimi K2.5 model to cortecs
2026-02-28 15:13:53 +01:00
dpuyosa
4289b59a04
[models] Update model pricing and limits
...
- Update Claude Sonnet 4-6 pricing and output limit
- Update Grok 41 Fast pricing and context/limits
2026-02-28 14:24:16 +01:00
Atte Viitanen
b986313f42
feat: [deepseek]: update official deepseek model details
2026-02-28 13:22:26 +02:00
yanismiraoui
28cfd4cab6
naming mercury 2 and mercury edit for inception provider
2026-02-27 17:45:01 -08:00
yanismiraoui
b93e62fc3a
Add Inception Mercury 2 and Mercury Edit models
2026-02-27 17:41:00 -08:00
Aiden Cline
e23b5ab010
Merge pull request #1026 from ryot/venice
...
Venice: Add GPT-5.3 Codex
2026-02-28 06:30:13 +05:00
Aiden Cline
b5f6024868
Merge pull request #1020 from jerome-benoit/feat/sap-ai-core-add-models
...
feat(sap-ai-core): Add GPT-4.1, Gemini 2.5 Flash Lite, Perplexity Sonar, and Claude 4.6 models
2026-02-28 06:29:28 +05:00
Aiden Cline
07db15e984
Merge pull request #1029 from dpuyosa/veniceScript
...
Venice: Remove interactive API key prompt & use new maxCompletionTokens field
2026-02-28 06:28:33 +05:00
Aiden Cline
74abf8851a
Merge pull request #1042 from SomeoneWithOptions/dev
...
add gemini 3.1 pro preview custom tools for openrouter
2026-02-28 06:27:56 +05:00
Aiden Cline
7e13ecdfd9
Merge pull request #1056 from xezpeleta/fix/azure-gpt-5-3-codex
...
fix(azure): add gpt-5.3-codex model
2026-02-28 06:27:34 +05:00
Aiden Cline
6ad2c28b2d
Merge pull request #1048 from dpuyosa/feat/add-venice-models
...
Venice: Add NVIDIA Nemotron 3 Nano and Qwen 3 Coder Turbo models
2026-02-28 06:27:20 +05:00
Frank
a124036692
update zen models
2026-02-27 16:16:37 -05:00
Xabi Ezpeleta
d37d362cc8
fix(azure): add gpt-5.3-codex model
2026-02-27 16:41:11 +01:00
Shantanu Goel
c387f94c8e
Add Gemini 3.1 Flash Image Preview
2026-02-27 20:03:41 +05:30
laiiihz
45457c34d8
update xiaomi models detail
2026-02-27 14:56:16 +08:00
spiffytech
44774ec3d6
Ollama Cloud removed support for Gemini 3 Pro
2026-02-26 17:18:34 -05:00
spiffytech
c8fdcf80dd
Updated Ollama Cloud generator to delete old models, only write out files if they changed
2026-02-26 17:18:33 -05:00
fhennerkes
9d33b6409c
Merge branch 'anomalyco:dev' into dev
2026-02-26 12:04:52 -08:00
Matt Silverlock
c76586a174
Cloudflare: add codex models to AI Gateway ( #1050 )
...
* add gpt-5.2-codex
* add gpt-5.3-codex
* Update gpt-5.2-codex.toml
* Update gpt-5.3-codex.toml
2026-02-26 14:41:12 -05:00
Jérôme Benoit
2a267614aa
feat(sap-ai-core): add Claude Opus 4.6 and Sonnet 4.6 models
2026-02-26 17:58:02 +01:00
PandaSt0rm
aac62378b2
Update MiniMax-M2.5 guidance per Alibaba docs
2026-02-26 17:20:58 +02:00
David Hill
56cc5f71bf
fix(ui): opencode zen logo update
2026-02-26 11:09:25 +00:00
David Hill
df2c87d32a
fix(ui): opencode go logo
2026-02-26 11:09:13 +00:00
heimoshuiyu
ff41c2b6c3
fix: mark GLM-5 as open weights
...
GLM-5 is an open-source model, but several provider config files
incorrectly had open_weights set to false. This commit corrects
all GLM-5 configurations to properly reflect its open-source status.
Affected providers:
- zhipuai
- zhipuai-coding-plan
- zai
- zai-coding-plan
- zenmux
- vercel
- siliconflow
- siliconflow-cn
- meganova
2026-02-26 18:43:28 +08:00
dpuyosa
1d137e2f1f
[venice] Add NVIDIA Nemotron 3 Nano and Qwen 3 Coder models
...
- Add NVIDIA Nemotron 3 Nano 30B A3B model configuration
- Add Qwen 3 Coder 480B A35B Instruct Turbo model configuration
2026-02-26 10:53:00 +01:00
dpuyosa
16720bcd1a
[venice] Use maxCompletionTokens for output limit
...
- Add optional maxCompletionTokens field to model spec schema
- Use maxCompletionTokens when calculating output token limit instead of checking existing limit
2026-02-26 10:24:43 +01:00
Xinrui
1feaf76749
aihubmix add models
2026-02-26 16:12:51 +08:00
shrwnsan
080ef5cc9e
fix(openrouter): remove auto router and add missing limit.input
...
- Remove auto router (cost varies, doesn't fit schema)
- Add limit.input = 200_000 to free.toml (schema requirement)
OpenRouter's auto router has 'pricing varied' - it charges based on the
routed model. This doesn't fit the numeric cost schema required by
models.dev, so we're removing it. The free router is retained as it
genuinely costs $0.
2026-02-26 14:40:38 +08:00
Ryo Tulman
f8121c8dc3
Update Venice GPT 5.3 Codex output limit
2026-02-26 00:32:13 -06:00
shrwnsan
d2d5c5a7cc
feat: add openrouter free and auto routers
2026-02-26 10:51:55 +08:00
SomeoneWithOptions
09d9e91d83
add gemini 3.1 pro preview custom tools for openrouter
2026-02-25 15:13:43 -05:00
Robert Holak
930d6a94b8
Add sonnet 4.6 and opus 4.6 to abacus model list
2026-02-25 12:31:06 -06:00
mulder
b9217aff8e
Add Clarifai Model Provider
...
Add Clarifai as a new provider with 11 models:
- GPT OSS 20B, GPT OSS 120B High Throughput
- Ministral 3 14B/3B Reasoning 2512
- Qwen3 Coder 30B, Qwen3 30B Instruct/Thinking 2507
- MiniMax-M2.5 High Throughput
- Trinity Mini, DeepSeek OCR, MM Poly 8B
Also adds 'mm-poly' family to family.ts for the Clarifai multimodal model.
2026-02-25 12:28:40 -05:00
Lucas Almeida
c240bce614
fix: adding missing structured_output parameter
2026-02-25 11:19:54 -03:00
Lucas Almeida
843a1d182a
feat: adding gpt-5.3-codex for Azure Foundry
2026-02-25 11:09:10 -03:00
PandaSt0rm
443c06ca03
fix MiniMax M2.5 modalities in Alibaba Coding Plan
...
- set MiniMax-M2.5 input modalities to text-only
- keep output modality as text
- validate with bun validate
2026-02-25 13:10:58 +02:00
PandaSt0rm
84466021ca
add MiniMax M2.5 to Alibaba Coding Plan and align third-party limits
...
- add MiniMax-M2.5 model config under providers/alibaba-coding-plan/models
- update GLM-4.7 limits to 202,752 context / 16,384 output
- update GLM-5 limits to 202,752 context / 16,384 output
- update Kimi K2.5 output limit to 32,768
- validate with bun validate
2026-02-25 13:05:18 +02:00
mingholy.lmh
b995e90cf5
fix: update context and output limits for alibaba-coding-plan-cn models
...
Update model limits:
- qwen3-coder-plus: context 1_048_576 → 1_000_000
- glm-5: output 131_072 → 16_384
- glm-4.7: output 131_072 → 16_384
- kimi-k2.5: output 65_536 → 32_768
Co-authored-by: Qwen-Coder <qwen-coder@alibabacloud.com >
2026-02-25 17:44:49 +08:00
dpuyosa
291e2eefe9
[venice] Remove interactive API key prompt
...
- Remove readline import and promptForApiKey function
- Remove prompt fallback, rely on CLI arg or env var only
- Update README to reflect change
2026-02-25 09:51:50 +01:00
Sewer56
0428299773
Added: Qwen3.5-397B natively supports image, MM2.5 No Image as it was a mistake.
2026-02-25 08:10:42 +00:00
Frank
96e9537b34
update zen models
2026-02-25 01:05:35 -05:00
Ryo Tulman
09dc7060ac
Venice: Add GPT-5.3 Codex
2026-02-24 23:37:03 -06:00
Colby Gilbert
a463717783
chore(firmware): remove gpt-5
2026-02-24 20:33:49 -08:00
Colby Gilbert
6d721dd32d
feat(firmware): gpt-5.3-codex
2026-02-24 20:32:17 -08:00
liuchang-reolink
3ae513785a
add step-3-5-flash for nvidia
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-02-25 12:05:19 +08:00
Colby Gilbert
e8d667b628
Merge branch 'anomalyco:dev' into dev
2026-02-24 20:01:26 -08:00
liuchang-reolink
7616a65e63
add qwen3.5-397b-a17b for nvidia
...
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-opencode )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-02-25 11:40:15 +08:00
yinxulai
a5b3da12c1
chore(qiniu-ai): update provider config
2026-02-25 10:20:55 +08:00
yinxulai
a05fbc3604
feat(qiniu-ai): add new model configurations
2026-02-25 10:16:40 +08:00
Aiden Cline
2189030e57
Merge pull request #1022 from armishra/feat/add-minimax-m2.5-baseten
...
feat(provider): Add MiniMax-M2.5 for baseten
2026-02-24 17:28:04 -06:00
Aiden Cline
c7ecc08442
Merge pull request #1019 from dpuyosa/venice
...
Venice: Update gemini-3-1-pro-preview config
2026-02-24 17:27:46 -06:00
Aiden Cline
830046e45e
Merge pull request #1021 from sylviezhang37/update-vercel-models-20260224-2134
...
Update Vercel models
2026-02-24 17:27:19 -06:00
Aiden Cline
c7b26477b9
Update cache_read value in gemini-3.1-pro-preview.toml
2026-02-25 04:26:56 +05:00
Aiden Cline
9e60f516fa
Update cost input and output values in TOML file
2026-02-25 04:26:18 +05:00
PandaSt0rm
7ccbb58c5f
add Alibaba Coding Plan provider and model configs
2026-02-25 01:10:42 +02:00
Archit Mishra
da906a0816
feat(provider): Add MiniMax-M2.5 for baseten
2026-02-24 14:33:41 -08:00
github-actions[bot]
36c9f82905
chore(vercel): update Vercel model definitions
...
Auto-generated by weekly workflow from Vercel AI Gateway API.
Co-Authored-By: github-actions[bot] <41898282+github-actions[bot]@users.noreply.github.com>
2026-02-24 21:34:20 +00:00
Frank
978214e143
update zen models
2026-02-24 15:23:24 -05:00
Jérôme Benoit
6aa69460a1
feat(sap-ai-core): add GPT-4.1, Gemini 2.5 Flash Lite, and Sonar models
...
Add 5 new model definitions for SAP AI Core provider:
- gpt-4.1: OpenAI GPT-4.1 (1M context, 32K output)
- gpt-4.1-mini: OpenAI GPT-4.1 Mini (1M context, 32K output)
- gemini-2.5-flash-lite: Google Gemini 2.5 Flash Lite (1M context, 65K output)
- sonar: Perplexity Sonar (128K context, 4K output)
- sonar-pro: Perplexity Sonar Pro (200K context, 8K output)
All specs verified against official provider documentation.
2026-02-24 20:50:10 +01:00
fhennerkes
0f16fcf231
poe: add GPT-5.3-Codex model
2026-02-24 11:39:47 -08:00
fhennerkes
e583c700f9
Merge branch 'anomalyco:dev' into dev
2026-02-24 11:36:20 -08:00
dpuyosa
96d278932c
[venice] Update gemini-3-1-pro-preview config
...
- Reduce output token limit from 250K to 65K
2026-02-24 11:27:18 +01:00
Sewer56
eee3303df0
Add missing synthetic.new models
...
Add configuration for hf:Qwen/Qwen3.5-397B-A17B and hf:MiniMaxAI/MiniMax-M2.5
to the synthetic provider, based on API specs from synthetic.new.
Note: API reports image support but these models may not natively support
images (likely rerouted/proxied through vision-capable infrastructure).
2026-02-24 09:19:57 +00:00
RioPlay
838416044f
add: newer MiniMax, GLM, and Kimi models to DeepInfra
2026-02-23 22:21:28 -06:00
Frank
51441f47d9
update zen models
2026-02-23 15:08:28 -05:00
Nacho F. Lizaur
7fc2c6154d
fix(amazon-bedrock): correct Claude Opus 4.6 context window from 1M to 200K
2026-02-23 20:10:58 +01:00
Colby Gilbert
1439781a76
feat(firmware): add deepseek 3.2, glm 5, kimi k2.5, minimax m2.5
2026-02-22 21:29:51 -08:00
minpeter
8fc0d87742
add friendli minimax m2.5 model config
2026-02-23 13:18:17 +09:00
Colby Gilbert
f660955784
feat(firmware): add grok models
2026-02-22 15:37:53 -08:00
Colby Gilbert
eb11c327b8
feat(firmware): gemini 3.1 pro, sonnet reasoning
2026-02-21 23:43:25 -08:00
Bryan Nie
0dfde60c14
models: alibaba-cn: add MiniMax-M2.5
2026-02-22 01:09:24 +08:00
Wendell Misiedjan
bab7727bad
Add AWS_BEARER_TOKEN_BEDROCK to Amazon Bedrock provider env
...
The @ai-sdk/amazon-bedrock package supports Bearer token authentication
via the AWS_BEARER_TOKEN_BEDROCK environment variable as an alternative
to IAM SigV4 auth. This uses Bedrock API keys for simplified access.
2026-02-21 16:27:45 +01:00
Mikal Sande
4b4a2364c6
Append (latest) to Mistral models that refer to the latest version.
2026-02-21 09:19:13 +01:00
Frank
c36b8e9433
update zen models
2026-02-20 23:20:24 -05:00
xiaojie.zj
5b8e983e7c
feat: add Gemini 3.1 Pro Preview for ZenMux provider
2026-02-21 10:39:30 +08:00
Frank
0d2a52dd9d
update zen models
2026-02-20 20:41:52 -05:00
Frank
b667ab78ac
update zen models
2026-02-20 20:19:33 -05:00
fhennerkes
e2da96cde4
poe: add Gemini-3.1-Pro and update Claude Sonnet 4.6
2026-02-20 11:28:18 -08:00
Jake Jia
9d042ac986
Update providers/zenmux/models/openai/gpt-5.2-pro.toml
...
Co-authored-by: Copilot <175728472+Copilot@users.noreply.github.com >
2026-02-21 01:22:04 +08:00
Phoen1xCode
7192dc0ba8
feat(openai): add GPT-5.2-Pro model via zenmux provider
...
Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com >
2026-02-21 01:07:58 +08:00
Phoen1xCode
05ea56a12a
fix(minimax): remove duplicated provider prefix from MiniMax M2.5 Lightning name
...
Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com >
2026-02-21 01:07:47 +08:00
Aiden Cline
beb449a417
Merge pull request #994 from davidfph/fix/qwen3.5-release-date
...
fix(qwen): update Qwen3.5 release_date and last_updated to 2026-02-16
2026-02-20 10:29:15 -06:00
Aiden Cline
8f76f9b217
Merge pull request #988 from MeganovaAI/fix-meganova-logo
...
Update Meganova logo to official brand icon
2026-02-20 10:29:02 -06:00
Aiden Cline
18ddcde669
fix: azure & cognitive model distinctions
2026-02-20 10:28:22 -06:00
David Fu
ff36ded35f
fix(qwen): update Qwen3.5 release_date and last_updated to 2026-02-16
2026-02-20 20:33:08 +08:00
Aiden Cline
b0b8074a94
Merge pull request #990 from kailiu42/dev
...
models: siliconflow-cn: add new models
2026-02-20 03:01:45 -06:00
Kai Liu
4c30a522f7
models: siliconflow-cn: add new models
...
New models per the latest list: https://cloud.siliconflow.cn/me/models
- Pro/MiniMaxAI/MiniMax-M2.5
- deepseek-ai/DeepSeek-OCR
- PaddlePaddle/PaddleOCR-VL
- PaddlePaddle/PaddleOCR-VL-1.5
Signed-off-by: Kai Liu <kraml.liu@gmail.com >
2026-02-20 16:26:26 +08:00
Aiden Cline
2d63d713de
Merge pull request #992 from anomalyco/fix-azure-models
...
fix: ensure that anthropic models on azure providers have correct urls
2026-02-20 02:19:13 -06:00
Aiden Cline
a1ad90a9b1
Merge pull request #991 from zainhas/dev
...
[Together AI] add qwen3.5
2026-02-20 01:02:23 -06:00
Zain Hasan
8a6e0dd917
add qwen3.5
2026-02-19 21:38:46 -08:00
Aiden Cline
2da10b739c
Merge pull request #989 from propilideno/fix/adding_missing_azure_foundry_model
...
Add missing GPT-5.2 metadata for Azure Cognitive Services
2026-02-19 18:47:56 -06:00
Aiden Cline
c17e0b9d0f
Merge pull request #987 from dpuyosa/venice
...
Venice: Add Gemini 3.1 Pro Preview and update model configs
2026-02-19 18:47:48 -06:00
Lucas Almeida
877a1175f4
chore: replacing by symbolic link like the other ones
2026-02-19 21:21:05 -03:00
Boqian
1bc83abeeb
Update Meganova logo to official brand icon
2026-02-19 18:57:17 -05:00
dpuyosa
70caedba86
[venice] Add Gemini 3.1 Pro Preview and update model configs
...
- Add new Gemini 3.1 Pro Preview model configuration
- Update Claude Sonnet 4.6 release dates
- Enable open_weights for MiniMax M25
2026-02-19 22:54:59 +01:00
Aiden Cline
60c90a27a0
Merge pull request #985 from sylviezhang37/update-vercel-model-gen-script
...
feat(provider): exclude image/video models
2026-02-19 15:49:25 -06:00
Aiden Cline
2bd0d5446e
Merge pull request #986 from riasvdv/add-gemini-3.1-pro
...
Add Gemini 3.1 Pro Preview to copilot models
2026-02-19 15:49:12 -06:00
Aiden Cline
5f135517b1
Remove audio and video from input modalities
2026-02-19 15:48:42 -06:00
Aiden Cline
e2af7819b4
Rename gemini-3.5-pro-preview.toml to gemini-3.1-pro-preview.toml
2026-02-19 15:47:25 -06:00
Rias
ca6c251b3a
Add Gemini 3.1 Pro Preview to copilot models
2026-02-19 22:43:26 +01:00
Sylvie Zhang
7b1b590d10
exclude image/video gen models
2026-02-19 13:24:11 -08:00
Aiden Cline
5097a1e954
Merge pull request #966 from mhkok/mkok/feat/add-evroc-provider
...
add evroc provider + models
2026-02-19 14:15:51 -06:00
Aiden Cline
05959a83b6
Update font family in Kimi-K2.5 configuration
2026-02-19 14:15:07 -06:00
Aiden Cline
1492e067a4
Merge pull request #976 from too-green/patch-2
...
Add Qwen3 Coder Next model for openrouter
2026-02-19 14:01:07 -06:00
Matthijs Kok
029522aa96
fix family names
2026-02-19 08:48:43 +01:00
Ahmed
482ed2e833
Add Qwen3 Coder Next model for openrouter
...
Added model configuration for Qwen3 Coder Next
2026-02-19 12:36:55 +05:00
Lucas Almeida
c82b08d778
fix: adding missing gpt-5.2 model on azure foundry
2026-02-18 23:38:10 -03:00
Matthijs Kok
7fa0eeebf9
add evroc provider + models
2026-02-18 19:57:44 +01:00
Daltonganger
33f4e77ee4
fix(nano-gpt): normalize family enums for model validation
2026-02-18 12:53:40 +01:00
Daltonganger
51ef2e51ae
finalize nano-gpt model sync and release metadata
2026-02-18 12:44:36 +01:00
Ruben Beuker
20abb5b8df
preserve curated release dates for key nano-gpt models
...
Keep existing curated release and last-updated values for models where NanoGPT API uses the generic created timestamp baseline.
2026-02-17 22:09:03 +01:00
Ruben Beuker
8cb462f29b
sync nano-gpt models with live API catalog
...
Refresh NanoGPT model files to match the current /api/v1/models output, remove stale entries, and add newly available models while preserving path-based IDs.
Also ignore local TokenSpeed sqlite artifacts so private monitoring data is not shown or committed.
2026-02-17 22:01:54 +01:00