Adam
c0ed0e62df
feat(web): redesign model-centric navigation
2026-06-04 19:16:30 -05:00
Adam
69d1cb7773
feat(web): remove benchmarks column
2026-06-04 13:03:45 -05:00
Adam
32d509dd3b
Normalize base-model inheritance ( #2007 )
2026-06-04 12:53:42 -05:00
Aiden Cline
8e6d393c01
Merge pull request #1984 from anomalyco/feat/sarvam-reasoning-options
...
feat(sarvam): add reasoning options
2026-06-04 11:59:21 -05:00
Aiden Cline
ee8104f0a1
Merge pull request #2005 from zainhas/dev
...
[Together AI] add nemotron 3 ultra
2026-06-04 11:51:16 -05:00
Aiden Cline
10155a62eb
Merge pull request #2004 from anomalyco/fix/sync-reasoning-options
...
fix(sync): preserve reasoning options
2026-06-04 11:51:03 -05:00
Zain Hasan
afa81334ce
[Together AI] add nemotron 3 ultra
2026-06-04 09:49:59 -07:00
Aiden Cline
cc956258d6
fix(sync): preserve reasoning options
2026-06-04 11:49:16 -05:00
Aiden Cline
dfcf5ba1cf
Merge pull request #1999 from houtanb/fix-together-deepseek-models
...
fix(together): DeepSeek-R1, DeepSeek-V3 name & release date
2026-06-04 11:40:33 -05:00
Aiden Cline
637f7e3eec
Merge pull request #1997 from dpuyosa/update-models
...
Venice: Add Qwen 3.7 Plus and update models
2026-06-04 11:40:20 -05:00
Aiden Cline
a4b39711da
Merge pull request #1996 from anomalyco/automation/sync-models-openrouter
...
chore(sync): update OpenRouter model catalog
2026-06-04 11:39:34 -05:00
Aiden Cline
8f5d9e7b82
Merge pull request #1998 from chenkuilinckl-ai/fix/qwen3.7-plus-vision-1m-context
...
fix(alibaba): qwen3.7-plus GA adds vision (image+video) and 1M context
2026-06-04 11:39:04 -05:00
Aiden Cline
7beeab6cef
Merge pull request #2003 from BlockListed/cortecs-gpt-5-4
...
add gpt 5.4 to cortecs
2026-06-04 11:37:46 -05:00
BlockListed
0a03c18424
add gpt 5.4 to cortecs
2026-06-04 18:13:57 +02:00
Adam
c8a52d19c1
feat(models): add coding benchmarks and weights ( #2000 )
...
* feat(models): add coding benchmarks and weights
* feat(models): add more coding benchmarks
* feat(models): add agent benchmark scores
* feat(models): normalize benchmark metadata
2026-06-04 11:09:53 -05:00
Adam
859bb31ffa
refactor(chutes): use base_model for tee wrappers ( #2002 )
2026-06-04 11:05:45 -05:00
Adam
a672fe4b08
feat(web): surface model metadata links ( #2001 )
2026-06-04 11:03:53 -05:00
Frank
72fc81a4da
update zen models
2026-06-04 11:31:06 -04:00
github-actions[bot]
2f77d0afac
chore(sync): update OpenRouter model catalog
2026-06-04 15:25:27 +00:00
Houtan Bastani
03333d9aa4
fix(together): DeepSeek-R1, DeepSeek-V3 name & release date
...
See timestamps:
* "Release DeepSeek-R1": https://github.com/deepseek-ai/DeepSeek-R1/commit/23807ced51627276434655dd9f27725354818974
* "Release DeepSeek-V3": https://github.com/deepseek-ai/DeepSeek-V3/commit/4c2fdb8f55e049553b9f4f1a3241f86d739c8cf8
2026-06-04 13:16:14 +02:00
chenkuilinckl-ai
977e2f0155
fix(alibaba): qwen3.7-plus GA adds vision (image+video) and 1M context
2026-06-04 17:29:42 +08:00
dpuyosa
963e9a6bab
[venice] Add Qwen 3.7 Plus and update models
...
- Added qwen-3-7-plus
- Update google-gemma-4-31b-it cache_read
- Update minimax-m3 output limit, modalities
2026-06-04 10:01:57 +02:00
Aiden Cline
9e8cad9ac9
Merge pull request #1988 from anomalyco/feat/stepfun-reasoning-options
...
feat(stepfun): add reasoning effort options
2026-06-04 00:38:32 -05:00
Aiden Cline
f6c1036a86
Merge pull request #1985 from anomalyco/feat/xiaomi-reasoning-options
...
feat(xiaomi): add reasoning toggles
2026-06-04 00:37:58 -05:00
Aiden Cline
1926832a9d
fix(sarvam): expose null reasoning effort
2026-06-04 00:17:23 -05:00
Aiden Cline
cb22ad5625
Merge pull request #1992 from anomalyco/fix/sync-base-model-output
...
fix(sync): preserve base model output
2026-06-04 00:02:46 -05:00
Aiden Cline
909db75087
fix(sync): preserve base model output
2026-06-04 00:00:46 -05:00
Aiden Cline
b551552f14
Merge pull request #1983 from anomalyco/feat/cohere-reasoning-options
...
feat(cohere): add reasoning options
2026-06-03 23:34:36 -05:00
Aiden Cline
0919062b40
Merge pull request #1989 from anomalyco/feat/google-gemini-reasoning-options
...
feat(google): add Gemini reasoning options
2026-06-03 23:05:44 -05:00
Aiden Cline
8662c63313
fix(google): retain deprecated Gemini reasoning metadata
2026-06-03 23:03:08 -05:00
Aiden Cline
36b808691a
Merge pull request #1991 from shzdehmd/dev
...
chore(fireworks): update qwen3p6-plus limits to 262K context / 65K output
2026-06-03 22:21:52 -05:00
Ahmad Shahzad
c8feebb7cd
chore(fireworks): update qwen3p6-plus limits to 262K context / 65K output
2026-06-04 07:05:29 +05:00
Aiden Cline
259801fba5
fix(google): exclude unavailable Gemini 3 Pro preview
2026-06-03 18:28:25 -05:00
Aiden Cline
ce35e18561
feat(stepfun): add reasoning effort options
2026-06-03 17:49:21 -05:00
Aiden Cline
133a0b0126
feat(google): add Gemini reasoning options
2026-06-03 17:49:09 -05:00
Aiden Cline
99406ae7df
feat(xiaomi): add reasoning toggles
2026-06-03 17:48:22 -05:00
Aiden Cline
4f25170be2
feat(cohere): add reasoning options
2026-06-03 17:48:05 -05:00
Aiden Cline
6ddf935238
feat(sarvam): add reasoning options
2026-06-03 17:48:00 -05:00
Aiden Cline
6ae56b00a8
Merge pull request #1981 from anomalyco/feat/glm-coding-plan-reasoning-toggle
...
feat(glm): add Zhipu and coding plan reasoning toggles
2026-06-03 17:31:39 -05:00
Aiden Cline
2cb0d28e17
feat(zhipuai): add reasoning toggles
2026-06-03 17:27:57 -05:00
Aiden Cline
03d90aeacc
Merge pull request #1982 from anomalyco/fix/sync-model-catalog-matrix
...
fix(sync): restore model catalog workflow
2026-06-03 17:26:22 -05:00
Aiden Cline
eb7dbead75
fix(sync): restore model catalog workflow
2026-06-03 17:20:01 -05:00
Aiden Cline
36cbbfc577
feat(glm): add coding plan reasoning toggles
2026-06-03 16:56:42 -05:00
Aiden Cline
d84e883ede
Merge pull request #1980 from anomalyco/feat/deepseek-reasoning-options
...
feat(deepseek): add reasoning options
2026-06-03 16:09:22 -05:00
Aiden Cline
dd9d6ff54e
feat(deepseek): add reasoning options
2026-06-03 16:08:09 -05:00
Aiden Cline
d7e19c7627
Merge pull request #1979 from anomalyco/feat/moonshot-reasoning-toggle
...
feat(moonshot): add reasoning toggle options
2026-06-03 15:46:39 -05:00
Aiden Cline
ca3eb39fbd
feat(moonshot): add reasoning toggle options
2026-06-03 15:45:15 -05:00
Aiden Cline
d136e7b036
Merge pull request #1955 from eliasaronson/chore/mark-deprecated-models
...
chore: mark retired models as deprecated
2026-06-03 15:25:27 -05:00
Adam
f6c6f04367
feat(models): add model metadata ( #1974 )
...
* feat(models): add model metadata
* feat(models): rename model metadata namespaces
2026-06-03 15:13:54 -05:00
Aiden Cline
1e9e4bbdab
Merge pull request #1977 from Ardakilic/feat/nano-gpt-20260603
...
chore: sync nano-gpt models: 20260603
2026-06-03 15:02:42 -05:00
Arda Kılıçdağı
efb553d287
chore: sync nano-gpt models: 20260603
2026-06-03 21:31:28 +03:00
Jack
d9b83ca9cd
Merge pull request #1975 from anomalyco/update/opencode-go-qwen3.7-plus
...
feat(opencode-go): add Qwen3.7 Plus model
2026-06-04 01:30:11 +08:00
Aiden Cline
25592e361a
Merge pull request #1976 from jerome-benoit/feat/sap-ai-core-gpt-5.5
...
feat(sap-ai-core): add GPT-5.5
2026-06-03 12:28:07 -05:00
Jérôme Benoit
012928800c
fix(sap-ai-core): align GPT-5.4 release_date with upstream
...
SAP AI Core routes to OpenAI gpt-5.4; release_date should reflect
the actual model release (2026-03-05) rather than the SAP catalog
availability date (2026-04-27).
2026-06-03 19:09:54 +02:00
Jérôme Benoit
ee60c2e4d4
feat(sap-ai-core): add GPT-5.5
...
SAP AI Core routes to OpenAI gpt-5.5; specs mirror the canonical
provider/openai/gpt-5.5 with the established sap-ai-core wrapper
adjustments (lowercase name, drop [[cost.tiers]], drop
[experimental.modes.fast]).
2026-06-03 19:06:07 +02:00
Jack
7e76cde0b1
fix(opencode-go): correct Qwen3.7 Plus dates
2026-06-04 00:59:42 +08:00
Jack
bbd2479e12
feat(opencode-go): add Qwen3.7 Plus model
2026-06-04 00:57:42 +08:00
Aiden Cline
eeb17ccab4
Merge pull request #1971 from coder-wangbin/feat/alibaba-cn-qwen3.7-plus
...
feat: add Qwen3.7 Plus model for alibaba-cn provider
2026-06-03 10:00:39 -05:00
Aiden Cline
46ffeee012
Merge pull request #1972 from Phosmachina/feature/update-deepinfra-deepseek-and-mimo
...
Feature/update deepinfra deepseek and mimo
2026-06-03 09:56:37 -05:00
Aiden Cline
8c1e45007d
Merge pull request #1967 from anomalyco/automation/sync-models-openrouter
...
chore(sync): update OpenRouter model catalog
2026-06-03 09:54:14 -05:00
Frank
364628cb93
update zen models
2026-06-03 10:32:41 -04:00
github-actions[bot]
50f5d40b4e
chore(sync): update OpenRouter model catalog
2026-06-03 14:02:08 +00:00
Michel Paronnaud
707be517c6
chore(deepinfra): adjust prices and limit for DeepSeek models
2026-06-03 12:40:49 +02:00
Michel Paronnaud
657598e609
fix(deepinfra): path for mimo models
2026-06-03 12:28:08 +02:00
wangbin
0999a475ba
feat: add Qwen3.7 Plus model for alibaba-cn provider
...
- Add base definition in providers/alibaba/models/qwen3.7-plus.toml
- Add extends reference in providers/alibaba-cn/models/qwen3.7-plus.toml
- Release date: 2026-06-02
- Context: 131K tokens, Output: 16K tokens
- Pricing: $0.50/$3.00 per 1M tokens (input/output)
- Supports reasoning and tool calling
2026-06-03 17:40:12 +08:00
Frank
e1ea0254ed
update zen models
2026-06-02 22:51:19 -04:00
Aiden Cline
7fad18b054
fix
2026-06-02 16:36:29 -05:00
Aiden Cline
710bbc5375
Merge pull request #1961 from anomalyco/feat/minimax-m3-reasoning-toggle
...
feat(minimax): add M3 reasoning toggle
2026-06-02 16:34:19 -05:00
Aiden Cline
0b9fa62153
Merge pull request #1963 from nicholasgriffintn/open-mistral-nemo
...
fix: update mistral nemo
2026-06-02 14:34:54 -05:00
Aiden Cline
126a481a70
Merge pull request #1965 from nicholasgriffintn/update-devstral-models
...
chore: update devstral models
2026-06-02 14:34:38 -05:00
Aiden Cline
b080cf1fe4
Merge pull request #1952 from anomalyco/automation/sync-models-xai
...
chore(sync): update xAI model catalog
2026-06-02 14:03:02 -05:00
Nicholas Griffin
042002d3a3
chore: update devstral models
2026-06-02 19:47:17 +01:00
Nicholas Griffin
a192b51e84
chore: undo
2026-06-02 19:46:13 +01:00
Nicholas Griffin
b870c796ef
chore: undo
2026-06-02 19:45:25 +01:00
Nicholas Griffin
35e4981aa9
chore: update
2026-06-02 19:40:02 +01:00
Frank
970b660070
sync
2026-06-02 14:24:55 -04:00
Nicholas Griffin
bb03a8c207
fix: update mistral nemo
2026-06-02 19:22:36 +01:00
github-actions[bot]
b6dbb68c32
chore(sync): update xAI model catalog
2026-06-02 18:21:19 +00:00
Aiden Cline
b34a7c21f3
feat(minimax): add M3 reasoning toggle
2026-06-02 12:43:17 -05:00
Aiden Cline
4b169b8cb6
Merge pull request #1960 from anomalyco/feat/zai-reasoning-toggle
...
feat(zai): add reasoning toggle options
2026-06-02 12:38:15 -05:00
Aiden Cline
f8ecf4daec
feat(zai): add reasoning toggle options
2026-06-02 12:02:54 -05:00
Frank
eb68be3fbc
stats
2026-06-02 12:38:49 -04:00
Aiden Cline
55fb058cc5
tweak: update available options
2026-06-02 11:35:26 -05:00
Aiden Cline
5545a98866
tweak: handle missing omits gracefully
2026-06-02 11:25:55 -05:00
Aiden Cline
1a7b103bba
fix: omit
2026-06-02 11:22:52 -05:00
Aiden Cline
f341858398
Merge pull request #1896 from stevenyeung/add/alibaba-token-plan
...
feat: add Alibaba Token Plan provider with 15 models
2026-06-02 11:20:42 -05:00
Aiden Cline
f2d1e42582
Merge pull request #1951 from anomalyco/automation/sync-models-openrouter
...
chore(sync): update OpenRouter model catalog
2026-06-02 11:03:27 -05:00
Aiden Cline
2f9d2d8d78
Merge pull request #1957 from kapelame/fix/minimax-m3-prune
...
fix(minimax): correct MiniMax-M3 pricing and max output
2026-06-02 10:53:58 -05:00
Aiden Cline
bf6b0e32d7
Merge pull request #1959 from Tavernari/feat/claudinio-add-audio-video
...
feat(claudinio): add audio and video input modalities
2026-06-02 10:53:25 -05:00
github-actions[bot]
5b4d2ea2e4
chore(sync): update OpenRouter model catalog
2026-06-02 15:41:53 +00:00
Victor Carvalho Tavernari
ee4badab7b
fix(claudinio): update cache_read price to 0.150 per 1M tokens
2026-06-02 16:34:05 +01:00
Victor Carvalho Tavernari
6c05a6b519
Merge branch 'dev' into feat/claudinio-add-audio-video
2026-06-02 16:32:12 +01:00
Victor Carvalho Tavernari
91ae9a48e3
feat(claudinio): add audio and video input modalities
2026-06-02 16:24:06 +01:00
kapelame
85d02711bd
fix(minimax): correct MiniMax-M3 pricing and max output
...
The M3 entries added in #1940 copied M2.7's cost values. Correct them to
the official M3 pricing and limits:
- minimax / minimax-cn (pay-as-you-go): input 0.30 -> 0.60,
output 1.20 -> 2.40, cache_read 0.06 -> 0.12, and remove cache_write
(M3 has no active prompt-cache-write tier).
- max output 131072 -> 128000 across all four providers.
- coding-plan variants keep their subscription-plan zero pricing; only
max output is corrected.
Context (512K), modalities, and the other flags are unchanged.
2026-06-02 21:03:45 +08:00
Elias H Aronsson
1ba404612d
chore: mark retired models as deprecated
...
Add status = "deprecated" to models that are past their provider's
shutdown/retirement date (no longer served by the public API).
Google Gemini (4):
gemini-2.0-flash, gemini-2.0-flash-lite, gemini-3-pro-preview,
gemini-3.1-flash-lite-preview
Anthropic Claude (7):
claude-3-sonnet-20240229, claude-3-5-sonnet-20240620,
claude-3-5-sonnet-20241022, claude-3-opus-20240229,
claude-3-7-sonnet-20250219, claude-3-5-haiku-20241022,
claude-3-haiku-20240307
OpenAI (2):
o1-preview, o1-mini
Sources:
https://ai.google.dev/gemini-api/docs/deprecations
https://platform.claude.com/docs/en/about-claude/model-deprecations
https://developers.openai.com/api/docs/deprecations
2026-06-02 10:06:18 +02:00
Aiden Cline
f91dd4ad0b
Merge pull request #1950 from anomalyco/fix/reasoning-options-inheritance
...
fix: do not inherit reasoning options
2026-06-01 23:24:13 -05:00
Aiden Cline
449b926f40
fix: do not inherit reasoning options
2026-06-01 23:22:04 -05:00
Aiden Cline
2546ffe570
Merge pull request #1940 from matstrange/add-minimax-m3
...
Add MiniMax-M3 model (#1933 )
2026-06-01 22:57:45 -05:00
Hex Agent
f33ff9ba78
Add MiniMax-M3 model to 5 providers
...
MiniMax-M3 is MiniMax's new frontier multimodal coding model: 1M context
window (512K minimum on ollama-cloud), native text/image/video input,
tool calling, reasoning, and open weights.
Adds the model to all five providers where it should be available:
- minimax (pay-as-you-go)
- minimax-cn (pay-as-you-go, China)
- minimax-coding-plan (token plan subscription)
- minimax-cn-coding-plan (token plan subscription, China)
- ollama-cloud
Closes #1933 .
Notes for reviewers:
- Cost fields on minimax/minimax-cn match M2.7; M3 docs state the
pricing is unchanged from M2.7.
- The 1M/512K context divergence on ollama-cloud is intentional —
ollama advertises 1M with a 512K minimum, so 512K is the safe floor
that won't surprise opencode users with mid-request rejections.
- output = 131072 is inherited from the existing M2.7 files; MiniMax's
published M3 docs only advertise the 1M input context, not a separate
output cap.
- The minimax-coding-plan variant has been verified end-to-end in
opencode against the MiniMax token plan API.
2026-06-01 22:54:17 -05:00
Aiden Cline
98bf80cd77
Merge pull request #1949 from anomalyco/fix/github-copilot-context-limits-complete
...
fix(github-copilot): preserve model-specific limits
2026-06-01 22:49:00 -05:00
Aiden Cline
68660cad83
Merge pull request #1938 from anyapi-ai/dev
...
Add AnyAPI provider
2026-06-01 22:47:26 -05:00
Aiden Cline
78e6c1a2dd
fix(github-copilot): preserve model-specific limits
2026-06-01 22:46:56 -05:00
Aiden Cline
062ba1fbb7
Merge pull request #1939 from anomalyco/automation/sync-models-openrouter
...
chore(sync): update OpenRouter model catalog
2026-06-01 22:46:18 -05:00
Aiden Cline
477ccffee8
Merge pull request #1948 from anomalyco/feat/anthropic-reasoning-options
...
feat(anthropic): add reasoning options
2026-06-01 22:41:23 -05:00
github-actions[bot]
6ce2f88eac
chore(sync): update OpenRouter model catalog
2026-06-02 03:26:49 +00:00
Aiden Cline
4501589a20
feat(anthropic): add reasoning options
2026-06-01 22:17:07 -05:00
Aiden Cline
0b643efd17
Merge pull request #1947 from LYY/fix/github-copilot-context-limits
...
fix(github-copilot): correct limits for 8 models with verified CAPI data (partial, refs #1946 )
2026-06-01 22:16:07 -05:00
LYY
8a0ea15aca
fix(github-copilot): correct limits for models with verified CAPI data
...
The github-copilot model stubs use [extends] and inherit the upstream
base model's context/input/output limits (~1M), which do not match the
limits GitHub Copilot actually enforces via api.githubcopilot.com/models.
Override the limit fields for the 8 models whose live CAPI values were
verified first-hand. Models that were not enabled on the test account
(no CAPI data available) are intentionally left unchanged.
Refs anomalyco/models.dev#1946
2026-06-02 11:08:51 +08:00
Aiden Cline
b3676fd77d
Merge pull request #1934 from yukoba/github-copilot-2026-06
...
June 2026 changes of GitHub Copilot
2026-06-01 16:10:45 -05:00
Aiden Cline
f97a02078c
Merge pull request #1932 from eliasto/ovhcloud/update-models-qwen
...
feat(ovhcloud): Add new Qwen models and sync mode
2026-06-01 14:27:58 -05:00
Aiden Cline
2a7153fcef
Merge pull request #1943 from peculiarnewbie/fix/crof-models-update
...
fix: update crof.ai models to match current API
2026-06-01 13:41:23 -05:00
Aiden Cline
3396f15854
Merge pull request #1942 from anomalyco/feat/reasoning-options-schema
...
feat: add reasoning options schema
2026-06-01 13:41:10 -05:00
Aiden Cline
ccfd99b5e9
Merge pull request #1941 from KTibow/fix-llmgateway-deepseek-v3-2-price
...
fix(llmgateway): update model pricing
2026-06-01 13:40:20 -05:00
bolt
ef6eb88216
fix: update crof.ai models to match current API
2026-06-02 00:52:14 +07:00
Aiden Cline
94e128244a
feat: add reasoning options schema
2026-06-01 12:48:11 -05:00
KTibow
ec6d07d41c
fix(llmgateway): use weighted pricing selection
2026-06-01 10:41:23 -07:00
KTibow
4efbe58af8
fix(llmgateway): update model pricing
2026-06-01 10:24:05 -07:00
Christina
3b55102a45
Add AnyAPI provider with 30 models
...
Adds AnyAPI (https://anyapi.ai ) as a new provider. Models reuse existing
canonical entries through `extends`. Cost fields are omitted as AnyAPI uses
a credit-based pricing system. Validated locally with `bun validate`.
Models (30):
- openai: gpt-5.4, gpt-5.2, gpt-5.1, gpt-5, gpt-5-mini, gpt-4.1, gpt-4.1-mini, o4-mini, o3, o3-mini
- anthropic: claude-opus-4-7, claude-opus-4-6, claude-sonnet-4-6, claude-sonnet-4-5, claude-haiku-4-5
- google: gemini-2.5-pro, gemini-2.5-flash, gemini-2.5-flash-lite, gemini-3-pro-preview, gemini-3-flash-preview
- deepseek: deepseek-v4-pro, deepseek-v4-flash, deepseek-chat, deepseek-r1
- mistralai: mistral-large-2512, devstral-2512
- perplexity: sonar-pro, sonar-reasoning-pro
- cohere: command-r-plus-08-2024
- xai: grok-4.3
2026-06-01 17:14:24 +02:00
Aiden Cline
a4fe8fc36b
fmt
2026-06-01 09:53:16 -05:00
Aiden Cline
b24b44ca0a
Merge pull request #1920 from anomalyco/automation/sync-models-openrouter
...
chore(sync): update OpenRouter model catalog
2026-06-01 09:39:43 -05:00
Aiden Cline
466d664bf9
Merge pull request #1930 from dpuyosa/dev
...
Venice: Add MiniMax M3 model
2026-06-01 09:39:27 -05:00
Aiden Cline
340ef131e9
Merge pull request #1936 from JoshuaDietz/dev
...
feat(ollama-cloud): add minimax m3
2026-06-01 09:38:47 -05:00
Aiden Cline
6ce69a706a
Merge pull request #1937 from KTibow/fix/hpc-ai-pricing
...
fix: update HPC-AI model pricing
2026-06-01 09:38:34 -05:00
KTibow
3505674a1b
fix: update HPC-AI model pricing
2026-06-01 07:29:43 -07:00
github-actions[bot]
cfeaef8f77
chore(sync): update OpenRouter model catalog
2026-06-01 14:25:41 +00:00
Joshua Dietz
f73a5d571b
feat(ollama-cloud): add minimax m3
2026-06-01 16:14:49 +02:00
Yu Kobayashi
62d31abff8
June 2026 changes of GitHub Copilot
2026-06-01 22:03:44 +09:00
Elias TOURNEUX
ee71ea7181
feat(ovhcloud): Add qwen 3.6 27b
2026-06-01 11:57:58 +02:00
Elias TOURNEUX
b038f40b64
feat(ovhcloud): Add OVHcloud sync mode
2026-06-01 10:28:46 +02:00
Elias TOURNEUX
4b7bdd9692
feat(ovhcloud): Add new Qwen models
2026-06-01 10:16:54 +02:00
dpuyosa
b73ef1a634
[venice] Add MiniMax M3 model
...
- Add minimax-m3.toml with cost, limits, modalities
- Enable text/image input and text output support
- Set 500k context, 32k output, cache pricing
2026-06-01 09:38:12 +02:00
Aiden Cline
c11b5bb00b
Merge pull request #1927 from Jercik/codex/update-wafer-price-cuts
...
fix: update Wafer.ai pricing
2026-06-01 00:01:03 -05:00
Łukasz Jerciński
01bc1785b4
fix: update Wafer price cuts
2026-06-01 06:25:20 +02:00
Aiden Cline
36b5847fa4
Merge pull request #1925 from anomalyco/fix/vercel-claude-opus-4-1-symlink
...
fix vercel claude opus 4.1 symlink
2026-05-31 22:12:27 -05:00
Aiden Cline
040beaf818
fix vercel claude opus 4.1 symlink
2026-05-31 22:11:05 -05:00
Frank
9f8e1a3858
update zen models
2026-05-31 22:07:52 -04:00
Jack
ce7a1c0218
update context limit of M3 in Go
2026-06-01 09:45:46 +08:00
Jack
55e6c7a2ee
add image,video modalities to M3
2026-06-01 08:58:02 +08:00
Frank
8bacf496fe
update zen models
2026-05-31 19:49:41 -04:00
Frank
ac12cf1f43
update zen models
2026-05-31 19:47:33 -04:00
Frank
06c7ea2303
update zen models
2026-05-31 19:47:04 -04:00
Frank
92b6e9da53
update zen models
2026-05-31 19:43:09 -04:00
Frank
660ff026a0
update zen models
2026-05-31 19:40:57 -04:00
Aiden Cline
9a910a1fb3
Merge pull request #1917 from markgibaud/fix/eu-opus-bedrock-pricing
...
fix(bedrock): correct EU cross-region pricing for Claude Opus and Sonnet models
2026-05-31 17:37:45 -05:00
Aiden Cline
7dc5f787b0
Merge pull request #1915 from anomalyco/automation/sync-models-openrouter
...
chore(sync): update OpenRouter model catalog
2026-05-31 17:10:10 -05:00
Aiden Cline
c25ddce07c
Merge pull request #1916 from anomalyco/automation/sync-models-cloudflare-workers-ai
...
chore(sync): update Cloudflare Workers AI model catalog
2026-05-31 17:05:57 -05:00
github-actions[bot]
e73a82e44e
chore(sync): update OpenRouter model catalog
2026-05-31 21:36:49 +00:00
github-actions[bot]
7de9710b5a
chore(sync): update Cloudflare Workers AI model catalog
2026-05-31 21:36:49 +00:00
Frank
d1a8dbcdde
update zen models
2026-05-31 14:19:49 -04:00
markgibaud
556ef84d11
fix(bedrock): correct EU cross-region pricing for Claude Opus and Sonnet models
...
AWS Bedrock EU (Europe/London) cross-region inference has a 10% premium
over the base Anthropic pricing. The EU models were incorrectly using
the same pricing as the US/base models.
Affected models:
- eu.anthropic.claude-opus-4-6-v1
- eu.anthropic.claude-opus-4-7
- eu.anthropic.claude-opus-4-8
- eu.anthropic.claude-sonnet-4-5-20250929-v1:0
- eu.anthropic.claude-sonnet-4-6
Corrected Opus per 1M token prices:
- Input: $5.00 -> $5.50
- Output: $25.00 -> $27.50
- Cache read: $0.50 -> $0.55
- Cache write: $6.25 -> $6.875
Corrected Sonnet per 1M token prices:
- Input: $3.00 -> $3.30
- Output: $15.00 -> $16.50
- Cache read: $0.30 -> $0.33
- Cache write: $3.75 -> $4.125
Source: AWS Bedrock pricing page, Europe (London) region
2026-05-31 11:07:18 +01:00
Aiden Cline
f2020553ea
Merge pull request #1898 from Suat-B/codex/xpersona-frieren-1
...
Rename Xpersona Frieren display name
2026-05-30 16:44:39 -05:00
Aiden Cline
ca0a7e17cb
Merge pull request #1910 from anomalyco/automation/sync-models-openrouter
...
chore(sync): update OpenRouter model catalog
2026-05-30 16:44:22 -05:00
github-actions[bot]
d2468eafb2
chore(sync): update OpenRouter model catalog
2026-05-30 21:37:16 +00:00
Aiden Cline
96d0f9beff
Merge pull request #1913 from orangeclk/add-glm-4.6v-zhipuai-coding-plan
...
feat(zhipuai-coding-plan): add glm-4.6v symlink from zai
2026-05-30 11:52:50 -05:00
Aiden Cline
7989be9f34
Merge pull request #1912 from Jercik/codex/update-wafer-provider-models
...
feat: update Wafer provider models
2026-05-30 11:52:41 -05:00
Łukasz Jerciński
f30e44f4b5
feat: update Wafer provider models
2026-05-30 14:53:02 +02:00
OrangeCLK
1383177893
feat(zhipuai-coding-plan): add glm-4.6v symlink from zai
2026-05-30 18:42:09 +08:00
Aiden Cline
2e58165af9
Merge pull request #1902 from kameshsampath/feat/provider/snowflake-cortex
...
feat(snowflake-cortex): add Snowflake Cortex provider
2026-05-29 23:55:55 -05:00
Kamesh Sampath
f3466affc0
feat(snowflake-cortex): add Snowflake Cortex provider
...
Adds the snowflake-cortex provider which exposes Snowflake's Cortex
REST API (OpenAI Chat Completions-compatible endpoint) to opencode.
Provider details:
- npm: @ai-sdk/openai-compatible
- API: https://${SNOWFLAKE_ACCOUNT}.snowflakecomputing.com/api/v2/cortex/v1
- Auth: SNOWFLAKE_ACCOUNT + SNOWFLAKE_CORTEX_PAT (Programmatic Access Token)
Models (11, all with tool_call support):
- Anthropic: claude-opus-4-7 (beta/preview), claude-sonnet-4-6,
claude-sonnet-4-5, claude-haiku-4-5
- OpenAI: openai-gpt-5.4 (beta), openai-gpt-5.2, openai-gpt-5.1,
openai-gpt-5 (beta), openai-gpt-5-mini (beta), openai-gpt-5-nano (beta),
openai-gpt-4.1
Models without tool_call support (deepseek-r1, llama3.1-70b,
snowflake-llama-3.3-70b, mistral-large2) are excluded per Snowflake docs:
"Tool calling is supported for OpenAI and Claude models only."
Preview/not-GA models are marked with status = "beta".
Cost fields are intentionally omitted on all models. The Cortex REST API
is billed in USD (AI_INFERENCE service type) at rates defined in the
Snowflake Service Consumption Table.
Closes #1895
2026-05-30 09:04:43 +05:30
Aiden Cline
3697c99297
Merge pull request #1901 from anomalyco/automation/sync-models-openrouter
...
chore(sync): update OpenRouter model catalog
2026-05-29 20:50:33 -05:00
Aiden Cline
796dda2fd1
Merge pull request #1904 from sdnts/dev
...
feat(cloudflare-ai-gateway): add opus 4.8
2026-05-29 20:50:11 -05:00
Aiden Cline
1d363cbd7f
Merge pull request #1908 from anomalyco/fix/unique-provider-names
...
Ensure provider names are unique
2026-05-29 20:49:54 -05:00
Aiden Cline
3758725979
Ensure provider names are unique
...
Rename the China StepFun provider to "StepFun (China)" so it no longer
collides with the international "StepFun" provider, following the existing
Alibaba / Alibaba (China) naming convention.
Add a validation check in generate() that throws when two providers share
a (case-insensitive) name, preventing future duplicates.
Fixes #1906
2026-05-29 20:48:09 -05:00
github-actions[bot]
c039e82f12
chore(sync): update OpenRouter model catalog
2026-05-30 01:16:42 +00:00
Siddhant
1e90242cdb
feat(cloudflare-ai-gateway): add opus 4.8
2026-05-29 13:55:09 -04:00
Aiden Cline
277ac8577e
Merge pull request #1897 from mikeyp/update-digitalocean-models
...
Add deepseek-4-flash and claude-opus-4.8 for DigitalOcean
2026-05-29 10:20:26 -05:00
Aiden Cline
6e5f35546d
Merge pull request #1900 from ceyhanmolla/add-step-3.7-flash-v2
...
Add Step 3.7 Flash model to NVIDIA provider
2026-05-29 10:20:02 -05:00
ceyhanmolla
aa69da14d6
Add Step 3.7 Flash model to NVIDIA provider
...
StepFun AI's Step 3.7 Flash - sparse MoE multimodal reasoning model:
- 198B total params, ~11B active per token
- 256K context window with sliding window attention
- Text + image input, text output
- Reasoning, tool calling, and attachment support
- Apache 2.0 license
2026-05-29 12:55:13 +02:00
Jack
c7af5ba9be
update model in Go
2026-05-29 16:20:16 +08:00
Suat-B
d6e9e6735f
Rename Xpersona Frieren model display name
2026-05-29 03:00:21 -05:00
Mike Prasuhn
a987719410
Add deepseek-4-flash and claude-opus-4.8
2026-05-29 03:45:50 -04:00
Steven Yeung
3b792029c3
feat(alibaba-token-plan): add provider and 15 model TOMLs
...
Add Alibaba Cloud Model Studio Token Plan (Team Edition) provider with:
- Provider config (Singapore region, OpenAI-compatible endpoint)
- 10 models using extends pattern (zero-cost overrides from canonical providers)
- 5 models with full definitions (no canonical source available)
- Image generation models (qwen-image, wan2.7) with output=0 per convention
- deepseek-v3.2 with structured_output and corrected release date
Models: qwen3.7-max, qwen3.6-flash, qwen3.6-plus, kimi-k2.5, kimi-k2.6,
glm-5, glm-5.1, MiniMax-M2.5, deepseek-v4-pro, deepseek-v4-flash,
deepseek-v3.2, qwen-image-2.0, qwen-image-2.0-pro, wan2.7-image, wan2.7-image-pro
2026-05-29 15:12:27 +08:00
Aiden Cline
9528528f69
Merge pull request #1890 from anomalyco/automation/sync-models-openrouter
...
chore(sync): update OpenRouter model catalog
2026-05-28 23:42:03 -05:00
github-actions[bot]
fbad40d863
chore(sync): update OpenRouter model catalog
2026-05-29 03:26:05 +00:00
Aiden Cline
84559b598b
Merge pull request #1894 from anomalyco/add-siliconflow-cn-deepseek-v4-pro
...
feat(siliconflow-cn): add deepseek-ai/DeepSeek-V4-Pro
2026-05-28 19:10:18 -05:00
Aiden Cline
9d45730db9
Merge pull request #1889 from fhennerkes/dev
...
poe: add Claude-Opus-4.8 model
2026-05-28 19:10:06 -05:00
Aiden Cline
791089e4aa
Merge pull request #1891 from ticoombs/dev
...
feat(copilot): update all copilot models, add claude-opus-4.8
2026-05-28 19:09:52 -05:00
Aiden Cline
fcc4387187
feat(siliconflow-cn): add deepseek-ai/DeepSeek-V4-Pro
2026-05-28 19:09:15 -05:00
Tim C
6aa32ba566
fix(copilot): update all context,input,output with correct limits
2026-05-29 08:56:19 +10:00
Tim C
459a563d2b
feat(copilot): add claude-opus-4.8
2026-05-29 08:52:42 +10:00
Aiden Cline
739e5a7c8e
Merge pull request #1877 from aakash-gupte/add-merge-gateway-provider
...
Add Merge Gateway provider
2026-05-28 16:55:40 -05:00
Aiden Cline
efc87afa3a
Merge pull request #1888 from smakosh/add-llmgateway-opus-4-8
...
feat(llmgateway): add Claude Opus 4.8
2026-05-28 16:54:46 -05:00
Aiden Cline
f616aa6a0b
Merge pull request #1887 from dpuyosa/dev
...
Venice: Add Claude Opus 4.8 models
2026-05-28 16:54:35 -05:00
fhennerkes
f77e75a567
poe: add Claude-Opus-4.8 model
...
Add new Anthropic model from Poe API (released 2026-05-28).
Uses extends format inheriting from anthropic/claude-opus-4-8
with Poe-specific overrides (name format, markup pricing,
slightly different context limit).
Co-Authored-By: Claude Opus 4.7 <noreply@anthropic.com >
2026-05-28 14:39:52 -07:00
smakosh
651bd4d91a
feat(llmgateway): add Claude Opus 4.8
...
Extends anthropic/claude-opus-4-8, omitting the fast mode which LLM
Gateway does not expose, matching the existing 4.6/4.7 entries.
Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com >
2026-05-28 22:07:40 +01:00
dpuyosa
30b9170106
[venice] Add Claude Opus 4.8 models
...
- Add claude-opus-4-8.toml (standard pricing)
- Add claude-opus-4-8-fast.toml (fast variant)
- Both: 1M context, image input, Opus 4.8 family
2026-05-28 22:54:39 +02:00
Aiden Cline
81e42c653c
Merge pull request #1886 from vglafirov/add-gitlab-opus-4-8
...
feat: add GitLab Agentic Chat Opus 4.8
2026-05-28 15:25:06 -05:00
Vladimir Glafirov
11f2a38594
feat: add GitLab Agentic Chat Opus 4.8
2026-05-28 22:12:12 +02:00
Aiden Cline
6621be978f
Merge pull request #1880 from bas3line/update-routing-run-base-url
...
Update routing.run API base URL
2026-05-28 14:46:09 -05:00
Aiden Cline
c8d72661d6
Merge pull request #1885 from alaviss/bedrock-opus-4-8
...
feat: add cross-region inference entries for Bedrock for Opus 4.8
2026-05-28 14:45:43 -05:00
Hiếu Lê
21b7cacc33
feat: add cross-region inference entries for Bedrock for Opus 4.8
2026-05-28 12:38:19 -07:00
Aiden Cline
ac1dc14439
Merge pull request #1884 from calebboyd/update-vercel-models-latest
...
feat: add vercel ai gateway opus 4.8
2026-05-28 14:31:03 -05:00
Aiden Cline
63acb799da
Merge pull request #1883 from anomalyco/automation/sync-models-openrouter
...
chore(sync): update OpenRouter model catalog
2026-05-28 14:30:56 -05:00
github-actions[bot]
2e6ebb6750
chore(sync): update OpenRouter model catalog
2026-05-28 19:11:47 +00:00
calebboyd
d8c2e15632
feat: add vercel ai gateway opus 4.8
2026-05-28 13:33:37 -05:00
Frank
8a70160200
update zen model
2026-05-28 14:01:35 -04:00
Aiden Cline
7ea1d4e15c
Merge pull request #1882 from anomalyco/add-opus-4.8
...
feat: add opus 4.8
2026-05-28 12:04:51 -05:00
Aiden Cline
1c71b1bf17
feat: add opus 4.8
2026-05-28 12:04:14 -05:00
Aiden Cline
acc3126194
Merge pull request #1879 from Ardakilic/chore/ci-fork-ensurance
...
CI Scheduled fork upstream ensurance
2026-05-28 11:27:47 -05:00
Aiden Cline
1f67addc01
Merge pull request #1878 from anomalyco/automation/sync-models-openrouter
...
chore(sync): update OpenRouter model catalog
2026-05-28 11:27:07 -05:00
github-actions[bot]
bb011e71a7
chore(sync): update OpenRouter model catalog
2026-05-28 15:28:44 +00:00
bas3line
3df181411f
fix(routing-run): update API base URL
2026-05-28 08:58:33 +05:30
Arda Kilicdagi
a1f588463c
chore: prevent forks to run cronjobbed sync commands
2026-05-28 06:25:45 +04:00
Arda Kilicdagi
cb414c6886
chore: prevent forks to run cronjobbed sync commands
2026-05-28 05:31:19 +04:00
Arda Kilicdagi
4095e819c9
chore: prevent forks to run cronjobbed sync commands
...
chore: prevent forks to run cronjobbed sync commands
2026-05-28 05:26:15 +04:00
Aakash Gupte
9367cc1825
Add Merge Gateway logo
...
Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com >
2026-05-27 17:22:44 -04:00
Aakash Gupte
96b9a04c4b
Add Merge Gateway provider
...
Merge Gateway (https://merge.dev ) is an LLM gateway exposing an
OpenAI/Anthropic-compatible API across many providers, using the
published `merge-gateway-ai-sdk-provider` npm package. Models are
defined via `extends` from existing canonical entries.
Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com >
2026-05-27 14:32:21 -04:00
Jack
d9de217dc4
update zen models
2026-05-28 01:46:38 +08:00
Aiden Cline
64ea80d416
Merge pull request #1869 from anomalyco/fix/xiaomi-token-plan-models
...
fix Xiaomi Token Plan model catalog
2026-05-27 12:02:22 -05:00
Aiden Cline
d3859517d1
Merge pull request #1870 from anomalyco/automation/sync-models-openrouter
...
chore(sync): update OpenRouter model catalog
2026-05-27 12:01:59 -05:00
Aiden Cline
985a144c5f
Merge pull request #1866 from tarikko/patch-1
...
Update Mistral Small model details in TOML file
2026-05-27 11:43:52 -05:00
Jack
5b05070eb4
Merge pull request #1875 from anomalyco/update/opencode-go-mimo-v2-5-pricing
...
update opencode go MiMo V2.5 pricing
2026-05-28 00:39:53 +08:00
Aiden Cline
dad245c5e1
Merge pull request #1871 from oskarkocol/chore/update-groq-cache-prices
...
chore: update groq cache pricing
2026-05-27 11:39:22 -05:00
Aiden Cline
bcff9bd6b2
Merge pull request #1874 from Alex-yang00/codex/sync-novita-models
...
Sync NovitaAI models
2026-05-27 11:39:03 -05:00
Aiden Cline
561bdaa546
Merge pull request #1876 from sebastiand-cerebras/deprecate-cerebras-qwen235b-llama8b-20260527
...
Remove deprecated Cerebras Qwen 3 235B model
2026-05-27 11:38:26 -05:00
Jack
f4f210986c
fix opencode go MiMo pricing conflict resolution
2026-05-28 00:36:48 +08:00
Jack
f4206f4eaa
Merge branch 'dev' into update/opencode-go-mimo-v2-5-pricing
2026-05-28 00:34:02 +08:00
Seb Duerr
52e5b67bc4
Remove deprecated Cerebras Qwen 3 235B model
2026-05-27 09:13:50 -07:00
Jack
d11b0e448c
update opencode go MiMo V2.5 pricing
2026-05-27 23:54:58 +08:00
github-actions[bot]
591745690e
chore(sync): update OpenRouter model catalog
2026-05-27 15:28:06 +00:00
Codex
5a4bc9b383
Add selected NovitaAI models
2026-05-27 21:30:38 +08:00
Tarik
a14171bcd4
Update mistral-small.toml
...
the only difference is the price so I removed the other fields
2026-05-27 14:00:33 +01:00
oskar
0741c5a53c
chore: update kimi cache rate
2026-05-27 13:09:20 +07:00
oskar
7a1ec8737b
chore: update groq cache pricing
2026-05-27 13:05:24 +07:00
Aiden Cline
7c37c92c2d
fix Xiaomi Token Plan model catalog
2026-05-27 00:39:17 -05:00
Aiden Cline
ec4ec6d441
Merge pull request #1867 from anomalyco/automation/sync-models-openrouter
...
chore(sync): update OpenRouter model catalog
2026-05-27 00:00:17 -05:00
Aiden Cline
d96da5ece6
Merge pull request #1868 from Ardakilic/chore/sync-kilo-models-20260527-1
...
Sync Kilo models with upstream gateway
2026-05-26 23:59:26 -05:00
github-actions[bot]
0342c79d03
chore(sync): update OpenRouter model catalog
2026-05-27 03:26:29 +00:00
Arda Kilicdagi
37136fc2f3
chore: sync upstream kilo api gateway models
2026-05-27 02:27:53 +04:00
Tarik
131eca9c5e
Update Mistral Small model configuration
...
use [extends] syntax
2026-05-26 20:44:22 +01:00
Tarik
3e38d46c4c
Update Mistral Small model details in TOML file
...
The details on helicone website are false,
accurate details are pulled from deepinfra website
https://www.helicone.ai/model/mistral-small
https://deepinfra.com/mistralai/Mistral-Small-3.2-24B-Instruct-2506
2026-05-26 19:34:07 +01:00
Aiden Cline
989939773b
Merge pull request #1863 from anomalyco/automation/sync-models-openrouter
...
chore(sync): update OpenRouter model catalog
2026-05-26 13:04:26 -05:00
Frank
4f1d5c511a
update go models
2026-05-26 13:54:19 -04:00
github-actions[bot]
34a21e07f2
chore(sync): update OpenRouter model catalog
2026-05-26 17:26:26 +00:00
Aiden Cline
979f5da0ca
Merge pull request #1864 from oskarkocol/update/openai-cache-rates
...
chore: update openai cache rates
2026-05-26 12:06:15 -05:00
oskar
73b357e2ab
update cache rates
2026-05-26 23:28:11 +07:00
Aiden Cline
97c71f1663
Merge pull request #1862 from anomalyco/automation/sync-models-openrouter
...
chore(sync): update OpenRouter model catalog
2026-05-26 09:11:54 -05:00
github-actions[bot]
59c270bc85
chore(sync): update OpenRouter model catalog
2026-05-26 13:23:20 +00:00
Aiden Cline
a4c18f88ed
Merge pull request #1860 from bas3line/sync-routing-run-models-2
...
Sync routing.run model catalog
2026-05-25 23:44:24 -05:00
bas3line
3343a585ac
fix(routing-run): sync model catalog
2026-05-26 09:51:14 +05:30
Aiden Cline
f23db95550
Merge pull request #1858 from smakosh/fix/llmgateway-qwen3.7-max-id
...
fix(llmgateway): correct Qwen3.7 Max model id to qwen3.7-max
2026-05-25 17:22:56 -05:00
Aiden Cline
556d9a9045
Merge pull request #1859 from Suat-B/update-xpersona-frieren-coder-limits
...
Update Xpersona Frieren Coder limits
2026-05-25 17:22:46 -05:00
Frank
49991c8f8f
update zen models
2026-05-25 17:56:52 -04:00
Suat-B
20c1e810ce
Update Xpersona Frieren Coder limits
2026-05-25 13:23:59 -05:00
smakosh
15d08b4540
fix(llmgateway): correct Qwen3.7 Max model id to qwen3.7-max
...
Rename qwen37-max.toml to qwen3.7-max.toml so the model id matches
the canonical alibaba/qwen3.7-max definition (name: Qwen3.7 Max).
Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com >
2026-05-25 18:48:12 +01:00
Aiden Cline
14bbc303e5
Merge pull request #1852 from anomalyco/automation/sync-models-openrouter
...
chore(sync): update OpenRouter model catalog
2026-05-25 09:49:26 -05:00
Aiden Cline
02da63ed48
Merge pull request #1853 from dpuyosa/chore/venice-pricing
...
Venice: Update pricing and limits for 4 models
2026-05-25 09:49:09 -05:00
github-actions[bot]
07daef8c54
chore(sync): update OpenRouter model catalog
2026-05-25 13:27:22 +00:00
dpuyosa
b6ebe8696c
[venice] Update pricing, limits, and last_updated for 4 models
...
- Reduce input/output/cache prices for gemini-3-5-flash, google-gemma-4-31b-it, and qwen-3-7-max
- Lower max output tokens from 65,536 to 16,384 for qwen3-5-35b-a3b
- Bump last_updated to 2026-05-25 for all 4 models
2026-05-25 10:07:32 +02:00
Aiden Cline
ad654e71da
Merge pull request #1850 from huxeon/dev
...
fix: modify the deepseek v4 flash/pro price
2026-05-25 00:18:18 -05:00
Aiden Cline
508f4d48e1
Merge pull request #1847 from ceyhanmolla/poolside/laguna-direct
...
Add Poolside provider with Laguna M.1 and XS.2 models
2026-05-24 23:26:34 -05:00
opencode-agent[bot]
cd7c70b4fe
revert: remove opencode-go deepseek-v4-pro price changes
...
Keep only the deepseek provider price updates as intended.
2026-05-25 04:26:14 +00:00
Aiden Cline
10e752ea84
Merge pull request #1845 from yukoba/vultr
...
Update Vultr models
2026-05-24 23:26:12 -05:00
huxeon
4cdb4c700b
fix: modify the deepseek v4 flash/pro price in provider deepseek and opencode-go
2026-05-24 13:33:56 +08:00
Aiden Cline
d497a446eb
Merge pull request #1849 from technoabsurdist/add-wafer-ai-qwen3.6-35b-a3b-and-kimi-k2.6
...
providers/wafer.ai: add Qwen3.6-35B-A3B and Kimi-K2.6
2026-05-23 16:41:24 -05:00
Emilio Andere
745cf557b3
feat(wafer.ai): add Qwen3.6-35B-A3B and Kimi-K2.6
...
Both models are public serverless on pass.wafer.ai/v1/models but were
missing from the wafer.ai provider in models.dev, so OpenCode and other
tools that pull from the registry could not discover them.
- Qwen3.6 35B A3B: compact MoE, 32K context, vision-capable, $0.19/M in,
$1.25/M out (NVFP4 on AMD MI355X — see wafer.ai/blog/qwen36-mi355x).
- Kimi K2.6: 1T sparse MoE, 262K context, vision-capable, $1.10/M in,
$4.80/M out (NVFP4 on Blackwell — see wafer.ai/blog/kimi-k26-nvfp4).
Co-authored-by: Cursor <cursoragent@cursor.com >
2026-05-23 13:26:18 -04:00
Yu Kobayashi
db295d6842
Refactor to use the extends syntax
2026-05-24 01:20:31 +09:00
ceyhanmolla
47c0eed41e
Add Poolside provider with Laguna M.1 and XS.2 models
2026-05-23 17:09:33 +02:00
Yu Kobayashi
e690980857
Update Vultr models
2026-05-23 16:25:11 +09:00
Aiden Cline
f5f7d1a167
Merge pull request #1840 from anomalyco/automation/sync-models-openrouter
...
chore(sync): update OpenRouter model catalog
2026-05-22 16:58:23 -05:00
Aiden Cline
cfbbb55c4d
Merge pull request #1839 from anthraxx/alibaba-qwen3.7-plus
...
Alibaba: Add Qwen 3.7 Max and 3.6 Flash to all regions and plans
2026-05-22 16:58:08 -05:00
github-actions[bot]
1956762cdf
chore(sync): update OpenRouter model catalog
2026-05-22 21:41:35 +00:00
Levente Polyak
28571433f7
feat(alibaba): add Qwen3.7 Max model configuration to all regions
...
Link: https://bailian.console.alibabacloud.com/cn-beijing?tab=model#/model-market/detail/qwen3.7-max?serviceSite=asia-pacific-china
2026-05-22 19:00:29 +02:00
Levente Polyak
aef5e48bae
feat(alibaba): add Qwen3.6 Flash model configuration to all regions
...
Link: https://bailian.console.alibabacloud.com/cn-beijing?tab=model#/model-market/detail/qwen3.6-flash?serviceSite=asia-pacific-china
2026-05-22 18:54:22 +02:00
Aiden Cline
8ba19639a7
Merge pull request #1825 from shzdehmd/dev
...
update(fireworks): sync models and pricing with current offerings
2026-05-22 11:45:45 -05:00
Aiden Cline
0f9b4c9edc
Merge pull request #1834 from monotykamary/chore/update-neuralwatt-qwen3.6-pricing
...
fix(neuralwatt): update Qwen3.6 pricing to match API
2026-05-22 09:13:02 -05:00
Aiden Cline
2a9b6256dc
Merge pull request #1836 from PierreLeGuen/nearai-provider
...
Add current NEAR AI Cloud models
2026-05-22 09:12:36 -05:00
Aiden Cline
51633fe106
Merge pull request #1838 from Quentinchampenois/fix/update-models-scaleway
...
fix: update Scaleway provider models list
2026-05-22 09:12:26 -05:00
Aiden Cline
169b5c4331
Merge pull request #1833 from NicoAvanzDev/add-copilot-gemini-3-5-flash
...
[GitHub Copilot] add Gemini 3.5 Flash
2026-05-22 09:09:25 -05:00
Aiden Cline
9d0f0d6d56
Merge pull request #1837 from fydrah/fix/google-vertex-gemini-3.5-flash
...
fix: missing google-vertex gemini 3.5 flash extend
2026-05-22 09:07:34 -05:00
Aiden Cline
1d2af0c97b
Merge pull request #1835 from dpuyosa/feat/venice-models
...
Venice: Add Gemini 3.5 Flash and Qwen 3.7 Max, update Grok Build 0.1 pricing
2026-05-22 09:07:03 -05:00
Aiden Cline
cc14ae4370
Merge pull request #1832 from anomalyco/automation/sync-models-openrouter
...
chore(sync): update OpenRouter model catalog
2026-05-22 09:06:16 -05:00
github-actions[bot]
ad9a3448d2
chore(sync): update OpenRouter model catalog
2026-05-22 14:03:32 +00:00
Quentin Champenois
c6f9cf58e4
fix(scaleway): add new models gemma-4-26b-a4b-it and qwen3.6-35b-a3b
2026-05-22 15:43:11 +02:00
Quentin Champenois
bba5809471
fix(scaleway): Extends existing models
2026-05-22 15:42:36 +02:00
Quentin Champenois
c0852f1b4f
fix(scaleway): Clear removed models from list
2026-05-22 15:11:44 +02:00
Flavien Hardy
0fc3a3f635
fix: missing google-vertex gemini 3.5 flash extend
2026-05-22 08:56:44 -04:00
Pierre LE GUEN
68936dc062
Add current NEAR AI Cloud models
2026-05-22 09:27:57 +00:00
dpuyosa
ade9760060
[venice] Add Gemini 3.5 Flash and Qwen 3.7 Max, update Grok Build 0.1 pricing
...
- Add Gemini 3.5 Flash model (1M context, multimodal input)
- Add Qwen 3.7 Max model (1M context, text-only)
- Update Grok Build 0.1 cost tiers and pricing
2026-05-22 10:57:38 +02:00
Tom X Nguyen
2a8b90b197
fix(neuralwatt): update Qwen3.6 pricing to match API
...
Updates the per-token cost for Qwen3.6-35B-A3B and qwen3.6-35b-fast from /bin/bash.05//bin/bash.10 to /bin/bash.29/.15 (input/output per million tokens), matching the actual Neuralwatt API pricing as reflected in pi-neuralwatt-provider commit f634286.
2026-05-22 15:19:51 +07:00
NicoAvanzDev
8569f0dfef
[GitHub Copilot] add Gemini 3.5 Flash
2026-05-22 07:57:10 +00:00
Aiden Cline
fc98ceb72e
fix sync workflow force lease
2026-05-21 23:52:34 -05:00
Aiden Cline
be7c5afc94
Merge pull request #1831 from anomalyco/fix-vertex-sonnet-4-6-limits
...
Fix Vertex Sonnet 4.6 token limits
2026-05-21 23:38:18 -05:00
Aiden Cline
57e62c43b0
fix vertex sonnet 4.6 limits
2026-05-21 23:37:30 -05:00
Aiden Cline
0898c35c9f
Merge pull request #1830 from zainhas/dev
...
[Together AI] add Qwen3.7 max
2026-05-21 21:00:51 -05:00
Zain Hasan
46b23fb313
Merge branch 'anomalyco:dev' into dev
2026-05-21 17:47:35 -07:00
Zain Hasan
05fedc76cc
[Together AI] add Qwen3.7
2026-05-21 17:47:19 -07:00
Aiden Cline
2738f81d1a
Merge pull request #1828 from anomalyco/refactor/sync-core-layout
...
refactor: move sync implementation into core src
2026-05-21 18:10:09 -05:00
Aiden Cline
89b834086a
refactor: move sync implementation into core src
2026-05-21 18:06:25 -05:00
Aiden Cline
1ab2ff8163
Merge pull request #1826 from smakosh/add-llmgateway-models
...
feat: add new LLM Gateway text models
2026-05-21 17:58:29 -05:00
Frank
9468676683
update zen models
2026-05-21 18:42:36 -04:00
Claude
6cdd2f054b
Merge upstream/dev into add-llmgateway-models; resolve gemini-3.5-flash conflict
...
# Conflicts:
# providers/google/models/gemini-3.5-flash.toml
2026-05-21 21:48:48 +00:00
Aiden Cline
b13abc9141
Merge pull request #1827 from anomalyco/update-xai-pricing
...
fix xAI long-context pricing
2026-05-21 16:45:43 -05:00
Aiden Cline
e5ba264751
fix xAI long-context pricing
2026-05-21 16:41:36 -05:00
smakosh
a7811fb522
refactor: use extends for gemini and qwen models
...
Add canonical google/gemini-3.5-flash and alibaba/qwen3.7-max defs and
have the llmgateway entries extend them, per PR review.
Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com >
2026-05-21 23:38:57 +02:00
smakosh
605fae75d9
feat: add new LLM Gateway text models
...
Add Grok 4.20 (reasoning/non-reasoning), Gemini 3.5 Flash, and Qwen3.7 Max to the llmgateway provider.
Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com >
2026-05-21 23:04:47 +02:00
Ahmad Shahzad
bcab0885bc
update(fireworks): sync models and pricing with current offerings
...
Removed — 11 deprecated models:
- deepseek-v3p1
- deepseek-v3p2
- glm-4p5
- glm-4p5-air
- glm-4p7
- glm-5
- kimi-k2-instruct
- kimi-k2-thinking
- minimax-m2p1
- routers/kimi-k2p5-turbo
Added — 2 new Turbo (routers) models:
- routers/glm-5p1-fast
- routers/kimi-k2p6-turbo
Modified — pricing fixes:
- deepseek-v4-pro: cache_read 0.15 → 0.145
- gpt-oss-120b: added cache_read = 0.015
- gpt-oss-20b: input 0.05 → 0.07, output 0.20 → 0.30, added cache_read = 0.035
- minimax-m2p7: cache_read 0.03 → 0.06
2026-05-22 01:55:31 +05:00
Aiden Cline
26b05268ae
Merge pull request #1824 from anomalyco/fix/vercel-gemini-35-flash
...
Add new Vercel AI Gateway models
2026-05-21 15:30:35 -05:00
Aiden Cline
1aee13d2e5
Add new Vercel AI Gateway models
2026-05-21 13:21:25 -05:00
Aiden Cline
0a924e6bb2
Merge pull request #1822 from anomalyco/automation/sync-models-xai
...
chore(sync): update xAI model catalog
2026-05-21 13:05:20 -05:00
Aiden Cline
9769b2b11d
Merge pull request #1823 from anomalyco/automation/sync-models-openrouter
...
chore(sync): update OpenRouter model catalog
2026-05-21 13:05:13 -05:00
github-actions[bot]
1b0db099cf
chore(sync): update OpenRouter model catalog
2026-05-21 17:56:13 +00:00
github-actions[bot]
16ed78587c
chore(sync): update xAI model catalog
2026-05-21 17:56:11 +00:00
Frank
0b88965165
update zen models
2026-05-21 13:41:55 -04:00
Aiden Cline
acc704ce39
Merge pull request #1797 from arnavchachra/add-crof-provider
...
add crof.ai provider with 21 models
2026-05-21 11:43:02 -05:00
Aiden Cline
51ad3b264e
Merge pull request #1821 from anomalyco/sync-provider-ci
...
chore: automate provider sync jobs
2026-05-21 11:23:49 -05:00
Aiden Cline
146b6c7084
Merge pull request #1819 from Inceptron-Software/add_inceptron_provider
...
Add Inceptron provider
2026-05-21 11:21:27 -05:00
Aiden Cline
0e3cbe3c64
chore: automate provider sync jobs
2026-05-21 11:21:12 -05:00
Aiden Cline
604d4d66a4
Merge pull request #1820 from Suat-B/codex/xpersona-image-input-20260521
...
Add image input modality to Xpersona model
2026-05-21 11:00:49 -05:00
SuatB
f5090028b8
Add image input modality to Xpersona model
2026-05-21 09:46:28 -05:00
Frank
4bad8faf29
update zen models
2026-05-21 09:05:13 -04:00
Oskar Gustafsson
0df2ccf586
Add Inceptron provider
2026-05-21 09:24:43 +02:00
Aiden Cline
bafdc00b45
Merge pull request #1812 from anomalyco/openrouter-extends-sync
...
Sync OpenRouter models with extends
2026-05-20 21:07:21 -05:00
Aiden Cline
49840c013b
Merge pull request #1814 from neonn0d/feat/stepfun-ai
...
feat(stepfun-ai): add international StepFun platform
2026-05-20 20:58:37 -05:00
Aiden Cline
eccae0b54e
sync openrouter models with extends
2026-05-20 20:32:11 -05:00
Aiden Cline
4cca29405f
Merge pull request #1817 from dpuyosa/dev
...
Venice: Remove Grok 4.1 Fast and add Grok Build 0.1
2026-05-20 20:26:45 -05:00
Aiden Cline
e40d9dd338
Merge pull request #1818 from anomalyco/cloudflare-sync-env
...
chore(sync): isolate cloudflare credentials
2026-05-20 20:26:22 -05:00
Aiden Cline
6a74991397
chore(sync): isolate cloudflare credentials
2026-05-20 20:19:47 -05:00
dpuyosa
035999cb58
[venice] Replace Grok 4.1 Fast with Grok Build 0.1
...
- Remove deprecated grok-41-fast model entry
- Add grok-build-0-1 with 200K token tiered pricing
- Update context to 256K and output limit to 65,536
2026-05-21 02:36:19 +02:00
Frank
cec56bf1bc
update zen models
2026-05-20 19:43:25 -04:00
Aiden Cline
85f0cdcb2f
Merge pull request #1816 from anomalyco/xai-sync
...
Add PDF input modality to Grok models
2026-05-20 18:12:47 -05:00
Aiden Cline
ef80d4df4e
Infer PDF modality for xAI image models
2026-05-20 18:12:11 -05:00
Aiden Cline
af0ef00109
Update xAI Grok PDF modalities
2026-05-20 18:06:33 -05:00
Aiden Cline
5fdcea6b36
Merge pull request #1815 from anomalyco/cloudflare-ai-gateway
...
chore(sync): add cloudflare workers ai sync
2026-05-20 18:02:28 -05:00
Aiden Cline
31e56480b4
chore(sync): add cloudflare workers ai sync
2026-05-20 16:55:29 -05:00
Aiden Cline
92a621594e
Merge pull request #1813 from anomalyco/sync-xai
...
chore(sync): add xai model sync
2026-05-20 16:02:38 -05:00
Aiden Cline
d2db353ceb
chore: ignore sync reports
2026-05-20 16:01:51 -05:00
Aiden Cline
900ae509d2
Merge pull request #1808 from ajussak/scaleway
...
Added Mistral Medium 3.5 128B from Scaleway
2026-05-20 15:58:00 -05:00
neo
9d60164243
feat(stepfun-ai): add international StepFun platform
...
StepFun runs two separate platforms with distinct accounts/keys:
platform.stepfun.com (China, already covered by providers/stepfun) and
platform.stepfun.ai (international). Keys are not interchangeable
across the two — .ai keys are rejected by api.stepfun.com as
invalid_api_key.
Stepfun's own opencode integration guide instructs users to point at
https://api.stepfun.ai/step_plan/v1 . This adds providers/stepfun-ai
for that endpoint, symlinking the shared chat models. Follows the
moonshotai / moonshotai-cn pattern.
2026-05-20 20:43:47 +02:00
Adrien Jussak
cd3e99025f
Update Mistral Medium 3.5 128B model configuration to extend from mistral-medium-2604 and adjust context window size.
2026-05-20 20:43:45 +02:00
Aiden Cline
1098981eb6
chore(sync): add xai model sync
2026-05-20 13:23:02 -05:00
Aiden Cline
27a151cf53
Merge pull request #1811 from anomalyco/xai-grok-build-model
...
Add xAI Grok Build model
2026-05-20 13:05:06 -05:00
Aiden Cline
41ff42ab7f
Merge pull request #1810 from fhennerkes/dev
...
poe: add Gemini-3.5-Flash model
2026-05-20 13:04:47 -05:00
Aiden Cline
adf1cbdecd
add xai grok build model
2026-05-20 13:04:13 -05:00
Frank
10ddc78ce0
update zen models
2026-05-20 14:02:01 -04:00
fhennerkes
7ae897e440
poe: add Gemini-3.5-Flash model
...
Add new Google model from Poe API (released 2026-05-19).
Uses extends format inheriting from google/gemini-3.5-flash with
Poe-specific overrides (name format, no temperature, markup pricing,
limited input modalities).
Co-Authored-By: Claude Opus 4.7 <noreply@anthropic.com >
2026-05-20 10:52:15 -07:00
Aiden Cline
02e452c2e8
Merge pull request #1805 from anomalyco/sync-google
...
sync google models
2026-05-20 10:53:57 -05:00
Aiden Cline
f05b63fff5
Merge pull request #1807 from anomalyco/automation/sync-models-aggregators
...
chore(sync): update aggregator model catalogs
2026-05-20 10:34:30 -05:00
Adrien Jussak
e277d60236
Add Mistral Medium 3.5 128B to Scaleway
2026-05-20 16:14:37 +02:00
github-actions[bot]
b01277a737
chore(sync): update aggregator model catalogs
2026-05-20 09:25:37 +00:00
Frank
a5da5aa429
update zen models
2026-05-20 04:14:28 -04:00
Aiden Cline
5ceac8a58b
updates
2026-05-19 23:43:01 -05:00
Aiden Cline
11e1d5623a
Merge pull request #1806 from Cahl-Dee/grid-model-updates-2026-05
...
Grid model updates 2026-05
2026-05-19 23:42:26 -05:00
Aiden Cline
f107afc57c
sync
2026-05-19 22:56:50 -05:00
Carl DiClementi
e789d7c1f2
Merge branch 'anomalyco:dev' into grid-model-updates-2026-05
2026-05-19 16:06:23 -05:00
Cahl-Dee
899668ad49
added new code and agent models and updated existing text models
2026-05-19 16:04:58 -05:00
Aiden Cline
8c677f0134
sync google models
2026-05-19 15:58:06 -05:00
Aiden Cline
462c7877d9
add gemini 3.5 flash
2026-05-19 15:43:38 -05:00
Aiden Cline
55871f9dca
Merge pull request #1804 from Adanlink/dev
...
feat: add deepseek-v4-flash to the fireworks-ai provider
2026-05-19 15:29:33 -05:00
Aiden Cline
4f7194a3c8
test
2026-05-19 15:09:11 -05:00
Aiden Cline
f65f0148da
add sync guide
2026-05-19 15:08:44 -05:00
Adán
14952f8855
Rename deepseek-v4-flash to deepseek-v4-flash.toml
2026-05-19 18:56:43 +02:00
Adán
d7c6d3ad12
Add deepseek-v4-flash model configuration
2026-05-19 18:53:48 +02:00
Aiden Cline
356bc79d08
Merge pull request #1637 from elvexai/fix/amazon-bedrock-kimi-token-limits
...
fix: Token limits for Amazon Bedrock Kimi K2 models
2026-05-19 09:42:30 -05:00
Aiden Cline
a89b1ed726
Merge pull request #1801 from bas3line/sync-routing-run-models
...
Sync routing.run model catalog
2026-05-19 09:41:38 -05:00
bas3line
a998576773
fix(routing-run): match live model metadata
2026-05-19 10:38:58 +05:30
bas3line
fbe842bbea
fix(routing-run): expose reasoning metadata
2026-05-19 08:39:51 +05:30
bas3line
6c0c3d1b10
fix(routing-run): sync model catalog
2026-05-19 07:37:53 +05:30
Aiden Cline
db0a7cf611
Merge pull request #1798 from anomalyco/rework-sync-logic
...
sync: centralize aggregator model updates
2026-05-18 20:12:50 -05:00
Aiden Cline
d775e37e3b
Merge pull request #1800 from jerome-benoit/feat/sap-ai-core-gpt-5.4
...
feat(sap-ai-core): add GPT-5.4
2026-05-18 20:12:21 -05:00
Jérôme Benoit
36753063d9
feat(sap-ai-core): add GPT-5.4
...
Add gpt-5.4 with availability date from official SAP source.
Drop [[cost.tiers]] from gemini-2.5-pro pending SAP-side tiered
pricing confirmation; sap-ai-core now declares no per-model tiers
(SAP Note 3437766 is login-gated and authoritative for capacity
unit conversion rates).
2026-05-19 02:58:29 +02:00
Aiden Cline
5ee955297a
sync: drop vercel catalog updates
2026-05-18 19:07:14 -05:00
Aiden Cline
1b77511903
Merge pull request #1799 from vglafirov/add-gitlab-gpt-5-5
...
feat(gitlab): add Agentic Chat (GPT-5.5) model
2026-05-18 15:24:18 -05:00
Aiden Cline
8896ead7bf
sync: fix vercel pricing tiers
2026-05-18 14:52:40 -05:00
Vladimir Glafirov
eb96594d47
feat(gitlab): add Agentic Chat (GPT-5.5) model
...
Adds duo-chat-gpt-5-5 to the GitLab provider. The GitLab AI Gateway
proxies this model to OpenAI's gpt-5.5-2026-04-23 backend with a
1.05M token context window (922k input + 128k output).
Source: gitlab-org/modelops/applied-ml/code-suggestions/ai-assist
models.yml (gpt_5_5 entry with proxy_provider: openai).
The gitlab-ai-provider npm package exposes this model id starting in
v6.7.0.
2026-05-18 20:44:34 +02:00
Aiden Cline
327332efe3
Merge pull request #1794 from bas3line/add-routing-run-provider
...
Add routing.run provider
2026-05-18 12:32:34 -05:00
Aiden Cline
5020951745
sync: refresh openrouter after dev merge
2026-05-18 12:30:29 -05:00
Aiden Cline
cb6f97774e
Merge remote-tracking branch 'origin/dev' into rework-sync-logic
2026-05-18 12:29:28 -05:00
Aiden Cline
7f8b493b0c
Merge pull request #1795 from delafthi/delafthi/lxxqxzktnozv
...
fix(providers/novita-ai): use lowercase model names
2026-05-18 12:28:32 -05:00
Aiden Cline
d65a862533
sync: centralize aggregator model updates
2026-05-18 12:12:15 -05:00
arnavchachra
9420048dfe
fix crof model limits and reasoning flag to match Crof API
2026-05-18 21:51:19 +05:30
arnavchachra
8db6c27634
add crof provider with 21 models
2026-05-18 21:40:11 +05:30
Victor Navarro
8e710e19ea
bring back old bick-pickle
...
Added interleaved section with reasoning_content field and removed provider section.
2026-05-18 11:44:14 +02:00
Frank
36c6896e97
update zen models
2026-05-17 22:58:06 -04:00
Aiden Cline
a8be548a5d
Merge pull request #1416 from Luew2/add-lilac-provider
...
Add Lilac provider
2026-05-17 19:23:07 -05:00
Luew2
8c2fae4ab0
Keep exact Lilac Gemma model name
2026-05-17 17:19:53 -07:00
Luew2
5e7fad350d
Align Lilac Gemma display name
2026-05-17 17:19:04 -07:00
Luew2
feb85ef2c9
Align Lilac provider with registry conventions
2026-05-17 17:15:40 -07:00
Luew2
d4161ebf24
Follow models.dev conventions for Lilac provider
2026-05-17 17:11:03 -07:00
Luew2
b91ab02e2b
Add Lilac MiniMax M2.7 model
2026-05-17 17:07:44 -07:00
Luew2
dd09d07f75
Update Lilac Kimi model to K2.6
2026-05-17 17:07:44 -07:00
Luew2
4515f85d47
Add Lilac cache pricing
2026-05-17 17:07:44 -07:00
Luew2
97572240e1
Add Gemma 4 31B IT model
...
Adds google/gemma-4-31b-it to the Lilac provider ($0.11/M input,
$0.35/M output, 262K context, native multimodal with image/video).
Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com >
2026-05-17 17:07:44 -07:00
Luew2
80cc04e9e5
Fix Kimi K2.5 output limit to 262,144 tokens
2026-05-17 17:07:44 -07:00
Luew2
d532ebb89b
Use purple gradient for Lilac logo (brand colors #6451dc → #b6a6f9)
2026-05-17 17:07:44 -07:00
Luew2
9bae3e887f
Replace placeholder logo with Lilac icon mark (currentColor)
2026-05-17 17:07:44 -07:00
Luew2
0be69bf872
Add Lilac provider
...
Add Lilac as an OpenAI-compatible provider serving:
- z-ai/glm-5.1: Z.ai's flagship agentic model (754B MoE, 202.8K context)
- moonshotai/kimi-k2.5: Moonshot AI's multimodal reasoning model (1T MoE, 262K context)
API: https://api.getlilac.com/v1
Docs: https://docs.getlilac.com
2026-05-17 17:07:44 -07:00
Thierry Delafontaine
0ed38cbecf
fix(providers/novita-ai): use lowercase model names
...
Mixed case naming causes conflicts on case-insensitive filesystems like macOS.
2026-05-17 21:17:46 +02:00
bas3line
e65703382a
feat: add routing.run provider
2026-05-17 20:27:56 +05:30
Aiden Cline
748754c99b
Merge pull request #1792 from monotykamary/fix/neuralwatt-context-limits
...
fix(neuralwatt): sync context window and output limits with upstream API
2026-05-16 13:14:05 -05:00
Tom X Nguyen
ec9c12d0fc
fix(neuralwatt): sync context window and output limits with upstream API
...
Updates all 14 neuralwatt model TOML files with corrected context window
and max output token values as reported by the Neuralwatt API:
- Devstral-Small-2-24B-Instruct-2512: 262,144 -> 262,128
- GLM-5/GLM-5.1 variants: 200,000 -> 202,736
- GPT-OSS-20B: 16,384 -> 16,368
- Kimi-K2.5/K2.6 variants: 262,144 -> 262,128
- MiniMax-M2.5: 196,608 -> 196,592
- Qwen3.5-397B variants: 262,144 -> 262,128
- Qwen3.6-35B variants: 131,072 -> 131,056
Also fixes the README: moves kimi-k2.6-fast from 'Reasoning Models' to
'Fast Variants' and removes incorrect claim that fast variants support
reasoning.
2026-05-16 23:16:34 +07:00
Aiden Cline
ac81822c89
Merge pull request #1777 from berget-ai/feat/berget-kimi-k2.6
...
feat: add Kimi K2.6 to berget.ai
2026-05-16 06:09:00 -05:00
Aiden Cline
d32ed764bd
Merge pull request #1791 from anomalyco/automation/sync-openrouter-models
...
Sync OpenRouter models
2026-05-16 06:08:09 -05:00
Christian Landgren
45fb951c42
feat: add Kimi K2.6 to berget.ai
...
Add Moonshot AI Kimi K2.6 model to berget.ai provider catalog.
- 262K context window
- 16K output tokens
- Text input/output
- Supports: reasoning, structured output, tool calling
- Pricing: /bin/zsh.83/M input, .85/M output (EUR-based)
- Open weights
2026-05-16 12:46:35 +02:00
github-actions[bot]
260d79b2d5
Sync OpenRouter models
2026-05-16 08:54:41 +00:00
Aiden Cline
746b9caf79
Merge pull request #1785 from jerome-benoit/feat/sap-ai-core-opus-4-7
...
feat(sap-ai-core): add Claude Opus 4.7 and sync model specs
2026-05-15 23:24:47 -05:00
Aiden Cline
8362b55503
Merge pull request #1786 from Ardakilic/chore/kilo-sync-20260516
...
providers(kilo): sync upstream
2026-05-15 23:24:34 -05:00
Aiden Cline
dde3953a9f
Merge pull request #1787 from Suat-B/codex/xpersona-www-api-url
...
Fix Xpersona API base URL
2026-05-15 23:24:06 -05:00
Aiden Cline
0a1695212c
Merge pull request #1788 from Jaaneek/xai-may-15-2026-retirement
...
xai: drop models retired May 15, 2026 + add Grok Imagine models
2026-05-15 23:23:52 -05:00
Jaaneek
89fbb6bb69
xai: drop models retired May 15, 2026 + add Grok Imagine models
2026-05-16 01:57:15 +01:00
SuatB
005fe0fb5a
Fix Xpersona provider API URL
2026-05-15 18:37:41 -05:00
Frank
e283875ce7
update zen models
2026-05-15 17:24:53 -04:00
Arda Kilicdagi
3598019251
providers(kilo): sync upstream
2026-05-16 01:11:25 +04:00
Jérôme Benoit
e9ad8b0a3f
feat(sap-ai-core): add Claude Opus 4.7 and sync model specs
2026-05-15 22:16:07 +02:00
Aiden Cline
0ee78eeda5
sync: openrouter models
2026-05-15 10:20:36 -05:00
Aiden Cline
0a5b33e518
Merge pull request #1778 from zhenjunchen-png/add-orcarouter
...
feat: add OrcaRouter as a new provider
2026-05-15 10:16:29 -05:00
Aiden Cline
22416dda64
Merge pull request #1783 from anomalyco/sync-openrouter
...
add sync script for openrouter, sync openrouter models
2026-05-15 10:12:02 -05:00
Aiden Cline
eba2702e3a
fix: families
2026-05-15 10:03:39 -05:00
Aiden Cline
c2c5cc8f21
add sync script for openrouter, sync openrouter models
2026-05-15 10:00:32 -05:00
zhenjun.chen
022b1b9946
feat(orcarouter): expand to 80 chat models and add brand logo
...
Adds 55 additional upstream-mirrored models alongside the existing 25,
covering the full OrcaRouter chat catalog as exposed by
https://www.orcarouter.ai/api/pricing (text-only chat — TTS, embeddings,
video, and image generation are filtered out).
Per-namespace upstream mappings used by [extends]:
OrcaRouter ns -> models.dev provider
---------------- + ---------------
openai -> openai
anthropic -> anthropic (dot version -> dash, e.g. opus-4.7 -> opus-4-7)
google -> google
deepseek -> deepseek
qwen -> alibaba
grok -> xai
kimi -> moonshotai
minimax -> minimax (minimax-m2.7 -> MiniMax-M2.7)
z-ai -> zai
OrcaRouter-specific aliases (dated snapshots like gpt-5-2025-08-07,
search-preview variants, qwen3-vl-* visual variants) are excluded from
v1 because their upstream canonical files do not yet exist in models.dev.
Also adds providers/orcarouter/logo.svg.
2026-05-15 14:48:31 +08:00
Aiden Cline
8269e04222
Merge pull request #1782 from Suat-B/codex/xpersona-limits-logo
...
Update Xpersona limits, cutoff, and logo
2026-05-15 00:29:32 -05:00
Aiden Cline
a75cf2ed1c
Merge pull request #1552 from aredridel/as/add-umans
...
feat(models): add umans.ai coding plan
2026-05-14 22:40:39 -05:00
Aria Stewart
7b00aafa79
feat(models): umans.ai definitions
2026-05-14 23:39:30 -04:00
Suat-B
e54e0fc8b5
Update Xpersona limits, cutoff, and logo
2026-05-15 03:22:16 +00:00
Aiden Cline
38611e75fa
Merge pull request #1781 from dpuyosa/feat/add-venice-claude-opus-4-7-fast-model
...
Venice: Add Claude Opus 4.7 Fast model
2026-05-14 22:08:14 -05:00
dpuyosa
e62c1e973e
[venice] Add Claude Opus 4.7 Fast model
...
- New pricing with 36/180 input/output per million tokens
- 1M context window with 128K output limit
- Text+image input, text-only output
2026-05-15 01:32:24 +02:00
Aiden Cline
c2b3c601e4
Merge pull request #1724 from isaachuangGMICLOUD/feat/add-gmicloud-provider
...
providers(gmicloud): add GMI Cloud provider
2026-05-14 17:26:50 -05:00
Aiden Cline
9ff1d36a21
Merge pull request #1476 from Vect0rM/feat/add-atomic-chat-provider
...
feat: add Atomic Chat provider
2026-05-14 17:18:21 -05:00
Aiden Cline
e0f4042ad1
Merge pull request #1776 from kapelame/docs/minimax-token-plan-rename
...
providers(minimax): rename Coding Plan → Token Plan in display labels
2026-05-14 17:16:28 -05:00
Frank
14736ba4b6
update zen models
2026-05-14 17:15:53 -04:00
Aiden Cline
9351d68731
Merge pull request #1772 from nearai/add-nearai
...
Add NEAR AI Cloud provider
2026-05-14 10:30:28 -05:00
Aiden Cline
7a5a1d2aff
Merge pull request #1779 from NameIsHiki/siliconflow-deepseek-v4
...
feat(siliconflow): add DeepSeek v4 models
2026-05-14 10:29:52 -05:00
zhenjun.chen
699284ce91
chore(orcarouter): drop oversize logo, fall back to models.dev default
...
The previously committed logo is ~100KB; existing wrapper-provider logos
(openrouter, llmgateway, kilo, aihubmix, ambient) are all 0.3-6KB and use
`currentColor`. Falling back to the default logo per README:
> If we don't have a provider's logo, a default logo is served instead.
A properly-sized currentColor logo will follow in a separate PR.
2026-05-14 21:37:48 +08:00
Hiki
21ce5c3ac8
Create deepseek-v4-flash.toml
2026-05-14 15:26:37 +02:00
Hiki
3485cf52d0
Create deepseek-v4-pro.toml
2026-05-14 15:22:53 +02:00
Frank
99e8f25c78
update zen models
2026-05-14 08:59:42 -04:00
zhenjun.chen
7102978cb4
feat: add OrcaRouter provider
...
OrcaRouter is an OpenAI-compatible meta-router aggregating 150+ LLMs
(OpenAI, Anthropic, Google, xAI, DeepSeek, Qwen, Kimi, MiniMax, ...)
behind a single API key, with a virtual orcarouter/auto smart-routing
entry that picks an upstream per request.
This initial scope covers 26 models (1 AUTO router + 25 upstream mirrors
using [extends]). Pricing computed from https://www.orcarouter.ai/api/pricing
on 2026-05-14: input = model_ratio * $2, output = model_ratio *
completion_ratio * $2 (USD per 1M tokens).
Disclosure: I'm an engineer on the OrcaRouter team.
2026-05-14 20:55:47 +08:00
kapelame
530f60c69c
providers(minimax): rename Coding Plan → Token Plan in display labels
...
The product was renamed from "Coding Plan" to "Token Plan" when its
scope expanded beyond coding to cover all MiniMax modalities (text,
speech, video, music, image). Per
https://platform.minimax.io/docs/token-plan/intro :
"Token Plan extends upon our former Coding Plan."
Updates display name and doc URL for the two affected provider
catalog entries. Provider IDs (minimax-coding-plan,
minimax-cn-coding-plan) are intentionally unchanged for backward
compatibility — anyone with these IDs in opencode.json or
elsewhere keeps working. Old /coding-plan/* URLs still 307-redirect
to the new /token-plan/* paths upstream.
Region disambiguation stays as the URL in parens (matching the
existing minimax / minimax-cn naming convention) — no "China" word
added, since the URL already conveys the region cleanly in the
provider picker.
2026-05-14 19:18:03 +08:00
Misha Skvortsov
2415c5be21
fix(atomic-chat): update logo.svg with new design
...
Replaces the existing logo.svg file with an updated design for the Atomic Chat provider. This change enhances the visual branding of the application.
2026-05-14 10:58:36 +03:00
Aiden Cline
85aba468cd
Merge pull request #1767 from Suat-B/codex/xpersona-provider-20260513
...
Add Xpersona provider
2026-05-13 23:20:20 -05:00
Aiden Cline
d1ec1ba777
Merge pull request #1773 from ambient-gregory/dev
...
feat: add Ambient provider with GLM-5.1 and Kimi K2.6
2026-05-13 19:20:20 -05:00
Gregory
0f94bf16ec
fix(ambient): shrink logo display size to match other providers
2026-05-13 19:33:26 -04:00
Aiden Cline
506e8f48a9
Merge pull request #1770 from EriDeLee/dev
...
chore(aihubmix): sync model catalog
2026-05-13 17:43:30 -05:00
Aiden Cline
3480bc5992
Merge pull request #1775 from michaelnchin/fix/amazon-bedrock-gpt-oss-tokens
...
fix: Output tokens for Bedrock GPT-OSS models
2026-05-13 17:36:42 -05:00
Michael Chin
fde97814ef
fix: Output tokens for Bedrock GPT-OSS models
2026-05-13 14:45:17 -07:00
Gregory
ff7eddcb70
feat: add Ambient provider with GLM-5.1 and Kimi K2.6
...
Adds the Ambient inference provider (api.ambient.xyz) with an initial
catalog of GLM-5.1 and Kimi K2.6, plus a generator script that pulls
from /v1/models so pricing and limits stay in sync with the upstream API.
Run `bun run ambient:generate` to refresh model TOMLs.
2026-05-13 11:30:45 -04:00
Evrard-Nil Daillet
5cbab85b8d
Add nearai logo.svg from cloud.near.ai favicon
2026-05-13 16:32:08 +02:00
Evrard-Nil Daillet
6f9820de9f
Add NEAR AI Cloud provider
...
Adds nearai as an OpenAI-compatible provider at https://cloud-api.near.ai/v1
serving 33 models. First-party mirrors (anthropic/openai/google) use `extends`;
NEAR-hosted open-weight models (Qwen, GLM-5.1-FP8, gpt-oss, whisper, FLUX) have
full definitions.
Pricing and context limits sourced from cloud-api.near.ai/v1/models.
2026-05-13 16:32:08 +02:00
EriDeLee
cdfb429098
chore(aihubmix): sync model catalog
2026-05-13 21:55:32 +08:00
Victor Navarro
1c2546af8a
perf: virtualize models table and other improvements
2026-05-13 12:29:27 +02:00
vimtor
2288a1626b
trim search index to essential fields and remove dead code
2026-05-13 12:21:52 +02:00
vimtor
b3ecfc3d70
improve row scanning
2026-05-13 11:20:42 +02:00
Suat-B
30b3e677fd
Add Xpersona provider
2026-05-13 00:39:09 -05:00
Suat-B
3b37eee86e
Add Xpersona provider
2026-05-13 00:39:08 -05:00
Suat-B
71f069670e
Add Xpersona provider
2026-05-13 00:39:07 -05:00
Aiden Cline
f401672689
Merge pull request #1766 from michaelnchin/fix/amazon-bedrock-structured-output-05122026-2
...
fix: add structured_output=True for more supported Bedrock models
2026-05-12 23:24:39 -05:00
Michael Chin
2a0d86a034
update structured_output for more Bedrock models
2026-05-12 20:52:36 -07:00
Aiden Cline
d9439cdf2f
Merge pull request #1762 from zxyaction/feat/add-auriko-provider
...
feat: add Auriko provider with 15 models
2026-05-12 22:16:02 -05:00
Aiden Cline
3d443d568d
Merge pull request #1765 from michaelnchin/fix/amazon-bedrock-structured-output-05122026
...
fix: update structured_output for Bedrock Claude 4.x models
2026-05-12 22:15:51 -05:00
Michael Chin
a76c8fe9dd
fix: update structured_output for Bedrock Claude 4.x models
2026-05-12 20:07:12 -07:00
Aiden Cline
d08e8d6cc1
Merge pull request #1763 from Tavernari/feat/add-claudinio-provider
...
feat: add Claudinio provider
2026-05-12 19:13:54 -05:00
Victor Carvalho Tavernari
4c06e44047
fix: use currentColor in logo SVG per contributing guidelines
2026-05-12 23:59:03 +01:00
Aiden Cline
5e344ded49
Merge pull request #1755 from anomalyco/correct-context-tracking
...
feat: add new context pricing tiers
2026-05-12 17:40:48 -05:00
Aiden Cline
458b7f4d1a
use Venice context tier thresholds
2026-05-12 17:39:40 -05:00
Aiden Cline
baf4432140
Merge pull request #1759 from NameIsHiki/deepinfra-xiaomi-mimo-models
...
feat(deepinfra): add Xiaomi MiMo v2.5 and v2.5 Pro
2026-05-12 17:10:59 -05:00
Aiden Cline
bbf72ea4e0
Merge pull request #1764 from anomalyco/add-anthropic-opus-4-7-fast-mode
...
Add fast mode for Anthropic Opus 4.7
2026-05-12 17:10:49 -05:00
Aiden Cline
8f9adc7567
fix generated tier change detection
2026-05-12 17:01:35 -05:00
Aiden Cline
addaf1c036
add fast mode for anthropic opus 4.7
2026-05-12 17:01:05 -05:00
Victor Carvalho Tavernari
55d16a58b6
feat: add claudinio provider (OpenAI-compatible, 256K ctx, $0.50/$2.00 per MTok)
2026-05-12 22:33:45 +01:00
Hiki
72a4deab66
Update mimo-v2.5.toml
2026-05-12 23:23:53 +02:00
Hiki
8dd829a187
Update mimo-v2.5-pro.toml
2026-05-12 23:23:01 +02:00
Aiden Cline
a671cc05d5
align cost tiers with model schema
2026-05-12 16:01:49 -05:00
Aiden Cline
656c6f08a7
Merge pull request #1761 from Sewer56/deprecate-wafer-models
...
providers/wafer.ai: Remove DeepSeek-V4-Pro and MiniMax-M2.7 models
2026-05-12 15:51:10 -05:00
Aiden Cline
8979741a32
Merge pull request #1760 from Ardakilic/fix/kilo/kimik26
...
Fix: Kimi k2.6 definition on Kilo Gateway
2026-05-12 15:50:53 -05:00
Frank
82851b9a3d
update zen models
2026-05-12 16:44:41 -04:00
zxy_action
ae511892d7
feat: add Auriko provider with 15 models
...
All models use [extends] to inherit from canonical definitions,
overriding only Auriko-specific pricing. Omits remove cost tiers
and features Auriko doesn't carry.
Models: claude-opus-4-{6,7}, claude-sonnet-4-6, deepseek-v4-{pro,flash},
gemini-{2.5-pro,2.5-flash,3.1-pro-preview}, grok-4.3, kimi-k2.{5,6},
minimax-m2-7{,-highspeed}, glm-5.1, qwen-3.6-plus
2026-05-12 13:34:34 -07:00
Sewer56
9312418242
Changed: Remove DSv4 Pro & MiniMax M2.7 from models.dev
2026-05-12 20:16:08 +01:00
vimtor
9828a0177d
improve empty row
2026-05-12 19:49:03 +02:00
Arda Kılıçdağı
122627a852
fix: Kimi k2.6 definition on Kilo Gateway
2026-05-12 20:35:58 +03:00
vimtor
4dee8d0d34
minor improvements
2026-05-12 19:10:21 +02:00
Hiki
6883e793ce
Create mimo-v2.5-pro.toml
2026-05-12 18:22:41 +02:00
Hiki
e69064709b
Update mimo-v2.5.toml
2026-05-12 18:19:52 +02:00
Hiki
bd8e582b96
Create mimo-v2.5.toml
2026-05-12 18:03:09 +02:00
Shoubhit Dash
21945db90f
Merge pull request #1662 from anomalyco/nxl/add-sarvam-provider
...
provider(sarvam): add chat models
2026-05-12 13:31:59 +05:30
Aiden Cline
1771e02be8
Merge pull request #1660 from Alex-wuhu/feat/novita-ai-sync-models
...
provider(novita-ai): sync latest models
2026-05-11 23:15:44 -05:00
Aiden Cline
bb08fc26e9
Merge pull request #1706 from rohita5l/rohit/addDatabricks
...
Add Databricks as a provider
2026-05-11 19:28:34 -05:00
Aiden Cline
2e015de42d
preserve generated tier thresholds
2026-05-11 17:07:19 -05:00
Rohit Agrawal
914a3d9d18
fix: restore bun.lock to use default registry instead of Databricks npm proxy
...
Co-Authored-By: Claude Sonnet 4.6 <noreply@anthropic.com >
2026-05-11 17:54:26 -04:00
Rohit Agrawal
bab01dd9ab
refactor: move databricks generate to packages/core/script following repo conventions
...
Addresses review feedback by removing AI SDK dependencies from package.json
and aligning with the Vercel/Helicone/Wandb pattern.
- Move generate-databricks.ts to packages/core/script/
- Add databricks:generate to root scripts
- Remove smoke test and runtime filtering (catalog should reflect what the
upstream API exposes; AI SDK compatibility is a downstream concern)
- Add --dry-run and --new-only flags
- Merge with existing TOMLs instead of nuking them; warn about orphans
- Restore databricks-gemini-3-pro and databricks-gemini-3-1-pro
- Drop @ai-sdk/openai-compatible, ai, zod from root dependencies
Co-Authored-By: Claude Sonnet 4.6 <noreply@anthropic.com >
2026-05-11 17:54:04 -04:00
Rohit Agrawal
bdaae956af
feat: add AI SDK compatibility test to generate script, remove incompatible models
...
Generate script now smoke-tests each model with streamText after writing TOMLs
and removes any that return empty responses (incompatible with @ai-sdk/openai-compatible).
Removes databricks-gemini-3-pro and databricks-gemini-3-1-pro which return content
as array with thoughtSignature that the AI SDK cannot parse.
Also adds test-databricks.ts for standalone smoke testing.
Co-Authored-By: Claude Sonnet 4.6 <noreply@anthropic.com >
2026-05-11 17:53:44 -04:00
Rohit Agrawal
6f118145c0
fix: inline gpt-oss model metadata instead of invalid extends path
...
openrouter models in subdirectories can't use extends (schema requires
provider/model format); resolve() now inlines the source TOML content directly.
Co-Authored-By: Claude Sonnet 4.6 <noreply@anthropic.com >
2026-05-11 17:53:44 -04:00
Rohit Agrawal
386eaed119
add databricks
2026-05-11 17:53:44 -04:00
Aiden Cline
151e9c9071
fix tiered cost generation
2026-05-11 16:43:32 -05:00
Isaac
ba99f1edce
Add GMI Cloud GLM models
2026-05-11 14:40:03 -07:00
Aiden Cline
b96b074a3b
fix long-context cost tier omissions
2026-05-11 16:31:07 -05:00
Aiden Cline
2593e131a1
wip
2026-05-11 16:11:19 -05:00
Aiden Cline
4aebbe5ca3
Merge pull request #1754 from BruceMacD/brucemacd/fix-ollama-kimi-k2-6-model-id
...
fix ollama cloud kimi k2.6 model id
2026-05-11 14:57:03 -05:00
Bruce MacDonald
b98495e9c8
fix ollama cloud kimi k2.6 model id
2026-05-11 12:40:17 -07:00
Frank
bc95b42ccd
update zen models
2026-05-11 11:26:08 -04:00
Aiden Cline
6139fb8c69
Merge pull request #1598 from mugnimaestra/feat/chutes-generate-script
...
feat(chutes): add API-driven model generator script
2026-05-11 09:30:38 -05:00
Aiden Cline
3070758007
Merge pull request #1678 from 5kahoisaac/chore/nvidia-models
...
Sync NVIDIA endpoint model catalog
2026-05-11 09:29:54 -05:00
Frank
5525e83de4
update zen models
2026-05-11 10:00:09 -04:00
Frank
359fd879b8
update zen models
2026-05-10 03:54:03 -04:00
Frank
01b5a1a656
update zen models
2026-05-10 02:52:44 -04:00
Frank
b1958be099
update zen models
2026-05-10 02:42:44 -04:00
Aiden Cline
08aa068523
Temporarily remove kiro provider and models
2026-05-10 01:19:47 -05:00
Aiden Cline
f31ad0b02f
Merge pull request #1738 from mattiacerutti/chore/remove-gh-copilot-deprecated
...
chore(copilot): mark deprecated models
2026-05-09 15:42:19 -05:00
Aiden Cline
585aa7fa1b
Merge pull request #1741 from EriDeLee/dev
...
Update aihubmix models
2026-05-09 15:42:03 -05:00
Aiden Cline
c42a327b3e
Merge pull request #1745 from mads-digitial-solutions/patch-1
...
Update Google provider docs url from pricing page to models page
2026-05-09 15:41:51 -05:00
mads-digitial-solutions
92ebbfb5c4
Update provider.toml
...
Update Google provider docs URL from the pricing page to the models page
2026-05-09 19:49:19 +01:00
Aiden Cline
535fe8c971
Merge pull request #1744 from OpeOginni/fix/bedrock-model-ids
...
chore(bedrock): Getting rid of legacy Amazon Bedrock model offerings
2026-05-09 13:48:52 -05:00
OpeOginni
a3b4bfc16c
fix(bedrock): remove uneeded model configurations
2026-05-09 20:35:34 +02:00
OpeOginni
d0fcd6f11f
fix(bedrock): align models with current docs
2026-05-09 20:29:11 +02:00
OpeOginni
e55cd54218
fix(bedrock): remove legacy model entries
2026-05-09 20:21:33 +02:00
OpeOginni
0d73b82b9f
fix(bedrock): restore regional model IDs
2026-05-09 20:18:59 +02:00
Aiden Cline
83c7e2b63f
Merge pull request #1742 from Adam8234/add-firepass-provider
...
feat: add Fireworks (Firepass) provider
2026-05-09 12:45:09 -05:00
Adam
83ae4cf813
feat: add Fireworks (Firepass) provider
...
Adds the Fireworks AI Firepass subscription provider.
- Provider uses a dedicated FIREPASS_API_KEY
- Uses @ai-sdk/openai-compatible SDK
- Includes Kimi K2.6 Turbo (accounts/fireworks/routers/kimi-k2p6-turbo)
- Zero per-token cost since covered by subscription
2026-05-09 12:22:13 -05:00
EriDeLee
77eae6eef7
Update aihubmix models
2026-05-09 20:53:57 +08:00
Aiden Cline
8cbf6ed10e
Merge pull request #1736 from vercel/update-vercel-models-1778258030
...
Update Vercel models
2026-05-08 21:58:40 -05:00
Mattia Cerutti
06d87e4411
chore(copilot): remove deprecated models
2026-05-09 00:19:50 +02:00
Frank
2cb3832618
update zen models
2026-05-08 17:11:05 -04:00
vimtor
950a0446d4
bring back svg loading
2026-05-08 20:18:12 +02:00
vimtor
41fbfb1a17
change sst mention
2026-05-08 20:13:09 +02:00
vimtor
cf5045a90b
move copy button next to name
2026-05-08 20:12:39 +02:00
vimtor
ef739220de
lock virtualized table column widths to prevent scroll jitter
2026-05-08 19:07:25 +02:00
github-actions[bot]
df960d1a90
chore(vercel): update Vercel model definitions
...
Auto-generated by weekly workflow from Vercel AI Gateway API.
Co-Authored-By: github-actions[bot] <41898282+github-actions[bot]@users.noreply.github.com>
2026-05-08 16:33:52 +00:00
vimtor
7294efee15
move row-render.ts to shared.ts
2026-05-08 18:30:14 +02:00
vimtor
1e537147eb
remove table row tuple optimization
2026-05-08 18:24:18 +02:00
Aiden Cline
8f2f83ef61
Merge pull request #1735 from slpdy/dev
...
Create DeepSeek-V4-Pro.toml
2026-05-08 10:59:33 -05:00
Aiden Cline
b133426465
Merge pull request #1734 from oskarkocol/chore/update-novita-deepseek-prices
...
chore: update novita deepseek-v4-pro prices
2026-05-08 10:59:25 -05:00
Aiden Cline
dafff5a770
Merge pull request #1730 from smakosh/add-llmgateway-models
...
Add new LLM Gateway text models (gpt-5.5, grok-4-3, gemini-3.1-flash-lite, qwen3.6, MiMo v2)
2026-05-08 10:58:20 -05:00
smakosh
dd894f077f
Merge remote-tracking branch 'upstream/dev' into add-llmgateway-models
...
# Conflicts:
# providers/google/models/gemini-3.1-flash-lite.toml
2026-05-08 17:46:52 +02:00
smakosh
91590874e7
Revert "fix(models): use canonical entries for qwen3.6-max-preview and grok-4.3"
...
This reverts commit 70ac6fccda .
2026-05-08 17:43:07 +02:00
vimtor
e9f4cecc54
extract shared row rendering module
2026-05-08 17:29:04 +02:00
vimtor
fb1ac09883
virtualize model table
2026-05-08 17:07:28 +02:00
Jj
a436236146
Create DeepSeek-V4-Pro.toml
...
Added DeepSeek-v4-Pro model to Nebius provider
2026-05-08 11:03:16 +01:00
oskar
1415b4be97
chore: update novita deepseek prices
2026-05-08 14:21:27 +07:00
Aiden Cline
1437da86e7
Merge pull request #1685 from 8023/dev
...
Add kimi-k2.6/deepseek-v4 model and EmpirioLabs AI integration for poe.com
2026-05-07 22:38:51 -05:00
Aiden Cline
e7d57885d1
Merge pull request #1733 from chl-0537/feature/add-tencent
...
add model by openrouter
2026-05-07 22:21:07 -05:00
Aiden Cline
8157916515
fix: attachment
2026-05-07 22:12:11 -05:00
Aiden Cline
2e5b87a9c2
Merge pull request #1731 from mikeyp/chore/update-digitalocean-models
...
Add script to generate/update Digitalocean models
2026-05-07 16:42:09 -05:00
Aiden Cline
92e19432d3
add google gemini 3.1 flash lite
2026-05-07 15:44:25 -05:00
smakosh
70ac6fccda
fix(models): use canonical entries for qwen3.6-max-preview and grok-4.3
...
Apply the canonical TOML provided by the LLM Gateway team for the
Qwen3.6 Max Preview and Grok 4.3 parent definitions, replacing the
upstream-merged variants whose dates and pricing did not match.
Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com >
2026-05-07 22:34:01 +02:00
smakosh
dc3283417d
Merge remote-tracking branch 'upstream/dev' into add-llmgateway-models
...
# Conflicts:
# providers/alibaba/models/qwen3.6-max-preview.toml
# providers/llmgateway/models/qwen3.6-max-preview.toml
2026-05-07 22:28:23 +02:00
smakosh
34fd6673e5
chore(llmgateway): add new text models from llmgateway catalog
...
Add gemini-3.1-flash-lite, grok-4-3, gpt-5.5, gpt-5.5-pro, qwen3.6
and MiMo v2 models that exist in llmgateway.io but were missing
from models.dev. Adds parent definitions for grok-4-3,
gemini-3.1-flash-lite, and qwen3.6-max-preview where they did not
already exist.
Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com >
2026-05-07 22:20:52 +02:00
Mike Prasuhn
5df9293314
Add script to generate/update Digitalocean models
2026-05-07 15:28:07 -04:00
Aiden Cline
a81a9559d7
Merge pull request #1726 from Ardakilic/chore/sync-kilo-models
...
Chore: Sync Kilo models with upstream gateway
2026-05-07 13:04:51 -05:00
Aiden Cline
4197bf57a6
Merge pull request #1725 from dpuyosa/chore/venice-grok-costs
...
Venice: Update grok-4-20 pricing
2026-05-07 13:04:26 -05:00
Aiden Cline
d0ac772507
Merge pull request #1717 from juls0730/refactor/mimo-extends/token-plan
...
refactor(xiaomi-token-plan): extends xiaomi base provider for xiaomi token plans
2026-05-07 13:04:05 -05:00
Aiden Cline
1afaa053e7
Merge pull request #1727 from sergeykonkin/update-nebius-models-2026-05
...
chore: update nebius provider models
2026-05-07 13:03:46 -05:00
Aiden Cline
9b77ce1c9e
Merge pull request #1729 from Sewer56/add-minimax-m27-wafer
...
Added: MiniMax-M2.7 model for wafer.ai
2026-05-07 12:59:49 -05:00
Aiden Cline
68dc6d1820
Merge pull request #1728 from arafatkatze/codex/openrouter-qwen-cache-pricing
...
Add OpenRouter Qwen cache pricing
2026-05-07 12:59:29 -05:00
Arafatkatze
a1eb5eece3
Add OpenRouter Qwen cache pricing
2026-05-07 10:29:48 -07:00
Sewer56
f7ec2c517f
Added: wafer.ai/MiniMax-M2.7 model
2026-05-07 18:12:50 +01:00
Sergey Konkin
f98e8ec793
chore: update nebius provider models
2026-05-07 15:37:18 +02:00
Arda Kilicdagi
d210066793
chore: sync kilo gw models
2026-05-07 14:17:15 +03:00
Frank
06908cbf36
update zen models
2026-05-07 04:47:22 -04:00
dpuyosa
b0614d2088
[venice] Update grok-4-20 pricing
...
- Lower grok-4-20 and multi-agent input/output costs to latest Venice pricing
2026-05-07 10:29:48 +02:00
mickalchen
32c1c45c52
add openrouter model
2026-05-07 11:22:11 +08:00
Frank
7c033f27e6
update zen models
2026-05-06 23:00:15 -04:00
mickalchen
d648e63499
Merge remote-tracking branch 'origin/dev' into feature/add-tencent
2026-05-07 10:54:56 +08:00
Zoe
1353f965b9
refactor(xiaomi-token-plan): extends xiaomi base provider for xiaomi token plan
2026-05-06 18:53:49 -05:00
Isaac Huang
175082d43f
Add GMI Cloud provider
2026-05-06 15:12:07 -07:00
Aiden Cline
bba0a9c3f4
Merge pull request #1312 from NachoFLizaur/feat/kiro-provider
...
feat: add Kiro provider with 12 models
2026-05-06 12:16:47 -05:00
Frank
4bdb195178
update zen models
2026-05-06 12:56:58 -04:00
Aiden Cline
12706b7652
Merge pull request #1716 from juls0730/refactor/mimo-extends/zenmux
...
refactor(zenmux): extends mimo models from xiaomi provider
2026-05-06 10:45:01 -05:00
Aiden Cline
d8c76d0c67
Merge pull request #1715 from juls0730/refactor/mimo-extends/qiniu-ai
...
refactor(qiniu-ai): extends mimo models from xiaomi provider
2026-05-06 10:30:46 -05:00
Aiden Cline
3963dd13d5
Merge pull request #1714 from juls0730/refactor/mimo-extends/openrouter
...
refactor(openrouter): extends mimo models from xiaomi provider
2026-05-06 10:30:32 -05:00
Aiden Cline
8749a56efa
Merge pull request #1711 from juls0730/refactor/mimo-extends/kilo
...
refactor(kilo/xiaomi): extends mimo models from xiaomi provider
2026-05-06 10:29:31 -05:00
Aiden Cline
25de2ee27d
Merge pull request #1718 from juls0730/refactor/mimo-extends/xiaomi
...
refactor(xiaomi): round xiaomi models to powers of 2 & fix small errors
2026-05-06 10:27:36 -05:00
Aiden Cline
438df7f03c
Merge pull request #1723 from CloudFerro/fix/cloudferro-sherlock/minimax-m2.5
...
fix: cloudferro sherlock - update context values for minimax m2.5
2026-05-06 10:25:54 -05:00
Aiden Cline
7d18558aa3
Merge pull request #1722 from dpuyosa/chore/venice-model-updates
...
Venice: Remove deprecated models, enable reasoning on gpt-oss-120b
2026-05-06 10:25:36 -05:00
Jan Szypulski
c82f736fcc
fix: cloudferro sherlock - update context values for minimax m2.5
2026-05-06 10:52:39 +02:00
dpuyosa
6a436805b3
[venice] Remove deprecated models, enable reasoning on gpt-oss-120b
...
- Remove kimi-k2-thinking, qwen3-coder-480b-a35b-instruct, and venice-uncensored models
- Enable reasoning capability on openai-gpt-oss-120b
2026-05-06 09:44:24 +02:00
Jack
ce7823f073
Merge pull request #1720 from anomalyco/fix/opencode-go-kimi-k26-pricing
...
fix(opencode-go): restore kimi k2.6 pricing
2026-05-06 12:33:24 +08:00
Jack
033efdb7d4
fix(opencode-go): restore kimi k2.6 pricing
2026-05-06 12:32:06 +08:00
Alex-wuhu
0f1855c0a7
provider(novita-ai): use extends for kimi k2.6
2026-05-06 10:41:46 +08:00
Zoe
bd9e0c2677
refactor(xiaomi): round xiaomi models to powers of 2 & fix small errors
2026-05-05 18:40:05 -05:00
Zoe
38545d63a5
refactor(zenmux): extends mimo models from xiaomi provider
2026-05-05 18:17:21 -05:00
Zoe
014be328d6
refactor(qiniu-ai): extends mimo models from xiaomi provider
2026-05-05 18:05:39 -05:00
Zoe
c139147540
refactor(openrouter): extends mimo models from xiaomi provider
2026-05-05 17:49:39 -05:00
Zoe
59cd93cafc
refactor(kilo/xiaomi): extends mimo models from xiaomi provider
2026-05-05 17:13:06 -05:00
Aiden Cline
e91db96d83
Merge pull request #1710 from Spherrrical/add-digitalocean-kimi-2-6
...
feat(digitalocean): add kimi-k2.6 model
2026-05-05 14:07:50 -05:00
Spherrrical
d70dd8dcdc
feat(digitalocean): add kimi-k2.6 model
2026-05-05 12:05:16 -07:00
Frank
b18e681457
update deepseek flash on deepinfra
2026-05-05 14:24:03 -04:00
Aiden Cline
b73a6a2130
Merge pull request #1709 from xiaomochn/fix/xiaomi-mimo-v2.5-modalities
...
fix(xiaomi): swap modalities for MiMo-V2.5 and MiMo-V2.5-Pro
2026-05-05 11:36:43 -05:00
Aiden Cline
153c1cc420
Merge pull request #1614 from Yashwanth-Kumar-26/patch-1
...
Add Qwen 3.6 27B model configuration
2026-05-05 11:22:43 -05:00
Aiden Cline
ca0b30569e
update google vertex to include all anthropic models
2026-05-05 11:12:23 -05:00
xiaomochn
8345bfbd06
fix(xiaomi): correct modalities for MiMo V2.5 models across providers
...
Issues fixed:
1. MiMo-V2.5 and MiMo-V2.5-Pro had their modalities swapped in the
xiaomi canonical source (affects OpenRouter/ZenMux via extends)
2. Removed 'pdf' from V2.5 models — not a supported input modality
3. Fixed vercel provider: V2.5-Pro incorrectly marked as multimodal
4. Fixed opencode-go and vercel V2.5: added missing 'video', removed pdf
Correct modalities:
- MiMo-V2.5: input = ["text", "image", "audio", "video"]
- MiMo-V2.5-Pro: input = ["text"]
Affected providers: xiaomi, opencode-go, vercel, openrouter (extends),
zenmux (extends)
Fixes #1708
2026-05-05 22:46:44 +08:00
Shoubhit Dash
28c0d9ce23
fix(sarvam): correct output limits
2026-05-05 15:30:47 +05:30
Aiden Cline
c16f3da694
Merge pull request #1707 from deathbeam/revert-1664-fix/glm-qwen-ollama-output-limit
...
Revert "fix(ollama): set glm-5.1 and qwen3.5:397b output limits to match context"
2026-05-04 23:51:53 -05:00
Tomas Slusny
6aa6e55d60
fix(ollam): use correct output limit for qwen3.5:397b
...
{"error":"max_tokens (262144) exceeds model's maximum output tokens (65536) for model qwen3.5:397b (ref: 8554a681-e6a8-45d7-9fdd-433785eb6c67)"}
Signed-off-by: Tomas Slusny <slusnucky@gmail.com >
2026-05-05 01:31:20 +02:00
Tomas Slusny
9efaf754a5
Revert "fix(ollama): set glm-5.1 and qwen3.5:397b output limits to match context"
2026-05-05 01:04:11 +02:00
Aiden Cline
104e4bdc1f
Merge pull request #1684 from TheBaconWizard/add-clarifai-kimi-k2.6
...
provider(clarifai): add Kimi-K2.6 (moonshotai/chat-completion)
2026-05-04 12:00:33 -05:00
Aiden Cline
db16c113f7
Merge pull request #1704 from stylings/stylings/add-openrouter-grok-4-3
...
feat: add OpenRouter Grok 4.3
2026-05-04 12:00:08 -05:00
Alex
64a122a13d
fix: update OpenRouter Grok 4.3 file
2026-05-04 12:44:07 -04:00
Alex
fcc6521d1d
fix: simplify OpenRouter Grok 4.3 file
2026-05-04 12:41:56 -04:00
Alex
c623b4a55f
feat: add OpenRouter Grok 4.3
2026-05-04 12:36:33 -04:00
Aiden Cline
1d730fea16
Merge pull request #1696 from cgilly2fast/dev
...
chore(frogbot): convert firmware provider to frogbot
2026-05-04 10:26:36 -05:00
Aiden Cline
5457215e29
Merge pull request #1650 from PedroACosta/feat/add-dinference-models
...
feat(dinference): add GLM-5.1 and MiniMax-M2.5 models
2026-05-04 10:26:02 -05:00
Aiden Cline
b3ab45990f
Merge pull request #1697 from rocuevas9511/feat/add-deepinfra-gemma4
...
add gemma4 26b a4b and 31b to deepinfra
2026-05-04 10:25:31 -05:00
Aiden Cline
d906a07e31
Merge pull request #1703 from dpuyosa/feat/venice-grok-4-3
...
Venice: Add Grok 4.3 model configuration
2026-05-04 10:25:16 -05:00
dpuyosa
4b1f6edd52
[venice] Add Grok 4.3 model configuration
...
- Add Grok 4.3 model with 1M context and 32K output
- Configure standard and >200K cost tiers
- Enable text+image input with text output modalities
2026-05-04 09:54:01 +02:00
Aiden Cline
a92a2cfe3d
Merge pull request #1702 from langyo/fix/glm-5v-turbo-naming
...
fix: use proper GLM family casing for GLM-5V-Turbo
2026-05-03 16:53:44 -05:00
Aiden Cline
70891f58e5
Merge pull request #1701 from kaeltrn/add-perplexity-agent-opus-4-7-gpt-5-5
...
Add Claude Opus 4.7 and GPT-5.5 models for perplexity-agent
2026-05-03 16:53:26 -05:00
Aiden Cline
1600c827fc
Merge pull request #1698 from tim-mcdonald/add-kimi-k2.6-nvidia
...
Add Kimi K2.6 model for NVIDIA provider
2026-05-03 16:53:01 -05:00
Aiden Cline
5851cdc136
Merge pull request #1700 from JDinABox/dev
...
Add Synthetic Kimi-K2.6 model configuration
2026-05-03 16:52:46 -05:00
langyo
4fd0e38c58
fix: use proper GLM family casing for GLM-5V-Turbo
...
- Rename glm-5v-turbo to GLM-5V-Turbo in zai, zhipuai, and 302ai providers
- Add GLM-5V-Turbo back to zhipuai-coding-plan (removed in #1589 )
Ref: #1589
2026-05-04 00:51:54 +08:00
Pedro
dd685ea42c
refactor(dinference): use extends format for GLM and MiniMax models
2026-05-03 17:54:18 +02:00
kaeltrn
7835298241
Add Claude Opus 4.7 and GPT-5.5 models for perplexity-agent
2026-05-03 20:09:14 +07:00
Isaac Ng
c5fbcc2c9b
📦 CHORE: remove senera
2026-05-03 16:17:37 +08:00
Isaac Ng
ab2eb51b4e
chore(nvidia): align Nemotron endpoint slugs
...
Replace stale NVIDIA Nemotron entries with the live Build catalog slugs so the local provider catalog matches current free and partner endpoints.
2026-05-03 15:47:19 +08:00
Isaac Ng
8e19ec580c
📦 CHORE: sync latest nvidia model
2026-05-03 15:19:17 +08:00
Isaac Ng
3aecc94c46
chore(nvidia): sync endpoint model catalog
...
Update NVIDIA model TOMLs to match the live Build endpoint list by removing stale entries and adding missing ones.
This keeps the provider catalog aligned with the current free and partner endpoint inventory.
2026-05-03 14:47:30 +08:00
JD Crawford
2ac7912ee7
use extends format
2026-05-03 01:04:53 -04:00
JD Crawford
9ac5b5f625
Add Synthetic Kimi-K2.6 model configuration
2026-05-03 00:17:37 -04:00
Aiden Cline
8c4d9f4696
Merge pull request #1699 from cfal/qwen3.6-max-preview
...
providers/alibaba/models/qwen3.6-max-preview.toml: add qwen 3.6 max
2026-05-02 21:35:02 -05:00
cfal
c4bb0b4b11
providers/alibaba/models/qwen3.6-max-preview.toml: add qwen 3.6 max
2026-05-03 09:25:31 +08:00
Tim McDonald
db4d03c171
Add Kimi K2.6 model for NVIDIA provider
2026-05-02 16:09:37 -06:00
rocuevas9511
60d1d4df77
add gemma4 26b a4b and 31b to deepinfra
2026-05-02 14:26:19 -06:00
Colby Gilbert
31654fc2ef
chore(frogbot): convert firmware provider to frogbot
2026-05-02 12:36:20 -07:00
Aiden Cline
01e56b1e1e
Merge pull request #1682 from varunrandery/poolside/laguna-openrouter
...
Add Poolside Laguna series models (OpenRouter)
2026-05-02 14:30:20 -05:00
Aiden Cline
4a7d9275d6
Merge pull request #1695 from hgraca/cortecs
...
Cortecs
2026-05-02 14:04:00 -05:00
Aiden Cline
978fe11e0c
Merge pull request #1694 from cgilly2fast/dev
...
feat(firmware): add deepseek v4, remove gemini 3 pro add gpt 5.4 min,…
2026-05-02 14:03:14 -05:00
Herberto Graca
0b4198955f
refactor(cortecs): use extends for models with canonical bases
...
Convert 7 Cortecs models to extends format, inheriting from their
canonical provider definitions (deepseek, llama, mistral, alibaba).
Reduces duplication by ~55 lines while preserving Cortecs-specific
cost overrides.
2026-05-02 20:56:11 +02:00
Herberto Graca
66b568b681
Add deepseek-v4-pro model for cortecs
2026-05-02 20:56:11 +02:00
Herberto Graca
589fddeed7
Add codestral-2508 model for cortecs
2026-05-02 20:56:10 +02:00
Herberto Graca
6f75266d42
Add deepseek-v3.2 model for cortecs
2026-05-02 20:56:10 +02:00
Herberto Graca
0f03115ee0
Add deepseek-r1-0528 model for cortecs
2026-05-02 20:56:09 +02:00
Herberto Graca
ae2564edc2
Add mixtral-8x7B-instruct-v0.1 model for cortecs
2026-05-02 20:56:09 +02:00
Herberto Graca
629856cf41
Add hermes-4-70b model for cortecs
2026-05-02 20:56:08 +02:00
Herberto Graca
567bc13ab4
Add llama-3.3-70b-instruct model for cortecs
2026-05-02 20:56:08 +02:00
Herberto Graca
5eac173bc5
Add qwen3-235b-a22b-instruct-2507 model for cortecs
2026-05-02 20:56:07 +02:00
Herberto Graca
cb0d640afd
Add qwen3-coder-30b-a3b-instruct model for cortecs
2026-05-02 20:56:07 +02:00
Herberto Graca
b5e27e652f
Add nemotron-3-super-120b-a12b model for cortecs
2026-05-02 20:56:06 +02:00
Herberto Graca
650c27411f
Add qwen3.5-122b-a10b model for cortecs
2026-05-02 20:56:06 +02:00
Herberto Graca
98c55325bf
Add mistral-large-2512 model for cortecs
2026-05-02 20:56:05 +02:00
Herberto Graca
2fdd33b8a7
Add qwen3.5-397b-a17b model for cortecs
2026-05-02 20:55:57 +02:00
Colby Gilbert
e4dea0a3d8
feat(firmware): grok 4.3
2026-05-02 10:53:24 -07:00
Colby Gilbert
b4b3622bfd
feat(firmware): add deepseek v4, remove gemini 3 pro add gpt 5.4 min, add gpt 5.4 nano, add gpt 5.5, and minimax m2.7
2026-05-02 09:23:42 -07:00
Aiden Cline
3b35b5598a
Merge pull request #1693 from hgraca/add-deepseek-v4-flash-cortecs
...
Add deepseek-v4-flash model for cortecs
2026-05-02 10:46:57 -05:00
Aiden Cline
d2e16bab34
Merge pull request #1610 from monotykamary/feat/neuralwatt-provider
...
feat: add neuralwatt provider with 14 models
2026-05-02 10:45:50 -05:00
Tom X Nguyen
051dc6236a
fix(neuralwatt): sync model capabilities with provider API
2026-05-02 20:42:00 +07:00
Herberto Graca
5b9dee1f3d
Add deepseek-v4-flash model for cortecs
2026-05-02 12:50:59 +02:00
8023
426dfe7be5
fix validate error
2026-05-02 12:15:50 +08:00
8023
fddfb083d1
add EmpirioLabs AI and deepseek v4
2026-05-02 11:14:51 +08:00
8023
4eccfaba87
Add Kimi-K2.6 model configuration file
2026-05-02 11:08:32 +08:00
Jeff Lim
ea46d016d3
fix(clarifai): drop redundant cost override for Kimi-K2.6
...
Clarifai's published pricing ($0.95 input, $4.00 output) matches the
canonical moonshotai/kimi-k2.6, so the explicit [cost] block was just
duplicating upstream. Inherit it via extends instead.
Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com >
2026-05-01 18:20:39 -07:00
Jeff Lim
1bd1449ec5
provider(clarifai): add Kimi-K2.6 (moonshotai/chat-completion)
...
Extends moonshotai/kimi-k2.6 with Clarifai-specific cost and modalities
(text+image only on Clarifai; cache pricing not exposed).
Model URL: https://clarifai.com/moonshotai/chat-completion/models/Kimi-K2_6
Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com >
2026-05-01 18:14:19 -07:00
Varun Randery
7d334a51c9
Add Laguna models
2026-05-01 23:35:13 +01:00
Aiden Cline
10ae0b5a1d
Merge pull request #1664 from fernandoenzo/fix/glm-qwen-ollama-output-limit
...
fix(ollama): set glm-5.1 and qwen3.5:397b output limits to match context
2026-05-01 15:50:22 -05:00
Aiden Cline
d25cce170d
Merge pull request #1665 from Ardakilic/chore/cleanup-kilo-provider
...
Chore: Sync Kilo Gateway provider models with upstream API changes and add Owl Alpha model.
2026-05-01 15:49:53 -05:00
Aiden Cline
3e097a5d89
Merge pull request #1671 from hgraca/add-qwen-2.5-72b-instruct-cortecs
...
Add qwen-2.5-72b-instruct model for cortecs
2026-05-01 15:49:28 -05:00
Aiden Cline
074b38eb98
Merge pull request #1663 from fernandoenzo/fix/minimax-m2.7-ollama-context-output-limit
...
fix(ollama): set minimax-m2.7 context and output limits to match Ollama API
2026-05-01 14:24:35 -05:00
Aiden Cline
d474922588
Merge pull request #1679 from Spherrrical/add-digitalocean-deepseek-v4-pro
...
feat(digitalocean): add deepseek-v4-pro model
2026-05-01 13:46:50 -05:00
Spherrrical
56ddb9017c
feat(digitalocean): add deepseek-v4-pro model
2026-05-01 11:20:09 -07:00
Rohan Taneja
a565bbebc0
Merge pull request #1677 from vercel/update-vercel-models-1777652649
2026-05-01 10:47:39 -07:00
github-actions[bot]
5b97fca592
chore(vercel): update Vercel model definitions
...
Auto-generated by weekly workflow from Vercel AI Gateway API.
Co-Authored-By: github-actions[bot] <41898282+github-actions[bot]@users.noreply.github.com>
2026-05-01 16:24:16 +00:00
Arda Kilicdagi
aec3f94081
feat: owl alpha, chore: sync kilo code upstream
...
feat: owl alpha, chore: sync kilo code upstream
2026-05-01 14:09:51 +03:00
Fernando Guarini
e5c63d671e
fix(ollama): set glm-5.1 and qwen3.5:397b output limits to match context
2026-05-01 11:29:49 +02:00
Fernando Guarini
22e87ab31b
fix(ollama): set minimax-m2.7 context and output limits to match Ollama API
2026-05-01 11:29:41 +02:00
Shoubhit Dash
90515f1913
provider(sarvam): add chat models
2026-05-01 14:41:47 +05:30
Herberto Graca
97d93676bc
Add qwen-2.5-72b-instruct model for cortecs
2026-05-01 09:36:16 +02:00
Alex-wuhu
6c3c4a721c
provider(novita-ai): sync latest models
2026-05-01 14:19:17 +08:00
Aiden Cline
692fbd0f19
Merge pull request #1628 from fanweixiao/dev
...
provider(vivgrid): remove GLM-5, add GPT-5.5 and DeepSeek-v4-Pro model
2026-04-30 23:57:52 -05:00
C.C. Fan
c3bd263a67
provider(vivgrid): add deepseek-v4-pro model
...
Co-Authored-By: Claude Sonnet 4.6 <noreply@anthropic.com >
2026-05-01 11:46:39 +08:00
Aiden Cline
9b9685c297
Merge pull request #1658 from v1gnesh/dev
...
Create grok-4.3.toml
2026-04-30 22:46:08 -05:00
C.C. Fan
b3f063da79
provider(vivgrid): use extends for gpt-5.5 instead of duplicating fields
...
Co-Authored-By: Claude Sonnet 4.6 <noreply@anthropic.com >
2026-05-01 11:45:13 +08:00
C.C.
a33d0aea81
Merge branch 'anomalyco:dev' into dev
2026-05-01 11:38:42 +08:00
v1gnesh
ca380a136e
Create grok-4.3.toml
2026-05-01 08:17:12 +05:30
Arda Kılıçdağı
dfa186a768
feat: owl alpha
2026-05-01 01:47:25 +03:00
Aiden Cline
d63ffa53e8
Merge pull request #1654 from zainhas/dev
...
[Together AI] add qwen 3.6 plus
2026-04-30 16:34:28 -05:00
Aiden Cline
284def86ef
Merge pull request #1653 from smakosh/feat/llmgateway-add-gpt-5-5-and-qwen3-6
...
feat(llmgateway): add gpt-5.5, gpt-5.5-pro, qwen3.6-35b-a3b, qwen3.6-plus, qwen3.6-max-preview
2026-04-30 16:34:18 -05:00
Siddharth Dhulipalla
c19936565d
Remove Fire Pass from Fireworks Kimi K2.5 Turbo description ( #1655 )
2026-04-30 17:28:36 -04:00
Zain Hasan
73b872e141
output 500_000
2026-04-30 14:22:41 -07:00
Zain Hasan
7373bb5878
[Together AI] add qwen 3.6 plus
2026-04-30 14:21:40 -07:00
smakosh
d11c151c25
feat(llmgateway): add gpt-5.5, gpt-5.5-pro, qwen3.6-35b-a3b, qwen3.6-plus, qwen3.6-max-preview
...
Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com >
2026-04-30 22:05:40 +02:00
Aiden Cline
f3f4fea66c
Merge pull request #1645 from stylings/stylings/add-mistral-medium-3-5
...
feat: add Mistral Medium 3.5
2026-04-30 14:37:16 -05:00
Aiden Cline
90116256a6
Merge pull request #1652 from Spherrrical/add-digitalocean-provider
...
feat(digitalocean): sync model catalog
2026-04-30 14:32:07 -05:00
Spherrrical
5008df8bcc
feat(digitalocean): sync model catalog with /v1/models API
...
Add 16 new models (anthropic, openai, alibaba, deepseek, google,
meta, mistral, nvidia, baai, intfloat, fal-hosted) to match the
current DigitalOcean Gradient AI Platform catalog, and rename
openai-gpt-5-2-pro to openai-gpt-5.2-pro to match the API id.
2026-04-30 12:08:57 -07:00
Alex
696aa80dec
revert(openrouter): remove Mistral Medium 3.5 stub
2026-04-30 14:59:56 -04:00
Aiden Cline
7d00d863dc
Merge pull request #1647 from dpuyosa/chore/venice-kimi-pricing
...
Venice: Update Kimi K2.5 and K2.6 pricing and dates
2026-04-30 11:29:08 -05:00
Aiden Cline
ad9eb83b8c
Merge pull request #1648 from Snat3r/patch-1
...
Fix casing in model name MiniMax m2.7 foor cortects provider
2026-04-30 11:28:55 -05:00
Aiden Cline
467da9a82c
Merge pull request #1649 from berget-ai/feat/berget-mistral-medium-3.5
...
feat: add Mistral Medium 3.5 128B to berget.ai
2026-04-30 11:28:43 -05:00
Pedro
22786bcf4b
feat(dinference): add GLM-5.1 and MiniMax-M2.5 models
2026-04-30 13:43:44 +02:00
Christian Landgren
c8d258b7cb
feat: add Mistral Medium 3.5 128B to berget.ai
2026-04-30 12:31:28 +02:00
Snat3r
70309829ca
Fix casing in model name and update output limit
2026-04-30 11:59:44 +02:00
dpuyosa
71e00f193b
[venice] Update Kimi K2.5 and K2.6 pricing and dates
...
- Bump kimi-k2-5 cache_read from 0.11 to 0.22
- Bump kimi-k2-6 input from 0.7448 to 0.85 and cache_read from 0.1463 to 0.22
- Update last_updated dates to 2026-04-30
2026-04-30 11:18:54 +02:00
Alex
d44e724170
feat(openrouter): add Mistral Medium 3.5
2026-04-29 18:58:26 -04:00
Alex
414695db9f
feat(mistral): add Mistral Medium 3.5
2026-04-29 18:44:28 -04:00
Aiden Cline
e8b5a27723
Merge pull request #1642 from Sewer56/add-wafer-deepseek-v4-pro
...
feat(wafer.ai): add DeepSeek V4 Pro
2026-04-29 17:14:17 -05:00
Aiden Cline
6dc9d805db
Merge pull request #1643 from Sawyerb/patch-1
...
Delete providers/vercel/models/inception/mercury-coder-small.toml
2026-04-29 17:14:00 -05:00
Aiden Cline
bdf56c9111
add kimi k2.6 to azure cognitive services
2026-04-29 17:13:19 -05:00
Sewer56
293cc3e82f
feat(wafer.ai): add DeepSeek V4 Pro
2026-04-29 23:12:04 +01:00
Sawyer Birnbaum
7f5bd4f231
Delete providers/vercel/models/inception/mercury-coder-small.toml
...
Mercury Coder Small has been deprecated. People should use Mercury Edit 2 instead.
2026-04-29 14:50:35 -07:00
Aiden Cline
c4826babc5
Merge pull request #1497 from Lydanne/fix/302ai-models
...
Update 302ai model metadata and add GPT-5.4 configs
2026-04-29 14:26:01 -05:00
Rohan Taneja
91e8bb985e
Merge pull request #1641 from vercel/update-vercel-models-1777480545
...
Update Vercel models
2026-04-29 11:29:45 -07:00
Aiden Cline
56723051d4
Merge pull request #1615 from deaquino/dev
...
Add Qwen3.5-9B model configuration file to OVHCloud
2026-04-29 13:09:19 -05:00
Aiden Cline
0bb3e55c08
Merge pull request #1640 from dpuyosa/chore/venice-update-pricing
...
Venice: Update DeepSeek v4 and Qwen 3.6 model pricing and metadata
2026-04-29 13:02:40 -05:00
github-actions[bot]
080b7a328b
chore(vercel): update Vercel model definitions
...
Auto-generated by weekly workflow from Vercel AI Gateway API.
Co-Authored-By: github-actions[bot] <41898282+github-actions[bot]@users.noreply.github.com>
2026-04-29 16:35:47 +00:00
Misha Skvortsov
c59c4aae6c
feat(atomic-chat): add curated initial model list
...
Re-introduces a small curated list of models that ship preconfigured
in Atomic Chat, so opencode users get a working `models.dev` entry
out of the box instead of an empty `models: {}`.
Models (ids match the normalized form returned by Atomic Chat's
/v1/models endpoint, i.e. dots replaced with underscores):
- gemma-4-E4B-it-IQ4_XS
- gemma-4-E4B-it-MLX-4bit
- Qwen3_5-9B-Q4_K_M
- Qwen3_5-9B-MLX-4bit
- Meta-Llama-3_1-8B-Instruct-GGUF
Qwen 3.5 9B dates are taken from the verified providers/venice entry
for the same base model; quantization does not change release dates.
Made-with: Cursor
2026-04-29 17:52:56 +03:00
Frank
c5d696583e
update zen models
2026-04-29 09:38:00 -04:00
Mike Sukmanowsky
2cb5a99b98
fix: add model card links for Kimi K2 and Kimi K2.5
2026-04-29 09:29:36 -04:00
dpuyosa
70490d892f
[venice] Update DeepSeek v4 and Qwen 3.6 model pricing and metadata
...
- Reduce DeepSeek v4 Flash/Pro input and output pricing
- Add cache_read pricing for DeepSeek v4 models
- Fix Qwen 3.6 27B model name formatting
2026-04-29 09:39:11 +02:00
Aiden Cline
f858a85ba6
Merge pull request #1625 from xinrui-z/fix-aihubmix-2026-04-28
...
fix: sync AIHubMix models (2026-04-28)
2026-04-28 23:03:28 -05:00
Aiden Cline
44e4c92882
Merge pull request #1620 from kill74/add-zai-coding-plan-glm-5v-turbo
...
Add GLM-5V-Turbo to Z.ai coding plan
2026-04-28 19:28:48 -05:00
Aiden Cline
43eafcb258
Merge pull request #1581 from xiaojiezj/zenmux_0425
...
feat: add models for zenmux provider
2026-04-28 19:27:38 -05:00
Aiden Cline
b382ac7af9
Merge branch 'dev' into zenmux_0425
2026-04-28 19:08:09 -05:00
Tom X Nguyen
ca21110644
fix: sync neuralwatt models with updated API pricing and capabilities
...
The Neuralwatt API now returns accurate pricing and capabilities,
eliminating the need for manual patches (patch.json is now empty).
Changes:
- Update pricing for all 14 models from API (significant changes for
GLM, GPT-OSS, Qwen, and MiniMax models)
- Devstral Small 2 now supports image input (vision)
- kimi-k2.5-fast now supports image input (vision)
- kimi-k2.6-fast now supports reasoning + image input (was non-reasoning)
- Qwen3.6-35B-A3B now supports reasoning (was non-reasoning)
- GLM models context window: 202,752 → 200,000
- Rename fast variant model IDs to match API (dropped org prefix):
zai-org/glm-5-fast → glm-5-fast
zai-org/glm-5.1-fast → glm-5.1-fast
moonshotai/kimi-k2.5-fast → kimi-k2.5-fast
moonshotai/kimi-k2.6-fast → kimi-k2.6-fast
Qwen/qwen3.5-397b-fast → qwen3.5-397b-fast
Qwen/qwen3.6-35b-fast → qwen3.6-35b-fast
2026-04-29 07:05:23 +07:00
Mike Sukmanowsky
0c2e47e8ba
Fix token limits for Amazon Bedrock Kimi K2 models
...
Correct context and output limits for moonshot.kimi-k2-thinking and
moonshotai.kimi-k2.5 on Amazon Bedrock:
- context: 256_000 → 262_143
- output: 256_000 → 16_000
2026-04-28 17:45:59 -04:00
Aiden Cline
6a0704574b
Merge pull request #1621 from YuzhongHuangCS/dev
...
feat(wandb): Add GLM-5.1
2026-04-28 15:51:43 -05:00
Aiden Cline
1cb1341516
Merge pull request #1635 from stylings/feat/nemotron-3-nano-omni
...
feat: add Nemotron 3 Nano Omni model
2026-04-28 15:04:07 -05:00
Alex
bc47e95427
fix: rename Nemotron Omni metadata
2026-04-28 15:20:24 -04:00
Aiden Cline
b071e8add8
Merge pull request #1619 from fernandoenzo/fix/deepseek-v4-pro-ollama-cloud
...
fix(ollama-cloud): correct deepseek-v4-pro model config
2026-04-28 14:00:42 -05:00
Aiden Cline
23e527753e
Merge pull request #1623 from itsnebulalol/dev
...
feat: add gpt-5.5 pro on openai and openrouter
2026-04-28 14:00:34 -05:00
Aiden Cline
e81c045ed0
Merge pull request #1636 from dsingal0/feat/openrouter-deepseek-v4
...
feat(baseten): add DeepSeek V4 Pro
2026-04-28 13:49:11 -05:00
Dhruv Singal
0c602ca936
feat(baseten): update DeepSeek V4 Pro pricing
2026-04-28 11:45:02 -07:00
Dhruv Singal
9866f84989
feat(baseten): add DeepSeek V4 Pro
2026-04-28 11:38:22 -07:00
Alex
6b397ffe37
fix: align nvidia output limit
2026-04-28 14:35:31 -04:00
Alex
bc21596889
fix: drop openrouter provider prefix
2026-04-28 14:22:04 -04:00
Alex
20abba3190
feat: add Nemotron 3 Nano Omni
2026-04-28 14:17:22 -04:00
Dominic Frye
5881bf98a0
fix: enable pdf input modality for gpt-5.5 pro
2026-04-28 13:33:31 -04:00
Aiden Cline
595f7d028c
Merge pull request #1632 from rocuevas9511/feat/deepinfra-deepseek-v4-pro
...
feat: add DeepSeek-V4-Pro to deepinfra
2026-04-28 12:07:22 -05:00
rocuevas9511
c81dec9c5d
feat: add DeepSeek-V4-Pro to deepinfra
2026-04-28 10:56:44 -06:00
Guiii
4d45ed25d2
Use extended GLM-5V-Turbo config
...
Removed various fields and added extends section.
2026-04-28 17:49:34 +01:00
Yuzhong Huang
03cf48de53
use extends instead
2026-04-28 09:19:26 -07:00
Aiden Cline
0d3a284395
Merge pull request #1622 from eduqr/feat/fireworks-ai-deepseek-v4-pro
...
feat(fireworks-ai): add deepseek-v4-pro
2026-04-28 10:52:54 -05:00
Aiden Cline
5b1bb0fc80
Merge pull request #1624 from shelvick/add-azure-kimi-k2-6
...
Add Kimi K2.6 to Azure
2026-04-28 10:38:07 -05:00
Aiden Cline
332ebb8811
Merge pull request #1627 from ceoAppsknight/kilo/add-mimo-models
...
Add Kilo Mimo v2.5 models
2026-04-28 10:37:40 -05:00
Aiden Cline
2111813bd4
Merge pull request #1629 from ndeybach/PR-azure-5.4-limits
...
fix(azure): correct GPT-5.4 series limits and cleanup
2026-04-28 10:37:01 -05:00
Nils DEYBACH
f5b8521af6
fix: use extends and not symlinks
2026-04-28 17:34:44 +02:00
Nils DEYBACH
79481cff40
fix(azure): update GPT-5.4 metadata
...
Use `extends` for Azure GPT-5.4 variants and keep Azure-specific overrides for
PDF input and omitted fast mode.
Validated with `bun validate`.
Azure runtime manual probing confirmed GPT-5.4 uses the documented 1.05M context /
922K input / 128K output limits.
2026-04-28 14:00:20 +02:00
Nils DEYBACH
40dc356d4c
fix(azure): correct GPT-5.4 and GPT-5.4 Pro limits (and convert to extend)
...
Correct Azure GPT-5.4 and GPT-5.4 Pro limits to `1_050_000` context,
`922_000` input, and `128_000` output based on Azure runtime results and
Microsoft Learn docs. Mini and Nano already matched and are unchanged.
The limits were tested directly (see script at : https://github.com/ndeybach/Azure_endpoint_limit_test_script )
2026-04-28 12:33:27 +02:00
C.C. Fan
f929fe89e7
provider(vivgrid): remove GLM-5, add GPT-5.5 model
2026-04-28 16:44:59 +08:00
Syed Assadullah Shah
cbd245d454
add Kilo Mimo v2.5 models
2026-04-28 13:04:36 +05:00
xinrui
e5289e9b3a
fix: sync AIHubMix models (2026-04-28)
2026-04-28 11:33:25 +08:00
Scott Helvick
bb623f3ff9
Add Kimi K2.6 to Azure
2026-04-28 02:28:46 +00:00
Dominic Frye
37fffafafc
feat: add gpt-5.5 pro on openai and openrouter
2026-04-27 22:24:32 -04:00
eduqr
d329310745
feat(fireworks-ai): add deepseek-v4-pro
2026-04-27 21:05:55 -05:00
Yuzhong Huang
8016a6c45a
Add GLM-5.1 to wandb provider
2026-04-27 17:35:38 -07:00
Guilherme Sales
1eeaa0b756
Add GLM-5V-Turbo to Z.ai coding plan
2026-04-28 00:47:15 +01:00
Frank
dd3533b4e0
update zen models
2026-04-27 19:31:24 -04:00
Fernando Guarini
3a5867834f
fix(ollama-cloud): correct deepseek-v4-pro model config
...
- Remove fields that don't belong in ollama-cloud: temperature, structured_output, knowledge, interleaved
- Set output = context (1048576) per ollama-cloud convention
- Set name to lowercase per ollama-cloud convention
- Reorder fields to match existing ollama-cloud model files
2026-04-28 00:24:54 +02:00
Aiden Cline
1e83bca7a3
Merge pull request #1617 from JoshuaDietz/dev
...
feat(ollama cloud): add deepseek v4 pro
2026-04-27 16:42:56 -05:00
Aiden Cline
cd8853f88b
Merge pull request #1616 from fhennerkes/dev
...
poe: add GPT-5.5 and GPT-5.5-Pro models
2026-04-27 16:08:36 -05:00
Joshua Dietz
23b290c6a8
fix(ollama cloud): fix model name
...
Model name was inconsistent with naming schema of flash model on ollama cloud
2026-04-27 21:47:30 +02:00
fhennerkes
4e7849cee7
poe: reduce omits in gpt-5.5 extends configs
...
Inherit family, knowledge, and structured_output from base models
instead of omitting them. Only omit fields that genuinely don't
apply to Poe (provider-specific pricing tiers, different context
limits, opencode-specific provider config).
Co-Authored-By: Claude Sonnet 4.6 <noreply@anthropic.com >
2026-04-27 12:44:49 -07:00
Joshua Dietz
e3e63a7247
feat(ollama cloud): add deepseek v4 pro
2026-04-27 21:42:10 +02:00
Aiden Cline
fb297153e4
Merge pull request #1572 from YoshiTabletopGamer/qwen3.5-3.6-alibaba-open
...
[alibaba] Add remaining open Qwen 3.5 and 3.6 models, fix Qwen-3.5 397B-A17B
2026-04-27 14:36:51 -05:00
fhennerkes
e8dd06e0ce
poe: use extends format for gpt-5.5-pro
...
Address PR review comment to use extends format. Inherit from
opencode/gpt-5.5-pro since no openai/gpt-5.5-pro base exists yet.
Co-Authored-By: Claude Sonnet 4.6 <noreply@anthropic.com >
2026-04-27 10:49:36 -07:00
fhennerkes
bdee3d438b
poe: use extends format for gpt-5.5
...
Address PR review comment to use extends format and inherit from
openai/gpt-5.5 base model.
Co-Authored-By: Claude Sonnet 4.6 <noreply@anthropic.com >
2026-04-27 10:36:10 -07:00
fhennerkes
b410bc3ef2
poe: add GPT-5.5 and GPT-5.5-Pro models
2026-04-27 10:34:52 -07:00
Aiden Cline
870e3d1d26
Merge pull request #1612 from hanouticelina/add-deepseek-v4-for-huggingface
...
feat(huggingface): add DeepSeek V4 Pro
2026-04-27 11:20:12 -05:00
Aiden Cline
5f21f3b603
Merge pull request #1586 from fernandoenzo/add-ollama-cloud-deepseek-v4-flash
...
feat(ollama-cloud): add deepseek-v4-flash model
2026-04-27 10:42:53 -05:00
Celina Hanouti
27dad05e25
extend deepseek/deepseek-v4-pro
2026-04-27 16:38:36 +01:00
Aiden Cline
2ed88cbcc0
Merge pull request #1604 from ndeybach/PR-gpt-5.5
...
feat(azure): add GPT-5.5 model metadata
2026-04-27 10:20:32 -05:00
Nils DEYBACH
4d199c932e
fix: base azure-cognitive-services model not on azure
...
extend of extend does not seem to be supported
2026-04-27 17:04:20 +02:00
Jaime de Aquino
2d142f920c
Add Qwen3.5-9B model configuration file
2026-04-27 16:46:04 +02:00
Celina Hanouti
47dea9e551
fix
2026-04-27 15:40:35 +01:00
Celina Hanouti
0d25c3dcac
use extends
2026-04-27 15:37:31 +01:00
Aiden Cline
3d7f9256cb
Merge pull request #1583 from abliteration-ai/codex/add-abliteration-provider
...
Add abliteration.ai provider
2026-04-27 09:29:26 -05:00
Aiden Cline
6aa1ebd4be
Merge pull request #1595 from Contraboi/contra/add-openrouter-nano-banana-2
...
feat(openrouter): add Gemini 3.1 flash image preview (Nano Banana 2)
2026-04-27 09:28:23 -05:00
Aiden Cline
f9ebebaffd
Merge pull request #1601 from shikbupt/alibaba-deepseek
...
add alibaba-cn deepseek-v4
2026-04-27 09:28:08 -05:00
sk
7b3fe83c09
use extend format
2026-04-27 21:48:26 +08:00
Yashwanth Kumar
0a06b3efc2
Update Qwen model configuration in TOML file
2026-04-27 16:40:28 +05:30
Yashwanth Kumar
90dcbbbcc5
Add Qwen3.6 27B model configuration
2026-04-27 16:29:56 +05:30
Yashwanth Kumar
b17f5fd8ae
Delete providers/openrouter/models/qwen/qwen-3.6-27b.toml
2026-04-27 16:28:19 +05:30
Yashwanth Kumar
68691ac3f9
Update providers/openrouter/models/qwen/qwen-3.6-27b.toml
...
Co-authored-by: Copilot <175728472+Copilot@users.noreply.github.com >
2026-04-27 16:27:56 +05:30
Yashwanth Kumar
78fe905fcb
Add Qwen 3.6 27B model configuration
2026-04-27 16:22:24 +05:30
Nils DEYBACH
6b488802cf
fix: parsing error and add context over price
...
- adds the context over X price capability from parent (awaiting refactor to be correct on exact limit threashold)
- fix parsing since anything must be before extends.
2026-04-27 12:27:47 +02:00
Jack
4d0505b70e
Merge pull request #1611 from anomalyco/fix/opencode-go-deepseek-v4-flash-cache-read-20260427
...
fix(opencode-go): update deepseek v4 flash cache pricing in Go
2026-04-27 17:34:52 +08:00
Jack
b729923bd9
fix(opencode-go): correct deepseek v4 flash cache pricing
2026-04-27 17:31:41 +08:00
Celina Hanouti
34c7aa7dfe
update context limit
2026-04-27 09:33:57 +01:00
Celina Hanouti
09b4d3548b
add support for DeepSeek V4 Pro for Hugging Face provider
2026-04-27 09:31:55 +01:00
xiaojie.zj
f1cad8fdc0
feat: add zenmux models
2026-04-27 16:27:46 +08:00
Tom X Nguyen
b2f7f57f26
feat: add neuralwatt provider with 14 models
...
Add Neuralwatt as an OpenAI-compatible inference provider with
energy-aware GPU optimization. Includes 14 models across 6
sub-providers (Mistral, ZAI, OpenAI, Moonshot, MiniMax, Qwen).
Models include reasoning variants (Kimi K2.5/K2.6, GLM 5.1 FP8,
MiniMax M2.5, Qwen3.5 397B, GPT OSS 20B) and fast non-reasoning
variants (Kimi K2.5/K2.6 Fast, GLM 5/5.1 Fast, Qwen3.5/3.6 Fast),
plus Devstral Small 2 and Qwen3.6 35B A3B.
Logo derived from official Neuralwatt favicon (currentColor variant).
Pricing sourced from Neuralwatt's published rates.
2026-04-27 15:01:56 +07:00
Aiden Cline
925d4eba1f
Merge pull request #1536 from philipmat/add-openrouter-pareto-code-router
...
Adds support for openrouter/pareto-code
2026-04-26 23:52:04 -05:00
Aiden Cline
bc1e4b870b
Merge pull request #1608 from Alex-wuhu/dev
...
add deepseek-v4, qwen3.6 on novita
2026-04-26 23:16:56 -05:00
Alex-wuhu
ef913f9645
refactor: use extends format for novita deepseek v4 models
...
Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com >
2026-04-27 12:00:04 +08:00
Alex-wuhu
cec747ade3
fix: add cache_read pricing for deepseek v4 models
2026-04-27 11:09:15 +08:00
Aiden Cline
2ced32d52f
Merge pull request #1596 from NathanDrake2406/add-cf-ai-gateway-gpt-5.5
...
feat(cloudflare-ai-gateway): add openai/gpt-5.5
2026-04-26 21:57:13 -05:00
Alex-wuhu
92cc5a79df
feat: add deepseek-v4, qwen3.6 on novita
2026-04-27 10:49:13 +08:00
Aiden Cline
4cf8661f92
Merge pull request #1602 from shikbupt/alibaba-qwen3.6-max
...
add alibaba-cn qwen3.6 max
2026-04-26 17:48:14 -05:00
Aiden Cline
ebe431c0dc
Merge pull request #1606 from LightAndy1/dev
...
Add gemini-3.1-flash-preview for google-vertex
2026-04-26 17:30:05 -05:00
LightAndy
8b2c5f30a0
✏️ Fix typo
2026-04-26 21:48:54 +03:00
Nils DEYBACH
0fc92e1c47
fix: align limit on base 5.5 model
...
now that limits were fixed in base, we align azure on it
2026-04-26 20:48:00 +02:00
Nils DEYBACH
d73ca9024a
Merge remote-tracking branch 'upstream/dev' into PR-gpt-5.5
2026-04-26 20:45:20 +02:00
LightAndy
bff48f9fde
Merge branch 'anomalyco:dev' into dev
2026-04-26 21:43:13 +03:00
Aiden Cline
c5c803f415
fix: ensure openai gpt-5.5 limits are exact
2026-04-26 13:38:56 -05:00
Nils DEYBACH
cd69509b48
fix: simplify by extending the azure model from opeani
2026-04-26 20:04:27 +02:00
Aiden Cline
83d15fd756
Merge pull request #1574 from juls0730/dev
...
feat: add mimo v2.5/pro to xiaomi and openrouter providers
2026-04-26 13:55:57 -04:00
LightAndy
b3c09451d9
Add gemini-3.1-flash-preview model configuration
2026-04-26 20:27:17 +03:00
Frank
d98bdb5eff
Merge pull request #1560 from TigerBeanst/patch-1
...
fix: opencode go mimo-v2.5 context limit to 1,000,000
2026-04-26 13:16:33 -04:00
Frank
ea205913ce
update zen models
2026-04-26 12:49:37 -04:00
Frank
3532801639
update zen models
2026-04-26 11:48:32 -04:00
Nils DEYBACH
1bb50141c2
feat(azure): add GPT-5.5 model metadata
...
## Summary
Adds GPT-5.5 metadata for:
- Azure
- Azure Cognitive Services
The Azure Cognitive Services entry mirrors the existing local convention of full TOML model definitions.
## Sources
- Microsoft Learn lists `gpt-5.5` for Azure OpenAI / Microsoft Foundry with version `2026-04-24`, `1,050,000` context, `922,000` input, `128,000` output, structured outputs, tools, image input, and December 2025 training data.
- Microsoft’s Azure GPT-5.5 announcement lists pricing at `$5.00` input, `$0.50` cached input, and `$30.00` output per 1M tokens.
- Azure Responses API docs list PDF input support for vision-capable models and include `gpt-5.5` version `2026-04-24`.
## Notes
This intentionally does not add `gpt-5.5-pro`, since Azure Learn currently lists `gpt-5.5` but not `gpt-5.5-pro` in the Azure model catalog.
This also intentionally omits OpenAI-specific `context_over_200k` and `experimental.modes.fast` metadata because the Azure sources confirm the standard pricing and limits, but not those OpenAI-specific fields.
2026-04-26 17:45:25 +02:00
sk
fa8bdd63fb
add alibaba-cn qwen3.6 max
2026-04-26 19:39:41 +08:00
sk
56f64577ce
add alibaba-cn deepseek-v4
2026-04-26 19:23:04 +08:00
Nathan Nguyen
f726af5767
refactor(cloudflare-ai-gateway): use [extends] for openai/gpt-5.5
...
The model entry duplicated every field from providers/openai/models/gpt-5.5.toml,
so any future change to the upstream OpenAI definition would silently drift here.
Switch to the `[extends] from = "openai/gpt-5.5"` form already used by sibling
providers (openrouter, requesty), omitting `experimental.modes.fast` since the
gateway does not surface the OpenAI priority service tier. Validation output is
byte-identical to the prior expanded form.
2026-04-26 13:33:39 +10:00
Zoe
4838e3cb9b
feat: add mimo v2.5/pro to xiaomi and openrouter providers
2026-04-25 21:32:24 -05:00
Muhammad Mugni Hadi
dbe92646c3
chore(chutes): add header comments to generated TOML files
...
Each generated TOML now includes a comment noting which fields are
auto-managed vs manually overridable on re-run.
2026-04-26 06:52:21 +07:00
Muhammad Mugni Hadi
4717c67054
feat(chutes): add API-driven model generator script
...
Add generate-chutes.ts that fetches models from https://llm.chutes.ai/v1/models
and generates/updates TOML files, following the same pattern as generate-vercel.ts.
Supports --dry-run, --new-only, and --keep-orphans flags. Auto-deletes TOML files
for models no longer in the API (with empty directory cleanup).
Preserves manually-set fields (family, knowledge, interleaved, status) when merging
with API data. Also syncs current models from the API.
2026-04-26 06:51:16 +07:00
Nathan Nguyen
d4c77c14fd
feat(cloudflare-ai-gateway): add openai/gpt-5.5
...
Mirrors the existing direct openai/gpt-5.5 entry under the
cloudflare-ai-gateway provider so opencode and other consumers can
route GPT-5.5 traffic through Cloudflare AI Gateway without hitting
ProviderModelNotFoundError.
Pricing, limits, modalities, and dates copied from
providers/openai/models/gpt-5.5.toml; provider stanza follows the
sibling gpt-5.4 entry (npm = "ai-gateway-provider").
2026-04-26 05:22:29 +10:00
Selmir Nedzibi
96b3d65307
feat(openrouter): add Gemini 3.1 flash image preview (Nano Banana 2)
2026-04-25 21:14:44 +02:00
Aiden Cline
b491c29cf9
Merge pull request #1573 from zainhas/dev
...
[Together AI] add deepseek-v4
2026-04-25 13:47:14 -04:00
Aiden Cline
d937abd849
Merge pull request #1539 from manascb1344/fix-xiaomi-provider-ids
...
feat: add MiMo-V2.5 and MiMo-V2.5-Pro to xiaomi-token-plan providers
2026-04-25 13:33:42 -04:00
Aiden Cline
df52175b0c
Merge pull request #1580 from LeGazeon/add-nvidia-deepseek-v4-pro/flash
...
Add NVIDIA DeepSeek-V4 models
2026-04-25 13:32:41 -04:00
Aiden Cline
bee8339c07
Merge pull request #1589 from MiyakoMeow/feat/restrict-zai-zhipuai-coding-plan-models
...
rm: unavailable models in zai/zhipuai coding plan
2026-04-25 13:30:39 -04:00
Aiden Cline
181bf96fa3
Merge pull request #1585 from saju01/add-copilot-gpt-5.5
...
feat(github-copilot): add gpt-5.5
2026-04-25 13:30:16 -04:00
Aiden Cline
9d49d2fd52
Merge pull request #1587 from smakosh/claude/rebase-add-llmgateway-models-yqKLn
...
feat(llmgateway): add deepseek-v4-pro, deepseek-v4-flash, kimi-k2.6
2026-04-25 13:29:39 -04:00
Aiden Cline
648776aa85
Merge pull request #1590 from dpuyosa/feat/venice-models
...
Venice: Add GPT-5.5 and Qwen3.6 model configs
2026-04-25 13:29:03 -04:00
Aiden Cline
421cb099b0
Merge pull request #1591 from dpuyosa/fix/venice-deepseek-family
...
Venice: Fix DeepSeek V4 Flash family classification
2026-04-25 13:28:54 -04:00
Aiden Cline
f458b19994
Merge pull request #1592 from MiyakoMeow/feat/deepseek-1m-context
...
fix(deepseek): all has 1M context / 384k output / adjusted price
2026-04-25 13:28:45 -04:00
MiyakoMeow
d347093b03
feat(deepseek): 1M context / 384k output
2026-04-25 18:58:12 +08:00
MiyakoMeow
3328712262
feat: restrict zai/zhipuai coding plan models to glm-5.1, glm-5-turbo, glm-4.7, glm-4.5-air only
...
Based on official documentation:
- ZAI DevPack Coding Plan: https://docs.z.ai/devpack/overview
- Zhipu AI BigModel Coding Plan: https://docs.bigmodel.cn/cn/coding-plan/overview
Both providers only officially support the following GLM models for coding plans:
- glm-5.1
- glm-5-turbo
- glm-4.7
- glm-4.5-air
Removed unsupported models from zai-coding-plan:
- glm-4.5, glm-4.5-flash, glm-4.5v
- glm-4.6, glm-4.6v
- glm-4.7-flash, glm-4.7-flashx
- glm-5, glm-5v-turbo
Removed unsupported models from zhipuai-coding-plan:
- glm-4.5, glm-4.5-flash, glm-4.5v
- glm-4.6, glm-4.6v, glm-4.6v-flash
- glm-4.7-flash, glm-4.7-flashx
- glm-5, glm-5v-turbo
2026-04-25 18:49:10 +08:00
dpuyosa
60edc1b52d
[venice] Add GPT-5.5 and Qwen3.6 model configs
...
- Add OpenAI GPT-5.5 with 1M context window and tiered pricing
- Add OpenAI GPT-5.5 Pro with premium pricing and 128K output limit
- Add Qwen3.6 27B with text, image, and video input modalities
2026-04-25 12:41:31 +02:00
dpuyosa
eee44cd080
[venice] Fix DeepSeek V4 Flash family classification
...
- Correct family from "deepseek" to "deepseek-flash" for accurate model categorization
2026-04-25 12:36:06 +02:00
smakosh
048a3235e8
feat(llmgateway): add deepseek-v4-pro, deepseek-v4-flash, kimi-k2.6
2026-04-25 12:19:40 +02:00
Fernando Guarini
1db03ec1e6
feat(ollama-cloud): add deepseek-v4-flash model
2026-04-25 11:24:12 +02:00
Saju Sarangdharan
6d283349ad
feat(github-copilot): add gpt-5.5
...
GitHub Copilot now serves gpt-5.5 (verified via GET https://api.githubcopilot.com/models with a Copilot Enterprise token). Adding the catalog row so downstream consumers (e.g. pi-ai) can route requests.
2026-04-25 10:31:21 +02:00
Abliteration.ai
7e07302ecd
add abliteration.ai provider
2026-04-24 22:46:54 -07:00
LeGazeon
8bc407a617
chore: remove deepseek-v4-pro config (duplicated by #1578 )
...
The Pro model configuration was already added via #1578 which
has been merged. Removing the duplicate from this branch to
keep only the Flash variant.
2026-04-25 13:17:12 +08:00
LeGazeon
5305d2bae2
refactor: extend flash config from deepseek base
...
Remove duplicated fields by inheriting common settings
from providers/deepseek base config via [extends].
This addresses the review comment in #1580
2026-04-25 13:11:26 +08:00
Aiden Cline
fee96c27b9
Merge pull request #1578 from panwar-stack/dev
...
feat(nvidia): add DeepSeek V4 model
2026-04-25 00:41:03 -04:00
Aiden Cline
66520adbc6
Merge pull request #1577 from ezShroom/dev
...
add openrouter gpt-5.5
2026-04-25 00:40:34 -04:00
Zain Hasan
40714995cc
Add interleaved section to DeepSeek-V4-Pro.toml
2026-04-24 18:59:45 -07:00
LeGazeon
66c4896003
Add NVIDIA DeepSeek-V4 models
...
Add model entries for DeepSeek V4 Pro and DeepSeek V4 Flash to the NVIDIA NIM provider.
## Changes
- Added `providers/nvidia/deepseek-v4-pro.toml`
- Added `providers/nvidia/deepseek-v4-flash.toml`
## Data Sources
- NVIDIA NIM Model Cards:
- DeepSeek V4 Pro: https://build.nvidia.com/deepseek-ai/deepseek-v4-pro/modelcard
- DeepSeek V4 Flash: https://build.nvidia.com/deepseek-ai/deepseek-v4-flash/modelcard
2026-04-25 09:57:43 +08:00
panwar-stack
31091f3d4e
Rename deepseek-v4.toml to deepseek-v4-pro.toml
2026-04-24 17:09:12 -07:00
panwar-stack
5fe512c1b1
Follow extends pattern
...
Follow extends pattern
2026-04-24 17:08:49 -07:00
panwar-stack
63efa131e7
feat(nvidia): add DeepSeek V4 model
...
add DeepSeek V4 model
2026-04-24 17:05:12 -07:00
Shroom
29cd503070
Add gpt-5.5.toml configuration file
2026-04-25 00:08:04 +01:00
Rohan Taneja
a9b704c656
Merge pull request #1575 from vercel/update-vercel-models-1777063875
2026-04-24 15:29:39 -07:00
Aiden Cline
0a88e412e5
Merge pull request #1576 from dsingal0/feat/openrouter-deepseek-v4
...
Add OpenRouter DeepSeek V4 models
2026-04-24 17:42:48 -04:00
Dhruv Singal
b46e29ccd5
fix(openrouter): use DeepSeek reasoning content field
2026-04-24 14:06:05 -07:00
Jerilyn Zheng
a181b770d6
Update kimi-k2.6.toml
2026-04-24 13:55:27 -07:00
Jerilyn Zheng
fa71201f20
Update deepseek-v4-pro.toml
2026-04-24 13:54:56 -07:00
Jerilyn Zheng
3ac17aefb6
Enable open_weights in deepseek-v4-flash configuration
2026-04-24 13:54:34 -07:00
Jerilyn Zheng
1f2ceb91a5
Update qwen-3.6-max-preview.toml
2026-04-24 13:53:57 -07:00
github-actions[bot]
d82681900c
chore(vercel): update Vercel model definitions
...
Auto-generated by weekly workflow from Vercel AI Gateway API.
Co-Authored-By: github-actions[bot] <41898282+github-actions[bot]@users.noreply.github.com>
2026-04-24 20:51:22 +00:00
Zain Hasan
7e296f4cbd
ds recommend 384,000
2026-04-24 12:39:09 -07:00
Zain Hasan
f202ff51e4
Reduce output limit from 512000 to 300000
2026-04-24 12:13:53 -07:00
Zain Hasan
4085a0b536
Merge branch 'anomalyco:dev' into dev
2026-04-24 12:12:28 -07:00
Zain Hasan
935fdeca65
[Together AI] Add deepseekv4 pro
2026-04-24 12:08:33 -07:00
Frank
cef8828fbe
update zen models
2026-04-24 14:51:59 -04:00
Frank
9277a23a29
update zen models
2026-04-24 14:50:50 -04:00
Frank
d7bfb16b0f
update zen models
2026-04-24 14:42:25 -04:00
Frank
37ae6fe7c8
update zen models
2026-04-24 12:11:33 -04:00
Aiden Cline
87ce527476
Merge pull request #1571 from cgilly2fast/dev
...
fix(firmware): glm 5.1 name
2026-04-24 11:49:51 -04:00
Aiden Cline
34bed30fa1
Merge pull request #1567 from dsingal0/feat/openrouter-deepseek-v4
...
feat(openrouter): add DeepSeek V4 Pro and V4 Flash
2026-04-24 11:49:35 -04:00
Dhruv Singal
8a0c3cb75f
refactor(openrouter): extend official DeepSeek V4 Pro/Flash
...
Use [extends] from deepseek/ with OpenRouter-specific overrides
(attachment, interleaved reasoning_details, limits).
Made-with: Cursor
2026-04-24 08:43:38 -07:00
Colby Gilbert
f65460ac16
fix(firmware): glm 5.1 name
2026-04-24 08:40:19 -07:00
YoshiTabletopGamer
8ea92aed9a
[alibaba] Add remaining open Qwen 3.5 models, fix Qwen-3.5 397B-A17B, add open Qwen 3.6 models
...
- Added Qwen 3.5 122B-A10B
- Added Qwen 3.5 27B
- Added Qwen 3.5 35B-A3B
- Fixed Qwen 3.5 397B-A17B (see below)
- Added Qwen 3.6 27b
- Added Qwen 3.6-35B-A3B
I was not able to find a reliable source for the knowledge cutoff of any of these models.
2025-04 was already set as the cutoff for Qwen 3, and Qwen 3.5 is newer.
All data is from the ModelStudio webpage.
It seems to not include audio, but the ModelStudio page clearly has an audio symbol and the model is capable of this.
And I found no data for a price for reasoning tokens in particular, unlike what was in the file for Qwen 3.5 397B-A17B.
The models are all capable of structured output.
2026-04-24 12:38:52 -03:00
Dhruv Singal
f340d82fc3
feat(openrouter): add DeepSeek V4 Pro and V4 Flash
...
Add model configs aligned with OpenRouter pricing and limits
(1M context, 384K max output, cache read rates from provider page).
Made-with: Cursor
2026-04-24 08:25:17 -07:00
Frank
c7431ae24c
update zen models
2026-04-24 10:53:10 -04:00
Frank
3d1888b7b5
update zen models
2026-04-24 10:24:34 -04:00
Misha Skvortsov
16a8fa5c20
improve(atomic-chat): drop hardcoded model list per maintainer feedback
...
Made-with: Cursor
2026-04-24 17:20:26 +03:00
Aiden Cline
dcd37ccdbb
add deepseek v4 flash
2026-04-24 08:34:03 -04:00
Aiden Cline
d18c3f910c
Merge pull request #1562 from dpuyosa/update/venice-kimi-pricing
...
Venice: Update kimi-k2-6 pricing
2026-04-24 08:08:23 -04:00
Aiden Cline
2cec5a492c
Merge pull request #1563 from dpuyosa/feat/venice-deepseek-v4
...
Venice: Add DeepSeek V4 Flash and Pro models
2026-04-24 08:08:13 -04:00
dpuyosa
61dd0ec489
[venice] Add DeepSeek V4 Flash and Pro models
...
- Add DeepSeek V4 Flash with 1M context, reasoning, and tool support
- Add DeepSeek V4 Pro with 1M context, reasoning, and tool support
- Set pricing and interleaved reasoning_content field for both
2026-04-24 12:21:07 +02:00
dpuyosa
c7758204b5
[venice] Update kimi-k2-6 pricing
...
- Update input, output, and cache_read costs to current rates
- Update last_updated timestamp to 2026-04-24
2026-04-24 12:17:51 +02:00
manascb1344
ed91520aa2
feat: add MiMo-V2.5 and MiMo-V2.5-Pro to xiaomi-token-plan providers
2026-04-24 15:38:55 +05:30
Frank
3e82669a82
Merge pull request #1561 from wenbindu/dev
...
add deepseek new moels
2026-04-24 03:04:52 -04:00
Frank
1cc0c9c074
sync
2026-04-24 03:03:06 -04:00
TigerBeanst
d73d7f6453
fix: opencode go mimo-v2.5 context limit to 1,000,000
...
https://platform.xiaomimimo.com/docs/pricing
2026-04-24 12:48:34 +08:00
Aiden Cline
afb59f86ee
Merge pull request #1557 from seffhunnn/dev
...
feat: add AU Sonnet and Opus models for Amazon Bedrock
2026-04-24 00:30:51 -04:00
wenbindu
05242f68d4
add deepseek new moel
2026-04-24 12:06:19 +08:00
Mohd Saif
c1b029dcc1
feat: add AU Opus model for Amazon Bedrock
2026-04-24 03:26:28 +05:30
Mohd Saif
bc2dd5137a
feat: add AU Sonnet model for Amazon Bedrock
2026-04-24 03:25:38 +05:30
Aiden Cline
99ec4900c7
Merge pull request #1555 from brentdurksen/add-azure-claude-sonnet-4-6
...
feat(azure): add Claude Sonnet 4.6 model
2026-04-23 17:33:44 -04:00
Brent Durksen
a8c124ac9e
refactor: use extends to inherit from anthropic/claude-sonnet-4-6
2026-04-23 15:16:38 -06:00
Aiden Cline
0d20a363a9
Merge pull request #1556 from fhennerkes/dev
...
poe: add GPT-Image-2 model
2026-04-23 17:12:20 -04:00
fhennerkes
3ac613678b
poe: add GPT-Image-2 model
2026-04-23 12:38:22 -07:00
Brent Durksen
e2ead1b4e6
feat(azure): add Claude Sonnet 4.6 model
2026-04-23 13:37:50 -06:00
Aiden Cline
be53c33588
Merge pull request #1550 from BlockListed/cortecs-kimi-k2.6
...
Add kimi k2.6 to cortecs
2026-04-23 15:26:58 -04:00
Aiden Cline
55cf5fa310
Merge pull request #1554 from mattyatea/add-gpt-5-5
...
[codex] Add GPT-5.5
2026-04-23 15:17:20 -04:00
mattyatea
3e0fe362f2
add gpt-5.5 model
2026-04-24 04:13:51 +09:00
BlockListed
89d06ae31f
add kimi k2.6 to cortecs
2026-04-23 19:49:18 +02:00
Aiden Cline
c994b116ae
Merge pull request #1542 from u007/patch-1
...
Add Chutes: Kimi K2.6 TEE
2026-04-23 12:47:49 -04:00
Aiden Cline
833e8f7a66
Merge pull request #1548 from fernandoenzo/fix/gemma4-ollama-output-limit
...
fix(ollama): set gemma4:31b output limit to match context
2026-04-23 12:46:00 -04:00
Aiden Cline
32bd1427fb
Merge pull request #1545 from Alex-wuhu/dev
...
Add deepseek, gemma, ling, llama, kimi on NovitaAI
2026-04-23 12:36:06 -04:00
Frank
ae7672b87e
update zen models
2026-04-23 11:11:30 -04:00
Fernando Guarini
9b27cc5a76
fix(ollama): set gemma4:31b output limit to match context
...
Ollama does not impose official output limits. The existing convention for Gemma models on Ollama (gemma3:4b, gemma3:12b, gemma3:27b) is to set output equal to context. gemma4:31b was the only exception with output=8192 vs context=262144.
Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-openagent )
Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai >
2026-04-23 11:42:30 +02:00
Alex-wuhu
7014e3e318
feat: add missing Novita AI model configurations
...
Add 6 models served by Novita:
- deepseek/deepseek-r1-distill-qwen-14b
- deepseek/deepseek-r1-distill-qwen-32b
- google/gemma-3-12b-it
- inclusionai/ling-2.6-1t
- meta-llama/llama-3.2-3b-instruct
- moonshotai/kimi-k2.6
Capabilities, pricing, context, and modalities sourced from Novita's
/v1/models API; family slugs and release dates aligned with existing
same-model entries in the repo.
2026-04-23 14:58:43 +08:00
Frank
e0e153e8d6
update zen models
2026-04-23 02:51:28 -04:00
mickalchen
a0fbef93e4
revert
2026-04-23 14:41:05 +08:00
mickalchen
e4a89229be
add model by openrouter
2026-04-23 14:23:06 +08:00
mickalchen
9ad15e97cd
add model by openrouter
2026-04-23 14:18:51 +08:00
James
e3b4df4e87
Update Kimi-K2.6-TEE.toml
...
fix reasoning
2026-04-23 13:53:37 +08:00
Aiden Cline
e3e2066c83
Merge pull request #1544 from GodTamIt/deepinfra/kimi-k2.6
...
deepinfra: Add Kimi-K2.6 support
2026-04-23 00:40:50 -04:00
Aiden Cline
208dcd12a2
Merge pull request #1533 from qychen2001/dev
...
Add Kimi-K2.6 and Qwen3.6-35B-A3B, update Kimi-K2.5 config for siliconflow and siliconflow-cn
2026-04-23 00:40:40 -04:00
Aiden Cline
d166444aa1
Merge pull request #1541 from zainhas/dev
...
[Together AI] add Kimi k2.6 support
2026-04-23 00:40:15 -04:00
Aiden Cline
cd35f96e73
Merge pull request #1528 from seffhunnn/dev
...
Fix incorrect model ID for Gemma 4 26B (Google provider)
2026-04-23 00:39:12 -04:00
Christopher Tam
2b0c21e0b4
deepinfra: Add Kimi-K2.6 support
2026-04-22 23:39:42 -04:00
James
787b0bf9e9
Update Kimi K2.5 TEE to Kimi K2.6 TEE
2026-04-23 10:58:39 +08:00
Zain Hasan
77fbf02ff6
remove interleaved
2026-04-22 16:19:10 -07:00
Zain Hasan
89973fd122
[Together AI] add Kimi k2.6 support
2026-04-22 16:16:08 -07:00
Frank
e458a9f5b8
Merge pull request #1540 from dsingal0/feat/baseten-kimi-k2.6
...
feat(baseten): add Kimi K2.6
2026-04-22 16:59:51 -04:00
Dhruv Singal
61d95afae2
feat(baseten): add Kimi K2.6
2026-04-22 13:53:46 -07:00
Jack
4a2df5e008
Merge pull request #1537 from anomalyco/feat/opencode-go-mimo-v2.5
...
Feat/opencode go mimo-v2.5-pro & mimo-v2.5
2026-04-23 00:51:49 +08:00
Aiden Cline
3db907f3ce
Merge pull request #758 from regolo-ai/dev
...
Add Regolo-ai Provider
2026-04-22 12:33:02 -04:00
Jack
70d8f9cc6e
update mimo v2 output limits to 128k
2026-04-22 23:32:16 +08:00
Jack
b9b354ada0
update mimo v2.5 output limits to 128k
2026-04-22 23:17:40 +08:00
Jack
95632ad376
remove mimo-v2.5-omni (renamed to mimo-v2.5)
2026-04-22 23:15:00 +08:00
Philip M
7153506989
Adds support for openrouter/pareto-code
...
The Pareto Router is a way to have OpenRouter always pick a strong coding model for your needs without committing to a specific one. You express a single min_coding_score preference between 0 and 1, and the router routes your request to a coding model that meets that bar.
The Pareto Router is tuned for coding use cases. Under the hood it keeps a curated shortlist of strong coding models currently available on OpenRouter. The exact shortlist and selection logic evolve over time as new models land and benchmarks shift.
2026-04-22 10:14:46 -05:00
Jack
c73bab2e7e
providers(opencode-go): rename mimo-v2.5-omni to mimo-v2.5
2026-04-22 23:14:27 +08:00
Jack
a783dc808d
providers(opencode-go): add mimo v2.5 models and separate v2 families
2026-04-22 23:07:40 +08:00
Daniele Scasciafratte
58e72802bc
feat(models): update
2026-04-22 16:05:48 +02:00
Mohd Saif
19233d93c4
fix: remove unnecessary id field
2026-04-22 15:19:47 +05:30
QiyuanChen
7eea45e078
feat(siliconflow-cn): add Kimi-K2.6, Qwen3.6-35B-A3B and update Kimi-K2.5 config
2026-04-22 13:38:34 +08:00
QiyuanChen
001ec226f6
feat(siliconflow): add Kimi-K2.6 and update Kimi-K2.5 config
2026-04-22 13:37:59 +08:00
Jack
32461d5b44
Merge pull request #1532 from chl-0537/feature/add-tencent
...
Remove tencent token plan
2026-04-22 13:12:06 +08:00
mickalchen
9f078294c0
Remove tencent token plan
2026-04-22 13:07:03 +08:00
Aiden Cline
c885ed49cd
Merge pull request #1531 from zhiyuan1024/zhiyuan/alibaba-cn_kimi-k2.6
...
feat(alibaba-cn): add Kimi K2.6 model configuration
2026-04-21 23:41:29 -04:00
Aiden Cline
2fc434062f
Merge pull request #1529 from compumike/compumike/fix-openrouter-openai-gpt-5.4-pricing
...
Fix pricing for openrouter/openai gpt-5.4-[mini,nano] off by 10^6
2026-04-21 23:40:51 -04:00
Aiden Cline
dbcb7e6d69
Merge pull request #1530 from cgilly2fast/dev
...
feat(firmware): kimi k2.6 model
2026-04-21 23:40:22 -04:00
Zhiyuan Hou
b58392fc62
feat(alibaba-cn): add Kimi K2.6 model configuration
...
Signed-off-by: Zhiyuan Hou <zhiyuan2048@outlook.com >
2026-04-22 10:37:37 +08:00
Frank
a4818c90ca
update zen models
2026-04-21 20:18:35 -04:00
Colby Gilbert
49fbdba49f
feat(firmware): kimi k2.6 model
2026-04-21 17:14:46 -07:00
Mike Robbins
f2dd4da7f9
Fix pricing for openrouter/openai gpt-5.4-[mini,nano] off by 10^6
2026-04-21 18:41:02 -04:00
Frank
a3ed215038
update zen models
2026-04-21 17:43:38 -04:00
Mohd Saif
d45df0530b
fix: correct Gemma 4 26B model ID for Google provider
...
Updated model ID from gemma-4-26b-it to gemma-4-26b-a4b-it to match actual Gemini API. Also added missing id field and renamed the file accordingly.
2026-04-22 01:51:23 +05:30
Jack
990531258b
Merge pull request #1527 from anomalyco/feat/opencode-go-kimi-k2.6-3x-name
...
providers(opencode-go): rename kimi k2.6
2026-04-21 22:59:26 +08:00
Jack
cd2c9e3b62
providers(opencode-go): rename kimi k2.6
2026-04-21 22:54:53 +08:00
Aiden Cline
a114991278
Merge pull request #1505 from rocuevas9511/feat/deepinfra-qwen-3.5-35b
...
feat: add Qwen 3.5 35B A3B to deepinfra
2026-04-21 10:02:27 -04:00
Aiden Cline
5ff2035cee
Merge pull request #1520 from Marenz/add-deepinfra-qwen3.6-35b-a3b
...
Add Qwen3.6-35B-A3B to Deep Infra
2026-04-21 10:01:56 -04:00
Aiden Cline
a5993cc140
Merge pull request #1515 from llc1123/chore/zenmux-update
...
providers(zenmux): add support for kimi k2.6
2026-04-21 10:00:06 -04:00
Aiden Cline
aebe4b6cd0
Merge pull request #1517 from otterDeveloper/kimi2.6-pull
...
add Firework's kimi k2.6
2026-04-21 09:59:54 -04:00
Aiden Cline
214adb1154
Merge pull request #1522 from ceoAppsknight/kilo/kimi-k2.6
...
Added kilo/kimi-k2.6
2026-04-21 09:59:33 -04:00
Aiden Cline
b301c1f8b6
Merge pull request #1523 from sk0x0y/feature/nanogpt-kimi-k2.6-qwen-3.6
...
feat(nano-gpt): add Kimi K2.6 and Qwen 3.6 models
2026-04-21 09:58:45 -04:00
Aiden Cline
b81c8b385b
Merge pull request #1507 from rocuevas9511/feat/deepinfra-qwen-3.5-397b
...
feat: add Qwen 3.5 397B A17B to deepinfra
2026-04-21 09:58:28 -04:00
rocuevas9511
1caa438b3e
fix: remove id field (per Marenz feedback)
2026-04-21 07:35:52 -06:00
rocuevas9511
673bc92f4d
fix: remove id field (per Marenz feedback)
2026-04-21 07:35:38 -06:00
Jack
0159eaa158
Merge pull request #1526 from anomalyco/feat/moonshotai-cn-kimi-k2.6
...
providers(moonshotai-cn): add kimi k2.6
2026-04-21 20:39:05 +08:00
Jack
efdc7b9a54
providers(moonshotai-cn): add kimi k2.6
2026-04-21 20:26:19 +08:00
Jack
b08721206d
Merge pull request #1525 from anomalyco/feat/moonshotai-kimi-k2.6
...
providers(moonshotai): add kimi k2.6
2026-04-21 19:35:13 +08:00
Jack
fe0d4cd9fa
providers(moonshotai): add kimi k2.6
2026-04-21 19:33:06 +08:00
sk0x0y
738ad6ed70
feat(nano-gpt): add Kimi K2.6 and Qwen 3.6 model family
2026-04-21 19:22:38 +09:00
Syed Assadullah Shah
19b66fbaf2
Added kilo/kimi-k2.6
2026-04-21 15:19:17 +05:00
Mathias L. Baumann
c416836497
Add Qwen3.6-35B-A3B to Deep Infra
...
35B-total / 3B-active MoE (256 experts, 8 routed + 1 shared).
262K native context, vision + video input, thinking mode, tool calls.
Apache 2.0, $0.20 in / $1.00 out per 1M tokens.
2026-04-21 11:58:04 +02:00
rocuevas9511
059cc4b91f
fix: update model id to match DeepInfra API
2026-04-21 00:33:20 -06:00
rocuevas9511
116bd328a6
fix: update model id to match DeepInfra API
2026-04-21 00:31:22 -06:00
Frank
aa30ce3ef2
update zen models
2026-04-21 02:12:27 -04:00
Frank
9533a47906
update zen models
2026-04-21 01:20:47 -04:00
Miguel Medina
ad18f780f0
add firework's kimi 2.6
2026-04-20 23:12:11 -06:00
粒粒橙
a48519e557
providers(zenmux): add support for kimi k2.6
2026-04-21 10:13:22 +08:00
Aiden Cline
23f5e74392
Merge pull request #1514 from mfbalestra/add/kimi-k2.6-ollama-cloud
...
providers/ollama-cloud: add kimi-k2.6:cloud
2026-04-20 21:47:51 -04:00
Aiden Cline
7a50ea28a1
Merge pull request #1510 from dpuyosa/feat/venice-add-kimi-k2-6
...
Venice: Add Kimi K2.6 model configuration
2026-04-20 21:45:49 -04:00
mfbalestra
5944f94197
providers(ollama-cloud): add kimi-k2.6:cloud
2026-04-20 22:45:01 -03:00
Aiden Cline
98732b7d76
Merge pull request #1512 from SomeoneWithOptions/dev
...
add kimi-K2.6 for OpenRouter provider
2026-04-20 21:44:39 -04:00
SomeoneWithOptions
c6412d7e59
add kimi-K2.6 for OpenRouter provider
2026-04-20 18:59:49 -05:00
dpuyosa
b5a7a6e974
[venice] Add Kimi K2.6 model configuration
...
- Add new model definition for Venice provider
- Include cost, limits, and modality specs
- Enable reasoning, tool calling, and image input
2026-04-21 00:40:19 +02:00
Aiden Cline
ba7c3d7b0b
Merge pull request #1509 from kostiak/patch-1
...
Add support for Kimi-K2.6 in Kimi For Coding provider
2026-04-20 18:02:04 -04:00
Aiden Cline
53a9a2a36d
Merge pull request #1508 from hanouticelina/add-kimi-k2.6-modeling
...
feat(huggingface): add Kimi K2.6
2026-04-20 18:01:32 -04:00
kostiak
3b6bca9b90
Add support for Kimi-K2.6 for Kimi For Coding provider
2026-04-21 00:24:51 +03:00
Celina Hanouti
f5a048060f
add support for Kimi-K2.6 for Hugging Face provider
2026-04-20 21:41:14 +01:00
rocuevas9511
9e6178a5e3
fix: update cost for Qwen 3.5 397B A17B
2026-04-20 13:38:45 -06:00
rocuevas9511
e4150361b3
fix: update cost for Qwen 3.5 35B A3B
2026-04-20 13:38:06 -06:00
rocuevas9511
8de0fc059d
add Qwen 3.5 397B A17B to deepinfra
2026-04-20 13:34:56 -06:00
rocuevas9511
054733884f
add Qwen 3.5 35B A3B to deepinfra
2026-04-20 13:32:54 -06:00
Aiden Cline
3d09981eda
Merge pull request #1499 from rovo89/patch-1
...
[google] Fix cache_read cost in gemini-2.5-flash model
2026-04-20 14:43:53 -04:00
Aiden Cline
6877af7770
Merge pull request #1501 from mchenco/kimi-k2.6
...
Add Kimi K2.6 to Workers AI and AI Gateway
2026-04-20 14:41:58 -04:00
Nacho F. Lizaur
802985f76c
feat: update kiro provider to use kiro-acp-ai-provider, add opus 4.7
2026-04-20 20:13:43 +02:00
mchen
794993fd48
Add Kimi K2.6 to Workers AI and AI Gateway
2026-04-20 13:54:31 -04:00
Jack
00b53a422a
separate opencode-go kimi k2 families
2026-04-21 01:00:50 +08:00
Jack
2ccecb6011
Merge pull request #1500 from chl-0537/feature/add-tencent
...
feat: rename model
2026-04-20 22:49:09 +08:00
mickalchen
9b7e3abf00
rename model
2026-04-20 22:02:37 +08:00
Robert Vollmer
27d6a3d503
[google] Fix cache_read cost in gemini-2.5-flash model
...
https://ai.google.dev/gemini-api/docs/pricing#gemini-2.5-flash
There's no cache_read_audio, is there?
2026-04-20 15:02:54 +02:00
Lyda
6beb1f8be2
feat(302ai): standardize Claude model metadata and capabilities
...
- Add family field for all Claude models (claude-haiku, claude-opus, claude-sonnet)
- Standardize knowledge cutoff dates to full date format (YYYY-MM-DD)
- Enable reasoning capability for Claude Opus 4.x and Sonnet 4.x series models
- Add PDF input modality support for claude-opus-4-1-20250805
- Update claude-opus-4-7 context limit to 1,000,000 tokens
2026-04-20 16:10:52 +08:00
Lyda
e0c1124fe2
feat(302ai): update GPT model capabilities and specifications
...
- Add structured_output capability for GPT-4.1, GPT-4o, and GPT-5 series models
- Enable reasoning capability and disable temperature for GPT-5 series models
- Update context limits: GPT-4.1 series to 1,047,576 tokens, GPT-5.4 series to 1,050,000 tokens
- Add input token limits for GPT-5 series models (272,000 or 922,000 tokens)
- Update knowledge cutoffs across GPT-5 series (2024-05-30 to 2025-08-31)
- Add PDF input modality support for GPT-4
2026-04-20 15:58:32 +08:00
Lyda
bf4ceb7baa
feat(302ai): update GLM model capabilities and knowledge cutoffs
...
- Enable reasoning capability for GLM-4.5-air, GLM-4.5, GLM-4.5V, and GLM-4.6V models
- Update knowledge cutoff to 2025-04 for GLM-4.5, GLM-4.5V, GLM-4.6, GLM-4.6V, and GLM-4.7
- Add video input modality support for GLM-4.5V and GLM-4.6V
- Add structured_output capability for GLM-5-turbo and GLM-5.1
- Add interleaved reasoning_content field for GLM-4.7, GLM-5, GLM-5-turbo, GLM-5.1, and GLM-5V-turbo
2026-04-20 15:48:29 +08:00
Aiden Cline
1a41934e55
Merge pull request #1448 from Lydanne/dev
...
feat(302ai): supplement commonly missing models
2026-04-19 22:19:18 -05:00
Aiden Cline
aeb4caec9f
Merge pull request #1488 from Sewer56/add-wafer-provider
...
Add wafer.ai provider
2026-04-19 22:19:06 -05:00
Aiden Cline
ccb8dcc65f
Merge pull request #1493 from rovo89/patch-1
...
Add context_over_200k for gemini-2.5-pro and adjust cache_read costs
2026-04-19 22:17:45 -05:00
Lyda
435ec1df7b
feat(302ai): add claude-opus-4-7 model
2026-04-20 11:14:27 +08:00
Robert Vollmer
7c9d609143
Add context_over_200k for gemini-2.5-pro and adjust cache_read costs
...
https://ai.google.dev/gemini-api/docs/pricing#gemini-2.5-pro
https://cloud.google.com/vertex-ai/generative-ai/pricing#gemini-models-2.5 (rounds 0.125 to 0.13)
2026-04-20 00:09:11 +02:00
Aiden Cline
add7947164
Merge pull request #1492 from dpuyosa/feat/venice-add-gemma4-uncensored
...
Venice: Add Gemma 4 and Venice Uncensored 1.2 models
2026-04-19 16:52:28 -05:00
Aiden Cline
dd0c1af12a
Merge pull request #1491 from dpuyosa/chore/venice-pricing-update
...
Venice: Update pricing for Grok 4.20 and Qwen3.5 9B
2026-04-19 16:52:15 -05:00
Aiden Cline
0c93cc03be
Merge pull request #1485 from anomalyco/more-extends-cases
...
migrate more providers to extends format
2026-04-19 16:51:58 -05:00
dpuyosa
7f16117bba
[venice] Add Gemma 4 and Venice Uncensored 1.2 models
...
- Add Gemma 4 Uncensored with 256K context, image support
- Add Venice Uncensored 1.2 with 128K context, image support
- Both models support tool calls and structured output
2026-04-19 23:41:17 +02:00
dpuyosa
2b96a2d3d6
[venice-models] Update pricing for Grok 4.20 and Qwen3.5 9B
...
- Update cache_read pricing for Grok 4.20 context_over_200k (0.23 → 0.45)
- Update input cost for Qwen3.5 9B (0.05 → 0.1)
2026-04-19 23:38:20 +02:00
Sewer56
0568b412aa
Add wafer.ai provider with GLM-5.1 and Qwen3.5-397B-A17B models
2026-04-19 03:02:04 +01:00
Misha Skvortsov
7336b3619c
atomic-chat: add provider with initial blessed models
...
Adds Atomic Chat as a local OpenAI-compatible provider at
http://127.0.0.1:1337/v1 . Includes logo and three curated models:
- unsloth/Qwen3.5-9B-IQ4_XS (id: Qwen3_5-9B-IQ4_XS)
- unsloth/gemma-4-E4B-it-IQ4_XS (id: gemma-4-E4B-it-IQ4_XS)
- unsloth/MiniMax-M2.5-UD-TQ1_0 (id: MiniMax-M2_5-UD-TQ1_0)
Model ids match the normalized form returned by Atomic Chat's
/v1/models endpoint (dots replaced with underscores).
Made-with: Cursor
2026-04-17 13:03:53 +03:00
Lyda
ac7e35af4e
feat(302ai): supplement commonly missing models
2026-04-15 17:32:18 +08:00
Nacho F. Lizaur
8f340e1eb2
feat: update kiro provider to use kiro-ai-provider npm package
2026-04-13 23:12:11 +02:00
Nacho F. Lizaur
62da5b0cbd
feat: enable reasoning on Kiro Claude models
2026-04-13 19:49:25 +02:00
Nacho F. Lizaur
4e49abc10c
feat: add Kiro provider
2026-04-13 19:49:25 +02:00
massaindustries
6ecd9ec509
add qwen-next-coder-2
2026-02-05 09:11:08 +00:00
massaindustries
bb42d0b855
add qwen-next-coder
2026-02-05 09:09:30 +00:00
massaindustries
a46966efef
add-regolo-02
2026-01-29 15:14:02 +00:00
massaindustries
19ace4436f
add-regolo-01
2026-01-29 12:36:44 +00:00