pashpashpash
89ad9a241c
resolving merge conflicts
2025-03-04 15:12:46 -08:00
loupzeur and Saoud Rizwan
ccba2ed8f8
feat(ollama): use official ollama library ( #1859 )
...
Co-authored-by: Saoud Rizwan <7799382+saoudrizwan@users.noreply.github.com >
2025-03-04 14:46:26 -08:00
pashpashpash
29c16ea7c8
reconciling merge conflicts
2025-03-01 18:15:41 -08:00
Saoud Rizwan
6a2172e156
Add gemini models to vertex
2025-03-01 18:03:30 -08:00
Pao and Saoud Rizwan
3ca529dc62
feat: Vertex Gemini Flash 2.0 support ( #1853 )
...
* feat: Vertex Gemini Flash 2.0 support
* fix: Add changeset for PR
* Moved gemini model
* Resolve conflicts
* Prettier fix
---------
Co-authored-by: Saoud Rizwan <7799382+saoudrizwan@users.noreply.github.com >
2025-03-01 17:55:43 -08:00
d103f8ee62
feat: add asksage support ( #2011 )
...
* feat: add asksage support
* chore: create changeset for asksage support
* Fix Typo
Fix Typo
* chore: fix lint
* Fixes
* Validate asksage API key
---------
Co-authored-by: Dennis Bartlett <bartlett.dc.1@gmail.com >
Co-authored-by: Saoud Rizwan <7799382+saoudrizwan@users.noreply.github.com >
2025-03-01 15:20:58 -08:00
Leonid Bugaev and Saoud Rizwan
65f582bb0d
Feat: Add support for AWS Bedrock Anthropic prompt caching ( #2034 )
...
* Add Bedrock prompt caching support (optional)
This feature protected under checkbox because it is not yet rolled out
to everyone, and if you will try to send cache headers, and its not
enabled for you, you will get error
* Add changeset
* Update supported models
* Fix copy
---------
Co-authored-by: Saoud Rizwan <7799382+saoudrizwan@users.noreply.github.com >
2025-03-01 14:09:10 -08:00
Saoud Rizwan
f555b83c7c
Revert "Sonnet: Make Separate Thinking Model ( #2031 )"
...
This reverts commit 29c01eb7d3 .
2025-02-28 23:04:20 -08:00
Evan
29c01eb7d3
Sonnet: Make Separate Thinking Model ( #2031 )
...
* make thinking model separate and remove UI slider
* use constant for min budget tokens
* changeset
2025-02-28 22:36:02 -08:00
Saoud Rizwan
eaa76512fc
Remove requesty polling; fix deepseek cost calculation changes; fix preferred language parsing
2025-02-28 22:23:24 -08:00
Evan
df7b458229
Sonnet, Give Me a Reason ( #1961 )
...
* wip
* added slider for setting reasoning budget tokens
* refactor out generic slider component; improve styling; add debounce
* added setting validation
* styling and adding reasoning level
* changeset
* make change to trigger test rerun
* revert useless comment
2025-02-28 18:16:56 -08:00
Dennis Bartlett
bb902f9b6b
Fix conflict
2025-02-28 14:24:51 -08:00
Dennis Bartlett
dfeb57d3af
Fix Lint
2025-02-28 14:06:00 -08:00
Dennis Bartlett
9232d752b2
Merge remote-tracking branch 'upstream/main' into dev
2025-02-28 14:05:27 -08:00
Daniel Trugman and Dennis Bartlett
d1a097cbd4
Requesty dynamic model selection ( #1836 )
...
* Extract reuseable ModelDescriptionMarkdown from OpenRouter model picker
* Requesty: Add model picker component
* Refactor readOpenRouterModels to allow any dynamic list filename
* Extract parsePrice to allow reuse by other providers
* Simplify model display name switch case
* Requesty: Add dynamic model list fetching from API
* Requesty: Add default model selection
* Requesty: Specify max_tokens when sending request
* Add changeset
---------
Co-authored-by: Dennis Bartlett <bartlett.dc.1@gmail.com >
2025-02-28 13:29:34 -08:00
Andrew Monostate
d7e9ead730
feat: add X AI provider integration
2025-02-28 12:01:00 -08:00
Daniel Trugman
d6184e9dac
OpenAI & DeepSeek cost calculation ( #1864 )
...
* Add OpenAI compatible cost calculation
* Requesty: Prepare for correct price calculation
* Native OpenAI: Update model caching info
According to [OpenAI's
website](https://platform.openai.com/docs/guides/prompt-caching ),
gpt-4o, gpt-4o-mini, o1-preview and o1-mini support caching.
For gpt-4o, even though gpt-4o-2024-05-13 and
chatgpt-4o-latest do no support caching, users will see there are no
cached tokens, which will help avoid confusion.
* Native OpenAI: Call getModel once
* Native OpenAI: Extract yield usage into method
* Native OpenAI: Add caching and cost info to task header
* DeepSeek: Add cost info to task header
* Add changeset
2025-02-28 11:25:17 -08:00
Saoud Rizwan
1913b69ca5
Move GPT-4.5
2025-02-27 18:49:19 -08:00
Saoud Rizwan
aa714f28e1
Add GPT-4.5
2025-02-27 18:48:00 -08:00
Dennis Bartlett
10cde3671e
Feature/privacy policy ( #1994 )
...
* Update Privacy to point to website. Add ToS to point to website.
* Lint Fix
2025-02-27 17:19:26 -08:00
Andrei Edell and Andrei Edell
43e939666d
Hugelung/gpt 4.5 ( #1999 )
...
* Add gpt-4.5-preview
* changeset
---------
Co-authored-by: Andrei Edell <andrei@nugbase.com >
2025-02-27 16:29:19 -08:00
pashpashpash
39980e9be7
Revert "Better Streaming Support ( #1980 )" ( #1993 )
...
This reverts commit 93856ab3b0 .
2025-02-27 15:50:02 -08:00
brownrw8 and ellipsis-dev[bot]
9b5cefd69d
Better Streaming Support ( #1980 )
...
* feat: enterprise support
* Update src/api/providers/enterprise.ts
Co-authored-by: ellipsis-dev[bot] <65095814+ellipsis-dev[bot]@users.noreply.github.com>
* refactor + anthropic
* fix
* update chunking
* fix enterprise providers
* minor refactor for chunking
* comments
* defaults
* tests
* remove imports
* suggested fixes
* decouple message type
* upgrade libs, prompt caching no longer in beta for claude models
* updates for tests
* enterprise -> streaming provider
* update tests
* remove section from anthropic.ts
* handle specific GCP Vertex invalid_grant error
* finish comment
* add caching stats to chunking
---------
Co-authored-by: ellipsis-dev[bot] <65095814+ellipsis-dev[bot]@users.noreply.github.com>
2025-02-27 15:46:01 -08:00
Doug Daniels
a229b8ce9d
feat(vertex): Add prompt caching support for Claude on Vertex AI ( #1885 )
...
* feat(vertex): Add prompt caching support for Claude on Vertex AI
* Remove countTokens update claude 3.7
* claude-3-7-sonnet@20250219 support in Vertex AI as default model
2025-02-26 22:28:02 -08:00
Minhao-Zhang
7bc02926b4
Fix Official DeepSeek-V3 cost ( #1944 )
...
* update the api price for deepseek (discount period is over)
* update the api price for deepseek (discount period is over)
2025-02-26 19:05:19 -08:00
watany
5e65ea04c9
fix: Anthropic's default model is 3.7 ( #1971 )
...
* fix: Anthropic's default model is 3.7
* changeset
2025-02-26 18:42:39 -08:00
Saoud Rizwan
408c0887ac
Update default model ID
2025-02-24 18:17:35 -08:00
Saoud Rizwan
1c9da770a8
Revert "Temporarily revert default openrouter model until API is fixed"
...
This reverts commit 45b13dd775 .
2025-02-24 18:15:33 -08:00
Saoud Rizwan
45b13dd775
Temporarily revert default openrouter model until API is fixed
2025-02-24 11:14:32 -08:00
Saoud Rizwan
1127e2a33c
Add Claude 3.7 Sonnet
2025-02-24 11:07:23 -08:00
pashpashpash
32dcdcf038
reconciling merge conflicts with main
2025-02-19 20:12:13 -08:00
ee06c0811c
add all qwen2.5 coder models ( #1797 )
...
* add alibaba qwen-max qwen-plus qwen-turbo qwen-coder-plus stable/latest models
* add alibaba qwen-max qwen-plus qwen-turbo qwen-coder-plus stable/latest models
* Provide the api line choice for international user
* Remove redundant code
* Copy fixes
* Create dry-socks-talk.md
* fix problem what is when you use Qwen api provider and then you want to change the api provider ,the apiline dropdown will obscure your api provider drop-down options
* feat: add qwen2.5-coder models
Description
Add new models as list:
qwen2.5-coder-32b-instruct
qwen2.5-coder-14b-instruct
qwen2.5-coder-7b-instruct
qwen2.5-coder-3b-instruct
Doc: https://help.aliyun.com/zh/model-studio/getting-started/models#9f8890ce29g5u
* add changeset
* feat: add all alibaba qwen2.5 coder models
---------
Co-authored-by: yaojunWang <gooqle.com.hk@gmail.com >
Co-authored-by: Saoud Rizwan <7799382+saoudrizwan@users.noreply.github.com >
Co-authored-by: 刘耸 <song.liu@yo-star.com >
2025-02-14 11:10:37 -08:00
wen-jy and 执无
8e12bdb00d
Add support for qwen vl models ( #1776 )
...
* feat: add support for qwen vl models
* feat: add support for qwen vl models
* feat: updated the price of the Qianwen model following the BaiLian platform documentation
---------
Co-authored-by: 执无 <jiyong.wjy@alibaba-inc.com >
2025-02-13 18:03:20 -08:00
Hiroki Nakashima
283d7d6928
fix: adjust litellm default context window settings ( #1774 )
2025-02-12 23:01:01 -08:00
0434b5c772
Advanced configuration for OpenAI Compatible Providers ( #1737 )
...
* feat: advanced configuration for OpenAI Compatible Providers
* Update .changeset/thirty-eyes-appear.md
Co-authored-by: ellipsis-dev[bot] <65095814+ellipsis-dev[bot]@users.noreply.github.com>
* Update webview-ui/src/components/settings/ApiOptions.tsx
Co-authored-by: ellipsis-dev[bot] <65095814+ellipsis-dev[bot]@users.noreply.github.com>
* dropdown menu
* Show pricing if user entered model info
---------
Co-authored-by: ellipsis-dev[bot] <65095814+ellipsis-dev[bot]@users.noreply.github.com>
Co-authored-by: Saoud Rizwan <7799382+saoudrizwan@users.noreply.github.com >
2025-02-12 14:09:55 -08:00
Sanjaykumar S and Saoud Rizwan
e534c3dbc7
Add api key for litellm api provider #1766 ( #1767 )
...
* Add api key for litellm api provider
* Added Changeset
* Update litellm.ts
* Update ApiOptions.tsx
---------
Co-authored-by: Saoud Rizwan <7799382+saoudrizwan@users.noreply.github.com >
2025-02-12 10:56:13 -08:00
Waylon and fine
e8a2e88fa0
feat: qwen platform adds deepseek-r1/v3 support ( #1729 )
...
Co-authored-by: fine <fine_951111@163.com >
2025-02-10 20:04:16 -08:00
pashpashpash
0f7b726b70
cleaned up all old logic for extension server side auth persistence
2025-02-09 02:31:41 -08:00
pashpashpash
5a3a9f1392
resolving all merge conflicts
2025-02-08 17:49:27 -08:00
Evan and Saoud Rizwan
4449b51e2c
Let's reason together ( #1597 )
...
* let's reason together
* typo
* changeset
* Fix type error
---------
Co-authored-by: Saoud Rizwan <7799382+saoudrizwan@users.noreply.github.com >
2025-02-07 14:32:14 -08:00
brownrw8
84f017c98e
feat:Add Together API Provider ( #1698 )
2025-02-07 13:55:48 -08:00
Daniel Steigman
076b1e39e6
Nighttrek/bedrock credential manager ( #1667 )
...
* added initial AWS bedrock support
* fixed formatting
* updated to fix persistence
* added changeset
2025-02-06 23:36:38 -08:00
brownrw8 and Saoud Rizwan
3cacd57949
Add dedicated Requesty provider ( #1677 )
...
* feat: Add dedicated Requesty provider
* Update ExtensionStateContext.tsx
---------
Co-authored-by: Saoud Rizwan <7799382+saoudrizwan@users.noreply.github.com >
2025-02-06 20:45:52 -08:00
Saoud Rizwan
9bddd9a846
Change default OpenRouter model to non-moderated ( #1683 )
...
* Change default OR model to non-moderated
* Create breezy-bobcats-change.md
2025-02-06 18:43:38 -08:00
aicc and Saoud Rizwan
548338d39e
Add alibaba qwen models plus/max/coder-plus/turbo both stable and latest to use. ( #1648 )
...
* add alibaba qwen-max qwen-plus qwen-turbo qwen-coder-plus stable/latest models
* add alibaba qwen-max qwen-plus qwen-turbo qwen-coder-plus stable/latest models
* Provide the api line choice for international user
* Remove redundant code
* Copy fixes
* Create dry-socks-talk.md
---------
Co-authored-by: Saoud Rizwan <7799382+saoudrizwan@users.noreply.github.com >
2025-02-05 23:23:33 -08:00
Saoud Rizwan
0795b046e1
Prepare for release
2025-02-05 12:49:58 -08:00
Saoud Rizwan
b2e4623559
Add new gemini models ( #1656 )
...
* Update gemini models
* Update changeset
2025-02-05 10:49:55 -08:00
omercelik
0b7fac0c9d
Add Gemini 2.0 Pro ( #1655 )
2025-02-05 10:43:55 -08:00
pashpashpash
787863f7b4
merge conflicts resolved
2025-02-05 02:51:17 -08:00
Daniel Steigman
f108f20466
updated the model list to remove the embedding model ( #1646 )
2025-02-05 01:11:08 -08:00