Commit Graph
372 Commits
Author SHA1 Message Date
YeforiandJoel 443d929137 feat: add function calling for deepseek models (#6990) 2024-08-07 11:33:19 +08:00
小羽andJoel aeda8869bc feat:nvidia add nemotron4-340b and microsoft/phi-3 (#6973) 2024-08-07 11:33:19 +08:00
takatostandGitHub 6da14c2d48 security: fix api image security issues (#6971) 2024-08-05 20:21:08 +08:00
Pedro GomesandGitHub a34285196b Revise the wrong pricing of certain LLM models. (#6967) 2024-08-05 18:41:44 +08:00
takatostandGitHub ea30174057 chore: optimize streaming tts of xinference (#6966) 2024-08-05 18:23:23 +08:00
141e4e0276 fix: restore xinference secret field (#6941)
Co-authored-by: liuzhenghua-jk <[email protected]>
2024-08-04 22:32:24 +08:00
WeaxsandGitHub 5e634a59a2 compatible xinference reranker server (#6927) 2024-08-04 13:49:38 +08:00
JuHyung SonandGitHub 2e941bb91c add new provider Solar (#6884) 2024-08-02 20:48:09 +08:00
sinoandGitHub 8166a8caf5 feat: update llama3.1 parameters for openrouter (#6901) 2024-08-02 13:13:34 +08:00
灰灰andGitHub 56af1a0adf pref: change ollama embedded api request (#6876) 2024-08-02 12:04:47 +08:00
dufeiandGitHub f8617db012 fix tongyi tool calls (#6896) 2024-08-02 10:03:43 +08:00
WeaxsandGitHub cc4785f094 fix: xinference reranker return_documents (#6888) 2024-08-01 19:57:53 +08:00
chenxu9741andGitHub a9cd6df97e Remove tts (blocking call) (#6869) 2024-08-01 14:50:22 +08:00
呆萌闷油瓶andGitHub f31142e758 Azure 4o mini options (#6873) 2024-08-01 14:04:18 +08:00
crazywoolaandGitHub 792f908afb Revert "feat:Azure gpt4o mini" (#6870) 2024-08-01 13:32:03 +08:00
呆萌闷油瓶andGitHub 14367ddc09 feat:Azure gpt4o mini (#6866) 2024-08-01 13:03:08 +08:00
Charlie.WeiGitHubluowei <glpat-EjySCyNjWiLqAED-YmwM>crazywoolacrazywoola
cbf7f21ade Add azure gpt4omini (#6862)
Co-authored-by: luowei <glpat-EjySCyNjWiLqAED-YmwM>
Co-authored-by: crazywoola <[email protected]>
Co-authored-by: crazywoola <[email protected]>
2024-08-01 12:57:52 +08:00
WeaxsandGitHub f6e8e120a1 support xinference tts (#6746) 2024-08-01 11:59:15 +08:00
JoeandGitHub 08f922d8c9 fix: anthropic max token NoneType error (#6858) 2024-08-01 11:30:00 +08:00
小羽andGitHub 56b43f62d1 feat: nvidia add llama3.1 model (#6844) 2024-07-31 21:24:02 +08:00
4b410494b3 Add model parameter enable_enhance for hunyuan llm model (#6847)
Co-authored-by: sun <[email protected]>
2024-07-31 20:04:43 +08:00
JoeandGitHub df9bd36cab fix: claude-3-5-sonnet-20240620 max token error (#6843) 2024-07-31 18:34:44 +08:00
longzhihunandGitHub 9ce5cea911 feat: bedrock invoke enhancement (#6808) 2024-07-30 21:57:18 +08:00
SiliconFlow, IncandGitHub 3e18d32ce5 add deepseek-coder-v2 in siliconflow (#6149) 2024-07-29 18:45:19 +08:00
CharlesandGitHub 94d68b6a08 upgrade deepseek params (#6744) 2024-07-29 18:31:56 +08:00
c9ff0e3961 Add model hunyuan-embedding (#6657)
Co-authored-by: sun <[email protected]>
2024-07-29 18:30:52 +08:00
Bowen LiangandGitHub 20268708cc chore: improve position map conversion and tolerate empty position yaml file (#6541) 2024-07-29 10:32:11 +08:00
-LAN-andGitHub 83af50368f fix(api/core/model_runtime/model_providers/azure_openai/llm/llm.py): Try to skip if delta.delta is None. (#6727)
Signed-off-by: -LAN- <[email protected]>
2024-07-27 00:05:21 +08:00
JoeandGitHub e4542215cc fix: tongyi empty tool_calls is not supported in message (#6719) 2024-07-26 18:10:13 +08:00
3d3677e912 Feat/model provider novita (#6717)
Co-authored-by: takatost <[email protected]>
2024-07-26 17:37:21 +08:00
chenxu9741andGitHub 6b50bb0fe6 issues #6655 Open ai tts issues (#6696) 2024-07-26 14:55:49 +08:00
longzhihunandGitHub c5ac004f15 [seanguo] fix: unsupported filename in windows & add Mistral Large 2 (#6679) 2024-07-25 19:26:46 +08:00
RookieAgentandGitHub 78a339a794 modify llama3-1 yaml filename to support Windows pull operations (#6677) 2024-07-25 18:58:55 +08:00
ca696fe94c Add support of tool-call for model provider "hunyuan" (#6656)
Co-authored-by: sun <[email protected]>
2024-07-25 11:27:58 +08:00
longzhihunandGitHub 9815aab7a3 [seanguo] feat: add llama 3.1 support in bedrock (#6645) 2024-07-25 11:20:37 +08:00
5af2df0cd5 fix: qwen fc error (#6620)
Co-authored-by: dufei <[email protected]>
2024-07-24 16:56:06 +08:00
takatostandGitHub 4c85393a1d feat: add GroqCloud llama3.1 series models support (#6596) 2024-07-24 00:41:58 +08:00
sinoandGitHub d5c2680fde feat: support llama3.1 series models for openrouter provider (#6595) 2024-07-24 00:37:48 +08:00
JoeandGitHub 8123a00e97 feat: update prompt generate (#6516) 2024-07-23 19:52:14 +08:00
Lance MaoandGitHub 7c55c39085 feat: add tencent asr (#6091) 2024-07-23 16:38:39 +08:00
5e6fc58db3 Feat/environment variables in workflow (#6515)
Co-authored-by: JzoNg <[email protected]>
2024-07-22 15:29:39 +08:00
4f9f175f25 fix: correct gpt-4o-mini max token (#6472)
Co-authored-by: crazywoola <[email protected]>
2024-07-19 18:24:58 +08:00
sinoandGitHub 9e168f9d1c feat: support gpt-4o-mini for openrouter provider (#6447) 2024-07-19 13:09:41 +08:00
WeaxsandGitHub ea45496a74 update ernie models (#6454) 2024-07-19 13:08:39 +08:00
Richards TuandGitHub 8e49146a35 [EMERGENCY] Fix Anthropic header issue (#6445) 2024-07-19 07:38:15 +08:00
takatostandGitHub dad3fd2dc1 feat: add gpt-4o-mini (#6442) 2024-07-19 01:53:43 +08:00
4a026fa352 Enhancement: add model provider - Amazon Sagemaker (#6255)
Co-authored-by: Yuanbo Li <[email protected]>
Co-authored-by: crazywoola <[email protected]>
2024-07-18 19:32:31 +08:00
themanforfreeandGitHub ba181197c2 feat: api_key support for xinference (#6417)
Signed-off-by: themanforfree <[email protected]>
2024-07-18 18:58:46 +08:00
forrestlinfengandGitHub 3b5b548af3 Add Stepfun LLM Support (#6346) 2024-07-18 07:47:18 +08:00
Richards TuandGitHub 4782fb50c4 Support new Claude-3.5 Sonnet max token limit (#6335) 2024-07-18 07:47:06 +08:00