fdb02983rhy 05d43a4074 Fix: Correct the max tokens of Claude-3.5-Sonnet-20241022 for Bedrock and VertexAI (#10508) vor 10 Monaten
..
__base e61752bd3a feat/enhance the multi-modal support (#8818) vor 10 Monaten
anthropic 1e8457441d fix(model_runtime): remove vision from features for Claude 3.5 Haiku (#10360) vor 10 Monaten
azure_ai_studio 574c4a264f chore(lint): Use logging.exception instead of logging.error (#10415) vor 10 Monaten
azure_openai f6fecb957e fix azure chatgpt o1 parameter error (#10067) vor 10 Monaten
baichuan b90ad587c2 refactor: move the embedding to the rag module and abstract the rerank runner for extension (#9423) vor 10 Monaten
bedrock 05d43a4074 Fix: Correct the max tokens of Claude-3.5-Sonnet-20241022 for Bedrock and VertexAI (#10508) vor 10 Monaten
chatglm 40fb4d16ef chore: refurbish Python code by applying refurb linter rules (#8296) vor 1 Jahr
cohere b90ad587c2 refactor: move the embedding to the rag module and abstract the rerank runner for extension (#9423) vor 10 Monaten
deepseek 153807f243 fix: response_format label (#8326) vor 1 Jahr
fireworks b90ad587c2 refactor: move the embedding to the rag module and abstract the rerank runner for extension (#9423) vor 10 Monaten
fishaudio 62051d5171 Corrected type annotation to "Any" from "any" all files in "model_providers" folder (#9135) vor 11 Monaten
gitee_ai 2aa171c348 Using a dedicated interface to obtain the token credential for the gitee.ai provider (#10243) vor 10 Monaten
google 12adcf8925 fix: gemini model use some tools raise error (#9993) vor 10 Monaten
gpustack 76b0328eb1 feat: add gpustack model provider (#10158) vor 10 Monaten
groq b92504bebc Added Llama 3.2 Vision Models Speech2Text Models for Groq (#9479) vor 10 Monaten
huggingface_hub b90ad587c2 refactor: move the embedding to the rag module and abstract the rerank runner for extension (#9423) vor 10 Monaten
huggingface_tei 1e829ceaf3 chore: format get_customizable_model_schema return value (#9335) vor 10 Monaten
hunyuan 92a3898540 fix: resolve the incorrect model name of hunyuan-standard-256k (#10052) vor 10 Monaten
jina b90ad587c2 refactor: move the embedding to the rag module and abstract the rerank runner for extension (#9423) vor 10 Monaten
leptonai 2cf1187b32 chore(api/core): apply ruff reformatting (#7624) vor 1 Jahr
localai 1e829ceaf3 chore: format get_customizable_model_schema return value (#9335) vor 10 Monaten
minimax b90ad587c2 refactor: move the embedding to the rag module and abstract the rerank runner for extension (#9423) vor 10 Monaten
mistralai 5ddb601e43 add MixtralAI Model (#8517) vor 11 Monaten
mixedbread b90ad587c2 refactor: move the embedding to the rag module and abstract the rerank runner for extension (#9423) vor 10 Monaten
moonshot 1b5adf40da fix: moonshot response_format raise error (#9847) vor 10 Monaten
nomic b90ad587c2 refactor: move the embedding to the rag module and abstract the rerank runner for extension (#9423) vor 10 Monaten
novita 2cf1187b32 chore(api/core): apply ruff reformatting (#7624) vor 1 Jahr
nvidia b90ad587c2 refactor: move the embedding to the rag module and abstract the rerank runner for extension (#9423) vor 10 Monaten
nvidia_nim 2cf1187b32 chore(api/core): apply ruff reformatting (#7624) vor 1 Jahr
oci b90ad587c2 refactor: move the embedding to the rag module and abstract the rerank runner for extension (#9423) vor 10 Monaten
ollama b90ad587c2 refactor: move the embedding to the rag module and abstract the rerank runner for extension (#9423) vor 10 Monaten
openai 1e829ceaf3 chore: format get_customizable_model_schema return value (#9335) vor 10 Monaten
openai_api_compatible 70ddc0ce43 openai compatiable api usage and id (#9800) vor 10 Monaten
openllm 1e829ceaf3 chore: format get_customizable_model_schema return value (#9335) vor 10 Monaten
openrouter 5a9448245b fix: remove unsupported vision in OpenRouter Haiku 3.5 (#10364) vor 10 Monaten
perfxcloud b90ad587c2 refactor: move the embedding to the rag module and abstract the rerank runner for extension (#9423) vor 10 Monaten
replicate b90ad587c2 refactor: move the embedding to the rag module and abstract the rerank runner for extension (#9423) vor 10 Monaten
sagemaker d45d90e8ae chore: lazy import sagemaker (#10342) vor 10 Monaten
siliconflow 1e829ceaf3 chore: format get_customizable_model_schema return value (#9335) vor 10 Monaten
spark d0e0111f88 fix:Spark's large language model token calculation error #7911 (#8755) vor 11 Monaten
stepfun 1e829ceaf3 chore: format get_customizable_model_schema return value (#9335) vor 10 Monaten
tencent 40fb4d16ef chore: refurbish Python code by applying refurb linter rules (#8296) vor 1 Jahr
togetherai 2cf1187b32 chore(api/core): apply ruff reformatting (#7624) vor 1 Jahr
tongyi 033ab5490b feat: support LLM understand video (#9828) vor 10 Monaten
triton_inference_server 1e829ceaf3 chore: format get_customizable_model_schema return value (#9335) vor 10 Monaten
upstage b90ad587c2 refactor: move the embedding to the rag module and abstract the rerank runner for extension (#9423) vor 10 Monaten
vertex_ai 05d43a4074 Fix: Correct the max tokens of Claude-3.5-Sonnet-20241022 for Bedrock and VertexAI (#10508) vor 10 Monaten
vessl_ai aa895cfa9b fix: [VESSL-AI] edit some words in vessl_ai.yaml (#10417) vor 10 Monaten
volcengine_maas 1e829ceaf3 chore: format get_customizable_model_schema return value (#9335) vor 10 Monaten
voyage b90ad587c2 refactor: move the embedding to the rag module and abstract the rerank runner for extension (#9423) vor 10 Monaten
wenxin 4d5546953a add llm: ernie-4.0-turbo-128k of wenxin (#10135) vor 10 Monaten
x bf9349c4dc feat: add xAI model provider (#10272) vor 10 Monaten
xinference 1e829ceaf3 chore: format get_customizable_model_schema return value (#9335) vor 10 Monaten
yi e0846792d2 feat: add yi custom llm intergration (#9482) vor 10 Monaten
zhinao 2cf1187b32 chore(api/core): apply ruff reformatting (#7624) vor 1 Jahr
zhipuai 033ab5490b feat: support LLM understand video (#9828) vor 10 Monaten
__init__.py d069c668f8 Model Runtime (#1858) vor 1 Jahr
_position.yaml fb49413a41 feat: add voyage ai as a new model provider (#8747) vor 11 Monaten
model_provider_factory.py 4e7b6aec3a feat: support pinning, including, and excluding for model providers and tools (#7419) vor 1 Jahr