summaryrefslogtreecommitdiffstats
path: root/src/client/ollama.rs
AgeCommit message (Collapse)AuthorLines
2024-06-01refactor: rename some client structs and methods (#555)sigoden-11/+15
* rename `Completeion*` to `ChatCompletions*` * rename `send_message*` to `chat_completions*` * rename `request_builder` to `chat_completions_builder` * rename `build_body` to `build_chat_completions_body` * rename `extract_completion` to `extract_chat_completions` * format * remove unused config fields
2024-05-30refactor: rename `SendData` to `CompletionData` (#553)sigoden-5/+9
2024-05-29refactor: use `json_stream` for ollama to improve reliability (#549)ProjectMoon-11/+11
* Use JSON stream for ollama to improve reliability. Fixes #548. * remove unused import * fix clippy error * format --------- Co-authored-by: sigoden <sigoden@gmail.com>
2024-05-22feat: allow patching req body with client config (#534)sigoden-3/+4
2024-05-18feat: support function calling (#514)sigoden-5/+18
* feat: support function calling * fix on Windows OS * implement multi-steps function calling * fix on Windows OS * add error for client not support function calling * refactor message data structure and make claude client supporting function calling * support reuse previous call results * improve error handling for function calling * use prefix `may_` as indicator for `execute` type fucntions
2024-05-08refactor: model pass_max_tokens (#493)sigoden-1/+1
2024-04-30feat: openai-compatible platforms share the same client (#469)sigoden-2/+2
2024-04-30refactor: improve code qualitysigoden-3/+1
2024-04-29refactor: check res status codesigoden-2/+2
2024-04-29refactor: prompts for generating config file (#463)sigoden-3/+5
2024-04-29refactor: rename some structs (#457)sigoden-5/+5
2024-04-29feat: non-streaming returns completion stats (#456)sigoden-5/+5
2024-04-28refactor: rename ollama config field api_key => api_auth (#453)sigoden-6/+6
2024-04-28refactor: extract prelude models to models.yaml (#451)sigoden-1/+0
2024-04-26refactor: simplify impl client trait (#445)sigoden-22/+3
2024-04-25refactor: extract common catch_error (#437)sigoden-10/+2
2024-04-25refactor: rewrite list models of all clients (#436)sigoden-7/+6
2024-04-24feat: support customizing `top_p` parameter (#434)sigoden-0/+4
2024-04-24refactor: handling of response error (#433)sigoden-5/+12
2024-04-24refactor: handling of system message (#432)sigoden-5/+3
2024-04-23feat: builtin models can be overwrited by models config (#429)sigoden-7/+3
2024-04-23feat: customize model's max_output_tokens (#428)sigoden-19/+12
2024-04-23refactor: more async code (#427)sigoden-3/+3
2024-04-11refactor: all clients use openai token counter (#402)sigoden-5/+1
2024-03-06refactor: rename model's `max_tokens` to `max_input_tokens` (#339)sigoden-3/+3
BREAKING CHANGE: rename model's `max_tokens` to `max_input_tokens`
2024-02-05fix: do not attempt to deserialize zero byte chunks in ollama stream (#303)Joseph Goulden-0/+3
2024-01-30feat: add `extra_fields` to models of localai/ollama clients (#298)Kelvie Wong-1/+4
* Add an "extra_fields" config to localai models Because there are so many local AIs out there with a bunch of custom parameters you can set, this allows users to send in extra parameters to a local LLM runner, such as, e.g. `instruction_template: Alpaca`, so that Mixtral can take a system prompt. * support ollama --------- Co-authored-by: sigoden <sigoden@gmail.com>
2024-01-13feat: supports model capabilities (#297)sigoden-13/+6
1. automatically switch to the model that has the necessary capabilities. 2. throw an error if the client does not have a model with the necessary capabilities
2023-12-25refactor: ollam api_base configuration (#285)sigoden-1/+1
2023-12-20feat: support ollama (#276)sigoden-0/+209