summaryrefslogtreecommitdiffstats
path: root/src/client/gemini.rs
AgeCommit message (Collapse)AuthorLines
2024-07-27feat: support patching request url, headers and body (#756)sigoden-28/+15
2024-07-27feat: change model patch structure (#754)sigoden-1/+1
2024-06-21feat: support rerank (#620)sigoden-1/+1
2024-06-05feat: support RAG (#560)sigoden-11/+62
* feat: support RAG * support more embeddings models and implement concurrent embedding api * show the progress of addings paths * ignore embedding context when saving message * embedding model max_chunk_size => default_chunk_size * support pdf and pandoc formats (docx, epub, ipynb)
2024-06-01refactor: rename some client structs and methods (#555)sigoden-9/+7
* rename `Completeion*` to `ChatCompletions*` * rename `send_message*` to `chat_completions*` * rename `request_builder` to `chat_completions_builder` * rename `build_body` to `build_chat_completions_body` * rename `extract_completion` to `extract_chat_completions` * format * remove unused config fields
2024-05-30refactor: rename `SendData` to `CompletionData` (#553)sigoden-3/+7
2024-05-22feat: allow patching req body with client config (#534)sigoden-3/+7
2024-05-18feat: support function calling (#514)sigoden-3/+3
* feat: support function calling * fix on Windows OS * implement multi-steps function calling * fix on Windows OS * add error for client not support function calling * refactor message data structure and make claude client supporting function calling * support reuse previous call results * improve error handling for function calling * use prefix `may_` as indicator for `execute` type fucntions
2024-05-06feat: extract vertexai-claude client (#485)sigoden-4/+3
2024-04-30feat: openai-compatible platforms share the same client (#469)sigoden-2/+2
2024-04-30refactor: improve code qualitysigoden-3/+1
2024-04-28refactor: extract prelude models to models.yaml (#451)sigoden-9/+0
2024-04-26refactor: simplify impl client trait (#445)sigoden-25/+8
2024-04-25refactor: extract common catch_error (#437)sigoden-4/+4
2024-04-25refactor: rewrite list models of all clients (#436)sigoden-8/+9
2024-04-23feat: builtin models can be overwrited by models config (#429)sigoden-13/+6
2024-04-23feat: customize model's max_output_tokens (#428)sigoden-2/+2
2024-04-23refactor: more async code (#427)sigoden-2/+2
2024-04-11refactor: all clients use openai token counter (#402)sigoden-4/+1
2024-04-10refactor: update gemini/vertexai models list (#395)sigoden-4/+4
2024-03-25feat: support customizing gemini safeSettings (#375)sigoden-1/+4
2024-03-25refactor: reorder models (#372)sigoden-0/+1
The capable ones come first.
2024-03-06refactor: rename model's `max_tokens` to `max_input_tokens` (#339)sigoden-4/+5
BREAKING CHANGE: rename model's `max_tokens` to `max_input_tokens`
2024-02-16refactor: update vertexai/gemini/ernie clients (#309)sigoden-159/+3
2024-02-15feat: support vertexai (#308)sigoden-3/+3
2024-02-13feat: update openai/qianwen/gemini models (#306)sigoden-2/+1
2024-01-13feat: supports model capabilities (#297)sigoden-9/+8
1. automatically switch to the model that has the necessary capabilities. 2. throw an error if the client does not have a model with the necessary capabilities
2023-12-19feat: support qianwen:qwen-vl-plus (#275)sigoden-8/+11
2023-12-19feat: support gemini (#273)sigoden-0/+241