文件历史

提交图

169 次代码提交

作者 SHA1 备注 提交日期
Sukhbinder Singh ac3d0089d0 Fix windows bug where llm doesn't run <<llm chat>> on Windows issue #495 (#646)
* Fix windows bug where llm doesn't run <<llm chat>> on Windows issue #495

* Applied Black

---------

Co-authored-by: Sukhbinder Singh <sukhbindersingh@gmail.com>
Co-authored-by: Simon Willison <swillison@gmail.com>
2024-12-01 15:57:24 -08:00
Simon Willison cfb10f4afd Log input tokens, output tokens and token details (#642)
* Store input_tokens, output_tokens, token_details on Response, closes #610
* llm prompt -u/--usage option
* llm logs -u/--usage option
* Docs on tracking token usage in plugins
* OpenAI default plugin logs usage
2024-11-19 20:21:59 -08:00
Simon Willison 4a059d722b Log --async responses to DB, closes #641
Refs #507
2024-11-19 18:11:52 -08:00
Simon Willison ba75c674cb llm.get_async_model(), llm.AsyncModel base class and OpenAI async models (#613)
- https://github.com/simonw/llm/issues/507#issuecomment-2458639308

* register_model is now async aware

Refs https://github.com/simonw/llm/issues/507#issuecomment-2458658134

* Refactor Chat and AsyncChat to use _Shared base class

Refs https://github.com/simonw/llm/issues/507#issuecomment-2458692338

* fixed function name

* Fix for infinite loop

* Applied Black

* Ran cog

* Applied Black

* Add Response.from_row() classmethod back again

It does not matter that this is a blocking call, since it is a classmethod

* Made mypy happy with llm/models.py

* mypy fixes for openai_models.py

I am unhappy with this, had to duplicate some code.

* First test for AsyncModel

* Still have not quite got this working

* Fix for not loading plugins during tests, refs #626

* audio/wav not audio/wave, refs #603

* Black and mypy and ruff all happy

* Refactor to avoid generics

* Removed obsolete response() method

* Support text = await async_mock_model.prompt("hello")

* Initial docs for llm.get_async_model() and await model.prompt()

Refs #507

* Initial async model plugin creation docs

* duration_ms ANY to pass test

* llm models --async option

Refs https://github.com/simonw/llm/pull/613#issuecomment-2474724406

* Removed obsolete TypeVars

* Expanded register_models() docs for async

* await model.prompt() now returns AsyncResponse

Refs https://github.com/simonw/llm/pull/613#issuecomment-2475157822

---------

Co-authored-by: github-actions[bot] <github-actions[bot]@users.noreply.github.com>
2024-11-13 17:51:00 -08:00
Simon Willison bc96e1c739 Ruff fix for #626 2024-11-13 06:37:31 -08:00
Simon Willison 330e171e86 Fix for not loading plugins during tests, refs #626 2024-11-12 21:42:49 -08:00
Simon Willison dff53a9cae Better --help for llm keys get, refs #623 2024-11-11 09:53:24 -08:00
Simon Willison 561784df6e llm keys get command, refs #623 2024-11-11 09:47:13 -08:00
Simon Willison 5d1d723d4b Special case treat audio/wave as audio/wav, closes #603 2024-11-07 17:13:54 -08:00
Simon Willison 98d2c19876 Promote alternative model providers in llm --help 2024-11-06 06:38:53 -08:00
Simon Willison 7146fe82d1 Back to the deprecated Pydantic thing to get tests passing
Refs #612
2024-11-05 23:03:31 -08:00
Simon Willison 12df1a3b2a Show attachment types in llm models --options, closes #612 2024-11-05 22:49:26 -08:00
Simon Willison 0cc4072bcd Support attachments without prompts, closes #611 2024-11-05 21:27:18 -08:00
Simon Willison f0ed54abf1 Docs for CLI attachments, refs #587 2024-10-28 15:41:34 -07:00
Simon Willison 286cf9fcd9 attachments= keyword argument, tests pass again - refs #587 2024-10-28 15:41:34 -07:00
Simon Willison a9bc1c7329 llm logs --json for attachments, refs #587 2024-10-28 15:41:34 -07:00
Simon Willison 2384fd52a0 Tighter log output for binary content 2024-10-28 15:41:34 -07:00
Simon Willison bb5b802d4f llm logs markdown support for attachments, refs #587 2024-10-28 15:41:34 -07:00
Simon Willison dff5b456fd Got llm --continue to work with images, refs #587 2024-10-28 15:41:34 -07:00
Simon Willison c0fe719df6 Store prompt attachments in attachments and prompt_attachments tables
Refs https://github.com/simonw/llm/issues/587#issuecomment-2439791231
2024-10-28 15:41:34 -07:00
Simon Willison 6df00f92ff First working prototype of new attachments feature, refs #587 2024-10-28 15:41:34 -07:00
Simon Willison 6deed8f976 get_model() improvement, get_default_model() / set_default_wodel() now documented
Refs #553
2024-08-18 17:37:31 -07:00
Simon Willison 0ef5037b00 Remove obsolete DEFAULT_EMBEDDING_MODEL, closes #537 2024-07-18 12:22:17 -07:00
Simon Willison a83421607a Switch default model to gpt-4o-mini (from gpt-3.5-turbo), refs #536 2024-07-18 11:57:19 -07:00
Simon Willison 964f4d9934 Fix for llm logs -q plus -m bug, closes #515 2024-06-16 14:35:38 -07:00
Simon Willison fb63c92cd2 llm logs -r/--response option, closes #431 2024-03-04 13:29:07 -08:00
Simon Willison 9119b03a07 Chmod 600 keys.json on creation, refs #351 2024-01-26 13:18:13 -08:00
Simon Willison 1a4853d80e left/right arrow key readline bindings in chat, closes #376 2024-01-26 13:12:55 -08:00
Simon Willison 214fcaaf86 Upgrade to run against OpenAI >= 1.0
* strategy: fail-fast: false - to help see all errors
* Apply latest Black

Refs #325
2024-01-25 22:00:44 -08:00
Simon Willison 5cc8efe98f readline support for llm chat, closes #355 2023-11-22 16:55:54 -08:00
Simon Willison 8b78ac6099 Fix for bug where embed did not use default model, closes #317 2023-10-31 21:19:59 -07:00
e. alvarez 839b4d7161 Fix issues: #274, #280 (#282)
* Fix issue with reading directories in `iterate_files()` (#280)
* Add directory checking logic in `iterate_files()` (#274)
* Added tests for #282, #274, #280

---------

Co-authored-by: Simon Willison <swillison@gmail.com>
2023-09-18 23:14:30 -07:00
Simon Willison 9c7792dce5 Removed rogue raise 2023-09-18 22:05:00 -07:00
Simon Willison 4fea46113f logprobs support for OpenAI completion models, refs #284 2023-09-18 22:04:28 -07:00
Simon Willison 33dee4762e llm embed-multi --batch-size option, closes #273 2023-09-13 16:33:27 -07:00
Simon Willison b9478e6a17 batch_size= argument to embed_multi(), refs #273 2023-09-13 16:24:04 -07:00
Simon Willison 4952a8d119 llm similar --binary, closes #269 2023-09-12 11:23:31 -07:00
Simon Willison 9e529bb36a Wrap !multi in single quotes, for consistency with exit/quit 2023-09-12 10:45:02 -07:00
Simon Willison 9c33d30843 llm chat !multi support, closes #267 2023-09-12 09:31:20 -07:00
Simon Willison 22a59f795e llm collections defaults to llm collections list, close #265 2023-09-11 23:08:11 -07:00
Simon Willison 52cec1304b Binary embeddings (#254)
* Binary embeddings support, refs #253
* Write binary content to content_blob, with tests - refs #253
* supports_text and supports_binary embedding validation, refs #253
2023-09-11 18:58:44 -07:00
mhalle f07b291a78 change type of cli "embed -c" to be STR not FILE (#263)
* change type of cli "embed -c" to be STR not FILE

The type of the "content" argument was erroneously set to be FILE with a "type=click.Path(...)" argument, which content is a string.

---------

Co-authored-by: Simon Willison <swillison@gmail.com>
2023-09-11 18:17:34 -07:00
Simon Willison 5ba34dbe36 llm embed-db is now llm collections, refs #229 2023-09-10 14:24:27 -07:00
Simon Willison 6012f31e36 llm plugins --all option, closes #259 2023-09-10 14:18:16 -07:00
Alexis Métaireau df32d7685d Updated error message for invalid or missing embedding model (#257)
* Updated error message for missing embedding model

---------

Co-authored-by: Simon Willison <swillison@gmail.com>
2023-09-10 11:56:29 -07:00
Simon Willison 5912bd47c0 --no-stream option for llm chat, closes #248 2023-09-10 11:47:38 -07:00
Simon Willison ae7f4f6de7 llm chat -o/--option - refs #244 2023-09-10 11:14:28 -07:00
Simon Willison 17e6402908 Fix bug with llm chat -c and API keys, closes #247 2023-09-05 19:26:21 -07:00
Simon Willison 5495112d9f Initial tests for llm chat, refs #231 2023-09-04 23:36:25 -07:00
Simon Willison 969a5d3364 Initial prototype of llm chat, refs #231 2023-09-04 23:36:25 -07:00