Files
discourse-ai/spec/shared
Rafael dos Santos Silvaandwozulong 3b8f900486 FIX: Handle unicode on tokenizer (#515)
* FIX: Handle unicode on tokenizer

Our fast track code broke when strings had characters who are longer in tokens than
in UTF-8.

Admins can set `DISCOURSE_AI_STRICT_TOKEN_COUNTING: true` in app.yml to ensure token counting is strict, even if slower.


Co-authored-by: wozulong <[email protected]>
2024-03-14 17:33:30 -03:00
..