ggml-org/llama.cpp
feature request - disabling tokenizer in conversion / inference
已关闭
#1,765 创建于 2023年6月8日
enhancementgood first issuehelp wanted
仓库指标
- 星标
- (124,043 个星标)
- PR 合并指标
- (平均合并 6天 8小时) (30 天内合并 389 个 PR)
描述
in #1764 i asked if it'd be possible to add a Huggingface tokenizer. but - HF tokenizers are quite flexible and officially supporting them in llama.cpp (or ggml?) might be a lot of hassle.
a much easier workaround would be allowing to disable tokenizers in both model conversion and inference. this means the users are supposed to encode(text)/decode(ids) in their implementation for using llama.cpp. in my case, for example, i'll use a python GUI and a wrapper anyway.
i'd like to work on it, but honestly i don't think i understand enough to be able to do this. i'd appreciate very much if anyone's interested in it.