mirror of
https://github.com/tiennm99/DocsGPT.git
synced 2026-10-05 16:13:51 +00:00
The chunker caches only a model's tokenizer.json in the embeddings cache. FastEmbed counts any cached snapshot as the model, so it never downloaded the ONNX graph and every load failed with NO_SUCHFILE. The loader now checks the files FastEmbed needs and fetches them first when they are missing; offline it leaves them for FastEmbed to report.