docs: document self-hosted v0.0.5 model mixing bug and v0.0.7 resolution (#1450) - #1606
Merged
Dhravya merged 1 commit intoAug 27, 2026
Conversation
This file contains hidden or bidirectional Unicode text that may be interpreted or compiled differently than what appears below. To review, open the file in an editor that reveals hidden Unicode characters.
Learn more about bidirectional Unicode characters
Sign up for free
to join this conversation on GitHub.
Already have an account?
Sign in to comment
Add this suggestion to a batch that can be applied as a single commit.This suggestion is invalid because no changes were made to the code.Suggestions cannot be applied while the pull request is closed.Suggestions cannot be applied while viewing a subset of changes.Only one suggestion per line can be applied in a batch.Add this suggestion to a batch that can be applied as a single commit.Applying suggestions on deleted lines is not supported.You must change the existing code in this line in order to create a valid suggestion.Outdated suggestions cannot be applied.This suggestion has been applied or marked resolved.Suggestions cannot be applied from pending reviews.Suggestions cannot be applied on multi-line comments.Suggestions cannot be applied while the pull request is queued to merge.Suggestion cannot be applied right now. Please check back later.
Resolves #1450.
Description
Documents the self-hosted
v0.0.5embedding model mixing bug in the embeddings documentation, explaining the root cause of why Japanese memories query and profile retrieval returned empty results, and notes that upgrading tov0.0.7resolves the issue by enforcing a locked embedding plan.Root Cause
In version
v0.0.5, the embedding configuration was not locked. The server could mix different embedding models between document write paths (e.g. OpenAItext-embedding-3-smallwith 1536 dimensions) and memory query paths (e.g. local English defaultbge-base-en-v1.5), causing vector search cosine similarities to drop to near-zero.{"results":[],"total":0}because both the vector path (due to model mixing) and the lexical path (due to tokenization mismatch) failed.Resolution
This was resolved in
v0.0.7by locking the embedding plan uniformly across all write and query paths.