Reverse the earlier 'default to LiteLLM' recommendation: per user intent, cognee runs its LLM on the Kimi subscription via cognee-llm (the reason the wrapper exists). Gate-5 latency is an accepted tradeoff; LiteLLM stays a documented fallback. Embeddings remain on LiteLLM nomic-embed. Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com> Claude-Session: https://claude.ai/code/session_01LeqyaxJF2nbRXJtae2kNB2