mirror of
https://github.com/CopilotKit/CopilotKit.git
synced 2026-09-14 16:26:20 +08:00
6d5d407624
The bundled tei embedder image is amd64-only; under arm64 emulation the Candle backend is unavailable and TEI falls back to the ONNX/ORT backend, which needs onnx/model.onnx files Qwen3-Embedding-0.6B doesn't publish (404) -> crash-loop. A fresh clone on Apple Silicon therefore couldn't stand up the embedder, so memory save/recall were dead. Ports the proven pattern from the Intelligence repo's docker-compose.deps.yml + demos/splat-demo/run-demo.sh into this demo: - docker-compose.yml: gate the bundled `tei` behind the `cpu-fallback` profile, so a bare `docker compose up` skips the crash-looping emulated image. amd64/CI opt back in with `--profile cpu-fallback`. (intelligence's tei dep is required:false, so it starts fine without it, using MEMORY_EMBEDDINGS_URL.) - run-demo.sh: one-command cold start. On Apple Silicon it runs a native Metal TEI on :7067 (same 1.9.3 + Qwen3-Embedding-0.6B => byte-identical embeddings, ~20x faster) and points app-api at it; on amd64/CI it uses the docker tei via the profile. Mints a dev license if .env lacks one, then starts `pnpm dev`. - README: correct the failure description (emulation->ONNX crash-loop, not OOM), document run-demo.sh as the recommended start, and the profile-gated manual path. All CopilotKit-repo-only (banking's compose is standalone); no Intelligence changes. Validated: shellcheck clean, compose valid, bare `up` skips tei and keeps intelligence healthy, memory save/recall verified through the native TEI. Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>
10 KiB
10 KiB