backend/cpp/bonsai/patches/README.md
The bonsai backend reuses backend/cpp/llama-cpp/grpc-server.cpp (written against
LocalAI's pinned upstream llama.cpp) but compiles it against the PrismML prism fork,
which branched from upstream some commits earlier. Any upstream API change that the shared
gRPC server depends on, but that the fork does not yet carry, is back-ported here as a
*.patch file and applied to the cloned fork checkout by ../apply-patches.sh.
CI treats both this directory and backend/cpp/llama-cpp/ as Bonsai inputs, since
the wrapper copies and builds the shared llama.cpp backend sources.
Rules:
NNNN-short-description.patch.git apply from the fork's checkout root.apply-patches.sh fails fast if a patch stops applying cleanly — that is the signal the
fork has caught up (or diverged), so re-cut or drop the patch.