ggml-org/llama.cpp adds Longcat-Flash support for testing; PR 19182 seeks tests on larger models with GGUF excerpt from HuggingFace
Read the original at old.reddit.com→This PR should be ready for testing now. I tested with a very small (8B params) sub-model extracted from the original one. Appreciate if someone can test with the bigger model. GGUF(for testing) from PR: (Please...
Original headline: "model: support Longcat-Flash (need testing) by ngxson · Pull Request #19182 · ggml-org/llama.cpp"
Coverage timeline
- Aug 8, 07:28 UTC r/LocalLLaMA lead source model: support Longcat-Flash (need testing) by ngxson · Pull Request #19182 · ggml-org/llama.cpp