llama.cpp b10999

Change max context length for auto-fitting with unified KV ( #28849 ) Website: https://llama.app Attestations:…

Change max context length for auto-fitting with unified KV ( #28849 ) Website: https://llama.app Attestations: https://github.com/ggml-org/llama.cpp/attestations/47906437 macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel…

Read the original source — github.com

release · Shared by tscosj

0 comments

No comments yet.