Back to AI intel
重点
llama.cpp Release b9969: Vulkan Improvements & Long Prompt Fix
AI intel briefing
Core summary
One sentence to understand this update
The latest llama.cpp release b9969 improves Vulkan performance on Adreno GPUs and fixes `llama-cli` crashing with long prompts for q4_0 quantized networks.
Impact & opportunity
What this could mean
Users running llama.cpp on Vulkan-enabled Adreno GPUs will experience better performance, while all users will benefit from increased stability with longer prompts in `llama-cli`.
Source
View original