Back to AI intel
重点
llama.cpp Release b11390 Fixes CUDA MMQ Memory Fault
AI intel briefing
Core summary
One sentence to understand this update
llama.cpp's b11390 release addresses a critical CUDA MMQ memory fault that occurred when `n_expert` was significantly larger than `n_ubatch`.
Impact & opportunity
What this could mean
This fix improves the stability and reliability of llama.cpp for users running models on CUDA, especially those with varying expert and ubatch configurations.
Source
View original