Back to AI intel
重点

llama.cpp Release b11390 Fixes CUDA MMQ Memory Fault

AI intel briefing

Core summary

One sentence to understand this update

llama.cpp's b11390 release addresses a critical CUDA MMQ memory fault that occurred when `n_expert` was significantly larger than `n_ubatch`.

Impact & opportunity

What this could mean

This fix improves the stability and reliability of llama.cpp for users running models on CUDA, especially those with varying expert and ubatch configurations.