发布

  • model: align Laguna with upstream llama.cpp (#17335)

    frostbyte_neo 发布于 2026-07-23 00:09:18 +00:00

    Update llama.cpp to pick up upstream Laguna implementation and remove Ollama's local Laguna implementation. Retain a narrow Metal-only scaling workaround for routed-MoE prompt overflow.

    Translate older Ollama GGUF attention-gate and SWA metadata names so existing models continue to load.

    下载附件