Models/shipped/2026-09-30

llama.cpp integrates support for GLM-5.3-Flash and Qwen4Exp models

The `llama.cpp` open-source project has expanded its model compatibility by integrating support for GLM-5.3-Flash (GLM5-Next) and Qwen4Exp models. These additions were reported by the r/LocalLLaMA community, referencing specific pull requests on the `ggml-org/llama.cpp` GitHub repository. The integration of these models is confirmed by the existence of the pull requests and associated commit activity.

3 articles from 1 outlet covered this story. The underlying claim is sourced from a shipped.

What do all outlets agree on?

1 outlet covered “llama.cpp integrates support for GLM-5.3-Flash and Qwen4Exp models”. All of them report the following:

  • llama.cpp repository received updates
  • Support for GLM-5.3-Flash (GLM5-Next) was added via Pull Request #27773
  • Support for Qwen4Exp with MTP was added via Pull Request #29761
  • Updates are associated with specific pull requests on ggml-org/llama.cpp
  • Reported by the r/LocalLLaMA community

Which outlets covered this?

All 3 articles found on this story, grouped by the stance of the piece. Every link goes to the original publisher.

What related stories are there?

Get the week in AI in one email

What happened, which outlets reported it, and where their coverage differed. One issue a week.

The first issue hasn’t gone out yet. Subscribe and it’s the one you’ll get.

We’ll send the digest and nothing else. One-click unsubscribe. Privacy.