Live page ยท Day archive

ggml-org/llama.cpp v0.6.0

Read the original at GitHub Releases

Summary

llama.cpp v0.6.0 introduces an extended batch API, adds support for the 320B GLM-5.3-Flash model, and updates the Web UI with a Hugging Face Hub data layer.

Carried by: GitHub Releases. First seen: .