/
← Accept All   週ごとのアーカイブ
llama.cpp releasesTools

v0.4.0

9月4日
v0.4.0

Overview llama.cpp 0.4.0 adds initial Qwen3.8-Flash-Next and Nemotron-3-Puzzle support, on-demand tensor reading, per-slot server context limits, video input options, and a ggml update to 0.23.0 with major sparse flash a

Models & releases
llama.cpp releasesで読む ↗

関連する記事

llama.cpp releasesの他の記事