61llama.cpp Release b10287 Fixes Unlimited-OCR Converter BugGITHUB · github-actions[bot] · github.com · Aug 05, 05:04 PM29d
62llama.cpp Releases Version b10286 with Grammar Repetition AdjustmentGITHUB · github-actions[bot] · github.com · Aug 05, 04:23 PM29d
63llama.cpp Release b10284 Fixes Memory Allocation for Multi-Token Prediction LayersGITHUB · github-actions[bot] · github.com · Aug 05, 02:37 PM29d
64Qwen3-TTS Voice Cloning Support Merged into Mainline llama.cppREDDIT · /u/BTA_Labs · reddit.com · Aug 05, 07:47 AM30d
65llama.cpp Releases Version b10276 with Minor Security UpdateGITHUB · github-actions[bot] · github.com · Aug 05, 12:21 AM30d
66llama.cpp b10273 Release Fixes Sampler Initialization in Backend SamplingGITHUB · github-actions[bot] · github.com · Aug 04, 08:31 PM30d
67llama.cpp Release b10271 Introduces UI Improvements for Agent Working DirectoriesGITHUB · github-actions[bot] · github.com · Aug 04, 06:56 PM30d
68llama.cpp Release b10270 Adds Support for Qwen3-TTS Voice CloningGITHUB · github-actions[bot] · github.com · Aug 04, 06:03 PM30d
69llama.cpp b10251 Adds Multi-Token Prediction Support for GLM-4.7-FlashGITHUB · github-actions[bot] · github.com · Aug 04, 03:06 AM31d
70Llama.cpp Releases Official macOS App and Simplified 'llama serve' CommandREDDIT · /u/rm-rf-rm · reddit.com · Aug 02, 08:44 PM32d
71llama.cpp Release b10231 Introduces Support for DSpark Speculative SidecarsGITHUB · github-actions[bot] · github.com · Aug 02, 06:14 PM32d
72llama.cpp Release b10224 Adds WebGPU f16 Repeat SupportGITHUB · github-actions[bot] · github.com · Aug 02, 07:03 AM33d
73llama.cpp b10219 Release Persists Reasoning Content in CLI Chat HistoryGITHUB · github-actions[bot] · github.com · Aug 01, 04:46 PM33d
74llama.cpp b10217 Released with Tool Calling in Thinking Phase for DS4 ModelsGITHUB · github-actions[bot] · github.com · Aug 01, 06:50 AM34d
75llama.cpp b10215 Released with Vulkan Driver Checks for Intel GPUsGITHUB · github-actions[bot] · github.com · Jul 31, 09:31 PM34d
76llama.cpp Release b10205 Introduces AMD ZenDNN Backend OptimizationsGITHUB · github-actions[bot] · github.com · Jul 31, 01:26 PM34d
77llama.cpp Release b10199 Adds Input Embedding Support for Token GenerationGITHUB · github-actions[bot] · github.com · Jul 30, 08:33 PM35d
78llama.cpp Releases Build b10196 with Async Copy Sync FixGITHUB · github-actions[bot] · github.com · Jul 30, 06:29 PM35d
79llama.cpp Releases Build b10197 with Alternative Convolution Layout Test SupportGITHUB · github-actions[bot] · github.com · Jul 30, 06:58 PM35d
80llama.cpp Release b10188 Fixes Metal Memory Leak on macOSGITHUB · github-actions[bot] · github.com · Jul 30, 08:56 AM36d