01llama.cpp Release b10794 Refactors SYCL MKL FlashAttention FlagGITHUB · github-actions[bot] · github.com · Sep 04, 03:20 AM7h
02llama.cpp Release b10795 Adds SYCL Kernel Fusion OptimizationsGITHUB · github-actions[bot] · github.com · Sep 04, 04:30 AM7h
03llama.cpp Release b10793 Fixes Build System Incremental Compilation IssueGITHUB · github-actions[bot] · github.com · Sep 03, 10:18 PM13h
04llama.cpp Release b10790 Tunes CUDA Kernel Crossover for NVIDIA Orin (SM87)GITHUB · github-actions[bot] · github.com · Sep 03, 07:52 PM15h
05Local AI Demonstrates Superior Reliability During Major Cloud OutagesREDDIT · /u/jacek2023 · reddit.com · Sep 03, 03:20 PM20h
06Real-Time Ngram Knowledge Injector for Qwen Models Created for llama.cppREDDIT · /u/ortegaalfredo · reddit.com · Sep 03, 11:40 AM1d
07llama.cpp Adds Support for NVIDIA's Nemotron-3-Puzzle-75B Hybrid ModelREDDIT · /u/jacek2023 · reddit.com · Sep 03, 07:35 AM1d
08llama.cpp Release b10772 Adds F16 ABS Unary Operation to Qualcomm Hexagon BackendGITHUB · github-actions[bot] · github.com · Sep 03, 12:48 AM1d
09llama.cpp Release b10771 Refactors Multimodal Tokenization APIGITHUB · github-actions[bot] · github.com · Sep 03, 12:24 AM1d
10llama.cpp Release b10769 Fixes Apple Metal Low-Memory Query BugGITHUB · github-actions[bot] · github.com · Sep 02, 11:37 PM1d
11llama.cpp Release b10767 Updates ROCm Support to Version 10.0.0GITHUB · github-actions[bot] · github.com · Sep 02, 10:59 PM1d
12llama.cpp Release b10766 Fixes Vision Input Support for DeepSeek V4GITHUB · github-actions[bot] · github.com · Sep 02, 08:33 PM1d
13llama.cpp b10762 Adds Support for DeepSeek-V4-Flash-Vision-Exp ModelGITHUB · github-actions[bot] · github.com · Sep 02, 06:37 PM1d
14llama.cpp Release b10758 Brings Hexagon DSP Matrix Fusion and Memory OptimizationsGITHUB · github-actions[bot] · github.com · Sep 02, 09:29 AM2d
15llama.cpp b10759 Released with KleidiAI Buffer Initialization FixGITHUB · github-actions[bot] · github.com · Sep 02, 10:22 AM2d
16Google's Android Studio Uses llama.cpp to Power Native Local Gemma ModelsREDDIT · /u/DrBattletoad · reddit.com · Sep 02, 07:23 AM2d
17llama.cpp Release b10752 Adds Metal Library Support for XCFrameworksGITHUB · github-actions[bot] · github.com · Sep 02, 12:13 AM2d
18llama.cpp Release b10743 Introduces Metal Optimizations for Apple M2 ProGITHUB · github-actions[bot] · github.com · Sep 01, 06:18 PM2d
19llama.cpp Release b10749 Adds Context Autoscaling for YaRN ScalingGITHUB · github-actions[bot] · github.com · Sep 01, 06:57 PM2d
20llama.cpp b10739 Released with Metal Flash-Attention Vector Tuning for M2 MaxGITHUB · github-actions[bot] · github.com · Sep 01, 03:15 PM2d