Logo
Explore Help
Sign In
mikhail/llama.cpp-mtp-turboquant
1
0
Fork 0
You've already forked llama.cpp-mtp-turboquant
Code Issues Pull Requests Actions 30 Packages Projects Releases Wiki Activity
Files
58062860afb88e555857c1266d3a17e1b65b5eb9
llama.cpp-mtp-turboquant/ggml
T
History
Aadeshveer Singh 58062860af ggml : use WARP_SIZE/2 for argmax reduction offset (#18092)
2025-12-17 11:47:01 +08:00
..
cmake
ggml: Skip backend library linking code when GGML_BACKEND_DL=ON (#15094)
2025-08-07 13:45:41 +02:00
include
llama: automatically set parameters not set by the user in such a way that maximizes GPU utilization (#16653)
2025-12-15 09:24:59 +01:00
src
ggml : use WARP_SIZE/2 for argmax reduction offset (#18092)
2025-12-17 11:47:01 +08:00
.gitignore
vulkan : cmake integration (#8119)
2024-07-13 18:12:39 +02:00
CMakeLists.txt
cmake : set CMAKE_RUNTIME_OUTPUT_DIRECTORY for non standalone build (ggml/1394)
2025-12-14 08:33:51 +02:00
Powered by Gitea Version: 1.26.1 Page: 152ms Template: 1ms
Auto
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API