This website requires JavaScript.
Explore
Help
Sign In
mikhail
/
llama.cpp-mtp-turboquant
Watch
1
Star
0
Fork
0
You've already forked llama.cpp-mtp-turboquant
Code
Issues
Pull Requests
Actions
27
Packages
Projects
Releases
Wiki
Activity
Files
84ab83cc0b4b7e769451ee48e4c7d1acef91ef25
llama.cpp-mtp-turboquant
/
ggml
/
src
/
ggml-musa
T
History
Johannes Gäßler
7a6e91ad26
CUDA: replace GGML_CUDA_F16 with CUDA arch checks (
#15433
)
2025-08-20 16:58:49 +02:00
..
CMakeLists.txt
CUDA: replace GGML_CUDA_F16 with CUDA arch checks (
#15433
)
2025-08-20 16:58:49 +02:00
mudnn.cu
musa: Upgrade MUSA SDK version to rc4.0.1 and use mudnn::Unary::IDENTITY op to accelerate D2D memory copy (
#13647
)
2025-05-21 09:58:49 +08:00
mudnn.cuh
musa: enable fp16 mma (all) and cublas on qy2 (
#13842
)
2025-06-26 12:11:59 +08:00