This website requires JavaScript.
Explore
Help
Register
Sign In
marfrit
/
rk-llama.cpp
Watch
1
Star
0
Fork
0
You've already forked rk-llama.cpp
Code
Issues
Pull Requests
Actions
34
Packages
Projects
Releases
Wiki
Activity
Files
c1e79e610fd28f2c3923539fee9313734bbf8cfa
rk-llama.cpp
/
ggml
/
src
T
History
Georgi Gerganov
0a57271ab6
CUDA : fix unused argument when USE_CUDA_GRAPH=OFF (
#18800
)
2026-01-13 12:25:53 +02:00
..
ggml-blas
cmake : update blas logic (
#18205
)
2026-01-10 18:00:54 +02:00
ggml-cann
…
ggml-cpu
…
ggml-cuda
CUDA : fix unused argument when USE_CUDA_GRAPH=OFF (
#18800
)
2026-01-13 12:25:53 +02:00
ggml-hexagon
…
ggml-hip
…
ggml-metal
metal : add MoE kernel specialization for ne20=5 (
#18667
)
2026-01-08 12:37:45 +02:00
ggml-musa
…
ggml-opencl
opencl: add SOFTPLUS op support (
#18726
)
2026-01-10 21:57:44 -08:00
ggml-rpc
…
ggml-sycl
…
ggml-vulkan
vulkan: change memory_logger to be controlled by an env var (
#18769
)
2026-01-12 13:32:55 +01:00
ggml-webgpu
Updates to webgpu get_memory (
#18707
)
2026-01-09 08:17:18 -08:00
ggml-zdnn
…
ggml-zendnn
…
CMakeLists.txt
…
ggml-alloc.c
…
ggml-backend-impl.h
llama: use host memory if device reports 0 memory (
#18587
)
2026-01-09 05:34:56 +08:00
ggml-backend-reg.cpp
…
ggml-backend.cpp
…
ggml-common.h
…
ggml-impl.h
…
ggml-opt.cpp
…
ggml-quants.c
…
ggml-quants.h
…
ggml-threading.cpp
…
ggml-threading.h
…
ggml.c
…
ggml.cpp
…
gguf.cpp
…