This website requires JavaScript.
Explore
Help
Register
Sign In
wylab
/
llama.cpp
Watch
0
Star
0
Fork
1
mirror of
https://github.com/ggml-org/llama.cpp.git
synced
2026-09-01 01:27:50 +02:00
Code
Issues
Packages
Projects
Releases
Wiki
Activity
Files
tmp-q4
llama.cpp
/
include
T
Add File
New File
Upload File
Apply Patch
Copy Permalink
Download directory as ZIP
Download directory as TAR.GZ
History
Xuan-Son Nguyen
and
GitHub
732707dff2
quantize: cap working memory size to avoid loading big tensors onto RAM (
#27795
)
2026-08-27 18:31:13 +02:00
..
llama-cpp.h
llama : re-enable manual LoRA adapter free (
#19983
)
2026-03-18 12:03:26 +02:00
llama.h
quantize: cap working memory size to avoid loading big tensors onto RAM (
#27795
)
2026-08-27 18:31:13 +02:00