Add precompiled llama-quantize binaries for flux Just-in-time GGUF quantization

These binaries are built from https://github.com/pollockjj/llama.cpp/tree/flux-quant-b3600

Linux binary (SHA256: 846c0bab3c7f7c6729f22b6229a2db5da2da63a4549ef8bf76546092f56fec15):
- Ubuntu 22.04
- gcc 13.3.0
- Debug build with flags: --config Debug -j10 --target llama-quantize

Windows binary (SHA256: 121c02184fcf30dc4ff3dcf4a69177e44aced9d7ba8c77f49cd138c2d73f7e6b):
- Windows 11
- MSVC 19.42.34436.0
- Debug build with flags: -G "Visual Studio 17 2022" -A x64 -DBUILD_SHARED_LIBS=OFF

Binaries are verified and released at:
https://github.com/pollockjj/llama.cpp/releases/tag/1.0.0
This commit is contained in:
John Pollock
2025-02-03 00:15:18 -06:00
parent df779bf820
commit 262ceea716
4 changed files with 12 additions and 0 deletions
+12
View File
@@ -1,2 +1,14 @@
# Python and IDE
__pycache__/
.vscode/settings.json
# Binary management
binaries/*
binaries/win64/*
binaries/linux/*
# Keep directory structure and rename file for Linux
!binaries/
!binaries/win64/
!binaries/linux/
!binaries/linux/rename_this_to_keep_user_binary.txt
Binary file not shown.
Binary file not shown.