Skip to content

Implement '--keep-split' to quantize model into several shards #10947

Implement '--keep-split' to quantize model into several shards

Implement '--keep-split' to quantize model into several shards #10947

windows-latest-cmake (noavx, -DLLAMA_NATIVE=OFF -DLLAMA_BUILD_SERVER=ON -DLLAMA_AVX=OFF -DLLAMA_A...

succeeded Apr 23, 2024 in 5m 42s