Georgi Gerganov
|
f1c9df5806
|
metal : sync ggml-metal (ref #1047)
|
2023-06-25 15:40:39 +03:00 |
|
Georgi Gerganov
|
6c25fae1c4
|
opencl : sync latest ggml-opencl
|
2023-06-25 15:38:30 +03:00 |
|
Georgi Gerganov
|
2b6a074305
|
extra : update ggml sync script
|
2023-05-14 10:01:52 +03:00 |
|
Georgi Gerganov
|
0bcb64b184
|
ggml : sync ggml (clBLAST + tensor names)
|
2023-05-02 21:24:18 +03:00 |
|
Georgi Gerganov
|
794b162a46
|
whisper : add integer quantization support (#540)
* whisper : add integer quantization support
* examples : add common-ggml + prepare to add "quantize" tool
* whisper : quantization tool ready
* whisper : fix F32 support
* whisper : try to fix shared lib linkage
* wasm : update quantized models to Q5
* bench.wasm : remove "medium" button
* bench.wasm : fix custom model button
* ggml : add Q5_0 and Q5_1 WASM SIMD
* wasm : add quantized models to all WASM examples
* wasm : bump DB version number to 2
* talk-llama : update example to latest llama.cpp
* node : increase test timeout to 10s
* readme : add information for model quantization
* wasm : add links to other examples
|
2023-04-30 18:51:57 +03:00 |
|
Georgi Gerganov
|
3eaeb030ff
|
extra : add sync-ggml.sh script
|
2023-04-29 12:32:28 +03:00 |
|