11 Commits (94ab428e3f77fdd9d9c833b369bb40980c65049a)

Author SHA1 Message Date
Jesse Gross 94ab428e3f ggml: Seperate tensor load from backend creation 1 year ago
Daniel Hiltgen ff80718e9c
fix crash in old clients with quantization progress (#10710) 1 year ago
Bruce MacDonald ad035ad595
convert: quantize from safetensors needs kv (#10675) 1 year ago
Daniel Hiltgen 424810450f
Move quantization to new backend (#10363) 1 year ago
Michael Yang 340448d2d1 explicitly decode maxarraysize 1024 1 year ago
Michael Yang 88738b357b create tempdir in models directory 1 year ago
Bruce MacDonald bebb6823c0
server: validate local path on safetensor create (#9379) 1 year ago
Michael Yang 58245413f4
next ollama runner (#7913) 2 years ago
Patrick Devine 2539f2dbf9
Fix absolute path names + gguf detection (#8428) 2 years ago
Patrick Devine 8bccae4f92
show a more descriptive error in the client if it is newer than the server (#8351) 2 years ago
Patrick Devine 86a622cbdc
Update the /api/create endpoint to use JSON (#7935) 2 years ago