You can not select more than 25 topics Topics must start with a letter or number, can include dashes ('-') and can be up to 35 characters long.
 
 
 
 
 
 
Daniel Hiltgen 1c6669e64c
Re-remove cuda v11 (#10694)
1 year ago
..
0001-ggml-backend-malloc-and-free-using-the-same-compiler.patch llama: update to commit de4c07f93 (#10655) 1 year ago
0002-pretokenizer.patch llama: update to commit de4c07f93 (#10655) 1 year ago
0003-embeddings.patch llama: update to commit de4c07f93 (#10655) 1 year ago
0004-clip-unicode.patch llama: update to commit de4c07f93 (#10655) 1 year ago
0005-solar-pro.patch chore: update mllama to use ollama engine (#10637) 1 year ago
0006-fix-deepseek-deseret-regex.patch chore: update mllama to use ollama engine (#10637) 1 year ago
0007-maintain-ordering-for-rules-for-grammar.patch chore: update mllama to use ollama engine (#10637) 1 year ago
0008-ensure-KV-cache-is-fully-defragmented.patch chore: update mllama to use ollama engine (#10637) 1 year ago
0009-sort-devices-by-score.patch chore: update mllama to use ollama engine (#10637) 1 year ago
0010-add-phony-target-ggml-cpu-for-all-cpu-variants.patch chore: update mllama to use ollama engine (#10637) 1 year ago
0011-remove-amx.patch chore: update mllama to use ollama engine (#10637) 1 year ago
0012-fix-string-arr-kv-loading.patch chore: update mllama to use ollama engine (#10637) 1 year ago
0013-ollama-debug-tensor.patch chore: update mllama to use ollama engine (#10637) 1 year ago
0014-add-ollama-vocab-for-grammar-support.patch chore: update mllama to use ollama engine (#10637) 1 year ago
0015-add-argsort-and-cuda-copy-for-i32.patch model: add Qwen2.5-VL support (#10385) 1 year ago
0016-graph-memory-reporting-on-failure.patch ggml: Report graph memory for failed allocations 1 year ago
0017-ggml-Export-GPU-UUIDs.patch Revert "Revert "ggml: Export GPU UUIDs" (#11115)" (#11117) 1 year ago
0018-temporary-prevent-rocm-cuda-mixed-loading.patch Re-remove cuda v11 (#10694) 1 year ago