4153 Commits (parth/python-tools-calling)
 

Author SHA1 Message Date
Michael Yang 26a26998fb
Merge pull request #9590 from ollama/mxyng/dump-pad 1 year ago
Michael Yang 9926eae015 fix: pad tensor item if ge zero 1 year ago
Vincent Koc 8585b7b151
docs: add opik to observability integrations (#9626) 1 year ago
Parth Sareen 7e34f4fbfa
sample: add numerical stability to temperature/softmax transform (#9631) 1 year ago
Michael Yang fe776293f7
Merge pull request #9569 from dwt/patch-1 1 year ago
frob d8a5d96b98
docs: Add OLLAMA_CONTEXT_LENGTH to FAQ. (#9545) 1 year ago
Xiaowei Zhu 757668c42f
docs: add SwiftChat (#9540) 1 year ago
Sam 96ec8afd09
docs(tool): add mcp-llm (#9537) 1 year ago
Jeffrey Morgan e093db92c4
sample: temporarily use grammars for constrained generation in new engine (#9586) 1 year ago
Jesse Gross a1cda80bcb model: Update encoder cache to use multimodal input processing handler 1 year ago
Jesse Gross 4614fafae0 ollamarunner: Don't panic for unimplemented features at runtime. 1 year ago
Jesse Gross 4100ed7bdd ml: Add support for quantized KV cache 1 year ago
Jesse Gross f52b2615ef kvcache: Set context for shift offsets 1 year ago
Jesse Gross 25f9b152f9 ggml-backend: Ensure allocation meet backend requirements 1 year ago
Jesse Gross 6da8b6a879 kvcache: Support non-causal attention 1 year ago
Jesse Gross 0daaaef8c9 ollamarunner: Quiet debug logging and panic on unimplemented features 1 year ago
Jesse Gross 98272fbd58 additional review comments 1 year ago
Michael Yang b27e8f3f10 ml/backend/ggml: use backend buffer type 1 year ago
Michael Yang 45df786f09 comments 1 year ago
Michael Yang daaf42e4a4 ml/backend/ggml: clean up 1 year ago
Michael Yang 2dc60d4620 ml/backend/ggml: offload vision to cpu 1 year ago
Michael Yang b5312f30e8 ml/backend/ggml: handle tensor split 1 year ago
Michael Yang 26c2e0bd35 ml/backend/ggml: handle user specified cpu offloading 1 year ago
Michael Yang bf920883d5 ml/backend/ggml: set cpu n_threads 1 year ago
Michael Yang 58b9ec1f6b kvcache: update tests 1 year ago
Michael Yang 7bae7fa5ce ml/backend/ggml: create tensor on specific backend 1 year ago
Michael Yang 764e199d67 kvcache: create cache ctx per layer 1 year ago
Michael Yang bfce55db3d model: load non-repeated tensors into multiple backends 1 year ago
Michael Yang bab6f34dc0 ml/backend/ggml: update model loading for hybrid/multi backends 1 year ago
Parth Sareen 0682dae027
sample: improve ollama engine sampler performance (#9374) 1 year ago
Breaker 1f6986e919
readme: add QwQ to the supported models list (#9565) 1 year ago
Jeffrey Morgan 4289c74359
llama: fix kv loading on snowflake-arctic-embed models (#9536) 1 year ago
‮rekcäH nitraM‮ 25248f4bd5
Better WantedBy declaration 1 year ago
Jesse Gross a7e63b82be ollamarunner: Improve multimodal input handling 1 year ago
Jesse Gross b70fc4d51e model: Don't unconditionally add special tokens 1 year ago
Blake Mizerany e2252d0fc6
server/internal/registry: take over pulls from server package (#9485) 1 year ago
Daniel Hiltgen cae5d4d4ea
Win: doc new rocm zip file (#9367) 1 year ago
Michael Yang 05a01fdecb ml/backend/ggml: consolidate system info logging 1 year ago
aritra saha 8fe6f69f28
docs: add granite-3.2 to the readme 1 year ago
Daniel Hiltgen 1fdb351c37
New engine: vision models and auto-fallback (#9113) 1 year ago
Blake Mizerany 7a01ad7614
server/internal/registry: reintroduce pruning on model deletion (#9489) 1 year ago
Blake Mizerany 55ab9f371a
server/.../backoff,syncs: don't break builds without synctest (#9484) 1 year ago
KindBrave fefbf8f74b
docs: add Ollama Android Chat community integration 1 year ago
Michael Yang b428ddd796 docker: use go version from go.mod 1 year ago
Michael Yang ba7d31240e fix: own lib/ollama directory 1 year ago
CYJiang d25efe3954
cmd: add default err return for stop (#9458) 1 year ago
Mark 36dfb906bb
docs: don't use self-closing tag for anchor element (#9456) 1 year ago
aritra saha a6f0f908b9
docs: update phi3-mini to phi4-mini (#9424) 1 year ago
İbrahim Çetin 3b1ddb2b3a
docs: add reins to community integrations (#9411) 1 year ago
Jeffrey Morgan 1579c4f06d
build: install binutils alongside gcc in Dockerfile (#9475) 1 year ago