Commits · e6f39bce8f3f122c9b7f9d444c8ea49f274accfe · OpenDAS / ollama

04 Aug, 2025 1 commit
- cuda graph · e6f39bce
  Michael Yang authored Jul 31, 2025
  
  e6f39bce
30 Jul, 2025 1 commit

mac: disable bf16 on unsupported OS versions (#11585) · 25911a6e

Daniel Hiltgen authored Jul 30, 2025

Support for bf16 was added in MacOS v14+ and attempting to enable
on older versions causes runtime failures.

25911a6e

29 Jul, 2025 1 commit

Increase performance for Gemma3n models on NVGPUs by enabling CUDA Graph execution (#11525) · ea85e27b

Oliver Simons authored Jul 29, 2025

* Enable CUDA Graphs for gemma3n.

Similar to
https://github.com/ggml-org/llama.cpp/pull/14741,
though ollama has a slightly different model graph
than llama.cpp which requires different workaround
checks.

* Remove residual check by reshaping differently in gemma3n model

This should make the heuristics more robust

ea85e27b

26 Jun, 2025 1 commit

add new gemma model (#11204) · 73b642e6

Michael Yang authored Jun 25, 2025

* update patches

* cherry pick metal mean kernel

* cherry pick cuda mean kernel

* gemma3n

73b642e6