Commits · 1c6669e64cc8a482fbf1e35c0249f17b35a4e87a · OpenDAS / ollama

23 Jun, 2025 1 commit

Daniel Hiltgen authored Jun 23, 2025

* Re-remove cuda v11

Revert the revert - drop v11 support requiring drivers newer than Feb 23

This reverts commit c6bcdc42.

* Simplify layout

With only one version of the GPU libraries, we can simplify things down somewhat.  (Jetsons still require special handling)

* distinct sbsa variant for linux arm64

This avoids accidentally trying to load the sbsa cuda libraries on
a jetson system which results in crashes.

* temporary prevent rocm+cuda mixed loading

1c6669e6

20 Jun, 2025 1 commit
- build speedups (#11142) · 65bff664
  Daniel Hiltgen authored Jun 20, 2025
```
Enable parallel building of the GPU architectures.
```
  65bff664
13 May, 2025 1 commit

Revert "remove cuda v11 (#10569)" (#10692) · c6bcdc42

Daniel Hiltgen authored May 13, 2025

Bring back v11 until we can better warn users that their driver
is too old.

This reverts commit fa393554.

c6bcdc42

07 May, 2025 1 commit

remove cuda v11 (#10569) · fa393554

Daniel Hiltgen authored May 06, 2025

This reduces the size of our Windows installer payloads by ~256M by dropping
support for nvidia drivers older than Feb 2023. Hardware support is unchanged.

Linux default bundle sizes are reduced by ~600M to 1G.

fa393554

25 Apr, 2025 1 commit
- ci: silence deprecated gpu targets warning · 0b9198bf
  Michael Yang authored Mar 21, 2025
  
  0b9198bf
27 Mar, 2025 1 commit
- Add gfx1200 & gfx1201 support on linux (#9878) · ead27aa9
  saman-amd authored Mar 27, 2025
  
  ead27aa9
17 Mar, 2025 1 commit
- Add support for ROCm gfx1151 (#9773) · 50b59620
  Daniel Hiltgen authored Mar 17, 2025
  
  50b59620
28 Feb, 2025 1 commit

build: add compute capability 12.0 to CUDA 12 preset (#9426) · a1491285

Jeffrey Morgan authored Feb 28, 2025

Focuses initial Blackwell support on compute capability 12.0
which includes the 50x series of GeForce cards. In the future
additional compute capabilities may be added

a1491285

26 Feb, 2025 1 commit

Add cuda Blackwell architecture for v12 (#9350) · e12af460

Daniel Hiltgen authored Feb 26, 2025

* Add cuda Blackwell architecture for v12

* Win: Split rocm out to separate zip file

* Reduce CC matrix

The 6.2 and 7.2 architectures only appear on Jetsons, so they were wasting space.
The 5.0 should be forward compatible with 5.2 and 5.3.

e12af460

25 Feb, 2025 1 commit
- build: support Compute Capability 5.0, 5.2 and 5.3 for CUDA 12.x (#8567) · a4993906
  Pavol Rusnak authored Feb 25, 2025
```
CUDA 12.x still supports Compute Capability 5.0, 5.2 and 5.3,
so let's build for these architectures as well
```
  a4993906
07 Feb, 2025 1 commit
- add gfx instinct gpus (#8933) · abb8dd57
  Michael Yang authored Feb 07, 2025
  
  abb8dd57
29 Jan, 2025 1 commit

next build (#8539) · dcfb7a10

Michael Yang authored Jan 29, 2025



* add build to .dockerignore

* test: only build one arch

* add build to .gitignore

* fix ccache path

* filter amdgpu targets

* only filter if autodetecting

* Don't clobber gpu list for default runner

This ensures the GPU specific environment variables are set properly

* explicitly set CXX compiler for HIP

* Update build_windows.ps1

This isn't complete, but is close.  Dependencies are missing, and it only builds the "default" preset.

* build: add ollama subdir

* add .git to .dockerignore

* docs: update development.md

* update build_darwin.sh

* remove unused scripts

* llm: add cwd and build/lib/ollama to library paths

* default DYLD_LIBRARY_PATH to LD_LIBRARY_PATH in runner on macOS

* add additional cmake output vars for msvc

* interim edits to make server detection logic work with dll directories like lib/ollama/cuda_v12

* remove unncessary filepath.Dir, cleanup

* add hardware-specific directory to path

* use absolute server path

* build: linux arm

* cmake install targets

* remove unused files

* ml: visit each library path once

* build: skip cpu variants on arm

* build: install cpu targets

* build: fix workflow

* shorter names

* fix rocblas install

* docs: clean up development.md

* consistent build dir removal in development.md

* silence -Wimplicit-function-declaration build warnings in ggml-cpu

* update readme

* update development readme

* llm: update library lookup logic now that there is one runner (#8587)

* tweak development.md

* update docs

* add windows cuda/rocm tests

---------
Co-authored-by: jmorganca <jmorganca@gmail.com>
Co-authored-by: Daniel Hiltgen <daniel@ollama.com>

dcfb7a10