Commits · dcfb7a105c455ae8d44a06b3380731d8b1ffcc22 · OpenDAS / ollama

29 Jan, 2025 1 commit

Michael Yang authored Jan 29, 2025



* add build to .dockerignore

* test: only build one arch

* add build to .gitignore

* fix ccache path

* filter amdgpu targets

* only filter if autodetecting

* Don't clobber gpu list for default runner

This ensures the GPU specific environment variables are set properly

* explicitly set CXX compiler for HIP

* Update build_windows.ps1

This isn't complete, but is close.  Dependencies are missing, and it only builds the "default" preset.

* build: add ollama subdir

* add .git to .dockerignore

* docs: update development.md

* update build_darwin.sh

* remove unused scripts

* llm: add cwd and build/lib/ollama to library paths

* default DYLD_LIBRARY_PATH to LD_LIBRARY_PATH in runner on macOS

* add additional cmake output vars for msvc

* interim edits to make server detection logic work with dll directories like lib/ollama/cuda_v12

* remove unncessary filepath.Dir, cleanup

* add hardware-specific directory to path

* use absolute server path

* build: linux arm

* cmake install targets

* remove unused files

* ml: visit each library path once

* build: skip cpu variants on arm

* build: install cpu targets

* build: fix workflow

* shorter names

* fix rocblas install

* docs: clean up development.md

* consistent build dir removal in development.md

* silence -Wimplicit-function-declaration build warnings in ggml-cpu

* update readme

* update development readme

* llm: update library lookup logic now that there is one runner (#8587)

* tweak development.md

* update docs

* add windows cuda/rocm tests

---------
Co-authored-by: jmorganca <jmorganca@gmail.com>
Co-authored-by: Daniel Hiltgen <daniel@ollama.com>

dcfb7a10

23 Jan, 2025 1 commit
- docs: remove reference to the deleted examples folder (#8524) · ca2f9843
  Daniel Jalkut authored Jan 23, 2025
  
  ca2f9843
21 Jan, 2025 1 commit
- docs: remove tfs_z option from documentation (#8515) · 294b6f5a
  frob authored Jan 21, 2025
  
  294b6f5a
20 Jan, 2025 1 commit
- docs: update suspend header in gpu.md (#8487) · 7bb356c6
  EndoTheDev authored Jan 20, 2025
  
  7bb356c6
15 Jan, 2025 1 commit
- docs: fix path to examples (#8438) · a041b4df
  Gloryjaw authored Jan 16, 2025
  
  a041b4df
14 Jan, 2025 1 commit
- add new create api doc (#8388) · ab39872c
  Patrick Devine authored Jan 13, 2025
  
  ab39872c
13 Jan, 2025 1 commit
- examples: remove codified examples (#8267) · 84a23144
  Parth Sareen authored Jan 13, 2025
  
  84a23144
29 Dec, 2024 1 commit
- docs: add /api/version endpoint documentation (#8082) · 103db421
  Anas Khan authored Dec 30, 2024
```
Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com>
```
  103db421
27 Dec, 2024 1 commit
- docs: add syntax highlighting on Go template code blocks (#8215) · b68e8e57
  CIIDMike authored Dec 28, 2024
  
  b68e8e57
20 Dec, 2024 1 commit
- remove tutorials.md which pointed to removed tutorials (#8189) · d8bab8ea
  Patrick Devine authored Dec 20, 2024
  
  d8bab8ea
13 Dec, 2024 1 commit

openai: return usage as final chunk for streams (#6784) · e28f2d49

Anuraag (Rag) Agrawal authored Dec 13, 2024



* openai: return usage as final chunk for streams

---------
Co-authored-by: ParthSareen <parth.sareen@ollama.com>

e28f2d49

11 Dec, 2024 1 commit
- llama: update vendored code to commit 40c6d79f (#7875) · 527cc978
  Jeffrey Morgan authored Dec 10, 2024
  
  527cc978
10 Dec, 2024 3 commits

all: fix typos in documentation, code, and comments (#7021) · abfdc471
Stefan Weil authored Dec 10, 2024

abfdc471
build: fix typo in override variable (#8031) · 82a02e18
Daniel Hiltgen authored Dec 10, 2024
```
The "F" was missing.
```
82a02e18

build: Make target improvements (#7499) · 4879a234

Daniel Hiltgen authored Dec 10, 2024

* llama: wire up builtin runner

This adds a new entrypoint into the ollama CLI to run the cgo built runner.
On Mac arm64, this will have GPU support, but on all other platforms it will
be the lowest common denominator CPU build.  After we fully transition
to the new Go runners more tech-debt can be removed and we can stop building
the "default" runner via make and rely on the builtin always.

* build: Make target improvements

Add a few new targets and help for building locally.
This also adjusts the runner lookup to favor local builds, then
runners relative to the executable, and finally payloads.

* Support customized CPU flags for runners

This implements a simplified custom CPU flags pattern for the runners.
When built without overrides, the runner name contains the vector flag
we check for (AVX) to ensure we don't try to run on unsupported systems
and crash.  If the user builds a customized set, we omit the naming
scheme and don't check for compatibility.  This avoids checking
requirements at runtime, so that logic has been removed as well.  This
can be used to build GPU runners with no vector flags, or CPU/GPU
runners with additional flags (e.g. AVX512) enabled.

* Use relative paths

If the user checks out the repo in a path that contains spaces, make gets
really confused so use relative paths for everything in-repo to avoid breakage.

* Remove payloads from main binary

* install: clean up prior libraries

This removes support for v0.3.6 and older versions (before the tar bundle)
and ensures we clean up prior libraries before extracting the bundle(s).
Without this change, runners and dependent libraries could leak when we
update and lead to subtle runtime errors.

4879a234

08 Dec, 2024 2 commits
- docs: remove comment regarding tool streaming in openai.md (#7960) · da09488f
  Yannick Gloster authored Dec 07, 2024
  
  da09488f
- docs: fix syntax error in openai.md (#7986) · 7f0ccc8a
  湛露先生 authored Dec 08, 2024
  
  7f0ccc8a
06 Dec, 2024 1 commit
- docs: update readmes for structured outputs (#7962) · f6e87fd6
  Parth Sareen authored Dec 06, 2024
  
  f6e87fd6
03 Dec, 2024 2 commits
- llm: introduce k/v context quantization (vRAM improvements) (#6279) · 1bdab9fd
  Sam authored Dec 04, 2024
  
  1bdab9fd
- docs: correct default num_predict value in modelfile.md (#7693) · 2b82c5a8
  owboson authored Dec 04, 2024
  
  2b82c5a8
02 Dec, 2024 1 commit
- docs: remove extra quote in modelfile.md (#7908) · 55c3efa9
  Tigran authored Dec 02, 2024
  
  55c3efa9
30 Nov, 2024 1 commit
- server: add warning message for deprecated context field (#7878) · d543b282
  Jeffrey Morgan authored Nov 30, 2024
  
  d543b282
21 Nov, 2024 2 commits
- docs: remove tutorials, add cloud section to community integrations (#7784) · 27d9c749
  Jeffrey Morgan authored Nov 21, 2024
  
  27d9c749
- docs: Link to AMD guide on multi-GPU guidance (#7744) · d8632982
  Daniel Hiltgen authored Nov 20, 2024
  
  d8632982
20 Nov, 2024 1 commit
- docs: fix minor typo in import.md (#7764) · 2f0a8c87
  rohitanshu authored Nov 20, 2024
```
change 'containg' to 'containing'
```
  2f0a8c87
19 Nov, 2024 1 commit
- update the docs (#7731) · 712d63c3
  Patrick Devine authored Nov 18, 2024
  
  712d63c3
17 Nov, 2024 1 commit
- docs: add customization section in linux.md (#7709) · b42a5964
  Jeffrey Morgan authored Nov 17, 2024
  
  b42a5964
15 Nov, 2024 1 commit
- build: fix arm container image (#7674) · a0ea067b
  Daniel Hiltgen authored Nov 14, 2024
```
Fix a rebase glitch from the old C++ runner build model
```
  a0ea067b
12 Nov, 2024 2 commits

doc: capture numeric group requirement (#6941) · ac07160c

Daniel Hiltgen authored Nov 12, 2024

Docker uses the container filesystem for name resolution, so we can't guide users
to use the name of the host group.  Instead they must specify the numeric ID.

ac07160c

docs: Capture docker cgroup workaround (#7519) · 6606e424

Daniel Hiltgen authored Nov 12, 2024

GPU support can break on some systems after a while.  This captures a
known workaround to solve the problem.

6606e424

11 Nov, 2024 1 commit
- docs: add mentions of Llama 3.2 (#7517) · 479d5517
  frances720 authored Nov 10, 2024
  
  479d5517
08 Nov, 2024 1 commit
- docs: update langchainpy.md with proper model name (#7527) · 771fab1d
  Edward J. Schwartz authored Nov 08, 2024
  
  771fab1d
06 Nov, 2024 2 commits
- docs: OLLAMA_NEW_RUNNERS no longer exists · 3020d2dc
  Jesse Gross authored Nov 06, 2024
  
  3020d2dc
- runner.go: Remove unused arguments · a9094176
  Jesse Gross authored Oct 30, 2024
```
Now that server.cpp is gone, we don't need to keep passing arguments
that were only ignored and only kept for compatibility.
```
  a9094176
30 Oct, 2024 4 commits

Soften windows clang requirement (#7428) · 712e99d4

Daniel Hiltgen authored Oct 30, 2024

This will no longer error if built with regular gcc on windows.  To help
triage issues that may come in related to different compilers, the runner now
reports the compier used by cgo.

712e99d4

Remove submodule and shift to Go server - 0.4.0 (#7157) · b754f5a6

Daniel Hiltgen authored Oct 30, 2024

* Remove llama.cpp submodule and shift new build to top

* CI: install msys and clang gcc on win

Needed for deepseek to work properly on windows

b754f5a6

Move windows app out of preview (#7347) · a805e594
Daniel Hiltgen authored Oct 30, 2024

a805e594

windows: Support alt install paths, fit and finish (#6967) · 91dfbb1b

Daniel Hiltgen authored Oct 30, 2024

* windows: Support alt install paths

Advanced users are leveraging innosetup's /DIR switch to target
an alternate location, but we get confused by things not existing in the LocalAppData dir.
This also hardens the server path lookup code for a future attempt to unify with a ./bin prefix

* Fit and finish improvements for windows app

Document alternate install location instructions for binaries and model.
Pop up progress UI for upgrades (automatic, with cancel button).
Expose non-default port in menu to disambiguate mutiple instances.
Set minimum Windows version to 10 22H2

91dfbb1b

29 Oct, 2024 1 commit

Switch windows to clang (#7407) · c9ca3861

Daniel Hiltgen authored Oct 29, 2024

* Switch over to clang for deepseek on windows

The patch for deepseek requires clang on windows. gcc on windows
has a buggy c++ library and can't handle the unicode characters

* Fail fast with wrong compiler on windows

Avoid users mistakenly building with GCC when we need clang

c9ca3861

26 Oct, 2024 1 commit

Better support for AMD multi-GPU on linux (#7212) · d7c94e0c

Daniel Hiltgen authored Oct 26, 2024

* Better support for AMD multi-GPU

This resolves a number of problems related to AMD multi-GPU setups on linux.

The numeric IDs used by rocm are not the same as the numeric IDs exposed in
sysfs although the ordering is consistent.  We have to count up from the first
valid gfx (major/minor/patch with non-zero values) we find starting at zero.

There are 3 different env vars for selecting GPUs, and only ROCR_VISIBLE_DEVICES
supports UUID based identification, so we should favor that one, and try
to use UUIDs if detected to avoid potential ordering bugs with numeric IDs

* ROCR_VISIBLE_DEVICES only works on linux

Use the numeric ID only HIP_VISIBLE_DEVICES on windows

d7c94e0c