1. 07 May, 2025 2 commits
  2. 25 Feb, 2025 1 commit
  3. 20 Feb, 2025 1 commit
    • Michael Yang's avatar
      ci: use clang for windows cpu builds · ba9ec3d0
      Michael Yang authored
      clang outputs are faster. we were previously building with clang via gcc
      wrapper in cgo but this was missed during the build updates so there was
      a drop in performance
      ba9ec3d0
  4. 18 Feb, 2025 1 commit
    • Michael Yang's avatar
      ci: set owner/group in tarball · 7b5d916a
      Michael Yang authored
      set owner and group when building the linux tarball so extracted files
      are consistent. this is the behaviour of release tarballs in version
      0.5.7 and lower
      7b5d916a
  5. 08 Feb, 2025 1 commit
    • Michael Yang's avatar
      ci: use windows-2022 to sign and bundle (#8941) · 1f766c36
      Michael Yang authored
      ollama requires vcruntime140_1.dll which isn't found on 2019. previously
      the job used the windows runner (2019) but it explicitly installs
      2022 to build the app. since the sign job doesn't actually build
      anything, it can use the windows-2022 runner instead.
      1f766c36
  6. 06 Feb, 2025 1 commit
    • Michael Yang's avatar
      ci: fix linux archive (#8862) · 1c198977
      Michael Yang authored
      the find returns intermediate directories which pulls the parent
      directories. it also omits files under lib/ollama.
      
      switch back to globbing
      1c198977
  7. 05 Feb, 2025 2 commits
  8. 04 Feb, 2025 2 commits
  9. 03 Feb, 2025 2 commits
  10. 31 Jan, 2025 2 commits
  11. 30 Jan, 2025 1 commit
  12. 29 Jan, 2025 1 commit
    • Michael Yang's avatar
      next build (#8539) · dcfb7a10
      Michael Yang authored
      
      
      * add build to .dockerignore
      
      * test: only build one arch
      
      * add build to .gitignore
      
      * fix ccache path
      
      * filter amdgpu targets
      
      * only filter if autodetecting
      
      * Don't clobber gpu list for default runner
      
      This ensures the GPU specific environment variables are set properly
      
      * explicitly set CXX compiler for HIP
      
      * Update build_windows.ps1
      
      This isn't complete, but is close.  Dependencies are missing, and it only builds the "default" preset.
      
      * build: add ollama subdir
      
      * add .git to .dockerignore
      
      * docs: update development.md
      
      * update build_darwin.sh
      
      * remove unused scripts
      
      * llm: add cwd and build/lib/ollama to library paths
      
      * default DYLD_LIBRARY_PATH to LD_LIBRARY_PATH in runner on macOS
      
      * add additional cmake output vars for msvc
      
      * interim edits to make server detection logic work with dll directories like lib/ollama/cuda_v12
      
      * remove unncessary filepath.Dir, cleanup
      
      * add hardware-specific directory to path
      
      * use absolute server path
      
      * build: linux arm
      
      * cmake install targets
      
      * remove unused files
      
      * ml: visit each library path once
      
      * build: skip cpu variants on arm
      
      * build: install cpu targets
      
      * build: fix workflow
      
      * shorter names
      
      * fix rocblas install
      
      * docs: clean up development.md
      
      * consistent build dir removal in development.md
      
      * silence -Wimplicit-function-declaration build warnings in ggml-cpu
      
      * update readme
      
      * update development readme
      
      * llm: update library lookup logic now that there is one runner (#8587)
      
      * tweak development.md
      
      * update docs
      
      * add windows cuda/rocm tests
      
      ---------
      Co-authored-by: default avatarjmorganca <jmorganca@gmail.com>
      Co-authored-by: default avatarDaniel Hiltgen <daniel@ollama.com>
      dcfb7a10
  13. 11 Dec, 2024 2 commits
  14. 10 Dec, 2024 1 commit
    • Daniel Hiltgen's avatar
      build: Make target improvements (#7499) · 4879a234
      Daniel Hiltgen authored
      * llama: wire up builtin runner
      
      This adds a new entrypoint into the ollama CLI to run the cgo built runner.
      On Mac arm64, this will have GPU support, but on all other platforms it will
      be the lowest common denominator CPU build.  After we fully transition
      to the new Go runners more tech-debt can be removed and we can stop building
      the "default" runner via make and rely on the builtin always.
      
      * build: Make target improvements
      
      Add a few new targets and help for building locally.
      This also adjusts the runner lookup to favor local builds, then
      runners relative to the executable, and finally payloads.
      
      * Support customized CPU flags for runners
      
      This implements a simplified custom CPU flags pattern for the runners.
      When built without overrides, the runner name contains the vector flag
      we check for (AVX) to ensure we don't try to run on unsupported systems
      and crash.  If the user builds a customized set, we omit the naming
      scheme and don't check for compatibility.  This avoids checking
      requirements at runtime, so that logic has been removed as well.  This
      can be used to build GPU runners with no vector flags, or CPU/GPU
      runners with additional flags (e.g. AVX512) enabled.
      
      * Use relative paths
      
      If the user checks out the repo in a path that contains spaces, make gets
      really confused so use relative paths for everything in-repo to avoid breakage.
      
      * Remove payloads from main binary
      
      * install: clean up prior libraries
      
      This removes support for v0.3.6 and older versions (before the tar bundle)
      and ensures we clean up prior libraries before extracting the bundle(s).
      Without this change, runners and dependent libraries could leak when we
      update and lead to subtle runtime errors.
      4879a234
  15. 04 Nov, 2024 3 commits
  16. 02 Nov, 2024 1 commit
  17. 30 Oct, 2024 2 commits
  18. 24 Sep, 2024 1 commit
  19. 21 Sep, 2024 1 commit
  20. 20 Sep, 2024 3 commits
  21. 17 Sep, 2024 1 commit
  22. 16 Sep, 2024 2 commits
  23. 12 Sep, 2024 1 commit
    • Daniel Hiltgen's avatar
      Optimize container images for startup (#6547) · cd5c8f64
      Daniel Hiltgen authored
      * Optimize container images for startup
      
      This change adjusts how to handle runner payloads to support
      container builds where we keep them extracted in the filesystem.
      This makes it easier to optimize the cpu/cuda vs cpu/rocm images for
      size, and should result in faster startup times for container images.
      
      * Refactor payload logic and add buildx support for faster builds
      
      * Move payloads around
      
      * Review comments
      
      * Converge to buildx based helper scripts
      
      * Use docker buildx action for release
      cd5c8f64
  24. 20 Aug, 2024 1 commit
    • Daniel Hiltgen's avatar
      Split rocm back out of bundle (#6432) · a017cf2f
      Daniel Hiltgen authored
      We're over budget for github's maximum release artifact size with rocm + 2 cuda
      versions.  This splits rocm back out as a discrete artifact, but keeps the layout so it can
      be extracted into the same location as the main bundle.
      a017cf2f
  25. 19 Aug, 2024 4 commits