ollama

Commit Graph

Author	SHA1	Message	Date
Inforithmics	04fba9ba09	revert debugging changes	2025-09-20 11:03:09 +02:00
Inforithmics	2098e6a8e3	trying to use version 1.4.313	2025-09-20 11:00:37 +02:00
Inforithmics	fe47191720	add some more extra	2025-09-20 10:53:43 +02:00
Inforithmics	6f546457de	try again	2025-09-20 10:49:24 +02:00
Inforithmics	19bc49de5f	try without version number	2025-09-20 10:48:18 +02:00
Inforithmics	a7557cf1a8	trying again	2025-09-20 10:39:05 +02:00
Inforithmics	3ccc18f1e1	try again	2025-09-20 10:36:48 +02:00
Inforithmics	79a0f526b1	fixed vulkan-sdk name	2025-09-20 10:33:23 +02:00
Inforithmics	0f86789808	fix version	2025-09-20 10:31:44 +02:00
Inforithmics	62a8d66002	trying again	2025-09-20 10:30:31 +02:00
Inforithmics	26df69a025	trying again	2025-09-20 10:24:31 +02:00
Inforithmics	475d2c2583	trying to fix	2025-09-20 10:15:29 +02:00
Inforithmics	c91b494a8b	fix version	2025-09-20 10:10:10 +02:00
Inforithmics	af50fd5af7	try again linux build	2025-09-20 10:08:24 +02:00
Inforithmics	236c274017	temporarly disable cuda and rocm	2025-09-20 10:00:14 +02:00
Inforithmics	e29bb17613	trying to build vulkan for linux	2025-09-20 09:58:31 +02:00
Inforithmics	a0389785c7	revert windows-latest	2025-09-20 09:45:36 +02:00
Inforithmics	b244c9f9f3	revert debugging changes (vulkan builds on windows)	2025-09-20 09:44:09 +02:00
Inforithmics	6e310d1cb6	fixed install command	2025-09-20 09:37:25 +02:00
Inforithmics	b4595f0022	correct vulkan silent install	2025-09-20 09:31:58 +02:00
Inforithmics	7e161f1dbf	correct vulkan install	2025-09-20 09:16:54 +02:00
Inforithmics	d1125ea349	comment out cude for faster turnaround	2025-09-20 09:14:02 +02:00
Inforithmics	c972cf6d46	set vulkan path	2025-09-20 09:12:14 +02:00
Inforithmics	45f7850e75	temporarly commenting out rocm	2025-09-20 09:04:30 +02:00
Inforithmics	e2b38c391b	commenting out error action stop	2025-09-20 09:02:55 +02:00
Inforithmics	ed03bb7928	reenable cpu	2025-09-20 09:01:25 +02:00
Inforithmics	c84ac53579	Commenting out other presets to build vulkan	2025-09-20 09:00:26 +02:00
Inforithmics	a4461bc0d4	use temporarly windows-latest for build	2025-09-20 08:46:59 +02:00
Inforithmics	6bbc054705	temporarly comment out gate to run windows task	2025-09-20 08:35:58 +02:00
Inforithmics	0f543fdb1e	Vulkan on Windows Test	2025-09-20 08:04:11 +02:00
Daniel Hiltgen	61fb912ca4	CI: fix windows cuda build (#12246 ) * ci: adjust cuda component list v13 has a different breakdown of the components required to build ollama * review comments	2025-09-11 12:25:26 -07:00
Daniel Hiltgen	17a023f34b	Add v12 + v13 cuda support (#12000 ) * Add support for upcoming NVIDIA Jetsons The latest Jetsons with JetPack 7 are moving to an SBSA compatible model and will not require building a JetPack specific variant. * cuda: bring back dual versions This adds back dual CUDA versions for our releases, with v11 and v13 to cover a broad set of GPUs and driver versions. * win: break up native builds in build_windows.ps1 * v11 build working on windows and linux * switch to cuda v12.8 not JIT * Set CUDA compression to size * enhance manual install linux docs	2025-09-10 12:05:18 -07:00
Daniel Hiltgen	405d2f628f	ci: rocm parallel builds on windows (#11187 ) The preset CMAKE_HIP_FLAGS isn't getting used on Windows. This passes the parallel flag in through the C/CXX flags, along with suppression for some log spew warnings to quiet down the build.	2025-06-24 15:27:09 -07:00
Daniel Hiltgen	c85c0ebf89	CI: switch windows to vs 2022 (#11184 ) * CI: switch windows to vs 2022 * ci: fix regex match	2025-06-24 13:26:55 -07:00
Daniel Hiltgen	1c6669e64c	Re-remove cuda v11 (#10694 ) * Re-remove cuda v11 Revert the revert - drop v11 support requiring drivers newer than Feb 23 This reverts commit `c6bcdc4223`. * Simplify layout With only one version of the GPU libraries, we can simplify things down somewhat. (Jetsons still require special handling) * distinct sbsa variant for linux arm64 This avoids accidentally trying to load the sbsa cuda libraries on a jetson system which results in crashes. * temporary prevent rocm+cuda mixed loading	2025-06-23 14:07:00 -07:00
Daniel Hiltgen	c6bcdc4223	Revert "remove cuda v11 (#10569 )" (#10692 ) Bring back v11 until we can better warn users that their driver is too old. This reverts commit `fa393554b9`.	2025-05-13 13:12:54 -07:00
Daniel Hiltgen	fa393554b9	remove cuda v11 (#10569 ) This reduces the size of our Windows installer payloads by ~256M by dropping support for nvidia drivers older than Feb 2023. Hardware support is unchanged. Linux default bundle sizes are reduced by ~600M to 1G.	2025-05-06 17:33:19 -07:00
Jeffrey Morgan	943464ccb8	llama: update to commit 71e90e88 (#10192 )	2025-04-16 15:14:01 -07:00
Blake Mizerany	76e903cf9d	.github/workflows: swap order of go test and golangci-lint (#9389 ) The linter is secondary to the tests, so it should run after the tests, exposing test failures faster.	2025-02-26 23:03:48 -08:00
Jeffrey Morgan	a5272130c4	ml/backend/ggml: follow on fixes after updating vendored code (#9388 ) Fixes sync filters and lowers CUDA version to 11.3 in test.yaml	2025-02-26 22:33:53 -08:00
Blake Mizerany	0d694793f2	.github: always run tests, and other helpful fixes (#9348 ) During work on our new registry client, I ran into frustrations with CI where a misspelling in a comment caused the linter to fail, which caused the tests to not run, which caused the build to not be cached, which caused the next run to be slow, which caused me to be sad. This commit address these issues, and pulls in some helpful changes we've had in CI on ollama.com for some time now. They are: * Always run tests, even if the other checks fail. Tests are the most important part of CI, and should always run. Failures in tests can be correlated with failures in other checks, and can help surface the root cause of the failure sooner. This is especially important when the failure is platform specific, and the tests are not platform independent. * Check that `go generate` is clean. This prevents 'go generate' abuse regressions. This codebase used to use it to generate platform specific binary build artifacts. Let's make sure that does not happen again and this powerful tool is used correctly, and the generated code is checked in. Also, while adding `go generate` the check, it was revealed that the generated metal code was putting dates in the comments, resulting in non-deterministic builds. This is a bad practice, and this commit fixes that. Git tells us the most important date: the commit date along with other associated changes. * Check that `go mod tidy` is clean. A new job to check that `go mod tidy` is clean was added, to prevent easily preventable merge conflicts or go.mod changes being deferred to a future PR that is unrelated to the change that caused the go.mod to change. * More robust caching. We now cache the go build cache, and the go mod download cache independently. This is because the download cache contains zips that can be unpacked in parallel faster than they can be fetched and extracted by tar. This speeds up the build significantly. The linter is hostile enough. It does not need to also punish us with longer build times due to small failures like misspellings.	2025-02-25 14:28:07 -08:00
Daniel Hiltgen	e91ae3d47d	Update ROCm (6.3 linux, 6.2 windows) and CUDA v12.8 (#9304 ) * Bump cuda and rocm versions Update ROCm to linux:6.3 win:6.2 and CUDA v12 to 12.8. Yum has some silent failure modes, so largely switch to dnf. * Fix windows build script	2025-02-25 13:47:36 -08:00
Blake Mizerany	348b3e0983	server/internal: copy bmizerany/ollama-go to internal package (#9294 ) This commit copies (without history) the bmizerany/ollama-go repository with the intention of integrating it into the ollama as a replacement for the pushing, and pulling of models, and management of the cache they are pushed and pulled from. New homes for these packages will be determined as they are integrated and we have a better understanding of proper package boundaries.	2025-02-24 22:39:44 -08:00
Michael Yang	5b446cc815	chore: update gitattributes (#8860 ) * chore: update gitattributes * chore: add build info source	2025-02-05 16:37:18 -08:00
Michael Yang	dcfb7a105c	next build (#8539 ) * add build to .dockerignore * test: only build one arch * add build to .gitignore * fix ccache path * filter amdgpu targets * only filter if autodetecting * Don't clobber gpu list for default runner This ensures the GPU specific environment variables are set properly * explicitly set CXX compiler for HIP * Update build_windows.ps1 This isn't complete, but is close. Dependencies are missing, and it only builds the "default" preset. * build: add ollama subdir * add .git to .dockerignore * docs: update development.md * update build_darwin.sh * remove unused scripts * llm: add cwd and build/lib/ollama to library paths * default DYLD_LIBRARY_PATH to LD_LIBRARY_PATH in runner on macOS * add additional cmake output vars for msvc * interim edits to make server detection logic work with dll directories like lib/ollama/cuda_v12 * remove unncessary filepath.Dir, cleanup * add hardware-specific directory to path * use absolute server path * build: linux arm * cmake install targets * remove unused files * ml: visit each library path once * build: skip cpu variants on arm * build: install cpu targets * build: fix workflow * shorter names * fix rocblas install * docs: clean up development.md * consistent build dir removal in development.md * silence -Wimplicit-function-declaration build warnings in ggml-cpu * update readme * update development readme * llm: update library lookup logic now that there is one runner (#8587) * tweak development.md * update docs * add windows cuda/rocm tests --------- Co-authored-by: jmorganca <jmorganca@gmail.com> Co-authored-by: Daniel Hiltgen <daniel@ollama.com>	2025-01-29 15:03:38 -08:00
Jeffrey Morgan	527cc97899	llama: update vendored code to commit 40c6d79f (#7875 )	2024-12-10 19:21:34 -08:00
Jeffrey Morgan	aed1419c64	ci: skip go build for tests (#7899 )	2024-12-04 21:22:36 -08:00
Daniel Hiltgen	636a743c2b	CI: give windows lint more time (#7635 ) It looks like 8 minutes isn't quite enough and we're seeing sporadic timeouts	2024-11-12 11:22:39 -08:00
Daniel Hiltgen	b8d5036e33	CI: omit unused tools for faster release builds (#7432 ) This leverages caching, and some reduced installer scope to try to speed up builds. It also tidies up some windows build logic that was only relevant for the older generate/cmake builds.	2024-11-02 13:56:54 -07:00
Daniel Hiltgen	712e99d477	Soften windows clang requirement (#7428 ) This will no longer error if built with regular gcc on windows. To help triage issues that may come in related to different compilers, the runner now reports the compier used by cgo.	2024-10-30 12:28:36 -07:00

1 2 3

108 Commits