ollama

Commit Graph

Author	SHA1	Message	Date
Sheikh	6df4208836	docs: fix metal gpu section header (#13045 )	2025-11-10 21:51:22 -08:00
Eva Ho	9d615cdaa0	fix test	2025-11-10 20:13:50 -05:00
Eva Ho	6a818b8a09	clean up	2025-11-10 19:08:42 -05:00
Eva Ho	2aaf29acb5	app/ui: do not send to prevent errors with cloud provider	2025-11-10 19:05:00 -05:00
Eva H	a42f826acb	app/ui: using streamdown AI elements for markdown rendering	2025-11-10 12:05:59 -05:00
Bruce MacDonald	e10a3533a5	app/docs: remove out of date storybook instructions (#13006 )	2025-11-08 13:28:18 -08:00
Patrick Devine	91ec3ddbeb	bugfix: don't include both consolidated.safetensors and model-*.safetensors (#13010 )	2025-11-07 22:41:57 -08:00
Parth Sareen	755ac3b069	docs: update n8n URL for Ollama (#12994 )	2025-11-07 20:07:26 -08:00
Daniel Hiltgen	60b8973559	doc: re-add login autostart faq and GPU updates (#12975 ) * doc: re-add login autostart faq This appears to have been accidentally dropped during the doc migration. * docs: GPU updates lost on the doc update * review comments: improve windows login disable instructions	2025-11-07 11:21:44 -08:00
Tomoya Fujita	d2ef679d42	docs: fix 404 link to modelfile documentation (#12996 )	2025-11-07 10:06:46 -08:00
Thomas Stocker	d4e0da0890	Remove unnecessary MacOs 13 and lower Patches (#12656 ) * Remove unnecessary macos 13 Patch * Remove unnecessary MacOs Version Guard patch * rename patchesw * remove again macos13 patch * rename files	2025-11-06 15:52:56 -08:00
Jeffrey Morgan	565b802a6b	openai: fix tool call ID mapping (#12988 )	2025-11-06 15:26:25 -08:00
Saifeddine ALOUI	6c79e6c09a	readme: add security tools section and Ollama fortress to community integrations (#12981 )	2025-11-06 15:21:13 -08:00
breatn	780762f9d2	server: fix duplicate 'is' typo in comment (#12985 )	2025-11-06 14:44:44 -08:00
Jeffrey Morgan	30fcc71983	api: add omitempty to required tool function parameter type (#12989 )	2025-11-06 14:08:55 -08:00
Eva Ho	3501a4bdf9	address comment	2025-11-06 16:49:22 -05:00
Eva H	73a0cafc1e	Merge pull request #12973 from macarronesc/main feat: add support for WebP images in Ollama's app	2025-11-06 16:31:46 -05:00
Eva Ho	e309c80474	address comments	2025-11-06 13:49:59 -05:00
Daniel Hiltgen	544b6739dd	ggml update to b6840 (#12791 )	2025-11-06 10:19:22 -08:00
Daniel Alejandro Coll Tejeda	a4a53692f8	refactor: remove GIF support from image validation tests and logging	2025-11-06 09:09:51 +00:00
7394112478	c4ba257c64	readme: remove 404 link (#11351 )	2025-11-05 23:36:59 -08:00
mags0ft	342e58ce4f	readme: add hle-eval-ollama to list of terminal community integrations (#11371 )	2025-11-05 23:04:30 -08:00
Saifeddine ALOUI	47b2585cfd	readme: add lollms and lollms WebUI to community integrations (#11981 )	2025-11-05 22:48:43 -08:00
Vincent Koc	4111db013f	app: fix macOS file picker to support Uniform Type Identifiers (#12965 )	2025-11-05 21:37:17 -08:00
Eva Ho	536c987c39	address comment	2025-11-05 20:19:34 -05:00
Eva Ho	a534d4e9e1	fixing thinking not scrolling issue	2025-11-05 16:06:55 -05:00
Eva Ho	74586aa9df	address comments	2025-11-05 16:06:55 -05:00
Eva Ho	8c74f5ddfd	ui: using streamdown AI elements for markdown rendering	2025-11-05 16:06:55 -05:00
Daniel Hiltgen	80d34260ea	ci: re-enable signing (#12974 )	2025-11-05 12:33:01 -08:00
Daniel Alejandro Coll Tejeda	bddfa2100f	feat: add support for WebP images in Ollama's app	2025-11-05 21:23:20 +01:00
nicole pardal	1ca608bcd1	embeddings: added embedding command for cl (#12795 ) Co-authored-by: A-Akhil <akhilrahul70@gmail.com> This PR introduces a new ollama embed command that allows users to generate embeddings directly from the command line. Added ollama embed MODEL [TEXT...] command for generating text embeddings Supports both direct text arguments and stdin piping for scripted workflows Outputs embeddings as JSON arrays (one per line)	2025-11-05 11:58:03 -08:00
Daniel Hiltgen	6aa7283076	mac: fix stale VRAM data (#12972 ) The scheduler updates free VRAM based on current loaded models. This was mutating the persisted list of GPUs, and when coupled with the non-refreshing logic for Metal that lead to stale low VRAM reporting after unload. The fix is to make sure the GPU discovery always returns a copy so the schedulers GPU list is in fact ephemeral and doesn't leak any temporary adjustments back into the persistent list.	2025-11-05 11:55:17 -08:00
Patrick Devine	f89fc1cadd	bugfix: show connection string for interactive cli usage (#12930 )	2025-11-05 11:55:04 -08:00
Daniel Hiltgen	97e05d2a6b	win: revert CPU discovery logic to 0.12.3 (#12969 ) The behavior change in 0.12.4 is the most likely the root cause of hangs some users are seeing. This reverts to the 0.12.3 code, with some added trace logging.	2025-11-05 10:32:38 -08:00
Youdon	8bbc7395db	readme: Add handy-ollama to community integrations (#8601 )	2025-11-05 09:56:14 -08:00
Daniel Hiltgen	408c2f99d0	log: trace logging for scheduler (#12961 )	2025-11-05 08:12:15 -08:00
Grace	809b9c68fa	Add Tool Call ID (#12956 ) * routes/types: add tool call id --------- Co-authored-by: ParthSareen <parth.sareen@ollama.com>	2025-11-04 16:43:33 -08:00
Daniel Hiltgen	ba8c035846	log: instrument CPU discovery timing (#12960 )	2025-11-04 16:23:37 -08:00
Daniel Hiltgen	27f1fde413	discovery: only retry AMD GPUs (#12894 ) * discovery: only retry AMD GPUs CUDA and Vulkan don't crash on unsupported devices, so retry isn't necessary. This also refactors the code to shift the Library specific logic into the ml package. * review comments	2025-11-04 15:33:46 -08:00
virajwad	220e133fca	vulkan: Add memory detection for Intel GPU using DXGI+PDH (#12664 ) * PDH free memory skeleton * Add PDH printing * Add LUID support for Vulkan * wire luid from ggml-vulkan to mem-dxgi-pdh file * Fix to ggml-impl * Continue skeleton * Implemented ggml_dxgi_pdh_get_device_memory * fix comments * Fix - change value GB to bytes * add ifdefs to only support windows and not linux * modify error codes * Finished ggml_dxgi_pdh_init() function * completed ggml_dxgi_pdh_release() * Formatting changes, add static to functions * fix build errors * fix go build error * fix luid - now should match between dxgi and vulkan * Fix the free memory reporting (was using copy by value, change to reference) * keep only dxgi1_2.h * Modifications based on PR feedback * fix merge conflicts (2) and fix desc1.description printout * move dxgi + pdh api calls to before the vendor specific library calls * change from 3 samples to 1 sample for PDH * modify when old_mode is set * add fix for building MacOS * fix release and returns for other vendors * add patch file	2025-11-04 14:11:55 -08:00
Daniel Hiltgen	d3b4b9970a	app: add code for macOS and Windows apps under 'app' (#12933 ) * app: add code for macOS and Windows apps under 'app' * app: add readme * app: windows and linux only for now * ci: fix ui CI validation --------- Co-authored-by: jmorganca <jmorganca@gmail.com>	2025-11-04 11:40:17 -08:00
Daniel Hiltgen	a4770107a6	vulkan: enable flash attention (#12937 ) Also adjusts the vulkan windows build pattern to match recent changes in other backends so incremental builds are faster.	2025-11-04 10:31:22 -08:00
Jesse Gross	ef549d513c	ggml: Increase maximum graph size The initial implementation of qwen3-vl:235b exceeded the maximum graph size based on the number of tensors. Although this was later fixed through the use of the mrope operation, we are close to the limit in some cases. This updates to track the current llama.cpp usage of GGML.	2025-11-03 16:05:37 -08:00
Rajath Bail	d2158ca6f4	readme: add Hillnote to community integrations (#12929 )	2025-11-03 12:55:04 -08:00
Michael Yang	ce3eb0a315	chore(gptoss): cleanup dead code (#12932 )	2025-11-03 11:27:15 -08:00
Ryan Coleman	60829f7ec6	readme: add Strands Agents to community integrations (#11740 )	2025-11-02 16:01:28 -08:00
Attogram Project	9a50fd584c	readme: add Ollama Bash Lib to community integrations (#12235 )	2025-11-02 15:44:56 -08:00
Jesse Gross	392a270261	ggml: Avoid cudaMemsetAsync during memory fitting We pass invalid pointers when we check the size of the required compute graph before fitting. Some CUDA APIs validate these pointers but we can just skip them during this phase. cudaMemsetAsync is one of these that we weren't skipping but never took the code path that used it. Now that we have enabled op_offload, we can hit it in memory pressured situations.	2025-10-31 15:23:28 -07:00
Daniel Hiltgen	3bee3af6ed	cpu: always ensure LibOllamaPath included (#12890 ) In CPU only setups the LibOllamaPath was omitted causing us not to load the ggml-cpu-XXX libraries during inference.	2025-10-31 14:37:29 -07:00
Daniel Hiltgen	83537993d7	logs: catch rocm errors (#12888 ) This will help bubble up more crash errors	2025-10-31 09:54:25 -07:00

... 2 3 4 5 6 ...

4922 Commits All Branches Search

4922 Commits

All Branches