ollama

Files

Michael Yang 9f3a37fd36 fix: model load for unsupported embedding models (#12311 )

with #12181, there's now support for embeddings in ollama engine.
this is done by mutating the architecture and adding _embed when it
detects an embedding model. however this introduced a bug where if
an embedding model was run based on an existing ollama engine model
without an embedding implementation, e.g. llama4, it will pass the
initial arch support check but fail when actually loaded.

there's currently two entrypoints to creating a model. previously this
second entrypoint was necessary because calling model.New would also
load the model. since #11818, this is no longer th case so merge them
to reduce complexity

2025-09-18 16:11:08 -07:00

imageproc

imageproc mllama refactor (#7537 )

2024-12-14 19:50:15 -08:00

input

batch: use tensors for outputs (#12185 )

2025-09-15 14:33:06 -07:00

models

feat: qwen3 embed (#12301 )

2025-09-18 15:50:32 -07:00

parsers

address comments