ollama
7 stories and discussions about ollama, aggregated from every source we track.
Ollaya downloads and serves open decision models on your own machine. Typed, calibrated answers in milliseconds — private and open source.
Apple's container CLI 1.4.1 cannot boot the official debian:13 image as a machine, because it has no /sbin/init. A small Dockerfile fixes that and gives you a persistent Debian 13 VM with systemd, your Mac user and your home folder. The VM has no GPU, so Ollama runs on macOS and the VM calls it at 192.168.64.1:8000. From inside Debian, gemma4:e2b answered at 45.2 tok/s, loaded 100% on the M3 GPU.
A 9B decision model from Bespoke Labs for fast, typed classification.
open-source speech AI platform for organizations that cannot send sensitive conversations to a third party - nanosamurai/nanosamurai
Ollama now supports decision models, based on TypeSafe's Jev API for fast, typed decisions. Decision models can now be run at no cost with low latency. Based on text, decision models answer yes-or-no questions, choices…
Ollama 0.12 added cloud models behind the same localhost API, so local-only inference became a setting instead of a guarantee. We replaced Ollama with a built-in llama.cpp engine so documents never leave the machine by…