Ollama 0.12 added cloud models behind the same localhost API, so local-only inference became a setting instead of a guarantee. We replaced Ollama with a built-in llama.cpp engine so documents never leave the machine by…

1 points•michall9k•2 days ago•0 comments•

0 comments

No comments yet.

Read the full thread on Hacker News →

Related stories