To run a System One model locally you pick a runtime as well as a model. Ollaya serves ten open decision models behind a Jev-compatible API, laya-mlx runs...

2 points•gosen•4 days ago•1 comment•

1 comment

gosen4 days ago
To run a System One model locally you pick a runtime as well as a model. Ollaya serves ten open decision models behind a Jev-compatible API, laya-mlx runs Laya on Apple silicon, and llama.cpp needs a shim. Which numbers are measured, and when hosted Jev still wins.

Read the full thread on Hacker News →

Related stories