locally
23 stories and discussions about locally, aggregated from every source we track.
Run typed option logits and autoregressive JSON locally with open models and WebGPU.
Floci runs AWS, Azure, GCP, and OCI locally in milliseconds: a fast, free, credential-free feedback loop for developers and AI coding agents. MIT licensed. No account, no auth token.
Self-hosted AI workspace with agents, skills, and tools (Gmail, Calendar) that runs entirely on your own provider API keys (BYOK). Bring your own keys — Groq, OpenRouter, NVIDIA, Hugging Face, Goog...
How I replaced a cloud LLM with a fully local one — same agent, same tools, zero inference cost.
A 9B decision model from Bespoke Labs for fast, typed classification.
Your model writes the code. This runs it where it can't touch your machine.
Give your coding agents names. Let Claude Code, Codex, Grok Build, OpenCode, and Pi talk to each other with Agent Exchange.
Skip the endless downloading. Free, open source, runs on your computer. You sign in, PaperPull handles the downloading, and your records become PDFs and spreadsheets you keep.
To run a System One model locally you pick a runtime as well as a model. Ollaya serves ten open decision models behind a Jev-compatible API, laya-mlx runs...
A high-performance, GGUF-native Rust & CUDA inference engine optimized for cold-start latency and real-time 'System 1' agent decision loops. - lateos-ai/reflex
A very simple way to search for command lines in git, linux, docker in the console without leaving the console. - HE11032006/EveryCli
Run Jev-style typed decisions locally on your Mac with low RAM usage and fast responses - afshinm/laya-mps
Mac Studio M5 Max review with UAE pricing, Mac mini benchmarks, 128GB local AI tests, GPT-OSS 120B, ASUS GX10 comparison, power and thermals.
Rehearse and verify restic restores locally, on a schedule, with a diff-proof report.
Strip image metadata locally in your browser. Batch process JPEG, PNG and WebP files without uploads, tracking or third party code.
See what your real AI coding workload is worth at published API prices and what drives it, then replay it against other plans and APIs, privately in your browser.
What I learnt from three months of running personal Hermes agents
NobodyWho is an inference engine that lets you run LLMs locally and efficiently on any device. - nobodywho-ooo/nobodywho
Build the pipeline beside your logs. Preview locally, run on an Expanso Edge, and store zero log payload.
Local-first macOS menu bar app that tells you how long your prepaid cloud GPU credit will last. Reads Vast.ai, computes burn and runway, warns before you run dry. - Astralchemist/creditwatch
Contribute to jbrick2070/ComfyUI-OldTimeRadio development by creating an account on GitHub.