The local AI paddock · Updated 2026-10-01

Run more
with less.

A short directory for people who run AI on their own hardware. Who to follow, which models fit your card, who builds the tools, and where the community meets next.

Next on the calendar

All events →

Fits in 24 GB

All models →

Start your feed here

All people →

From the garage

All rigs →

Why this site exists

Everyone running models at home hits the same wall sooner or later. The model you want is a few gigabytes too big for the card you own. So you drop to a smaller quant, trim the context, push a few layers to system RAM, and watch the tokens per second fall.

That wall is where the interesting work happens. People fit 70B models on two used 3090s. Others run image and video pipelines on a laptop. Quantizers publish GGUF files a few hours after a release, and the inference engines get faster every month.

This site keeps track of the people doing that work, the models worth downloading, the companies building for local inference and the events where everyone meets. The list is short on purpose. Each entry is here because it helps someone who runs models on their own machine.

Missing someone? Suggest an entry.