You can now run a 90M conversational LLM on the Sony PSP (hardware from 2004). Doesn't get more local than this.
Github link: https://github.com/thatblend/LLMPSP
I wanted to see what the PSP can theoretically handle and I got my answer - a 90M model is about the max it can do without atrocious inference speeds. It's running around 0.5 - 0.6 tokens per second, which is very slow, but it's useable. Maybe 1-3 minutes for a reply.
The model is actually fairly impressive for 90M parameters, it's not really useful in any real metric, but it can generate crappy poems, short stories, write non-functional code and sometimes it gets things right if you ask it what company makes macbooks, what is an LLM etc, while other times it just hallucinates a crazy answer. Fun.
scorecomments19 sightings
first seen 2026-10-02 11:33 UTClast seen 2026-10-07 04:40 UTCscore then 1443score now 1443gained 0sightings 19
open on reddit ↗
💬 110