I turned my old phone into a local LLM server, and it handles productivity tasks better than I expected
Despite my initial scepticism, locally-hosted LLM models have gotten a lot of oomph to their reasoning capabilities as of late. Newer Mixture-of-Experts models, for example, can run at respectable token rates on my outdated Pascal-era cards, and with a …