I'm not talking only about Fable level models, since we are obviously already close with stuff like Qwen. But I'm also wondering about being able to run them on consumer-end hardware. I remember using a local model 2-3 years ago and had to wait around 2-3 minutes for a basic answer to be printed.
Source: [Hacker News](https://news.ycombinator.com/item?id=49814142)