Earlier quoted context omitted.
It is much slower than nVidia, but for a lot of personal-use LLM scenarios, it's very workable. And it doesn't need to be anywhere near as fast considering it's really the only viable (affordable) option for private, local inference, besides building a server like this, which is no faster: https://news.ycombinator.com/item?id=42897205
It's fast enough for me to cancel monthly AI services on a mac mini m4 max.
What model do you find fast enough and smart enough?