In my regular research (behind a paywall), I have been saying for a while that I think the future of AI is not large language models (LLM), but small language models (SLM) run on local desktop computers or even mobile phones.
I wouldn’t say it’s totally legacy. (v)ram bandwidth does matter and while the Apple M-series chips do well with their unified memory architecture don’t forget it’s fixed because it’s part of the CPU chip.
I wouldn’t say it’s totally legacy. (v)ram bandwidth does matter and while the Apple M-series chips do well with their unified memory architecture don’t forget it’s fixed because it’s part of the CPU chip.