| ▲ | kristianp 2 hours ago | |
LLMs token generation is memory bandwidth constrained. If the M7 has double the memory bandwidth as some speculate [1], then it will help with LLM performance. There is also expectation (possibly unfounded) that the GPU will have improved matrix performance to help prompt processing as well. [1] "LPDDR6 is coming." - https://news.ycombinator.com/item?id=49436849 | ||
| ▲ | Dylan16807 an hour ago | parent [-] | |
14GT/s is a goal but it might be a while, I think 11-12 is more likely in that time frame. 50% more pins... we'll see. If they can reasonably make that fit then even more shame upon the traditional desktop CPU makers for sticking with 128 bits for so long. | ||