Remix.run Logo
kristianp 2 hours ago

LLMs token generation is memory bandwidth constrained. If the M7 has double the memory bandwidth as some speculate [1], then it will help with LLM performance. There is also expectation (possibly unfounded) that the GPU will have improved matrix performance to help prompt processing as well.

[1] "LPDDR6 is coming." - https://news.ycombinator.com/item?id=49436849

Dylan16807 an hour ago | parent [-]

14GT/s is a goal but it might be a while, I think 11-12 is more likely in that time frame.

50% more pins... we'll see. If they can reasonably make that fit then even more shame upon the traditional desktop CPU makers for sticking with 128 bits for so long.