| ▲ | tedd4u 2 hours ago | |
Agreed, seems like the M5 has already made steps in that direction, with 4x prompt processing / prefill performance vs. M4. [1] And token generation also got a 10% boost. The data below is only for Pro & Max but I think the base M5 got the same relative boosts vs. M4 base.
[1] https://www.hardware-corner.net/m5-pro-m5-max-local-llm-4x-f... | ||