It's too conservative, on my Air M4 I run gemma-4-12B-it-qat-UD-Q4_K_XL fine with llama.cpp
The app states "Small Models (1-3B)".