| ▲ | foobar_______ 3 hours ago | |||||||
Hard to believe numbers. I don't mean that as a critique, but literally I am so impressed. Even if the model is a few percent lower for performance but is 80+% cheaper than competitors and is a US company hosted on US based hyperscaler clouds this is kind of a no brainer. Hard for most businesses to justify otherwise. | ||||||||
| ▲ | rpdillon 3 hours ago | parent | next [-] | |||||||
This is exactly the model that DeepSeek V4 Flash followed, and it's been insanely successful as a result, even though it's not frontier. | ||||||||
| ▲ | ignoramous 2 hours ago | parent | prev [-] | |||||||
DeepSeek v4 Pro & MiMo v2.5 Pro (Opus 4.6 quality models for code) are insanely cheap for agent-driven work due to their super low cached-input prices ($0.0036/mtok) [0]. For Luna, the cached-input price drop isn't disclosed in TFA, but the pricing page puts it at $0.02/mtok, & that's 5x more expensive. [0] I am constantly surprised how much work pay-as-you-go with DeepSeek / MiMo will get done. I've barely crossed $2 each in a month of use (~200m tokens). | ||||||||
| ||||||||