| ▲ | Whitespace 5 hours ago | |||||||
I should not trust their "vibe-coded productivity/token cost saving hacks" but I should trust yours?
Why should I trust that what you're peddling isn't snakeoil? | ||||||||
| ▲ | aeneas_ory 4 hours ago | parent | next [-] | |||||||
I literally say you should take benchmarks with a grain of salt :) > Of course, it's always dependent on statistical noise + host system load, and running sufficiently large benchmarks is simply too expensive, so take em with a grain of salt. And the savings listed are coming from a benchmark harness that implements different OSS bugs one time with and one without lumen - in those cases the % saved are reproducible (caveat: it was on older models, Opus 4.6 I believe). Also I explain WHY it saves tokens - because the model doesn’t have to brute force different terms until it finds the match it needs, but uses semantic „distance“ so the embedding does it for the model. | ||||||||
| ||||||||
| ▲ | icantevenhold 5 hours ago | parent | prev | next [-] | |||||||
Only way to find out is to do some testing yourself i think. I’m using less tokens with Lumen but I also use a bunch of other tokens hacks/skills; it’s hard to measure the impact exactly but it feels significant | ||||||||
| ▲ | 4 hours ago | parent | prev | next [-] | |||||||
| [deleted] | ||||||||
| ▲ | huflungdung 4 hours ago | parent | prev [-] | |||||||
[dead] | ||||||||