Remix.run Logo
Kimi K3-256k(kimi.com)
91 points by monneyboi 44 minutes ago | 12 comments
illithid0 4 minutes ago | parent | next [-]

This was posted 38 minutes ago, and as of 20 minutes ago, several Anthropic services are now designated as having a "major outage".

Doubt these are related, but it made me laugh a little.

lukan 4 minutes ago | parent | prev | next [-]

Since Claude is the first time for me really, really out (TIL against my wished about https://status.claude.com/), I am now interested enough to see what else works. But ... when I click pricing, I see "Join a waitlist". Wtf? Are they really that good, so were totally surprised and overwhelmed by the requests, is this a marketing stunt, or do they just don't have the hardware being in china?

dgritsko 17 minutes ago | parent | prev | next [-]

This isn't quantized, right? Just a smaller context?

DSingularity 7 minutes ago | parent [-]

Its 256k context window. Quantization is orthogonal. We cant really tell directly so it could be quantized.

madihaa 32 minutes ago | parent | prev | next [-]

That's actually nice! I usually try to stay below 200k context anyway.

giancarlostoro 28 minutes ago | parent [-]

For me the sweet spot is somewhere under 500k depending on how extensive I want to get. You can build up a sizable effort project in half a million tokens with Claude, with Claude having all the context from ground 0 to wherever you're off at.

cyanydeez 24 minutes ago | parent [-]

I'm always curious what you guys are working on; every git repo I've run a local model on and stick below <100k to increase speed seems effective enough to scope patches and changes.

jdoe1337halo 2 minutes ago | parent [-]

They are just talking to the model in CC, while staying in a single thread. Doubt they have any actual coding knowledge to compartmentalize different problems in the codebase.

hawtads 25 minutes ago | parent | prev | next [-]

This is just an API level change right? The model itself should be the same I think.

wxw 18 minutes ago | parent | prev | next [-]

> k3-256k is now available. Within 256k context, it delivers the same results. k3 (1M) consumes about twice as much quota as k3-256k.

ibuildproducts 14 minutes ago | parent | prev | next [-]

omg! new model!!

superloika 31 minutes ago | parent | prev [-]

The bells tolled today, but nobody came to church.