| ▲ | baalimago 7 hours ago |
| New Deepseek models are like Christmas for me. Really big fan of low cost API models, noone does it better than DS. Until VRAM price is low enough to run models locally, this is the way to go. The subsidized subscription model won't last, API pricing "feels" closer to a true sustainable business model. |
|
| ▲ | dpacmittal 4 hours ago | parent | next [-] |
| Can't wait for China to catchup on hardware (memory and compute) and absolutely crush American companies in both price and performance. |
| |
| ▲ | hilios 4 hours ago | parent | next [-] | | I'm not looking forward to it, them being strapped for resources provides a huge incentive to develop and release these smaller models. Even if they'd still release their models once they are able to comfortably service all potential customers via their cloud, running them locally would be almost impossible due to their size. | | |
| ▲ | seanmcdirmid 4 hours ago | parent | next [-] | | Chinese are pretty pragmatic. Even if they produce more expensive chips and memory, they are still going to focus on value. But who knows, maybe India will step up again like they did for Y2K 3 decades ago. | | |
| ▲ | shwetanshu21 12 minutes ago | parent | next [-] | | Unfortunately, India's focus is on everything but models. Hope I am wrong but the VCs there are now very risk free and are only investing in Snacks, Beauty and similar companies making risk free money due to growing income of population. All the good talent moved to other countries due to this. | |
| ▲ | budsniffer952 an hour ago | parent | prev [-] | | Ah, positive racism. So Americans aren't pragmatic, generally? | | |
| |
| ▲ | LoganDark 2 hours ago | parent | prev [-] | | I agree. A meeting transcript posted recently from DeepSeek's founder suggests they're going to reach for larger models as soon as they can, and not look back. The model is already bigger than my machine can handle, though (I cannot put up with 25 t/s after getting used to over 100), so I'm not impacted. |
| |
| ▲ | xXSLAYERXx 3 minutes ago | parent | prev | next [-] | | You're clearly not an American and they will never catch up on hardware. I'm mostly curious why you are so excited for China to "crush" American companies? | |
| ▲ | baalimago 4 hours ago | parent | prev | next [-] | | Personally I'm hoping for EU to step up a bit | | |
| ▲ | chronogram 2 hours ago | parent | next [-] | | We'll make sure there will be enough regulations to have neither the electricity to power it or the companies to build it. | | |
| ▲ | fy20 an hour ago | parent [-] | | Well quite a few nuclear powerplants have had to shut down recently due to droughts... So we are well on the way for that! |
| |
| ▲ | speedgoose 4 hours ago | parent | prev | next [-] | | Sorry we a focusing on leather handbags there. | | |
| ▲ | seanmcdirmid 4 hours ago | parent [-] | | The leather bag designs are shipped from Paris or Milan to China for production. | | |
| ▲ | serial_dev 2 hours ago | parent | next [-] | | But some worker in Italy bolts on a zipper in the end so it’s actually Italian, believe it or not. | | | |
| ▲ | speedgoose an hour ago | parent | prev [-] | | It depends on the bags. I know people who work in handbags production, in France. |
|
| |
| ▲ | tao_oat 4 hours ago | parent | prev | next [-] | | Seems incredibly unlikely, doesn't it? (I hope for the same thing but it feels almost like wishful thinking!) | |
| ▲ | m00dy 2 hours ago | parent | prev [-] | | EU is out of game and can't be back anytime soon. |
| |
| ▲ | Saline9515 4 hours ago | parent | prev | next [-] | | Good luck when the last non-Chinese frontier labs will have closed and the CCP will ask to stop sharing models open source. | | |
| ▲ | InsideOutSanta 4 hours ago | parent | next [-] | | We'll never have fewer open-weight models than exist now. They won't suddenly disappear when labs stop publishing new ones. In fact, people will keep improving them and will keep distilling new frontier models into existing open-weight models. | |
| ▲ | codedokode an hour ago | parent | prev | next [-] | | Then non-Chinese labs can reopen, or they could offer lower prices right now not waiting for this event. But imagine if Chinese labs would lose. Then there will be no open models, and the prices for closed models would be raised to the maximum. | |
| ▲ | donquichotte 3 hours ago | parent | prev | next [-] | | My conspiracy theory is that this is the new space race, and the CCP encourages this to show the world what Chinese engineers are capable of, and tank the Anthropic/OpenAI valuation bubble as a desirable side effect. | | |
| ▲ | throwa356262 3 hours ago | parent | next [-] | | Half of the engineers at OAI and Anthropic are Asian, I don't think China does all these for signaling. | | | |
| ▲ | u8080 2 hours ago | parent | prev [-] | | Oh no, 1kkk market country with top-tier research labs developing its own technology, must be evil! | | |
| ▲ | arjie an hour ago | parent [-] | | Well, no, they're not aligned with our interests therefore we're concerned with their mastery here. You have the clause the wrong way around. |
|
| |
| ▲ | slopinthebag an hour ago | parent | prev [-] | | This will never happen btw. |
| |
| ▲ | segmondy 3 hours ago | parent | prev | next [-] | | :-( | |
| ▲ | 4 hours ago | parent | prev | next [-] | | [deleted] | |
| ▲ | dominotw 4 hours ago | parent | prev [-] | | [flagged] | | |
| ▲ | dpacmittal 4 hours ago | parent | next [-] | | I'm neither pro China, nor pro US. I'm pro open weights models, and I'm pro cheaper hardware. At this point I don't see any american frontier labs releasing SOTA open weights model, and I don't see ASML/Nvidia/Samsung monopoly getting any competition from anywhere apart from China in the near future. | | |
| ▲ | dominotw 4 hours ago | parent [-] | | > I'm neither pro China, nor pro US. I'm pro open weights models, and I'm pro cheaper hardware. yea i got that from your first comment ( although you removed crush American companies in _price_ ). you are pro cheapness at any cost even if its from your country's state funded direct geopolitical enemy. China can always count on first order greed to win | | |
| |
| ▲ | VulgarExigency 3 hours ago | parent | prev [-] | | And what is their nefarious plan after we get "addicted to cheap shit"? |
|
|
|
| ▲ | Flere-Imsaho 5 hours ago | parent | prev | next [-] |
| Indeed. My fellow software engineers keep complaining about using up all their Claude tokens within an hour... Whilst I'll be rocking DS flash for the entire day. Sure it gets a few things wrong here and there, but that's when you pull out the Claude models or whatever for those tricky tasks. |
| |
| ▲ | ljosifov 5 hours ago | parent | next [-] | | Same. And I have come to use OMP (oh my pi) agent /advisor mode to put a 2nd model on the case (also mid-size one), reading everything. It can not block anything or change anything - just inserts comments in the text stream with 1 turn delay. Good portion of the time it's quiet. I'd say 1/2 of the time it's got something to say. About 2/3-rd of that the 'advice' is insubstantial or about something not-quite wrong. The good thing is the main model is confident - checks and then it stands its ground. Have not noticed it turning a right into a wrong b/c of advisor false alarm. And in 1/3-rd of the advice, it's a genuine defect teh advisor noticed, the main model works out a fix. This is my approximate feeling just observing the process, have not got collected the data. Afaik only OMP has advisor mode. Agent pi has plugin pi-omplike-advisor. For agent Hermes I had them code me an /advisor plugin (for now -0.1 old v0.18.x; yet to upgrade it to latest). | |
| ▲ | jmartrican 4 hours ago | parent | prev | next [-] | | The problem is picking between models. I do not want to spend my time switching models and trying to decipher which model should be used for what. Maybe that's just a me problem that I need to figure out. | | |
| ▲ | wongarsu 3 hours ago | parent [-] | | DeepSeek v4 is honestly good enough that I'm fine throwing it at everything in my hobby projects. I guess now I'll be switching from V4 Pro-Preview to V4 Flash. My only real complaint is that they can't do images, which limits their ability to autonomously debug some kinds of issues Of course you can get more bang for your buck by being more deliberate. But that's equally true with US frontier models. You can optimize your work by choosing between Opus, Fable, Sonnet, Sol, Luna and Terra for each task. Some people seem to prefer to let Opus code and Sol review, for example. And then there is the whole debate whether $current_version is actually better (Some people stay on Claude 4.8 because they dislike how 5.0 is sometimes doing stupid things, just as many opted out of dynamic reasoning when they still could) | | |
| ▲ | rented_mule 2 hours ago | parent [-] | | CodeWhale is a coding agent that auto-routes requests to Flash / Pro based on complexity, as determined by Flash. It's also tuned for DeepSeek's caching behavior, making things even more inexpensive. I'm retired, but I've been using it for just over two months at about the rate I would use it if I were working half-time, and I've spent $19 total. https://github.com/Hmbown/CodeWhale |
|
| |
| ▲ | felixgallo 5 hours ago | parent | prev | next [-] | | what plan are your 'fellow software engineers' using? I have a hard time even using up the Fable part of my allowance in a week of coding. | | |
| ▲ | Flere-Imsaho 4 hours ago | parent | next [-] | | I'm not sure to be fair, but they do have constant "token anxiety", which I simply don't have anymore since using v4 flash. | | | |
| ▲ | stavros an hour ago | parent | prev [-] | | I'm on the Premium business plan and I can easily use up my entire week's allocation in two days of coding. I only use Claude for planning, too, the rest is done by Deepseek and GPT. Claude is very spendy. |
| |
| ▲ | ReptileMan 4 hours ago | parent | prev [-] | | usually I am - Codex, make pi with deepseek to do something. |
|
|
| ▲ | VulgarExigency 5 hours ago | parent | prev | next [-] |
| I would bet that Deepseek API pricing is still more cost effective per token than the subscriptions. With the increase in quality Deepseek Flash just got (in my personal testing so far, it seems to have improved a lot at following instructions, and has become more proactive), there really isn’t anything that can match it in terms of cost effectiveness. |
| |
| ▲ | rapind 4 hours ago | parent | next [-] | | The issue for me is the data privacy if you're using their hosted prices, because you cannot opt out of data collection (and I'm not sure how you'd legally follow up on it if they did offer it, but didn't actually follow through). That means all of my code is being retained for training. I've used it in some open source code though, and loved how fast it was. My mind is changing on how valuable my code actually is though... it's the complete picture, how it's put together, the design, the UI, the attention to detail that's the real value. | |
| ▲ | nodja 2 hours ago | parent | prev [-] | | I'm not a heavy user of agentic coding, but still use them quite a bit for some automation here and there. I've been going around shopping all the ~$10 subscriptions and I finally settled on openrouter + ds4 pro. The more intensive days cost me $1 and I set a $2 weekly limit which I've never hit the past 3 weeks, to me it's way cheaper than most subscriptions and I don't have to worry about maximizing my weekly quota/resets. |
|
|
| ▲ | singingtoday 2 hours ago | parent | prev | next [-] |
| I have one of my projects on DeepSeek and it blows my mind how I can use tens of millions of tokens for pennies. I don't find it suitable for everything, but there's some tasks it crushes for what feels like almost free. |
|
| ▲ | 5 hours ago | parent | prev | next [-] |
| [deleted] |
|
| ▲ | refulgentis 2 hours ago | parent | prev | next [-] |
| FWIW, looking at pricing myself, it needs to 5x the tokens to get the same results as GPT 5.6 Luna at a much lower TPS. (doesn't obviate the fact that the pressure causes incumbents to push and push to optimize, thank you Deepseek!) |
|
| ▲ | 4 hours ago | parent | prev | next [-] |
| [deleted] |
|
| ▲ | jingpostmedia 5 hours ago | parent | prev [-] |
| [flagged] |