If compute is not the bottleneck, memory is easy-ish to produce (the hard part is mostly on the fab side); what stops a Chinese NVIDIA (huawei) from being 10x cheaper?
They are usually the same family, LPDDR is used for amd and macs, but the fabs are the same as the most expesive HBM memory, if they have a choice they are going to produce the ones that they can sell for more $$.
I'm writing my own inference engine for Strix Halo and the same model. I already have 30%+ performance plus a more graceful decay over long contexts; that said, their point stands: memory bandwidth is what you really want.
> This is very literally what already happens, it's called a EULA.
Yes, but they "reserve the right" to update whenever, making it pointless
> "In favor of the customer over anything else" is not a legally viable clause.
I'm sure that legislators could put the principle down in a much clearer way. What's lacking is the will.
Yep, that's me. the only real blocker is that American companies don't trust Chinese providers, but i could just find a good American provider that hosts DeepSeek and/or GLM. I would at least be able to choose my own agent instead of a quite mediocre one that wastes time and output nonsense verbs in a pathetic attempt to gain sympathy.
The only reason that stopped me from doing it is the absence of a subscription, and I did believe I couldn't get the same value with API pricing, but I'm starting to see that it's a blatant lie and true only for anthropic and openai...
A simple law: everything the customer buys must always behave *in favor of the customer over anything else*. If the product/service contradicts this, it must be fully stated before the purchase and cannot be updated. <= This would be a sane balance.
It isn't a promotion, it's 2x the parameters of opus and we are paying with 2x the consumption rate.
They just want to get rid of the subscription model.
> Try separating politics from the advancement of humanity, you'll feel better.
can't. won't.