49 points perelin 1 hour ago 43 comments
N_Lens 57 minutes ago | parent
Apple got caught out because it had a shorter contract that ended in the beginning of Q3, and they've tried to source their RAM from Chinese CXMT who declined because all capacity was already locked in contracts. They've gone with a lesser known company Kioxia.
Overall the consumer segment is completely neglected, companies don't care about end users in the current market conditions.
josephg 50 minutes ago | parent
I wonder if they'll, at some point, have enough RAM? Or is this is the new normal? Will models keep scaling with the amount of ram chips openai and anthropic own?
bionhoward 38 minutes ago | parent
epicureanideal 16 minutes ago | parent
fragmede 4 minutes ago | parent
ohyes 34 minutes ago | parent
I’d guess no. Past a certain point the model has all the capabilities it can possibly usefully offer and honestly we may already be past that. The next gen model just doesn’t seem like as clear a step up as it once was.
seanmcdirmid 46 minutes ago | parent
The US Government also wouldn't let Apple use Chinese RAM anyways, even for product that was just going to be used in China. China does have capacity issues though, and focusing on local brands first is probably the right call. Hopefully they can ramp even if they can't get the fancy lithography machines from the Netherlands.
duskwuff 42 minutes ago | parent
Formerly known as Toshiba's memory division. They spun off the business as Kioxia in 2017.
rasz 35 minutes ago | parent
Kioxia is Toshiba, The most known company from the list, and it doesnt make any ram
georgemcbay 34 minutes ago | parent
Mirroring trends throughout the entire economy (not just RAM and other inputs for AI).
Everyone is chasing the top 10% or higher of the K-shaped economy for all goods and services, servicing everyone else isn't seen as being worth the investment.
This will just keep getting worse and worse everywhere for everything as long as we allow income inequality to keep exploding, which seems to be the plan.
dendrite9 32 minutes ago | parent
dingaling 25 minutes ago | parent
Because the end users keep rushing to use the megacorps' latest AI models, weaving them into their work and life. So the megacorps keep locking in contracts to build more compute.
If you use LLMs, you're responsible for this, there's no way to pass the buck. "I only use it as a companion for learning about history" - it's still your fault. "I only use it to help guide my solitions, not for vibecoding" - it's still your fault.
azan_ 21 minutes ago | parent
georgemcbay 14 minutes ago | parent
They'd try it if we weren't living in an age of persistent supply chain problems across every industry. But we haven't lived in that world since pre-2020.
In the world we do live in, good luck trying to capture the low-end of any market when you are competing with those serving the high-end for whatever your inputs are.
discordance 31 minutes ago | parent
"CXMT currently has two 12-inch DRAM fabrication plants — or fabs - in Hefei and one in Beijing, with a combined capacity of about 300,000 wafers per month.
With the new Shanghai facility and other new capacity, CXMT will double its DRAM wafer output to approximately 600,000 wafers per month, all three sources added."
https://www.reuters.com/world/china/chinas-cxmt-wins-3-billi...
rdedev 23 minutes ago | parent
quux0r 13 minutes ago | parent
bigglebear 28 minutes ago | parent
It's almost like these companies WANT a dystopia with a centralized winner-takes-all power structure. Anything in the name of profits, who gives a fuck about humanity and distribution of rights or freedoms.
brcmthrowaway 16 minutes ago | parent
whatever1 14 minutes ago | parent
isomorphic 4 minutes ago | parent
MiroslavPokorny 48 minutes ago | parent
shoobiedoo 41 minutes ago | parent
cryptoegorophy 35 minutes ago | parent
lioeters 4 minutes ago | parent
bigglebear 30 minutes ago | parent
Otherwise, there will be no end to this. There are no hard limits on the speed of a parallel bruteforce. It's an infinite complexity problem class. The more parallel bruteforce power you have, the more likely you are to be able to solve a problem. So there is no world where demand ends. So if something isn't done about this, we'll have million dollar GPUs and RAM sticks because they've priced everyone out of the market and are the only ones able to afford them. Say goodbye to owning your own hardware at that point.
ollysb 18 minutes ago | parent
deathanatos 15 minutes ago | parent
https://en.wikipedia.org/wiki/Cornering_the_market
Right now, the RAM market is effectively cornered. Anti-trust enforcement is what I'd look to, but the GOP do not believe in ensuring competitive markets / enforcement. (The current admin is pretty clearly 110% pro corporation.) Vote in November, but in all likelihood, there won't be government intervention earlier than 2028, and even that is optimistic. One hopes the AI bubble pops, but I think this market can remain irrational for far longer than I can keep old hardware alive.
xbmcuser 28 minutes ago | parent
AmazingTurtle 10 minutes ago | parent
Barrin92 26 minutes ago | parent
smy20011 23 minutes ago | parent
p0w3n3d 23 minutes ago | parent
SillyUsername 21 minutes ago | parent
- memory compression algorithms
- alternative LLM architectures that don't rely on memory or GPUs
- compatibility hardware (like DDR3 to DDR4 boards)
- distributed computing improvements, both at local GPU and networking levels (SLI for AI)
- GPU hacks to add more memory or support older architectures
I'm personally looking forward to the new LLM architectures that don't require as much compute, e.g. DLLMs, which can be good enough for CPU usage but lack the accuracy of frontier models currently.
When this happens the bottom will fall out of the GPU and memory markets, putting a glut of cheap hardware out there.
Doom mongering like this never seems to include these as viable future alternatives, which is standard market adjustments, I wonder who the doom narrative helps? :)
YuechenLi 17 minutes ago | parent
Interesting effect is that since DRAM production tooling has been switched from DDR/GDDR to HBM, we may finally see the proliferation of HBM in consumer GPUs/accelerators after the bust cycle starts.
Catloafdev 6 minutes ago | parent
There's no indication that this will change, and every indication this will continue to grow. This is not some temporary thing. Demand is already exponentially higher than what is possible to produce, and there's no current reason to believe it has any ceiling.
Catloafdev 16 minutes ago | parent
uejfiweun 5 minutes ago | parent
belZaah 1 minute ago | parent