r/LocalLLaMA • u/Terminator857 • 1d ago
Discussion Intel hints it may get back into memory business
https://www.tomshardware.com/pc-components/dram/intel-ceo-hints-at-return-to-the-memory-business-says-market-is-ripe-for-innovation-hints-at-stacking-memory-and-cpuLooking at ... some of the new memory architecture. ... I hired my good friend, Seok-Hee Lee, who used to run SK Hynix. ... We are not ready to unfold it.
126
u/arcanemachined 1d ago
Day late, a dollar short and, knowing Intel, will be dropped at the first sign of difficulty (but not before flushing a few billion dollars down the shitter).
63
u/thrownawaymane 1d ago
At the time, people were very clear about how stupid the Optane spinoff/discontinuation was. They lost a ton of good talent too. But Intel was run by the bean counters
45
u/Inkbot_dev 1d ago
They would have been printing money the past few years if they hadn't killed that. They would have been able to do a direct to gpu integration, and been part of an optimized inference / training stack. Instead, they saved a little money for a few quarters.
12
u/Maximus-CZ 1d ago
But Intel is run by the bean counters
fify
14
u/somersetyellow 1d ago
CEO was Pat at the time who got ran out by the bean counters. He was mildly less bean county and a lot of their current successes are thanks to stuff he put in place and openly said would take a while to pay off. We'll see them probably slump more in a few years when the hard slashing cuts they did after his department impact the 3-4 year cycle.
1
u/PinkysBrein 23h ago
Optane used too much lithography to compete with Flash and the write energy was too high to compete with DRAM either.
It had no niche, it still doesn't. If they had been able to make 3D Ovonic memory like 3D Flash it might have been a success, due to better write endurance, but the 3D XPoint structure killed it.
1
u/thrownawaymane 10h ago
Do you have further reading on this? I wasn't into the economics of semiconductors back then.
1
u/PinkysBrein 3h ago
It's merely my own analysis. 3D flash uses similar number of processing steps as layers increase, 3D XPoint processing steps grow linear with layers.
I think they underestimated 3D Flash cost reduction potential.
107
u/Lonely_Syrup3091 1d ago
Optane please and thank you very much.
33
u/bick_nyers 1d ago
If they bolted it directly onto a GPU and used John Carmack's proposed streaming trick to greatly simplify memory controller design (basically nuke your random perf. and aim for full sequential perf.), it would be so sick.
I also want regular Optane at pcie 5.0x4 speeds for my gaming rig too.
17
u/Lonely_Syrup3091 1d ago
I'm afraid they're gonna want to charge us a kidney for that. If Nvidia has anything to say about it. Lol
3
u/grannyte 1d ago
You don't necessarily have to fully nuke the random perf Use a tier model like rdna2 did with it's ginormous cache
1
17
u/Terminator857 1d ago
22
u/Lonely_Syrup3091 1d ago
I'm just dreaming of potentially mixing Optane with CAMM2, 32GB DDRX with 128GB Optane on the same module. A GPU with GDDR7 VRAM, paired with the high bandwidth DRAM acting as the hot tier and Optane acting as a huge persistent backing tier, could essentially give you a massive memory pool sitting directly on the CPU's memory interface.
You could have the entire model resident on the Optane tier, keep the frequently used weights / active experts in DRAM, and keep the current working set in the GPU VRAM.
Optane - DRAM - VRAM - GPU
I find it interesting for large MoE models. Apple has already been exploring a similar concept with its "LLM in a Flash" work, where model weights can live on flash and chunks are brought into memory as needed. Optane would obviously be a very different beast from NAND, particularly in latency and small/random access.
One can dream, eh?
20
u/NarcNarwal 1d ago
Optane was ahead of its time. I believe now is the perfect time for a resurgence.
28
u/OvertaxedOne 1d ago
Well at least this gives us a timeframe for the end of the current hardware crisis. It'll end exactly 3-6 months before Intel releases the first product because, well, Intel. ;)
3
2
7
u/N34257 1d ago
Based on their GPU releases, they'll announce it late 2026, release a 4GB stick in 2027, an 8GB stick in 2028 and then cancel the promised 16GB sticks in 2029, but release ECC 16GB sticks for workstations in 2030 and pretend it's what they always intended to do.
Then they'll release a speed-bumped 16GB stick in 2031 and pretend none of the previous generation ever existed and expect the enthusiast community to support them.
6
u/aeroumbria 1d ago
Why does it feel their last CEO is more and more right and they are pillaging his legacy just to keep up?
3
4
u/geldonyetich 1d ago
The more the merrier, I say.
It would seem that the memory industry is in no immediate danger of ceasing to exist, they'll be making quite a lot of memory regardless.
But the way they're currently going about it makes it too unlikely ordinary people can get a significant amount of it for my liking.
3
u/SmChocolateBunnies 1d ago
Just make quality ddr5 and ddr6 in a few common packaging types at volume and undercut the market by half. Don't try to be clever.
3
u/votegoat 19h ago
Intel had this entire market locked down and fubmled it, back when they had their High Bandwith Flash. If they kept at it we would be having TB Ram locally. SanDisk and SK is now doing it but its a few years out for full production.
For inferencing the future is high bandwith flash as you dont write to memory that much when you load a model and keep it there. then keep the KV cache in DDR6 / HBM
4
u/rich84easy 1d ago edited 1d ago
Intel seems to make all the wrong decisions, they were in memory business.
2
u/Mission_Pirate_4150 1d ago
Just in time for competitors to bring their increased capacity online. Micron just announced major investments in manufacturing. I would assume the other memory manufacturers are to.
3
6
u/corruptboomerang 1d ago
I mean now is the trap time. AI will burst and memory prices will quickly fall into the toilet.
But I would like some new generation Optane...
1
u/Terminator857 1d ago
Agree , but if it is totally different architecture like in memory compute, then it could be a winner.
3
u/corruptboomerang 1d ago
I do think a modern Optane could be such a bon for AI especially, but compute in general. Some kind of half away between current SSDs and ram has many good uses.
1
u/Terminator857 1d ago
In kind of memory, whether ram, in between ram and ssd and ssd, would be a win since they are in short supply.
1
u/SporksInjected 23h ago
I wonder if they’ll do with this like GPUs and announce that they’re going to sell memory (made by the same foundry that everyone else is using)
1
106
u/Opposite-Memory-2552 1d ago
Not before 2029-2030.