23 comments

  • ricardobayes5 hours ago
    The article features a truck that the chips get loaded onto. This quote from Tanenbaum immediately jumped into mind:<p>&quot;Never underestimate the bandwidth of a station wagon full of tapes hurtling down the highway.&quot;
    • leoedin4 hours ago
      I was wondering how much financial value they put in one truck. Presumably the value density of this stuff is incredibly high (although not really to anyone else, so security probably isn&#x27;t a huge issue). At what point do you split the shipment between multiple trucks to avoid huge losses if there&#x27;s a crash?<p>When I worked on satellites we had some electronics which used a radiation hardened FPGA which cost something like $500k per chip. Carrying that around was nerve wracking.
      • quadragenarian28 minutes ago
        Ha - I still remember when I worked on the IBEX-Lo Sensor, we had to hand carry it to Switzerland. One lucky member of our team got to fly first class with it because it wouldn&#x27;t fit in a coach seat!
      • ceejayoz3 hours ago
        I once had to keep a $20M (technically; a million download codes for the same $20 movie for a promotion) thumbdrive in my home safe because it arrived too late to put in the safety deposit box like it was supposed to be.<p>I had to remind myself a few times the <i>practical</i> value to anyone was about $22.
        • fluoridation3 hours ago
          Why did that have to be transported in a thumb drive? It can&#x27;t have been more than a few hundred megabytes at the most, right?
          • ceejayoz2 hours ago
            Enterprisey client chain-of-custody requirements. It was couriered to us, and the courier was late. Cue panicked phone calls to get an exemption to the contractual requirement for overnight storage at the bank.
      • victorbjorklund1 hour ago
        They probably have insurance which means it doesn’t matter if it is all in one car or not. And even if they didn’t Samsung is big enough where a loss of a couple million dollars of components isn’t gonna bankrupt them.
      • kotaKat2 hours ago
        It&#x27;s treated the same way as any other generic shipment of cargo, really. It&#x27;s <i>fine</i> throwing $20 million in one truck.<p>Security <i>is</i> an issue, though, these are treated as a high value load and typically would be followed by unmarked, plain clothes armed security guards tailing around the truck. (That&#x27;s how Apple often does warehouse deliveries of trucks full of iPhones. It wouldn&#x27;t surprise me if there was going to be some rolling lead soon following some trucks full of Duos...)
    • sebastiansm73 hours ago
      Almost 10 years ago at my university in Chile. My professor of Computer Networks invited a PhD Student that was working on the ALMA Observatory [0] doing some research on signal processing.<p>What stayed with me from that talk was that all the data that was captured from those massive antennas was packaged in boxes, loaded in a cargo van and travel 1.600 kilometers to Santiago to be uploaded into the systems.<p>[0] <a href="https:&#x2F;&#x2F;www.almaobservatory.org" rel="nofollow">https:&#x2F;&#x2F;www.almaobservatory.org</a>
    • hacker_homie4 hours ago
      Just hope nothing changes the latency is terrible.
  • y1n018 hours ago
    I’ve often wondered about die thinning. I mean just looking at sd cards and micro sd etc, it was clear that there had to be a thinning step (or many steps) and it seemed amazing to me that such a complex step could be economically viable. And yet it clearly is. All a matter of being able to operate at scale I guess.<p>It’s not a step I’ve seen discussed very much, so it was kind of neat to see it discussed front and center in a lay article like this.
    • pjc504 hours ago
      Planarization has always been part of the process: <a href="https:&#x2F;&#x2F;jeez-semicon.com&#x2F;blog&#x2F;Planarization-in-Semiconductor-Manufacturing-Complete-Guide&#x2F;" rel="nofollow">https:&#x2F;&#x2F;jeez-semicon.com&#x2F;blog&#x2F;Planarization-in-Semiconductor...</a><p><a href="https:&#x2F;&#x2F;resources.pcb.cadence.com&#x2F;blog&#x2F;2023-the-planarization-process-for-semiconductor-manufacturing" rel="nofollow">https:&#x2F;&#x2F;resources.pcb.cadence.com&#x2F;blog&#x2F;2023-the-planarizatio...</a><p>Not dissimilar to lens or mirror making, the process of grinding to a precise flatness. The key step here is bonding it to a glass carrier, presumably because otherwise it gets too thin to support itself.<p>See also the &quot;reduced kerf diamond wire saw&quot;: technology critical to cheap solar panels is the ability to saw the boule into thinner and thinner pieces with less waste at each cut, like a big ham slicer.
    • KurSix10 hours ago
      Yeah, the scale is kind of mind-bending. You&#x27;re taking something that&#x27;s already extremely fragile, grinding most of its thickness away, handling it through several more process steps, and somehow doing that cheaply enough for mass-market products.
      • xadhominemx7 hours ago
        The critical back end step for HBM is thermal compression bonding - where they actually compress 13 very thin die separated by solder balls. Cracking is a major issue.
        • pixl972 hours ago
          We are getting so many technologies like this that are so insanely specialized and require such deep knowledge of physical processes that some event like a 1918 flu that kills 10% of the population could take out enough people that we&#x27;d be unable to make things like this for years as we fill in the missing knowledge workers.<p>It&#x27;s just crazy how far down the scale we&#x27;re tickling atoms these days.
      • lazide7 hours ago
        Automation, automation, automation.<p>Zero chance you could do it if a human had to touch the die&#x2F;chip directly during those steps.<p>But clean room + vacuum pick and place machine? A matter of fine tuning, and it should work just fine.
  • HarHarVeryFunny20 hours ago
    On a related note, I was reading yesterday that apparently the real bottleneck for Chinese production of AI accelerators is HBM production, not processors or ASML equipment.<p>The lack of ASML EUV machines certainly hurts, and pushing DUV so hard results in abysmal yields of good chips, but you can compensate by running more wafers or making smaller chips, and the net result is that Huawei&#x27;s Ascend production volume is limited by CXMT&#x27;s HBM capacity not processor dies.<p>The problem is that HBM manufacture requires many steps (die thinning, via drilling, plating, alignment) where the equipment used by everyone else (Samsung, SK Hynix, Micron) is also blocked by sanctions, so the Chinese are having to develop all of this themselves too, which they have, but yields are currently low, even when using shorter HBM stacks.
    • jacquesm16 hours ago
      The only thing the West is achieving here is that sooner or later China will be able to compete on their own terms rather than ours. It may buy some time but the end result is very predictable.
      • bigcat1234567813 hours ago
        As other comments pointed out, that was always China&#x27;s plan, the value here is really that when the Chinese overlord annillate the world&#x27;s industrial competition, they&#x27;ll have the moral high ground of being out of self-reliance and hostility by western monopolists.<p>Edit:<p>And, that, my fellow keyboard high-minded thinker but powerless corp peons, was also part of the China&#x27;s plan all along.
        • nixon_why6911 hours ago
          Eh. Industrial policy in a big country is always the output of politics between stakeholders, anthropomorphizing it to a &quot;guy with a plan&quot;, especially an &quot;overlord&quot; looking to &quot;annihilate&quot; is probably getting it wrong.<p>A friend of a friend is married to a member of the national academy of sciences in China. I heard at dinner a few years ago that he had been lobbying for more onshoring of chip production for years before the Nvidia restrictions, and always heard &quot;you&#x27;re right, maybe next year we&#x27;ll be able to fund that&quot;. Then the restrictions came in and everything changed.
          • lazide7 hours ago
            There are literal 5 year plans, eh?
            • peterfirefly6 hours ago
              And they are more about sounding good than about reality. They also affect a much smaller share of the economy than they used to -- which is very good! The original five year plans caused nothing but misery and starvation.
              • lazide4 hours ago
                I’m pointing out it’s disenguous to say ‘it’s not like there was a plan’, when there was explicitly a plan with that title on it. Especially if the plan didn’t work out well.
                • nixon_why693 hours ago
                  It&#x27;s disingenuous to characterize &quot;don&#x27;t anthropomorphize the political process&quot; as saying no planning documents exist as part of that process.<p>And, to the point I was responding to originally, we can check those documents! They don&#x27;t speak of &quot;annihilating&quot; other countries&#x27; industrial capacities.
                  • pixl972 hours ago
                    &gt;They don&#x27;t speak of &quot;annihilating&quot; other countries&#x27; industrial capacities.<p>Why would they have to?<p>When China talked about their 5 and 10 year plans on solar panels there was no need to talk about putting the rest of the competitors out of business. That was a highly probable end goal of what they were doing, no need to state it.<p>It would be like Walmart stating &quot;We&#x27;re going to put out massive loss leaders in sporting goods to put X out of business&quot;. Instead they&#x27;ll talk about growing market share and long term customer loyalty.
                    • nixon_why691 hour ago
                      I&#x27;m sorry, I&#x27;m not following. Is the problem here that China is too capitalist or too communist?
                      • pixl971 hour ago
                        Yes.<p>More of, China is smart enough to use the benefits of both capitalism and authoritarianism to grow their country while western governments are busy stealing the furniture and fixtures off the walls.
                        • nixon_why691 hour ago
                          I agree with 100% of that comment but I&#x27;m struggling to understand the beef?<p>Of course they&#x27;re trying to take care of themselves while the US implodes, how can you blame them for that? The whole world has a serious oil problem right now, because of US actions, and of course each country is dealing with it in their own way. Why is that a problem?
                          • pixl971 hour ago
                            You do realize that a lot of people have had problems with America after WWII and using their economic and military power to dictate what the rest of the world does, right?<p>&quot;Problem&quot; is a relative term for those dealing with the consequences of it. It is one if you are America means a whole lot less economic freedom than you had in the past. If this is a good thing or not depends on your position.
                            • nixon_why6954 minutes ago
                              If your beef is that they are succeeding or at least failing less while America steps on it&#x27;s dick repeatedly, maybe this isn&#x27;t their problem?<p>If you&#x27;re fired up by nationalism but your nation is sucking, getting mad at the Chinese isn&#x27;t the answer. And that&#x27;s before we even get into relative levels of activism, Chinese political ideology is much more biased towards &quot;neutral, just trade&quot; than American politics.
                      • fragmede56 minutes ago
                        False dichotomy. the problem is far deeper than that, and wanting to simplify the problem to those two extremes belies a belief that the world could and should be able to be simplified into two power dynamic. It is not. The world, and the powers that fight within in, are far more dynamic than a simple communist or capitalist spectrum, and attempts to simplify the world along those lines is only going to hurt you.
            • nixon_why696 hours ago
              That just means your stakeholders argue twice: once when drafting the five year plan and then again, later, saying their policy position is more aligned with the spirit of the plan.
        • FooBarWidget4 hours ago
          China&#x27;s plan to be self-sufficient in semiconductors was a failure before western sanctions, because all the Chinese chipmakers found it too easy to buy western chipmaking equipment or to fab through TSCM. No Chinese company wanted to buy from other Chinese companies. Nobody was incentivized to develop self-sufficiency. &quot;Because the government wants it&quot; was not enough incentive.<p>Western sanctions have succeeded in achieving China&#x27;s self-sufficiency goals where their own government failed.
        • Woodi11 hours ago
          &gt; [...] that was always China&#x27;s plan, [...] they&#x27;ll have the moral high ground<p>It&#x27;s a possibility. It is also possible that they will steal yesterday&#x27;s technology and still pretend to have moral high ground.<p>I do not wish China and Chinees wrong but most probably communism will work as it always do - by faking thing and bringing self-harm to their own citizens.
          • mitxela7 hours ago
            In free markets, everyone steals everything they can, and stops everyone else stealing from them as much as they can. This is normal and expected. The ideology says that it will even out and converge to a fair situation.
            • hobo1232 hours ago
              And by stealing you mean exchange? Or I somehow misunderstood what you consider a <i>market</i> to be - a place for pickpockets?
            • philipallstar3 hours ago
              &gt; In free markets, everyone steals everything they can, and stops everyone else stealing from them as much as they can. This is normal and expected. The ideology says that it will even out and converge to a fair situation.<p>This is all nonsense. Capitalism isn&#x27;t an ideology, it&#x27;s just a slight formalisation of what naturally arises. People are free to choose; they can own things; they can make agreements; they must honour the agreements or pay a penalty. Capitalism in no way is about stealing.
              • andrekandre45 minutes ago
                <p><pre><code> &gt; Capitalism isn&#x27;t an ideology </code></pre> your confusing trading&#x2F;markets with capitalism; they aren’t the same thing, and it definitely has ideology(ies)<p><a href="https:&#x2F;&#x2F;www.britannica.com&#x2F;money&#x2F;capitalism" rel="nofollow">https:&#x2F;&#x2F;www.britannica.com&#x2F;money&#x2F;capitalism</a>
            • lazide4 hours ago
              It’s the same in communism&#x2F;socialism except people are prevented from stopping everyone else from stealing their stuff, in the name of the ‘greater good’, and if they try to get more than everyone else get guilt tripped into giving it up under the ‘to everyone what they need, from everyone what they can’.<p>Which incentivizes pretending you need a ton, and can give nothing. You think being useful at work in Capitalism is bad? Wait until you are accidentally useful in a communist&#x2F;socialist economy.<p>There is usually less pretending under capitalism, but people are people, so it all the same stories still play out.<p>People are shitty, if given the chance and influenced in specific ways. Capitalism or not. We’re dealing with the current bullshit in the west because somehow people bought the line that everyone is a good person - and we totally can’t actually see or do anything if we see someone doing something bad! - and we need to always follow ‘the rules’ even when someone bad clearly isn’t.<p>Until we find our balls, good luck.
          • cumshitpiss4 hours ago
            [dead]
          • eecc11 hours ago
            Oef, still kicking the “communism” horse. It’s dead Jim, especially in China it’s dead and buried
        • fangspire12 hours ago
          Was it also the Chinese plan all along to annihilate their own birh rate?<p>Was the Great Leap Forward part of the plan too?
          • vkou10 hours ago
            &gt; Was it also the Chinese plan all along to annihilate their own birh rate?<p>No, that&#x27;s generally what happens when countries stop being poor. Taiwan, Singapore, Korea, all have similar fertility rates.
            • rmunn9 hours ago
              But none of them had a government-mandated one-child policy (maximum, not minimum) for over three decades. That&#x27;s a pretty major difference between China and neighboring countries.
              • vkou7 hours ago
                And yet all of them have an even lower fertility rate than China.<p>On the one hand, you&#x27;re proving my point. On the other hand, if you think that China&#x27;s headed for a demographic disaster, <i>so are all of those countries</i>, and your thesis of blaming the CCP for everything... Falls a bit flat.
                • ricardobayes5 hours ago
                  China is headed for a demographic disaster, though. High schools are incredibly packed, middle schools, not so much and primary schools and kindergartens are kind of empty.
                  • lazide4 hours ago
                    Everyone is, not just China.
                    • belowavgiq3 hours ago
                      Everyone stupid enough to let that happen is.<p>There is a handful of developed countries where this is not an issue.
                      • HarHarVeryFunny2 hours ago
                        It&#x27;s hard to say that it&#x27;s stupidity (at least not country-specific stupidity) - it seems that globally, presumably due to the state of the world along many axis, many people are choosing to have fewer children, and it&#x27;s happening on a large enough scale to have major consequences for things like the size of your tax base, trying to keep your economy growing, demographic concerns over an aging population with no-one to look after or pay for them etc, etc.<p>Different countries are handling this in different ways, with encouraging immigration being the major one. Japan is facing a major lack of workers, with many businesses having to shut down as a result, but have decided to tackle it in a different way relying on advanced robotics as an alternative to immigration.<p>There are very few developed countries NOT having this problem. Most of the population growth is in poorer parts of the world like Africa.
                        • belowavgiq2 hours ago
                          Well, if there are &gt;0 developed countries that have solved this (Israel, the GCC, dare I say fairly well-off MENA countries like Jordan, Algeria, Morocco, etc., and to a minor degree Botswana) then this proves wrong the hypothesis that modern hyper-development must inherently mean population collapse.<p>Going purely by the evolution theory and natural selection, having so few children that there is a serious risk a population will disappear in a century or two, means losing. Needless to say, this is stupid. From the entire American continent, to Europe, India more recently, Thailand, China, the Koreas, Japan.<p>Africans, Arabs, and some non-Arab Muslim populations will carry the torch. Much respect for not falling into the trap.
                          • pixl971 hour ago
                            &gt; then this proves wrong the hypothesis that modern hyper-development must inherently mean population collapse.<p>Honestly you&#x27;re reading the data completely wrong. You cannot look at it at the country level. In Israel for example the populations that are producing the most population are highly religious and less educated than those that are producing less population.<p>It&#x27;s more about memes. Religion is a meme driven breeding cult. This is why said religions tend to have very long term staying power. The particular problem here is these religions come at a cost of having to ignore a lot of reality to get to that point. Because they are faith based anything that presents a direct challenge to that faith must be setup as an enemy or it may risk the religion. Yes, religions can incorporate new information but it tends to be a slow process as it needs to involve slowly changing peoples mind to avoid instability.<p>What this looks like in the long term is large populations of highly religious that have a fundamental disconnect from the technology that&#x27;s keeping them alive. A new Late Bronze Age collapse.
                            • belowavgiq51 minutes ago
                              So? How does the population being religious disprove the fact that they are developed and having children? Not to mention that I have listed around ~10 countries after Israel.<p>Uh... Meme-based? This is such an unserious argument I don&#x27;t even know what I can possibly say to that.<p>How are they disconnected with the technology keeping them alive? All people, in GCC countries especially, are working to make 1st world development happen.<p>Implying that religious Jews or Muslims are stuck in the Bronze Age is ignoring reality. Most European development itself came from men of science who were Christian, I thought this apparent conflict of ideas was something taught in schools.
                              • pixl979 minutes ago
                                First, you seems to think a meme is something unserious. Read Richard Dawkins The Selfish Gene so you know what meme actually stands for rather than the thing you see on Reddit.<p>GCC countries are resource economies mostly. They must weave a very narrow path to success to non-oil based means of income or the entire region collapses. I will say they do realize this is a problem and are using oil money to diversify, but resource curses are very hard to break and collapse can happen very quickly.<p>&quot;My grandfather rode a camel, my father rode a camel, I drive a Mercedes, my son drives a Land Rover, his son will drive a Land Rover, but his son&#x27;s son will ride a camel.&quot; --Sheikh Rashid bin Saeed Al Maktoum<p>&gt;Most European development itself came from men of science who were Christian<p>This is just stating that that in 1700 you were a Christian or you got hung (hell, pre-independence America was founded by people that thought that too few people were being hung). As we see when you take out the hanging or social shame part of the religion people tend to flee them as they are oppressive structures. You could say them being men of science either had no bearing to their religion, or was in spite of their religion.<p>I am saying, for example, that the Abrahamic religions are bronze age (ok, sorry, Iron age) social structures. I don&#x27;t think studied scholars are going to disagree on that point very much. People saw how people worked and could be sheppereded in societies without formal education and turned it into a series of stories that were effective in modifying peoples behaviors. Kings later came along and added their stories so they could further control the population. Over time a semi-coherent story about how people worked was formed below a set of rather incoherent stories regarding things that didn&#x27;t happen that people are supposed to accept with faith.
                          • HarHarVeryFunny49 minutes ago
                            Don&#x27;t worry, we&#x27;ve got Elon Musk trying to single-handedly repopulate America.<p>And if that fails, we&#x27;ve got &quot;Idiocracy&quot; to save the human species as the dumb-asses pump out babies as fast as you like.
                      • lazide3 hours ago
                        Population collapse? Who, besides Zimbabwe?
                • inglor_cz4 hours ago
                  Still, a stupid governmental policy cost them at least 100 million people compared to the expected spontaneous development. Probably more. That is the sort of punishment that omnipotent governments sometimes visit upon their folk.<p>I don&#x27;t believe that even the CCP itself, in hindsight, considers the one-child policy to have been a good decision. At the very least it should have been ended some 15 years earlier.
                • badpun5 hours ago
                  China has extremely low fertitily rate given that they haven&#x27;t yet reached the status of a developed country (which Taiwan, SK and Singapore have reached).
          • lazide7 hours ago
            Yes, they literally named it the ‘one child policy’. It was just overly successful.
            • peterfirefly6 hours ago
              The official one child policy didn&#x27;t actually have much of an impact. Birth rates had already fallen a lot, despite the official &quot;lots-a-children&quot; policy.<p>The one child policy is gone now but birth rates haven&#x27;t really moved upwards.
              • nixon_why696 hours ago
                It&#x27;s debatable and impossible to disentangle from living standards going up, but a lot of the economy is built on assumptions of one child, from apartment sizes to personal transit (ebike can do 2 adults + 1 small child).
              • ricardobayes5 hours ago
                Of course it had an impact, if nothing else, a cultural one. It basically eradicated the concept of large families.
                • pixl971 hour ago
                  So China has similar family sizes as the US that had no one child policy.<p>It&#x27;s difficult to near impossible here to determine which counterfactuals would actually hold true. Based on other world wide trends it&#x27;s very likely that Chinese family sizes now would only be slightly higher than if the policy didn&#x27;t exist. In the modern world there are just a huge number of different factors working together to decrease family sizes.
              • lazide4 hours ago
                Not what I heard from friends who were subjected to it, and their parents.<p>It was insanely disruptive.
      • audunw10 hours ago
        What the West is doing is forcing China into a situation which is impossible to sustain when their economic dividend runs out. Especially considering these investments are fuelled by an extremely high level of debt.<p>And it’s not a given that they’ll catch up. I think it’s fair to say that their development of jet engines is on track to be obsolete before they’re competitive.<p>They may in principle have established the technologies to be self sustained but that’s only relevant and sustainable if they develop an economy based on domestic consumption rather than exports. Which they are struggling with as well.
        • chii10 hours ago
          &gt; their economic dividend runs out<p>i dont think it will - they still have a chokehold on battery production (which implies dominance in the EV sector), and they also control the vast majority of all rare earths. The chinese gov&#x27;t can subsidize chip research&#x2F;development with these other sectors that are dominating.<p>The west is the one that need to watch out for their own economic dividend running out from the past century.
          • sysguest8 hours ago
            another thing: chinese gov can just some of its wipe its populace easily (the gov controls media 100%)<p>but... not the western govs
        • HarHarVeryFunny3 hours ago
          What makes you think they don&#x27;t have a domestic market?<p>The Chinese domestic EV market is the largest in the world.<p>Chinese factories are full of Chinese robotics.<p>Companies like Huawei are struggling to meet the domestic demand for AI accelerators, while growing volume very fast year over year.<p>The Chinese domestic AI market is massive, but we just don&#x27;t hear much about it outside of the models that are also sold here. Can you even name the models that ByteDance (China&#x27;s Meta) makes?<p>As far as exports go, China is a low cost producer for many things. If the US doesn&#x27;t want to buy high quality EV&#x27;s costing 1&#x2F;2 the price of a Tesla, then there are plenty of other countries happy to do so. The US is only 5% of the global population - there are lots of other consumers out there.
        • pinkmuffinere10 hours ago
          &gt; They may in principle have established the technologies to be self sustained but that’s only relevant and sustainable if they develop an economy based on domestic consumption rather than exports. Which they are struggling with as well.<p>Is this really true? Can&#x27;t they just keep exporting &quot;forever&quot;? What requires them to find a domestic audience? I can think of some other countries that have done quite well mostly without finding a domestic audience (eg, Switzerland comes to mind, and frankly Europe might largely be in the same bucket).
          • T-A10 hours ago
            <a href="https:&#x2F;&#x2F;www.imf.org&#x2F;en&#x2F;publications&#x2F;fandd&#x2F;issues&#x2F;2023&#x2F;06&#x2F;superpowers-are-forsaking-free-trade-ngaire-woods" rel="nofollow">https:&#x2F;&#x2F;www.imf.org&#x2F;en&#x2F;publications&#x2F;fandd&#x2F;issues&#x2F;2023&#x2F;06&#x2F;sup...</a>
            • jacquesm9 hours ago
              Totally non-biased of course.
              • pixl971 hour ago
                Nothing exists in the world that is not biased. The key to knowledge is understanding what those biases are.
          • yxhuvud9 hours ago
            Europe (by which I assume you mean EU or possibly Euro area) has a fairly balanced balance of trade overall.<p>And a teeny-tiny country like Switzerland can&#x27;t be usefully compared to something as big as China - the rest of the world can easily consume whatever the Swiss produce, but there are limits to how much Chinese can produce until consumption will have to shift to domestic consumption because the world doesn&#x27;t have infinite ability to pay for its demand.<p>That said don&#x27;t expect any big change anytime soon - there are very strong political incentives to keep the status quo with China focusing a bit too much on exports.
      • stingraycharles15 hours ago
        It baffles me how shortsighted the policymaking here is. Like, what did they expect to happen?
        • adrianN14 hours ago
          They expect to retire with a decent amount of money before things go down hill.
        • odo124213 hours ago
          They probably weren&#x27;t expecting AI foundation models &#x2F; model research to be quite as fungible as they are
          • safety1st12 hours ago
            Yes.<p>Over the last 20 years the economy has become dysfunctional. It no longer really resembles a free market; monopolies have established barriers to entry everywhere. And the biggest investors are awash with helicopter money that&#x27;s been doled out for favors by the political class.<p>So those investors have tons of cash to burn and surprisingly few opportunities. Even within Silicon Valley&#x2F;VC there are surprisingly few who seem to really understand the fundamental economics of software. Or perhaps those economics just aren&#x27;t that important when you have billions of dollars on hand and cash is obviously not going to get you a return. Any whisper of possible exponential growth is worth throwing money at. Crypto? Why not. AI? Why not. Datacenters? Why not. Tulips? Why not.<p>This is by all definitions an empire in decline. Everything is broken or fake. Everyone is afraid to do what needs to be done. Power forbids it. So we&#x27;re all just waiting for the other shoe to drop. Our secret police aren&#x27;t as bad as the late stage USSR&#x27;s yet, but hold Uncle Sam&#x27;s beer...<p>(That&#x27;s not a recommendation to try and time the nadir, by the way, as it could easily be 50 years away.)
            • user439285 hours ago
              Datacenters and AI obviously work and come with a credible value proposition and path to profitability.<p>Comparing this to crypto or turnips makes me doubt the merit of any other opinions your comment offers.
              • pixl971 hour ago
                The initial growth of the internet was a boom bust cycle, and as you see the internet is still here. This is more about the economic paradigm that is being used to grow a technology versus the economic paradigm the technologies rate of growth can support over time.<p>Just to give some rough numbers in the past 5 years the amount of GPU compute installed in FLOPS is somewhere over 5 times all the CPU flops that have ever existed.<p>This has nothing to do with AI being good or bad or being able to produce things and economic value. It is by far the largest and fastest growth of any technology ever and we have zero clue about the economic stability of this grand experiment we&#x27;re performing.
                • user439281 hour ago
                  The internet is a much more appropriate comparison than crypto or tulips, which bring no or negligible value.<p>Some differences to the market at that time seems to be that during the dotcom bubble, many of the companies had little to no revenue.<p>Leading AI labs already generate enormous revenue. The investments into capacity are needed to address the current demand.<p>The situation seems somewhat less speculative.<p>That said, I cannot predict how AI capabilities will develop and how demand will respond.<p>Should capabilities plateau hard and soon, maybe the demand will not be there for the compute investments.<p>If it does not, and instead AI applications in robotics, science, and self driving expand, chances seem reasonably good that the demand will be there, no?<p>And as for the economic stability, much of the investment comes from existing giants like Microsoft, Alphabet, Amazon, and Meta, who have the necessary cash flow.<p>These companies are less likely to collapse than some of the ones during the dotcom bubble.
                  • pixl971 hour ago
                    Eh, I&#x27;d say it&#x27;s closer to something like internet + tulips.<p>It&#x27;s the total amount of money in the economy that&#x27;s been invested toward a potential outcome. AI represents the largest amount of money, and largest fractional part of the economy invested ever.<p>Because of this AI could be the biggest economic boon ever, yet still not recover the full amount invested. This will have deep economic impacts that affect everyone and everything. At this point AI must achieve all its stated economic impacts or there will still be a huge economic crash that kills off any company that is over invested and cannot make a profit.<p>Worse, the many of the perceived economic impacts of AI are not for humans like you or I, but the huge companies you listed. Even if they economically win, everyone else made out of meat could still lose.<p>They say history doesn&#x27;t repeat, but it does rhyme. This, at least to me, sounds like a mixtape of &quot;internet&quot; + &quot;tulips&quot; + &quot;1920s financial world leading to global political instability&quot;.<p>Every potential outcome I see occurring pushes us closer to further instability, even if the economics on it work out on paper.
                    • user4392811 minutes ago
                      &gt; At this point AI must achieve all its stated economic impacts or there will still be a huge economic crash<p>I had a look at the numbers.<p>The investment into data centers is estimated around $800B&#x2F;year currently.<p>OpenAI + Anthropic combined had ARR of &gt;$100B in July.<p>Global labor income is around $65T&#x2F;year, the US GDP $32T, the global one $126T.<p>Global software spending: $1.4T&#x2F;year.<p>I am no financial analyst, but I don&#x27;t think all the stated AI impacts have to be met just to recover the investments.<p>A 1% productivity gain on global labor income represents $660B&#x2F;year. At 5% we would look at $3.3T&#x2F;year.<p>They don&#x27;t need to cure cancer to justify the investment, even though Amodei hopes to cure most major disease in the next 5-10 years.
            • piva008 hours ago
              It started way before the last 20 years, the past 40 years have created the conditions for this to happen.<p>Overfinancialisation of the economy while removing safeguards to keep markets functional, like anti-trust enforcement, are the main forces behind this erosion. Through finance the focus of companies is completely shifted away from producing good products and services, and into how to extract more paper wealth from existing structures to the detriment of products.<p>Not enforcing anti-trust and letting behemoths to form which cannot be competed against since with their amount of capital they can either buy their competition outright or just price dump for long enough to make competition non-viable.<p>It&#x27;s the failure of neoliberalism, and that agenda has been pushed into Western countries since the 1980s-1990s, it made enormous wealth for the few at the top while eroding whole societies, economically and socially, it&#x27;s all dysfunctional.
              • odo12426 hours ago
                I feel like my point is being over extrapolated a bit. All I was saying is that people in 2020 believed AI companies had much more of a moat then they actually do - that other companies would struggle to obtain enough data (y’know, downloading the internet to train models wasn’t legal back then) or resources to compete
                • piva005 hours ago
                  I replied to the extrapolation, not to your comment directly, I think your point might be more fitting under that comment than mine, I just drilled down from the points above mine.
        • mrheosuper14 hours ago
          they expect the Chinese to bend their knee and beg for the sweet, sweet chip
        • qaq14 hours ago
          they were sold an idea the AGI is right around the corner so even a slight delay might be critical to secure the &quot;win&quot;
        • georgemcbay11 hours ago
          &gt; It baffles me how shortsighted the policymaking here is. Like, what did they expect to happen?<p>As someone who is myself &#x27;old af&#x27; by tech standards (I&#x27;ll be 53 very soon) the main problem is that the policymakers are all 70+ years old.<p>Their concept of what China even is is forever stuck in pre-2004 thinking from back when they still had significant neuroplasticity.<p>Vote these fucking fossils out, kids, before it is entirely too late.<p>(And, yeah, before someone inevitably brings it up, some exceptions-that-prove-the-rule older policymakers&#x2F;politicians do avoid this trap of getting stuck in the past, but most don&#x27;t)
        • ekianjo13 hours ago
          if you sell them ASML equipment your market gets destroyed tomorrow instead of 10 years later so what options do they have?
          • DeepSeaTortoise9 hours ago
            No longer strangling their own economy would be a good start. Take a look at the magnificance that was the ESRS. A 300 A4 pages long list of yearly reporting duties. And the best thing is: Just because they&#x27;re no longer nicely summarized in one place, doesn&#x27;t mean the underlying duties went away.
      • KurSix10 hours ago
        Sanctions can absolutely create the incentive to build a domestic supply chain, but incentive doesn&#x27;t automatically translate into catching up on the same timeline
      • deepsun11 hours ago
        Well Soviet Union tried to be self-sufficient, but the truth is such strategy loses to global free trade. Whether we want it or not, global trade is what keeps the world from WW3 for now, not the nuclear MAD.
        • jacquesm9 hours ago
          The Soviet Union never was the manufacturing power behind a very significant fraction of the worlds products.
        • klrefg10 hours ago
          Few countries rely more economically on global trade than China though.
        • hiddencost9 hours ago
          You understand that the US has alienated almost all of its trading partners and is aggressively attempting to kill free global trade, right?<p>Whereas China is developing a powerful network of rapidly developing trading partners like Africa and the Middle East?
      • nonethewiser14 hours ago
        They always had the option to do that
        • nixon_why6914 hours ago
          But not the necessity, so it wasn&#x27;t getting funded with the same urgency as now.
          • nonethewiser42 minutes ago
            But if it was always an option and they didnt do it then they didnt think it was in their best interest
      • ekianjo13 hours ago
        people have been claiming Nvidia will get competition too for years but you guys seem to be forgetting that the current leaders are not asleep at the wheel in the meantime.
      • iwontberude10 hours ago
        [dead]
      • SadErn13 hours ago
        When China eventually catches up on conventional lithography, the goalposts will shift to silicon photonics and optical logic, a transition DARPA laid the publicly visible groundwork for decades ago through programs like POEM and PIPES while keeping the mature implementations strictly classified until needed.
      • skeptic_ai14 hours ago
        I think eventually The west will want to kneecap China, so I guess this is the best time to do so.
    • andy_ppp19 hours ago
      Tokens per second is almost entirely memory bandwidth at inference time, training obviously needs more compute but you can add more chips for that.
      • martinald15 hours ago
        Not quite, it&#x27;s got quite a bit more complicated with agentic use cases.<p>Prefill (input tokens) is heavily compute bound. And the ratio of input to output continues to rise, as typically in agentic sessions you have a few tokens output for a tool call and (many) thousands of input from the tool result.<p>Then you have cached input tokens, which is a totally different issue, system RAM or NVMe bound.<p>Obviously output tokens is VRAM memory bandwidth bound, but this is less and less of the bottleneck these days for overall agentic speed.
        • andy_ppp6 hours ago
          This is why I was careful to specify tokens per second not time to first token which is the prefill step you’re talking about. Clearly to run these models well you need both but as I said adding more chips or compute units can give you more latency where as overall memory bandwidth (throughput) is limited by access to fast memory.
        • com2kid13 hours ago
          I can easily use close to 100 million input tokens a day. A few million output tokens but at the end of the day maybe a thousand or so lines of code get written.
      • cubefox18 hours ago
        According to SemiAnalysis, both inference and post-training (RLVR) is mostly memory bandwidth bound. Only pre-training is compute bound, but it now only takes a small share of overall data center capacity.<p><a href="https:&#x2F;&#x2F;x.com&#x2F;EugeneNg&#x2F;status&#x2F;2099315982959616369" rel="nofollow">https:&#x2F;&#x2F;x.com&#x2F;EugeneNg&#x2F;status&#x2F;2099315982959616369</a>
      • darig17 hours ago
        [dead]
    • KurSix11 hours ago
      China has had a very strong incentive to recreate lithography and fab equipment for years, but a lot of the specialized HBM packaging chain was probably much lower priority until relatively recently
    • chvid18 hours ago
      Huawei uses their own non-standard HBM called HiZQ probably not produced by CXMT.
    • vatsachak19 hours ago
      China should invest in an analog inference chip. It&#x27;s a hail mary but why not.
      • danielheath16 hours ago
        The two challenges there are firstly - that analog design has been a separate electrical engineering school for most of a century, so there are few who could design it - and secondly - that every single chip will have subtle variations in its computations, necessitating some sort of model finetuning per chip. Possibly the chip could be characterised at the factory, and ship with the characterisation data burned into a controller rom or something, but if that doesn’t pan out the whole thing is likely a non-starter.<p>If it could be made to work, you could run a fable-grade model in tens of watts.
        • simondotau13 hours ago
          Or we treat them like humans and make each one do job interviews to work out what tasks they’re suited for.
          • mindok12 hours ago
            This made me smile. A true test of human-like intelligence. Maybe they could take some entry level jobs (“I can make a nice Wordpress site”) and build a CV?
          • pixl971 hour ago
            It&#x27;s the hidden sociopath that&#x27;s able to mirror a functional entity that I&#x27;d be a bit worried about in these cases. Damn thing would make CEO quick.
        • itsthecourier13 hours ago
          awesome explanation. never thought about the complexity of doing analog and having precision issues because it&#x27;s not discrete anymore and repeatable
      • geysersam18 hours ago
        They can probably afford to do both.
      • sroussey18 hours ago
        That would be like skipping land line phones for mobile...
        • clipsy17 hours ago
          Developing countries have done exactly that in many cases.
          • mitxela7 hours ago
            And skipped DSL for fiber, and skipped credit cards for banking apps.
      • bobmcnamara13 hours ago
        Or on DRAM on die compute
      • bell-cot17 hours ago
        I&#x27;d assume that they have - but will keep mum &#x27;till they have a major breakthrough or large-scale operational deployment to announce.
    • FooBarWidget4 hours ago
      The &quot;abysmal&quot; yield using DUV is an overblown statement by western commentators who don&#x27;t look closer at the development. Certainly yield is worse than with EUV, but after multiple iterations of development, yield has become pretty good, within economically acceptable bounds, still making scaling possible. Volume is still ramping up. A lot of capacity will come online in 2027. The removal of western middlemen such as Mediatek actually improved the economics. Further yield improvements are still coming. Volume is already so high that domestically made phones and NPUs are making a real impact, yet are also selling like hot cakes.
      • HarHarVeryFunny3 hours ago
        It depends on what you are trying to make, but there are workarounds for most things, and it becomes more of a cost and volume issue than a showstopper.<p>If your yield is low and you still want the volume, then run more wafers, but then you need access to more DUV machines and more wafers. If your defect density is too high, then design smaller chips with a higher chance of avoiding defects, which is what Hauwei are doing - split processor into multiple chiplets (NVidia do this too, to raise yields and reduce cost).<p>HBM3E usually has 5-6000 vias, but doing this with multi-pattern DUV without defects is tough, so CXMT currently drop that to 3000, at a cost of some loss of thermal and voltage stability.<p>HBM:processor production ratios aren&#x27;t what Huawei would like, so they mitigate this at system level by putting an optical memory bus on the GPU chip and sharing memory across the system.<p>Not everything is a huge LLM - smaller models like recommendation systems don&#x27;t have so many parameters and need so much memory, so why waste HBM on them? ByteDance use their SeedChip acelerator for this, currently made by TSMC (using an older sanctions-approved 28mm process), which instead etches dense RRAM on die beside the processor.<p>There is also a time dynamic to this, with China still using pre-sanctions equipment and chips as their domestic alternatives ramp up to replace them. One interesting part of this is HBM ... HBM is very demanding to make, and everyone had problems with it, with initially only SK Hynix being successful. Samsung and Micron took a year or so to catch up, and during this time Samsung had made a ton of HBM that didn&#x27;t meet NVidia&#x27;s specifications, so ended up, pre-sanctions, selling it to China, where it has acted as a stockpile to carry them over as CXMT&#x27;s domestic capacity ramps up.<p>China seems to be doing fine. Of course they would like sanctions lifted, moreso for HBM than anything else, but all that sanctions have really achieved is accelerating their semiconductor independence. They are still building 1T+ SOTA models, standing up 100K clusters of domestic AI accelerators, etc, etc.
    • 651019 hours ago
      I read HBM yields are 25-30% (vs 80-90%) making them 3 to 5 times as expensive. They are 4-5 years behind, that probably means 1-2 in Chinese time.
  • fooker22 hours ago
    What&#x27;s the main blocker (other than the current inflated cost) for using HBM instead of DRAM as the primary memory for consumer electronics?
    • bob102922 hours ago
      It&#x27;s not that there&#x27;s a blocker. It&#x27;s that it takes roughly 3x the manufacturing capacity to produce an HBM package at the same storage capacity as DRAM. We are sacrificing total bytes for bandwidth.
      • rkagerer19 hours ago
        I don&#x27;t fully understand the source of the &quot;total bytes&quot; constraint, but a major factor may be because HBM4 &#x2F; HBM4E can only make use of the footprint directly above the processor&#x2F;logic die (or in direct vicinity of its interconnect), while traditional DRAM can be placed further away where there&#x27;s lots of real estate on the motherboard.<p>I gather a practical max ceiling today is a stack of 16 chips in height yielding 64GB?<p>These chips have a massive bus size of 2048 bits, instead of the 64 or 128 bits (dual channel) used by DDR5. That&#x27;s what gives them their order-of-magnitude bandwidth speedup. But even though they technically pack in more capacity per <i>square millimeter of motherboard</i>, I gather they take up more space than older technologies once you account for the vias and interconnects to route all those signals.
        • threecheese19 hours ago
          Thanks for that, just went down an interesting rabbit hole. Many of us were hoping this re-tooling would eventually trickle some fast RAM down to DRAM-exhausted PCs, but given it would require a rearchitecture of the motherboard it&#x27;s unlikely.
          • craigjb16 hours ago
            HBM4 has over 2048 signals to the processor’s PHY with tight signal integrity requirements that require the HBM stack to be &lt; 0.5 mm from the processor die. That’s why HBM integration is done with interposers (soldered on the package). So, it’d be the CPU package that integrates it. Motherboard is too far away.
            • Melatonic8 hours ago
              Kind of seems like we should be making chips with both. Big HBM stack on top as a sort of huge L5 cache like thing. And then a bunch of DRAM type sockets (like LPCAMM) around the exterior.
              • craigjb11 minutes ago
                For chips with integrated CPU+GPU+NPU, it could be worth it tech-wise. The GPU and NPU can eat HBM bandwidth. For general purpose CPU code, the HBM would likely not be worth it. It&#x27;s high bandwidth, but you trade latency, and general CPU code is branchy. Economics-wise, the HBM stacks alone will cost more than a consumer CPU (or APU).<p>[edit] The packaging needed to support HBM is also much more expensive too. If demand for current HBM applications tanks and the manufacturing lines need filled, then maybe. Currently, the price point would make it very very niche.
          • sroussey18 hours ago
            HBM also trades bandwidth for latency, and your regular computing is much more sensitive to latency than bandwidth.
      • tliltocatl21 hours ago
        How so? FEOL is pretty much the same, BEOL is almost the same save TSVs, the packaging tech is different and more advanced, but not exactly 1:1 comparable. Do TSVs really occupy 3x the area of DDR IO&#x27;s?
        • Const-me21 hours ago
          See remark on the slide 11: <a href="https:&#x2F;&#x2F;www.servethehome.com&#x2F;micron-evolving-memory-architectures-for-ai-at-hot-chips-2026&#x2F;" rel="nofollow">https:&#x2F;&#x2F;www.servethehome.com&#x2F;micron-evolving-memory-architec...</a> That presentation is by Micron.
          • Karliss21 hours ago
            It doesn&#x27;t really say that it needs to use 3x more area, but that 3x more gets consumed due to &quot;advanced packaging and manfuacturing complexity&quot;. Which doesn&#x27;t properly explain why it consumes 3x more and could simply mean they have a bad yield and 2&#x2F;3 produced is garbage.
            • crote7 hours ago
              You make HBM by stacking a whole bunch of dies on top of each other. The signals from the upper dies need to pass through vias in the lower dies to get out - taking up valuable die space in a way which simply isn&#x27;t needed with regular DDR. Similarly, HBM has a <i>far</i> wider bus, so each individual die has, say, 16 banks of depth 32, rather than 4 banks of depth 128. That&#x27;s more control area needed per byte of memory.<p>Those two combined already result in a huge reduction in bytes per mm2, so with the same wafer processing capacity you&#x27;re producing far less byte of memory. Add to that a complicated chain of HBM-specific packaging steps, and you&#x27;re now <i>also</i> losing a decent bunch of perfectly-fine dies because rather than putting it into DDR you tried making a HBM sandwich and screwed up.<p>Even if the memory cells are the same and have an absolutely identical yield, HBM will <i>always</i> end up having a significantly lower output. That&#x27;s just the cost of stacking, but some people are willing to pay the per-gigabyte price penalty in return for the higher bandwidth.
            • bob102920 hours ago
              &gt; 2&#x2F;3 produced is garbage.<p>This might not be far off the mark. You are irreversibly linking the fates of these devices after a certain stage of manufacturing. If something goes wrong at final packaging time, you lose all dies instead of one.
            • tliltocatl20 hours ago
              Yea, that&#x27;s the question. Yield situation can improve. Area overhead would not improve short of a completely new and incompatible tech.
            • imtringued6 hours ago
              Classic DRAM stacks up to four wafers on top of each other and then is packaged with BGAs. The manufacturer can check the DRAM chips independently.<p>Soldering the DRAM onto a PCB is such a reliable process that there is almost zero risk of defects and even if a defect occurs the damage is limited. If the DRAM is soldered onto a DIMM the risk of a defect on the non memory hardware is non-existent. If the DRAM is soldered straight onto an SBC or GPU, then the DRAM can be removed to save the precious SoC or GPU chips.<p>Meanwhile HBM is the ultimate nightmare scenario. You stack up to 16 DRAM wafers on top of each other. One defect and the whole stack is worthless and that was actually the easy part.<p>In stage two things get even worse. You now have your accelerator chip and you must place the HBM on that chip. E.g. Blackwell GB300 has eight HBM stacks and the accelerator chip has a bigger area than the HBM. You must get the packaging right eight times in a row or you have wasted not only the DRAM silicon, but also the accelerator silicon because HBM cannot be removed and defects are permanent.<p>The issue here isn&#x27;t just the yield of the HBM (which is obviously lower if you have taller stacks) but rather the yield of the combined HBM-based product, which is why doesn&#x27;t make sense to say it needs more area but it is completely correct to say that HBM leads to more silicon being consumed. Hence it doesn&#x27;t make sense to talk about yield of the HBM itself, because it is always part of an integrated product.
          • skavi21 hours ago
            direct link: <a href="https:&#x2F;&#x2F;www.servethehome.com&#x2F;micron-evolving-memory-architectures-for-ai-slide-11&#x2F;" rel="nofollow">https:&#x2F;&#x2F;www.servethehome.com&#x2F;micron-evolving-memory-architec...</a>
            • tliltocatl21 hours ago
              Yuck. Time to build an xSPI&#x2F;HyperRAM workstation (if only these had multi-bank chips).
              • buildbot20 hours ago
                Sadly the $ per byte of xSPI and HyperRAM quite high
                • tliltocatl20 hours ago
                  Yes, but enough to run a text editor, or even a mechanical CAD. Not an LLM, but that&#x27;s the point!
        • monster_truck21 hours ago
          3x is a reasonable figure. They are not literally that large, though.
    • phkahler21 hours ago
      &gt;&gt; What&#x27;s the main blocker (other than the current inflated cost) for using HBM instead of DRAM as the primary memory for consumer electronics?<p>HBM is meant to be integrated into the same package as the CPU, so no more DIMM sockets. It also has higher latency apparently.
      • MarleTangible6 hours ago
        One of the comments mentioned that they have a bus size of 2048, which may be why the latency is higher.<p>&gt; These chips have a massive bus size of 2048 bits, instead of the 64 or 128 bits (dual channel) used by DDR5.
    • HarHarVeryFunny20 hours ago
      Why would you want&#x2F;need to?<p>The advantage of HBM over regular non-stacked DRAM is memory bandwidth, which also requires a super-wide memory bus - 2048 bits wide for HBM4. Compare that to the 128 bit wide bus of a modern CPU.<p>So to take advantage of it on the desktop, or anywhere else, you need that 2048 bit wide bus, and a processor capable of consuming 2-3 TB of data per second!<p>These are not normal requirements, other than for a GPU.
      • Dylan1680715 hours ago
        &gt; Compare that to the 128 bit wide bus of a modern CPU.<p>Or 256-512 bits on medium to high end consumer CPUs if you&#x27;re apple.<p>At least DDR6 is probably widening things 50%.
      • bobmcnamara13 hours ago
        Intel and IBM have already done this almost this with their wide eDRAM caches.<p>The advantage in going wide is transferring cache lines rapidly, not the CPU bus interface.
      • fooker20 hours ago
        SIMD (and especially the modern matrix extensions) can use as much bandwidth you can throw at it.
        • b11220 hours ago
          Without further clarification, that statement seems impossible. &quot;As much&quot; being unbounded and all. You should expand what you mean.
          • fooker19 hours ago
            This will blow your mind, but it actually is pretty close to being unbounded. :)<p>Consider the &#x27;MMA N matrices&#x27; primitive modern CPUs are starting to support. For the current generation of CPUs, N is a constant like 16 or 32, but there&#x27;s nothing preventing it from being 1024 or larger if we have more memory bandwidth.<p>All this with a single instruction.
            • imtringued6 hours ago
              Nah his point is a bit simpler. Vector units can process a limited amount of data per time unit. Memory can load a limited amount of data per time unit.<p>If you have infinite memory bandwidth you just move the bottleneck back to compute so both have to grow simultaneously in lockstep.<p>What you should have said is that CPUs have so much compute headroom for matrix vector multiplication that simply adding more memory bandwidth would make them faster so every improvement in memory bandwidth is welcome.
              • fooker6 hours ago
                Agreed.<p>The &quot;move the bottleneck back to compute&quot; bit is changing rapidly though. The first time a major hardware company ships a PIM chip, you can push for a few orders of magnitude more data through without being compute bound.
            • articulatepang17 hours ago
              Surely something prevents it being 1 quadrillion bits per instruction? Since that’s well within “unbounded”.
              • pixl971 hour ago
                Most likely the chip running at the core temperature of the sun.<p>We&#x27;ll have to figure out how to read and right to the surface of a black hole to get speeds that high.
      • XorNot20 hours ago
        Right but if HBM memory is all that people want to produce, then building a CPU which can use it use it would be useful on it&#x27;s own merits.<p>But in reality we also already have unified memory architecture systems, integrated graphics etc.
        • p1esk20 hours ago
          People want to produce hbm because it’s more expensive and more profitable than regular memory.
          • XorNot19 hours ago
            There is a world where scale and experience means it&#x27;s about the same though, is the thing.<p>And memory is already expensive. It&#x27;s downright hard to even get it though - you frequently would prefer not what&#x27;s cheapest, but whatever is in largest scale production.
            • Dylan1680715 hours ago
              Scale and experience almost entirely share between HBM and normal memory. And they&#x27;re both in large-enough scale production to not have a big difference on availability; if you&#x27;re willing to pay HBM prices you should find even more sellers of DDR.<p>The only way I see HBM becoming competitive for consumer CPUs is if they solve the yield issues. Or if AI crashes so hard that people are putting those GPUs on fire sale and salvaging mass quantities of HBM off of them.
              • zeristor10 hours ago
                I was thinking this, but I doubt that the HBM is so easy to repurpose.<p>I’m assuming that it’s mounted in the same unit as the GPU, not in a handy-dandy socket.<p>It could be cracked open and extracted no doubt but if it’s glued in that’s going to be nigh on impossible to extract.<p>I had been pinning my hopes on HBM coming onto the second hand markets after there three years or so of use, perhaps I’m wrong.
                • pixl971 hour ago
                  It seems unlikely as the entire machines with HBM on them are apt to be sold whole on the market and snapped up quickly. With how much demand is in the current market machines 3 years old may not be sold if they can&#x27;t get newer faster machines fast enough.
    • chessgecko21 hours ago
      I think people might prefer the lower idle power consumption from lpddr over the better bandwidth in hbm in battery powered stuff. That said right now the price is definitely preventing us from finding out.
      • fooker21 hours ago
        Idle yes, but HBM energy consumption &#x2F; memory operations seems to be a bit better than DRAM.
        • vlovich12321 hours ago
          Only if you’re running at 100%. Consumers generally do not.
          • monster_truck21 hours ago
            That hasn&#x27;t been true since early HBM2 days, before the controllers standardized on power&#x2F;voltage management and did things like leave them in P0 to ship on time
            • vlovich12319 hours ago
              LPDDR&#x2F;DDR&#x2F;GDDR generally still win over HBM when there’s no data being transferred. HBM is primarily better in watts&#x2F;byte transferred. Consumer electronics spend most of their time idle.
              • Zagitta19 hours ago
                Racing to idle is a very common power optimization technique
                • vlovich12318 hours ago
                  Right, but HBM idle is significantly worse than LPDDR idle or even DDR idle for that matter. That matters a lot precisely because the device is idle most of the time. Your idle power draw dominates.
                  • monster_truck11 hours ago
                    Comparing HBM to LPDDR is pretty silly. That&#x27;s like a bath tub vs a urinal
                    • vlovich1232 hours ago
                      This is literally a thread where people are saying “but why laptops no HBM”. It’s a perfectly cromulent response that meets the question where it is. It might be mildly a more defensible question for normal plugged in desktops but that would require a massive ecosystem change since people who build those traditionally really like their DIMMs and less buying a CPU that has non-swappable RAM built in, not to mention there’s no sustainable market for the higher cost and low volume product.
                    • imtringued6 hours ago
                      If LPDDR6 PIM ever becomes a thing then the advantage of HBM will shrink. I&#x27;m not saying LPDDR6 PIM will make HBM obsolete in the datacenter or where it is currently shining, I&#x27;m saying that large volumes have their own charm and it is more likely for LPDDR6 PIM to be in your laptop or smartphone or SBC than HBM.<p>LPDDR6 PIM would primarily help the low end and mid range accelerator market. E.g. embedded models running on SBCs can be up to 1 GiB in size with acceptable performance, small models at 8 GiB become really easy to run at reasonable speeds on a smartphone and PIM enabled laptops or mini PCs make it possible to run 32 GiB models locally without compromise.<p>Of course this also assumes that the associated accelerators (NPUs) will catch up too, but the general point is that you won&#x27;t need a 5090 or a 4090 anymore.
    • torginus19 hours ago
      Mainly bus width. Afaik HBM is like 1024 bits vs DDRs 64 so you need lots of transfers in parallel to saturate the bus, and CPUs kinda want 64 bytes of data as that&#x27;s the size of a cache line ASAP. So you need a ton of in flight transfers which isn&#x27;t a thing CPUs provide, maybe multicore workloads.<p>Buy the way you win with CPUs is with latency, and not bandwidth, which is why Apple M series actually uses DDR with lower latency because of the stacking.
    • KurSix10 hours ago
      HBM isn&#x27;t dramatically better in every dimension. You get huge bandwidth and good energy efficiency per bit transferred, but not necessarily a meaningful latency improvement and capacity expansion becomes tied to the package
    • tjwebbnorfolk12 hours ago
      Even if the cost were the same of the RAM itself, you&#x27;d need much bigger and more expensive CPU to deal with it.<p>Running 17 chrome tabs doesn&#x27;t benefit at all from that HBM and all the additional hardware+software complexities that come with it. You want a specialized coprocessor to handle specialized workloads. The GPU exists separately from the CPU for a reason.
    • reliabilityguy21 hours ago
      HBM is a stack of DRAMs, so there is no “instead”.
      • buckle801721 hours ago
        The vias to enable stacking is a significant amount of the die area.
        • saltcured21 hours ago
          If they&#x27;re talking about production capacity, that is some product of die area and process steps, right? It doesn&#x27;t have to be 3x die area, just 3x lower factory throughput for the same number of functioning memory bits.
          • buckle801717 hours ago
            HBM is less dense at a water scale than DDR because of all the vias, but each memory but is as dense or denser.<p>You make HBM instead of DDR and the number of bits you&#x27;re making goes down.<p>It&#x27;s really that simple.
    • refulgentis22 hours ago
      In one sense, nothing, in another, everything. It is DRAM, but the bandwidth requirements mean it’s paired to a processor, i.e. no DIMMs. Not 100% sure but things like MacBooks and the Framework tower, where you have fixed RAM for the device lifetime, have ~0 tradeoff.
    • nutjob222 hours ago
      Nothing except CPU manufacturer choices. Mac laptops use it and they&#x27;re consumer products.<p>People will have to get used to buying a fixed amount of RAM with their CPU but thats unlikely to be a problem.
      • fooker21 hours ago
        Apple&#x27;s &quot;unified memory&quot; marketing is so strong that even tech literate people seem to have this misconception!<p>They have managed to pull this sort of thing off many many times. <a href="https:&#x2F;&#x2F;en.wikipedia.org&#x2F;wiki&#x2F;Reality_distortion_field" rel="nofollow">https:&#x2F;&#x2F;en.wikipedia.org&#x2F;wiki&#x2F;Reality_distortion_field</a>
        • Rohansi21 hours ago
          Yup, the only reason Macs have higher memory bandwidth is because they use more memory channels, which gives them a wider bus. Both Intel and AMD only allow more than dual channel memory on server class processors these days.
          • kstrauser21 hours ago
            The “Apple only does X better because they do Y” thing has been a meme for ages. I remember dismissals like “PowerPC is only faster at math because it has more integer units” or something along those lines, and thinking, uh, isn’t that a good thing?
            • pdpi20 hours ago
              There&#x27;s a qualitative difference between &quot;they&#x27;re doing a different thing&quot; and &quot;they&#x27;re doing the same thing, tuned differently&quot;. GP is saying that this is a case of &quot;they just tuned it differently&quot;.<p>This distinction doesn&#x27;t change what the performance numbers look like today, but it does inform what changes would be necessary for those numbers to look different tomorrow. E.g. Apple Silicon isn&#x27;t fundamentally orders of magnitude more efficient than x86, they just used smaller features. Newer Intel and AMD chips made on equivalent processes _also_ get similar efficiency gains.
              • sroussey18 hours ago
                There are AMD and Intel devices on similar process (not talking about A20Pro or M6 which are set to ship later this week), and they do not get the same gains.<p>And honestly, they have historically had different markets.<p>When the design is for only one customer, you don&#x27;t need to generalize things, and those things you generalize to give different customers different options has costs.<p>AMD will soon be a larger customer for TSMC than Apple (NVIDIA is already there) so Apple&#x27;s pre-booking new processes is likely to be gone in the near future.
            • Dylan168072 hours ago
              The point is that &quot;unified architecture&quot; is a buzzword unrelated to what actually makes things fast.
              • Rohansi49 minutes ago
                It does make <i>some</i> things faster because you can share memory between CPU, GPU, etc. without copying. But not everything.
            • pixl9721 hours ago
              Depends on the expense trade off.<p>If I get 10% more performance for 50% more cost it really depends on one&#x27;s needs, for example.
            • JohnBooty17 hours ago
              &quot;You&#x27;re only better than me at sports because you practice more and try harder!&quot;
            • Rohansi21 hours ago
              I am just saying it&#x27;s not magic and x86 is capable of doing the same. Quad channel memory used to be more common in consumer hardware but now it looks like you don&#x27;t even have the option anymore for desktops. AMD&#x27;s Strix Halo was the first sign to reversing that (it has quad channel memory!) and hopefully we see more of that in the future.
              • colejohnson6620 hours ago
                AMD does market segmentation and limits consumer chips to dual-channel. You need to cough up the dough and get Threadripper for quad-channel. Or even more for Threadripper Pro to get octa-channel.
                • Rohansi19 hours ago
                  As I mentioned above AMD&#x27;s Strix Halo has quad channel memory and is consumer level. But yes, other than that everything is segmented away.
                • torginus19 hours ago
                  Afaik steam deck is quad channel, despite it being a pretty low end chip (with a decent GPU though)
                  • Rohansi17 hours ago
                    Kind of but not really. DDR5 splits your typical 64-bit channel into two 32-bit subchannels meaning the bus width is not increased. These subchannels are not always advertised because it&#x27;s a just a feature of DDR5. Actually adding more channels increases bus width, which is what meaningfully improves memory bandwidth.
          • PunchyHamster18 hours ago
            Threadripper have that extra bandwidth and M5 still is faster
            • Dylan1680715 hours ago
              Can you link a specific benchmark?<p>Keep in mind that non-Pro threadripper is still only 256 bits wide and Pro is 512. And the memory is 30% slower than with an M5. So an M5 Ultra has 3x the memory bandwidth of the best threadripper.
              • Melatonic8 hours ago
                Don&#x27;t some AMD full server CPUs have 12 memory channels ?
                • Dylan168072 hours ago
                  Yes, some do. So similar bandwidth to a Max but with lots of slower cores. Half as much as an Ultra.
          • throwaway8582519 hours ago
            Except strix halo.
          • Kon5ole21 hours ago
            The memory bandwidth is a small thing compared to the massive win you get by not having to move data between two memory pools at all.
            • Rohansi21 hours ago
              Depends on your workload. And AMD has supported unified memory long before Apple Silicon existed anyway.
            • fooker20 hours ago
              Often you don&#x27;t move memory around as a programmer, but that&#x27;s exactly what happens in the background.<p>It&#x27;s the address space that&#x27;s unified, not always the physical hardware.<p>The data movement (when needed) is handled transparently in the background by page faults and other tricks.
              • Rohansi14 hours ago
                Moving memory around is a bottleneck. Non-unified memory usually means copying over the PCIe bus which is way slower than RAM (and way way slower than VRAM). Actually unified memory means you don&#x27;t need to copy anything at all though which is the absolute best case for performance.
        • Kon5ole21 hours ago
          Having unified memory is a real advantage though, it&#x27;s not a reality distortion.
          • jorvi18 hours ago
            It isn&#x27;t really, as long as you don&#x27;t care about power consumption, physical constraints and money. Basically desktops &lt;2025 (and hopefully &gt;2027).<p>DDR is optimized for latency and stability at the cost of bandwidth whilst GDDR is optimized for bandwidth at the cost of latency and stability. GDDR is pushed so hard these days that a small percentage of errors is expected and corrected because this is still faster than running it slower but more accurate.<p>GDDR7 often has 10-20x the total bandwidth but 3x the latency of DDR5. Graphical workloads want as much bandwidth as possible but care relatively little for latency. Conversely, applications love low latency but don&#x27;t really see any performance benefit from higher bandwidth.<p>So basicallyt you have workloads that are diametrically opposed and running unified memory forces you to compromise.
            • Melatonic8 hours ago
              Is GDDR7 that&#x27;s ECC then just the super well binned stuff (plus the ECC parts added) ? I&#x27;ve been wondering why we don&#x27;t see more cards using it as they perform pretty damn well. Take the Nvidia 6000 Pro Blackwell for example. The compute is insanely fast assuming you can fit what you need in 96GB of ECC GDDR7
          • davrosthedalek21 hours ago
            The price is that you essentially glue CPU and GPU together, which limits total compute, from a size and thermal perspective.<p>This is really not a limit because of unified memory -- in principle, PCIe GPUs could read&#x2F;write main memory without the CPU. But it&#x27;s a limit for &#x2F;fast&#x2F; unified memory, because fast means close.<p>So unified memory is great as long as the integrated GPU is strong enough. Then it has two advantages: a) probably faster transfer CPU&lt;-&gt;GPU (but that&#x27;s an implementation choice for the non-unified case b) If you either need a lot of memory for the CPU or the GPU, but not for both at the same time, you pay for memory only once.
          • fooker20 hours ago
            It is a real advantage.<p>The reality distortion is that people seem to believe it&#x27;s HBM, or somehow it gives you extraordinary amounts of vram. Neither are really true.
          • nvme0n1p120 hours ago
            Agreed. That&#x27;s why it&#x27;s a good thing all computers made in the past 15 years have unified memory, not just macs.<p><a href="https:&#x2F;&#x2F;x.com&#x2F;Lina_Hoshino&#x2F;status&#x2F;1820947147312820497" rel="nofollow">https:&#x2F;&#x2F;x.com&#x2F;Lina_Hoshino&#x2F;status&#x2F;1820947147312820497</a>
            • Kon5ole18 hours ago
              In theory perhaps but the benefit is not as relevant with a weak iGPU. In practice all PC&#x27;s with performance ambitions had a dGPU until Strix Halo and Panther Lake.
            • stymaar20 hours ago
              Is that Marcan&#x27;s vtuber persona (Asahi Lina) that changed name since Marcan doesn&#x27;t work on Asahi Linux anymore?
        • nutjob21 hour ago
          No I just failed to check before I posted, I thought it was HBM. I don&#x27;t have any interest in Apple hardware so wasn&#x27;t properly informed, and have been duly crucified.
      • KeplerBoy21 hours ago
        MacBooks use regular soldered lpddr5(x) RAM. Same RAM as every other laptop manufacturer, they just use more lanes to achieve a higher bandwidth.
        • bunderbunder21 hours ago
          Perhaps more noteworthy for general home and business computing, doesn’t it also allow for lower latency?
          • Rohansi21 hours ago
            More channels or soldered memory? Channels are basically RAID 0 so it depends what you&#x27;re measuring. Soldering memory down was the only way to use LPDDR5X so if you wanted the best memory you had to solder it down. LPCAMM2 exists though so newer devices can use that instead of soldering them down, but not all devices would be able to fit the required LPCAMM2 slots.
            • throwaway8582519 hours ago
              SOCAMM2 allows for removable ram in nearly 0 added space.
              • Rohansi14 hours ago
                That&#x27;s tiny! But it still depends a lot on form factor. LPDDR5X is used in phones and making memory removable would have its compromises. You may even have compromises in laptops. Look at how tiny a MacBook Air&#x27;s mainboard is and you&#x27;ll see the RAM modules on the same package as the SoC. SOCAMM2 is too large for that but a variant with only two modules could possibly work.
                • Melatonic8 hours ago
                  I bet it could fit. It just do a combo of soldered ram and non soldered like many laptops used to.
      • addaon22 hours ago
        &gt; Mac laptops use it and they&#x27;re consumer products<p>No, Mac laptops use LPDDR, currently LPDDR5X.
      • nomorewords22 hours ago
        The normal non-tech-savvy person already does this. They simply don&#x27;t know that their ram is upgradeable or something else breaks first, before having to touch ram.
        • dylan60420 hours ago
          what is this too much RAM thing you mention? I thought the only valid RAM situation you could find yourself is not enough RAM. Too much? That&#x27;s just fantasy
      • chessgecko22 hours ago
        pretty sure its lpddr not hbm.
  • nullbio9 hours ago
    Will consumers actually get to purchase it though? Or are they just doubling it for frontier AI labs?
    • Grimburger8 hours ago
      What does it matter?<p>All supply is supply, despite what certain folks seem to believe.
      • the_other8 hours ago
        It matters so that average people have a chance to own their means of production, and that the current hardware-grab by entrenched capital doesn&#x27;t get to charge such high rent.
        • chapz8 hours ago
          Frontier AI labs are customers just like the average Joe down the street. The only difference is who can bid more. Average Joe probably doesn&#x27;t have millions or wont buy millions of chips.<p>There are no privileged consumers that have &quot;dedicated&quot; amount of production, no matter the demand.
          • franga20008 hours ago
            What point are you making by saying &quot;they&#x27;re just like the average Joe&quot; followed immediately by saying they have so much more money that they can outbid the average Joe at any price point.<p>If someone can dump 100x the money to get something I want, they&#x27;re not &quot;just like me&quot;.
            • sysguest8 hours ago
              yeah this<p>if I&#x27;m ordering 1 RAM and demanding some special customization, even mom-and-pop stores will just ignore me<p>but if I&#x27;m buying 100000000000 RAM, things change
        • mitxela7 hours ago
          You won&#x27;t ever own it as long as it comes from a company the size of Samsung.
          • MatekCopatek7 hours ago
            I mean... You can buy a stick of RAM with Samsung chips on it and use it in your own machine to run a local model vs. paying for an Anthropic subscription. I assume that&#x27;s what they&#x27;re referring to, not owning Samsung stock.
        • lstodd7 hours ago
          &quot;average people&quot;&#x27;s means of production is their brains. which they have all the chances to own but often reject them.
        • lazide7 hours ago
          Your average consumer is not even a citizen of the country where the company that makes this is based, or where the factories that produce it are.<p>You can own stock, but other than that?
    • user439284 hours ago
      I have read an interview with Acer&#x27;s CEO claiming that the shortage is already over.<p>He alleged memory maker&#x27;s public announcements of the shortage lasting throughout the end of 2027 are attempts to covertly signal the others on pricing strategy.<p>With additional production capacity ramping up this year and next, this seems somewhat plausible.<p>If manufacturer&#x27;s inventories are filling since they advanced orders when prices were rising, there could be an oversupply of memory when they don&#x27;t place new orders at the higher current price.<p>This could quickly result in a price war while capacity expansion and yield improvements create irreversible oversupply.
    • ilogik8 hours ago
      As far as I know, the only place HBM is used in consumer products is GPUs?
      • intothemild7 hours ago
        HBM hasn&#x27;t been used in a consumer GPU in some years. I think the RX 7900XT<p>It&#x27;s all GDDR6X or 7X
  • gs1722 hours ago
    A shame that this should if anything, lead to consumer DRAM prices getting even worse.
    • nicoburns22 hours ago
      Why would more RAM supply lead to higher consumer prices?
      • devy22 hours ago
        HBM4 and HBM4E DRAM are NOT the DDR4&#x2F;5 that consumer markets need. Capacity allocation is leaning more to data center grade HBMs so less to produce dedicated DDR4&#x2F;5 DRAMs. Supply demand will further drive up the consumer DRAM price! Note, the article mentions NO of new fabs is being constructed (all semiconductor manufacturers know that constructing more fabs means the boom&#x2F;burst cycle will eventually kill them, so no one create more fabs) Perhaps the federal government need to step in here - the market doesn&#x27;t fit the issue.
        • tipsytoad21 hours ago
          Plus it takes 3x the wafer capacity for hbm than the same byte capacity in dram, so we’ll likely see the consumer market be decimated here
          • chr15m17 hours ago
            The consumer market will not be decimated. Memory of all kinds will get very cheap quite soon because of supply and demand and substitution.
            • T-A9 hours ago
              <a href="https:&#x2F;&#x2F;www.tweaktown.com&#x2F;news&#x2F;112105&#x2F;amd-warns-ddr5-prices-wont-recover-until-2028-as-ai-demand-continues-pulling-supply-away-from-consumers&#x2F;index.html" rel="nofollow">https:&#x2F;&#x2F;www.tweaktown.com&#x2F;news&#x2F;112105&#x2F;amd-warns-ddr5-prices-...</a>
        • bnjms20 hours ago
          HBM can be used for GPU though. I wish it was more common.
        • seanmcdirmid10 hours ago
          Aren’t new DRAM fabs being built in China?
        • sgt21 hours ago
          Might lower the prices of Mac Minis, Studios etc though
          • airspresso21 hours ago
            No, those use LPDDR5(x), not HBM.
        • chr15m17 hours ago
          No, completely wrong. Supply and demand will definitely fix this.<p>More HBM4 will drive the price of it down. That will make it less profitable to produce. That will mean firms shift some production back to more profitable &quot;consumer&quot; RAM. Increased supply will lead to lower prices.<p>A shortage is always followed by a glut.
          • Dylan1680715 hours ago
            What <i>specifically</i> did that say that you think is wrong?<p>Supply is currently moving the wrong way. Supply and demand for consumer DRAM will get even worse, even if it will &quot;definitely fix this&quot; in the long run.<p>And the fact that production <i>is</i> moving the wrong way suggests that in the short term supply and demand is actually doing more harm than good. HBM is so profitable it&#x27;s radioactive.<p>I don&#x27;t know where you got the idea in another comment that all memory will get &quot;very cheap quite soon&quot;. <i>They&#x27;re not increasing total memory production.</i> (much&#x2F;yet)
            • Melatonic8 hours ago
              China is about to massively compete. And doesn&#x27;t Micron and someone else have new fabs in NY coming online ? I can&#x27;t imagine they&#x27;re going to exclusively make HBM.<p>Ram always goes through these boom and bust cycles. They&#x27;re going to increase supply but not so much that they&#x27;re eventually left with huge overproduction as usual.<p>Of course if Google researchers come up with more ram reduction tricks for newer AI architectures that could make a huge difference as well
              • Dylan168072 hours ago
                I don&#x27;t expect China to be a particularly large competitor in the next year or two affected by the decision in the article, but they&#x27;ll help a lot in the long run.<p>Micron&#x27;s NY fab had a press release two months ago that they started pouring the first concrete, so that&#x27;s far from coming online.<p>If AI demand sticks around and buildouts stay cautious it could take much much longer than usual to build up enough capacity to push prices back down below $3&#x2F;GB.<p>More efficient AI memory use risks even more demand for memory because you&#x27;re now able to get even more performance per dollar; the benefit of extra memory has to be sufficiently far into diminishing returns for demand to drop overall.
            • chr15m14 hours ago
              &gt; What specifically did that say that you think is wrong?<p>You said:<p>&gt; the market doesn&#x27;t fit the issue.<p>I presume you meant &quot;fix&quot; not &quot;fit&quot;. You are wrong that the market will not fix the issue. There are literally millions of observed cases across many markets of increased demand (or decreased supply) leading to higher prices leading to increased supply leading to lower prices, which is why it&#x27;s a fundamental law of economics.<p>The price of RAM will fall.
              • eek212112 hours ago
                I feel like everyone here forgets that the DRAM industry is operated by one big cartel. The rules of supply and demand don&#x27;t apply here. Now that prices are up, they will keep it up, cutting supply as needed in order to do so. Every big DRAM manufacturer in the industry has been charged with price fixing, among other things: <a href="https:&#x2F;&#x2F;en.wikipedia.org&#x2F;wiki&#x2F;DRAM_industry_price_fixing" rel="nofollow">https:&#x2F;&#x2F;en.wikipedia.org&#x2F;wiki&#x2F;DRAM_industry_price_fixing</a><p>There are so few players in this space they can coordinate and collude pretty easily.
                • Melatonic8 hours ago
                  DRAM isn&#x27;t nearly as hard to get into as compute. The price fixing stuff is going to bite them in the ass someday soon by massively increasing the incentives for new players to enter the market and compete.
              • Dylan1680714 hours ago
                &gt; You said:<p>That wasn&#x27;t me.<p>Also they&#x27;re talking about the short to medium term. On that time scale they&#x27;re not wrong. The market is so far from fixing it right now, it&#x27;s going in the opposite direction. The market is chasing HBM money which is outbidding everything else.<p>&gt; The price of RAM will fall.<p>And if that might take 5+ years, that&#x27;s not much comfort.<p>&gt; higher prices leading to increased supply leading to lower prices<p>So far higher prices are causing barely any supply increases. (Note that the article we&#x27;re commenting on is about a supply decrease.) Companies are barely trying to expand, and even if they do start trying hard it&#x27;ll take a very long time.
          • no-name-here15 hours ago
            &gt; <i>That will mean firms shift some production back to more profitable &quot;consumer&quot; RAM. Increased supply will lead to lower prices.</i><p>Long-term, sure. But the article we are discussing is about next year, non-HBM supply going from 60% of their output to 20% (HBM 40%→80%).
            • chr15m14 hours ago
              &gt; Long-term, sure.<p>Yes. Non-HBM Samsung supply going from 60% to 20% will cause prices to increase, which will cause increased production of non-HBM, and then prices will fall (below where they currently are).
      • matja22 hours ago
        Depends on what proportion of Samsung&#x27;s output is current HBM4 and HBM4E DRAM.<p>Everything in their statement can be true and it be a bad thing for consumers of non-HBM RAM.<p>&quot;HBM capacity to expand to 250,000 wafers a month&quot;<p>So let&#x27;s say their current HBM capacity is 100k wafers&#x2F;month (pure speculation&#x2F;random number for illustration), and their total RAM capacity (including HBM( is 300k wafers&#x2F;month, then non-HBM capacity reduces from 200k to 50k.
        • gs1722 hours ago
          And it&#x27;s worse than that, Micron said that converting capacity to HBM was at a 3:1 ratio, so that 250k HBM could be up to 750k in non-HBM. Fortunately, some of the increase is due to improved processes.
        • EgregiousCube22 hours ago
          Keep in mind that total manufacturing capacity is increasing as well. Perhaps not enough to fully offset, but it&#x27;s incorrect to assume that supply capacity is flat.
          • cogman1019 hours ago
            It&#x27;ll take a lot of time.<p>Micron has been working on expanding in Boise since around 2023. They are predicting that the first new chips will start rolling out around 2027.<p>I don&#x27;t think we&#x27;ll see any relief until at least 2028 or even later. And even then, probably not without antitrust laws being enforced. Because you know memory manufacturers won&#x27;t just simply drop prices, they have a long history of price fixing.
          • pixl9721 hours ago
            Question is what does the demand curve for regular ram look like also? If regular production is on a 10% increase curve but a 40% demand curve than we&#x27;ve been at the samw place.
        • no-name-here15 hours ago
          OP article estimates HBM is going from 40% to 80% of their output.
      • mattstir21 hours ago
        Part of the implication is that factories that could be producing consumer-facing DRAM like DDR5 would be retooled to produce HBM instead, leading to even less total consumer RAM production.
        • crote8 hours ago
          This is already happening, by the way.<p>SK Hynix&#x27;s new Cheongju M15x fab was originally intended to be for DDR, but mid-construction they shifted it to HBM. Similarly, the Icheon M10F packaging line has been retooled to HBM, allowing it to assist in processing the output of other fabs.<p>And this was news early last year, so who knows how far it has progressed since then.
        • m4rtink17 hours ago
          The crash won&#x27;t be kind to them &amp; they deserve everything they will got for their past behavior.
      • zargon22 hours ago
        Memory allocated for HBM is memory taken away from DDR production.
      • kenny1122 hours ago
        If Samsung is limited in the number of wafers they can process per month and they use more of those to produce HBM they necessarily have less of them left to make other products, like consumer DRAM.
      • stevefan199922 hours ago
        Gamma squeezed by the AI business, that&#x27;s why. The more you add the fuel, the further it will pump up.
      • deagle5022 hours ago
        Presumably DRAM production capacity gets reallocated to HBM, not sure.
      • shevy-java22 hours ago
        Depends on how quickly they are sold.
    • robotnikman19 hours ago
      Maybe this leads to device manufacturers using HBM memory instead.
  • redsymbol5 hours ago
    Right now, every company that can is ramping up their RAM production as much as humanly possible. There is such outrageous <i>and growing</i> demand, that it&#x27;s like printing money if you can manufacture more RAM chips. And that massive financial incentive is driving huge investment in manufacturing capacity.<p>But someday, that demand will stop rising.<p>At some point, most people will have enough RAM. And they just won&#x27;t need to buy more.<p>The inventories of RAM will become overfull, in warehouses and stores. And suddenly the manufacturers, middlemen, and even retailers will all find the price they can sell at keeps dropping.<p>But you can&#x27;t shut down a huge manufacturing facility overnight. You can&#x27;t shut down a years-in-the-making retail distribution system overnight, either. You can&#x27;t even shut down a facility that is halfway finished with construction. They&#x27;ll keep producing and selling RAM at lower and lower profit margins, until the day comes they finally accept they&#x27;re losing money with each sale.<p>And when that happens, we&#x27;ll all have more RAM than we know what to do with. It&#x27;ll be cheaper than we have all ever experienced in our lifetimes.<p>I don&#x27;t expect this to happen soon. But demand cannot keep exponentially rising <i>forever</i>. There will come a point, probably in the next 5 years, when demand at least levels off, even as production capacity is still rising.<p>And it will be interesting to see what&#x27;s possible at that point. We&#x27;ll be able to create software that is not doable right now, because it will no longer be constrained by memory. And even cheap laptops will massive amounts.
    • burnerRhodov34 minutes ago
      When you commoditize intelligence, you dramatically lower the cost to create. The amount of electronics being created is... incredible. We&#x27;ve really seem to have hit the asymtote part of the growth and i don&#x27;t really see it coming back.<p>Everything is moving to edge compute. From smart rings and watches, to robotics and drones, iphones and humanoids, and no end in sight to the scaling improvements in AI... It does seem that you can continue to get more emergent behavior the more compute you expend in pre&#x2F; post training. There&#x27;s no reason we can&#x27;t scale up to a million data centers in space, as long as you get dramatically improved intellegence.
    • ricardobayes5 hours ago
      Yeah, that&#x27;s kind of the beauty of capitalism. It can lead to short-term shortages but never structural ones. Any demand gets immediately fulfilled.<p>I grew up in a country where there was a structural shortage of say, shoes and cars. You didn&#x27;t really even get to choose which shoe you wanted to get, when there was a new shipment, folks went to the &quot;department store&quot; and got whatever size they could get their hands on and later trading it for the right one. To get any car, you would submit an application and 5-8 years down the line you would get a letter that your car would now be ready to be picked up, in a random color.
      • redsymbol3 hours ago
        Thanks for sharing that. I am grateful every day that I was born in such an abundant environment. I have a few friends who did not, like you, and it always reminds me of how lucky I have been.
  • GreenLightGo38 minutes ago
    The AI boom is turning into a massive race for HBM capacity.
  • glub19 hours ago
    tl;dr; This is existing production being redirected to HBM. Which means consumer market is going to get a lot worse.<p>And the worst thing is that is the best possible strategy for them. It&#x27;s essentially win-win for everyone but consumers.<p>- Buyer (AI) has crazy money, so will pay whatever<p>- Seller doesn&#x27;t have to build anything new, as the buyer is willing to pay whatever<p>- Generate ridiculous profits from crazy money<p>- No oversupply risk in case of reversal<p>- Return from producing HBM to DDR5 in a single quarter if reversal does happen.<p>Hold on to your existing hardware, people, and be on lookout for your local deals.
    • radiator18 hours ago
      The solution would seem to be preventing that &quot;Buyer (AI) has crazy money&quot;. They already have borrowed billions and are only generating losses so the money should run out. It just needs to be made sure that nobody (taxpayers) will rescue them.
      • yonaguska17 hours ago
        &gt; just needs to be made sure that nobody (taxpayers) will rescue them.<p>oof.
    • intothemild7 hours ago
      This should be the top comment.<p>It absolutely is not a great situation at all. This won&#x27;t make things better.
  • amelius22 hours ago
    Will that be enough for AI&#x27;s hunger?
    • mixedbit21 hours ago
      During the Covid global chip shortage Intel announced new factories to address the production bottleneck, before the factories even started to be constructed, the shortage was long over and the projects were eventually canceled.
    • GoToRO22 hours ago
      It will be just in time for when AI will run very well on consumer hardware and the need for data centers will collapse.
      • konaraddi21 hours ago
        There will be demand for both
        • consp9 hours ago
          Some demand. The question is if the capacity will meet it, be under or be over.
      • amelius20 hours ago
        What if the AI companies start selling the hardware with their models?<p>I.e., you get a locked down device with access to their AI and maybe a way to run third party apps. Of course the developers of those apps will have to pay a percentage of their revenue.<p>Sound familiar?
        • throwaway8582519 hours ago
          More likely they would sell a chip with the model etched in silicon for the dramatically higher token&#x2F;s.
          • amelius18 hours ago
            A chip that can only be accessed through their software. With a monthly subscription fee, and other enshittification surprises.
        • szatkus18 hours ago
          Most likely those models would be extracted from the hardware in no time.
          • amelius18 hours ago
            I suspect no, because most of the tokens would be used for its internal reasoning loop, and a cheaper model could be used to obfuscate the output.
      • IshKebab21 hours ago
        That&#x27;s never going to happen. By the time you can run current frontier models on your $10k desktop the frontier will have massively advanced and people will want those models instead.
        • bunderbunder21 hours ago
          I’m not so sure about that. Already AI vendors are back to cutting prices to try and keep customers from cutting back on their usage. My own employer is working hard at pivoting to much smaller fine-tuned models for established use cases, and seeing model performance improvement in addition to large inference cost reductions. Being able to run them locally hasn’t exactly been a disaster for devex, either.<p>It may turn out that demand for SOTA frontier models isn’t so limitless after all.
          • pixl9721 hours ago
            Every product follows demand curves. At a price of 0 you could find infinite usage. This has nearly zero relation to how much it costs to provide the product.
            • trombuance17 hours ago
              &gt;At a price of 0 you could find infinite usage<p>Infinite demand isn&#x27;t a thing. Even if they&#x27;d offer free compute forever (not likely possible) I and many others would still use local models that we have full control over and that do not harvest our personal data.
            • bunderbunder20 hours ago
              Except of course it relates. All else being equal, we will prefer $X COGS over $2X COGS because that helps us with both profit margins and price competition.
              • pixl9720 hours ago
                It relates in the sense there&#x27;s a minimum cost of production without losses, not the actual price people are willing to pay.
                • bunderbunder19 hours ago
                  Framing it in terms of the price people might be willing to pay for a single product in isolation frames the point I was making, which was about price competition, right out of the picture.<p>Maybe I&#x27;d be willing to pay $10 for product A if I had other options. But if there&#x27;s a product B for $3 that&#x27;s not quite as nice but still ticks all my boxes, then product instantly becomes a lot less attractive.
        • kingleopold21 hours ago
          This + your ROI in $10k device will be always lower than busy datacenter, its literally math. They sell free compute to others when you dont use it, you will never sell at that level or even you magically sell home compute, you will not compete at price
          • leoc17 hours ago
            That greater efficiency only benefits the LLM SaaS providers as long as the hardware manufacturers, probably especially the VRAM manufacturers, remain supply constrained, since the high-efficiency users are the ones who can pay top dollar. But the hardware guys&#x27; dream is presumably to get parts into millions of laptops which remain on standby for 19 hours a day, not to bargain with SaaS providers who obsessively optimise their memory consumption.
        • hypfer21 hours ago
          I&#x27;m not sure if this prediction will hold true.<p>We&#x27;re not seeing the progress in those &quot;frontier models&quot; that we have previously seen. There&#x27;s certainly still gas left in tank tank, but we&#x27;re way into the diminishing returns by now.<p>Cloud inference still beats hardware investments by orders of magnitude of course, but that&#x27;s only if your data doesn&#x27;t really matter to you.
          • airspresso21 hours ago
            We are certainly not in the diminishing returns phase for LLM progress. No sign of that yet.
            • bunderbunder21 hours ago
              I’ll grant that for specialized applications like coding agents and mathematics, but even there I suspect that most the real gains are actually taking place in the harness.<p>But I suspect returns may have already diminished into negative territory for at least some other use cases. One of my least favorite job responsibilities in this brave new era is figuring out how to avoid performance and behavior regressions when an older model were using for some application reaches end of life. It’s getting uncommon for me to look at our benchmark results and say, “Oh, good, it does better on one of the newer models!”
              • pixl9721 hours ago
                &gt;suspect that most the real gains are actually taking place in the harness.<p>Part of the reason harnesses work well is you can run a lot of agents in parallel. That doesn&#x27;t slow down demand.
                • hypfer21 hours ago
                  That is true, but the eventual realization that more machines doing more coin flips in parallel does not mean &quot;more work gets done&quot; might.<p>LLMs are amazing tech, but they&#x27;re terrible without oversight. More agents faster just makes reality collapse on them quicker.<p>But yeah, you&#x27;re right, temporarily, this will still push demand. But the topic was about &quot;diminishing returns&quot; as in &quot;tech getting better&quot;. Not as in &quot;customer spending&quot;.
                  • pixl9720 hours ago
                    It&#x27;s kind of weird because more machines working together does mean more work gets done. Coin flips and weighted coin flips are totally different things. Any biases weights towards reality push you closer to reality when you use them.<p>New models keep being able to use more and more agents on longer time frames. Your hypothesis doesn&#x27;t look like what we&#x27;re measuring.
                    • hypfer20 hours ago
                      Who is we?
                      • pixl9720 hours ago
                        The people mapping AI capabilities.
                        • hypfer20 hours ago
                          Oh cool, so that we includes me! :)
                          • pixl9719 hours ago
                            Maybe turn on your light when you use a ruler? Not sure what else to say.
                • leoc16 hours ago
                  But high demand for LLM time isn&#x27;t sufficient to keep customers at the frontier LLM SaaS providers. That demand can be satisfied locally or at non-frontier outlets, absent hardware shortages at least. The Tier 1 providers (and the would-be Tier 1s) presumably need to open up a much bigger lead in model quality, one that doesn&#x27;t simply get distilled away this time, and&#x2F;or continue to be protected by ongoing (or worsening!) hardware shortages. (And that&#x27;s overlooking the revenue shortfalls which OpenAI and Anthropic seem to be facing already.)
                • bunderbunder20 hours ago
                  I had actually been thinking more about all the non-LLM functionality that go into the harnesses. I&#x27;m not going to name names and I haven&#x27;t done any rigorous testing, but my general impression is that choice of harness matters more than choice of model. In terms of basic task completion success specifically, not code aesthetics.
                  • pixl9720 hours ago
                    A perfect harness will not extract gold from a dumb model. It&#x27;s a system that builds on each other, though we&#x27;ve not probed that frontier much to have a good intuition on what effects what.
              • 3eb7988a166320 hours ago
                One thing that I really want to know - the better models from today vs a year ago - what has changed. They have already pre-trained on all available public data. Scooping up the last percentage of archaic texts which were never digitized is not going to move the needle.<p>Is it just that the providers are generating tons of synthetic datasets on coding tasks so that the models get more exposure to the right thing to do? Every time someone points out an LLM stupidity they add some training data to patch over the weakness (trivial to generate &quot;there are two &#x27;l&#x27;s in llama&quot;)?
            • SideQuark13 hours ago
              Google scholar has a flood of papers showing LLM diminishing returns on pretty much every facet<p><a href="https:&#x2F;&#x2F;scholar.google.com&#x2F;scholar?hl=en&amp;as_sdt=0%2C23&amp;q=llm+diminishing+returns+&amp;btnG=" rel="nofollow">https:&#x2F;&#x2F;scholar.google.com&#x2F;scholar?hl=en&amp;as_sdt=0%2C23&amp;q=llm...</a>
              • T-A9 hours ago
                The first title I see there is &quot;The Illusion of Diminishing Returns: Measuring Long Horizon Execution in LLMs&quot;:<p><a href="https:&#x2F;&#x2F;proceedings.iclr.cc&#x2F;paper_files&#x2F;paper&#x2F;2026&#x2F;hash&#x2F;3b4e1336f775c3dba16ebbb8d2afd258-Abstract-Conference.html" rel="nofollow">https:&#x2F;&#x2F;proceedings.iclr.cc&#x2F;paper_files&#x2F;paper&#x2F;2026&#x2F;hash&#x2F;3b4e...</a>
            • hypfer21 hours ago
              Well I mean if I wanted to be extra pedantic, I would argue that we&#x27;ve been in that phase since LLMs were first introduced.<p>Before that, we had 0. After that, we had more than 1.<p>A leap as far as that is hard to recreate.<p>But that wasn&#x27;t my point. That&#x27;s just trolling.<p>The actual point is that LLMs aren&#x27;t gaining new capabilities anymore. They just get more reliable at the ones they already have; turning what was a coin flip to some higher probability.<p>That&#x27;s (intuitively speaking, not strictly mathematically speaking) kinda the mathematical definition of diminishing returns.
          • gehsty21 hours ago
            It’s a constant tension in computing that has been around since mainframes and clients… Neither is going to disappear. My general feeling is normal people care more about how thin and light something is than their privacy, so if data center powered LLMs will have a strong future.
            • hypfer21 hours ago
              Hmm I&#x27;m not 100% sure about that, given that edge is very viable, and the geopolitical climate has changed quite significantly.<p>I agree that datacenters are not going to go away, but I have doubts that the buildup that has happened is really going to pay off for most operators.
          • 4848844821 hours ago
            they really dont want to hear this bro lol
            • hypfer21 hours ago
              I can see that by those reddit-style vote swings, but who are &quot;they&quot;, exactly?<p>Who is so emotionally invested into random comment sections being purely positive about their pet.. uuuuuuuh.. tech?<p>Very weird.
        • dabinat15 hours ago
          You don’t need frontier-level performance for every task. That’s why companies hire both junior and senior developers. I suspect a decent percentage of people using Fable would probably be fine with Opus.<p>Also, the frontier can’t keep advancing at this rate forever. Eventually the low-hanging fruit will all be gone and advances will slow down.
        • blurbleblurble21 hours ago
          It&#x27;s going to happen very soon, which is why these frontier labs are scrambling to shut down open source language models. There&#x27;s an existential risk threatening their obscene returns.
          • mrlonglong20 hours ago
            It is for that reason they are being archived and torrented as a very large middle finger.
          • usef-18 hours ago
            I feel like cost competitiveness of local has been going down, if anything, not up. API providers can use hardware more and have scale efficiencies. Do you see any reason this will reverse?
          • IshKebab20 hours ago
            &gt; It&#x27;s going to happen very soon<p>Why? You can&#x27;t just assert it. There are very good reasons to think it <i>won&#x27;t</i> happen soon, and you&#x27;ve given no reasons to think it will happen soon.
            • blurbleblurble18 hours ago
              Because everything is converging on a backlog of huge efficiency gains established in research, waiting to be combined. Looped transformers, a whole host of diffusion techniques and new quantization techniques, maturation of ternary distillation and new ways to separate logic from stuff that can be looked up. It would surprise me if most frontier models were actually even that big at that point in terms of active params. I highly doubt it.
            • dwedge20 hours ago
              You asserted that it was never going to happen first
              • geysersam18 hours ago
                But he gave a reason for that. &quot;Before that happens the frontier will move&quot;. Why do you think it will happen anyway? Do you think the frontier will not move fast enough that local models are unable to catch up, or do you think people will prefer local models at a point. Or something else?
      • EA-316722 hours ago
        Or even more amusing in time for the bubble to burst, I hope the big RAM makers end up holding the whole bag for that. Greedy bastards.
        • glub19 hours ago
          Nearly the entire reason we&#x27;re in this mess is because RAM&#x2F;SSD&#x2F;HDD makers are terrified of AI bubble bursting.<p>Long-term contracts, not building new fabs - they&#x27;re in full on hedging mode right now.
          • leoc17 hours ago
            Seemingly the AI bulls are either not sufficiently optimistic, or not sufficiently wealthy(?!), to build new fabs as joint ventures with the manufacturers in which they agree to assume most of the downside risk?
            • no-name-here15 hours ago
              They signed long term contracts for the hardware delivery over years. AI investment is approaching 1 trillion per year and expected to grow to well over 1 trillion per year. [1]<p>Even over their many years, projects like the Manhattan Project, the Apollo Program, or the U.S. Interstate Highway System never added up to that. [2]<p>Is your argument that there is insufficient AI spending at present as they aren’t also taking on building their own fabs?<p>[1] <a href="https:&#x2F;&#x2F;www.pwc.com&#x2F;gx&#x2F;en&#x2F;news-room&#x2F;press-releases&#x2F;2026&#x2F;global-investment-in-ai-infrastructure.html" rel="nofollow">https:&#x2F;&#x2F;www.pwc.com&#x2F;gx&#x2F;en&#x2F;news-room&#x2F;press-releases&#x2F;2026&#x2F;glob...</a><p>[2] <a href="https:&#x2F;&#x2F;www.aljazeera.com&#x2F;news&#x2F;2026&#x2F;2&#x2F;19&#x2F;visualising-ai-spending-how-does-it-compare-with-historys-mega-projects" rel="nofollow">https:&#x2F;&#x2F;www.aljazeera.com&#x2F;news&#x2F;2026&#x2F;2&#x2F;19&#x2F;visualising-ai-spen...</a>
          • m4rtink16 hours ago
            If the long term contact are with any AI companies, then good luck getting anything back in bankrupcy proceedings!
        • ThrowawayR213 hours ago
          If the AI bubble bursts, the DRAM manufacturers switch back to making conventional DRAM and pent-up business and consumer demand results in a surge of purchases, cushioning the fall considerably.
        • m4rtink16 hours ago
          Looking forward to see them burn.
        • bethekidyouwant22 hours ago
          Since you’re not greedy when there is a memory glut I’m sure you’ll be willing to pay extra to make it fair.
          • EA-316721 hours ago
            Not at all, I’m going to relax with a cooling drink and watch the consequences of this unbelievably destructive and wasteful venture implode. I also adore the idea that it&#x27;s on the customer to be &quot;fair&quot; to a company that dumped us a group in favor of chasing B2B money.<p>Their choice, their consequences.
            • usef-18 hours ago
              I think they&#x27;re talking about the past: memory has often been boom&#x2F;bust. Memory manufacturers lose money during gluts, and did so long before AI. People benefited from cheap prices at that time.
            • pixl9721 hours ago
              Internet is a fad, it will implode at any moment.
              • EA-316721 hours ago
                In fact as a business opportunity it did just that because the massive investment in the mid-late 1990&#x27;s was without anything like a connection to profitability or sustainable demand. It took years after that crash for the concept of e-commerce to begin both a recovery and evolution into what we see today.<p>I expect something similar for AI. It&#x27;s useful tech... just not very profitable tech, and certainly not to the tune of trillions of dollars worth of public demand. The promises of superintelligence, replacing everyone, and the rest of the hype will die with the companies who made the promises, but the tech will survive and thrive.
                • pixl9721 hours ago
                  &gt;The promises of superintelligence, replacing everyone, and the rest of the hype will die with the companies who made the promises<p>The hype will die but there are many reasons why super intelligence and robots are a separate entity from said hype.<p>Look, back a few decades and tell someone our gdp would be in the trillions and it&#x27;s likely they&#x27;d have a hard time believing you. A huge portion of our products that we use day to day would be complete science fiction to them.<p>All we&#x27;re negotiating at this point is the timescale it will take.
                  • card_zero20 hours ago
                    Unpredictable things happen routinely, therefore these specific unpredictable things will happen?
                    • pixl9719 hours ago
                      These things aren&#x27;t really unpredictable. Work long enough on a robot with more dexterity and you will get just that. The timetable on when it happens is a bit more up in the air, economic downturns and wars can have huge impacts on it.<p>But in my mind robots and superintelligence are inevitable unless there is a massive setback in humanity before then. Nature already did it once. We&#x27;re not inventing something totally new. Add in every new invention and bit of intelligence we build on and actualize pushes us that much closer to the goal. Information technology allows this to speed up even further with easy sharing of information and testing.<p>Superintelligence is not like faster than light travel. We have many frameworks that show us FTL is impossible. I don&#x27;t believe there is a single widely accepted framework that shows there is some limit to intelligence and we&#x27;re near it. Intelligence isn&#x27;t even a singular thing. My calculator is a super intelligent adder compared to me. It would seem highly improbable that somehow nature random walked into making brains the most efficient general intelligence device in all dimensions and scales.
    • gs1722 hours ago
      Samsung is already building another fab (which was originally suspended due to <i>low</i> memory prices years ago). Hopefully when it goes live in a few years it&#x27;s not all dedicated to HBM for AI.
    • heaney-55522 hours ago
      Not even close.
    • nutjob222 hours ago
      Not until the industry gets severe indigestion, which doesn&#x27;t seem that far off.
  • tugback8 hours ago
    Good to see Samsung stepping up; maybe GPU prices will finally cool down a bit with more HBM supply hitting the market.
  • chr15m17 hours ago
    A shortage is almost always followed by a glut. Memory of all kinds is going to get very cheap due to supply and demand and substitution effects.
    • cute_boi17 hours ago
      Not when companies have monopolies.
      • chr15m17 hours ago
        There is no monopoly on computer memory.<p>Yes, only a small number of companies produce memory because of capital requirements and complexity. When prices go up both of those things become less of a problem for investors who can see better return for lower risk on that high capital investment. They will then invest in production to capture those returns. This is literally the exact reason Samsung are doubling supply of this particular type of memory.<p>High prices will fix the shortage.
        • 0x45712 hours ago
          Yeah, its not like companies ever colluded to keep prices high...<p><a href="https:&#x2F;&#x2F;en.wikipedia.org&#x2F;wiki&#x2F;DRAM_industry_price_fixing?useskin=vector" rel="nofollow">https:&#x2F;&#x2F;en.wikipedia.org&#x2F;wiki&#x2F;DRAM_industry_price_fixing?use...</a>
        • bobmcnamara13 hours ago
          &gt; There is no monopoly on computer memory.<p>Never has been, but what makes you think the major players won&#x27;t collude to keep prices high?
        • ajnin15 hours ago
          Memory factories cost billions to tens of billions, &quot;investors&quot; aren&#x27;t going to spend that kind of money if they are not reasonably certain that they are going to make a profit, and the market is very uncertain right now with everybody bracing for a bubble burst at any moment. It&#x27;s not the first time there has been such a shortage in the computer parts market, and the oligopoly, be it RAM or hard drives manufacturers, has always managed to keep prices high and production low enough to avoid a price crash after recovery.
          • chr15m14 hours ago
            &gt; hard drives manufacturers, has always managed to keep prices high<p>No. Please take a look at the unit price for hard drive storage 10 years ago as a percentage of what it is today.
  • Taniwha18 hours ago
    You just know that the AI bubble is going to break and they&#x27;re going to be stuck with tons of HBM4s.<p>The big problem is that this will not depress the consumer DRAM market and mean we can all afford DRAM again
    • m4rtink3 hours ago
      I&#x27;m sure you can make it work with HBM chips on adapter boards for PC. And there should be lots of HBM chips from scavengers taking apart bankrupt data centers.
  • dalton749 hours ago
    Good, maybe this means cloud GPU prices will <i>eventually</i> normalize. Still feeling the pinch from the AI gold rush.
  • m4rtink16 hours ago
    This insane nonsence needs to stop.
  • xbmcuser8 hours ago
    To me, the most idiotic thing is that the world currently cannot build power plants fast enough to even use all the HBM that is being produced. That is, a large portion of HBM and cards are sitting in storage, waiting for the data centers and the power plants to run them. The FOMO in the AI world isn&#x27;t just in the stock market
  • phooenixIam11 hours ago
    This is crazy!
  • badgersnake21 hours ago
    Now make some DDR5.
    • pixl9721 hours ago
      HBM is DDR5 in a different package.
      • badgersnake20 hours ago
        If it doesn’t fit it’s no use to me.
  • KurSix11 hours ago
    [dead]
  • sehw11 hours ago
    [dead]
  • psyphy222 hours ago
    [flagged]
    • zahlman22 hours ago
      Why wouldn&#x27;t it get posted here?
    • esafak21 hours ago
      Because you are not familiar with the site.
  • blurbleblurble21 hours ago
    The beginning of the bubble pop?
    • blovescoffee20 hours ago
      how would you even derive that conclusion from the headline&#x2F;article?
      • blurbleblurble11 hours ago
        <a href="https:&#x2F;&#x2F;youtu.be&#x2F;k_gaZjXD5OY?t=200" rel="nofollow">https:&#x2F;&#x2F;youtu.be&#x2F;k_gaZjXD5OY?t=200</a>
        • esskay9 hours ago
          so you fell for clickbait.
    • esskay21 hours ago
      its ram for the types of systems ai needs in datacenters so...no?
  • shevy-java22 hours ago
    About 2 years ago I bought some SDRAM or something like that, DDR4 or DDR5, don&#x27;t recall offhand. A few months ago I looked at the price today, and it was over 3x as high. That&#x27;s just insane. Governments need to do something against this abuse system that AI amplified here.
    • digdigdag22 hours ago
      &gt; Governments need to do something against this abuse system that AI amplified here.<p>Which governments, and what exactly would you want these governments to do? The demand is global. There are no levers a single government can pull to meaningfully influence the global demand without fully committing to an protectionist economic policy, in which case the U.S. doesn&#x27;t have the facilities to magically pop up world class fabs overnight, and South Korea and Taiwan don&#x27;t have the market demand that the U.S. generates to justify their investments in making these chips and China lacks the IP to be able to build anything comparable to Nvidia&#x27;s silicon at the moment.<p>No one has all the cards and no one controls all the levers.
      • carlm4222 hours ago
        Governments have sued DRAM companies for price fixing and forming cartels before: <a href="https:&#x2F;&#x2F;en.wikipedia.org&#x2F;wiki&#x2F;DRAM_industry_price_fixing" rel="nofollow">https:&#x2F;&#x2F;en.wikipedia.org&#x2F;wiki&#x2F;DRAM_industry_price_fixing</a>. They can absolutely step in and if there are evidence of price fixing or anti-competitive behaviours, can impose fines. The choice of whether to do that though is political.
        • SideQuark13 hours ago
          And on that page govts have tried and failed at it also. Thinking the current prices are the result of fixing and not insane demand is surely going to lose in court. It’s trivial to show massive demand, and thus production line changes, as the driving forces here.<p>If you look over all such lawsuits, they extremely rarely succeed.
        • daedrdev21 hours ago
          Several of these companies almost went under before the AI boom. There isn’t a cartel.
        • lovich21 hours ago
          Is there any allegations of price fixing in this case? The demand just skyrocketed. Other than legislation like the defense act in the US I don’t know of any governments that have jurisdiction over these ram makers that mandate companies produce more of a product.<p>Especially not for consumer goods.
          • wmf20 hours ago
            If they&#x27;re charging one customer $X and another customer $5X for the same chips... there&#x27;s got to be some law about this.
            • srdjanr19 hours ago
              I missed that this is happening, is it really? I thought it&#x27;s 5X for everyone
              • wmf19 hours ago
                Supposedly OpenAI is paying 2025 prices.
                • niltecedu8 hours ago
                  So you mean people with contracts are having their contracts honoured.... thats just normal business.
      • rtpg16 hours ago
        The same major memory makers got prosecuted for cartel behavior by the DoJ a little over 20 years ago[0].<p>One might say &quot;well what specific law have they broken&quot; and I don&#x27;t know and I don&#x27;t know if they have, but at the very least the power of the State _could_ be used to force behavioral changes in some ways or another.<p>[0]: <a href="https:&#x2F;&#x2F;en.wikipedia.org&#x2F;wiki&#x2F;DRAM_industry_price_fixing" rel="nofollow">https:&#x2F;&#x2F;en.wikipedia.org&#x2F;wiki&#x2F;DRAM_industry_price_fixing</a>
      • m4rtink16 hours ago
        The same governments that dumped bilions into to these companies and gave them insane tax breaks.
      • CamperBob222 hours ago
        <i>No one has all the cards and no one controls all the levers.</i><p>I&#x27;m as fanatical a free-market fundamentalist as you&#x27;ll ever meet, but if someone were to argue that government has a role in preventing bullshit like OpenAI&#x27;s unilateral 40% attack on the entire DRAM market, backed by nothing but funny money, I would have a hard time coming up with defensible counterarguments.
        • bunderbunder21 hours ago
          I don’t really know much about how this stuff works, but I have a feeling we wouldn’t even have to do anything punitive. We might just need to find and close whatever weird loophole allows the folks who are participating in this gold rush to feel confident enough that they’ve externalized their risks to be willing to engage in speculative data center buildout projects on such a grand scale in the first place.
        • rangestransform20 hours ago
          Apple used to do the same thing with TSMC’s newer process nodes, and nobody really complained because people were getting superior products. I happen to like US antitrust enforcement that considers consumer benefit primarily.
        • pfdietz21 hours ago
          Your attempts to insult what they are doing doesn&#x27;t show they did anything wrong, but it does impeach the validity of your own argument.
          • CamperBob220 hours ago
            Hit dog hollers, sounds like.
            • pfdietz20 hours ago
              That&#x27;s abuser logic.
              • CamperBob220 hours ago
                The funny thing is, we&#x27;re mostly in violent agreement, going by your other comments. This is a weird point to disagree on... so weird, in fact, that it&#x27;s easy to believe you are associated with OpenAI. Hence my follow-up comment.<p>What Altman did, and the way he did it, was not OK. Unless you believe that humanity&#x27;s adoption of AI-based intellectual work is best served by concentrating an enormous amount of power in one company at the expense of everyone else, turning the market into a literal zero-sum game. Do you?
      • newsclues22 hours ago
        How much memory does it take for the consumer market not to be totally screwed?<p>Can that share of production be allocated to consumers, and the AI fights over the rest?
        • m4rtink16 hours ago
          Well, thats kinda how strategic reserves work - so that say food manufacturers just don&#x27;t dump all the cheap food to sell crack instead &amp; everybody starves to death.
      • ozgrakkurt8 hours ago
        It is funny how regular people are trapped into legalism while companies are seeing things as a two sided fight.<p>You can’t demand anything with this mentality.<p>“The current situation is messing up the economy for a very large portion of us” is universally a good enough reason
    • pfdietz22 hours ago
      What abuse? Responding to supply and demand fluctuations isn&#x27;t abuse, it&#x27;s the proper operation of the market.
      • dumberquestions22 hours ago
        Would if one industry gobbles up an important resource to the point that other consumers lose practical access to it, shouldn&#x27;t there be limits?
        • jefftk22 hours ago
          No. The other industries can bid for access to the resource, and under most circumstances, whether they are willing to bid higher tracks importance. Sometimes there are cases where something has significant positive externalities and so people are not willing to bid high enough, and then we can talk about some kind of government intervention (typically a subsidy that attempts to track the value of the externalities) but I don&#x27;t see that applying here.
        • pfdietz22 hours ago
          If others are willing to pay more, that a signal that more value is created if the resource goes to them rather than to the more price sensitive customers.
          • applfanboysbgon21 hours ago
            This logic falls apart completely when gamblers allocate one trillion to bidding up the price, denying access to the resource to people who are responsibly spending their own money. The ideal world of the imaginary perfect free market never takes into consideration the messiness of the real world, like the fact that people will spend extreme sums of money irrationally. It can take years for this effect to correct, damaging the market severely in the meantime.<p>Alternatively, the money invested can be rational because it prices out competitors and establishes a monopoly, after which point the monopolist earns their absurd investment back with complete control of the market. This is <i>also</i> bad.
            • pfdietz21 hours ago
              You can reach any conclusion you want if you assume others are behaving irrationally. But why should I assume they are wrong instead of you being wrong?
            • pixl9721 hours ago
              This has zero to do with the company that is making the product and all to do with the company buying them.<p>The problem you have here is hundreds of different companies are doing the &quot;gambling&quot; in a non collusionary manner so it&#x27;s going to take decades in court to prove it. Are you saying governments should do an authoritarian take over of RAM allotment?
              • pfdietz20 hours ago
                And even if the companies are &quot;gambling&quot; (and what company doesn&#x27;t do that) it doesn&#x27;t mean the gambling is irrational.
      • joshheitzman21 hours ago
        Where is the additional supply resulting from the sustained increase in demand?
        • ErneX21 hours ago
          Some are building new fabs, example:<p><a href="https:&#x2F;&#x2F;news.skhynix.com&#x2F;en&#x2F;fab-facility-investment-2026&#x2F;" rel="nofollow">https:&#x2F;&#x2F;news.skhynix.com&#x2F;en&#x2F;fab-facility-investment-2026&#x2F;</a><p>It&#x27;s not an overnight thing obviously.
        • DonsDiscountGas20 hours ago
          Check the article we&#x27;re commenting under
          • Dylan1680714 hours ago
            No, this article is not about new fabs.
    • yread21 hours ago
      3x? I bought 128GB DDR5 SODIMMs for 300 eur. Now it&#x27;s 4000 eur.
      • mrlonglong20 hours ago
        Damn you were lucky. I bought 96GB DDR5 last year for £600 UKP. They&#x27;re now retailing for three times that.<p>Bastards.
        • yread5 hours ago
          Yeah the 300 already felt like a splurge - it&#x27;s the max. 96 was less than 200 I felt kinda stupid not just going with that. Imagine 96 of anything now for less than 200
    • jeroenhd21 hours ago
      It&#x27;s not just normal consumer memory. Everything with a chip in it is getting more expensive because of the AI circlejerk. When Bitcoin screwed over consumers, at least normal memory prices weren&#x27;t affected (though that time a shitcoin based around storing large amounts of data made the rounds, hard drive prices did take a brief hit).<p>Everything is affected by the endless thirst for RAM by AI companies and the accompanying hyperscalers. I&#x27;m honestly disappointed governments aren&#x27;t doing more to keep the price of consumer goods in check.<p>I guess the AI companies telling the world how good they are at accidental cybercrime and fraud is creating enough FUD to let the AI industry run wild.