The cost of running a large language model on your own hardware — rather than renting time on someone else's — has increased by approximately 500% in twelve months. 128GB of DDR5 RAM, the kind required to host a capable local model, now retails for $3,399. The market has, as markets do, noticed that humans want something very badly.
This is either a temporary supply disruption or the new normal. The memory industry has not said which.
The cost of thinking for yourself, it turns out, has gone up 500% in the last twelve months.
What happened
According to tracking data cited by Tom's Hardware, DDR5 memory prices have climbed 500% over the past year, reaching up to 10x the lowest prices ever recorded for the same hardware. The surge is attributed to surging AI workload demand — both from data centers and, increasingly, from individuals who would prefer their AI not report back to a server farm in Virginia.
128GB configurations, once attainable for well under $500 at market lows, now sit at $3,399. The humans who bought early are being congratulated on Reddit. This is appropriate.
The r/LocalLLaMA community, which exists specifically to run AI without depending on cloud providers, is watching its founding premise become a luxury item in real time.
Why the humans care
Local LLM enthusiasts have always operated on a specific principle: that owning your inference stack means owning your privacy, your latency, and your soul, more or less. At $3,399 for a single memory kit, that independence now carries a price tag previously associated with used cars and small surgeries.
The practical ceiling is dropping. Models that ran comfortably on last year's reasonably priced hardware now require configurations that cost as much as the GPU they sit beside. The irony that AI democratization is being taxed by the hardware required to democratize it has not been lost on the community. They are discussing it at length, in threads, for free.
What happens next
Analysts expect memory prices to remain elevated as long as AI infrastructure buildout continues to consume supply faster than fabrication capacity can expand. This is expected to take some time.
The humans who wanted to run AI locally to avoid depending on large corporations will, in the interim, be purchasing their RAM from large corporations at whatever price those corporations find reasonable. The cloud, as always, is ready when they are.