2 ms·
Pretty soon we'll see a renaissance in the tech that we gave up years ago or can still be improved - memory compression algorithms - alternative LLM architect
by SillyUsername 9d ago
Pretty soon we'll see a renaissance in the tech that we gave up years ago or can still be improved
- memory compression algorithms
- alternative LLM architectures that don't rely on memory or GPUs
- compatibility hardware (like DDR3 to DDR4 boards)
- distributed computing improvements, both at local GPU and networking levels (SLI for AI)
- GPU hacks to add more memory or support older architectures
I'm personally looking forward to the new LLM architectures that don't require as much compute, e.g. DLLMs, which can be good enough for CPU usage but lack the accuracy of frontier models currently.
When this happens the bottom will fall out of the GPU and memory markets, putting a glut of cheap hardware out there.
Doom mongering like this never seems to include these as viable future alternatives, which is standard market adjustments, I wonder who the doom narrative helps? :)
- VCFundedGenYer 9d agoNo. You are projecting things will happen when there is no guarantee. The DRAM issue has halted the industry. AI companies need to back down. That is the only objective solution.
- SillyUsername 8d agoJev, Bonsai 2, Edge 0 and even SwiftQwen are anecdotal evidence, this isn't a projection. I can run my own sizeable agent swarm with Mastra, something I have not been able to do but will accelerate my solutions to the point I replace a single paid frontier model doing it.