@wholookshere@lemmy.blahaj.zone
2026-09-23 12:33 UTC
Replies (3)
-
@piccolo@sh.itjust.works 2026-09-23 13:43
NPUs are already a thing, specialized cpus for nerual networks. The problem is the models are just databases, and you need lots of fast memory in order to feed the NPUs data, and thats the current bottleneck.
-
@trebach@sh.itjust.works 2026-09-23 13:51
They aren’t comparable. GPS is deterministic and simple so as you said it could be reduced to an FPGA or ASIC. LLMs are partially matrix math and probabilities but they’re also more complex than that and require a large amount of RAM to run even the first time. Each query added to the context increases the RAM needed further.
-
@merc@sh.itjust.works 2026-09-23 18:03
LLMs run on matrix math and probabilities Yes, and the silicon to handle that math already exists. It’s the “GPUs” being pumped out by nVidia. Those are basically now highly specialized matrix math machines.