768GB of cheap Intel Optane DIMM memory sticks used to run 1-trillion-parameter LLM on a system with a single GPU — local Kimi K2.5 install achieved roughly 4 tokens per second
Running a One Trillion-Parameter LLM Locally on AMD Ryzen AI Max+ Cluster
China Trained a 1-Trillion-Parameter LLM Using Only Domestic Chips