I've been testing NVIDIA GPU for local LLM just because I had a 3090 TI in my desktop, which also has 32GB System RAM... has anyone run into a situation where when using the system for LLM use only they needed more system RAM or a specific ratio to VRAM?
I'm not talking about Openclaw or another management platform that can eat system ram by doing other things...
context- debating dropping my 5090 system from 64gb ddr5 to 32gb
additional question... Has anyone compared other system components and how they affect LLM performance when utilizing NVIDIA GPU?
IE: Is my AMD 7900 CPU & desktop motherboard ram limitations a potential bottleneck for a 5090, and if so in which scenarios? or is it a 1 t\s kind of thing...
What about specific models ie: MOE vs dense -- is it noticeable to have a 8 channel RAM system vs consumer desktop, and assuming that increases performance > on MOE model vs dense? Noticeable for the faster RAM or not worth the huge expense? RTX 6000 relative to system replacement is very similar in price. Then again a new server-class system can handle >1 RTX6000 which FOR SURE has MUCH greater (and quicker) ROI than multiple 5090 systems for example.
For personal use, not such an issue but as we think of scaling this out to multiple users, and ROI all of these things start to matter, especially considering the power requirements and potential consolidation of larger systems vs many more medium\smaller ones.
If you don't have answers but experiences in comparisons please share
For home use, and niche image\video work it's hard to beat the Mac's power usage, and idle power draw!!
I'm not talking about Openclaw or another management platform that can eat system ram by doing other things...
context- debating dropping my 5090 system from 64gb ddr5 to 32gb
additional question... Has anyone compared other system components and how they affect LLM performance when utilizing NVIDIA GPU?
IE: Is my AMD 7900 CPU & desktop motherboard ram limitations a potential bottleneck for a 5090, and if so in which scenarios? or is it a 1 t\s kind of thing...
What about specific models ie: MOE vs dense -- is it noticeable to have a 8 channel RAM system vs consumer desktop, and assuming that increases performance > on MOE model vs dense? Noticeable for the faster RAM or not worth the huge expense? RTX 6000 relative to system replacement is very similar in price. Then again a new server-class system can handle >1 RTX6000 which FOR SURE has MUCH greater (and quicker) ROI than multiple 5090 systems for example.
For personal use, not such an issue but as we think of scaling this out to multiple users, and ROI all of these things start to matter, especially considering the power requirements and potential consolidation of larger systems vs many more medium\smaller ones.
If you don't have answers but experiences in comparisons please share
For home use, and niche image\video work it's hard to beat the Mac's power usage, and idle power draw!!



