The repaste worked. The VRAM got 12 °C hotter.
Follows on from Putting an RTX 3090 in an HP ProLiant ML350 Gen9. The card is a second-hand blower-style 3090 in an enterprise tower, power-capped to 275 W.
Read entryTag
2 observations with this tag
Follows on from Putting an RTX 3090 in an HP ProLiant ML350 Gen9. The card is a second-hand blower-style 3090 in an enterprise tower, power-capped to 275 W.
Read entryWe have been running low-bit quantization experiments on CPU only. It works, and it is slow enough that the feedback loop hurts — a context-depth sweep is most of an afternoon. The GPUs we already own are doing production work, so borrowing one was not an option. What we did have spare was an older dual-socket ML350 Gen9 tower with 512 GB of RAM, which is a good shape for MoE expert offload and had no GPU at all. So: a second-hand RTX 3090, blower-style, into an enterprise tower that was never designed for one.
Read entry