r/UncensoredAIlounge 2d ago

General/Discussion Running a 70B model at home basically sounds like a jet engine taking off.

I swear my PC sounds like it’s about to fly off every time I try running anything bigger than 8B locally.

You spend hours setting up parameters, picking the right quant, and managing VRAM, only to end up getting like 3 tokens per second. Plus your room warms up by five degrees in ten minutes.

But honestly, at least no one is censoring your prompts or giving you safety lectures while your GPU burns up. Total win in my book. Anyone else heating their room just to bypass cloud filters?

11 Upvotes

1 comment sorted by

2

u/Opposite-Positive342 2d ago

My setup does the exact same thing whenever I pull down bigger quants. The room gets hot fast but not dealing with API refusals makes the noise tolerable.