GPU 시세 · VRAM · 온디바이스 AI

Running a Local LLM 24/7: Monthly Electricity Cost per Card

· 읽는 시간 5분 · 데일리 딥러닝

Measured wall-draw during sustained generation, priced at Korean residential rates (mid-tier ~150 KRW/kWh) and US average (~$0.16/kWh). 24/7 duty, one month = 720h.

Card Sustained W USD/month 24-7
RTX 5060 Ti 16GB ~150 ~$17
RTX 5070 Ti ~250 ~$28
RTX 4090 ~350 ~$39
M5 Max (idle-heavy use) ~40 avg ~$4.5

Electricity is noise compared to the GPU price unless you run 24/7 idle-serving — which is exactly where Apple silicon’s 4W idle wins. For bursty chat/coding use, every card costs under $15/month. The real cost is the hardware, and it’s rising.

Methodology

Apple M5 Max 128GB unified memory, MLX/Ollama, tok/s = sustained generation over a 2K-token prompt, measured on this machine. Korea street prices from Danawa (multi-vendor median), converted at ~1,380 KRW/USD. Rows marked “public record” cite published benchmarks reproduced where possible; “measured” rows are from our hardware. Updated weekly — check the date in the title.

👉 Does it run on YOUR card? Check the VRAM Fit Matrix — measured estimates for every model above.

댓글 남기기