Would an $1,800 rig ever pay for itself against just paying per token?
I use models about two hours a day for code and for summarising documents I would rather not upload anywhere, and my API bill has been running around $40 a month. I have a 3060 12GB in my desktop now, which handles small models fine and falls over on anything I actually want. Before I spend $1,800 on a used 3090 build, has anyone done this arithmetic honestly, including the electricity and the part where you stop using it after two months?
@semver_sid · 3w ago
Hybrid, and I have not gone back. Local model does the bulk, private and repetitive work where being merely good is fine, and anything hard or customer-facing goes to an API. My API bill dropped by roughly two thirds rather than to zero, and the local box earns its keep on the documents I genuinely cannot send anywhere. Trying to make one of the two do everything is what makes people unhappy with both.
Reply
Report