

4·
1 month agoHow are you running a 34B model without a GPU? You must be getting one token an hour! How much RAM do you have in the LLM box?
Don’t algorithm me bro.


How are you running a 34B model without a GPU? You must be getting one token an hour! How much RAM do you have in the LLM box?


Oh boy, do I have bad news about 90% of the internet for you…
So dumped into boxes, shoved into the attic and forgotten, got it thanks 👍.