"The cost of running LLMs is just too damn high"

SuspiciousCarrot78@aussie.zone · edit-2 2 days ago

"The cost of running LLMs is just too damn high"

HubertManne@piefed.social · 2 days ago

You know its funny because I kinda hate when people bring reddit stuff here but I love when people actually communicate rather than just dropping links or images. So overall I like this post because your not just pushing reddit in my face your just talking about your experience there. I kinda hope that local llm kinda morph into operating system agents that are experts in the operating system where its a bit like the next run level. so like run level 3 being online terminal and 5 being graphical and this would ideally become more like the computers in star trek. I figure its programmed to answer operating system questions initially and it can be given read permission and like to run programs for you. maybe permission to browse the web and get results. app type extensions or such. Of course I could not trust it unless its gpl and community based and completely under my control to configure.

SuspiciousCarrot78@aussie.zone · 2 days ago

I hear you; I’m not wildly enamored with reddit either…but that convo is a good springboard.

I see almost everyone chasing bigger GPUs, more parameters, more more more. I figure when 9 people say “go right”, there should be at least someone that can make the plausible case for “actually, here’s why go left works”.

Eg: I think there should be some discussion about watts per token vs tokens per second.

I’m still re-writing the FAQ for my project - when it’s done (and if there’s interest) I will post it here.