this post was submitted on 23 Nov 2024
444 points (96.4% liked)
Technology
59594 readers
3330 users here now
This is a most excellent place for technology news and articles.
Our Rules
- Follow the lemmy.world rules.
- Only tech related content.
- Be excellent to each another!
- Mod approved content bots can post up to 10 articles per day.
- Threads asking for personal tech support may be deleted.
- Politics threads may be removed.
- No memes allowed as posts, OK to post as comments.
- Only approved bots from the list below, to ask if your bot can be added please contact us.
- Check for duplicates before posting, duplicates may be removed
Approved Bots
founded 1 year ago
MODERATORS
you are viewing a single comment's thread
view the rest of the comments
view the rest of the comments
I have read the comments here and all I understand from my small brain is that, because we are using bigger models which are online, for simple tasks, this huge unnecessary power consumption is happening.
So, can the on-device NPUs we are getting on flagship mobile phones solve these problems, as we can do most of those simple tasks offline on-device?
I’ve run an LLM on my desktop GPU and gotten decent results, albeit not nearly as good as what ChatGPT will get you.
Probably used less than 0.1Wh per response.
Yes, kind of… when those businesses making money out of the subscriptions are willing to ship with the OS for free which something only Apple has the luxury to do instead of OpenAI who doesn’t ship hardware or software (like Windows) beyond an app that’s less than 100MB. Servers would still be needed but not for general cases like help me solve this math or translation. Stable Diffusion or Flux is one example where you only need the connection to internet when downloading a certain model like you wouldn’t necessarily want to download every kind of game in the world when the intention is to play games arises.