mysteryhumpf

joined 1 year ago
[–] mysteryhumpf@feddit.org -1 points 1 week ago (2 children)

You know nothing about the situation yet you judge.

[–] mysteryhumpf@feddit.org 16 points 1 week ago (9 children)

Yes it means you need help, not a murderbot

[–] mysteryhumpf@feddit.org 44 points 1 week ago (19 children)

None of us is dealing with a „full set of cards“ we all depend on others to be there for us in times of need and not direct us to kill ourselves. Did you read the article?

[–] mysteryhumpf@feddit.org 1 points 1 week ago

Both can be true at the same time

[–] mysteryhumpf@feddit.org 1 points 1 week ago (1 children)

If I buy a rig I might as well host it on the internet to get back some of the investment and sell the compute to others … wait a minute that’s cloud!

It’s always cheaper to have the same hardware serve multiple people than just one.

[–] mysteryhumpf@feddit.org 3 points 2 weeks ago (1 children)

I also hope that don’t get me wrong, but as I said: Waiting for the LLM agent to finish coding is currently a bottleneck in software development, they don’t pay high salaries for watching the AI code, they will prefer faster agents even if they are expensive, because they are not only paying the AI Company but also the software engineer overseeing them.

[–] mysteryhumpf@feddit.org 4 points 2 weeks ago (1 children)

I ran Gemma 4 31 B quantized so it fits in my RAM. The decoding speed was decent, but if you look at the newest models for example Gemini flash 3.5 they have a decoding speed of 280 token per second, they generate an entire page before my Mac locally generates a sentence.

[–] mysteryhumpf@feddit.org 3 points 2 weeks ago (3 children)

Yes ofc I ran Gemma 4 for example, but compare that to the speed of Gemini in the cloud the difference is massive.

[–] mysteryhumpf@feddit.org 3 points 3 weeks ago

Alternative stores exist on iOS, just nobody knows about them.

view more: next ›