[-] eager_eagle@lemmy.world 5 points 2 hours ago

well, that's just for spying. I still think ww2 was worse, when they would feed dogs under tanks, so that when deployed in the battlefield, the dogs with explosives strapped would get under enemy tanks and boom

[-] eager_eagle@lemmy.world 2 points 5 hours ago* (last edited 5 hours ago)

I don't know you, but it doesn't sound like a high standards issue to me, sounds like a lack of process. I've been a thorough reviewer before AI, at least thorough in the ways that mattered, not nitpicking formatting. And I'll tell you an automated reviewer today can catch more things before I have time to confirm the first item I find. There are still false positives, but it's still more thorough than I have time to be. Bc of that it really helps including a round of automated review before looping in a human, regardless of who/what wrote the code.

The main kind of review problems AI still struggles with are the project direction ones: "does it make sense to implement this/like this?", "should this be a new package instead?", and things involving tacit knowledge that often goes undocumented "last time we did this, someone had to access prod on a Sunday" - so that's what I focus my reviews on. And the other area is if you're writing UI code, whether it's a web app or a game, it'll also struggle to determine what "feels" good to use, so it'll need a human earlier in the loop.

[-] eager_eagle@lemmy.world 1 points 5 hours ago* (last edited 5 hours ago)

This kind of specification applies to the agent behavior or the agent behavior in a project. You don't write it on every conversation, you write/review it once for the project or agent and let the harness include it on every conversation. You can even ask the agent to scan the codebase and pull desirable patterns out of it for future sessions.

The entire point of using an agent is to not have to write everything yourself, so that when you write "implement feature X", the feature gets implemented in a way that makes sense in that project, is reviewed, refactored, and tested by agents, and it's good enough to bring in a human reviewer.

[-] eager_eagle@lemmy.world 1 points 1 day ago

See, that's the issue. People letting it "drive itself" get worse results. You should be the one holding the standards and guiding the model, otherwise you will be frustrated.

[-] eager_eagle@lemmy.world 8 points 1 day ago* (last edited 1 day ago)

browsers offload inactive tabs from memory anyway, unless you're actively switching among 15+ tabs, they won't be using as much ram

[-] eager_eagle@lemmy.world 0 points 1 day ago

My honest opinion is that it's bad because a lot of people using LLMs have no standards and push the first thing that seems to work. Be mad at who's at the driving wheel, not the car.

You absolutely can generate crap with agents/LLMs, and like a humans writing, the first draft will probably be subpar or maybe complete garbage. Every new session is a clean slate, that's why putting effort in the documents guiding it is so important.

[-] eager_eagle@lemmy.world 0 points 1 day ago* (last edited 1 day ago)

I don't think there was a lot of people working on the rewrite. Most PRs are from bots. Original estimates for a manual rewrite were a small team working for a year or so, which puts total costs over $1M. Even doubling the token cost estimates, it was still cheaper than doing it manually by a factor of 2x-3x

[-] eager_eagle@lemmy.world -3 points 1 day ago* (last edited 1 day ago)

I've seen plenty of code in my life, from humans and AI.

For the past year or so, these agents can code just fine most of the time, as long as they are given enough context (or have the tools to get it). Regardless of how many downvotes I get here, they really are capable of generating decent code. I'm sorry you couldn't make it work yet.

[-] eager_eagle@lemmy.world 2 points 1 day ago* (last edited 1 day ago)

It was higher sure, but anthropic also has some of the most expensive models out there. That cost could be 5x-10x less just by going with cheaper model providers (if one were to pay the API costs, not the case for bun).

bun 1.4 was released 3 weeks ago, btw

[-] eager_eagle@lemmy.world 116 points 3 days ago* (last edited 3 days ago)

wtf are they smoking at united lmao

are they planning a re-enactment for next year?

29
submitted 1 month ago by eager_eagle@lemmy.world to c/privacy@lemmy.ml
5
submitted 1 month ago by eager_eagle@lemmy.world to c/signal@lemmy.ml
1562

If you haven't seen it yet, we recently made the announcement that starting July 1, 2026, the price of "Jellyfin Premium+ One Super Unlimited (with Ads)" will increase to $0.00 USD*. There has been a lot of enthusiasm regarding charge backs, and we're simply blown away by the community's response.

As we've had a high volume of inquiries, I'd ask if you could please wait until I'm off the support email shift to reach out about this issue. I've attached our schedule so you'll know when it is safe to reach out.

Thanks, and happy streaming!

*Example price in USD. Exact pricing in other currencies may vary.

118
submitted 4 months ago* (last edited 4 months ago) by eager_eagle@lemmy.world to c/selfhosted@lemmy.world

Update your nginx instances

cross-posted from: https://lemmy.world/post/46851448


CVE - Common Vulnerabilities and Exposures system
RCE - Remote Code Execution
PoC - Proof of Concept

11
submitted 4 months ago* (last edited 4 months ago) by eager_eagle@lemmy.world to c/cybersecurity@sh.itjust.works
67
19
submitted 4 months ago by eager_eagle@lemmy.world to c/linux@lemmy.ml
21
132
74

Any experiences with a self-hosted assistant like the modern Google Assistant? Looking for something LLM-powered that is smarter than older assistants that would just try to call 3rd party tools directly and miss or misunderstand requests half of the time.

I'd like integration with a mobile app to use it from the phone and while driving. I see Home Assistant has an Android Auto integration. Has anyone used this, or another similar option? Any blatant limitations?

48
submitted 1 year ago* (last edited 1 year ago) by eager_eagle@lemmy.world to c/privacy@lemmy.ml

Another banger from Benn Jordan exposing a really concerning reality in the US.

61
view more: next ›

eager_eagle

0 post score
0 comment score
joined 3 years ago