robsteranium

joined 1 year ago
[–] robsteranium@lemmy.world 3 points 1 week ago (1 children)

We spin up a VPS on demand then tear it down again once they've finished playing.

Some months we accrue so little use that Hetzner doesn't bother issuing an invoice!

We've shared the provisioning scripts on Codeberg if you want to try it ou: LuanParty.

[–] robsteranium@lemmy.world 6 points 2 weeks ago

Yeah the intelligence is still in the model. The promise of symbolic AI is about logic programming/ formal semantics not recursive loops.

To a large extent the idea has failed because it proved too hard to get non-experts to represent systems formally.

I still think there's potential value in a hybrid approach - e.g. get language models to do the representation then let them use formal reasoners/ verification instead of hallucinating.

[–] robsteranium@lemmy.world 7 points 2 weeks ago

That sounds about right to me! Trying to go full steam ahead for 8 hours a day isn't sustainable or even desirable.

I found Cal Newport's book Slow Productivity really helped me get some perspective on work. It has three main lessons: Do Fewer Things, Work at a Natural Pace and Obsess over Quality.

[–] robsteranium@lemmy.world -1 points 2 weeks ago

That's not necessarily true either. You can make inference deterministic by setting the RNG seed and having the model select the highest likelihood token/ sequence instead of choosing at random.

[–] robsteranium@lemmy.world 0 points 3 weeks ago (4 children)

That's demonstrably incorrect. Different instructions give different outcomes.

[–] robsteranium@lemmy.world 2 points 3 weeks ago

Perhaps not fine-tuning weights but they can certainly use feedback from PR comments to optimise guardrails etc. And it's definitely possible to have an agent experiment with that. And yes the Devs are the guinea pigs in that context!

[–] robsteranium@lemmy.world 1 points 3 weeks ago (6 children)

You don't need to train a foundation model from scratch. You can fine tune an existing coding model or only a LORA layer on a budget.

You could also just use the feedback to optimise system prompts, skills/ guidance or verification tools.

[–] robsteranium@lemmy.world 5 points 3 weeks ago (10 children)

I presume they're gathering this as training data and hope to later get rid of OP.

That's why cursor went for $60bn. Anyone can make a harness. The real value is that it gathers data on which changes are accepted/ integrated.

[–] robsteranium@lemmy.world 1 points 3 weeks ago

Can't believe the strength of opinion on this post! I assume it's people frustrated that you dare post a pun in a (non-American) dialect... or perhaps you've touched a raw nerve in Oldham!

[–] robsteranium@lemmy.world 1 points 3 weeks ago (1 children)

Ah I see. I've only played Mindustry (2D Java/ Android equivalent) and if I left that running overnight I'd wake up to complete devastation!

[–] robsteranium@lemmy.world 1 points 3 weeks ago (3 children)

Hmm... I was going off Wikipedia which says 2020 - I guess the earlier release was an alpha.

In any case having your factories running in the background (overnight?) would quickly bump up those hours. Not sure it'd make for a very interesting stream though!

[–] robsteranium@lemmy.world 2 points 3 weeks ago

I never really got into quake in the same way but I did enjoy some quake 2 mods. I'd gone to uni by then so didn't have as much free time to while away!

view more: next ›