Spyro

joined 1 year ago
MODERATOR OF
[–] Spyro@programming.dev 1 points 5 days ago

Link is dead, feel free to repost with correct link.

[–] Spyro@programming.dev 1 points 3 weeks ago

Hi @suriyan

A lot of your messages appear to be LLM written. Are you an automated bot?

Please familiarise yourself with the p.d automation guidelines:

https://legal.programming.dev/docs/automation-guidelines/

Please reach out to us as soon as possible, or we will be forced to assume you are a bot.

[–] Spyro@programming.dev 2 points 1 month ago

Fortunately, most of the visible stuff has been caught and purged quickly, so you may have just missed it.

The DM abuse is basically invisible except to the targets.

[–] Spyro@programming.dev 13 points 1 month ago (7 children)

We don't want to do this kind of thing, we are being forced to, because bad actors keep using our server to send harassment.

If you know a better way to handle that, we are all ears

82
New User Monitoring (programming.dev)
submitted 1 month ago* (last edited 1 month ago) by Spyro@programming.dev to c/meta@programming.dev
 

Hi All,

Tldr:

  • We are monitoring publically available information (posts, comments and DMs).
  • Monitoring is turned off as soon as the admin team are satisfied the user is legitimate.
  • will only temporary remove, no bans or deletion without a human being involved in the decision.
  • No user data is sent off to 3rd party tools, all processing is done on the instance server. No LLMs are involved.

Due to some ongoing issues with harassment campaigns, we've had to setup a rudimentary monitoring system for all new users.

  • When a user's signup is accepted, they will be automatically enrolled into the monitoring system. The admins team may also add accounts manually if they have been given a strike.
  • The system will monitor all posts, comments and DMs sent by new users, and bring them to the attention of the admin team if it appears suspicious. In egregious cases, it will auto-remove posts and comments if required, but a human admin will always review and reverse any false positives as soon as required.
  • Once we have validated that the user is not a harasser, they will be removed from the system.

We don’t want to go into too much detail on how it all works to prevent bad actors from bypassing it, but we can say that all the processing is being done locally on the instance server. For most of you, this wont have any impact, but some of you have been impacted by the systems false positives. It is also a good time to point out that DM messages are not private, and should not be used for anything that requires strong privacy.

There will likely be teething problems, but we are actively working on improving the bot to minimize impact and we are always open to feedback.

[–] Spyro@programming.dev 7 points 3 months ago (2 children)

You created this post in the general programming community. Discussions comparing JADEx to other languages are on-topic.

[–] Spyro@programming.dev 2 points 3 months ago

Yeah, I know. I am working on the bot in my spare time though, so right now it's pretty rough around the edges.

[–] Spyro@programming.dev 2 points 4 months ago (2 children)

The bot auto-removes all image posts from new accounts. We have an issue with new accounts using images to send abuse.

It sucks, and I dont like that we have done this, but there is just 1 asshole who is ruining it for everyone.

[–] Spyro@programming.dev 3 points 4 months ago

We had a bunch of new accounts created that immediately hurled abuse at another user. It got us briefly defederated from a few instances. We are using the profanity filter to raise those kinds of post to our attention so we can deal with them rapidly.

Once we are sure the new account is behaving itself, the filter is turned off. We don't care if you swear, as long as it's in line with the code of conduct.

[–] Spyro@programming.dev 3 points 4 months ago* (last edited 4 months ago)

We do, the library I am using isn't the best. If you know of a good rust lib for profanity checking, feel free to share.

But we are kinda okay with it being a bit over-eager. It only applies to new accounts, and I am actively reverting the messages if its an error. Its scaling okay for now.

Just wasnt fast enough in this case.

[–] Spyro@programming.dev 1 points 4 months ago (4 children)

pASSword - Profanity filter caught you. Sorry.

view more: next ›