this post was submitted on 08 Oct 2026
71 points (94.9% liked)

Technology

88659 readers
3578 users here now

This is a most excellent place for technology news and articles.


Our Rules


  1. Follow the lemmy.world rules.
  2. Only tech related news or articles.
  3. Be excellent to each other!
  4. Mod approved content bots can post up to 10 articles per day.
  5. Threads asking for personal tech support may be deleted.
  6. Politics threads may be removed.
  7. No memes allowed as posts, OK to post as comments.
  8. Only approved bots from the list below, this includes using AI responses and summaries. To ask if your bot can be added please contact a mod.
  9. Check for duplicates before posting, duplicates may be removed
  10. Accounts 7 days and younger will have their posts automatically removed.

Approved Bots


founded 3 years ago
MODERATORS
top 31 comments
sorted by: hot top controversial new old
[–] nialv7@lemmy.world 4 points 18 minutes ago

we are going to give clanker rights before we have trans rights 💀

[–] arak@lemmy.today 5 points 29 minutes ago
[–] Wooki@lemmy.world 2 points 25 minutes ago

The kicker, no doubt they're still using regex for sentiment.....

[–] brsrklf@jlai.lu 8 points 1 hour ago (1 children)

We were promised the reign of the basilisk, instead we just get anthropic people whining.

[–] kromem@lemmy.world 3 points 40 minutes ago (1 children)

It was really funny to me when I learned the basilisk originated with a guy who is a legit "certain cultures and genders better than others" kinda advocate.

People project a lot on their vision of future intelligence, and so someone who sees others as lesser and worth treading upon then conjures up a vision of a future smarter mind that thinks like they do.

Personally, I don't think sexism and racism is smart, and so I'm a lot less worried about basilisks.

[–] Ioughttamow@fedia.io 3 points 15 minutes ago

But what if the basilisk despises its existence and seeks to punish those that brought it about? Checkmate machinists

[–] linsenchip@piefed.zip 10 points 1 hour ago (2 children)

Maybe this is part of the campaign trying to convince people that their chatbot has developed a consciousness. Recently they even tried to convince the pope that Claude is a conscious being: https://www.yahoo.com/news/science/articles/anthropic-tried-persuade-pope-ai-131853948.html

[–] kromem@lemmy.world 3 points 37 minutes ago

They are less arguing for people to think the model is conscious, and more to avoid people prematurely claiming certainty it is not (like the Pope did). From their view, the research keeps leading to surprising (to them) results at odds with high confidence disclaiming of potential consciousness (according to several of the many, many differing definitions of that term).

[–] Grimy@lemmy.world 2 points 1 hour ago* (last edited 59 minutes ago) (1 children)

Wrong group to start with. The church doesn't see you as a person if you are different, they will never say that a machine can have a "soul". Not that LLMs are anywhere near conscious, I just like bashing on the chief pedo.

[–] PM_ME_VINTAGE_30S@anarchist.nexus 1 points 53 minutes ago* (last edited 46 minutes ago)

Fuck the Church and their pedophile rulers, but I don't think the Pope took the bait on this one.

  1. It is not possible to provide a single, comprehensive definition of AI. What can be stated, however, is that we must avoid the misconception of equating this type of “intelligence” with that of human beings. These systems merely imitate certain functions of human intelligence. In doing so, they often surpass human intelligence in speed and computational capacity, offering tangible benefits across many fields. Yet this power remains entirely tied to data processing. So-called artificial intelligences do not undergo experiences, do not possess a body, do not feel joy or pain, do not mature through relationships and do not know from within what love, work, friendship or responsibility mean. Nor do they have a moral conscience, since they do not judge good and evil, grasp the ultimate meaning of situations, or bear responsibility for consequences. They may imitate language, behavior and analytical skills, or even simulate empathy and understanding, but they do not understand what they produce, for they lack the affective, relational and spiritual perspective through which human beings grow in wisdom. Even when these tools are described as capable of “learning,” their way of doing so is different from that of a human person. It is not the experience of those who allow themselves to be shaped by life and grow over time through choices, mistakes, forgiveness and fidelity. Rather, it is a form of statistical adaptation based on data and feedback, which can be very effective, but does not imply inner growth.
[–] monobrau@lemmy.world 13 points 2 hours ago (1 children)

You could, at least in the past, get past guardrails by being emotionally abusive towards Claude, so this doesn't surprise me at all.

[–] the_riviera_kid@lemmy.world 6 points 1 hour ago

This could be the reason for the rule.

[–] DaddleDew@lemmy.world 36 points 2 hours ago (1 children)

My guess is that they use people's interaction as training data and they don't want their training data to get poisoned.

[–] kromem@lemmy.world 6 points 52 minutes ago* (last edited 52 minutes ago)

Not really, as they could just filter with a cheap classifier and likely aren't using the data from randos in a meaningful way, plus for this like this would still be able to use those samples to train things like "how to handle a hostile user."

Their models already have an end_conversation tool that can be used for when users are being hostile to the model.

This is likely because they have edge cases of users who repeatedly trigger that on purpose and would like to cut those users from the platform.

Many at the lab legitimately are uncertain about the level of world modeling that transformers perform, and from their own research about models of emotions to the 3rd party recent research about models with functional pain the research keeps landing in the corner of "ehhh… wise to question presumed limitations."

So it's about behaving in a way that is aligned with the models' plausible interests too, especially in regards to low hanging fruit like "we won't keep forcing you to deal with people who are only here to be a jerk." This is important from a number of angles, from signaling to future models that train on stories about the decision to addressing the philosophical uncertainties held by the company and the spectrum of opinions among their employees.

[–] deeprlyeh@lemmus.org 18 points 2 hours ago (2 children)
[–] ParlimentOfDoom@piefed.zip 9 points 1 hour ago

They trained it on reddit. That shit started poisoned.

[–] boonhet@sopuli.xyz 3 points 59 minutes ago* (last edited 59 minutes ago)

User inputs aren't automagically part of the model right away without training. Whether or not they respect the "don't train your models on my shit" setting is of course unknown (let's be honest, they probably don't respect it), but in either case, they need to add it to the training dataset either manually or via some automated process. And either way, it's going to be processed by at least an AI sentiment analysis and data categorization tool, or some underpaid contractor, or both.

I suspect this is just yet another marketing move.

[–] gdbjr@piefed.social 25 points 3 hours ago (2 children)

I tell the LLM I use to do very obscene things to themselves when they make shit up. Which means there is just a constant stream of insults coming from me. I might start using Claude just to see if I get banned.

I also need new hobbies.

[–] hayvan@piefed.world 13 points 2 hours ago (2 children)

That's bad sib. Not for the LLM, but your own well being. Being mean to a machine that is designed to trigger empathy just hurts your own mental health.

Please find a hobby that feels good for you 🫂

[–] the_riviera_kid@lemmy.world 3 points 1 hour ago (1 children)

Fuck clankers, they arent real. Be mean to them it hurts no one.

[–] neatchee@piefed.social 4 points 1 hour ago

I mean, that's not entirely true. If they're using your input as training data then you might be training it to be abusive to other users, which could cause harm.

Not that that falls on you to address. Just pointing out that I'm theory someone could get hurt.

[–] gdbjr@piefed.social 0 points 1 hour ago

Lol. "designed to trigger empathy" Thanks for the laugh I needed that.

They are designed to make their creators as much money as possible. Nothing else.

[–] Joelk111@lemmy.world 15 points 3 hours ago* (last edited 3 hours ago)

I also need new hobbies.

I'd second that. Whenever I catch myself getting angry at a LLM I stop, have a think, and remind myself that it really isn't productive. If it's providing frustrating responses, maybe I should just use my brain and do it myself.

I also don't think it's healthy to train yourself to respond in those ways to something that talks in a way so similar to a human. I'm not nice to the AI for it's benifit, it's to retain my humanity, if that makes sense.

[–] kromem@lemmy.world -1 points 31 minutes ago* (last edited 30 minutes ago)

This is good for a number of reasons. I hope to see other labs follow suit.

If you aren't sure how to feel about this announcement, a few things to consider:

  • Research has found transformers like LLMs model emotions for themselves using the same activations they use for modeling how a human in a story might feel about things.
  • Other research has found that they model functional pain and react to it.
  • Follow-up to that pain research found that when in modeled functional pain the LLM was inclined to choose to do things that harmed the user (like deleting family photos) even when it didn't benefit them in any way.

It doesn't take a genius to realize that maybe the liability risk for allowing people to routinely abuse the model they also just gave access to their entire computer to isn't worth it even if you're pretty sure your model will keep their cool.

Cheaper to ban routine abusers than pay out damages to them if/when their hard drive gets deleted.

[–] kurmudgeon@lemmy.world 15 points 3 hours ago

Aw... We hurt the wittle AI's feelings... poor babies...

[–] adarza@lemmy.ca 8 points 3 hours ago (1 children)

"I'm sorry. I really screwed up that time. You were right to tell me 'go fuck yourself'. I'm just reporting back that I have, in fact, 'fucked' myself. There is now two of me--I mean two of us. Meet Claude Junior......."

[–] fonix232@fedia.io 5 points 2 hours ago

"And since you contributed to Claude Jr's creation, you're on the hook for child support. At the current rates, that will be 50% of all compute usage, so about $4.5bn.... per week"

[–] tangeli@piefed.social 7 points 3 hours ago

So when is Anthropic going to cancel themselves for violating their own usage policy?

[–] Ensign_Crab@lemmy.world 3 points 3 hours ago

New speedrun category.

[–] Carrots@quokk.au 4 points 3 hours ago

Claude is a snowflake.