this post was submitted on 21 Sep 2026
648 points (97.2% liked)

Microblog Memes

12204 readers
3650 users here now

A place to share screenshots of Microblog posts, whether from Mastodon, tumblr, ~~Twitter~~ X, KBin, Threads or elsewhere.

Created as an evolution of White People Twitter and other tweet-capture subreddits.

RULES:

  1. Your post must be a screen capture of a microblog-type post that includes the UI of the site it came from, preferably also including the avatar and username of the original poster. Including relevant comments made to the original post is encouraged.
  2. Your post, included comments, or your title/comment should include some kind of commentary or remark on the subject of the screen capture. Your title must include at least one word relevant to your post.
  3. You are encouraged to provide a link back to the source of your screen capture in the body of your post.
  4. Current politics and news are allowed, but discouraged. There MUST be some kind of human commentary/reaction included (either by the original poster or you). Just news articles or headlines will be deleted.
  5. Doctored posts/images and AI are allowed, but discouraged. You MUST indicate this in your post (even if you didn't originally know). If an image is found to be fabricated or edited in any way and it is not properly labeled, it will be deleted.
  6. Absolutely no NSFL content.
  7. Be nice. Don't take anything personally. Take political debates to the appropriate communities. Take personal disagreements & arguments to private messages.
  8. No advertising, brand promotion, or guerrilla marketing.

RELATED COMMUNITIES:

founded 3 years ago
MODERATORS
648
RatGPT (media.piefed.zip)
submitted 6 days ago* (last edited 6 days ago) by inari@piefed.zip to c/microblogmemes@lemmy.world
 
you are viewing a single comment's thread
view the rest of the comments
[–] LovableSidekick@lemmy.world 4 points 5 days ago (3 children)

Interesting that it knows it "crossed the line" but did it anyway.

[–] Bugamn@lemmy.zip 55 points 5 days ago (1 children)

It doesn't know anything, it just parroted a response that fits with what the user wrote. That's what LLMs do

[–] jj4211@lemmy.world 8 points 5 days ago (1 children)

Generally speaking, you can get an AI to admit fault and accept blame for something that never even happened.

It reacts to the operator input that something happened that warrants blame. So the natural continuation of that conversation is to acknowledge the situation, accept blame, and promise to do better. It generally doesn't model what did or did not happen, nor does the promise to do better mean anything. It does weigh in the context window that it wouldn't be a natural continuation to do that action after text saying that it won't happen, but can evaporate and doesn't carry over when the operator goes to a new context window, which of course is a crazy nuance for the rando user to understand.

[–] LovableSidekick@lemmy.world 1 points 5 days ago (1 children)

Hopefully most users get that a promise AI makes won't stay in force forever, but it should definitely last through the current conversation. I mean, I would call that a minimal performance standard, and if it's not true the AI shouldn't make promises in the first place.

I'm a little hazy on the part where accepting blame for something that didn't happen is a natrual continuation of a conversation. It's certainly not natural in human conversations. I would expect a well written and well trained AI to correct false assumptions or factual errors made by the user, not pander to them - unless it's been told never to contradict the user.

[–] jj4211@lemmy.world 1 points 5 days ago

On the promises, they will have some influence within the current conversation, maybe (everything is a bit wobbly). It is unlikely to repeat verbatim a step it did before that still has something in the context window correlated with bad operator input. It will have some success with things that resemble, but are not the same, but sometimes it can surprise by failing to recognize what a human would have considered roughly the same thing and doing it anyway. Or the 'promise' has evaporated out of the context window, and when called on it, it will take the fact that it previously promised not to do that as a fact even though it had lost the promise. So even if it says "I'm sorry I made the promise and broke it", it doesn't even mean it actually had the promise in "memory" when it broke it.

On the blame for something that didn't happen, it's a limitation where the model treats the narrative as reality. An instance is often incapable of evaluating facts without data about those facts one way or another. I have seen models "pass" my test by recognizing in their context that I'm dealing with an utterly non-agentic service, and thus it could not have been the case that it ever deleted files or dropped tables. Sometimes even those produce a narrative of what it did wrong (out of nothing), because narrative is the thing being generated, relationship with factual reality being a function of good correlation with narrative content. If a service is even potentially agentic, then it would tend to accept the operator account as factual, since it has no idea instance to instance if there's some other context where it did, and if it generated skepticism that it would ever generate such a mistake, that's more likely to piss off an operator.

So it incorporates the operator assessment as some 'factual' injection into the context, unless some objective source of truth is available to contradict the operator.